Skip to main content
London, United Kingdom

Hi, I'm Vikramaditya.

Data Engineer,
Building Pipelines that Scale.

I build dependable data pipelines and analytics systems that turn complex data into clear, useful insights. I care about practical architecture, reliable delivery, and tools that help people make better decisions.

Vikramaditya Tatke - Data Engineer

Technical skills

A practical data engineering stack.

Technologies I use to build and operate data platforms, grouped by where they fit in the system.

  1. 01

    Ingestion & orchestration

    Batch jobs, event streams, vendor APIs, and object storage.

    • Python
    • Apache Airflow
    • Apache Kafka
    • AWS Lambda
    • AWS Glue
    • Amazon S3
    • REST APIs
  2. 02

    Transformation & compute

    Data models, validation, distributed processing, and files.

    • SQL
    • dbt
    • Apache Spark
    • Polars
    • Apache Parquet
  3. 03

    Databases & query engines

    Columnar analytics, relational workloads, and document data.

    • ClickHouse
    • PostgreSQL
    • MongoDB
    • DuckDB
  4. 04

    Analytics & applications

    Dashboards, internal tools, and browser-native analytics.

    • Power BI
    • Tableau
    • Streamlit
    • TypeScript
    • React
    • Next.js
    • ECharts

Platform & reliability

The layer underneath every workload.

  • AWS
  • Docker
  • Linux
  • GitHub Actions
  • Datadog
  • CI/CD
  • TDD

Certifications

Professional certifications validating expertise in data engineering and analytics technologies

All certifications are verifiable through their respective issuing organizations

Data Pipeline Stages

Follow the data as it flows from raw sources to transformed output. Click the 'Load' stage to query it!

Stage 1: Source

Raw Data Ingestion

Data loaded

Stage 2: Transform

SQL

Transform complete
Preparing data

💾 Idle (0%)

Data loads silently so analytics are instant when you need them

Featured Projects

Explore my data engineering portfolio

F1 ETL and Analysis

End-to-end ETL pipeline for Formula 1 race data with comprehensive analytics and visualizations. Built with modern data engineering practices.

PythonPandasAWSSQL
View on GitHub

Streamlit Admin Panel

Interactive admin dashboard built with Streamlit for data management and visualization. Features real-time updates and user authentication.

StreamlitPythonPostgreSQL
View on GitHub

Polars GPU Test

Performance benchmarking and testing suite for Polars GPU capabilities. Demonstrates GPU-accelerated data processing for large datasets.

PolarsCUDAPythonGPU
View on GitHub

MWAA Kafka Project

Managed Apache Airflow with Kafka integration for real-time data streaming. Orchestrates complex data workflows in the cloud.

AirflowKafkaAWS MWAAPython
View on GitHub

TCP Flink Kafka Connector

Custom TCP connector for Apache Flink and Kafka integration. Enables real-time data ingestion from TCP sources into Kafka streams.

Apache FlinkKafkaJavaTCP
View on GitHub

Coming Soon

Exciting new project in development. Stay tuned for updates on innovative data engineering solutions.

TBD

Get in Touch

Have a project in mind or want to discuss data engineering? Drop me a message.