aa.

DATA ENGINEER · MUNICH, GERMANY

Arpit Angraj.

I like figuring out how things work.
And making them work better.

I’m a data engineer with 6+ years of experience in consulting, manufacturing, and fintech. I’ve built a Databricks platform from scratch at AGCO and run Payla’s data platform as its sole data engineer.

I enjoy getting into the details—why a pipeline is slow, where a number comes from, or what someone actually needs from a report. Most days that means working with Python, SQL, Spark, and AWS, and talking to the people who use the data.

6+ years of experienceDatabricks certified
Arpit Angraj
01

Core skills

The tools I use to build, ship, and look after data systems.

Data platforms & processing01

Building batch and streaming pipelines with a focus on reliability, performance, and clear data models.

DatabricksApache SparkPySparkDelta LakePolarsPythonSQL
Streaming & integration02

Moving operational data into the lake through event streams, change data capture, and API integrations.

KafkaKafka ConnectAmazon MSKStructured StreamingPostgreSQLMongoDBREST APIs
Cloud & delivery03

Owning deployments, infrastructure, orchestration, and the operational work that keeps pipelines running.

AWSS3LambdaDockerKubernetesTerraformDagsterGitLab CI/CDGitHub ActionsAWS GlueRedshift
Quality & governance04

Making data trustworthy with validation, reconciliation, access controls, and production monitoring.

Unity CatalogData modelingReconciliationPytestGrafanaLoki
Applied AI & team delivery05

Building practical tools for engineering teams, running Databricks workshops, and working directly with finance, risk, and business stakeholders.

GitHub CopilotCustom agentsDatabricks GenieTechnical workshopsClient delivery
02

Experience

From consulting for manufacturing teams to owning a fintech data platform.

MAR 2025 — PRESENT

Munich, Germany

Data Engineer @ Payla Services GmbH

Running the data platform behind payments, reporting, and financial operations.

Since June 2025, I’ve owned the platform as the sole data engineer, working directly with leadership and the Finance, Risk, Integration, and Site Reliability Engineering teams.

What I worked on

  • Built PySpark and Polars pipelines for batch, streaming, change data capture, reconciliation, and near-real-time financial processing.
  • Delivered one-time and daily bank reporting files to AWS S3 using Dagster and Databricks.
  • Resolved production incidents across Amazon MSK, Kafka Connect, MongoDB change data capture, Databricks, and Kubernetes during migrations and reporting cycles.
  • Tuned Kubernetes CPU and memory allocation for more than 50 bronze tables to reduce infrastructure costs.
  • Automated releases with GitLab CI/CD and merge trains, publishing Python wheels to S3 and managing versioned deployments.
  • Built a GitHub Copilot custom agent to identify memory-heavy Databricks jobs and propose resource changes, changelog updates, and validation steps.
100+

Unit and integration tests

≈95%

Coverage across transformations, orchestration, and SDK integrations

DatabricksPySparkPolarsKafkaDagsterAWSKubernetesPytest

APR 2024 — FEB 2025

Marktoberdorf, Germany

Data Engineer @ AGCO GmbH

Building the Enterprise Data Hub on Databricks from scratch.

  • Helped migrate AWS-native workloads to Databricks on AWS, establishing Unity Catalog, access management, jobs, Git integration, Databricks Asset Bundles, and GitHub Actions.
  • Migrated and optimized Spark, SQL, Kafka, and Databricks workloads: improved reporting speed by 30%, reduced latency by 25%, and improved Unity Catalog query performance by up to 3×.
  • Proposed and delivered Salesforce ingestion with AWS AppFlow, replacing long-running custom REST API jobs in Glue.
  • Ran Databricks workshops for data engineers and business teams to support adoption of the new platform.
>90%

Reduction in Salesforce ingestion costs

Up to 3×

Improvement in Unity Catalog query performance

DatabricksUnity CatalogAWS AppFlowAWS GlueSparkGitHub Actions

DEC 2019 — MAR 2024

Infosys Ltd. · Munich, Germany

Data Engineering Consultant @ Infosys

Client: AGCO / Fendt, Marktoberdorf

Worked across client engineering and business teams to deliver cloud and data solutions. Translated requirements into user stories, estimated work, and delivered across Scrum sprints.

Improved Spark and SQL processing performance by 20% through caching, partitioning, and workflow tuning.

Technology AnalystJan 2023 – Mar 2024
Senior Systems EngineerJan 2022 – Dec 2022
Systems EngineerDec 2019 – Dec 2021
AWSSparkSQLConsultingScrumManufacturing
03

Education & credentials

JUL 2015 — JUL 2019

Bachelor of Engineering

Information Science

Visvesvaraya Technological University
Bengaluru, India

CERTIFICATION

Databricks Certified

Data Engineer Associate

English — fluent
German — A1, currently learning B1

04

What people say

Excerpts from recommendations shared by colleagues on LinkedIn. Spacing and punctuation lightly edited for readability.

View recommendations on LinkedIn ↗
05

Let’s talk

Have a role, a project, or a data problem in mind? Get in touch.