Ajeet Kumar

Ajeet Kumar

Data Engineer with 4+ years of experience building scalable data platforms and distributed ETL/ELT pipelines

Bangalore, India

#OpenToWork

About

I am a Data Engineer with over 4 years of experience building scalable data platforms and distributed ETL/ELT pipelines using Python, SQL, PySpark, Hive, Trino, and AWS. Most recently at GlobalLogic India Private Limited, I have worked on data ingestion, transformation, and large-scale data processing across FinTech, Retail, and B2B SaaS domains. My work spans designing and enhancing distributed ETL pipelines, resolving application security vulnerabilities using SAST and secure coding practices, and delivering production pull requests with tools like Claude Code and Devin. I have built backend services and APIs with Python and AWS Lambda, optimized SQL queries for performance, and developed financial data workflows and migration frameworks using PostgreSQL, PySpark, and AWS infrastructure.

What I'm looking for

Looking for Data Engineer / Senior Data Engineer opportunities where I can leverage my 4+ years of experience in Python, SQL, PySpark, AWS, ETL/ELT, Big Data, data pipelines, data modeling, and cloud data platforms. Interested in building scalable data solutions, distributed data processing systems, and modern data platforms using technologies such as Snowflake and dbt. Open to opportunities in Data Engineering, Data Platform, and Big Data roles.

Experience

GlobalLogic India Private Limited

Data Engineer

GlobalLogic India Private Limited

Apr 2026 – Present

• Designed and enhanced distributed ETL pipelines using Python, PySpark, PyFlink, Hive, Trino, SQL, MySQL, and AWS S3, delivering end-to-end ingestion for the Ziplabs Leadership pipeline integrating external datasets into the Redux → MPD platform. • Resolved 20+ application security vulnerabilities including SQL Injection, Weak Hash, Weak Random, and Path Traversal across Python services using SAST, taint-flow analysis, secure coding practices, parameterized SQL, and cryptographically secure random generation. • Performed Software Composition Analysis (SCA) by assessing CVEs, validating exploitability, upgrading vulnerable dependencies, implementing automated data quality checks, and collaborating with Security and Data Engineering teams to resolve security findings. • Delivered 14+ production pull requests covering security remediation, dependency upgrades, parser enhancements, and ETL improvements while leveraging Claude Code, Devin, and Windsurf for development, debugging, code reviews, and technical documentation.

PythonPySparkAWS

Data Engineer

More Retail Pvt. Ltd.

Nov 2025 – Mar 2026

• Developed scalable backend services and REST APIs using Python, AWS Lambda, and Aurora PostgreSQL to support real-time retail operations. • Designed APIs and data models for the Store Control Dashboard, enabling real-time operational monitoring and business analytics. • Optimized backend processing and SQL queries, improving execution performance by nearly 40% through query and processing optimizations. • Integrated GraphQL, Aurora PostgreSQL, Amazon Redshift, and AWS services to build scalable backend and data solutions for retail operations.

PythonAWS LambdaPostgreSQL

Data Engineer

Nineleaps Technology Solutions Pvt. Ltd.

Jul 2022 – Nov 2025

• Client: Uber Technologies — Uber for Business Jan 2025 – Nov 2025 ∗ Built end-to-end data ingestion pipelines using Python, Pandas, SFTP, and Amazon S3 to ingest partner CSV datasets and process JSON-based configurations for employee transportation workflows. ∗ Developed Python, SQL, and Trino workflows to clean, transform, validate, and reconcile partner data into system-compatible formats, ensuring data quality and consistency. ∗ Processed transportation datasets through downstream business models to support cab location allocation, minimizing route deviation and improving operational efficiency across employee transportation services. ∗ Automated report generation and email delivery using Google Drive and Mail APIs, supporting production onboarding, troubleshooting, and launches across 10+ cities. • Client: Saison Omni — Loan Management System (FinTech) Jul 2022 – Jan 2025 ∗ Built scalable ETL pipelines and financial data workflows using Python, PySpark, SQL, PostgreSQL, and AWS for enterprise Loan Management Systems. ∗ Led large-scale financial data migration with automated data cleansing, validation, and reconciliation, achieving 99.93% migration accuracy. ∗ Designed data validation and partner onboarding frameworks to improve data quality, reporting accuracy, reconciliation, and month-end financial book closure. ∗ Optimized PostgreSQL queries, AWS infrastructure, and reporting workflows, improving system reliability and reducing manual effort across distributed financial systems.

PythonSQLPostgreSQL

Education

Indian Institute of Technology (ISM), Dhanbad

B.Tech · Electronics and Instrumentation Engineering

2018 – 2022

Skills

SnowflakeKafkaSparkData ReconciliationData QualityData ValidationDistributed ComputingData LakesData WarehousingData PipelinesETLData ModelingMySQLPostgreSQLAWSTrinoHivePySparkSQLPython

Languages

English (Professional working proficiency)Hindi (Native or bilingual proficiency)