Ashwani Kumar
Data Engineer
Haryana, India
#OpenToWork
About
Data Engineer with nearly 5 years of experience designing cloud-based data solutions, high-throughput pipelines, and automated processing workflows.
Proven expertise in SQL, Python, and AWS to build scalable data ingestion frameworks, optimize database systems, and implement agentic AI tools that support business reporting and analytics layers.
What I'm looking for
Looking for Forward Deployed Engineering or Senior Data Engineering opportunities. I thrive in roles requiring deep backend optimization (Python, SQL, Redis, Docker/K8s), automated CI/CD pipelines, and scaling data integration engines. I am looking for a culture that values engineering velocity, architectural stability, and direct business results.
Experience
Data Engineer
Adeptmind.ai
May 2024 – May 2026
* Process Automation: Contributed to the design of an automated agentic AI scraping system using Langgraph for data collection, successfully reducing manual development time required for data gathering by 75%.
* Data Processing: Overhauled database validation scripts and backend workflows, shrinking data validation processing cycles from 48 hours to just 2 hours.
* Database Optimization: Resolved application memory issues and database slowdowns during heavy traffic spikes by refactoring pipelines with Python generators and deploying PgBouncer connection pooling.
* Infrastructure Support: Migrated internal monitoring utilities to ControlPlane microservices and configured automated build/test pipelines using GitHub Actions CI/CD.
PythonAgentic AIControlPlanePgBouncerGit
Python Developer
Turbolab Technologies Pvt. Ltd.
Sep 2021 – Apr 2024
* Data Ingestion: Architected a stable, multi-tenant data harvesting infrastructure that automated on-demand data feeds for multiple external business clients.
* Pipeline Management: Built and managed data pipelines processing over 3 million records across the Healthcare, E-commerce, and News sectors while maintaining strict data quality standards.
* Team Support: Developed shared data-cleaning and manipulation libraries using NumPy and Pandas, improving pipeline
development speed across the engineering team by 40%.
PythonNumPyPandasHTMLXML
Data Science Intern
The Sparks Foundation
Jan 2021 – Feb 2021
* Pipeline Development: Formulated automated web scraping extraction routines integrated with text-processing and sentiment
analysis workflows to ingest unstructured data arrays.
* Data Modeling: Developed and optimized a Decision Tree Classifier using Scikit-learn to systematically categorize structural damage patterns across regional datasets.
PythonScikit-learnpandasnumpy
Education
UIET, Kurukshetra University
Bachelor of Technology
2014 – 2018
Skills
CI/CDDatabase OptimizationAgentic AIData Processing PipelinesExcelGitHubGitRedisPgBouncerControlPlaneKubernetesDockerAWSXMLHTMLLanggraphCore JavaPythonSQL