Hi, I'm Satyajeet Dharmadhikari
ETL Developer - Data Engineer • VOIS
Passionate Data Engineer with 3+ years of hands-on experience designing robust ETL pipelines, high-volume data staging, and server-to-server legacy codebase migrations. Google Cloud Certified Professional Data Engineer and completed M.Tech in Cloud Computing at BITS Pilani.
Bridging Data Infrastructure & Cloud Intelligence
Transforming complex big data challenges into high-efficiency, reliable, and scalable automated pipelines.
I am an ETL Developer - Data Engineer based in Pune, India, currently driving enterprise data solutions as a Senior Executive at VOIS.
My expertise spans the entire data lifecycle: from high-throughput batch ingestion using Ab Initio and Teradata staging environments, to complex SQL query optimization, Unix/Linux server automation, and cloud modernization on Google Cloud Platform and AWS.
I recently spearheaded a crucial server-to-server legacy codebase migration, modernizing data pipelines with zero operational downtime. Completed my M.Tech in Cloud Computing at BITS Pilani in January 2026 while continually expanding my cloud-native lakehouse and distributed engineering capabilities.
Professional Data Engineer
Issued by Google Cloud • Verified on Credly
Demonstrates proven proficiency to design, build, operationalize, secure, and monitor data processing systems on Google Cloud Platform. Certified in architecting robust streaming & batch systems, big data analytics, and machine learning model integration.
Work Experience
Delivering resilient data systems and enterprise infrastructure for global telecom scale.
Senior Executive — ETL & Data Engineering
VOIS (Vodafone Intelligent Solutions)
- Core ETL Development: Architect and maintain production-grade data pipelines processing enterprise-scale telecommunications datasets using Ab Initio and Teradata.
- Server Migration Leadership: Successfully executed server-to-server legacy codebase migration, validating multi-terabyte data staging environments with zero data loss and minimal pipeline downtime.
- Performance Tuning & SQL Optimization: Optimized complex Teradata SQL queries, indexing, and data staging schemas, resulting in reduced batch execution windows.
- Pipeline Automation: Engineered custom Python and Unix Shell automation scripts for automated job dependency checks, data reconciliation, and alerting.
Graduate Engineer Trainee
VOIS (Vodafone Intelligent Solutions)
- ETL Unit Testing & Validation: Designed comprehensive ETL test suites and verified data transformations across staging and warehouse layers.
- Perl & Shell Scripting: Built automated test-validation scripts in Perl and Bash to automate file format verifications, delimiter checks, and log monitoring.
- Production Support & Issue Resolution: Monitored daily/weekly batch schedules on UNIX servers, analyzed failure logs, and delivered swift defect resolutions.
Skills & Architecture Matrix
Technologies and tools I leverage to build scalable, resilient data pipelines.
ETL & Data Pipelines
Designing, scheduling, and orchestrating massive batch & stream data flows.
- Ab Initio (GDE, Co>Operating System) Advanced
- Data Staging & Ingestion Expert
- Data Warehousing (EDW Architecture) Advanced
- Server-to-Server Codebase Migration Specialist
Cloud & Big Data
Cloud-native data architecture, managed services, and distributed storage.
- Google Cloud Platform (BigQuery, Storage) Certified
- Cloud Dataflow & Cloud Dataproc Proficient
- AWS Cloud Computing (BITS Pilani) Proficient
- Docker & Containerization Basics Intermediate
Databases & SQL
Writing performant queries, schema design, and analytical warehousing.
- Advanced SQL Query Optimization Expert
- Teradata Database & Utilities (BTEQ, FastLoad) Advanced
- Google BigQuery (Serverless Analytics) Advanced
- PostgreSQL & MySQL Proficient
Languages & Scripting
Automating workflows, data parsing, testing frameworks, and tooling.
- Python (Data Analytics, Scripting, Automation) Advanced
- Linux / UNIX Shell Scripting (Bash) Advanced
- Perl Scripting & Text Processing Proficient
- Core Java & C++ Foundation
Key Projects & Systems
Open-source streaming lakehouses, data lineage architectures, and mission-critical enterprise migrations.
Financial Market Lakehouse & Streaming Platform
Real-time financial market data platform architected for ultra-low latency analytics. Integrates Apache Beam streaming pipelines, PySpark distributed compute, Apache Iceberg table format via PyIceberg, Arrow Flight high-throughput transport, and embedded DuckDB analytics.
Data Lineage & Pipeline Tracking with PySpark
End-to-end data lineage, metadata collection, and observability system for PySpark pipelines. Implements OpenLineage standards to trace schema drift, dataset dependencies, and execution runtimes for enhanced pipeline governance.
Gold Price Prediction using Holt-Winters Exponential Smoothing
Comprehensive statistical time-series modeling project applying Holt-Winters Exponential Smoothing (HWES) on multi-year historical commodity data. Includes trend decomposition, seasonality modeling, hyperparameter optimization, and MAPE evaluation.
Amazon Reviews Sentiment Analysis Web Application
Natural Language Processing web service that processes unstructured e-commerce product reviews to classify customer sentiments into positive, neutral, and negative categories with text tokenization and TF-IDF feature extraction.
Server-to-Server Legacy ETL Codebase Migration
Spearheaded the migration of legacy Ab Initio ETL graphs, Unix wrapper scripts, and staging data across enterprise server clusters at VOIS. Designed automated checksum reconciliation to ensure 100% data fidelity with zero production downtime.
Automated Data Validation & Staging Reconciliation
Engineered a custom automated data reconciliation tool to validate millions of staging records against target warehouse tables. Replaced manual unit-testing routines with Perl and Python validation scripts, cutting staging cycle times by over 60%.
Education & Credentials
Solid foundations in computer science, software architecture, and modern cloud technologies.
Ready to build something impactful?
Whether you have an exciting data engineering role, a cloud lakehouse project, or want to discuss tech — let's connect!