Senior Data Engineer | AWS, Python, PySpark | Building scalable data pipelines and ETL solutions | Bangalore, India
- Bangalore, India
Pinned Loading
- account-360-analytics-platform
account-360-analytics-platform PublicEnterprise 5-layer Medallion pipeline on AWS | 88% ETL improvement (50 min → 6 min) | AWS Glue · PySpark · Redshift · Athena
- clinical-data-integration-platform
clinical-data-integration-platform PublicBi-directional CTMS integration: CluePoints RBQM ↔ Veeva (Parexel, PPD) | Serverless AWS pharma pipeline replacing manual SharePoint workflows
- enterprise-colleague-reporting-platform
enterprise-colleague-reporting-platform PublicEnterprise HR analytics pipeline | 500K+ daily records · 3,000+ stores · 99.9% reliability | Azure Databricks · Delta Lake · PySpark
- realtime-iot-vehicle-monitoring
realtime-iot-vehicle-monitoring PublicProduction-grade streaming pipeline for real-time vehicle fleet monitoring | Kafka · Flink SQL · Azure · PySpark · Sub-second latency
Python 1
- snowflake-dbt-retail-analytics
snowflake-dbt-retail-analytics PublicEnd-to-end retail analytics platform on Snowflake + dbt | Staging → Marts · SCD Type 2 · 45% query time reduction · GitHub Actions CI/CD
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
Uh oh!
There was an error while loading. Please reload this page.