Curriculum Vitae & Career Milestones
14+ years of proven leadership in enterprise data engineering, cloud-native Lakehouse platforms, and distributed distributed systems optimization.
SAGAR D. SHINGARE
Senior Data Engineer & Data Architect | Big Data Tech Lead
Professional Summary
Senior Data Engineer and Data Architect with overall 14+ years and 8+ years of end-to-end experience delivering high-scale data platforms, cloud-native ETL/ELT systems, and distributed processing solutions across Life Sciences, CPG, Retail, Insurance, BFSI, and Enterprise technology domains. Specialized in Spark (Scala/Python), Databricks, Hadoop ecosystem, BigQuery, AWS EMR, Delta Lake, and modern Lakehouse architectures. Adept at building 3 TB+ pipelines, improving workload efficiency by 60%+, and engineering reusable frameworks that strengthen data reliability, performance, and analytical readiness. Recognized for strong data modeling expertise (ERD/Dimensional), orchestration design, and cross-functional collaboration across complex enterprise environments.
Core Competencies & Architecture Toolkit
Big Data Ecosystem
Spark (Core, SQL, Streaming), Hive, HDFS, YARN, MapReduce, Sqoop
Cloud Platforms
AWS (EMR, S3, Lambda, Batch), GCP (BigQuery, Dataflow, Composer), Azure (ADF, Databricks)
Data Architecture
Delta Lake, Apache Iceberg, Medallion Architecture, CDC, SCD Type 2, Data Quality Frameworks
Programming & Languages
Scala, Python (PySpark), Advanced SQL, Core Java, Linux Shell Scripting
Data Modeling & Warehousing
Dimensional (Kimball/Inmon), ERD Design, Star/Snowflake Schemas, Snowflake, PostgreSQL, MySQL
Orchestration & FinOps
Apache Airflow, Cloud Composer, dbt, Git, Docker, CI/CD, Terraform, Cloud FinOps Cost Tuning
Professional Experience
Senior Big Data Tech Lead & Data Architect
Global IT & Financial Services Practice (BFSI)Leading enterprise BFSI data engineering delivery, architecting cloud-native ETL and scalable analytics platforms for premier banking and financial services clients. Driving high-throughput data pipelines, real-time transaction streaming, fraud detection pipelines, and regulatory reporting platforms under stringent SLAs.
Data Architect
Enterprise Data & AI PracticeArchitected and implemented enterprise-grade data engineering solutions for Life Sciences, Retail, Insurance, Consumer Tech, and AI/ML programs. Designed 3TB+ daily scalable ingestion, ELT/ETL, and Medallion Lakehouse transformations using Spark, Databricks, BigQuery, and cloud orchestration tools. Developed standardized, reusable data engineering frameworks.
Big Data Developer
Global Digital & Technology ServicesDelivered high-volume data workflows for CPG forecasting and Life Sciences analytics using Azure Databricks, AWS EMR, and Cloudera. Built ingestion and analytics pipelines for a global Sales & Marketing Promotional AI platform using PySpark. Integrated 26+ diverse Life Sciences data sources for Oncology analytics and automated ETL pipelines with Python.
Big Data Technology Specialist
Enterprise Technology ConsultingEngineered core components of large-scale enterprise product data lakes using Hadoop ecosystem technologies (Spark, Hive, Sqoop). Managed distributed cluster performance, validated and harmonized 10TB+ ingested data feeds, and migrated relational Oracle workloads to distributed platforms.
Software Developer
EdTech Enterprise SolutionsDelivered backend systems, mobile applications, and automation solutions for educational institutions. Enhanced enterprise student information systems (SIS) database reliability and optimized relational data operations across MySQL and Linux server environments. Built shell scripts for automated operations and deployed customized solutions across 30+ institutions.
Academic Credentials
Specialized in advanced database systems, algorithms, distributed computing architectures, and software engineering principles.
Foundation in programming languages (Java, C, C++), relational database design, data structures, and operating systems.