s
shivp634

Shivam Pandey

@shivp634

Senior Data Engineer

Inde
Anglais, Hindi
Certaines informations sont présentées en anglais.
À propos de moi
I am a Senior Data Engineer with over 3 years of experience building and leading AWS-based data engineering solutions. I specialize in data lake architectures using AWS S3, Athena, Glue, and Python, with a focus on scalability and cost efficiency. I currently own end-to-end delivery from solution architecture to ETL pipeline automation.... Plus d’infos

Compétences

s
shivp634
Shivam Pandey
hors ligne • 

Voir mes services

Conseil en ingénierie des données
I will design and build your AWS data lake architecture on AWS
Conseil en ingénierie des données
I will optimize your slow AWS athena or spark sql queries

Expérience professionnelle

Senior Software Engineer

Not Found • Temps plein

Jan 2023 - Present3 yrs 7 mos

Work History Senior Data Engineer & Team Lead | Leading IT Services Company | 2023 – Present Data Lake & Analytics — BFSI (Banking & Financial Services) Client Lead a data engineering team building and maintaining a multi-tier AWS data lake (bronze → stage → standard model → marts → analytics) for a major banking/mutual fund client Own end-to-end SQL/Spark pipeline architecture handling terabytes of financial and institutional reporting data Act as primary technical point of contact for data pipeline issues, architecture decisions, and cross-team coordination Key Achievements - Led Data Lake architecture project — architected and delivered a large-scale AWS data lake serving mutual fund and institutional reporting needs, now a core production system for the client Large-scale data refresh design — built a historical RM-level attribution refresh pipeline handling 100+ GB of partitioned S3/Parquet data, using EMR with broadcast joins to keep fortnightly refreshes performant Legacy SQL modernization — converted a large, complex SQL Server stored procedure into Athena-compatible, CTE-based queries — improved both maintainability and query performance Cross-account infrastructure debugging — diagnosed and resolved Glue Catalog access issues between EMR Spark SQL and Athena in a cross-account AWS setup, unblocking a critical reporting workflow Automated financial reporting — built AWS Lambda-based pipelines generating multi-sheet Excel sales reports and institutional AUM/AAUM summaries, replacing manual reporting processes Pipeline consolidation — simplified a multi-step Scala/Spark channel-derivation pipeline into a clean, lookup-table-driven design, reducing complexity and maintenance overhead