|
Hi , I hope this email finds you well and that you are having a phenomenal week!
I am reaching out to introduce a true Architect of Modern Data Ecosystems, who is actively seeking his next strategic engagement.
With a solid decade of experience under his belt, He is not just a Data Engineer; he is a catalyst for operational excellence.
While many candidates can build pipelines, He lies in designing future-proof, enterprise-scale Lakehouse architectures and operationalizing cutting-edge AI/BI tools like Databricks Genie and Microsoft Fabric.
Professional Summary:
Certified Azure Data Engineer (DP-203) and SnowPro Core professional with 10 years of elite experience in architecting scalable data pipelines, ingestion frameworks, and analytics platforms. Proficient in Python, PySpark, SQL, and REST API development, with deep-domain expertise in Microsoft Fabric, Azure Databricks, and Snowflake. A master of the Medallion Architecture and Unity Catalog, Pranay excels at solving complex data governance challenges, ensuring ironclad security, and driving needle-moving business insights through innovative, cloud-native solutions.
Clients Worked With:
- GM Financial (Sr. Data Engineer)
- McKesson (Sr. Data Engineer)
- DXC Technology / GM Financial (Sr. Data Engineer)
- Opel Vauxhall Finance (Sr. Data Engineer)
- MetLife Insurance (Data Engineer)
- Phoenix Life Insurance (Data Engineer/ETL Developer)
Recent Work History & Impact:
Company: GM Financial | Position: Senior Data Engineer | Duration: April 2024 – Present
Description: Spearheaded the design and development of enterprise-scale ETL/ELT pipelines and cloud-native Lakehouse architectures to ingest and process massive financial datasets, directly supporting executive analytics and AI/ML initiatives.
- Architected a robust, cloud-native Lakehouse ecosystem utilizing Microsoft Fabric, OneLake, Azure Databricks, and ADLS Gen2, ensuring scalable storage, ironclad governance, and seamless analytics across structured and semi-structured financial data.
- Engineered metadata-driven ingestion frameworks in Python, drastically reducing onboarding time for new data sources and elevating overall pipeline maintainability.
- Implemented the Medallion Architecture (Bronze, Silver, Gold) using Delta Lake, distinctly separating raw ingestion from curated business datasets to maximize scalability, traceability, and data quality.
- Optimized heavy Spark workloads by fine-tuning partitions, joins, caching, and executor configurations, achieving a massive leap in processing performance and slashing compute costs.
- Published governed Lakehouse datasets to Power BI semantic models, empowering executive dashboards with trusted business metrics and enabling frictionless self-service analytics.
- Automated the deployment of Fabric notebooks and ADF pipelines via Azure DevOps CI/CD, ensuring consistent, error-free environment provisioning and delivery.
Company: McKesson | Position: Senior Data Engineer | Duration: July 2023 – March 2024
Description: Led the modernization of healthcare and pharmacy data pipelines, focusing heavily on data governance, real-time streaming, and next-gen AI/BI self-service tools within a strictly HIPAA-compliant environment.
- Pioneered the configuration of Databricks Genie experiences, leveraging trusted datasets, business metrics, and strict guardrails to deliver highly accurate, governed, natural-language self-service insights to business stakeholders.
- Operationalized Databricks Unity Catalog to centralize data discovery, metadata management, lineage, and Role-Based Access Control (RBAC), ensuring secure, compliant access across dev and prod environments.
- Implemented Delta Live Tables (DLT) to automate data quality validation, enforce schemas, and guarantee reliable, idempotent pipeline execution for multi-terabyte healthcare datasets.
- Built high-throughput, event-driven data pipelines utilizing Azure Event Hubs, Azure Functions, and Data Factory to achieve near real-time processing for critical pharmacy datasets.
- Developed incremental data loading strategies utilizing Delta Lake MERGE operations and advanced Change Data Capture (CDC) techniques, ensuring zero data loss and optimal performance.
- Orchestrated complex CI/CD deployments for Databricks and Snowflake objects using GitHub Actions, Azure DevOps, and Terraform, significantly improving deployment consistency.
I am confident that his expertise in Microsoft Fabric and Databricks Genie will make him an absolute favorite with your client’s technical interviewers.
Please let me know if you’d like to schedule a quick screening call or if you need the formatted Word version of his resume.
Also, don’t forget to send over the roles you are working on—I’ve got the perfect candidates for those as well!
Looking forward to partnering with you and making this a win-win!
Thank you,
Bhaskar
Team Lead
✉️: Bhaskar@peachitpros.com
☎️: +1(732) 344-4258
www.peachitpros.com
|