Pipeline Automation & Data Ingestion: Design, build, and optimize automated data feeds for external CPG vendor cost updates, competitor market pricing, and store spatial layout data. Modern Data Stack Governance: Manage and structure modular pipelines using dbt, Databricks Unity Catalog, and Python, ensuring high performance, proper data governance, and schema control. Complex Data Transformations: Write advanced, production-grade SQL and Python scripts to clean, structure, and stitch disparate vendor datasets into scalable data models. CI/CD & Code Quality: Enforce continuous integration, unit testing, version control, and automated deployment practices using GitHub. Cloud Architecture Optimization: Monitor and tune compute performance, data storage, and pipeline execution within AWS and Databricks.