Sr. Data Engineer
Coca-Cola
- Location
- Mexico Mexico City
- Work model
- On-Site
- Level
- Senior
- H-1B history
- 3 approvals (FY2023)
- Posted
- 16h ago
Skills
About this role
Job Description
Summary: At The Coca ‑ Cola Company, we believe data is the foundation for creating personalized experiences and powering sustainable growth in today’s digital-first world. As consumers engage with our brands in more connected ways than ever before, we are investing in advanced data and analytics platforms to unlock deeper insights, accelerate decision-making, and enable seamless innovation at scale. In this role as a Sr. Data Engineer , you will play a critical part in shaping how data flows across our enterprise ecosystem. By leveraging your deep expertise in Spark, Scala development, and JVM-based engineering, you will design and implement integration frameworks that power both batch and streaming workloads on Microsoft Fabric and Synapse Analytics.
Core Responsibilities
Design and develop data integration pipelines and backend services for structured and unstructured data, using Apache Spark with Scala. Build and optimize distributed computing applications that deliver scalable batch and streaming workloads within Microsoft Fabric environments. Perform advanced Spark performance tuning, including partitioning strategies, memory configuration, query optimization, and workload balancing across large datasets. Develop REST-based microservices and APIs on the JVM to enable seamless interoperability between internal and external systems. Engineer modular, testable, and maintainable code applying functional programming principles and industry-best patterns for reliability and performance. Optimize workflows through data partitioning strategies, storage formats, and platform configurations to enable high-performance querying and cost efficiency. Engineering Rigor: Champion the "You Build It, You Run It" philosophy. Apply strict functional programming and SOLID principles to write clean, modular, and testable Scala code. Performance Optimization: Conduct deep Spark performance tuning - managing memory, query plans, partitioning strategies, and capacity optimization to deliver highly efficient batch and streaming workloads within Microsoft Fabric. CI/CD & Automation: Architect and maintain robust CI/CD pipelines using GitHub Actions, Maven, and SBT to automate testing, deployment, and configuration of data artifacts. Required Qualifications & Experience Strong proficiency in Scala programming with demonstrated knowledge of functional programming concepts, concurrency models, and modular system design. Spark Expertise ( 3 + years): Deep, hands-on experience building, debugging and optimizing distributed computing applications using Apache Spark (Scala/Java preferred) . Experience designing data models (STAR, Vault, Tabular) for analytics. Experience building REST based microservices on JVM. Synapse/Fabric Experience (3+ years of combined experience); Designing and transforming large data using Azure Data Services (Synapse, ADF, Azure Functions, ARM Templates) . DevOps & CI/CD: Proven ability to build automated release cycles. Hands-on expertise with Git, GitHub Actions, and build tools (Maven/SBT, UV are required ). Demonstrated experience working with big data ecosystems such as Hadoop, Delta Lake, or similar technologies within cloud platforms (AWS, Azure, or GCP). Delivery: Experience in Agile development (Scrum, lean techniques) and collaborating with global engineering teams.
Skills
Location(s): Mexico City/Cities: Mexico Travel Required: 00% - 25% Relocation Provided: No Job Posting End Date: August 29, 2026 Our Purpose and Growth Culture: We are taking deliberate action to nurture an inclusive culture that is grounded in our company purpose, to refresh the world and make a difference. We act with a growth mindset, take an expansive approach to what’s possible and believe in continuous learning to improve our business and