Senior Cloud Infrastructure Development Engineer - Istio Service Mesh
Expedia Group
- Location
- USA - Illinois - Chicago
- Work model
- On-Site
- Level
- Senior
- Posted
- 1d ago
Skills
About this role
At Expedia Group, we help travelers explore the world, one journey at a time. As a global travel company powered by passionate people, trusted partnerships, and leading technology, we connect travelers, partners, and advertisers through our consumer brands, B2B network, and travel advertising business. Here, you'll do meaningful work that helps millions of people discover, book, and experience travel with more ease, confidence, and joy. Our five Behaviors-Traveler First, Think Big, Operate with Excellence, Ownership Mindset, and Succeed Together-help foster a supportive environment where people can grow their careers and have the flexibility, benefits, and support to do their best work. Join us and build for travelers everywhere. Introduction to Team Our Technology Team partners with teams across Expedia Group to create innovative products, services, and tools to deliver high-quality experiences for travelers, partners, and our employees. A singular technology platform powered by data and machine learning provides secure, differentiated, and personalized experiences that drive loyalty and traveler satisfaction. This Senior Cloud Infrastructure Development Engineer role is part of the cloud infrastructure team which sits within our technology division. The cloud infrastructure team designs, builds, and operates the foundational cloud platforms, tools, and services that power Expedia Group’s products at global scale, ensuring they are secure, reliable, and cost-efficient. As a Senior Software Development Engineer, you will lead the design and delivery of robust, scalable cloud solutions that enable product teams to ship features faster and more safely while maintaining high standards of availability and performance. In this role, you will: Architect and evolve secure, scalable platform capabilities across multi-cloud and hybrid environments, with a focus on connectivity, automation, and infrastructure reliability. Design, implement, and scale orchestration and infrastructure automation solutions using tools such as Terraform, KubeFed, and related Kubernetes ecosystem technologies. Configure, implement, and enhance service mesh capabilities, including ingress controller patterns and Istio-based solutions, to support secure service-to-service communication and workload connectivity. Drive observability, capacity planning, system and service performance analysis, and environment tuning, while debugging issues across production and pre-production environments. Advance continuous delivery practices by automating application and infrastructure deployments, integrating infrastructure as code into CI/CD pipelines, and detecting and remediating deployment issues. Take ownership of high-pressure operational scenarios by applying calm, data-driven decision making, advocating for operational excellence through resiliency, scalability, testing, and service-level practices, participating in on-call rotations, and exploring new technologies, including AI-driven tools and workflows, that improve engineering outcomes.
Minimum Qualifications
Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience. 8+ years (Bachelor's) or 6+ years (Master's) of professional software engineering experience 3+ years of infrastructure automation, configuration management or container orchestration Several years of experience designing, building, and operating Kubernetes-based infrastructure and cloud-native services in production environments, with ownership of multi-service or platform-level systems. Demonstrated proficiency with at least one major public cloud provider, container orchestration concepts, networking, and security controls, including hands-on experience with infrastructure-as-code and automation. Proven ability to design and support APIs, system integrations, and data models for platform services, ensuring reliability, scalability, and maintainability across multiple teams or