Sr Systems Reliability Engineer
T-Mobile
- Location
- Overland Park Kansas
- Work model
- On-Site
- Level
- Senior
- H-1B history
- 142 approvals (FY2023)
- Posted
- 22h ago
Skills
About this role
At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees enjoy multiple wealth-building opportunities through our annual stock grant, employee stock purchase plan, 401(k), and access to free, year-round money coaches. That’s how we’re UNSTOPPABLE for our employees! Are you ready for the next chapter in your Uncarrier journey? The Sr. System Reliability Engineer (SRE) guides and mentors other SREs and improves and protects the software and systems behind all of T-Mobile's IT services, including management of scalability, availability, latency, performance, security, and capacity, and delivering software faster, better, and cheaper. From designing & maintaining CICD Pipelines to building the next generation of T-Mobile applications on cloud native platforms, the SRE's enable great customer experience and product innovation by continuous improvement of operational support. We pride ourselves on encouraging a culture of innovation, advocating for agile methodologies, and promoting transparency in all that we do. Join us in embodying the spirit of the 'Un-carrier' and make a tangible impact! The Compliance Campaign Management (CCM) team within T-Mobile's Identity Security and Access Management (ISAM) organization is leading a major transformation in logical access compliance. What was once a predominantly manual User Access Review process is rapidly evolving into a sophisticated, automation-driven platform powered by cloud technologies, APIs, workflow orchestration, and artificial intelligence. Our team designs and builds solutions using Power Platform, Azure DevOps Pipelines, Microsoft Graph APIs, and AI-powered automation to streamline compliance operations while expanding risk visibility across the enterprise. Every innovation we deliver helps reduce manual effort, strengthen control effectiveness, and improve audit readiness. As a Senior Systems Reliability Engineer (SRE), you will play a critical leadership role in shaping and operating the platform that powers this transformation. CCM is building an environment where human expertise and AI agents work side by side to execute compliance operations with greater speed, consistency, and precision. You will own the reliability, scalability, observability, and continuous improvement of this compliance automation ecosystem, ensuring it remains highly available, performant, secure, and audit ready as it grows across the enterprise. This is more than a traditional operations role. You'll drive engineering excellence across the platform, establish reliability standards, champion automation-first practices, and lead efforts to improve resilience, monitoring, incident response, and operational maturity. Working closely with software engineers, compliance experts, security teams, and AI-enabled solutions, you'll help define the next generation of compliance operations and governance technology. Join a team that's not just supporting compliance but engineering the future of access governance. You'll have the opportunity to solve complex enterprise-scale challenges, influence strategic technology decisions, mentor engineers, and help build intelligent automation capabilities that redefine how access risk is identified, reviewed, and remediated at T-Mobile. If you're energized by the intersection of cloud engineering, automation, AI, cybersecurity, and operational excellence, this is your chance to make a measurable impact on one of the company's most innovative compliance modernization initiatives. Why This Role Is Unique Lead the reliability strategy for a rapidly growing compliance automation platform. Partner with engineers, compliance professionals, and AI solutions to modernize access governance. Build and enhance observability, resilience, and automation across cloud-native services and