yoinka

Senior Site Reliability Engineer - CTJ - Poly

Microsoft

United States, Virginia, Reston; United States, Maryland, Annapolis Junction; United States, Washington, RedmondSeniorH-1B sponsor company
Sign in to applyVerified 55m ago
Location
United States, Virginia, Reston; United States, Maryland, Annapolis Junction; United States, Washington, Redmond
Work model
On-Site
Level
Senior
H-1B history
2,066 approvals (FY2023)
Posted
4h ago

Skills

Azure

About this role

Overview

Do you want to be at the heart of cloud computing? The Compute team is at the core of Azure and is growing incredibly fast. We build and manage fault tolerant distributed systems on top of commodity datacenter hardware, to deliver an infrastructure for hosting customer applications. The platform is at the core of Azure that provides millions of virtual machines for customers to run their workload in the cloud. Our team fosters a collaborative environment and builds upon each other’s ideas, to deliver world-class customer value at a rapid pace. We empower engineers to deliver creative solutions through bottoms-up innovation. This is a fun environment and a great opportunity to work on something highly strategic to Microsoft and extremely relevant in the industry. We’re looking for a Senior Site Reliability Engineer and a leader passionate about delivering value to customers in mission critical environments, who enjoys a growth hacking culture, and is eager to play a part in one of the most important long games for Microsoft.   Microsoft’s mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of psychological safety where everyone can thrive at work.

Responsibilities

Acts as a Designated Responsible Individual (DRI) and guides other engineers by developing and following the playbook, working on call to monitor system/product/service for degradation, downtime, or interruptions, alerting stakeholders about status and initiates actions to restore system/product/service for simple and complex problems when appropriate. Proactively seeks new knowledge and adapts to new trends, technical solutions, and patterns that will improve the availability, reliability, efficiency, observability, and performance of service fabric services while also driving consistency in monitoring and operations at scale Drives development of design documents for a product, application, service, or platform. Creates, implements, optimizes, debugs, refactors, and reuses code to establish and improve performance and maintainability, effectiveness, and return on investment (ROI). Leverages subject-matter expertise of product features and partners with appropriate stakeholders (e.g., project managers) to drive a workgroup's project plans, release plans, and work items. Take full ownership of assigned services, actively contributing to its enhancement across all cloud environments. Ensure the service maintains parity with the commercial cloud, delivering high support standards for customers. Participate in the service lifecycle, including design, development, deployment, and maintenance. Collaborate with cross-functional teams to uphold the highest standards of quality and performance. Engage in continuous improvement initiatives to enhance the service's capabilities and user experience. Identify opportunities for automation and optimization within the cloud to better support customers. This includes evaluating current processes and workflows to pinpoint inefficiencies and areas for enhancement. Design and implement automation solutions to streamline operations, reduce manual effort, and improve overall service delivery. Focus on optimizing existing systems and processes to boost performance and customer satisfaction. Embody our culture and values .

Qualifications

Required Qualifications: Master's Degree in Computer Science, Information Technology, or related field AND 2+ years technical experience in software engineering, network engineering, or systems administration OR Bachelor's Degree in Computer Science, Information Technology, or related field AND 4+ years technical experience in software engineering, network engineering, or systems administration OR equivalent experience.  Other

Senior Site Reliability Engineer - CTJ - Poly at Microsoft, United States, Virginia, Reston; United States, Maryland, Annapolis Junction; United States, Washington, Redmond | Yoinka