Data Center Technical Operations Engineer, AWS Infrastructure Operations
Amazon
- Location
- AU, Sydney
- Employment
- Full Time
- Work model
- On-Site
- Level
- Mid
- Posted
- 4h ago
Skills
About this role
AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we’re looking for talented people who want to help. You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion. The Infrastructure Operations (Data Center) team is the backbone of AWS, supporting the rapidly growing AWS business and customers 24/7. We are committed to maintain the physical infrastructure of AWS, ensuring the standards for operational performance in the areas of safety, security, availability, productivity, capacity, efficiency, and cost. We are looking for a Data Center Engineering Operations (DCEO) Engineer with experience in critical facilities management, a result-driven individual with strong technical (electrical/mechanical) understanding and the drive and vision to take our data center engineering operations to the next level. The role will report to the DCEO Manager and be responsible for sustaining availability, cost management, risk assessment and mitigation, review corrective and preventative maintenance of critical infrastructure and metric reporting. Key job responsibilities - Participate in a 24/7 rotating shift pattern - Undertake and support management of both routine maintenance and emergency service of a variety of critical systems such as switch-gear, generators, UPS systems, power distribution equipment, chillers, cooling towers, computer room air handlers, building monitoring systems, fire systems, etc. - Installation of new rack capacity and provisioning of power/cooling - On-site management of contractors and vendors, ensuring that all work performed, is in accordance with established practices, procedures & local legislation - Establish performance benchmarks, conduct analyses and prepare reports and documentation of all aspects of critical facility infrastructure operations and maintenance. Manage the facility’s assets - Generate change management requests & incident management tickets for all facility related events (mechanical, electrical, plumbing) and assist in the resolution of infrastructure engineering and services issues - Collaborate with data center IT managers and other business leaders to coordinate day to day operations, projects, manage capacity, and optimize plant safety, performance, reliability, sustainability, and efficiency - Drive & implement projects to increase current facility capacity, efficiency, sustainability & reliability as well as assist in the design implementation, commissioning and build out of new data center facilities. Additional role requirements: - This role may involve bending, lifting, stretching, and reaching, ascending and descending ladders, stairs, and gangways safely - This role has a rostered on-call requirement to provide 24x7 support. This may change based on the future needs of the business - Due to the location of our sites, candidates should possess a valid driver's license and their own vehicle. A day in the life At today's large-scale data centers, you will be involved in: - Work schedule changes depending on specific site needs. Shifts are 12-hours and will rotate on a predefined schedule - Assist in troubleshooting of facility and rack-level events within