Manager, Core Infrastructure Engineering
Oracle
- Location
- Nashville, TN, United States
- Work model
- On-Site
- Level
- Mid
- Posted
- 21h ago
About this role
For a single team delivering components of distributed systems. Translates goals into a 1–2 quarter execution plan, sets coding, testing, and scalability practices, and provides hands-on oversight of performance tuning and load/perf testing. Guides the team in building fault-tolerant, in-service-upgradable components (redundancy, replication, failover) and in applying resiliency patterns (retries, circuit breakers, timeouts). Ensures robust observability (tests, alarms, dashboards, telemetry) and operational readiness via reviewed runbooks and standard procedures. Manages delivery of scoped features and correctness testing (including fault-injection/brownouts), and directs implementation of data replication/synchronization to maintain integrity and availability. Leads team incident response and root-cause efforts, enforces no-customer-downtime practices, and drives use of automation/IaC for troubleshooting. Oversees team security implementation (encryption, access controls), tracks remediation plans, verifies compliance documentation, and coaches adherence to change-management plans for safe patching, updates, and rollbacks.