Site Reliability, Staff
Synopsys
- Location
- Posted 23-Jul-2026
- Work model
- On-Site
- Level
- Staff
- H-1B history
- 112 approvals (FY2023)
Skills
About this role
Descriptions & Requirements
Job Description and Requirements
We Are Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis solutions, and design services. We partner closely with our customers across a wide range of industries to maximize their R&D capability and productivity, powering innovation today that ignites the ingenuity of tomorrow. You Are You have spent years keeping production systems running, not the kind that serve web pages, but the kind that power semiconductor R&D. Massive HPC grids, petabytes of shared storage, GPU clusters behind firewalls you cannot just SSH through. You know that uptime is not luck, it is runbooks, monitoring that actually fires, and the discipline to fix the root cause instead of restarting the service one more time. You are comfortable walking into a customer data center in Korea, troubleshooting an air-gapped NetApp array that lost quorum at 2am, and explaining what happened to both the local IT team and a global incident bridge in English an hour later. You do not need someone to tell you what to check next. You have done this enough times to know where the problem hides. Containers, orchestration, LLM gateways, vector databases, these are not buzzwords to you. They are just the next layer of infrastructure you need to operate reliably in restricted, high-security environments where "just spin up a cloud instance" is not an option. You script, you automate, you document, and you hand off cleanly to the next region when your shift ends. At Synopsys, you will work on the infrastructure that enables both traditional EDA workloads and the Synopsys.ai platform for semiconductor customers across Korea and APAC.
What You'll Be Doing
Maintain and operate Synopsys Korea engineering compute and data center environments, including server lifecycle, enterprise storage (NetApp, NFS), and shared filesystems used by EDA and AI platform workloads Administer HPC clusters running LSF or Slurm, troubleshoot scheduler integration, InfiniBand where deployed, GPU compute, and performance incidents that affect engineering teams Deploy, configure, and support Synopsys.ai platform components in customer-hosted and air-gapped environments, including container platforms, API/AI gateways, LLM gateway integration, and vector database services Build and maintain observability, alerting, and authentication integration (OIDC, SSO, OAuth) for distributed platform services; write runbooks and operational documentation Participate in shared on-call rotation (weekday nights, weekends, holidays); lead or support incident response, produce post-incident reports in English Support customer on-site deployments and break-fix at air-gapped or high-security facilities; travel within Korea and APAC as needed (typically 10 to 20 percent of time) Automate operational tasks using scripting, infrastructure as code, or AI-assisted workflows to reduce toil and improve handoffs across the APAC team The Impact You Will Have Improve availability and utilization of business-critical EDA and Synopsys.ai infrastructure that directly enables R&D and customer engineering teams across Korea Reduce operational risk for customer air-gapped and on-prem deployments through strong runbooks, monitoring, and incident practices Enable Synopsys tools and GenAI services to run predictably at scale by connecting customer HPC schedulers, shared storage, GPU compute, and platform gateways Extend APAC shared-support coverage with peers in Japan, Vietnam, Taiwan, and China, improving regional response time and knowledge transfer Close recurring operational gaps documented in runbooks and contribute measurable improvements to monitoring, documentation, or automation Build trusted working relationships with Korea engineering users, customer IT contacts, and global platform support teams that make future