Site Reliability, Staff
Technology, Data & Digital · IT Infrastructure & Security · DevOps · Systems Engineering · Software Engineering
In short
Synopsys is looking for a Staff Site Reliability Engineer in Seongnam-si, South Korea, to maintain and operate critical IT infrastructure for Electronic Design Automation (EDA) and AI platform workloads. The role involves managing HPC clusters, enterprise storage, and customer-hosted deployments, requiring strong Linux system administration, datacenter experience, and proficiency in both Korean and English.
Responsibilities
- Maintain and operate Synopsys Korea engineering compute and data center environments, including servers, enterprise storage (NetApp, NFS), and shared filesystems.
- Administer HPC clusters (LSF or Slurm), troubleshoot scheduler integration, InfiniBand, GPU compute, and performance incidents.
- Deploy, configure, and support Synopsys.ai platform components in customer-hosted and air-gapped environments, including container platforms and LLM gateways.
- Build and maintain observability, alerting, and authentication integration for distributed platform services.
- Participate in shared on-call rotation and lead/support incident response.
- Support customer on-site deployments and break-fix at air-gapped or high-security facilities, with travel within Korea and APAC.
- Automate operational tasks using scripting or infrastructure as code to reduce toil.
Requirements
- 5 or more years in Linux system administration, infrastructure engineering, or production network and storage support.
- Enterprise data center experience including servers, storage, networking, monitoring, and operational processes.
- Hands-on NetApp, NFS, and storage troubleshooting in production environments.
- HPC or EDA engineering compute experience with Linux clusters, workload schedulers (LSF or Slurm), and production support.
- Ability to support GPU-enabled compute and containerized platform services in customer-controlled environments.
- Familiarity with virtualization, remote access, observability stacks (Prometheus, Grafana), and infrastructure automation.
- Understanding of secure multi-tier deployments including gateways, service-to-service authentication, and operational telemetry.
- Korean language proficiency for local user support, vendor coordination, and customer communication.
- English sufficient for global IT collaboration, incident documentation, and cross-regional handoffs.
- Legal authorization to work in Korea and willingness to travel for on-site support.
Desired Qualifications
- Experience supporting Korea enterprise or semiconductor customer environments is a plus.
- Prior exposure to Synopsys.ai, LLM gateway, or GenAI platform operations in on-prem or air-gapped settings is a plus.
- EDA toolchain or fab and design-house HPC support, MCP-based tool integration, Kubernetes or container platforms, vector database operations, cloud AI service interfaces, vLLM or similar on-prem inference serving, or RHCE and other infrastructure certifications are strong differentiators.
Benefits
- Comprehensive medical and healthcare plans.
- ETO and FTO Programs for time away.
- Maternity and paternity leave, parenting resources, adoption and surrogacy assistance.
- ESPP to purchase Synopsys common stock at a 15% discount.
- Retirement plans.
- Competitive salaries.
#IT#Information Technology#HPC#storage#GPU#Containers#AI/ML#GenAI#Cloud#EDA