Site Reliability, Staff

Teknik, data och digitalt · IT-infrastruktur och säkerhet · DevOps · Systemteknik · Mjukvaruutveckling

I korthet

Synopsys is looking for a Staff Site Reliability Engineer in Seongnam-si, South Korea, to maintain and operate critical IT infrastructure for Electronic Design Automation (EDA) and AI platform workloads. The role involves managing HPC clusters, enterprise storage, and customer-hosted deployments, requiring strong Linux system administration, datacenter experience, and proficiency in both Korean and English.

Ansvarsområden

  • Maintain and operate Synopsys Korea engineering compute and data center environments, including servers, enterprise storage (NetApp, NFS), and shared filesystems.
  • Administer HPC clusters (LSF or Slurm), troubleshoot scheduler integration, InfiniBand, GPU compute, and performance incidents.
  • Deploy, configure, and support Synopsys.ai platform components in customer-hosted and air-gapped environments, including container platforms and LLM gateways.
  • Build and maintain observability, alerting, and authentication integration for distributed platform services.
  • Participate in shared on-call rotation and lead/support incident response.
  • Support customer on-site deployments and break-fix at air-gapped or high-security facilities, with travel within Korea and APAC.
  • Automate operational tasks using scripting or infrastructure as code to reduce toil.

Krav

  • 5 or more years in Linux system administration, infrastructure engineering, or production network and storage support.
  • Enterprise data center experience including servers, storage, networking, monitoring, and operational processes.
  • Hands-on NetApp, NFS, and storage troubleshooting in production environments.
  • HPC or EDA engineering compute experience with Linux clusters, workload schedulers (LSF or Slurm), and production support.
  • Ability to support GPU-enabled compute and containerized platform services in customer-controlled environments.
  • Familiarity with virtualization, remote access, observability stacks (Prometheus, Grafana), and infrastructure automation.
  • Understanding of secure multi-tier deployments including gateways, service-to-service authentication, and operational telemetry.
  • Korean language proficiency for local user support, vendor coordination, and customer communication.
  • English sufficient for global IT collaboration, incident documentation, and cross-regional handoffs.
  • Legal authorization to work in Korea and willingness to travel for on-site support.

Önskade kvalifikationer

  • Experience supporting Korea enterprise or semiconductor customer environments is a plus.
  • Prior exposure to Synopsys.ai, LLM gateway, or GenAI platform operations in on-prem or air-gapped settings is a plus.
  • EDA toolchain or fab and design-house HPC support, MCP-based tool integration, Kubernetes or container platforms, vector database operations, cloud AI service interfaces, vLLM or similar on-prem inference serving, or RHCE and other infrastructure certifications are strong differentiators.

Förmåner

  • Comprehensive medical and healthcare plans.
  • ETO and FTO Programs for time away.
  • Maternity and paternity leave, parenting resources, adoption and surrogacy assistance.
  • ESPP to purchase Synopsys common stock at a 15% discount.
  • Retirement plans.
  • Competitive salaries.
#IT#Information Technology#HPC#storage#GPU#Containers#AI/ML#GenAI#Cloud#EDA
Synopsys Inc Logo

Företag

Synopsys Inc

Publicerade jobb

för 5 dagar sedan

Anställningstyp

Heltid

Arbetsform

På plats

Erfarenhetsnivå

Senior

Platser

Seongnam-si, South Korea

Sökande

Ansök tidigt