
Real job — pulled straight from Gauss Labs’s careers page · Verified July 26, 2026 · No reposts.
Job description
Gauss Labs is hiring a Senior Site Reliability Engineer — a full-time, based in Yeoksam, Seoul role. Apply directly on Gauss Labs's careers page below.
Senior Site Reliability Engineer (KR)
Team: Infrastructure Engineering
Location: Yeoksam, Seoul
Commitment: Full-time
Workplace Type: hybrid
Responsibilities
- Platform reliability and operations: Own platform-layer reliability across both environments. In our internal cloud environment: full ownership — cluster health, resource management (CPU/memory/OOM), scheduling, autoscaling, Kubernetes/EKS lifecycle. In the customer environment: operate directly at the application-namespace level and for the customer-controlled cluster/node layer, diagnose and clearly communicate what's needed, and operate the platform within their setup, decisions, and constraints.
- Monitoring and Alerting: Build and maintain robust monitoring and alerting for the infrastructure and platform layer to proactively identify and resolve issues before they impact the platform.
- Incident Response: Own incident first-response for the platform layer and participate in the on-call rotation to minimize downtime and restore service quickly.
- Automation: Develop automation tools and scripts to streamline operations, reduce manual effort, and enable engineering teams to operate their own services safely.
- Capacity Planning: Forecast resource needs, optimize resource utilization, and ensure the platform infrastructure can handle increasing workloads.
- Deployment infrastructure: Build and maintain CI/CD pipelines and deployment infrastructure for the platform.
- Continuous Improvement: Drive a culture of continuous improvement by identifying opportunities to enhance platform reliability, performance, and efficiency.
Basic Qualifications
- Bachelor's degree in computer science, engineering, or a related discipline
- 5+ years of industry experience as a Site Reliability Engineer or in platform/infrastructure engineering
- Hands-on experience operating Kubernetes in production (EKS preferred): cluster lifecycle, scheduling, autoscaling, resource management
- Experience with cloud platforms (AWS preferred) and containerization technologies (Docker, Kubernetes)
- Experience with observability and alerting tools (Prometheus, Grafana, ElasticSearch, Jaeger)
- Experience with scripting languages (Python, Bash)
- Working knowledge of GitHub, GitHub Actions, and CI/CD concepts
- Strong problem-solving and troubleshooting skills
- Working proficiency in English for internal documentation and technical coordination
Preferred Qualifications
- Knowledge of AI/ML infrastructure and workloads.
- Knowledge of database technologies (MongoDB, PostgreSQL)
- Experience operating software in customer-managed (on-prem or customer-cloud) environments
- Exposure to manufacturing, semiconductor, or enterprise B2B customer environments
Get Site Reliability Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What skills are required for Senior Site Reliability Engineer at Gauss Labs?
The required skills for Senior Site Reliability Engineer at Gauss Labs include: Kubernetes, AWS, Docker, Prometheus, Grafana, Elasticsearch, Python, Bash, GitHub Actions, CI/CD, MongoDB, PostgreSQL, EKS.
What is the seniority level for Senior Site Reliability Engineer at Gauss Labs?
Senior Site Reliability Engineer at Gauss Labs is a Senior level position.
How do I apply for Senior Site Reliability Engineer at Gauss Labs?
You can view the full description and apply for Senior Site Reliability Engineer at Gauss Labs on EchoJobs: https://echojobs.io/job/gauss-labs-senior-site-reliability-engineer-kr-gll53.