
Real job — pulled straight from TechDome’s careers page · Verified September 15, 2026 · No reposts.
Job description
TechDome is hiring a Site Reliability Engineer (Remote) — a full-time, based in Hyderabad, India role. Apply directly on TechDome's careers page below.
Site Reliability Engineer (SRE)
Location: Hyderabad, India; Indore, India
Department: Engineering
Experience: 2+ years
- Define SLIs/SLOs, own the error budget, and decide when to slow down shipping to protect it
- Build zero-downtime CI/CD pipelines supporting Blue-Green, Canary, and Rolling releases
- Manage all environments through Terraform and Ansible — no manual console changes
- Instrument systems with Prometheus, Grafana, ELK, Datadog, and OpenTelemetry so alerts are actionable, not noisy
- Lead incident response, drive root cause analysis, and ensure postmortem action items are closed
- Right-size infrastructure and forecast capacity proactively, rather than reacting to billing
- Apply AI to alert triage, incident summarization, and automated runbooks
- Participate in a shared on-call rotation as a dependable, trusted responder
- 2+ years of production ownership experience as an SRE, DevOps, Platform, or Cloud Engineer
- Hands-on, production-grade experience with AWS, Azure, or GCP
- Real-world Docker/Kubernetes experience under production load
- Daily use of Terraform and Ansible (or equivalent IaC tools)
- Experience building at least one CI/CD pipeline from scratch (Jenkins, GitHub Actions, GitLab CI, or similar)
- Strong fundamentals in Linux, networking, and distributed systems
- Proficiency in Python, Go, or Bash for automation and scripting
- Proven experience shipping Blue-Green, Canary, and Rolling deployments in production
- Prior work in FinTech, Payments, Healthcare, or another high-availability domain
- Practical, everyday use of AI tools such as Copilot, Claude, Cursor, or ChatGPT
- Experience building AI-powered operations tooling — triage bots, incident auto-summarization, or reliable anomaly detection
- Fluency in SLOs, error budgets, and chaos engineering as core practice, not theory
- Direct reporting line to founders and senior engineering leadership
- Infrastructure decisions ship the same week they're made — no bureaucratic delay
- You'll protect systems handling real healthcare data and real financial transactions
- A team that shares on-call and builds a culture where the SRE on call at 2am is someone people trust
Get Site Reliability Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What skills are required for Site Reliability Engineer (Remote) at TechDome?
The required skills for Site Reliability Engineer (Remote) at TechDome include: AWS, Azure, GCP, Docker, Kubernetes, Terraform, Ansible, Jenkins, GitHub Actions, GitLab CI, Linux, Networking, Python, Go, Bash, Prometheus, Grafana, Elasticsearch, Datadog, OpenTelemetry, CI/CD, AI.
What is the seniority level for Site Reliability Engineer (Remote) at TechDome?
Site Reliability Engineer (Remote) at TechDome is a Senior / Mid Level level position.
How do I apply for Site Reliability Engineer (Remote) at TechDome?
You can view the full description and apply for Site Reliability Engineer (Remote) at TechDome on EchoJobs: https://echojobs.io/job/techdome-site-reliability-engineer-sre-1eyno.