
Site Reliability Engineer, Multi-Cloud Infrastructure
Real job — pulled straight from Skit.ai’s careers page · Verified July 12, 2026 · No reposts.
Job description
Skit.ai is hiring a Site Reliability Engineer, Multi-Cloud Infrastructure — a full-time, based in Bangalore, India role. Apply directly on Skit.ai's careers page below.
Site Reliability Engineer — Multi-Cloud Infrastructure
Location: Bangalore, India
Department: Technology
Experience: 5+ years
- Multi-cloud operations. Provision, operate, and keep healthy compute, networking, storage, and identity across AWS, GCP, and Azure — with sensible consistency instead of three snowflakes.
- Infrastructure as code. Manage the estate through Terraform (or equivalent) and version control — reproducible environments, reviewed changes, no undocumented hand-tweaks.
- CI/CD and delivery. Keep build and deploy pipelines fast and reliable so engineers ship safely and often.
- Clusters and workloads. Run Kubernetes/container platforms and the supporting services (databases, queues, caches, internal tooling) that everything depends on.
- Monitoring and on-call. Maintain monitoring and alerting for infrastructure health, take a turn in the rotation, and respond to and mitigate incidents with clear communication and blameless follow-up.
- Cost and hygiene. Keep an eye on cloud spend, rightsizing, and waste; own the unglamorous but essential hygiene — patching, backups, secrets, and access.
- Automation and toil reduction. Replace manual, repetitive operations with automation and self-service so the team scales without headcount scaling with it.
- First 90 days. Learn the estate across all three clouds. Take a turn on call. Close the most obvious gaps in monitoring, backups, and access hygiene.
- By 6 months. More of the estate under consistent infrastructure-as-code. Reliable, reviewed CI/CD. A clearer, quieter alerting setup and documented runbooks for the common incidents.
- By 12 months. Measurably less manual toil through automation and self-service. Sensible cost controls in place. Provisioning and environment setup that's repeatable rather than tribal knowledge.
- A few years in SRE, DevOps, or infrastructure operations for production systems, including on-call.
- Hands-on experience across at least two of AWS, GCP, and Azure (all three is a strong plus).
- Kubernetes and containers in production.
- Infrastructure-as-code (Terraform or similar) and CI/CD pipelines.
- Monitoring and alerting practice (e.g. Prometheus/Grafana) and structured incident handling.
- A scripting/programming language for automation (Python, Go, or Bash beyond one-liners).
- Solid Linux systems and networking fundamentals.
- All three clouds run in production, and comfort designing for consistency across them.
- Cost optimization / FinOps.
- Secrets management, security hardening, and compliance/data-residency contexts.
- PostgreSQL and other stateful-service operations at scale.
- Some exposure to real-time or voice infrastructure — enough to back up the platform SRE on call.
Get Site Reliability Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What skills are required for Site Reliability Engineer, Multi-Cloud Infrastructure at Skit.ai?
The required skills for Site Reliability Engineer, Multi-Cloud Infrastructure at Skit.ai include: AWS, GCP, Azure, Kubernetes, Terraform, CI/CD, Prometheus, Grafana, Python, Go, Bash, Linux, Networking, PostgreSQL, GitHub Actions.
What is the seniority level for Site Reliability Engineer, Multi-Cloud Infrastructure at Skit.ai?
Site Reliability Engineer, Multi-Cloud Infrastructure at Skit.ai is a Senior level position.
How do I apply for Site Reliability Engineer, Multi-Cloud Infrastructure at Skit.ai?
You can view the full description and apply for Site Reliability Engineer, Multi-Cloud Infrastructure at Skit.ai on EchoJobs: https://echojobs.io/job/skit-ai-site-reliability-engineer-multi-cloud-infrastructure-5go14.