
Real job — pulled straight from soundhound’s careers page · Verified August 8, 2026 · No reposts.
Job description
soundhound is hiring a Staff Site Reliability Engineer — a full-time, remote role. Apply directly on soundhound's careers page below.
Staff Site Reliability Engineer
Location: Toronto, Canada (remote)
Department: Restaurants and Retail
Location Type: REMOTE
Employment Type: FULL_TIME
The Opportunity
What You'll Do
- Design, build, and maintain highly available and scalable infrastructure on Google Cloud Platform.
- Architect and automate CI/CD pipelines to ensure rapid, reliable deployments.
- Implement robust monitoring, alerting, and observability strategies to proactively identify and resolve system issues.
- Partner with engineering teams to optimize performance, cost, and reliability of backend services.
- Drive incident response, post-mortem analysis, and long-term remediation efforts.
- Identify and eliminate sources of toil, promoting operational maturity and self-service capabilities.
- Collaborate with cross-functional teams to ensure alignment on infrastructure roadmaps and security standards.
- Lead department wide compliance (PCI, SOC) initiatives.
What You'll Bring
- 12+ years of software engineering experience, with significant experience in Site Reliability Engineering or DevOps roles.
- Expert-level experience with Google Cloud Platform (GCP) services (e.g., GKE, Compute Engine, Cloud Run, Pub/Sub).
- Proficient in Infrastructure as Code (IaC) tools like Terraform or Pulumi.
- Deep experience with Kubernetes, container orchestration, and service mesh architectures.
- Strong background in monitoring and observability tools (e.g., Datadog, Prometheus, Grafana, Cloud Monitoring).
- Experience designing and managing high-throughput, distributed systems.
- Strong problem-solving skills and a growth mindset—comfortable with ambiguity and making high-stakes technical trade-offs.
- Excellent communication skills and a demonstrated ability to mentor engineers.
Preferred Qualifications
- Experience working in a high-velocity, customer-focused environment.
- Familiarity with functional programming paradigms (e.g., Clojure/ClojureScript).
- Prior experience in the restaurant technology, hospitality, or AI-driven SaaS space.
- Experience implementing security and compliance best practices in the cloud.
Workplace & Compensation
Get Site Reliability Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Senior Cloud Architect & DevOps Manager
Frequently asked questions
Is Staff Site Reliability Engineer at soundhound a remote job?
Yes, Staff Site Reliability Engineer at soundhound is a remote position. Candidates in Toronto, Canada may be preferred.
What skills are required for Staff Site Reliability Engineer at soundhound?
The required skills for Staff Site Reliability Engineer at soundhound include: SRE, DevOps, GCP, Terraform, Kubernetes, Datadog, Prometheus, Grafana, PCI DSS, Clojure.
What is the seniority level for Staff Site Reliability Engineer at soundhound?
Staff Site Reliability Engineer at soundhound is a Staff level position.
How do I apply for Staff Site Reliability Engineer at soundhound?
You can view the full description and apply for Staff Site Reliability Engineer at soundhound on EchoJobs: https://echojobs.io/job/soundhound-staff-site-reliability-engineer-ldejx.