Soundhound logo

Staff Site Reliability Engineer

soundhound

Remote
Full-time
Staff
12+ yrs
Salary not listedPosted 4d ago

Real job — pulled straight from soundhound’s careers page · Verified August 8, 2026 · No reposts.

Job description

soundhound is hiring a Staff Site Reliability Engineer — a full-time, remote role. Apply directly on soundhound's careers page below.

Staff Site Reliability Engineer

Location: Toronto, Canada (remote)

Department: Restaurants and Retail

Location Type: REMOTE

Employment Type: FULL_TIME

The Opportunity

We’re looking for a Staff Software Engineer (SRE) to join our Retail and Restaurants AI team. You will be responsible for the reliability, scalability, and performance of our infrastructure, with a deep focus on Google Cloud Platform (GCP). You will architect and maintain high-availability systems, automate operational tasks, and ensure our services can handle the demands of millions of voice AI interactions.

What You'll Do

  • Design, build, and maintain highly available and scalable infrastructure on Google Cloud Platform.
  • Architect and automate CI/CD pipelines to ensure rapid, reliable deployments.
  • Implement robust monitoring, alerting, and observability strategies to proactively identify and resolve system issues.
  • Partner with engineering teams to optimize performance, cost, and reliability of backend services.
  • Drive incident response, post-mortem analysis, and long-term remediation efforts.
  • Identify and eliminate sources of toil, promoting operational maturity and self-service capabilities.
  • Collaborate with cross-functional teams to ensure alignment on infrastructure roadmaps and security standards.
  • Lead department wide compliance (PCI, SOC) initiatives.


What You'll Bring

  • 12+ years of software engineering experience, with significant experience in Site Reliability Engineering or DevOps roles.
  • Expert-level experience with Google Cloud Platform (GCP) services (e.g., GKE, Compute Engine, Cloud Run, Pub/Sub).
  • Proficient in Infrastructure as Code (IaC) tools like Terraform or Pulumi.
  • Deep experience with Kubernetes, container orchestration, and service mesh architectures.
  • Strong background in monitoring and observability tools (e.g., Datadog, Prometheus, Grafana, Cloud Monitoring).
  • Experience designing and managing high-throughput, distributed systems.
  • Strong problem-solving skills and a growth mindset—comfortable with ambiguity and making high-stakes technical trade-offs.
  • Excellent communication skills and a demonstrated ability to mentor engineers.


Preferred Qualifications

  • Experience working in a high-velocity, customer-focused environment.
  • Familiarity with functional programming paradigms (e.g., Clojure/ClojureScript).
  • Prior experience in the restaurant technology, hospitality, or AI-driven SaaS space.
  • Experience implementing security and compliance best practices in the cloud.



Workplace & Compensation

This role is available throughout Canada.

Compensation includes salary, equity, comprehensive healthcare, paid time off, and other benefits. Our recruiting team will provide a specific salary range based on location and years of experience.

#LI-MQ1 #LI-REMOTE

Get Site Reliability Engineer jobs like this

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
Jack Links logo

Site Reliability Engineering Manager

$135k–$138kMinong, WI
✓ From careers page· 42m ago
ThoughtSpot logo

Senior Systems Reliability Engineer

Bengaluru, India
✓ From careers page· 4h ago
JLL logo

JLL

New

Senior Cloud Architect & DevOps Manager

$212k–$259kRemote · US-eligible
✓ From careers page· 5h ago

Frequently asked questions

Is Staff Site Reliability Engineer at soundhound a remote job?

Yes, Staff Site Reliability Engineer at soundhound is a remote position. Candidates in Toronto, Canada may be preferred.

What skills are required for Staff Site Reliability Engineer at soundhound?

The required skills for Staff Site Reliability Engineer at soundhound include: SRE, DevOps, GCP, Terraform, Kubernetes, Datadog, Prometheus, Grafana, PCI DSS, Clojure.

What is the seniority level for Staff Site Reliability Engineer at soundhound?

Staff Site Reliability Engineer at soundhound is a Staff level position.

How do I apply for Staff Site Reliability Engineer at soundhound?

You can view the full description and apply for Staff Site Reliability Engineer at soundhound on EchoJobs: https://echojobs.io/job/soundhound-staff-site-reliability-engineer-ldejx.