
Real job — pulled straight from Félix’s careers page · Verified September 7, 2026 · No reposts.
Job description
Félix is hiring a Site Reliability Engineer — a full-time, based in Mexico role. Apply directly on Félix's careers page below.
Senior Site Reliability Engineer
Location: Mexico
Department: Engineering
Location Type: REMOTE
Employment Type: FULL_TIME
- Manage and optimize our infrastructure on Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
- Automate provisioning and configuration using Terraform, Helm, and scripting languages such as Go, Python, and Bash.
- Build, maintain, and improve monitoring and alerting systems using OpenTelemetry standards
- Participate in on-call rotations, incident response, and post-mortem analyses, ensuring rapid recovery and continuous learning from failures.
- Define and track SLOs/SLIs and error budgets to monitor service health and performance.
- Implement cloud security best practices to protect sensitive data and maintain the integrity of our systems.
- Collaborate across Engineering, Security, and Product teams to embed reliability and automation in every phase of development and deployment.
- Contribute to GKE cost optimization and resource management strategies to enhance efficiency and control operational spend.
- 4+ years of experience as a SRE/Platform Engineer.
- Strong hands-on experience with GCP and GKE.
- Proficiency in Kubernetes (architecture, deployments, networking, and troubleshooting).
- Solid programming or scripting skills in Go, Python, or Bash.
- Proficiency with Docker and Linux
- Experience with Terraform
- Experience with Helm
- Experience with GitHub Actions
- Strong understanding of monitoring and observability using Prometheus, Grafana, and logging frameworks.
- Familiarity with incident management, on-call operations, and post-mortem processes.
- Knowledge of network fundamentals (TCP/IP, DNS, Load Balancing).
- Experience with PostgreSQL or distributed databases.
- Awareness of FinOps and cloud cost management principles.
- Excellent problem-solving, communication, and collaboration skills, with a proactive mindset.
- GCP certifications, such as Professional DevOps Engineer or Cloud Architect.
- Certified Kubernetes Administrator (CKA).
- Experience in FinOps, cloud security, or regulated industries.
- Familiarity with PagerDuty or similar incident management tools.
- Background implementing SLOs/SLIs and error budgets in production environments.
- These are the applicable requisites, although equivalent competencies in any of the above will also be considered.
- Competitive salary
- Initial stock options grant
- Annual performance bonus
- Health, dental, and vision plans
- Remote work environment, although we have offices in Miami and México City and would love to work in hybrid model if you are up to it.
- Continuous learning opportunities
- Unlimited PTO
- Paid parental leave
- Empowering opportunities for growth in a dynamic entrepreneurial environment
Get Site Reliability Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs


Senior Software Engineer, DGX Cloud Orchestration (Remote)


Senior Solutions Architect, Cloud Partners
Frequently asked questions
What skills are required for Site Reliability Engineer at Félix?
The required skills for Site Reliability Engineer at Félix include: GCP, Kubernetes, Go, Python, Bash, Docker, Linux, Terraform, Helm, GitHub Actions, Prometheus, Grafana, PostgreSQL, TCP/IP, DNS, OpenTelemetry.
What is the seniority level for Site Reliability Engineer at Félix?
Site Reliability Engineer at Félix is a Senior level position.
How do I apply for Site Reliability Engineer at Félix?
You can view the full description and apply for Site Reliability Engineer at Félix on EchoJobs: https://echojobs.io/job/f-lix-senior-site-reliability-engineer-xp81m.