Carbon3 AI logo

Site Reliability Engineer

Carbon3 AI

Hybrid
United Kingdom
Full-time
Mid Level
Salary not listedPosted 45m ago

Real job — pulled straight from Carbon3 AI’s careers page · Verified August 21, 2026 · No reposts.

Job description

Carbon3 AI is hiring a Site Reliability Engineer — a full-time, based in United Kingdom role. Apply directly on Carbon3 AI's careers page below.

Site Reliability Engineer

Location: United Kingdom: (Occasional office visit required)

Department: Executive & Operations



Role Summary:

We’re hiring SRE/Platform engineers with an automation bias to help build Era4’s operations capability from the ground up. You’ll turn runbooks, alerts and operational workflows into safe, auditable automation and internal tooling that improves reliability across our AI infrastructure and datacentre platform.

 

This is a Platform / SRE role with software engineering, not an AI model-building role. You’ll work closely with operations, platform and engineering teams to reduce manual toil, improve alert quality, and speed up incident response.

 

Key Responsibilities:

  • Build Python-based automation for incident triage, runbook execution, and routine operational tasks.
  • Integrate observability, ITSM and infrastructure APIs to enrich alerts and automate workflows.
  • Improve monitoring signal quality through correlation, enrichment, suppression and deduplication.
  • Build internal tools and self-service capabilities such as CLI utilities, ChatOps integrations and dashboards.
  • Maintain version-controlled runbook-as-code and automation libraries.
  • Translate post-incident learnings into better tooling, automation and operational standards.
  • Support safe, auditable automation for higher-risk actions with appropriate approval controls.

 

Essential Experience:

  • Experience in SRE, Platform Engineering, or production infrastructure operations.
  • Hands-on experience with observability/monitoring tooling (for example Prometheus, Grafana or similar).
  • Exposure to incident management / on-call and converting manual runbooks into automation.
  • Experience with Python for automation, APIs and integrations.

 

Nice To Have:

  • GPU, datacentre or colocation infrastructure experience.
  • ITSM integrations (, Halo, Jira Service Management or similar).
  • ChatOps tooling (Slack or Microsoft Teams bots).
  • OpenTelemetry, logging or distributed tracing experience.
  • DCIM, IPAM or hypervisor-control-plane integrations.
  • Experience with LLM-assisted or agent-based operational automation.

 

Why Join Era4:

You’ll be joining a mission-driven start-up building critical national infrastructure, where operational excellence directly enables growth. This role offers high visibility with leadership, real autonomy, and the chance to shape how a next-generation company operates at scale.

 

Diversity & Inclusion:

Era4 is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.

 

About the Company

Era4 develops, owns and operates AI infrastructure across the UK, powered by renewable energy. Converting legacy industrial and energy sites into modern data-centre facilities, Era4 is combining brownfield regeneration opportunities with cleaner, efficient, scalable compute capacity for healthcare, research, finance, enterprise, and public-sector organisations

Get Site Reliability Engineer jobs like this

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
Boston Dynamics logo

Controls Reliability Engineer

$100k–$120kWaltham, MA
✓ From careers page· 8m ago
PNC Financial Services logo

Software Developer Lead

$86k–$173kRemote · US-eligible
✓ From careers page· 2h ago
Kyndryl Canada logo

Squad Leader Infrastructure

Mexico City, DF
✓ From careers page· 2h ago
PNC Financial Services logo

Senior System Reliability and Support Specialist

$75k–$138kPittsburgh, PA
✓ From careers page· 3h ago

Frequently asked questions

What skills are required for Site Reliability Engineer at Carbon3 AI?

The required skills for Site Reliability Engineer at Carbon3 AI include: SRE, Python, Prometheus, Grafana, API, OpenTelemetry, LLM, Slack, Microsoft Teams.

What is the seniority level for Site Reliability Engineer at Carbon3 AI?

Site Reliability Engineer at Carbon3 AI is a Mid Level level position.

How do I apply for Site Reliability Engineer at Carbon3 AI?

You can view the full description and apply for Site Reliability Engineer at Carbon3 AI on EchoJobs: https://echojobs.io/job/carbon3-ai-site-reliability-engineer-fvbr4.