
Senior Software Engineer, Platform Operation and Reliability
Real job — pulled straight from AT&T’s careers page · Verified August 28, 2026 · No reposts.
Job description
AT&T is hiring a Senior Software Engineer, Platform Operation and Reliability — a full-time, based in Bengaluru, KA role. Apply directly on AT&T's careers page below.
Sr Specialist, Software Engineer - Engineer – Platform Operation and Reliability
Location: IND:KA:Bengaluru / Innovator Building, Itpb, Whitefield Rd - Adm: Intl Tech Park, Innovator Bldg
Remote Type: Office Worker (NOT Remote)
Time Type: Full time
Job Description
Experience: 8+ years
Key Responsibilities
Production Support Leadership
Lead day-to-day operations of production support across Tier 1 and Tier 2 teams.
Ensure platform availability, reliability, and performance across environments.
Manage support operations for multiple modules and instances.
Define and maintain support processes, workflows, and escalation models.
Oversee shift operations including 24x7 support structures.
Active Participation in Platform Clone/Upgrade Activity.
Incident & Major Incident Management
Act as the primary escalation point for P1 and P2 incidents.
Lead Major Incident bridges and ensure timely resolution.
Drive Root Cause Analysis (RCA) and corrective actions.
Track incident trends and implement preventive measures.
Ensure SLA and OLA adherence across all support tiers.
Service Delivery & Governance
Own overall support service delivery.
Define KPIs, SLAs, and operational metrics.
Conduct service reviews with stakeholders and leadership.
Ensure compliance with governance standards and best practices.
Drive operational maturity and service improvements.
Team Leadership & People Management
Lead and mentor Tier 1 and Tier 2 support teams.
Manage team capacity planning and resource allocation.
Conduct performance reviews and skill development planning.
Identify training needs and drive capability building.
Foster a collaborative and high-performance team culture.
Platform Stability & Continuous Improvement
Identify recurring issues and drive permanent fixes.
Promote automation and platform optimization initiatives.
Improve platform performance and operational efficiency.
Drive knowledge management and documentation maturity.
Implement proactive monitoring strategies.
Stakeholder & Vendor Management
Act as the primary point of contact for business stakeholders.
Coordinate with infrastructure, security, and integration teams.
Manage vendor relationships and support transitions if applicable.
Provide executive-level reporting on platform health and performance.
Release & Change Governance
Oversee change and release activities across environments.
Ensure smooth production deployments and stabilization.
Review high-risk changes and approve deployment readiness.
Support release planning and execution strategies.
Required Skills
Technical Skills
Strong hands-on experience with platform administration and development
Expertise in core modules such as:
ITSM
Service Catalog
CMDB
Knowledge Management
Change and Incident Management
Good understanding of:
Integrations (REST/SOAP APIs)
MID Servers
Performance tuning
Data management
Experience supporting multi-instance environments
Knowledge in Performance management, resiliency engineering
Knowledge of release and deployment processes
Experience in NowAssist and AI Capabilities
Leadership Skills
Strong team leadership and mentoring capabilities
Experience managing large, distributed teams
Strong incident leadership experience
Strategic thinking and operational planning ability
Stakeholder management and executive communication
Soft Skills
Strong decision-making ability under pressure
Excellent communication and reporting skills
Ability to manage multiple priorities
Strong analytical and problem-solving mindset
High accountability and ownership
Preferred Qualifications
Certified System Administrator (CSA) – Required
Certified Application Developer (CAD) – Preferred
ITIL Foundation / ITIL Intermediate – Preferred
Experience managing enterprise-scale support environments
Experience supporting global 24x7 support models
Exposure to platform governance frameworks
Exposure to AWS, Azure
Team Scope Typically Managed
Typical team size:
15–30+ team members across shifts
Ideal Candidate Profile
Has deep production support experience
Comfortable handling high-pressure incident situations
Strong people leader who can balance technical depth with delivery ownership
Experienced in building and scaling support teams
Good at driving operational maturity and automation
Weekly Hours:
40Time Type:
RegularLocation:
IND:KA:Bengaluru / Innovator Building, Itpb, Whitefield Rd - Adm: Intl Tech Park, Innovator BldgAT&T and its subsidiaries are committed to equal employment opportunity. All hiring, promotion, and other employment decisions remain merit-based and free from discrimination on the basis of race, color, religion, religious creed, national origin, ancestry, age, sex, sexual orientation, gender, gender identity, gender expression, physical disability, mental disability, pregnancy, medical condition, genetic information, marital status, citizenship status, military status, veteran status, or any other characteristic protected by federal, state, or local laws. In addition, AT&T will provide reasonable accommodations to qualified individuals with disabilities. AT&T is a fair chance employer and does not initiate a background check until an offer is made. Click here to learn more or request an application accommodation here.
Get Software Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs
Frequently asked questions
What skills are required for Senior Software Engineer, Platform Operation and Reliability at AT&T?
The required skills for Senior Software Engineer, Platform Operation and Reliability at AT&T include: REST, SOAP, API, ITIL, AWS, Azure.
What is the seniority level for Senior Software Engineer, Platform Operation and Reliability at AT&T?
Senior Software Engineer, Platform Operation and Reliability at AT&T is a Senior / Staff level position.
How do I apply for Senior Software Engineer, Platform Operation and Reliability at AT&T?
You can view the full description and apply for Senior Software Engineer, Platform Operation and Reliability at AT&T on EchoJobs: https://echojobs.io/job/at-t-sr-specialist-software-engineer-servicenow-engineer-platform-operation-and-reliability-0eir6.

