أبلاي إيدج ابدأ البحث عن عمل

SRE Practice Lead

Coforge · Noida, Uttar Pradesh, India

قدّم وتابع مع أبلاي إيدج
Job Title: SRE Practice Lead – Engineering ServicesSkills: SRE, Reliability Engineering Frameworks, SLI / SLO / Error Budget Management, Incident Management & RCA, Service Resilience & Availability Engineering, Capacity Planning, Cloud & Platform Engineering, DevOps & Automation, Observability & AIOps, COE Setup, Competency DevelopmentExperience: 18+ YearsLocation: Greater Noida Job Summary:We are seeking a seasoned SRE Practice Lead to drive Reliability Engineering transformation across enterprise clients through a combination of consulting, solutioning, presales, and delivery leadership.This role requires a leader who has built and scaled SRE capabilities within an engineering services or system integration environment and has successfully partnered with clients to modernize operations through cloud-native engineering, observability, automation, platform engineering, and AI-enabled reliability practices. The ideal candidate should possess strong customer-facing experience, hands-on solution architecture capabilities, and the ability to lead end-to-end transformation initiatives, from presales and consulting through execution and value realization.Note: Candidates from engineering services, digital engineering, cloud transformation, infrastructure modernization, or system integration organizations will be strongly preferred. Candidates whose experience is primarily limited to governance, enterprise architecture, or captive/shared-services environments may not be suitable unless they demonstrate substantial presales and delivery ownership.Key Responsibilities:Practice Leadership: Define and scale enterprise-wide SRE strategies, frameworks, operating models, and best practicesLead transformation engagements that move clients from traditional support-centric operations to engineering-led, automation-first reliability modelsEstablish reusable assets, accelerators, reference architectures, and service offerings aligned to modern managed services and digital engineeringDrive adoption of AI-enabled SRE practices, predictive operations, intelligent automation, and self-healing platformsSolutioning & Presales Leadership:Partner with sales and account teams to develop winning SRE, observability, platform engineering, and cloud transformation solutionsLead customer workshops, discovery sessions, assessments, and executive-level discussionsOwn solution architecture, effort estimation, commercials, response to RFPs/RFIs, and technical proposalsPresent solution strategies, transformation roadmaps, business cases, and value realization models to customer stakeholdersSupport deal pursuits through Proof of Concepts, demonstrations, solution reviews, and technical due diligenceEngineering & Delivery Leadership:Lead large-scale SRE and platform engineering programs across cloud-native and distributed environmentsDesign highly resilient, scalable, self-healing systems leveraging Kubernetes, containers, cloud services, observability platforms, and modern DevOps toolchainsDrive implementation of reliability metrics including SLIs, SLOs, Error Budgets, MTTR reduction, and operational excellence initiativesGovern delivery quality, transformation outcomes, stakeholder management, and customer successDrive automation across provisioning, deployment, incident response, monitoring, and operational workflowsCustomer Consulting: Advise customers on SRE maturity assessments, operating model transformation, cloud adoption, observability strategy, and reliability roadmapsArticulate business outcomes around availability, performance, productivity, operational efficiency, and cost optimizationAct as a trusted advisor to engineering leadership, architecture teams, and executive stakeholdersPeople & Capability Development:Build and mentor high-performing SRE and platform engineering teamsDefine capability roadmaps, learning paths, certifications, and engineering standardsFoster a culture of reliability, automation, innovation, and continuous improvementRole Competencies:15+ years of experience across SRE, Platform Engineering, DevOps, Cloud Engineering, or Infrastructure TransformationStrong experience working within global engineering services, digital engineering, or system integration organizationsProven experience leading client-facing solutioning, architecture discussions, consulting engagements, and presales pursuitsDemonstrated ownership of large transformation programs from proposal stage through delivery executionStrong expertise in AWS, Azure, and/or GCP environmentsDeep understanding of SRE principles including SLIs, SLOs, Error Budgets, Incident Management, Observability, and Reliability EngineeringExperience designing cloud-native platforms using Kubernetes, containers, IaC, CI/CD, and automation frameworksPreferred Skills:Terraform, Ansible, GitOps, CI/CD toolchainsPrometheus, Grafana, Dynatrace, Datadog, Splunk, ELK, OpenTelemetryPlatform Engineering and Internal Developer Platform experienceAI-driven operations, AIOps, predictive monitoring, and intelligent automationExecutive stakeholder management and consulting skills