Support Team Lead
encora10
If you're ready to move into leadership early in your career, this role puts you in charge of a real team working on real problems. You'll lead application support engineers who keep customer-facing systems running 24/7, which means your decisions directly affect how well the business performs. It's a step up from individual contributor work, but the company expects you to stay hands-on technically.
Your day-to-day: manage a team of Application Support SREs, mentor them on skills and career growth, and handle the big incidents when they happen. You'll own how the team handles incidents, problems, and changes. You'll set team goals aligned with what the business needs, plan capacity for round-the-clock coverage, and lead blameless postmortems so everyone learns from what goes wrong. You're also the senior escalation point when things get urgent.
You need strong technical foundations in SRE or systems engineering, plus the ability to lead without being defensive. If you've worked on reliability, incident response, or operations teams, you have the right background. You should be comfortable with process ownership and can talk clearly to engineers and non-technical stakeholders alike.
To apply, submit your resume and a short note about a time you helped fix a critical incident or improved a process. Use the application on CareerJumpShip to send it to Encora10.
About this role
Coforge is seeking an experienced Site Reliability Engineering (SRE) Team Lead to guide our Application Support SRE function and manage a high‑performing team responsible for ensuring the performance, availability, and reliability of mission‑critical customer‑facing applications. This role combines hands‑on technical leadership with people management, process ownership, and operational excellence. What You'll Lead & Oversee People Leadership Manage, mentor, and coach a team of Application Support SREs; support career progression and skills development. Oversee team performance, capacity planning, and staffing for a 24x7 support model. Serve as the senior escalation point during major incidents and high-severity events. Foster a culture of accountability, blameless postmortems, continuous learning, and operational excellence. Establish team OKRs, KPIs, and reliability goals aligned with business objectives. Operational & Process Ownership Own and mature SRE processes including incident management, problem management, change management, and service readiness. Lead major incident response, coordinate cross-functional teams, ensure communication excellence, and drive root cause analysis. Define and enforce SLOs, SLIs, and error budgets for supported applications. Implement preventative solutions and systemic fixes that reduce incident recurrence. Enhance observability practices across Splunk, OpenTelemetry, AppDynamics, Datadog, and similar tools. Improve dashboards, alerting strategies, and telemetry coverage. Technical Leadership Provide insights and recommendations for reliability, scalability, and performance across AWS-hosted applications, Mulesoft APIs, and Kubernetes-based services. Collaborate with development and architecture teams to integrate SRE principles early in the lifecycle. Champion automation to reduce toil—CI/CD optimization, deployment improvements, self-healing mechanisms, and runbooks. Analyze logs, performance issues, and code behavior to support Tier 2/Tier 3 escalations. Recommend initiatives to expand Splunk automation and AI-driven insights. Qualifications Required 5–8+ years in SRE, DevOps, or production engineering roles. 2–4+ years in a technical lead or people management capacity. Strong experience supporting AWS-based applications, microservices, or API-driven environments. Advanced troubleshooting skills. Hands-on experience with observability stacks (either opensource or splunk) Familiarity with ITIL and incident frameworks. Preferred Bachelor's or Master's degree in Computer Science or related field. Certifications in ITIL, AWS, Azure, or GCP. Experience with Mulesoft, Postman, and API testing. Proficiency with Kubernetes Strong cloud-native networking knowledge. What We Value A proactive, ownership-driven mindset. Strong communication and stakeholder management. Ability to lead during major incidents. Passion for continuous improvement and operational rigor.
Ready to apply?
Related roles
Similar remote openings sourced directly from company career pages.
- encora10
Senior SRE Engineer
USA, USsenior - sentilink
Senior Platform Development Engineer
Remotesenior - sentilink
Senior Systems Software Engineer
Remotesenior - sentilink
Senior Software Engineer, Performance & Reliability
Remotesenior - Bright Vision Technologies
Reliability Engineer
Remotesenior$100K – $150K