Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Senior DevOps Engineer

TechGrove by Banyan Software
CompanyTechGrove by Banyan Software
CategoryEngineering
LocationBengaluru
RemoteOn-site (inferred)
EmploymentNot stated
LevelSenior
SalaryNot stated by the employer
Posted13 Jul 2026
Last verified30 Jul 2026
SourceEmployer career page (greenhouse)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
TechGrove is the Centre of Excellence for Banyan Software, based in Chennai, India. It plays a key role in supporting Banyan’s global businesses through technology, security, and software development. TechGrove brings together India’s deep pool of technical talent with Banyan’s long-term approach to growth, creating a trusted, developer-focused environment where people can do their best work. Job Title: Senior DevOps Engineer – Modernized Application Operations & SRE Overview We are seeking a highly experienced and hands-on Senior DevOps Engineer to own the operational excellence of the modernized SaaS applications produced by the Banyan AI Factory. This is not a role focused on building the factory itself; instead, you will run the cloud operations and reliability of the modernized applications the factory delivers to our Operating Companies (OpCos). You will join a team that provides 24x7 coverage with rotating on-call responsibilities, serving as Tier 1 Site Reliability Engineering (SRE) for our OpCos’ modernized SaaS platforms. Day to day, that means cloud operations, performance and availability monitoring/observability, disaster recovery, and security incident response across our two target clouds — Amazon Web Services (AWS) and Microsoft Azure. The ideal candidate pairs deep multi-cloud mastery with a strong commitment to operational excellence and a track record of keeping secure, highly available production systems running at scale. Key Responsibilities 24x7 Operations & On-Call: Operate as part of a team providing round-the-clock coverage of OpCo modernized SaaS platforms, participating in a rotating on-call schedule to ensure continuous availability and rapid response. Tier 1 SRE & Cloud Operations: Serve as Tier 1 SRE for the modernized applications produced by the factory, managing day-to-day cloud operations across our two target clouds — AWS and Azure — to keep production systems healthy, performant, and secure. Performance & Availability Monitoring/Observability: Implement and maintain robust cloud-native observability tooling (monitoring, logging, tracing) to track performance and availability, proactively detect degradation, and drive down mean-time-to-detect and mean-time-to-resolve. Security Incident Response: Respond to security incidents and operational events affecting OpCo SaaS platforms, executing established runbooks, coordinating remediation, and leading post-incident reviews. Disaster Recovery & Resilience: Plan, test, and execute disaster recovery procedures that protect the availability and integrity of modernized applications, ensuring backup, failover, and restoration capabilities meet defined RTO/RPO targets. Infrastructure-as-Code & Automation: Use Infrastructure-as-Code (Terraform) and CI/CD pipelines (e.g., GitHub Actions, GitLab CI) to manage, deploy, and automate the operational environments of modernized applications, reducing toil and improving consistency. AI Agents & DevSecOps Scale: Build scale in our DevSecOps practice by designing, building, and operating AI agents that automate SRE tasks and incident response, reducing toil and accelerating detection, triage, and remediation. Hands-on Problem Solving: Serve as a technical escalation point for complex operational challenges, applying strong analytical skills to resolve infrastructure, network, and automation issues across distributed, multi-tenant SaaS environments while navigating technical ambiguity. Mentorship & Coaching: Mentor and coach junior DevOps/SRE engineers, fostering a culture of operational ownership, continuous improvement, and knowledge sharing in cloud-native and reliability best practices. Required Qualifications & Experience Experience: 5–7 years of progressive experience in Software Engineering, DevOps, and/or Site Reliability Engineering, with a focus on operating production SaaS systems.
HOUSE AD986,449 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →