Senior DevOps Engineer
TechGrove by Banyan Software
| Company | TechGrove by Banyan Software |
| Category | Engineering |
| Location | Bengaluru |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Senior |
| Salary | Not stated by the employer |
| Posted | 13 Jul 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (greenhouse) |
Description
TechGrove is the Centre of Excellence for Banyan Software, based in Chennai, India. It plays a key role in supporting Banyan’s global businesses through technology, security, and software development. TechGrove brings together India’s deep pool of technical talent with Banyan’s long-term approach to growth, creating a trusted, developer-focused environment where people can do their best work.
Job Title: Senior DevOps Engineer – Modernized Application Operations & SRE
Overview
We are seeking a highly experienced and hands-on Senior DevOps Engineer to own the operational excellence of the modernized SaaS applications produced by the Banyan AI Factory. This is not a role focused on building the factory itself; instead, you will run the cloud operations and reliability of the modernized applications the factory delivers to our Operating Companies (OpCos).
You will join a team that provides 24x7 coverage with rotating on-call responsibilities, serving as Tier 1 Site Reliability Engineering (SRE) for our OpCos’ modernized SaaS platforms. Day to day, that means cloud operations, performance and availability monitoring/observability, disaster recovery, and security incident response across our two target clouds — Amazon Web Services (AWS) and Microsoft Azure. The ideal candidate pairs deep multi-cloud mastery with a strong commitment to operational excellence and a track record of keeping secure, highly available production systems running at scale.
Key Responsibilities
24x7 Operations & On-Call: Operate as part of a team providing round-the-clock coverage of OpCo modernized SaaS platforms, participating in a rotating on-call schedule to ensure continuous availability and rapid response.
Tier 1 SRE & Cloud Operations: Serve as Tier 1 SRE for the modernized applications produced by the factory, managing day-to-day cloud operations across our two target clouds — AWS and Azure — to keep production systems healthy, performant, and secure.
Performance & Availability Monitoring/Observability: Implement and maintain robust cloud-native observability tooling (monitoring, logging, tracing) to track performance and availability, proactively detect degradation, and drive down mean-time-to-detect and mean-time-to-resolve.
Security Incident Response: Respond to security incidents and operational events affecting OpCo SaaS platforms, executing established runbooks, coordinating remediation, and leading post-incident reviews.
Disaster Recovery & Resilience: Plan, test, and execute disaster recovery procedures that protect the availability and integrity of modernized applications, ensuring backup, failover, and restoration capabilities meet defined RTO/RPO targets.
Infrastructure-as-Code & Automation: Use Infrastructure-as-Code (Terraform) and CI/CD pipelines (e.g., GitHub Actions, GitLab CI) to manage, deploy, and automate the operational environments of modernized applications, reducing toil and improving consistency.
AI Agents & DevSecOps Scale: Build scale in our DevSecOps practice by designing, building, and operating AI agents that automate SRE tasks and incident response, reducing toil and accelerating detection, triage, and remediation.
Hands-on Problem Solving: Serve as a technical escalation point for complex operational challenges, applying strong analytical skills to resolve infrastructure, network, and automation issues across distributed, multi-tenant SaaS environments while navigating technical ambiguity.
Mentorship & Coaching: Mentor and coach junior DevOps/SRE engineers, fostering a culture of operational ownership, continuous improvement, and knowledge sharing in cloud-native and reliability best practices.
Required Qualifications & Experience
Experience: 5–7 years of progressive experience in Software Engineering, DevOps, and/or Site Reliability Engineering, with a focus on operating production SaaS systems.
986,449 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →