Systems Development Engineer (SRE/DevOps)
modeln
| Company | modeln |
| Category | Engineering |
| Location | Hyderabad |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Not stated |
| Salary | Not stated by the employer |
| Posted | 28 Jul 2026 |
| Last verified | 12 Aug 2026 |
| Source | The employer's own careers page (company_site) |
Description
This company seeks a Systems Development Engineer specializing in SRE/DevOps to design, build, and maintain automated CI/CD pipelines, infrastructure, and Kubernetes environments on AWS. The role combines site reliability engineering practices with platform engineering, focusing on observability, automation, and operational excellence.
What You'll Do
• Design, build, and maintain automated CI/CD pipelines using Harness, GitHub Actions, and ArgoCD
• Develop infrastructure and configuration as code using CloudFormation, Terraform, and Ansible
• Administer and optimize AWS environments including networking, security, and architecture for availability and cost efficiency
• Manage Kubernetes clusters and containerized workloads, including configuration, scaling, and upgrades
• Design and evolve end-to-end monitoring and observability frameworks using OpenTelemetry, CloudWatch, Datadog, Prometheus, or similar platforms
• Create dashboards, logs, traces, and automated alerting systems; lead incident response and root cause analysis for production incidents
• Automate operational tasks, runbooks, and incident remediation workflows to reduce toil and improve reliability
What You Need
• 2–4 years of experience designing, implementing, and maintaining CI/CD pipelines
• Hands-on experience with Terraform, CloudFormation, and Ansible for Infrastructure as Code and Configuration as Code
• Strong understanding of Infrastructure as Code principles and modern SRE/DevOps practices
• AWS administration and architecture experience including networking, security, IAM, and core services
• Experience operating Kubernetes clusters (EKS or other distributions) and containerized workloads
• Deep experience with monitoring and observability tools such as OpenTelemetry, CloudWatch, Datadog, or Prometheus
• Proficiency in Linux administration, system configuration, troubleshooting, and performance tuning
• Programming/scripting skills in Python, Go, or Rust for automation and tooling
• Experience troubleshooting complex distributed systems and supporting incident response
Nice to Have
• Experience building unified observability platforms or standardized dashboards for multiple services and teams
• Experience with GitOps workflows and tools for declarative infrastructure and application delivery
• Background in incident command and post-mortem frameworks
• Experience integrating observability and reliability practices into microservices and serverless architectures
• Experience integrating testing, security, and compliance checks into CI/CD pipelines