AI Native, Tech Ops Engineer
ConsumerAffairs
| Company | ConsumerAffairs |
| Category | Engineering |
| Location | United States |
| Remote | Remote |
| Employment | Full-time |
| Level | Not stated |
| Salary | Not stated by the employer |
| Posted | 22 Jul 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (workable) |
Description
The Tech Ops Engineer serves as a force multiplier for operational excellence, combining infrastructure expertise with AI-powered automation to create self-healing, highly observable, and scalable technology systems. This role harnesses machine intelligence, predictive monitoring, and automated workflows to anticipate issues before they affect users, accelerate incident response, and continuously improve platform performance. Through close partnership with engineering, security, and business teams, the Tech Ops Engineer helps build an AI-native operational environment that maximizes reliability, efficiency, and business agility. Responsibilities System Monitoring and Maintenance: Monitor and maintain the organization’s infrastructure, including servers, networks, storage systems, and applications. Perform routine system checks and preventive maintenance to ensure optimal performance and uptime. Respond to system alerts and incidents, diagnosing and resolving issues promptly to minimize downtime. Troubleshooting and Support: Provide technical support to resolve infrastructure-related issues, working closely with other technical teams. Troubleshoot and resolve hardware, software, and network issues, escalating to higher-level support when necessary. Maintain detailed documentation of issues, solutions, and processes to improve the team’s knowledge base. System Upgrades and Patching: Plan and execute system upgrades, patches, and configuration changes, ensuring minimal disruption to business operations. Test and validate updates in development environments before deploying them to production. Ensure that all systems comply with security standards and best practices. Automation and Optimization: Identify opportunities to automate routine tasks and processes, improving operational efficiency and reducing manual workload. Implement scripts, automation tools, and AI skills to streamline system management and monitoring. Continuously evaluate and optimize infrastructure performance, capacity, and resource utilization. Disaster Recovery and Backup: Support the development and execution of disaster recovery plans to ensure business continuity in case of system failures. Manage backup and restore processes for critical systems and data, ensuring data integrity and availability. Participate in regular disaster recovery testing and drills. Infrastructure Lifecycle Management: Plan and execute decommissioning of legacy infrastructure, including EC2 instances, VPCs, and load balancers, coordinating Terraform state cleanup and DNS cutover. Collaboration and Communication: Work closely with development, network, and security teams to ensure alignment and effective communication on infrastructure projects. Provide input on infrastructure design and architecture to support new projects and initiatives. Communicate effectively with non-technical stakeholders, providing updates on system status and issues. Requirements Minimum Qualifications & Credentials Bachelor’s degree in Computer Science, Information Technology, or a related field, or equivalent work experience. 5+ years of experience in system administration, or a similar role. 5+ years of professional experience in Linux administration, managing AWS resources, developing CI/CD and server orchestration pipelines, scripting and monitoring. Hard/Technical Skills You are an expert in: Cloud-based production systems at scale Amazon Web Services (EC2, VPC, EFS, S3, EKS etc.) Production experience running workloads in Kubernetes (EKS), including ArgoCD GitOps deployments Infrastructure as Code tools, primarily Terraform You have experience with: Working in a Python and JavaScript-centric codebase and are familiar with their related best-practices Creating CI/CD pipelines with Jenkins, Concourse or other CI/CD implementation Monitoring tools, like Datadog or Prometheus Scripting for server side automat
You found the opening. Now track it.Tracker, radar and AI drafts in one place.erioun.com →