Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Site Reliability & Infrastructure Engineer

Instructure
CompanyInstructure
CategoryEngineering
LocationBudapest
RemoteHybrid
EmploymentNot stated
LevelNot stated
SalaryNot stated by the employer
Posted30 Jun 2026
Last verified31 Jul 2026
SourceEmployer career page (ashby)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
At Instructure, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning and personal development, facilitate meaningful relationships, and inspire people to go further in their education and careers. We do this by giving smart, creative, passionate people opportunities to create awesome. And that's where you come in: This is a foundational role for an experienced Site Reliability & Infrastructure Engineer to build and lead the reliability practice for our entire portfolio of custom integrations. You will be the first SRE dedicated to the post-deployment success of our custom solutions, moving beyond project-based delivery to a model of continuous, proactive operational excellence. This is a high-impact position to design and implement a modern SRE practice that integrates into our existing applications and workstreams, establishing a foundation for continued success of our custom development organization. KEY RESPONSIBILITIES 1. Reliability & Performance Engineering: - Design & Implement Observability: Build and manage a centralized monitoring, logging, and alerting strategy for all custom integrations. - Define Service Level Objectives (SLOs): Work with stakeholders to define SLOs and Service Level Indicators (SLIs) that align with customer expectations and business impact. - Architect for Reliability: Partner with Solution Architects and Developers to establish and enforce best practices for integration architecture, ensuring solutions are built for scalability, resiliency, and performance from day one. - Troubleshoot & Remediate: Serve as the primary escalation point for critical incidents related to custom integrations. Lead troubleshooting efforts and perform hands-on development work to resolve complex, high-stakes issues. 2. Infrastructure & Automation: - Manage Integration Infrastructure: Own the cloud infrastructure (AWS, Azure, etc.) that hosts our custom solutions, focusing on security, cost-optimization, and scalability. - Champion Infrastructure as Code (IaC): Partner with our core engineering team to align on best practices and systems to manage deployments and infrastructure. - Automate Everything: Develop and manage CI/CD pipelines for the safe and efficient deployment of integration code and infrastructure changes. 3. Cross-Functional Collaboration & Knowledge Sharing: - Bridge Team Gaps: Create and document clear operational handoffs and processes between teams to ensure a seamless flow from development to production support. - Lead Post-Mortems: Foster a blameless post-mortem culture to analyze incidents, identify root causes, and drive actionable improvements to prevent recurrence. - Create a Knowledge Hub: Develop runbooks, architectural diagrams, and best-practice guides to empower all of Professional Services with the knowledge to better support our solutions. QUALIFICATIONS & EXPERIENCE Required: - You have experience in a Site Reliability Engineering (SRE), DevOps, or Cloud Infrastructure role. - Expertise with AWS. - Hands-on with Infrastructure as Code (Terraform, CloudFormation). - Experience with observability platforms (Datadog, New Relic, Prometheus, Grafana). - Proficient in at least one scripting or programming language, preferably Ruby. - Experience building and managing CI/CD pipelines (e.g., GitLab CI, GitHub Actions, Jenkins). - A systems-thinker with a passion for troubleshooting complex problems and improving processes. Preferred: - Experience working within a client-facing Professional Services or technical consulting organization. - Background in managing the reliability of APIs, middleware, and complex data integrations. - Experience with containerization and orchestration (Docker, Kubernetes). - Bachelor's degree in Computer Science or a related technical field. Get in on all the awesome at Instructure! We o
HOUSE ADYou found the opening. Now track it.Tracker, radar and AI drafts in one place.erioun.com →
Site Reliability & Infrastructure Engineer — Instructure · Job Opportunities API