Senior Cloud & Kubernetes Engineer
InterSystems
| Company | InterSystems |
| Category | Engineering |
| Location | Dublin |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Senior |
| Salary | Not stated by the employer |
| Posted | 10 Jul 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (greenhouse) |
Description
InterSystems powers some of the world's most mission-critical healthcare, financial services, and data platforms. Our technology runs where reliability, security, and performance are not optional, and where the work directly supports organizations that depend on trusted data every day.
As Managed Services and SaaS continue to modernize, cloud engineering and Kubernetes are becoming core foundations for how we deliver resilient, automated, and scalable services globally. This role offers the opportunity to work on real production platforms, improve how services are built and operated, and help shape a modern cloud-native operating model.
This is a chance to do hands-on engineering work that matters: cloud automation, Kubernetes operations, reliability, security, and platform improvement in an environment where technical quality has meaningful business and client impact.
About the Role
We are seeking a Senior Cloud and Kubernetes Engineer to serve as a senior hands-on contributor for the build, operation, automation, and continuous improvement of InterSystems Kubernetes-based managed services and SaaS platforms.
This role is more advanced than the Kubernetes Engineer position and is intended for someone who can independently own complex platform work, lead production troubleshooting, improve engineering standards, and mentor other engineers while partnering closely with Kubernetes architects and SRE teams.
The Senior Cloud and Kubernetes Engineer will help translate architecture into reliable implementation, bring operational discipline to Kubernetes environments, and drive automation, reliability, security, and platform maturity across public cloud, private cloud, and on-premises infrastructure.
Key Responsibilities
Senior Platform Engineering and Operations
Independently build, operate, upgrade, and improve production Kubernetes clusters across public cloud, private cloud, and on-premises environments.
Lead complex cluster lifecycle activities including version upgrades, node pool changes, platform patching, capacity planning, and workload migration.
Serve as an escalation point for Kubernetes issues involving cluster health, networking, ingress, storage, certificates, scheduling, resource pressure, and workload failures.
Identify systemic platform risks and drive corrective actions that improve reliability, security, and operability.
Partner with architects to implement platform standards, reference designs, and approved operating patterns.
Automation, GitOps, and Infrastructure as Code
Design and maintain reusable Terraform, Helm, and GitOps automation for repeatable Kubernetes platform deployment and configuration.
Improve CI/CD and GitOps workflows using tools such as Argo CD, Flux, GitHub Actions, GitLab, Jenkins, or Azure DevOps.
Drive automation that reduces manual operations, improves consistency, and lowers operational risk.
Review infrastructure and platform changes for quality, reliability, security, and maintainability.
Contribute to automation standards, code review practices, reusable modules, and deployment patterns.
Security, Governance, and Compliance
Implement and improve Kubernetes RBAC, namespace isolation, network policies, secrets management, certificate lifecycle management, and workload security controls.
Partner with security teams to remediate vulnerabilities, enforce secure platform patterns, and support audit or compliance requirements.
Support image registry controls, container scanning, admission controls, and policy-as-code practices where applicable.
Help define practical security standards that can be consistently implemented by engineering and application teams.
Reliability, Observability, and Incident Leadership
Lead troubleshooting of complex production incidents and coordinate with application, infrastructure, network, security, and database teams to restore service.
Build and improve observability using Prometheus, Grafana
You found the opening. Now track it.Tracker, radar and AI drafts in one place.erioun.com →