Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Senior Production Engineer

CoreWeave Europe
CompanyCoreWeave Europe
CategoryEngineering
LocationWarsaw
RemoteOn-site (inferred)
EmploymentNot stated
LevelSenior
SalaryNot stated by the employer
Posted27 Oct 2025
Last verified3 Aug 2026
SourceEmployer career page (greenhouse)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at  www.coreweave.com .   We're proud to be a Living Wage accredited Employer.   What You'll Do: As a Production Engineer, you will play a key role in maintaining the reliability and stability of CoreWeave’s cloud infrastructure. You will work closely with the Production Engineer Team Lead and other engineers to support incident response, platform reliability, and operational improvements. You will be part of a dynamic team responsible for monitoring the health of our infrastructure, troubleshooting issues, and participating in both day-to-day operational tasks and incident resolution efforts. About the role:   You will assist in incident response efforts by helping identify and resolve service disruptions quickly while documenting root cause analysis (RCA) and post-incident reviews (PIRs). You will monitor system performance and health using tools like Prometheus and Grafana to identify potential incidents and implement automation to reduce manual intervention. This position involves collaborating across teams to improve platform reliability and resilience while refining incident response playbooks. As you gain experience, you will take on more complex responsibilities in incident management and system reliability. Who You Are: 5+ years of experience in cloud operations, site reliability engineering (SRE), or related technical roles. Strong understanding of cloud platforms (e.g., Kubernetes, AWS, GCP) and cloud infrastructure. Expertise in scripting or using automation tools such as Python, Bash, Terraform, or Ansible. Good familiarity with incident management practices and frameworks like ITIL or SRE best practices. Experience with monitoring and alerting tools including Prometheus and Grafana. Strong communication skills with the ability to work in a fast-paced, high-pressure environment. Preferred: Experience working with Kubernetes, containerization, and distributed systems. Knowledge of change management processes and post-incident analysis. Experience with automated systems or self-healing infrastructure. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You love to maintain the reliability and stability of high-scale cloud infrastructure. You're curious about automation and process improvements to enhance incident detection. You're an expert in cloud operations and incident management frameworks. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the organization's growth opportunities ar
HOUSE ADYour CV gets thirty seconds.CV writing and honest review. English & Greek.kaeros.app →