GPU Performance and Benchmarking Engineer
Vultr
| Company | Vultr |
| Category | Engineering |
| Location | Remote - United States |
| Remote | Remote |
| Employment | Not stated |
| Level | Not stated |
| Salary | Not stated by the employer |
| Posted | 13 Jul 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (ashby) |
Description
WHO WE ARE
Vultr is on a mission to make high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators around the world. With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December 2024 Vultr announced an equity financing at a $3.5 billion valuation. Founded by David Aninowsky and self-funded for over a decade, Vultr has grown to become the world’s largest privately-held cloud infrastructure company.
Vultr Cares
- 100% company-paid insurance premiums for employee medical, dental and vision plans.
- 401(k) plan that matches 100% up to 4%, with immediate vesting
- Professional Development Reimbursement of $2,500 each year
- 11 Holidays + Paid Time Off Accrual + Rollover Plan
- Commitment matters to Vultr! Increased PTO at 3 year and 10 year anniversary + 1 month paid sabbatical every 5 years + Anniversary Bonus each year
- $500 stipend for remote office setup in first year + $400 each following year
- Internet reimbursement up to $75 per month
- Gym membership reimbursement up to $50 per month
- Company paid Wellable subscription
JOIN VULTR
Vultr is seeking a highly skilled and experienced GPU Performance and Benchmarking Engineer to drive performance validation and optimization of GPU infrastructure through rigorous benchmarking and systematic tuning of AI workloads. The ideal candidate is deep hands-on experience with GPU performance analysis, AI/ML workload characterization, and identifying and applying the parameters that maximize training and inference throughput. This is a highly visible role in a high-growth technology company, which will require strong analytical skills, familiarity with GPU profiling tools, and the ability to translate benchmark results into actionable tuning recommendations. This is your opportunity to join our fast growing team and leave your mark on Vultr and the future of Cloud Infrastructure.
Key Responsibilities
- Design and execute performance benchmarks for AI training and inference workloads
- Profile and characterize GPU workloads to identify performance bottlenecks and optimization opportunities
- Systematically tune workload parameters (batch size, precision, parallelism, memory, etc.) to maximize throughput
- Establish and maintain performance baselines and success criteria across GPU platforms
- Develop benchmarking tools, scripts, and automation for repeatable performance validation
- Analyze and report benchmark results with actionable recommendations for engineering teams
- Validate performance of new GPU hardware platforms before production deployment
- Collaborate with GPU Engineers and Fabric Engineers to correlate performance with system-level health
- Track and evaluate emerging GPU architectures and software releases for performance impact
- Document benchmarking methodologies, tuning playbooks, and performance best practices
Qualifications
- 3–7 years of experience in GPU performance engineering, benchmarking, or HPC
- Hands-on experience with GPU profiling and benchmarking tools (e.g., NVIDIA Nsight, DCGM, nccl-tests, MLPerf, GPU-Burn)
- Strong understanding of AI/ML workload performance characteristics (training vs. inference, batch sizing, precision modes)
- Experience tuning performance parameters across GPU, CUDA, and framework layers
- Familiarity with major GPU platforms (NVIDIA, AMD) and their performance tooling
- Proficiency in Python for benchmark scripting and data analysis
- Basic understanding of Linux systems and server hardware
- Familiarity with high-speed networking concepts (InfiniBand, RoCE, NCCL) and their impact on distributed performance
- Strong analytical skills with the ability to translate data into clear recommendations
- Excellent written