Inference Optimization Intern – Performance Modeling
Ifm Us
| Company | Ifm Us |
| Category | Engineering |
| Location | Sunnyvale |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Intern |
| Salary | Not stated by the employer |
| Posted | 24 Jun 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (lever) |
Description
About the Institute of Foundation Models
The Institute of Foundation Models is dedicated to advancing the science and engineering of large-scale AI systems. Our researchers and engineers develop cutting-edge foundation models while pushing the limits of high-performance computing and efficient AI inference. By combining deep expertise in machine learning, systems engineering, and hardware optimization, we build scalable AI solutions that drive scientific discovery and real-world impact.
As part of the team, interns work alongside world-class researchers and performance engineers to optimize the execution of large-scale foundation models on next-generation NVIDIA GPU architectures. This internship provides hands-on experience in low-level GPU performance analysis, kernel optimization, and hardware-aware inference acceleration.
995,367 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →