Back to search

Inference Optimization Intern – Performance Modeling

ifm-us

For faster consideration
Prepare with AI feedback
10 hits6 practices3 interviews
Sunnyvale, CAIntern | FallEngineeringAI/ML

About the Institute of Foundation Models

The Institute of Foundation Models is dedicated to advancing the science and engineering of large-scale AI systems. Our researchers and engineers develop cutting-edge foundation models while pushing the limits of high-performance computing and efficient AI inference. By combining deep expertise in machine learning, systems engineering, and hardware optimization, we build scalable AI solutions that drive scientific discovery and real-world impact.

As part of the team, interns work alongside world-class researchers and performance engineers to optimize the execution of large-scale foundation models on next-generation NVIDIA GPU architectures. This internship provides hands-on experience in low-level GPU performance analysis, kernel optimization, and hardware-aware inference acceleration.

AI enhanced job description

Source: leverRecruiter: Rekroot
f03e1363-85c2-479d-b38f-0b199883d5c1
Inference Optimization Intern – Performance Modeling at ifm-us — HudsonSignals