Inference Optimization Intern – Performance Modeling
Ifm Us
Institute of Foundation Models (IFM) at Ifm Us focuses on advancing large-scale AI systems through high-performance computing and efficient inference. As an Inference Optimization Intern you will develop analytical performance models, build simulators, and profile GPU kernels to optimize transformer inference on NVIDIA Hopper and Blackwell GPUs.