MogiMogiJobsPowered by MobiusEngineLet Mogi apply

Inference.ai

Senior Software Engineer - Model Performance

San Francisco · 8 months ago

On-siteFull-time

$220,000 to $320,000 a year

About the job

Help us make inference blazingly fast. If you love squeezing every last drop of performance out of GPUs, diving deep into CUDA kernels, and turning optimization techniques into production systems, we'd love to meet you. About Inference.net http://Inference.net Inference.net http://Inference.net trains and hosts specialized language models for companies that need frontier-quality AI at a fraction of the cost. The models we train match GPT-5 accuracy but are smaller, faster, and up to 90% cheaper. Our platform handles everything end-to-end: distillation, training, evaluation, and planet-scale hosting. We are a well-funded ten-person team of engineers who work in-person in downtown San...