Jobs › US jobs › Software Engineer jobs in San Francisco › Software Engineer, Inference - Performance Optimization
Openai · San Francisco
About the Team Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks and drive faster, cheaper inference. We combine systems profiling, benchmarking, and analysis to understand where time and cost are spent, then turn that understanding into performance optimizations and models that project performance and capacity needs for future launches.
Read the full advert and apply on Openai's site →
Collected from Openai's own careers site (Ashby). Posted 25 Apr 2026. TUNAI shows an excerpt and the facts it read from the advert; the employer's page has the full description and the application.
Does your CV match this job? Paste both and see exactly which of these skills it shows.
Check my CV against this job →