OpenAI is an AI research and deployment company dedicated to developing and safely deploying general-purpose artificial intelligence. The Software Engineer, Model Runtime will build a production-grade LLM inference runtime for frontier models on OpenAI’s custom silicon, optimizing execution across latency, throughput, utilization, and reliability. The role involves designing distributed execution, collaborating across hardware and software layers, and developing tools for profiling, observability, and performance analysis.