ChatGPT · inference infrastructure
About the Team
The Search Product Infrastructure team builds the systems that power search experiences across ChatGPT. We partner with teams developing models, operating inference infrastructure, building specialized search experiences, and maintaining search indexes to bring advances in models and retrieval into production. Our work spans search orchestration, model serving, experimentation, and distributed systems, with direct impact on answer quality, responsiveness, reliability, and efficiency at ChatGPT scale.
About the Role
As a senior engineer on the Search Product Infrastructure team, you will design, build, and operate the systems that connect models with search at ChatGPT scale. You will tackle challenges in search orchestration, inference efficiency, experimentation, and production reliability, making tradeoffs that directly shape answer quality and user experience. Working closely with researchers and partner engineering teams, you will own projects from technical design through launch and iteration, evolving the architecture as models, product capabilities, and demand grow.
This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.
In this role, you will
- Design and evolve the services that coordinate search classification, retrieval, ranking, and model inference, working closely with researchers and partner engineering teams to bring new capabilities into production.
- Improve end-to-end latency, throughput, and infrastructure efficiency through profiling, ca