Software Engineer, API Multimodal
Backend Engineer · Senior · Full Time
Opens OpenAI's application page
Role
What you'll do.
Join OpenAI's API Multimodal team to design and operate high-scale, low-latency developer-facing APIs powering image generation, speech transcription, and real-time voice interactions. This role combines backend engineering excellence with systems depth to transform frontier AI capabilities into reliable, production-grade experiences for millions of developers worldwide. You'll partner directly with researchers to integrate cutting-edge multimodal models while owning availability, scalability, and performance across distributed systems at unprecedented scale.
Responsibilities
- Design and ship developer-facing APIs: Architect and implement backend services and REST/WebSocket APIs that expose OpenAI's frontier multimodal models (image generation, speech transcription, real-time voice) to external developers with optimal developer experience and performance characteristics.
- Build low-latency streaming systems: Design distributed streaming architectures for real-time interactions with audio, voice, and image models, ensuring sub-second latencies and handling concurrent requests from thousands of developers globally while maintaining reliability and cost efficiency.
- Architect model integration infrastructure: Work directly with Research and Inference teams to integrate new model capabilities into production systems, design the systems around them, and build the operational frameworks that allow models to scale reliably from laboratory to production environments.
- Own service reliability and performance: Take end-to-end ownership of availability, latency optimization, horizontal scalability, and cost efficiency for the backend services and APIs you build, making operational tradeoff decisions aligned with business and user needs.
- Lead complex projects from conception to production: Own technically complex projects from ambiguous requirements, through technical design and implementation, to successful launch and iterative improvement, while establishing strong engineering standards and best practices that elevate team capabilities.
- Incorporate real-world developer feedback: Gather and synthesize feedback from external developers and customers to continuously improve API design, performance, and reliability, ensuring product decisions are grounded in actual user needs and pain points.
Qualifications
What we look for.
Technical
Backend programming languages
Strong proficiency in one or more general-purpose backend languages such as Python, Go, Rust, or TypeScript, with demonstrated ability to write production-grade, performant, maintainable code.
Distributed systems design
Expertise in architecting reliable, scalable distributed systems with deep understanding of asynchronous patterns, message queues, load balancing, failover mechanisms, and reasoning about system behavior under failure conditions.
API design and development
Expertise in designing intuitive, well-documented, and maintainable APIs that balance developer experience with operational constraints, including REST, WebSocket, or gRPC protocols and API versioning strategies.
Cloud infrastructure and orchestration
Hands-on experience with cloud platforms, containerization (Docker/Kubernetes), and infrastructure-as-code tools that enable rapid scaling and reliable deployment of distributed applications in production environments.
Observability and debugging
Strong background in implementing comprehensive observability (logging, metrics, tracing), debugging production issues at scale, and establishing monitoring systems that enable rapid issue detection and root cause analysis.
Education
Computer Science or related field
Bachelor's degree in Computer Science, Computer Engineering, Mathematics, Physics, or equivalent professional software engineering experience demonstrating mastery of computer science fundamentals.
Experience
7+ years backend/infrastructure engineering
Demonstrate a track record of seven or more years of professional software engineering experience (excluding internships) in backend, infrastructure, platform, or product-focused roles, with significant exposure to production systems and distributed architecture decisions.
Production backend and API experience
Proven track record designing, building, implementing, and operating production-grade backend services, developer-facing APIs, or distributed systems that have scaled to support significant traffic and business requirements in real-world environments.
Leadership of technically complex projects
Demonstrated ability to lead technically ambitious projects from ambiguous or undefined requirements through design, implementation, launch, and operational maturity, with successful outcomes in production environments.
Distributed systems expertise
Deep understanding of distributed system design patterns, including expertise in concurrency models, latency analysis and optimization, observability and monitoring, session management, and the operational tradeoffs between consistency, availability, and scalability.
Collaborative cross-functional partnerships
Proven ability to communicate clearly and work collaboratively with diverse stakeholders including researchers, product managers, designers, infrastructure engineers, and external customers, translating between technical and non-technical contexts.
Skills
Required
Python
Production-grade Python expertise with strong fundamentals in asynchronous programming, performance optimization, and building scalable backend services.
Go or Rust
Proficiency in systems programming languages (Go for distributed systems, Rust for performance-critical components) with understanding of concurrency primitives and memory safety considerations.
Distributed systems design
Deep expertise in designing systems that handle fault tolerance, eventual consistency, load distribution, and high concurrency at scale.
API design
Expert-level ability to design developer-friendly APIs that balance performance, usability, backwards compatibility, and operational requirements.
Production backend systems
Proven experience building and operating backend infrastructure that serves high-volume, low-latency traffic with strong SLAs and observability.
System design and architecture
Ability to reason about and design large-scale distributed systems including database design, caching strategies, async processing, and infrastructure topology decisions.
Preferred
Real-time streaming systems
Nice to haveExperience building real-time streaming architectures for audio, video, or WebSocket-based interactions with emphasis on latency optimization and concurrent session handling.
Audio processing and speech systems
Nice to haveFamiliarity with audio codecs, speech recognition/synthesis pipelines, or real-time voice communication systems that inform API design for speech products.
Image generation systems
Nice to haveExperience working with image generation models, computer vision systems, or image processing pipelines that informs understanding of async request patterns and resource optimization.
Multimodal AI applications
Nice to haveBackground working with multimodal AI systems that combine text, audio, images, or video, providing context for designing APIs that expose complex model capabilities intuitively.
TypeScript/Node.js
Nice to haveExperience with TypeScript and Node.js for building backend services, particularly for JavaScript-based tooling or full-stack development contexts.
Product engineering mindset
Nice to haveTrack record of balancing technical depth with product judgment, prioritizing developer experience and user feedback in technical decisions.
Tech stack
Languages
Frameworks
Databases
Tools
Other
Compensation
Pay and benefits.
Base·USD 293,000 – 385,000
Equity·Stock options
Benefits
Health and wellness coverage
Comprehensive medical, dental, and vision insurance with low deductibles and dependents coverage, plus mental health support and fitness stipends.
401(k) retirement plan
Competitive employer-matched 401(k) retirement savings plan with immediate vesting to help you plan for long-term financial security.
Stock options and equity
Significant equity grants allowing you to participate in OpenAI's success and growth as a private company, with standard vesting schedules and option pool management.
Unlimited paid time off
Flexible paid time off policy emphasizing work-life balance, with additional paid holidays and encouragement to take meaningful vacation time.
Professional development
Learning stipends, conference attendance budget, and opportunities to work on cutting-edge AI infrastructure and scale challenges at the forefront of the industry.
Remote flexibility
Flexible work arrangements with opportunities for remote work, though collaboration and team cohesion remain priorities as location was not specified.
Parental leave
Generous parental leave policies for primary and secondary caregivers, with gradual return-to-work options supporting family planning.
Commuter benefits
Transit subsidies, parking support, or commuter benefits for San Francisco Bay Area-based employees, with flexibility for remote-first arrangements.
Process
Interview steps.
- 01
Initial screening call
30-minute phone conversation with a recruiter or senior engineer focused on understanding your background, motivation for joining OpenAI, and high-level alignment with the role's requirements and team dynamics.
- 02
Technical deep dive interview
60-90 minute interview with a senior backend engineer covering distributed systems design, API architecture decisions, concrete examples from your past work, and your approach to solving ambiguous technical problems at scale.
- 03
System design exercise
Interview focused on designing a large-scale distributed system relevant to multimodal APIs, such as architecting a real-time streaming system or designing an API for a complex model capability, including discussion of tradeoffs and operational considerations.
- 04
Code review and implementation
Technical discussion of your past code, design decisions you've made, or optional coding exercise focused on backend systems design rather than algorithmic puzzle-solving, emphasizing production-grade thinking and code quality standards.
- 05
Cross-functional collaboration interview
Conversation with product, research, or infrastructure team members to assess your ability to work across teams, translate between technical and non-technical contexts, and demonstrate collaborative problem-solving in ambiguous situations.
- 06
Leadership and culture fit discussion
Final-round conversation with a manager or director focusing on your sense of ownership, comfort with ambiguity, ability to raise engineering standards, and alignment with OpenAI's mission-driven culture and values.
Full posting
Original listing.
About the Team
API Multimodal builds the developer-facing products and infrastructure that bring OpenAI’s image, audio, and real-time model capabilities into the world. We are responsible for high-scale APIs for image generation, speech transcription, speech generation, and low-latency voice interactions. We partner closely with Research and Inference to bring frontier model capabilities to developers and use customer feedback to improve our models.
About the Role
As a software engineer on API Multimodal, you will build and operate the products and distributed systems behind OpenAI’s image, audio, and real-time APIs. You will work across model integration, API design, and production infrastructure to turn new research capabilities into reliable developer experiences. This hands-on role combines backend and systems depth with product judgment: you will own projects end to end, partner with Research, Inference, and Safety, and help make multimodal AI useful at scale. Model training experience is not required.
In this role, you will:
Design, build, and ship developer-facing APIs and backend services that serve frontier models.
Architect low-latency streaming, request, session, and model integration systems that make complex multimodal interactions reliable and intuitive at scale.
Work directly with Research to bring new model capabilities into production, shape the systems around them, and incorporate feedback from real-world developers and customers.
Own the availability, latency, scalability, and cost efficiency of the services you build.
Own projects from technical design and implementation through launch and ongoing iteration, while raising the team’s engineering standards.
Your background might look something like:
7+ years of professional experience, excluding internships, in backend, infrastructure, platform, or product engineering roles.
A track record of designing, building, and operating production backend services, developer-facing APIs, or distributed systems.
Strong software engineering and systems fundamentals, with experience leading technically complex projects from ambiguous ideas to production.
Proficiency in one or more general-purpose backend languages, such as Python, Go, Rust, or TypeScript.
Experience building reliable, scalable systems and reasoning about distributed architecture, concurrency, latency, observability, and operational tradeoffs.
Product judgment and developer empathy, including an ability to turn complex model or infrastructure capabilities into clear, intuitive APIs.
Clear communication and a collaborative approach to working with researchers, product managers, designers, infrastructure engineers, and customers.
A strong sense of ownership, comfort with ambiguity, and a bias toward learning directly from users while continuously improving engineering quality.
Experience with real-time streaming, audio processing, speech systems, image generation, computer vision, or multimodal AI applications is helpful, but not required.
About OpenAI
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.
For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.
Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.
To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.
We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.
OpenAI Global Applicant Privacy Policy
At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
Redirects to OpenAI's application page.
Other roles
More at OpenAI.
Software Engineer, API Agents
Senior
Manager, Forward Deployed Engineer (FDE), Life Sciences
Manager
Software Engineer, Ads Integrity
Senior
Forward Deployed Engineer - Zurich
Senior
Manager, Forward Deployed Engineer - Tokyo
Manager