Software Engineer, API Multimodal

Backend Engineer · Senior · Full Time

San FranciscoUSD 293k – 385k3d ago
Apply for this role

Opens OpenAI's application page

Role

What you'll do.

Join OpenAI's API Multimodal team to design and operate high-scale, low-latency developer-facing APIs powering image generation, speech transcription, and real-time voice interactions. This role combines backend engineering excellence with systems depth to transform frontier AI capabilities into reliable, production-grade experiences for millions of developers worldwide. You'll partner directly with researchers to integrate cutting-edge multimodal models while owning availability, scalability, and performance across distributed systems at unprecedented scale.

Responsibilities

  • Design and ship developer-facing APIs: Architect and implement backend services and REST/WebSocket APIs that expose OpenAI's frontier multimodal models (image generation, speech transcription, real-time voice) to external developers with optimal developer experience and performance characteristics.
  • Build low-latency streaming systems: Design distributed streaming architectures for real-time interactions with audio, voice, and image models, ensuring sub-second latencies and handling concurrent requests from thousands of developers globally while maintaining reliability and cost efficiency.
  • Architect model integration infrastructure: Work directly with Research and Inference teams to integrate new model capabilities into production systems, design the systems around them, and build the operational frameworks that allow models to scale reliably from laboratory to production environments.
  • Own service reliability and performance: Take end-to-end ownership of availability, latency optimization, horizontal scalability, and cost efficiency for the backend services and APIs you build, making operational tradeoff decisions aligned with business and user needs.
  • Lead complex projects from conception to production: Own technically complex projects from ambiguous requirements, through technical design and implementation, to successful launch and iterative improvement, while establishing strong engineering standards and best practices that elevate team capabilities.
  • Incorporate real-world developer feedback: Gather and synthesize feedback from external developers and customers to continuously improve API design, performance, and reliability, ensuring product decisions are grounded in actual user needs and pain points.

Qualifications

What we look for.

Technical

  • Backend programming languages

    Strong proficiency in one or more general-purpose backend languages such as Python, Go, Rust, or TypeScript, with demonstrated ability to write production-grade, performant, maintainable code.

  • Distributed systems design

    Expertise in architecting reliable, scalable distributed systems with deep understanding of asynchronous patterns, message queues, load balancing, failover mechanisms, and reasoning about system behavior under failure conditions.

  • API design and development

    Expertise in designing intuitive, well-documented, and maintainable APIs that balance developer experience with operational constraints, including REST, WebSocket, or gRPC protocols and API versioning strategies.

  • Cloud infrastructure and orchestration

    Hands-on experience with cloud platforms, containerization (Docker/Kubernetes), and infrastructure-as-code tools that enable rapid scaling and reliable deployment of distributed applications in production environments.

  • Observability and debugging

    Strong background in implementing comprehensive observability (logging, metrics, tracing), debugging production issues at scale, and establishing monitoring systems that enable rapid issue detection and root cause analysis.

Education

  • Computer Science or related field

    Bachelor's degree in Computer Science, Computer Engineering, Mathematics, Physics, or equivalent professional software engineering experience demonstrating mastery of computer science fundamentals.

Experience

  • 7+ years backend/infrastructure engineering

    Demonstrate a track record of seven or more years of professional software engineering experience (excluding internships) in backend, infrastructure, platform, or product-focused roles, with significant exposure to production systems and distributed architecture decisions.

  • Production backend and API experience

    Proven track record designing, building, implementing, and operating production-grade backend services, developer-facing APIs, or distributed systems that have scaled to support significant traffic and business requirements in real-world environments.

  • Leadership of technically complex projects

    Demonstrated ability to lead technically ambitious projects from ambiguous or undefined requirements through design, implementation, launch, and operational maturity, with successful outcomes in production environments.

  • Distributed systems expertise

    Deep understanding of distributed system design patterns, including expertise in concurrency models, latency analysis and optimization, observability and monitoring, session management, and the operational tradeoffs between consistency, availability, and scalability.

  • Collaborative cross-functional partnerships

    Proven ability to communicate clearly and work collaboratively with diverse stakeholders including researchers, product managers, designers, infrastructure engineers, and external customers, translating between technical and non-technical contexts.

Skills

Required

  • Python

    Production-grade Python expertise with strong fundamentals in asynchronous programming, performance optimization, and building scalable backend services.

  • Go or Rust

    Proficiency in systems programming languages (Go for distributed systems, Rust for performance-critical components) with understanding of concurrency primitives and memory safety considerations.

  • Distributed systems design

    Deep expertise in designing systems that handle fault tolerance, eventual consistency, load distribution, and high concurrency at scale.

  • API design

    Expert-level ability to design developer-friendly APIs that balance performance, usability, backwards compatibility, and operational requirements.

  • Production backend systems

    Proven experience building and operating backend infrastructure that serves high-volume, low-latency traffic with strong SLAs and observability.

  • System design and architecture

    Ability to reason about and design large-scale distributed systems including database design, caching strategies, async processing, and infrastructure topology decisions.

Preferred

  • Real-time streaming systems

    Nice to have

    Experience building real-time streaming architectures for audio, video, or WebSocket-based interactions with emphasis on latency optimization and concurrent session handling.

  • Audio processing and speech systems

    Nice to have

    Familiarity with audio codecs, speech recognition/synthesis pipelines, or real-time voice communication systems that inform API design for speech products.

  • Image generation systems

    Nice to have

    Experience working with image generation models, computer vision systems, or image processing pipelines that informs understanding of async request patterns and resource optimization.

  • Multimodal AI applications

    Nice to have

    Background working with multimodal AI systems that combine text, audio, images, or video, providing context for designing APIs that expose complex model capabilities intuitively.

  • TypeScript/Node.js

    Nice to have

    Experience with TypeScript and Node.js for building backend services, particularly for JavaScript-based tooling or full-stack development contexts.

  • Product engineering mindset

    Nice to have

    Track record of balancing technical depth with product judgment, prioritizing developer experience and user feedback in technical decisions.

Tech stack

Languages

PythonGoRustTypeScript

Frameworks

FastAPI or FlaskgRPCWebSocketApache Kafka or RabbitMQ

Databases

PostgreSQLRedisDynamoDB or similar NoSQL

Tools

KubernetesDockerPrometheus and GrafanaELK Stack or DatadogJaeger or similar tracing

Other

REST API design principlesLoad balancing strategiesConcurrency patternsSystem performance optimization

Compensation

Pay and benefits.

Base·USD 293,000 – 385,000

Equity·Stock options

Benefits

  • Health and wellness coverage

    Comprehensive medical, dental, and vision insurance with low deductibles and dependents coverage, plus mental health support and fitness stipends.

  • 401(k) retirement plan

    Competitive employer-matched 401(k) retirement savings plan with immediate vesting to help you plan for long-term financial security.

  • Stock options and equity

    Significant equity grants allowing you to participate in OpenAI's success and growth as a private company, with standard vesting schedules and option pool management.

  • Unlimited paid time off

    Flexible paid time off policy emphasizing work-life balance, with additional paid holidays and encouragement to take meaningful vacation time.

  • Professional development

    Learning stipends, conference attendance budget, and opportunities to work on cutting-edge AI infrastructure and scale challenges at the forefront of the industry.

  • Remote flexibility

    Flexible work arrangements with opportunities for remote work, though collaboration and team cohesion remain priorities as location was not specified.

  • Parental leave

    Generous parental leave policies for primary and secondary caregivers, with gradual return-to-work options supporting family planning.

  • Commuter benefits

    Transit subsidies, parking support, or commuter benefits for San Francisco Bay Area-based employees, with flexibility for remote-first arrangements.

Process

Interview steps.

  1. 01

    Initial screening call

    30-minute phone conversation with a recruiter or senior engineer focused on understanding your background, motivation for joining OpenAI, and high-level alignment with the role's requirements and team dynamics.

  2. 02

    Technical deep dive interview

    60-90 minute interview with a senior backend engineer covering distributed systems design, API architecture decisions, concrete examples from your past work, and your approach to solving ambiguous technical problems at scale.

  3. 03

    System design exercise

    Interview focused on designing a large-scale distributed system relevant to multimodal APIs, such as architecting a real-time streaming system or designing an API for a complex model capability, including discussion of tradeoffs and operational considerations.

  4. 04

    Code review and implementation

    Technical discussion of your past code, design decisions you've made, or optional coding exercise focused on backend systems design rather than algorithmic puzzle-solving, emphasizing production-grade thinking and code quality standards.

  5. 05

    Cross-functional collaboration interview

    Conversation with product, research, or infrastructure team members to assess your ability to work across teams, translate between technical and non-technical contexts, and demonstrate collaborative problem-solving in ambiguous situations.

  6. 06

    Leadership and culture fit discussion

    Final-round conversation with a manager or director focusing on your sense of ownership, comfort with ambiguity, ability to raise engineering standards, and alignment with OpenAI's mission-driven culture and values.

Full posting

Original listing.

About the Team

API Multimodal builds the developer-facing products and infrastructure that bring OpenAI’s image, audio, and real-time model capabilities into the world. We are responsible for high-scale APIs for image generation, speech transcription, speech generation, and low-latency voice interactions. We partner closely with Research and Inference to bring frontier model capabilities to developers and use customer feedback to improve our models.

About the Role

As a software engineer on API Multimodal, you will build and operate the products and distributed systems behind OpenAI’s image, audio, and real-time APIs. You will work across model integration, API design, and production infrastructure to turn new research capabilities into reliable developer experiences. This hands-on role combines backend and systems depth with product judgment: you will own projects end to end, partner with Research, Inference, and Safety, and help make multimodal AI useful at scale. Model training experience is not required.

In this role, you will:

  • Design, build, and ship developer-facing APIs and backend services that serve frontier models.

  • Architect low-latency streaming, request, session, and model integration systems that make complex multimodal interactions reliable and intuitive at scale.

  • Work directly with Research to bring new model capabilities into production, shape the systems around them, and incorporate feedback from real-world developers and customers.

  • Own the availability, latency, scalability, and cost efficiency of the services you build.

  • Own projects from technical design and implementation through launch and ongoing iteration, while raising the team’s engineering standards.

Your background might look something like:

  • 7+ years of professional experience, excluding internships, in backend, infrastructure, platform, or product engineering roles.

  • A track record of designing, building, and operating production backend services, developer-facing APIs, or distributed systems.

  • Strong software engineering and systems fundamentals, with experience leading technically complex projects from ambiguous ideas to production.

  • Proficiency in one or more general-purpose backend languages, such as Python, Go, Rust, or TypeScript.

  • Experience building reliable, scalable systems and reasoning about distributed architecture, concurrency, latency, observability, and operational tradeoffs.

  • Product judgment and developer empathy, including an ability to turn complex model or infrastructure capabilities into clear, intuitive APIs.

  • Clear communication and a collaborative approach to working with researchers, product managers, designers, infrastructure engineers, and customers.

  • A strong sense of ownership, comfort with ambiguity, and a bias toward learning directly from users while continuously improving engineering quality.

  • Experience with real-time streaming, audio processing, speech systems, image generation, computer vision, or multimodal AI applications is helpful, but not required.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. 

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Redirects to OpenAI's application page.

Other roles

More at OpenAI.

View all 126 roles