Software Engineer, API Safety

Backend Engineer · Senior · Full Time

San FranciscoUSD 293k – 385k5d ago
Apply for this role

Opens OpenAI's application page

Role

What you'll do.

Join OpenAI's API Safety team as a Senior Software Engineer to design and build critical safeguards for frontier AI models in production. This role bridges safety research and developer experience, requiring expertise in backend systems, distributed architecture, and API design to ensure responsible deployment of transformative AI capabilities at scale.

Responsibilities

  • Design and Build Safety Infrastructure: Architect and implement dashboards and APIs that expose safety controls and provide customer-facing observability into model behavior and usage patterns. Create intuitive interfaces that translate complex safety and infrastructure capabilities into clear, accessible APIs for developers building on OpenAI's platform.
  • Develop Scalable Safeguard Systems: Engineer scalable backend systems that extend trusted safety capabilities to new use cases, customer segments, and deployment environments. Focus on building systems that can reliably handle rapid scaling while maintaining safety guarantees and performance characteristics.
  • Collaborate on Risk Mitigation: Partner closely with Safety Research and Integrity teams to translate emerging risk insights into robust safeguards and technical controls. Contribute to the design and implementation of systems that proactively mitigate frontier AI risks at the platform level.
  • Ensure Production Reliability and Performance: Take ownership of the availability, latency, and scalability of safeguard systems across high-volume API traffic. Make informed architectural decisions balancing developer experience, system latency, reliability, and risk posture in production environments serving millions of requests.
  • Lead Full Lifecycle Project Ownership: Own projects end-to-end from technical design and implementation through launch, iteration, and ongoing optimization. Raise engineering standards across the team by demonstrating best practices in distributed systems, testing, monitoring, and operational excellence.
  • Build Developer-Centric APIs: Apply product judgment and deep developer empathy to create intuitive, well-designed APIs that help developers understand safety events, share usage context, and apply risk-appropriate safeguards. Focus on APIs that balance power with usability for diverse developer skill levels.

Qualifications

What we look for.

Technical

  • Backend Language Proficiency

    Advanced proficiency in one or more general-purpose backend languages such as Python, Go, Rust, or TypeScript. Ability to write performant, maintainable code and make informed language selection decisions based on architectural requirements.

  • Distributed Systems Architecture

    Deep understanding of distributed systems principles including consistency models, fault tolerance, scalability patterns, and service communication protocols. Ability to design systems that operate reliably across multiple availability zones and geographic regions.

  • API Design and Developer Experience

    Expertise in designing robust, intuitive APIs that serve as the contract between systems. Understanding of API versioning, backward compatibility, rate limiting, authentication, and error handling. Ability to create developer experiences that minimize friction and maximize clarity.

  • Production Systems and Observability

    Hands-on experience with logging, metrics, tracing, alerting, and monitoring systems. Ability to instrument systems effectively for visibility and troubleshooting. Familiarity with on-call responsibilities and incident response in mission-critical environments.

  • Scalability and Performance Engineering

    Experience optimizing systems for performance, throughput, and latency at scale. Ability to profile code, identify bottlenecks, and make architectural decisions that support 10x or 100x growth in traffic and data volume.

Education

  • Computer Science or Related Field

    Bachelor's degree in Computer Science, Software Engineering, or related technical field, or equivalent professional experience demonstrating mastery of foundational computer science concepts including algorithms, data structures, and system design.

Experience

  • Senior Backend or Infrastructure Engineering

    7+ years of professional software engineering experience (excluding internships) in backend, infrastructure, platform, or product engineering roles. Demonstrated track record of designing, building, and operating production-grade backend services, developer-facing APIs, or distributed systems in high-scale environments.

  • Distributed Systems and Production Operations

    Proven expertise leading technically complex projects from ambiguous requirements through production deployment and ongoing iteration. Strong ability to reason about distributed architecture, concurrency patterns, latency optimization, observability solutions, and operational tradeoffs in production systems.

  • Systems Design and Technical Complexity

    Demonstrated ability to tackle and resolve technically complex challenges at scale. Experience designing systems that balance multiple competing requirements and can evolve as requirements change or scale increases significantly.

  • Safety, Security, or Compliance Systems (Preferred)

    Experience working with observability platforms, abuse prevention systems, trust and safety infrastructure, identity systems, security controls, privacy frameworks, compliance systems, or risk management platforms is highly valued. Understanding of how safety and compliance requirements impact system design is a strong differentiator.

Skills

Required

  • Python or Go

    Strong proficiency in Python, Go, or similar general-purpose backend languages. Ability to write production-quality code that is performant, maintainable, and well-tested.

  • Distributed Systems Design

    Deep expertise in designing and reasoning about distributed systems, including understanding of trade-offs in consistency models, availability, fault tolerance, and scalability at large scale.

  • Backend Service Architecture

    Proven experience designing and building microservices, REST APIs, or platform services. Understanding of service boundaries, inter-service communication, resilience patterns, and dependency management.

  • Production Reliability Engineering

    Hands-on experience ensuring systems are reliable, observable, and maintainable in production. Familiarity with monitoring, alerting, logging, and incident response in mission-critical environments.

  • API Design

    Demonstrated ability to design clear, intuitive APIs that balance developer needs with technical constraints. Understanding of REST principles, error handling, versioning, and creating delightful developer experiences.

  • Software Engineering Fundamentals

    Strong grasp of core software engineering principles including testing strategies, code review practices, design patterns, and technical decision-making. Ability to communicate technical concepts clearly to diverse audiences.

Preferred

  • Safety and Compliance Systems

    Nice to have

    Experience designing or implementing safety systems, abuse prevention platforms, content moderation infrastructure, identity verification systems, or compliance frameworks. Understanding of how policy requirements translate into technical systems.

  • High-Scale Backend Systems

    Nice to have

    Experience building systems that handle billions of requests or petabyte-scale data volumes. Demonstrated ability to optimize for latency, throughput, and cost at massive scale while maintaining reliability.

  • Observability and Monitoring

    Nice to have

    Expertise with observability platforms like Datadog, New Relic, Prometheus, or ELK stack. Ability to design comprehensive monitoring and alerting strategies that enable rapid incident detection and diagnosis.

  • Rust or Go

    Nice to have

    Experience with systems-level languages like Rust or Go that enable high-performance, concurrent applications. Understanding of memory safety, concurrency primitives, and performance optimization at a lower level.

  • Machine Learning Systems

    Nice to have

    Familiarity with ML pipelines, model serving infrastructure, or AI platform development. Understanding of how model behavior impacts system design and operational concerns.

  • API Platform Development

    Nice to have

    Experience building API platforms or developer tools that serve as infrastructure for other teams. Understanding of platform reliability, documentation, developer onboarding, and long-term API evolution.

  • Cloud Infrastructure and Deployment

    Nice to have

    Proficiency with cloud platforms (AWS, GCP, Azure), containerization (Docker, Kubernetes), and infrastructure-as-code approaches. Experience managing infrastructure at scale and optimizing for cost and performance.

Tech stack

Languages

PythonGoRustTypeScript

Frameworks

FastAPI or FlaskgRPCAsync/Await Patterns

Databases

PostgreSQLRedisCloud Data Warehouses

Tools

KubernetesDockerDatadog or Similar APMGit and GitHubCI/CD Systems

Other

API Safety and Content Moderation SystemsLoad Testing and Chaos EngineeringIncident Response and Postmortems

Compensation

Pay and benefits.

Base·USD 293,000 – 385,000

Equity·Stock options

Benefits

  • Comprehensive Health Coverage

    Medical, dental, and vision insurance with plans covering preventive care, specialists, and major treatments. Coverage extends to dependents with employer subsidies for premiums.

  • Equity Compensation

    Meaningful equity grants that allow you to participate in OpenAI's growth and success. Equity vests over time, aligning long-term incentives with company performance.

  • Retirement Savings

    401(k) plan with employer matching contributions, enabling tax-advantaged retirement savings and wealth building throughout your career.

  • Generous Paid Time Off

    Flexible paid time off policy supporting work-life balance, mental health, and personal development. Separate pool for sick leave and parental leave.

  • Parental Leave

    Comprehensive parental leave program for birth parents and non-birth parents, supporting family planning and bonding time with newborns.

  • Professional Development

    Learning budget, conference attendance support, and opportunities to develop new skills. Access to educational resources and mentorship from experienced engineers.

  • Home Office Setup

    Equipment stipends for setting up a productive home office, including ergonomic furniture, monitors, and other technology essentials.

  • Wellness Programs

    Gym memberships, mental health resources, meditation apps, and wellness coaching. Support for physical and mental well-being.

  • Life Insurance and Disability Coverage

    Life insurance and short/long-term disability coverage providing financial protection for you and your family.

  • Commuter and Transportation Benefits

    Pre-tax commuter benefits and support for public transportation or parking. Subsidies for electric vehicles at San Francisco headquarters.

  • Free Meals and Snacks

    Complimentary meals and beverages at headquarters, supporting employee nutrition and fostering community spaces for collaboration.

  • Flexible Work Arrangements

    While based in San Francisco, flexible scheduling and remote work options for focused work. Collaboration spaces and team presence expectations calibrated for productivity.

Full posting

Original listing.

About the Team

OpenAI's mission is to ensure that artificial general intelligence (AGI) benefits all of humanity. The API Platform turns frontier research into reliable capabilities that developers use to build transformative products and services for people around the world.

API Safety's goal is to ensure safe deployment of frontier models in the API. We design APIs and systems that help developers share usage context, understand safety events, and apply safeguards tailored to the risk profile of the applications they are building. This work is critical to our frontier model launches and partners closely with teams across API, Integrity, and Safety Research.

About the Role

We're looking for product-minded software engineers to join a team that is addressing emerging risks at the frontier of model development while building novel solutions for real-world AI deployment. The day-to-day work ranges from solving production challenges to designing new product experiences and safeguards. The right candidate is comfortable balancing tradeoffs across developer experience, latency, reliability, and risk.

In this role, you will:

  • Design and build dashboards and APIs for safety controls and customer-facing observability.

  • Develop scalable systems that extend trusted safety capabilities to new use cases, customers, and deployment environments.

  • Partner with Safety Research and Integrity to build safeguards that mitigate emerging risks.

  • Be responsible for the availability, latency, and scalability of safeguards across high-volume API traffic.

  • Own projects from technical design and implementation through launch and ongoing iteration, while raising the team’s engineering standards

Your background might look something like:

  • 7+ years of professional experience, excluding internships, in backend, infrastructure, platform, or product engineering roles.

  • A track record of designing, building, and operating production backend services, developer-facing APIs, or distributed systems.

  • Strong software engineering and systems fundamentals, with experience leading technically complex projects from ambiguous ideas to production.

  • Proficiency in one or more general-purpose backend languages, such as Python, Go, Rust, or TypeScript.

  • Experience building reliable, scalable systems and reasoning about distributed architecture, concurrency, latency, observability, and operational tradeoffs.

  • Product judgment and developer empathy, including an ability to turn complex model or infrastructure capabilities into clear, intuitive APIs.

  • Clear communication and a collaborative approach to working with researchers, product managers, designers, infrastructure engineers, and customers.

  • Experience with observability, abuse prevention, trust and safety, identity, security, privacy, compliance, or risk systems is a strong plus.

  • Interest in helping frontier AI reach more developers and businesses responsibly.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. 

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Redirects to OpenAI's application page.

Other roles

More at OpenAI.

View all 107 roles