Software Engineer, Core Services
Backend Engineer · Senior · Full Time
Opens OpenAI's application page
Role
What you'll do.
A Senior Software Engineer role at OpenAI's Core Services team building foundational infrastructure services including caching systems, workflow orchestration, and distributed metadata stores that serve as the backbone for OpenAI's AI products. The role focuses on designing highly reliable, scalable backend platforms using technologies like Redis, FoundationDB, and Temporal in San Francisco with a hybrid work model.
Responsibilities
- Infrastructure Design: Design and build shared infrastructure services including caching layers, workflow orchestration systems, metadata stores, and file storage services
- System Architecture: Architect highly reliable, scalable, and performant distributed systems that serve as the backbone for OpenAI's AI products
- Cross-team Collaboration: Collaborate with product engineering teams to provide scalable, reliable primitives that abstract distributed systems complexities
- Performance Optimization: Improve performance, resilience, and scalability of core services powering customer-facing AI applications
- API Development: Create well-designed APIs and abstractions that accelerate product development across teams
- System Maintenance: Maintain and monitor critical backend platforms ensuring high availability and performance
- Technical Leadership: Provide technical guidance on distributed systems best practices and infrastructure decisions
Qualifications
What we look for.
Technical
Distributed Systems
Extensive experience designing and implementing distributed systems with focus on consistency models and replication strategies
Caching Infrastructure
Hands-on experience with caching systems like Redis, Memcached, or similar technologies
Workflow Orchestration
Experience with workflow orchestration platforms such as Temporal or Cadence
Metadata Storage
Experience with distributed databases like FoundationDB or similar metadata storage solutions
Cloud Platforms
Proven experience running containerized services in cloud environments (AWS, GCP, or Azure)
CI/CD
Experience integrating services into automated build, test, and release workflows
Multi-region Systems
Understanding of trade-offs in consistency models and performance optimization in multi-region deployments
Education
Bachelor's Degree
Bachelor's degree in Computer Science, Software Engineering, or related technical field
Advanced Degree
Master's degree in Computer Science or equivalent experience preferred
Experience
Backend Development
5+ years of experience in backend software development with focus on distributed systems
Infrastructure Engineering
3+ years of experience building and maintaining large-scale infrastructure services
Production Systems
Experience operating high-availability production systems serving millions of users
Cross-functional Collaboration
Proven track record of collaborating effectively with product teams and stakeholders
Skills
Required
Distributed Systems Design
Deep understanding of distributed systems architecture, consistency models, and fault tolerance
Backend Programming
Strong proficiency in backend programming languages like Python, Go, or Java
Database Systems
Experience with both SQL and NoSQL databases, particularly distributed database systems
Containerization
Hands-on experience with Docker and Kubernetes for service deployment
API Design
Expertise in designing RESTful APIs and RPC interfaces
System Monitoring
Experience with observability tools and monitoring distributed systems
Preferred
AI/ML Infrastructure
Nice to haveExperience building infrastructure supporting AI/ML workloads
Temporal Workflows
Nice to haveHands-on experience with Temporal or similar workflow orchestration platforms
FoundationDB
Nice to haveExperience with FoundationDB or similar ACID-compliant distributed databases
Service Mesh
Nice to haveKnowledge of service mesh technologies like Istio or Linkerd
Infrastructure as Code
Nice to haveExperience with Terraform, Pulumi, or similar IaC tools
Performance Engineering
Nice to haveExpertise in performance optimization and capacity planning for large-scale systems
Tech stack
Languages
Frameworks
Databases
Tools
Other
Compensation
Pay and benefits.
Base·USD 230,000 – 385,000
Equity·Stock options
Benefits
Equity Compensation
Competitive equity package in one of the world's leading AI companies
Health Insurance
Comprehensive medical, dental, and vision coverage
Relocation Assistance
Full relocation support for new employees moving to San Francisco
Hybrid Work Model
Flexible hybrid work arrangement with 3 days in-office per week
Professional Development
Access to cutting-edge AI research and learning opportunities
Retirement Plans
401(k) retirement savings plan with company matching
Parental Leave
Generous parental leave policy for new parents
Wellness Programs
Mental health support and wellness benefits
Process
Interview steps.
- 01
Application Review
Initial screening of resume and portfolio focusing on distributed systems experience
- 02
Technical Phone Screen
45-minute technical discussion covering distributed systems concepts and past experience
- 03
System Design Interview
Design a distributed caching layer or workflow orchestration system similar to OpenAI's infrastructure
- 04
Coding Interview
Live coding session focusing on backend algorithms and data structures
- 05
Technical Deep Dive
In-depth discussion about a complex distributed system you've built or maintained
- 06
Team Fit Interview
Behavioral interview with hiring manager and team members
- 07
Final Interview
Leadership interview focusing on technical leadership and cross-functional collaboration
Full posting
Original listing.
About the Team
The Core Services team is responsible for building and managing foundational services. It acts as the bridge between core infrastructure (e.g. compute, storage, networking) and product engineering teams, and enables product teams to move fast, build reliably, and scale efficiently.
About the Role
As a software engineer in the core services team, you will design and operate critical backend platforms such as caching systems, workflow orchestration, metadata stores, and file services. You’ll focus on building highly reliable, scalable, and performant systems that serve as the backbone of our products.
We’re looking for people who are passionate about building infrastructure that empowers product teams, love working on distributed systems challenges, and enjoy creating well-designed APIs and abstractions that accelerate development.
This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.
In this role, you will:
Design, build, and maintain shared infrastructure services such as caching layers, workflow orchestration (Temporal), metadata stores, and file storage services.
Collaborate with product teams to provide scalable, reliable primitives that abstract the complexities of distributed systems.
Improve performance, resilience, and scalability of core services that power customer-facing applications.
You might thrive in this role if you:
Have experience with distributed systems, caching infrastructure (e.g., Redis, Memcached), metadata storage (e.g., FoundationDB), or workflow orchestration (e.g., Temporal, Cadence).
Have experience running containerized services in cloud environments and integrating them into automated build/test/release (CI/CD) workflows.
Understand trade-offs in consistency models, replication strategies, and performance optimization in multi-region systems.
Excel at communication and collaboration with cross-functional teams, and are obsessed with delivering customer success.
About OpenAI
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.
For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.
Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.
To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.
We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.
OpenAI Global Applicant Privacy Policy
At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
Redirects to OpenAI's application page.
Other roles
More at OpenAI.
Manager, Forward Deployed Engineer (FDE), Life Sciences
Manager
Manager, Forward Deployed Engineer - Tokyo
Manager
Software Engineer, Ads Integrity
Senior
Forward Deployed Engineer - Zurich
Senior
Software Engineer, Conversion Measurement
Senior