Platform Infrastructure Engineer
Senior · Full Time
Opens Snowflake's application page
Role
What you'll do.
Join Snowflake's Cloud Infrastructure Engineering team to build and scale the Data Cloud Platform supporting 3.9 billion queries daily across multiple cloud providers. As a Platform Infrastructure Engineer, you'll design highly resilient, performant infrastructure systems while providing technical leadership on large-scale projects that solve complex cloud infrastructure challenges. This role requires strong software engineering fundamentals combined with cloud platform expertise, offering the opportunity to mentor junior engineers and drive adoption of infrastructure platforms at exceptional scale.
Responsibilities
- Design and Build Scalable Infrastructure Systems: Develop highly scalable, resilient, and performant infrastructure services that support Snowflake's globally deployed platform spanning multiple cloud providers. Build and optimize systems managing hundreds of thousands of VMs while maintaining exceptional reliability standards to support 3.9 billion daily queries and 515 million data workloads.
- Establish Deep Infrastructure Understanding: Develop comprehensive knowledge of Snowflake's cloud infrastructure architecture, service dependencies, and operational characteristics. This foundational expertise enables you to make informed architectural decisions and guide the team's technical direction across cloud infrastructure domains.
- Provide Technical Leadership on Complex Projects: Lead large-scale infrastructure engineering initiatives focused on building platforms, tools, automation, and cloud infrastructure solutions. Demonstrate both infrastructure expertise and software engineering excellence by driving projects from conception through production deployment while managing technical complexity and cross-team dependencies.
- Drive Platform Adoption and Evangelize Infrastructure Services: Champion infrastructure platform adoption across development teams by demonstrating business value, facilitating integration, and gathering requirements. Evangelize best practices and new capabilities to enable other teams to build upon simplified infrastructure abstractions while meeting organizational business goals.
- Mentor Junior Engineers and Establish Best Practices: Mentor and support junior team members by modeling excellent code quality, comprehensive documentation, and software development best practices. Lead by example in implementing reliable infrastructure patterns, conducting thorough code reviews, and fostering a culture of operational excellence and continuous learning.
- Maintain High Service Level Objectives: Champion infrastructure reliability and performance by actively maintaining and improving SLOs across infrastructure services. Work to protect Snowflake's platform from cloud provider issues, optimize infrastructure efficiency, and manage complex infrastructure changes at massive scale with strong operational oversight.
- Troubleshoot and Resolve Complex Technical Issues: Investigate and resolve challenging infrastructure problems spanning cloud platforms, networking, containerization, and operating system layers. Apply systematic debugging approaches and deep technical knowledge to understand root causes and implement resilient solutions that prevent recurrence.
Qualifications
What we look for.
Technical
Container Platform Architecture
Deep knowledge of containerization technologies, orchestration platforms, and microservices deployment patterns. Experience designing and managing container infrastructure at scale, including container networking, storage, and lifecycle management across production environments.
Cloud Infrastructure and Multi-Cloud Expertise
Minimum 1+ years hands-on experience with AWS, Azure, or GCP cloud platforms. Proficiency in cloud resource provisioning, networking, security, cost optimization, and multi-cloud architecture patterns to manage Snowflake's globally distributed infrastructure across providers.
Infrastructure as Code and Automation
Production experience with Infrastructure as Code tools such as Terraform and Pulumi. Strong ability to define infrastructure through code, manage configuration at scale, implement CI/CD pipelines, and automate infrastructure deployment and management processes.
System Design and Software Architecture
Demonstrated expertise in designing large-scale distributed systems, building resilient service architecture, implementing fault tolerance patterns, and managing complex infrastructure state. Ability to evaluate tradeoffs between reliability, performance, cost, and operational simplicity.
Networking and Operating Systems Knowledge
Strong understanding of networking fundamentals including DNS, TCP/IP, load balancing, and service mesh technologies. Proficiency with Linux operating systems, kernel concepts, system performance optimization, and OS-level configuration management for production infrastructure.
Site Reliability Engineering Principles
Comprehensive knowledge of SRE practices including SLO definition, monitoring, alerting, incident response, and blameless postmortem culture. Experience building observability systems, implementing runbooks, and establishing operational frameworks for highly available services.
Education
Computer Science Degree
Bachelor's degree in Computer Science, or equivalent professional experience. Advanced degree (Master's in Computer Science) preferred to demonstrate advanced technical knowledge and research capability.
Experience
Cloud Infrastructure Platform Experience
Minimum 2+ years of hands-on experience building and supporting mission-critical services and infrastructure within cloud-native SaaS environments. Experience designing for scale, implementing reliability improvements, and managing infrastructure across production deployments.
Software Engineering Fundamentals
Strong foundational software engineering expertise including proficiency in multiple programming languages (Golang, Java, Python, or C), solid grasp of data structures and algorithms, design patterns, and software development lifecycle best practices such as testing, documentation, and code review processes.
Leadership and Mentorship
Proven ability to lead technical projects, make architectural decisions independently, and mentor junior team members. Experience communicating technical concepts across teams, driving adoption of new technologies, and influencing organizational technical direction.
Skills
Required
Golang
Proficiency in Go programming language for building efficient, concurrent infrastructure tooling and services. Experience with goroutines, channels, and Go's approach to systems programming for cloud infrastructure development.
Terraform
Hands-on expertise with Terraform for Infrastructure as Code, including state management, modules, and managing complex infrastructure configurations at scale across multiple cloud providers.
Kubernetes or Container Orchestration
Deep knowledge of Kubernetes architecture, pod networking, storage orchestration, and cluster management. Experience deploying and operating containerized workloads in production environments with focus on reliability and performance.
AWS, Azure, or GCP
Production experience with at least one major cloud platform, including compute services, networking, storage, security, and cost optimization. Understanding of cloud-specific infrastructure patterns and managed services.
Linux Systems Administration
Strong proficiency with Linux operating systems, shell scripting, system configuration, performance tuning, and troubleshooting. Experience managing Linux servers in production environments and optimizing system-level performance.
System Design and Architecture
Ability to design scalable, reliable, and maintainable infrastructure systems. Experience with distributed systems concepts, resilience patterns, and making architectural tradeoffs for production systems at scale.
Troubleshooting and Debugging
Advanced ability to systematically diagnose complex technical issues across infrastructure layers. Proficiency with debugging tools, log analysis, and methodical problem-solving approaches for production incident resolution.
Preferred
Java
Nice to haveExperience with Java for systems development and infrastructure tooling. Knowledge of Java's concurrency models and ecosystem for building scalable backend services.
Python
Nice to haveProficiency in Python for infrastructure automation, scripting, and building operational tools. Experience with automation frameworks and infrastructure orchestration tooling.
Pulumi
Nice to haveExperience with Pulumi as an alternative Infrastructure as Code approach, enabling infrastructure definition using general-purpose programming languages.
Service Mesh Technologies
Nice to haveFamiliarity with service mesh platforms such as Istio for managing service-to-service communication, traffic management, and observability in microservices architectures.
Multi-Cloud Architecture
Nice to haveExperience designing and operating infrastructure across multiple cloud providers, understanding cloud-specific differences, and implementing portable infrastructure patterns.
Observability and Monitoring
Nice to haveExperience building comprehensive monitoring, logging, and tracing infrastructure. Knowledge of observability platforms, metrics collection, and alerting systems for production environments.
Configuration Management
Nice to haveExperience with configuration management tools such as Ansible, Chef, or Puppet for managing infrastructure state and ensuring consistent deployments across environments.
CI/CD Pipeline Development
Nice to haveExperience designing and implementing continuous integration and deployment pipelines for infrastructure code, including testing automation and deployment orchestration.
Tech stack
Languages
Frameworks
Databases
Tools
Other
Compensation
Pay and benefits.
Base·USD 165,000 – 245,000
Benefits
Competitive Health and Wellness Benefits
Comprehensive health insurance coverage including medical, dental, and vision plans. Access to wellness programs, fitness benefits, and mental health support services to ensure employee wellbeing.
Retirement Planning and Financial Security
401(k) retirement plan with employer matching contributions to support long-term financial security and retirement planning for eligible employees.
Generous Time Off and Work-Life Balance
Flexible paid time off policies, paid holidays, and parental leave programs designed to promote work-life balance and support major life events.
Professional Development and Learning Opportunities
Career development programs, professional certifications, training budgets, and internal mentorship opportunities to support continuous skill advancement and career growth within Snowflake.
Equity Compensation Program
Stock option packages aligned with company success, providing employees with ownership stake and incentive to drive long-term value creation at Snowflake.
Collaborative and Innovative Work Environment
Dynamic, fast-moving team culture emphasizing experimental mindset, AI-native thinking, and low-ego collaboration. Access to cutting-edge cloud infrastructure technologies and opportunities to influence the future of data cloud platforms.
Global Scale Impact Opportunities
Direct involvement in solving infrastructure challenges at exceptional scale, supporting 3.9 billion daily queries across multiple cloud providers managing hundreds of thousands of VMs globally.
Full posting
Original listing.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.
Snowflake customers run more than 3.9 billion queries and 515 million data workloads each day. To support this workload, our globally deployed infrastructure spans multiple cloud providers and manages hundreds of thousands of VMs. Join the Snowflake team to build the Data Cloud Platform at an exceptional scale.
We are hiring talented Software Engineers for our Cloud Infrastructure Engineering teams. If you are passionate about solving complex infrastructure challenges by using software engineering expertise, this opportunity may be for you. Our Cloud Infrastructure Engineering Teams are focused on things like:
simplifying our large-scale infrastructure for the other development teams to use
building services that improve the reliability of Snowflake by protecting us from the cloud providers issues
enabling Snowflake services to operate globally across multiple cloud providers
managing complicated infrastructure changes at the scale a few companies can match
keeping the SLOs of our infrastructure and services at a very high bar
defining the OS and containerization layer for all of Snowflake’s services
All Cloud Infrastructure Engineering Teams aim to optimize and evolve our infrastructure's reliability, availability, serviceability, performance, and cost efficiency.
AS A SOFTWARE ENGINEER - CLOUD INFRASTRUCTURE ENGINEERING AT SNOWFLAKE, YOU WILL:
Will build a deep understanding of Snowflake’s infrastructure and services
Demonstrate both infrastructure and software engineering expertise by contributing to the team charter to build and operate highly scalable, resilient, and performant infrastructure.
Provide technical leadership on large and complex projects that build infrastructure, platforms, tools, and automation in the cloud.
Evangelize and drive adoption of the Platform to meet business goals.
Mentor and support more junior team members, leading by example with excellent code quality, documentation, and software development best practices
OUR IDEAL SOFTWARE ENGINEER - CLOUD ENGINEERING AT SNOWFLAKE, WILL HAVE:
BS/CS, MS/CS, or equivalent.
At least 2+ years experience in a platform or cloud team building spent on supporting mission-critical services and infrastructure in a SaaS environment.
Strong software engineering fundamentals, coding skills, and knowledge of software development best practices.
At least 1+ years in cloud computing (AWS, Azure, or GCP).
Fluent in one or more languages (Golang, Java, Python, C).
Expertise in at least one of the following areas: container platforms, automation, networking, operating systems, site reliability, config management, and infrastructure as code solutions such as Pulumi and Terraform.
Tremendous attention to detail and ability to build reliable and scalable software systems.
Effective communication and collaboration skills.
Ability to troubleshoot and resolve complex technical issues.
A strong work ethic, ability to self-manage and drive project success, and a passion for problem-solving.
Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.
How do you want to make your impact?
Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.
How do you want to make your impact?
For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com
Redirects to Snowflake's application page.
Other roles
More at Snowflake.
Software Engineer, Full Stack - Marketplace
Mid
Staff Software Engineer - Snowhouse
Staff
Principal Industry Architect, FINS
Principal
Senior Manager - Engineering Systems AI Developer Experience
Manager
Engineering Manager - Cost Intelligence
Manager