# Senior Infrastructure Engineer
**Company:** [Tennr](https://scaleengineer.com/companies/tennr)
As a Senior Infrastructure Engineer at Tennr, you'll architect and maintain the cloud-native infrastructure foundation that powers an AI-driven healthcare automation platform. This high-ownership role combines strategic infrastructure design with hands-on execution, requiring expertise in Kubernetes (EKS), Terraform, AWS, and observability platforms to support a rapidly scaling system processing critical healthcare data.
**Role:** Infrastructure Engineer
**Seniority:** Senior
**Locations:** New York City Office
**Salary:** 200000–230000 USD
[Apply](https://jobs.ashbyhq.com/tennr/c5928e76-7a96-45bb-989b-9f7207031a8b)
Canonical: https://scaleengineer.com/jobs/tennr/senior-infrastructure-engineer
---
## Responsibilities

- Architect and Build EKS Foundation: Design, implement, and evolve the Kubernetes (EKS) cluster architecture that serves as the foundation for all engineering teams. Create reusable modules and patterns that establish best practices for cluster configuration, networking, and security, enabling teams across the organization to inherit consistent, production-grade standards.
- Own Multi-Environment Cluster Management: Manage the complete lifecycle of dev, staging, and production Kubernetes environments end-to-end. Ensure cluster health, performance optimization, capacity planning, and smooth operations across all environments while maintaining reliability and cost efficiency as infrastructure scales.
- Lead Infrastructure Migrations: Plan, execute, and unblock infrastructure migrations with minimal disruption to production systems. Take ownership of complex migration strategies, staged rollouts, risk mitigation, and communication to ensure smooth transitions while maintaining system stability and data integrity throughout the process.
- Consolidate and Standardize Observability: Lead the migration and consolidation of infrastructure monitoring and observability onto Datadog. Establish comprehensive monitoring standards, dashboards, alerting strategies, and logging practices that the entire engineering organization adopts, creating visibility into system health and performance.
- Manage AWS Infrastructure as Code: Own the design, provisioning, and maintenance of core AWS infrastructure including networking (VPCs, security groups, load balancing), identity and access management (IAM policies and roles), and essential services. Implement all infrastructure through code (Terraform or similar IaC tools) with emphasis on security-by-default configurations and cost optimization.
- Drive Security and Cost Optimization: Implement and enforce security best practices across the AWS footprint, ensuring compliance with healthcare data regulations (HIPAA/PHI where applicable). Continuously monitor and optimize cloud spending, rightsizing resources and identifying cost reduction opportunities without compromising performance or reliability.

## Requirements

### education

- {"name":"Bachelor's Degree in Computer Science or Related Field","description":"Formal education in Computer Science, Computer Engineering, Information Systems, or equivalent discipline providing foundational knowledge in software systems and infrastructure."}
- {"name":"Professional Certifications (Preferred)","description":"AWS Solutions Architect Associate/Professional, Certified Kubernetes Administrator (CKA), or Terraform Associate certifications demonstrate formal commitment to cloud and infrastructure expertise."}

### technical

- {"name":"Kubernetes (EKS) Administration","description":"Demonstrated hands-on experience deploying, managing, and scaling production Kubernetes clusters, preferably with AWS EKS. Should have deep knowledge of cluster networking, container orchestration, resource management, and operational patterns."}
- {"name":"Infrastructure as Code","description":"Proficiency with Terraform or equivalent IaC tools (CloudFormation, CDK) for defining and managing cloud infrastructure. Experience writing modular, maintainable, and well-documented infrastructure code that scales across multiple environments."}
- {"name":"AWS Cloud Architecture","description":"Strong foundational knowledge of AWS services including VPC networking, IAM policies and roles, EC2, RDS, S3, and other core compute and storage services. Understanding of AWS security best practices and ability to design secure, scalable infrastructure."}
- {"name":"Observability and Monitoring","description":"Hands-on experience with Datadog or comparable observability platforms (Prometheus, New Relic, Splunk). Ability to design monitoring strategies, create meaningful dashboards, configure alerting, and interpret metrics to drive operational decisions."}
- {"name":"Linux and Container Fundamentals","description":"Solid understanding of Linux operating systems, container technology (Docker), and container registry management. Knowledge of system administration, process management, networking protocols, and troubleshooting."}

### experience

- {"name":"Production Infrastructure Ownership","description":"5-8 years of professional experience in infrastructure engineering, platform engineering, or DevOps roles with significant responsibility for production systems. Demonstrated track record of managing mission-critical infrastructure and making architectural decisions that impact engineering velocity."}
- {"name":"Kubernetes Production Operations","description":"Proven experience running Kubernetes in production environments, including cluster upgrades, security patching, troubleshooting performance issues, and managing stateful and stateless workloads at scale."}
- {"name":"Autonomous Project Execution","description":"Experience executing complex infrastructure projects independently with minimal oversight. Ability to own projects end-to-end, from design through implementation to documentation and knowledge transfer."}
- {"name":"Cross-functional Collaboration","description":"Experience working effectively with product engineering, security, and operations teams. Ability to establish standards and best practices that other teams adopt and build upon."}

## Skills

### required

- {"name":"Kubernetes and Container Orchestration","description":"Expert-level knowledge of Kubernetes architecture, deployments, services, ingress, persistent volumes, and operational patterns. Experience with EKS specifically preferred."}
- {"name":"Terraform and Infrastructure as Code","description":"Proficiency writing production-grade Terraform modules, managing state, implementing CI/CD for infrastructure changes, and organizing code for scalability and reusability."}
- {"name":"AWS Services and Architecture","description":"Deep knowledge of VPC networking, IAM, security groups, load balancing, DNS, and core compute services. Ability to architect secure, scalable infrastructure on AWS."}
- {"name":"System Design and Architecture","description":"Ability to design resilient, scalable infrastructure systems with consideration for high availability, disaster recovery, performance, and cost optimization."}
- {"name":"Datadog or Equivalent Observability","description":"Hands-on experience configuring monitoring, creating dashboards, setting up alerting, and using observability data to troubleshoot and optimize systems."}
- {"name":"Linux System Administration","description":"Proficiency with Linux operating systems, command-line tools, networking concepts, security fundamentals, and troubleshooting production systems."}
- {"name":"Autonomous Problem Solving","description":"Demonstrated ability to identify problems, research solutions, execute without extensive guidance, and learn new tools and technologies independently."}

### preferred

- {"name":"Healthcare or Regulated Industry Experience","description":"Prior experience in healthcare, finance, or other regulated industries with exposure to compliance requirements like HIPAA, PHI, or data protection regulations."}
- {"name":"CI/CD Pipeline Experience","description":"Hands-on experience implementing and managing CI/CD pipelines using tools like GitHub Actions, GitLab CI, Jenkins, or ArgoCD for automated infrastructure and application deployments."}
- {"name":"Disaster Recovery and Business Continuity","description":"Experience designing, implementing, and testing disaster recovery strategies, backup solutions, and failover mechanisms for production systems."}
- {"name":"Cost Optimization Expertise","description":"Proven ability to analyze cloud spending, identify optimization opportunities, implement cost-saving measures, and establish chargeback or accountability models."}
- {"name":"Security and Compliance Implementation","description":"Experience implementing security controls, conducting vulnerability assessments, managing secrets and credentials, and ensuring infrastructure meets compliance standards."}
- {"name":"Cloud Migration Experience","description":"Experience planning and executing large-scale infrastructure migrations between cloud providers or from on-premises to cloud with minimal downtime."}
- {"name":"Prometheus and Advanced Monitoring","description":"Familiarity with Prometheus, Grafana, or other open-source observability stacks for metrics collection, alerting, and performance analysis."}

## Tech stack

### tools

- {"name":"AWS CLI","description":"Command-line interface for managing AWS services programmatically and automating cloud operations."}
- {"name":"Datadog","description":"Observability platform for monitoring infrastructure, applications, and user experience with metrics, logs, and traces."}
- {"name":"Git and GitHub","description":"Version control system and collaborative platform for managing infrastructure code, enabling code review and collaboration."}
- {"name":"Docker","description":"Container runtime and tooling for building, managing, and deploying containerized applications on Kubernetes."}
- {"name":"ArgoCD or similar GitOps Tools","description":"GitOps tools for declarative, version-controlled infrastructure and application deployment management."}

### others

- {"name":"AWS VPC and Networking","description":"Virtual Private Cloud setup, subnet design, security groups, network ACLs, and advanced networking patterns for multi-environment infrastructure."}
- {"name":"IAM and Access Management","description":"AWS Identity and Access Management for implementing least-privilege access, role-based access control (RBAC), and secure credential management."}
- {"name":"HIPAA and Healthcare Compliance","description":"Understanding of healthcare industry compliance requirements, data protection standards (PHI), and regulated infrastructure design patterns."}
- {"name":"Load Balancing and DNS","description":"Application load balancers, network load balancers, Route53 DNS management, and traffic management strategies for high-availability infrastructure."}
- {"name":"Infrastructure Automation and Scripting","description":"Automation of repetitive infrastructure tasks, infrastructure testing frameworks, and operational runbooks using modern DevOps practices."}

### databases

- {"name":"Amazon RDS","description":"AWS managed relational database service for production databases. Experience with provisioning, scaling, and managing RDS instances."}
- {"name":"Amazon DynamoDB","description":"AWS NoSQL database service for high-scale applications. Understanding of how applications interact with managed databases on AWS."}
- {"name":"Redis/ElastiCache","description":"In-memory data store and caching layer. Experience with AWS ElastiCache or self-managed Redis in Kubernetes environments."}

### languages

- {"name":"Bash/Shell Scripting","description":"Shell scripting for automation, infrastructure tooling, and operational tasks. Essential for Linux administration and infrastructure engineering."}
- {"name":"Python","description":"Python for infrastructure automation, tooling, Terraform custom providers, and operational scripts that enhance cloud operations."}
- {"name":"Go","description":"Go language experience valuable for working with cloud-native tools, Kubernetes controllers, and modern infrastructure tooling written in Go."}

### frameworks

- {"name":"Kubernetes","description":"Container orchestration platform for managing containerized workloads at scale. Core to Tennr's infrastructure foundation and cluster management."}
- {"name":"Helm","description":"Package manager for Kubernetes enabling templated deployments, version management, and standardized application deployment patterns."}
- {"name":"Terraform","description":"Infrastructure as Code tool for provisioning and managing AWS resources in a declarative, version-controlled manner."}

## Benefits

### benefits

- {"name":"100% Paid Employee Health Benefits","description":"Comprehensive health insurance coverage with employer funding full employee premium for medical, dental, and vision coverage options."}
- {"name":"Unlimited PTO","description":"Flexible time-off policy with unlimited paid time off to support work-life balance and personal well-being without tracking limits."}
- {"name":"Employer-Funded 401(k) Match","description":"Retirement savings plan with employer matching contributions, enabling competitive long-term wealth building and financial security."}
- {"name":"Competitive Parental Leave","description":"Generous parental leave policy supporting both primary and secondary caregivers with paid time off to support family planning and new parents."}
- {"name":"Modern Office at 345 Hudson Street","description":"Beautiful, newly designed office space in Hudson Square, NYC with full amenities and collaborative work environment for in-person collaboration."}
- {"name":"Complimentary Lunch and Snacks","description":"Free lunch provided daily plus fully stocked office pantry with snacks and beverages supporting employee wellness and convenience."}
- {"name":"Equity Compensation","description":"Participation in company equity program enabling employees to share in long-term company success and growth as early-stage founders."}

## Compensation

- **max:** 220000
- **min:** 160000
- **currency:** USD
- **stockOptions:** true

## Interview process

### steps

- {"name":"Initial Screening Call","description":"Introductory conversation with Tennr's recruiting team to discuss background, experience with Kubernetes and AWS, and alignment with the Senior Infrastructure Engineer role and company mission."}
- {"name":"Technical Infrastructure Assessment","description":"Comprehensive technical interview covering infrastructure design, Kubernetes architecture, AWS services, Terraform code review, and real-world problem-solving scenarios relevant to platform engineering."}
- {"name":"Infrastructure Systems Design Round","description":"Deep-dive technical discussion on designing scalable, secure infrastructure systems. Expected to discuss trade-offs, architectural patterns, observability strategy, and experience with multi-environment cluster management."}
- {"name":"Team and Leadership Conversation","description":"Meeting with infrastructure team lead and engineering leadership to discuss collaboration style, team dynamics, mentorship approach, and how you establish standards and best practices for cross-functional teams."}
- {"name":"Founder or Executive Alignment","description":"Conversation with company founders or executive leadership to discuss company vision, product roadmap, scaling challenges, and how infrastructure decisions impact the healthcare automation mission."}
- {"name":"Offer and Equity Discussion","description":"Final stage covering compensation details, equity vesting schedule, benefits, start date, and any final questions about the role or company culture at Tennr."}

## Full description
**Company Description**

Today, when you go to your doctor and get referred to a specialist, your doctor sends out a referral and tells you, “They’ll be in touch soon.” So you wait. And wait. Sometimes days, weeks, or even months. Why? Because too often providers are overwhelmed with the painstakingly tedious work required to get paid by insurance companies. Powered by proprietary models, Tennr handles the complex paperwork that gets patients through the door and providers paid, helping operators get patients the right care, at the right time, in the right setting.

**Role Description**

You’ll be the senior builder on Tennr’s Infrastructure team, owning the cloud, clusters, and environments that everything else runs on. This is a hands-on, high-ownership role: you’ll help design our Kubernetes foundation, run dev, staging, and production end to end, and keep our AWS footprint clean, secure, and reliable as we scale.

We’re looking for someone who executes fast but doesn’t just take orders. The patterns you set and the guardrails you establish are the ones the rest of engineering will inherit.

**Responsibilities**

* Help build out our EKS cluster module and the patterns other teams inherit from it.
* Own cluster management and the dev, staging, and production build-out end to end.
* Drive and unblock infrastructure migrations, keeping them safe, staged, and low-drama.
* Lead the consolidation of our infrastructure observability onto Datadog and set the standards other teams build on.
* Own AWS building and maintenance - networking, IAM, and core services as infrastructure-as-code. Keeping the environment secure-by-default and cost-aware.

**Candidate Qualifications**

* 5–8 years in infrastructure, platform, or DevOps engineering, with real ownership of production systems.
* Hands-on depth with Kubernetes (EKS preferred) and Terraform or comparable infrastructure-as-code.
* Solid AWS fundamentals: networking, IAM, and the core services a modern platform runs on.
* Experience with Datadog or a comparable observability stack, and a point of view on what good looks like.
* Comfortable executing autonomously
* A pragmatic bias toward simple, durable systems, and toward shipping over gold-plating
* Bonus: experience in a regulated or healthcare (HIPAA / PHI) environment.

**Why Tennr?**

* **Drive Impact:** one of our company values is Cowboy, meaning you set the pace. You won’t just talk about things, you’ll get them done. And feel the impact.
* **Develop Operational Expertise:** learn the inner workings of scaling systems, tools, and infrastructure.
* **Innovate with Purpose:** we’re not just doing this for fun (although we do have a lot of fun). At Tennr, you’ll join a high-caliber team maniacally focused on reducing patient delays across the U.S. healthcare system.
* **Build Relationships:** collaborate and connect with like-minded, driven individuals in our Hudson Square office 4 days/week.
* **Free lunch!** Plus a pantry full of snacks.

**Benefits**

* Beautiful new office at 345 Hudson Street
* Unlimited PTO
* 100% paid employee health benefit options
* Employer-funded 401(k) match
* Competitive parental leave
