UiPath

Senior Software Engineer- Site Reliability

UiPath3 days ago
Location

Bangalore

Type

Full Time

Salary

USD 165,000 – 240,000

Level

Senior

Role

Site Reliability Engineer

Posted

Jul 22, 2026

Full TimeSenior

The role

Summary

Join UiPath's Site Reliability Engineering team to design and build mission-critical SRE platforms that empower the entire engineering organization. This Senior Software Engineer role focuses on architecting large-scale distributed systems, driving livesite incident response, and shipping AI-powered reliability solutions that other teams depend on in their critical path. You'll combine deep distributed systems expertise with product thinking to eliminate reliability gaps and accelerate customer success.

What you'll do

Design and Build SRE Platforms: Architect, engineer, and ship scalable SRE platform systems with AI integration that other engineering teams depend on in their critical path. Treat these systems as products by conducting early user research, gathering feedback from adopting teams, and iterating based on real-world usage patterns and outcomes rather than theoretical requirements.
Livesite Incident Management: Participate in on-call monitoring rotations, handle production escalations, and drive effective incident mitigations to minimize customer impact. Conduct thorough postmortems that identify root causes and systemic improvements, not just surface-level fixes, to prevent recurrence across the platform.
Drive Reliability Improvements: Lead availability, scalability, and performance enhancements informed by livesite learnings. Identify architectural bottlenecks and design systematic solutions that embed best practices directly into platform infrastructure, eliminating the need for documentation-driven compliance across teams.
Ensure Technical Excellence: Maintain and exceed standards for reliability, scalability, quality, and performance in all deliverables. Identify and champion architectural changes that significantly improve these dimensions at scale, holding yourself accountable to measurable outcomes across distributed systems.
Drive Platform Adoption: Onboard other engineering teams onto your platforms by actively removing integration friction—writing integration code yourself, pair programming with teams, and addressing usability challenges. Measure success by adoption metrics and business outcomes, not handoff completion or documentation quality.
Rapid Iteration and Feedback: Ship platform capabilities early and frequently, aggressively seeking feedback from internal users. Treat every user complaint or friction point as a design input, prioritize iteration cycles, and continuously refine the developer experience based on real usage.
Technical Planning and Execution: Drive task planning, effort estimation, project scheduling, and resource allocation for SRE platform initiatives. Balance technical depth with delivery velocity while maintaining architectural integrity across large-scale distributed systems.
Engineering Best Practices Leadership: Participate in and influence process improvements, reliability best practices, and engineering standards across the organization. Contribute to organizational learning through incident analysis, technical discussions, and architectural reviews with peer engineers.

What we look for

Technical

Distributed Systems ArchitectureProven expertise designing large-scale, complex distributed commercial applications and services at enterprise scale. Deep understanding of system design principles including scalability, fault tolerance, load balancing, and eventual consistency patterns in production environments.
Internal Platform DevelopmentTrack record building large-scale, complex internal platforms that multiple teams depend on in critical paths—systems that have proven durable over time rather than prototypes. Experience with platform adoption strategies, developer experience optimization, and infrastructure-as-a-service frameworks.
Object-Oriented ProgrammingProficiency in one or more object-oriented languages including C#, C++, Go, or Python. Strong computer science fundamentals including data structures, algorithms, memory management, and performance optimization.
Cloud Architecture and DevOpsHands-on experience with cloud providers (Azure, AWS, or GCP) and managed services (AKS, GKE, ECS). Strong understanding of CI/CD pipelines, infrastructure-as-code, containerization, and modern DevOps practices. Production Kubernetes experience is highly valuable.
Microservices and Service-Oriented ArchitectureDeep experience building service-oriented and microservice-based architectures. Proficiency in HTTP applications, REST/gRPC web services development, API design, and managing inter-service dependencies and communication patterns.
Concurrency and Asynchronous ProgrammingStrong understanding of multithreading, synchronization primitives, race conditions, deadlock prevention, and asynchronous patterns. Experience implementing and debugging complex concurrent systems in production environments.
Database Design and Data EngineeringHands-on experience with modern database backends including relational databases (Azure SQL, MySQL), NoSQL systems (MongoDB, CosmosDB), data warehousing (Azure Data Lake, DynamoDB), and analytical platforms (Power BI). Understanding of data modeling, query optimization, and scaling strategies.
AI-Powered SystemsDemonstrated experience building, deploying, and maintaining AI-powered applications in production. Understanding of model integration, inference optimization, monitoring, and reliability considerations for AI systems at scale.

Education

Computer Science or Related FieldBachelor's degree in Computer Science, Engineering, or related discipline. Strong foundational knowledge of algorithms, data structures, systems design, and computer architecture is essential.
Continuous LearningDemonstrated commitment to staying current with emerging cloud technologies, reliability engineering practices, AI/ML advancements, and distributed systems research. Evidence of self-directed learning through certifications, open-source contributions, or technical publications.

Experience

Senior-Level Software EngineeringMinimum 6+ years architecting and delivering world-class, production-grade distributed applications and services. Track record of ensuring customer success through reliable, scalable system design and execution.
Platform Engineering at ScaleSignificant experience building internal platforms adopted by multiple engineering teams. Demonstrated ability to drive adoption through hands-on integration work, removing friction for users, and measuring success by outcomes rather than features shipped.
Production Systems LeadershipExperience identifying and closing reliability gaps across complex systems. Proven ability to move from gap identification through design, implementation, and organization-wide adoption of solutions.
Globally Distributed CollaborationExperience working effectively with geographically distributed engineering teams, managing time zones, asynchronous communication, and coordinating across organizational boundaries.

Skills

Required skills

C#, C++, Go, or PythonExpert-level proficiency in at least one object-oriented language with strong performance optimization and production debugging capabilities.
Distributed Systems DesignDeep expertise in designing scalable, reliable, fault-tolerant distributed systems including load balancing, replication, consensus mechanisms, and failure handling.
Cloud Infrastructure (Azure, AWS, or GCP)Hands-on proficiency with at least one major cloud provider and their managed services. Strong understanding of cloud-native architecture patterns and resource management.
Kubernetes and Container OrchestrationProduction experience operating and managing Kubernetes clusters, including deployment strategies, resource management, networking, and troubleshooting at scale.
Microservices ArchitectureProven ability designing and implementing microservice-based systems including service discovery, communication patterns, API design, and managing distributed tracing.
Database SystemsHands-on experience with relational databases (SQL), NoSQL systems (MongoDB, CosmosDB), and data warehousing solutions. Understanding of schema design, indexing, query optimization, and scaling strategies.
CI/CD and DevOpsStrong experience with modern CI/CD pipelines, infrastructure-as-code, automated testing frameworks, deployment automation, and observability infrastructure.
System Reliability EngineeringDeep understanding of SRE principles including SLOs/SLIs, error budgets, incident management, postmortem culture, observability, and reliability-focused architecture design.

Nice to have

Production Kubernetes OperationsExperience managing production Kubernetes infrastructure at scale, including cluster upgrades, multi-cluster deployments, and managing complex storage or networking requirements.
Machine Learning SystemsExperience building, deploying, and maintaining AI/ML-powered applications in production including model serving, monitoring, and inference optimization.
Observability and MonitoringHands-on experience designing monitoring strategies, building dashboards, implementing distributed tracing, log aggregation, and anomaly detection systems.
Cost OptimizationExperience with cloud cost governance, resource optimization, capacity planning, and implementing cost awareness across engineering organizations.
Access Control and SecurityFamiliarity with modern access control systems, identity management, authentication/authorization patterns, and security compliance in cloud environments.
Incident Response LeadershipExperience leading complex incident investigations, conducting effective postmortems, and driving systemic improvements from livesite learnings.
Open Source ContributionsActive participation in open-source projects, particularly in infrastructure, Kubernetes, observability, or reliability engineering domains.
Technical Architecture ReviewExperience participating in or leading architecture review boards, designing RFCs, and influencing technical direction across engineering organizations.

Compensation & benefits

Salary

USD 165,000 – 240,000 (annual)

Stock options

Available


Apply for this position

You'll be redirected to the company's application page