UiPath
Posted 2mo ago

Principal Software Engineer, Site Reliability

UiPath
Bucharest, Bucharest, Romania
OnsiteFull Time
Responsibilities
  • designing platforms
  • monitoring rotations
  • driving availability
Requirements
  • 10+ years in architecting and engineering large-scale distributed applications
  • Strong OO languages
  • Cloud and AI-powered systems
  • Kubernetes and multiple cloud providers
  • Cross-team collaboration
Technical tools mentioned
C#C++JavaPythonKubernetesAzureAWSGCPAKSGKECosmosDBAzure SQLMongoDBMySQLDynamoDBPower BI

Job description

Life at UiPath

The people at UiPath believe in the transformative power of automation to change how the world works. We’re committed to creating category-leading enterprise software that unleashes that power.

To make that happen, we need people who are curious, self-propelled, generous, and genuine. People who love being part of a fast-moving, fast-thinking growth company. And people who care—about each other, about UiPath, and about our larger purpose.

Could that be you?

Your mission

At UiPath’s Site Reliability team, we build the platforms and systems that the entire company depends on to deliver on our compliance and SLA promises to customers. This spans monitoring, alerting, cloud infrastructure, access management, standardized synthetics, performance and test validation, incident detection and status reporting, incident management, automated remediation, structured post-mortems, customer communications, repair item tracking, and assertion of engineering best practices across UiPath. We are scaling each of these pillars - and building the next generation of capabilities on each, increasingly powered by AI.


This is a software engineering role.


You will not be the person who identifies a reliability gap and files a ticket for another team to fix. You are the engineer who identifies the gap, designs the system that closes it, builds it, ships it, and drives its adoption - often by doing the integration work yourself rather than asking other teams to come to you. You build platforms that other engineers depend on in their critical path, and you hold yourself accountable to outcomes, not outputs.
You treat every system you ship as a product: you put it in front of users early, seek feedback, and iterate until it delivers real results. If adoption is slow, you don’t blame the docs - you sit with the team, understand the friction, and remove it.

What you'll do at UiPath

  • Design, engineer, and build SRE platform systems and capabilities with cutting-edge AI, treating them as products that other engineering teams depend on in their critical path.

  • Participate in livesite monitoring rotations, handle escalations, and drive effective mitigations - reducing customer impact through broad, detailed, and effective post-mortems.

  • Drive availability, scalability, and performance improvements based on livesite learnings. Generate (or codify existing) best practices and ensure they are followed widely across UiPath - not by publishing guidance, but by embedding them into the systems you build.

  • Ensure technical deliverables meet or exceed expectations on reliability, scalability, quality, and performance. Identify and drive architectural changes that significantly move the needle on these dimensions.

  • Onboard other teams onto your platforms by driving outcomes yourself - writing the integrations, pairing with their engineers, removing friction - rather than handing off documentation and waiting.

  • Ship early, seek feedback relentlessly, and iterate fast. Treat every user complaint as a design input, not a support ticket.

  • Drive task planning, estimation, scheduling, and staffing.

  • Mentor Software Engineers to develop their skills and knowledge through hands-on coaching, advice, and training opportunities.

  • Participate in and influence process improvements and best practices across the engineering organization.

What you'll bring to the team

  • Proven track record (10+ years) of architecting and engineering world-class, large-scale, distributed commercial applications and services, and ensuring customer success.

  • Experience building large-scale, complex internal platforms adopted by 10+ teams in their critical path at a large company — systems that have stood the test of time, not prototypes that were handed off or abandoned.

  • Demonstrated ability to drive adoption of your systems by doing the hard work yourself: writing integrations, removing friction for other teams, and measuring success by outcomes delivered - not features shipped.

  • Experience building and maintaining complex AI-powered applications in production.

  • Proficiency in one or more object-oriented languages (such as C#, C++, Java, or Python), backed by solid computer science fundamentals.

  • Deep understanding of data structures, algorithms, multithreading, synchronization, asynchronous patterns, and cloud programming.

  • Experience with service-oriented and microservice-based architectures, HTTP applications, and web services development.

  • Familiarity with modern engineering practices including agile development, CI/CD, and DevOps. Ability to work with globally distributed teams.

  • Experience working with or managing production Kubernetes infrastructure is a plus.

  • Experience with cloud providers (Azure, AWS, GCP) and managed services (AKS, GKE, etc.) is a plus.

  • Experience with database backends (e.g., Azure SQL, CosmosDB, Azure Data Lake, Power BI, MongoDB, MySQL, DynamoDB, etc.).

Maybe you don’t tick all the boxes above—but still think you’d be great for the job? Go ahead, apply anyway. Please. Because we know that experience comes in all shapes and sizes—and passion can’t be learned.

Many of our roles allow for flexibility in when and where work gets done. Depending on the needs of the business and the role, the number of hybrid, office-based, and remote workers will vary from team to team. Applications are assessed on a rolling basis and there is no fixed deadline for this requisition. The application window may change depending on the volume of applications received or may close immediately if a qualified candidate is selected.

We value a range of diverse backgrounds, experiences and ideas. We pride ourselves on our diversity and inclusive workplace that provides equal opportunities to all persons regardless of age, race, color, religion, sex, sexual orientation, gender identity, and expression, national origin, disability, neurodiversity, military and/or veteran status, or any other protected classes. Additionally, UiPath provides reasonable accommodations for candidates on request and respects applicants' privacy rights. To review these and other legal disclosures, visit our privacy policy.

About UiPath

Provides robotic process automation software for enterprise workflow automation.

Similar jobs

Site Reliability Engineer roles near Bucharest, Bucharest
18h
Save
Mark Applied
Hide
Senior Site Reliability Engineer - PAM
Bucharest, Bucharest, Romania
HybridFull Time
Expleo
Expleo: Global provider of engineering, technology, and consulting services.
4+ YOERequires 4–5 years of PAM experience, CyberArk or comparable platform expertise, DevOps, networking, Windows/Linux, monitoring, Terraform, Puppet, troubleshooting, and strong English communication skills.
CyberArk, Delinea Secret Server, BeyondTrust, LDAP, Active Directory, Red Hat Linux, Prometheus, Grafana, Datadog, PagerDuty, New Relic, JIRA, Confluence, Service Line, SOAP, REST, PowerShell, Toad, Terraform, Puppet, CyberArk Vault, PVWA, CPM, PSM, AWS, GCP
1d
Save
Mark Applied
Hide
Site Reliability Engineer
Bucharest, Bucharest, Romania
HybridFull Time
Worldline
WorldlineEuronext Paris: WLN: Global provider of payment and digital transaction services.
Requires SQL, UNIX/Linux, Oracle and PostgreSQL, monitoring, scripting, ITIL, XML/XSD, Agile, integration, cloud, CI/CD, security, and fluent English; experience with AI automation preferred.
SQL, UNIX, Linux, Oracle, PostgreSQL, Grafana, Dynatrace, JIRA, Confluence, GitHub, REST, SOAP, Java, CI/CD, XML, XSD
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (remote within EMEA)
Warsaw or Kyiv or Bucharest or Tallinn or Barcelona or Riga
RemoteFull Time
FYUL
FYUL: A platform powering global on-demand eCommerce merchandise production.
Several years of production infrastructure/SRE experience with AWS, Kubernetes, Terraform, Python, Linux, CI/CD, observability, incident response, and senior-level technical ownership and mentoring.
Linux, Python, AWS, Amazon EKS, IAM, VPC, RDS, S3, SQS, Terraform, Terragrunt, ArgoCD, Grafana, Prometheus, Loki, Tempo, Mimir, Postgres, MySQL, MongoDB, Aurora, Jenkins, GitHub Actions, Helm, Cilium, ECR, Kafka, AWS MSK, PHP, Symfony, Node.js, TypeScript, Angular, Redis, Atlantis, Postman, Git, GitHub Copilot, PhpStorm, Kibana, Jira, Miro, Google Workspace, Slack
3w
Save
Mark Applied
Hide
Site Reliability Developer 3
Bucharest or Ia\u000219i
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOERomania resident with BS/MS in CS or related field,3+ years SRE/cloud ops experience,proficient in Python/Go/Java/Bash,Linux,networking,container tech,monitoring,incident response,automation,and mentoring.
Python, Go, Java, Bash, Docker, Kubernetes, Codex, GitHub Copilot, Cursor, Claude Code, Large Language Models (LLMs)
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Bucharest, Bucharest, Romania
OnsiteFull Time
London Stock Exchange Group
London Stock Exchange GroupLondon Stock Exchange: LSEG: Provides financial market infrastructure and global data analytics services.
Extensive SRE/platform reliability experience, proven SLO/SLI design, incident response leadership, mentoring skills, and ability to influence cross-team engineering practices.
OpenTelemetry, Grafana, ClickHouse, Cribl, Datadog, BigPanda, Redis, PromQL, ClickHouse SQL
1mo
Save
Mark Applied
Hide
Enabling SRE — Bucharest Sovereign Cloud Hub
Bucharest, Bucharest, Romania
HybridFull Time
Thales
ThalesEuronext Paris: HO: Develops electronics and digital systems for aerospace and defense.
3+ YOE3+ years SRE/Platform/DevOps experience with strong Kubernetes, Linux, scripting (Python/Go/Bash), observability or automation knowledge, English proficiency, and willingness to operate in a follow-the-sun on-call model.
Kubernetes, NixOS, Grafana, Loki, Tempo, Mimir, ELK, ArgoCD, FluxCD, Terraform, CI/CD, Python, Go, Bash, GKE, Vertex AI, Prometheus, Borg, Colossus, Spanner, Google Cloud, GitOps
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (12 months contract)
Bucharest, /, Romania
HybridContract
Electronic Arts
Electronic ArtsNASDAQ: EA: Develops and publishes video games and interactive entertainment software.
5+ YOE5+ years building SRE practices; experience with cloud (AWS, Azure), monitoring/observability, IaC, automation, on-call rotations, mentoring, and incident response.
Prometheus, Grafana, Datadog, ELK, Terraform, Ansible, AWS CloudFormation, GitLab CI/CD, Python, Bash, Kubernetes, EKS, AKS, GKE, AWS, Azure
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Bucharest, Bucharest, Romania
HybridFull Time
Resideo
ResideoNYSE: REZI: Manufacturing and distributing home comfort and security solutions.
6+ YOE6+ years SRE or cloud infrastructure experience; 3+ years with a major public cloud (Azure/AWS/GCP); 2+ years with Terraform or similar IaC; experience with containers (Docker/Kubernetes); strong automation and incident management skills.
Microsoft Azure, AWS, GCP, Terraform, ARM Templates, Ansible, Chef, Helm, Kubernetes, Git, Git Actions, Jenkins, Docker, Grafana, Prometheus, Elastic, PowerShell, Bash, Python, Windows, Linux