Personio logo

Senior Site Reliability Engineer

Personio

RemoteFull timeMid levelPosted today
Apply with JobAssist

About the role

Who you are

  • Bachelor’s degree in Computer Science, a related field, or equivalent practical experience
  • Proven software engineering experience in Java or Kotlin, including algorithm development and coding
  • 6+ years of experience with SaaS software development in distributed systems using languages such as Kotlin/Java, Typescript, Python, and technologies like IaC, Docker, and Kubernetes
  • 1+ years’ experience as an SRE or similar role designing, operating, analyzing and troubleshooting distributed systems in agile environments
  • Act as a Datadog subject matter expert, assisting with observability stack design, dashboard creation, and training peers in best practices
  • Hands-on experience running Kafka at scale including configuration, operational failure modes and reliable recovery/runbooks
  • Systematic problem solving and debugging skills with a strong sense of ownership and bias towards establishing mechanisms which can scale across the entire company
  • Excellent written, verbal, and documentation skills
  • Collaborative team player, able to communicate effectively across disciplines
  • Experience with CI/CD tooling (GitHub Actions/GitOps tools)
  • Experience tuning JVM-based services and Node.js runtimes
  • Experience with AWS MSK Connect

What the job involves

  • Join us to shape the future of software in the underserved and high-impact HR technology industry
  • Your work will have a direct and tangible impact on customers, offering ownership and the chance to make a meaningful difference
  • As we prepare for significant growth, you'll face exciting challenges and have the opportunity to influence our path toward becoming one of the world's leading tech companies
  • Personio is seeking an experienced Engineer to design, build, operate, monitor and scale our infrastructure through automated solutions
  • You’ll empower engineering teams by sharing cloud platform expertise, developing tools and establishing company wide mechanisms to ensure reliability, scalability and uptime
  • Our ideal candidate combines strong technical expertise with a collaborative mindset, working closely with other engineering teams to build, scale and enhance their applications on our platform
  • Engage in and improve the full service lifecycle from initial design through deployment, operation, and continuous improvement
  • Prepare services for production by taking part in system design reviews, developing shared frameworks and platforms, planning capacity and conducting launch assessments
  • Operate, monitor, and maintain live services, designing observability stacks and dashboards to track key metrics and improve operational insight
  • Ensure sustainable scalability through automation, actively contributing to continuous improvement for reliability and delivery speed
  • Collaborate with product and engineering teams to define SLOs, error budgets and ensure services are reliable, scalable and observable
  • Support incident management processes, including on-call rotations, assisting with outage response, and contributing to post-mortems and root cause analysis
  • Identify and reduce toil through process automation, creating playbooks and automated runbooks to reduce MTTR
  • Support resilience strategies and help implement chaos testing to proactively uncover weaknesses and validate recovery strategies
  • Own and maintain the reliability of our event streaming and Change Data Capture (CDC) stack
  • Mentor and train peers on reliability best practices and tooling, contributing to community growth

Benefits

  • Receive a competitive reward package – reevaluated each year – that includes salary, benefits, and pre-IPO equity
  • Enjoy 28 days of paid vacation, plus an additional day after 2 and 4 years (because we love what we do, but we also love vacation!)
  • Make an impact on the environment and society with Impact Days
  • Receive generous family leave, child support, mental health support, and sabbatical opportunities with PersonioCares
  • Connect with your fellow Personios at regular company and team events like All Company Culture Week and local year-end celebrations
  • Engage in a high-impact working environment with flat hierarchies and short decision-making processes
  • Find your best way to work with our office-led, and remote-friendly PersonioFlex!

Millions of jobs, with real people getting hired every day

20,000+
New jobs added daily
7,000,000+
Verified job listings
500,000+
Tailored applications submitted
FAQ

Questions, answered

Click "Apply with JobAssist" – we tailor your resume and application to this role and submit it for your approval.

Yes. This role at Personio was screened before publishing – we confirmed the employer before listing it.

The employer didn't disclose a salary range for this listing. JobAssist shows pay whenever it's available.

This position can be done from anywhere, with no in-office requirement.

Yes – every application is tailored from your profile and this job's requirements, and you can review and edit before it's sent.