Senior Software Engineer, Reliability (f/m/d)

Product, Technology & Design
Full Time
Munich, Berlin, Dublin

Personio's intelligent HR platform helps small and medium-sized organizations unlock the power of people by making complicated, time-consuming tasks simple and efficient. Our team of 1,500 Personios is building user-friendly products that delight our 15,000+ customers and their 1.5 million employees. Ready to make an impact from day one?

The Role

This role is available in Munich, Berlin, Dublin, and Remote (Germany or Ireland).

Join us to shape the future of software in the underserved and high-impact HR technology industry. Your work will have a direct and tangible impact on customers, offering ownership and the chance to make a meaningful difference. As we prepare for significant growth, you'll face exciting challenges and have the opportunity to influence our path toward becoming one of the world's leading tech companies.

Personio is seeking an experienced Engineer to design, build, operate, monitor and scale our infrastructure through automated solutions. You’ll empower engineering teams by sharing cloud platform expertise, developing tools and establishing company wide mechanisms to ensure reliability, scalability and uptime. Our ideal candidate combines strong technical expertise with a collaborative mindset, working closely with other engineering teams to build, scale and enhance their applications on our platform.

What You’ll Do

  • Design and build the platforms and automated tooling, including AI-assisted agents, that enable engineering teams to scale their infrastructure on demand and embed reliability practices directly into how they work

  • Engage in and improve the full service lifecycle from initial design and inception through to deployment, operation and continuous improvement

  • Prepare services for production by engaging in system design reviews, developing shared frameworks and platforms, planning capacity and conducting launch assessments

  • Operate, monitor and maintain live services by tracking key metrics such as availability, latency and overall system health

  • Design and implement observability stacks and dashboards (e.g. Datadog, CloudWatch) to improve operational insight across logs, metrics, traces and alerts

  • Collaborate with product and engineering teams to define SLOs and error budgets to ensure services are reliable, scalable and observable

  • Identify and reduce toil through process automation, creating playbooks and building automated runbooks to reduce MTTR

  • Partner on security posture and compliance metric reviews related to operational reliability

  • Mentor, train and grow the community. Guide engineers across teams in reliability best practices and tooling

What You Need To Succeed

  • Bachelor’s degree in Computer Science, a related field, or equivalent practical experience

  • 5+ years of experience in SaaS software development, building and shipping production systems in a distributed-systems context, with strong proficiency in at least one of: Kotlin, Java, TypeScript or Python

  • 2+ years working in a backend/platform engineering role in agile environments

  • 2+ years of experience in designing, analyzing, and troubleshooting distributed systems

  • Good knowledge of modern application/infrastructure monitoring concepts (Datadog experience advantageous)

  • Comfortable reading, debugging and contributing to backend service code, not only infrastructure configuration

  • Systematic problem solving/debugging skills, coupled with a strong sense of ownership

  • Excellent written and verbal communication and documentation skills

  • Ability to work independently in a result-oriented way

Nice to Have/Bonus:

  • Expertise in designing, analyzing, and troubleshooting large-scale distributed systems.

  • Ability to debug, optimize code and automate routine tasks

  • Experience reducing on-call toil for engineers by automating repetitive operational work, tuning alerts to improve signal-to-noise or building self-healing mechanisms

  • Experience with Kafka/Debezium Connectors

  • Experience with PostgreSQL databases

Why Personio

Personio is an equal opportunities employer, committed to building an integrative culture where everyone feels welcomed and supported. We embrace uniqueness and understand that our diverse, values-driven culture makes us stronger. We are proud to have an inclusive workplace environment that will foster your development no matter your gender, civil status, family status, sexual orientation, religion, age, disability, education level, or race.

At Personio, we value in-person collaboration while also offering flexibility. This role is office-based, with 2 required in your contracted office location. The remaining days can be worked from home or in the office if you prefer. In addition, you’ll have 20 Flex Days per year to work remotely from other locations.

Aside from our people, culture, and mission, check out some of the other benefits that make Personio a great place to work:

  • Receive a competitive reward package – reevaluated each year – that includes salary, benefits, and pre-IPO equity.

  • Enjoy 28 days of paid vacation, plus an additional day after 2 and 4 years.

  • Make an impact on the environment and society with 1 (fully paid) Impact Day.

  • Receive generous family leave, child support, mental health support, and sabbatical opportunities.

  • We enjoy gathering for meals, cultural initiatives, and events like local Summer Sessions and year-end celebrations. There's also healthy snacks, drinks, and a weekly catered lunch.

Apply now