CronJobs

backend jobs

Senior Software Engineer, Backend - Reliability

Camunda · Remote

remotesenior$143,800–$231,900Posted Jul 24, 2026JavaGoKubernetesHelmPrometheusGrafanaChaos EngineeringLoad Testing

Apply on the employer site

About this role

## Senior Software Engineer, Backend — Reliability ### About Camunda Camunda is the enterprise platform for agentic orchestration—helping organizations coordinate AI agents, people, and systems across complex, end-to-end business processes. With governance, auditability, and human oversight, Camunda enables teams to move from pilots to production safely and at scale. This role is fully remote and global, as Camunda transforms into an AI-first organization built on its own platform. --- ### About the role As a **Senior Software Engineer (Backend), Reliability**, you’ll take ownership of **automated reliability testing** and **chaos engineering** at the core of **Camunda 8**. You’ll work in a culture that values curiosity, experimentation, and relentless improvement—**breaking things (on purpose!) in safe environments** to strengthen the platform before customers are impacted. You’ll also collaborate with talented colleagues who live our **FAITH** values: **Focus, Ambition, Integrity, Talent, Humor**. --- ### What you’ll be doing - **Design, implement, and execute** automated chaos experiments and reliability tests to validate Camunda under real-world scenarios. - **Investigate, root-cause, and debug** potential failures or performance regressions in **Java-based products**. - **Continuously improve** load testing, chaos engineering, and observability infrastructure to keep systems reliable and performant. - Work with multiple technologies, including **Java, Go, Kubernetes, Prometheus, Grafana**, and more. - Introduce new tooling and approaches for reliability testing, driving measurable improvements to operational experience and user outcomes. - Collaborate closely with **QA and engineering** (including cross-functional teams), sharing learnings and influencing roadmaps based on experiment findings. - Champion pragmatic, autonomous software design and learn new principles and technologies (e.g., Kubernetes, Helm, Grafana, Java, Go). - Advocate for **user-centric reliability** by using and understanding Camunda products firsthand. **Example resources (optional):** - CamundaCon talk: https://www.camundacon.com/event-session/camundacon-amsterdam-2025/one-exporter-to-rule-them-all-exploring-camunda-exporter/?on_demand=true - Chaos experiments blog: https://camunda.com/blog/2023/08/automate-chaos-experiments/ - zbchaos (fault injection): https://camunda.com/blog/2022/09/zbchaos-a-new-fault-injection-tool-for-zeebe/ - Team blog: https://camunda.github.io/zeebe-chaos/ --- ### Defining success (examples) - **After 3 months:** Design and deliver a way to run **quick, reproducible load tests** to reduce engineer feedback loops and fit into the development lifecycle. - **After 3 months:** Design and implement a realistic, automated, holistic load testing framework for **Optimize**, in close collaboration with Senior Engineers and Field teams—performing deep-dive analysis, sharing results, and documenting findings in public-facing documentation. --- ### What you bring - **5+ years** of backend software engineering experience (**Java**). - Proven drive to **experiment**, learn new tech, and conduct automated reliability and chaos testing in distributed environments. - Strong enthusiasm for improving **performance** and **fault-tolerance** in production systems. - Autonomous, pragmatic problem-solving with ability to guide and influence others. - Ability/willingness to use Camunda products (accounts available here: https://accounts.cloud.

Listing freshness

CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord