🎁 Before you apply, rehearse this interview. Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card. I want my training →
Job Description Experience: 7+ Years (Camunda administration & platform operations) Certification: Mandatory Valid Camunda Certification (attach certificate ID/details with application) Work Mode: Offshore / Remote (work independently with onshore stakeholders) Role Overview We are looking for a highly experienced Camunda Administrator to own the stability, performance, security, and upgrades of our Camunda BPM platform in a production workplace. This role is ideal for someone who is hands-on, technically strong, and comfortable driving platform execution end-to-endfrom installation and monitoring to incident troubleshooting and version upgradeswhile coordinating clearly with onshore teams and client stakeholders. You will be the go-to person for Camunda platform reliability, supporting development/support teams, handling upgrades (including Camunda 8.8+), and ensuring the platform runs smoothly across Kubernetes/containerized environments with solid observability and security practices. Key Responsibilities Platform Installation, Configuration & Administration Install, configure, and administer the Camunda BPM platform in dev/test/prod environments Manage setting setup, configuration, secrets, certificates (TLS), access policies, and runtime parameters Ensure high availability and correct deployment practices in Kubernetes/Docker environments Monitoring, Performance & Reliability Proactively monitor platform health, uptime, latency, and throughput Troubleshoot issues across platform components (runtime, workflow engine, indexing/search, integrations) Perform capacity planning, scaling, and performance tuning (CPU/memory, cluster sizing, storage planning) Drive root-cause analysis (RCA) and implement preventive fixes Operational Support & Incident Management Provide L2/L3 support for production issues, including on-call support (as required by project) Handle incident triage, logs analysis, and recovery actions Create and maintain runbooks, SOPs, and .