🎁 Before you apply, rehearse this interview. Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card. I want my training →
About Uffizio Uffizio Technology is a SaaS product company based in Valsad, India, and part of the Uffizio Group. For over 25 years, we've been building software platforms trusted by B2B customers in more than 100 countries. We develop innovative SaaS products based on telematics and IoT solutions, helping businesses improve productivity, operational efficiency, and sustainability. Our products are offered as both white-label solutions for technology partners and ready-to-deploy platforms for businesses worldwide. We foster a collaborative, innovation-driven culture where talented people build impactful technology and grow their careers. About The Role We are looking for a DevOps Team Lead who can build and scale a highly reliable infrastructure for globally deployed SaaS platforms. This is not a maintenance-only role. The expectation is to architect systems that improve uptime, deployment speed, scalability, observability, security, and operational efficiency across multiple products and business units. The role requires someone who can think beyond servers and CI/CD pipelines. We need a systems thinker who understands how infrastructure impacts customer experience, release velocity, support load, product scalability, and business growth. You'll own the automation, deployment, and reliability of our systems end to end — from CI/CD pipelines to production monitoring — and work closely with development and product teams to ship faster without compromising stability. The candidate will lead the DevOps function across multiple SaaS products handling: Real-time telematics workloads IoT device communication High-ingestion APIs Live tracking systems Video and sensor-based platforms Multi-tenant SaaS deployments Key Responsibilities Infrastructure & Cloud Management Design, provision, and manage cloud and on-premise infrastructure (AWS) using Infrastructure as Code (Terraform, CloudFormation, or Ansible). Ensure high availability, scalability, redundancy, and disaster recovery planning. Manage Linux-based production environments. Optimize infrastructure cost without compromising reliability — through right-sizing, reserved capacity planning, and usage monitoring. Handle scaling strategies for increasing device load and customer growth. CI/CD & Release Engineering Build and maintain robust CI/CD pipelines (Jenkins, GitLab CI, GitHub Actions, or similar) to automate build, test, deployment, rollback, and environment provisioning processes. Reduce deployment risks and deployment time. Standardize deployment practices across teams and products. Monitoring & Reliability Establish strong monitoring, logging, and alerting/observability systems (CloudWatch, Prometheus, Grafana, ELK/EFK stack, Zabbix, Datadog). Implement proactive incident detection and root cause analysis. Reduce downtime and improve platform stability. Drive SRE-oriented operational maturity. Participate in on-call rotation, lead incident response, and drive root-cause analysis and post-mortems. Security & Compliance Implement infrastructure security best practices — IAM policies, network security groups, secrets management, SSL, firewall policies, backups, and vulnerability handling. Manage access control and compliance requirements. Ensure infrastructure hardening and operational compliance. Containerization & Orchestration Deploy, manage, and scale containerized workloads using Docker and Kubernetes (EKS/AKS/GKE or self-managed clusters). Improve deployment consistency and environment portability. Support microservices architecture where applicable. Database & Performance Optimization Work Closely With Backend And Database Teams On Performance tuning Query optimization support Load balancing Caching strategies Replication and failover systems Team Leadership Lead and mentor DevOps engineers. Create operational SOPs and infrastructure standards. Build accountability, documentation culture, and ownership within the team. Coordinate with Development, QA, Support, and Product teams to improve deployment velocity, system reliability, and developer experience. Incident Management & Business Continuity Handle production incidents with urgency and ownership. Build escalation systems and incident response frameworks. Conduct postmortem analysis and preventive planning. Maintain disaster recovery, backup, and business continuity plans for production systems. Required Technical Skills Strong Expertise In Linux Server Administration (patching, configuration management, access control) AWS / GCP / Azure Infrastructure as Code (Terraform, CloudFormation, or Ansible) Docker & Kubernetes (EKS/AKS/GKE or self-managed clusters) Jenkins / GitHub Actions / GitLab CI Nginx / Apache, Load Balancers & Reverse Proxies Networking & Security (DNS, VPN, firewalls) Monitoring Tools (Prometheus, Grafana, ELK/EFK, CloudWatch, Zabbix, Datadog) Infrastructure Automation Shell Scripting / Python / Bash Good Understanding Of High-availability architecture Distributed systems Scaling real-time applications Database replication and clustering Message brokers (RabbitMQ, Kafka, Redis Streams, etc.) API infrastructure SSL, DNS, VPN, CDN, WAF Cloud security best practices (IAM, encryption, secrets management) Nice to Have Experience in IoT or telematics platforms Experience managing large-scale real-time tracking systems SRE practices Cost optimization at scale Multi-region deployment experience LEADERSHIP EXPECTATIONS This role is not for someone who only executes tickets. We expect the person to: Think proactively instead of reactively Build systems before problems become incidents Create operational leverage through automation Reduce dependency on manual intervention Build infrastructure that supports aggressive business growth Create visibility and measurable operational KPIs KPIS / SUCCESS METRICS The DevOps TL Will Be Evaluated On Platform uptime Deployment frequency & stability MTTR (Mean Time to Recovery) Infrastructure scalability Security incident reduction Alert quality and monitoring maturity Automation coverage Infrastructure cost efficiency Team efficiency and operational discipline Experience Required 8+ years in DevOps / Infrastructure Engineering / Site Reliability Engineering / Cloud Infrastructure roles. 2+ years leading teams or handling critical production infrastructure. Experience managing production SaaS environments at scale. Strong working knowledge of at least one major cloud platform (AWS), with proven hands-on experience in Terraform, Docker, and Kubernetes. IDEAL CANDIDATE PROFILE We are not looking for a server administrator. We are looking for someone who: Understands business impact of infrastructure decisions Can scale systems under uncertainty Handles pressure calmly during outages Builds processes, not heroics Has a strong ownership mindset Can challenge poor engineering practices Thinks in terms of reliability engineering, not firefighting WHY THIS ROLE MATTERS For most SaaS companies, DevOps becomes a support function. For us, it is a growth constraint or growth accelerator. A Weak DevOps Team Creates Slow releases Customer dissatisfaction Downtime Engineering bottlenecks Support overload Revenue risk A strong DevOps function compounds the effectiveness of every other department. That is why this role is strategically important
India alone has 1,433 of them. Here's how to pick your next move.
4.658Jobs
1.433IN
70%EN
The devops job market is concentrated. India has 1,433 open roles, Poland has 560, and the US has 432. If you're based elsewhere or willing to relocate, knowing where the volume is helps you focus your search instead of applying blind.
Three-quarters of the ads are written in English—3,280 out of 4,658. This tells you about the ad itself, not whether the role requires you to speak English on the job. Check each posting carefully for language expectations.
A handful of employers are hiring hard right now. FullStack has 69 open devops positions, Link Group has 57, Upvanta has 36, and FlexBoard has 28. If you're building a target list, these four are worth researching deeply—they're actively scaling.
When you interview, expect questions about your incident response process. Hiring managers want to know how you've handled a real outage: what monitoring alerted you, what you did first, and what you'd do differently next time. Prepare a story with those three pieces.