← WorkMundi · 1M+ jobs from around the world, liveSign inCreate free account

Resilience And Reliability Architect Manager Noida (India)

EY · Noida

📅 09/08/2026
🔔 Alert me about jobs like this
No password, no sign-up. Just the email — and you can leave the list anytime.
🔓 Apply — free →
Opens this job on WorkMundi. The account is free and takes under a minute.

See the other 113,931 jobs in India →

🎁 Before you apply, rehearse this interview. Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card. I want my training →
Site Reliability Engineering (SRE) Architect / Consultant (M)Description Site Reliability Engineering (SRE) is a contemporary way of delivering IT Operations by imbibing Software engineering principles in Service Delivery to reduce IT Risk to business, improve business resilience, attain predictability & reliability, optimize cost of IT Infra and Ops An SRE Architect / Consultant will help in designing the roadmap to SRE for Enterprise IT They will also implement various SRE Solutions across the enterprise / line-of-businesses They will be able to assess SRE Maturity of an IT Organization and provide strategy and roadmap to achieve higher maturity levels Responsibilities Defining SLA/SLO/SLI for a product / service Engineering in resilient design and implementation practices into solutions as they go through the product life cycle Designing & implementing Observability Solutions to track, report, and measure SLA adherence Engineering out manual effort (Toil) through the development of automated processes and services (e.g., Automated Management of Systems, CI/CD improvements) Optimize Cost of IT Infra & Operations - FinOps Typical Skills and Background 12+ years of experience in software product engineering principles, processes and systems Hands-on experience in Java / J2EE, one of web server (Apache Tomcat or IBM HTTP Server), one of the application servers (Tomcat/WebSphere), and any major RDBMS like Oracle Hands-on experience in at least one CI-CD (Azure DevOps, GitLab CI/CD, Jenkins) and IaC tools (Terraform, AWS CloudFormation, Ansible etc.) Experience in at least one cloud technology (AWS/Azure/GCP etc. and Docker, Pivotal, Kubernetes, OpenShift etc.) and its reliability tools (Azure AppInsight, CloudWatch, Azure Monitor etc.) Experience in Observability - APM tools (Dynatrace, AppDynamics etc.), metrics / log consolidation (Splunk) and ELK Stack Experience in Linux (RHEL) operating system performance monitoring parameters and their interpretation, commands used for monitoring Knowledge on queuing models used, thread pools, request servicing processes etc. Knowledge of application design patterns, J2EE application architectures, Microservices, Spring boot & Cloud native architectures Knowledge at least one automation scripting language like Python .
Read the rest of the job →

Similar jobs

Job on WorkMundi — the world's largest job board. See more jobs from every continent, updated live.

📢
🎁

Before you apply, rehearse this interview.

Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card.

I want my training →