Tata Consultancy Services

DevOps & Site Reliability Engineer (Digital)

Tata Consultancy Services$110K — $150K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of experience in full stack Java development
  • Expertise in Observability platforms, particularly Dynatrace
  • Strong knowledge of Spring Framework components
  • Hands-on experience with various databases like Oracle and MySQL
  • Familiarity with Retail Point of Sale and Payment Systems
  • Proficient in Azure cloud services and architecture
  • Experience in CI/CD pipeline management using Azure DevOps

Responsibilities

  • Act as SRE Technical Architect for retail platforms, focusing on reliability and stability
  • Define SRE standards and best practices across teams
  • Architect scalable and highly available cloud environments
  • Lead the implementation of CI/CD and Infrastructure as Code (IaC) practices
  • Establish proactive monitoring and incident prevention systems
  • Oversee major incident management and root cause analysis
  • Collaborate with cross-functional teams to enhance reliability by design
  • Mentor DevOps and SRE engineers to improve skills and processes

Benefits

  • Opportunities for professional development and training
  • Mentorship for career growth
  • Engagement in high-impact projects
  • Flexible work culture to enhance work-life balance
  • Involvement in the latest cloud technologies and practices
Full Job Description
Must Have Technical/Functional Skills

  • Technology and Programming (Expert Level)
  • Strong proficiency in Java full stack developer
  • Object-Oriented programming principles and concepts
  • Hands-on experience on Observability platform Dynatrace
  • Hands-on experience with Spring Framework (Spring Boot, Spring MVC, Spring Security)
  • Knowledge if RESTful API development
  • Experience with database like Oracle, DB2, MySQL
  • Proficiency in Payment Switch BASE24 EPS, C++, AS400 and Python is also added advantage
  • Domain, Cloud & Platform Engineering
  • Must have domain experience on Retail Point of Sale/Payment Systems/Merchandising/Inventory/Logistics area
  • Expertise in Microsoft Azure, including:
  • Compute (VMs, App Services, Azure Container Apps)
  • Containers & Orchestration (AKS, Docker)
  • Storage, Azure Key Vault, Azure Monitor, Log Analytics
  • Proven experience designing enterprise grade, highly available cloud platforms


DevOps & Engineering Excellence

  • Advanced experience with Azure DevOps and CI/CD pipeline architecture
  • Strong scripting skills (PowerShell, Bash)
  • GitOps concepts, branching strategies, release orchestration


Site Reliability Engineering:

  • Ownership of platform reliability, resiliency, and performance


Definition and governance of:

  • SLIs, SLOs, SLAs
  • Error budgets and reliability metrics


Advanced observability strategy, designing and implementation:

  • Metrics, logs, traces, alerts, dashboards using Dynatrace
  • Incident response leadership, RCA facilitation, and long term remediation planning
  • Experience operating 99.9%99.99% availability systems


Security, Compliance & Cost

  • Secure cloud design using Key Vault, managed identities, RBAC
  • Cost optimization (FinOps mindset) across cloud infrastructure


Roles & Responsibilities

  • Act as SRE Technical Architect (Should be interested to work on Implementations) for client's Retail platforms, owning reliability and stability outcomes
  • Define and enforce SRE standards, best practices, and operating models
  • Architect and govern highly available, scalable cloud platforms
  • Lead the design and implementation of CI/CD and IaC strategies
  • Establish proactive monitoring, alerting, and incident prevention mechanisms
  • Own major incident leadership, RCA execution, and corrective action tracking
  • Partner with application, security, and architecture teams to build reliability by design
  • Drive automation to reduce toil and improve operational efficiency
  • Mentor and coach SRE and DevOps engineers across teams
  • Influence roadmap decisions with a reliability, scalability, and cost lens


#LI-KR3

Salary Range-$110,000-$150,000 a year

About Tata Consultancy Services

Tata Consultancy Services (TCS) is an Indian multinational information technology (IT) services and consulting company, headquartered in Mumbai, Maharashtra, India. It is a subsidiary of Tata Group and operates in 149 locations across 46 countries. TCS is the largest Indian company by market capitalization and is ranked 11th on the Forbes Global 2000 list of the world's biggest public companies. TCS is also the second-largest IT services company in the world by revenue and the largest employer of women in India. The company provides services in areas including IT, consulting, and business solutions.
Learn more about Tata Consultancy Services
Size
469,261 employees
Industry

Similar Jobs

More Jobs at Tata Consultancy Services

More Information Technology Jobs

Find similar DevOps & Site Reliability Engineer (Digital) jobs: