We are currently seeking a Site Reliability Engineer- REMOTE - Onsite Training to join our team in Memphis, Tennessee (US-TN), United States (US).
Day-to Day Responsibilities: - Implement and support CI/CD tools and pipelines across the organization (GitLab preferred).
- Experience with software and platform design, implementation, CI/CD pipelines, EKS, Bamboo, Jenkins, Docker, Splunk, and Datadog or similar monitoring tools.
- Monitor production/non-production systems and help solve problems, using tools including Datadog.
- Maintain containerized applications with Docker or similar technologies.
- Work with the product and development teams to continue to scale our application and infrastructure while ensuring performance and high availability.
- Collaborate with the development team on defining Service Level Indicators (SLIs) that represent the health of their service.
- Develop systems and software that increase site reliability and performance and be part of building SRE competency within the Digital organization.
- Help evolve the CI/CD pipeline leveraging SLIs and other monitorability works. Help with automating testing strategies to determine quality gates for production deployment.
- Build long term automation solutions using scripting and programming (Groovy, Shell, Python, Terraform, or Java/JavaScript).
- Provide technical leadership and implementation guidance for UI, APIs, and microservices.
- Participate in a scheduled on-call rotation supporting production systems.
Basic Qualifications: - 8+ years of experience in DevOps/SRE environments, including substantial experience with CI/CD platforms, containerization technologies, and Kubernetes-based deployments.
- 5+ years' experience in troubleshooting using APM (Application Performance Management) tools like Datadog and Dynatrace.
- 4+ years' experience with Unix/Linux shell scripting and supporting NoSQL databases such as Couchbase.
- 3+ years' experience with static code analysis tools such as Checkmarx and SonarQube.
Preferred Skills: - Strong log analysis using tools like Splunk.
- Experience in working with Nexus Repository.
- Ability to diagnose, troubleshoot complex issues.
- Demonstrated advanced ability in one or more of the required technical areas.
- Demonstrated commitment to continuous improvement and operational excellence.
- Experience with ServiceNow and Jira for change management.
- OWASP knowledge
Degree: Bachelor's degree in Computer Science or equivalent work experience.Nice to Have; (But not a must)NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this role is $87,120 - 151,250. Actual compensation will depend on a number of factors, including the candidate's relevant experience, technical skills, and other qualifications.
This position may also be eligible for incentive compensation based on individual and/or company performance.
This position is eligible for company benefits including medical, dental, and vision insurance with an employer contribution, flexible spending or health savings account, life and AD&D insurance, short and long term disability coverage, paid time off, employee assistance, participation in a 401k program with company match, and additional voluntary or legally-required benefits.
#LI-NorthAmerica