Druva

Cloud Operations Engineer

Druva • $108K — $151K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Minimum 2 years of relevant experience or advanced degree in a related field.
  • Proficient in AWS and Azure services, or other cloud providers.
  • Experience with containerization technologies such as ECS, Kubernetes, Docker.
  • Familiar with observability tools like Grafana and OpenSearch.
  • Knowledge of Infrastructure as Code using Terraform or CloudFormation.
  • Basic coding skills in at least one programming language, preferably Python or Go.
  • Knowledgeable in Linux operating systems, with experience in deployment and patching.

Responsibilities

  • Participate in weekly team meetings, collaborating with development, operations, and QA teams.
  • Monitor and manage cloud infrastructure, including overseeing production updates and OS patches during deployments.
  • Prioritize and troubleshoot server, network, and storage issues by addressing incident tickets.
  • Provide escalated support and ensure incident resolution meets SLA requirements.
  • Develop and maintain Infrastructure as Code and automation scripts for testing and deployments.
  • Configure and deploy monitoring and observability tools to enhance proactive oversight of the environment.
  • Communicate effectively about project updates and collaborate with cross-functional teams for seamless operations.

Benefits

  • Comprehensive health and wellness benefits program.
  • 401(k) retirement plan with company matching.
  • Life and disability insurance coverage.
  • Potential for discretionary annual bonuses and commissions.
  • Participation in equity opportunities provided to employees.
Full Job Description
About the Role

As a Cloud Operations Engineer at Druva, the industry leader in Cloud Data Protection and Management, you will play a key role in supporting, maintaining, and scaling cloud infrastructure within a fast-paced, complex, and FedRAMP-certified environment. Working under general guidance, your focus will be on developing professional expertise in cloud operations, addressing tasks of moderate scope, proactive monitoring, and assisting cross-functional teams to deliver reliable, highly available solutions for our growing customer community.

Responsibilities:
  • Participate in weekly team meetings, working closely with devops, support, and quality assurance
  • Monitor and manage all cloud infrastructure, including Friday evening deployments of production updates and OS patches
  • Support various AWS and Azure services being utilized and gain foundational expertise across primary infrastructure components

Incident Response and Troubleshooting:
  • Prioritize and address tickets related to server, network, and storage issues
  • Provide escalated support for problems and incidents
  • Ensure incident resolution is within the defined SLA (Service Level Agreement)

Automation and Monitoring:
  • Develop, maintain, and enhance Infrastructure as Code (e.g., Terraform, CloudFormation) and automation scripts for testing and deployment.
  • Configure and deploy supporting tools for proactive monitoring, observability, and security (e.g., Grafana, Prometheus/Thanos, Wiz, OpenSearch)

Communication and Collaboration:
  • Effectively communicate project updates and tasks to the Operations team and stakeholders in other departments
  • Collaborate with cross-functional teams to ensure seamless cloud operations

Expectations:
  • Support production deployments scheduled on Fridays between 5pm - 11pm PDT on a rotating basis
  • Take responsibility for identifying and resolving vulnerabilities in the FedRAMP environment
  • Develop and monitor dashboards to track real-time metrics

Experience & Skills:
  • Minimum of 2 years of related experience (or an advanced degree without experience, or equivalent work experience)
  • Proficiency in AWS and Azure services (or another Cloud Provider).
  • Experience with containerized applications (ECS, Kubernetes, Docker)
  • Experience with observability and monitoring (Grafana, OpenSearch)
  • Experience with an SIEM (Security Incident and Event Management) application
  • Familiarity with one programming language (Python, Go)
  • Familiar with Infrastructure as Code (Terraform, CloudFormation)
  • Knowledge of Linux operating systems deployments & patching (Ubuntu, RHEL)
  • U.S. citizenship is required to support Federal customer engagements. Existing or ability to obtain a U.S. government security clearance (Public Trust, Secret, or higher) is a strong plus.

Certifications (Bonus):
  • AWS, Azure, RHEL, CCNA
What We Offer

The pay range for this position is expected to be between $108,000 and $151,333/year; however, base pay offered may vary depending on multiple individualized, non-discriminatory factors, including market location, job-related knowledge, skills, and experience. The total compensation package for this position may also include other incentive compensation opportunities in the form of discretionary annual bonus or commissions, and equity. Additionally, full-time employees are eligible to participate in our comprehensive benefits program, including health and wellness benefits, 401(k) retirement plan, life and disability insurance coverages, and other benefits the Company may offer from time to time.

About Druva

Druva is a cloud data protection and management company that was founded in 2008. The company is headquartered in Sunnyvale, California and has offices in India, the United Kingdom, Germany, and Singapore. Druva's platform provides backup, disaster recovery, and archiving services for cloud applications and data. The company has received numerous awards for its technology, including recognition as a Leader in the 2020 Gartner Magic Quadrant for Enterprise Backup and Recovery Software Solutions.
Learn more about Druva
Size
1,000 employees
Industry
Founded
2008

Similar Jobs

More Jobs at Druva

More Information Technology Jobs

Find similar Cloud Operations Engineer jobs: