Cisco

Site Reliability Engineer

Cisco$165K — $241K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree plus 7 years of related experience, Master's degree plus 4 years, or PhD plus 1 year required.
  • Solid conceptual and practical knowledge in primary technical job family and related technical fields.
  • Experience with a range of technologies.
  • Ability to work independently with minimal guidance.
  • Strong communication skills to gather input and establish relationships across teams.

Responsibilities

  • Evaluate scalability, resiliency, performance, and security of production environments.
  • Support uptime of services through On-Call rotation and monitoring for SLAs and SLOs.
  • Conduct Disaster Recovery drills to ensure reliable incident processes.
  • Investigate incidents, implement remediation strategies, and learn from past incidents.
  • Automate repetitive tasks to reduce operational expenses and improve reliability.
  • Plan, design, and implement local and wide-area network solutions.
  • Integrate coding expertise to enhance service reliability across different systems.

Benefits

  • Participation in an On-Call rotation for operational support.
  • Opportunities for professional development and guidance for less experienced colleagues.
  • Involvement in cross-functional projects to enhance collaboration.
  • Access to advanced tools and technologies for monitoring and automating systems.
Full Job Description
The application window is expected to close on:
Job posting may be removed earlier if the position is filled or if a sufficient number of applications are received.

The successful applicant may be performing work in FedRAMP High or IL-5 environments, and therefore, must be a U.S. Person (i.e. U.S. citizen, U.S. national, lawful permanent resident, asylee, or refugee). This position may also perform work that the U.S. government has specified can only be performed by a U.S. citizen on U.S. soil.

From a reliability standpoint, this role involves evaluating the scalability, resiliency, performance, and security properties and techniques used in production environments. It supports the uptime of production services through an On-Call rotation, which includes monitoring and alerting to meet internal Service Level Objectives (SLOs) and customer-facing Service Level Agreements (SLAs). Ensuring reliable incident processes is achieved by conducting Disaster Recovery drills. The role also focuses on improving reliability through incident management by investigating incidents, implementing remediation strategies, and learning from past incidents to make improvements. It involves determining the reliability and security requirements of components and systems to meet the reliability objectives of the company, customers, and any relevant governmental agencies. Additionally, the role aims to reduce operational expenses through automation, by identifying and mitigating failure points, and automating repetitive and resource-intensive tasks. It also involves developing new acceleration techniques and analytical tools to ensure the early identification of potential issues with new products, packaging, processes, and overall product reliability.

Ensures a reliable and scalable network, manages network / cloud infrastructure and storage systems supporting business operations, and responds to planned maintenance, real-time outages, and issues.

Plans, designs and implements local and wide-area network solutions between multiple platforms and protocols.

What You'll Do:
• 7+ yrs experienced professional using best practices and knowledge of internal or external business issues to improve products or services.
• Works independently, but receives minimal guidance and direction from leader then determines best approach to accomplish work.
• Acts as a resource for colleagues with less experience.
• Understands project and/or department needs and establishes relationships with appropriate cross-functional stakeholders to gather input, collect information, and complete work steps.
• Designs and deploys small to mid-size or moderately complex solutions to optimize reliability, availability, latency, and performance.
• Integrates knowledge of design, automation, and deployment with expertise in coding to improve service reliability for existing or new systems and adapts for regions, countries, or customers.
• Designs and tests high availability and disaster recovery measures for our services to ensure automation is improving reliability, scalability, and velocity.
• Forecasts and builds reports to determine at what point resources will be at capacity.
• Designs and implements tools that provide visibility into performance and reliability of our infrastructure.
• Builds automated platforms.
• Monitors the environment and works with Developers and Ops to identify problems and develop monitoring tools that provide visibility into performance and reliability, serves as on-call SRE, leads post mortems, and writes root cause analysis.
• Builds and ensures security controls are in place in regards to architectural design, collaborates with security in designing or providing input to security controls, and may actively contribute in security incident response.

Minimum Qualifications:
• Bachelors + 7 years of related experience, or Masters + 4 years of related experience, or PhD + 1 year of related experience.
• Requires solid conceptual and practical knowledge in primary technical job family and knowledge of related technical job families; has worked with a range of technologies.

About Cisco

Cisco Careers

Join the vibrant team at Cisco, a global leader in networking and cybersecurity solutions, where innovation and leadership thrive. Cisco offers a plethora of job opportunities that cater to a range of skills and experiences, making it an ideal place for both seasoned professionals and those seeking an internship to jumpstart their career. Work You’ll Do At Cisco, you’ll be part of a culture that values diversity, leadership, and professional growth. Engage in work that matters with a team that combines technology, creativity, and the power of human connection to redefine networking. Cisco’s commitment to innovation isn’t just about technology, but also about transforming the way we work and collaborate. Cisco’s employment philosophy supports career advancement and nurtures a leadership pipeline that is equipped with diversity training and opportunities for growth. Whether you’re applying your skills to drive our latest innovations or using our vast networking capabilities to solve complex problems, at Cisco, every role is impactful. Join Our Dynamic Team Explore job opportunities in areas ranging from engineering to marketing, sales to cybersecurity. Cisco is hiring individuals who are passionate, curious, and ready to drive change. Positions at Cisco offer competitive benefits, a supportive culture, and the chance to work with cutting-edge technology. Internship Programs Kickstart your career with a Cisco internship. Gain invaluable industry experience, enhance your resume, and build professional networks that last a lifetime. Our internships provide hands-on experience and the chance to work on projects that matter. Leadership and Development Cisco is committed to fostering leadership skills and providing employees with the training needed to succeed. Our leadership programs help you develop new skills, manage teams effectively, and lead with confidence. Cisco’s commitment to professional development ensures that your career path is as dynamic as our technologies. Benefits and Culture Cisco understands the importance of a balanced life. Our benefits package is designed to ensure that our team members are healthy, happy, and secure. At Cisco, you’ll find a supportive culture that encourages open communication, teamwork, and mutual respect. Stay Connected Join Cisco’s Talent Network Stay informed about new positions that match your skills and interests. At Cisco, we value the curiosity and unique perspectives of our team members. Subscribe to receive personalized job alerts and insider tips directly from our hiring managers. Explore Cisco Jobs Ready to advance your career at Cisco? Search open positions, prepare your resume, and get ready for an interview that could lead to your next big opportunity. At Cisco, we’re not just filling positions—we’re investing in leaders. Keep Up to Date Stay ahead with career tips, insider perspectives, and industry-leading insights you can put to use today—all from the people who work here. READ CAREERS BLOG Job Alert Emails Customize your subscription to receive job alerts, the latest news, and insider tips tailored to your preferences. Discover the exciting and rewarding career opportunities that await you at Cisco.
Learn more about Cisco
Size
79,500 employees
Market Cap
$194.5 billion
Industry
Net Income
$10.1 billion
Founded
2014
5 Year Trend
+1.4%
Revenue
$48 billion
NASDAQ

Similar Jobs

More Jobs at Cisco

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: