Loves Travel Stops & Country Stores

Cloud Systems Engineer II

Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 2-4+ years supporting cloud/infrastructure in production
  • Experience with AWS platforms at scale
  • Track record in resolving complex infrastructure issues
  • Hands-on with Infrastructure-as-Code (Terraform preferred)
  • Familiarity with hybrid/distributed environments including edge compute
  • Practical experience applying AI tools in infrastructure operations

Responsibilities

  • Operate AWS infrastructure including EC2, S3, and VPC
  • Resolve Tier II/III incidents impacting cloud systems
  • Troubleshoot complex issues using logs and metrics
  • Implement root cause analysis for incident prevention
  • Automate operational processes to improve efficiency
  • Monitor system performance and optimize cloud spend
  • Support edge compute and hybrid infrastructure environments

Benefits

  • Company-funded tuition assistance
  • Paid Time Off
  • 401(k) with 100% match up to 5%
  • Medical/Dental/Vision insurance after 30 days
  • Career development opportunities
Full Job Description
Req ID: 492316

Benefits: * Fuel Your Growth with Love's - company funded tuition assistance * Paid Time Off * 401(k) - 100% Match up to 5% * Medical/Dental/Vision Insurance after 30 days * Competitive Pay * Career Development *

This position will be an on-site position at Love's Corporate Office in Oklahoma City, Oklahoma.

The Cloud Platform Operations Engineer II is responsible for the day-to-day operation, stability, and continuous improvement of a primarily AWS-based cloud platform, along with a distributed set of edge compute environments. This role operates in a highly dynamic, ticket-driven environment, ensuring rapid incident resolution, high system reliability, and consistent service delivery.

This position goes beyond traditional support by leveraging automation and AI-enabled tools to enhance troubleshooting, streamline operational workflows, and proactively identify and resolve systemic issues. The role partners across infrastructure, engineering, and security teams to ensure the platform is scalable, secure, and optimized for performance and cost.

ESSENTIAL DUTIES

Cloud Operations & Support
  • Operate and support AWS infrastructure, including EC2, S3, IAM, VPC, and associated services in a production environment
  • Respond to and resolve Tier II/III incidents impacting cloud and edge systems, ensuring timely restoration of service
  • Troubleshoot complex issues across compute, storage, networking, and access layers using logs, metrics, and system data


Incident Management & Continuous Improvement
  • Perform root cause analysis and implement corrective actions to prevent recurrence
  • Identify recurring operational issues and translate them into scalable solutions through automation or design improvements
  • Maintain and enhance monitoring, alerting, and observability practices to improve system reliability


Automation & Infrastructure as Code
  • Design, implement, and maintain Infrastructure-as-Code solutions (e.g., Terraform, CloudFormation)
  • Develop and maintain automation scripts to reduce manual intervention and improve operational efficiency
  • Apply an automation-first mindset to all repeatable operational processes


AI-Enabled Operations
  • Leverage AI tools (e.g., LLMs, copilots, log analysis platforms) to accelerate diagnostics, summarize system behavior, and recommend remediation actions
  • Contribute to the development of AI-enabled operational capabilities, including knowledge retrieval, intelligent runbooks, and workflow automation
  • Identify opportunities to integrate AI into operational processes to improve speed, accuracy, and scalability


Platform Performance, Cost & Reliability
  • Monitor system performance, availability, and cost utilization, proactively addressing anomalies
  • Optimize cloud spend by identifying underutilized or misconfigured resources and implementing cost controls
  • Support reliability engineering practices to improve uptime and service resilience


Edge & Hybrid Environment Support
  • Support edge compute and hypervisor environments, including connectivity, synchronization, and integration with cloud platforms
  • Troubleshoot hybrid infrastructure issues spanning on-premise/edge and cloud environments


Documentation & Operational Excellence
  • Develop and maintain runbooks, knowledge articles, and standard operating procedures
  • Ensure documentation reflects current-state architecture and operational practices
  • Contribute to a culture of operational discipline, knowledge sharing, and continuous improvement


SKILLS & QUALIFICATIONS

Experience
  • 2-4+ years of experience supporting cloud or infrastructure environments in a production setting
  • Demonstrated experience operating and supporting AWS-based platforms at scale
  • Proven track record of resolving complex, real-world infrastructure issues end-to-end (compute, storage, networking, access)
  • Hands-on experience implementing Infrastructure-as-Code solutions in a production environment (Terraform preferred)
  • Experience working in ticket-driven, operational support environments with defined SLAs and incident management processes
  • Exposure to hybrid or distributed environments, including edge compute or virtualization platforms
  • Practical experience applying AI tools to infrastructure operations (e.g., log analysis, incident triage, workflow automation)


Hard Skills
  • Cloud Platforms: AWS (EC2, S3, IAM, VPC, networking fundamentals)
  • Infrastructure as Code: Terraform (preferred), CloudFormation
  • Configuration Management: Ansible or similar tools
  • Systems Administration: Linux and/or Windows server environments
  • Troubleshooting: Root cause analysis across compute, storage, networking, and access layers
  • Monitoring & Observability: Experience with logging, metrics, and alerting tools
  • Networking Fundamentals: DNS, routing, firewalls, and connectivity troubleshooting
  • Automation & Scripting: Ability to automate repetitive operational tasks (e.g., Python, Bash, or similar)
  • Cloud Security Fundamentals: IAM, access controls, and security best practices
  • AI Tool Application: Use of LLMs, copilots, or log analysis platforms to support diagnostics and operational efficiency
  • Edge/Hybrid Infrastructure (Preferred): Virtualization, edge compute, and cloud synchronization concepts


Soft Skills
  • Operational Excellence: Strong sense of ownership, accountability, and urgency in maintaining system stability and performance
  • Analytical Thinking: Structured problem-solving with the ability to diagnose and resolve complex issues under pressure
  • Automation Mindset: Proactively identifies opportunities to eliminate manual effort and improve efficiency
  • Adaptability: Effective in fast-paced, ticket-driven environments with shifting priorities
  • Collaboration: Works effectively across infrastructure, engineering, and security teams to drive outcomes
  • Communication: Translates technical issues into clear, actionable insights for both technical and non-technical stakeholders
  • Continuous Improvement: Actively seeks opportunities to enhance processes, tools, and platform reliability
  • Learning Agility: Stays current with evolving cloud technologies, automation practices, and AI capabilities
  • Enterprise Mindset & Team Contribution: Proactively supports broader team and enterprise priorities by contributing to initiatives beyond core responsibilities; steps in where needed to drive outcomes and ensure collective success


Education:
  • Bachelor's degree in Computer Science or a related discipline such as Information Technology, Software Engineering, or Computer engineering is required.


WORK ENVIRONMENT
  • Prolonged sitting, some bending and stooping
  • Eye strain (screen use)
  • Manual dexterity sufficient to operate a computer keyboard and calculator
  • Occasional lifting of up to 25 pounds
  • Requires normal range of hearing and vision
  • Additional hours may be necessary.


Note: The items identified above are representative of those commonly associated with this position but are not exhaustive. Employees may encounter additional or unforeseen responsibilities in the course of their duties.
  • This job description should not be construed to imply that these requirements are the exclusive standards of the position. All employees may be required to follow any other instructions, cross train in other positions, and perform other duties as required by workloads.

About Loves Travel Stops & Country Stores

Love's Travel Stops & Country Stores is a family-owned chain of gas stations and convenience stores. The company was founded in 1964 and has since expanded to over 500 locations across 41 states in the United States. Love's offers a variety of services including fuel, food, and merchandise. The company is known for its commitment to customer service and has won numerous awards for its efforts. Love's is also involved in philanthropic efforts, supporting organizations such as Children's Miracle Network Hospitals and the United Way.
Learn more about Loves Travel Stops & Country Stores
Size
32,000 employees
Industry
Founded
1964

Similar Jobs

More Jobs at Loves Travel Stops & Country Stores

More Information Technology Jobs

Find similar Cloud Systems Engineer II jobs: