Submitting more applications increases your chances of landing a job.

Here’s how busy the average job seeker was last month:

Opportunities viewed

Applications submitted

Keep exploring and applying to maximize your chances!

Looking for employers with a proven track record of hiring women?

Click here to explore opportunities now!
We Value Your Feedback

You are invited to participate in a survey designed to help researchers understand how best to match workers to the types of jobs they are searching for

Would You Be Likely to Participate?

If selected, we will contact you via email with further instructions and details about your participation.

You will receive a $7 payout for answering the survey.


User unblocked successfully
Thank you. Your report has been submitted and will be reviewed shortly.
https://bayt.page.link/Db1bPiX7doBfDESD9
Back to the job results

Site Reliability Engineer - SRE Fleet

4 days ago 2026/11/21 ·Application closes in 115 days
Remote
Other Business Support Services
Create a job alert for similar positions
Job alert turned off. You won’t receive updates for this search anymore.

Job description

Meet the Team

The SRE Fleet team is responsible for maintaining the stability, scalability, and efficiency of the infrastructure that powers our global cloud platform. As a team of six engineers distributed across the US, Canada, and the UK, we combine deep infrastructure expertise with a strong focus on automation, reliability, and operational excellence. We are one of several SRE teams working together to support a platform that serves more than 500,000 customers and manages over 18 million devices worldwide.



The team operates with a high degree of autonomy, giving engineers the opportunity to drive both critical initiatives and grassroots improvements that solve real operational challenges. One of our most exciting areas of focus is expanding our ability to build highly automated regional, sovereign, and isolated cloud environments that support new markets and evolving regulatory requirements.



Everyone on the team has a voice, and engineers are encouraged to identify problems, propose solutions, and take ownership of improvements that make the platform more reliable and easier to operate.




Your Impact

Develop and maintain automation solutions that improve the reliability, scalability, and operational efficiency of infrastructure spanning more than 2,000 machines across global cloud environments. Design and enhance deployment pipelines, testing frameworks, and operational tooling to support the continued growth of a platform serving millions of managed devices worldwide. Troubleshoot complex infrastructure and distributed systems issues to ensure high availability while helping teams identify and address performance and scalability challenges.



Contribute to critical projects such as cluster build out by building automation that enables the rapid and repeatable deployment of new sovereign, regional, and purpose-built cloud environments. Partner with other engineering teams, product management, and business partners across multiple teams and time zones to understand platform dependencies, seek opportunities for improvement, and deliver solutions that enhance reliability and reduce operational overhead.




Minimum Qualifications
  • 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments.
  • Experience developing and maintaining infrastructure automation using Ansible.
  • Experience programming in Ruby and developing automated tests using RSpec or comparable testing frameworks.
  • Experience administering and troubleshooting Linux-based systems and distributed infrastructure environments.
  • Experience designing, implementing, and maintaining CI/CD pipelines, including GitLab CI.
  • Experience supporting large-scale infrastructure environments consisting of hundreds or thousands of systems.

Preferred Qualifications
  • Familiarity with AWS or other public cloud platforms and hybrid infrastructure environments.
  • Knowledge of monitoring, observability, and reliability engineering practices and tooling.
  • Familiarity with Kubernetes concepts and containerized application platforms.
  • Experience leveraging AI-assisted development tools to improve software development, automation, operational analysis, and engineering productivity.

Why Cisco? 

At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.



Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere. 



We are Cisco, and our power starts with you. 





This job post has been translated by AI and may contain minor differences or errors.
You’ve reached the maximum limit of 15 job alerts. To create a new alert, please delete an existing one first.
Job alert created for this search. You’ll receive updates when new jobs match.
Are you sure you want to unapply?

You'll no longer be considered for this role and your application will be removed from the employer's inbox.