Talent.com
Site Reliability / Gitops Engineer

Site Reliability / Gitops Engineer

CanonicalWashington, DC, United States
30+ days ago
Job type
  • Full-time
Job description

Join to apply for the Site Reliability / Gitops Engineer role at Canonical

1 day ago Be among the first 25 applicants

Join to apply for the Site Reliability / Gitops Engineer role at Canonical

Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's leading public cloud and silicon providers, and industry leaders in many sectors. The company is a pioneer of global distributed collaboration, with 1200+ colleagues in 75+ countries and very few office-based roles. Teams meet two to four times yearly in person, in interesting locations around the world, to align on strategy and execution.

The company is founder-led, profitable, and growing.

We are hiring a Site Reliability / Gitops Engineer to our Information Systems (IS) team. This role is an opportunity for an "automation-first" technologist with a passion for Linux to build a career with Canonical and drive the success with those leveraging Ubuntu and open source products. If you have experience of IT operations automation, Infrastructure as Code and a passion for technology, then you will enjoy working with some of the best people in the industry at Canonical.

Job Summary

The IS team at Canonical supports and maintains all of Canonical's IT production services. The team is in charge of running services used by over 60 million Ubuntu users.

As an SRE & Gitops engineer you'll be in a unique position to drive operations automation to the next level, both in our own private clouds as well as in the public clouds. We do this by utilizing the best of open source infrastructure as code software, software development practices such as CI / CD pipelines, and Canonical's leading products for software operation automation.

In addition to defining the infrastructure as code, you will improve Canonical products and the open-source technologies they're based on by providing critical feedback to developers on how their products operate at scale. This is done by submitting bugs (and sometimes writing pull requests) and collaborating on design and implementations with other teams within the company.

You'll be part of a global team of SREs that work together and support each other to provide the best possible services to our company, Canonical's customers and the Ubuntu Community.

Location : This role is available remotely in any timezone.

As a Site Reliability / Gitops Engineer engineer you will

  • Apply your experience of IaC to develop infrastructure as code practice within IS by constantly increasing automation and improving IaC processes
  • Automate software operations for re-usability and consistency across private and public clouds, taking into consideration the complexities of distributed systems
  • Develop new features and improve the resilience and scalability of the existing cloud and container portfolio at Canonical
  • Maintain operational responsibility for all of Canonical's core services, networks, and infrastructure
  • Develop skills in troubleshooting, capacity planning, and performance investigation, Setting up, maintaining and using observability tools such as Prometheus, Grafana, and Elasticsearch; design, implement and maintain monitoring and alerting for various systems and services
  • Collaborate with development teams to design service architecture, documentation, playbooks, policies and operational procedures
  • Provide assistance and work with globally distributed engineering, operations, and support peers
  • Be given uninterrupted development time to focus on larger projects and automation of manual tasks
  • Share your experience, know-how and best practices with other team members in design sessions, mentorship and 'doing work together'
  • Carry final responsibility for time-critical escalations

What we are looking for in you

  • A deep experience of, and knowledge to define operations in code, using version control, peer review and CI / CD to roll out changes both to applications and infrastructure
  • Strong modern engineering background (peer-review, unit testing, SCM, CI / CD, Agile)
  • Python software development experience, with large projects
  • Practical knowledge of Linux networking, routing, and firewalls
  • Affinity with various forms of Linux storage, from Ceph to Databases
  • Hands-on experience administering enterprise Linux servers
  • Extensive knowledge of cloud computing concepts and technologies
  • Bachelor's degree or greater, preferably in computer science or related engineering field
  • Able to communicate clearly and effectively in English over email, chat, video or voice calls and in-person
  • Motivated and able to troubleshoot from kernel to web, and willing to ask others when appropriate
  • A willingness to be flexible and able to learn new things quickly
  • Be inspired by the needs of fast-changing environments
  • Happy to work within distributed teams
  • Be passionate and familiarized about open-source, especially Ubuntu or Debian
  • About Canonical

    Canonical is a pioneering tech firm at the forefront of the global move to open source. As the company that publishes Ubuntu, one of the most important open-source projects and the platform for AI, IoT, and the cloud, we are changing the world of software. We recruit on a global basis and set a very high standard for people joining the company. We expect excellence; in order to succeed, we need to be the best at what we do. Most colleagues at Canonical have worked from home since our inception in 2004. Working here is a step into the future and will challenge you to think differently, work smarter, learn new skills, and raise your game.

    Canonical is an equal opportunity employer

    We are proud to foster a workplace free from discrimination. Diversity of experience, perspectives, and background create a better work environment and better products. Whatever your identity, we will give your application fair consideration.

    Seniority level

    Seniority level

    Mid-Senior level

    Employment type

    Employment type

    Full-time

    Job function

    Job function

    Engineering and Information Technology

    Industries

    Software Development

    Referrals increase your chances of interviewing at Canonical by 2x

    Get notified about new Site Reliability Engineer jobs in Spokane County, WA .

    Senior Site Reliability / Gitops Engineer

    Python and Kubernetes Software Engineer - Data, AI / ML & Analytics

    Software Engineer (Python / Linux / Packaging)

    Python and Kubernetes Software Engineer - Data, Workflows, AI / ML & Analytics

    Python Software Engineer - Ubuntu Hardware Certification Team

    Software Engineer - Solutions Engineering

    Software Engineer, Ceph & Distributed Storage

    Distributed Systems Software Engineer, Python / Go

    Graduate Software Engineer, Open Source and Linux, Canonical Ubuntu

    Golang System Software Engineer - Containers / Virtualisation

    System Software Engineer - Ubuntu Networking

    Software Engineer - packaging - optimize Ubuntu Server

    Software Engineer - Cross-platform C++ - Multipass

    Software Engineer - packaging - optimize Ubuntu Server for public clouds

    Software Engineer - packaging - optimize Ubuntu Server

    Embedded Linux Senior Software Engineer - Optimisation

    Software Engineer - packaging - optimize Ubuntu Server for public clouds

    Software Engineer - packaging - optimize Ubuntu Server for public clouds

    Software Engineer - packaging - optimize Ubuntu Server for public clouds

    Software Engineer - packaging - optimize Ubuntu Server

    We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.

    #J-18808-Ljbffr

    Create a job alert for this search

    Site Reliability Engineer • Washington, DC, United States

    Related jobs
    • Promoted
    Staff Site Reliability Engineer

    Staff Site Reliability Engineer

    VisaAshburn, VA, United States
    Full-time
    Visa is a world leader in payments and technology, with over 259 billion payments transactions flowing safely between consumers, merchants, financial institutions, and government entities in more t...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Leidos IncReston, VA, United States
    Full-time
    The Multi Domain Solutions Division at Leidos is looking for a.This role involves supporting the delivery of comprehensive IT and support services to ensure mission success while adhering to DoD st...Show moreLast updated: 16 days ago
    Site Reliability Engineer

    Site Reliability Engineer

    Tax AnalystsFalls Church, VA, US
    Full-time
    Quick Apply
    Tax Analysts is seeking a Site Reliability Engineer (SRE) to help establish and shape our reliability engineering practice from the ground up. This is a unique opportunity to join a mission-driven o...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer III

    Site Reliability Engineer III

    VerisignReston, VA, United States
    Full-time
    Verisign helps enable the security, stability, and resiliency of the internet.We are a trusted provider of internet infrastructure services for the networked world and deliver unmatched performance...Show moreLast updated: 30+ days ago
    • Promoted
    Lead Site Reliability Engineer

    Lead Site Reliability Engineer

    Federated ITWashington, DC, United States
    Full-time
    Bridge Defense is redefining how modern defense technology is delivered.Department of Defense, the Intelligence Community, and federal law enforcement agencies. We provide full-spectrum national sec...Show moreLast updated: 5 days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Improvix TechnologiesWashington, DC, United States
    Full-time
    Site Reliability Engineer (SRE).We are seeking a Site Reliability Engineer (SRE) with strong GitLab expertise to support and enhance enterprise platforms. This role will focus primarily on GitLab wh...Show moreLast updated: 1 day ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    CanonicalWashington, DC, United States
    Full-time
    Canonical is a leading provider of open source software and operating systems.Our platform, Ubuntu, is used across enterprise initiatives in public cloud, data science, AI, engineering innovation, ...Show moreLast updated: 4 days ago
    • Promoted
    Senior Reliability Engineer

    Senior Reliability Engineer

    The Johns Hopkins University Applied Physics LaboratoryLaurel, MD, United States
    Full-time
    Are you passionate about applying reliability and system engineering principles to analyze and assess the resilience of future strategic weapon systems?. Do you have a strong technical background in...Show moreLast updated: 7 days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    CSCI ConsultingQuantico, VA, United States
    Full-time
    CSCI Consulting is looking for a.Site Reliability Engineer (SRE).This role combines deep systems engineering knowledge with DevOps automation, proactive monitoring, and incident response practices....Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Powder River IndustriesWashington, DC, United States
    Full-time
    Conduct analysis of alternatives for configuration tools, make recommendations, work with team to design, develop, test, implement, and maintain tool choice. Responsible for the administration, moni...Show moreLast updated: 4 days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    EngFlowWashington, DC, United States
    Full-time
    Join to apply for the Site Reliability Engineer role at EngFlow.At EngFlow, we help developers save time by accelerating software builds and tests. Our cloud-based, distributed service optimizes dev...Show moreLast updated: 4 days ago
    • Promoted
    Principal Site Reliability Engineer (SRE) at Jobgether Washington DC

    Principal Site Reliability Engineer (SRE) at Jobgether Washington DC

    JobgetherWashington, DC, United States
    Full-time
    Principal Site Reliability Engineer (SRE) job at Jobgether.This position is posted by Jobgether on behalf of.We are currently looking for a. Principal Site Reliability Engineer (SRE).Join a high-imp...Show moreLast updated: 30+ days ago
    • Promoted
    Deployment Site Reliability Engineer - Connected Warfare

    Deployment Site Reliability Engineer - Connected Warfare

    Anduril Industries, Inc.Washington, DC, United States
    Full-time
    Senior Deployed Site Reliability Engineer, Connected Warfare.Washington, District of Columbia, United States.Anduril Industries is a defense technology company with a mission to transform U.By brin...Show moreLast updated: 4 days ago
    • Promoted
    Gitlab Site Reliability Engineer

    Gitlab Site Reliability Engineer

    2Prod Technologies Corp.Washington, DC, United States
    Full-time
    Site Reliability Engineer (SRE) with strong GitLab expertise to support and enhance enterprise platforms.This role will focus primarily on GitLab while also maintaining Jira and Confluence in a sec...Show moreLast updated: 4 days ago
    • Promoted
    • New!
    Site Reliability Engineer — Scale mission-critical platforms

    Site Reliability Engineer — Scale mission-critical platforms

    Anduril IndustriesWashington, DC, United States
    Full-time
    A defense technology company is seeking a Site Reliability Engineer in Washington, DC.The role involves solving challenges in networking and systems integration while working with cross-functional ...Show moreLast updated: 7 hours ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    CapeWashington, DC, United States
    Full-time
    Cape was founded in early 2022 by Palantir and Anduril alums with deep expertise in privacy and national security.While running Palantir’s US national security business, our CEO became passionate a...Show moreLast updated: 4 days ago
    • Promoted
    Lead Site Reliability Engineer

    Lead Site Reliability Engineer

    Bridge DefenseWashington, DC, United States
    Full-time
    Bridge Defense is redefining how modern defense technology is delivered.Department of Defense, the Intelligence Community, and federal law enforcement agencies. We provide full-spectrum national sec...Show moreLast updated: 4 days ago
    • Promoted
    Staff Site Reliability Engineer (Federal)

    Staff Site Reliability Engineer (Federal)

    Okta for DevelopersWashington, DC, United States
    Full-time
    Staff Site Reliability Engineer (Federal).Be among the first 25 applicants.Okta is The World’s Identity Company.We free everyone to safely use any technology, anywhere, on any device or app.Our fle...Show moreLast updated: 4 days ago