A Site Reliability GitOps Engineer at Canonical is responsible for automating and maintaining the company's IT production services using infrastructure as code, enhancing system resilience and scalability, and collaborating with global teams to improve open source cloud and container technologies.
Key Responsibilities
Develop and implement infrastructure as code to automate IT operations across private and public clouds.
Automate software operations to ensure reusability and consistency in distributed systems.
Enhance the resilience and scalability of cloud and container services.
Maintain operational responsibility for core services, networks, and infrastructure.
Set up, monitor, and troubleshoot systems using observability tools like Prometheus, Grafana, and Elasticsearch.
Collaborate with development teams to design service architecture, documentation, and operational procedures.
Support and coordinate with globally distributed engineering, operations, and support teams.
Lead and resolve time-critical escalations related to system operations.
Requirements
Deep experience of, and knowledge to define operations in code, using version control, peer review and CI CD to roll out changes both to applications and infrastructure
Strong modern engineering background including peer-review, unit testing, SCM, CI CD, and Agile methodologies
Python software development experience, with large projects
Practical knowledge of Linux networking, routing, and firewalls
Affinity with various forms of Linux storage, from Ceph to Databases
Hands-on experience administering enterprise Linux servers
Extensive knowledge of cloud computing concepts and technologies
Bachelor’s degree or greater, preferably in computer science or related engineering field
Able to communicate clearly and effectively in English over email, chat, video or voice calls and in-person
Motivated and able to troubleshoot from kernel to web, and willing to ask others when appropriate
A willingness to be flexible and able to learn new things quickly
Be inspired by the needs of fast-changing environments
Happy to work within distributed teams
Be passionate and familiarized about open-source, especially Ubuntu or Debian
Benefits & Perks
Remote work available in any timezone
Uninterrupted development time to focus on larger projects and automation
Supportive global team collaboration
Opportunities for skill development in troubleshooting, capacity planning, and performance investigation
Work in a company that values open-source contributions and community engagement
Part of a pioneering tech firm at the forefront of open source innovation
Work environment that fosters diversity and equal opportunity
Ready to Apply?
Join Canonical and make an impact in renewable energy