Principal Site Reliability Engineer
3 settimane fa
Principal Site Reliability Engineer - Azure Red Hat OpenShift in Madrid or Remote The Red Hat Site Reliability Engineering (SRE) team is looking for a Principal Site Reliability Engineer to join us. In this role, you will develop, scale, and operate our OpenShift managed cloud services. OpenShift is Red Hat's enterprise Kubernetes distribution. As an SRE you will contribute to running OpenShift at scale by enabling customer self‐service, making our monitoring system more sustainable, and eliminating work through automation. On the SRE team, you will have the opportunity to influence the complex challenges of scale which are unique to Red Hat Managed Cloud Services, while using your skills in coding, operations, and large‐scale distributed system design. What you will do Lead the development of code and automation scripts to optimize the scalability, reliability, and performance of services. Conduct thorough code reviews and implement best practices in software development to maintain a high‐quality codebase. Mentor and guide junior engineers, fostering a culture of continuous learning and improvement within the team. Design and implement advanced monitoring and alerting systems to proactively detect and resolve issues. Coordinate and lead complex incident response procedures, ensuring timely resolution and thorough post‐mortems. Interface with internal stakeholders and external cloud providers to architect, build, and maintain fault‐tolerant systems. Manage large‐scale, distributed systems, focusing on minimizing downtime and improving system resilience. Participate in an on‐call rotation and provide leadership during critical incidents to ensure 24x7x365 production support. Lead the continuous enhancement of the SRE team's processes, tools, and methodologies to support the evolving needs of the service. What you will bring 5 years of software engineering experience using object‐oriented languages; Golang is preferred. Extensive experience managing Linux‐based systems in a public cloud such as AWS, GCP, or Azure. Proficient experience with enterprise systems monitoring; knowledge of Prometheus is preferred. Extensive experience with enterprise configuration management such as Ansible, Puppet, or Chef. 1 year experience with container‐related technologies like Docker or Kubernetes. Experience with containers on Linux. Solid understanding of standard TCP/IP networking and common protocols like DNS and Good verbal and written communication skills in English. About Red Hat Red Hat is the world's leading provider of enterprise open source software solutions, using a community‐powered approach to deliver high‐performing Linux, cloud, container, and Kubernetes technologies. Spread across 40 countries, our associates work flexibly across work environments, from in‐office, to office‐flex, to fully remote, depending on the requirements of their role. Red Hatters are encouraged to bring their best ideas, no matter their title or tenure. We're a leader in open source because of our open and inclusive environment. We hire creative, passionate people ready to contribute their ideas, help solve complex problems, and make an impact. Inclusion at Red Hat Red Hat's culture is built on the open source principles of transparency, collaboration, and inclusion, where the best ideas can come from anywhere and anyone. When this is realized, it empowers people from different backgrounds, perspectives, and experiences to come together to share ideas, challenge the status quo, and drive innovation. Our aspiration is that everyone experiences this culture with equal opportunity and access, and that all voices are not only heard but also celebrated. We hope you will join our celebration, and we welcome and encourage applicants from all the beautiful dimensions that compose our global village. Equal Opportunity Policy (EEO) Red Hat is proud to be an equal opportunity workplace and an affirmative action employer. We review applications for employment without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, citizenship, age, veteran status, genetic information, physical or mental disability, medical condition, marital status, or any other basis prohibited by law. J-18808-Ljbffr
-
Principal Site Reliability Engineer
3 settimane fa
Piemonte, Italia SaaS Industry A tempo pienoPrincipal Site Reliability Engineer - Azure Red Hat OpenShift in Madrid or Remote The Red Hat Site Reliability Engineering (SRE) team is looking for a Principal Site Reliability Engineer to join us. In this role, you will develop, scale, and operate our OpenShift managed cloud services. As an SRE you will contribute to running OpenShift at scale by enabling...
-
Site Reliability Engineer
2 settimane fa
Piemonte, Italia Canonical A tempo pienoOverviewJoin to apply for the Site Reliability Engineer role at Canonical . Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is widely used in breakthrough enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our...
-
Senior Site Reliability Engineer
2 settimane fa
Piemonte, Italia Canonical A tempo pienoOverviewJoin to apply for the Senior Site Reliability Engineer role at Canonical . Location: Globally remote role. Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is widely used in breakthrough enterprise initiatives such as public cloud, data science, AI,...
-
Senior Site Reliability Engineer
3 settimane fa
Piemonte, Italia Remotely A tempo pienoLocation LATAM, ERUOPE CloudDevs works with fast-moving, venture-backed startups across the US. We're building a pool of world-class Site Reliability Engineers for current roles and for upcoming opportunities. You will either be placed directly into one of our partner startups or added to our vetted SRE network for future projects. This role is ideal for...
-
Principal SRE: OpenShift on Azure
3 settimane fa
Piemonte, Italia SaaS Industry A tempo pienoA leading technology firm is looking for a Principal Site Reliability Engineer to develop, scale, and operate OpenShift managed cloud services. The role emphasizes high-quality code development, system reliability, and mentoring junior engineers. Candidates should have strong experience in software engineering, Linux systems, and automation practices. This...
-
Senior Site Reliability Engineer
3 settimane fa
Piemonte, Italia SaaS Industry A tempo pienoSenior Site Reliability Engineer - OpenShift-Based Platform in Madrid or Remote Red Hat is looking for a Platform Engineer to join its Platform Engineering team! In this role, you will help architect, implement, improve, and support the OpenShift-based platform that runs many of Red Hat's most important multi-tenant Software-as-a-Service (SaaS) and...
-
Site Reliability Engineer
2 settimane fa
Piemonte, Italia Immobiliare.it A tempo pienoAzienda società insurtech, sede principale a Londra 50 dipendenti opportunità di lavoro in full-remote Offerta Sviluppo di modelli predittivi per il rischio Attività: Progettare e sviluppare modelli di... L’AI/LLM Engineer contribuisce allo sviluppo di prodotti software innovativi integrando modelli linguistici avanzati (locali e cloud) all’interno...
-
Site reliability engineer
2 settimane fa
Piemonte, Italia Immobiliare.it A tempo pienoImmobiliare.it S.p. A. è un gruppo italiano composto da società specializzate in servizi Digital Tech per la compravendita e l'affitto di immobili, rivolti a privati, professionisti del real estate, istituti bancari e operatori del settore finanziario. [...] Immobiliare.it Insights, la proptech della società, offre servizi digitali di advisory, insights e...
-
Site Reliability Engineer
2 settimane fa
Torrazza Piemonte (TO), Italia Immobiliare.it A tempo pienoAzienda società insurtech, sede principale a Londra 50 dipendenti opportunità di lavoro in full-remote Offerta Sviluppo di modelli predittivi per il rischio Attività: Progettare e sviluppare modelli di L'AI/LLM Engineer contribuisce allo sviluppo di prodotti software innovativi integrando modelli linguistici avanzati (locali e cloud) all'interno...
-
Site Reliability Engineering Manager
3 settimane fa
Piemonte, Italia Canonical A tempo pienoJoin to apply for the Site Reliability Engineering Manager role at Canonical 1 day ago Be among the first 25 applicants Join to apply for the Site Reliability Engineering Manager role at Canonical Get AI-powered advice on this job and more exclusive features. Canonical is a leading provider of open-source software and operating systems for global enterprise...