DevOps Engineer – Cloud Native AI
Salva questo lavoro e mantieni la tua ricerca organizzata
Crea un account gratuito per salvare lavori, creare avvisi e tornare a questa inserzione dalla tua dashboard.
Employment Type: Full-Time, preferably Freelance
Department: Engineering / AI Infrastructure
Language Requirement: Italian & English
AIVB isn't just a startup; we are a high-velocity scaleup redefining how AI solutions reach the market.
Hyper-Growth: 300% YoY growth, jumping from zero to £1.4M in our first full year.
Proven Pedigree: Founded by the pioneers behind Deep Learning Italia, the gold standard for AI education in Italy.
Global Ambition: Backed by £4.5M in seed funding, we are preparing for a major US-led Series A in Q4 2026 to dominate international markets.
We are building the \"Operating System\" for AI ventures. We don't just build products; we build companies.
About AIVBFounded in 2024 by Marco Tognoni and Matteo Testi, AIVB is an AI Venture Builder designed to scale innovation at speed. Our team of 40 is anchored by an elite engineering core, primarily AI PhDs, who transform complex research into vertical‑specific powerhouses.
We operate through a Centralized Powerhouse that develops strategic blueprints for specialized spinoffs in high‑impact sectors such as Fashion, Real Estate, and CyberDefense. We are \"International by Design\", English is our language, and global scale is our target.
Why this role mattersWe are building cutting‑edge AI solutions within the boundaries of the Italian National Cloud Strategy. You won't just be clicking buttons in AWS; you will be architecting compliant, self‑hosted AI platforms on sovereign infrastructure. This is a role for an engineer who understands how to build cloud‑native reliability without relying on public hyperscalers.
Role SummaryAs a DevOps Engineer for Sovereign AI, you will design and maintain the infrastructure that powers our AI/ML workloads on Italian National Cloud providers (e.g., SeeWeb, Netalia, Aruba). Your focus will be on self‑managed Kubernetes, strict Data Sovereignty, and implementing open‑source MLOps toolchains that function independently of US‑based public clouds.
Key Responsibilities1. Sovereign Infrastructure & Orchestration
Deploy and manage production‑grade Kubernetes clusters on private cloud or national provider infrastructure (using Kubernetes / Rancher).
Ensure high availability and disaster recovery within the specific zones/regions of the national provider.
2. Self-Hosted MLOps
Since we cannot use managed services, you will architect and maintain a self‑hosted MLOps stack using tools like MLflow, or Kubeflow.
Configure and optimize MinIO‑like S3‑compatible object storage (e.g. SeaweedFS, Garage) to handle large training datasets locally.
Manage container registries (Harbor) located strictly within Italian borders.
3. Compliance & Security (GDPR/AGID)
Strictly enforce Data Sovereignty principles; ensure no data egresses outside of Italy/EU.
Implement security standards compliant with AGID (Agenzia per l'Italia Digitale) and ACN (Agenzia per la Cybersicurezza Nazionale) guidelines.
4. GPU & Hardware Optimization
Configure NVIDIA vGPU or PCI passthrough on virtualized national cloud instances.
Optimize the AI stack (CUDA drivers, Container Toolkit) for maximum performance on constrained infrastructure.
Serverless GPU usage experience
The environment is \"Cloud Native\" but relies heavily on open‑source and self‑hosted equivalents of public cloud services.
Domain & Related Technology:
Cloud Environment - Italian National Cloud (Seeweb, Netalia)
Orchestration - Kubernetes, Rancher
Virtualization - es. Proxmox, OpenStack, KVM, VMware
Storage - SeaweedFS, Garage, Vast
AI/ML Platform - Kubeflow, MLflow, JupyterHub
CI/CD - Github actions, ArgoCD
Observability - Prometheus, Grafana, Loki (PLG Stack)
Required:
2+ years of experience in roles such as: DevOps Engineer, DevSecOps Engineer, Cloud Architect, Infrastructure Engineer, Release Engineer, Platform Engineer, MLOps
Strong expertise in Kubernetes: you must be able to deploy, manage, and troubleshoot K8s clusters without relying on external support (e.g., cloud providers)
Hands‑on experience with containerization and orchestration tools (Docker, Helm)