Site Reliability Engineer (SRE) - Early Talent
Nebius
- Location
- Amsterdam, Netherlands
- Work model
- On-Site
- Level
- Mid
- Posted
- 2h ago
Skills
About this role
About Nebius
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.
Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
Summary
Location: Amsterdam Duration: 3-6 months Start date: January 2027 Compensation: Paid Eligibility: Current University student (Computer Science or related field), Recent Graduate or Early Career specialist Work authorization: Permitted to work in the job’s location
Why work at Nebius Nebius is leading a new era in cloud computing to serve the global AI economy. We create the tools and resources our customers need to solve real-world challenges and transform industries, without massive infrastructure costs or the need to build large in-house AI/ML teams. Our employees work at the cutting edge of AI cloud infrastructure alongside some of the most experienced and innovative leaders and engineers in the field.
Your responsibilities will include
• Operation:
• assist in day-to-day SRE operations tasks
• Project work:
• execute tasks from the backlog
• together or instead: work on small and well-defined SRE project from backlog
• create tests for their changes/project (if applicable)
• write technical documentation about their work
• track and update Jira tasks according to their progress
• Learning and growth:
• actively study system fundamentals: Linux ( networking, cgroups, ebpf, etc.)
• learn internal tools, workflows and operational standards
• learn container and platform fundamentals: Kubernetes (core architecture and networking), Helm, Terraform
• learn configuration management and automation basics
What we expect you to have
• demonstrate strong interest and foundational knowledge in cloud providers, operations, networking, hardware, and software
• programming experience in any of Python, Go, C++ or Java
• basic understanding of the Network (ethernet, ip, tcp/udp, routing fundamentals)
• familiarity with Git
• basic understanding of containers, Kubernetes and Gitops approach
What we offer
• Mentorship from experienced AI, ML, and cloud infrastructure professionals
• Hands-on experience with real customer workloads and production systems
• Opportunities for professional growth within Nebius
• Opportunity to be considered for a full-time role after the Early Talent Program
We’re growing and expanding our products every day. If you’re up to the challenge and are excited about AI and ML as much as we are, join us!
Benefits & Perks
• Competitive compensation
• Career growth and learning opportunities
• Flexibility and ownership
• Collaborative and innovative culture
• Opportunity to work on impactful AI