LogoKode$word
Nebius Group logo
Verified Tech Organization

Careers at Nebius Group

Browse and filter through all verified positions currently open at Nebius Group.

Total Company Roles175
Matching Filter175
nebius.com/companyHQ: Schiphol, NLCEO: Arkady Volozh1543 employees

Nebius Group N.V. is a technology company dedicated to developing comprehensive infrastructure to serve the global artificial intelligence industry. Its operations encompass several key areas. Central to its mission is Nebius, an AI-focused cloud platform engineered to handle demanding AI workloads. This division constructs end-to-end AI infrastructure, featuring extensive GPU computing clusters, robust cloud platforms, and essential tools and services for developers. The group also includes Toloka AI, which functions as a data solutions provider, assisting with various phases of generative AI development. TripleTen operates as an educational technology venture, focused on equipping individuals with new skills for careers in the tech sector. Furthermore, Avride specializes in pioneering autonomous driving technologies for self-driving vehicles and delivery robots. Founded in 1989, the company was previously known as Yandex N.V. until its rebranding to Nebius Group N.V. in August 2024. Its headquarters are located in Amsterdam, the Netherlands, with additional research and development facilities spread across Europe, North America, and Israel.

Sector:Software Application

All Openings (175)

Ordered by most recently published

Senior Software Engineer (Managed PostgreSQL)

On-sitefull timeSeniorAmsterdam, Netherlands
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We're looking for a Senior Software Engineer to help us build the best Managed PostgreSQL for AI workloads. The service is already in production; the mission now is to make it the obvious choice for teams running AI applications on Postgres — first-class vector search, painless migrations from RDS / Cloud SQL / self-managed clusters, and the operational quality of a mature managed database. You'll work across the stack: control plane and lifecycle automation in Go, Postgres internals and performance, migration tooling, and the customer escalations where it actually matters. You're welcome to work from our offices in Amsterdam or London, hybrid or remotely from EU timezones. Your responsibilities : Develop the control plane and lifecycle automation for Managed PostgreSQL — provisioning, HA, failover, backups, PITR, version upgrades, zero-downtime maintenance. Tune and harden PostgreSQL itself — replication, WAL, vacuum, query planner, connection pooling, extensions — and turn that into product features and sane defaults customers don't have to think about. Build migration tooling that gets customers off AWS RDS, Google Cloud SQL, Azure Database for PostgreSQL, and self-managed clusters with minimal downtime. Drive the AI-Postgres story end to end — vector search (pgvector, pgvectorscale), hybrid retrieval, integration with the wider Nebius AI Cloud stack. Run the service like an SRE — define SLOs, build observability, lead incident response, feed every postmortem back into the platform. Work directly with customers on architecture reviews, performance escalations, and the production problems that don't fit a ticket template. Must-haves : 5+ years of professional software engineering experience, with significant time spent building or operating production PostgreSQL at scale (TB+ datasets, replication topologies, real failure modes you've debugged in production). Strong software engineering skills in Go or another backend/systems language, with a willingness to work primarily in Go. Deep knowledge of PostgreSQL internals: MVCC, WAL, replication (physical and logical), vacuum, query planning, extensions, partitioning. Hands-on experience with the surrounding ecosystem: Patroni / Stolon / pg_auto_failover, pgBackRest / WAL-G, pgBouncer / PgCat, logical replication tooling. The instincts of an ex-DBA — you can read EXPLAIN ANALYZE fluently, reason about lock behaviour, and know how to handle database corruptions. Ability to write reliable code and dig into complex problems. Teamwork-oriented approach. Nice - to - have s : Experience with pgvector and pgvectorscale — vector search at scale, index choice, recall vs. latency trade-offs. Background building managed-database control planes at a cloud provider (RDS, Cloud SQL, Aiven, Crunchy, Timescale, Supabase, Neon, or similar). Experience writing Kubernetes operators (Go + controller-runtime / kubebuilder). Contributions to PostgreSQL itself, popular extensions, or the surrounding OSS ecosystem. AI/ML workload experience — RAG pipelines, embedding stores, GPU-resident workloads. We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We're looking for a Senior Software Engineer to help us build the best Managed PostgreSQL for AI workloads. The service is already in production; the mission now is to make it the obvious choice for teams running AI applications on Postgres — first-class vector search, painless migrations from RDS / Cloud SQL / self-managed clusters, and the operational quality of a mature managed database. You'll work across the stack: control plane and lifecycle automation in Go, Postgres internals and performance, migration tooling, and the customer escalations where it actually matters. You're welcome to work from our offices in Amsterdam or London, hybrid or remotely from EU timezones. Your responsibilities : Develop the control plane and lifecycle automation for Managed PostgreSQL — provisioning, HA, failover, backups, PITR, version upgrades, zero-downtime maintenance. Tune and harden PostgreSQL itself — replication, WAL, vacuum, query planner, connection pooling, extensions — and turn that into product features and sane defaults customers don't have to think about. Build migration tooling that gets customers off AWS RDS, Google Cloud SQL, Azure Database for PostgreSQL, and self-managed clusters with minimal downtime. Drive the AI-Postgres story end to end — vector search (pgvector, pgvectorscale), hybrid retrieval, integration with the wider Nebius AI Cloud stack. Run the service like an SRE — define SLOs, build observability, lead incident response, feed every postmortem back into the platform. Work directly with customers on architecture reviews, performance escalations, and the production problems that don't fit a ticket template. Must-haves : 5+ years of professional software engineering experience, with significant time spent building or operating production PostgreSQL at scale (TB+ datasets, replication topologies, real failure modes you've debugged in production). Strong software engineering skills in Go or another backend/systems language, with a willingness to work primarily in Go. Deep knowledge of PostgreSQL internals: MVCC, WAL, replication (physical and logical), vacuum, query planning, extensions, partitioning. Hands-on experience with the surrounding ecosystem: Patroni / Stolon / pg_auto_failover, pgBackRest / WAL-G, pgBouncer / PgCat, logical replication tooling. The instincts of an ex-DBA — you can read EXPLAIN ANALYZE fluently, reason about lock behaviour, and know how to handle database corruptions. Ability to write reliable code and dig into complex problems. Teamwork-oriented approach. Nice - to - have s : Experience with pgvector and pgvectorscale — vector search at scale, index choice, recall vs. latency trade-offs. Background building managed-database control planes at a cloud provider (RDS, Cloud SQL, Aiven, Crunchy, Timescale, Supabase, Neon, or similar). Experience writing Kubernetes operators (Go + controller-runtime / kubebuilder). Contributions to PostgreSQL itself, popular extensions, or the surrounding OSS ecosystem. AI/ML workload experience — RAG pipelines, embedding stores, GPU-resident workloads. We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer (Managed PostgreSQL)

On-sitefull timeSeniorUnited Kingdom
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We're looking for a Senior Software Engineer to help us build the best Managed PostgreSQL for AI workloads. The service is already in production; the mission now is to make it the obvious choice for teams running AI applications on Postgres — first-class vector search, painless migrations from RDS / Cloud SQL / self-managed clusters, and the operational quality of a mature managed database. You'll work across the stack: control plane and lifecycle automation in Go, Postgres internals and performance, migration tooling, and the customer escalations where it actually matters. You're welcome to work from our offices in Amsterdam or London, hybrid or remotely from EU timezones. Your responsibilities : Develop the control plane and lifecycle automation for Managed PostgreSQL — provisioning, HA, failover, backups, PITR, version upgrades, zero-downtime maintenance. Tune and harden PostgreSQL itself — replication, WAL, vacuum, query planner, connection pooling, extensions — and turn that into product features and sane defaults customers don't have to think about. Build migration tooling that gets customers off AWS RDS, Google Cloud SQL, Azure Database for PostgreSQL, and self-managed clusters with minimal downtime. Drive the AI-Postgres story end to end — vector search (pgvector, pgvectorscale), hybrid retrieval, integration with the wider Nebius AI Cloud stack. Run the service like an SRE — define SLOs, build observability, lead incident response, feed every postmortem back into the platform. Work directly with customers on architecture reviews, performance escalations, and the production problems that don't fit a ticket template. Must-haves : 5+ years of professional software engineering experience, with significant time spent building or operating production PostgreSQL at scale (TB+ datasets, replication topologies, real failure modes you've debugged in production). Strong software engineering skills in Go or another backend/systems language, with a willingness to work primarily in Go. Deep knowledge of PostgreSQL internals: MVCC, WAL, replication (physical and logical), vacuum, query planning, extensions, partitioning. Hands-on experience with the surrounding ecosystem: Patroni / Stolon / pg_auto_failover, pgBackRest / WAL-G, pgBouncer / PgCat, logical replication tooling. The instincts of an ex-DBA — you can read EXPLAIN ANALYZE fluently, reason about lock behaviour, and know how to handle database corruptions. Ability to write reliable code and dig into complex problems. Teamwork-oriented approach. Nice - to - have s : Experience with pgvector and pgvectorscale — vector search at scale, index choice, recall vs. latency trade-offs. Background building managed-database control planes at a cloud provider (RDS, Cloud SQL, Aiven, Crunchy, Timescale, Supabase, Neon, or similar). Experience writing Kubernetes operators (Go + controller-runtime / kubebuilder). Contributions to PostgreSQL itself, popular extensions, or the surrounding OSS ecosystem. AI/ML workload experience — RAG pipelines, embedding stores, GPU-resident workloads. We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer, Observability

On-sitefull timeSeniorUnited States
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The Role Nebius is hiring a Senior Software Engineer to design, build, and own backend systems that power metrics, monitor large-scale infrastructure, and develop a comprehensive infrastructure maintenance platform. This role requires strong production experience, sound system design judgment, and the ability to operate and improve critical services. Your responsibilities will include: Design and build services and agents that provide deep visibility into large-scale server fleets and data center engineering systems Evolve metrics, aggregation, and alerting pipelines, with a focus on signal quality and reliability Design and operate maintenance and remediation systems that enable safe, predictable fleet-wide changes and keep infrastructure healthy Investigate production incidents hands-on, including on-host Linux debugging, and drive root-cause fixes Collaborate closely with hardware, networking, and data center operations teams to improve reliability What we expect you to have: 5+ years of professional software engineering experience Strong production experience with Python and Go, or the ability to ramp up quickly Solid Linux fundamentals and comfort debugging live systems Ability to write reliable, maintainable code and dig into complex, ambiguous problems Experience building and operating production systems at scale It will be an added bonus if you have: Ubuntu experience, including internal tooling and packaging workflows (e.g., building Debian packages) CCNA (Cisco Certified Network Associate) or equivalent networking experience Key employee benefits: Health insurance: 100% company-paid medical, dental, and vision coverage for employees and families. 401(k) plan: up to 4% company match with immediate vesting. Parental leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers. Remote work reimbursement: up to $85/month for mobile and internet. Disability & life insurance: company-paid short-term, long-term and life insurance coverage. Compensation We offer competitive salaries, ranging from $130k- $170k base + quarterly performance bonuses. Join Nebius Today! Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer (Serverless)

On-sitefull timeSeniorAmsterdam, Netherlands
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We're looking for a Senior Software Engineer to help us build Nebius Serverless AI — our GPU-native platform for deploying inference endpoints, batch jobs, and AI workloads without managing infrastructure. Customers run production AI on it today; the mission now is to make it the fastest, cheapest, most reliable place to ship GPU workloads on the planet. This is a senior, high-ownership position on a fast-evolving service. You'll set the technical bar: own the architecture of the parts you ship, drive the hard design decisions, and raise the standard of everything around you through code reviews, design reviews, and the way you operate. The work is end-to-end — from cold-start latency in the runtime to the API contract customers integrate against. You'll work from our office in Amsterdam or London, hybrid. In this position, your responsibility will be to: Design and build core components of the Serverless platform — the control plane, scheduler, runtime, autoscaler, and the APIs customers integrate against. Own the hardest engineering problems on the service: cold-start latency, GPU scheduling under contention, multi-tenant isolation, fair-share quotas, request routing at the edge. Set the technical direction for the areas you own — drive architecture decisions, write the design docs, and align the team behind the chosen approach. Raise the engineering bar across the team through code reviews, design reviews, and example — the kind of senior presence that compounds. Run the service like an SRE: define SLOs, build observability, lead incident response, and make sure every postmortem changes the platform, not just the runbook. Work directly with customers on architecture reviews, performance escalations, and the production problems that don't fit a ticket template. Partner closely with Product, GTM, and infrastructure teams to translate customer needs into a coherent technical roadmap. We expect you to have: 7+ years of professional software engineering experience, with a track record of shipping production distributed systems at scale. Excellent knowledge of Golang, or you are ready to quickly switch to it. Deep experience with Kubernetes and container orchestration — you've operated it in anger, not just deployed it. Strong distributed-systems instincts: consistency vs. availability trade-offs, queueing, backpressure, retries, idempotency, multi-tenancy. Experience designing and operating high-throughput, low-latency services — you know where the milliseconds go and how to get them back. A history of being the engineer others look to on hard problems — driving design discussions, unblocking teammates, and shipping the thing nobody else wanted to touch. Ability to write reliable code and dig into complex problems. Teamwork-oriented approach. It would be an added bonus if you had: Experience building serverless or function-as-a-service platforms (Knative, AWS Lambda, GCP Cloud Run, Cloudflare Workers, Modal, Replicate, Together, Fireworks, Anyscale, or similar). GPU scheduling experience — Kubernetes device plugins, MIG, MPS, time-slicing, NVIDIA GPU Operator. ML inference experience — vLLM, TensorRT-LLM, Triton Inference Server, SGLang, model loading and warm-pool strategies. Cold-start optimization at the runtime, image, or snapshot level (FireCracker, gVisor, checkpoint/restore, image streaming). Experience writing Kubernetes operators (Go + controller-runtime / kubebuilder). Contributions to relevant open-source projects in the serverless, scheduling, or inference ecosystems. We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer (Serverless)

On-sitefull timeSeniorLondon, United Kingdom
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We're looking for a Senior Software Engineer to help us build Nebius Serverless AI — our GPU-native platform for deploying inference endpoints, batch jobs, and AI workloads without managing infrastructure. Customers run production AI on it today; the mission now is to make it the fastest, cheapest, most reliable place to ship GPU workloads on the planet. This is a senior, high-ownership position on a fast-evolving service. You'll set the technical bar: own the architecture of the parts you ship, drive the hard design decisions, and raise the standard of everything around you through code reviews, design reviews, and the way you operate. The work is end-to-end — from cold-start latency in the runtime to the API contract customers integrate against. You'll work from our office in Amsterdam or London, hybrid. In this position, your responsibility will be to: Design and build core components of the Serverless platform — the control plane, scheduler, runtime, autoscaler, and the APIs customers integrate against. Own the hardest engineering problems on the service: cold-start latency, GPU scheduling under contention, multi-tenant isolation, fair-share quotas, request routing at the edge. Set the technical direction for the areas you own — drive architecture decisions, write the design docs, and align the team behind the chosen approach. Raise the engineering bar across the team through code reviews, design reviews, and example — the kind of senior presence that compounds. Run the service like an SRE: define SLOs, build observability, lead incident response, and make sure every postmortem changes the platform, not just the runbook. Work directly with customers on architecture reviews, performance escalations, and the production problems that don't fit a ticket template. Partner closely with Product, GTM, and infrastructure teams to translate customer needs into a coherent technical roadmap. We expect you to have: 7+ years of professional software engineering experience, with a track record of shipping production distributed systems at scale. Excellent knowledge of Golang, or you are ready to quickly switch to it. Deep experience with Kubernetes and container orchestration — you've operated it in anger, not just deployed it. Strong distributed-systems instincts: consistency vs. availability trade-offs, queueing, backpressure, retries, idempotency, multi-tenancy. Experience designing and operating high-throughput, low-latency services — you know where the milliseconds go and how to get them back. A history of being the engineer others look to on hard problems — driving design discussions, unblocking teammates, and shipping the thing nobody else wanted to touch. Ability to write reliable code and dig into complex problems. Teamwork-oriented approach. It would be an added bonus if you had: Experience building serverless or function-as-a-service platforms (Knative, AWS Lambda, GCP Cloud Run, Cloudflare Workers, Modal, Replicate, Together, Fireworks, Anyscale, or similar). GPU scheduling experience — Kubernetes device plugins, MIG, MPS, time-slicing, NVIDIA GPU Operator. ML inference experience — vLLM, TensorRT-LLM, Triton Inference Server, SGLang, model loading and warm-pool strategies. Cold-start optimization at the runtime, image, or snapshot level (FireCracker, gVisor, checkpoint/restore, image streaming). Experience writing Kubernetes operators (Go + controller-runtime / kubebuilder). Contributions to relevant open-source projects in the serverless, scheduling, or inference ecosystems. We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Staff Software Engineer (Hardware Infrastructure)

On-sitefull timeLead / StaffUnited States
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The team The Hardware Automation team builds the internal platforms and tooling that power how Nebius operates its data center infrastructure at scale. Our mission is to eliminate manual effort, reduce human error, and give every team in the Hardware Infrastructure department real-time visibility and control over the systems they own. We operate as a product engineering team embedded within hardware infrastructure — meaning we don't just write requirements and hand them off. We own the full stack: from requirements gathering with data center operations and hardware engineering, through design and build, all the way to rollout and ongoing reliability. The role Nebius is looking for a Staff Software Engineer in Hardware Infrastructure team. Hardware Infrastructure team designs, develops and supports systems involved in the data center lifecycle. Your responsibilities will include: Design and develop services that automate work with a large server fleet We expect you to have: 5+ years of professional software engineering experience Excellent knowledge of Python or Golang, or you are ready to quickly switch to these programming languages Ability to write reliable code and dig into complex problems Strong interest in being engaged in DevOps processes It will be an added bonus if you have: Solid understanding of modern server architecture and its components Good knowledge of computer networks Experience designing, developing and running high-load distributed syste Experience with cloud platforms or infrastructure-level systems Knowledge of Kubernetes, containers, or service meshes Background in performance optimization or low-latency systems Familiarity with observability tools (metrics, logging, tracing) We expect Staff Engineers to : Manage large-scale projects involving multiple stakeholders Break down complex tasks and guide both their own work and that of more junior colleagues Be experts in specific technologies and write high-quality code that can serve as a reference Assess task priority and focus on high-impact work, avoiding low-value efforts Have strong architectural thinking and contribute to system design Be involved in hiring and actively contribute to interviews Be willing to share knowledge and mentor others Key employee benefits: Health insurance: 100% company-paid medical, dental, and vision coverage for employees and families. 401(k) plan: up to 4% company match with immediate vesting. Parental leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers. Remote work reimbursement: up to $85/month for mobile and internet. Disability & life insurance: company-paid short-term, long-term and life insurance coverage. Join Nebius Today! Pay Transparency We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law. Base Compensation Range $175,000 — $225,000 USD Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. Why work at Nebius Nebius is leading a new era in cloud computing to serve the global AI economy. We create the tools and resources our customers need to solve real-world challenges and transform industries, without massive infrastructure costs or the need to build large in-house AI/ML teams. Our employees work at the cutting edge of AI cloud infrastructure alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and listed on Nasdaq, Nebius has a global footprint with R&D hubs across Europe, North America, and Israel. The team of over 1400 employees includes more than 400 highly skilled engineers with deep expertise across hardware and software engineering, as well as an in-house AI R&D team. The role We are looking for a Technical Project Manager to join our team and manage different projects and changes in data center. You will work with modern GPU clusters, responsible for execution of projects around IT infrastructure and equipment: GPU servers, high-speed network, compute and storage systems. Acting as a key link between the stakeholders, project and technical teams, you will drive operational activities across your responsibility area. Your responsibilities will include: Management and delivery of IT infrastructure projects, change requests in data centers Proactively identify and manage resolution of technical and operational issues Managing risks and dependencies in live environments Maintain clear operational documentation, including plans, risk logs, and issue trackers. Lead communication with all stakeholders: provide regular progress updates, escalate blockers, and manage expectations. Lead meetings and technical discussions with a focus on decision-making, clear next steps, and timely execution. Collaborate with cross-functional teams. Engage directly with data center engineers, vendors, and operations teams. Participate in related project tasks, including automation, process improvements, and design support as needed. We expect you to have: Proven experience in data center infrastructure operations, preferably in high-performance or AI-oriented environments. Solid understanding of data center design and operations: racking, cabling, power planning, cooling considerations. Hands on experience working with IT hardware in data centers. Be comfortable in data center environments and during live operations. Experience in project delivery, team management, cross team collaboration Strong problem-solving skills; ability to troubleshoot technical and workflow issues in an IT stack. Ability to document processes, maintain tracking systems, and produce clear deployment reports. Activity, responsibility and purposefulness Ready for occasional business trips It will be an added bonus if you have: Familiarity with structured project management or delivery frameworks (e.g., PRINCE2, PMI, Agile) Relevant technical certifications In depth technical knowledge of GPU servers, compute nodes, high-speed networking Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Cloud, DevOps & SREVia Greenhouse
Verified28 days ago

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. We give you the opportunity to work with cutting-edge technologies in data operations, cloud computing and infrastructure management. As global data center operations grow, there will be ample opportunities for career progression. Working in the data center directly impacts performance, customer satisfaction and efficiency, with the opportunity to contribute to new data center projects.You’ll collaborate with experts in AI data center development and operations, gaining insights from leaders in the field. This environment fosters innovation, and allows you to work on solutions that exceed industry standards in design and deployment. The role We are looking for a Technical Project Manager to join our team and manage different projects and changes in data center. You will work with modern GPU clusters, responsible for execution of projects around IT infrastructure and equipment: GPU servers, high-speed network, compute and storage systems. Acting as a key link between the stakeholders, project and technical teams, you will drive operational activities across your responsibility area. Your responsibilities will include: Management and delivery of IT infrastructure projects, change requests in data centers Proactively identify and manage resolution of technical and operational issues Managing risks and dependencies in live environments Maintain clear operational documentation, including plans, risk logs, and issue trackers. Lead communication with all stakeholders: provide regular progress updates, escalate blockers, and manage expectations. Lead meetings and technical discussions with a focus on decision-making, clear next steps, and timely execution. Collaborate with cross-functional teams. Engage directly with data center engineers, vendors, and operations teams. Participate in related project tasks, including automation, process improvements, and design support as needed. We expect you to have: Proven experience in data center infrastructure operations, preferably in high-performance or AI-oriented environments. Solid understanding of data center design and operations: racking, cabling, power planning, cooling considerations. Hands on experience working with IT hardware in data centers. Be comfortable in data center environments and during live operations. Experience in project delivery, team management, cross team collaboration Strong problem-solving skills; ability to troubleshoot technical and workflow issues in an IT stack. Ability to document processes, maintain tracking systems, and produce clear deployment reports. Activity, responsibility and purposefulness Ready for occasional business trips It will be an added bonus if you have: Familiarity with structured project management or delivery frameworks (e.g., PRINCE2, PMI, Agile) Relevant technical certifications In depth technical knowledge of GPU servers, compute nodes, high-speed networking Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Cloud, DevOps & SREVia Greenhouse
Verified28 days ago

Vulnerability Operation Center Lead

Remotefull timeLead / StaffWorldwide (Remote)
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. Vulnerability Operation Center The team maintains the full vulnerability management lifecycle - from detection and triage through remediation tracking and reporting - across Nebius's cloud infrastructure, product, and hardware stack. The team's focus is on improving vulnerability triaging capabilities and vulnerability data quality, driving effective remediation, and serving as the first line of response for the most critical zero-day vulnerabilities. The role We're building a Vulnerability Operations Center from the ground up and need a senior practitioner to lead it. You'll own the full vulnerability management lifecycle - from detection through triage to remediation tracking and reporting - across Nebius' cloud infrastructure, product and hardware stack. This is a high-impact role with big ownership: you'll define processes, select tooling, work directly with engineering teams, and set the standard for how Nebius responds to vulnerabilities. What You'll Do Improve and maintain automated vulnerability triaging capabilities and vulnerability data quality through data enrichment (internal and external intelligence feeds), quality metrics, AI-assisted triage, context-aware risk scoring, and vulnerability correlation/chaining. Hands-on validation and triaging of most critical/zero-days vulnerabilities Work closely with vulnerability data consumers and stakeholders across the organization to meet engineering, compliance, and regulatory requirements by delivering high-quality, actionable findings and driving effective remediation. Contribute to the development of internal platforms, including the Unified Vulnerability Management (UVM) system and the security orchestration platform, to automate core vulnerability management workflows and reduce manual effort. Identify current deficiencies in patch management, drive and oversee improvements across engineering teams and business units to improve remediation efficiency and reduce organizational risk Own vulnerability intake from all sources - scanners, bug bounty, threat intel feeds, and penetration tests Prioritize findings using risk-based frameworks (CVSS, EPSS, SSVC, asset criticality, exploitability context, business impact) and reduce false positive noise Drive remediation accountability across infrastructure, platform, and product engineering teams Identify, track and report vulnerability management metrics, present KPIs and trends to security leadership Coordinate response to critical/zero-day vulnerabilities, acting as the primary point of contact across security, engineering, and operations Define and maintain integration between VOC tooling and the broader security ecosystem. What We're Looking For 5–8 years in information security with at least 3 years focused on vulnerability management or security operations Deep familiarity with vulnerability scanning tools and their strengths/limitations in cloud-native environments Strong grasp of CVE/NVD, CVSS scoring, EPSS, SSVC and how to apply them to real-world prioritization Experience managing vulnerabilities across IaaS/cloud infrastructure (AWS, GCP, Azure, or private cloud) - experience with GPU/HPC environments is a plus Deep understanding of vulnerability sources limitations and corner cases for different vulnerability classes Able to write and review code (ability or willingness to develop in Golang) - well enough to assess exploitability, validate fixes, and build lightweight automation (e.g., variant-detection scripts, triage tooling, data pipelines for trend analysis) Strong grasp of common vulnerability classes at the code and infrastructure level. Comfortable feeding well-documented vulnerability patterns into AI-assisted code review workflows and critically evaluating the output Track record of working cross-functionally with engineering teams and holding stakeholders accountable to timelines without being a bottleneck Strong written and verbal communication - able to translate technical risk into business impact for non-security audiences Nice to Have Experience building vulnerability management program from scratch Familiarity with container and Kubernetes security Experience with supply chain security (SBOMs, dependency scanning) Bug bounty triage experience Public talks or research articles Why this role at Nebius Build a Platform Security Vulnerability program from scratch and get to build and lead your own team while doing it The potential to evolve an in-house AI-powered vulnerability management platform into a cloud security offering for Nebius customers Work alongside world-class engineers on infrastructure that powers frontier AI. Competitive compensation with equity upside in a Nasdaq-listed, high-growth company. Flexible, remote-first culture #LI-CP1 Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
CybersecurityVia Greenhouse
Verified28 days ago

Page 12 of 18