LogoKode$word
Nebius Group logo
Verified Tech Organization

Careers at Nebius Group

Browse and filter through all verified positions currently open at Nebius Group.

Total Company Roles175
Matching Filter175
nebius.com/companyHQ: Schiphol, NLCEO: Arkady Volozh1543 employees

Nebius Group N.V. is a technology company dedicated to developing comprehensive infrastructure to serve the global artificial intelligence industry. Its operations encompass several key areas. Central to its mission is Nebius, an AI-focused cloud platform engineered to handle demanding AI workloads. This division constructs end-to-end AI infrastructure, featuring extensive GPU computing clusters, robust cloud platforms, and essential tools and services for developers. The group also includes Toloka AI, which functions as a data solutions provider, assisting with various phases of generative AI development. TripleTen operates as an educational technology venture, focused on equipping individuals with new skills for careers in the tech sector. Furthermore, Avride specializes in pioneering autonomous driving technologies for self-driving vehicles and delivery robots. Founded in 1989, the company was previously known as Yandex N.V. until its rebranding to Nebius Group N.V. in August 2024. Its headquarters are located in Amsterdam, the Netherlands, with additional research and development facilities spread across Europe, North America, and Israel.

Sector:Software Application

All Openings (175)

Ordered by most recently published

Senior Software Developer: Models Team (Token Factory)

On-sitefull timeSeniorAmsterdam, Netherlands
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About the Product Token Factory is focused on building a next-generation platform that enables companies to seamlessly integrate AI into their products and workflows. Our vision is to create a powerful, open, and scalable alternative for deploying and managing AI systems—making advanced AI infrastructure more accessible to both fast-growing startups and large enterprises. We work with a wide range of customers, from AI-first companies to established technology organisations, helping them run AI workloads reliably at scale. Our goal is to become a leading platform for high-performance AI inference, delivering predictable latency, strong reliability, and the ability to scale to meet demanding production needs. Customer feedback plays a central role in how we build—our development process is highly iterative and closely aligned with real-world use cases. About the Team The Models Team is responsible for onboarding state-of-the-art (SOTA) open-source models into Nebius TokenFactory, including models such as DeepSeek V4 Pro, GLM 5.1, Kimi K2.6, and Minimax M2.7. A major focus of the team is serving large-scale AI models efficiently and reliably in production. To achieve this, we work on advanced inference and systems optimization techniques, including: Cache-aware routing NUMA-aware deployments KV-cache offloading Disaggregated serving architectures Autoscaling with high-speed model loading over InfiniBand / RoCE The team maintains and extends forks of leading inference frameworks such as vLLM and TRT-LLM . We have deep expertise in production-scale model serving and regularly support the Solutions Architects team on the most demanding customer PoCs. To operate efficiently at scale, we invest heavily in tooling and automation. Examples include: Performance, quality, and smoke-testing frameworks Hyperparameter optimization for inference framework configurations Gibberish detection systems Automated rollout pipelines for inference framework upgrades Diagnostics and observability tooling Traffic replay systems Automated search for optimal serverless deployment configurations We collaborate closely with model builders, open-source communities, Nebius Cloud teams, and hardware vendors to continuously improve our serving infrastructure. The team is highly goal-oriented and outcome-driven, with a strong focus on delivering results rather than following rigid processes. Team Structure We are currently a team of eight engineers distributed across Europe, with members based in the Netherlands, the United Kingdom, Germany, and Latvia. Our workflows are optimized for remote collaboration. At the same time, we meet in person every one to two months at one of our locations to work together, brainstorm new ideas, and plan upcoming milestones. While many team members joined without extensive AI/ML experience, we have rapidly developed strong expertise in large-scale model serving and AI infrastructure. Technology Our work is deeply integrated with the broader cloud and infrastructure ecosystem. We primarily use Go and Python to build and scale backend systems. We collaborate closely with teams working on cloud infrastructure, observability, reliability, fault tolerance, and platform engineering. The challenges we solve sit at the intersection of distributed systems, high-performance computing, and modern AI infrastructure. We expect you to have: Experience serving LLMs in production Strong Python and/or Go programming skills Experience designing and operating highly scalable, highly available distributed services Nice to have: Contributions to vLLM, SGLang, TRT-LLM, or NVIDIA ecosystem open-source projects Deep understanding of KV cache management, speculative decoding, and quantization Experience with LLM evaluation frameworks Hands-on experience with performance benchmarking and optimization Deep understanding of Kubernetes Familiarity with distributed serving architectures and autoscaling Knowledge of InfiniBand, RoCE, or high-performance networking Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Developer: Models Team (Token Factory)

On-sitefull timeSeniorLondon, United Kingdom
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About the Product Token Factory is focused on building a next-generation platform that enables companies to seamlessly integrate AI into their products and workflows. Our vision is to create a powerful, open, and scalable alternative for deploying and managing AI systems—making advanced AI infrastructure more accessible to both fast-growing startups and large enterprises. We work with a wide range of customers, from AI-first companies to established technology organisations, helping them run AI workloads reliably at scale. Our goal is to become a leading platform for high-performance AI inference, delivering predictable latency, strong reliability, and the ability to scale to meet demanding production needs. Customer feedback plays a central role in how we build—our development process is highly iterative and closely aligned with real-world use cases. About the Team The Models Team is responsible for onboarding state-of-the-art (SOTA) open-source models into Nebius TokenFactory, including models such as DeepSeek V4 Pro, GLM 5.1, Kimi K2.6, and Minimax M2.7. A major focus of the team is serving large-scale AI models efficiently and reliably in production. To achieve this, we work on advanced inference and systems optimization techniques, including: Cache-aware routing NUMA-aware deployments KV-cache offloading Disaggregated serving architectures Autoscaling with high-speed model loading over InfiniBand / RoCE The team maintains and extends forks of leading inference frameworks such as vLLM and TRT-LLM . We have deep expertise in production-scale model serving and regularly support the Solutions Architects team on the most demanding customer PoCs. To operate efficiently at scale, we invest heavily in tooling and automation. Examples include: Performance, quality, and smoke-testing frameworks Hyperparameter optimization for inference framework configurations Gibberish detection systems Automated rollout pipelines for inference framework upgrades Diagnostics and observability tooling Traffic replay systems Automated search for optimal serverless deployment configurations We collaborate closely with model builders, open-source communities, Nebius Cloud teams, and hardware vendors to continuously improve our serving infrastructure. The team is highly goal-oriented and outcome-driven, with a strong focus on delivering results rather than following rigid processes. Team Structure We are currently a team of eight engineers distributed across Europe, with members based in the Netherlands, the United Kingdom, Germany, and Latvia. Our workflows are optimized for remote collaboration. At the same time, we meet in person every one to two months at one of our locations to work together, brainstorm new ideas, and plan upcoming milestones. While many team members joined without extensive AI/ML experience, we have rapidly developed strong expertise in large-scale model serving and AI infrastructure. Technology Our work is deeply integrated with the broader cloud and infrastructure ecosystem. We primarily use Go and Python to build and scale backend systems. We collaborate closely with teams working on cloud infrastructure, observability, reliability, fault tolerance, and platform engineering. The challenges we solve sit at the intersection of distributed systems, high-performance computing, and modern AI infrastructure. We expect you to have: Experience serving LLMs in production Strong Python and/or Go programming skills Experience designing and operating highly scalable, highly available distributed services Nice to have: Contributions to vLLM, SGLang, TRT-LLM, or NVIDIA ecosystem open-source projects Deep understanding of KV cache management, speculative decoding, and quantization Experience with LLM evaluation frameworks Hands-on experience with performance benchmarking and optimization Deep understanding of Kubernetes Familiarity with distributed serving architectures and autoscaling Knowledge of InfiniBand, RoCE, or high-performance networking Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Backend Engineer

On-sitefull timeSeniorPrague, Czech Republic
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We’re looking for a Senior Backend Software Engineer to help develop our hyperscaling platform. We have a lot of tasks on Golang and little less on Java and Python. You’re welcome to work from our office in Prague, Berlin, London or in Amsterdam, hybrid or remotely. In this position, your responsibility will be to: Developing fault-tolerant, reliable cloud services: cloud infrastructure, managed databases, managed Kubernetes, billing, monitoring systems, ML service and more. We expect you to have: 5+ years of professional software engineering experience Excellent knowledge of Golang or you are ready to quickly switch to this programming language Ability to write reliable code and dig into complex problems Teamwork-oriented approach It would be an added bonus if you had: Experience designing high-load and scaling services We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About the Product Token Factory is focused on building a next-generation platform that enables companies to seamlessly integrate AI into their products and workflows. Our vision is to create a powerful, open, and scalable alternative for deploying and managing AI systems—making advanced AI infrastructure more accessible to both fast-growing startups and large enterprises. We work with a wide range of customers, from AI-first companies to established technology organisations, helping them run AI workloads reliably at scale. Our goal is to become a leading platform for high-performance AI inference, delivering predictable latency, strong reliability, and the ability to scale to meet demanding production needs. Customer feedback plays a central role in how we build—our development process is highly iterative and closely aligned with real-world use cases. About the Team The Models Team is responsible for onboarding state-of-the-art (SOTA) open-source models into Nebius TokenFactory, including models such as DeepSeek V4 Pro, GLM 5.1, Kimi K2.6, and Minimax M2.7. A major focus of the team is serving large-scale AI models efficiently and reliably in production. To achieve this, we work on advanced inference and systems optimization techniques, including: Cache-aware routing NUMA-aware deployments KV-cache offloading Disaggregated serving architectures Autoscaling with high-speed model loading over InfiniBand / RoCE The team maintains and extends forks of leading inference frameworks such as vLLM and TRT-LLM . We have deep expertise in production-scale model serving and regularly support the Solutions Architects team on the most demanding customer PoCs. To operate efficiently at scale, we invest heavily in tooling and automation. Examples include: Performance, quality, and smoke-testing frameworks Hyperparameter optimization for inference framework configurations Gibberish detection systems Automated rollout pipelines for inference framework upgrades Diagnostics and observability tooling Traffic replay systems Automated search for optimal serverless deployment configurations We collaborate closely with model builders, open-source communities, Nebius Cloud teams, and hardware vendors to continuously improve our serving infrastructure. The team is highly goal-oriented and outcome-driven, with a strong focus on delivering results rather than following rigid processes. Team Structure We are currently a team of eight engineers distributed across Europe, with members based in the Netherlands, the United Kingdom, Germany, and Latvia. Our workflows are optimized for remote collaboration. At the same time, we meet in person every one to two months at one of our locations to work together, brainstorm new ideas, and plan upcoming milestones. While many team members joined without extensive AI/ML experience, we have rapidly developed strong expertise in large-scale model serving and AI infrastructure. Technology Our work is deeply integrated with the broader cloud and infrastructure ecosystem. We primarily use Go and Python to build and scale backend systems. We collaborate closely with teams working on cloud infrastructure, observability, reliability, fault tolerance, and platform engineering. The challenges we solve sit at the intersection of distributed systems, high-performance computing, and modern AI infrastructure. We expect you to have: Experience serving LLMs in production Strong Python and/or Go programming skills Experience designing and operating highly scalable, highly available distributed services Nice to have: Contributions to vLLM, SGLang, TRT-LLM, or NVIDIA ecosystem open-source projects Deep understanding of KV cache management, speculative decoding, and quantization Experience with LLM evaluation frameworks Hands-on experience with performance benchmarking and optimization Deep understanding of Kubernetes Familiarity with distributed serving architectures and autoscaling Knowledge of InfiniBand, RoCE, or high-performance networking Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Developer: Models Team (Token Factory)

Remotefull timeSeniorUnited States (Remote)
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About the Product Token Factory is focused on building a next-generation platform that enables companies to seamlessly integrate AI into their products and workflows. Our vision is to create a powerful, open, and scalable alternative for deploying and managing AI systems—making advanced AI infrastructure more accessible to both fast-growing startups and large enterprises. We work with a wide range of customers, from AI-first companies to established technology organisations, helping them run AI workloads reliably at scale. Our goal is to become a leading platform for high-performance AI inference, delivering predictable latency, strong reliability, and the ability to scale to meet demanding production needs. Customer feedback plays a central role in how we build—our development process is highly iterative and closely aligned with real-world use cases. About the Team The Models Team is responsible for onboarding state-of-the-art (SOTA) open-source models into Nebius TokenFactory, including models such as DeepSeek V4 Pro, GLM 5.1, Kimi K2.6, and Minimax M2.7. A major focus of the team is serving large-scale AI models efficiently and reliably in production. To achieve this, we work on advanced inference and systems optimization techniques, including: Cache-aware routing NUMA-aware deployments KV-cache offloading Disaggregated serving architectures Autoscaling with high-speed model loading over InfiniBand / RoCE The team maintains and extends forks of leading inference frameworks such as vLLM and TRT-LLM . We have deep expertise in production-scale model serving and regularly support the Solutions Architects team on the most demanding customer PoCs. To operate efficiently at scale, we invest heavily in tooling and automation. Examples include: Performance, quality, and smoke-testing frameworks Hyperparameter optimization for inference framework configurations Gibberish detection systems Automated rollout pipelines for inference framework upgrades Diagnostics and observability tooling Traffic replay systems Automated search for optimal serverless deployment configurations We collaborate closely with model builders, open-source communities, Nebius Cloud teams, and hardware vendors to continuously improve our serving infrastructure. The team is highly goal-oriented and outcome-driven, with a strong focus on delivering results rather than following rigid processes. Team Structure We are currently a team of eight engineers distributed across Europe, with members based in the Netherlands, the United Kingdom, Germany, and Latvia. Our workflows are optimized for remote collaboration. At the same time, we meet in person every one to two months at one of our locations to work together, brainstorm new ideas, and plan upcoming milestones. While many team members joined without extensive AI/ML experience, we have rapidly developed strong expertise in large-scale model serving and AI infrastructure. Technology Our work is deeply integrated with the broader cloud and infrastructure ecosystem. We primarily use Go and Python to build and scale backend systems. We collaborate closely with teams working on cloud infrastructure, observability, reliability, fault tolerance, and platform engineering. The challenges we solve sit at the intersection of distributed systems, high-performance computing, and modern AI infrastructure. We expect you to have: Experience serving LLMs in production Strong Python and/or Go programming skills Experience designing and operating highly scalable, highly available distributed services Nice to have: Contributions to vLLM, SGLang, TRT-LLM, or NVIDIA ecosystem open-source projects Deep understanding of KV cache management, speculative decoding, and quantization Experience with LLM evaluation frameworks Hands-on experience with performance benchmarking and optimization Deep understanding of Kubernetes Familiarity with distributed serving architectures and autoscaling Knowledge of InfiniBand, RoCE, or high-performance networking Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer (Agentic Search) - Billing

On-sitefull timeSeniorNew York, United States
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About Tavily We're building the infrastructure layer for agentic web interaction at scale. Our API is designed from the ground up to power Retrieval-Augmented Generation (RAG) and real-time reasoning in AI systems. By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that are not only intelligent — but also informed. We work with some of the most innovative teams in AI — from small startups shaping the ecosystem to the largest enterprises deploying AI at scale. Whether it's powering sales assistants, research copilots, or internal knowledge tools, we're the missing link between LLMs and the real world. The Role We are looking for a Senior Software Engineer to design and build the billing platform behind a novel search engine tailored for agentic AI consumption. This role is based in our New York office, hybrid with the NYC team. Our APIs are used by millions of developers and the AI agents acting on their behalf, all paying based on what they consume. You will own the systems that turn product usage into accurate, trustworthy revenue, supporting flexible billing models—usage-based pricing, subscriptions, prepaid balances with auto top-up, and enterprise contracts. Correctness is non-negotiable: every event must be measured precisely and every invoice must be right to the cent. ‌ In this position, your responsibility will be to: Design and operate the billing platform end to end, from usage events to invoices Build metering and usage-aggregation pipelines that turn high-volume events into billable amounts Design and operate prepaid credit/wallet systems with auto top-up and balance management Support enterprise/contract billing with custom pricing, negotiated terms, and scheduled invoicing Build reconciliation, idempotency, and invoice-accuracy safeguards to ensure revenue data is always correct Lead the migration to the next generation of our billing system with no revenue loss or downtime Define observability and quality metrics for billing correctness, freshness, and throughput Ensure billing systems meet audit and compliance requirements ( SOX , financial audit) ‌ What we're looking for : 6+ years building production backend systems, of which 3+ years actively building and operating billing or metering systems in production Hands-on experience with usage-based / metered billing: rating, aggregation, and proration logic Hands-on experience with subscription / recurring billing alongside usage-based models Hands-on experience with prepaid credits / wallet systems and auto top-up mechanics Hands-on experience with enterprise / contract billing: custom pricing and negotiated terms End-to-end integration of a third-party billing platform (Stripe Billing, Metronome, Orb, or Zuora) Strong expertise in Python and TypeScript/Node (Go a plus) Production experience with transactional databases (Postgres or equivalent) for financial-grade data Built high-volume event / usage pipelines into a data warehouse (Kafka or Kinesis → Snowflake or BigQuery) Owned billing correctness in production: reconciliation, idempotency, invoice accuracy Product-Led Growth (PLG) experience: worked on billing for high-volume, self-serve users Hands-on with Auth0, Okta, or similar identity systems in production Shipped revenue systems under SOX or financial audit constraints Nice to have: Led a billing system migration or re-platform with zero revenue loss or downtime Built hybrid billing combining usage, subscriptions, and prepaid in one system Worked across both a payments layer (Stripe) and a metering / rating layer (Metronome / Orb) Billed a developer-facing / API product with pay-as-you-go and API-key metering Used Snowflake for usage aggregation and revenue analytics Why Tavily ? Full ownership — small team, you own the entire infrastructure, not a slice of it Real scaling challenges — bursty scraping workloads, cache invalidation, multi-region, millions of daily requests AI-native company — your infra directly powers AI agents used by leading companies in the space. Key employee benefits in the US: Health insurance: 100% company-paid medical, dental, and vision coverage for employees and families. 401(k) plan: Up to 4% company match with immediate vesting. Parental leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers. Remote work reimbursement: Up to $85/month for mobile and internet. Disability & life insurance : Company-paid short-term, long-term and life insurance coverage. Pay Transparency We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law. Base Compensation Range $147,200 — $224,300 USD Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer (Capacity and Quota Management)

On-sitefull timeSeniorAmsterdam, Netherlands
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role: We are looking for a Senior Software Engineer to join the team building the Capacity and Quota Management control plane — the system that decides who gets compute, when, and how much. Idle GPUs are waste. Overallocated ones are a contract you can't keep. Our platform consists of two core components: a capacity reservation engine managing guaranteed GPU reservations and their lifecycle, including rebalancing and auto-reclamation logic; and a Quotas Control Plane that manages quotas across all Nebius services, handles customer quota requests, and runs the algorithms that distribute resources. Both are exposed through APIs that give external customers, internal users, and downstream services visibility and control over their limits. This is a high-ownership role. You will define how Capacity and Quota Management matures, from the technical foundations to the tradeoffs that outlast any single feature. We expect you to have: 5+ years of professional software engineering experience Strong knowledge of Python or willingness to quickly become productive with our technology stack Experience building distributed backend systems and microservice architectures Understanding of distributed transactions, consistency models, and reliable asynchronous workflows Strong communication and ownership mindset Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role: We are looking for a Senior Software Engineer to join the team building the Capacity and Quota Management control plane — the system that decides who gets compute, when, and how much. Idle GPUs are waste. Overallocated ones are a contract you can't keep. Our platform consists of two core components: a capacity reservation engine managing guaranteed GPU reservations and their lifecycle, including rebalancing and auto-reclamation logic; and a Quotas Control Plane that manages quotas across all Nebius services, handles customer quota requests, and runs the algorithms that distribute resources. Both are exposed through APIs that give external customers, internal users, and downstream services visibility and control over their limits. This is a high-ownership role. You will define how Capacity and Quota Management matures, from the technical foundations to the tradeoffs that outlast any single feature. We expect you to have: 5+ years of professional software engineering experience Strong knowledge of Python or willingness to quickly become productive with our technology stack Experience building distributed backend systems and microservice architectures Understanding of distributed transactions, consistency models, and reliable asynchronous workflows Strong communication and ownership mindset Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role: We are looking for a Senior Software Engineer to join the team building the Capacity and Quota Management control plane — the system that decides who gets compute, when, and how much. Idle GPUs are waste. Overallocated ones are a contract you can't keep. Our platform consists of two core components: a capacity reservation engine managing guaranteed GPU reservations and their lifecycle, including rebalancing and auto-reclamation logic; and a Quotas Control Plane that manages quotas across all Nebius services, handles customer quota requests, and runs the algorithms that distribute resources. Both are exposed through APIs that give external customers, internal users, and downstream services visibility and control over their limits. This is a high-ownership role. You will define how Capacity and Quota Management matures, from the technical foundations to the tradeoffs that outlast any single feature. We expect you to have: 5+ years of professional software engineering experience Strong knowledge of Python or willingness to quickly become productive with our technology stack Experience building distributed backend systems and microservice architectures Understanding of distributed transactions, consistency models, and reliable asynchronous workflows Strong communication and ownership mindset Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Senior Software Engineer (Data Platform, C++)

On-sitefull timeSeniorLondon, United Kingdom
Apply Now

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role #LI-MM1 We’re looking for a Software Engineer with strong C++ expertise to join the team building and operating Nebius Data Platform — a distributed storage and a processing platform that acts as the company’s “source of truth” and the backbone of many internal (and some external) products. Nebius Data Platform is a single multi-tenant ecosystem based on YTsaurus — instead of running separate HDFS/Kafka/HBase-style systems, we provide storage, compute, and analytics capabilities inside one platform. Built on top of the open-source YTsaurus ecosystem, we run and extend our own Nebius distribution and develop significant in-house functionality (core and platform-level). We can design, implement, and roll out features end-to-end on our clusters without waiting for upstream approvals and contribute upstream when it makes sense. At scale today, this includes ~500 servers, ~20k CPU cores and ~10 PB of compressed data in our largest production cluster, supporting workloads ranging from business-critical pipelines and financial transactions to large-scale ML/LLM training datasets and compute. What’s inside the platform You’ll work on a system that includes (and ties together): Distributed Storage (Cypress) : transactional semantics, tiered storage, erasure coding, replication, and strong reliability expectations. Compute & ETL : a cluster-wide job scheduler (tens of thousands of cores), MapReduce, YQL for SQL-like data processing, and SPYT (Spark over YTsaurus) for modern data engineering. Interactive analytics (CHYT) : ClickHouse® instances spun up directly on compute nodes for fast SQL over data in-place. Dynamic Tables : low-latency NoSQL KV with distributed ACID transactions for OLTP-style workloads and feature stores. Orchestracto : workflow orchestration deeply integrated with the platform (Airflow-like, but platform-native). What you’ll do We’re looking for engineers who combine strong systems skills with product sense : understanding who uses the platform, why certain capabilities matter, and making pragmatic trade-offs to maximize impact. On our team, engineering work is expected to be connected to real users and outcomes — you’ll regularly align with internal stakeholders, clarify requirements, and help drive prioritization. In this role, you will: Design and implement new functionality in YTsaurus core (C++) with production reliability in mind. Build and evolve platform-level capabilities: platform architecture and operating model—multi-cluster growth, shared primitives, and a consistent experience that scales with new teams and use cases. Improve end-to-end platform experience for internal (and external-facing) users: APIs, guardrails, debugging workflows, and automation. Own production quality: incident response / on-call rotation , root cause analysis, and turning learnings into durable fixes. Example projects Roll out sharded YTsaurus masters (incl. Kubernetes operator support) and build automatic balancing of metadata across master cells (consensus groups) to remove control-plane bottlenecks and unlock 10–100x cluster growth . Make CHYT interactive SQL faster and more predictable at high load via performance work like data-skipping / min-max-style indexes and improved execution introspection. Turn Orchestracto into a platform product by defining the building blocks, developer experience, and governance for how teams create and share workflows. Scale and harden Parquet-on-S3 for native YTsaurus workloads by tackling replication/movement, consistent lifecycle semantics, and master-server metadata optimizations for performance and reliability. Design and ship complete, trustworthy audit trails for data changes (who/what/when) across heterogeneous storage and compute paths. Tech stack Core: modern C++ (C++20, async + multithreaded primitives) Services & tooling: Go and Python (microservices, utilities, integration tests) What we expect 5+ years of software engineering experience. Strong C++ skills (you’ll write core code). Working knowledge of Python and/or Go (you don’t have to be expert, but should be comfortable navigating them). Experience developing and/or operating high-load, distributed services . Production mindset: ability to use SSH, read logs/metrics/traces , and debug distributed systems behavior. Solid CS fundamentals: algorithms, data structures, concurrency basics. Nice to have Experience with Big Data systems (YTsaurus/Hadoop/Spark/ClickHouse/Kafka-like ecosystems). Experience with multi-tenant platforms, schedulers, resource isolation, quotas, and reliability engineering. Strong performance engineering skills (profiling, lock contention, latency/throughput tradeoffs). We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

View more...
Software EngineeringVia Greenhouse
Verified28 days ago

Page 11 of 18