Verified Tech Jobs & Hiring Companies, Updated Every 24 Hours
Direct career links to high-growth tech startups and Fortune 500 engineering teams across the United States, Europe, and Worldwide. We audit careers daily to ensure zero ghost listings and zero expired apply links.
All Verified Employers (626)
Filtered and verified against live career portals
The Verified Direct-Apply Tech Job Board
Landing a high-compensation software engineering, data, AI, or product role should not require fighting through zombie job posts, recruiter agency reposts, or expired links. Kodesword indexes verified tech career openings by connecting directly with official corporate career portals. Every single role featured on this platform is active and routes straight to the hiring company's career page.
Popular Tech Roles
Top Tech Hubs
Why Tech Candidates Use Kodesword vs. Traditional Aggregators
- 100% Direct Corporate Links: Zero middleman recruiter reposts.
- Continuous 24h Pruning: Expired and filled listings removed daily.
- Comprehensive Salary Data: Compensation extracted from verified JDs.
- Zero Paywalls or Registration: Browse and apply completely free.
Frequently Asked Questions
- How often are tech job openings updated on Kodesword?
- Our systems sync directly with official company career portals every 24 hours. Expired or filled roles are pruned daily to prevent ghost job listings.
- Are these direct job applications or recruiter agency reposts?
- Every role links directly to the official corporate careers portal. There are zero intermediary recruiters, no paywalls, and no sponsored spam.
- What kinds of tech roles are listed on Kodesword?
- We index white-collar software engineering, AI/Machine Learning, DevOps, SRE, Cloud Infrastructure, Data Engineering, Cyber Security, and Technical Product Management roles across US hubs and remote companies.
Nebius Group
Actively Hiring179 open positions matching criteria
AI Full-Stack Developer
Office of CoS
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. Summary: We are seeking an AI Full-Stack Developer to build and scale automation solutions across internal company processes using AI, LLMs, and agent-based systems. This role focuses on delivering practical automation solutions, integrating AI capabilities with internal tools, and orchestrating workflows across multiple systems, with a dynamic range of tasks that offer both quick wins and complex challenges. Responsibilities: Build AI-driven automation solutions using LLMs, APIs, and agent frameworks. Develop systems to automate document processing, data extraction, and operational workflows. Design and implement multi-step automated workflows connecting AI models and internal services. Integrate automation solutions with internal systems such as HR, finance, legal, dashboards, and ticketing tools. Deliver and iterate practical automation solutions swiftly, from simple GPT/AI integrations to complex workflows. Train and support internal teams on newly built or optimized systems for strong adoption and self-sufficiency. Collaborate with Technical Team Lead, Product Manager, and IT teams on architecture and system integration. Communicate clearly with internal stakeholders and external partners or vendors. Required Skills: 5+ years of experience working as a Full-Stack Developer Strong programming skills in Python, TypeScript, JavaScript, and React. Experience building backend services and API integrations. Practical experience with LLM APIs or AI-based services. Strong familiarity with AI developer ecosystems. Proven ability to use AI-assisted coding tools effectively. Ability to work independently and efficiently. Familiarity with workflow automation systems or distributed architectures. Preferred Qualifications: Experience building AI agents or multi-step AI workflows. Experience with prompt engineering or LLM orchestration frameworks. Experience integrating with enterprise systems or internal business tools. Familiarity with event-driven architectures, queues, and task orchestration systems. Experience in teaching non-technical users how to utilize developed systems effectively. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Application Integration Developer
IT Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. Your responsibilities will include: Develop and maintain Python-based API integrations between our HR information system (HRIS) and business systems, including finance, travel, expense management, and recruiting platforms. Build and maintain cloud-native services, web applications, and containerized workloads on Microsoft Azure. Implement automation for system workflows, onboarding/offboarding, access-related processes, and data synchronization. Create and maintain deployment pipelines and Infrastructure-as-Code with Terraform. Own integrations through implementation, testing, deployment, and production support; respond to alerts and troubleshoot issues across APIs, Azure services, and data sources within agreed SLAs. Implement safeguards for sensitive employee data, including validation, audit trails, and controls around mass changes. Work with business stakeholders to clarify requirements and maintain technical documentation as part of each change. We expect you to have: 3+ years of integration or backend development experience, including hands-on experience deploying and supporting production workloads in Microsoft Azure. Practical experience with Azure compute and integration services. Our stack includes Azure Functions, App Service, Container Apps, Container Instances, Service Bus, and API Management. Practical Python development skills, including maintaining and debugging existing services and automations. Experience integrating REST APIs, webhooks, and Microsoft Graph API, with an understanding of pagination, rate limits, retries, and idempotency. Experience working with CI/CD pipelines and Terraform for Azure deployments. Good understanding of OAuth 2.0, Microsoft Entra ID, and managed identities for secure authentication and service-to-service access. Experience designing and working with relational databases (PostgreSQL, MSSQL, etc.). Experience with automated testing and diagnosing production issues through logs, monitoring, and alerts. It will be an added bonus if you have: Experience integrating HRIS or other business systems, such as ERP, expense management, or recruiting platforms. Azure certifications (AZ-204, AZ-900, AZ-104). Hands-on experience with Application Insights and Log Analytics. Experience using AI-assisted development tools, with the ability to critically evaluate their output. #LI-RK1 Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Critical Infrastructure Engineer
Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. Why work at Nebius Nebius is leading a new era in cloud computing to serve the global AI economy. We create the tools and resources our customers need to solve real-world challenges and transform industries, without massive infrastructure costs or the need to build large in-house AI/ML teams. Our employees work at the cutting edge of AI cloud infrastructure alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and listed on Nasdaq, Nebius has a global footprint with R&D hubs across Europe, North America, and Israel. Our teams bring together deep expertise across hardware, software, networking, data center infrastructure, and AI to build and operate the infrastructure behind large-scale GPU computing. The team You will join our Data Center Infrastructure organization, supporting the critical environments that power Nebius GPU clusters and AI cloud infrastructure. Our team works across the boundary between traditional IT infrastructure and the electrical, mechanical, and cooling systems that keep high-density compute environments online. We partner closely with Data Center IT, Network Engineering, infrastructure providers, colocation partners, and internal leadership to ensure our facilities deliver the capacity, resilience, and operational performance required by our customers. This is an opportunity to develop broad expertise across both IT and critical infrastructure while helping establish the operational standards that support Nebius as our North American data center footprint continues to scale. The role We are seeking a Critical Infrastructure Engineer to help ensure the availability, resilience, and operational readiness of the critical systems supporting Nebius data center IT infrastructure. The primary objective of this role is uptime . You will provide technical oversight across the electrical and mechanical infrastructure responsible for delivering reliable power and cooling to our GPU and IT environments. Rather than serving primarily as a maintenance technician, you will verify that critical infrastructure is operated safely, consistently, and in accordance with established SLAs, engineering standards, change-control procedures, and operational best practices. You will also act as an important bridge between IT infrastructure teams and electrical/mechanical specialists. The ideal candidate understands how servers, networking equipment, racks, and GPU systems operate inside a data center while also having enough exposure to critical facilities systems to understand—and challenge when necessary—the infrastructure supporting them. The position combines technical analysis, provider governance, change management, incident response, and hands-on familiarity with data center IT environments. Your responsibilities will include: Critical Infrastructure & Uptime Help ensure the availability and operational readiness of the electrical and mechanical infrastructure supporting production data halls and high-density GPU environments. Monitor critical infrastructure performance against contractual SLAs, operational requirements, and established reliability standards. Develop a strong understanding of the complete power and cooling path supporting IT equipment and identify conditions that could introduce operational risk. Review infrastructure capacity, redundancy, and operating conditions to ensure the environment can reliably support current and planned compute deployments. Identify infrastructure risks and work with service providers and internal teams to drive corrective actions before they impact production. Support infrastructure planning for data center expansions, capacity increases, and new GPU deployments. Power & Electrical Infrastructure Provide technical oversight of data center electrical infrastructure, including generator plants, automatic transfer switches (ATS), UPS systems, battery banks, switchgear, breakers, busbars, bus plugs, PDUs, and related power distribution equipment. Understand electrical distribution from facility-level infrastructure through rack-level delivery and IT equipment. Participate in technical reviews involving power capacity, electrical distribution, equipment sizing, redundancy, and infrastructure design. Work with electrical engineers and infrastructure providers to evaluate proposed changes and ensure appropriate engineering validation is completed before production implementation. Cooling & Mechanical Infrastructure Understand the cooling architecture supporting high-density GPU and IT environments, including water and glycol loops, rear-door heat exchangers (RDHx), evaporative systems, coolant distribution systems, facility water systems, dry coolers, and chillers. Evaluate how cooling infrastructure interacts with GPU systems and high-density racks to maintain required operating conditions. Partner with mechanical engineers and service providers to review system performance, capacity constraints, and proposed infrastructure changes. Identify potential thermal or cooling risks that could affect compute availability or future capacity. Provider Governance & Change Control Provide technical oversight of third-party critical infrastructure and colocation service providers. Ensure provider activities comply with Nebius policies, approved procedures, contractual SLAs, and operational requirements. Review and approve change requests involving critical infrastructure supporting production environments. Challenge incomplete or high-risk work plans and ensure appropriate testing, rollback procedures, risk analysis, and stakeholder communication are in place before work begins. Maintain strong governance around maintenance and infrastructure changes that could affect production availability. Hold service providers accountable for corrective actions, operational performance, and agreed service levels. Incident Response & Operational Risk Participate in critical infrastructure incidents and coordinate technical response with providers, Data Center IT, networking, and engineering teams. Support root-cause analysis following power, cooling, or infrastructure-related incidents. Review incident findings and ensure corrective and preventive actions are documented, assigned, and completed. Help develop and continuously improve emergency response procedures, escalation paths, change-control standards, and operational documentation. Identify recurring infrastructure risks and drive improvements that increase reliability and reduce the likelihood of customer impact. IT & Critical Infrastructure Integration Work closely with Data Center IT teams to understand how critical infrastructure conditions affect servers, networking equipment, GPU clusters, and other production systems. Apply practical knowledge of data center IT operations, including racks, servers, fiber, cabling, network equipment, and hardware deployment. Support cross-functional troubleshooting where the root cause may span IT equipment and facility infrastructure. Help create stronger operational alignment between IT infrastructure and electrical/mechanical teams. Reporting & Stakeholder Communication Translate complex infrastructure conditions, incidents, risks, and provider performance into clear information for technical and business leadership. Develop reports, dashboards, presentations, and operational analyses related to uptime, infrastructure performance, capacity, incidents, and service-provider performance. Participate in technical and leadership meetings as a subject-matter resource for data center critical infrastructure. Use operational data to identify trends, communicate risk, and drive measurable improvements in reliability and provider performance. We expect you to have: Experience working in data center, cloud infrastructure, colocation, critical facilities, or other mission-critical environments. Practical understanding of IT infrastructure, including servers, racks, networking equipment, structured cabling, and fiber. Working knowledge of data center electrical infrastructure such as UPS systems, generators, switchgear, PDUs, batteries, breakers, and power distribution. Exposure to data center mechanical and cooling systems, including chilled-water, glycol, liquid-cooling, or comparable thermal-management environments. Ability to understand how electrical and mechanical infrastructure directly impacts IT equipment availability and performance. Experience participating in infrastructure change management, incident response, operational risk management, or maintenance governance. Ability to review technical plans, ask detailed engineering questions, identify risk, and work effectively with electrical and mechanical subject-matter experts. Strong analytical skills with experience using Excel for reporting, data analysis, and operational metrics. Strong written and verbal communication skills with the ability to communicate effectively with engineers, vendors, service providers, and senior leadership. A proactive, ownership-driven approach with the ability to operate effectively in a high-availability production environment. Nice to have: Experience supporting high-density GPU, AI, HPC, or hyperscale data center environments. Experience with direct-to-chip liquid cooling or other advanced cooling technologies used for high-density compute. Experience managing colocation or third-party critical infrastructure providers against contractual SLAs. Familiarity with Tier III data center environments and high-availability infrastructure principles. Experience developing or implementing change-management, incident-response, or emergency-response procedures. Experience supporting infrastructure capacity planning, expansion projects, or new data center deployments. Relevant electrical, mechanical, data center, or critical facilities certifications. Key Employee Benefits in the US: Health Insurance: 100% company-paid medical, dental, and vision coverage for employees and families. 401(k) Plan: Up to 4% company match with immediate vesting. Parental Leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers. Disability & Life Insurance: Company-paid short-term, long-term, and life insurance coverage. Join Nebius Today! Pay Transparency We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law. Base Compensation Range $85,000 — $140,000 USD Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Critical Infrastructure Engineer
Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. Infrastructure Engineer Independently operates, maintains, troubleshoots, and takes end-to-end ownership of critical power, critical cooling, liquid cooling, and DCIM/BMS infrastructure systems; participates in incident management and commissioning activities; and mentors Associate Infrastructure Engineers. Role Overview The Infrastructure Engineer is a critical role within Nebius data center operations. A fully proficient Infrastructure Engineer who independently manages and troubleshoots critical power and critical cooling systems across the site and takes ownership of assigned engineering tasks from start to finish. Participates in incident management activities, supports site commissioning and build reviews, and contributes to ensuring all engineering work meets Nebius SLAs and standards. Enforces safe working practices, collaborates with facilities, network, and technician teams, and mentors Associate Infrastructure Engineers toward independent operation. Core Responsibilities Monitor and contribute to the tracking of critical environment maintenance and repair for assigned service lines to Nebius SLAs; identify and escalate developing faults before they impact service uptime. Independently operate, monitor, and troubleshoot critical power distribution infrastructure including UPS systems, PDUs, RPPs, and in-rack busbar power systems. Manage and maintain critical cooling infrastructure including CRAHs, CRACs, CDUs, and RDHx; perform capacity checks and leakage inspections. Participate in incident management for infrastructure-impacting events; support root-cause analyses, document findings, and contribute to CAPA execution under the direction of the Senior Infrastructure Engineer. Operate and administer DCIM and BMS platforms; build and maintain dashboards, alerts, capacity reports, and infrastructure records. Support deployment and commissioning of GB-scale liquid-cooled rack infrastructure including direct liquid cooling (DLC) systems, manifolds, and CDU connections. Perform and own preventive maintenance tasks for all critical power and critical cooling systems; maintain accurate maintenance logs and compliance records. Participate in site reviews, design reviews, and commissioning activities; prepare and review technical reports to document findings and communicate results. Enforce Nebius critical power and critical cooling safety procedures; act as safety lead for engineering work orders on live infrastructure. Collaborate with facilities, network, and technician teams on infrastructure changes, capacity expansions, and major deployments. Support adherence to Nebius standards and policies through documentation review and active participation in commissioning and design activities. Contribute to vendor and contractor coordination by supporting scheduling, site access, and execution of work per Nebius expectations and safe-working practices. Actively mentors and trains Associate Associate Infrastructure Engineers; guides them through systems operation, safe working practices, and structured skill development toward independent operation. Expected Capabilities / Expectations Working proficiency across all critical power systems: UPS, PDU, RPP, in-rack busbar, and generator interfacing. Hands-on experience with critical cooling systems: CRAH/CRAC, CDU, RDHx, and water-side infrastructure including leak detection. Competent DCIM/BMS administration: alert management, capacity modeling, and dashboard reporting. Developing knowledge of GB rack liquid cooling technology, DLC manifolds, and heat rejection infrastructure. U.S. Data Center Career Framework - Technical Career Framework Ability to participate in and support incident management activities including RCA and CAPA documentation. Strong documentation discipline: change records, maintenance logs, capacity data, and incident write-ups. Clear and confident communication with facilities, network, and vendor teams during changes and incidents. Core Physical Requirements Ability to stand and remain on your feet for extended periods, typically four or more consecutive hours during active shift operations. Ability to safely lift, carry, and position equipment weighing up to 50 pounds unassisted, and heavier loads with appropriate team-lift protocols or mechanical aids. Comfortable working at heights including ascending and descending ladders, raised platform equipment, and elevated data center infrastructure. Ability to work in confined spaces such as under raised floors, within enclosed rack enclosures, and in cable management pathways. Comfortable working in environments with variable temperatures, including cold aisle containment zones and active cooling infrastructure. Manual dexterity sufficient to handle small form-factor components, precision cabling, and fine connector installations. Visual acuity sufficient to read equipment labels, small-form-factor interface indicators, and detailed wiring diagrams in variable lighting conditions. Ability to push or pull equipment carts, server sleds, and wheeled infrastructure weighing up to 500 pounds on level surfaces. On-Call Requirement This role includes on-call participation to respond to after-hours critical power and critical cooling events, infrastructure alarms, and urgent maintenance requiring prompt engineering response. Pay Transparency We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law. Base Compensation Range $49 — $54 USD Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Data Center IT Infrastructure Engineer
Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We’re looking for a IT infrastructure engineer to troubleshoot and solve data center IT hardware issues, process the RMA. This is a position for a technical expert working at the intersection of multiple technical and operational domains. Your responsibilities will include Solve the most challenging firmware and hardware related issues with servers, involving in-depth knowledge of system architecture and advanced troubleshooting Execute workarounds and solutions for IT hardware issues Act as a subject matter expert and point of escalation for L1 and L2 technicians Create new processes and documentation for IT hardware team Collaborate with related departments to im Collaborate with vendors on warranty replacements (RMA), create requests and manage. Improve support processes, documentation and training materials Requirements Knowledge of datacenters, and server equipment Deep knowledge of IT hardware and practical experience of troubleshooting Skills working with the Unix/linux operating system and command line Experience with equipment monitoring, data analysis and presentation Proactiveness and sense of responsibility High proficiency in spoken and written English It will be a bonus if you have Skills of repairing electronics at the component level (SMD) Knowledge of network equipment and troubleshooting Driving license type B Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Data Center IT Infrastructure Engineer
Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role We’re looking for a IT infrastructure engineer to troubleshoot and solve data center IT hardware issues, process the RMA. This is a position for a technical expert working at the intersection of multiple technical and operational domains. Your responsibilities will include Solve the most challenging firmware and hardware related issues with servers, involving in-depth knowledge of system architecture and advanced troubleshooting Execute workarounds and solutions for IT hardware issues Act as a subject matter expert and point of escalation for L1 and L2 technicians Create new processes and documentation for IT hardware team Collaborate with related departments to im Collaborate with vendors on warranty replacements (RMA), create requests and manage. Improve support processes, documentation and training materials Requirements Knowledge of datacenters, and server equipment Deep knowledge of IT hardware and practical experience of troubleshooting Skills working with the Unix/linux operating system and command line Experience with equipment monitoring, data analysis and presentation Proactiveness and sense of responsibility High proficiency in spoken and written English It will be a bonus if you have Skills of repairing electronics at the component level (SMD) Knowledge of network equipment and troubleshooting Driving license type B Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Data Center - QA Engineer
Product & Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About the Role We are looking for a technically strong, hands-on QA Engineer to join our hardware team on-site at ODM factories in Taiwan. This is not a checklist job - we're looking for someone who enjoys digging deep into technical issues, investigating root causes, and taking ownership of complex hardware problems.You'll be the key person ensuring the quality of our servers and racks before they ship, but more importantly, you'll play a critical role in debugging failures, analyzing test data, and working closely with RnD, logistics, and factory teams to continuously improve the process and the product.This is a deeply technical role that blends hardware validation, manufacturing QA, and problem-solving - perfect for someone who understands how servers are built and tested, and wants to make sure every unit that leaves the factory is production-grade. What You'll Own Technical Investigation & Debugging Investigate complex problems (e.g., high GPU failure rate, power-related test failures), gather logs, run diagnostics, and escalate with context to RnD when needed. Drive root cause analysis across factory teams and internal engineering groups. Document findings and help define preventive actions for recurring problems. Act as the first line of technical escalation for hardware issues discovered during factory QA or internal testing.Engineering Support Participate in new platform bring-up sessions together with the visiting RnD teams during on-site trips to ODM labs. Provide technical support, coordination, and hands-on assistance during the bring-up process. Help ensure early-stage hardware behaves as expected, and escalate integration or platform issues to the relevant teams.On-Site Product QA Perform visual inspections of completed products (servers, racks) before packaging. Define and maintain QA checklists and inspection procedures tailored to different product lines. Verify inventory records at the factory against internal system data (part numbers, serials, configurations). Oversee the product packaging process for compliance with defined standards. Supervise pickup operations: ensure outbound trucks meet shipment conditions and schedules.Failure Rate Monitoring & Analytics Collect failure data from vendor-side burn-in and our own test systems. Analyze failure trends and estimate spare part needs for future datacenter deployments. Use dashboards and structured reporting to communicate insights with QA, engineering, and supply chain teams.Feedback Loop & Quality Improvement Gather and process feedback from datacenters on each delivered batch of equipment: * Report on packaging issues, impact sensor triggers, shipping anomalies. * Assess rack-level build quality: cabling, bracket alignment, labeling. * Log systemic hardware issues (design flaws, infant mortality, recurring failures). Forward the feedback to the teams: logistics, ODM partners, hardware RnD, QA.Test Infrastructure & Validation Assist with deployment and maintenance of test infrastructure on-site. Ensure Nebius post-manufacturing hardware validation tests run smoothly (uptime, monitoring, coordination with support team). Coordinate real-time issue escalation and basic triage with factory and internal teams.Local Insight & Communication Communicate relevant local risks and context (e.g., typhoons, holidays, factory-specific constraints) to our global logistics and hardware teams. Maintain productive relationships with factory staff, logistics providers, and internal stakeholders. Working Conditions & Tools During production peaks, issues may arise that require fast, hands-on debugging and resolution on-site. Flexibility is expected: you may need to stay late to investigate failures in freshly built batches or arrive early to verify and unblock outbound truck shipments. Rapid response and clear communication with engineering and factory teams are critical during these high-pressure periods. Occasional international travel may be expected to Nebius headquarters in Amsterdam or to datacenters in Europe and the US. Daily work tools involve: * Managing workflows and escalation via Jira * Writing and maintaining technical documentation in Confluence * Using Grafana dashboards for monitoring test environments and system health * Operating with several internal inventory and test control systems What You'll Bring Strong technical background in hardware or systems engineering, able to independently investigate and troubleshoot complex issues with server systems. 5+ years of experience in hardware QA, manufacturing supervision, or server validation. A strong background in R&D is a significant plus. Solid understanding of server and rack hardware: components, layout, cabling, power/cooling, diagnostics. Ability to read and interpret technical documentation (e.g., datasheets, system specs, debug manuals). Solid knowledge of electrical engineering fundamentals (e.g., power specs, grounding, signal integrity). Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...Senior Data Engineer
Technology
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role The Data Engineering team builds and operates the data platform that powers analytics, business intelligence, operational reporting, and data-driven products across Nebius. We ingest data from internal and external systems, develop reliable transformation pipelines and data models, and make trusted datasets available to business and product teams. We are looking for a Senior Data Engineer to own substantial parts of our data platform and deliver complex data products end to end. You will turn ambiguous business needs into pragmatic technical solutions, make architecture and implementation trade-offs, and improve the reliability, scalability, and usability of our data ecosystem. You will work closely with product, platform, and business teams, helping them use data effectively while ensuring that our systems remain maintainable and trustworthy as Nebius grows. Your responsibilities : Own the design, delivery, and operation of complex data pipelines, datasets, and platform components. Translate business and analytical requirements into scalable data models, reliable data products, and clear technical plans. Design and evolve data architecture, storage, processing, and orchestration patterns for large-scale workloads. Improve data quality, observability, lineage, and incident response for critical datasets and pipelines. Investigate and resolve challenging performance, reliability, and data-correctness issues in production. Establish reusable tools, conventions, and automation that improve engineering productivity and reduce operational risk. Work with product teams and business stakeholders to define data contracts, priorities, and success criteria. Contribute to technical direction through design reviews, thoughtful trade-offs, and documentation. Support and mentor other engineers through reviews, pairing, and knowledge sharing. Participate in the on-call rotation and take ownership of improving the operational health of the systems you support. Must-haves : 5+ years of experience in data engineering, backend engineering, or a related role; or equivalent experience delivering and operating production data systems. Proven experience independently delivering complex data pipelines or data-platform capabilities from problem definition through production operation. Strong Python and SQL skills, including writing maintainable production code and optimizing non-trivial queries. Hands-on experience with workflow orchestration tools such as Airflow, Prefect, or Dagster. Strong understanding of data modeling, including designing maintainable analytical models and data contracts for multiple consumers. Solid knowledge of data architectures and storage systems, including the trade-offs between different processing and storage approaches. Experience designing for reliability: testing, monitoring, data-quality validation, alerting, debugging, and incident resolution. Ability to make technical decisions under ambiguity, explain trade-offs clearly, and collaborate effectively with engineers and non-technical stakeholder Nice - to - have s : Experience with real-time or event-driven data platforms and streaming technologies. Experience building or operating cloud-native services with Docker and Kubernetes. Familiarity with Infrastructure as Code, particularly Terraform. Experience with data governance, access control, privacy, and compliance requirements such as GDPR or SOC 2. Experience with data observability and quality tools or frameworks, such as Great Expectations. Experience improving engineering standards through shared libraries, platform tooling, technical documentation, or mentoring. We conduct coding interviews as part of the process. Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
View more...

