Verified Tech Jobs & Hiring Companies, Updated Every 24 Hours
Direct career links to high-growth tech startups and Fortune 500 engineering teams across the United States, Europe, and Worldwide. We audit careers daily to ensure zero ghost listings and zero expired apply links.
All Verified Employers (633)
Filtered and verified against live career portals
The Verified Direct-Apply Tech Job Board
Landing a high-compensation software engineering, data, AI, or product role should not require fighting through zombie job posts, recruiter agency reposts, or expired links. KodeSword indexes verified tech career openings by connecting directly with corporate Applicant Tracking Systems (ATS) including Greenhouse, Lever, Ashby, and Workday. Every single role featured on this platform is active and routes straight to the hiring company’s career page.
Popular Tech Roles
Top Tech Hubs
Why Tech Candidates Use KodeSword vs. Traditional Aggregators
- 100% Direct Corporate Links: Zero middleman recruiter reposts.
- Continuous 24h Pruning: Expired and filled listings removed daily.
- Comprehensive Salary Data: Compensation extracted from verified JDs.
- Zero Paywalls or Registration: Browse and apply completely free.
Frequently Asked Questions
- How often are tech job openings updated on KodeSword?
- Our crawlers sync with official company Applicant Tracking Systems (ATS) including Greenhouse, Lever, Workday, and Ashby every 24 hours. Expired or filled roles are pruned daily to prevent ghost job listings.
- Are these direct job applications or recruiter agency reposts?
- Every role links directly to the official corporate careers portal. There are zero intermediary recruiters, no paywalls, and no sponsored spam.
- What kinds of tech roles are listed on KodeSword?
- We index white-collar software engineering, AI/Machine Learning, DevOps, SRE, Cloud Infrastructure, Data Engineering, Cyber Security, and Technical Product Management roles across US hubs and remote companies.

Lambda
Actively Hiring18 open positions matching criteria
Senior Cloud Infrastructure Engineer, Cloud Foundations
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco/San Jose/Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. Cloud Foundations manages the account, access, and network foundation that Lambda's cloud infrastructure runs on, implementing governance and access controls in line with company policies. In practice that means AWS accounts and IAM, the VPC and transit network structure, DNS, and the Terraform execution platform that every change flows through. This is a new team, and this role helps define it. Working closely with Security, GRC, IT, and the engineering teams that consume the foundation, this position turns policy into working controls and paved paths, so that hundreds of engineers can safely self-serve instead of filing requests. What You’ll Do Own and evolve the IaC execution platform (Terraform, Atlantis): how infrastructure changes are proposed, reviewed, and applied across the company Design and operate Lambda's cloud account, access, and IAM structure: organizational hierarchy, provisioning, and permission models Own the cloud network foundation: VPCs, Transit Gateway, and DNS, including the patterns other teams build on Drive AWS cloud governance: tagging standards, compliance guardrails, and cost visibility and accountability Implement least-privilege access and compliance automation in line with company requirements and policies, replacing manual access and provisioning work with self-service Help shape how this new team operates: its scope, priorities, and service boundaries Participate in an on-call rotation and keep the systems you own well-documented, reliable, and observable Who You Are 6+ years of experience in cloud infrastructure, cloud platform, or cloud security engineering, or a closely related discipline Deep experience managing AWS at scale: multi-account organizations, IAM and permission boundaries, SCPs, and resource organization Strong foundation in infrastructure-as-code (Terraform) and the tooling around it: modules, state management, and automated plan/apply workflows Solid cloud networking fundamentals: VPC design, transit and hybrid connectivity, DNS, and edge/CDN configuration Comfortable writing the automation and tooling that runs the foundation (Go and/or Python), and treating it as a product: real users, real quality standards, and things that actually get adopted Use AI tools as a natural part of how you work, and hold their output to the same bar as your own: you review it, verify it, and own what ships Build for the many, not for one team or one use case Comfortable with ambiguity and clarifying ownership Distill the right problem out of the noise: ask the why behind the why Own what you build past the point it ships: monitor it, measure whether it actually helped, and act on what you learn Sweat the details and aren't satisfied until things run smoothly Nice to Have Experience building agentic workflows that automate meaningful parts of infrastructure operations or governance Experience implementing controls for a compliance program (SOC 2, ISO 27001) without turning it into manual toil Cloud cost management and FinOps experience Hands-on experience with Azure, GCP, or OCI alongside AWS Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Senior Cloud Infrastructure Engineer, Cloud Foundations
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco/San Jose/Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. Cloud Foundations manages the account, access, and network foundation that Lambda's cloud infrastructure runs on, implementing governance and access controls in line with company policies. In practice that means AWS accounts and IAM, the VPC and transit network structure, DNS, and the Terraform execution platform that every change flows through. This is a new team, and this role helps define it. Working closely with Security, GRC, IT, and the engineering teams that consume the foundation, this position turns policy into working controls and paved paths, so that hundreds of engineers can safely self-serve instead of filing requests. What You’ll Do Own and evolve the IaC execution platform (Terraform, Atlantis): how infrastructure changes are proposed, reviewed, and applied across the company Design and operate Lambda's cloud account, access, and IAM structure: organizational hierarchy, provisioning, and permission models Own the cloud network foundation: VPCs, Transit Gateway, and DNS, including the patterns other teams build on Drive AWS cloud governance: tagging standards, compliance guardrails, and cost visibility and accountability Implement least-privilege access and compliance automation in line with company requirements and policies, replacing manual access and provisioning work with self-service Help shape how this new team operates: its scope, priorities, and service boundaries Participate in an on-call rotation and keep the systems you own well-documented, reliable, and observable Who You Are 6+ years of experience in cloud infrastructure, cloud platform, or cloud security engineering, or a closely related discipline Deep experience managing AWS at scale: multi-account organizations, IAM and permission boundaries, SCPs, and resource organization Strong foundation in infrastructure-as-code (Terraform) and the tooling around it: modules, state management, and automated plan/apply workflows Solid cloud networking fundamentals: VPC design, transit and hybrid connectivity, DNS, and edge/CDN configuration Comfortable writing the automation and tooling that runs the foundation (Go and/or Python), and treating it as a product: real users, real quality standards, and things that actually get adopted Use AI tools as a natural part of how you work, and hold their output to the same bar as your own: you review it, verify it, and own what ships Build for the many, not for one team or one use case Comfortable with ambiguity and clarifying ownership Distill the right problem out of the noise: ask the why behind the why Own what you build past the point it ships: monitor it, measure whether it actually helped, and act on what you learn Sweat the details and aren't satisfied until things run smoothly Nice to Have Experience building agentic workflows that automate meaningful parts of infrastructure operations or governance Experience implementing controls for a compliance program (SOC 2, ISO 27001) without turning it into manual toil Cloud cost management and FinOps experience Hands-on experience with Azure, GCP, or OCI alongside AWS Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Senior Cloud Infrastructure Engineer, Cloud Foundations
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco/San Jose/Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. Cloud Foundations manages the account, access, and network foundation that Lambda's cloud infrastructure runs on, implementing governance and access controls in line with company policies. In practice that means AWS accounts and IAM, the VPC and transit network structure, DNS, and the Terraform execution platform that every change flows through. This is a new team, and this role helps define it. Working closely with Security, GRC, IT, and the engineering teams that consume the foundation, this position turns policy into working controls and paved paths, so that hundreds of engineers can safely self-serve instead of filing requests. What You’ll Do Own and evolve the IaC execution platform (Terraform, Atlantis): how infrastructure changes are proposed, reviewed, and applied across the company Design and operate Lambda's cloud account, access, and IAM structure: organizational hierarchy, provisioning, and permission models Own the cloud network foundation: VPCs, Transit Gateway, and DNS, including the patterns other teams build on Drive AWS cloud governance: tagging standards, compliance guardrails, and cost visibility and accountability Implement least-privilege access and compliance automation in line with company requirements and policies, replacing manual access and provisioning work with self-service Help shape how this new team operates: its scope, priorities, and service boundaries Participate in an on-call rotation and keep the systems you own well-documented, reliable, and observable Who You Are 6+ years of experience in cloud infrastructure, cloud platform, or cloud security engineering, or a closely related discipline Deep experience managing AWS at scale: multi-account organizations, IAM and permission boundaries, SCPs, and resource organization Strong foundation in infrastructure-as-code (Terraform) and the tooling around it: modules, state management, and automated plan/apply workflows Solid cloud networking fundamentals: VPC design, transit and hybrid connectivity, DNS, and edge/CDN configuration Comfortable writing the automation and tooling that runs the foundation (Go and/or Python), and treating it as a product: real users, real quality standards, and things that actually get adopted Use AI tools as a natural part of how you work, and hold their output to the same bar as your own: you review it, verify it, and own what ships Build for the many, not for one team or one use case Comfortable with ambiguity and clarifying ownership Distill the right problem out of the noise: ask the why behind the why Own what you build past the point it ships: monitor it, measure whether it actually helped, and act on what you learn Sweat the details and aren't satisfied until things run smoothly Nice to Have Experience building agentic workflows that automate meaningful parts of infrastructure operations or governance Experience implementing controls for a compliance program (SOC 2, ISO 27001) without turning it into manual toil Cloud cost management and FinOps experience Hands-on experience with Azure, GCP, or OCI alongside AWS Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Staff Software Engineer - Managed Kubernetes
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. About the Role Lambda is building the AI Cloud of the future. We are seeking a Staff Engineer to help our development of our Managed Kubernetes platform. Think GKE, but purpose-built for AI workloads and running on bare metal. This is a foundational technical leadership role where you will shape the infrastructure that powers the next generation of AI training and inference at scale. As a Staff Engineer on our Orchestration team, you will collaborate to help drive the technical vision for Lambda's managed orchestration services, including Managed Kubernetes, Managed Slurm on Kubernetes, and higher-level platform services for inference and AIOps. You'll work at the intersection of distributed systems, GPU-accelerated computing, and Cloud Native infrastructure to build systems that are reliable, performant, and elegantly simple for our customers. This is not a role for someone who just operates Kubernetes; it is a technical leadership role for an engineer who has synthesized the core domains of infrastructure (compute, network, storage, security) and can design holistic solutions across all of them. You'll be working closely with NVIDIA's open-source ecosystem, and partnering with internal teams across the stack to deliver a world-class managed platform. What You'll Do Product Engineering Drive technical vision for Lambda's Managed Kubernetes bare-metal platform, including control plane scalability, multi-tenancy, cluster lifecycle management, and high availability Integrate and extend NVIDIA's open-source ecosystem: GPU Operator, Network Operator, DCGM, NCCL, and emerging projects like AICR and Topograph for topology-aware scheduling and placement Design GPU-aware orchestration systems Lead development of services that power our managed services Inform on and help with networking solutions for AI workloads: CNI integration (Cilium, Multus), high-performance fabrics (InfiniBand, RoCE), RDMA, and GPUDirect. You will work closely with our Network team to define and drive requirements Inform and help with storage architecture requirements for AI workloads. You will partner with Storage teams on what managed K8s, Slurm, and future services need Build the foundation for Managed Slurm on Kubernetes, enabling traditional HPC workloads to run seamlessly alongside Kubernetes workload Design higher-level platform services for inference, including model serving infrastructure, autoscaling based on inference load, and multi-model deployment patterns Design self-healing systems and automation for incident response, root cause analysis, and platform resilience Lead chaos engineering efforts to validate system behavior under failure conditions at scale Establish operational excellence for a managed service: upgrade automation, security patching, and zero-downtime maintenance Cross-Functional Infrastructure Leadership Serve as the technical bridge between Orchestration and other infrastructure teams (Network, Storage, Security), translating platform requirements into actionable specifications Drive infrastructure-wide decisions that enable successful managed services. You’re someone who understands what's needed end-to-end, not just at the Kubernetes layer. Provide input on bare-metal provisioning, network topology, and storage systems to ensure they meet the needs of managed the services being built by the Orchestration organization Champion consistency and standardization across Lambda's infrastructure stack Work directly with customers and internal teams to understand existing deployments and chart a path to the managed platform Technical Leadership Set technical direction for Kubernetes services across the Orchestration team, influencing roadmap and prioritization Drive reviews and design sessions, ensuring we build systems that are scalable, maintainable, and aligned with customer needs Mentor and grow engineers, establishing best practices for Kubernetes development, distributed systems, and Cloud Native engineering Collaborate cross-functionally with Network, Storage, Security, and Customer Success teams Engage with NVIDIA and the open-source community to stay current on GPU orchestration technologies and contribute back where appropriate Represent Lambda externally through technical blog posts, conference talks, and strategic customer engagements Shape our AIOps vision: design intelligent systems for automated capacity planning, anomaly detection, and predictive maintenance of cloud infrastructure Who You Are You are a creative, innovative engineer who operates at high velocity. You don't just solve problems. You find elegant solutions and ship them quickly. You embrace modern tools and AI-assisted development (like Claude Code) to accelerate your productivity and multiply your impact. You're energized by building new things, not maintaining the status quo. Required Qualifications 10+ years of experience in software engineering, platform engineering, or SRE, with at least 5 years focused on Kubernetes at scale Expert-level understanding of Kubernetes internals: API machinery, controllers, schedulers, operators, CRDs, CSI, CNI, and the extension patterns that make Kubernetes powerful Holistic infrastructure expertise: you've synthesized knowledge across compute, networking, storage, and security, not just Kubernetes in isolation. You can build solutions that span the full stack. Strong software engineering skills in Go (required) and Python; you write production-quality code, not just scripts Deep experience with GPU orchestration in Kubernetes: NVIDIA GPU Operator, device plugins, DCGM, MIG, time-slicing, and GPU-aware scheduling. Familiarity with NVIDIA Network Operator and GPUDirect is strongly preferred. Proven track record of technical leadership: driving design decisions across teams, mentoring engineers, and influencing infrastructure direction beyond your immediate scope Deep experience designing and operating managed services or multi-tenant platforms. You understand what it takes to run infrastructure for external customers Strong understanding of distributed systems principles: consensus, fault tolerance, consistency models, and graceful degradation Experience with observability at scale: Prometheus, Grafana, distributed tracing, and building actionable alerting systems Solid knowledge of Linux systems and networking (L2-L7), including high-performance networking concepts (RDMA, InfiniBand, RoCE) Experience with infrastructure-as-code and GitOps workflows Preferred Qualifications Experience building and operating managed Kubernetes services (GKE, EKS, AKS, or similar) or working on Kubernetes control plane components Hands-on experience with NVIDIA's open-source ecosystem beyond GPU Operator: Network Operator, NCCL tuning, Topograph, AICR, or similar emerging projects Familiarity with HPC and traditional job schedulers (Slurm) and Kubernetes-native batch scheduling (KAI, Volcano, Kueue) Background in confidential computing Experience migrating customers or workloads from legacy/bespoke infrastructure to standardized platforms Contributions to CNCF projects, Kubernetes SIGs, or NVIDIA open-source projects Familiarity with security and compliance in multi-tenant environments: RBAC, Pod Security Standards, network policies, workload isolation Background in ML infrastructure: training clusters, inference serving, simulation Why Lambda Lambda is building the essential infrastructure for the AI era. We're not just another cloud provider: we're a company founded by ML practitioners, for ML practitioners. Our customers include leading AI research labs and enterprises pushing the boundaries of what's possible with artificial intelligence. What makes this role special: You'll be building core platform services the world’s largest AI companies will consume NVIDIA partnership: Deep integration with NVIDIA's GPU and networking stack, working with cutting-edge open-source tooling Real technical challenges: Massive scale GPU clusters and the unique demands of AI workloads Cross-stack influence: Shape not just Kubernetes, but the network, storage, and compute infrastructure that supports it Direct impact: Your work enables AI breakthroughs. Every model trained on Lambda benefits from systems you build World-class team: Work alongside engineers with deep expertise in ML, systems, and infrastructure Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Senior Software Engineer - Managed Kubernetes
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. About the Role We are seeking a Senior Software Engineer to join our Managed Kubernetes (Mk8s) team. You will play a crucial role in shaping the architecture, reliability, and automation of our Kubernetes-based infrastructure, which powers mission-critical workloads across our global platform. Lambda is building the AI Cloud of the future. We are seeking a Senior Software Engineer to help our development of our Managed Kubernetes platform. Think GKE, but purpose-built for AI workloads and running on bare metal. In this role, you will help build the infrastructure that powers the next generation of AI training and inference at scale. As a Senior Engineer on our Orchestration team, you will contribute to Lambda's managed orchestration services, including Managed Kubernetes, Managed Slurm on Kubernetes, and higher-level platform services for inference and AIOps. You'll work at the intersection of distributed systems, GPU-accelerated computing, and Cloud Native infrastructure to build systems that are reliable, performant, and elegantly simple for our customers. This is not a role for someone who just operates Kubernetes; it's a role for an engineer who understands how compute, network, storage, and security interact, and can build solutions that account for that context — even while focused primarily on the orchestration layer. You'll be working closely with NVIDIA's open-source ecosystem, and partnering with internal teams across the stack to deliver a world-class managed platform. What You’ll Do Design, build, and maintain scalable control plane services, operators, and custom Kubernetes controllers; develop automation in Go/Python for end-to-end cluster lifecycle management — provisioning, upgrades, patching, and deletion Build GPU-aware orchestration systems, working within the platform architecture to support GPU scheduling and resource allocation Partner with the Network team on networking solutions for AI workloads: CNI integration (Cilium, Multus), high-performance fabrics (InfiniBand, RoCE), RDMA, and GPUDirect Write resilient systems that handle failure gracefully — timeouts, retries, backoff, and degraded-mode operation — across large-scale distributed environments Develop platform services for inference: model serving infrastructure, autoscaling based on inference load, and multi-model deployment patterns Build internal tools and CLIs that let ML/AI teams deploy and monitor their own inference services Support and debug production issues through on-call rotation Required Qualifications Have 6+ years of experience in software engineering, with a track record of owning significant technical scope within a team (e.g., driving a project from design through production, or acting as a de facto tech lead on a workstream) Deep understanding of Kubernetes internals: controllers, schedulers, operators, CRDs, CSI, CNI, and the extension patterns that make Kubernetes powerful Solid grasp of distributed systems fundamentals — fault tolerance, graceful degradation, and failure handling in large-scale environments Experience operating the control plane and low-level pieces of large-scale Kubernetes clusters Experience with observability at scale: Prometheus, Grafana, distributed tracing, and building actionable alerting systems Strong programming skills in Go and Python; ability to collaborate effectively on shared codebases Solid knowledge of Linux systems, networking, containers, and cloud infrastructure Take pride in owning and delivering core components of products and platforms Preferred Qualifications Experience building and operating managed Kubernetes services (GKE, EKS, AKS, or similar) or working on Kubernetes control plane components Hands-on experience with NVIDIA's GPU/networking ecosystem: GPU Operator, device plugins, DCGM, MIG, Network Operator, NCCL tuning, or similar Familiarity with HPC and traditional job schedulers (Slurm) and Kubernetes-native batch scheduling (KAI, Volcano, Kueue) Familiarity with GPU, InfiniBand, RDMA, or high-performance computing on Kubernetes Exposure to storage architecture for AI/ML workloads Past contributions to CNCF projects or Kubernetes SIGs a plus If you don’t meet all of these requirements but believe you may be a good fit, please still apply and provide a cover letter that helps us understand your experience and readiness for this role. Why Lambda Lambda is building the essential infrastructure for the AI era. We're not just another cloud provider: we're a company founded by ML practitioners, for ML practitioners. Our customers include leading AI research labs and enterprises pushing the boundaries of what's possible with artificial intelligence. What makes this role special: You'll be building core platform services the world's largest AI companies will consume NVIDIA partnership: Deep integration with NVIDIA's GPU and networking stack, working with cutting-edge open-source tooling Real technical challenges: Massive scale GPU clusters and the unique demands of AI workloads Cross-stack exposure: Work at the intersection of Kubernetes, networking, storage, and compute — gaining depth across the full infrastructure stack that powers AI workloads, not just the orchestration layer Direct impact: Your work enables AI breakthroughs. Every model trained on Lambda benefits from systems you build World-class team: Work alongside engineers with deep expertise in ML, systems, and infrastructure Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Senior Software Engineer - Core Cloud Platform
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco/San Jose/Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. About the Role As a Senior or Staff Software Engineer in Lambda’s Cloud Services Engineering organization, you will build and operate the distributed systems that power Lambda’s GPU cloud. Our teams own platform capabilities across compute control planes, managed Kubernetes, cloud APIs, identity and access, usage metering and billing, capacity and orchestration, reliability, and developer-facing infrastructure. You will turn large-scale GPU infrastructure into reliable, secure, customer-facing cloud services by building APIs, workflows, stateful controllers, schedulers, and operational tooling. You will be full cycle engineer, owning systems through design, deployment, on-call, incident follow-through, and continuous improvement. This role is a strong fit for engineers who enjoy cloud infrastructure, distributed systems, operational excellence, and solving ambiguous problems across software and infrastructure boundaries. Senior engineers lead complex work within a team or domain; Staff engineers additionally shape cross-team architecture and make other teams more effective. What You’ll Do Design, build, and operate services, APIs, control planes, and platform capabilities that power Lambda’s AI cloud. Solve distributed-systems problems involving state, consistency, concurrency, scheduling, failure recovery, and safe lifecycle management. Own the full engineering lifecycle: problem framing, architecture, implementation, testing, rollout, observability, on-call, and continuous improvement. Improve system availability, latency, throughput, efficiency, security, and operability as Lambda grows by orders of magnitude. Turn incidents and near misses into durable engineering improvements, including better automation, testing, guardrails, and backstops. Work across product, infrastructure, networking, storage, security, and SRE teams to resolve dependencies and deliver the right outcome for customers. Use AI-assisted development tools with judgment: accelerate exploration and implementation while independently verifying correctness, security, and maintainability. Contribute to technical standards, design and code reviews, and mentorship; at Staff level, lead cross-team architecture and raise the technical ceiling of the organization. What we’re looking for 7 or more years of professional software engineering experience, or equivalent evidence of impact building production systems. Depth in at least one general-purpose language, we work primarily in Go and Python, and candidates interview in the language they know best. We look for someone who can reason about concurrency, error handling, and testing in that language, not someone who has used it. Experience designing, building, and operating backend services, distributed systems, infrastructure, or platform capabilities at meaningful scale. Practical understanding of system design, data models, APIs, failure modes, performance, and the tradeoffs required to run reliable software in production. A track record of owning complex work through delivery and operation, including testing, staged rollout, monitoring, incident response, and root-cause improvement. Proven track record of aligning cross functional partners and gaining consensus around decisions and tradeoffs. Nice to Have 2+ years of experience building cloud services or platform infrastructure, or operating large-scale production systems on AWS, GCP, Azure, or a comparable cloud platform. Experience with Kubernetes, container orchestration, schedulers, controllers, or cloud control-plane systems. Depth in one or more cloud infrastructure or platform domains, such as compute, storage, networking, identity and access, developer platforms, container orchestration, usage metering and billing, databases, or fleet management. Experience with infrastructure automation, durable workflow systems, event-driven architectures, or infrastructure as code. Experience designing highly available, multi-region, or rapidly scaling distributed systems. Familiarity with GPU infrastructure, HPC environments, or large-scale AI/ML training and inference workloads. Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Senior Security Engineer - Enterprise Security
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our Bellevue, San Francisco, or San Jose office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. About the Role Lambda Security protects some of the world's most valuable digital assets: invaluable training data, model weights representing immense computational investments, and the sensitive inputs required to leverage best of breed AI models. We're responsible for securing every byte that powers breakthrough artificial intelligence. Reporting to the CISO, you'll be a key member of the Enterprise Security team responsible for securing Lambda's SaaS applications, corporate IT infrastructure, identity and access management (IAM), and endpoint ecosystems. Your work will focus on establishing secure-by-default configurations, automating access governance, and protecting corporate environments as Lambda scales rapidly. You will directly impact the security posture of our enterprise operations by implementing robust Identity Governance and Administration (IGA), hardening corporate SaaS tools, automating joiner-mover-leaver (JML) workflows, and deploying zero-trust network access controls. Leveraging Lambda's hosted LLMs, you'll pioneer AI-driven enterprise security automation—such as intelligent access reviews, automated policy enforcement, and SaaS risk monitoring. If you're excited about building enterprise security controls that balance strong protection with developer and employee productivity, we want to talk. We value diverse backgrounds, experiences, and skills, and we are excited to hear from candidates who can bring unique perspectives to our team. If you do not exactly meet this description but believe you may be a good fit, please still apply and help us understand your readiness for a Sr. Security Engineer role. Your application is not a waste of our time. What You’ll Do Identity & Access Governance: Design, deploy, and maintain Identity and Access Management (IAM) architectures, SSO/IdP integrations (Okta, Entra ID), and Privileged Access Management (PAM) across corporate and SaaS ecosystems Automate Joiner-Mover-Leaver (JML) workflows and lifecycle management to enforce least-privilege access and zero-trust principles across all internal tooling Build automated access certification and review tooling to streamline compliance and governance controls Implement strong multi-factor authentication (MFA) standards, passwordless solutions, and contextual access policies SaaS Security & Corporate Infrastructure: Establish SaaS Security Posture Management (SSPM) and conduct security posture reviews for corporate SaaS applications (Google Workspace, Slack, GitHub, Jira, Salesforce) Define and enforce baseline security configurations, patch management policies, and MDM/EDR security controls across corporate endpoints (macOS, Windows, Linux) Partner with IT Infrastructure and Operations teams to architect Zero Trust Network Access (ZTNA), VPN/SASE solutions, and secure remote worker access frameworks Automation & Strategic Innovation: Develop custom security automation tools and scripts in Python or Go to eliminate manual IT/Enterprise security toil Pioneer AI-powered enterprise security workflows leveraging Lambda’s hosted LLMs to automate third-party risk analysis, access request approvals, and anomaly detection in corporate logs Assess third-party vendor risks and build automated tooling to evaluate SaaS vendor security postures Governance & Cross-Functional Alignment: Collaborate with IT, Legal, HR, and Compliance teams to ensure corporate systems align with frameworks such as SOC 2, ISO 27001, and HIPAA Establish internal security documentation, end-user security awareness guidelines, and self-service security options for employees Define strategic roadmap initiatives for corporate security, driving measurable security posture improvements across all business units What We Think a Candidate Needs to Demonstrate to Succeed 5+ years of dedicated security engineering experience, with a focus on enterprise security, corporate security, or IAM infrastructure Proven experience designing and scaling Identity & Access Management architectures (e.g., Okta, Entra ID, SAML/OIDC, SCIM, PAM) Hands-on technical expertise in securing corporate SaaS applications and managing SaaS Security Posture Management (SSPM) tools Strong programming and scripting capabilities (Python, Go, or PowerShell) to build automations for identity lifecycles and security operations Solid understanding of Zero Trust architecture, Endpoint Detection & Response (EDR), Mobile Device Management (MDM), and corporate network security Strong communication and stakeholder management skills, with a track record of partnering effectively with IT, Legal, and HR Ability to balance robust enterprise security controls with employee experience and productivity Thrives in fast-moving, high-growth startup environments where building scalable structure is required Nice to Have Excitement about leveraging our direct access to state-of-the-art LLMs to revolutionize enterprise security—imagine AI-assisted access reviews, automated SaaS risk assessments, and intelligent policy generation. Experience building enterprise security and IAM programs at high-growth tech companies or cloud providers Familiarity with compliance frameworks (SOC 2, ISO 27001, FedRAMP, HIPAA) and audit preparation Experience implementing Infrastructure as Code (Terraform) for managing enterprise identity and SaaS configurations Deep technical background allowing hands-on contribution when needed Experience with both build and buy decisions for security tooling Experience driving or providing significant evidence for compliance audits, such as SOC 2, ISO 27001, PCI-DSS, HIPAA/HITECH, or FedRAMP. Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...Senior Security Engineer - Enterprise Security
Data Center Business
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our Bellevue, San Francisco, or San Jose office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. About the Role Lambda Security protects some of the world's most valuable digital assets: invaluable training data, model weights representing immense computational investments, and the sensitive inputs required to leverage best of breed AI models. We're responsible for securing every byte that powers breakthrough artificial intelligence. Reporting to the CISO, you'll be a key member of the Enterprise Security team responsible for securing Lambda's SaaS applications, corporate IT infrastructure, identity and access management (IAM), and endpoint ecosystems. Your work will focus on establishing secure-by-default configurations, automating access governance, and protecting corporate environments as Lambda scales rapidly. You will directly impact the security posture of our enterprise operations by implementing robust Identity Governance and Administration (IGA), hardening corporate SaaS tools, automating joiner-mover-leaver (JML) workflows, and deploying zero-trust network access controls. Leveraging Lambda's hosted LLMs, you'll pioneer AI-driven enterprise security automation—such as intelligent access reviews, automated policy enforcement, and SaaS risk monitoring. If you're excited about building enterprise security controls that balance strong protection with developer and employee productivity, we want to talk. We value diverse backgrounds, experiences, and skills, and we are excited to hear from candidates who can bring unique perspectives to our team. If you do not exactly meet this description but believe you may be a good fit, please still apply and help us understand your readiness for a Sr. Security Engineer role. Your application is not a waste of our time. What You’ll Do Identity & Access Governance: Design, deploy, and maintain Identity and Access Management (IAM) architectures, SSO/IdP integrations (Okta, Entra ID), and Privileged Access Management (PAM) across corporate and SaaS ecosystems Automate Joiner-Mover-Leaver (JML) workflows and lifecycle management to enforce least-privilege access and zero-trust principles across all internal tooling Build automated access certification and review tooling to streamline compliance and governance controls Implement strong multi-factor authentication (MFA) standards, passwordless solutions, and contextual access policies SaaS Security & Corporate Infrastructure: Establish SaaS Security Posture Management (SSPM) and conduct security posture reviews for corporate SaaS applications (Google Workspace, Slack, GitHub, Jira, Salesforce) Define and enforce baseline security configurations, patch management policies, and MDM/EDR security controls across corporate endpoints (macOS, Windows, Linux) Partner with IT Infrastructure and Operations teams to architect Zero Trust Network Access (ZTNA), VPN/SASE solutions, and secure remote worker access frameworks Automation & Strategic Innovation: Develop custom security automation tools and scripts in Python or Go to eliminate manual IT/Enterprise security toil Pioneer AI-powered enterprise security workflows leveraging Lambda’s hosted LLMs to automate third-party risk analysis, access request approvals, and anomaly detection in corporate logs Assess third-party vendor risks and build automated tooling to evaluate SaaS vendor security postures Governance & Cross-Functional Alignment: Collaborate with IT, Legal, HR, and Compliance teams to ensure corporate systems align with frameworks such as SOC 2, ISO 27001, and HIPAA Establish internal security documentation, end-user security awareness guidelines, and self-service security options for employees Define strategic roadmap initiatives for corporate security, driving measurable security posture improvements across all business units What We Think a Candidate Needs to Demonstrate to Succeed 5+ years of dedicated security engineering experience, with a focus on enterprise security, corporate security, or IAM infrastructure Proven experience designing and scaling Identity & Access Management architectures (e.g., Okta, Entra ID, SAML/OIDC, SCIM, PAM) Hands-on technical expertise in securing corporate SaaS applications and managing SaaS Security Posture Management (SSPM) tools Strong programming and scripting capabilities (Python, Go, or PowerShell) to build automations for identity lifecycles and security operations Solid understanding of Zero Trust architecture, Endpoint Detection & Response (EDR), Mobile Device Management (MDM), and corporate network security Strong communication and stakeholder management skills, with a track record of partnering effectively with IT, Legal, and HR Ability to balance robust enterprise security controls with employee experience and productivity Thrives in fast-moving, high-growth startup environments where building scalable structure is required Nice to Have Excitement about leveraging our direct access to state-of-the-art LLMs to revolutionize enterprise security—imagine AI-assisted access reviews, automated SaaS risk assessments, and intelligent policy generation. Experience building enterprise security and IAM programs at high-growth tech companies or cloud providers Familiarity with compliance frameworks (SOC 2, ISO 27001, FedRAMP, HIPAA) and audit preparation Experience implementing Infrastructure as Code (Terraform) for managing enterprise identity and SaaS configurations Deep technical background allowing hands-on contribution when needed Experience with both build and buy decisions for security tooling Experience driving or providing significant evidence for compliance audits, such as SOC 2, ISO 27001, PCI-DSS, HIPAA/HITECH, or FedRAMP. Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda Founded in 2012, with 500+ employees, and growing fast Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG Our values are publicly available: https://lambda.ai/careers We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
View more...




