Verified Tech Jobs & Hiring Companies, Updated Every 24 Hours
Direct career links to high-growth tech startups and Fortune 500 engineering teams across the United States, Europe, and Worldwide. We audit careers daily to ensure zero ghost listings and zero expired apply links.
All Verified Employers (643)
Filtered and verified against live career portals
The Verified Direct-Apply Tech Job Board
Landing a high-compensation software engineering, data, AI, or product role should not require fighting through zombie job posts, recruiter agency reposts, or expired links. KodeSword indexes verified tech career openings by connecting directly with corporate Applicant Tracking Systems (ATS) including Greenhouse, Lever, Ashby, and Workday. Every single role featured on this platform is active and routes straight to the hiring company’s career page.
Popular Tech Roles
Top Tech Hubs
Why Tech Candidates Use KodeSword vs. Traditional Aggregators
- 100% Direct Corporate Links: Zero middleman recruiter reposts.
- Continuous 24h Pruning: Expired and filled listings removed daily.
- Comprehensive Salary Data: Compensation extracted from verified JDs.
- Zero Paywalls or Registration: Browse and apply completely free.
Frequently Asked Questions
- How often are tech job openings updated on KodeSword?
- Our crawlers sync with official company Applicant Tracking Systems (ATS) including Greenhouse, Lever, Workday, and Ashby every 24 hours. Expired or filled roles are pruned daily to prevent ghost job listings.
- Are these direct job applications or recruiter agency reposts?
- Every role links directly to the official corporate careers portal. There are zero intermediary recruiters, no paywalls, and no sponsored spam.
- What kinds of tech roles are listed on KodeSword?
- We index white-collar software engineering, AI/Machine Learning, DevOps, SRE, Cloud Infrastructure, Data Engineering, Cyber Security, and Technical Product Management roles across US hubs and remote companies.
Togetherai
Actively Hiring29 open positions matching criteria
About the Role Together AI runs one of the largest GPU fleets in the world. The Infra Agent Systems team builds the software systems that power and automate that infrastructure. We develop production AI agents that diagnose hardware failures, investigate incidents, correlate signals across the fleet, and automate operational workflows. Alongside these agents, we build the platform they run on, including knowledge graphs, retrieval systems, orchestration frameworks, and developer tooling. You’ll work across two areas: Infrastructure Agent Systems — Build production AI agents that help operate our GPU fleet by diagnosing failures, investigating incidents, gathering evidence from live systems, and assisting with remediation. These agents are used every day by our infrastructure and datacenter teams through APIs, CLI, dashboards, and Slack. Core Agent Platform — Build the platform that powers these agents, including knowledge graphs, search and retrieval, orchestration, evaluation, and the tooling that enables agents to reason, act, and continuously improve. We’re working on something that hasn’t really been done before: building knowledge graphs and self-improving AI agents that understand, operate, and continuously improve large-scale AI infrastructure. This is an opportunity to work at the intersection of AI agents, distributed systems, infrastructure, and automation , solving challenging engineering problems with real production impact. There’s an enormous amount to build, learn, and shape as we define the future of autonomous infrastructure. responsible for delivering the software but also for operating and supporting it in production. Why this Role You’ll work on two hard problems at the same time: making AI agents trustworthy enough to operate production infrastructure, and building the knowledge, retrieval, and distributed systems that make those agents effective. You’ll have the opportunity to build foundational systems from the ground up, work on infrastructure at massive scale, and help define how self-improving AI agents operate real-world AI infrastructure. Remote based in the UK Responsibilities Design and build production AI agent systems that diagnose, investigate, and remediate infrastructure issues across one of the world’s largest GPU fleets. Build the distributed services, orchestration framework, knowledge graph, and retrieval systems that power infrastructure agents. Develop fleet intelligence systems that combine telemetry, infrastructure state, operational knowledge, and historical incidents to help agents make better decisions. Integrate with observability, incident management, ticketing, fleet inventory, source control, chat, and internal infrastructure systems through well-designed APIs. Own services end to end, including architecture, implementation, testing, deployment, observability, and production operations. Improve agent performance through evaluations, retrieval improvements, better tools, and production feedback loops. Turn what agents learn in production into reliable, reviewed software and automation. Requirements 5+ years of experience building production backend systems, distributed systems, or infrastructure platforms. Strong systems design skills and experience owning significant systems from design through production. Depth in at least one of the following: AI agent systems, orchestration, tool use, evaluation, or grounding Knowledge graphs or graph data modeling Search, retrieval, ranking, RAG, or semantic search systems Strong backend engineering experience, including API design, service boundaries, data modeling, and integrations across complex systems. Experience with Kubernetes, GitOps such as ArgoCD, infrastructure-as-code, and cloud platforms. Comfortable working across languages such as Go, TypeScript, Python, or Rust. Experience in the following is a plus: GPU infrastructure, datacenters, bare-metal systems, hardware failure modes, BMC/IPMI, or cluster schedulers Graph databases Event-driven systems and messaging platforms such as NATS or Kafka Observability platforms such as Prometheus and Grafana Building evaluation frameworks or improving the quality and reliability of LLM-powered systems About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...About the Role Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. The AI Infrastructure team at Together AI is at the forefront of building and scaling the foundational systems that power our generative AI platform. The storage and observability team is crucial for designing, implementing, and maintaining robust distributed storage solutions, ensuring seamless data access and management. They are also responsible for developing comprehensive observability platforms, providing critical insights into system performance and GPU utilization, and proactively identifying and resolving issues. Responsibilities Design and implement a scalable observability platform (metrics, logs, traces) using tools like Prometheus, Grafana, ClickHouse, ClickStack, and OpenTelemetry, including telemetry data pipelines and log aggregation workflows. Develop automated monitoring, alerting, and anomaly detection systems, including SLIs/SLOs, runbooks, and predictive analytics for critical services. Build and deploy custom observability tools and infrastructure-as-code using Go, Python, Terraform, Ansible, and Helm. Collaborate with engineering teams to enhance distributed tracing and application monitoring, and lead incident response with post-mortem analysis. Define observability best practices. Requirements Expertise in observability platforms (Prometheus, Grafana, ClickStack, OpenTelemetry) and cloud-native monitoring services (AWS, GCP, Azure). Strong programming skills in Go, Python, or similar languages, with proficiency in infrastructure-as-code tools (Terraform, Ansible, Helm). Experience designing, operating, and scaling large-scale distributed systems and pipelines for high-volume data ingestion and real-time querying. Deep understanding of containerization (Docker) and orchestration (Kubernetes). Knowledge of microservices architecture, service mesh technologies, CI/CD pipelines, and GitOps workflows. Expertise in managing databases (PostgreSQL, MongoDB, Redis) and time-series databases with high-cardinality data. Preferred Experience monitoring AI/ML infrastructure, GPU clusters, and custom metrics for model performance and training pipelines. Background in high-frequency, low-latency systems monitoring, chaos engineering, and reliability testing. Contributions to open-source observability projects. Familiarity with security monitoring and compliance frameworks. About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Compensation We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $200,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...About the Role Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior Backend Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal StaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world. Some of what you’ll work on: Work on a distributed GPU scheduling system for the on-demand clusters product, Instant Clusters. Build out a global management plane for managing our data center compute, networking, and storage. Design and build new customer-facing cloud platform services, delivering killer enterprise AI cloud features. Responsibilities Identify, design, and develop foundational backend services that power Together’s cloud platform Analyze and improve the robustness and scalability of existing distributed systems, APIs, databases, and infrastructure Partner with product teams to understand functional requirements and deliver solutions that meet business needs Write clear, well-tested, and maintainable software and IaC for both new and existing systems Conduct design and code reviews, create developer documentation, and develop testing strategies for robustness and fault tolerance Participate in an on-call rotation to address critical incidents when necessary Requirements 5+ years of demonstrated experience in building large scale, fault tolerant, distributed systems and API microservices Experience designing, analyzing and improving efficiency, scalability, and stability of various system resources Excellent communication skills – able to write clear design docs and work effectively with both technical and non-technical team members Demonstrated experience with building and operating high-performance and/or globally distributed microservice architectures across one or more cloud providers (AWS, Azure, GCP) Strong systems knowledge across compute, networking, and storage, including concurrency, memory management, performant I/O, and scale Experience developing against and managing a relational database, such as PostgreSQL Expert-level programmer in one or more of programming language (Golang preferred) Proficiency in version control practices and integrating IaC with CI/CD pipelines. Experience with Kubernetes and containers preferred Experience building and operating data infrastructure (Kinesis, Airflow, Kafka, etc) a plus Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical field, or equivalent practical experience About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Compensation We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $160,000 - $230,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...Software Development In Test Intern
Engineering
Role Overview As a Software Development in Test (SDET) Intern, you’ll have the opportunity to be a key player in setting a high quality bar for our users. You’ll work on designing and implementing automated testing processes while getting exposure to a key function at Together AI. Across teams like Cluster Management and Inference Platform, our work centers on automating and testing the critical flows behind our infrastructure. This is a rare opportunity to gain deep insight into how AI infrastructure is provisioned, managed, and scaled or to get hands-on experience benchmarking the newest cutting edge open source models. This role is on-site at our HQ in San Francisco, CA. Responsibilities Developing automated test scripts for functionality, performance, and reliability testing across the website and services Write clean, efficient, and well-documented code with a focus on long-term maintainability Extend and improve test automation frameworks to increase efficiency and overall coverage across the platform. Collaborate with engineering and product teams to understand project requirements and contribute to defining test plans and quality standards Minimum Qualifications Actively pursuing a degree in Computer Science, Software Engineering, or a related field, earned or expected by Summer 2028 Excellent programming skills in Typescript, Go or Python Knowledge of automation testing methodologies, tools, and best practices Ability to solve problems creatively and communicate trade-offs effectively Preferred Qualifications Prior SDET experience through internships, hackathons, or projects Experience in API Testing, AI infrastructure, and/or Git workflows and CI automation. Experience with Playwright or Cypress About Together AI Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, Mamba, FlexGen, Petals, Mixture of Agents, and RedPajama. Internship Program Details Our internship program runs 12 to 14 weeks, giving you the opportunity to work alongside industry-leading engineers and researchers across multiple teams. This cohort's internship dates span either May 17th to August 6th or June 14th to September 3rd. Compensation We offer competitive compensation, housing stipends, and other competitive benefits. The estimated US hourly rate for this role is $58 an hour. Our hourly rates are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...Software Engineer Intern
Engineering
Role Overview As a Software Engineer Intern, you’ll have the opportunity to work on a variety of projects, from designing scalable systems, building product features, to optimizing performance-critical code. You’ll collaborate with cross-functional teams to build robust, user-focused solutions and contribute to our mission of delivering cutting-edge technology. This role is on-site at our HQ in San Francisco, CA and you’ll be placed into one of our engineering teams spanning Platform Engineering, Infrastructure, Inference, and more! Responsibilities Design, develop, and maintain high-quality software across our tech stack from low-level system to customer-facing UIs Collaborate with product managers, designers, and engineers to deliver features and improvements Write clean, efficient, and well-documented code with a focus on maintainability Participate in code reviews, debugging, and performance optimization Contribute ideas to shape the direction of our products and technical infrastructure Qualifications Actively pursuing a degree in Computer Science, Software Engineering, or a related field, earned or expected by Summer 2028 Excellent programming skills Experience with version control systems (e.g., Git) and collaborative development workflows Ability to solve problems creatively and communicate trade-offs effectively Bonus: Experience with web development, databases, distributed systems, or cloud platforms, OSS contributions. About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Internship Program Details: Our internship program spans over 12 weeks where you’ll have the opportunity to work with industry-leading engineers building a cloud from the ground up and possibly contribute to influential open source projects. Our internship dates are May 17th to August 6th or June 14th to September 3rd. Compensation We offer competitive compensation, housing stipends, and other competitive benefits. The estimated US hourly rate for this role is $58/hr. Our hourly rates are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...Software Engineer Intern
Engineering
Role Overview As a Software Engineer Intern, you’ll have the opportunity to work on a variety of projects, from designing scalable systems, building product features, to optimizing performance-critical code. We’re looking for engineers looking to build meaningful projects end-to-end from initial scope into production. You’ll collaborate with cross-functional teams to build robust, user-focused solutions and contribute to our mission of delivering cutting-edge technology. This role is on-site at our HQ in San Francisco, CA and you must be available for our Winter Internship Dates between January to April. You’ll also be placed into one of our engineering teams spanning Platform Engineering, Infrastructure, SREs, and Inference. Responsibilities Design, develop, and maintain high-quality software across our tech stack from low-level system to customer-facing UIs Collaborate with product managers, designers, and engineers to deliver features and improvements Write clean, efficient, and well-documented code with a focus on maintainability Participate in code reviews, debugging, and performance optimization Contribute ideas to shape the direction of our products and technical infrastructure Qualifications Actively pursuing a degree in Computer Science, Software Engineering, or a related field, earned or expected by Summer 2028 Excellent programming skills with experience in Python, Go, or Typescript Ability to solve problems creatively and communicate trade-offs effectively Prior experiences through internships, hackathons, or projects Bonus: Experience with web development, databases, distributed systems, or cloud platforms, OSS contributions. About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Internship Program Details: Our internship program spans over 12 to 14 weeks where you’ll have the opportunity to work with industry-leading engineers building a cloud from the ground up and possibly contribute to influential open source projects. Our internship dates span between January 4th to April 9th. Compensation We offer competitive compensation, housing stipends, and other competitive benefits. The estimated US hourly rate for this role is $58 an hr. Our hourly rates are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...Software Engineer, New Grad
Engineering
About the Role As an early career Software Engineer, you’ll work on a variety of projects, from designing scalable systems, building product features, to optimizing performance-critical code. You’ll collaborate with cross-functional teams to build robust, user-focused solutions and contribute to our mission of delivering cutting-edge technology. This role is ideal for recent graduates with a strong foundation in software development and a passion for learning. This role is on-site at our HQ in San Francisco, CA and you’ll be placed into one of our engineering teams spanning Machine Learning, Platform Engineering, Infrastructure, and Inference. Responsibilities Design, develop, and maintain high-quality software across our tech stack from low-level system to customer-facing UIs. Collaborate with product managers, designers, and engineers to deliver features and improvements. Take high autonomy over your work - own projects end to end Participate in code reviews, debugging, and performance optimization. Contribute ideas to shape the direction of our products and technical infrastructure. Work on real, meaningful projects and learn from experienced professionals along the way. Requirements Actively pursuing a degree in Computer Science, Software Engineering, or a related field, earned or expected by Summer 2027 Solid CS fundamentals (data structures, algorithms, systems thinking) Past hands-on experience through internships, personal projects, or previous roles Experience with version control systems (e.g., Git) and collaborative development workflows. Ability to solve problems creatively and communicate trade-offs effectively. Enthusiasm for learning and thriving in a fast-paced, collaborative environment. Bonus: Experience with web development, databases, distributed systems, or cloud platforms, OSS contributions. About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Compensation We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $150,000 - $160,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a Staff ML Engineer to drive the model serving layer for voice workloads. You'll work hands-on with inference engines like TRT-LLM and SGLang to optimize how we serve models like Whisper, Parakeet, Orpheus, and Kokoro — pushing latency and throughput to the frontier. You'll profile GPU utilization, design batching strategies for streaming audio, and ensure new model architectures can go from research to production quickly. This is a foundational hire on a small, high-impact team. Voice inference has unique challenges — streaming audio, tokenization, real-time latency budgets — that require dedicated ML engineering focus. You'll shape how Together serves voice models as the industry moves from pipeline architectures (ASR → LLM → TTS) toward end-to-end speech-to-speech. Own the model serving stack that powers Together's voice platform across STT, TTS, and speech-to-speech. Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice model inference. Collaborate with model partners (Cartesia, Deepgram, Rime, and others) to bring their models to production on Together's infrastructure. Build quality evaluation frameworks that guide model selection for customers and inform the roadmap. Join a small, early-stage team with outsized impact on a fast-growing product area. Responsibilities Own the voice inference roadmap end-to-end — define and execute the technical strategy for optimizing STT, TTS, and speech-to-speech models across Together's infrastructure, with a clear-eyed view of where the field is heading and how to position the platform ahead of it. Drive best-in-class inference performance — architect and implement systems targeting leading TTFB, throughput, and GPU utilization for voice workloads; set the performance bar others in the industry measure against, not just catch up to. Lead productionization of voice models at scale — design the serving architecture for serverless and dedicated endpoints, including batching strategies, streaming inference pipelines, and memory management tailored to real-time audio; own reliability and latency SLAs. Build the voice evaluation platform — design a rigorous, extensible evaluation framework covering WER across accents, languages, and noise conditions for STT; naturalness, latency, and pronunciation fidelity for TTS; establish the internal benchmark methodology that informs model selection and roadmap decisions. Shape the architecture for next-generation model support — anticipate and enable emerging model paradigms — audio-native LLMs, codec-based architectures (SNAC, Encodec), and end-to-end speech-to-speech systems — before they're mainstream, not after. Serve as the technical DRI for model partner integrations — lead deep collaboration with partners such as Cartesia, Deepgram, and Rime; own the full lifecycle from integration to optimization to ongoing performance accountability. Diagnose and resolve the hardest performance problems in the stack — conduct systematic profiling and root-cause analysis from GPU kernel behavior to framework-level bottlenecks; drive shipped improvements with documented, measurable impact. Influence platform architecture across the organization — partner with platform engineering leadership to ensure the serving layer is built for the latency and reliability demands of real-time voice APIs; your technical decisions should raise the ceiling for the whole team. Define and scale voice fine-tuning capabilities — lead the technical direction for enabling customers to fine-tune STT and TTS models on Together's infrastructure, establishing the primitives for differentiated voice experiences. Lay technical foundations for a category-defining product surface — architect systems with enough foresight that they support multiple new voice products with minimal rework; think in terms of platforms, not point solutions. Requirements 8+ years of ML engineering experience, with a demonstrated focus on model serving, inference optimization, or ML infrastructure at production scale — including systems you've owned from design through live traffic. Deep, practical expertise in LLM serving engines (vLLM, SGLang, TensorRT-LLM, or equivalent) — you've modified engine internals, debugged edge cases under load, and contributed improvements back; you don't stop at the API surface. Expert-level Python and PyTorch proficiency, with a strong command of GPU optimization — CUDA kernels, memory hierarchies, profiling toolchains — and a track record of turning that knowledge into shipped latency or throughput wins. Proven system design judgment — you've made architectural decisions that held up at scale and influenced how a team or platform evolved; you can articulate the tradeoffs you made and why. Strong technical leadership — you operate with high autonomy, define the right problems before solving them, and raise the bar for engineering quality around you without requiring process overhead. Sharp product intuition for developer tooling — you understand what voice application developers actually need to ship great products, and you let that shape your technical priorities, not just the other way around. Proven ability to move fast in ambiguous environments — you've thrived on early-stage or platform teams where scope is wide, ownership is deep, and the roadmap you build is the one you execute. Strong foundation in speech and audio ML (ASR/TTS architectures, audio signal processing) — directly relevant experience is strongly preferred; exceptional ML engineering fundamentals with genuine curiosity about the domain is also considered. Familiarity with audio codec and tokenization schemes (SNAC, Encodec, DAC) is a meaningful plus at this level. Experience training or fine-tuning speech models at scale is a significant advantage. Bachelor's or Master's in Computer Science, Electrical Engineering, or related field — or equivalent depth demonstrated through your work. About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Compensation We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $220,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy
View more...



