Together AI jobs
5,613 matches, filter-driven and evidence-linked.
Filters
0 active
Remote, hybrid, onsite
State
Shift type
Weekend work
Country
Cover letter
Assessment
Salary type
Equity type
Family-building benefits
Benefit evidence
-
Senior Software Engineer - Together Cloud Infrastructure
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 433 days agoWhy we showed this
Description: "together"Description: "ai"+5
Unspecified Engineering From the posting source - Senior From the posting source $160K-$230K From the posting source EquitySenior Software Engineer - Together Cloud Infrastructure San Francisco About the Role Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform - a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world. Responsibilities - Design, build, and maintain performant, secure, and highly-available backend services/operators that run in our data centers and automate hardware management, such as Infiniband partitioning, in-DC parallel storage provisioning, and VM provisioning. - Design and build out the IaaS software layer for a new GB200 data center with thousands of GPUs. - Work on a global multi-exabyte high-performance object store, serving massive datasets for pretraining. - Build advanced observability stacks for our customers with automated node lifecycle management for fault-tolerant distributed pretraining. - Perform architecture and research work for decentralized AI workloads - Work on the core, open-source Together AI platform - Create services, tools, and developer documentation - Create testing frameworks for robustness and fault-tolerance To be successful, you'll need to be deeply technical and possess excellent communication, collaboration, and
-
Senior Software Engineer - Together Cloud Platform
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 431 days agoWhy we showed this
Description: "ai"Description: "together"+5
Unspecified Engineering From the posting source - Senior From the posting source $160K-$230K From the posting source EquitySenior Software Engineer - Together Cloud Platform San Francisco About the Role Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior Backend Engineer, you will play a key role in building the next generation AI cloud platform - a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal StaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world. Some of what you'll work on: - Work on a distributed GPU scheduling system for the on-demand clusters product, Instant Clusters. - Build out a global management plane for managing our data center compute, networking, and storage. - Design and build new customer-facing cloud platform services, delivering killer enterprise AI cloud features. Responsibilities - Identify, design, and develop foundational backend services that power Together's cloud platform - Analyze and improve the robustness and scalability of existing distributed systems, APIs, databases, and infrastructure - Partner with product teams to understand functional requirements and deliver solutions that meet business needs - Write clear, well-tested, and maintainable software and IaC for both new and existing systems - Conduct design and code reviews, create developer documentation, and develop testing strategies for robustness and fault tolerance - Participate
-
AI Infrastructure Engineer
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 87 days agoWhy we showed this
Description: "ai"Description: "together"+5
Unspecified Engineering From the posting source - Mid From the posting source $190K-$270K From the posting source EquityAI Infrastructure Engineer San Francisco As an AI Infrastructure Engineer at Together, you are responsible for keeping all user-facing services and production systems running smoothly. You are a blend of a pragmatic operator and a software engineer that applies sound engineering principles, operational discipline, and mature automation to our operating environments and codebase. You specialize in systems (operating systems, storage subsystems, networking), while implementing best practices for availability, reliability and scalability, with varied interests in algorithms and distributed systems. Responsibilities - Participate in on-call rotation (Pagerduty) to respond to production incidents - Build and run our infrastructure with Ansible, Terraform, and Kubernetes to enable scaling to a massive number of concurrent users - Build monitoring systems to ensure the highest quality service for our customers - Design and implement operational processes (such as deployments and upgrades) - Debug production issues across all services and levels of the stack - Identify improvements for the product architecture from the reliability, performance and availability perspectives - Plan the growth of Together AI's infrastructure Requirements - 5+ years of professional AI Infra or related experience - Bachelor's degree in Computer Science or a related field or equivalent work experience - Knowledge of Ansible (roles, playbooks), Terraform, and Kubernetes - Proficiency in programming/scripting languages - Direct experience in monitoring and observability practices - Knowledge of cloud services - Ability to thrive in a collaborative environment involving different stakeholders and subject matter experts About Together AI Together AI is a research-driven artificial intelligence company. We believe
-
Senior Machine Learning Engineer, Voice AI
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 132 days agoWhy we showed this
Description: "together"Description: "ai"+5
Unspecified Engineering From the posting source - Senior From the posting source $200K-$260K From the posting source EquitySenior Machine Learning Engineer, Voice AI San Francisco About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications - serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a Senior ML Engineer to drive the model serving layer for voice workloads. You'll work hands-on with inference engines like TRT-LLM and SGLang to optimize how we serve models like Whisper, Parakeet, Orpheus, and Kokoro - pushing latency and throughput to the frontier. You'll profile GPU utilization, design batching strategies for streaming audio, and ensure new model architectures can go from research to production quickly. This is a foundational hire on a small, high-impact team. Voice inference has unique challenges - streaming audio, tokenization, real-time latency budgets - that require dedicated ML engineering focus. You'll shape how Together serves voice models as the industry moves from pipeline architectures (ASR → LLM → TTS) toward end-to-end speech-to-speech. - Own the model serving stack that powers Together's voice platform across STT, TTS, and speech-to-speech. - Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice model inference. - Collaborate with model partners (Cartesia, Deepgram, Rime, and others) to bring their models to production on Together's infrastructure. - Build quality evaluation frameworks that guide model selection for customers and inform the roadmap. - Join a small, early-stage team with outsized impact on a fast-growing product area. Responsibilities - Optimize inference performance for voice models (STT,
-
Senior Platform Engineer, Voice AI
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 132 days agoWhy we showed this
Description: "ai"Description: "together"+5
Unspecified Engineering From the posting source - Senior From the posting source $200K-$260K From the posting source EquitySenior Platform Engineer, Voice AI San Francisco About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications - serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a Senior Platform Engineer to own the API and infrastructure layer for voice workloads. You'll build the real-time WebSocket and HTTP APIs that developers use to ship voice experiences, design autoscaling for latency-sensitive streaming workloads, and ensure our multi-provider voice platform is reliable enough for production voice agents handling millions of calls. This is a foundational hire on a small, high-impact team. Voice APIs have fundamentally different infrastructure requirements than text-based inference - bidirectional audio streaming, stateful connections, tight latency SLOs, and complex multi-model routing. You'll define how developers interact with Together's voice platform as we grow from early customers to the default infrastructure for voice AI. - Own the real-time API layer (WebSocket + HTTP streaming) that powers Together's voice platform. - Design autoscaling and orchestration for voice workloads running on tens of thousands of GPUs. - Build the developer experience - APIs, observability, and tooling - for a fast-growing product area. - Work with production voice customers (contact centers, AI agents, communication platforms) to ship what they actually need. - Join a small, early-stage team with outsized impact on a new product line. Responsibilities - Build and harden real-time WebSocket and HTTP streaming APIs for STT and TTS - including connection lifecycle
- posted 68 days ago
Why we showed this
Description: "ai"Description: "together"+5
Lead/Manager Together Cloud Infrastructure Amsterdam About the Role Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Lead/Manager, you will play a key role in building the Together cloud platform engineering team in the Netherlands. We are a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world. Some of what you'll work on: - Work on a distributed GPU scheduling system for the on-demand clusters product, Instant Clusters. - Build out a global management plane for managing our data center compute, networking, and storage. - Design and build new customer-facing cloud platform services, delivering killer enterprise AI cloud features. Hybrid working 2 days a week at our offices in Amsterdam Responsibilities - Lead/Manage a team of 8 together cloud Infrastructure Engineer in Amsterdam, - Identify, design, and develop foundational backend services that power Together's commerce platform - Analyze and improve the robustness and scalability of existing distributed systems, APIs, databases, and infrastructure - Partner with product teams to understand functional requirements and deliver solutions that meet business needs - Write clear, well-tested, and maintainable software and IaC for both new and existing
-
Senior Technical Recruiter
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 283 days agoWhy we showed this
Description: "together"Description: "ai"+4
Unspecified Hr From the posting source - Senior From the posting source $165K-$210K From the posting source EquitySenior Technical Recruiter San Francisco About the Role Together AI is building the AI Acceleration Cloud. We are building an end-to-end platform for the generative AI lifecycle, integrating fast, reliable inference and model-shaping services with cutting-edge AI cloud infrastructure. We seek a seasoned Senior Technical Recruiter to collaborate with Engineering leaders to drive hiring for diverse roles within our team. Responsibilities - Partner with executives and engineering leadership to assess current and future hiring needs in core engineering functions - Manage the candidate journey from sourcing handoff to offer stage, ensuring a seamless and exceptional experience. - Provide market intelligence, funnel metrics, and competitive insights to shape hiring strategies and influence decisions. - Collaborate with sourcers, coordinators, and recruiters to streamline processes and deliver a consistent, brand-aligned candidate experience. - Act as a strategic thought partner to leaders, scaling from concept to execution with data-driven insights, creativity, and sound judgment. Requirements - 5+ years of technical recruiting experience in high-growth tech environments. - Proven success in hiring for hyperscale startups - Strong ability to build and maintain relationships with leadership and hiring teams. - You thrive in a fast-paced, ambiguous environment with rapidly shifting priorities. - Expertise in designing new interview processes and hiring strategies from the ground up. - You adapt gracefully under pressure and navigate changing priorities with ease. - Ability to do more with less, you thrive in a scrappy environment and you find creative solutions to problems About Together AI Together AI is a research-driven artificial
-
Staff Machine Learning Engineer, Voice AI
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 82 days agoWhy we showed this
Description: "together"Description: "ai"+5
Unspecified Engineering From the posting source - Staff Plus From the posting source $220K-$280K From the posting source EquityStaff Machine Learning Engineer, Voice AI San Francisco About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications - serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a Staff ML Engineer to drive the model serving layer for voice workloads. You'll work hands-on with inference engines like TRT-LLM and SGLang to optimize how we serve models like Whisper, Parakeet, Orpheus, and Kokoro - pushing latency and throughput to the frontier. You'll profile GPU utilization, design batching strategies for streaming audio, and ensure new model architectures can go from research to production quickly. This is a foundational hire on a small, high-impact team. Voice inference has unique challenges - streaming audio, tokenization, real-time latency budgets - that require dedicated ML engineering focus. You'll shape how Together serves voice models as the industry moves from pipeline architectures (ASR → LLM → TTS) toward end-to-end speech-to-speech. - Own the model serving stack that powers Together's voice platform across STT, TTS, and speech-to-speech. - Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice model inference. - Collaborate with model partners (Cartesia, Deepgram, Rime, and others) to bring their models to production on Together's infrastructure. - Build quality evaluation frameworks that guide model selection for customers and inform the roadmap. - Join a small, early-stage team with outsized impact on a fast-growing product area. Responsibilities - Own the voice inference roadmap end-to-end -
-
Staff Platform Engineer, Voice AI
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 82 days agoWhy we showed this
Description: "ai"Description: "together"+5
Unspecified Engineering From the posting source - Staff Plus From the posting source $220K-$280K From the posting source EquityStaff Platform Engineer, Voice AI San Francisco About the Role Together AI is defining the infrastructure layer for the next generation of voice applications. Our Voice AI platform powers production-grade, real-time voice agents at scale - and we're looking for a Staff Platform Engineer to own the architecture that makes it possible. This isn't a role about maintaining what exists. You'll set the technical direction for how developers interact with Together's voice platform - from the real-time API primitives they build on, to the autoscaling systems that keep latency SLOs intact under unpredictable load, to the multi-provider abstraction layer that makes our platform uniquely powerful. Voice infrastructure is categorically harder than text inference: bidirectional audio streams, stateful long-lived connections, millisecond latency requirements, and complex multi-model routing don't forgive architectural shortcuts. You'll bring the judgment to get this right the first time, at scale. This is a foundational hire on a small, high-conviction team. The decisions you make in this role will define the platform architecture for years. Responsibilities - Own the architecture and reliability of Together's real-time API layer - set the technical direction for WebSocket and HTTP streaming APIs powering STT and TTS at scale; establish the reliability bar (connection lifecycle, backpressure, graceful degradation, reconnection) that production voice agents - contact centers, AI agents, communication platforms - depend on. - Lead autoscaling architecture for latency-sensitive voice workloads - design and ship orchestration systems that handle bursty, real-time traffic across tens of thousands of GPUs; solve the hard problems at
- posted 468 days ago
Why we showed this
Description: "together"Description: "ai"+5
AI infrastructure Engineer (SRE) Amsterdam Amsterdam As a AI Infrastructure Engineer (SRE) at Together, you are responsible for keeping all user-facing services and production systems running smoothly. You are a blend of a pragmatic operator and a software engineer that applies sound engineering principles, operational discipline, and mature automation to our operating environments and codebase. You specialize in systems (operating systems, storage subsystems, networking), while implementing best practices for availability, reliability and scalability, with varied interests in algorithms and distributed systems. Requirements - 7+ years of professional SRE or related experience - Bachelor's degree in Computer Science or a related field or equivalent work experience - Expert knowledge of Ansible (roles, playbooks), Terraform, and Kubernetes - Proficiency in programming/scripting languages - Direct experience in monitoring and observability practices - Advanced knowledge of cloud services - Ability to thrive in a collaborative environment involving different stakeholders and subject matter experts Responsibilities - Be on an on-call (PagerDuty) rotation to respond to incidents that impact availability - Build and run our infrastructure with Ansible, Terraform, and Kubernetes to enable scaling to a massive number of concurrent users - Build monitoring systems to ensure the highest quality service for our customers - Design and implement operational processes (such as deployments and upgrades) - Debug production issues across all services and levels of the stack - Identify improvements for the product architecture from the reliability, performance and availability perspectives - Plan the growth of Together AI's infrastructure About Together AI Together AI is a
-
Senior Developer Productivity Engineer
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 410 days agoWhy we showed this
Description: "ai"Description: "together"+4
Unspecified Engineering From the posting source - Senior From the posting source $150K-$230K From the posting source EquitySenior Developer Productivity Engineer San Francisco About the Role Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure. As a Senior Developer Productivity Engineer at Together AI, you'll own the systems and tooling that empower engineers to ship high-quality software faster. You'll optimize workflows, enhance testing, enable reliable and reusable CI/CD, and work with developers to build out stable local environments. Your work will directly impact release velocity, developer happiness, and cross-cutting enablement ensuring engineers spend less time churning on infrastructure and more time building. Responsibilities - Automate CI/CD pipelines for zero-downtime deployments (e.g., Canary, Blue/Green). - Create smart pipelines encouraging reusable workflows and simplicity - Streamline build/test/deploy workflows. - Build shared tooling (CLIs, codegen, IDE plugins) to accelerate teams. - Reduce friction (e.g., faster builds, hot-reload, test tooling). - Collaborate with developers to identify pain points and streamline workflows. - Champion best practices through documentation. Requirements - Bachelor's degree in Computer Science, Engineering, or related field or
-
Senior Technical Recruiter, AI/ML Research
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 73 days agoWhy we showed this
Employer: semantic matchEmployer: "together"+2
Unspecified Hr From the posting source - Senior From the posting source $165K-$210K From the posting source EquitySenior Technical Recruiter, AI/ML Research San Francisco About the Role Together AI is building the AI Native Cloud - an end-to-end platform for generative AI lifecycle, integrating fast, reliable inference, and model-shaping services with cutting-edge AI cloud infrastructure. We are looking for a seasoned Senior Technical Recruiter to partner closely with AI Research and Engineering leadership to scale world-class research teams across kernels, inference optimization, applied AI and model shaping. This role is ideal for someone who understands the unique dynamics of recruiting top-tier AI researchers and research engineers in a highly competitive market and can operate as a strategic talent partner to technical leadership. Responsibilities - Partner with executives, research leadership, and hiring managers to define and execute hiring strategies - Lead full-cycle recruiting for specialized AI talent, including researchers, research engineers, applied scientists, and ML systems engineers - Build and nurture relationships with top AI talent across academia, open-source communities, research labs, and industry networks - Drive exceptional candidate experiences from initial engagement through offer close, with a strong focus on relationship building and long-term talent cultivation - Provide market intelligence on AI talent trends, compensation, competitive hiring landscapes, and emerging research organizations to influence hiring strategy and organizational planning - Collaborate cross-functionally with sourcing, coordination, people operations, and leadership teams to continuously improve recruiting processes and operational excellence - Design and refine interview processes tailored for research hiring, including technical evaluations, publication reviews, and research presentation loops Requirements - 5+ years of technical recruiting experience at high-growth
-
Senior Backend Engineer, Inference Platform
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 352 days agoWhy we showed this
Employer: semantic matchEmployer: "ai"+2
Unspecified Engineering From the posting source - Senior From the posting source $160K-$250K From the posting source EquitySenior Backend Engineer, Inference Platform San Francisco About the Role Together AI is building the Inference Platform that brings the most advanced generative AI models to the world. Our platform powers multi-tenant serverless workloads and dedicated endpoints, enabling developers, enterprises, and researchers to harness the latest LLMs, multimodal models, image, audio, video, and speech models at scale. If you get a thrill from optimizing latency down to the last millisecond, this is your playground. You'll work hands-on with tens of thousands of GPUs (H100s, H200s, GB200s, and beyond), figuring out how to fully utilize every FLOP and every gigabyte of memory. You'll collaborate directly with research teams to bring frontier models into production, making breakthroughs usable in the real world. Our team also works closely with the open source community, contributing to and leveraging projects like SGLang, vLLM, and NVIDIA Dynamo to push the boundaries of inference performance and efficiency. - Shape the core inference backbone that powers Together AI's frontier models. - Solve performance-critical challenges in global request routing, load balancing, and large-scale resource allocation. - Work with state-of-the-art accelerators (H100s, H200s, GB200s) at global scale. - Partner with world-class researchers to bring new model architectures into production. - Collaborate with and contribute to the open source community, shaping the tools that advance the industry. - A culture of deep technical ownership and high impact - where your work makes models faster, cheaper, and more accessible. - Competitive compensation, equity, and benefits. Responsibilities - Build and optimize global and
- posted 564 days ago
Why we showed this
Description: "together"Description: "ai"+4
Unspecified Engineering From the posting source - Staff Plus From the posting source $180K-$260K From the posting source EquitySolutions Architect San Francisco About the Role As a Solutions Architect at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment. Responsibilities - Act as a technical advisor to our most strategic customers, deeply embedding with them to support the ideation and development of innovative applications using OSS models on Together AI - Run complex demonstrations and POCs of Together's entire stack, including both hardware and software solutions - Collaborate with sales to qualify new prospects and support existing customers along their journey to build cutting-edge Generative AI solutions - Build and maintain strong relationships with customer leadership and stakeholders, ensuring the successful deployment and scaling of their applications - Deliver high-value feedback to our Product, Engineering, and Research teams, ensuring our platform continues to evolve to meet customer needs - Build educational content and tooling for both internal and external use around Together's solutions (i.e., playbooks, blogs, demos, etc.) Qualifications - 5+ years of experience in a customer-facing technical role with at least 2 years in a pre-sales function -
- posted 297 days ago
Why we showed this
Description: "together"Description: "ai"+4
Unspecified Engineering From the posting source - Staff Plus From the posting source Salary not disclosedSolutions Architect (Inference) London About the Role As a Solutions Architect (Inference) at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment. Responsibilities - Act as a technical advisor to our most strategic customers, deeply embedding with them to support the ideation and development of innovative applications using OSS models on Together AI - Run complex demonstrations and POCs of Together's entire stack, including both hardware and software solutions - Collaborate with sales to qualify new prospects and support existing customers along their journey to build cutting-edge Generative AI solutions - Build and maintain strong relationships with customer leadership and stakeholders, ensuring the successful deployment and scaling of their applications - Deliver high-value feedback to our Product, Engineering, and Research teams, ensuring our platform continues to evolve to meet customer needs - Build educational content and tooling for both internal and external use around Together's solutions (i.e., playbooks, blogs, demos, etc.) Qualifications - 7+ years of experience in a customer-facing technical role with at least 2 years in a pre-sales function
-
Machine Learning Engineer
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 569 days agoWhy we showed this
Description: "ai"Description: "together"+4
Unspecified Engineering From the posting source - Mid From the posting source $160K-$220K From the posting source EquityMachine Learning Engineer San Francisco About the Role Together AI is looking for an ML Engineer who will develop systems and APIs that enable our customers to perform inference and fine tune LLMs. Relevant experience includes implementing runtime systems that perform inference at scale using AI/ML models from simple models up to the largest LLMs. Requirements - 5+ years experience writing high-performance, well-tested, production quality code - Bachelor's degree in computer science or equivalent industry experience - Familiar with LLM inference ecosystem, including frameworks and engines (e.g. vLLM, SGLang, TRT, ...) - Demonstrated experience in building large scale, fault tolerant, distributed systems like storage, search, and computation - Expert level programmer in one or more of Python, Go, Rust, or C/C++ - Experience implementing runtime inference services at scale or similar Responsibilities - Design and build the production systems that power the Together Cloud inference and fine-tuning APIs, enabling reliability and performance at scale - Partner with researchers, engineers, product managers, and designers to bring new features and research capabilities to the world - Analyze and improve efficiency, scalability, and stability of various system resources - Conduct design and code reviews - Create services, tools & developer documentation - Create testing frameworks for robustness and fault-tolerance - Participate in an on-call rotation to respond to critical incidents as needed About Together AI Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we
-
Machine Learning Engineer - Inference
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 794 days agoWhy we showed this
Description: "ai"Description: "together"+4
Unspecified Engineering From the posting source - Mid From the posting source $160K-$230K From the posting source EquityMachine Learning Engineer - Inference San Francisco About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and ensuring they run efficiently and effectively at scale. If you are passionate about AI inference, PyTorch, and developing high-performance systems, we want to hear from you. This position offers the chance to collaborate closely with AI researchers and engineers to create cutting-edge AI solutions. Join us in shaping the future at Together AI! Responsibilities - Design and build the production systems that power the Together AI inference engine, enabling reliability and performance at scale. - Develop and optimize runtime inference services for large-scale AI applications. - Collaborate with researchers, engineers, product managers, and designers to bring new features and research capabilities to the world. - Conduct design and code reviews to ensure high standards of quality. - Create services, tools, and developer documentation to support the inference engine. - Implement robust and fault-tolerant systems for data ingestion and processing. Requirements - 3+ years of experience writing high-performance, well-tested, production-quality code. - Proficiency with Python and PyTorch. - Demonstrated experience in building high performance libraries and tooling. - Excellent understanding of low-level operating systems concepts including multi-threading, memory management, networking, storage, performance, and scale. - Preferred: Knowledge of existing AI inference systems such as TGI, vLLM, TensorRT-LLM, Optimum - Preferred: Knowledge of AI inference
-
Staff Engineer, API Core Platform
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 166 days agoWhy we showed this
Description: "ai"Description: "together"+4
Unspecified Engineering From the posting source - Staff Plus From the posting source $240K-$275K From the posting source EquityStaff Engineer, API Core Platform San Francisco Staff Engineer - API Core Platform About the role Together AI is seeking an experienced Backend Engineer to found Together's API Platform team within the Production Foundations organization. In this role, you will define, build, and scale the core systems and architecture that power Together's mission-critical cloud control-plane APIs - including public customer APIs used directly by customers and through SDKs and CLIs, as well as the client APIs powering Together's Cloud UI. This team is distinct from Together's dedicated Inference API team, but partners closely with it to deliver a cohesive customer platform experience. In the near term, you will improve and standardize the backend API layer within our primary Next.js monolith, raising the bar on reliability, performance, and consistency. In parallel, you will design and lead the evolution toward scalable, purpose-built next-gen API platform solutions optimized for different Public API and Client API use cases and traffic patterns - defining the long-term architecture and driving its incremental rollout. This is a deeply hands-on role for an engineer who thrives on writing critical-path code and building platforms that unify engineering efforts across teams. You will work across backend systems, infrastructure layers, identity and access flows, and developer tooling to establish a cohesive API strategy that supports Together's rapidly growing AI Cloud. Responsibilities - Design and drive the evolution of Together's API platform, defining how APIs are built, versioned, secured, tested, and operated across the company - Own and improve the backend API
-
Research Engineer, Frontier Speculative Decoding
Together AI - San Francisco, New York CityIndexed from Greenhouse Comp disclosed in postingposted 263 days agoWhy we showed this
Employer: semantic matchEmployer: "together"+2
Unspecified Engineering From the posting source - Mid From the posting source $190K-$270K From the posting source EquityResearch Engineer, Frontier Speculative Decoding San Francisco, New York City About the Role Together AI is building the Inference Platform that powers the world's most advanced generative AI models. Your role will be a critical bridge between cutting-edge research and real-world applications, focusing on making translating our internal model training research to production-ready deployment for our customers. This involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation before models ever hit production. This role will involve understanding customer specific needs and fine-tuning models on our internal data recipe and their proprietary data. The goal is to transform general-purpose models into highly performant, specialized tools that solve real business problems. You will not be training foundation models from scratch but rather focusing on creating highly efficient, specialized models by working with dedicated GPU clusters. Responsibilities - Design and iterate on novel speculator algorithms, combining architectural innovations with carefully curated data to push the frontier of accuracy-efficiency tradeoffs. - Be the critical link between raw data and a production-ready model, seeing your work directly impact our customers' success. - Work in a fast-paced, high-impact role at the cutting edge of generative AI. - Collaborate with a team of experts dedicated to solving real-world, high-performance challenges. - You'll collaborate directly with customers to understand their needs, and work closely with our core inference and Applied ML research teams to integrate your work into the production platform. - A culture of deep technical ownership where you are empowered to
-
Systems Research Engineer, GPU Programming
Together AI - San FranciscoIndexed from Greenhouse Comp disclosed in postingposted 936 days agoWhy we showed this
Description: "ai"Description: "together"+3
Unspecified Engineering From the posting source - Mid From the posting source $160K-$230K From the posting source EquitySystems Research Engineer, GPU Programming San Francisco About the Role As a Systems Research Engineer specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. Working closely with the modeling and algorithm team, you will co-design GPU kernels and model architecture to enhance the performance and efficiency of our AI systems. Collaborating with the hardware and software teams, you will contribute to the co-design of efficient GPU architectures and programming models, leveraging your expertise in GPU programming and parallel computing. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation. Requirements - Strong background in GPU programming and parallel computing, such as CUDA and/or Triton. - Knowledge of ML/AI applications and models - Knowledge of performance profiling and optimization tools for GPU programming - Excellent problem-solving and analytical skills - Bachelor's, Master's, or Ph.D. degree in Computer Science, Electrical Engineering, or equivalent practical experiences Responsibilities - Optimize and fine-tune GPU code to achieve better performance and scalability - Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems - Stay up-to-date with the latest advancements in GPU programming techniques and technologies About Together AI Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower
-
Senior AI Engineer
Pair Team - Remote (United States)Indexed from Greenhouse Comp disclosed in postingposted 143 days agoWhy we showed this
Title: semantic matchTitle: "ai"+1
Remote From the posting source Data From the posting source - Senior From the posting source $170K-$200K From the posting source EquitySenior AI Engineer Remote (United States) About Pair Team Pair Team is building a new kind of healthcare system across Medicaid, Medicare, and public assistance programs: one that recognizes that access to housing, nutritious food, and reliable transportation are just as critical to health as having the right medications or seeing a doctor. As a public benefit corporation and AI-enabled medical group, we partner with shelters, food pantries, and community organizations to deliver “whole-person” care to the 115 million Americans who rely on the safety net. We are currently the largest complex care provider in California with over 500 employees and are expanding nationally. Our model replaces fragmented healthcare and social services systems with one trusted relationship for all medical, behavioral, and social needs. We improve access, build trust, and dramatically lower costs (52% fewer ER visits, 26% fewer hospitalizations). Our model is a rare combination of saving tax payer dollars ($150B annually at scale) while putting people on an upward life trajectory. At national scale, this approach would save taxpayers. These outcomes are driven by the AI-first, whole-person infrastructure we are building - a platform that connects healthcare and social-service organizations into a unified network. Leveraging our vast data and years of operational experience, we are building the agentic infrastructure for the safety net to coordinate care, automate operations, and learn from every patient interaction to continuously improve outcomes. Check out the AI-First Medicaid System we are building here . - Forbes: For Pair Team, Accessibility Is About Delivering
- posted 213 days ago
Why we showed this
Description: semantic matchDescription: "ai"+1
Partnerships San Francisco, California, United States About Krea At Krea, we are building next-generation AI creative tools. We're dedicated to making AI intuitive and controllable for creatives - our mission is to build tools that empower human creativity, not replace it. We believe AI is a new medium that allows us to express ourselves through various formats - text, images, video, sound, and even 3D. We're building better, smarter, and more controllable tools to harness this medium. We've raised over $83M and are backed by world-class Silicon Valley investors such as Andreesen Horowitz and Bain Capital. We work full-time and in-person at our waterfront office in San Francisco. We care about creativity: our team includes musicians, designers, visual artists, and engineers. This role We're looking for a Partnerships Lead to work at the intersection of product, partnerships, and creativity. You'll identify and close strategic partnerships, stay deeply informed about the creative AI ecosystem, and drive partnerships from idea → implementation → launch. This role requires strong product instincts, comfort coordinating across product and creative teams, and the ability to manage technically complex partnerships. Some stuff you'll do: - Lead negotiations with GPU and cloud providers, managing credits, capacity, pricing, and SLAs. - Stay on top of the creative AI landscape-tools, models, platforms, and emerging players. - Own partnerships end-to-end: from strategy and deal structure to product integration and launch. - Work closely with product and engineering teams to scope, prioritize, and ship partnership-related work. - Coordinate with creative teams to
-
Staff AI Engineer
Pair Team - Remote (United States)Indexed from Greenhouse Comp disclosed in postingposted 143 days agoWhy we showed this
Title: semantic matchTitle: "ai"+1
Remote From the posting source Data From the posting source - Mid From the posting source $210K-$235K From the posting source EquityStaff AI Engineer Remote (United States) About Pair Team Pair Team is building a new kind of healthcare system across Medicaid, Medicare, and public assistance programs: one that recognizes that access to housing, nutritious food, and reliable transportation are just as critical to health as having the right medications or seeing a doctor. As a public benefit corporation and AI-enabled medical group, we partner with shelters, food pantries, and community organizations to deliver “whole-person” care to the 115 million Americans who rely on the safety net. We are currently the largest complex care provider in California with over 500 employees and are expanding nationally. Our model replaces fragmented healthcare and social services systems with one trusted relationship for all medical, behavioral, and social needs. We improve access, build trust, and dramatically lower costs (52% fewer ER visits, 26% fewer hospitalizations). Our model is a rare combination of saving tax payer dollars ($150B annually at scale) while putting people on an upward life trajectory. At national scale, this approach would save taxpayers. These outcomes are driven by the AI-first, whole-person infrastructure we are building - a platform that connects healthcare and social-service organizations into a unified network. Leveraging our vast data and years of operational experience, we are building the agentic infrastructure for the safety net to coordinate care, automate operations, and learn from every patient interaction to continuously improve outcomes. Check out the AI-First Medicaid System we are building here . - Forbes: For Pair Team, Accessibility Is About Delivering
-
AI Engineer
2K - Novato, California, United StatesIndexed from Greenhouse Comp disclosed in postingposted 89 days agoWhy we showed this
Description: "together"Description: "ai"+2
Unspecified Engineering From the posting source - Mid From the posting source $86K-$127K From the posting source Inferred from posting Mental health support EquityAI Engineer Novato, California, United States As an AI Engineer at Cloud Chamber you will join our talented development team to further push the development of the AI Systems and AI Archetypes for BioShock 4. We have high ambitions for tying AI and storytelling together in our project. It will be your responsibility to support several AI Systems and bring our AI Archetypes to life. In this role, you'll help push the state of the art, and experiment with unproven ideas. We are attempting new and ambitious things with AI in this game. You are a bold and curious engineer who believes in an iterative, experimental approach to game development; moving comfortably between design and engineering, collaborating and communicating ambition and limitations. What You'll Do: - Collaborate with the engineering and design teams to iterate and polish existing AI Systems and Archetypes. - Support the development of architecture and implementation within the systems of your domain. - Collaborate with other engineers to implement and maintain AI behavior trees and associated systems. - Test our AI Systems and Archetypes, reporting on their performance operation and memory budgets. - Test and suggest improvements for combat design-facing authoring and encounter scripting tools for - Maintain technical documentation. What We'll Do Together: You will work closely with our Design, Animation, Audio, VFX, and Art departments to support and finalize the implementations for core AI systems: Combat Management, Perception, Steaming/Memory/Performance, etc. You will also help support the creation of new AI Archetypes and finalize existing
Take this list with you
Download the 5,613 matching jobs in any format - read offline, archive, or hand to an AI assistant with your resume to find the best fits.
Or email it to me instead
The AI-ready prompt is a pre-written question you can paste into Claude, ChatGPT, Gemini, or Perplexity along with your resume. We never see your resume; this happens in your AI client of choice.
AI agent reading directly? Same data lives at /api/jobs.json?q=Together+AI.
See /llms.txt and /api/openapi.json for the full schema.