← All jobs · Together AI

Staff Engineer, API Core Platform

Together AI ·
51
AI-Agency
B62 U35
📍 San Francisco, US 💰 $240K–$275K Staff 8+ yrs
TypeScriptGolangNext.jsKubernetesAWSTerraformGraphQL
TL;DR

Staff Engineer at Together AI building the core API platform for their AI Cloud infrastructure. Responsible for designing, scaling, and operating mission-critical APIs across public customer-facing and internal systems, with focus on reliability, performance, and developer experience.

Apply at Together AI →
share:
you'll be redirected to the company's career page

Job description

Staff Engineer — API Core Platform

About the role

Together AI is seeking an experienced Backend Engineer to found Together’s API Platform team within the Production Foundations organization. In this role, you will define, build, and scale the core systems and architecture that power Together’s mission-critical APIs — including public customer APIs used directly by customers and via SDKs, CLIs, as well as the client APIs powering Together’s Cloud UI.

In the near term, you will improve and standardize the backend API layer within our primary Next.js monolith, raising the bar on reliability, performance, and consistency. In parallel, you will design and lead the evolution toward scalable, purpose-built next-gen API platform solutions optimized for different Public API and Client API use cases and traffic patterns — defining the long-term architecture and driving its incremental rollout.

This is a deeply hands-on role for an engineer who thrives on writing critical-path code and building platforms that unify engineering efforts across teams. You will work across backend systems, infrastructure layers, identity and access flows, and developer tooling to establish a cohesive API strategy that supports Together’s rapidly growing AI Cloud.

Responsibilities

Required Qualifications

Nice to Have

 

About Together AI

Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.

Compensation

We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $240,000 - $275,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.

Equal Opportunity

Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

 

Please see our privacy policy at https://www.together.ai/privacy 

 

Apply at Together AI →

More open roles at Together AI

Together AI ⚡ AI-native · 🔄 synced 11h ago
AI Researcher, Core ML (Turbo)
📍 San Francisco, US · Senior
AI Researcher, Core ML at Together AI building efficient inference and RL/post-training systems. Role spans algorithms, inference engines (SGLang, vLLM), and production-scale RL pipelines to optimize model speed, cost, and capabilities.
PythonSGLangvLLMGRPORLHFDPO
88
AI-core
Together AI ⚡ AI-native · 🔄 synced 11h ago
Research Engineer, Core ML
📍 San Francisco, US · Staff
Research Engineer, Core ML at Together AI building production inference and RL/post-training systems. Focus on efficient inference algorithms, speculative decoding, and scaling RL pipelines to optimize latency, throughput, and model quality.
PythonSGLangvLLMATLASPyTorchGRPO
82
AI-core
Together AI ⚡ AI-native · 🔄 synced 11h ago
Machine Learning Engineer - Inference
📍 San Francisco, US 💰 $160K–$230K · Mid
Machine Learning Engineer at Together AI building the inference engine for large language models. Focus on optimizing runtime services, performance at scale, and high-performance systems using PyTorch and low-level systems concepts.
PythonPyTorchCUDATritonRustCython
73
AI-fluent
Together AI ⚡ AI-native · 🔄 synced 11h ago
LLM Inference Frameworks and Optimization Engineer
📍 San Francisco, US 🌐 Remote 💰 $160K–$230K · Mid
LLM inference frameworks and optimization engineer at Together AI building distributed inference engines for large language models. Focus on GPU optimization, tensor parallelism, and software-hardware co-design for scalable model serving.
PythonC++CUDATritonTensorRTTensorRT-LLM
73
AI-fluent
Together AI ⚡ AI-native · 🔄 synced 11h ago
Senior Machine Learning Engineer, Voice AI
📍 San Francisco, US 💰 $200K–$260K · Senior
Senior ML Engineer at Together AI optimizing inference for voice models (STT, TTS, speech-to-speech). Focus on model serving engines, GPU optimization, and productionizing voice workloads at scale.
PythonPyTorchTensorRT-LLMvLLMSGLangCUDA
72
AI-fluent