1. Home
  2. AI Companies
  3. Together AI
Together AI logoTA

About

Together AI operates a purpose-built GPU cloud platform for training, fine-tuning, and deploying generative AI models. The infrastructure is designed without vendor lock-in, serving developers and organizations that need to run open-source models at scale. The engineering work centers on distributed systems, model optimization, and AI infrastructure - areas where trade-offs between throughput, latency, and operational complexity define production viability.

The company maintains active contributions to open-source projects including FlashAttention, Mamba, and RedPajama. Engineers and researchers work in close proximity, with new hires taking ownership of substantial technical challenges from the start. The tech stack spans PyTorch, CUDA, TensorRT, TensorRT-LLM, vLLM, SGLang, and TGI, reflecting the requirement to support multiple inference backends and optimization paths. Work involves designing distributed inference engines and developing model architectures where performance characteristics - memory bandwidth utilization, kernel fusion opportunities, multi-GPU coordination overhead - directly impact what models can run economically in production.

Technical problems include optimizing inference for various model architectures across heterogeneous GPU clusters, managing the reliability and cost trade-offs in serving large language models, and building tooling that makes open-source AI accessible without sacrificing control over deployment parameters. The platform must handle the operational complexity of supporting diverse workloads: training runs with different parallelization strategies, fine-tuning jobs with varying dataset sizes, and inference deployments where tail latency matters.

Open roles at Together AI

Explore 56 open positions at Together AI and find your next opportunity.

Together AI logoTA

Staff Software Engineer, GPU Infrastructure Lifecycle Management

Together AI

San Francisco, California, United States (On-site)

$240–280K Yearly1 mo. ago
Together AI logoTA

Data Center Operations Coordinator

Together AI

San Francisco, California, United States (On-site)

$150–200K Yearly18 hr. ago
Together AI logoTA

Solutions Architect

Together AI

London, England, United Kingdom (On-site)

18 hr. ago
Together AI logoTA

Customer Support Engineer (Inference), India

Together AI

India (On-site)

18 hr. ago
Together AI logoTA

Senior Software Engineer - Together Cloud Infrastructure

Together AI

San Francisco, California, United States (On-site)

$160–230K Yearly18 hr. ago
Together AI logoTA

Systems Research Engineer Intern - GPU Programming (Fall 2026)

Together AI

San Francisco, California, United States (On-site)

$58.00 – $63.00 Hourly18 hr. ago
Together AI logoTA

Lead Product Designer

Together AI

San Francisco, California, United States (On-site)

$200–240K Yearly18 hr. ago
Together AI logoTA

Solutions Architect

Together AI

San Francisco, California, US or Remote (United States)

$180–260K Yearly18 hr. ago
Together AI logoTA

Product Marketing Director

Together AI

San Francisco, California, United States (Hybrid)

$250–295K Yearly18 hr. ago
Together AI logoTA

Machine Learning Engineer - Inference

Together AI

San Francisco, California, United States (On-site)

$160–230K Yearly18 hr. ago
Together AI logoTA

Research Engineer, Core ML

Together AI

San Francisco, California, United States (On-site)

$200–280K Yearly18 hr. ago
Together AI logoTA

Senior Software Engineer - Together Cloud Platform

Together AI

San Francisco, California, United States (On-site)

$160–230K Yearly18 hr. ago
Together AI logoTA

Senior Backend Engineer - Together Cloud

Together AI

Amsterdam, North Holland, Netherlands (Hybrid)

18 hr. ago
Together AI logoTA

Strategic Account Executive

Together AI

San Francisco, California, United States (Hybrid)

$150–185K Yearly18 hr. ago
Together AI logoTA

Strategic Finance Senior Associate - Revenue

Together AI

San Francisco, California, United States (On-site)

$138–175K Yearly18 hr. ago
Together AI logoTA

AI Infrastructure Engineer

Together AI

San Francisco, California, United States (On-site)

$190–270K Yearly18 hr. ago
Together AI logoTA

Director of Technical Accounting

Together AI

San Francisco, California, United States (On-site)

$245–300K Yearly18 hr. ago
Together AI logoTA

Senior University Recruiting Program Manager

Together AI

San Francisco, California, United States (On-site)

$155–195K Yearly18 hr. ago
Together AI logoTA

Commercial Counsel-Infrastructure and GTM

Together AI

San Francisco, California, United States (On-site)

$200–230K Yearly6 days ago

Similar companies

OpenAI logoOP

OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity.

677 jobs
Baseten logoBA

Baseten

Baseten is an AI infrastructure platform providing the tooling, expertise, and hardware needed to deploy and scale AI models in production.

62 jobs
d-Matrix logoD-

d-Matrix

d-Matrix builds purpose-built AI inference computing platforms to make generative AI commercially viable, efficient, and sustainable through digital in-memory compute technology.

33 jobs
Modal logoMO

Modal

Modal is a serverless compute platform for AI and data teams that enables running compute-intensive workloads like ML inference, fine-tuning, and batch jobs with instant GPU access and usage-based pricing.

29 jobs
Runpod logoRU

Runpod

RunPod provides cloud infrastructure for AI developers, offering GPU computing services for training, deploying, and scaling AI models.

23 jobs
Inferact logoIN

Inferact

Inferact commercializes vLLM, an open-source LLM inference engine built by its founders, to reduce inference latency, cost, and serving complexity at scale.