Together AI is seeking a Senior Backend Engineer for its Inference Platform to build and optimize global request routing, auto-scaling, and multi-tenant traffic shaping. The role requires 5+ years of distributed systems experience and expertise in Rust, Go, Python, or TypeScript, with a focus on low-latency serving of LLMs.
Together AI is seeking a Senior Machine Learning Engineer to drive the model serving layer for voice workloads, optimizing inference for STT/TTS models on their Voice AI platform. This role involves working with TRT-LLM, SGLang, and GPU optimization, and building evaluation frameworks. The position is based in San Francisco with a salary range of $200k-$260k.
Together AI's Model Shaping team seeks a Research Engineer to develop a platform for customizing open-source models. You'll work on fine-tuning, reinforcement learning, and evaluation services, and optimize inference engines for post-training workloads.
Together AI is seeking a Machine Learning Engineer to join its Inference Engine team, focusing on optimizing AI inference systems for large language models. The role involves building production systems, developing runtime inference services, and collaborating with researchers to create cutting-edge AI solutions.
Together AI seeks a Lead Product Designer to shape AI development tools and lay the foundation for its growing design organization. This role involves leading UX initiatives, evolving the design system, and partnering with engineering and product teams. Requires 7+ years of experience in product-driven environments.
Together AI seeks an Infrastructure Design Engineer to own whitespace design for AI data centers, including rack layout, power, cooling, and cabling. The role involves collaborating with engineering teams and contractors to ensure high-density GPU cluster deployments meet specifications.
Together AI is seeking a Head of Hyperscaler Partnerships to lead end-to-end partnership development with major cloud providers, structuring complex commercial deals across model licensing, inference, and cloud distribution. This principal-level role requires deep experience inside hyperscalers and a track record of closing multi-surface agreements.
Together AI is hiring a Senior Software Engineer, Observability to design and implement a scalable observability platform for their AI Acceleration Cloud. The role involves building monitoring, alerting, and telemetry systems using tools like Prometheus, Grafana, and OpenTelemetry, with a focus on GPU utilization and system performance. Requires expertise in Go/Python, Kubernetes, and distributed systems.
Join Together AI's Turbo team as an AI Researcher to work at the intersection of efficient inference and RL-driven post-training. You'll design and optimize production-scale systems—from algorithms to kernels—and contribute to frontier model development.
Together AI is seeking a Data Center Operations Coordinator to manage break/fix activities across multiple data centers, coordinating vendors, tickets, and reporting to ensure uptime and fast resolution.
Together AI is hiring a Senior Technical Program Manager to lead cross-functional teams in building and scaling global GPU infrastructure for AI. The role owns the product roadmap, coordinates with research and engineering, and ensures reliable operations of distributed systems.
Together AI is hiring a Staff Software Engineer to build the infrastructure that provisions and manages GPU clusters for AI inference. The role involves designing state machines, APIs, and self-healing systems to turn bare metal into running inference clusters.
Together AI is hiring a Director of Data Center Operations to own the operational foundation of its growing data center portfolio across the US and Asia. This role involves designing and commissioning white space deployments, building a break-fix team from scratch, and managing multiple sites. Requires deep technical knowledge of power and cooling systems, plus leadership experience.
Research internship at Together AI on frontier agents, focusing on building, aligning, and scaling AI systems for complex tasks. Collaborate with leading researchers on training recipes, dataset curation, and infrastructure for agentic AI.
Together AI seeks a Commercial Counsel to support infrastructure and go-to-market agreements. The role involves negotiating GPU and cloud capacity deals, managing enterprise commercial agreements, and building scalable legal workflows. Requires a JD and 5+ years of commercial transactions experience.
Together AI is hiring a Senior Developer Productivity Engineer to own the systems and tooling for CI/CD, release infrastructure, local dev environments, and testing systems. The role is high-leverage and build-heavy, focused on enabling engineers to ship quickly and reliably.
Together AI seeks a Product Marketing Director to lead and scale the product marketing function, owning positioning, messaging, and GTM strategy for its AI cloud platform.
Together AI is scaling its compute infrastructure and needs an analytical backbone to support capacity planning, compute sourcing, vendor evaluation, and site selection. This role partners with Infra Eng and Finance to drive data-backed decisions.
Together AI is seeking an IT Engineer to provide hands-on support for in-office and remote employees, manage MDM fleets, administer Okta and SaaS tools, and support procurement and asset management. The role requires experience with Okta, Google Workspace, macOS, and automation.
Together AI is seeking a Systems Research Engineer Intern specialized in GPU programming to develop and optimize GPU-accelerated kernels for ML/AI applications. You will collaborate with cross-functional teams, contribute to cutting-edge AI infrastructure, and work on influential open-source projects. This fall internship offers competitive compensation and a unique opportunity to work with industry-leading engineers.