Forward Deployed Engineer focused on inference optimization and post-training, working with strategic customers to deploy and optimize LLM inference engines and fine-tuning pipelines. Requires 5+ years experience and Mandarin speaking. Based in Singapore.
Research internship at Together AI on frontier agents, focusing on building, aligning, and scaling AI systems for complex tasks. Collaborate with leading researchers on training recipes, dataset curation, and infrastructure for agentic AI.
Together AI seeks a Senior Network Engineer to design, deploy, and operate global network infrastructure for high-performance AI compute environments. The role involves hands-on engineering, troubleshooting across Linux, Kubernetes, and automation, and requires deep networking expertise.
Together AI is seeking a Senior Software Engineer to build the next generation AI cloud platform, focusing on infrastructure, Kubernetes, and high-performance compute. This hybrid role in Amsterdam involves designing and operating services that virtualize cutting-edge ML hardware.
Together AI is seeking a Head of Hyperscaler Partnerships to lead end-to-end partnership development with major cloud providers, structuring complex commercial deals across model licensing, inference, and cloud distribution. This principal-level role requires deep experience inside hyperscalers and a track record of closing multi-surface agreements.
Together AI seeks an Infrastructure Design Engineer to own whitespace design for AI data centers, including rack layout, power, cooling, and cabling. The role involves collaborating with engineering teams and contractors to ensure high-density GPU cluster deployments meet specifications.
Together AI seeks a Lead Product Designer to shape AI development tools and lay the foundation for its growing design organization. This role involves leading UX initiatives, evolving the design system, and partnering with engineering and product teams. Requires 7+ years of experience in product-driven environments.
Together AI is seeking a Machine Learning Engineer to join its Inference Engine team, focusing on optimizing AI inference systems for large language models. The role involves building production systems, developing runtime inference services, and collaborating with researchers to create cutting-edge AI solutions.
Together AI's Model Shaping team seeks a Research Engineer to develop a platform for customizing open-source models. You'll work on fine-tuning, reinforcement learning, and evaluation services, and optimize inference engines for post-training workloads.
Together AI is seeking a Senior Backend Engineer for its Inference Platform to build and optimize global request routing, auto-scaling, and multi-tenant traffic shaping. The role requires 5+ years of distributed systems experience and expertise in Rust, Go, Python, or TypeScript, with a focus on low-latency serving of LLMs.
Together AI is seeking a Senior Machine Learning Engineer to drive the model serving layer for voice workloads, optimizing inference for STT/TTS models on their Voice AI platform. This role involves working with TRT-LLM, SGLang, and GPU optimization, and building evaluation frameworks. The position is based in San Francisco with a salary range of $200k-$260k.
Together AI is hiring a Senior Software Engineer, Observability to design and implement a scalable observability platform for their AI Acceleration Cloud. The role involves building monitoring, alerting, and telemetry systems using tools like Prometheus, Grafana, and OpenTelemetry, with a focus on GPU utilization and system performance. Requires expertise in Go/Python, Kubernetes, and distributed systems.
Together AI is hiring a Senior Technical Program Manager to lead cross-functional teams in building and scaling global GPU infrastructure for AI. The role owns the product roadmap, coordinates with research and engineering, and ensures reliable operations of distributed systems.
Together AI is hiring a Staff Software Engineer to build the infrastructure that provisions and manages GPU clusters for AI inference. The role involves designing state machines, APIs, and self-healing systems to turn bare metal into running inference clusters.
Together AI is seeking a Technical Program Manager to own the compute qualification process, screening prospective compute providers for performance and reliability. This remote role coordinates cross-functional engineering teams and drives go/no-go decisions for new GPU clusters.
Together AI is hiring a Director of Data Center Operations to own the operational foundation of its growing data center portfolio across the US and Asia. This role involves designing and commissioning white space deployments, building a break-fix team from scratch, and managing multiple sites. Requires deep technical knowledge of power and cooling systems, plus leadership experience.
Together AI seeks an LLM Inference Frameworks and Optimization Engineer to design and optimize distributed inference engines for large language models. The role focuses on GPU/accelerator optimizations, software-hardware co-design, and high-performance serving. Requires experience with CUDA, TensorRT, vLLM, and distributed systems.
Join Together AI's Turbo team as an AI Researcher to work at the intersection of efficient inference and RL-driven post-training. You'll design and optimize production-scale systems—from algorithms to kernels—and contribute to frontier model development.
Together AI seeks a Director of Data Center Strategy and Site Selection to own the scaling of its AI cloud infrastructure. The role involves site evaluation, commercial negotiations, and cross-functional collaboration to balance cost, risk, and speed across regions.
Together AI seeks a Solutions Architect to work with customers and prospects on Generative AI applications. The role involves technical advisory, demos, POCs, and collaboration with sales. Requires 5+ years in customer-facing technical roles.