{"jobs":[{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5180155007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4652948007,"location":{"name":"Bangalore, India"},"metadata":null,"id":5180155007,"updated_at":"2026-08-04T10:39:34-04:00","requisition_id":"273","title":"AI infrastructure System Engineer Bangalore ","company_name":"Together AI","first_published":"2026-07-13T12:32:15-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h1\u0026gt;\u0026lt;strong\u0026gt;AI Infrastructure Systems Engineer\u0026lt;/strong\u0026gt;\u0026lt;/h1\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Build the infrastructure powering the next generation of AI.\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk.\u0026lt;/p\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;strong\u0026gt;What You’ll Build\u0026lt;/strong\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and build \u0026lt;strong\u0026gt;fleet automation systems\u0026lt;/strong\u0026gt; that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build \u0026lt;strong\u0026gt;AI Infrastructure Agents\u0026lt;/strong\u0026gt; that automate deployment, root-cause failures, incident triage, and autonomous remediation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop \u0026lt;strong\u0026gt;Fleet Intelligence\u0026lt;/strong\u0026gt; platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build software that maximizes \u0026lt;strong\u0026gt;GPU availability, utilization, performance, and reliability\u0026lt;/strong\u0026gt; across thousands of accelerators.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Continuously improve deployment velocity, reliability, and operational efficiency through automation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;strong\u0026gt;What We’re Looking For\u0026lt;/strong\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;3+ years\u0026lt;/strong\u0026gt; building distributed systems, infrastructure platforms, or large-scale backend software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering skills in \u0026lt;strong\u0026gt;Python, Go, or Rust\u0026lt;/strong\u0026gt;.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building platforms, automation systems, or developer infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems thinking with the ability to understand problems across hardware and software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A passion for solving complex infrastructure challenges through software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;An \u0026lt;strong\u0026gt;automation-first mindset\u0026lt;/strong\u0026gt;—if a task is repeated, your instinct is to build a system to eliminate it.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Bonus Experience\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;GPU infrastructure, CUDA, NCCL, NVLink/NVSwitch\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;InfiniBand or RoCE networking\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bare-metal provisioning and lifecycle management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Large-scale AI training or inference clusters\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hardware health monitoring and predictive failure detection\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Distributed storage systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;AI agents and autonomous infrastructure operations\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;strong\u0026gt;You’ll Thrive Here If You\u0026lt;/strong\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Love building systems that replace repetitive operational work.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Think of infrastructure as a software engineering problem.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Enjoy solving hard problems with no existing playbook.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Care deeply about performance, reliability, and scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Want to build technology that powers frontier AI models.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Our mission is simple:\u0026lt;/strong\u0026gt; build AI infrastructure that largely runs itself—where intelligent systems deploy, monitor, diagnose, optimize, and heal GPU fleets at massive scale. Every system you build will directly improve the speed, efficiency, and reliability of one of the world’s most advanced AI compute platforms.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong style=\u0026quot;font-size: 14px;\u0026quot;\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5180155007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5138540007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4125519007,"location":{"name":"San Francisco"},"metadata":null,"id":5138540007,"updated_at":"2026-08-04T14:43:48-04:00","requisition_id":"R\u0026D-ENG-229","title":"AI Infrastructure Systems Engineer","company_name":"Together AI","first_published":"2026-05-14T23:04:22-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Build the infrastructure powering the next generation of AI.\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;You’ll Thrive Here If You:\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Love building systems that replace repetitive operational work.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Think of infrastructure as a software engineering problem.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Enjoy solving hard problems with no existing playbook.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Care deeply about performance, reliability, and scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Want to build technology that powers frontier AI models.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Our mission is simple:\u0026lt;/strong\u0026gt; build AI infrastructure that largely runs itself—where intelligent systems deploy, monitor, diagnose, optimize, and heal GPU fleets at massive scale. Every system you build will directly improve the speed, efficiency, and reliability of one of the world’s most advanced AI compute platforms.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and build \u0026lt;strong\u0026gt;fleet automation systems\u0026lt;/strong\u0026gt; that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build \u0026lt;strong\u0026gt;AI Infrastructure Agents\u0026lt;/strong\u0026gt; that automate deployment, root-cause failures, incident triage, and autonomous remediation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop \u0026lt;strong\u0026gt;Fleet Intelligence\u0026lt;/strong\u0026gt; platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build software that maximizes \u0026lt;strong\u0026gt;GPU availability, utilization, performance, and reliability\u0026lt;/strong\u0026gt; across thousands of accelerators.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Continuously improve deployment velocity, reliability, and operational efficiency through automation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;3+ years\u0026lt;/strong\u0026gt; building distributed systems, infrastructure platforms, or large-scale backend software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering skills in \u0026lt;strong\u0026gt;Python, Go, or Rust\u0026lt;/strong\u0026gt;.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building platforms, automation systems, or developer infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems thinking with the ability to understand problems across hardware and software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A passion for solving complex infrastructure challenges through software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;An \u0026lt;strong\u0026gt;automation-first mindset\u0026lt;/strong\u0026gt;—if a task is repeated, your instinct is to build a system to eliminate it.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Bonus Experience\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;GPU infrastructure, CUDA, NCCL, NVLink/NVSwitch\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;InfiniBand or RoCE networking\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bare-metal provisioning and lifecycle management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Large-scale AI training or inference clusters\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hardware health monitoring and predictive failure detection\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Distributed storage systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;AI agents and autonomous infrastructure operations\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $190,000 - $270,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5138540007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4555544007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4338587007,"location":{"name":"Amsterdam "},"metadata":null,"id":4555544007,"updated_at":"2026-08-04T03:36:55-04:00","requisition_id":"73","title":"AI Infrastructure Systems Engineer (Amsterdam \u0026 London)","company_name":"Together AI","first_published":"2025-04-28T19:00:19-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h1\u0026gt;\u0026lt;strong\u0026gt;AI Infrastructure Systems Engineer\u0026lt;/strong\u0026gt;\u0026lt;/h1\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Hybrid at our office in Amsterdam or Remote in the UK.\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Build the infrastructure powering the next generation of AI.\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk.\u0026lt;/p\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;strong\u0026gt;What You’ll Build\u0026lt;/strong\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and build \u0026lt;strong\u0026gt;fleet automation systems\u0026lt;/strong\u0026gt; that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build \u0026lt;strong\u0026gt;AI Infrastructure Agents\u0026lt;/strong\u0026gt; that automate deployment, root-cause failures, incident triage, and autonomous remediation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop \u0026lt;strong\u0026gt;Fleet Intelligence\u0026lt;/strong\u0026gt; platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build software that maximizes \u0026lt;strong\u0026gt;GPU availability, utilization, performance, and reliability\u0026lt;/strong\u0026gt; across thousands of accelerators.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Continuously improve deployment velocity, reliability, and operational efficiency through automation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;strong\u0026gt;What We’re Looking For\u0026lt;/strong\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;3+ years\u0026lt;/strong\u0026gt; building distributed systems, infrastructure platforms, or large-scale backend software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering skills in \u0026lt;strong\u0026gt;Python, Go, or Rust\u0026lt;/strong\u0026gt;.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building platforms, automation systems, or developer infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems thinking with the ability to understand problems across hardware and software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A passion for solving complex infrastructure challenges through software.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;An \u0026lt;strong\u0026gt;automation-first mindset\u0026lt;/strong\u0026gt;—if a task is repeated, your instinct is to build a system to eliminate it.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Bonus Experience\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;GPU infrastructure, CUDA, NCCL, NVLink/NVSwitch\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;InfiniBand or RoCE networking\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bare-metal provisioning and lifecycle management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Large-scale AI training or inference clusters\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hardware health monitoring and predictive failure detection\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Distributed storage systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;AI agents and autonomous infrastructure operations\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;strong\u0026gt;You’ll Thrive Here If You\u0026lt;/strong\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Love building systems that replace repetitive operational work.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Think of infrastructure as a software engineering problem.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Enjoy solving hard problems with no existing playbook.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Care deeply about performance, reliability, and scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Want to build technology that powers frontier AI models.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Our mission is simple:\u0026lt;/strong\u0026gt; build AI infrastructure that largely runs itself—where intelligent systems deploy, monitor, diagnose, optimize, and heal GPU fleets at massive scale. Every system you build will directly improve the speed, efficiency, and reliability of one of the world’s most advanced AI compute platforms.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029319007,"name":"Europe","location":"Amsterdam, North Holland, Netherlands","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4555544007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4187681007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4125275007,"location":{"name":"San Francisco"},"metadata":null,"id":4187681007,"updated_at":"2026-07-10T19:42:33-04:00","requisition_id":"4","title":"AI Researcher, Core ML (Turbo)","company_name":"Together AI","first_published":"2024-01-16T09:59:47-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;The Turbo team sits at the intersection of efficient inference (algorithms, architectures, engines) and post‑training / RL systems. We build and operate the systems behind Together’s API, including high‑performance inference and RL/post‑training engines that can run at production scale.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Our mandate is to push the frontier of efficient inference and RL‑driven training: making models dramatically faster and cheaper to run, while improving their capabilities through RL‑based post‑training (e.g., GRPO‑style objectives). This work lives at the interface of algorithms and systems: asynchronous RL, rollout collection, scheduling, and batching all interact with engine design, creating many knobs to tune across the RL algorithm, training loop, and inference stack. Much of the job is modifying production inference systems—for example, SGLang‑ or vLLM‑style serving stacks and speculative decoding systems such as ATLAS—grounded in a strong understanding of post‑training and inference theory, rather than purely theoretical algorithm design.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You’ll work across the stack—from RL algorithms and training engines to kernels and serving systems—to build and improve frontier models via RL pipelines. People on this team are often spiky: some are more RL‑first, some are more systems‑first. Depth in one of these areas plus appetite to collaborate across (and grow toward more full‑stack ownership over time) is ideal.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We don’t expect anyone to check every box below. People on this team typically have deep expertise in one or more areas and enough breadth (or interest) to work effectively across the stack. The closer you are to full‑stack (inference + post‑training/RL + systems), the stronger the fit—but being spiky in one area and eager to grow is absolutely okay.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You might be a good fit if you:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Have strong expertise in at least one of the following, and are excited to collaborate across (and grow into) the others:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Systems‑first profile: Large‑scale inference systems (e.g., SGLang, vLLM, FasterTransformer, TensorRT, custom engines, or similar), GPU performance, distributed serving.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;RL‑first profile: RL / post‑training for LLMs or large models (e.g., GRPO, RLHF/RLAIF, DPO‑like methods, reward modeling), and using these to train or fine‑tune real models.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Model architecture design for Transformers or other large neural nets.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Distributed systems / high‑performance computing for ML.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Are comfortable working from algorithms to engines:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong coding ability in Python\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience profiling and optimizing performance across GPU, networking, and memory layers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Able to take a new sampling method, scheduler, or RL update and turn it into a production‑grade implementation in the engine and/or training stack.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Have a solid research foundation in your area(s) of depth:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Track record of impactful work in ML systems, RL, or large‑scale model training (papers, open‑source projects, or production systems).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Can read new RL / post‑training papers, understand their implications on the stack, and design minimal, correct changes in the right layer (training engine vs. inference engine vs. data / API).\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Operate well as a full‑stack problem solver:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;You naturally ask: “Where in the stack is this really bottlenecked?”\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;You enjoy collaborating with infra, research, and product teams, and you care about both scientific quality and user‑visible wins.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Minimum qualifications\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience working on ML systems, large‑scale model training, inference, or adjacent areas (or equivalent experience via research / open source).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advanced degree in Computer Science, EE, or a related field, or equivalent practical experience.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience owning complex technical projects end‑to‑end.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;If you’re excited about the role and strong in some of these areas, we encourage you to apply even if you don’t meet every single requirement.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Advance inference efficiency end‑to‑end\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and prototype algorithms, architectures, and scheduling strategies for low‑latency, high‑throughput inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Implement and maintain changes in high‑performance inference engines (e.g., SGLang‑ or vLLM‑style systems and Together’s inference stack), including kernel backends, speculative decoding (e.g., ATLAS), quantization, etc.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Profile and optimize performance across GPU, networking, and memory layers to improve latency, throughput, and cost.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Unify inference with RL / post‑training\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and operate RL and post‑training pipelines (e.g., RLHF, RLAIF, GRPO, DPO‑style methods, reward modeling) where 90+% of the cost is inference, jointly optimizing algorithms and systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Make RL and post‑training workloads more efficient with inference‑aware training loops—for example, async RL rollouts, speculative decoding, and other techniques that make large‑scale rollout collection and evaluation cheaper.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Use these pipelines to train, evaluate, and iterate on frontier models on top of our inference stack.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Co‑design algorithms and infrastructure so that objectives, rollout collection, and evaluation are tightly coupled to efficient inference, and quickly identify bottlenecks across the training engine, inference engine, data pipeline, and user‑facing layers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run ablations and scale‑up experiments to understand trade‑offs between model quality, latency, throughput, and cost, and feed these insights back into model, RL, and system design.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own critical systems at production scale\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Profile, debug, and optimize inference and post‑training services under real production workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive roadmap items that require real engine modification—changing kernels, memory layouts, scheduling logic, and APIs as needed.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish metrics, benchmarks, and experimentation frameworks to validate improvements rigorously.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Provide technical leadership (Staff level)\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Set technical direction for cross‑team efforts at the intersection of inference, RL, and post‑training.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Mentor other engineers and researchers on full‑stack ML systems work and performance engineering.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $200,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4187681007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5152193007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4639518007,"location":{"name":"San Francisco"},"metadata":null,"id":5152193007,"updated_at":"2026-08-20T17:43:32-04:00","requisition_id":"259","title":"Associate, Infrastructure Strategy \u0026 Operations","company_name":"Together AI","first_published":"2026-06-02T15:35:43-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About The Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is rapidly scaling its compute infrastructure across multiple sites and deployment types. The Associate, Infrastructure Strategy \u0026amp;amp; Operations will be the analytical backbone of the Infrastructure Strategy team, powering the research, benchmarking, and operational analysis across the team\u0026#39;s core workstreams, including capacity planning, compute sourcing, vendor evaluation, and site selection.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll be the person who keeps the team organized and well-informed as infrastructure scales, partnering closely with Infra Eng and Finance to ensure timely, data-backed decisions.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Support the Infrastructure Strategy team\u0026#39;s planning process by gathering data, running comparisons, and preparing materials and recommendations for leadership reviews.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build internal trackers and dashboards that help the team monitor infrastructure deployments, vendor pipelines, and compute allocations across products and customers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain and improve the team\u0026#39;s sourcing comparison frameworks across location, vendor, and site evaluation workstreams, in partnership with Finance and Infra Eng.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run analyses to support capacity allocation decisions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Research and evaluate data center sites and energy sourcing options, comparing power availability, connectivity, permitting timelines, deployment readiness, and reliability.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Champion process improvements across the Infrastructure Strategy function, collaborating cross-functionally with Engineering, Data, and Finance to design AI-native workflows that streamline operations and automate repetitive analysis.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Take on ad hoc analytical projects as priorities evolve, operating with speed and minimal direction.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience in management consulting, business operations, strategy, or a similar analytically rigorous role.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in Excel/Google Sheets, familiarity with SQL or data visualization tools, and comfort with AI productivity tools (e.g., Claude Code, Codex).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong quantitative skills with a data-driven approach to problem-solving and comfort building analyses from scratch.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to learn new domains quickly and operate effectively in unfamiliar territory.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Organized and detail-oriented, with the ability to manage multiple workstreams and keep information current.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Clear communicator who can present data and comparisons to both technical and non-technical audiences.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Prior experience at a high-growth startup or AI company.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Exposure to cloud infrastructure, data center strategy, or GPU/compute procurement.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Interest or background in power markets, energy procurement, and/or hardware.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an AI-native cloud company building the infrastructure to make AI faster, cheaper, and more accessible. We\u0026#39;re rapidly scaling our GPU footprint: signing our own data center leases, building large-scale clusters, and expanding toward a global owned-infrastructure presence. Our research team has contributed to breakthroughs like FlashAttention, Hyena, and RedPajama, and we co-design across software, hardware, and algorithms to push the frontier of AI efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $140-170K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4044336007,"name":"Business Operations","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5152193007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5182998007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4654339007,"location":{"name":"San Francisco"},"metadata":null,"id":5182998007,"updated_at":"2026-07-10T19:42:37-04:00","requisition_id":"276","title":"Commercial Counsel-Infrastructure and GTM","company_name":"Together AI","first_published":"2026-07-10T17:55:56-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is looking for a Commercial Counsel to support both the agreements that secure our infrastructure and the deals that drive our growth. You\u0026#39;ll work alongside the Infrastructure, Sales, Finance, Security, and Procurement teams as the legal partner who helps each of them move quickly while keeping risk in view.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Our business runs on a large, multi-vendor compute footprint, and the terms we lock in early — covering capacity, pricing, service levels, and the freedom to change providers — shape our resilience for years. At the same time, what we sell sits at the core of our customers\u0026#39; own products, so our commercial deals call for a real understanding of how that technology works. This role suits a lawyer who wants to operate across both of those worlds rather than within a narrow contracts lane.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Infrastructure\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Lead negotiations for GPU and cloud capacity, colocation, power, data center space, networking, and related infrastructure-services agreements across our many suppliers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Negotiate terms to match commitments we make in customer contracts — keeping capacity, service levels, security, and data protection obligations aligned and ensuring compliance.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Handle export and trade-control requirements, such as end-use and end-user conditions, as they flow into compute and hardware deals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build the templates, playbooks, and scalable workflows that make infrastructure dealmaking repeatable — standard forms, negotiation playbooks, and clear escalation paths that keep capacity and vendor deals moving as our compute footprint grows\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;GTM\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Manage enterprise commercial agreements\u0026amp;nbsp; — master agreements, order forms, data processing terms, security exhibits, and NDAs — as the legal point of contact for Sales across each deal\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support go-to-market and partnership deals — including collaboration, reseller, and co-sell structures.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build the templates, playbooks, and scalable workflows that make Sales deals repeatable — standard forms, negotiation playbooks, and clear escalation paths\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Procurement\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Run procurement transactions across the company and stand up the request, review, and approval processes, baseline terms, and tooling that let Procurement and internal teams handle routine deals on their own and at pace\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop and maintain Together\u0026#39;s library of commercial and vendor templates and default negotiating positions to build scalable and repeatable processes\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;A JD and an active bar license in good standing\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;At least 5 years in commercial and technology transactions, including substantial in-house time at an infrastructure company\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience supporting deals for purchasing compute and critical infrastructure, partner and strategic alliances agreements, and enterprise commercial deals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A working command of data privacy (GDPR, CCPA, HIPAA) and regulatory (EU AI Act, EU Data Act, DORA) regimes and security terms (data processing agreements, security addenda)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with export controls and the regulatory requirements that apply to partners and counterparties, such as the FCPA and other anti-bribery and anti-corruption regimes\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A plain-spoken style that makes legal and technical points accessible to non-lawyers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfort working independently in a small, fast-moving legal team, with sound instincts for what to own, what to escalate, and when to bring in specialized outside counsel\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Desire for learning and building with AI tools so you can operate self-sufficiently and at speed\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $200-230K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062686007,"name":"Legal","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5182998007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5123203007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4625336007,"location":{"name":"San Francisco"},"metadata":null,"id":5123203007,"updated_at":"2026-07-14T19:43:35-04:00","requisition_id":"S\u0026M-CSX-026","title":"Customer Success Engineer (CSE), GPU Cluster","company_name":"Together AI","first_published":"2026-04-30T13:27:42-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Customer Success Engineer at Together AI, you will serve as the named technical owner for one of our most strategic customer relationships. You will be the primary technical point of contact across all infrastructure domains — compute, networking, storage, and facilities — ensuring flawless delivery and operational health of large-scale GPU deployments. This role sits at the intersection of deep infrastructure expertise and high-stakes customer partnership, making you a critical driver of both customer success and company growth.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Serve as the named technical point of contact for a dedicated strategic customer, owning the end-to-end technical relationship across compute, networking, storage, and facilities\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Drive structured engagement through regular cadences — status reporting, technical steering meetings, quarterly business reviews (QBRs), and executive business reviews (EBRs) — spanning both operational and strategic levels\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Translate customer operational feedback into actionable input for Engineering, Product, and Infrastructure roadmaps\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Lead issue lifecycle management, escalation, and RCA authorship across all infrastructure domains in partnership with Support, SRE, DC Ops, and Engineering teams\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own end-to-end RMA coordination and hardware lifecycle management, including acceptance testing, spare inventory management, and hardware health reporting for large-scale GPU deployments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain deep technical expertise across the customer\u0026#39;s infrastructure stack — GPU compute, high-speed fabric, and large-scale storage systems — advising on configuration, operational best practices, and incident resolution\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own the observability strategy for the customer estate, including alert policy definition, dashboard development, and proactive health management across all infrastructure layers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Coordinate DC operations and facilities events in partnership with internal teams and hosting providers, ensuring SLA compliance and cluster availability\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Act as project manager for all capacity expansions, owning the full node deployment lifecycle from freight receipt through production acceptance\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years in a customer-facing technical role, with 2+ years in dedicated technical account management or solutions architecture for large-scale AI or HPC infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep expertise in GPU infrastructure — GPU health diagnostics, RMA workflows, and hardware acceptance testing\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience with large-scale Ethernet and InfiniBand fabric architecture\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Working knowledge of enterprise storage systems, including high-density NVMe, parallel file systems, and metadata infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with DC operations, facilities coordination, and hosting provider SLA management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong ownership mindset for incident management, RCA authorship, and executive-level customer communication\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in infrastructure monitoring and observability tooling (Prometheus, Grafana, or equivalent)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to manage multiple concurrent workstreams with hyperscaler-level rigor and communication standards\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in Python, Bash, or infrastructure automation tools preferred\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers on our journey in building the next generation of AI infrastructure.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $260-290K OTE + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Location\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;San Francisco, CA (Hybrid) or New York, NY (Hybrid)\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5123203007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5160139007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4643353007,"location":{"name":"San Francisco"},"metadata":null,"id":5160139007,"updated_at":"2026-07-10T19:42:36-04:00","requisition_id":"264","title":"Data Center Operations Coordinator","company_name":"Together AI","first_published":"2026-06-11T13:47:00-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;div\u0026gt;We’re looking for a detail-oriented Data Center Operations professional to manage and track all break/fix activities across multiple data center locations. This role acts as the central point of coordination for hardware incidents, vendor dispatches, ticket management, asset tracking, and operational reporting to ensure maximum uptime and fast issue resolution.\u0026lt;/div\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Track and manage all break/fix incidents across multiple data centers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Monitor ticket queues and ensure SLA compliance for incident response and resolution\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Coordinate with on-site technicians, remote hands teams, vendors, and engineering groups\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain accurate records of failed hardware, replacements, RMAs, and repair status\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Escalate critical outages and recurring infrastructure issues to leadership and engineering teams\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Schedule and oversee maintenance windows and emergency repair activities\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Provide daily/weekly operational status reports and incident summaries\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ensure all work follows data center operational procedures and change management policies\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Identify trends in hardware failures and recommend process improvements\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience working in data center operations, IT infrastructure, or hardware support\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong understanding of server, storage, and networking hardware\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with ticketing systems such as ServiceNow, Jira, or Remedy\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to manage multiple priorities across several sites simultaneously\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and organizational skills\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with SLA management and incident escalation processes\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency with Excel, reporting dashboards, and inventory tracking tools\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;Preferred Qualifications\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience supporting enterprise or hyperscale data centers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Knowledge of remote hands operations and vendor management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Understanding of ITIL processes and change management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;CompTIA Server+, Network+, or similar certifications\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven AI infrastructure company on a mission to dramatically lower the cost of modern AI by co-designing software, hardware, algorithms, and models. We believe open and transparent AI systems create the best outcomes for society — and we\u0026#39;re building the physical and computational foundation to make that real. Our team has been behind landmark advances including FlashAttention, Hyena, FlexGen, and RedPajama.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $150,000-200,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5160139007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5101202007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4615017007,"location":{"name":"San Francisco"},"metadata":null,"id":5101202007,"updated_at":"2026-07-10T19:42:36-04:00","requisition_id":"R\u0026D-ENG-152","title":"Director, Data Center Operations","company_name":"Together AI","first_published":"2026-04-07T20:20:52-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is scaling its physical AI infrastructure rapidly — and we\u0026#39;re looking for a Director of Data Center Operations to help us build it right. This is a ground-floor opportunity to own the operational foundation of Together\u0026#39;s growing data center portfolio across the US and Asia.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll be responsible for designing and commissioning white space deployments — taking pre-built environments and fitting them out with the power distribution, cooling distribution, and systems infrastructure needed to run high-density GPU workloads at scale. At the same time, you\u0026#39;ll be building the break-fix and smart hands team from scratch: hiring, defining the playbook, and standing up the function that keeps our sites running around the clock.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is not a steady-state operations role. It\u0026#39;s a builder role. You\u0026#39;ll be joining a small but fast-moving team, with real ownership over outcomes and the autonomy to shape how Together AI operates its physical infrastructure for years to come. If you\u0026#39;ve scaled data center infrastructure through hypergrowth before and want to do it again with more ownership — this is that opportunity.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the design, fit-out, and commissioning of white space sites across the US and Asia, with a focus on power distribution (PDUs), cooling distribution (CDUs), and IT-adjacent infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and lead a ~20-person break-fix and smart hands team from scratch — define the operating model, hire the initial team, and establish the processes and playbooks that keep sites running\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage a portfolio of 5+ sites across two regions in various stages of deployment and live operation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with vendors, contractors, and equipment suppliers to drive site deployments to schedule and quality\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish operational standards, runbooks, and escalation processes for a nascent but rapidly growing infrastructure function\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Serve as the technical authority on data center infrastructure decisions — from white space evaluation through to live production operation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Travel to Asia periodically to oversee regional site deployments and support the local team\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Deep, hands-on technical knowledge of data center power and cooling systems at the IT-adjacent layer — PDUs, CDUs, power distribution, and white space fit-out\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven experience designing and commissioning data center infrastructure, from evaluation through to live operation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience operating data center infrastructure at meaningful scale, supporting large production workloads\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;People leadership experience — you\u0026#39;ve hired, developed, and led technical operations teams\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A builder\u0026#39;s instinct — you\u0026#39;re comfortable standing up teams and functions from scratch, writing the playbook where none exists, and operating with ambiguity\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong vendor and contractor management skills — you know how to hold external partners accountable to timeline and quality\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Nice to have: experience with GPU or AI infrastructure deployments, multi-site or multi-region portfolio management, or familiarity with Asian markets\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven AI infrastructure company on a mission to dramatically lower the cost of modern AI by co-designing software, hardware, algorithms, and models. We believe open and transparent AI systems create the best outcomes for society — and we\u0026#39;re building the physical and computational foundation to make that real. Our team has been behind landmark advances including FlashAttention, Hyena, FlexGen, and RedPajama.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $250,000 - $300,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5101202007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5098697007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4613876007,"location":{"name":"San Francisco"},"metadata":null,"id":5098697007,"updated_at":"2026-07-10T19:42:36-04:00","requisition_id":"G\u0026A-FIN-016","title":"Director of Tax","company_name":"Together AI","first_published":"2026-04-07T11:47:12-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We’re seeking an exceptional tax leader to establish and scale our tax function from the ground up. This is a rare opportunity to define global tax strategy across multiple jurisdictions within a high-growth, dynamic environment, partnering with a world-class team.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Tax Strategy \u0026amp;amp; Planning\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Develop and execute a comprehensive global tax strategy aligned with business objectives\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advise leadership on tax implications of corporate structure, financing, and expansion\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Optimize effective tax rate through strategic planning and structuring\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Compliance \u0026amp;amp; Reporting\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Lead all federal, state, and international tax compliance (income, sales/use, VAT, payroll tax etc.)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ensure accurate and timely tax filings and provisions under U.S. GAAP (ASC 740) and applicable international local laws\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage audits and inquiries from tax authorities\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Indirect Tax\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Lead global indirect tax strategy for Together AI\u0026#39;s AI infrastructure platform, APIs, and usage-based services\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advise on U.S. and international indirect tax matters, including sales tax and VAT/GST across key jurisdictions, with a focus on AI, cloud, and compute offerings\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Assess the tax implications of fixed-fee subscription or usage-based AI services\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Product, Finance, and Legal to structure enterprise contracts and bundled offerings in a tax-efficient and compliant manner, including input on pricing, invoicing, and go-to-market decisions\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Data Center and Infrastructure Operations\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Advise on tax treatment of data center operations, including leasing and colocation arrangements\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage sales and use tax on infrastructure and hardware purchases, including applying for tax exemption status across geographies\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Identify and support data center tax incentives and credits\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Oversee property tax exposure for data center assets\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Monitor tax developments impacting data center and infrastructure operations\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;International Expansion\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Support market entry into new jurisdictions, including entity structuring and transfer pricing\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and maintain transfer pricing policies and documentation\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Process \u0026amp;amp; Systems\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Build scalable tax processes and leverage automation where possible\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with various teams\u0026amp;nbsp; on tax provision processes and systems integration\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Bachelor\u0026#39;s degree in Accounting, Finance, or a related field required; CPA or Master\u0026#39;s in Taxation (MST) preferred\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;10+ years of progressive tax experience in public accounting and/or in-house roles, including experience building tax functions from the ground up\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong knowledge of U.S. and international direct and indirect tax laws, including transfer pricing and global structuring\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience in high-growth technology or AI infrastructure companies; familiarity with indirect tax in cloud-based and hardware-inclusive business models\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;IPO or mid-to-late-stage private company experience, with the ability to operate effectively in a fast-paced startup environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with emerging regulatory frameworks impacting AI infrastructure companies\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $259k - $310k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4046768007,"name":"Finance","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5098697007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5210729007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4667940007,"location":{"name":"San Francisco"},"metadata":null,"id":5210729007,"updated_at":"2026-08-13T12:57:11-04:00","requisition_id":"289","title":"Director of Technical Accounting","company_name":"Together AI","first_published":"2026-08-13T12:57:11-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re seeking an exceptional technical accounting leader to establish and scale our technical accounting function as we grow. This is a rare opportunity to define accounting policy and financial reporting standards across a high-growth, dynamic environment, partnering with a world-class team, while ensuring rigorous quarterly, annual, and ad hoc investor reporting as a private company.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Technical Accounting\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Develop and maintain U.S. GAAP policies across complex areas, including revenue recognition (ASC 606), leases (ASC 842), fixed assets (ASC 360), debt/financing (ASC 470/835), business combinations (ASC 805), equity/stock compensation (ASC 718, ASC 505), and collaborative arrangements (ASC 808)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Serve as subject matter expert on complex transactions, including structured financing, equipment financing, and customer contracts\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Revenue Accounting on revenue recognition for usage-based, subscription, and enterprise contracts, and evaluate the accounting treatment for collaborative arrangements, including co-development and joint go-to-market deals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Assist with lease accounting and capitalization policy for data center, colocation, and equipment arrangements\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advise on equity transactions (stock issuances, SAFEs, convertible instruments, stock-based compensation) and strategic investments/joint ventures\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Draft technical accounting memos for significant transactions and non-standard deal structures\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Monitor and implement new accounting standards\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Legal, Tax, FP\u0026amp;amp;A, Strategic Finance, and other teams on accounting implications of new contracts and transactions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support the monthly, quarterly, and annual close process for technical accounting areas, including review of complex and judgmental journal entries and account reconciliations\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Investor Reporting and Financial Statement Audit\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Lead quarterly, annual and ad hoc investor financial reporting packages\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with FP\u0026amp;amp;A, Legal, and Investor Relations to ensure clear, accurate disclosures\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support ad hoc reporting tied to fundraising, board reporting, or strategic transactions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lead preparation of annual consolidated financial statements and manage the annual audit process, including coordination with external auditors\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;SOX \u0026amp;amp; Internal Controls\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Help lay the foundation for the Company\u0026#39;s future SOX program, including initial risk assessment and control design\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop internal control processes to support long-term audit readiness\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with external auditors on control-related discussions and early readiness efforts\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Foster awareness of strong internal controls across Accounting, FP\u0026amp;amp;A, and other teams\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Process, Systems \u0026amp;amp; Ad Hoc Projects\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Build scalable technical accounting processes and documentation standards\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Core Accounting, FP\u0026amp;amp;A, and Systems teams on close processes and automation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lead or support ad hoc projects, including new initiatives and process improvements\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Bachelor\u0026#39;s degree in Accounting, Finance, or related field required; CPA required\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;12+ years of progressive accounting experience, ideally a mix of Big 4 and in-house, with experience building/scaling a technical accounting function\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep expertise across ASC 606, 842, 360, 470/835, 805, 718, 505, and 808\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with investor-facing reporting, ideally at a high-growth or late-stage private company\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience in high-growth tech or AI infrastructure preferred; familiarity with usage-based/subscription models, hardware-inclusive business models, and data center compute infrastructure a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfortable operating in a fast-paced, evolving private company environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong written communication skills, with experience drafting technical memos and investor-ready disclosures\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with emerging accounting issues for AI infrastructure companies (e.g., compute capitalization, data center leasing, partnership arrangements)\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $245k - $300k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4046768007,"name":"Finance","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5210729007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5199993007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4662728007,"location":{"name":"Singapore"},"metadata":null,"id":5199993007,"updated_at":"2026-08-13T12:58:17-04:00","requisition_id":"282","title":"Forward Deployed Engineer (Inference \u0026 Post-Training) - Mandarin Speaking","company_name":"Together AI","first_published":"2026-08-04T18:25:34-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Forward Deployed Engineer (FDE) focused on Inference \u0026amp;amp; Post-Training, you will be a hands-on technical partner to our most strategic customers — production AI teams looking to leverage high quality models and do inference at scale. For us, FDE is not a replacement for a Solutions Architect; you will partner with our SAs as a deep-domain specialist in inference optimization, fine-tuning pipelines, and production deployment. As key contributors to both the CX, Engineering, and Sales organizations, FDEs add tremendous value by ensuring we can meet the requirements of our most complex POCs, facilitate successful platform adoption, and guide tailored optimization efforts — directly impacting customer success, company growth, and the hardening of our core platform.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Must be a permanent resident or citizen of Singapore.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Inference Engine Optimization: Select, configure, and optimize inference engine based on hardware, model architecture, and workload profile\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Configuration \u0026amp;amp; Performance Tuning: Develop configuration updates to win critical POCs, benchmarks, and optimize customer deployments; tune KV cache, apply speculative decoding, determine optimal tensor parallelism, and determine quantization strategy to hit throughput and latency targets.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Post-Training \u0026amp;amp; Fine-Tuning: Drive hands-on RL training runs and optimize system design; guide customers through LoRA, SFT, DPO, RLHF, and GRPO pipelines from experimentation through production.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strategic Customer Alignment: Act as the primary technical point of contact for aligned strategic accounts — monitoring and optimizing endpoint configurations, helping customers get the most out of the platform, and collaborating to ensure we hit critical milestones.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Opinionated Onboarding: Establish direct alignment with strategic customers at onboarding; ensure the right inference and post-training configurations are in place from day one to improve time-to-value.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Product Feedback Loop: Directly influence our software and model roadmap by surfacing insights from the field. Contribute back to the product where needed to support customer requirements or drive a better experience. Drive early feature and research adoption with strategic logos.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;Requirements\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience: 5+ years in a technical role, with a strong focus on inference systems, open-source LLM deployment, or post-training workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Inference Engine Depth: Expert-level, hands-on experience with inference engines (e.g., vLLM, TensorRT-LLM, SGLang); ability to diagnose and resolve performance issues at the engine level.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Inference Optimization: Deep knowledge of KV cache tuning, speculative decoding, tensor parallelism, pipeline parallelism, and quantization techniques\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Post-Training Knowledge: Hands-on experience with fine-tuning and post-training pipelines, including LoRA, SFT, DPO, RLHF, and GRPO; ability to advise on system design\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Model Landscape Awareness: Broad knowledge of state-of-the-art open-source models and strong judgment on model selection for specific customer use cases, hardware profiles, and performance targets.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Coding Proficiency: Strong Python skills; comfortable working in production environments\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers on our journey in building the next generation of AI infrastructure.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5199993007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5178712007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4652312007,"location":{"name":"San Francisco"},"metadata":null,"id":5178712007,"updated_at":"2026-07-10T19:42:37-04:00","requisition_id":"272","title":"GTM Engineer","company_name":"Together AI","first_published":"2026-07-02T13:09:44-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We are seeking a skilled GTM Engineer to join our team. In this role, you will architect an AI-native tech stack with seamless integrations and build a cohesive, automated engine to drive measurable revenue outcomes. As a builder and systems thinker, you will play a vital role in automating processes, maintaining data integrity, and turning GTM strategy into scalable systems.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Architect the GTM Tech Stack: Build an AI native tech stack with seamless integrations.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Centralize Data: Build a cohesive, automated engine that centralizes data in CRM and eventually data warehouse (Salesforce, Hubspot, Snowflake, Clay, Claude Cowork).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own Data Integrity: Build and maintain the data flows, enrichment logic, and sync architecture that keep our systems accurate, deduped, and trustworthy end-to-end.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Turn Strategy into Systems: Partner with leadership to translate GTM strategy into automated playbooks, dynamic scoring models, and signal-based workflows that drive measurable revenue outcomes.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lead CPQ System Design: Lead system design for Salesforce CPQ, including product catalog architecture, pricing and discount rules, quote templates, and approval workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Optimize Conversions: Optimize quote-to-order conversion ensuring accurate data flow into downstream billing and ERP systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Implement Governance: Implement governance on pricing, discounting, and quote approval thresholds to maintain margin and compliance integrity.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Improve Deal Velocity: Improve deal velocity by streamlining approvals, leveraging Salesforce CPQ (Revenue Cloud).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Integrate CLM Tools: Integrate CLM (e.g., Ironclad, Conga, or DocuSign CLM) with Salesforce to automate quote-to-contract and contract-to-order workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Automate Bottlenecks: Identify repetitive work across the GTM org - account research, CRM hygiene, post-call follow-ups - and replace it with intelligent automation using APIs, webhooks, and AI agents.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience: Have 3–5+ years of experience operating in a Growth, GTM Engineer, or Revenue Operations role. Open to candidates that have a technical background that have worked with a GTM team in other capacities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Systems Thinking: Are a builder and systems thinker: You want to build moonshot ideas, not just toggle settings. You see the GTM stack as an interconnected machine, not a collection of point solutions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Salesforce Expertise: Have Direct experience with Salesforce — you know the data model, you\u0026#39;ve built on it, and you\u0026#39;re comfortable owning CRM-level changes.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;GTM Fluency: Are GTM-fluent: You understand the GTM funnel - not just the data model, but the logic behind every stage. You know what actually converts pipeline because you\u0026#39;ve worked in or alongside GTM teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Quote to Cash Configurations: Strong working knowledge of quoting tool configurations and CLM (pricing rules, product bundles, approval chains).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Execution \u0026amp;amp; Ownership: Constant iteration: Be a Kaizen methodology champion by building, shipping and iterating fast. You thrive in ambiguity, operate with high ownership, and move fast without waiting for perfect requirements. But you know when to slow down and ensure you’re shipping the highest quality work possible.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Technical Capability: Are technically capable: Comfortable working with APIs, webhooks, data integrations and can navigate technical documentation without hand-holding.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Preferred Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Prior experience at a fast-moving B2B SaaS / AI startup or strong familiarity with the AI tooling ecosystem.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience with Salesforce and Clay, bonus points for building data warehouses.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Prior quoting tool and CLM implementation experience.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing enrichment and outbounding process and managing complex data sourcing across multiple providers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience centralizing selling signals from multiple data sources into a single location.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;You have technical familiarity with GTM \u0026amp;amp; automation tools (Salesforce, HubSpot, CPQ, Ironclad, Clay, Gong, Apollo, Clari, APIs, webhooks, etc.).\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;As our ideal candidate, you will be part engineer, part hacker, with a passion for cybersecurity and a drive to protect our systems from evolving threats.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $160K-190K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4062691007,"name":"Revenue Operations","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5178712007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5171124007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4648733007,"location":{"name":"San Francisco"},"metadata":null,"id":5171124007,"updated_at":"2026-07-10T19:42:37-04:00","requisition_id":"S\u0026M-PTR-003_New","title":"Head of Hyperscaler Partnerships","company_name":"Together AI","first_published":"2026-06-26T14:19:02-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI-native cloud — the fastest inference infrastructure on the planet, paired with a model ecosystem, orchestration layer, and data center footprint that enterprises and frontier labs depend on. As we deepen our relationships with the world\u0026#39;s largest cloud platforms and technology ecosystems, we\u0026#39;re hiring a Head of Hyperscaler Partnerships to lead these deals in our partner portfolio.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a principal-level role for a seasoned deal-maker who has navigated some of the most complex partnership structures in enterprise technology — across model licensing, software integrations, inference and model serving, and scaled cloud distribution. You will sit at the intersection of commercial strategy, product, and finance, owning end-to-end partnership cycles with hyperscalers, neoclouds, and platform partners that shape how Together AI\u0026#39;s technology reaches the market.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You will bring deep experience inside a hyperscaler or have structured major deals with one. You understand how these organizations work from the inside — how decisions get made, which stakeholders matter, and how to unlock joint commercialization at scale. You will operate with significant autonomy, reporting into the VP of Strategic Partnerships and working closely with our CEO, CFO, CRO, and legal teams on deals that require board-level judgment.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities \u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own the Full Deal Cycle for Hyperscaler Partnerships:\u0026lt;/strong\u0026gt; Lead end-to-end partnership development with major cloud service providers. You will manage relationship-building through complex commercial negotiations, launch, and long-term expansion.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Navigate Complex Orgs with Precision:\u0026lt;/strong\u0026gt; Map and develop relationships across the full stakeholder matrix at partner organizations — from product and engineering to alliance managers, procurement, legal, and C-suite executives.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Drive Commercial Structures That Create Durable Value:\u0026lt;/strong\u0026gt; Design and negotiate deal structures across multiple surfaces, including revshare, marketplace private offers, and model licensing. You will partner with Finance on revenue modeling and with Legal on IP provisions and liability.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Be the Internal Champion for Partnership Strategy:\u0026lt;/strong\u0026gt; Translate partner needs and ecosystem signals into actionable internal recommendations. Collaborate with Partner Marketing on co-marketing initiatives and joint announcements, showcasing real-world enterprise value through joint case studies.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements \u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Strategic Experience:\u0026lt;/strong\u0026gt; 10+ years in strategic partnerships, business development, or alliance management, with meaningful time spent at a hyperscaler or structuring significant deals with them.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Deal-Making Track Record:\u0026lt;/strong\u0026gt; Proven ability to close complex, multi-surface agreements involving cloud distribution, model licensing, and inference infrastructure SLAs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Hyperscaler Fluency:\u0026lt;/strong\u0026gt; Deep understanding of marketplaces, private offers, consumption-based revenue models, and cloud commitment vehicles.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Technical Fluency:\u0026lt;/strong\u0026gt; Sufficient depth to engage with engineering on APIs, model serving, and software stacks.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;High EQ \u0026amp;amp; Collaborative:\u0026lt;/strong\u0026gt; Partner-first orientation with the ability to build trust quickly and find the \u0026quot;win-win\u0026quot; in competitive environments.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Preferred Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Direct experience with hyperscaler / CSP strategic deal mechanics\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience structuring partnerships across multiple product surface areas with differing priorities and stakeholders.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Background working with AI infrastructure, ideally on both compute and token / model serving side\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $ 300 - 350K OTE + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/h3\u0026gt;","departments":[{"id":4049151007,"name":"Sales","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5171124007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5214243007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4669618007,"location":{"name":"Amsterdam, Netherlands "},"metadata":null,"id":5214243007,"updated_at":"2026-08-20T11:34:34-04:00","requisition_id":"301","title":"HR Coordinator- Amsterdam","company_name":"Together AI","first_published":"2026-08-20T11:34:34-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026amp;nbsp;\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI\u0026#39;s Amsterdam office is growing — and so is our hiring across Europe and India. With several active searches underway and interview volume increasing, we\u0026#39;re building out local People support in Amsterdam. That\u0026#39;s where you come in.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a generalist role spanning three areas: recruiting coordination for our Europe and India hiring, HR administration (onboarding, offboarding, benefits, and special projects), and day-to-day office and workplace operations for our growing Amsterdam team and office. You\u0026#39;ll be the first person to hold this combined role in Amsterdam, which means you\u0026#39;ll have real ownership over how the work gets done — building process and structure, not just following someone else\u0026#39;s playbook.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll report to our Talent Operations Partner (based in San Francisco), which means you\u0026#39;ll operate with a high degree of day-to-day autonomy. For the office/workplace side of the role, you\u0026#39;ll also have a dedicated local mentor on our International Finance \u0026amp;amp; Operations team, so you\u0026#39;re never fully on your own. As our Amsterdam office continues to grow, there\u0026#39;s real potential for this role to grow with it.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Own end-to-end interview scheduling for our Europe and India hiring, working closely with the recruiting team\u0026amp;nbsp; to manage high interview volume across multiple time zones and seniority levels.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Manage offer letters and support onboarding and offboarding processes, handling sensitive and confidential information with care.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Own daily operations and employee support for our Amsterdam office, including office maintenance and food, beverage \u0026amp;amp; employee events.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Support other HR special projects as they come up.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Build and improve processes across all three areas — as the first dedicated person in this role, you\u0026#39;ll help define how this work gets done at Together AI.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Partner closely with our recruiting team, People team, and local Amsterdam colleagues to keep everything running smoothly as the office scales.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Proven experience coordinating high-volume interview scheduling across multiple teams and time zones, with minimal oversight — you know how to stay accurate and organized even when the volume is high.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; A genuine self-starter mindset — you\u0026#39;re comfortable building process where none exists yet, and you don\u0026#39;t need a lot of hand-holding to figure out how to move forward.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Strong attention to detail and time management, even under pressure and competing deadlines.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Sound judgment handling sensitive, confidential situations — things like offboarding conversations and offer letter details.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; An empathetic, approachable demeanor and a genuine interest in helping others.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Strong communication skills, and the flexibility to adapt in a fast-changing, high-growth environment.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Comfortable working onsite in our Amsterdam office four-five days a week.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Experience with applicant tracking systems or recruiting tools (e.g., Greenhouse, Gem, or similar) is a strong plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Prior exposure to HR administration or office/workplace operations\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026amp;nbsp; Dutch proficiency preferred, not required.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;em\u0026gt;If you don\u0026#39;t meet every requirement on this list but believe you\u0026#39;d be a great fit, we\u0026#39;d still love to hear from you.\u0026lt;/em\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033062007,"name":"Human Resources","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5214243007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5135876007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4631492007,"location":{"name":"San Francisco"},"metadata":null,"id":5135876007,"updated_at":"2026-07-10T19:42:36-04:00","requisition_id":"252","title":"Infrastructure Design Engineer","company_name":"Together AI","first_published":"2026-05-13T18:11:16-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About The Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building its infrastructure footprint at scale, and this role is central to making that happen. As an Infrastructure Design Engineer, you will own the design, planning, and technical execution of whitespace environments (where servers, storage, and network equipment are deployed) across our AI data center portfolio. You are the in-house expert who ensures that rack layouts, power distribution, cooling strategy, structured cabling, and physical infrastructure design are all built to support the density, redundancy, network and reliability requirements of large-scale AI GPU clusters.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You will serve as the lead engineer across our DC portfolio, creating white space designs, reviewing partner and contractor designs, and ensuring plans are executed to spec. You will work closely with the Infrastructure Strategy, Infrastructure Engineering, and Operations teams, as well as external MEP consultants, general contractors, and data center partners. This is a technical role on a small, high-accountability team where your judgment directly shapes our ability to bring capacity online on time and to spec.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Architect HPC clusters by designing whitespace layouts, including rack placement, aisle configuration, hot/cold aisle containment, equipment density, and airflow strategy for high-density GPU deployments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with electrical and mechanical engineers to integrate power and cooling infrastructure into whitespace environments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with Network Engineering to define and validate physical layer requirements (structured cabling, pathway planning, port density) for high-speed AI cluster interconnects, ensuring design compatibility with both physical and logical network architectures.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advise Data Center build teams/ contractors to ensure data center build out matches design and architecture specifications.\u0026amp;nbsp; Provide direction to optimize performance, scalability and cost-effectiveness\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop and maintain CAD/BIM drawings, schematics, capacity planning models, and technical documentation to support site design, construction, operations, and audits of data center white space\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own capacity planning within the white space through modeling that incorporates growth, utilization, and infrastructure scaling\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Validate electrical and mechanical load distribution across whitespace fit-outs, including review of striping plans and upstream load balancing under normal and failover conditions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Infrastructure Strategy and Operations by providing design review and technical guidance during the white space fit-out phase.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive continuous improvement in design standards, build processes, and quality assurance, including identifying opportunities to improve infrastructure reliability, reduce time-to-rack, and apply new technologies such as liquid cooling and high-density power architectures.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ensure all designs comply with applicable standards including TIA-942, Uptime Institute Tier guidelines, ASHRAE thermal recommendations, and relevant local codes and safety regulations.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ years of experience in data center design, critical facilities engineering, or infrastructure delivery, with deep technical knowledge across the full stack (power, cooling, network) as it pertains to white space design and implementation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience serving as an owner\u0026#39;s engineer or resident engineer, including reviewing consultant drawings, managing contractor compliance, and interpreting construction specifications and submittal documents\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong working knowledge of how facility electrical (UPS, switchgear) and cooling (CRAC/CRAH) infrastructures interface with, and support, high-density white space systems (PDUs, in-row cooling, CDUs, RPPs)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency with CAD/BIM tools such as AutoCAD, Revit, or equivalent; ability to produce and review technical drawings, schematics, and capacity models\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Working knowledge of industry standards including TIA-942, Uptime Institute Tier classifications, ASHRAE thermal guidelines, and ANSI/BICSI; familiarity with local building codes and safety regulations\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to manage multiple concurrent projects across different sites and work effectively with MEP engineers, general contractors, and equipment vendors in fast-moving environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Willingness to travel to data center sites as required for technical due diligence, design inspections, and final white space deployment sign-off\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience with liquid cooling systems for high-density GPU clusters, including direct liquid cooling (DLC), immersion cooling, or rear-door heat exchangers, and familiarity with associated power and plumbing infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Background in hyperscale or AI-native infrastructure deployments, including reference architecture validation, and the development of final design acceptance criteria at scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with DCIM platforms, telemetry and monitoring systems, and infrastructure-as-code tooling for capacity and utilization tracking\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an AI-native cloud company building the infrastructure to make AI faster, cheaper, and more accessible. We’re rapidly scaling our GPU footprint: signing our own data center leases, building large-scale clusters, and expanding toward a global owned-infrastructure presence. Our research team has contributed to breakthroughs like FlashAttention, Hyena, and RedPajama, and we co-design across software, hardware, and algorithms to push the frontier of AI efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $210-250K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4044336007,"name":"Business Operations","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5135876007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5200773007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4663125007,"location":{"name":"San Francisco"},"metadata":null,"id":5200773007,"updated_at":"2026-08-09T22:51:18-04:00","requisition_id":"284","title":"IT Engineer","company_name":"Together AI","first_published":"2026-08-03T16:00:23-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As an IT Engineer, you will be working across identity, devices, SaaS platforms, and supporting end users. You should be organized, collaborative, and eager to learn while working within established systems and processes.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Provide reliable, hands-on IT support for our in-office and remote employees by efficiently resolving Help Desk tickets while delivering excellent customer service.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support Together’s MDM fleet comprising macOS and some Linux devices.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Provide in-office support by troubleshooting basic network issues and our A/V systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Help manage and optimize Okta and our core SaaS tooling by administering and improving integrations and role-based access controls.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support daily procurement and asset management operations by coordinating hardware purchasing and tracking while maintaining accurate asset inventory and device assignments throughout their lifecycle (purchase to decommission).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Help maintain clear, organized, and accessible IT documentation by creating easy-to-understand technical documentation for both technical and non-technical audiences.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contribute to IT initiatives by supporting project execution, managing priorities, and ensuring clear communication with stakeholders through regular updates.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Help maintain a secure environment by supporting basic security hygiene to include access reviews and basic identity security tasks.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support onboarding and offboarding of Together employees, interns, and contractors.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Support the IT team by creating reliable automations that reduce manual work and improve operational efficiency.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborates well with a team, but also able to work autonomously when needed: working in tandem with other locations and having consistency in practices, documentation, asset mgmt.,etc.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Practical experience administering Okta, including users, groups, and basic app integrations (Terraform a plus).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong knowledge of Google Workspace.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on MDM experience with macOS, ideally with FleetDM.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to support macOS, Linux, iOS, and Android devices.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience participating in procurement and asset management processes.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Clear, concise technical writing skills for documentation and knowledge base content.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and customer service abilities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Basic understanding of security principles and best practices.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with low‑code automation tools (Serval, Okta Workflows, Gumloop).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Scripting familiarity with common languages (Bash, Python).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Understanding of API integrations and webhooks and how to use them effectively.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with AWS provisioning and Identity Center\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Preferred Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Exposure to security and compliance frameworks (SOC 2, ISO 27001, NIST, CIS, FedRAMP).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with AI tools, agents, MCPs, automations, and prompting.\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with Notion and Linear for documentation and project management.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;HRIS integration experience with systems like HiBob and Workday.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $140K - $220k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5200773007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5090473007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4610092007,"location":{"name":"Amsterdam "},"metadata":null,"id":5090473007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"271","title":"Lead/Manager Site Reliability Engineering Team (Amsterdam) ","company_name":"Together AI","first_published":"2026-04-01T04:22:42-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Lead a team of Site Reliability Engineer (SRE) at Together based out of our office in Amsterdam, you \u0026amp;nbsp;and the SRE team are responsible for keeping all user-facing services and production systems running smoothly. You are a blend of a pragmatic operator and a software engineer that applies sound engineering principles, operational discipline, and mature automation to our operating environments and codebase.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You specialize in systems (operating systems, storage subsystems, networking), while implementing best practices for availability, reliability and scalability, with varied interests in algorithms and distributed systems.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Be on an on-call (PagerDuty) rotation to respond to incidents that impact availability\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage, develop and coach the SRE Team.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and run our infrastructure with Ansible, Terraform, and Kubernetes to enable scaling to a massive number of concurrent users\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build monitoring systems to ensure the highest quality service for our customers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and implement operational processes (such as deployments and upgrades)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Debug production issues across all services and levels of the stack\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Identify improvements for the product architecture from the reliability, performance and availability perspectives\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Plan the growth of Together AI’s infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ years of professional SRE or related experience\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ideally 2 years as a Lead SRE\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor\u0026#39;s degree in Computer Science or a related field or equivalent work experience\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Expert knowledge of Ansible (roles, playbooks), Terraform, and Kubernetes\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in programming/scripting languages\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Direct experience in monitoring and observability practices\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advanced knowledge of cloud services\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to thrive in a collaborative environment involving different stakeholders and subject matter experts\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029319007,"name":"Europe","location":"Amsterdam, North Holland, Netherlands","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5090473007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5145183007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4636050007,"location":{"name":"Amsterdam"},"metadata":null,"id":5145183007,"updated_at":"2026-07-10T19:42:36-04:00","requisition_id":"257","title":"Lead/Manager Together Cloud Infrastructure ","company_name":"Together AI","first_published":"2026-06-02T06:05:37-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As a Lead/Manager, you will play a key role in building the Together cloud platform engineering team in the Netherlands.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We are a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Some of what you’ll work on:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Work on a distributed GPU scheduling system for the on-demand clusters product, Instant Clusters.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build out a global management plane for managing our data center compute, networking, and storage.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and build new customer-facing cloud platform services, delivering killer enterprise AI cloud features.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Hybrid working 2 days a week at our offices in Amsterdam\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Lead/Manage a team of 8 together cloud Infrastructure Engineer in Amsterdam,\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Identify, design, and develop foundational backend services that power Together’s commerce platform\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Analyze and improve the robustness and scalability of existing distributed systems, APIs, databases, and infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with product teams to understand functional requirements and deliver solutions that meet business needs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Write clear, well-tested, and maintainable software and IaC for both new and existing systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Conduct design and code reviews, create developer documentation, and develop testing strategies for robustness and fault tolerance\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Participate in an on-call rotation to address critical incidents when necessary\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026amp;nbsp;\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Ideally 1-2 years of leading the Infrastructure team.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrable background in acquiring talent and retaining talent.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;5+ years of demonstrated experience in building large scale, fault tolerant, distributed systems and API microservices\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing, analyzing and improving efficiency, scalability, and stability of various system resources\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication skills – able to write clear design docs and work effectively with both technical and non-technical team members\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience with building and operating high-performance and/or globally distributed microservice architectures across one or more cloud providers (AWS, Azure, GCP)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems knowledge across compute, networking, and storage, including concurrency, memory management, performant I/O, and scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience developing against and managing a relational database, such as PostgreSQL\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Expert-level programmer in one or more of programming language (Golang preferred)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in version control practices and integrating IaC with CI/CD pipelines.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Kubernetes and containers preferred\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building and operating data infrastructure (Kinesis, Airflow, Kafka, etc) a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical field, or equivalent practical experience\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4051033007,"name":"Amsterdam","location":"Amsterdam, North Holland, Netherlands","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5145183007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5062829007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4597662007,"location":{"name":"San Francisco"},"metadata":null,"id":5062829007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"R\u0026D-PROD-024","title":"Lead Product Designer","company_name":"Together AI","first_published":"2026-02-26T14:25:29-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;The product team is seeking a Lead Product Designer with advanced expertise in crafting exceptional user experiences for technical tools. In this strategic role, you\u0026#39;ll shape AI development tools while laying the foundation for our growing design organization.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You bring a proven ability to lead complex UX initiatives and elevate design quality across complex product surfaces. Your technical design sensibility will elevate tools used daily by AI engineers. And your leadership experience will help establish processes and standards for future team growth.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You are a natural collaborator and an excellent communicator, able to develop and present design ideas in a larger cross-functional team context: Engineering, Product, and Marketing. You thrive in highly collaborative environments and work closely with Engineering and Product to ensure strong execution quality. As part of a scaling product team, We hold a high bar for craft across interaction, visual design, usability, and systems thinking, and expect this role to raise that bar across the product.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Naturally, you\u0026#39;ll be able to make key strategic decisions with us, besides your everyday work: Craft and map elegant user flows, prototype smooth interactions, launch new product features. You will also contribute to and evolve our design system to ensure consistency and scalability across surfaces.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Ideally, you are biased toward long-term and keep your eyes on priorities, are used to self-managing projects you take ownership over from start to finish and know how to involve others outside your role. You are a fast learner who can quickly ramp in complex technical domains and independently build context around AI and developer workflows.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ of experience in a professional software product-driven environment.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advanced skills in UX and interaction design\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building developer tools, IDEs, infrastructure platforms, or data products\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience building, scaling, or significantly contributing to a production design system, including component architecture and cross-product consistency\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experienced in user research\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Technical enough to learn about the needs of our users\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Can consume and interpret user interaction data\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing complex software used daily by technical users\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Love thinking through complex interaction design problems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Get excited about being immersed in the AI ecosystem and defining how tooling for AI services should look and behave\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Enjoy tackling ambiguous problems and shaping them into a clear vision\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and storytelling skills.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Creative problem-solving. Preferably with a systematic, long-term approach.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with data-informed design.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;High standards for visual, interaction, and product craftsmanship\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to ramp quickly in new technical environments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing enterprise software\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive design vision for AI development tooling ecosystem\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with exec team on strategic product decisions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop scalable design practices for team growth\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Foster cross-functional collaboration at senior levels\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Champion user-centered culture across organization\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Set and maintain a high bar for design quality and craft across teams\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Promote consistent design standards and systems thinking across products\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Contribute to the overall design and direction of Together products\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and ship high-quality products and improvements, from early concepts to high-fidelity prototypes and visuals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner closely with engineering, product management, AI research, and design peers to define both long-term strategy and short-term tactics\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Engage in user research to better understand our users and refine our products\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contribute to and evolve our design system\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own and improve shared components within the design system in collaboration with Engineering\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ensure strong implementation quality by partnering closely with Engineering throughout delivery\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain consistency and cohesion across product surfaces\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, and other competitive benefits. The base salary range is $200,000 - $240,000. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033061007,"name":"Product","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5062829007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4687884007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4402726007,"location":{"name":"San Francisco, Singapore, Amsterdam "},"metadata":null,"id":4687884007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"116","title":"LLM Inference Frameworks and Optimization Engineer","company_name":"Together AI","first_published":"2025-03-25T18:23:25-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;At Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize inference frameworks, algorithms, and infrastructure, pushing the boundaries of performance, scalability, and cost-efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We are seeking an\u0026lt;strong\u0026gt; \u0026lt;/strong\u0026gt;Inference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines that support multimodal and language models at scale. This role will focus on low-latency, high-throughput inference, GPU/accelerator optimizations, and software-hardware co-design, ensuring efficient large-scale deployment of LLMs and vision models.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This role offers a unique opportunity to shape the future of LLM inference infrastructure, ensuring scalable, high-performance AI deployment across a diverse range of applications. If you\u0026#39;re passionate about pushing the boundaries of AI inference, we’d love to hear from you!\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;h4\u0026gt;\u0026lt;strong\u0026gt;Inference Framework Development and Optimization\u0026lt;/strong\u0026gt;\u0026lt;/h4\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and develop fault-tolerant, high-concurrency distributed inference engine for text, image, and multimodal generation models.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Implement and optimize distributed inference strategies, including Mixture of Experts (MoE) parallelism, tensor parallelism, pipeline parallelism for high-performance serving.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Apply CUDA graph optimizations, TensorRT/TRT-LLM graph optimizations, and PyTorch-based compilation (torch.compile), and speculative decoding to enhance efficiency and scalability.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h4\u0026gt;\u0026lt;strong\u0026gt;Software-Hardware Co-Design and AI Infrastructure\u0026lt;/strong\u0026gt;\u0026lt;/h4\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with hardware teams on performance bottleneck analysis, co-optimize inference performance for GPUs, TPUs, or custom accelerators.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work closely with AI researchers and infrastructure engineers to develop efficient model execution plans and optimize E2E model serving pipelines.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Must-Have:\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Experience:\u0026lt;/strong\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience in deep learning inference frameworks, distributed systems, or high-performance computing.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Technical Skills:\u0026lt;/strong\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Familiar with at least one LLM inference frameworks (e.g., TensorRT-LLM, vLLM, SGLang, TGI(Text Generation Inference)).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Background knowledge and experience in at least one of the following: GPU programming (CUDA/Triton/TensorRT), compiler, model quantization, and GPU cluster scheduling.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep understanding of \u0026lt;span class=\u0026quot;notion-enable-hover\u0026quot; data-token-index=\u0026quot;1\u0026quot;\u0026gt;KV cache\u0026lt;/span\u0026gt; systems like \u0026lt;span class=\u0026quot;notion-enable-hover\u0026quot; data-token-index=\u0026quot;3\u0026quot;\u0026gt;Mooncake\u0026lt;/span\u0026gt;, \u0026lt;span class=\u0026quot;notion-enable-hover\u0026quot; data-token-index=\u0026quot;5\u0026quot;\u0026gt;PagedAttention\u0026lt;/span\u0026gt;, or custom in-house variants.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Programming:\u0026lt;/strong\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Proficient in Python and C++/CUDA for high-performance deep learning inference.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Optimization Techniques:\u0026lt;/strong\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Deep understanding of Transformer architectures and LLM/VLM/Diffusion model optimization.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Knowledge of inference optimization, such as workload scheduling, CUDA graph, compiled, efficient kernels\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Soft Skills:\u0026lt;/strong\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong analytical problem-solving skills with a performance-driven mindset.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent collaboration and communication skills across teams.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Nice-to-Have:\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience in developing software systems for large-scale data center networks with RDMA/RoCE\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiar with distributed filesystem(e.g., 3FS, HDFS, Ceph)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiar with open source distributed scheduling/orchestration frameworks, such as Kubernetes (K8S)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contributions to open-source deep learning inference projects.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $160,000 - $230,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4687884007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4385540007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4245330007,"location":{"name":"San Francisco "},"metadata":null,"id":4385540007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"32","title":"Machine Learning Engineer - Inference","company_name":"Together AI","first_published":"2024-06-06T15:12:19-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is seeking a Machine Learning Engineer to join our\u0026lt;strong\u0026gt; \u0026lt;/strong\u0026gt;Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and ensuring they run efficiently and effectively at scale. If you are passionate about AI inference, PyTorch, and developing high-performance systems, we want to hear from you. This position offers the chance to collaborate closely with AI researchers and engineers to create cutting-edge AI solutions. Join us in shaping the future at Together AI!\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and build the production systems that power the Together AI inference engine, enabling reliability and performance at scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop and optimize runtime inference services for large-scale AI applications.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with researchers, engineers, product managers, and designers to bring new features and research capabilities to the world.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Conduct design and code reviews to ensure high standards of quality.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create services, tools, and developer documentation to support the inference engine.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Implement robust and fault-tolerant systems for data ingestion and processing.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience writing high-performance, well-tested, production-quality code.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency with Python and PyTorch.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience in building high performance libraries and tooling.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent understanding of low-level operating systems concepts including multi-threading, memory management, networking, storage, performance, and scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Preferred: Knowledge of existing AI inference systems such as TGI, vLLM, TensorRT-LLM, Optimum\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Preferred: Knowledge of AI inference techniques such as speculative decoding.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Preferred: Knowledge of CUDA/Triton programming.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Nice to have: Knowledge of Rust, Cython and compilers.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society. Together, we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI. Our team has been behind technological advancements such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey to build the next-generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other competitive benefits. The US base salary range for this full-time position is $160,000 - $230,000 + equity + benefits. Our salary ranges are determined by location, level, and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunities to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;div\u0026gt;\n\u0026lt;div class=\u0026quot;job__description body\u0026quot;\u0026gt;\n\u0026lt;div\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;/div\u0026gt;\n\u0026lt;/div\u0026gt;\n\u0026lt;/div\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4385540007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5211627007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4668385007,"location":{"name":"San Francisco or NYC"},"metadata":null,"id":5211627007,"updated_at":"2026-08-18T20:02:43-04:00","requisition_id":"295","title":"Manager, International Cloud Sourcing","company_name":"Together AI","first_published":"2026-08-18T19:57:54-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is looking for a Manager to own the commercial relationships with our international cloud compute suppliers and the market intelligence that drives our compute supply strategy. You will manage the lifecycle of international cloud service provider (CSP) contracts and renewals, run RFP processes for new capacity, build pricing benchmarks across the global GPU cloud landscape, and partner with infrastructure engineering, finance, and GTM to secure the right capacity at the best economics as we rapidly scale our GPU fleet globally.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Manage the lifecycle of international CSP leasing agreements, from sourcing and RFPs through contract execution and renewals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with infrastructure engineering to translate technical requirements into commercial specifications, and coordinate technical evaluations during vendor selection\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop and maintain relationships with cloud compute suppliers to expand Together\u0026#39;s vendor ecosystem into new international markets\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and maintain market intelligence to stay ahead of pricing trends, supply availability, and the evolving international cloud compute vendor landscape\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop negotiation strategies that improve our cost position and terms as we scale, accounting for local contracting norms and market dynamics\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive continuous improvement on Total Cost of Ownership (TCO) and actively mitigate supply chain and commercial risks across the GPU fleet\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish and lead regular, structured vendor performance reviews, including Quarterly Business Reviews (QBRs), with international cloud and infrastructure compute suppliers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Track the country-level factors that shape international capacity decisions, including power availability, local regulatory and data residency requirements, currency, and tax and duty implications\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years in cloud sourcing, strategic sourcing, or infrastructure commercial roles, with exposure to GPU or high-performance compute environments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Track record negotiating cloud or infrastructure contracts (multi-million dollar deal sizes), including with suppliers outside the US\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Working knowledge of cloud infrastructure, including GPU configurations, networking, and cluster architecture, with the ability to translate technical requirements into commercial terms\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong analytical skills; experience building pricing models or financial analyses that inform sourcing decisions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication skills; comfortable working across engineering, finance, and vendor executives, including international suppliers, and available for calls outside standard US business hours\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to apply FinOps principles (e.g., cost allocation, rightsizing, commitment strategy) to cloud sourcing and contract negotiations to drive demonstrable cost efficiency across the GPU fleet.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience in GPU/AI compute sourcing specifically, or at a GPU cloud provider or hyperscaler\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with the AI infrastructure vendor ecosystem and current supply/demand dynamics, particularly outside the US\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building sourcing processes from scratch rather than operating within established systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with data center economics, including power, cooling, and site selection considerations, and with country-level factors such as data residency, currency, and duties\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an AI-native cloud company building the infrastructure to make AI faster, cheaper, and more accessible. We’re rapidly scaling our GPU footprint: signing our own data center leases, building large-scale clusters, and expanding toward a global owned-infrastructure presence. Our research team has contributed to breakthroughs like FlashAttention, Hyena, and RedPajama, and we co-design across software, hardware, and algorithms to push the frontier of AI efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $190-220K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our Privacy Policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4044336007,"name":"Business Operations","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5211627007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4790243007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4455044007,"location":{"name":"San Francisco "},"metadata":null,"id":4790243007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"140","title":"Platform Engineer, Model Shaping","company_name":"Together AI","first_published":"2025-07-16T15:44:06-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;The Model Shaping team at Together AI works on products and research for tailoring open foundation models to downstream applications. We build services that allow machine learning developers to choose the best models for their tasks and further improve these models using domain-specific data. In addition to that, we develop new methods for more efficient model training and evaluation, drawing inspiration from a broad spectrum of ideas across machine learning, natural language processing, and ML systems.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As a Platform Engineer in Model Shaping, you will work at the intersection of backend engineering and infrastructure, building the foundational layers of Together’s platform for model customization and evaluation. You will design, develop, and operate both the backend services and the underlying systems that enable us to sustainably and reliably scale production workflows launched by our users, as well as internal research experiments.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You will operate in a cross-functional environment, collaborating with other engineers and researchers in the team to improve the infrastructure based on the needs of projects they work on. You will also interact with other engineering teams at Together (such as Commerce, Data Engineering, and Cloud Infrastructure) to integrate the services developed by Model Shaping with systems developed by those teams.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and build Together’s systems and infrastructure for model customization, including user-facing features and internal improvements\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contribute to reliability improvements for the platform, participating in an on-call rotation and improving processes for incident response\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create and improve internal tooling for deployment, continuous integration, and observability\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build a job orchestration platform spanning multiple datacenters, supporting a highly heterogeneous hardware landscape\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with teams developing internal services, co-designing these services and incorporating them in systems built within Together\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience in building infrastructure or backend components of production services\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Extensive experience designing, operating, and troubleshooting production Linux environments and Kubernetes-based platforms\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering background in Python or Go\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experienced with infrastructure automation tools (Terraform, Ansible), monitoring/observability stacks (Prometheus, Grafana), and CI/CD pipelines (GitHub Actions, ArgoCD)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Cloud environment (e.g., AWS/GCP/Azure) administration experience, preferably with a hybrid bare-metal/cloud environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong communication skills, be willing to document systems and processes and collaborate with peers of varying technical expertise\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfortable operating across the stack, from cluster operations and infrastructure automation to backend service development\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Experience in any of the following will make you stand out:\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Developing large-scale production systems with high reliability requirements\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Pipeline orchestration frameworks (e.g., Kubeflow, Argo Workflows, Flyte)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Managing GPU workloads on HPC clusters, ideally with hands-on experience in operating NVIDIA’s networking stack (e.g., NCCL, Mellanox firmware, GPUDirect RDMA)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deployment of services for AI training or inference\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Networking fundamentals, including TCP/IP, DNS, routing, load balancing, TLS, and network debugging tools\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintaining or contributing to open-source projects\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, RedPajama, SWARM Parallelism, and SpecExec. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is $200,000 - $290,000. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4790243007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5172169007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4649241007,"location":{"name":"San Francisco"},"metadata":null,"id":5172169007,"updated_at":"2026-07-10T19:42:37-04:00","requisition_id":"270","title":"Product Manager, AI Infrastructure","company_name":"Together AI","first_published":"2026-06-24T17:23:11-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About the Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Our product surface is expanding fast - GPU clusters, managed storage, networking, and observability - and we\u0026#39;re adding a Product Manager to the Together Cloud team to own the day-to-day product work that keeps these AI infrastructure products moving. You\u0026#39;ll start across the full surface, with an early focus on observability and GPU Clusters, partnering closely with engineering to ship the feature work that customers feel every day.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a role for someone who runs toward problems. You won\u0026#39;t be handed a backlog and asked to coordinate it - you\u0026#39;ll find the issues others haven\u0026#39;t spotted yet, drive them to resolution, and use data and experimentation to decide what to build. You\u0026#39;ll have real autonomy from day one, and the breadth of working across compute, storage, and observability that few PM roles offer.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;It\u0026#39;s also a role with a clear path. Within roughly nine months, the goal is for you to grow into full ownership of a complete product area — observability or storage as your own. You\u0026#39;ll report directly to a Staff Product Manager, and our product leadership (including our CPO) is closely involved with this team. If you want to build infrastructure that the AI ecosystem runs on, and earn ownership quickly by proving you can operate, this is that seat.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;Responsibilities\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the day-to-day product and feature work across Together\u0026#39;s AI infrastructure products - GPU Clusters, Managed Storage, and observability - with an early focus on observability and storage.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Find and drive problems to resolution with minimal guidance, including issues that aren\u0026#39;t yet on anyone\u0026#39;s radar.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run structured, hypothesis-driven experiments - reading the data yourself and driving the data collection and instrumentation needed to answer open questions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner across engineering, product, and partner teams to ship improvements and unblock work without waiting for permission.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Translate a deep understanding of customers operating in fast-moving, high-ambiguity markets into product decisions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Juggle multiple priorities and workstreams across a broad infrastructure surface.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Grow into end-to-end ownership of a complete product area (observability or storage) as you ramp.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;Requirements\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Deep empathy for customers operating in high-ambiguity, fast-evolving markets; we work with AI Native startups and model labs that move at high pace\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong, hands-on data and analytics skills with a hypothesis-driven approach - you design experiments, read the data yourself, and can drive the data collection needed to get answers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;High agency: you find problems, push through resistance, and drive resolution without waiting to be told what to do.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A real technical foundation -whether from an AI/infrastructure background, hands-on software development, or strong working knowledge of cloud and infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A track record of collaborating across orgs and product lines to get things done.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfortable managing multiple priorities and workstreams at once.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with the AI stack, especially at the infrastructure layer, is a strong plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Cloud platform experience (AWS, Azure, GCP, or similar) - with extra weight if you\u0026#39;ve helped \u0026lt;em\u0026gt;build\u0026lt;/em\u0026gt; such products.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bonus: working knowledge of the Nvidia/AMD GPU stack (InfiniBand, NCCL, GPU operator, etc.), or experience with AI training or inference workloads.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $ 175,000 - 220,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033061007,"name":"Product","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5172169007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4970113007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4551961007,"location":{"name":"San Francisco"},"metadata":null,"id":4970113007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"176","title":"Product Marketing Director","company_name":"Together AI","first_published":"2025-11-05T14:09:54-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About the Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a frontier AI cloud, which has been built bottoms up to cater to the demand for the new generation of AI applications and agents. The company has seen tremendous growth with 20X customer growth and 6X ARR growth in the financial year.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As we continue to drive product innovation, we are also investing deeply in GTM. We are looking for a product marketing leader to continue to build and scale our PMM function. This role will own the platform as well as all product level value propositions and define how the messaging flows downstream across all channels. They will partner closely with the product management team to build and execute our product launch calendar and GTM plans to deliver adoption and user growth for our key products. This role will report into the head of marketing and is expected to lead our current PMM organization and continue to build a bar-raising PMM function.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Build and maintain detailed buyer personas and ideal customer profiles to guide segmentation, messaging, and campaign strategies.​\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop compelling product positioning and messaging that clearly differentiate us in a competitive landscape\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner closely with Product Management to influence roadmap priorities based on market insights, customer feedback, and competitive analysis​\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own the go-to-market strategy for new product launches and major updates, managing the cross-functional coordination needed for success.​\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lead creation of sales enablement tools—pitch decks, battlecards, and case studies—to empower sales and customer success teams.​\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive consistent storytelling across all customer touchpoints—website, campaigns, and events\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with PR, demand generation, field marketing, and web teams to ensure alignment between GTM campaigns and core value propositions.​\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage, coach and scale a bar-raising team of product marketers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Set up, measure and report on key OKRs for the PMM function\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026amp;nbsp;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;10+ years of PMM experience in enterprise software, preferably in AI, AI natives, Digital Natives or Cloud\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;5+ years as a team leader in the PMM function\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven success leading and scaling high-performing product marketing teams in fast-paced growth environments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong ability to translate complex technical features into business-oriented messaging for diverse audiences\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience building GTM strategies, launching new products, and achieving measurable awareness, adoption or pipeline growth​\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfortable operating cross-functionally with Sales, Product, and Engineering to align market strategy with execution\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Exceptional analytical skills with a data-driven approach to decision-making and reporting\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor’s degree in engineering and MBA preferre\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $250-295k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. This is a hybrid role based in the Bay Area.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062690007,"name":"Marketing","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4970113007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4384627007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4244909007,"location":{"name":"San Francisco"},"metadata":null,"id":4384627007,"updated_at":"2026-07-10T19:42:33-04:00","requisition_id":"30","title":"Research Engineer, Core ML","company_name":"Together AI","first_published":"2026-02-18T12:09:06-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;This is a research engineering role with direct production impact. You won’t be publishing ideas in isolation—you will translate new RL algorithms, scheduling methods, and inference optimizations into production-grade systems that power Together’s API. Success in this role means shipping measurable improvements in latency, throughput, cost, and model quality at scale. We are looking for researchers who enjoy owning systems end-to-end and turning frontier ideas into robust infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;The Core ML (Turbo) at Together AI team sits at the intersection of efficient inference (algorithms, architectures, engines) and post‑training / RL systems. We build and operate the systems behind Together’s API, including high‑performance inference and RL/post‑training engines that can run at production scale.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Our mandate is to push the frontier of efficient inference and RL‑driven training: making models dramatically faster and cheaper to run, while improving their capabilities through RL‑based post‑training (e.g., GRPO‑style objectives). This work lives at the interface of algorithms and systems: asynchronous RL, rollout collection, scheduling, and batching all interact with engine design, creating many knobs to tune across the RL algorithm, training loop, and inference stack. Much of the job is modifying production inference systems—for example, SGLang‑ or vLLM‑style serving stacks and speculative decoding systems such as ATLAS—grounded in a strong understanding of post‑training and inference theory, rather than purely theoretical algorithm design.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You’ll work across the stack—from RL algorithms and training engines to kernels and serving systems—to build and improve frontier models via RL pipelines. People on this team are often spiky: some are more RL‑first, some are more systems‑first. Depth in one of these areas plus appetite to collaborate across (and grow toward more full‑stack ownership over time) is ideal.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Advance inference efficiency end‑to‑end\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and prototype algorithms, architectures, and scheduling strategies for low‑latency, high‑throughput inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Implement and maintain changes in high‑performance inference engines (e.g., SGLang‑ or vLLM‑style systems and Together’s inference stack), including kernel backends, speculative decoding (e.g., ATLAS), quantization, etc.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Profile and optimize performance across GPU, networking, and memory layers to improve latency, throughput, and cost.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Unify inference with RL / post‑training\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and operate RL and post‑training pipelines (e.g., RLHF, RLAIF, GRPO, DPO‑style methods, reward modeling) where 90+% of the cost is inference, jointly optimizing algorithms and systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Make RL and post‑training workloads more efficient with inference‑aware training loops—for example, async RL rollouts, speculative decoding, and other techniques that make large‑scale rollout collection and evaluation cheaper.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Use these pipelines to train, evaluate, and iterate on frontier models on top of our inference stack.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Co‑design algorithms and infrastructure so that objectives, rollout collection, and evaluation are tightly coupled to efficient inference, and quickly identify bottlenecks across the training engine, inference engine, data pipeline, and user‑facing layers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run ablations and scale‑up experiments to understand trade‑offs between model quality, latency, throughput, and cost, and feed these insights back into model, RL, and system design.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own critical systems at production scale\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Profile, debug, and optimize inference and post-training services under real production workloads, taking research ideas all the way to stable, measurable improvements in deployed systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive roadmap items that require real engine modification—changing kernels, memory layouts, scheduling logic, and APIs as needed.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish metrics, benchmarks, and experimentation frameworks to validate improvements rigorously.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Provide technical leadership (Staff level)\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Set technical direction for cross‑team efforts at the intersection of inference, RL, and post‑training.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Mentor other engineers and researchers on full‑stack ML systems work and performance engineering.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We don’t expect anyone to check every box below. People on this team typically have deep expertise in one or more areas and enough breadth (or interest) to work effectively across the stack. The closer you are to full‑stack (inference + post‑training/RL + systems), the stronger the fit—but being spiky in one area and eager to grow is absolutely okay.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You might be a good fit if you:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Have a bias toward implementation and shipping\u0026lt;/strong\u0026gt;—you are excited to modify real engines and services, not just prototype in research code.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Have strong expertise in at least one of the following, and are excited to collaborate across (and grow into) the others:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Systems‑first profile: Large‑scale inference systems (e.g., SGLang, vLLM, FasterTransformer, TensorRT, custom engines, or similar), GPU performance, distributed serving.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;RL‑first profile: RL / post‑training for LLMs or large models (e.g., GRPO, RLHF/RLAIF, DPO‑like methods, reward modeling), and using these to train or fine‑tune real models.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Model architecture design for Transformers or other large neural nets.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Distributed systems / high‑performance computing for ML.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Are comfortable working from algorithms to engines:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong coding ability in Python\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience profiling and optimizing performance across GPU, networking, and memory layers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Able to take a new sampling method, scheduler, or RL update and turn it into a production‑grade implementation in the engine and/or training stack.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Have a solid research foundation in your area(s) of depth:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Track record of impactful work in ML systems, RL, or large‑scale model training (papers, open‑source projects, or production systems).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Can read new RL / post‑training papers, understand their implications on the stack, and design minimal, correct changes in the right layer (training engine vs. inference engine vs. data / API).\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Operate well as a full‑stack problem solver:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;You naturally ask: “Where in the stack is this really bottlenecked?”\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;You enjoy collaborating with infra, research, and product teams, and you care about both scientific quality and user‑visible wins.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Minimum qualifications\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience working on ML systems, large‑scale model training, inference, or adjacent areas (or equivalent experience via research / open source).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advanced degree in Computer Science, EE, or a related field, or equivalent practical experience.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience owning complex technical projects end‑to‑end.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;If you’re excited about the role and strong in some of these areas, we encourage you to apply even if you don’t meet every single requirement.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $200,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4384627007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5199554007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4662516007,"location":{"name":"San Francisco"},"metadata":null,"id":5199554007,"updated_at":"2026-07-30T12:52:05-04:00","requisition_id":"281","title":"Research Engineer, Large-Scale Training","company_name":"Together AI","first_published":"2026-07-30T12:52:05-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;About the Role\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p class=\u0026quot;isSelectedEnd\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;The Model Shaping team at Together AI works on products and research for tailoring open foundation models to downstream applications. We build services that allow machine learning developers to choose the best models for their tasks and further improve these models using domain-specific data. In addition, we develop new methods for more efficient model training and evaluation, drawing inspiration from a broad spectrum of ideas across machine learning, natural language processing, and ML systems.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p class=\u0026quot;isSelectedEnd\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;As a Research Engineer on the Scaling Team within Model Shaping, you will turn cutting-edge research on efficient foundation model training into robust, high-performance systems. You will profile and optimize Together\u0026#39;s training infrastructure, identify performance bottlenecks across the stack, and implement state-of-the-art techniques from both the research literature and our own scientists in production environments.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p class=\u0026quot;isSelectedEnd\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Your work will directly shape the fine-tuning experience of Together\u0026#39;s customers. You will rapidly bring newly released open-source models onto the Model Shaping platform, ensuring they train efficiently and reliably across diverse customer workloads. Working closely with Research Scientists, you will also build the experimental infrastructure that accelerates research and enables validated ideas to be deployed reliably at scale.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Responsibilities\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul data-spread=\u0026quot;false\u0026quot;\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Design, implement, and optimize core components of Together\u0026#39;s large-scale training infrastructure.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Integrate new model architectures, validate training correctness and convergence, and optimize performance for production fine-tuning workloads.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Profile distributed training workloads to identify and eliminate bottlenecks across compute, memory, and communication.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Design and execute experiments to validate performance hypotheses and benchmark new approaches against state-of-the-art methods.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Partner closely with Research Scientists to productionize novel training methods and contribute to publications and open-source releases.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Rapidly enable support for newly released open-source foundation models on the Together platform.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Build and maintain experimental infrastructure that accelerates research while ensuring production-quality reliability and scalability.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Requirements\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul data-spread=\u0026quot;false\u0026quot;\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Demonstrated ability to independently take ambiguous performance or infrastructure problems from investigation through deployment.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong programming skills in Python and PyTorch, with an emphasis on writing efficient, maintainable code.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Hands-on experience training or fine-tuning large neural networks in multi-GPU or multi-node environments.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Solid understanding of ML systems fundamentals, including GPU architecture, mixed-precision training, and distributed training paradigms such as data, tensor, pipeline, or expert parallelism.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong communication skills and the ability to collaborate effectively with both researchers and engineers.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Passion for staying current with advances in AI research and applying them to real-world systems.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Excitement about translating cutting-edge research into production systems that deliver customer impact.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3 class=\u0026quot;isSelectedEnd\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul data-spread=\u0026quot;false\u0026quot;\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience writing optimized NVIDIA GPU kernels using CUDA or Triton, or implementing communication collectives with technologies such as NCCL or NVSHMEM.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience with large-scale training frameworks such as FSDP, DeepSpeed, Megatron-LM, or custom distributed training systems.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience optimizing distributed training for compute efficiency, memory efficiency, or scalability.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience running and managing large-scale GPU experiments, including scheduling, monitoring, and fault tolerance.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Contributions to widely used open-source ML or ML systems projects.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience building or operating ML products or managed services used by external customers.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, ATLAS, RedPajama, and Mamba. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits. The US base salary range for this full-time position is $200,000 - $290,000. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Please see our privacy policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5199554007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5179372007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4652603007,"location":{"name":"San Francisco"},"metadata":null,"id":5179372007,"updated_at":"2026-07-19T11:24:48-04:00","requisition_id":"R\u0026D-RES-110_New","title":"Research Engineer, Post-Training Inference","company_name":"Together AI","first_published":"2026-07-06T14:21:40-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;The Model Shaping team at Together AI works on products and research focused on tailoring open foundation models to downstream applications. We build services that enable machine learning developers to choose the best models for their tasks and further improve these models using domain-specific data. In addition, we develop new methods for more efficient model training and evaluation, drawing inspiration from a broad range of ideas across machine learning, natural language processing, and ML systems.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As a Research Engineer within Model Shaping, you will develop a platform that enables users to customize open-source models with their own data. Working across the training and inference stacks, you will build and improve our Fine-Tuning, Reinforcement Learning, and Evaluation services – from ensuring a seamless path from post-training to production serving, to optimizing the inference engine for RL training workloads. You will collaborate closely with our product, research, and engineering teams to keep the API reliable, performant, and well integrated into the company\u0026#39;s technical infrastructure. Above all, you will help build the foundational layer of the open-source AI ecosystem, enabling developers around the world to efficiently create high-quality models tailored to their specific applications.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and build Together’s systems for customizing open-source models\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build integrations between the Model Shaping and Inference platforms to ensure a seamless path from post-training to serving production workloads\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Add features to inference engines for large-scale post-training experiments, including optimizations for RL workloads\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Make sure the service is stable and robust, participating in an on-call rotation and ensuring 24/7 availability of our platform\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Have 2+ years of experience building and deploying machine learning-based services in a production environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Have hands-on experience with modern inference engines, such as SGLang, vLLM, and TensorRT-LLM\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Are familiar with the latest methods for fine-tuning LLMs and other AI models\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Have a strong software engineering background in Python or Go\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Stay up to date with the latest advances and trends in the machine learning community\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Experience in any of the following will make you stand out\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Serving low-precision (FP4/FP8) models, multiple LoRA adapters within one model instance (Multi-LoRA), or models distributed across several GPU nodes\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Optimizing the performance of RL training workloads\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Developing CUDA/Triton/CuTE DSL kernels for inference\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Developing large-scale and high-load production systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintaining or contributing to open-source ML projects\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Managing machine learning workloads on Kubernetes clusters\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, ATLAS, RedPajama, and Mamba. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits. The US base salary range for this full-time position is $200,000 - $290,000. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5179372007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4835763007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4478121007,"location":{"name":"San Francisco"},"metadata":null,"id":4835763007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"153","title":"Senior Backend Engineer, Inference Platform","company_name":"Together AI","first_published":"2025-08-22T14:40:32-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the Inference Platform that brings the most advanced generative AI models to the world. Our platform powers multi-tenant serverless workloads and dedicated endpoints, enabling developers, enterprises, and researchers to harness the latest LLMs, multimodal models, image, audio, video, and speech models at scale.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;If you get a thrill from optimizing latency down to the last millisecond, this is your playground. You’ll work hands-on with tens of thousands of GPUs (H100s, H200s, GB200s, and beyond), figuring out how to fully utilize every FLOP and every gigabyte of memory.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You’ll collaborate directly with research teams to bring frontier models into production, making breakthroughs usable in the real world. Our team also works closely with the open source community, contributing to and leveraging projects like SGLang, vLLM, and NVIDIA Dynamo to push the boundaries of inference performance and efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Shape the core inference backbone that powers Together AI’s frontier models.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Solve performance-critical challenges in global request routing, load balancing, and large-scale resource allocation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work with state-of-the-art accelerators (H100s, H200s, GB200s) at global scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with world-class researchers to bring new model architectures into production.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with and contribute to the open source community, shaping the tools that advance the industry.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A culture of deep technical ownership and high impact — where your work makes models faster, cheaper, and more accessible.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Competitive compensation, equity, and benefits.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026amp;nbsp;\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Build and optimize global and local request routing, ensuring low-latency load balancing across data centers and model engine pods.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop auto-scaling systems to dynamically allocate resources and meet strict SLOs across dozens of data centers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design systems for multi-tenant traffic shaping, tuning both resource allocation and request handling — including smart rate limiting and regulation — to ensure fairness and consistent experience across all users.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Engineer trade-offs between latency and throughput to serve diverse workloads efficiently.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Optimize prefix caching to reduce model compute and speed up responses.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with ML researchers to bring new model architectures into production at scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Continuously profile and analyze system-level performance to identify bottlenecks and implement optimizations.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;Requirements\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of demonstrated experience building large-scale, fault-tolerant, distributed systems and API microservices.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong background in designing, analyzing, and improving efficiency, scalability, and stability of complex systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent understanding of low-level OS concepts: multi-threading, memory management, networking, and storage performance.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Expert-level programming in one or more of: Rust, Go, Python, or TypeScript.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Knowledge of modern LLMs and generative models and how they are served in production is a plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience working with the open source ecosystem around inference is highly valuable; familiarity with SGLang, vLLM, or NVIDIA Dynamo will be especially handy.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Kubernetes or container orchestration is a strong plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with GPU software stacks (CUDA, Triton, NCCL) and HPC technologies (InfiniBand, NVLink, MPI) is a plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or related field, or equivalent practical experience.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $160,000 - $250,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4835763007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4774159007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4446864007,"location":{"name":"San Francisco"},"metadata":null,"id":4774159007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"136","title":"Senior Developer Productivity Engineer","company_name":"Together AI","first_published":"2025-06-25T22:35:15-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;At Together AI, you\u0026#39;ll own the systems and tooling that let our engineers ship quickly and reliably: CI/CD pipelines, release infrastructure, local dev environments, and the testing systems the whole engineering team depends on.\u0026lt;br\u0026gt;This is a high-leverage, build-heavy role. You\u0026#39;ll take on big projects and drive them end to end: building a full-stack release service with deployment orchestration, rebuilding the front-end build pipeline, standing up testing infrastructure that kills flaky builds. The goal is straightforward: engineers should spend their time building products, not fighting their tooling.\u0026lt;br\u0026gt;This is an engineering role focused on building systems, not a QA or manual testing role. You build the infrastructure that makes manual effort unnecessary.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Build and own the production systems engineers rely on: release services with deployment orchestration (canary, blue/green), reusable CI/CD pipelines, and self-service tooling.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build testing infrastructure: the frameworks and harnesses that make tests fast and reliable, and track down flaky tests and slow builds at the systems level.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and maintain local dev environments and containerized workflows so engineers can spin up and iterate fast.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create starter templates and shared tooling (CLIs, codegen, IDE integrations) so teams can stand up new services in minutes.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Be the go-to engineer for GitHub Actions and GitOps: defining problems, untangling pipeline issues, and improving how the org works with them.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Improve front-end build tooling and developer workflows alongside product engineers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work across the org to find the biggest sources of friction and fix them at the root. Document what you build so others can use it.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years building software in production, with a strong track record of owning systems or services end to end, from design through operation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering fundamentals. You write production-quality code and have built and operated real services, not just scripts or test automation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in Python, Go, and JavaScript/TypeScript for tooling, automation, and enablement.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with CI/CD systems (GitHub Actions, ArgoCD, GitOps) and building scalable, reusable pipelines.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with containerized development workflows and local dev tooling (e.g., Skaffold).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building starter templates and scaffolding, with engineering teams, to accelerate new service creation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience with AI coding tools (e.g., Claude Code, Cursor, Copilot) and a habit of using them to move faster and ship more.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A builder\u0026#39;s mindset and strong ownership. You like creating the tools and systems other people depend on.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Rigor in diagnosing systemic issues: flaky tests, build latency, deployment failures.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Nice to have\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Kubernetes expertise (EKS, K3s) and experience optimizing containerized builds.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Infrastructure as Code (Terraform, Ansible, Pulumi).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Front-end tooling familiarity (React, Next.js, Jest) to optimize front-end dev workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Monitoring and observability (Prometheus, Grafana, Honeycomb) to debug bottlenecks.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $180,000 - $250,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at https://www.together.ai/privacy\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4774159007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5088817007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4609398007,"location":{"name":"San Francisco"},"metadata":null,"id":5088817007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"R\u0026D-ENG-082","title":"Senior Machine Learning Engineer, Voice AI ","company_name":"Together AI","first_published":"2026-03-30T15:36:00-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re looking for a Senior ML Engineer to drive the model serving layer for voice workloads. You\u0026#39;ll work hands-on with inference engines like TRT-LLM and SGLang to optimize how we serve models like Whisper, Parakeet, Orpheus, and Kokoro — pushing latency and throughput to the frontier. You\u0026#39;ll profile GPU utilization, design batching strategies for streaming audio, and ensure new model architectures can go from research to production quickly.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a foundational hire on a small, high-impact team. Voice inference has unique challenges — streaming audio, tokenization, real-time latency budgets — that require dedicated ML engineering focus. You\u0026#39;ll shape how Together serves voice models as the industry moves from pipeline architectures (ASR → LLM → TTS) toward end-to-end speech-to-speech.\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the model serving stack that powers Together\u0026#39;s voice platform across STT, TTS, and speech-to-speech.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice model inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with model partners (Cartesia, Deepgram, Rime, and others) to bring their models to production on Together\u0026#39;s infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build quality evaluation frameworks that guide model selection for customers and inform the roadmap.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Join a small, early-stage team with outsized impact on a fast-growing product area.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Optimize inference performance for voice models (STT, TTS, speech-to-speech) — targeting best-in-class TTFB, throughput, and GPU utilization across our curated model set.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Productionize voice models on serverless and dedicated endpoints, including batching strategies, streaming inference, and memory management tailored to audio workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and maintain a voice model evaluation framework — measuring WER across accents, languages, and noise conditions for STT; naturalness, latency, and pronunciation accuracy for TTS.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Enable new model architectures in our serving stack as the field evolves, including audio-native LLMs, codec-based models (SNAC), and speech-to-speech systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with model partners to integrate and optimize their models (Cartesia, Deepgram, Rime, and others) running on Together\u0026#39;s infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Profile and debug performance across the full inference stack — from GPU kernels to framework-level bottlenecks — and ship measurable improvements.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work with the platform engineering side of the team to ensure the serving layer meets the latency and reliability requirements of real-time voice APIs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contribute to voice model fine-tuning capabilities (STT and TTS) as we enable customers to build differentiated voice experiences on Together.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lay the groundwork for multiple new products down the line.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of experience in ML engineering, with a focus on model serving, inference optimization, or ML infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience with LLM serving engines (vLLM, SGLang, TensorRT-LLM, or similar) — comfortable reading and modifying engine internals, not just using APIs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong proficiency in Python and PyTorch; experience with GPU profiling and optimization (CUDA, memory management, kernel-level debugging).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Track record of shipping ML systems to production with measurable performance improvements.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong product sense — you think about what developers building voice apps actually need, not just what\u0026#39;s technically interesting.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfort working on a small, early-stage team where you\u0026#39;ll wear multiple hats and move fast.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with speech and audio ML (ASR, TTS architectures, audio signal processing) is a strong plus but not required — you can learn this quickly if you have strong ML engineering fundamentals.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with audio codecs and tokenization schemes (SNAC, Encodec, DAC) is a plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience training or fine-tuning speech models is a plus.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor\u0026#39;s or Master\u0026#39;s degree in Computer Science, Electrical Engineering, or related field, or equivalent practical experience\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $200,000 - $260,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5088817007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5180977007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4373901007,"location":{"name":"San Francisco"},"metadata":null,"id":5180977007,"updated_at":"2026-07-21T23:45:23-04:00","requisition_id":"89","title":"Senior Network Engineer","company_name":"Together AI","first_published":"2026-07-07T00:36:15-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is looking for a Senior Network Engineer to design, deploy, and operate the global network infrastructure supporting our production services and high-performance AI compute environments.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on engineering role for someone with deep networking expertise who can also troubleshoot across Linux, Kubernetes, automation, and application boundaries. You will work on large-scale, multi-vendor data center networks and help ensure they remain highly available, reliable, scalable, and performant.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;The ideal candidate has strong networking fundamentals, experience operating complex networks at scale, and a structured, evidence-based approach to troubleshooting. You should be comfortable owning problems from initial investigation through root cause and resolution, including situations where the issue may extend beyond the network itself.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;8+ years of professional experience designing, building, and supporting large-scale production data center, cloud, service-provider, or high-performance computing networks (excluding enterprise networks).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep understanding of TCP/IP and strong experience with technologies such as BGP, OSPF, VXLAN, EVPN, ECMP, and QoS.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing and supporting multi-tenant network environments using technologies such as VRFs, VLANs, overlays, and policy-based segmentation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience deploying and troubleshooting network platforms from vendors such as Arista, Cisco, Juniper, and NVIDIA.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong troubleshooting skills using tools such as Wireshark, tcpdump, MTR, curl, nmap, and standard Linux networking utilities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to diagnose connectivity, latency, packet-loss, routing, and performance issues across the network, host, and application layers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience developing or maintaining network automation using Python, Ansible, or similar tools.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience working through a Git-based software development lifecycle, including branching, code review, validation, linting, testing, CI/CD, deployment, and rollback.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Working knowledge of Kubernetes networking, including pods, services, CNIs, and basic connectivity troubleshooting.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Foundational knowledge of RDMA networking and technologies such as RoCE or InfiniBand.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with cloud networking in AWS, GCP, or Azure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong Linux administration and troubleshooting skills.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design, deploy, operate, and maintain global, multi-vendor, multi-protocol networks supporting high-performance AI compute infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Troubleshoot complex network and application-connectivity issues, identify root causes, and drive problems through resolution.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Analyze telemetry, packet captures, logs, and performance data to identify network degradation, congestion, packet loss, and capacity constraints.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Participate in architecture and design reviews to ensure solutions meet requirements for performance, availability, scalability, security, and operational supportability.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop and maintain automation, validation, and operational tooling that improves network reliability and reduces manual effort.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Evaluate network hardware, software, optics, and emerging technologies for use in production environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish standards and operational best practices for network design, deployment, monitoring, change management, and incident response.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lead projects addressing complex technical challenges and contribute directly to the network engineering roadmap.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with infrastructure, systems, security, and application teams to troubleshoot issues that cross traditional ownership boundaries.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Preferred\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience deploying or operating RoCE and/or InfiniBand fabrics.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience supporting GPU clusters, HPC environments, distributed storage, or other high-bandwidth and latency-sensitive workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Understanding of AI training and inference traffic patterns and the demands they place on network infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience operating networks spanning thousands of devices, multiple data centers, and multiple geographic regions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with AI-assisted engineering tools and the ability to validate, test, and safely deploy AI-generated automation or code.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $190,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5180977007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5026002007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4580359007,"location":{"name":"Amsterdam"},"metadata":null,"id":5026002007,"updated_at":"2026-07-22T13:40:20-04:00","requisition_id":"204","title":"Senior Network Engineer (Amsterdam)","company_name":"Together AI","first_published":"2026-01-20T17:36:03-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;About the Role\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is looking for a Senior Network Engineer to design, deploy, and operate the global network infrastructure supporting our production services and high-performance AI compute environments.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on engineering role for someone with deep networking expertise who can also troubleshoot across Linux, Kubernetes, automation, and application boundaries. You will work on large-scale, multi-vendor data center networks and help ensure they remain highly available, reliable, scalable, and performant.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;The ideal candidate has strong networking fundamentals, experience operating complex networks at scale, and a structured, evidence-based approach to troubleshooting. You should be comfortable owning problems from initial investigation through root cause and resolution, including situations where the issue may extend beyond the network itself.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;8+ years of professional experience designing, building, and supporting large-scale production data center, cloud, service-provider, or high-performance computing networks (excluding enterprise networks).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep understanding of TCP/IP and strong experience with technologies such as BGP, OSPF, VXLAN, EVPN, ECMP, and QoS.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing and supporting multi-tenant network environments using technologies such as VRFs, VLANs, overlays, and policy-based segmentation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on experience deploying and troubleshooting network platforms from vendors such as Arista, Cisco, Juniper, and NVIDIA.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong troubleshooting skills using tools such as Wireshark, tcpdump, MTR, curl, nmap, and standard Linux networking utilities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to diagnose connectivity, latency, packet-loss, routing, and performance issues across the network, host, and application layers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience developing or maintaining network automation using Python, Ansible, or similar tools.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience working through a Git-based software development lifecycle, including branching, code review, validation, linting, testing, CI/CD, deployment, and rollback.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Working knowledge of Kubernetes networking, including pods, services, CNIs, and basic connectivity troubleshooting.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Foundational knowledge of RDMA networking and technologies such as RoCE or InfiniBand.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with cloud networking in AWS, GCP, or Azure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong Linux administration and troubleshooting skills.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design, deploy, operate, and maintain global, multi-vendor, multi-protocol networks supporting high-performance AI compute infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Troubleshoot complex network and application-connectivity issues, identify root causes, and drive problems through resolution.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Analyze telemetry, packet captures, logs, and performance data to identify network degradation, congestion, packet loss, and capacity constraints.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Participate in architecture and design reviews to ensure solutions meet requirements for performance, availability, scalability, security, and operational supportability.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop and maintain automation, validation, and operational tooling that improves network reliability and reduces manual effort.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Evaluate network hardware, software, optics, and emerging technologies for use in production environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish standards and operational best practices for network design, deployment, monitoring, change management, and incident response.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Lead projects addressing complex technical challenges and contribute directly to the network engineering roadmap.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with infrastructure, systems, security, and application teams to troubleshoot issues that cross traditional ownership boundaries.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Requirements\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Must have hands-on experience deploying or operating RoCE and/or InfiniBand fabrics.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience supporting GPU clusters, HPC environments, distributed storage, or other high-bandwidth and latency-sensitive workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Understanding of AI training and inference traffic patterns and the demands they place on network infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience operating networks spanning thousands of devices, multiple data centers, and multiple geographic regions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with AI-assisted engineering tools and the ability to validate, test, and safely deploy AI-generated automation or code.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4051033007,"name":"Amsterdam","location":"Amsterdam, North Holland, Netherlands","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5026002007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4624894007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4373009007,"location":{"name":"San Francisco"},"metadata":null,"id":4624894007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"85","title":"Senior Product Engineer","company_name":"Together AI","first_published":"2025-01-15T14:21:11-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together is hiring a Senior Product Engineer to join our central Product Engineering team and embed day-to-day with the Commerce engineering team. You will own the product UI surface for Together AI’s most revenue-critical user flows: enterprise spend controls and governance, usage-based pricing displays and metering dashboards, payment method management, and enterprise contract surfaces.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on role for someone who enjoys operating at a high-quality, efficient pace: breaking ambiguous work into clear increments, shipping useful improvements quickly, and finding pragmatic ways to move customer and revenue impact forward. You will work alongside dedicated product, design, and backend engineers partners and use strong product judgment to make reasonable UX decisions within established patterns.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;4+ years of frontend engineering experience building and shipping productions features to real users\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong React fundamentals in production, including complex, stateful UIs that users depend on\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong TypeScript proficiency, with the ability to design clean typed interfaces across components, API boundaries, and application state\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong observability habits: you instrument what you ship, monitor whether it works, and use data to guide follow-up improvements\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to absorb complex product and business requirements and translate them into correctly sequenced frontend implementations\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Fluency consuming APIs: you can read schemas, understand data shapes, handle loading and error states, and build typed web app interfaces without being blocked by backend engineers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Next.js, Tailwind, and shadcn/ui experience is a meaningful plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with billing, payments, or usage-based pricing UIs is a strong differentiator, including Stripe SDKs, metering displays, subscription states, checkout flows, or enterprise invoicing\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own Commerce’s frontend layer end-to-end: from consuming service APIs and building typed interfaces to delivering polished, accessible, user experiences, jumping in to backend systems when necessary to unblock product work in a fast-paced environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Break down complex product requirements and business logic into an iterative roadmap, sequence implementations correctly, and execute without needing step-by-step direction\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Champion the customer experience across commerce surfaces: notice friction, identify product and UX gaps, and turn recurring customer issues into clear opportunities for the team to improve\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build reusable, well-typed React components on top of Together’s UI Platform and component library, making future Commerce work faster to ship\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Instrument, measure, and verify the success of your own features using data-driven metrics and signals including product analytics, web vitals, and error rates\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Catch and address edge cases, error states, and failure modes that matter in commerce contexts, including payment failures, quota overages, upgrade blockers, and inconsistent billing states\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $160,000 - $230,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4624894007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5192898007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4659249007,"location":{"name":"San Francisco"},"metadata":null,"id":5192898007,"updated_at":"2026-07-28T00:46:10-04:00","requisition_id":"278","title":"Senior Product Engineer, Fullstack","company_name":"Together AI","first_published":"2026-07-27T17:04:54-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI\u0026#39;s product is what developers touch every day — Playground, Model Garden, Fine-Tuning, File Management, Batch Processing, Model Evals. These surfaces are how users experience the platform, and they need to be fast, intuitive, and something developers genuinely love using. That\u0026#39;s what this role is about.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re growing the Product Engineering team and looking for a fullstack engineer who\u0026#39;s motivated by shipping things users genuinely value — not features that just look good in a slide deck. You\u0026#39;ll plug directly into new and existing workstreams and start shipping quickly, working across our core product surfaces to improve existing capabilities and build new ones.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a high-ownership role on a small, fast-moving team, where one strong engineer moves the needle visibly. You\u0026#39;ll see your work in production, instrument and measure it, and use that to shape what comes next.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Ship improvements and new capabilities across Together AI\u0026#39;s core product surfaces — Playground, Model Garden, Fine-Tuning, File Management, Batch Processing, and Model Evals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Own projects and features end-to-end, from data model and API integration through the user interface\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate closely with other engineers to build clean, typed, reusable interfaces in our web application\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Instrument and measure what you ship — add product analytics, monitor system health, and prove users are getting value with data\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work across design, product, and engineering to ship things that unlock user value\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Raise the bar on frontend and fullstack quality across the team through code review and collaborative architecture decisions\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;4–7 years of experience building and shipping real product to real users\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong JavaScript/TypeScript fundamentals; TypeScript experience highly valued\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong React expertise in a large-scale, production environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Full-stack fluency: you connect the dots from data model to API design to state management to the UI, and know how to make the right call at each layer in service of clean, maintainable code and an intuitive user experience\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A pragmatic approach to engineering — you know when to move fast and when to slow down, and you make that call based on user value\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Curiosity about AI and LLMs — you want to understand the models and products you\u0026#39;re building on top of, not just the code around them\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Adept at building and improving agentic workflows, with good judgment for when to lean on AI and when not to\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Cross-functional \u0026quot;we ship it\u0026quot; mentality — no narrow ownership, no finger-pointing\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Next.js, Tailwind, and shadcn/ui experience is a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with product analytics (e.g., Amplitude) and systems monitoring (e.g., Grafana) is a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with CI/CD workflows is a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with a backend language like Go or Python\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $160,000 - 230,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5192898007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5210951007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4668049007,"location":{"name":"San Francisco"},"metadata":null,"id":5210951007,"updated_at":"2026-08-17T13:58:36-04:00","requisition_id":"292","title":"Senior Product Manager, Model APIs \u0026 Developer Experience","company_name":"Together AI","first_published":"2026-08-17T13:58:36-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together serves a broad portfolio of open-weight models, backed by infrastructure designed to run them quickly and efficiently. The interfaces around those models shape the entire customer experience: how easily a developer can migrate an application, how reliably an agent can use a model, and how a team runs large asynchronous workloads.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As Senior Product Manager for Model APIs and Developer Experience, you will own those interfaces. Your scope will extend beyond chat to include the APIs and developer experience for image, video, voice, and passthrough models. You will also own compatibility with the wider developer and agent ecosystem, the voice consumption experience and associated model partnerships, and the evolution of our batch inference product.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;These areas are connected by a single goal: make Together\u0026#39;s models easy to adopt and reliable to build on for developers and autonomous agents.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on product role. You will read API specifications, test integrations, inspect code when useful, and build prototypes or demo applications to sharpen your thinking. You will work closely with engineering and research, own the product direction, make thoughtful tradeoffs, ship, and learn from how customers use what we build.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Set the product direction and roadmap across chat and multimodal APIs, ecosystem compatibility, voice, and batch inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Close the API and behavioral gaps that make it harder for customers to move workloads from proprietary model providers to Together.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop a grounded view of how Together\u0026#39;s APIs perform in the developer and agent ecosystem, and turn the most important gaps into clear product priorities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Define an intuitive API and consumption experience for real-time and agentic voice applications.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build productive relationships with voice model partners and align internal and external teams around a strong joint product experience.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Reimagine batch inference across the job lifecycle, developer experience, completion guarantees, pricing, and packaging.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build lightweight prototypes and demo applications to test product ideas and reduce uncertainty before committing significant engineering resources.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work closely with engineering and research to understand technical constraints, make product tradeoffs, and deliver reliable customer experiences.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Talk with customers and study product usage to separate isolated requests from patterns that should shape the platform.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create a clear, active roadmap and communicate why we are making specific investments and what we expect them to change.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Meaningful experience with APIs and developer tools, whether you designed them, built them, or owned them as a product manager.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong product judgment and the ability to make clear decisions in an evolving market.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A willingness to get close to the work by reading specifications and code, testing integrations directly, and building working prototypes with AI development tools.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A track record of earning trust with engineering and external partners through preparation, sound technical judgment, clear communication, and follow-through.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A bias toward experimentation and iteration. You know how to gather enough evidence to make a decision, ship, and adjust based on what you learn.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong communication and relationship-building skills, including the ability to align teams and move work forward across company boundaries..\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $200 - 280k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033061007,"name":"Product","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5210951007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5070981007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4601398007,"location":{"name":"Remote"},"metadata":null,"id":5070981007,"updated_at":"2026-07-20T14:48:26-04:00","requisition_id":"G\u0026A-BOPS-008","title":"Senior Program Manager, Data Center Delivery","company_name":"Together AI","first_published":"2026-03-10T16:36:52-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About The Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is scaling its own data center infrastructure to power the next generation of AI workloads. We are looking for a Program Manager, Data Center Delivery to serve as our representative across our physical infrastructure builds. In this role, you will manage critical data center builds, coordinate infrastructure deployment, and drive projects from contract negotiation through commissioning and data center fit out with vendors and strategic partners. You will own the full deployment lifecycle, ensuring builds are delivered on time, on budget, and to the quality standards required for high-density AI compute environments.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Serve as Together’s representative on data center construction and expansion projects, managing data center providers, sub-contractors, compute integration partners, and strategic infrastructure alliances with cloud providers and technology partners from inception through deployment.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Oversee full-stack deployment, from MEP installation, white space commissioning, facility readiness, white space fit out, and turnover to operations teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage the performance of integration partners (VARs/CI’s) regarding staging and validation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage partner contracts and collaborative technical partners, service level agreements (SLAs), schedules, and change orders; track progress against milestones and hold partners accountable to performance, cost, and timeline commitments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Conduct diligence on vendor and partner operational readiness.\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Coordinate with external partners, including data center providers, owner’s reps, GCs, design consultants, and commissioning agents, to resolve field issues and keep builds on track.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Track equipment deliveries from vendors, coordinating delivery schedules with construction milestones to ensure materials arrive on time.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Provide regular project status reporting to leadership, including schedule updates, risk registers, budget forecasts, and key decision points.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Drive process improvements across the data center build-out and infrastructure deployment program, establishing scalable workflows for tracking, documentation, and vendor management as the portfolio of builds grows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ensure all deployment activities comply with codes, security policies, and safety standards.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ years of experience in technical project or program management or construction management, with at least 5 years focused on data center, mission-critical, or high-density infrastructure builds.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Direct experience managing data center providers or general contractors on behalf of an owner or developer, including contract administration, schedule oversight, and scope amendment.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience managing high-density compute deployments, including cluster-level cabling, rack-level power distribution, and OEM/ODM staging workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong understanding of data center MEP systems, data hall fit out, structured cabling, and white space delivery.\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience managing multi-million infrastructure budgets and tracking costs and timelines across multiple concurrent projects.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication skills with the ability to translate technical construction and delivery details for non-technical stakeholders and senior leadership.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with high-density GPU compute environments, liquid cooling systems, or AI infrastructure requirements.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor’s degree in Construction Management, Engineering, Architecture, Business or a related field.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;PMP certification or equivalent project management credential.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience with fast-track or modular delivery methods for data center or mission-critical facilities.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an AI-native cloud company building the infrastructure to make AI faster, cheaper, and more accessible. We’re rapidly scaling our GPU footprint: signing our own data center leases, building large-scale clusters, and expanding toward a global owned-infrastructure presence. Our research team has contributed to breakthroughs like FlashAttention, Hyena, and RedPajama, and we co-design across software, hardware, and algorithms to push the frontier of AI efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $190-240K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4044336007,"name":"Business Operations","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5070981007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4971774007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4552947007,"location":{"name":"San Francisco"},"metadata":null,"id":4971774007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"177","title":"Senior Software Engineer, Observability","company_name":"Together AI","first_published":"2025-11-08T12:19:30-05:00","language":"en","application_deadline":null,"content":"\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;The AI Infrastructure team at Together AI is at the forefront of building and scaling the foundational systems that power our generative AI platform. The storage and observability team is crucial for designing, implementing, and maintaining robust distributed storage solutions, ensuring seamless data access and management. They are also responsible for developing comprehensive observability platforms, providing critical insights into system performance and GPU utilization, and proactively identifying and resolving issues.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design and implement a scalable observability platform (metrics, logs, traces) using tools like Prometheus, Grafana, ClickHouse, ClickStack, and OpenTelemetry, including telemetry data pipelines and log aggregation workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop automated monitoring, alerting, and anomaly detection systems, including SLIs/SLOs, runbooks, and predictive analytics for critical services.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and deploy custom observability tools and infrastructure-as-code using Go, Python, Terraform, Ansible, and Helm.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with engineering teams to enhance distributed tracing and application monitoring, and lead incident response with post-mortem analysis.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Define observability best practices.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Expertise in observability platforms (Prometheus, Grafana, ClickStack, OpenTelemetry) and cloud-native monitoring services (AWS, GCP, Azure).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong programming skills in Go, Python, or similar languages, with proficiency in infrastructure-as-code tools (Terraform, Ansible, Helm).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing, operating, and scaling large-scale distributed systems and pipelines for high-volume data ingestion and real-time querying.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep understanding of containerization (Docker) and orchestration (Kubernetes).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Knowledge of microservices architecture, service mesh technologies, CI/CD pipelines, and GitOps workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Expertise in managing databases (PostgreSQL, MongoDB, Redis) and time-series databases with high-cardinality data.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Preferred\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience monitoring AI/ML infrastructure, GPU clusters, and custom metrics for model performance and training pipelines.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Background in high-frequency, low-latency systems monitoring, chaos engineering, and reliability testing.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contributions to open-source observability projects.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with security monitoring and compliance frameworks.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;Compensation\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $200,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4971774007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4749787007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4125288007,"location":{"name":"San Francisco"},"metadata":null,"id":4749787007,"updated_at":"2026-08-06T13:46:11-04:00","requisition_id":"5","title":"Senior Software Engineer - Together Cloud Infrastructure","company_name":"Together AI","first_published":"2025-06-02T20:23:43-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About the Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI Native Cloud, an end-to-end platform for the full\u0026lt;br\u0026gt;generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art\u0026lt;br\u0026gt;AI cloud infrastructure. The Together Cloud team builds the Together GPU Clusters\u0026lt;br\u0026gt;product, which provides high-performance, AI-ready GPU clusters through a self-serve\u0026lt;br\u0026gt;cloud console and is the virtualized infrastructure layer powering Together’s inference,\u0026lt;br\u0026gt;RL, and fine-tuning products.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;As a Senior Software Engineer in Together Cloud Infrastructure, you will play a key role\u0026lt;br\u0026gt;in building the next generation AI cloud platform – a highly available, global, blazing-fast\u0026lt;br\u0026gt;cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s,\u0026lt;br\u0026gt;BlueField DPUs). You\u0026#39;ll enable state-of-the-art ML practitioners with self-serve AI cloud\u0026lt;br\u0026gt;services, such as on-demand + managed Kubernetes and Slurm clusters, for both our\u0026lt;br\u0026gt;internal SaaS products (inference, fine-tuning, RL) and our external cloud customers,\u0026lt;br\u0026gt;spanning dozens of data centers across the world.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design, build, and maintain performant, secure, and highly-available backend\u0026amp;nbsp;services/operators that run in our data centers and automate hardware\u0026amp;nbsp;management, such as Infiniband partitioning, in. DC parallel storage provisioning,\u0026amp;nbsp;and VM provisioning.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and build out the IaaS software layer for a new GB200 data center with\u0026amp;nbsp;thousands of GPUs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and build distributed GPU scheduling and the global management plane\u0026amp;nbsp;that power on-demand and managed clusters across dozens of data centers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop infrastructure that powers our internal inference, RL, and fine-tuning\u0026amp;nbsp;products in addition to external cloud customers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and build systems that scale per-cluster capacity limits and automate the\u0026amp;nbsp;onboarding of new capacity\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work on a global multi-exabyte high-performance object store, serving massive\u0026amp;nbsp;datasets for pretraining and model weights for large-scale inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build advanced observability stacks for our customers with automated node\u0026amp;nbsp;lifecycle management for fault-tolerant distributed pretraining and large-scale\u0026amp;nbsp;inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Perform architecture and research work for decentralized AI workloads\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work on the core, open-source Together AI platform\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create services, tools, and developer documentation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create testing frameworks for robustness and fault-tolerance\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;To be successful, you’ll need to be deeply technical and possess excellent\u0026amp;nbsp;communication, collaboration, and diplomacy skills. You have strong fundamental\u0026amp;nbsp;software development skills. In addition, you have strong systems knowledge and\u0026amp;nbsp;troubleshooting abilities.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of professional software development experience and proficiency in at\u0026amp;nbsp;least one backend programming language (Golang desired)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;5+ years experience writing high-performance, well-tested, production quality code, and demonstrated ownership of large-scale projects driven end-to-end to completion\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience with building and operating high-performance and/or\u0026amp;nbsp;globally distributed micro-service architectures across one or more cloud\u0026amp;nbsp;providers (AWS, Azure, GCP)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication skills – able to write clear design docs and work\u0026amp;nbsp;effectively with both technical and non-technical team members\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems knowledge across compute, networking, and storage, including\u0026amp;nbsp;concurrency, memory management, performant I/O, and scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building and operating reliable, customer-facing production systems\u0026amp;nbsp;at scale with infrastructure automation tools (Terraform, Ansible),\u0026amp;nbsp;monitoring/observability stacks (Prometheus, Grafana), and CI/CD pipelines\u0026amp;nbsp;(GitHub Actions, ArgoC\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Preferred Qualifications:\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with Kubernetes internals, such as implementing non-trivial\u0026amp;nbsp;Kubernetes operators, device/storage/network plugins, custom schedulers, or\u0026amp;nbsp;patches to Kubernetes itself\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with VMs/hypervisors, such as QEMU/KVM, cloud-hypervisor,\u0026amp;nbsp;VFIO, virtio, PCIE passthrough, Kubevirt, SR-IOV\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with DC networking tech + solutions, such as VLAN, VXLAN,\u0026amp;nbsp;VPN, VPC, OVS/OVN\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Cluster API or similar\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience working on high-performance compute, networking, and/or storage\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience virtualizing GPUs and/or Infiniband\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building IaaS or PaaS systems at scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with DPUs/SmartNICs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;GPU programming, NCCL, CUDA knowledge\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $220,000 - $290,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4749787007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5028862007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4581587007,"location":{"name":"Amsterdam"},"metadata":null,"id":5028862007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"207","title":"Senior Software Engineer Together Cloud Infrastructure ","company_name":"Together AI","first_published":"2026-01-20T15:19:04-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About the Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Hybrid working two days a week in the Amsterdam office.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Design, build, and maintain performant, secure, and highly-available backend services/operators that run in our data centers and automate hardware management, such as Infiniband partitioning, in-DC parallel storage provisioning, and VM provisioning.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and build out the IaaS software layer for a new GB200 data center with thousands of GPUs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work on a global multi-exabyte high-performance object store, serving massive datasets for pretraining.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build advanced observability stacks for our customers with automated node lifecycle management for fault-tolerant distributed pretraining.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Perform architecture and research work for decentralized AI workloads\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work on the core, open-source Together AI platform\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create services, tools, and developer documentation\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Create testing frameworks for robustness and fault-tolerance\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;To be successful, you’ll need to be deeply technical and possess excellent communication, collaboration, and diplomacy skills. You have strong fundamental software development skills. In addition, you have strong systems knowledge and troubleshooting abilities.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of professional software development experience and proficiency in at least one backend programming language (Golang desired)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;5+ years experience writing high-performance, well-tested, production quality code\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience with building and operating high-performance and/or globally distributed micro-service architectures across one or more cloud providers (AWS, Azure, GCP)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication skills – able to write clear design docs and work effectively with both technical and non-technical team members\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with Kubernetes internals a big plus, such as implementing non-trivial Kubernetes operators, device/storage/network plugins, custom schedulers, or patches thereon or Kubernetes itself\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with VMs/hypervisors a big plus, such as QEMU/KVM, cloud-hypervisor, VFIO, virtio, PCIE passthrough, Kubevirt, SR-IOV\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep experience with DC networking tech + solutions a big plus, such as VLAN, VXLAN, VPN, VPC, OVS/OVN\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Cluster API or similar a big plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience working on high-performance compute, networking, and/or storage a big plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience virtualizing GPUs and/or Infiniband a big plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems knowledge across compute, networking, and storage, including concurrency, memory management, performant I/O, and scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with infrastructure automation tools (Terraform, Ansible), monitoring/observability stacks (Prometheus, Grafana), and CI/CD pipelines (GitHub Actions, ArgoCD)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building IaaS or PaaS systems at scale a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with DPUs/SmartNICs a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;GPU programming, NCCL, CUDA knowledge a plus\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4051033007,"name":"Amsterdam","location":"Amsterdam, North Holland, Netherlands","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5028862007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4753072007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4435017007,"location":{"name":"San Francisco"},"metadata":null,"id":4753072007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"129","title":"Senior Software Engineer - Together Cloud Platform","company_name":"Together AI","first_published":"2025-06-04T13:04:51-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About the Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;As a Senior Backend Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware (GB200s/GB300s, BlueField DPUs) and enables state-of-the-art ML practitioners with self-serve AI cloud services, such as on-demand + managed Kubernetes and Slurm clusters. This platform serves both our internal StaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Some of what you’ll work on:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Work on a distributed GPU scheduling system for the on-demand clusters product, Instant Clusters.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build out a global management plane for managing our data center compute, networking, and storage.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Design and build new customer-facing cloud platform services, delivering killer enterprise AI cloud features.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Identify, design, and develop foundational backend services that power Together’s cloud platform\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Analyze and improve the robustness and scalability of existing distributed systems, APIs, databases, and infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with product teams to understand functional requirements and deliver solutions that meet business needs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Write clear, well-tested, and maintainable software and IaC for both new and existing systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Conduct design and code reviews, create developer documentation, and develop testing strategies for robustness and fault tolerance\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Participate in an on-call rotation to address critical incidents when necessary\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of demonstrated experience in building large scale, fault tolerant, distributed systems and API microservices\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience designing, analyzing and improving efficiency, scalability, and stability of various system resources\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication skills – able to write clear design docs and work effectively with both technical and non-technical team members\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated experience with building and operating high-performance and/or globally distributed microservice architectures across one or more cloud providers (AWS, Azure, GCP)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong systems knowledge across compute, networking, and storage, including concurrency, memory management, performant I/O, and scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience developing against and managing a relational database, such as PostgreSQL\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Expert-level programmer in one or more of programming language (Golang preferred)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in version control practices and integrating IaC with CI/CD pipelines.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Kubernetes and containers preferred\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building and operating data infrastructure (Kinesis, Airflow, Kafka, etc) a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical field, or equivalent practical experience\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $160,000 - $230,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4753072007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5135941007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4631531007,"location":{"name":"San Francisco"},"metadata":null,"id":5135941007,"updated_at":"2026-07-10T19:42:36-04:00","requisition_id":"253","title":"Senior Technical Recruiter, AI/ML Research","company_name":"Together AI","first_published":"2026-05-28T13:08:56-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3 style=\u0026quot;line-height: 1;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;About the Role\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is building the AI Native Cloud — an end-to-end platform for generative AI lifecycle, integrating fast, reliable inference, and model-shaping services with cutting-edge AI cloud infrastructure. We are looking for a seasoned Senior Technical Recruiter to partner closely with AI Research and Engineering leadership to scale world-class research teams across kernels, inference optimization, applied AI and model shaping.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p style=\u0026quot;line-height: 1;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;This role is ideal for someone who understands the unique dynamics of recruiting top-tier AI researchers and research engineers in a highly competitive market and can operate as a strategic talent partner to technical leadership.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3 style=\u0026quot;line-height: 1;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Responsibilities\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p style=\u0026quot;line-height: 1;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Partner with executives, research leadership, and hiring managers to define and execute hiring strategies\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Lead full-cycle recruiting for specialized AI talent, including researchers, research engineers, applied scientists, and ML systems engineers\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Build and nurture relationships with top AI talent across academia, open-source communities, research labs, and industry networks\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Drive exceptional candidate experiences from initial engagement through offer close, with a strong focus on relationship building and long-term talent cultivation\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Provide market intelligence on AI talent trends, compensation, competitive hiring landscapes, and emerging research organizations to influence hiring strategy and organizational planning\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Collaborate cross-functionally with sourcing, coordination, people operations, and leadership teams to continuously improve recruiting processes and operational excellence\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Design and refine interview processes tailored for research hiring, including technical evaluations, publication reviews, and research presentation loops\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3 style=\u0026quot;line-height: 1;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Requirements\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;5+ years of technical recruiting experience at high-growth technology companies, with a proven track record hiring exceptional AI researchers, research engineers, applied scientists, or ML infrastructure talent at leading AI labs, hyper-scale startups, or frontier technology companies\u0026lt;br\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience recruiting from both industry and academic pipelines, including familiarity with publications, conferences, open-source contributions, and research communities\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Ability to build trusted partnerships with senior technical leaders and influence hiring decisions through data and market expertise\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Resourceful, adaptable, and creative in solving recruiting challenges, with comfort operating in fast-paced, ambiguous, high-growth environments with rapidly evolving priorities and hiring needs\u0026lt;br\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Demonstrated ability to design recruiting processes and talent strategies from the ground up for emerging technical organizations\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;line-height: 1; font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong execution mindset with the ability to navigate complex searches and close highly competitive candidates in constrained talent markets\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $165,000 - $210,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026amp;nbsp;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Please see our Privacy Policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033062007,"name":"Human Resources","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5135941007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5209804007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4667491007,"location":{"name":"San Francisco"},"metadata":null,"id":5209804007,"updated_at":"2026-08-13T15:20:31-04:00","requisition_id":"288","title":"Senior University Recruiting Program Manager","company_name":"Together AI","first_published":"2026-08-13T14:55:21-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We are looking for a Senior University Recruiting Program Manager to design, build, and scale the programs that power our talent acquisition engine. In this role, you will own the strategy and execution of cross-functional recruiting initiatives, design programs and events that build pipeline, and grow our employer brand presence. You will partner closely with our University Recruiting team, People Team, and other cross-functional partners to ensure our hiring programs are efficient, data-driven, and deliver an exceptional candidate experience.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a high-impact role at the intersection of program management and talent acquisition. We’re looking for someone who gets energized by building from the ground up.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;University Recruiting \u0026amp;amp; Programs (60%)\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Building Recruiting Programs: \u0026lt;/strong\u0026gt;Map and execute recruiting program strategies that grow our recruiting pipelines and builds experiences. Examples include our employee referral and intern program.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Partner closely with our University Recruiting Team \u0026lt;/strong\u0026gt;on recruiting, initiatives, and events that expand the program and provide best-in-class experiences.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Data that drives programs: \u0026lt;/strong\u0026gt;Use data-driven insights to inform decisions, identify bottlenecks, and report on program effectiveness.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Support on ad-hoc programs \u0026amp;amp; projects\u0026lt;/strong\u0026gt; for the recruiting team as we continue to scale.\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Employer Branding (40%)\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Partner with our Marketing team \u0026lt;/strong\u0026gt;on building recruiting marketing initiatives, social media strategy, and event programming.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own Together AI’s employer brand\u0026lt;/strong\u0026gt; across LinkedIn, Careers Page, and blogs to attract top talent.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage all aspects of our recruiting presence at events \u0026amp;amp; conferences.\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of relevant experience in University Recruiting, Recruiting Program Management, or a highly cross-functional role in a fast-moving environment\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven track record of designing and scaling recruiting programs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfortable with ambiguity and complexity, as we’ll be building from scratch\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong analytical instincts - you use data to evaluate program health, spot patterns, and make informed decisions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Willingness to travel to conferences and events to represent Together AI\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have:\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with employer branding, recruiting marketing, or large-scale event management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Past experience building relationships with external organizations to grow partnerships\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits. The US base salary range for this full-time position is: $155,000 - $195,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4033062007,"name":"Human Resources","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5209804007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5196805007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4661178007,"location":{"name":"San Francisco"},"metadata":null,"id":5196805007,"updated_at":"2026-07-30T12:31:23-04:00","requisition_id":"279","title":"Software Engineer, Customer Insights","company_name":"Together AI","first_published":"2026-07-30T12:31:23-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together.ai is looking for a Software Engineer to join the Customer Insights team — a great role for a full-stack or backend engineer who wants to grow into event-driven systems, analytics, and the customer visibility space. Customer Insights owns the customer-facing visibility layer of Together\u0026#39;s Cloud: the historical analytics, activity history, audit logs, event timelines, notifications, and investigation workflows that help customers understand their activity and govern their AI workloads with confidence.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Every time a customer wants to know what happened, who acted, when it happened, and whether they need to act, they rely on the systems we build. Whether it\u0026#39;s a developer checking their recent API activity, an enterprise team reviewing an audit trail, or an operator investigating an anomaly, we make that visibility clear and trustworthy — simple for everyday cases while providing robust, enterprise-grade capabilities for complex organizational needs. Our work turns today\u0026#39;s fragmented visibility patterns into coherent product and platform foundations, and we\u0026#39;re actively building toward the next generation of insight workflows that summarize activity, explain anomalies, and correlate events across surfaces for both human operators and autonomous agents.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;You\u0026#39;ll own well defined pieces of work end-to-end, shipping work that lands in front of customers. You\u0026#39;ll get guidance as you take on unfamiliar problem spaces, with plenty of room to grow toward more autonomy over time. We pair, review each other\u0026#39;s code, and learn in the open - it\u0026#39;s a strong environment to level up in.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Build and ship features for Together\u0026#39;s Customer Insights platform, including customer-visible events, activity history, audit logs, timelines, notifications, and historical analytics and dashboards views.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Partner with Data Platform and Observability engineering teams to leverage Together’s core data stack and capabilities to build out Customers Insights data collection and processing systems.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Own well-defined pieces of work end-to-end, from implementation through testing and rollout\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Write and maintain critical-path backend and product code used by multiple teams and customer-facing surfaces.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Surface blockers early and collaborate with the team to work through them\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Help maintain the freshness, completeness, and correctness checks that keep customer-facing visibility systems trustworthy\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;2-5 years of experience building and operating large-scale distributed systems, product platforms, or customer-facing backend systems in production environments.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Good backend engineering skills in one or more of TypeScript, Go, Python, Java, C++, or similar production languages.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Genuine enthusiasm to learn event-driven systems, analytics pipelines, and the customer visibility space - deep prior experience isn\u0026#39;t required, and we\u0026#39;ll support you in getting there\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Developing data modeling instincts - comfort working with schemas, queries, and relational and/or non-relational data\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Writes clean, well-organized, well-tested code, and treats code review as a way the team grows together\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Comfortable owning scoped work and surfacing blockers early\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Curious, open to feedback, and willing to propose new approaches and make mistakes\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Interest in, or some exposure to, stream processing, data pipelines, notification systems, or analytical systems (this is a great team to learn them on)\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $150k - 200k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5196805007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4627491007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4374326007,"location":{"name":"San Francisco"},"metadata":null,"id":4627491007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"92","title":"Solutions Architect ","company_name":"Together AI","first_published":"2025-01-22T18:40:51-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Solutions Architect at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Act as a technical advisor to our most strategic customers, deeply embedding with them to support the ideation and development of innovative applications using OSS models on Together AI\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run complex demonstrations and POCs of Together’s entire stack, including both hardware and software solutions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with sales to qualify new prospects and support existing customers along their journey to build cutting-edge Generative AI solutions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and maintain strong relationships with customer leadership and stakeholders, ensuring the successful deployment and scaling of their applications\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deliver high-value feedback to our Product, Engineering, and Research teams, ensuring our platform continues to evolve to meet customer needs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build educational content and tooling for both internal and external use around Together’s solutions (i.e., playbooks, blogs, demos, etc.)\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years of experience in a customer-facing technical role with at least 2 years in a pre-sales function\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to consult with new and existing customers to map business needs to technical solutions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong technical background, with knowledge of AI, ML, GPU technologies and their integration into high-performance computing (HPC) environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong understanding of training, fine-tuning and inference in the context of open source LLMs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in Python and JavaScript, with experience building and delivering prototypes on API platforms.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible), container infrastructure Docker), and scripting and programming languages (Python, Javascript)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong sense of ownership and willingness to learn new skills to ensure both team and customer success.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to operate in dynamic environments, adept at managing multiple projects, and comfortable with frequent context switching and prioritization.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US salary range for this full-time position is: $180-260K OTE + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4627491007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4946442007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4537316007,"location":{"name":"London"},"metadata":null,"id":4946442007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"169","title":"Solutions Architect (Inference)","company_name":"Together AI","first_published":"2025-10-16T07:34:01-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role \u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Solutions Architect (Inference) at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Act as a technical advisor to our most strategic customers, deeply embedding with them to support the ideation and development of innovative applications using OSS models on Together AI\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run complex demonstrations and POCs of Together’s entire stack, including both hardware and software solutions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with sales to qualify new prospects and support existing customers along their journey to build cutting-edge Generative AI solutions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and maintain strong relationships with customer leadership and stakeholders, ensuring the successful deployment and scaling of their applications\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deliver high-value feedback to our Product, Engineering, and Research teams, ensuring our platform continues to evolve to meet customer needs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build educational content and tooling for both internal and external use around Together’s solutions (i.e., playbooks, blogs, demos, etc.)\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ years of experience in a customer-facing technical role with at least 2 years in a pre-sales function\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Must have a demonstrable background in building and delivering technical solutions within a Solutions Consulting / Pre-Sales environment, with strong experience leveraging the latest open-source models, ideally within inference-driven use cases.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to consult with new and existing customers to map business needs to technical solutions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong technical background, with knowledge of AI, ML, GPU technologies and their integration into high-performance computing (HPC) environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong understanding of training, fine-tuning and inference in the context of open source LLMs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in Python and JavaScript, with experience building and delivering prototypes on API platforms.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible), container infrastructure Docker), and scripting and programming languages (Python, Javascript)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong sense of ownership and willingness to learn new skills to ensure both team and customer success.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to operate in dynamic environments, adept at managing multiple projects, and comfortable with frequent context switching and prioritization.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4051035007,"name":"London","location":"London, England, United Kingdom","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4946442007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4661776007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4390483007,"location":{"name":"San Francisco"},"metadata":null,"id":4661776007,"updated_at":"2026-07-10T19:42:34-04:00","requisition_id":"R\u0026D-PROD-027","title":"Sr. Technical Program Manager (TPM)","company_name":"Together AI","first_published":"2025-02-25T14:39:24-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About the Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Senior Infrastructure Technical Program Manager (TPM) at Together AI, you will be at the core of building, optimizing, and scaling the global GPU resources needed for a pioneering AI infrastructure company. Your role is crucial in ensuring that the backbone of our AI models, thousands of GPUs distributed around the world, operates efficiently and reliably, enabling cutting-edge AI advancements that democratize access to AI technology globally.\u0026amp;nbsp; You will drive cross-functional excellence by streamlining critical workflows and enhancing communication across internal and external teams. Join top engineers, researchers, and innovators to shape the future of AI infrastructure and power the next generation of AI-driven solutions.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Product Development:\u0026lt;/strong\u0026gt; Design and build products for AI researchers, developers, and enterprise customers, translating technical requirements into product features and collaborating with research, engineering, and design teams. Develop and execute strategic plans for the Observability, Storage, Network Engineering, and Security infrastructure teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;End-to-End Product Ownership:\u0026lt;/strong\u0026gt; Own a comprehensive product roadmap, detailing key features, enhancements, and releases. Drive end-to-end product development, manage development and testing, and lead launches.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Stakeholder Engagement\u0026lt;/strong\u0026gt;: Engage with stakeholders to understand their needs, pain points, and feedback. Drive initiatives to enhance customer satisfaction and loyalty through product improvements and innovative solutions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Cross-Functional Execution:\u0026lt;/strong\u0026gt; Lead and align diverse cross-functional teams — including Research, Engineering, DevOps, SRE, and Go-to-Market — to ensure seamless project delivery and organizational success.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;ML Product or Infrastructure Experience\u0026lt;/strong\u0026gt;: 5+ years of experience building and scaling AI/ML-powered products and infrastructure, specifically collaborating with research and engineering teams.\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Proven experience with large-scale technology deployments, including cloud computing platforms, decentralized cloud infrastructure, and distributed systems (e.g., containerization and orchestration tools).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with the technical domains of Observability, Storage, Network Engineering, and Security for infrastructure.Experience with cloud computing platforms, decentralized cloud infrastructure, and/or similar large-scale technology deployments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with cloud-based technologies (e.g., AWS, Google Cloud, or Azure)\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Technical Foundation:\u0026lt;/strong\u0026gt; Bachelor\u0026#39;s or Master\u0026#39;s degree in Machine Learning, Computer Science, Engineering, or a related field.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Exceptional analytical and problem-solving skills, with a demonstrated ability to identify and proactively mitigate technical risks\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience using AI tools, such as ClaudeCode or similar, to accelerate analytical progress.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Executive and Organizational Acumen\u0026lt;/strong\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to thrive in a fast-paced, ambiguous startup environment, prioritizing complex tasks and managing multiple simultaneous projects.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong organizational abilities to build cross-functional alignment and establish clear, focused priorities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A proactive and collaborative team-oriented approach, demonstrating a willingness to drive necessary outcomes across the company.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and program management skills for effective collaboration with both internal stakeholders and external vendors.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $225k to 265k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. This is a hybrid role based in the Bay Area.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026amp;nbsp;\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033061007,"name":"Product","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4661776007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5155722007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4563210007,"location":{"name":"San Francisco"},"metadata":null,"id":5155722007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"182","title":"Staff Engineer, Distributed Storage and HPC \u0026 AI Infrastructure","company_name":"Together AI","first_published":"2026-06-04T12:37:35-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;In this role, you will operate, scale, and optimize multi-petabyte storage systems purpose-built for the world’s largest AI training and inference workloads. You’ll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You will also build Kubernetes-native storage operators and self-service platforms that provide automated provisioning, strict multi-tenancy, performance isolation, and quota enforcement at cluster scale. Day-to-day, you’ll optimize end-to-end data paths for 10-50 GB/s per node, design multi-tier caching architectures, implement intelligent prefetching and model-weight distribution, and tune parallel filesystems for AI workloads.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Architect and implement the technical strategy and storage roadmap for Together AI, driving high-performance architectural decisions as we scale our GPU fleet.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Engineer and scale multi-petabyte AI/ML storage systems by integrating Vast, Weka, and Ceph while executing deep cost optimization through automated tiering and lifecycle policies.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Develop intelligent caching and tiered storage architectures to achieve extreme IOPS and cluster-wide throughput at GPU scale for training and inference workloads.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Tune storage isolation at the L2/L3 network layers to ensure secure, production-grade multi-tenancy for storage clients.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Code Kubernetes storage operators and controllers to enable automated provisioning, self-service abstractions, and quota enforcement.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Engineer end-to-end data paths to achieve 10+ GB/s per GPU node; architect multi-tier caching for model weights and datasets; tune parallel filesystems using advanced profiling; and scale storage infrastructure across thousands of nodes.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Optimize end-to-end data paths through advanced benchmarking and profiling, contributing high-impact code to open-source storage projects and internal tooling.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;8+ years in storage engineering, managing distributed storage at multi-petabyte scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven track record deploying and operating high-performance storage for GPU/HPC clusters\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep Kubernetes and cloud-native storage experience in production environments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong coding skills in Go and Python with demonstrated ability to build production-grade systems and tooling\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;BS/MS in Computer Science, Engineering, or equivalent practical experience\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;History of technical leadership: designing systems that significantly improved performance, reliability (99.999%+ uptime), or cost efficiency\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Distributed Storage Systems: Deep expertise in either of Ceph, WekaFS, Lustre, Vast, GPFS, or similar parallel filesystems at multi-petabyte scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Object Storage: Production experience with S3, MinIO, Ceph, or R2 including performance optimization and cost management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Kubernetes Storage: CSI drivers, StatefulSets, PersistentVolumes, storage operators, and custom controllers\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Storage optimization for GPU workloads, RDMA/InfiniBand networking, parallel filesystem optimization (TB/s aggregate cluster throughput - line saturation)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Programming: Go and Python for automation, operators, and tooling\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Infrastructure as Code: Terraform, Ansible, Helm, GitOps (ArgoCD)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Linux Storage Stack: Advanced knowledge of filesystems (ext4, xfs), LVM, NVMe optimization, RAID configurations\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Observability: Prometheus, Grafana, Thanos architecture and operations\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have Skills\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;GPU Direct Storage (GDS), NVMe-oF, storage networking, RDMA implementations\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;ML/AI storage patterns (model weights, checkpointing, dataset caching)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Storage benchmarking and profiling tools (fio, iperf3, iostat, blktrace).\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $250,000 - $300,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5155722007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5140763007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4609398007,"location":{"name":"San Francisco"},"metadata":null,"id":5140763007,"updated_at":"2026-07-10T19:42:35-04:00","requisition_id":"R\u0026D-ENG-082","title":"Staff Machine Learning Engineer, Voice AI ","company_name":"Together AI","first_published":"2026-05-19T14:19:46-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re looking for a Staff ML Engineer to drive the model serving layer for voice workloads. You\u0026#39;ll work hands-on with inference engines like TRT-LLM and SGLang to optimize how we serve models like Whisper, Parakeet, Orpheus, and Kokoro — pushing latency and throughput to the frontier. You\u0026#39;ll profile GPU utilization, design batching strategies for streaming audio, and ensure new model architectures can go from research to production quickly.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a foundational hire on a small, high-impact team. Voice inference has unique challenges — streaming audio, tokenization, real-time latency budgets — that require dedicated ML engineering focus. You\u0026#39;ll shape how Together serves voice models as the industry moves from pipeline architectures (ASR → LLM → TTS) toward end-to-end speech-to-speech.\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the model serving stack that powers Together\u0026#39;s voice platform across STT, TTS, and speech-to-speech.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice model inference.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with model partners (Cartesia, Deepgram, Rime, and others) to bring their models to production on Together\u0026#39;s infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build quality evaluation frameworks that guide model selection for customers and inform the roadmap.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Join a small, early-stage team with outsized impact on a fast-growing product area.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p class=\u0026quot;font-claude-response-body break-words whitespace-normal leading-[1.7]\u0026quot;\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p class=\u0026quot;font-claude-response-body break-words whitespace-normal leading-[1.7]\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul class=\u0026quot;[li_\u0026amp;amp;]:mb-0 [li_\u0026amp;amp;]:mt-1 [li_\u0026amp;amp;]:gap-1 [\u0026amp;amp;:not(:last-child)_ul]:pb-1 [\u0026amp;amp;:not(:last-child)_ol]:pb-1 list-disc flex flex-col gap-1 pl-8 mb-3\u0026quot;\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Own the voice inference roadmap end-to-end — define and execute the technical strategy for optimizing STT, TTS, and speech-to-speech models across Together\u0026#39;s infrastructure, with a clear-eyed view of where the field is heading and how to position the platform ahead of it.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Drive best-in-class inference performance — architect and implement systems targeting leading TTFB, throughput, and GPU utilization for voice workloads; set the performance bar others in the industry measure against, not just catch up to.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Lead productionization of voice models at scale — design the serving architecture for serverless and dedicated endpoints, including batching strategies, streaming inference pipelines, and memory management tailored to real-time audio; own reliability and latency SLAs.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Build the voice evaluation platform — design a rigorous, extensible evaluation framework covering WER across accents, languages, and noise conditions for STT; naturalness, latency, and pronunciation fidelity for TTS; establish the internal benchmark methodology that informs model selection and roadmap decisions.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Shape the architecture for next-generation model support — anticipate and enable emerging model paradigms — audio-native LLMs, codec-based architectures (SNAC, Encodec), and end-to-end speech-to-speech systems — before they\u0026#39;re mainstream, not after.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Serve as the technical DRI for model partner integrations — lead deep collaboration with partners such as Cartesia, Deepgram, and Rime; own the full lifecycle from integration to optimization to ongoing performance accountability.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Diagnose and resolve the hardest performance problems in the stack — conduct systematic profiling and root-cause analysis from GPU kernel behavior to framework-level bottlenecks; drive shipped improvements with documented, measurable impact.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Influence platform architecture across the organization — partner with platform engineering leadership to ensure the serving layer is built for the latency and reliability demands of real-time voice APIs; your technical decisions should raise the ceiling for the whole team.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Define and scale voice fine-tuning capabilities — lead the technical direction for enabling customers to fine-tune STT and TTS models on Together\u0026#39;s infrastructure, establishing the primitives for differentiated voice experiences.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Lay technical foundations for a category-defining product surface — architect systems with enough foresight that they support multiple new voice products with minimal rework; think in terms of platforms, not point solutions.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p class=\u0026quot;font-claude-response-body break-words whitespace-normal leading-[1.7]\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul class=\u0026quot;[li_\u0026amp;amp;]:mb-0 [li_\u0026amp;amp;]:mt-1 [li_\u0026amp;amp;]:gap-1 [\u0026amp;amp;:not(:last-child)_ul]:pb-1 [\u0026amp;amp;:not(:last-child)_ol]:pb-1 list-disc flex flex-col gap-1 pl-8 mb-3\u0026quot;\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;8+ years of ML engineering experience, with a demonstrated focus on model serving, inference optimization, or ML infrastructure at production scale — including systems you\u0026#39;ve owned from design through live traffic.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Deep, practical expertise in LLM serving engines (vLLM, SGLang, TensorRT-LLM, or equivalent) — you\u0026#39;ve modified engine internals, debugged edge cases under load, and contributed improvements back; you don\u0026#39;t stop at the API surface.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Expert-level Python and PyTorch proficiency, with a strong command of GPU optimization — CUDA kernels, memory hierarchies, profiling toolchains — and a track record of turning that knowledge into shipped latency or throughput wins.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Proven system design judgment — you\u0026#39;ve made architectural decisions that held up at scale and influenced how a team or platform evolved; you can articulate the tradeoffs you made and why.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Strong technical leadership — you operate with high autonomy, define the right problems before solving them, and raise the bar for engineering quality around you without requiring process overhead.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Sharp product intuition for developer tooling — you understand what voice application developers actually need to ship great products, and you let that shape your technical priorities, not just the other way around.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Proven ability to move fast in ambiguous environments — you\u0026#39;ve thrived on early-stage or platform teams where scope is wide, ownership is deep, and the roadmap you build is the one you execute.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Strong foundation in speech and audio ML (ASR/TTS architectures, audio signal processing) — directly relevant experience is strongly preferred; exceptional ML engineering fundamentals with genuine curiosity about the domain is also considered.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Familiarity with audio codec and tokenization schemes (SNAC, Encodec, DAC) is a meaningful plus at this level.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Experience training or fine-tuning speech models at scale is a significant advantage.\u0026lt;/li\u0026gt;\n\u0026lt;li class=\u0026quot;font-claude-response-body whitespace-normal break-words pl-2\u0026quot;\u0026gt;Bachelor\u0026#39;s or Master\u0026#39;s in Computer Science, Electrical Engineering, or related field — or equivalent depth demonstrated through your work.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $220,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5140763007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5180690007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4653195007,"location":{"name":"San Francisco"},"metadata":null,"id":5180690007,"updated_at":"2026-07-10T19:42:37-04:00","requisition_id":"274","title":"Staff Platform Engineer, Service Infrastructure","company_name":"Together AI","first_published":"2026-07-10T18:06:31-04:00","language":"en","application_deadline":null,"content":"\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-size: 12pt;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI is hiring a Staff Platform Engineer to join the Product Foundations engineering organization and drive its service infrastructure strategy.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Product Foundations builds and operates Together’s mission-critical product platforms that support all cloud products, including API Platform (non-Inference), web UI Platform, Billing, and customer-facing IAM. These services sit on the critical path for customers and internal systems.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on Staff role focused on evolving Product Foundations’ core infrastructure strategy from the inside: understanding service team needs, turning repeated infrastructure problems into reusable patterns, and coordinating across platform owners so Product Foundations services are reliable, repeatable, and built on the right company-wide foundations.\u0026lt;/p\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;span style=\u0026quot;font-size: 12pt;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the technical direction for service infrastructure within Product Foundations, including Kubernetes, AWS, Terraform, CDNs, ALBs, DNS, IAM, service networking, and related operational patterns.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Up-level existing Product Foundations services by improving reliability, operability, deployment safety, infrastructure consistency, and production readiness.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner deeply with API Platform and UI Platform on networking, DNS, CDN, load balancing, delivery, and gateway patterns for critical customer-facing interfaces.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Work closely with Infrastructure, Networking, and Security teams to bring company-wide platform standards into Product Foundations and contribute PF requirements back into shared frameworks.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Help drive cross-company infrastructure initiatives that Product Foundations depend on or help maintain, including Terraform CI/CD, Kubernetes networking, zero-trust service communication, policy-as-code, and cross-DC/provider networking.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build and evolve reusable service infrastructure primitives, including Helm charts, Terraform modules, GitHub Actions/GitOps workflows, service scaffolding, runbooks, and documentation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Establish durable technical standards through design docs, architecture reviews, mentorship, and hands-on implementation that help Together scale services across teams, regions, and cloud environments.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;span style=\u0026quot;font-size: 12pt;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ years of professional experience in platform engineering, service infrastructure, SRE, distributed systems, cloud infrastructure, or related roles.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep production experience with Kubernetes, including EKS, Helm, ArgoCD/Argo Rollouts, ingress, autoscaling, secrets, service identity, networking, and progressive delivery.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong Terraform experience, including module design, infrastructure CI/CD, policy enforcement, production applies, and safe self-service workflows.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience operating networking and edge infrastructure such as CDNs, ALBs/NLBs, DNS, TLS, ingress/egress controls, and traffic management.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proficiency in one or more programming languages used for infrastructure tooling and automation, such as Go, Python, TypeScript, or similar.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;AWS experience, ideally including EKS, IAM, VPC networking, load balancing, Route 53, CloudFront, ECR, and related service infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Direct experience with observability systems, including metrics, logs, traces, dashboards, alerting, SLOs, and incident response.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to lead cross-functional technical initiatives across product engineering, infrastructure, networking, and security teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong written communication skills, with experience producing clear design docs, migration plans, operational guidance, and technical standards.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Staff-level judgment: you can define ambiguous problems, make pragmatic tradeoffs, influence without authority, and leave both systems and teams better than you found them.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h2\u0026gt;\u0026lt;span style=\u0026quot;font-size: 12pt;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h2\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience building internal developer platforms or paved-path service frameworks used by many engineering teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience embedding infrastructure best practices into product engineering teams at scale.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with service mesh or zero-trust infrastructure such as mTLS, SPIFFE/SPIRE, Cilium, Istio, Linkerd, Envoy, or similar.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with OPA, Gatekeeper, Kyverno, Sentinel, or other policy-as-code systems.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with multi-region, multi-cluster, hybrid-cloud, or cross-provider service networking.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with supply-chain security, image signing, SBOMs, vulnerability management, or compliance automation.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $240,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5180690007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5213322007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4669209007,"location":{"name":"Amsterdam"},"metadata":null,"id":5213322007,"updated_at":"2026-08-19T12:12:12-04:00","requisition_id":"299","title":"Staff Software Engineer, Inference / Compute Infrastructure Engineering","company_name":"Together AI","first_published":"2026-08-19T12:12:12-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re looking for a Software Engineer to build the systems that treat infrastructure as software. This role owns the software state machines that provision hardware, bring it into service, and manage its full lifecycle — turning racks of GPUs into \u0026lt;strong\u0026gt;running inference clusters\u0026lt;/strong\u0026gt; without a human touching a runbook. The Research and Inference team is your customer: today they file tickets and wait; the target state is that they issue a single API call to stand up, scale, or tear down a cluster, and the system takes care of the rest. The platform is manifest-driven such that teams declare the desired state of a cluster or host — shape, topology, software stack — and the system is responsible for reconciling reality to that manifest, continuously, through every stage of its lifecycle. You will design the engines that manifest the schema, the engines that execute against it, and the workflows that carry a piece of hardware or a cluster from one state to the next—taking it from bare metal to a fully functioning AI cluster for training or inference.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll write production code which is typed, tested, versioned, and deployed through CI/CD that models infrastructure state and reconciles it, the same way a Kubernetes controller reconciles a cluster\u0026#39;s desired state. Success looks like eliminating manual provisioning work, not documenting it better.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;A product mindset - you\u0026#39;ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;strong\u0026gt;You build it, you own it.\u0026lt;/strong\u0026gt; You are not only responsible for delivering the software but also for operating and supporting it in production.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the provisioning state machine: \u0026lt;/strong\u0026gt;design and implement the software that models the full lifecycle of a physical host from discovery, inference bring-up to GPU driver/CUDA stack, health validation, and decommission/RMA — as explicit, versioned states and transitions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the self-service API: \u0026lt;/strong\u0026gt;design declarative APIs and a control plane so the inference team can request, scale, and tear down inference clusters with one API call — no ticket, no human in the loop.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Automate self-healing: \u0026lt;/strong\u0026gt;detect degraded or failed nodes, drain them safely, trigger repair or replacement, and reintroduce healthy capacity into the pool automatically.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own reliability of the pipeline: \u0026lt;/strong\u0026gt;idempotency, retries, rollback, and drift detection so the provisioning system is as dependable as any other production service.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Partner with the inference/ML platform team: \u0026lt;/strong\u0026gt;understand the cluster shapes they need — topology, interconnect, scheduling constraints — and encode them as first-class abstractions in the platform.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Engineer it like software: \u0026lt;/strong\u0026gt;strong typing, automated tests, code review, versioning, and CI/CD for infrastructure code — this is a product, not a collection of Ansible playbooks.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Core requirements (all levels):\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering background in Go, Python, Rust, or similar — you write and test real software for a living.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent to run long-lived, manifest-driven workflows that survive failures and resume mid-execution.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building software control planes or orchestration systems that model state and reconcile it over time (e.g., Kubernetes controllers/operators, custom reconciliation loops, workflow engines).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with event-driven systems — designing and building software around message queues, event streams, or pub/sub (e.g., Kafka, NATS, SQS) rather than polling or cron-driven scripts.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A product mindset. You’ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Nice to have:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design), or GPU/accelerator infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Systems programming in Rust or Go.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;\u0026amp;nbsp;https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4051033007,"name":"Amsterdam","location":"Amsterdam, North Holland, Netherlands","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5213322007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5186628007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4656135007,"location":{"name":"San Francisco"},"metadata":null,"id":5186628007,"updated_at":"2026-07-22T15:43:17-04:00","requisition_id":"277","title":"Staff Software Engineer, Inference / Compute Infrastructure Engineering","company_name":"Together AI","first_published":"2026-07-16T12:58:09-04:00","language":"en","application_deadline":null,"content":"\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re looking for a Software Engineer to build the systems that treat infrastructure as software. This role owns the software state machines that provision hardware, bring it into service, and manage its full lifecycle — turning racks of GPUs into \u0026lt;strong\u0026gt;running inference clusters\u0026lt;/strong\u0026gt; without a human touching a runbook. The Research and Inference team is your customer: today they file tickets and wait; the target state is that they issue a single API call to stand up, scale, or tear down a cluster, and the system takes care of the rest. The platform is manifest-driven such that teams declare the desired state of a cluster or host — shape, topology, software stack — and the system is responsible for reconciling reality to that manifest, continuously, through every stage of its lifecycle. You will design the engines that manifest the schema, the engines that execute against it, and the workflows that carry a piece of hardware or a cluster from one state to the next—taking it from bare metal to a fully functioning AI cluster for training or inference.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll write production code which is typed, tested, versioned, and deployed through CI/CD that models infrastructure state and reconciles it, the same way a Kubernetes controller reconciles a cluster\u0026#39;s desired state. Success looks like eliminating manual provisioning work, not documenting it better.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;A product mindset - you\u0026#39;ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;strong\u0026gt;You build it, you own it.\u0026lt;/strong\u0026gt; You are not only responsible for delivering the software but also for operating and supporting it in production.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the provisioning state machine: \u0026lt;/strong\u0026gt;design and implement the software that models the full lifecycle of a physical host from discovery, inference bring-up to GPU driver/CUDA stack, health validation, and decommission/RMA — as explicit, versioned states and transitions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the self-service API: \u0026lt;/strong\u0026gt;design declarative APIs and a control plane so the inference team can request, scale, and tear down inference clusters with one API call — no ticket, no human in the loop.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Automate self-healing: \u0026lt;/strong\u0026gt;detect degraded or failed nodes, drain them safely, trigger repair or replacement, and reintroduce healthy capacity into the pool automatically.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own reliability of the pipeline: \u0026lt;/strong\u0026gt;idempotency, retries, rollback, and drift detection so the provisioning system is as dependable as any other production service.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Partner with the inference/ML platform team: \u0026lt;/strong\u0026gt;understand the cluster shapes they need — topology, interconnect, scheduling constraints — and encode them as first-class abstractions in the platform.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Engineer it like software: \u0026lt;/strong\u0026gt;strong typing, automated tests, code review, versioning, and CI/CD for infrastructure code — this is a product, not a collection of Ansible playbooks.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Core requirements (all levels):\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering background in Go, Python, Rust, or similar — you write and test real software for a living.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent to run long-lived, manifest-driven workflows that survive failures and resume mid-execution.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building software control planes or orchestration systems that model state and reconcile it over time (e.g., Kubernetes controllers/operators, custom reconciliation loops, workflow engines).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with event-driven systems — designing and building software around message queues, event streams, or pub/sub (e.g., Kafka, NATS, SQS) rather than polling or cron-driven scripts.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A product mindset. You’ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Nice to have:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design), or GPU/accelerator infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Systems programming in Rust or Go.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $240,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5186628007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5214645007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4669837007,"location":{"name":"London "},"metadata":null,"id":5214645007,"updated_at":"2026-08-20T04:03:52-04:00","requisition_id":"303","title":"Staff Software Engineer, Inference / Compute Infrastructure Engineering","company_name":"Together AI","first_published":"2026-08-20T04:03:52-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re looking for a Software Engineer to build the systems that treat infrastructure as software. This role owns the software state machines that provision hardware, bring it into service, and manage its full lifecycle — turning racks of GPUs into \u0026lt;strong\u0026gt;running inference clusters\u0026lt;/strong\u0026gt; without a human touching a runbook. The Research and Inference team is your customer: today they file tickets and wait; the target state is that they issue a single API call to stand up, scale, or tear down a cluster, and the system takes care of the rest. The platform is manifest-driven such that teams declare the desired state of a cluster or host — shape, topology, software stack — and the system is responsible for reconciling reality to that manifest, continuously, through every stage of its lifecycle. You will design the engines that manifest the schema, the engines that execute against it, and the workflows that carry a piece of hardware or a cluster from one state to the next—taking it from bare metal to a fully functioning AI cluster for training or inference.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll write production code which is typed, tested, versioned, and deployed through CI/CD that models infrastructure state and reconciles it, the same way a Kubernetes controller reconciles a cluster\u0026#39;s desired state. Success looks like eliminating manual provisioning work, not documenting it better.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;A product mindset - you\u0026#39;ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;strong\u0026gt;You build it, you own it.\u0026lt;/strong\u0026gt; You are not only responsible for delivering the software but also for operating and supporting it in production.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the provisioning state machine: \u0026lt;/strong\u0026gt;design and implement the software that models the full lifecycle of a physical host from discovery, inference bring-up to GPU driver/CUDA stack, health validation, and decommission/RMA — as explicit, versioned states and transitions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the self-service API: \u0026lt;/strong\u0026gt;design declarative APIs and a control plane so the inference team can request, scale, and tear down inference clusters with one API call — no ticket, no human in the loop.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Automate self-healing: \u0026lt;/strong\u0026gt;detect degraded or failed nodes, drain them safely, trigger repair or replacement, and reintroduce healthy capacity into the pool automatically.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own reliability of the pipeline: \u0026lt;/strong\u0026gt;idempotency, retries, rollback, and drift detection so the provisioning system is as dependable as any other production service.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Partner with the inference/ML platform team: \u0026lt;/strong\u0026gt;understand the cluster shapes they need — topology, interconnect, scheduling constraints — and encode them as first-class abstractions in the platform.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Engineer it like software: \u0026lt;/strong\u0026gt;strong typing, automated tests, code review, versioning, and CI/CD for infrastructure code — this is a product, not a collection of Ansible playbooks.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Core requirements (all levels):\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering background in Go, Python, Rust, or similar — you write and test real software for a living.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent to run long-lived, manifest-driven workflows that survive failures and resume mid-execution.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building software control planes or orchestration systems that model state and reconcile it over time (e.g., Kubernetes controllers/operators, custom reconciliation loops, workflow engines).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with event-driven systems — designing and building software around message queues, event streams, or pub/sub (e.g., Kafka, NATS, SQS) rather than polling or cron-driven scripts.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A product mindset. You’ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Nice to have:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design), or GPU/accelerator infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Systems programming in Rust or Go.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4051035007,"name":"London","location":"London, England, United Kingdom","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5214645007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5213325007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4669210007,"location":{"name":"India "},"metadata":null,"id":5213325007,"updated_at":"2026-08-19T12:29:46-04:00","requisition_id":"300","title":"Staff Software Engineer, Inference / Compute Infrastructure Engineering","company_name":"Together AI","first_published":"2026-08-19T12:15:34-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We\u0026#39;re looking for a Software Engineer to build the systems that treat infrastructure as software. This role owns the software state machines that provision hardware, bring it into service, and manage its full lifecycle — turning racks of GPUs into \u0026lt;strong\u0026gt;running inference clusters\u0026lt;/strong\u0026gt; without a human touching a runbook. The Research and Inference team is your customer: today they file tickets and wait; the target state is that they issue a single API call to stand up, scale, or tear down a cluster, and the system takes care of the rest. The platform is manifest-driven such that teams declare the desired state of a cluster or host — shape, topology, software stack — and the system is responsible for reconciling reality to that manifest, continuously, through every stage of its lifecycle. You will design the engines that manifest the schema, the engines that execute against it, and the workflows that carry a piece of hardware or a cluster from one state to the next—taking it from bare metal to a fully functioning AI cluster for training or inference.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You\u0026#39;ll write production code which is typed, tested, versioned, and deployed through CI/CD that models infrastructure state and reconciles it, the same way a Kubernetes controller reconciles a cluster\u0026#39;s desired state. Success looks like eliminating manual provisioning work, not documenting it better.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;A product mindset - you\u0026#39;ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;strong\u0026gt;You build it, you own it.\u0026lt;/strong\u0026gt; You are not only responsible for delivering the software but also for operating and supporting it in production.\u0026lt;br\u0026gt;\u0026lt;br\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Remote in India\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the provisioning state machine: \u0026lt;/strong\u0026gt;design and implement the software that models the full lifecycle of a physical host from discovery, inference bring-up to GPU driver/CUDA stack, health validation, and decommission/RMA — as explicit, versioned states and transitions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Build the self-service API: \u0026lt;/strong\u0026gt;design declarative APIs and a control plane so the inference team can request, scale, and tear down inference clusters with one API call — no ticket, no human in the loop.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Automate self-healing: \u0026lt;/strong\u0026gt;detect degraded or failed nodes, drain them safely, trigger repair or replacement, and reintroduce healthy capacity into the pool automatically.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Own reliability of the pipeline: \u0026lt;/strong\u0026gt;idempotency, retries, rollback, and drift detection so the provisioning system is as dependable as any other production service.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Partner with the inference/ML platform team: \u0026lt;/strong\u0026gt;understand the cluster shapes they need — topology, interconnect, scheduling constraints — and encode them as first-class abstractions in the platform.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;\u0026lt;strong\u0026gt;Engineer it like software: \u0026lt;/strong\u0026gt;strong typing, automated tests, code review, versioning, and CI/CD for infrastructure code — this is a product, not a collection of Ansible playbooks.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Core requirements (all levels):\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong software engineering background in Go, Python, Rust, or similar — you write and test real software for a living.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent to run long-lived, manifest-driven workflows that survive failures and resume mid-execution.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience building software control planes or orchestration systems that model state and reconcile it over time (e.g., Kubernetes controllers/operators, custom reconciliation loops, workflow engines).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with event-driven systems — designing and building software around message queues, event streams, or pub/sub (e.g., Kafka, NATS, SQS) rather than polling or cron-driven scripts.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A product mindset. You’ve built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;Nice to have:\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design), or GPU/accelerator infrastructure.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE).\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Systems programming in Rust or Go.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;.\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033058007,"name":"Engineering","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5213325007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4543435007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4332507007,"location":{"name":"San Francisco, New York City"},"metadata":null,"id":4543435007,"updated_at":"2026-07-15T13:02:26-04:00","requisition_id":"66","title":"Strategic Account Executive","company_name":"Together AI","first_published":"2025-01-23T18:17:08-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We’re hiring a Strategic Account Executive to own a portfolio of high-priority AI-native and enterprise accounts.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You’ll be responsible for generating pipeline, closing new business, and expanding customers across Together’s platform. You’ll work with technical teams, business leaders, and executives to understand their AI strategy and identify where Together can support production workloads, and also partner closely with Solutions, Engineering, and Research to run POC evaluations and move customers from initial use cases to broader deployments.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This role requires experience managing complex sales cycles, technical fluency, and a track record of building pipeline and closing large deals.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Close deals and drive revenue: own the full sales cycle for strategic AI-native and enterprise accounts, from prospecting through close and expansion.\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage and prioritize across multiple complex sales cycles simultaneously. Be able to keep scope defined, while also being flexible in dealing with new challenges and priorities.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Generate a strong pipeline, identifying high value customer segments/use cases and strategically positioning Together’s offerings.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Solutions Architecture, Research, and Engineering to design solutions/POCs and prove value across inference and post-training engagements.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Inform product needs by conveying customer needs, using these insights to enhance Together’s offering and market positioning.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Help improve Together’s sales process, positioning, and overall best practices.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;7+ years of B2B technology sales experience, including experience with strategic startup or enterprise accounts.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A track record of closing and expanding strategic, complex deals – especially with deep technical/commercial requirements requiring tailored solutions.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Technical and business fluency: ability to work effectively with both technical and business stakeholders, speaking with depth on technical ML/infrastructure topics while also connecting with business/C-level needs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to drive external and internal alignment and a “roll up your sleeves\u0026#39;\u0026#39; mentality to problem solving. Comfort working in a fast-moving environment where the product and market evolve rapidly month over month.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to build and maintain lasting relationships over time with top strategic accounts – both from an executive and IC level, across technical and commercial orgs.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;A passion and curiosity for AI infrastructure. Be on the frontlines of inference and post-training, and shape the trajectory and narrative moving forward.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Preferred Qualifications\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience selling to both AI-native companies and large enterprises\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with developer/AI/infra products preferred\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience at a high-growth technology company\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is building the AI Native Cloud, a full-stack platform for companies developing and running production AI. Our platform spans inference, fine-tuning, and GPU clusters, giving customers the performance, control, and economics to build with open-source and custom models.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Our technology is grounded in research across the AI stack, including work behind FlashAttention, Hyena, FlexGen, and RedPajama. Following our $800 million Series C, we are scaling the platform and infrastructure used by thousands of customers, including Cursor, Cognition, Decagon, Salesforce, The Washington Post, ElevenLabs, and Suno.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $300K - 370K OTE + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. This is a hybrid role based in the Bay Area.\u0026amp;nbsp;\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4049151007,"name":"Sales","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4543435007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5212018007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4668553007,"location":{"name":"San Francisco"},"metadata":null,"id":5212018007,"updated_at":"2026-08-17T12:31:16-04:00","requisition_id":"298","title":"Strategic Finance Senior Associate - Compute","company_name":"Together AI","first_published":"2026-08-17T12:31:16-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Strategic Finance Senior Associate at Together, you will own the cost side of our data center build-outs. Reporting to the Director of Strategic Finance, Infrastructure, you will be the financial counterpart to our Infrastructure Strategy team: building and maintaining the cost base for every deployment, validating that build-outs are tracking to plan, and helping evaluate the economics of the sites we take on next.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on, high-ownership role at the center of the largest capital decisions the company makes. The ideal candidate has done this before: inside a hyperscaler, neocloud, colocation provider, or data center developer, or covering data centers from the investing side, and pairs real infrastructure fluency with exceptional financial rigor.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the cost base and cost forecast for Together\u0026#39;s data center build-outs, spanning data center non-recurring costs, hardware capex across GPUs, networking, and storage, and ongoing operating costs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Validate that build-outs are tracking to plan: maintain budget-versus-actual against approved estimates, flag variances early with a clear read on the driver, and work with Infrastructure Strategy and Infrastructure Engineering to close them out\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain the all-in unit cost of owned capacity, on a dollars-per-GPU-hour basis by GPU type, as the source of truth used for pricing, margin, and capital allocation decisions\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with Infrastructure Strategy to evaluate new data center opportunities, building the economics for candidate sites including prepays, financing requirements, tax treatment and incentives, and total cost of ownership over the life of the site\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Pressure-test hardware cost assumptions and surface savings opportunities with Infrastructure Engineering and Infrastructure Strategy, from equipment choices to storage sizing and cluster configuration trade-offs at different scales\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Translate hardware delivery timelines and payment milestones maintained by Infrastructure Strategy into the cash forecast, and own tracking of prepays to hardware and data center suppliers, so that cash forecasting reflects real commitments\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build the reporting layer: dashboards and recurring processes that give leadership a current view of committed capital, spend to date, and forecast to complete, by site\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;4+ years of experience with substantive exposure to data center or large-scale infrastructure capex, via one of two paths:\u0026lt;/li\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Operating side: finance, capex planning, or infrastructure strategy/FP\u0026amp;amp;A at a hyperscaler, neocloud, colocation provider, data center developer, or another operator running large build-out programs\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Investing side: investment banking, infrastructure investing, or project finance with a data center / digital infrastructure focus, ideally paired with time in-house at an operator\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;li\u0026gt;Working knowledge of the cost stack of a data center build: non-recurring costs, hardware capex (GPUs, networking, storage), power and cooling, and ongoing operating costs, and how each flows through COGS, depreciation, and cash\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Exceptional financial modeling skills, including site-level TCO and scenario analysis, with the ability to distill a complex build into the handful of drivers that actually move the outcome\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated ownership of budget-versus-actual for large capital projects, with the discipline to catch variances early and the credibility to drive them to resolution\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Comfort working directly with technical partners: able to hold a substantive conversation with infrastructure engineers, translate it into a cost number, and communicate the result clearly to non-technical and non-financial audiences\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Exceptional business judgment and intuition, with the creativity and problem-solving instincts to tackle novel questions where there is no established playbook\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong work ethic and self-directed, with the ability to manage multiple workstreams under tight timelines and close attention to detail\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven AI infrastructure company building the go-to platform for developers to build and deliver their AI applications. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join us in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI\u0026#39;s investors include leading venture and growth investors such as General Catalyst, Kleiner Perkins, Salesforce Ventures, NVIDIA, Emergence Capital and Lux Capital.\u0026amp;nbsp; Together is one of the fastest growing AI startups and has been named to the Forbes AI 50 list of Top Artificial Intelligence Startups.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $138k - $175k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at\u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt; https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4046768007,"name":"Finance","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5212018007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5210933007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4668041007,"location":{"name":"San Francisco"},"metadata":null,"id":5210933007,"updated_at":"2026-08-13T12:56:08-04:00","requisition_id":"290","title":"Strategic Finance Senior Associate - Revenue","company_name":"Together AI","first_published":"2026-08-13T12:56:08-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About The Role\u0026lt;/strong\u0026gt;\u0026amp;nbsp;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Strategic Finance Senior Associate, you will own the company’s end-to-end revenue forecast as we scale. You will model a dynamic revenue base shaped by customer usage, new GPU deployments, product mix, and other key business drivers, translating these inputs into actionable insights for leadership. Beyond forecasting, you will proactively identify opportunities to accelerate revenue growth and improve business performance. You will serve as a strategic thought partner to teams across the company, including GTM, Product, and Infrastructure Engineering, helping to evaluate tradeoffs and drive execution against Together’s growth priorities.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;This is a hands-on, high ownership role. You will be part of a lean Strategic Finance team, with the opportunity to help drive strategic initiatives across the business as well as shape the strategic finance function of a rapidly growing AI infrastructure startup.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own the revenue model and forecast across sales-led and self-serve motions, partnerships, new business, renewals, churn, and new hardware deployments - spanning both committed contracts and consumption/usage-based revenue\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Prepare clear, concise analyses and reporting for leadership on revenue performance, forecast risk, and the drivers behind variances\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build expertise to deeply understand the company\u0026#39;s product offerings, product portfolio and associated nuances\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Go beyond reporting the number: proactively identify levers to grow revenue and work with Product, GTM, and Infrastructure Engineering to execute them\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Assist with long-range planning (LRP) and building the company\u0026#39;s overall financial strategy.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build repeatable processes and tooling (dashboards, models, reconciliations) that make revenue forecasting faster and more reliable each quarter\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Be ready to roll up your sleeves and assist with other strategic finance initiatives and priorities to drive business impact and help the company grow\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;2 – 4 years of experience, ideally 2 years in investment banking or consulting paired with 1 – 2 years in private equity, growth equity, or venture capital; strategic finance or FP\u0026amp;amp;A experience at a high-growth company is also a strong fit\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Exceptional financial modeling skills and close attention to detail, with the discipline to build models that are auditable and the instinct to catch errors before they reach leadership\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advanced Excel / Google Sheets skills; familiarity with SQL is a plus\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to build relationships with cross-functional partners in GTM, Infrastructure Engineering, and Product\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Exceptional business judgment and intuition, with the creativity and problem-solving instincts to tackle novel questions where there is no established playbook\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong work ethic and self-directed, with the ability to manage multiple workstreams under tight timelines\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with usage-based or consumption revenue models is a plus\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven AI infrastructure company building the go-to platform for developers to build and deliver their AI applications. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join us in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Together AI\u0026#39;s investors include leading venture and growth investors such as General Catalyst, Kleiner Perkins, Salesforce Ventures, NVIDIA, Emergence Capital and Lux Capital. Together is one of the fastest growing AI startups and has been named to the Forbes AI 50 list of Top Artificial Intelligence Startups.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $138k – $175k + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4046768007,"name":"Finance","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5210933007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4188119007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4125604007,"location":{"name":"San Francisco"},"metadata":null,"id":4188119007,"updated_at":"2026-07-10T19:42:33-04:00","requisition_id":"15","title":"Systems Research Engineer, GPU Programming","company_name":"Together AI","first_published":"2024-01-16T09:59:54-05:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Systems Research Engineer specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. Working closely with the modeling and algorithm team, you will co-design GPU kernels and model architecture to enhance the performance and efficiency of our AI systems. Collaborating with the hardware and software teams, you will contribute to the co-design of efficient GPU architectures and programming models, leveraging your expertise in GPU programming and parallel computing. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Strong background in GPU programming and parallel computing, such as CUDA and/or Triton.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Knowledge of ML/AI applications and models\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Knowledge of performance profiling and optimization tools for GPU programming\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent problem-solving and analytical skills\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Bachelor\u0026#39;s, Master\u0026#39;s, or Ph.D. degree in Computer Science, Electrical Engineering, or equivalent practical experiences\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Optimize and fine-tune GPU code to achieve better performance and scalability\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Stay up-to-date with the latest advancements in GPU programming techniques and technologies\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $160,000 - $230,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;Please see our privacy policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt; \u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026amp;nbsp;\u0026lt;/p\u0026gt;","departments":[{"id":4033059007,"name":"Research","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4188119007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5181912007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4653813007,"location":{"name":"San Francisco"},"metadata":null,"id":5181912007,"updated_at":"2026-08-20T17:41:16-04:00","requisition_id":"275","title":"Technical Compute Qualification Manager","company_name":"Together AI","first_published":"2026-07-20T15:00:46-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;About The Role\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is growing its compute footprint, and making sure new capacity meets our technical standards is an important priority for the company. Every new cluster has to clear a technical bar before it carries customer workloads, and this role owns that bar. As Technical Compute Qualification Manager, you will run the process that screens and qualifies prospective compute providers, taking each prospective\u0026amp;nbsp; deployment through a structured evaluation across compute, networking, storage, power, cooling, and operations.\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;You will coordinate various engineering partners through validation, review provider specifications and test results, and produce clear go/no-go recommendations on whether new capacity meets our standards. It is a high-impact, process-driven role for someone technical enough to know when a spec sheet does not add up, and additional diligence needs to be completed, and organized enough to drive many evaluations to closure in parallel. \u0026quot;You will deep-dive into critical hardware performance metrics, proactively identifying potential bottlenecks in cluster architecture before they impact our end customers training or inference workloads.\u0026quot; Conduct diligence and work with engineering teams to make assessments regarding technical and operational resilience.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Own and continuously improve the end-to-end qualification process for new compute capacity, from initial provider intake through final go/no-go recommendation.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Run multiple provider evaluations in parallel, setting timelines, tracking status, and keeping every stakeholder aligned on what is needed and by when.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Partner with infrastructure engineering, network engineering, data center engineering, and SRE teams to plan and coordinate technical validation, then translate their findings into clear decisions for leadership.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Review provider technical specifications and questionnaire responses for completeness and accuracy, flagging gaps, inconsistencies, and risks that warrant follow-up.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Conduct first-pass analysis of provider data yourself: compare specifications across suppliers , sanity-check performance claims, and surface issues before deeper engineering review.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain the standards, templates, and documentation that define what meets spec across compute, networking, storage, power, cooling, and operational support.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Build a structured, auditable record of evaluation outcomes that informs sourcing decisions and scales the qualification function as the team grows.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;5+ years in technical program or project management, infrastructure program management, or a comparable technical operations role, ideally involving hardware, data center, or large-scale compute environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to run multiple complex, cross-functional workstreams to deadline, with strong organization and stakeholder management.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Working technical fluency across data center infrastructure: server and GPU hardware, high-performance networking (InfiniBand or Ethernet fabrics), storage, and power and cooling fundamentals; enough depth to read a detailed technical specification and know what to question.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Hands-on comfort with data: able to write scripts or queries (for example, Python or SQL) to compare, validate, and analyze provider specifications and test results independently.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent written and verbal communication; able to turn dense technical detail into clear recommendations for both engineers and executives.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Willingness to travel to provider and data center sites as needed.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Nice to Have\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Experience qualifying, commissioning, or accepting GPU clusters or HPC infrastructure against defined performance and reliability standards.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with AI training and inference infrastructure, including interconnect topologies, cluster bring-up, and acceptance testing.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience in AI/HPC cluster design.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Background working directly with hardware vendors, colocation providers, or cloud capacity providers.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an AI-native cloud company building the infrastructure to make AI faster, cheaper, and more accessible. We’re rapidly scaling our GPU footprint: signing our own data center leases, building large-scale clusters, and expanding toward a global owned-infrastructure presence. Our research team has contributed to breakthroughs like FlashAttention, Hyena, and RedPajama, and we co-design across software, hardware, and algorithms to push the frontier of AI efficiency.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $200-250K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4044336007,"name":"Business Operations","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5181912007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/4840844007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4480571007,"location":{"name":"Pune or Bangalore, India"},"metadata":null,"id":4840844007,"updated_at":"2026-08-11T21:23:51-04:00","requisition_id":"155","title":"Technical Support Engineer (GPU Cluster), India ","company_name":"Together AI","first_published":"2025-08-28T21:59:02-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the Role \u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Customer Support Engineer at a pioneering AI company, you\u0026#39;ll be the first line of defense to support customers as they build out training, fine tuning, and inference solutions with Together AI. You\u0026#39;ll dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. As a part of the Customer Experience organization, you will collaborate closely with product and sales, driving continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Engage directly with customers to tackle and resolve complex technical challenges involving our cutting-edge Kubernetes GPU clusters; ensure swift and effective solutions every time.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Become a product expert in our GPU Cluster service, serving as the last line of technical defense before issues are escalated to Engineering and Product teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate seamlessly across Engineering, Research, and Product teams to address customer concerns; collaborate with senior leaders both internally and externally to ensure the highest levels of customer satisfaction.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Transform customer insights into action by identifying patterns in support cases and working with Engineering and Go-To-Market teams to drive Together’s roadmap (e.g., future models to support)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain detailed documentation of system configurations, procedures, troubleshooting guides, and FAQs to facilitate knowledge sharing with team and customers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Be flexible in providing support coverage during holidays, nights and weekends as required by business needs to ensure consistent and reliable service for our customers.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;3+ years of experience in a customer-facing technical role with at least 1 year in a support function in AI or supporting a mission-critical API in SaaS\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong technical background, with knowledge of AI, ML, GPU technologies and their integration into high-performance computing (HPC) environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible) high-performance network fabrics, NFS-based storage management, container infrastructure, and scripting and programming languages.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Foundational understanding in the installation, configuration, administration, troubleshooting, and securing of compute clusters.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Complex technical problem solving and troubleshooting, with a proactive approach to issue resolution\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to work cross-functionally with teams such as Sales, Engineering, Support, Product and Research to drive customer success.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong sense of ownership and willingness to learn new skills to ensure both team and customer success.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to operate in dynamic environments, adept at managing multiple projects, and comfortable with frequent context switching and prioritization.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong data-start=\u0026quot;522\u0026quot; data-end=\u0026quot;543\u0026quot;\u0026gt;Working Schedule\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;This position operates on a non-standard workweek to support business needs. The schedule will be either\u0026amp;nbsp;Saturday–Wednesday or Wednesday–Sunday, with two weekly off days.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work for the respective hiring region. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/4840844007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5202015007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"education":"education_optional","internal_job_id":4663730007,"location":{"name":"Remote"},"metadata":null,"id":5202015007,"updated_at":"2026-08-11T13:37:27-04:00","requisition_id":"286","title":"Technical Support Engineer (GPU Clusters) - US Weekends","company_name":"Together AI","first_published":"2026-08-04T18:24:34-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About the role\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;As a Technical Support Engineer at a pioneering AI company, you\u0026#39;ll be the first line of defense to support customers as they build out training, fine tuning, and inference solutions with Together AI. You\u0026#39;ll dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. As a part of the Customer Experience organization, you will collaborate closely with product and sales, driving continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Required hours\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;This is a fulltime position working US daytime hours. The role will work both weekend days (Saturday and Sunday) as well as two additional weekdays.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;This is a 4-day shift, 10 hours per day, with 2 additional hours of on-call coverage on Saturdays and Sundays.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;The role would start as a Monday to Friday role for the first few months to allow for ramping up and learning from teammates. After being considered fully ramped, the role would transition to the 4-day weekend shift.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Engage directly with customers to tackle and resolve complex technical challenges involving our cutting-edge Kubernetes GPU clusters; ensure swift and effective solutions every time.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Act as a customer facing SRE to ensure our customer’s Kubernetes clusters remain healthy and stable\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Become a product expert in our GPU Cluster service, serving as the last line of technical defense before issues are escalated to Engineering and Product teams.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Monitor GPU cluster health and proactively communicate hardware issues to customers (thermal throttling, BMC failures, missing GPUs, and NVLink/InfiniBand degradation) with clear remediation steps\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Operate and maintain production infrastructure for enterprise GPU customers, including fleet rebalancing, Slurm cluster maintenance, node repair/migration, and Kubernetes-based workload management\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Investigate and resolve storage and networking issues such as Weka filesystem degradation, InfiniBand link failures, and bandwidth anomalies on bare-metal and VM environments\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Collaborate seamlessly across Engineering, Research, and Product teams to address customer concerns; collaborate with senior leaders both internally and externally to ensure the highest levels of customer satisfaction.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Transform customer insights into action by identifying patterns in support cases and working with Engineering and Go-To-Market teams to drive Together’s roadmap (e.g., future models to support)\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Maintain detailed documentation of system configurations, procedures, troubleshooting guides, and FAQs to facilitate knowledge sharing with team and customers.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Be flexible in providing support coverage during holidays, nights and weekends as required by business needs to ensure consistent and reliable service for our customers.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;3+ years of experience in a customer-facing technical role with at least 1 year in a support function for an AI service or supporting a mission-critical API in SaaS\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience as an SRE or DevOps engineer working with Kubernetes\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong technical background, with knowledge of AI, ML, GPU technologies and their integration into high-performance computing (HPC) environments.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Advanced knowledge with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible) high-performance network fabrics, NFS-based storage management, container infrastructure, and scripting and programming languages.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience with HPC/Slurm cluster environments — node draining, job scheduling, maintenance workflows\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Familiarity with high-speed networking concepts — InfiniBand, RDMA, network interface diagnostics\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience with distributed storage systems (e.g., Weka, NFS) and troubleshooting I/O and bandwidth issues\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Foundational understanding in the installation, configuration, administration, troubleshooting, and securing of compute clusters.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Complex technical problem solving and troubleshooting, with a proactive approach to issue resolution\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Ability to work cross-functionally with teams such as Sales, Engineering, Support, Product and Research to drive customer success.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong sense of ownership and willingness to learn new skills to ensure both team and customer success.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Ability to operate in dynamic environments, adept at managing multiple projects, and comfortable with frequent context switching and prioritization.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $160K - $230K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5202015007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5069532007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4600753007,"location":{"name":"Pune or Bangalore, India"},"metadata":null,"id":5069532007,"updated_at":"2026-08-11T21:25:16-04:00","requisition_id":"S\u0026M-CSX-028","title":"Technical Support Engineer (Inference) - India Weekends","company_name":"Together AI","first_published":"2026-03-10T16:37:31-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About the role\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;As a Technical Support Engineer at a pioneering AI company, you\u0026#39;ll be the first line of defense to support customers as they build out training, fine tuning, and inference solutions with Together AI. You\u0026#39;ll dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. As a part of the Customer Experience organization, you will collaborate closely with product and sales, driving continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Required hours\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;This is a fulltime position working India daytime hours. The role will work both weekend days (Saturday and Sunday) as well as two additional weekdays.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;This is a 4-day shift, 10 hours per day, with 2 additional hours of on-call coverage on Saturdays and Sundays.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;The role would start as a Monday to Friday role for the first few months to allow for ramping up and learning from teammates. After being considered fully ramped, the role would transition to the 4-day weekend shift.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;Engage directly with customers to tackle and resolve complex technical challenges involving our cutting-edge GPU clusters and our inference and fine-tuning services; ensure swift and effective solutions every time.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Act as a customer facing SRE to ensure our customer’s Inference endpoints (running on Kubernetes) remain healthy, stable, and performant\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Become a product expert in all of our Gen AI solutions, serving as the last line of technical defense before issues are escalated to Engineering and Product teams.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Assist with hardware and platform migrations by validating system health and traffic routing. Monitor dashboards to detect anomalies and escalate with data-backed analysis\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Manage customer-facing communications during incidents and degradations; translate deep technical findings (latency regressions, provider issues, network reachability drops) into clear, evidence-backed updates without exposing platform internals\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Contribute infrastructure changes for model deployment, capacity rebalancing, and cluster configuration. You will execute infrastructure changes via pull requests (infra-as-code) for tasks such as endpoint configuration, model bringup/bringdown, and capacity scaling\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Flag engine-level bugs with logs and reproduction steps for engineering\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Collaborate seamlessly across Engineering, Research, and Product teams to address customer concerns; collaborate with senior leaders both internally and externally to ensure the highest levels of customer satisfaction.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Transform customer insights into action by identifying patterns in support cases and working with Engineering and Go-To-Market teams to drive Together’s roadmap (e.g., future models to support)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Maintain detailed documentation of system configurations, procedures, troubleshooting guides, and FAQs to facilitate knowledge sharing with team and customers.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Be flexible in providing support coverage during holidays, nights and weekends as required by business needs to ensure consistent and reliable service for our customers.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li\u0026gt;6+ years of experience in a customer-facing technical role, SRE, DevOps, or infrastructure engineering, with at least 1 year in a support role for an AI service\u0026amp;nbsp;\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong technical background, with knowledge of AI, ML, GPU technologies and their integration into high-performance computing (HPC) environments.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Advanced, production-level experience with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible) high-performance network fabrics, NFS-based storage management, and container infrastructure\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Familiarity with operating storage systems in HPC environments such as Vast and Weka\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Proven ability to diagnose complex network-layer issues and read traces\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong knowledge of Python, TypeScript, and/or JavaScript with testing/debugging experience using curl and Postman-like tools\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Demonstrated expertise with observability tooling (e.g., Prometheus, Grafana) at scale\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Deep familiarity with REST API debugging and HTTP semantics\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with LLM inference frameworks and LoRA fine-tuning and common training failure modes\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Experience with Infrastructure as Code and Git-based workflows\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Background in GPU cluster management\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Cloud platform experience (AWS, GCP, and/or Azure)\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Foundational understanding in the installation, configuration, administration, troubleshooting, and securing of compute clusters.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Complex technical problem solving and troubleshooting, with a proactive approach to issue resolution\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to work cross-functionally with teams such as Sales, Engineering, Support, Product and Research to drive customer success.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Strong sense of ownership and willingness to learn new skills to ensure both team and customer success.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.\u0026lt;/li\u0026gt;\n\u0026lt;li\u0026gt;Ability to operate in dynamic environments, adept at managing multiple projects, and comfortable with frequent context switching and prioritization.\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work for the respective hiring region. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029318007,"name":"Remote","location":"Remote","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5069532007/ai_opt_out"},{"absolute_url":"https://job-boards.greenhouse.io/togetherai/jobs/5202021007","data_compliance":[{"type":"gdpr","requires_consent":false,"requires_processing_consent":false,"requires_retention_consent":false,"retention_period":null,"demographic_data_consent_applies":false}],"internal_job_id":4663734007,"location":{"name":"Remote"},"metadata":null,"id":5202021007,"updated_at":"2026-08-09T22:50:48-04:00","requisition_id":"287","title":"Technical Support Engineer (Inference) - US Weekends","company_name":"Together AI","first_published":"2026-08-04T18:21:50-04:00","language":"en","application_deadline":null,"content":"\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About the role\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;As a Technical Support Engineer at a pioneering AI company, you\u0026#39;ll be the first line of defense to support customers as they build out training, fine tuning, and inference solutions with Together AI. You\u0026#39;ll dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. As a part of the Customer Experience organization, you will collaborate closely with product and sales, driving continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Required hours\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;This is a fulltime position working US daytime hours. The role will work both weekend days (Saturday and Sunday) as well as two additional weekdays.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;This is a 4-day shift, 10 hours per day, with 2 additional hours of on-call coverage on Saturdays and Sundays.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;The role would start as a Monday to Friday role for the first few months to allow for ramping up and learning from teammates. After being considered fully ramped, the role would transition to the 4-day weekend shift.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Responsibilities\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Engage directly with customers to tackle and resolve complex technical challenges involving our cutting-edge GPU clusters and our inference and fine-tuning services; ensure swift and effective solutions every time.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Act as a customer facing SRE to ensure our customer’s Inference endpoints (running on Kubernetes) remain healthy, stable, and performant\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Become a product expert in all of our Gen AI solutions, serving as the last line of technical defense before issues are escalated to Engineering and Product teams.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Assist with hardware and platform migrations by validating system health and traffic routing. Monitor dashboards to detect anomalies and escalate with data-backed analysis\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Manage customer-facing communications during incidents and degradations; translate deep technical findings (latency regressions, provider issues, network reachability drops) into clear, evidence-backed updates without exposing platform internals\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Contribute infrastructure changes for model deployment, capacity rebalancing, and cluster configuration. You will execute infrastructure changes via pull requests (infra-as-code) for tasks such as endpoint configuration, model bringup/bringdown, and capacity scaling\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Flag engine-level bugs with logs and reproduction steps for engineering\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Collaborate seamlessly across Engineering, Research, and Product teams to address customer concerns; collaborate with senior leaders both internally and externally to ensure the highest levels of customer satisfaction.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Transform customer insights into action by identifying patterns in support cases and working with Engineering and Go-To-Market teams to drive Together’s roadmap (e.g., future models to support)\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Maintain detailed documentation of system configurations, procedures, troubleshooting guides, and FAQs to facilitate knowledge sharing with team and customers.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Be flexible in providing support coverage during holidays, nights and weekends as required by business needs to ensure consistent and reliable service for our customers.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Requirements\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;ul\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;6+ years of experience in a customer-facing technical role, SRE, DevOps, or infrastructure engineering, with at least 1 year in a support role for an AI service\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience as an SRE or DevOps engineer working with Kubernetes\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong technical background, with knowledge of AI, ML, GPU technologies and their integration into high-performance computing (HPC) environments.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Advanced, production-level experience with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible) high-performance network fabrics, NFS-based storage management, and container infrastructure\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Familiarity with operating storage systems in HPC environments such as Vast and Weka\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Proven ability to diagnose complex network-layer issues and read traces\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong knowledge of Python, TypeScript, and/or JavaScript with testing/debugging experience using curl and Postman-like tools\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Demonstrated expertise with observability tooling (e.g., Prometheus, Grafana) at scale\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Deep familiarity with REST API debugging and HTTP semantics\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience with LLM inference frameworks and LoRA fine-tuning and common training failure modes\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Experience with Infrastructure as Code and Git-based workflows\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Background in GPU cluster management\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Cloud platform experience (AWS, GCP, and/or Azure)\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Foundational understanding in the installation, configuration, administration, troubleshooting, and securing of compute clusters.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Complex technical problem solving and troubleshooting, with a proactive approach to issue resolution\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Ability to work cross-functionally with teams such as Sales, Engineering, Support, Product and Research to drive customer success.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Strong sense of ownership and willingness to learn new skills to ensure both team and customer success.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;li style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Ability to operate in dynamic environments, adept at managing multiple projects, and comfortable with frequent context switching and prioritization.\u0026lt;/span\u0026gt;\u0026lt;/li\u0026gt;\n\u0026lt;/ul\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;About Together AI\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.\u0026amp;nbsp;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Compensation\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full-time position is: $160K - $230K + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;h3\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;\u0026lt;strong\u0026gt;Equal Opportunity\u0026lt;/strong\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/h3\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;\n\u0026lt;p\u0026gt;\u0026lt;span style=\u0026quot;font-family: helvetica, arial, sans-serif;\u0026quot;\u0026gt;Please see our Privacy Policy at \u0026lt;a href=\u0026quot;https://www.together.ai/privacy\u0026quot;\u0026gt;https://www.together.ai/privacy\u0026lt;/a\u0026gt;\u0026lt;/span\u0026gt;\u0026lt;/p\u0026gt;","departments":[{"id":4062689007,"name":"Customer Success","child_ids":[],"parent_id":null}],"offices":[{"id":4029317007,"name":"San Francisco","location":"San Francisco, California, United States","child_ids":[],"parent_id":null}],"ai_disclaimer":"\u003cp\u003eWe use Greenhouse’s AI-powered Talent Matching tool to compare your application against our job requirements.\u003c/p\u003e","include_ai_disclaimer":false,"ai_opt_out_request_url":"http://app7.greenhouse.io/ai_opt_out_request/job_post/5202021007/ai_opt_out"}],"meta":{"total":64}}