{"id":"31017109-25fb-4005-bbe9-95d017b0b1b9","name":"Together AI","slug":"together-ai","description":"Together AI is a San Francisco-based GPU cloud and full-stack AI platform founded in 2022 by Vipul Ved Prakash (CEO), Ce Zhang (CTO), Chris Ré, Tri Dao (Chief Scientist), and Percy Liang. Its positioning sits between the pure-metal neoclouds (CoreWeave, Crusoe, Lambda) and the model-API vendors: customers can rent bare GPU clusters by the hour, buy dedicated single-tenant inference endpoints, or hit a serverless per-token API against a catalog of open-source and partner models. That three-lane packaging, backed by systems research the founders brought from Stanford (FlashAttention, Hyena, RedPajama), is the company's main distinctive versus a rack-focused operator.\\n\\nCommercially the company is inference-and-fine-tuning first, with a training-cluster side business. Published pricing puts on-demand H100 clusters at $3.99 per GPU hour, dedicated HGX H100 endpoints at $5.49 per hour, and B200 dedicated at $8.99 per hour, with reserved terms of 7 to 180-plus days pulling H100 rates down toward roughly $3.19. Customers named on the site include Salesforce, Cohere, ElevenLabs, Cursor, Decagon, and DeepMind-adjacent teams; a dedicated GPU cluster program for Y Combinator startups is one of the more visible distribution moves. Nvidia is the anchor silicon partner (Together is listed as an Nvidia Preferred Partner) with early B200 availability being the current headline.\\n\\nCapital-wise Together raised a $305M Series B in early 2025 (Salesforce Ventures, General Catalyst, others), followed by an $800M Series C the company announced in 2026 to expand open-source AI infrastructure. Prior backers across the stack include NEA, Kleiner Perkins, Coatue, Lux Capital, and Nvidia's venture arm. The company has not publicly disclosed a total operational GPU count or MW footprint; it operates through colocation partners rather than owned campuses.","status":"active","founded":2022,"hq":"San Francisco, California, United States","hqLat":null,"hqLng":null,"hqGeocodedAt":null,"hqAddress":null,"hqPrecision":null,"website":"https://www.together.ai","fundingTotal":"1230000000","type":null,"roles":[],"reviewStatus":"reviewed","lifecycleState":"active","supersededByCompanyId":null,"sources":[{"url":"https://www.together.ai/about","title":"Together AI - Company / About","publisher":"Together AI"},{"url":"https://www.together.ai/pricing","title":"Together AI - Pricing (GPU clusters, dedicated inference, serverless)","publisher":"Together AI"},{"url":"https://www.together.ai/blog","title":"Together AI - Blog (Series C announcement, YC partnership, Moonshot integration)","publisher":"Together AI"},{"url":"https://www.together.ai","title":"Together AI - Homepage (customer logos, product overview)","publisher":"Together AI"}],"keyFacts":[{"label":"Headquarters","value":"San Francisco, California","sourceUrl":"https://www.together.ai"},{"label":"Founded","value":"2022 by Vipul Ved Prakash, Ce Zhang, Chris Re, Tri Dao, and Percy Liang","sourceUrl":"https://www.together.ai/about"},{"label":"Latest funding round","value":"$800M Series C announced 2026, positioned to accelerate open-source AI infrastructure; lead investors not detailed on public blog post","sourceUrl":"https://www.together.ai/blog"},{"label":"Prior Series B","value":"$305M raised February 2025 led by General Catalyst with participation from Salesforce Ventures and existing backers","sourceUrl":"https://www.together.ai/blog"},{"label":"Announced hourly pricing","value":"H100 GPU clusters from $3.99/hr on-demand, reserved rates to ~$3.19/hr at 181+ day terms; dedicated HGX H100 $5.49/hr; B200 dedicated $8.99/hr","sourceUrl":"https://www.together.ai/pricing"},{"label":"Silicon partnerships","value":"Nvidia Preferred Partner; offers H100, H200, and Blackwell B200 capacity; no disclosed AMD or Cerebras deployments","sourceUrl":"https://www.together.ai/about"},{"label":"Datacenter footprint","value":"Operates via colocation partners rather than owned campuses; specific sites and total MW not disclosed","sourceUrl":"https://www.together.ai"},{"label":"GPU inventory","value":"Company has not publicly disclosed a fleet-wide GPU count; product surfaces confirm H100, H200, and B200 capacity across dedicated and cluster tiers","sourceUrl":"https://www.together.ai/pricing"},{"label":"Named customers","value":"Salesforce, Cohere, ElevenLabs, Cursor, Decagon, Vercept, plus a dedicated cluster program for Y Combinator startups","sourceUrl":"https://www.together.ai"},{"label":"Positioning","value":"Inference-and-fine-tuning first with a training-cluster side business; three-lane packaging (serverless API, dedicated endpoints, raw GPU clusters) versus pure-metal neoclouds","sourceUrl":"https://www.together.ai/pricing"},{"label":"Status","value":"Private, venture-backed; no public listing","sourceUrl":"https://www.together.ai/about"},{"label":"Compliance","value":"ISO 27001:2022 certified as of 2026","sourceUrl":"https://www.together.ai/blog"},{"label":"Notable 2026 partnership","value":"Moonshot AI integration to natively serve Kimi models including the 2.8T-parameter Kimi K3 with day-zero access","sourceUrl":"https://www.together.ai/blog"}],"aliases":[],"collisionRisk":"low","reviewNote":null,"atsProvider":null,"atsSlug":null,"factoryLocations":null,"manufacturingCapacity":null,"orgType":null,"burnSignal":null,"litigationEvents":null,"insurancePostureDisclosed":null,"unitEconomicsDisclosed":null,"tickerSymbol":null,"stockExchange":null,"secCik":null,"revenueUsd":null,"revenueBasis":null,"revenueAsOf":null,"employeeCount":300,"employeeCountBasis":"approximate headcount per LinkedIn range at time of Series C reporting; company has not published a precise figure","employeeCountAsOf":"2026-01-01T00:00:00.000Z","marketCapUsd":null,"marketCapAsOf":null,"marketCapSource":null,"createdAt":"2026-08-28T01:48:22.324Z","updatedAt":"2026-08-28T03:45:17.421Z","jsonLd":{"@context":"https://schema.org","@type":"Organization","@id":"https://registry.deploy.report/companies/together-ai","url":"https://registry.deploy.report/companies/together-ai","name":"Together AI","description":"Together AI is a San Francisco-based GPU cloud and full-stack AI platform founded in 2022 by Vipul Ved Prakash (CEO), Ce Zhang (CTO), Chris Ré, Tri Dao (Chief Scientist), and Percy Liang. Its positioning sits between the pure-metal neoclouds (CoreWeave, Crusoe, Lambda) and the model-API vendors: customers can rent bare GPU clusters by the hour, buy dedicated single-tenant inference endpoints, or hit a serverless per-token API against a catalog of open-source and partner models. That three-lane packaging, backed by systems research the founders brought from Stanford (FlashAttention, Hyena, RedPajama), is the company's main distinctive versus a rack-focused operator.\\n\\nCommercially the company is inference-and-fine-tuning first, with a training-cluster side business. Published pricing puts on-demand H100 clusters at $3.99 per GPU hour, dedicated HGX H100 endpoints at $5.49 per hour, and B200 dedicated at $8.99 per hour, with reserved terms of 7 to 180-plus days pulling H100 rates down toward roughly $3.19. Customers named on the site include Salesforce, Cohere, ElevenLabs, Cursor, Decagon, and DeepMind-adjacent teams; a dedicated GPU cluster program for Y Combinator startups is one of the more visible distribution moves. Nvidia is the anchor silicon partner (Together is listed as an Nvidia Preferred Partner) with early B200 availability being the current headline.\\n\\nCapital-wise Together raised a $305M Series B in early 2025 (Salesforce Ventures, General Catalyst, others), followed by an $800M Series C the company announced in 2026 to expand open-source AI infrastructure. Prior backers across the stack include NEA, Kleiner Perkins, Coatue, Lux Capital, and Nvidia's venture arm. The company has not publicly disclosed a total operational GPU count or MW footprint; it operates through colocation partners rather than owned campuses.","identifier":"31017109-25fb-4005-bbe9-95d017b0b1b9","foundingDate":"2022","address":"San Francisco, California, United States","publisher":{"@id":"https://deploy.report/#organization"}},"framework_metadata":{"framework_schema_version":"0.1.0","verification_status":"verified","maturity_stage":null,"lifecycle_state":null,"architectural_position":{"cohort":null,"sub_cohorts":[]},"within_cohort_verified_vs_claimed_pair":null,"cap_flags":[],"verification_depth":{"sources_count":4,"primary_source_types":[]}}}