Together AI Headlines
Latest news and coverage for Together AI
Recent Headlines
32 headlinesNew York Times
As Companies Race for Cheaper A.I. Options, This Start-Up Pitches a Solution
NYT profile of Together AI, featuring interviews with the founding team
Together AI Blog
Announcing our $800M Series C to accelerate the shift to open-source AI
Together AI announces its Series C funding of $800 million from investors including Aramco Ventures, NVIDIA, Vista Equity, General Catalyst, Emergence Capital, SE Ventures, Pegatron, Salesforce Ventures, March Capital, DTCP Growth, Lux Capital, Geodesic, PSP Partners, and others.
TipRanks
Together AI Advances Open-Source Model Economics and Voice Agent Capabilities - TipRanks.com
Together AI highlighted new benchmarks showing open-source models can match closed systems at lower cost, and integrated Cartesia's Sonic 3.5 for real-time voice agents.
StartupHub.ai
Together AI Locks Down Enterprise Trust
Together AI achieved ISO 27001:2022 certification, enhancing security and trust for its enterprise AI platform.
BuzzRAG
AI Agents With 5M-Token Memory Raise Privacy | BuzzRAG
BuzzRAG analyzes Together AI's research on 5M-token context windows, discussing privacy implications of long-context AI agents.
SendTech Times
Together AI Taps Rumble for Dedicated Blackwell Cloud Capacity | SendTech Times
Together AI signed a multi-year agreement with Rumble for dedicated Nvidia HGX B300 capacity, expanding its Blackwell-class compute options.
AiThority
Rumble Signs Agreement with Together AI to Deploy NVIDIA Blackwell-Powered AI Compute as a Service
Rumble and Together AI announce a multi-year agreement for dedicated NVIDIA HGX B300 GPU capacity.
Data Center Dynamics
Together AI taps Rumble for Nvidia Blackwell GPU capacity - DCD
Together AI has signed a multi-year cloud capacity agreement with Rumble Inc. for dedicated Nvidia HGX B300 systems.
myMotherLode.com
Financial News | myMotherLode.com - Rumble Signs Agreement with Together AI to Deploy NVIDIA Blackwell-Powered AI Compute as a Service
Press release: Together AI and Rumble agree on multi-year GPU cloud capacity deal.
Web3Wire
Rumble Signs Agreement with Together AI to Deploy NVIDIA Blackwell-Powered AI Compute as a Service | Web3Wire
Together AI signs a multi-year agreement with Rumble for dedicated NVIDIA Blackwell GPU capacity.
FinancialContent
Rumble Signs Agreement with Together AI to Deploy NVIDIA Blackwell-Powered AI Compute as a Service | FinancialContent
Rumble and Together AI announced a multi-year agreement for dedicated NVIDIA Blackwell GPU capacity, as reported via GlobeNewswire.
Business Insider
Rumble Signs Agreement with Together AI to Deploy NVIDIA Blackwell-Powered AI Compute as a Service
Rumble and Together AI announced a multi-year agreement for Together AI to purchase dedicated GPU cloud capacity powered by NVIDIA HGX B300 systems.
Business Insider
Together AI commits to purchase dedicated GPU cloud capacity from Rumble
Rumble announced a multi-year agreement for Together AI to purchase dedicated GPU cloud capacity powered by NVIDIA HGX B300 systems.
Silicon Report
Together AI details Nvidia H100 cluster testing — Silicon Report
Together AI offers H100 clusters with acceptance testing for up to 2048 GPUs, emphasizing reliability and bring-up risk reduction.
MAXBIT
Together AI Claims Fastest Speech-to-Text Stack with Parakeet v3 – MAXBIT
Together AI claims fastest speech-to-text stack using NVIDIA Parakeet v3 and Whisper, achieving 20 hours of speech in under 10 seconds.
GogoAI News
Together AI OSCAR: 2-Bit LLM Inference Breakthrough - GogoAI News
Together AI open-sourced OSCAR, an INT2 KV cache quantization method for long-context LLMs, reducing memory usage significantly.
MarkTechPost
Together AI Open-Sources OSCAR: An Attention-Aware 2-Bit KV Cache Quantization System for Long-Context LLM Serving - MarkTechPost
Together AI open-sources OSCAR, a 2-bit KV cache quantization system that reduces memory usage and speeds up long-context LLM serving.
Hamidun News
Together AI introduced ATLAS: a speculator that speeds up LLMs 4x — Hamidun News
Together AI introduced ATLAS, an ML-based speculator that speeds up LLM inference 4x without manual tuning.
AIMultiple
Top 9 AI Providers Compared
An analysis comparing AI providers including Together AI, detailing its capabilities, limitations, and use cases.
Blockchain.News
Together AI Joins Pearl Labs to Cut AI Inference Costs With Blockchain
Together AI partners with Pearl Labs to reduce AI inference costs by 25% using blockchain-based Proof of Useful Work.
Digg
Together AI launches serverless inference endpoint for Gemma-4-31B-it-Pearl · Digg
Together AI launched a serverless inference endpoint for the Gemma-4-31B-it-Pearl model through a partnership with Pearl Research Labs, offering discounted pricing.
Blockchain.News
DeepSeek-V4 Tackles Million-Token Context on NVIDIA HGX B200 - Blockchain.News
Together AI's DeepSeek-V4 model introduces a 1 million-token context window with hybrid attention, running on NVIDIA HGX B200 hardware.
StartupHub.ai
Together AI Supercharges LLM Inference | StartupHub.ai
Together AI unveils ATLAS, an adaptive speculative decoding system that accelerates LLM inference up to 4x, addressing rising inference costs.
Blockchain.News
NVIDIA Nemotron 3 Nano Omni Launches on Together AI for Multimodal AI - Blockchain.News
Together AI launches NVIDIA's Nemotron 3 Nano Omni model, a multimodal AI model for reasoning across video, audio, and text.
Together AI Blog
Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0
Together AI announced the availability of NVIDIA Nemotron 3 Nano Omni on its platform on launch day, offering developers access to a multimodal AI model that reasons across video, images, audio, and text. The model uses a hybrid Mamba-Transformer MoE architecture with 30B total parameters, activating ~3B per token, and is designed for agentic production workloads.
COSS Weekly Newsletter
Stay up to date with the latest news, funding rounds, and announcements from the COSS universe.
Check out COSS Weekly on the webLatest Content from Chinstrap Community
View allCOSS Weekly – Week of July 27, 2026
This week in COSS: On the funding front, Databricks raised $3 billion at a $188 billion valuation, e...
COSS Weekly – Week of July 20, 2026
This week in COSS: Nous Research, the startup behind the OSS Hermes agent, is in talks to raise mew ...
Battle of the Software Acronyms: BYOC Is Beating SaaS, but the Winner is COSS
On June 15, 2026, Ververica, a company founded by the original creators of the open source Apache Fl...
COSS Weekly – Week of July 13, 2026
This week in COSS: Ollama raises a $65M Series B, Bespoke Labs announces a $40M Seed and Series A, a...
COSS Weekly – Week of July 6, 2026
This week in COSS: Together AI announced an $800M Series C to accelerate the shift to open-source AI...
COSS Weekly – Week of June 29, 2026
This week in COSS: The acquisition trend continued as Qualcomm agreed to acquire Modular for nearly ...
COSS Weekly – Week of June 22, 2026
This week in COSS: Databricks reported annualized revenue of $6.9 billion — up over 80% year-over-ye...
COSS Weekly – Week of June 15, 2026
This week in COSS: The recent flurry of COSS M&A activity continues as VoidZero was acquired by Clou...

