
Together AI News
Latest news, updates, and announcements from Together AI.
IBM and Together AI have entered a $240 million multi-year agreement to construct an AI inference cl...
IBM and Together AI have entered a $240 million multi-year agreement to construct an AI inference cluster on IBM Cloud, powered by Nvidia's HGX B300 systems. The deal highlights the growing demand for...

Together AI raises $800M at $8.3B valuation to expand its neocloud for open-source AI models
Together AI raises $800M at $8.3B valuation to expand its neocloud for open-source AI models Together AI, a neocloud provider renting Nvidia GPU clusters for open-source AI models, announced an $800 m...

Together AI is raising $1B at a $7.5B valuation, more than doubling its worth in just one year while...
Together AI is raising $1B at a $7.5B valuation, more than doubling its worth in just one year while revenue tripled to $1B annually. This massive funding round signals the neocloud market's rapid mat...

Together AI is raising $1B at a $7.5B valuation, more than doubling from $3.3B in early 2025, with a...
Together AI is raising $1B at a $7.5B valuation, more than doubling from $3.3B in early 2025, with annual revenue reaching approximately $1B. This Nvidia-backed company's meteoric rise reflects a broa...

Together AI's new ATLAS Adaptive Speculator is a major breakthrough, achieving up to a 400% speedup...
Together AI's new ATLAS Adaptive Speculator is a major breakthrough, achieving up to a 400% speedup for large language model inference by learning from real-time workloads. This adaptive speculative d...

Together AI’s new ATLAS adaptive speculator technique marks a major breakthrough in LLM deployment e...
Together AI’s new ATLAS adaptive speculator technique marks a major breakthrough in LLM deployment efficiency. The system delivers up to a 400% speedup in inference speed compared to existing systems...

Together AI just changed the economics of LLM deployment with the ATLAS adaptive speculator, deliver...
Together AI just changed the economics of LLM deployment with the ATLAS adaptive speculator, delivering a 400% speedup in inference. 🚀 This breakthrough combines optimizations like FP4 quantization a...