Together AI

2,648 posts
Square profile picture and Opens profile photo
Together AI
@togethercompute
Accelerate inference, model shaping, and pre-training on a research-optimized platform.
San Francisco, CAtogether.ai

Together AI’s posts

Pinned
Square profile picture
Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit from reliable inference for long-horizon coding and agentic workflows.
Square profile picture
Switching to Together AI flipped it: ⚡️Zero training-blocking failures ⚡️~50% cost savings vs. AWS ⚡️Issues resolved within hours via shared Slack They could finally focus on building the model.
On 5/5 and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session!
Quote
Together AI
@togethercompute
Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to 90%, enabling 1M-token context. We'll cover why that makes it great for agentic workflows, what it took to serve at scale, and how to build with it. Hear from @realDanFu @JueWANG26088228
Image
Most model trainings outside of frontier labs fail. 📈 Because of bad or insufficient data. 🚮 Or just data for what you want + not general capabilities. 🎯 Most builders give up + become elevated prompt engineers. Today, + start fixing that.
Quote
adaption
@adaption_ai
We believe that intelligence should not arrive preconfigured. @togethercompute is now available directly inside the Adaption platform, connecting Adaptive Data with large-scale training in a single workflow. One platform, end to end. Stop inheriting intelligence. Shape it.
0:04 / 0:29
Square profile picture
Fine-tuning quality starts before the training run. Adaptive Data helps teams analyze, adapt, and improve datasets; Together Fine-Tuning turns those shaped datasets into specialized open models.
Square profile picture
We believe that intelligence should not arrive preconfigured. is now available directly inside the Adaption platform, connecting Adaptive Data with large-scale training in a single workflow. One platform, end to end. Stop inheriting intelligence. Shape it.
$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in.
Quote
Together AI
@togethercompute
Introducing Qwen3.6-Plus from @Alibaba_Qwen, a 1M-context model built for real-world agents, agentic coding, and multimodal reasoning. AI natives can now use Qwen3.6-Plus on Together AI and benefit from reliable inference for production-scale agent workflows.
Image
Square profile picture
Highlights: 👉 1M context for long-horizon agentic workflows 👉 Stronger agentic coding across frontend, repo-level, and terminal-based tasks 👉 Multimodal reasoning across text, image, and video inputs 👉 Production-ready on the AI Native Cloud—serverless inference at $0.50
Square profile picture
Nemotron 3 Nano Omni is now on Together AI. Enterprise multimodal AI — video, audio, image, documents & text — optimized for speed and scale. ✅ ~3B active params, 9x higher throughput ✅ Fully managed, zero infra headache ✅ Secure, zero-trust architecture Build
Image