Command Palette
Search for a command to run...

DeepSeek V4 Pro on Together AI Ranks No. 1 on Artificial Analysis for Output Speed, Latency

aiai-infrastructureai-inference-platforms 4 posts · 3 accounts

DeepSeek V4 Pro on Together AI is now ranked No. 1 on Artificial Analysis for both output speed and latency, according to Together AI.

Together AI said serving V4 well is an inference-systems problem involving KV cache, prefix reuse, kernels and endpoint profiles, and tied the ranking to that systems work.

From the sources (4 posts)

@vipulved

DeepSeek V4 Pro on @togethercompute becomes #1 on both latency and speed.

@tri_dao

RT @vipulved: DeepSeek V4 Pro on @togethercompute becomes #1 on both latency and speed.

@togethercompute

RT @vipulved: DeepSeek V4 Pro on @togethercompute becomes #1 on both latency and speed.

@togethercompute

DeepSeek V4 Pro on Together AI is now #1 on Artificial Analysis for both output speed and latency. Serving V4 well is an inference systems problem: KV cache, prefix reuse, kernels, and endpoint profiles. We break down the systems work her

Preview built on a synthetic news corpus (16 weeks, Apr–Jul 2026). Impact calls are model reads, not price data.

About Archive