AI Devtools Radar

Together AI

Inference platformswww.together.ai13 published changes

2 sources under watch

Change timeline

deprecation

Llama Guard 4 12B model removed from Together AI pricing page

Users relying on Llama Guard 4 12B ($0.20/token) through Together AI will need to find alternative providers or models for content moderation/safety use cases.

Evidence
Llama Guard 4 12B $0.20
+No matching models

www.together.ai

other

Together AI published benchmark comparison of GLM-5.3 vs GLM-5.3 Flash models on DeepSWE

Developers can now evaluate the cost-performance tradeoff between GLM-5.3 and its faster Flash variant, with Flash offering 17x lower cost at the expense of ~5.6 points in pass@1 performance.

Evidence
+GLM-5.3 vs. GLM-5.3 Flash on DeepSWE: Cost, Coding, and Routing We ran 900 DeepSWE rollouts on GLM-5.3 and GLM-5.3 Flash. Flash gives up 5.6 points of pass@1 at 17x lower cost, and only 2.6 points at pass@4.

www.together.ai · also seen at 1 more source

feature

Added GLM-5.3 model to pricing list

New model option available to users; GLM-5.2 remains available as an alternative.

Evidence
GLM-5.2 $1.40 $0.26 (cached) $4.40
+GLM-5.3 $1.40 $0.26 (cached) $4.40 GLM-5.2 $1.40 $0.26 (cached) $4.40

www.together.ai

other

Pricing page announcement updated to highlight DeepSeek V4 Pro and GPT-5.6 Sol comparison

Marketing messaging changed from Series C funding announcement to model comparison focus, suggesting shift in promotional strategy

Evidence
💰 Announcing our Series C. Intelligence should be abundant, not expensive →
+🚀 DeepSeek V4 Pro 0813 vs. GPT-5.6 Sol on DeepSWE →

www.together.ai

featurehigh

New DeepSeek-V4 models added to pricing table, replacing R1 and V3 variants

Users can now access newer DeepSeek-V4 model variants (Flash 0731, Flash) at $6.00-$15.00 pricing, replacing older R1 and V3 series offerings

Evidence
DeepSeek-R1 DeepSeek-R1-0528 DeepSeek-V3 DeepSeek-V3-0324 DeepSeek-V3.1-Base
+DeepSeek-V4 Flash 0731 DeepSeek-V4 Flash $6.00 $15.00 $12.00

www.together.ai · also seen at 1 more source

pricinghigh

New DeepSeek V4 Pro 0813 model introduced with different pricing

Users now have access to a new DeepSeek V4 Pro 0813 variant with lower input pricing ($1.32 vs $1.74) and higher output pricing ($3.96 vs $3.48), with reduced cached token pricing ($0.13 vs $0.20).

Evidence
+DeepSeek V4 Pro 0813 $1.32 $0.13 (cached) $3.96

www.together.ai

featurehigh

Added Qwen3.8-2.4T-A95B model with pricing $2.50/$6.25 and cached token option

New model option available to users with standard and cached pricing tiers

Evidence
+Qwen3.8-2.4T-A95B $2.50 $0.50 (cached) $6.25

www.together.ai

other

New benchmark comparing DeepSeek-V4 Flash vs GPT-5.6 Luna on coding performance and cost efficiency

Developers can now compare cost-effectiveness of different LLMs for coding tasks; highlights DeepSeek's superior price performance for coding workloads

Evidence
+DeepSeek-V4 Flash 0731 vs GPT-5.6 Luna on DeepSWE: Cost and Coding We ran 900 DeepSWE rollouts on DeepSeek-V4 Flash and GPT-5.6 Luna. Luna leads pass@1 by 14 points; DeepSeek delivers 4.8x the solves per dollar.

www.together.ai

pricing

ByteDance Seedance 2.5 model pricing added at $0.115

New model option available for users, potential cost consideration for workloads using this model

Evidence
+ByteDance Seedance 2.5 $0.115

www.together.ai

feature

Added DeepSeek V4 Flash 0731 model with pricing and cache support

Users can now access DeepSeek V4 Flash 0731 model through Together AI's platform with input cost of $0.14/MTok, cached input cost of $0.03/MTok, and output cost of $0.28/MTok

Evidence
+DeepSeek V4 Flash 0731 $0.14 $0.03 (cached) $0.28

www.together.ai

feature

Kimi K3 model availability on Together AI platform

New model option available to users of Together AI's API for inference tasks.

Evidence
+Kimi K3 16,667 166,667 3,333 $0.05

www.together.ai · also seen at 1 more source

pricinghigh

Kimi K3 model pricing introduced with input/output rates and cache option

Developers can now use Kimi K3 model on Together AI platform with $3.00 input cost and $15.00 output cost, plus a $0.30 cached option for inputs.

Evidence
+Kimi K3 $3.00 $0.30 (cached) $15.00

www.together.ai

other

New blog post about Kimi K3 vs Claude Fable 5 performance and cost comparison

Users can now access comparative analysis of Kimi K3 and Claude Fable 5 models on DeepSWE benchmark, helping inform model selection decisions based on pass rates and cost-efficiency metrics.

Evidence
+Kimi K3 vs Claude Fable 5 on DeepSWE: Cost and Coding We ran 452 DeepSWE rollouts on Kimi K3 and Claude Fable 5. Fable leads pass@1 by 1.4 points; Kimi K3 wins pass@4 and delivers 2.8x the solves per dollar.

www.together.ai