We just signed one of the largest AI infrastructure deals for open source, period.
250MW data center. $5B+ in annualized revenue. Built with @HUMAIN in Saudi Arabia.
for september, we’re cutting the price of Dedicated Inference on H100s from $5.49/hr to $3.99/hr
new + existing deployments get the lower price automatically
deploy gemma 4, qwen3/3.5, gpt-oss, llama, nemotron 3.5 lightning models, or bring your own lora for a fine-tuned model
Excited to welcome @greptile, the AI code reviewer trusted by 11,000+ teams, as a Together AI customer.
Greptile's agents review and test pull requests with full context of your codebase, and running that at scale across thousands of teams takes serious production inference
glm-5.3 now beats gpt-5.6 sol and claude fable 5 on agentic benchmarks
5.3 flash is right behind it
glm-5.3 didn’t even need a new base model to get there
@Zai_org kept the glm-5.2 base and scaled post-training with more long-horizon environments, more diverse tasks, and more