npm install -g @ moonmath-ai/zro
zro login
zro launch prime --model kimi-k3
zro launch prime --model deepseek-v4-flash-0731
Inference HW Acceleration;
building @zroai_
- Today we added @AMD MI325X as a new serving backend for @zroai_. Kimi K3 is the first model running on it, powered by @sgl_project. We believe this is the first production support for the official Kimi K3 weights on CDNA3. With 256 GB of HBM per GPU, the full model fits on a
- Those are… unusually low prices 🧐 All three citations point to the same URL, and the quoted rates look more like heavily discounted or long-term reserved pricing than on-demand. Normalize using actual on-demand market prices, and the picture changes completely.
- We are here to serve
- The updated dsv4-Flash-0731 feels unfairly good :) We added day-@zroai_ support serving it from our EU infra. And, just like with Kimi K3, we cut the price 🤠🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇 🔷 The official V4-Flash now natively supports the





