We're publishing our first end-to-end benchmarks for Zyphra Inference on
@AMD Instinct MI355X.
Our inference optimizations strongly outperform the AMD baseline and narrows the gap between MI355X and B200 for serving Kimi K2.6, GLM 5.1, and DeepSeek V3.2 🧵