
🎉 We are thrilled to announce NextStep-1 has been selected as an ICLR 2026 Oral (Acceptance Rate: 1.18%)! 🏆 We are redefining Autoregressive Image Generation by effectively moving beyond Vector Quantization. NextStep-1 pairs a causal transformer with a lightweight Flow Matching head to predict continuous tokens directly. It is the first continuous autoregressive model to achieve SOTA performance comparable to diffusion models like Flux and SD3.5. 🔥 Code and blog for Nextstep-1.1 are coming soon! We are committed to open research to accelerate the future of AR generation. Paper: https://t.co/CcEhZWIcSB GitHub: https://t.co/DSEWk0olDs HuggingFace: https://t.co/eRRk5GvNZN
NextStep-1 predicts continuous image with a flow matching head on a causal , evidence that autoregressive generation can match diffusion models like Flux and SD3.5 without vector .
article👏🏻Congratulations!Step3-VL-10B was selected for HuggingFace Daily Papers…
postStep-Audio-R1.1 opens weights for a speech model that reasons in real timeChecking sign-in…
Loading comments…