Loading video…
Hook

Their other posts in the index, biggest breakout first.
Nvidia just open source a computer vision model that’s 10 times faster than top models and it's kind of insane. Most vision models today predict bounding boxes step by step, corner by corner, token by token. But this model changes that. It uses something called parallel box decoding. So instead of predicting pieces, it predicts the entire box at once. That's why it's up to 10 times faster than models like Qwen 3 VL. And they trained it on massive data. Over 100 million queries and hundreds of millions of boxes. It's fully open source on Hugging Face and GitHub. If you want to learn about AI news, tutorials, follow.