Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published 14 days ago • 83
view article Article Welcome Inkling by Thinking Machines +2 burtenshaw, merve, pcuenq, ariG23498 • 9 days ago • 116
SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer Paper • 2605.30409 • Published May 28 • 42
Cosmos3 Collection Omnimodal World Models for Physical AI • 20 items • Updated about 7 hours ago • 163
DINOv3 Collection DINOv3: foundation models producing excellent dense features, outperforming SotA w/o fine-tuning - https://arxiv.org/abs/2508.10104 • 15 items • Updated Mar 10 • 705
GraphLocator: Graph-guided Causal Reasoning for Issue Localization Paper • 2512.22469 • Published Dec 27, 2025 • 4
SpatialTree: How Spatial Abilities Branch Out in MLLMs Paper • 2512.20617 • Published Dec 23, 2025 • 44
view article Article Continuous batching from first principles +1 ror, ArthurZ, mcpotato • Nov 25, 2025 • 423
view article Article Building the Open Agent Ecosystem Together: Introducing OpenEnv +8 spisakjo, darktex, zkwentz, mortimerp9, Sanyam, Hamid-Nazeri, Pankit01, emre0, lewtun, reach-vb • Oct 23, 2025 • 166
The Well Collection A 15TB collection of physics simulation datasets. • 18 items • Updated Mar 24, 2025 • 54
MM Grounding DINO Collection See: https://github.com/huggingface/transformers/pull/37925 • 8 items • Updated Jun 26, 2025 • 5
LLMDet Collection See: https://github.com/huggingface/transformers/pull/37925 • 3 items • Updated Jun 26, 2025 • 3
SmolDocling datasets Collection Datasets used to train SmolDocling • 6 items • Updated Jul 31, 2025 • 31
D-FINE Collection State-of-the-art real-time object detection model with Apache 2.0 licence • 15 items • Updated May 5, 2025 • 56