MiMo-V2.6 Compressions Collection Collection of MiMo-V2.6 Compressions (GGUF / MLX / FP8) • 3 items • Updated about 16 hours ago • 1
view article Article tokenizers v1: encode, decode and scaling, measured +2 ArthurZ, sbrandeis, mcpotato, lysandre • 1 day ago • 61
view article Article Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem MultiverseComputingCAI • 1 day ago • 25
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 5 days ago • 64
EvoOntology: A Self-Evolving Ontology Layer for Data Agents Paper • 2609.15779 • Published 9 days ago • 128
IntBMoE: Integrating Block-Level Conditioning into Expert Composition for Full-Participation Mixture-of-Experts Paper • 2609.21346 • Published 5 days ago • 92
Qwen-Image-2.1-PE Compressions Collection Collection of Qwen-Image-2.1-PE (Prompt Enhancer) Compressions. • 6 items • Updated about 16 hours ago • 1
Merge Soup — Qwen3.5 Collection Collection of Experimental Model Merges • 8 items • Updated about 3 hours ago • 2
VisionGuardrail Evo2 Collection Collection of VisionGuardrail Multimodal Models! • 6 items • Updated about 16 hours ago • 1
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 13 days ago • 274
SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem Paper • 2609.07064 • Published 16 days ago • 146
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published 14 days ago • 329
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 15 days ago • 172
DriveZero: End-to-End Driving Beyond Human Demonstrations Paper • 2609.06055 • Published 18 days ago • 57
Scal3R: Learning Efficient Multi-Relative Pose Query for Scalable Online 3D Reconstruction Paper • 2609.04201 • Published 20 days ago • 50
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 20 days ago • 100
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 20 days ago • 186