Beyond Language Modeling: An Exploration of Multimodal Pretraining Paper ⢠2603.03276 ⢠Published Mar 3 ⢠109
Cosmos-Embed1 Collection Joint video-text embedding for physical AI ⢠5 items ⢠Updated Aug 11 ⢠9
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper ⢠2609.19969 ⢠Published 17 days ago ⢠201
deepseek-ai/DeepSeek-V4.1-Flash Image-Text-to-Text ⢠763B ⢠Updated 3 days ago ⢠788k ⢠⢠4.06k
On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Paper ⢠2608.30320 ⢠Published Aug 31 ⢠63
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper ⢠2609.00111 ⢠Published Aug 31 ⢠316
Cosmos-Reason2 Collection ā ļø This collection is archived. š https://huggingface.co/collections/nvidia/cosmos3 ⢠8 items ⢠Updated Aug 11 ⢠28
view article Article Welcome NVIDIA Cosmos 3: The First Open Omni-model for Physical AI Reasoning and Action nvidia ⢠Jun 1 ⢠90
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp Image-Text-to-Text ⢠305B ⢠Updated Sep 1 ⢠915k ⢠⢠930
Running Featured 84 QED-Nano: Teaching a Tiny Model to Prove Hard Theorems š 84 Who needs 1T parameters? Olympiad proofs with a 4B model