When to Switch: Reliable Action-Chunk Extension for Vision-Language-Action Models Paper • 2610.05719 • Published 1 day ago • 6
LEGO-Anything: Coding Agents for 3D Scene Reconstruction Paper • 2609.36380 • Published 8 days ago • 138
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-Coder-GGUF Image-Text-to-Text • 117B • Updated 7 days ago • 389k • 296
FLUX 3 Action Collection Open weights 7B world action model: base and shared encoders, SO-101 and DROID policies. Code: https://github.com/black-forest-labs/flux-action • 3 items • Updated 4 days ago • 41
nvidia/Nemotron-3-Diarization Voice Activity Detection • 99.2M • Updated 12 days ago • 55.5k • 701
view article Article **Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization** nvidia • 13 days ago • 67
primitive-ai/DeepSeek-V4-Flash-Vision-Exp-REAP-145B Text Generation • 146B • Updated 30 days ago • 572 • 8
Running 5 Navier Stokes Vortex WebGPU 🌀 5 Finite-time blowup in smoothly forced Navier–Stokes flow
NeoMME Collection Meet NeoMME: a family of 260M and 800M Multimodal-Native Multilingual Encoders • 12 items • Updated 8 days ago • 33
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • Sep 3 • 147
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF Image-Text-to-Text • 27B • Updated Sep 2 • 1.61M • 1.98k
Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published Aug 31 • 69