GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Paper • 2604.26752 • Published Apr 29 • 115
Inkling Collection Inkling is a versatile, customizable model that reasons over text, images, audio, with variable and efficient thinking effort. • 4 items • Updated Jul 27 • 59
Kairos Collection Experimental research on building a small multimodal model from scratch • 4 items • Updated 8 days ago • 4
GigaAM Multilingual: Foundation Model for Underrepresented Languages Paper • 2607.10371 • Published Jul 11 • 34
Krea 2 LoRAs Collection A collection of LoRAs for Krea 2 Turbo and Krea 2 Raw • 9 items • Updated Jun 23 • 54
InteractiveOmni: A Unified Omni-modal Model for Audio-Visual Multi-turn Dialogue Paper • 2510.13747 • Published Oct 15, 2025 • 33
view article Article Unlocking Agentic RL Training for GPT-OSS: A Practical Retrospective LinkedIn • Jan 27 • 81
Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator Paper • 2604.08121 • Published Apr 9 • 44
Harvey Collection A legal reasoning model specialized in Salvadoran jurisprudence • 4 items • Updated Apr 12 • 1
Think in Strokes, Not Pixels: Process-Driven Image Generation via Interleaved Reasoning Paper • 2604.04746 • Published Apr 8 • 74