Agentic Transaction: Towards ACID-Compliant Agent Systems Paper • 2608.13900 • Published 6 days ago • 25
How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks Paper • 2608.14905 • Published 6 days ago • 26
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 3 days ago • 141
view article Article Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers +1 tomaarsen, NohTow, raphaelsty • 1 day ago • 63
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • 10 days ago • 103
LettuceDetect v2 Collection SOTA hallucination detection for agentic workflows, multilingual, long context • 6 items • Updated 14 days ago • 3
Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation Paper • 2607.27372 • Published 22 days ago • 19
jina-reranker-v3.5: An Efficient Listwise Reranker with Hybrid Attention and Self-Distillation Paper • 2607.18152 • Published about 1 month ago • 4
jina-reranker-v3: Last but Not Late Interaction for Document Reranking Paper • 2509.25085 • Published Sep 29, 2025 • 12
Flux-OPD: On-Policy Distillation with Evolving Contexts Paper • 2607.28022 • Published 21 days ago • 44
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 21 days ago • 55
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering Paper • 2607.28568 • Published 21 days ago • 184
BM25 Wins at Scale: A Scaling Study of Retrieval-Augmented Generation Paradigms Paper • 2607.26497 • Published 21 days ago • 51
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 21 days ago • 58
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 21 days ago • 306
MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis Paper • 2607.27146 • Published 22 days ago • 29
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Paper • 2607.26784 • Published 22 days ago • 28