UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations Paper • 2608.15930 • Published 11 days ago • 47
Prefilling-dLLM: Predictive Prefilling for Long-Context Inference in Diffusion Language Models Paper • 2606.10537 • Published Jun 9
MMDeepResearch-Bench: A Benchmark for Multimodal Deep Research Agents Paper • 2601.12346 • Published Jan 18 • 52
From Verifiable Dot to Reward Chain: Harnessing Verifiable Reference-based Rewards for Reinforcement Learning of Open-ended Generation Paper • 2601.18533 • Published Jan 26
GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning Paper • 2604.20659 • Published Apr 22 • 1
What Makes Interaction Trajectories Effective for Training Terminal Agents? Paper • 2606.03461 • Published Jun 2 • 3
TeamHOI: Learning a Unified Policy for Cooperative Human-Object Interactions with Any Team Size Paper • 2603.07988 • Published Mar 9 • 2
In-Context Reinforcement Learning for Tool Use in Large Language Models Paper • 2603.08068 • Published Mar 9 • 43
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors Paper • 2602.22144 • Published Feb 25 • 1
When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing Models Paper • 2602.10179 • Published Feb 10 • 6
Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs Paper • 2502.12982 • Published Feb 18, 2025 • 19