Training, learning and inference: unified dynamics of neural systems Paper • 2608.20965 • Published 10 days ago • 2
J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data Paper • 2608.26582 • Published 4 days ago • 18
Code as Worlds: Agentic Discovery of Executable World Representations for Physical Reasoning Paper • 2608.27549 • Published 4 days ago • 20
Rubric-to-Code Credit Assignment for Reinforcement Learning Paper • 2608.27906 • Published 3 days ago • 3
Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge Paper • 2608.28478 • Published 3 days ago • 11
LMSM: LLM Security Framework Inspired by Linux Security Modules Paper • 2608.25697 • Published 5 days ago • 3
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL Paper • 2608.28476 • Published 3 days ago • 16
Rethinking Expressivity and Efficiency in Test-Time Training Paper • 2608.21308 • Published 10 days ago • 1
AgentWeave: Routing Before Reasoning for Efficient Function Calling in Tool-Rich Language Models Paper • 2608.23078 • Published 7 days ago • 1
Giga-Embeddings: Mixture-of-Experts Encoders for High-Throughput Text Embeddings Paper • 2608.23806 • Published 7 days ago • 2
Gated Activation Steering for Reducing Sycophancy & Hallucination in Medical Question Answering Paper • 2608.23666 • Published 7 days ago • 1
Tunable Tool-Call Rates in LLM Agents via Representation Steering Paper • 2608.25198 • Published 6 days ago • 1
Plans You Can Check: Verifier-Grounded Learning of an Open-Weight Planner for Executable Video-Editing Paper • 2608.25622 • Published 5 days ago • 1
RTPO: Reverse-Turn Policy Optimization for Stabilizing Agentic RL Training Paper • 2608.18682 • Published 12 days ago • 2
Semantic Overlays: Mitigating Prompt Injection with Annotations Beyond Tokens and Steering Vectors Paper • 2608.23873 • Published 7 days ago • 1
RecurSE: Bounded Recursive Self-Evaluation for LLM Rubric Judges Paper • 2608.24231 • Published 6 days ago • 1
TacForcing: Streaming Action Generation with Execution-Time Tactile Feedback Paper • 2608.25798 • Published 5 days ago • 5
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published 5 days ago • 19
From Seeing to Acting: Smart Glasses as First-Person Intelligence Platforms Paper • 2608.24877 • Published 6 days ago • 10
Super Star: Towards Streaming Real-time Interactive Agents for Digital Humans Paper • 2608.24909 • Published Jul 22 • 5