AdvSim2Real : Training Web Agents Against Adaptive Prompt Injection in a Web World Model Paper • 2610.08773 • Published 4 days ago • 15
What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 12 days ago • 138
Chinese-Jev: Bringing System One Model to Chinese-Language Tasks Paper • 2609.36965 • Published 12 days ago • 26
Post-Training Leaves Behavioral Shadows on Unrelated Decisions Paper • 2609.29233 • Published 17 days ago • 274
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures Paper • 2609.29429 • Published 17 days ago • 29
Neural Spectral Capacity: Measuring and Designing Architectures from Network Specification Alone Paper • 2609.23087 • Published 22 days ago • 11
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 19 days ago • 164
Transferring the Intelligence of VLMs to Robotic Control Paper • 2609.22966 • Published 22 days ago • 94
EvoOntology: A Self-Evolving Ontology Layer for Data Agents Paper • 2609.15779 • Published 27 days ago • 145