view article Article TutorMoments: Do AI tutors know when to help and when to hold back? allenai • 24 days ago • 31
Learning to Attack and Defend: Adaptive Red Teaming of Language Models via GRPO Paper • 2606.09701 • Published Jun 8 • 1
FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds Paper • 2608.01049 • Published 29 days ago • 13
GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization Paper • 2608.01492 • Published 29 days ago • 14
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Paper • 2608.01851 • Published 28 days ago • 12
MASS: Multiplayer World Models with Authoritative Shared State Paper • 2608.06257 • Published 21 days ago • 17
Invisible Shortcuts: Why Vision Encoders Know Your Camera Paper • 2608.05424 • Published 26 days ago • 17
KVAE: Family of Tokenizers for Multimodal Generative Models Paper • 2608.05798 • Published 25 days ago • 29
MameLoshnLM: Yiddish Language Model and Evaluation Benchmark Paper • 2608.05850 • Published 25 days ago • 23
SmartMage: Dynamic Modality Orchestration for 3D Scene Understanding Paper • 2608.05137 • Published 21 days ago • 27
Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains Paper • 2608.05138 • Published 26 days ago • 32
ChronoVision: Temporal Reasoning via Latent State Reconstruction Paper • 2608.05631 • Published 25 days ago • 40
EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning Paper • 2608.06197 • Published 25 days ago • 47
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published 25 days ago • 46
Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval Paper • 2608.01481 • Published 29 days ago • 71
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models Paper • 2607.28609 • Published Jul 30 • 73
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Paper • 2608.05987 • Published 25 days ago • 100
Seeing the Needle in the Haystack: Towards Weakly-Supervised Log Instance Anomaly Localization via Counterfactual Perturbation Paper • 2605.10988 • Published May 9 • 6