GEOID-Flood: A Large-Scale Multi-Modal Benchmark Dataset for Flood Segmentation Paper • 2608.02315 • Published 3 days ago • 1 • 2
InfiniSplat: Implicit Gaussian Decoding for Large-Baseline Monocular View Synthesis Paper • 2608.02437 • Published 3 days ago • 54 • 2
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Paper • 2607.29613 • Published 6 days ago • 23 • 7
VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation Paper • 2607.28590 • Published 7 days ago • 43 • 2
AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks? Paper • 2608.00155 • Published 6 days ago • 12 • 2
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published 3 days ago • 136 • 2
LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation Paper • 2608.00079 • Published 8 days ago • 14 • 2
GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning Paper • 2608.02585 • Published 3 days ago • 22 • 3
To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing Paper • 2607.28887 • Published 7 days ago • 19 • 2
MemSFT: Mitigating Alignment Tax with an External Parametric Memory Paper • 2607.25614 • Published 9 days ago • 20 • 2
ICDAR 2026 Competition on Information Extraction from Atomic Layer Deposition/Etching (ALD/E) Scientific Figures Paper • 2607.26848 • Published 8 days ago • 4 • 2
DreamTraj: Generating 6-DoF Object Trajectories by Reading Unrendered Video Diffusion Latents Paper • 2608.00486 • Published 5 days ago • 12 • 2
Zero-Mem: Zero-Token Memory Operations for LLM Agents Paper • 2607.29377 • Published 6 days ago • 7 • 2
Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models Paper • 2607.26326 • Published 9 days ago • 2 • 3
ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step Paper • 2608.02358 • Published 3 days ago • 10 • 2
SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space Paper • 2608.01397 • Published 4 days ago • 4 • 2
Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis Paper • 2608.00440 • Published 5 days ago • 10 • 4
Roomer: Reflective Object-Grounded Model Editing and Repair for 3D Indoor Layout Synthesis Paper • 2608.01973 • Published 3 days ago • 16 • 2