PAPER·May 14, 2026SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language ModelsarXiv
PAPER·May 13, 2026Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action ModelsarXiv
PAPER·May 13, 2026OpenAaaS: An Open Agent-as-a-Service Framework for Distributed Materials-Informatics ResearcharXiv
PAPER·May 13, 2026Phy-CoSF: Physics-Guided Continuous Spectral Fields Reconstruction and Super-Resolution for Snapshot Compressive ImagingarXiv
PAPER·May 13, 2026RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior ImitationarXiv
PAPER·May 13, 2026ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive MarginarXiv
PAPER·May 13, 2026Sleeper Channels and Provenance Gates: Persistent Prompt Injection in Always-on Autonomous AI AgentsarXiv