PAPER·July 27, 2026GOTS: Greedy Orthogonal Token Selection for High-Resolution Vision-Language ModelsarXiv
PAPER·July 27, 2026SimBEV2X: A Large-Scale Dataset and Data Generation Tool for Multi-Task Vehicle-to-Everything Cooperative PerceptionarXiv
PAPER·July 25, 2026ViTacWorld: Scaling Visuo-Tactile World Models for Contact-Rich Robot ManipulationarXiv
PAPER·July 25, 2026Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving SkillsarXiv
PAPER·July 25, 2026CausalForge: A Formally Grounded, Self-Improving Agentic Framework for Automated Research in Causal InferencearXiv
PAPER·July 25, 2026IR275K: A Benchmark for Infrared Multi-Frame Super-Resolution Toward Efficient Remote SensingarXiv
PAPER·July 24, 2026Neptuna: A Comprehensive Machine Learning Framework for Benchmarking Complex Multiphase FlowsarXiv
PAPER·July 24, 2026AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation AlignmentarXiv