Blog

告别冗长 PDF,一文读懂最新顶会与核心期刊的创新价值。

SPD Learn: A Geometric Deep Learning Python Library for Neural Decoding Through Trivialization
Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators
Learning to Drive is a Free Gift: Large-Scale Label-Free Autonomy Pretraining from Unposed In-The-Wild Videos
DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipulation
CRAG: Can 3D Generative Models Help 3D Assembly?
Training Agents to Self-Report Misbehavior
MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
Detection and Recognition: A Pairwise Interaction Framework for Mobile Service Robots
TESS Planet Occurrence Rates Reveal the Disappearance of the Radius Valley Around Mid-to-Late M Dwarfs
Iterative Closed-Loop Motion Synthesis for Scaling the Capabilities of Humanoid Control
A Bayesian approach to out-of-sample network reconstruction
Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG Training
How Do Latent Reasoning Methods Perform Under Weak and Strong Supervision?
ArchAgent: Agentic AI-driven Computer Architecture Discovery
DyaDiT: A Multi-Modal Diffusion Transformer for Socially Favorable Dyadic Gesture Generation
A data- and compute-efficient chest X-ray foundation model beyond aggressive scaling
SODA-CitrON: Static Object Data Association by Clustering Multi-Modal Sensor Detections Online
VRSL:Exploring the Comprehensibility of 360-Degree Camera Feeds for Sign Language Communication in Virtual Reality
Relativistic Tidal Dissipation and the Gravitational-wave Signal of a White Dwarf Orbiting an Intermediate-Mass Black Hole
PackUV: Packed Gaussian UV Maps for 4D Volumetric Video
Vibe Researching as Wolf Coming: Can AI Agents with Skills Replace or Augment Social Scientists?
MSJoE: Jointly Evolving MLLM and Sampler for Efficient Long-Form Video Understanding
Entropy-Controlled Flow Matching
Duel-Evolve: Reward-Free Test-Time Scaling via LLM Self-Preferences
Pixel2Catch: Multi-Agent Sim-to-Real Transfer for Agile Manipulation with a Single RGB Camera
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
Stable Adaptive Thinking via Advantage Shaping and Length-Aware Gradient Regulation
A Scaling Law for Bandwidth Under Quantization
Spherically Symmetric Gravity on a Graph I: Theoretical Foundations