WisPaper
WisPaper
学术搜索
功能
资源
价格
下载
工作空间
Blog
告别冗长 PDF,一文读懂最新顶会与核心期刊的创新价值。
用户分享
趋势
Eigenforms and graphs of Hecke operators with wild ramification
InCoder-32B: Code Foundation Model for Industrial Scenarios
The PokeAgent Challenge: Competitive and Long-Context Learning at Scale
Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
SWE-QA-Pro: A Representative Benchmark and Scalable Training Recipe for Repository-Level Code Understanding
Manifold-Matching Autoencoders
LoST: Level of Semantics Tokenization for 3D Shapes
Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding
Versatile Editing of Video Content, Actions, and Dynamics without Training
AgentFactory: A Self-Evolving Framework Through Executable Subagent Accumulation and Reuse
AR-CoPO: Align Autoregressive Video Generation with Contrastive Policy Optimization
Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control
V-Dreamer: Automating Robotic Simulation and Trajectory Synthesis via Video Generation Priors
Symphony: A Cognitively-Inspired Multi-Agent System for Long-Video Understanding
A brief introduction to Poisson geometry
Procedural Generation of Algorithm Discovery Tasks in Machine Learning
Memento-Skills: Let Agents Design Agents
CUBE: A Standard for Unifying Agent Benchmarks
Kestrel: Grounding Self-Refinement for LVLM Hallucination Mitigation
Reasoning over mathematical objects: on-policy reward modeling and test time aggregation
Promotion and rowmotion in rational Catalan combinatorics
VideoAtlas: Navigating Long-Form Video in Logarithmic Compute
DreamPlan: Efficient Reinforcement Fine-Tuning of Vision-Language Planners via Video World Models
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic Imagery
GUI-CEval: A Hierarchical and Comprehensive Chinese Benchmark for Mobile GUI Agents
Efficient Reasoning on the Edge
VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents
ProbeFlow: Training-Free Adaptive Flow Matching for Vision-Language-Action Models
S-VAM: Shortcut Video-Action Model by Self-Distilling Geometric and Semantic Foresight
←
1
...
26
27
28
...
69
→