Blog

告别冗长 PDF,一文读懂最新顶会与核心期刊的创新价值。

Observation interventions for motor skill learning and performance: an applied model for the use of observation
Assessment of offshore power potential in Zhoushan archipelago using a 45-year wind field product
Evaluation of planetary boundary layer schemes in WRF model for simulating sea-land breeze in Shanghai, China
Boundary layer dynamics and power performance of an offshore wind farm during unstable ocean-wave-atmosphere interactions
RankFlow: A Multi-Role Collaborative Reranking Workflow Utilizing Large Language Models
Deliberative Searcher: Improving LLM Reliability via Reinforcement Learning with constraints
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
MATHNET: A GLOBAL MULTIMODAL BENCHMARK FOR MATHEMATICAL REASONING AND RETRIEVAL
Through the Lens of Core Competency: Survey on Evaluation of Large Language Models
VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck
Energy Conversion and Management
SwitchMT: An Adaptive Context Switching Methodology for Scalable Multi-Task Learning in Intelligent Autonomous Agents
Large Language Models Do NOT Really Know What They Don't Know
Faithfulrag: Fact-level conflict modeling for context-faithful retrieval-augmented generation
To Be or Not To Be: The Impact of OCR Noise and Annotation Errors on Semi-Structured Extraction from Clinical Reports
An Information-Geometric Approach to Artificial Curiosity
AI-Enforced Ultra-Large Virtual Screening Discovers Potent CD28 Binders
Polarizations of Artin monomial ideals
Towards Ultra-High-Rate Quantum Error Correction with Reconfigurable Atom Arrays
Empowering Multi-Turn Tool-Integrated Reasoning with Group Turn Policy Optimization
Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning
Large Language Models Do NOT Really Know What They Don't Know
Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning
Is Agentic RAG worth it? An experimental comparison of RAG approaches
Scaling Laws for Reranking in Information Retrieval
Peerispect: Claim Verification in Scientific Peer Reviews
Align Documents to Questions: Question-Oriented Document Rewriting for Retrieval-Augmented Generation
Rethinking uncertainty estimation in natural language generation
Domain-oriented RAG Assessment (DoRA): Synthetic Benchmarking for RAG-based Question Answering on Defense Documents