WisPaper
WisPaper
Search
Features
Resources
Pricing
Download
Workspace
Blog
No more endless PDFs. Discover the core value of the latest top-tier research in one article.
User Shared
Trends
Autodata: An agentic data scientist to create high quality synthetic data
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors
Nemotron-TwoTower: Diffusion Language Modeling with Pretrained Autoregressive Context
SPIRAL: Learning to Search and Aggregate
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
Self-Improving Agents in the Era of Experience: A Survey of Self-to Meta-Evolution
Ask, Don't Judge: Binary Questions for Interpretable LLM Evaluation and Self-Improvement
The Red Queen Gödel Machine: Co-Evolving Agents and Their Evaluators
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs
Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding
Scalable GANs with Transformers
Are We Ready For An Agent-Native Memory System?
SWE-Together: Evaluating Coding Agents in Interactive User Sessions
Building Multi-Task Agentic LLMs via Two-Phase Distillation
MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training
DOPD: Dual On-policy Distillation
TacForeSight: Force-Guided Tactile World Model for Contact-Rich Manipulation
Representation Learning Enables Scalable Multitask Deep Reinforcement Learning
Still: Amortized KV Cache Compaction in a Single Forward Pass
On Training in Imagination
Orca: The World is in Your Mind
Random Reshuffling Dominates Stochastic Gradient Descent
AdaJEPA: An Adaptive Latent World Model
Manifold Bandits: Bayesian Curriculum Learning over the Latent Geometry of Large Language Models
Learning through Internalization
IS ONE LAYER ENOUGH? TRAINING A SINGLE TRANSFORMER LAYER CAN MATCH FULL-PARAMETER RL TRAINING PREPRINT
A Mathematical Introduction to Diffusion Models
A Hippocampus for Linear Attention: An Exact Memory for What the Recurrent State Forgets
AutoMem: Automated Learning of Memory as a Cognitive Skill
How much do language models memorize?
←
1
...
609
610
611
...
614
→