2154 summaries
Paper Graph
A reading graph of paper summaries across LLMs, RL training, inference systems and architectures — built for mobile commutes.2154summaries
73domains
16years
MOST RECENT
Adaptive Teacher Exposure for Self-Distillation in LLM Reasoning
How to use this
The Timeline is the fastest way to scan what landed recently. The Roadmap shows lineage between papers; tap any dot to open its summary. Everything is mobile-first — swipe left/right on a paper page to walk in-domain.
Browse by domain
Each domain is a sub-collection of papers, sorted newest first. Tap a card to open the timeline for that domain.
Your progress
- Training Methods0/306
- Architecture0/210
- Multimodal0/159
- Vision0/156
- Agents0/130
- Inference Optimization0/125
- Evaluation0/123
- RL Training0/118
- Pretraining0/117
- Reasoning0/99
- Alignment0/86
- LLM Systems0/69
- Data0/63
- Uncategorized0/58
- Serving0/47
- Diffusion0/46
- Safety0/42
- Scaling Laws0/36
- Code0/35
- Mixture of Experts0/32
- Distributed Training0/29
- Retrieval0/24
- Context Optimization0/16
- Prompting0/15
- Low Precision0/13
- cs lg0/0
- cs cv0/0
- cs cl0/0
- stat ml0/0
- cs ai0/0
- cs ir0/0
- cs ds0/0
- eess as0/0
- cs ro0/0
- cs hc0/0
- cs gt0/0
- cs si0/0
- cs cr0/0
- eess iv0/0
- cs sd0/0
- cs se0/0
- cs dc0/0
- cs db0/0
- cs cy0/0
- cs it0/0
- cs gr0/0
- cs pl0/0
- stat me0/0
- cs ne0/0
- cs lo0/0
- cs cc0/0
- cs ni0/0
- cs ma0/0
- stat ap0/0
- cs mm0/0
- cs ar0/0
- eess sp0/0
- eess sy0/0
- cs cg0/0
- stat co0/0
- cs dl0/0
- cs dm0/0
- cs ce0/0
- cs ms0/0
- cs pf0/0
- cs et0/0
- cs fl0/0
- cs oh0/0
- cs gl0/0
- stat ot0/0
- cs os0/0
- math st0/0
- cs sc0/0
Years covered
Topic threads
Cross-domain threads of papers that share a topic — RLHF, MoE, KV caches, agents, and more.
Tool use & function callingChain-of-thought reasoningRLHF / preference learningGRPO / RLVR group methodsSpeculative decodingKV cache / paged attentionMixture of ExpertsState-space / Mamba / linear attentionLong contextFP8 / low-precisionFlash attentionScaling lawsCoding agentsDistillationTransformer foundationsPretraining recipesRetrieval-augmented generationVision-language modelsReasoning via RLServing systemsSelf-improvement & synthetic dataWorld modelsMultimodal generation / diffusion