59 papers
Prompting
0/59Prompt engineering, CoT, in-context learning.
Progress0 of 59
2026
9- MaySkillOpt: Executive Strategy for Self-Evolving Agent Skillsagents2605.23904Microsoft ResearchMay 22, 2026~125 min
- MarPOLCA: Stochastic Generative Optimization with LLMcode2603.14769DeepmindMar 16, 2026score 9~113 min
- MarIn-Context Reinforcement Learning for Tool Use in Large Language Modelstraining-methods2603.08068National University of SingaporeMar 9, 2026~114 min
- MarUnderstanding the Challenges in Iterative Generative Optimization with LLMstraining-methods2603.23994DeepmindMar 25, 2026score 9~94 min
- FebP-GenRM: Personalized Generative Reward Model with Test-time User-based Scalingalignment2602.12116Tongyi-ConvAIFeb 12, 2026score 9~115 min
- FebCC-VQA: Conflict- and Correlation-Aware Method for Mitigating Knowledge Conflict in Knowledge-Based Visual Question Answeringmultimodal2602.23952Alibaba CloudFeb 27, 2026score 2~95 min
- JanExpSeek: Self-Triggered Experience Seeking for Web Agentsagents2601.08605TongyiLabJan 13, 2026score 5~99 min
- JanThe Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought Reasoningreasoning2601.06002ByteDanceJan 9, 2026score 8~152 min
- JanLess Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purificationrl-training2601.21244BAIDUJan 29, 2026score 8~98 min
2025
8- DecAligned but Stereotypical? The Hidden Influence of System Prompts on Social Bias in LVLM-Based Text-to-Image Modelsalignment2512.04981KAIST AIDec 4, 2025score 4~119 min
- OctBeyond Correctness: Evaluating Subjective Writing Preferences Across Culturesevaluation2510.14616ByteDance SeedOct 16, 2025score 8~104 min
- OctAgentic Context Engineering: Evolving Contexts for Self-Improving Language Modelsllm-systems2510.04618Oct 6, 2025score 9~107 min
- OctMultimodal Prompt Optimization: Why Not Leverage Multiple Modalities for MLLMsmultimodal2510.09201KAIST AIOct 10, 2025score 3~134 min
- OctThe End of Manual Decoding: Towards Truly End-to-End Language Modelsuncategorized2510.26697Oct 30, 2025~126 min
- JulA Survey of Context Engineering for Large Language Modelsprompting2507.13334Jul 17, 2025~88 min
- JunTransformers Meet In-Context Learning: A Universal Approximation Theoryinference-optimization2506.05200Jun 5, 2025~121 min
- FebChain of Draft: Thinking Faster by Writing Lessprompting2502.18600Zoom AIFeb 25, 2025score 8~88 min
2024
9- DecLearnLM: Improving Gemini for Learningtraining-methods2412.16429Dec 21, 2024score 7~111 min
- OctControllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirementsalignment2410.08968Oct 11, 2024score 8~113 min
- JunScaling Synthetic Data Creation with 1,000,000,000 Personasdata2406.20094Jun 28, 2024~88 min
- JunThe Prompt Report: A Systematic Survey of Prompting Techniquesprompting2406.06608Jun 6, 2024score 10~119 min
- FebPIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMsagents2402.07872DeepMindFeb 12, 2024score 8~119 min
- FebSelf-Discover: Large Language Models Self-Compose Reasoning Structuresprompting2402.03620DeepMindFeb 6, 2024score 9~88 min
- FebPremise Order Matters in Reasoning with Large Language Modelsreasoning2402.08939DeepMindFeb 14, 2024~116 min
- FebChain-of-Thought Reasoning without Promptingreasoning2402.10200Feb 15, 2024~113 min
- JanMeta-Prompting: Enhancing Language Models with Task-Agnostic Scaffoldingprompting2401.12954SlackJan 23, 2024~95 min
2023
15- DecThe Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learningalignment2312.01552Dec 4, 2023score 9~116 min
- NovInstruction-Following Evaluation for Large Language Modelsevaluation2311.07911Google ResearchNov 14, 2023score 9~89 min
- NovSystem 2 Attention (is something you might need too)prompting2311.11829Nov 20, 2023score 9~121 min
- NovEverything of Thoughts: Defying the Law of Penrose Triangle for Thought Generationreasoning2311.04254Nov 7, 2023score 9~116 min
- NovThe ART of LLM Refinement: Ask, Refine, and Trustreasoning2311.07961Nov 14, 2023score 9~98 min
- OctLarge Language Models as Analogical Reasonersprompting2310.01714DeepMindOct 3, 2023score 9~83 min
- SepScaling Catalog Attribute Extraction with Multi-modal LLMsprompting2309.03409InstacartSep 7, 2023~116 min
- AugRetroformer: Retrospective Large Language Agents with Policy Gradient Optimizationrl-training2308.02151Aug 4, 2023score 9~111 min
- JulSkeleton-of-Thought: Large Language Models Can Do Parallel Decodingllm-systems2307.15337Jul 28, 2023score 9~96 min
- JulSelf-consistency for open-ended generationsreasoning2307.06857Jul 11, 2023score 8~118 min
- JunRead Moreinference-optimization2306.17806EleutherAIJun 30, 2023~115 min
- MayTree of Thoughts: Deliberate Problem Solving with Large Language Modelsreasoning2305.10601May 17, 2023score 9~122 min
- MarReflexion: Language Agents with Verbal Reinforcement Learningagents2303.11366Mar 20, 2023~108 min
- MarRequirement Adherence: Boosting Data Labeling Quality Using LLMsreasoning2303.17651UberMar 30, 2023~96 min
- JanThe Flan Collection: Designing Data and Methods for Effective Instruction Tuningtraining-methods2301.13688Allen Institute for AIJan 31, 2023~111 min
2022
9- DecHyDE: Precise Zero-Shot Dense Retrieval without Relevance Labelsretrieval2212.10496DoorDashDec 20, 2022~94 min
- NovProgram of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasksprompting2211.12588Nov 22, 2022~106 min
- OctReAct: Synergizing Reasoning and Acting in Language Modelsagents2210.03629Qwen / Alibaba CloudOct 6, 2022~104 min
- OctScaling Instruction-Finetuned Language Modelsalignment2210.11416SalesforceOct 20, 2022~105 min
- MayFew-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learningalignment2205.05638May 11, 2022~115 min
- MayLarge Language Models are Zero-Shot Reasonersreasoning2205.11916May 24, 2022~102 min
- MarSelf-Consistency Improves Chain of Thought Reasoning in Language Modelsreasoning2203.11171Mar 21, 2022~99 min
- FebRethinking the Role of Demonstrations: What Makes In-Context Learning Work?prompting2202.12837Feb 25, 2022~119 min
- JanMonte Carlo, Puppetry and Laughter: The Unexpected Joys of Prompt Engineeringreasoning2201.11903InstacartJan 28, 2022~119 min
2021
6- NovOn Transferability of Prompt Tuning for Natural Language Processingprompting2111.06719Nov 12, 2021~117 min
- SepFinetuned Language Models Are Zero-Shot Learnerstraining-methods2109.01652Sep 3, 2021~98 min
- AprPrompt Tuning: The Power of Scale for Parameter-Efficient Prompt Tuningprompting2104.0869194 citesApr 18, 2021~88 min
- FebCalibrate Before Use: Improving Few-Shot Performance of Language Modelsprompting2102.09690Feb 19, 2021~102 min
- JanWhat Makes Good In-Context Examples for GPT-3?prompting2101.06804Jan 17, 2021~101 min
- JanPrefix-Tuning: Optimizing Continuous Prompts for Generationtraining-methods2101.00190Jan 1, 2021~85 min
2020
3- DecMaking Pre-trained Language Models Better Few-shot Learnerstraining-methods2012.15723Dec 31, 2020~117 min
- OctAutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Promptsprompting2010.15980Oct 29, 2020~100 min
- MayLanguage Models are Few-Shot Learnerspretraining2005.14165Meta AI / FAIRMay 28, 2020~119 min