414 papers
cs ai
0/02026
53- JulCoachable agents for interactive gameplayno summary yetcs-ai2607.00642Sony0 citesJul 1, 2026
- JulAutoMem: Automated Learning of Memory as a Cognitive Skillno summary yetcs-ai2607.01224Stanford0 citesJul 1, 2026
- JulAuto-FL-Research: Agentic Search for Federated Learning Algorithmsno summary yetcs-ai2607.01366NVIDIA0 citesJul 1, 2026
- JulProcedural Memory Distillation: Online Reflection for Self-Improving Language Modelsno summary yetcs-ai2607.01480Salesforce0 citesJul 1, 2026
- JulRobust Feasible Route Construction through Collaborative Partition Optimizationno summary yetcs-ai2607.03694CMU0 citesJul 4, 2026
- JulProgress- and Reliability-Oriented Group Policy Optimization for Agentic Reinforcement Learningno summary yetcs-ai2607.04242Princeton0 citesJul 5, 2026
- JulSovereignPA-Bench: Evaluating User-Owned Personal Agents under Evolving Intent, Platform Mediation, and Consent Constraintsno summary yetcs-ai2607.05363Stanford0 citesJul 6, 2026
- JulReason Less, Verify More: Deterministic Gates Recover a Silent Policy-Violation Failure Mode in Tool-Using LLM Agentsno summary yetcs-ai2607.07405MIT0 citesJul 8, 2026
- JulInstitutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safetyno summary yetcs-ai2607.07695MIT0 citesJul 8, 2026
- JulA First-Principles Theory of Slow Thinking and Active Perceptionno summary yetcs-ai2607.08196Princeton0 citesJul 9, 2026
- JulNorm Enforcement for AI Agents: Robustly Shaping Behavior in Multi-Agent Systemsno summary yetcs-ai2607.09766Berkeley0 citesJul 7, 2026
- JulPHITSBench: an execution-scored benchmark for AI-assisted PHITS radiation-transport input generation using natural languageno summary yetcs-ai2607.09789MIT0 citesJul 8, 2026
- JulDynamic Agent Skills: A Lifecycle Survey and Taxonomy of Evolving Skill Librariesno summary yetcs-ai2607.10113CMU0 citesJul 11, 2026
- JulInteraction Scaling: Grounding the Third Axis of Test-Time Computeno summary yetcs-ai2607.11598UW0 citesJul 13, 2026
- JulThink Through a Bottleneck: Hourglass Reasoning for Rigorous Inductionno summary yetcs-ai2607.11696Princeton0 citesJul 13, 2026
- JulThe Emerging Paradigm of Geospatial Foundation Models: From Pre-Training to Agentic Reasoningno summary yetcs-ai2607.12177Google Research0 citesJul 13, 2026
- JunDeployment-Centered Evaluation: Predicting Query-Level Rejection Risk in a Clinical LLM Systemno summary yetcs-ai2606.12702Stanford0 citesJun 10, 2026
- JunMARS: Margin-Adversarial Risk-controlled Stopping for Parallel LLM Test-time Scalingno summary yetcs-ai2606.12935Stanford0 citesJun 11, 2026
- JunTrust Between AI Agents: Measuring Formation, Breakage, and Recovery, with Implications for Governing Multi-Agent Systemsno summary yetcs-ai2606.14923MIT0 citesJun 12, 2026
- JunMetric Match: A Subset Selection Approach to Evaluating LLM Judge Reliabilityno summary yetcs-ai2606.15029Stanford0 citesJun 12, 2026
- JunCognitive Debt: AI as Intellectual Leverage and the Dynamics of Systemic Fragilityno summary yetcs-ai2606.15078NYU0 citesJun 13, 2026
- JunReward Hacking in Language Model Agents: Revisiting AI Safety Gridworldsno summary yetcs-ai2606.15385Berkeley0 citesJun 13, 2026
- JunArchitectural Wisdom: A Framework for Governing Optimization in AI Systemsno summary yetcs-ai2606.16319Stanford0 citesJun 15, 2026
- JunWhat Must Generalist Agents Remember?no summary yetcs-ai2606.18746CMU0 citesJun 17, 2026
- JunExit-and-Join Dynamics for Decentralized Coalition Formationno summary yetcs-ai2606.19683NYU0 citesJun 18, 2026
- JunOptimal Scheduling in a Question-Answering Forum of Knowledge Workersno summary yetcs-ai2606.19759CMU0 citesJun 18, 2026
- JunHuman-on-the-Loop Orchestration for AI-Assisted Legal Discoveryno summary yetcs-ai2606.19812Google Research0 citesJun 18, 2026
- JunA-Evolve-Training: Autonomous Post-Training of a 30B Modelno summary yetcs-ai2606.20657Amazon0 citesJun 9, 2026
- JunA Stackelberg Framework for Resource-Aware LLM Agents: Learning, Repair, and Conditional Guaranteesno summary yetcs-ai2606.23026Tencent0 citesJun 22, 2026
- JunSPIRAL: Learning to Search and Aggregateno summary yetcs-ai2606.23595Stanford0 citesJun 22, 2026
- JunCan Aggregate Invariants Accelerate Continuous Subgraph Matching? Limits, Laws, and a Dynamic Spectral Indexno summary yetcs-ai2606.24421Tencent0 citesJun 23, 2026
- JunThe Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agentsno summary yetcs-ai2606.24470UW0 citesJun 23, 2026
- JunAgentic System as Compressor: Quantifying System Intelligence in Bitsno summary yetcs-ai2606.25960Princeton0 citesJun 24, 2026
- JunGenerative Retrieval via Diffusion Transformer with Metric-Ordered Sequence Training and Hybrid-Policy Preference Optimizationno summary yetcs-ai2606.26899Princeton0 citesJun 25, 2026
- JunOdyssey: Constructing Verifiable Local Truth-Preserving Foundation Modelsno summary yetcs-ai2606.27593Adobe0 citesJun 25, 2026
- JunData and Evaluation Closed-Loop for Model Capability Enhancementno summary yetcs-ai2606.28471Baidu0 citesJun 26, 2026
- JunIMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversationsno summary yetcs-ai2606.28556Amazon0 citesJun 26, 2026
- JunAgent-Computer Observation Interfaces Enable Dynamic Computer Useno summary yetcs-ai2606.29472UW0 citesJun 28, 2026
- JunWhose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agentsno summary yetcs-ai2606.30383UW0 citesJun 29, 2026
- JunWhen Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Modelsno summary yetcs-ai2606.30852Stanford0 citesJun 29, 2026
- JunBeyond expert users: agents should help users construct preferences, not just elicit themno summary yetcs-ai2606.30863Stanford0 citesJun 29, 2026
- JunRoPoLL: Robust Panel of LLM Judgesno summary yetcs-ai2606.30931Amazon0 citesJun 29, 2026
- JunAdaptive Cluster-First Route-Second Decomposition for Industrial-Scale Vehicle Routingno summary yetcs-ai2606.31820CMU0 citesJun 30, 2026
- JunMnemosyne: Agentic Transaction Processing for Validating and Repairing AI-generated Workflowsno summary yetcs-ai2607.00269Stanford0 citesJun 30, 2026
- MayUnderstanding Annotator Safety Policy with Interpretabilityno summary yetcs-ai2605.05329Apple0 citesMay 6, 2026
- MayGenerative Auto-Bidding with Unified Modeling and Explorationno summary yetcs-ai2605.19457Alibaba0 citesMay 19, 2026
- MayHuman Decision-Making with AI Assistance under Correlated Featuresno summary yetcs-ai2606.20628CMU0 citesMay 27, 2026
- Apr2026 Roadmap on Artificial Intelligence and Machine Learning for Smart Manufacturingno summary yetcs-ai2605.00839MIT1 citesApr 5, 2026
- AprOn the Identifiability of User Adaptation in Co-Adaptive Neural Interfacesno summary yetcs-ai2606.20569Stanford0 citesApr 24, 2026
- FebAI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluationno summary yetcs-ai2602.05088Berkeley0 citesFeb 4, 2026
- FebDIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoningno summary yetcs-ai2602.08586Alibaba0 citesFeb 9, 2026
- JanConvoLearn: A Learning Sciences Grounded Dataset for Fine-Tuning Dialogic AI Tutorsno summary yetcs-ai2601.08950Stanford0 citesJan 13, 2026
- JanWhen to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Modelsno summary yetcs-ai2602.13215Stanford0 citesJan 22, 2026
2024
9- NovScaleViz: Scaling Visualization Recommendation Models on Large Datano summary yetcs-ai2411.18657Adobe0 citesNov 27, 2024
- SepSpaceBlender: Creating Context-Rich Collaborative Spaces Through Generative 3D Scene Blendingno summary yetcs-ai2409.1392611 citesSep 20, 2024score 1
- JulSequential Manipulation Against Rank Aggregation: Theory and Algorithmno summary yetcs-ai2407.01916Tencent5 citesJul 2, 2024
- MayA social path to human-like artificial intelligenceno summary yetcs-ai2405.15815DeepMind35 citesMay 22, 2024
- AprDesigning for Human-Agent Alignment: Understanding what humans want from their agentsno summary yetcs-ai2404.04289Google Research20 citesApr 4, 2024
- AprAugmenting Knowledge Graph Hierarchies Using Neural Transformersno summary yetcs-ai2404.08020Adobe0 citesApr 11, 2024
- FebDelivery Optimized Discovery in Behavioral User Segmentation under Budget Constraintno summary yetcs-ai2402.03388Adobe1 citesFeb 4, 2024
- FebJack of All Trades, Master of Some, a Multi-Purpose Transformer Agentno summary yetcs-ai2402.09844HuggingFace0 citesFeb 15, 2024
- JanQuantifying stability of non-power-seeking in artificial agentsno summary yetcs-ai2401.03529DeepMind0 citesJan 7, 2024
2023
12- DecLogic-Scaffolding: Personalized Aspect-Instructed Recommendation Explanation Generation using LLMsno summary yetcs-ai2312.14345Amazon4 citesDec 22, 2023
- NovModeling subjectivity (by Mimicking Annotator Annotation) in toxic comment identification across diverse communitiesno summary yetcs-ai2311.00203Google Research0 citesNov 1, 2023
- OctThe Impact of Explanations on Fairness in Human-AI Decision-Making: Protected vs Proxy Featuresno summary yetcs-ai2310.08617Microsoft Research2 citesOct 12, 2023
- SepSeMAnD: Self-Supervised Anomaly Detection in Multimodal Geospatial Datasetsno summary yetcs-ai2309.15245Apple2 citesSep 26, 2023
- Jul`It is currently hodgepodge'': Examining AI/ML Practitioners' Challenges during Co-production of Responsible AI Valuesno summary yetcs-ai2307.10221Google Research49 citesJul 14, 2023
- JunRhythm-controllable Attention with High Robustness for Long Sentence Speech Synthesisno summary yetcs-ai2306.02593Tencent1 citesJun 5, 2023
- JunExplainable AI using expressive Boolean formulasno summary yetcs-ai2306.03976Amazon0 citesJun 6, 2023
- MayGrowing and Serving Large Open-domain Knowledge Graphsno summary yetcs-ai2305.09464Apple3 citesMay 16, 2023
- MayAugmenting Autotelic Agents with Large Language Modelsno summary yetcs-ai2305.124874 citesMay 21, 2023score 3
- FebImproving Fairness in Adaptive Social Exergames via Shapley Banditsno summary yetcs-ai2302.09298Google Research4 citesFeb 18, 2023
- FebA Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPTno summary yetcs-ai2302.09419Salesforce153 citesFeb 18, 2023
- FebLabel Information Enhanced Fraud Detection against Low Homophily in Graphsno summary yetcs-ai2302.10407Baidu47 citesFeb 21, 2023
2022
11- NovMulti-Head Adapter Routing for Cross-Task Generalizationno summary yetcs-ai2211.03831Microsoft Research3 citesNov 7, 2022
- SepOn the Horizon: Interactive and Compositional Deepfakesno summary yetcs-ai2209.01714Microsoft Research17 citesSep 5, 2022
- JulMimetic Models: Ethical Implications of AI that Acts Like Youno summary yetcs-ai2207.09394Microsoft Research17 citesJul 19, 2022
- JunDisentangled Ontology Embedding for Zero-shot Learningno summary yetcs-ai2206.03739Alibaba22 citesJun 8, 2022
- MayAsking for Knowledge: Training RL Agents to Query External Knowledge Using Languageno summary yetcs-ai2205.06111Microsoft Research2 citesMay 12, 2022
- MayHierarchically Constrained Adaptive Ad Exposure in Feedsno summary yetcs-ai2205.15759Alibaba9 citesMay 31, 2022
- FebIntent Contrastive Learning for Sequential Recommendationno summary yetcs-ai2202.02519Salesforce371 citesFeb 5, 2022
- FebTriangle Graph Interest Network for Click-through Rate Predictionno summary yetcs-ai2202.02698Alibaba18 citesFeb 6, 2022
- FebDialFRED: Dialogue-Enabled Agents for Embodied Instruction Followingno summary yetcs-ai2202.13330Amazon46 citesFeb 27, 2022
- JanTaxoCom: Topic Taxonomy Completion with Hierarchical Discovery of Novel Topic Clustersno summary yetcs-ai2201.06771Google Research20 citesJan 18, 2022
- JanDiagnosing AI Explanation Methods with Folk Concepts of Behaviorno summary yetcs-ai2201.11239Google Research0 citesJan 27, 2022
2021
39- DecRole of Human-AI Interaction in Selective Predictionno summary yetcs-ai2112.06751DeepMind1 citesDec 13, 2021
- DecFilling gaps in trustworthy development of AIno summary yetcs-ai2112.07773OpenAI49 citesDec 14, 2021
- NovExploiting a Zoo of Checkpoints for Unseen Tasksno summary yetcs-ai2111.03628Baidu0 citesNov 5, 2021
- NovAcquisition of Chess Knowledge in AlphaZerono summary yetcs-ai2111.09259DeepMind99 citesNov 17, 2021
- OctDeep Synoptic Monte Carlo Planning in Reconnaissance Blind Chessno summary yetcs-ai2110.01810Google Research0 citesOct 5, 2021
- OctReward-Punishment Symmetric Universal Intelligenceno summary yetcs-ai2110.02450DeepMind1 citesOct 6, 2021
- OctPick Your Battles: Interaction Graphs as Population-Level Objectives for Strategic Diversityno summary yetcs-ai2110.04041DeepMind0 citesOct 8, 2021
- OctThink about it! Improving defeasible reasoning by first modeling the question scenariono summary yetcs-ai2110.12349AllenAI0 citesOct 24, 2021
- OctUniversal Decision Modelsno summary yetcs-ai2110.15431Adobe0 citesOct 28, 2021
- SepLearning with Holographic Reduced Representationsno summary yetcs-ai2109.02157Amazon0 citesSep 5, 2021
- SepMulti-modal Program Inference: a Marriage of Pre-trainedLanguage Models and Component-based Synthesisno summary yetcs-ai2109.02445Microsoft Research0 citesSep 3, 2021
- SepK-AID: Enhancing Pre-trained Language Models with Domain Knowledge for Question Answeringno summary yetcs-ai2109.10547Alibaba0 citesSep 22, 2021
- JulLeveraging Domain Agnostic and Specific Knowledge for Acronym Disambiguationno summary yetcs-ai2107.00316Alibaba7 citesJul 1, 2021
- JulLearning Altruistic Behaviours in Reinforcement Learning without External Rewardsno summary yetcs-ai2107.09598Google Research1 citesJul 20, 2021
- JunReward is enough for convex MDPsno summary yetcs-ai2106.00661Google Research1 citesJun 1, 2021
- JunAliCG: Fine-grained and Evolvable Conceptual Graph Construction for Semantic Search at Alibabano summary yetcs-ai2106.01686Alibaba5 citesJun 3, 2021
- JunProper Value Equivalenceno summary yetcs-ai2106.10316Google Research0 citesJun 18, 2021
- JunThe Option Keyboard: Combining Skills in Reinforcement Learningno summary yetcs-ai2106.13105Google Research38 citesJun 24, 2021
- MayLearning to Ask Appropriate Questions in Conversational Recommendationno summary yetcs-ai2105.04774Alibaba44 citesMay 11, 2021
- MayMapGo: Model-Assisted Policy Optimization for Goal-Oriented Tasksno summary yetcs-ai2105.06350Tencent3 citesMay 13, 2021
- MayTexture Generation with Neural Cellular Automatano summary yetcs-ai2105.07299Google Research6 citesMay 15, 2021
- MayExplicit Semantic Cross Feature Learning via Pre-trained Graph Neural Networks for CTR Predictionno summary yetcs-ai2105.07752Alibaba0 citesMay 17, 2021
- MayLink Prediction on N-ary Relational Facts: A Graph-based Approachno summary yetcs-ai2105.08476Baidu5 citesMay 18, 2021
- AprCapturing Row and Column Semantics in Transformer Based Question Answering over Tablesno summary yetcs-ai2104.08303Amazon0 citesApr 16, 2021
- AprRelational Learning with Gated and Attentive Neighbor Aggregator for Few-Shot Knowledge Graph Completionno summary yetcs-ai2104.13095Alibaba1 citesApr 27, 2021
- MarLearning Reasoning Paths over Semantic Graphs for Video-grounded Dialoguesno summary yetcs-ai2103.00820Salesforce6 citesMar 1, 2021
- MarTraining a First-Order Theorem Prover from Synthetic Datano summary yetcs-ai2103.03798Google Research4 citesMar 5, 2021
- MarCausal Analysis of Agent Behavior for AI Safetyno summary yetcs-ai2103.03938Google Research6 citesMar 5, 2021
- MarPolicy-Guided Heuristic Search with Guaranteesno summary yetcs-ai2103.11505DeepMind1 citesMar 21, 2021
- MarRobust Multi-Modal Policies for Industrial Assembly via Reinforcement Learning and Demonstrations: A Large-Scale Studyno summary yetcs-ai2103.11512DeepMind4 citesMar 21, 2021
- FebAgent Incentives: A Causal Perspectiveno summary yetcs-ai2102.01685DeepMind2 citesFeb 2, 2021
- FebDiscovering a set of policies for the worst case rewardno summary yetcs-ai2102.04323DeepMind0 citesFeb 8, 2021
- FebSequential Recommendation in Online Games with Multiple Sequences, Tasks and User Levelsno summary yetcs-ai2102.06950Tencent6 citesFeb 13, 2021
- FebReasoning Over Virtual Knowledge Bases With Open Predicate Relationsno summary yetcs-ai2102.07043Google Research7 citesFeb 14, 2021
- FebHow RL Agents Behave When Their Actions Are Modifiedno summary yetcs-ai2102.07716DeepMind5 citesFeb 15, 2021
- FebLifelong Learning based Disease Diagnosis on Clinical Notesno summary yetcs-ai2103.00165Tencent0 citesFeb 27, 2021
- JanScalable Anytime Planning for Multi-Agent MDPsno summary yetcs-ai2101.04788Microsoft Research3 citesJan 12, 2021
- JanShielding Atari Games with Bounded Prescienceno summary yetcs-ai2101.08153Amazon3 citesJan 20, 2021
- JanOn the Evaluation of Vision-and-Language Navigation Instructionsno summary yetcs-ai2101.10504Google Research11 citesJan 26, 2021
2020
68- DecLearning in two-player games between transparent opponentsno summary yetcs-ai2012.02671Google Research2 citesDec 4, 2020
- DecFairness Preferences, Actual and Hypothetical: A Study of Crowdworker Incentivesno summary yetcs-ai2012.04216Google Research0 citesDec 8, 2020
- DecModel-agnostic Fits for Understanding Information Seeking Patterns in Humansno summary yetcs-ai2012.04858Google Research0 citesDec 9, 2020
- DecRelative Variational Intrinsic Controlno summary yetcs-ai2012.07827DeepMind6 citesDec 14, 2020
- DecPredicting Events in MOBA Games: Prediction, Attribution, and Evaluationno summary yetcs-ai2012.09424Tencent17 citesDec 17, 2020
- DecGet It Scored Using AutoSAS -- An Automated System for Scoring Short Answersno summary yetcs-ai2012.11243Adobe0 citesDec 21, 2020
- DecLogic Tensor Networksno summary yetcs-ai2012.13635Sony7 citesDec 25, 2020
- NovOn the role of planning in model-based deep reinforcement learningno summary yetcs-ai2011.04021DeepMind26 citesNov 8, 2020
- NovAttentive Social Recommendation: Towards User And Item Diversitiesno summary yetcs-ai2011.04797Baidu5 citesNov 9, 2020
- NovWhat Did You Think Would Happen? Explaining Agent Behaviour Through Intended Outcomesno summary yetcs-ai2011.05064Amazon1 citesNov 10, 2020
- NovGame Plan: What AI can do for Football, and What Football can do for AIno summary yetcs-ai2011.09192DeepMind10 citesNov 18, 2020
- NovUsing Unity to Help Solve Intelligenceno summary yetcs-ai2011.09294Google Research6 citesNov 18, 2020
- NovSupervised Learning Achieves Human-Level Performance in MOBA Games: A Case Study of Honor of Kingsno summary yetcs-ai2011.12582Tencent54 citesNov 25, 2020
- NovTowards Playing Full MOBA Games with Deep Reinforcement Learningno summary yetcs-ai2011.12692Tencent43 citesNov 25, 2020
- NovResonance: Replacing Software Constants with Context-Aware Models in Real-time Communicationno summary yetcs-ai2011.12715Microsoft Research0 citesNov 23, 2020
- NovMeta-learning in natural and artificial intelligenceno summary yetcs-ai2011.13464DeepMind12 citesNov 26, 2020
- NovHuman-Agent Cooperation in Bridge Biddingno summary yetcs-ai2011.14124Google Research4 citesNov 28, 2020
- OctTemporal Difference Uncertainties as a Signal for Explorationno summary yetcs-ai2010.02255DeepMind7 citesOct 5, 2020
- OctEvaluating Tree Explanation Methods for Anomaly Reasoning: A Case Study of SHAP TreeExplainer and TreeInterpreterno summary yetcs-ai2010.06734Microsoft Research2 citesOct 13, 2020
- OctExplaining Creative Artifactsno summary yetcs-ai2010.07126Salesforce0 citesOct 14, 2020
- OctCausal Inference in the Presence of Interference in Sponsored Search Advertisingno summary yetcs-ai2010.07458Microsoft Research9 citesOct 15, 2020
- OctFormalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AIno summary yetcs-ai2010.07487AllenAI66 citesOct 15, 2020
- OctQBSUM: a Large-Scale Query-Based Document Summarization Dataset from Real-world Applicationsno summary yetcs-ai2010.14108Tencent16 citesOct 27, 2020
- OctBehavior Priors for Efficient Reinforcement Learningno summary yetcs-ai2010.14274OpenAI14 citesOct 27, 2020
- SepLearning to Infer User Hidden States for Online Sequential Advertisingno summary yetcs-ai2009.01453Alibaba1 citesSep 3, 2020
- SepAction and Perception as Divergence Minimizationno summary yetcs-ai2009.01791Google Research23 citesSep 3, 2020
- SepPhysically Embedded Planning Problems: New Challenges for Reinforcement Learningno summary yetcs-ai2009.05524Google Research6 citesSep 11, 2020
- SepReceptivity of an AI Cognitive Assistant by the Radiology Community: A Report on Data Collected at RSNAno summary yetcs-ai2009.06082Amazon1 citesSep 13, 2020
- SepThe Importance of Pessimism in Fixed-Dataset Policy Optimizationno summary yetcs-ai2009.06799OpenAI23 citesSep 15, 2020
- SepJob2Vec: Job Title Benchmarking with Collective Multi-View Representation Learningno summary yetcs-ai2009.07429Baidu10 citesSep 16, 2020
- SepLarge-Scale Intelligent Microservicesno summary yetcs-ai2009.08044Microsoft Research1 citesSep 17, 2020
- SepSplitting a Hybrid ASP Programno summary yetcs-ai2009.10236Google Research0 citesSep 22, 2020
- SepMachine Knowledge: Creation and Curation of Comprehensive Knowledge Basesno summary yetcs-ai2009.11564Amazon116 citesSep 24, 2020
- SepAliMe KG: Domain Knowledge Graph Construction and Application in E-commerceno summary yetcs-ai2009.11684Alibaba60 citesSep 24, 2020
- AugSpending Money Wisely: Online Electronic Coupon Allocation based on Real-Time User Intent Detectionno summary yetcs-ai2008.09982Alibaba0 citesAug 23, 2020
- AugLearning Models of Individual Behavior in Chessno summary yetcs-ai2008.10086Microsoft Research17 citesAug 23, 2020
- AugThe Advantage Regret-Matching Actor-Criticno summary yetcs-ai2008.12234DeepMind8 citesAug 27, 2020
- JulNeuro-Symbolic Generative Art: A Preliminary Studyno summary yetcs-ai2007.02171Adobe1 citesJul 4, 2020
- JulDeep Reinforcement Learning and its Neuroscientific Implicationsno summary yetcs-ai2007.03750DeepMind30 citesJul 7, 2020
- JulEvaluating the Apperception Engineno summary yetcs-ai2007.05367Google Research1 citesJul 9, 2020
- JulStructured Policy Iteration for Linear Quadratic Regulatorno summary yetcs-ai2007.06202DeepMind6 citesJul 13, 2020
- JulPolestar: An Intelligent, Efficient and National-Wide Public Transportation Routing Engineno summary yetcs-ai2007.07195Baidu2 citesJul 11, 2020
- JulLearning Compositional Neural Programs for Continuous Controlno summary yetcs-ai2007.13363DeepMind2 citesJul 27, 2020
- JunProbing Emergent Semantics in Predictive Agents via Question Answeringno summary yetcs-ai2006.01016DeepMind5 citesJun 1, 2020
- JunAligning Superhuman AI with Human Behavior: Chess as a Model Systemno summary yetcs-ai2006.01855Microsoft Research77 citesJun 2, 2020
- JunReinforcement Learning Under Moral Uncertaintyno summary yetcs-ai2006.04734OpenAI11 citesJun 8, 2020
- JunSurveys without Questions: A Reinforcement Learning Approachno summary yetcs-ai2006.06323Adobe1 citesJun 11, 2020
- JunCompositional Generalization by Learning Analytical Expressionsno summary yetcs-ai2006.10627Microsoft Research19 citesJun 18, 2020
- JunAutoKnow: Self-Driving Knowledge Collection for Products of Thousands of Typesno summary yetcs-ai2006.13473Amazon69 citesJun 24, 2020
- JunDoes the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performanceno summary yetcs-ai2006.14779Microsoft Research70 citesJun 26, 2020
- MayTransOMCS: From Linguistic Graphs to Commonsense Knowledgeno summary yetcs-ai2005.00206AllenAI4 citesMay 1, 2020
- MayNavigating the Landscape of Multiplayer Gamesno summary yetcs-ai2005.01642DeepMind21 citesMay 4, 2020
- MayAdaptive Dialog Policy Learning with Hindsight and User Modelingno summary yetcs-ai2005.03299Baidu6 citesMay 7, 2020
- MayFinding Game Levels with the Right Difficulty in a Few Trials through Intelligent Trial-and-Errorno summary yetcs-ai2005.07677Google Research9 citesMay 15, 2020
- MayPolicy-Driven Neural Response Generation for Knowledge-Grounded Dialogue Systemsno summary yetcs-ai2005.12529Amazon10 citesMay 26, 2020
- AprDynamicEmbedding: Extending TensorFlow for Colossal-Scale Applicationsno summary yetcs-ai2004.08366Google Research1 citesApr 17, 2020
- AprFirst return, then exploreno summary yetcs-ai2004.12919OpenAI197 citesApr 27, 2020
- AprPitfalls of learning a reward function onlineno summary yetcs-ai2004.13654DeepMind0 citesApr 28, 2020
- MarEnvironment-agnostic Multitask Learning for Natural Language Grounded Navigationno summary yetcs-ai2003.00443Amazon6 citesMar 1, 2020
- MarEcological Semantics: Programming Environments for Situated Language Understandingno summary yetcs-ai2003.04567AllenAI4 citesMar 10, 2020
- MarPlacement Optimization with Deep Reinforcement Learningno summary yetcs-ai2003.08445Google Research0 citesMar 18, 2020
- FebStimulating Creativity with FunLines: A Case Study of Humor Generation in Headlinesno summary yetcs-ai2002.02031Microsoft Research0 citesFeb 5, 2020
- FebDiversity and Inclusion Metrics in Subset Selectionno summary yetcs-ai2002.03256Google Research65 citesFeb 9, 2020
- FebRL agents Implicitly Learning Human Preferencesno summary yetcs-ai2002.06137Google Research0 citesFeb 14, 2020
- FebSimulation Pipeline for Traffic Evacuation in Urban Areas and Emergency Traffic Management Policy Improvements through Case Studiesno summary yetcs-ai2002.06198Google Research5 citesFeb 14, 2020
- FebAn Overview of Distance and Similarity Functions for Structured Datano summary yetcs-ai2002.07420Google Research11 citesFeb 18, 2020
- JanInducing Cooperative behaviour in Sequential-Social dilemmas through Multi-Agent Reinforcement Learning using Status-Quo Lossno summary yetcs-ai2001.05458Adobe1 citesJan 15, 2020
- JanCorrecting Knowledge Base Assertionsno summary yetcs-ai2001.06917Tencent12 citesJan 19, 2020
2019
45- DecAbstract Reasoning with Distracting Featuresno summary yetcs-ai1912.00569Google Research30 citesDec 2, 2019
- DecWhat Can Learned Intrinsic Rewards Capture?no summary yetcs-ai1912.05500DeepMind5 citesDec 11, 2019
- DecPrioritized Unit Propagation with Periodic Resetting is (Almost) All You Need for Random SAT Solvingno summary yetcs-ai1912.05906Google Research0 citesDec 4, 2019
- DecUncovering Relations for Marketing Knowledge Representationno summary yetcs-ai1912.08374Adobe1 citesDec 18, 2019
- NovCatch & Carry: Reusable Neural Controllers for Vision-Guided Whole-Body Tasksno summary yetcs-ai1911.06636DeepMind15 citesNov 15, 2019
- NovDecision Making for Autonomous Driving via Augmented Adversarial Inverse Reinforcement Learningno summary yetcs-ai1911.08044Google Research6 citesNov 19, 2019
- NovAlgorithmic Improvements for Deep Reinforcement Learning applied to Interactive Fictionno summary yetcs-ai1911.12511Google Research4 citesNov 28, 2019
- OctEnvironmental drivers of systematicity and generalization in a situated agentno summary yetcs-ai1910.00571Google Research22 citesOct 1, 2019
- OctMaking sense of sensory inputno summary yetcs-ai1910.02227DeepMind3 citesOct 5, 2019
- OctHow does the Mind store Information?no summary yetcs-ai1910.06718Google Research1 citesOct 3, 2019
- SepDiscovery of Useful Questions as Auxiliary Tasksno summary yetcs-ai1909.04607Google Research38 citesSep 10, 2019
- SepConversational AI : Open Domain Question Answering and Commonsense Reasoningno summary yetcs-ai1909.08258Microsoft Research0 citesSep 18, 2019
- SepA Human-Centered Data-Driven Planner-Actor-Critic Architecture via Logic Programmingno summary yetcs-ai1909.09209NVIDIA2 citesSep 18, 2019
- SepZero-shot Imitation Learning from Demonstrations for Legged Robot Visual Navigationno summary yetcs-ai1909.12971Google Research5 citesSep 27, 2019
- AugInverse Rational Control with Partially Observable Continuous Nonlinear Dynamicsno summary yetcs-ai1908.04696Google Research10 citesAug 13, 2019
- AugReward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspectiveno summary yetcs-ai1908.04734DeepMind22 citesAug 13, 2019
- AugThe many Shapley values for model explanationno summary yetcs-ai1908.08474Google Research36 citesAug 22, 2019
- AugContinuous Value Iteration (CVI) Reinforcement Learning and Imaginary Experience Replay (IER) for learning multi-goal, continuous action and state space controllersno summary yetcs-ai1908.10255Sony5 citesAug 27, 2019
- JunSearch on the Replay Buffer: Bridging Planning and Reinforcement Learningno summary yetcs-ai1906.05253Google Research39 citesJun 12, 2019
- JunEmbedding Biomedical Ontologies by Jointly Encoding Network Structure and Textual Node Descriptorsno summary yetcs-ai1906.05939Google Research5 citesJun 13, 2019
- JunModeling AGI Safety Frameworks with Causal Influence Diagramsno summary yetcs-ai1906.08663Google Research9 citesJun 20, 2019
- JunInductive general game playingno summary yetcs-ai1906.09627Google Research2 citesJun 23, 2019
- JunTraining an Interactive Helperno summary yetcs-ai1906.10165Google Research0 citesJun 24, 2019
- JunLearning to Interactively Learn and Assistno summary yetcs-ai1906.10187Google Research6 citesJun 24, 2019
- MayDeclarative Question Answering over Knowledge Bases containing Natural Language Text with Answer Set Programmingno summary yetcs-ai1905.00198AllenAI8 citesMay 1, 2019
- MayStay on the Path: Instruction Fidelity in Vision-and-Language Navigationno summary yetcs-ai1905.12255Google Research17 citesMay 29, 2019
- MayConstructing High Precision Knowledge Bases with Subjective and Factual Attributesno summary yetcs-ai1905.12807Google Research2 citesMay 28, 2019
- MayLearning Compositional Neural Programs with Recursive Tree Search and Planningno summary yetcs-ai1905.12941DeepMind13 citesMay 30, 2019
- AprNeural Logic Machinesno summary yetcs-ai1904.11694Google Research46 citesApr 26, 2019
- MarLearning To Follow Directions in Street Viewno summary yetcs-ai1903.00401DeepMind7 citesMar 1, 2019
- MarAutocurricula and the Emergence of Innovation from Social Interaction: A Manifesto for Multi-Agent Intelligence Researchno summary yetcs-ai1903.00742Google Research65 citesMar 2, 2019
- MarThe StreetLearn Environment and Datasetno summary yetcs-ai1903.01292Google Research49 citesMar 4, 2019
- MarInteraction Embeddings for Prediction and Explanation in Knowledge Graphsno summary yetcs-ai1903.04750Alibaba148 citesMar 12, 2019
- MarComputing Approximate Equilibria in Sequential Adversarial Games by Exploitability Descentno summary yetcs-ai1903.05614DeepMind18 citesMar 13, 2019
- MarProspection: Interpretable Plans From Language By Predicting the Futureno summary yetcs-ai1903.08309NVIDIA15 citesMar 20, 2019
- MarIteratively Learning Embeddings and Rules for Knowledge Graph Reasoningno summary yetcs-ai1903.08948Alibaba21 citesMar 21, 2019
- MarKnowledge Aware Conversation Generation with Explainable Reasoning over Augmented Graphsno summary yetcs-ai1903.10245Baidu11 citesMar 25, 2019
- FebA Generalized Framework for Population Based Trainingno summary yetcs-ai1902.01894Google Research6 citesFeb 5, 2019
- FebELF OpenGo: An Analysis and Open Reimplementation of AlphaZerono summary yetcs-ai1902.04522Baidu42 citesFeb 12, 2019
- FebEmergent Coordination Through Competitionno summary yetcs-ai1902.07151Google Research37 citesFeb 19, 2019
- FebWorld Discovery Modelsno summary yetcs-ai1902.07685Google Research12 citesFeb 20, 2019
- FebEntity Personalized Talent Search Models with Tree Interaction Featuresno summary yetcs-ai1902.09041LinkedIn16 citesFeb 25, 2019
- FebThe Termination Criticno summary yetcs-ai1902.09996DeepMind19 citesFeb 26, 2019
- JanReachability and Differential based Heuristics for Solving Markov Decision Processesno summary yetcs-ai1901.00921NVIDIA0 citesJan 3, 2019
- JanSelf-Monitoring Navigation Agent via Auxiliary Progress Estimationno summary yetcs-ai1901.03035Salesforce134 citesJan 10, 2019
2018
48- DecDeep Reinforcement Learning and the Deadly Triadno summary yetcs-ai1812.02648DeepMind111 citesDec 6, 2018
- DecEnhancing Person-Job Fit for Talent Recruitment: An Ability-aware Neural Network Approachno summary yetcs-ai1812.08947Baidu152 citesDec 21, 2018
- NovLogic Attention Based Neighborhood Aggregation for Inductive Knowledge Graph Embeddingno summary yetcs-ai1811.01399Tencent10 citesNov 4, 2018
- NovA Trustworthy, Responsible and Interpretable System to Handle Chit Chat in Conversational Botsno summary yetcs-ai1811.07600Microsoft Research4 citesNov 19, 2018
- NovHierarchical visuomotor control of humanoidsno summary yetcs-ai1811.09656Google Research31 citesNov 23, 2018
- OctNear-Optimal Representation Learning for Hierarchical Reinforcement Learningno summary yetcs-ai1810.01257Google Research63 citesOct 2, 2018
- OctNeural-Symbolic VQA: Disentangling Reasoning from Vision and Language Understandingno summary yetcs-ai1810.02338Google Research170 citesOct 4, 2018
- OctDexterous Manipulation with Deep Reinforcement Learning: Efficient, General, and Low-Costno summary yetcs-ai1810.06045Google Research19 citesOct 14, 2018
- OctOptimizing Agent Behavior over Long Time Scales by Transporting Valueno summary yetcs-ai1810.06721DeepMind17 citesOct 15, 2018
- SepImprobotics: Exploring the Imitation Game using Machine Intelligence in Improvised Theatreno summary yetcs-ai1809.01807Google Research1 citesSep 6, 2018
- SepLearning to Collaborate: Multi-Scenario Ranking via Multi-Agent Reinforcement Learningno summary yetcs-ai1809.06260Alibaba14 citesSep 17, 2018
- SepTalent Search and Recommendation Systems at LinkedIn: Practical Challenges and Lessons Learnedno summary yetcs-ai1809.06481LinkedIn1 citesSep 18, 2018
- SepIn-Session Personalization for Talent Searchno summary yetcs-ai1809.06488LinkedIn0 citesSep 18, 2018
- SepTStarBots: Defeating the Cheating Level Builtin AI in StarCraft II in the Full Gameno summary yetcs-ai1809.07193Tencent54 citesSep 19, 2018
- SepResilient Computing with Reinforcement Learning on a Dynamical System: Case Study in Sortingno summary yetcs-ai1809.09261Google Research0 citesSep 25, 2018
- AugMulti-Hop Knowledge Graph Reasoning with Reward Shapingno summary yetcs-ai1808.10568Salesforce37 citesAug 31, 2018
- JulColdRoute: Effective Routing of Cold Questions in Stack Exchange Sitesno summary yetcs-ai1807.00462Microsoft Research14 citesJul 2, 2018
- JulAttention Models in Graphs: A Surveyno summary yetcs-ai1807.07984Adobe59 citesJul 20, 2018
- JulActive Object Perceiver: Recognition-guided Policy Learning for Object Searching on Mobile Robotsno summary yetcs-ai1807.11174Adobe0 citesJul 30, 2018
- JunLearning to Understand Goal Specifications by Modelling Rewardno summary yetcs-ai1806.01946Google Research69 citesJun 5, 2018
- JunLearning-to-Ask: Knowledge Acquisition via 20 Questionsno summary yetcs-ai1806.08554Alibaba9 citesJun 22, 2018
- JunLearning Existing Social Conventions via Observationally Augmented Self-Playno summary yetcs-ai1806.10071Meta / FAIR5 citesJun 26, 2018
- JunRobust Neural Malware Detection Models for Emulation Sequence Learningno summary yetcs-ai1806.10741Microsoft Research5 citesJun 28, 2018
- MayPlanning and Learning with Stochastic Action Setsno summary yetcs-ai1805.02363Google Research5 citesMay 7, 2018
- MayFeedback-Based Tree Search for Reinforcement Learningno summary yetcs-ai1805.05935Tencent8 citesMay 15, 2018
- MayVisceral Machines: Risk-Aversion in Reinforcement Learning with Intrinsic Physiological Rewardsno summary yetcs-ai1805.09975Microsoft Research16 citesMay 25, 2018
- MayValue Propagation Networksno summary yetcs-ai1805.11199Google Research8 citesMay 28, 2018
- AprCompositional Obverter Communication Learning From Raw Visual Inputno summary yetcs-ai1804.02341Google Research36 citesApr 6, 2018
- AprEmergent Communication through Negotiationno summary yetcs-ai1804.03980Google Research47 citesApr 11, 2018
- AprEmergence of Linguistic Communication from Referential Games with Symbolic and Pixel Inputno summary yetcs-ai1804.03984Google Research33 citesApr 11, 2018
- AprIncomplete Contracting and AI Alignmentno summary yetcs-ai1804.04268OpenAI11 citesApr 12, 2018
- AprLearning Awareness Modelsno summary yetcs-ai1804.06318Google Research13 citesApr 17, 2018
- AprDemand-Weighted Completeness Prediction for a Knowledge Baseno summary yetcs-ai1804.11109Amazon5 citesApr 30, 2018
- MarDeep Reinforcement Learning for Sponsored Search Real-time Biddingno summary yetcs-ai1803.00259Alibaba9 citesMar 1, 2018
- MarA Dataset and Architecture for Visual Reasoning with a Working Memoryno summary yetcs-ai1803.06092Google Research13 citesMar 16, 2018
- MarIntPhys: A Framework and Benchmark for Visual Intuitive Physics Reasoningno summary yetcs-ai1803.07616Meta / FAIR11 citesMar 20, 2018
- MarA Review of Literature on Parallel Constraint Solvingno summary yetcs-ai1803.10981Adobe1 citesMar 29, 2018
- MarLearning to Navigate in Cities Without a Mapno summary yetcs-ai1804.00168Google Research52 citesMar 31, 2018
- FebMore Robust Doubly Robust Off-policy Evaluationno summary yetcs-ai1802.03493Adobe10 citesFeb 10, 2018
- FebNeural Program Search: Solving Programming Tasks from Description and Examplesno summary yetcs-ai1802.04335Google Research20 citesFeb 12, 2018
- FebM-Walk: Learning to Walk over Graphs using Monte Carlo Tree Searchno summary yetcs-ai1802.04394Microsoft Research19 citesFeb 12, 2018
- FebLearning to Search with MCTSnetsno summary yetcs-ai1802.04697Google Research28 citesFeb 13, 2018
- FebDiversity is All You Need: Learning Skills without a Reward Functionno summary yetcs-ai1802.06070Google Research97 citesFeb 16, 2018
- FebMachine Theory of Mindno summary yetcs-ai1802.07740Google Research69 citesFeb 21, 2018
- FebManipulating and Measuring Model Interpretabilityno summary yetcs-ai1802.07810Microsoft Research193 citesFeb 21, 2018
- FebBudget Constrained Bidding by Model-free Reinforcement Learning in Display Advertisingno summary yetcs-ai1802.08365Alibaba84 citesFeb 23, 2018
- JanFrom Eliza to XiaoIce: Challenges and Opportunities with Social Chatbotsno summary yetcs-ai1801.01957Microsoft Research66 citesJan 6, 2018
- JanDeep Reinforcement Fuzzingno summary yetcs-ai1801.04589Microsoft Research9 citesJan 14, 2018
2017
33- DecRecruitment Market Trend Analysis with Sequential Latent Variable Modelsno summary yetcs-ai1712.02975Baidu55 citesDec 8, 2017
- NovA Unified View of Piecewise Linear Neural Network Verificationno summary yetcs-ai1711.00455Google Research141 citesNov 1, 2017
- NovA Unified Game-Theoretic Approach to Multiagent Reinforcement Learningno summary yetcs-ai1711.00832Google Research142 citesNov 2, 2017
- NovCan Deep Reinforcement Learning Solve Erdos-Selfridge-Spencer Games?no summary yetcs-ai1711.02301Google Research8 citesNov 7, 2017
- NovBBQ-Networks: Efficient Exploration in Deep Reinforcement Learning for Task-Oriented Dialogue Systemsno summary yetcs-ai1711.05715Microsoft Research14 citesNov 15, 2017
- NovRecurrent Relational Networksno summary yetcs-ai1711.08028DeepMind15 citesNov 21, 2017
- NovCrossmodal Attentive Skill Learnerno summary yetcs-ai1711.10314Amazon0 citesNov 28, 2017
- OctPRM-RL: Long-range Robotic Navigation Tasks by Combining Reinforcement Learning and Sampling-based Planningno summary yetcs-ai1710.03937Google Research23 citesOct 11, 2017
- OctNeural Program Meta-Inductionno summary yetcs-ai1710.04157Microsoft Research1 citesOct 11, 2017
- OctDistributional Reinforcement Learning with Quantile Regressionno summary yetcs-ai1710.10044DeepMind150 citesOct 27, 2017
- SepProsocial learning agents solve generalized Stag Hunts better than selfish onesno summary yetcs-ai1709.02865Meta / FAIR25 citesSep 8, 2017
- SepLearning with Opponent-Learning Awarenessno summary yetcs-ai1709.04326OpenAI60 citesSep 13, 2017
- SepImproving Search through A3C Reinforcement Learning based Conversational Agentno summary yetcs-ai1709.05638Adobe0 citesSep 17, 2017
- SepNeural Optimizer Search with Reinforcement Learningno summary yetcs-ai1709.07417Google Research203 citesSep 21, 2017
- SepDeepTransport: Learning Spatial-Temporal Dependency for Traffic Condition Forecastingno summary yetcs-ai1709.09585Baidu20 citesSep 27, 2017
- AugFinding Streams in Knowledge Graphs to Support Fact Checkingno summary yetcs-ai1708.07239Amazon20 citesAug 24, 2017
- JulTrust-PCL: An Off-Policy Trust Region Method for Continuous Controlno summary yetcs-ai1707.01891Google Research31 citesJul 6, 2017
- JulTowards Zero-Shot Frame Semantic Parsing for Domain Scalingno summary yetcs-ai1707.02363Microsoft Research33 citesJul 7, 2017
- JunDAC-h3: A Proactive Robot Cognitive Architecture to Acquire and Express Knowledge About the World and the Selfno summary yetcs-ai1706.03661Amazon74 citesJun 12, 2017
- JunZero-Shot Task Generalization with Multi-Task Deep Reinforcement Learningno summary yetcs-ai1706.05064Google Research115 citesJun 15, 2017
- JunFrom Propositional Logic to Plausible Reasoning: A Uniqueness Theoremno summary yetcs-ai1706.05261Adobe4 citesJun 16, 2017
- MayLearning Hard Alignments with Variational Inferenceno summary yetcs-ai1705.05524Google Research2 citesMay 16, 2017
- MayReinforcement Learning with a Corrupted Reward Channelno summary yetcs-ai1705.08417DeepMind16 citesMay 23, 2017
- AprDeep Q-learning from Demonstrationsno summary yetcs-ai1704.03732DeepMind307 citesApr 12, 2017
- AprThe Reactor: A fast and sample-efficient Actor-Critic agent for Reinforcement Learningno summary yetcs-ai1704.04651Google Research25 citesApr 15, 2017
- MarHolStep: A Machine Learning Dataset for Higher-order Logic Theorem Provingno summary yetcs-ai1703.00426Google Research30 citesMar 1, 2017
- MarFeUdal Networks for Hierarchical Reinforcement Learningno summary yetcs-ai1703.01161Google Research252 citesMar 3, 2017
- MarEmergence of Grounded Compositional Language in Multi-Agent Populationsno summary yetcs-ai1703.04908OpenAI214 citesMar 15, 2017
- MarRobustFill: Neural Program Learning under Noisy I/Ono summary yetcs-ai1703.07469Microsoft Research109 citesMar 21, 2017
- FebStabilising Experience Replay for Deep Multi-Agent Reinforcement Learningno summary yetcs-ai1702.08887Microsoft Research335 citesFeb 28, 2017
- JanSpace-Time Graph Modeling of Ride Requests Based on Real-World Datano summary yetcs-ai1701.06635Microsoft Research7 citesJan 23, 2017
- JanDeep Network Guided Proof Searchno summary yetcs-ai1701.06972Google Research12 citesJan 24, 2017
- JanLearn&Fuzz: Machine Learning for Input Fuzzingno summary yetcs-ai1701.07232Microsoft Research43 citesJan 25, 2017
2016
25- DecInteraction Networks for Learning about Objects, Relations and Physicsno summary yetcs-ai1612.00222Google Research577 citesDec 1, 2016
- DecKnowledge Completion for Generics using Guided Tensor Factorizationno summary yetcs-ai1612.03871AllenAI0 citesDec 12, 2016
- DecCrowdsourced Outcome Determination in Prediction Marketsno summary yetcs-ai1612.04885Microsoft Research7 citesDec 14, 2016
- DecA Sparse Nonlinear Classifier Design Using AUC Optimizationno summary yetcs-ai1612.08633Microsoft Research1 citesDec 27, 2016
- NovNeuro-Symbolic Program Synthesisno summary yetcs-ai1611.01855Microsoft Research105 citesNov 6, 2016
- NovLearning to Navigate in Complex Environmentsno summary yetcs-ai1611.03673Google Research368 citesNov 11, 2016
- NovLink Prediction using Embedded Knowledge Graphsno summary yetcs-ai1611.04642Microsoft Research0 citesNov 14, 2016
- Nov#Exploration: A Study of Count-Based Exploration for Deep Reinforcement Learningno summary yetcs-ai1611.04717OpenAI344 citesNov 15, 2016
- NovThe Off-Switch Gameno summary yetcs-ai1611.08219OpenAI12 citesNov 24, 2016
- NovGuessWhat?! Visual object discovery through multi-modal dialogueno summary yetcs-ai1611.08481DeepMind405 citesNov 23, 2016
- NovNeural Combinatorial Optimization with Reinforcement Learningno summary yetcs-ai1611.09940Google Research278 citesNov 29, 2016
- OctIdentifying Unknown Unknowns in the Open World: Representations and Policies for Guided Explorationno summary yetcs-ai1610.09064Microsoft Research44 citesOct 28, 2016
- SepBayesian Reinforcement Learning: A Surveyno summary yetcs-ai1609.04436Adobe218 citesSep 14, 2016
- JulRobust Natural Language Processing - Combining Reasoning, Cognitive Semantics and Construction Grammar for Spatial Languageno summary yetcs-ai1607.05968Sony6 citesJul 20, 2016
- JunUnifying Count-Based Exploration and Intrinsic Motivationno summary yetcs-ai1606.01868DeepMind247 citesJun 6, 2016
- JunAssessing Human Error Against a Benchmark of Perfectionno summary yetcs-ai1606.04956Microsoft Research26 citesJun 15, 2016
- JunCompression of Neural Machine Translation Models via Pruningno summary yetcs-ai1606.09274Google Research50 citesJun 29, 2016
- MayBuilding a Large Scale Dataset for Image Emotion Recognition: The Fine Print and The Benchmarkno summary yetcs-ai1605.02677Snap81 citesMay 9, 2016
- MayLearning to Communicate with Deep Multi-Agent Reinforcement Learningno summary yetcs-ai1605.06676Google Research466 citesMay 21, 2016
- AprMoving Beyond the Turing Test with the Allen AI Science Challengeno summary yetcs-ai1604.04315AllenAI6 citesApr 14, 2016
- AprQuestion Answering via Integer Programming over Semi-Structured Knowledgeno summary yetcs-ai1604.06076AllenAI22 citesApr 20, 2016
- AprCompact-Table: Efficiently Filtering Table Constraints with Reversible Sparse Bit-Setsno summary yetcs-ai1604.06641Google Research0 citesApr 22, 2016
- FebLearning to Communicate to Solve Riddles with Deep Distributed Recurrent Q-Networksno summary yetcs-ai1602.02672Google Research86 citesFeb 8, 2016
- FebValue Iteration Networksno summary yetcs-ai1602.02867OpenAI228 citesFeb 9, 2016
- FebQ($λ$) with Off-Policy Correctionsno summary yetcs-ai1602.04951DeepMind7 citesFeb 16, 2016
2015
10- DecIncreasing the Action Gap: New Operators for Reinforcement Learningno summary yetcs-ai1512.04860DeepMind24 citesDec 15, 2015
- DecDeep Reinforcement Learning in Large Discrete Action Spacesno summary yetcs-ai1512.07679Google Research289 citesDec 24, 2015
- NovLearning Simple Algorithms from Examplesno summary yetcs-ai1511.07275OpenAI24 citesNov 23, 2015
- NovA Roadmap towards Machine Intelligenceno summary yetcs-ai1511.08130Meta / FAIR29 citesNov 25, 2015
- OctTime-Sensitive Bayesian Information Aggregation for Crowdsourcing Systemsno summary yetcs-ai1510.06335Microsoft Research1 citesOct 21, 2015
- JunQuizz: Targeted crowdsourcing with a billion (potential) usersno summary yetcs-ai1506.01062Google Research154 citesJun 2, 2015
- JunDeep Knowledge Tracingno summary yetcs-ai1506.05908Google Research631 citesJun 19, 2015
- MayMetareasoning for Planning Under Uncertaintyno summary yetcs-ai1505.00399Microsoft Research6 citesMay 3, 2015
- AprInformation Gathering in Networks via Active Explorationno summary yetcs-ai1504.06423Microsoft Research15 citesApr 24, 2015
- FebPolicy Gradient for Coherent Risk Measuresno summary yetcs-ai1502.03919Adobe35 citesFeb 13, 2015
2014
12- NovHow Many Workers to Ask? Adaptive Exploration for Collecting High Quality Labelsno summary yetcs-ai1411.0149Microsoft Research7 citesNov 1, 2014
- NovAutomatic Generation of Alternative Starting Positions for Simple Traditional Board Gamesno summary yetcs-ai1411.4023Microsoft Research3 citesNov 14, 2014
- NovCompress and Controlno summary yetcs-ai1411.5326DeepMind2 citesNov 19, 2014
- OctTowards a Model Theory for Distributed Representationsno summary yetcs-ai1410.5859Google Research7 citesOct 21, 2014
- OctA Statistical Decision-Theoretic Framework for Social Choiceno summary yetcs-ai1410.7856Google Research17 citesOct 29, 2014
- JunAlgorithms for CVaR Optimization in MDPsno summary yetcs-ai1406.3339Adobe67 citesJun 12, 2014
- MarOn Redundant Topological Constraintsno summary yetcs-ai1403.0613Baidu32 citesMar 3, 2014
- FebInformation Aggregation in Exponential Family Marketsno summary yetcs-ai1402.5458Microsoft Research4 citesFeb 22, 2014
- JanActive Tuples-based Scheme for Bounding Posterior Beliefsno summary yetcs-ai1401.3833Google Research7 citesJan 16, 2014
- JanA Utility-Theoretic Approach to Privacy in Online Servicesno summary yetcs-ai1401.3859Microsoft Research52 citesJan 16, 2014
- JanTopological Value Iteration Algorithmsno summary yetcs-ai1401.3910Google Research47 citesJan 16, 2014
- JanRobust Local Search for Solving RCPSP/max with Durational Uncertaintyno summary yetcs-ai1401.4595Google Research34 citesJan 18, 2014
2013
30- SepTighter Linear Program Relaxations for High Order Graphical Modelsno summary yetcs-ai1309.6848Microsoft Research8 citesSep 26, 2013
- MarDiagnosis of Multiple Faults: A Sensitivity Analysisno summary yetcs-ai1303.1463Microsoft Research0 citesMar 6, 2013
- MarCausal Independence for Knowledge Acquisition and Inferenceno summary yetcs-ai1303.1468Microsoft Research1 citesMar 6, 2013
- MarUtility-Based Abstraction and Categorizationno summary yetcs-ai1303.1469Microsoft Research0 citesMar 6, 2013
- MarA Synthesis of Logical and Probabilistic Reasoning for Program Understanding and Debuggingno summary yetcs-ai1303.1488Microsoft Research0 citesMar 6, 2013
- FebPerception, Attention, and Resources: A Decision-Theoretic Approach to Graphics Renderingno summary yetcs-ai1302.1547Microsoft Research49 citesFeb 6, 2013
- FebStructure and Parameter Learning for Causal Independence and Causal Interaction Modelsno summary yetcs-ai1302.1561Microsoft Research30 citesFeb 6, 2013
- FebDecision-Theoretic Troubleshooting: A Framework for Repair and Experimentno summary yetcs-ai1302.3563Microsoft Research5 citesFeb 13, 2013
- FebA Graph-Theoretic Analysis of Information Valueno summary yetcs-ai1302.3596Microsoft Research21 citesFeb 13, 2013
- FebEfficient Enumeration of Instantiations in Bayesian Networksno summary yetcs-ai1302.3605Microsoft Research6 citesFeb 13, 2013
- FebAutomating Computer Bottleneck Detection with Belief Netsno summary yetcs-ai1302.4932Microsoft Research14 citesFeb 20, 2013
- FebA Definition and Graphical Representation for Causalityno summary yetcs-ai1302.4956Microsoft Research16 citesFeb 20, 2013
- FebLearning Bayesian Networks: A Unification for Discrete and Gaussian Domainsno summary yetcs-ai1302.4957Microsoft Research126 citesFeb 20, 2013
- FebA Bayesian Approach to Learning Causal Networksno summary yetcs-ai1302.4958Microsoft Research25 citesFeb 20, 2013
- FebDisplay of Information for Time-Critical Decision Makingno summary yetcs-ai1302.4959Microsoft Research158 citesFeb 20, 2013
- FebReasoning, Metareasoning, and Mathematical Truth: Studies of Theorem Proving under Limited Resourcesno summary yetcs-ai1302.4960Microsoft Research17 citesFeb 20, 2013
- FebExploiting System Hierarchy to Compute Repair Plans in Probabilistic Model-based Diagnosisno summary yetcs-ai1302.4986Microsoft Research6 citesFeb 20, 2013
- FebLearning Gaussian Networksno summary yetcs-ai1302.6808Microsoft Research6 citesFeb 27, 2013
- FebA New Look at Causal Independenceno summary yetcs-ai1302.6814Microsoft Research0 citesFeb 27, 2013
- FebLearning Bayesian Networks: The Combination of Knowledge and Statistical Datano summary yetcs-ai1302.6815Microsoft Research4 citesFeb 27, 2013
- JanA Bayesian Approach to Tackling Hard Computational Problemsno summary yetcs-ai1301.2279Microsoft Research116 citesJan 10, 2013
- JanPerfect Tree-Like Markovian Distributionsno summary yetcs-ai1301.3834Microsoft Research12 citesJan 16, 2013
- JanA Decision Theoretic Approach to Targeted Advertisingno summary yetcs-ai1301.3842Microsoft Research50 citesJan 16, 2013
- JanDependency Networks for Collaborative Filtering and Data Visualizationno summary yetcs-ai1301.3862Microsoft Research21 citesJan 16, 2013
- JanConversation as Action Under Uncertaintyno summary yetcs-ai1301.3883Microsoft Research137 citesJan 16, 2013
- JanProceedings of the Nineteenth Conference on Uncertainty in Artificial Intelligence (2003)no summary yetcs-ai1301.4606Microsoft Research124 citesJan 19, 2013
- JanProceedings of the Seventeenth Conference on Uncertainty in Artificial Intelligence (2001)no summary yetcs-ai1301.4607Microsoft Research431 citesJan 19, 2013
- JanQuantifier Elimination for Statistical Problemsno summary yetcs-ai1301.6698Microsoft Research37 citesJan 23, 2013
- JanAttention-Sensitive Alertingno summary yetcs-ai1301.6707Microsoft Research311 citesJan 23, 2013
- JanThe Lumiere Project: Bayesian User Modeling for Inferring the Goals and Needs of Software Usersno summary yetcs-ai1301.7385Microsoft Research661 citesJan 30, 2013
2012
12- DecFinding Optimal Bayesian Networksno summary yetcs-ai1301.0561Microsoft Research112 citesDec 12, 2012
- DecFactorization of Discrete Probability Distributionsno summary yetcs-ai1301.0568Microsoft Research0 citesDec 12, 2012
- DecReduction of Maximum Entropy Models to Hidden Markov Modelsno summary yetcs-ai1301.0570Microsoft Research1 citesDec 12, 2012
- OctPractically Perfectno summary yetcs-ai1212.2503Microsoft Research3 citesOct 19, 2012
- SepTextual Features for Programming by Exampleno summary yetcs-ai1209.3811Microsoft Research0 citesSep 17, 2012
- JulPrediction, Expectation, and Surprise: Methods, Designs, and Study of a Deployed Traffic Forecasting Serviceno summary yetcs-ai1207.1352Microsoft Research73 citesJul 4, 2012
- JulOn the optimality of tree-reweighted max-product message-passingno summary yetcs-ai1207.1395Microsoft Research102 citesJul 4, 2012
- JulStructured Region Graphs: Morphing EP into GBPno summary yetcs-ai1207.1426Microsoft Research34 citesJul 4, 2012
- JulSiGMa: Simple Greedy Matching for Aligning Large Knowledge Basesno summary yetcs-ai1207.4525Microsoft Research27 citesJul 19, 2012
- JunCT-NOR: Representing and Reasoning About Events in Continuous Timeno summary yetcs-ai1206.3280Microsoft Research25 citesJun 13, 2012
- JunInference for Multiplicative Modelsno summary yetcs-ai1206.3296Microsoft Research6 citesJun 13, 2012
- JunStudies in Lower Bounding Probabilities of Evidence using the Markov Inequalityno summary yetcs-ai1206.5242Google Research12 citesJun 20, 2012
2011
5- SepAn Approximation of the Universal Intelligence Measureno summary yetcs-ai1109.5951DeepMind8 citesSep 27, 2011
- JunFinding a Path is Harder than Finding a Treeno summary yetcs-ai1106.1799Microsoft Research27 citesJun 9, 2011
- JunA Comprehensive Trainable Error Model for Sung Music Queriesno summary yetcs-ai1107.0054Microsoft Research16 citesJun 30, 2011
- MarRepresenting First-Order Causal Theories by Logic Programsno summary yetcs-ai1103.4558Google Research4 citesMar 23, 2011
- FebFrom Machine Learning to Machine Reasoningno summary yetcs-ai1102.1808Microsoft Research24 citesFeb 9, 2011