71 papers
Uncategorized
0/66Papers awaiting categorization.
Progress0 of 66
2026
15- AprHY-Embodied-0.5: Embodied Foundation Models for Real-World Agentsuncategorized2604.07430Tencent HunyuanApr 8, 2026score 9~127 min
- AprReducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularizationuncategorized2604.19079Apr 21, 2026~102 min
- MarMolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulationuncategorized2603.16861Ai2Mar 17, 2026score 9~120 min
- MarSIMART: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLMuncategorized2603.23386ByteDance SeedMar 24, 2026score 3~116 min
- MarLingshu-Cell: A generative cellular world model for transcriptome modeling toward virtual cellsuncategorized2603.25240DAMO AcademyMar 26, 2026score 2~128 min
- FebSingle-minus gluon tree amplitudes are nonzeroagents2602.12176OpenAIFeb 12, 2026score 1~121 min
- FebLIVE: Long-horizon Interactive Video World Modelinguncategorized2602.03747Microsoft ResearchFeb 3, 2026score 3~103 min
- FebDreamDojo: A Generalist Robot World Model from Large-Scale Human Videosuncategorized2602.06949NVIDIAFeb 6, 2026score 2~117 min
- FebBitDance: Scaling Autoregressive Generative Models with Binary Tokensuncategorized2602.14041ByteDanceFeb 15, 2026score 9~104 min
- FebWorld Action Models are Zero-shot Policiesuncategorized2602.15922NVIDIA Deep Imagination ResearchFeb 17, 2026score 2~101 min
- FebTOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Roboticsuncategorized2602.19313Ai2Feb 22, 2026score 3~97 min
- JanQuantifying the Effect of Test Set Contamination on Generative Evaluationsuncategorized2601.04301EleutherAIJan 7, 2026~128 min
- JanCosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planninguncategorized2601.16163NVIDIAJan 22, 2026score 2~115 min
- Jan360Anything: Geometry-Free Lifting of Images and Videos to 360°uncategorized2601.16192DeepmindJan 22, 2026score 2~106 min
- JanOpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understandingvision2601.09575NVIDIAJan 14, 2026score 2~107 min
2025
22- DecCosmos-H-Surgical: Learning Surgical Robot Policies from Videos via World Modelinguncategorized2512.23162NVIDIADec 29, 2025score 2~99 min
- DecNested Browser-Use Learning for Agentic Information Seekinguncategorized2512.23647TongyiLabDec 29, 2025score 9~115 min
- NovLearning Vision-Driven Reactive Soccer Skills for Humanoid Robotsuncategorized2511.03996ByteDance SeedNov 6, 2025score 2~136 min
- NovRynnVLA-002: A Unified Vision-Language-Action and World Modeluncategorized2511.17502DAMO AcademyNov 21, 2025score 6~109 min
- OctLanguage Models Model Languagereasoning2510.12766SnowflakeOct 14, 2025score 3~120 min
- OctCarleman Estimates for Backward Anisotropic Stochastic Parabolic Equations with General Dynamic Boundary Conditions and Applicationstraining-methods2510.12345Oct 14, 2025~132 min
- OctMemory Retrieval and Consolidation in Large Language Models through Function Tokensuncategorized2510.08203ByteDance SeedOct 9, 2025score 7~115 min
- OctHigh-Fidelity Simulated Data Generation for Real-World Zero-Shot Robotic Manipulation Learning with Gaussian Splattinguncategorized2510.10637DAMO AcademyOct 12, 2025score 3~114 min
- OctFrom Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priorsuncategorized2510.17439ByteDance SeedOct 20, 2025score 9~122 min
- OctLongCat-Video Technical Reportuncategorized2510.22200Meituan LongCatOct 25, 2025score 9~122 min
- SepRobix: A Unified Model for Robot Interaction, Reasoning and Planninguncategorized2509.01106ByteDance SeedSep 1, 2025score 9~108 min
- SepManipulation as in Simulation: Enabling Accurate Geometry Perception in Robotsuncategorized2509.02530ByteDance SeedSep 2, 2025score 2~106 min
- SepCanary-1B-v2 & Parakeet-TDT-0.6B-v3: Efficient and High-Performance Models for Multilingual ASR and ASTuncategorized2509.14128NVIDIASep 17, 2025~108 min
- AugPerch 2.0: The Bittern Lesson for Bioacousticsuncategorized2508.04665Aug 6, 2025~118 min
- AugREINA: Regularized Entropy Information-Based Loss for Efficient Simultaneous Speech Translationuncategorized2508.04946Roblox CorporationAug 7, 2025score 2~114 min
- AugTowards Affordance-Aware Robotic Dexterous Grasping with Human-like Priorsuncategorized2508.08896DAMO AcademyAug 12, 2025score 2~120 min
- AugFutureX: An Advanced Live Benchmark for LLM Agents in Future Predictionuncategorized2508.11987ByteDance SeedAug 16, 2025score 8~116 min
- JulFlexOlmo: Open Language Models for Flexible Data Usemoe2507.07024Jul 9, 2025~105 min
- JulFast and Simplex: 2-Simplicial Attention in Tritonscaling-laws2507.02754Jul 3, 2025score 10~112 min
- JulGR-3 Technical Reportuncategorized2507.15493ByteDance SeedJul 21, 2025score 5~104 min
- JunLog-Linear Attentionarchitecture2506.04761Jun 5, 2025~122 min
- MayZeroSearch: Incentivize the Search Capability of LLMs without Searchingtraining-methods2505.04588Alibaba DAMOMay 7, 2025~114 min
2024
7- OctA Survey of Small Language Modelsllm-systems2410.20011Oct 25, 2024~119 min
- Oct$π_0$: A Vision-Language-Action Flow Model for General Robot Controluncategorized2410.24164NVIDIAOct 31, 2024~111 min
- SepSortformer: A Novel Approach for Permutation-Resolved Speaker Supervision in Speech-to-Text Systemsuncategorized2409.06656NVIDIASep 10, 2024~113 min
- AugNEST: Self-supervised Fast Conformer as All-purpose Seasoning to Speech Processing Tasksuncategorized2408.13106NVIDIAAug 23, 2024~97 min
- MarAtP*: An efficient and scalable method for localizing LLM behaviour to componentsuncategorized2403.00745DeepMind0 citesMar 1, 2024score 3~123 min
- MarProsody for Intuitive Robotic Interface Design: It's Not What You Said, It's How You Said Ituncategorized2403.08144DeepMindMar 13, 2024~102 min
- FebInternLM-Math: Open Math Large Language Models Toward Verifiable Reasoninguncategorized2402.06332InternLM / Shanghai AI LabFeb 9, 2024score 9~97 min
2023
9- DecDistributional Bellman Operators over Mean Embeddingsuncategorized2312.07358DeepMindDec 9, 2023~126 min
- NovRoboVQA: Multimodal Long-Horizon Reasoning for Roboticsuncategorized2311.00899DeepMindNov 1, 2023score 5~140 min
- OctSum of the GL(3) Fourier Coefficients over Quadratics and Mixed Powersreasoning2310.11408Oct 17, 2023~132 min
- AugComputational Long Exposure Mobile Photographyno summary yetcs-cv2308.013790 citesAug 2, 2023score 2
- JulLine Search for Convex Minimizationuncategorized2307.16560DeepMindJul 31, 2023~135 min
- JunEstimating the Causal Effect of Early ArXiving on Paper Acceptancesafety2306.13891Jun 24, 2023~116 min
- MayHow Language Model Hallucinations Can Snowballuncategorized2305.13534May 22, 2023score 9~113 min
- MarEliciting Latent Predictions from Transformers with the Tuned Lensuncategorized2303.08112EleutherAIMar 14, 2023~106 min
- FebProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematicsuncategorized2302.12433EleutherAIFeb 24, 2023~96 min