Key papers π₯ Trending (default) π Most cited π Newest first π€ A β Z by title 60 papers Β· trending (default) numbers = π₯ heat
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing (2026) Xinjie Zhang et al.
13.98 AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents (2026) Kunlun Zhu et al.
11.84 Wan-Streamer v0.2: Higher Resolution, Same Latency (2026) Lianghua Huang et al.
10.59 ISO: An RLVR-Native Optimization Stack (2026) Hanqing Zhu et al.
8.86 PACE: A Proxy for Agentic Capability Evaluation (2026) Yueqi Song et al.
8.24 NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs (2026) Jiarong Zhao et al.
7.96 FilmWorld: Agentic Novel-to-Film Generation through Dynamic Cinematic World Modeling (2026) Jialong Zuo et al.
7.86 Multi-Agent LLMs Fail to Explore Each Other (2026) Hyeong Kyu Choi et al.
7.36 Out-of-Distribution Generalization in Time Series: A Survey (2025) Xin Wu et al.
7.35 Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning (2026) Lu Dai et al.
7.34 Deep Learning for Multivariate Time Series Imputation: A Survey (2024) Jun Wang et al.
7.10 Large-Language-Models-as-a-Judge in Theory-Agnostic Adaptive Metric-Alignment for Prototypical Networks in Personality Recognition (2026) Jing Jie Tan et al.
5.89 Evidence-Backed Video Question Answering (2026) Shijie Wang et al.
5.88 BEDTime: A Unified Benchmark for Automatically Describing Time Series (2025) Medhasweta Sen et al.
5.57 AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity (2026) Bhavya Gupta et al.
5.49 Decentralized Stochastic Subgradient-type Methods with Communication Compression for Nonsmooth Nonconvex Optimization (2026) Siyuan Zhang et al.
5.49 Full Bayesian Reinforcement Learning via LF-IBIS (2026) Stefano Masini et al.
5.01 TUDUM: A Turkish-Thinking Reasoning Pipeline for Qwen3.5-27B (2026) Baran Bingol et al.
5.01 OpenSafeIntent: Evaluating Intent-Calibrated Safe Completion Across Dual-Use Prompt Sets (2026) Rheeya Uppaal et al.
5.01 G-RRM: Guiding Symbolic Solvers with Recurrent Reasoning Models (2026) Timo Bertram et al.
5.01 COBART: Controlled, Optimized, Bidirectional and Auto-Regressive Transformer for Ad Headline Generation (2026) Yashal Shakti Kanungo et al.
5.01 ICDAR 2026 HIPE-OCRepair Competition on LLM-Assisted OCR Post-Correction for Historical Documents (2026) Maud Ehrmann et al.
5.01 ProsMAE: Multi-Source MAE Pretraining for ISUP Grade Classification (2026) Anna Jung et al.
5.01 Track2Map: Online Deformable SLAM with Motion-Aware Pose Optimization in Robotic Surgery (2026) Tianyi Song et al.
5.01 When Synthetic Speech Is All You Have: Better Call GRPO (2026) Shashi Kumar et al.
5.01 When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability (2026) Zongyou Yang et al.
5.01 A Practical Investigation of Training-free Relaxed Speculative Decoding (2026) Guoxuan Xia et al.
5.01 Conceptual Networks for Cross-Linguistic Idiomatic Expressions: A Feature-Based Graph Approach (2026) Kiran Pala et al.
5.01 STEC: Evidence Compression for Deep Search in Open-domain Multi-Hop QA (2026) Xinkang Li et al.
5.01 Distribution-First Population Simulation: Collapse, Calibration, and Recall in Non-WEIRD LLM Persona Modeling (2026) Gurkan Ozkan
5.01 AutoIndex: Learning Representation Programs for Retrieval (2026) Sam O'Nuallain et al.
5.01 Norm or Direction? Decoding Vision Mambas for High-Resolution Vision (2026) Jin Yu et al.
5.01 Evaluating medical AI under missing information: same-provider judges and human raters change apparent safety (2026) Koyar Afrasyab
5.01 Dual Adversarial Fine-tuning for Enhancing Robustness of Large Vision Language Model (2026) Sibo Wang et al.
5.01 Measuring Reward-Seeking via Contrastive Belief Updates (2026) Axel H{\o}jmark et al.
5.01 CoGoal3D: Collaborative 3D Object Detection with 3D-Aware Fusion and Refinement (2026) Zhihao Yang et al.
5.01 Prompt Design at Scale: How Format, Instruction Count, and Context Length Shape Instruction Adherence and Hallucination in Large Language Models (2026) Netanel Eliav
5.00 Active Sensing for RIS-Aided Tracking and Power Control: A Hybrid Neuroevolution and Supervised Learning Approach (2026) George Stamatelis et al.
4.39 From World Models to World Action Models: A Concise Tutorial for Robotics (2026) Xiaoxiong Zhang et al.
4.39 Beyond Detection: Redesigning Assessment and Governande of Generative AI at the Universidad Polit\'ecnica de Madrid (UPM) (2026) Jessica D\'iaz et al.
4.39 Benchmarking Federated Learning and Knowledge Distillation for Point Cloud Classification (2026) Aizierjiang Aiersilan
4.39 Black-Box Inference of LLM Architectural Properties with Restrictive API Access (2026) Christopher Ellis et al.
4.39 Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance (2026) Laxmipriya Ganesh Iyer
4.39 Evolutionary Feature Engineering for Structured Data (2026) Ege Onur Taga et al.
4.39 Reformalization of the Jordan Curve Theorem (2026) Simon Guilloud et al.
4.39 An Exploratory Study on LLM-Generated Code and Comments in Code Repositories (2026) Yongyi Ji et al.
4.39 ContextSniper: AntTrail's Token-Efficient Code Memory for Repository-Level Program Repair (2026) Chiwang Luk et al.
4.39 ContextNest: Verifiable Context Governance for Autonomous AI Agent (2026) Misha Sulpovar (PromptOwl et al.
4.39 Automated grading of Linux/bash examinations using large language models: a four-level cognitive taxonomy approach (2026) Manuel Alonso-Carracedo et al.
4.39 Automated Recommendation of Programming Learning Content Using Pattern-based Knowledge Components (2026) Muntasir Hoq et al.
4.39 KAT-Coder-V2.5 Technical Report (2026) Bo Huang et al.
4.39 Decision Protocols in Multi-Agent Large Language Model Conversations (2026) Lars Benedikt Kaesberg
4.39 EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems (2026) Kenneth Benavides et al.
4.39 Memory in the Loop: In-Process Retrieval as ExtendedWorking Memory for Language Agents (2026) Yusuf Khan et al.
4.39 From Closed-Loop Optimization to Open Decision Making: Coupled Digital Twins for Predictive and Autonomous Microscopy (2026) Yu Liu et al.
4.39 PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation (2026) Hyungseok Song et al.
4.39 DT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrail (2026) He Liu et al.
4.39 Danus: Orchestrating Mathematical Reasoning Agents with Fact-Graph Memory (2026) Jihao Liu et al.
4.39 FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games (2026) Chase McDonald et al.
4.39 Reaction-network reasoning with frontier models for experimentally confirmed catalyst-selectivity hypotheses (2026) Sutanay Choudhury et al.
4.39