Key papers π₯ Trending (default) π Most cited π Newest first π€ A β Z by title 60 papers Β· trending (default) numbers = π₯ heat
Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation (2026) Jasmine Brazilek et al.
5.87 Can AI agents conduct open-ended AI research? Early evidence from two case studies (2026) Peter Kirgis et al.
5.46 From Mind to Machine: The Rise of Manus AI as a Fully Autonomous Digital Agent (2025) Minjie Shen et al.
5.35 Analyzing the Ethical Logic of Eight Large Language Models (2025) W. Russell Neuman et al.
5.13 Geopolitical alignment: Endorsement effects in large language models (2026) Maxim Chupilkin
5.01 Practical Graph Optimisation and AI-Driven Models for Active Directory Security Hardening (2026) Huy Q. Ngo
5.01 Agent Security Needs Redefinition through a Holistic Framework (2026) Vincent Siu et al.
5.01 A Roadmap to Impactful Pluralistic Alignment Research (2026) Elinor Poole-Dayan et al.
5.01 The User Asks, Platforms Compete: How Agentic Recommendation Markets Take Shape (2026) Deyao Hong et al.
5.01 The Hitchhiker's Guide to Monoculture (2026) Gordon Burtch
4.39 Analyzing Curricular Pattern Complexity Using AI to Improve On-Time Graduation Rates (2026) Lynn Vonderhaar et al.
4.39 AI in Cyberpsychology: A systematic literature review of Cybersecurity enhancement by using AI for analyzing psychology of Victims, Attackers, and Defenders (2026) Georg Thamer Francis et al.
4.39 AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation (2026) Quanyan Zhu
4.39 Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Internal Audit Perspective (2026) Anupa Lodhi
4.39 AI advice suppresses people's willingness to say "I don't know", even when the advice is wrong and accuracy is incentivized (2026) Chiara Marcoccia et al.
4.39 Scaling Time Series Classification via XAI-Driven Data Reduction (2026) Davide Italo Serramazza et al.
4.39 AquaAugmentor: A Novel Feature Augmentation Algorithm for Water Potability Prediction (2026) Muntasir Tabasum et al.
4.39 When Not to Automate: A Formal Protocol for Human Preservation in AI-Optimized Organizations (2026) Jose Manuel de la Chica Rodriguez et al.
4.39 LLM-Powered Agentic AI for 5G/6G Networks: A Tutorial and Survey on Architectures, Protocols, and Standardization (2026) Mazene Ameur et al.
4.39 Operational Hallucination and Safety Drift in AI Agents (2026) Shasha Yu et al.
4.39 Governing Well in the Algorithmic Age: The Foundations of Digital Statecraft (2026) Zeynep Engin et al.
4.39 What General Intelligence Requires: Non-Reducible Constraints Across Levels of Description (2026) Subhomoy Bakshi
4.39 Assessment in Team Problem-Solving Exercises in Computing Education (2026) Valdemar \v{S}v\'abensk\'y et al.
4.39 The safety failures we are not instrumenting: a perspective on hidden safety-critical challenges in modern AI systems (2026) Gjergji Kasneci et al.
4.39 Analyzing Middle School Students' Dialogue and Behaviors during Collaborative AI Chatbot Development Using Ordered Network Analysis (2026) Shan Zhang et al.
4.39 From Obligation to Specification: A Survey on Validating EU AI Act Requirements in RE (2026) T. Y. Emmy Lai et al.
4.39 A Systematic Survey on Image Description Techniques for STEM Domains (2026) Marco Cardia et al.
4.39 Cloud-Native Evaluation-as-a-Service: A Microservices Architecture for Scalable AI Monitoring with Conformal Guarantees (2026) Lei Yang
4.39 Defining AI-Native Systems: Autonomy as Revision Authority (2026) Cheng Tan
4.39 Generative and multimodal AI for materials prediction and design: Progress, challenges, and perspectives (2026) Xianyuan Liu et al.
4.39 What AI Red-Team Evaluations Can and Cannot Prove (2026) Bandana Kaur
4.39 AI-Integrated Scientific Inquiry: A Practice-Centered Vision for Science Education (2026) Arne Bewersdorff et al.
4.39 How Do AI Coding Agents Contribute to Software Development? an Empirical Study of Agentic Pull Requests (2026) Iren Mazloomzadeh et al.
4.39 Multi-Agent System-driven Digital Twins for predictive maintenance: architectures, technologies and open research challenges (2026) Korota Ars\`ene Coulibaly et al.
4.39 Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for Industrial Evidence Integration (2026) Deshui Li et al.
4.39 AI4PLE: A Methodology for Integrating AI into Product Line Engineering (2026) Bedir Tekinerdogan
4.39 Do Agent Benchmarks Measure Capability? Protocol Validity in the Age of Agentic AI (2026) Jiaqi Shao et al.
4.39 Kutti AI: A Voice-First, Offline-Capable Learning Companion with Real-Time Struggle Detection for Visually-Impaired Children (2026) Kadharmoideen Fadurudeen
4.39 Agentic Root Cause Analysis through Evidence-Grounded Reasoning (2026) Amaury Wei et al.
4.39 Dynamic Capability Scoping for Enterprise AI Agents: A Synthetic Dataset and Three-Source Permission Architecture (2026) Halil Burak Noyan
4.39 Beyond Perspectives: A Trio-Ethnography of Interpretation Evolution in LLM-Supported Programming Education (2026) Jennie Ren et al.
4.39 Explainable Reinforcement Learning for assisting Air Traffic Controllers (2026) Anduel Mehmeti et al.
4.39 Where Is the Cost of Third-Party API Routers in Agentic Software Development? (2026) Donghao Fu et al.
4.39 Extremal Chowla sets and their linear analogues: A human-AI mathematical investigation using Co-Scientist (2026) Mohsen Aliabadi et al.
4.39 Observing sycophantic AI validate others reduces its appeal but not its persuasiveness (2026) Meryl Ye et al.
4.39 The Rising Unsustainability of AI Graphics Cards Production (2026) Cl\'ement Morand et al.
3.51 Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning (2026) Ajay Patel et al.
3.51 From Operations to Elderly Care Outcomes: A Thematic Review of Industrial Engineering and Decision-Support Approaches (2026) Shayan Farhang Pazhooh et al.
3.51 Validating the Single Item Kawaii Measure (2026) Katie Seaborn et al.
3.51 The Human Utility Factor: A Computable Welfare Metric That Reframes AI Governance as a Constrained Optimisation Problem (2026) Sivasathivel Kandasamy
3.51 PUDA: An AI-Native Hardware Harness for Self-Driving Laboratories (2026) Zekun Ren et al.
3.51 Introducing Human-Centeredness in AI-Assisted Lexicography (2026) Antonio San Martin et al.
2.00 BrainPilot: Automating Brain Discovery with Agentic Research (2026) Haoxuan Li et al.
2.00 ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D (2026) Lena Libon et al.
2.00 Large-Scale ChatBot Validation Through Customer Digital Twin Simulations (2026) Cristovao Iglesias et al.
2.00 Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual Disabilities (2026) Karly V. Coffey et al.
2.00 The Age of AI Agents Demands A New Scientific Paradigm To Sustain Trustworthy Science (2026) Belinda Mo
2.00 AI Security Priorities: A Field-Wide Agenda (2026) Gil Gekker et al.
2.00 Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness Levels (2026) Xinyu Yang et al.
2.00 When benchmark inferences do not compose: Projectibility in AI evaluation (2026) Brett Reynolds
2.00