AI Trustworthiness

Momentum

28 papers in the last four weeks, up 133% on the four weeks before. 0.3% of all new papers.

Jul 6Week of Sep 21

Latest papers 228

All topics
CardsList
  1. Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

    May 11, 2026Shijun Lei, Quang Nguyen, Swapneel S Mehta +7Multi-Agent SimulationsPersonalized Large Language Model Agents

  2. Trust or Abstain? A Self-Aware RAG Approach

    May 11, 2026Xi Zhu, Ziqi Wang, Kai Mei +5Large Language Model ReliabilityLarge Language Model Backbones

  3. The Granularity Mismatch in Agent Security: Argument-Level Provenance Solves Enforcement and Isolates the LLM Reasoning Bottleneck

    May 11, 2026Linfeng Fan, Ziwei Li, Yuan Tian +3Untrusted ContentAgentic Deployments

  4. Learning from Acceptance: Cumulative Regret in the Game of Coding

    May 10, 2026Hanzaleh Akbari Nodehi, Parsa Moradi, Mohammad Ali Maddah-AliSource-Channel CodingAcceptance

  5. Rethinking Ratio-Based Trust Regions for Policy Optimization in Multi-Agent Reinforcement Learning

    May 9, 2026Chulabhaya Wijesundara, Andrea Baisero, Zhongheng Li +3Multi-Agent Reinforcement LearningTrust Region

  6. When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning

    May 9, 2026Yann Berthelot, Philippe Preux, Riad AkrourExpertsAI Trustworthiness

  7. OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling

    May 8, 2026Yuxuan Lou, Yang YouSpectral NormScaling

  8. When to Trust Imagination: Adaptive Action Execution for World Action Models

    May 7, 2026Rui Wang, Yue Zhang, Canyang Chen +3Efficient World-Action ModelWorld Models

  9. Learned Neighbor Trust for Collaborative Deployment in Model-Agnostic Decentralized Learning

    May 6, 2026Michael Lanier, Luise Ge, Sastry Kompella +1Log-Ratio-Based Gaussian Trust WeightDecentralized Learning

  10. ITBoost: Information-Theoretic Trust for Robust Boosting

    May 6, 2026Ye Su, Longlong Zhao, Diego Garcia-Gil +4Noisy LabelsGradient Boosted Decision Tree

  11. TRACE: A Metrologically-Grounded Engineering Framework for Trustworthy Agentic AI Systems in Operationally Critical Domains

    May 5, 2026Serhii ZabolotniiTrustworthy Artificial IntelligenceAI Trustworthiness

  12. Multi-Agent Strategic Games with LLMs

    May 5, 2026Maxim ChupilkinStrategic ReasoningMulti-Agent Simulations

  13. When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

    May 4, 2026Javad Forough, Marios Kogias, Hamed HaddadiPoisoningAI Trustworthiness

  14. Trust, but Verify: Peeling Low-Bit Transformer Networks for Training Monitoring

    May 4, 2026Arian Eamaz, Farhang Yeganegi, Mojtaba SoltanalianLayer-WiseTransformer Architectures

  15. Query-Dependent Use of Generated Descriptions for Reliable Visual Question Answering

    May 3, 2026Zeshang Li, Shuoyang ZhangCaptionsAI Trustworthiness

  16. "I Don't Know" -- Towards Appropriate Trust with Certainty-Aware Retrieval Augmented Generation

    May 1, 2026Daan Di Scala, Maaike de Boer, Pınar YolumLarge Language Model ReliabilityAI Trustworthiness

  17. Auditing Frontier Vision-Language Models for Trustworthy Medical VQA: Grounding Failures, Format Collapse, and Domain Adaptation

    Apr 30, 2026Xupeng Chen, Binbin Shi, Chenqian Le +5Medical Vision-Language ModelsMedical Visual Question Answering

  18. TRUST: A Framework for Decentralized AI Service v.0.1

    Apr 29, 2026Yu-Chao Huang, Zhen Tan, Mohan Zhang +3Trustworthy Artificial IntelligenceAI Trustworthiness

  19. From Black-Box Confidence to Measurable Trust in Clinical AI: A Framework for Evidence, Supervision, and Staged Autonomy

    Apr 29, 2026Serhii Zabolotnii, Viktoriia Holinko, Olha AntonenkoTrustworthy Artificial IntelligenceAI Trustworthiness

  20. Authority Inversion in LLM-Mediated Ubiquitous Systems: When Models Trust Users Over Sensors

    Apr 28, 2026Long Zhang, Zi-bo Qin, Wei-neng ChenAuthorityAI Trustworthiness

  21. When AI reviews science: Can we trust the referee?

    Apr 26, 2026Jialiang Wang, Yuchen Liu, Hang Xu +7Peer ReviewReview

  22. How Adversarial Environments Mislead Agentic AI?

    Apr 20, 2026Zhonghao Zhan, Huichi Zhou, Zhenhao Li +3Adversarial RobustnessCyberattacks

  23. Decentralised Trust and Security Mechanisms for IoT Networks at the Edge: A Comprehensive Review

    Apr 19, 2026Khandoker Ashik Uz Zaman, Mahdi H. Miraz, Mohammed N. M. AliInternet Of ThingEdge Devices

  24. LLMs can persuade only psychologically susceptible humans on societal issues, via trust in AI and emotional appeals, amid logical fallacies

    Apr 18, 2026Alexis Carrillo, Salvatore Citraro, Ali Aghazhadeh Ardebili +5PersuasionHuman-Ai Interaction

  25. High-Risk AI Systems and the Problem of Identity in the European AI Act

    Apr 17, 2026Andrea FerrarioEu Artificial Intelligence ActGovernance

  26. A Systematic Study of Training-Free Methods for Trustworthy Large Language Models

    Apr 17, 2026Wai Man Si, Mingjie Li, Michael Backes +1Large Language Model ReliabilityLarge Language Model Training

  27. Robust Explanations for User Trust in Enterprise NLP Systems

    Apr 13, 2026Guilin Zhang, Kai Zhao, Jeffrey Friedman +3Large Language Model ReliabilityExplainability

  28. Learning When to Trust in Contextual Social Bandits

    Mar 9, 2026Majid Ghasemi, Mark CrowleyContextual Bandit FrameworkAI Trustworthiness

  29. Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise

    Mar 6, 2026Wenxin Li, Kunyu Peng, Di Wen +63D Semantic Occupancy PredictionRobotic Perception

  30. Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

    Mar 6, 2026Xisen Jin, Michael Duan, Qin Lin +4On-Chain AttestationsGuardrail

  31. CAT: Can Trust be Predicted with Context-Awareness in Dynamic Heterogeneous Networks?

    Dec 12, 2025Jie Wang, Zheng Yan, Jiahe Lan +2Log-Ratio-Based Gaussian Trust WeightGraph Neural Networks

  32. Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness

    Oct 6, 2025Amin Banayeeanzade, Ala N. Tak, Fatemeh Bahrani +5PersonalityPsychology

  33. Seven Security Challenges in Cross-domain Multi-agent LLM Systems

    May 28, 2025Ronny Ko, Jiseong Jeong, Shuyuan Zheng +4Multi-Agent Large Language Model SystemsLarge Language Model Safety

  34. Trust Under Siege: Label Spoofing Attacks against Machine Learning for Android Malware Detection

    Mar 14, 2025Tianwei Lan, Luca Demetrio, Farid Nait-Abdesselam +2MalwareAI Trustworthiness

  35. Trust-free Personalized Decentralized Learning

    Oct 15, 2024Yawen Li, Yan Li, Junping Du +3Federated LearningDecentralized Learning

  36. TRUST-SQL: Tool-Integrated Multi-Turn Reinforcement Learning for Text-to-SQL over Unknown Schemas

    Date pendingAi Jian, Xiaoyun Zhang, Eryu Guo +6Text-To-SqlSchema