Application Programming Interfaces

Momentum

9 papers in the last four weeks, up 200% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 57

All topics
CardsList
  1. Community-Driven API and AI Writer Design for Openly Scaling Community Notes

    Sep 30, 2026Brad Miller, Jay Baxter, Jiansong Chao +3Ai-Assisted WritingCrowdsourced Framework

  2. RACaP: Agentic Reasoning, Acting, and Coding as Policies for Evolvable Robot Learning

    Sep 24, 2026Zexi Li, Yehang Zhang, Haojian Huang +10Robot PoliciesScalable Robot Learning

  3. Pinocchio: Fast Uncertainty Estimates for Black-Box Language Models

    Sep 21, 2026Kevin David Hayes, Arka Pal, Haosong Zhang +2Large Language Model UncertaintyUncertainty

  4. Delphi Scanner: efficient and interpretable static malware detection via API sequence modeling

    Sep 17, 2026Bijied Brahimi, Vincent Cohadon, Gabriel Glazman +3MalwareMalicious Agents

  5. MAGS: Multi-agent Auto-formalization Guarantees Safety for Agentic Outputs

    Sep 16, 2026Albert Wu, Nicholas Roberts, Tzu-Heng Huang +5Probabilistic SafetyModel-Based Multi-Agent Systems

  6. API Benchmark Scores Do Not Reliably Transfer to Chatbot Interfaces

    Sep 8, 2026Jennifer Wang, Joachim Baumann, Daniel E. Ho +1Application Programming InterfacesScores

  7. ττ\tau^\tau-Bench: An Environment for End-To-End, Realistic Agent Construction

    Sep 7, 2026Quan Shi, Keshav Dhandhania, Karthik Narasimhan +1Agentic BenchmarksCoding Agents

  8. VAKRA: Evaluating Multi-Hop Reasoning Across APIs and Retrieval Under Tool-Use Policies

    Aug 12, 2026Ankita Rajaram Naik, Anupama Murthi, Benjamin Elder +6Agentic BenchmarksReasoning Benchmark

  9. Failure-Aware Long-Form Translation: Design and Implementation of a Recoverable LLM Translation System

    Aug 10, 2026Yanlin YuTranslation QualityLarge Language Models Fail

  10. ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

    Aug 8, 2026Yang Liu, Shiwei Hou, Xiyuan Chen +13Large Language Model AgentsEvaluation Agent

  11. Caliber: Cross-Architecture Extraction-Cost Control for Score-Returning APIs

    Aug 2, 2026Chi Wang, Hanwen Wang, Yu Xia +2Model-Agnostic DefenseCross-Architecture Generalization

  12. Hidden APIs in Language Models: Discovering Reusable Causal Interfaces from Forked Futures

    Jul 30, 2026SiYuan Ma, Yiqin Luo, Zhangji +8Hidden StatesApplication Programming Interfaces

  13. Domain-Adapted Power Curve for Cross-Farm Applications

    Jul 22, 2026Ahmadreza Chokhachian, V. Roshan Joseph, Yu DingOffshore Wind TurbineDomain Adaptation

  14. Casting Everything to Online API Services? A Survey of Integrating Localized Speech Recognition Models in Robotic Systems

    Jul 13, 2026Sheng Li, Jing Li, Felix Schijve +2End-To-End Automatic Speech Recognition ModelsAutomatic Speech Recognition

  15. Understanding the Impact of AI Code Assistants on Security API Usage: An Empirical Study

    Jul 13, 2026Zahra Mousavi, Chadni Islam, M. Ali Babar +2CopilotApplication Programming Interfaces

  16. Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

    Jun 9, 2026Zhixin Ma, Yutong Zhou, Yongqi Li +2Embodied Artificial IntelligenceMultimodal Large Language Models

  17. Data-aware Static Analysis: Improving Detection of Semantic Faults in Machine Learning Code Using Data Characteristics

    Jun 8, 2026Willem Meijer, Kristian Sandahl, Dániel VarróProgram AnalysisMle-Bench Lite

  18. Differentially Private Synthetic Data via APIs 4: Tabular Data

    Jun 6, 2026Toan Tran, Arturs Backurs, Zinan Lin +3Synthetic Tabular DataSynthetic Data

  19. Domain-Adapted Small Language Models with Hybrid Post-Processing: Achieving Cost-Efficient, Low-Latency Multi-Label Structured Prediction via LoRA Fine-Tuning on Scarce Data

    Jun 4, 2026Srinivasan Manoharan, Dilipkumar Nallusamy, Sachin Kumar +1Multi-Label ClassificationApplication Programming Interfaces

  20. ADK Arena: Evaluating Agent Development Kits via LLM-as-a-Developer

    Jun 4, 2026Jintao Huang, Xiaomin Li, Gaurav Mittal +1Agentic FrameworkChatbot Arena

  21. Self-Reflective APIs: Structure Beats Verbosity for AI Agent Recovery

    Jun 3, 2026Arquimedes Canedo, Grama ChethanRetryingFeedback

  22. Token Rankings are Unforgeable Language Model Signatures

    Jun 3, 2026Matthew Finlayson, Andreas Grivas, Xiang Ren +1RankingFingerprint

  23. Sophrosyne: Agentic Exploration of Relational Data Systems Needs Moderation

    May 29, 2026Madhav Jivrajani, Ramnatthan Alagappan, Aishwarya GanesanText-To-SqlAgentic Discovery

  24. CacheProbe: Auditing Prompt Cache Isolation in Gateway APIs

    May 28, 2026Ryan FaheyCachePrompt Engineering

  25. KBF: Knowledge Boundary as Fingerprint for Language Model and Black-Box API Auditing

    May 28, 2026Yijia Fang, Yiqing Feng, Bingyu Li +1Model AuditingApplication Programming Interfaces

  26. Multi-Agent LLM-based Metamorphic Testing for REST APIs

    May 27, 2026Shehroz Khan, Abdullah Mughees, Gaadha Sudheerbabu +2Application Programming InterfacesAgentic Workflows

  27. HarnessAPI: A Skill-First Framework for Unified Streaming APIs and MCP Tools

    May 21, 2026Edwin JoseApplication Programming InterfacesModel Context Protocol

  28. Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

    May 17, 2026Yuxuan Lu, Ziyi Wang, Yingzhou Lu +12CallLarge Language Model Tool Use

  29. Agentic Interpretation: Lattice-Structured Evidence for LLM-Based Program Analysis

    May 12, 2026Jacqueline L. Mitchell, Chao WangProgram AnalysisApplication Programming Interfaces

  30. Bye Bye Perspective API: Lessons for Building and Governing Measurement Infrastructure

    Apr 28, 2026David Hartmann, Manuel Tonneau, Angelie Kraft +7Large Language Model EvaluationToxicity

  31. SeqShield: A Behavioral Analysis Approach to Uncover Rootkits

    Apr 26, 2026Paras Ghodeshwar, Sandeep K Shukla, Anand Handa +1MalwareDetection Performance

  32. Behavioral Consistency and Transparency Analysis on Large Language Model API Gateways

    Apr 22, 2026Guanjie Lin, Yinxin Wan, Shichao Pei +3TransparencyApplication Programming Interfaces

  33. Real Money, Fake Models: Deceptive Model Claims in Shadow APIs

    Mar 2, 2026Yage Zhang, Yukun Jiang, Zeyuan Chen +3Application Programming InterfacesDeception

  34. Declarative by Design, Assistable Only by Convention: Benchmarking Multi-Agent Frameworks for AI-Assistability

    Feb 3, 2026Shafiuddin Rehan Ahmed, Sourabh DeshpandeAi-Assisted Programming TasksApplication Programming Interfaces

  35. Mind the Gap: Action Rebinding Attacks against Android GUI Agents

    Jan 18, 2026Yi Qian, Kunwei Qian, Xingbang He +7Graphical User Interface AgentsMalicious Agents

  36. Benchmarking Web API Integration Code Generation

    Sep 24, 2025Daniel Maninger, Leon Chemnitz, Amir Molzam Sharifloo +2Code GenerationApplication Programming Interfaces

  37. Beyond Semantic Similarity: Reducing Unnecessary API Calls via Behavior-Aligned Retriever

    Aug 20, 2025Yixin Chen, Ying Xiong, Shangyu Wu +4Large Language Model Tool UseRetrievers