Automated Software Testing

Latest papers 93

All topics
CardsList
  1. Multi-Label Perceptual Bug Detection in Video Games using Deep Learning on Gameplay Footage

    Oct 6, 2026Nahian Rifaat, Felix Morosov, Loutfouz ZamanAutomated Software TestingMulti-Label Classification

  2. FastCI: Efficient GPU-Intensive CI for LLM Training Frameworks

    Oct 1, 2026Tianshuo Qiao, Naiqian Zheng, Xiaopeng Liu +4Automated Software Testing

  3. Do Agent Benchmarks Do What They Say? An Executable-Contract Audit of Tool-Using Agent Environments

    Sep 29, 2026Rohith Reddy Bellibatlu, Zichong Wang, Wenbin ZhangTool-Use EvaluationAutomated Software Testing

  4. Scenario-Driven Neuroevolution: Using Models to Guide Test Generation for Games

    Sep 23, 2026Gijs van Cuyck, Patric Feldmeier, Jan Tretmans +1Automated Software TestingAutomated Test Generation

  5. FairTest: Search-Based Fairness Testing for Multi-Agent Reinforcement Learning Systems

    Sep 23, 2026Xiaotong Wang, Xuan XieAutomated Software TestingAlgorithmic Fairness

  6. IWC-Bench: Evaluating Web Application Generation from a Software Testing Perspective

    Sep 14, 2026Chenxu Liu, Zilu Zou, Peizhong Gao +9Automated Software TestingWeb Application Generation

  7. LLM-Assisted Dynamic Threat Analysis for Attacker-Reachable Software Weaknesses in Autonomous Vehicles

    Aug 13, 2026Md Wasiul Haque, Sagar Dasgupta, Mizanur Rahman +1Automated Software TestingSoftware Vulnerability Detection

  8. Directed Neuro-Symbolic Stochastic Execution for Verification of Distributed Parallel AI Programs

    Aug 8, 2026Gautham Koorma, Vikas Sharma, George Edwards +1Automated Software TestingFormal Verification

  9. WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader

    Aug 6, 2026Boshui Chen, Huiping Liu, Shaolei ZhangRL for Code GenerationAutomated Software Testing

  10. IcFuzz: Fuzzing Isaac Sim with Semantic Stage Guidance and Multi-level Mutation

    Aug 6, 2026Zhixiang Chen, Zhuangbin Chen, Ruoxi Jia +4Automated Software TestingRobotics Simulation

  11. Can LLMs Test Terminal User Interfaces?

    Aug 4, 2026Chao Peng, Ruida Hu, Ajitha Rajan +3Automated Software TestingAutomated Test Generation

  12. Coding Agents as Test-Suite Auditors: Finding What Official Suites Miss While Approaching What They Catch

    Aug 3, 2026Shuyang Xie, Shuxiao Xie, Feng Zhu +2Automated Software TestingBenchmark Auditing

  13. Validation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?

    Jul 30, 2026Xiaonan Xu, Wenjing WuAutomated Software TestingAutomated Program Repair

  14. DragonCrawl: A Generative, Intent-Based Framework for Scalable Mobile End-to-End Testing

    Jul 30, 2026Sowjanya Puligadda, Mengdie Zhang, Ali Zamani +3Mobile GUI AutomationAutomated Software Testing

  15. Agentic Method for Deterministic Validation of Legacy Code Migration

    Jul 30, 2026Andras Ferenczi, Jordan Docherty, Mariya Bessonov +2Code TranslationAutomated Software Testing

  16. Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

    Jul 28, 2026Carlos Celemin, Benedict Wilkins, Adrián Barahona-Ríos +2VLM EvaluationVLM Robustness

  17. KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models

    Jul 28, 2026Fuyuan Xia, Qixin Zhang, Chenhao Ying +5Automated Software TestingLLMs for Cybersecurity

  18. Do Coverage and Mutation Scores of LLM-Generated Test Suites Correlate with Their Effectiveness? (Replicability Study)

    Jul 24, 2026Junda Zhao, Shurui Zhou, Eldan CohenLLM EvaluationAutomated Software Testing

  19. Test Case Prioritization for DNNs via Neural Collapse Instability

    Jul 22, 2026Chunyu Liu, Mingyuan Li, Yang Li +4Automated Software Testing

  20. Beyond Fail-to-Pass: Iterative Hardening of Co-Generated Bug Reproduction Tests and Fixes

    Jul 22, 2026Yuhao Tan, Zhibang Yang, Fangkai Yang +9Automated Software TestingAutomated Program Repair

  21. Integrating High-Level Requirements to Low-Level Tests with Machine-Readable V&V Specifications

    Jul 20, 2026Mansur Arief, Nur Ahmad Khatim, Ali Akarma +1Requirements EngineeringAutomated Software Testing

  22. Faithful Autoformalization of Natural Language Assertions

    Jul 14, 2026Hongyi Liu, Madhusudan Parthasarathy, Adithya MuraliAutomated Software TestingLLM-Based Program Synthesis

  23. Fail-Aware and Explainable Test Oracle Prediction

    Jul 13, 2026Yue Zhao, Binish Tanveer, Jelena ZdravkovicAutomated Software TestingFault Detection

  24. An Empirical Study for Android-to-OpenHarmony GUI Test Migration

    Jul 13, 2026Yakun Zhang, Xinjia Chen, Yiyun Chen +6Mobile GUI AutomationAutomated Software Testing

  25. Specification Grounding Drives Test Effectiveness for LLM Code

    Jul 7, 2026Amin Haeri, Mahdi GhelichiLLM GroundingAutomated Software Testing

  26. LLM-Based Test Oracles: Source-of-Authority Taxonomy -- A Systematic Literature Review

    Jul 6, 2026Ali Hassaan Mughal, Muhammad BilalSoftware EngineeringAutomated Software Testing

  27. An Exploration of Agentic Information Fusion for Test Maintenance Prediction

    Jul 6, 2026Jingxiong Liu, Nasser Mohammadiha, Gregory GayAutomated Software TestingPredictive Maintenance

  28. Flow-A11y: Flow-Aware Accessibility Testing

    Jul 3, 2026Nasr Eddine Fliti, Leisan Kokorina, Florian Tambon +1Automated Software TestingAutomated Evaluation

  29. TestEvo-Bench: An Executable and Live Benchmark for Test and Code Co-Evolution

    Jul 2, 2026Jiale Amber Wang, Kaiyuan Wang, Pengyu NieAutomated Software TestingBenchmark Design

  30. Beyond Pixel Diffs: Benchmarking Image Change Captioning for Web UI Visual Regression Testing

    Jul 2, 2026Licheng Zhang, Bach Le, Pengtao Zhao +1Image CaptioningAutomated Software Testing

  31. MAS-Lab: A Specification-Driven Validation Framework for Reliable Multi-Agent Systems

    Jun 29, 2026Jordan Augé, Giovanna Carofiglio, Giulio Grassi +1Multi-Agent LLM SystemsAutomated Software Testing

  32. MANGO: Automated Multi-Agent Test Oracle Generation for Vision-Language-Action Models

    Jun 23, 2026Pablo Valle, Shaukat Ali, Aitor Arrieta +1Automated Software TestingVision-Language-Action Models

  33. Detecting and Understanding Vulnerabilities in Fully Homomorphic Encryption Frameworks

    Jun 21, 2026Yiteng Peng, Dongwei Xiao, Zhibo Liu +2Automated Software TestingSoftware Vulnerability Detection

  34. All Green, Still Broken: Real-Flow Verification Lessons from an LLM-Integrated, Multi-Market Web Application

    Jun 21, 2026Muhammad Bilal, Ali Hassaan MughalAutomated Software Testing

  35. Load Testing for Machine Learning Model Serving Systems at Scale

    Jun 20, 2026Amr S. Abdelfattah, Nakul Tirumalai, Indu Mohanan +4Automated Software Testing

  36. Library-Aware Doubles and Iterative Repair for Large Language Model-Generated Unit Tests in OpenSIL Firmware

    Jun 18, 2026Ma Toan Bach, Yuchi Zheng, Haingo Razafindranto +4Automated Software TestingAutomated Program Repair

  37. Code-Augur: Agentic Vulnerability Detection via Specification Inference

    Jun 17, 2026Zhengxiong Luo, Mehtab Zafar, Dylan Wolff +1Automated Software TestingSoftware Vulnerability Detection

  38. All Smoke, No Alarm: Oracle Signals in Agent-Authored Test Code

    Jun 16, 2026Dipayan Banik, Kowshik Chowdhury, Shazibul Islam ShamimAutomated Software TestingAI Coding Agents

  39. AIChilles: Automatically Uncovering Hidden Weaknesses in AI-Evolved Systems

    Jun 14, 2026Yajie Zhou, Ao Li, Ashwin Silla +2Automated Software TestingDifferential Testing

  40. AI-Driven Test Case Generation from Natural Language Requirements: A Survey of Techniques and Research Gaps

    Jun 4, 2026Orimoloye Folorunsho, Hassan RezaAutomated Software TestingAutomated Test Generation

  41. WebRISE: Requirement-Induced State Evaluation for MLLM-Generated Web Artifacts

    Jun 2, 2026Yuxin Meng, Yuhan Suo, Junjie Wang +9Automated Software TestingAutomated Evaluation

  42. FVSpec: Real-World Property-Based Tests as Lean Challenges

    May 31, 2026Quinn Dougherty, Max von Hippel, Simon Henniger +2Automated Software TestingAutomated Theorem Proving

  43. Inferring Code Correctness from Specification

    May 28, 2026Tambon Florian, Papadakis MikeAutomated Software TestingCode Generation Evaluation

  44. Multi-Agent LLM-based Metamorphic Testing for REST APIs

    May 27, 2026Shehroz Khan, Abdullah Mughees, Gaadha Sudheerbabu +2Automated Software TestingMetamorphic Testing

  45. HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML

    May 26, 2026Jiajun Wu, Jian Yang, Tuney Zheng +4Automated Software TestingAutomated Program Repair