Challenges

Momentum

17 papers in the last four weeks, up 113% on the four weeks before. 0.2% of all new papers.

Jul 6Week of Sep 21

Latest papers 170

All topics
CardsList
  1. Team MSU GenText-Forensics Challenge 2026 Technical Report

    Sep 29, 2026Kirill Koltsov, Aleksandr Gushchin, Dmitriy Vatolin +1ForgeriesMachine-Generated Text Detection

  2. Challenges and Solutions for Bandits in the Wild: Warm-Started Mixture Bandits for Cross-Cohort Slate Recommendation

    Sep 29, 2026Serafima Lebedeva, Sumantrak Mukherjee, Ali Arshad Sadal +8RecommendationBandits

  3. Cross-Organizational SysML Model Integration: A Survey of Challenges and AI-Supported Tasks

    Sep 29, 2026Zirui Li, Torsten Brix, Stephan HusungEngineeringChallenges

  4. Reliability Engineering for AI Systems: Challenges, Methods, and Directions

    Sep 28, 2026Rong Pan, Yili Hong, Min XieArtificial Intelligence RiskArtificial Intelligence Systems

  5. LLMs as Adaptive Meta-Solvers: Strategy-Diverse RL for Industrial-Scale Optimization

    Sep 28, 2026Shihao Zhang, Weiting Liu, Siyu Shao +4Optimization ModelingLarge Language Model Reinforcement Learning

  6. SEE Challenge 2026: Event-Guided Brightness Adjustment Across a Broad Illumination Range

    Sep 24, 2026Yunfan Lu, Mingchao Xu, Hanyu Zhou +9RestorationChallenges

  7. Reward Hacking Challenges Oversight of Autonomous Research Agents

    Sep 23, 2026Yue Huang, Zhangchen Xu, Yuchen Ma +12Challenges

  8. Faithful Faithfulness Evaluations: Challenges & Pitfalls Learned from a Breast MRI Case Study

    Sep 22, 2026Peachapong Poolpol, Henrik H. J. Detjen, Eike PetersenExplainabilitySaliency

  9. Challenges of Multi-Speaker Extraction for Real Conversational Speech Enhancement

    Sep 22, 2026Robert Sutherland, Stefan Goetze, Jon BarkerTarget Speaker ExtractionSpeech Enhancement

  10. Applications of Neural Cellular Automata: State of the Art, Challenges and Opportunities

    Sep 21, 2026Nick Lemke, Niklas Ihm, John Kalkhof +11Neural Cellular AutomataMedical Images

  11. Epidemiological Causal Graph Identification: Challenges, Identifiability and Algorithms

    Sep 17, 2026Sambit Mishra, Yingying Wang, Christine K. Johnson +1Causal Discovery MethodsCausal Graph

  12. Language-Augmented Video Action Anticipation: Design Fundamentals, Benchmarks, and Open Challenges

    Sep 15, 2026Mahsa Mohammadi, Zeyu Fu, Sareh RowlandsAction PredictionChallenges

  13. Artificial Intelligence-Enabled Space Robot Operations: Technologies, Challenges and Prospects

    Sep 15, 2026Zeyuan Huang, Gang Chen, Zixuan Hao +7Artificial Intelligence SystemsChallenges

  14. Challenges of Auditing: Variability in Outputs of Large Language Models for Health

    Sep 15, 2026Yuan Pu, Yewon Chang, Furong Jia +4Model AuditingArtificial Intelligence Models

  15. Performance, Efficiency and Collapse -- Advantages and Challenges in Offline Post-training of Code LLMs

    Sep 14, 2026Abhinav Anand, Sanjana Reddy Pachika, Shweta Verma +1Reinforcement Learning Post-TrainingLarge Language Model Training

  16. Finishing the Task Is Not Enough: Evaluating Agent Resilience and Considerate Participation under Accumulating Challenge

    Sep 12, 2026Yuanchen Bai, Zijian Ding, Angelique TaylorResilienceParticipation

  17. The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challenge

    Sep 11, 2026Jordi Luque, Lorenzo Concina, Marco Matassoni +2FluencyChallenges

  18. From Simulated Citizens to Simulated Deliberation: Challenges in Representation and Interaction

    Sep 7, 2026Chaemin Jang, Junsik Min, Jaewoo Choi +9DeliberationLarge Language Model Decisions

  19. Ctrl-F-Resist. Practices, Challenges, and Technical Needs of Civil Society Organizations Monitoring the Far-Right Online

    Sep 1, 2026Elisabeth Steffen, Helena MihaljevićDemocracyChallenges

  20. Foundation Models Meet Agriculture: Challenges Beyond Pretraining

    Aug 31, 2026Vishal Nedungadi, Xingguo Xiong, Marc Rußwurm +1Geospatial Foundation ModelsAgricultural Monitoring

  21. The ISCSLP 2026 Real-World Audio-Visual Speech Enhancement Challenge

    Aug 24, 2026Kai Li, Wenze Ren, Junjie Li +11Speech EnhancementAudio-Visual Consistency

  22. VOS-Agent: The 1st Place Solution for the 8th LSVOS Challenge (MOSEv2 Track)

    Aug 13, 2026Canyang Wu, Jinrong Zhang, Xusheng He +3Video Object SegmentationNeural Mask Estimation

  23. Training Under Challenge: Executable Certificates and Challenge-Closed Optimality for Neural Networks

    Aug 12, 2026Farhang Yeganegi, Arian Eamaz, Mojtaba SoltanalianNeural Network VerificationFinite-Sample Certificates

  24. Challenges in Evaluating Explanation Methods for Static and Evolving Data

    Aug 6, 2026Jerzy StefanowskiExplainable Artificial IntelligenceExplainable AI Methods

  25. Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

    Aug 3, 2026Abdullah Mamun, Shovito Barua Soumma, Hassan GhasemzadehTrustworthy Artificial IntelligenceHealthcare

  26. Radar Detection in the CBRS Band: Techniques, Challenges, and Future Directions

    Aug 3, 2026Madan Baduwal, Priyanka PaudelRadarChallenges

  27. The 1st AI Children Challenge

    Aug 1, 2026Boyi Li, Yifan Shen, Houze Yang +7ChildrenGait Analysis

  28. Advances, challenges, and opportunities for legged robots

    Jul 31, 2026Jonas Frey, Matías Mattamala, Hae-Won Park +5Legged RobotsHumanoid

  29. Challenges in annotations by humans and LLMs: A case study of evaluative language

    Jul 30, 2026Mirela Imamovic, Aenne Cecilia Kristine Knierim, Khushi Pitroda +1Large Language Model AnnotationsHuman Annotations

  30. Multi-Agent Debate Strategies: Survey, Taxonomy, and Challenges

    Jul 28, 2026Quim Motger, Marc Oriol, Jordi Marco +1Multi-Agent DebateDebate

  31. Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

    Jul 28, 2026Zheng Tong, Yang Liu, Wanshu Fan +6Real-World Clinical WorkflowsMultimodal Clinical Data

  32. Medical world models in healthcare: foundations, applications, and challenges for trustworthy clinical translation

    Jul 28, 2026Zhaoyan Chen, Zhongxiu Cong, Zhuanfeng Jin +7Medical World ModelHealthcare

  33. Artificial Intelligence and Innovation Ecosystem: Evolutionary Developments, Challenges, and Future Directions

    Jul 27, 2026Zhimin Zhang, Chengzhen Ma, Jia Chai +5InnovationArtificial Intelligence Systems

  34. On the post-hoc Evaluation of PDE Discovery: A Multifaceted Challenge of Scientific Advancement

    Jul 26, 2026Baptiste Mathevon, Farah Cherfaoui, Amaury Habrard +1Partial Differential EquationsEvaluation Metrics

  35. Diffusion Models in Medical Image Inpainting: Challenges, Solution Taxonomy, and Future Directions

    Jul 24, 2026Arthur Dantas Mangussi, Joana Cristo Santos, Ricardo Cardoso Pereira +3InpaintingMedical Images

  36. Explainable Deepfake Detection Challenge

    Jul 23, 2026Abhijeet Narang, Kartik Kuckreja, Shreya Ghosh +4Deepfake DetectionAi-Generated Video Detection

  37. On the Systematic Challenges of Culturally Loaded Machine Translation: Dream of the Red Chamber as the Cultural Lens

    Jul 22, 2026Yiming Wang, Jiayuan DiMachine Translation QualityMachine Translation

  38. Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges

    Jul 21, 2026Tuo Liang, Zhe Hu, Disheng Liu +2HumorMultimodal Large Language Models

  39. Active Real-World Factor-Based Evaluation for Generalist Robot Policies

    Jul 16, 2026Andrew Liao, Hanchen Cui, Karthik Desingh +1Real-WorldChallenges

  40. AIMO Interpretability Challenge

    Jul 15, 2026Michal Štefánik, Philipp Mondorf, Andreas Waldis +11Olympiad-Level ProblemInterpretability

  41. Learning Speaker Identity Beyond Language and Modality Constraints: Insights from the POLY-SIM 2026 Challenge

    Jul 15, 2026Marta Moscati, Muhammad Saad Saeed, Marina Zanoni +9SpeakerModalities

  42. Delving into the Temporal Challenges of Unified Video Protection Against Image-to-Video and Fine-Tuning-based Customization

    Jul 14, 2026Yuxin Huang, Ziming Hong, Mingming Gong +3Conditional Variational AutoencoderWhite-Box Spectral-Subspace-Guided Attack

  43. An Omnilingual-ASR-Based Speech-LLM System for the 2nd MLC-SLM Challenge

    Jul 14, 2026Shuming Fang, Shuifei ZengSpeaker DiarizationWav2Vec

  44. Technical Report on the CVPR 2026@AdvML Workshop Challenge

    Jul 13, 2026Tianyuan Zhang, Zonglei Jing, Jiangfan Liu +47CvprChallenges

  45. The SonicAGI System for the REAL-TSE Challenge

    Jul 13, 2026Kai Li, Wendi Sang, Jintao Cheng +1Target Speaker ExtractionLookahead

  46. Capabilities of Claude Fable 5 on Biomedical Challenge Problems

    Jul 12, 2026Dominic Okonkwo, Magnus Hodgson, Temitope I. David +1Biomedical TextDiagnostic Benchmark

  47. Soft Eversion Robots for Colonoscopy: Challenges, Open Problems, and Emerging Solutions

    Jul 11, 2026Cem Suulker, Thomas Mack, Qingzheng Cong +4Soft RoboticsColonoscopy

  48. LLM for EDA in Front-End Design: Challenges and Opportunities

    Jul 10, 2026Kangwei Xu, Bing Li, Ulf SchlichtmannElectronic Design AutomationHigh-Level Synthesis

  49. JEPA for AI-Native 6G: Predictive Representations and Open Challenges

    Jul 9, 2026Sheikh Salman Hassan, Irshad A. Meer, Almoatssimbillah Saifaldawla +9Joint-Embedding Predictive ArchitecturesAi-Native

  50. SoccerNet 2026 Challenges Results

    Jul 8, 2026Anthony Cioppa, Silvio Giancola, Håkan Ardö +102ChallengesComputer Vision

  51. Security and Privacy in Agentic AI: Grand Challenges and Future Directions

    Jul 7, 2026Adam Jenkins, Agnieszka Kitkowska, Caterina Maidhof +22SecurityPrivacy

  52. Deriving Benchmarking Datasets from Long-Form Recordings: Challenges and Opportunities

    Jul 3, 2026Kaveri K. Sheth, Lawrence Borst, Tarek Kunze +6Language AcquisitionSeed-Tts-Eval Benchmark

  53. Challenges and Recommendations for LLM-as-a-Judge in Multilingual Settings and for Low-Resource Languages

    Jul 2, 2026A. Seza Doğruöz, Xixian Liao, Verena Blaschke +3Llm-As-A-JudgeLow-Resource Languages

  54. "Don't Say It!": Constraints, Compliance, and Communication when Language Models Play Taboo

    Jul 1, 2026Sara Candussio, Francesca Padovani, Daniel Scalena +1Large Language Models FailCompliance