GPT 4
GPT-4, a large language model, is being extensively researched for its capabilities across diverse tasks, including translation, code analysis, educational assessment, and medical information extraction. Current research focuses on evaluating its performance against human benchmarks, exploring its limitations (e.g., susceptibility to prompt engineering and inconsistencies in complex reasoning), and developing methods to improve its reliability and efficiency, including the use of prompt engineering and ensemble methods with other machine learning models. These investigations are crucial for understanding GPT-4's strengths and weaknesses, informing its responsible deployment in various applications, and advancing the broader field of large language model development.
Papers
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4
Sachin Yadav, Tejaswi Choppa, Dominik Schlechtweg
GPT-4 vs. Human Translators: A Comprehensive Evaluation of Translation Quality Across Languages, Domains, and Expertise Levels
Jianhao Yan, Pingchuan Yan, Yulong Chen, Judy Li, Xianchao Zhu, Yue Zhang
CryptoGPT: a 7B model rivaling GPT-4 in the task of analyzing and classifying real-time financial news
Ying Zhang, Matthieu Petit Guillaume, Aurélien Krauth, Manel Labidi
Generative AI for Enhancing Active Learning in Education: A Comparative Study of GPT-3.5 and GPT-4 in Crafting Customized Test Questions
Hamdireza Rouzegar, Masoud Makrehchi
Putting GPT-4o to the Sword: A Comprehensive Evaluation of Language, Vision, Speech, and Multimodal Proficiency
Sakib Shahriar, Brady Lund, Nishith Reddy Mannuru, Muhammad Arbab Arshad, Kadhim Hayawi, Ravi Varma Kumar Bevara, Aashrith Mannuru, Laiba Batool
Is GPT-4 conscious?
Izak Tait, Joshua Bensemann, Ziqi Wang
Look Further Ahead: Testing the Limits of GPT-4 in Path Planning
Mohamed Aghzal, Erion Plaku, Ziyu Yao
Iterative Length-Regularized Direct Preference Optimization: A Case Study on Improving 7B Language Models to GPT-4 Level
Jie Liu, Zhanhui Zhou, Jiaheng Liu, Xingyuan Bu, Chao Yang, Han-Sen Zhong, Wanli Ouyang
A Two-dimensional Zero-shot Dialogue State Tracking Evaluation Method using GPT-4
Ming Gu, Yan Yang
Building another Spanish dictionary, this time with GPT-4
Miguel Ortega-Martín, Óscar García-Sierra, Alfonso Ardoiz, Juan Carlos Armenteros, Ignacio Garrido, Jorge Álvarez, Camilo Torrón, Iñigo Galdeano, Ignacio Arranz, Oleg Vorontsov, Adrián Alonso