2 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.
Oct 5, 2026·Jie Wang, Shiwei Luo, Qi Zhang +1Language Model ScalingByte-Level Language Model
School of Computer Science, East China Normal University, Shanghai, China · School of Computer Science, Fudan University, Shanghai, China
Oct 5, 2026·Lennart Carstens-Behrens, Holger FröhlichLanguage Model PretrainingMeta-Learning
Fraunhofer Institute for Algorithms and Scientific Computing SCAI · University Hospital Bonn
Sep 24, 2026·Gautam VeldandaTernary QuantizationQuantization-Aware Training
Independent Researcher
Sep 14, 2026·Kalyani Marathe, Artidoro Pagnoni, Tomasz Limisiewicz +4Decoder-Only Language ModelsLanguage Model Scaling Laws
University of Washington, Seattle · Work done at Meta FAIR · Meta FAIR
Aug 31, 2026·Lukas Edman, Alexander FraserEfficient Language Model TrainingByte-Level Language Model
School of Computation, Information and Technology, TU Munich · Munich Center for Machine Learning · Munich Data Science Institute
Aug 4, 2026·Mykola HaltiukByte-Level Language ModelTransfer Learning
Faculty of Computer Science AGH University of Krakow Krakow, Poland
Jul 31, 2026·Bo Liu, Muxuab Yu, Yu Zhang +2Sparse Mixture-of-ExpertsByte-Level Language Model
University of Bristol · School of Automation Science and Electrical Engineering,Beihang University, Beijing, China · University of Manchester
Jun 12, 2026·Sangwhan Moon, Daisuke Oba, Youmi Ma +2Byte-Level Language ModelLLM Reliability
Google LLC / Mountain View, CA, USA · Institute of Science Tokyo / Tokyo, Japan · Mohamed bin Zayed University of Artificial Intelligence / Abu Dhabi, United Arab Emirates
Jun 1, 2026·Florian Störtz, Catalin-Andrei Stan, Alexandru Dinu +4Malware ClassificationByte-Level Language Model
CrowdStrike · CrowdStrike U.K. · CrowdStrike Romania +1
May 28, 2026·Rohan ShravanLanguage Model PretrainingLLM Compression
The School of AI Bengaluru, India
May 10, 2026·Lin Zheng, Vasilisa Bashlovkina, Timothy Dozat +3Byte-Level Language ModelLanguage Modeling
1Google DeepMind · 2The University of Hong Kong
May 8, 2026·Julie Kallini, Artidoro Pagnoni, Tomasz Limisiewicz +5Speculative DecodingLLM Inference Acceleration
FAIR at Meta · Stanford University · University of Washington
May 4, 2026·Mullosharaf K. ArabovMultilingual Language ModelsTransformer
Institute of Computational Mathematics and Information Technologies Kazan Federal University Kazan, Russia
Feb 1, 2026·Zishuo Bao, Jiaqi Leng, Junxiong Wang +2LLM Fine-TuningByte-Level Language Model
Fuzhou University · NYU Shanghai · Fudan University +2