Large Models

Recent momentum

+31%

17 papers in the last 28 days · 0.3% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-21

8 new papers

A weekly snapshot of new work published in Large Models.

Period ending 2026-09-14

2 new papers

A weekly snapshot of new work published in Large Models.

Period ending 2026-09-07

8 new papers

A weekly snapshot of new work published in Large Models.

193 papers

Latest in Large Models

  1. When is Warmstarting Effective for Scaling Language Models?

    May 13, 2026Neeratyoy Mallik, Maciej Janowski, Johannes Hog +4Large ModelsCold-Start

  2. Task-Adaptive Embedding Refinement via Test-time LLM Guidance

    May 12, 2026Ariel Gera, Shir Ashury-Tahan, Gal Bloch +2RefinementLarge Models

  3. Tokens-per-Parameter Coverage Is Critical for Robust LLM Scaling Law Extrapolation

    May 8, 2026Joshua Shay Kricheli, Alexander Lawrence Reid, Soumajyoti Sarkar +2Scaling LawsLarge Models

  4. Compute Optimal Tokenization

    May 2, 2026Tomasz Limisiewicz, Artidoro Pagnoni, Srini Iyer +6Token CompressionLarge Models

  5. Low-Rank Adaptation Redux for Large Models

    Apr 23, 2026Bingcong Li, Yilang Zhang, Georgios B. GiannakisLow-Rank AdaptationLarge Models

  6. On the Rejection Criterion for Proxy-based Test-time Alignment

    Apr 17, 2026Ayoub Hammal, Pierre Zweigenbaum, Caio CorroRejectionLarge Models

  7. SCATR: Simple Calibrated Test-Time Ranking

    Apr 16, 2026Divya Shyamal, Marta Knežević, Lan Tran +3Test-Time ScalingScoring

  8. When Does Sparsity Mitigate the Curse of Depth in LLMs

    Mar 16, 2026Dilxat Muhtar, Xinyuan Song, Sebastian Pokutta +4Sparse Expert ActivationSparsity