Layer-Wise

Latest papers 104

All topics
CardsList
  1. Before It Fades: Reinforcing Temporal Representations at Inference Time in VideoLLMs

    Oct 1, 2026Youngwoo Shin, Yusung Ro, Minseo Kim +1Inference-TimeLayer-Wise

  2. Persistent Depth Ordering amid Shifting Block-Bypass Responses in Language Model Pretraining

    Oct 1, 2026Shengye Tao, Yinzhu Cheng, Haihua XieLayer-WiseResponses

  3. Increasing Width Allows Greedy Layer-wise Training to Rival End-to-End Backpropagation in Self-Supervised Learning

    Sep 30, 2026Syon Mansur, Joel ZylberbergBackpropagationLayer-Wise

  4. Right In-Place (RiP) Convolution: A Simple, General, and Near-Optimal Strategy for Memory-Efficient CNN Inference

    Sep 30, 2026Opegbemi Matthias Busoye, Tolulope Matthew Busoye, Eghonghon-aye EigbeNeural Network InferenceConvolutional

  5. Draft in Parallel, Condition Through Depth: Adjacent Causal Injection for Speculative Decoding

    Sep 28, 2026Haohui Zhang, Keyu Chen, Haocheng Sun +4Speculative DecodingDrafter

  6. LAYERSCOPE: A Layerwise Characterization of Video and Multimodal Learned Representations

    Sep 23, 2026Sandra Arcos-Holzinger, Debashish Chakraborty, Rohita Mocharla +7Multimodal RepresentationsFine-Grained Video Understanding

  7. Beyond Depth Truncation: Controlled Evaluation of Depth Utilization in Recursive Language Models

    Sep 17, 2026Ha Van Dau, Thanh Tung Khuat, Nguyen Thanh DungLayer-WiseTruncation

  8. Pre-PEFT Probing: Weight Statistics and Perturbation Robustness for Layer Selection in VLM Vision Encoders

    Sep 14, 2026Qingtao Xia, Jiahua Bao, Siyao Cheng +1Parameter-Efficient Fine-Tuning MethodsVision-Language Model Adaptation

  9. Temporal Recurrence Favors Fewer Layers

    Sep 14, 2026Ivan Anokhin, Johan Obando-Ceron, Irina Rish +1Recurrent ModelRecurrence

  10. Charts Are Beyond Pixels: Probing for Layer-Wise Chart Understanding and Editing

    Sep 8, 2026Xiaochuan Zhong, Yifan Hou, Chenxi Pang +1ChartLayer-Wise

  11. Some Emotions Run Deeper: Layer-wise Probing and Causal Intervention in Large Language Models

    Sep 1, 2026Tian Fang, Gaël Guibon, Davide BuscaldiEmotionLayer-Wise

  12. Toppling the Hierarchy in Byte-level Language Modeling

    Aug 31, 2026Lukas Edman, Alexander FraserLanguage ModelingLayer-Wise

  13. Analyzing Speech Condition Effects in Dysarthric ASR: A Layer-wise Probing Study

    Aug 3, 2026Darwin Jelestin Muthu, Navya Gupta, Wei Lin Tay +3Dysarthric SpeechAutomatic Speech Recognition

  14. Divergent large language model predictions from convergent representations in ambiguous word pairs

    Aug 3, 2026K. Jack Scott, Narun Pat, Veronica LiesaputraTransformer EncoderLayer-Wise

  15. Explicit Layer Modeling for Video Object Insertion and Video Layer Decomposition

    Jul 28, 2026Kyujin Han, Seungjoo Shin, Sunghyun ChoLayer-WiseDecomposition

  16. The Intruder Threshold: A Spectral Law for LoRA Fine-Tuning

    Jul 26, 2026Peng XieModel Fine-TuningLayer-Wise

  17. Rethinking Layer-Wise Information Allocation for Vision Foundation Model Adaptation

    Jul 24, 2026Yuqi Li, Xi Xiao, Yunbei Zhang +6Recent Vision Foundation ModelsFrozen Vision-Language Models

  18. LionVote: Per-Layer Learning Rate Adaptation for Lion

    Jul 10, 2026Kris AtallahBatchLayer-Wise

  19. When Synthetic Speech Is All You Have: Better Call GRPO

    Jul 9, 2026Shashi Kumar, Yanis Labrak, Hasindri Watawana +5Flow-GrpoSpeech-To-Text Alignment

  20. Prompt Compression via Activation Aggregation

    Jul 9, 2026Thibaud Ardoin, Semira Einsele, Evis Bregu +1Large Language Model CompressionInstruction-Tuned Models

  21. Understanding Layer Patching in Model Size Interpolation

    Jul 9, 2026Sara Kangaslahti, Jonathan Geuter, Nihal V. Nayak +3Model SizeLayer-Wise

  22. InsideSSL: Understanding Self-Supervised Speech Representations using a Model-Centric Perspective

    Jul 7, 2026Samir Sadok, Xavier Alameda-PinedaSelf-Supervised Speech ModelsWav2Vec

  23. Localized LoRA-MoE: Block-wise Low-Rank Experts With Adaptive Routing

    Jul 6, 2026Babak Barazandeh, Subhabrata Majumdar, Vinay Prithyani +1Layer-WiseLocalized Lora-Moe

  24. Layer-Parallel Inference Reduces Encrypted Nonlinear Depth in Transformers

    Jul 6, 2026Ligong Han, Kai Xu, Hao Wang +3Homomorphic EncryptionTransformer Architectures

  25. Dynamic Neural Graph Encoding of Inference Processes in Deep Weight Space

    Jul 2, 2026Di Wu, Huan Liu, Zhixiang Chi +3Neural RepresentationsNeural Network Inference

  26. Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

    Jul 1, 2026Zijian Zhang, Rizhen Hu, Athanasios Glentis +4Transformer ArchitecturesLayer-Wise

  27. Gradient Smoothing: Coupling Layer-wise Updates for Improved Optimization

    Jun 29, 2026Haoming Meng, Anton Sugolov, Vardan PapyanGradientLayer-Wise

  28. Layer-wise Probing of wav2vec 2.0 and Whisper for Consonant Cluster Reduction in African American English

    Jun 22, 2026Hamid Mojarad, Kevin TangWav2VecGrapheme-To-Phoneme

  29. Tapered Language Models

    Jun 22, 2026Reza Bayat, Ali Behrouz, Aaron CourvilleLanguage ModelingLayer-Wise

  30. How Well Do Self-Supervised Speech Models Encode Age and Gender in Children's Speech? A Layer-Wise Analysis Across Multiple Architectures

    Jun 20, 2026Abhijit Sinha, Hemant Kumar Kathania, Mohit Joshi +3Self-Supervised Speech ModelsWav2Vec

  31. Topological Neural Dynamics: A Neuron-wise Framework for Sequence Modeling

    Jun 19, 2026Borui Cai, Yao ZhaoNeural DynamicsSequence Modeling

  32. RegimeVGGT: Layer-Wise Spatially Preserving Redundancy Removal for Visual Geometry Grounded Transformer

    Jun 16, 2026Jinhao You, Shuo Lyu, Zhuohang Lyu +5Visual Geometry Grounded TransformerSelf-Supervised Vision Transformers

  33. Calibrated Sampling-Free Uncertainty Estimation in Bayesian Deep Learning

    Jun 15, 2026Tobias Jan Wieczorek, Leon de Andrade, Thomas Möllenhoff +1Bayesian Neural NetworksUncertainty

  34. Do Video Foundation Models Understand Intuitive Physics? A Layerwise Probing Analysis

    Jun 8, 2026Samuele Punzo, Niccolò Caselli, Ippokratis Pantelidis +3Video Foundation ModelsFoundation Model

  35. Trajectory Geometry of Transformer Representations Across Layers

    Jun 8, 2026Vishal Pandey, Gopal Singh, Yacine MahdidTransformer ArchitecturesLayer-Wise

  36. The Geometry of Last-Layer Model Stealing

    Jun 5, 2026Snigdha Chandan KhilarLayer-WiseTransformer Architectures

  37. Skip a Layer or Loop It? Learning Program-of-Layers in LLMs

    Jun 4, 2026Ziyue Li, Yang Li, Tianyi ZhouLLM Inference OptimizationLLM Reasoning Strategies

  38. Class-Specific Branch Attention for Mitigating Gradient Interference under Class Imbalance

    Jun 4, 2026Arush Singhal, Umang SoniImbalanced ClassificationClass Imbalance

  39. Dominant-Layer ZO: A Single Layer Dominates Zeroth-Order Fine-Tuning of LLMs

    Jun 3, 2026Wanhao Yu, Ziyan Wang, Zheng Wang +7Continual Fine-TuningZeroth-Order

  40. Depth-Attention: Cross-Layer Value Mixing for Language Models

    Jun 3, 2026Boyi Zeng, Yiqin Hao, Zitong Wang +7Layer-WiseCross-Layer Interactions

  41. PURGE: Projected Unlearning via Retain-Guided Erasure

    Jun 2, 2026Vedant Jawandhia, Daksh Ahuja, Ghufran Alam Siddiqui +3Machine UnlearningConcept Erasure

  42. Beyond Compression: Quantifying Spectral Accessibility in Vision Representations

    Jun 2, 2026Akayou A. Kitessa, Yijun ZhaoVision EncodersSpectral

  43. LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training

    May 28, 2026Minju Gwak, Minseo Kwak, Dongseok Lee +3Reinforcement Learning Post-TrainingRepresentation-Level Misalignment

  44. When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception

    May 28, 2026Vahideh ZolfaghariDeceptionLayer-Wise

  45. Model Merging by Output-Space Projection

    May 27, 2026Bethan Evans, Benjamin Etheridge, Stephen Roberts +1Continual Model MergingMulti--Task Learning

  46. Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations

    May 27, 2026Simardeep Singh, Paras ChopraPerceptuallyLayer-Wise

  47. Locality-Aware Redundancy Pruning for LLM Depth Compression

    May 27, 2026Vincent-Daniel Yun, Youngrae Kim, Woosang Lim +3Large Language Model CompressionLayer-Wise