Intermediate Decoder Layers

Momentum

4 papers in the last four weeks, with none the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 24

All topics
CardsList
  1. HiRAE: Hierarchical Representation Autoencoding with Residual Budgets

    Sep 29, 2026Xuanyu Zhu, Yan Bai, Yang Shi +5Autoregressive Image GenerationAutoencoder Architectures

  2. When Text Matters: Design Principles for Visual Token Pruning in Vision-Language Model

    Sep 28, 2026Minchan Kang, Kyeonghye Park, Seoyoung Cho +2Visual Token PruningRecent Vision-Language Models

  3. LightMIS: Ultra-Lightweight Medical Image Segmentation Without a Stage-Wise Decoder

    Sep 23, 2026Andrei Arhire, Mihaela-Elena Breabăn, Radu TimofteSemi-Supervised Medical Image SegmentationU-Net

  4. Accelerated Decoding of Centroid Positional Encoding for Instance Segmentation

    Sep 15, 2026Carmelo Scribano, Filippo Muzzini, Nedyalko Prisadnikov +6Intermediate Decoder LayersPositional Encoding

  5. Temperature-driven inversion and nonlinear dynamics in ChatGPT-like AIs

    Aug 2, 2026Neil F. Johnson, Frank Yingjie Huo, Bella Xinrui LiNonlinear DynamicsOpenai

  6. Lightweight Neural Networks for Affordance Segmentation: Enhancement of the Decoder Module

    Jul 31, 2026Simone Lugani, Edoardo Ragusa, Rodolfo Zunino +1AffordancesMobilenetv2

  7. Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

    Jul 29, 2026Mingkuan Feng, Zhengqi Wen, Jianhua TaoLong Visual-Token SequencesMultimodal Large Language Models

  8. Beyond Independent Labels: Schwartz-Geometry Decoding for Human Value Detection

    Jul 6, 2026Víctor Yeste, Paolo RossoMulti-Label ClassificationIntermediate Decoder Layers

  9. One Forward Beats Two: InnerZoom for Accurate and Efficient GUI Grounding

    Jun 29, 2026Chen Liu, Ling Chen, Hanzhang Zhou +5Graphical User InterfaceObject Localization

  10. Multi-task Learning is Not Enough: Representational Entanglement in Dual-output Second Language Speech Recognition

    Jun 4, 2026Seung Hwan Cho, Young-Min KimMultilingual Automatic Speech RecognitionDual-Encoder Architectures

  11. FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection

    May 30, 2026Yao Wei, Andrea Cavallaro, Changjae OhOpen-Vocabulary Object DetectionOpen-Vocabulary

  12. Dino-NestedUNet: Unlocking Foundation Vision Encoders for Pathology Tumor Bulk Segmentation via Dense Decoding

    Apr 27, 2026Tianyang Wang, Ziyu Su, Abdul Rehman Akbar +7Tumor SegmentationU-Net

  13. Addressing Image Authenticity When Cameras Use Generative AI

    Apr 23, 2026Umar Masud, Abhijith Punnappurath, Luxi Zhao +2Ai-Generated Image DetectionAuthenticity

  14. ASGNet: Adaptive Spectrum Guidance Network for Automatic Polyp Segmentation

    Apr 16, 2026Yanguang Sun, Hengmin Zhang, Jianjun Qian +2Polyp SegmentationColonoscopy

  15. Seeing the imagined: latent functional alignment in visual imagery decoding from fMRI data

    Apr 15, 2026Fabrizio Spera, Tommaso Boccato, Michal Olak +2Functional Magnetic Resonance ImagingVisual Perception

  16. pFedNavi: Structure-Aware Personalized Federated Vision-Language Navigation for Embodied AI

    Feb 16, 2026Qingqian Yang, Hao Wang, Sai Qian Zhang +6Vision-Language NavigationFederated Learning