cs.CVSep 28, 2026

Sparse-View Interpretable 3D Animal Behavior Representations for Neural Encoding and Decoding

Authors: Xinming Dai, Qihang Jin, Tianshu Tan, Baiyuan Chen, Hanrui Lyu, Lenny Aharon, Kyle Daruwalla, Xun Helen Hou, +3 more

Organizations: Columbia University · University of Science and Technology of China · Harvard University · University of Cambridge · Northwestern University · Cold Spring Harbor Laboratory

Abstract

A deeper understanding of brain function requires a precise, structured characterization of behavior. Yet, extracting behavioral representations from video in a form suitable for scientific analysis remains a fundamental challenge. Many prior studies represent behavior via pose estimation or nonlinear video embeddings. However, pose tracking discards rich information beyond predefined keypoints, while nonlinear video embeddings lack interpretability. We address this limitation with SABLE (Sparse-view Animal Behavior Latent Embeddings), a self-supervised framework that leverages a geometric inductive bias to learn behavior representations. By augmenting a multi-view transformer with priors from monocular depth and pose estimation, SABLE reconstructs 3D animal behavior from extremely sparse views while learning explicit 3D latent structure. Without ground-truth 3D labels, it reliably recovers 3D behavior from two-view videos, whereas state-of-the-art (SOTA) methods fail or yield degenerate solutions. Across the International Brain Lab and Cheese3D datasets, we demonstrate that SABLE learns 3D representations that match or exceed prior SOTA performance in neural encoding and decoding. Once pretrained across animals, SABLE serves as an off-the-shelf model that generalizes zero-shot to unseen animals without animal-specific calibration or retraining. Our method establishes 3D-aware video embeddings that capture complex behavior, opening new avenues for studying brain-behavior relationships.

Figures & tables

Appendix figures & tables11 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. BEAST3D: Animal behavioral analysis and neural encoding from multi-view video via Gaussian splatting

    Jun 1, 2026Yanchen Wang, Lenny Aharon, Wangshu Zhu +73D Representation

  2. Lightweight 3D Feature Pretraining by Bayesian Inversion of 2D Foundation Models

    Jun 19, 2026Marwane Hariat, Gianni Franchi, David Filliat +13D Scene Understanding3D Generation

  3. SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild

    May 8, 2026Xuyi Hu, Jin Lyu, Jiuming Liu +43D Reconstruction