cs.AIOct 8, 2026

BrickBench: Evaluating Agentic Brick Design

Authors: Peter Kulits, Yiqing Xu, R. Kenny Jones, Cordelia Schmid, Jiajun Wu

Organizations: Stanford University · Max Planck Institute for Intelligent Systems · Inria

Abstract

We propose BrickBench, a benchmark for agentic text-conditioned LEGO-set design. Given a prompt, an agent is tasked with producing an assembly that not only satisfies semantic and design criteria, but that can also be physically built. To do so, it must select parts from a discrete library and reason jointly about local and global constraints. We score validity, alignment, and design across three settings that vary in scale and part availability. We provide BrickAgent, an environment for coding agents to construct, inspect, and validate their designs. We find that leading agents largely satisfy verifiable physical and semantic requirements, but fall short of human designs. We release our benchmark and environment at http://www.brickben.ch

Figures & tables

Appendix figures & tables9 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. BrickNet: Graph-Backed Generative Brick Assembly

    Apr 24, 2026Peter Kulits, Cordelia SchmidConstrained Generative Modeling3D Asset Generation

  2. WorkBenchMark: A LEGO-Based Assembly Benchmark with an Assembly-by-Disassembly Baseline for the Smart Manufacturing League

    Jun 2, 2026Wenbo Ma, Daniel Swoboda, Matteo Tschesche +1Robotic ManipulationRobot Task Planning

  3. Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

    Jun 3, 2026Jiateng Liu, Bingxuan Li, Zhenhailong Wang +8Visual Spatial ReasoningBenchmark Construction