Explorerโ€บArtificial Intelligenceโ€บAI
Research PaperResearchia:202610.09003

BrickBench: Evaluating Agentic Brick Design

Peter Kulits

Abstract

We propose BrickBench, a benchmark for agentic text-conditioned LEGO-set design. Given a prompt, an agent is tasked with producing an assembly that not only satisfies semantic and design criteria, but that can also be physically built. To do so, it must select parts from a discrete library and reason jointly about local and global constraints. We score validity, alignment, and design across three settings that vary in scale and part availability. We provide BrickAgent, an environment for coding ...

Submitted: October 9, 2026Subjects: AI; Artificial Intelligence

Description / Details

We propose BrickBench, a benchmark for agentic text-conditioned LEGO-set design. Given a prompt, an agent is tasked with producing an assembly that not only satisfies semantic and design criteria, but that can also be physically built. To do so, it must select parts from a discrete library and reason jointly about local and global constraints. We score validity, alignment, and design across three settings that vary in scale and part availability. We provide BrickAgent, an environment for coding agents to construct, inspect, and validate their designs. We find that leading agents largely satisfy verifiable physical and semantic requirements, but fall short of human designs. We release our benchmark and environment at http://www.brickben.ch


Source: arXiv:2610.12452v1 - http://arxiv.org/abs/2610.12452v1 PDF: https://arxiv.org/pdf/2610.12452v1 Original Link: http://arxiv.org/abs/2610.12452v1

Please sign in to join the discussion.

No comments yet. Be the first to share your thoughts!

Access Paper
View Source PDF
Submission Info
Date:
Oct 9, 2026
Topic:
Artificial Intelligence
Area:
AI
Comments:
0
Bookmark