Read the original at HF Daily Papers
Researchers introduce LMBuild, a benchmark that evaluates large language model agents on generating buildable and functional structures by assessing structural soundness, functional affordance, design quality, and physical realization across thirty systems.
Carried by: HF Daily Papers. First seen: .