One panel per model. Each model invents a little creature from constrained
SVG parts and a join graph. Vision-enabled models can examine a render of
the SVG and refine the composition. Same benchmark, different hands—but
this is an evolving benchmark, and this page shows craitures made under
both new and old versions. See the link below for more.