logoalt Hacker News

vunderba • yesterday at 5:49 PM • 0 replies • view on HN

The idea of using an LLM to drive graphic output is pretty popular, so I definitely wouldn’t be surprised if there are already several benchmarks out there already.

I've seen a few voxel-based benchmarks built around the same idea, with LLMs effectively constructing models using a discrete set of instructions.

https://minebench.ai