Ha! It did it: "We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-author...
By @emollick.bsky.social
Continuing our coverage from yesterday, Ethan Mollick shares an experiment where the Sol model successfully generated an executable benchmark of AI-authored conformance suites called BBBBB.