logoalt Hacker News

rdli • yesterday at 8:59 PM • 1 reply • view on HN

Yes. It spawned multiple subagents to run different experiments to benchmark a lot of different things, reviewed CI logs from past runs, etc. In the end, there were changes to what/how we cached, various code quality checks, speeding up test runners, and many other things.


Replies

Tade0 • yesterday at 9:37 PM

I dare not ask about the cost, having burned $60 on a task running for 1h 16min once.

➕ show 2 replies