I will say that Fable + Codex & on-demand VMs honestly feels superhuman, lacking only the “spark of genius” heuristic (heuristic set?). The problem search spaces are nevertheless very large, so brute-force simply isn’t an option without some clever perspective.
I think also that people underestimate the utility of partial results and we need a way to verify and document these to avoid wasted and duplicated work.
So far, the big breakthroughs are coming from strong professionals asking LLMs to prove truly important results.
While the amateurs do get empowered quite a bit, right now the situation rewards high competence.
I will say that Fable + Codex & on-demand VMs honestly feels superhuman, lacking only the “spark of genius” heuristic (heuristic set?). The problem search spaces are nevertheless very large, so brute-force simply isn’t an option without some clever perspective.
I think also that people underestimate the utility of partial results and we need a way to verify and document these to avoid wasted and duplicated work.