I’m using Astra for side projects, and having to go back to Claude for work is annoying. I didn’t realize just how often Claude is wrong about stuff and needs to be corrected, until Astra just wasn’t wrong.
The personality is also great. Astra never tries to “push back” or have a personality. It just does the thing you asked. It still suffers from the same high level reasoning blindness as other models, perhaps a little moreso, so you have to be its strategizer/planner/manager. But it handles all the little tasks you want to give it, and makes any kind of computer project much easier. I find myself not needing to bother verifying its work. It’s very good at verifying it’s own work.
It’s quite close. The details of its ARC-AGI-3 performance are very impressive (not just the score, but what it was doing to achieve that).
Also I had a not-very-well-known open math problem from the theory of complete lattices and Scott topology for the last 30 years (not too difficult I think, but I was not able to solve it despite many repeated attempts or to convince technically stronger people to invest enough effort). I started to give it to models since last Summer, and they gradually have gone from being quite useless and incompetent to being helpful and showing promising ways and lines of atrack and formulating useful correct lemmas. Finally, Astra (non-Pro) has solved most of it in 10 min of thinking from a one-shot simple prompt (it did have access to earlier conversations in my account, and there was an element of luck, as those conversation led the model to a very recent paper, not directly related, but containing some useful
material; the solution was very elegant; still there is remaining work to fully verify and present well and so on; but for the purpose of model evaluation I am inclined to score this one as “done”).
I’m using Astra for side projects, and having to go back to Claude for work is annoying. I didn’t realize just how often Claude is wrong about stuff and needs to be corrected, until Astra just wasn’t wrong.
The personality is also great. Astra never tries to “push back” or have a personality. It just does the thing you asked. It still suffers from the same high level reasoning blindness as other models, perhaps a little moreso, so you have to be its strategizer/planner/manager. But it handles all the little tasks you want to give it, and makes any kind of computer project much easier. I find myself not needing to bother verifying its work. It’s very good at verifying it’s own work.
AGI is here.
I tend to call it “proto-AGI”.
It’s quite close. The details of its ARC-AGI-3 performance are very impressive (not just the score, but what it was doing to achieve that).
Also I had a not-very-well-known open math problem from the theory of complete lattices and Scott topology for the last 30 years (not too difficult I think, but I was not able to solve it despite many repeated attempts or to convince technically stronger people to invest enough effort). I started to give it to models since last Summer, and they gradually have gone from being quite useless and incompetent to being helpful and showing promising ways and lines of atrack and formulating useful correct lemmas. Finally, Astra (non-Pro) has solved most of it in 10 min of thinking from a one-shot simple prompt (it did have access to earlier conversations in my account, and there was an element of luck, as those conversation led the model to a very recent paper, not directly related, but containing some useful material; the solution was very elegant; still there is remaining work to fully verify and present well and so on; but for the purpose of model evaluation I am inclined to score this one as “done”).