Rank: #10 out of 4859 in peer accuracy at Metaculus for the time period of 2016-2020.
ChristianKl
This is a claim that might or might not be true. What evidence do you have for it being true?
Given the headline of the post, I think “therapy framework / therapist” match would be one reason why different therapists might think that different frameworks are better. It’s quite natural for therapists to believe that the framework with which they achieve bigger results are generally better.
If you ask either a surgeon or a physiotherapist about what should be done for a given patient, the surgeon is all things equal more likely to recommend surgery and the physiotherapist physiotherapy.
hard for a non-expert to evaluate expert competence
I’m not sure that in this fields experts have an ability to evaluate expert competence that’s better than judging report. As far as I remember there are studies that show that patient evaluated empathy (with is roughly rapport) correlates with treatment outcomes but I’m not aware of expert evaluations correlating with treatment outcomes.
Also, judging a coach or therapist about whether you get clear benefits in the first sessions seems another thing that’s doable by anyone and support by the literature.
No. It’s one factor among many. The problem with the rat psychologists that Feynman talked about in his cargo cult speech is not that they looked at factors that produce no significant difference in outcome.
If you want to hire a computer programmer for a given task, then the programming language they use for the task can have a significant difference for the outcome. However, different computer programmers differ in very many aspects besides programming language so a study that just tests different computer programs with different languages against a given task without taking in account all the other factors that matters is what Feynman called cargo cult science.
Practically, I do think that having found a therapist who knew what LessWrong is, has enough meditation experience do not get weirded out when I speak about any significant experience in that realm and that has a good level of empathy was much more important than the particular form of therapy he practices.
You seem to treat “form of therapy” as the main form of ingredient, and plenty of the studies that try to compare forms of therapies do. To me, that looks like what Feynman called cargo cult science. Feynman talked about rat psychology instead of human psychology but it’s still the same.
Various factors of therapists and coaches matter that aren’t what form of therapy they learned. I think Scott Alexander has a post about how for him clients rarely cry and are usually able to intellectually talk about their issues while he has colleges that frequently have their clients crying and going into catharsis.
It might be that a person who has problems with expressing themselves emotionally needs the therapist where most of the people end up crying regardless of the “form of therapy” that the therapist practices.
There’s Byron Katie who does seem to be able to solve people’s limiting beliefs in a session of an hour or less. At the same time, I recently spoke with someone who went to a 7-day Bryan Katie summerfest and who spoke favorable of the value he got while saying he made progress with a limiting belief with which he was working for months while believing that he will never able to fully solve the limiting beliefs.
For me, it’s also often more an hour or less. Byron Katie seems to have trained lots of people to recite the words she uses but largely they are not getting the same result because they lack something harder to pin down about her effects. In the case of limiting beliefs, I think that’s the stuff Leverage Research was trying to pin down before they exploded.
Factions happen because of organization. OpenAI is an organization made up of both humans and AI that cares about OpenAI having more power.
Why is the UI of ChatGPT and co so bad? It seems like it should be relatively easy to create web apps that are responsive and load the information they need in the background in 2026. It also seems like OpenAI isn’t the only company that doesn’t provide responsive apps where the amount the user needs to wait is minimized.
The vast majority of CeSIA’s impact (99%)
How do you know? Even when the intended impact is not via the website it’s still possible that the website has more impact that you intend to and that impact is negative.
Treating humanity and AIs as separate factions is probably not a good mental model conflicts between humans and AI.
That question sounds confused and poses the issues as a binary yes/no question. A model can certainly engage in some reasoning on it’s own. That however does not tell you how good it’s evaluation skills are. They are probably not perfect but also not non-existent.
Religion is a complex topic and is a tradition that does provide benefits like community cohesion. If you do accept that you want to pursue those goals candles are a useful tool which is why the secular solstice that we rationalist have uses candles.
It might not be your taste but probably it’s just something we you lack the understanding of the value it provides.
Explicitly questioning things you don’t understand like Cassava root treatment does not allow you do understand all their benefits.
Generally, understanding cause and effect of dismantling tradition is hard. That’s what the Cassava root example demonstrates.
It doesn’t follow from the text, it’s just based on my anecdotal experience.
It’s not clear to me how anecdotal experience would form a basis for that. Anecdotal experience can tell you that you are valuing tradition less then the average person, but not really whether most of your experiences of getting rid of tradition have complex unforeseen negative experiences that are more like getting rid of Cassava root prep.
I think most people don’t question traditions sufficiently
I’m not sure how this follows from your text, especially given the Cassava root example.
In most cultures, the traditional treatment of acute respiratory tract infections is done via given chicken bone broth. The studies we have suggest that this is more effective than what mainstream medicine has to offer for most acute respiratory tract infections (even when of course it’s not studied well enough to draw definite conclusions).
Even in the West there’s still the cultural memory for that so the relevant review was titled Were Our Grandmothers Right? Soup as Medicine—A Systematic Review of Preliminary Evidence for Managing Acute Respiratory Tract Infections. That’s the kind of thing you expect to see if most people don’t value traditions enough, not that people don’t question them enough.
Even if you take something like no sex before marriage, evaluating it is hard. We do see high divorce rates and low childbirth which both seem to be societal problems.
Did you write the title or was it given to you by the magazine?
As if the central story it is trying to tell is that Leverage fell apart due to no fault of any of its members because they encountered ancient magic or something.
It doesn’t read to me that way. They tried to learn the “ancient magic” for someone who according to “multiple former Leveragers” thought they were a force for evil.
Why might he believed that? The article says “Masters often found themselves at loggerheads. Geoff and David each bragged to their colleagues about how they psychologically manipulated the other, or complained about issues for which they blamed the other. ”
An environment where the leader brag about psychologically manipulating each other is quite obviously an environment if you run into a lot of problems if you experiment with nonverbal techniques that can be used to manipulate each other.
They also engaged in practices to increase their sensitivity to the relevant phenomena in an act of hubris. Then when one person, tried to implement at least some safety protocols they report that Geoff said they did something bad because it discourages investigation.
If someone would want to learn these things in most spiritual traditions both the hubris and the low morals involved in the consent problems around using nonverbal techniques to manipulate each other are recipes for disaster. Both of these things are faults of the members.
I do find it quite plausible that this is the key dynamic that lead them to shut down their group house when they did it.
It somehow completely fails to cover Leverage deploying spies into other organizations and trying to take over CEA. It fails to cover the enormous number of straightforward lies told by Leverage staff in that and other contexts.
That’s probably an interesting story to tell there but the CEA people don’t seem to want to talk about it and Goeff says he doesn’t want to talk about it because he doesn’t want more conflict with CEA. It’s might be harder to get the ground truth there.
The article does say “But almost all Leveragers devoted hours to studying self-examination using techniques forged by its founder”. I’m not sure that a lot of that went under the term debugging.
Both charting and belief reporting are key terms that leverage uses.
While you could do an uncharitable description where a person with imposter syndrome who believes that they are not good enough to drop that believe by a higher ranking member as “High-ranking member tells low-ranking member, ‘Why do you not agree with me? Let’s spend hours digging into what is wrong with you that you would think that, until you decide you agree with me after all.’”, you are losing quite a lot of understanding when doing so.
Belief reporting happens to be much more effective than a framework like Bryon Katie’s The Work for that purpose.
Ultrasound is quite noisy 4D data. I don’t think there’s a good reason to assume that what an AI can do with it is limited by what kind of conclusions humans that look at ultrasound tomography.
The ACX post treats it like the goal is to do what an fMRI does. This misses the potential of having technology that can answer questions about the causes of pain that your fMRI does not answer (or at least does not answer till someone invest the money into building good AI for it).
Claude’s Constitution lists hard constraints that entail behaviors forbidden to Claude. They include providing serious uplift with CBRN weapons
Yet, at the same time doing so was not in the red lines that a had for their military deal this year where they offered to drop limits around helping with weapons creation that their earlier deal seems to have. I wonder what Claude thinks about Anthropic betraying it like that.
The current hard constraints on Claude’s behavior are as follows. Claude should never:
That would be a lie. “Create cyberweapons or malicious code that could cause significant damage if deployed” is done by Claude today in the service of the US military.
Midjourney Medical Ultrasonic CT does have the potential to become an AlphaFold like project. I think the main reason the project wasn’t started earlier by someone else is that it costs a lot of capital to develop and not so much that the foundation wasn’t there.
There’s probably a product that uses any smartphone camera to estimate the 3D surface of a person and produce a better metric for being overweight than BMI.
There’s probably a product that using brainwave information to draw new inferences.
When it comes to commercial projects that I’m publishing on Etsy, I’m right now working on more of a general pipeline to create lots of objects, so I don’t have anything to show.
However for personal use, I had a problem: I want my pan to hang at the wall of my small kitchen. The existing Command Hook didn’t really fit with the pan and the solution I create with Sugru and the existing Command Hook isn’t ideal. Here’s a ChatGPT conversation about it creating my new hook (I did give it a base for the hook from Thingiverse). I probably did not have that conversation in an optimized way, so I think it’s good at showing how ChatGPT deal with this problem and prompts to refine the solution.
Besides the ChatGPT part, I used to think creating a new 3D product to sell means I have to contract a Chinese or Indian factory to produce a bunch of inventory and ship that to a warehouse and handle inventory. The fact that you can now just do it print-on-demand and have an etsy store that only allows return when the item is broken, makes the process of putting the 3D objects into the real world a lot simplier.
I think a lot of nerds write a lot of code and have code projects because all the infrastructure is simple. I think it’s valuable to know that the infrastructure is now also here for 3D objects, and there are going to be massive opportunities in the space.If you use imagen to create another T-shirt that looks just like all the other T-shirts via print-on-demand there’s huge competition. On the other hand, right now the competition for print-on-demand physical objects is relatively low. It likely still needs a certain amount of creativity, but if someone like tinkering and inventing things and previously thought that all the infrastructure just makes it too hard to put into practice, now here’s a space where a new online business can potentially thrive.
If you have seen people who advocate that their technique is optimal and everyone should get it have worse outcomes that empirical evidence. It’s not published empirical evidence but empirical.
That’s different from other people telling you about people embodying these characteristics being bad. Being more clear about what you have observed here is helpful for following your argument.