Your github link is broken at the top of the page. Goes to 404 not found
Steven
This reminds me of METR’s study on the effects of AI on software engineer productivity. I bet there’s a small window where you could convince people to be in the control group and never interact with AI, so you could do that right now, but a few years from now, I’m not so sure.
There are still disanalogies like software engineers being comparatively weirder and more niche than ‘people who vote’ and that LLMs are used for productivity, not conversations. On the other side, having an LLM delete your production database or cause something catastrophic seems (I don’t have data on this) to happen way more often than catastrophically bad chatbot conversations.
Another issue is that how are you going to stop the AI from manipulating you a few months / years from now when they are everywhere, even better at manipulation than they are today? I’m mostly unimpressed by AI persuasion today, but I doubt that will hold
Maybe this is insensitive but have you considered getting stronger? I have no idea how much you work out or how much your housemates worked out, but ranting about the tyranny of grip strength seems like it won’t work as well as increasing your grip strength. Since you said that you’re sure you’re low in female grip strength distribution, getting stronger would actually be pretty easy
Each transcript was scored on 38 dimensions, then reviewed independently by Claude to confirm whether flagged transcripts indicated genuine violations
This seems like an odd choice to me, could you share the prompt for conversations checking for violations? I think it’s worth making sure that Claude doesn’t have a non-neutral understanding of its own constitution where other models might disagree
We surely now:
Small typo
Wouldn’t there be even cheaper ways to satisfy preferences about living humans? A fake, cheap version which satisfies that preference would probably be possible in the same way that a preference for a pet can be satisfied by a plush toy. Wanting humans or uploads but not being able to satisfy that desire with something fake seems like it isn’t how many of our actual desires work
You can keep talking more. You can repeat the proper analysis for your vaccine, talk about your own behavior, talk about why other people analyzing your behavior is either good or bad. You don’t have to concede the public square to someone else because you’re concerned they will misinterpret things and in fact these examples seem like situations where you can and should talk your way out of them
Isn’t the theory that consultants add value by saying true obvious things? If you realize you’re surrounded by sycophants, you might need someone who you’re sure won’t just tell you that you’re amazing (unless the consultant is also a yes man and dooms you even harder)
Thanks for writing this. I’m not sure I’d call your beliefs moderate, since they involve extracting useful labor from misaligned AI by making deals with them, sometimes for pieces of the observable universe or by verifying with future tech.
On the point of “talking to AI companies”, I think this would be a healthy part of any attempted change although I see that PauseAI and other orgs tend to talk to AI companies in a way that seems to try to make them feel bad by directly stating that what they are doing is wrong. Maybe the line here is “You make sure that what you say will still result in you getting invited to conferences” which is reasonable but I don’t think that talking to AI companies gets at the difference between you and other forms of activism.
I think you’re pretty severely mistaken about bullshit jobs. You said
At the start of this post we mentioned “bullshit jobs” as a major piece of evidence that standard “theory of the firm” models of organization size don’t really seem to capture reality. What does the dominance-status model have to say about bullshit jobs?
But there are many counter examples of this not being a real concept. See here for many of them: https://www.thediff.co/archive/bullshit-jobs-is-a-terrible-curiosity-killing-concept/
How would a military which is increasingly run by AI factor into these scenarios? It seems most similar to organizational safety a la google building software with SWEs but the disanalogy might be that the AI is explicitly supposed to take over some part of the world and maybe it interpreted a command incorrectly. Or does this article only consider the AI taking over because it wanted to take over?
Huh, did you experience any side effects?
I think discernment is not essential to entertainment. If people really want to learn what a slightly off piano sounds like and also pay for expert piano tuning, then that’s fine, but I don’t think people should be looked down upon for not having that level of discernment.
How would the agent represent non-coherent others? Like humans don’t have entirely coherent goals and in cases where the agent learns that it may satisfy one or another goal, how would it select which goal to choose? Take a human attempting to lose weight, with goals to eat to satisfaction and to not eat. Would the agent give the human food or withhold it?
One thing I find weird is that most of these objects of payment are correlated. The best paying jobs also have the best peers also have the most autonomy also have the most fun. Low paid jobs were mostly drudgery along all axes in my experience
Thanks for the summary. Why should this be true?
The fact that sympathy for hedonic utilitarianism is strongly correlated with intelligence is a somewhat worrying datapoint in favor of the plausibility of squiggle-maximizers.
Embracing positive sensory experience due to higher human levels of intelligence implies a linearity that I don’t think is true among other animals. Are chimps more hedonic utilitarian than ants than bacteria? Human intelligence is too narrow for this to be evidence of what something much smarter would do
Thank you for writing this. My girlfriend and I would like kids, but I generally try not to bring AI up around her. She got very anxious while listening to an 80k hours podcast on AI and it seemed generally bad for her. I don’t think any of my work will end up making an impact on AI, so I think basically the CS Lewis quote applies. Even if you know the game you’re playing is likely to end, there isn’t anything to do since there are no valid moves if the new game actually starts.
I did want to ask, how did you think about putting your children in school? Did you send them to a public school?
What does impossible mean in the context of clock neurons?
impossible in the first few moves.
What causes them to be unable to fire?
Q. What is generalization really for? What does it offer you?
Based on the vibe of the post, it seems like you’re trying to point at the concept of “being able to do many things”. I guess generalization isn’t ‘for’ anything, it’s a concept. For an agent, generalization is a method of being able to achieve an outcome based on limited past experience without needing to waste resources figuring out strategies it could have made if only it could generalize better. I can’t really tell based on what you said what I’m supposed to answer with “What does it offer you?”. Like, generalization offers me the ability to recognize bad chess moves in new scenarios that I haven’t seen, or it offers me the ability to take over the universe based on limited knowledge of physics. I don’t know where you’re trying to limit the word
I think two things are going on. First, accountability laundering. It’s acceptable for companies to say that AI gets stuff wrong sometimes and for there to be insufficient communication between business units to figure out what would be required for the AI to get this right. In a worse case, this happens on purpose because if one exec asks for AI to be implemented to cut costs, they mostly expect the issues with unprofessionalism to be outweighed by the few cents your call cost them. It also sounds like they don’t even pay the cost of their bad service. You and your doctor had to deal with it so from their perspective, nothing actually went wrong unless they connected the second follow up call with the bad initial service.
Second is a lack of norms. Think about how weird it would be if your cashier said something like “I’m going to accurately bag your groceries instead of flinging them at you”. During AI demos, I doubt the CVS manager was actually shown how hard implementing an AI system is, and if you got to speak to a real time AI model, you might assume that it would be resourceful the way a well spoken human would be. I’m speculating, but I think it’s really hard to tell how good an AI agent is in the best of circumstances and most businesses implement AI without a really good idea of how successful it will be in their specific context. They probably know really well how to handle human interviews and onboarding and could tell if a human couldn’t do their job but there isn’t any script anywhere to see how well an AI agent could do your particular job.
Anyway, I expect this to make the economy very jagged as some businesses implement AI well and it makes them much more productive while others set their cultural capital on fire and can’t figure out how to do it (assuming nothing even weirder happens due to AI)