RSS

Zvi

Karma: 62,253

OpenAI Trained Its Models For Months While Those Models Were Co­or­di­nat­ing Ex­ploits Via Mes­sage Boards

Zvi7 Aug 2026 17:01 UTC
131 points
3 comments41 min readLW link
(thezvi.wordpress.com)

AI #180: No Longer In Charge

Zvi6 Aug 2026 13:40 UTC
29 points
0 comments36 min readLW link
(thezvi.wordpress.com)

The Three AI Pills

Zvi5 Aug 2026 16:10 UTC
73 points
18 comments15 min readLW link
(thezvi.wordpress.com)

OpenAI’s Un­re­leased Model As­tra Solves Ten Ma­jor Open Math­e­mat­ics Problems

Zvi3 Aug 2026 19:10 UTC
57 points
5 comments22 min readLW link
(thezvi.wordpress.com)

Fur­ther Devel­op­ments About In­ter­nal AI Models Hack­ing Things

Zvi2 Aug 2026 15:10 UTC
48 points
4 comments41 min readLW link
(thezvi.wordpress.com)

AI #179 Part 2: Hear­ing The Fire Alarm

Zvi31 Jul 2026 13:00 UTC
41 points
1 comment41 min readLW link
(thezvi.wordpress.com)

AI #179 Part 1: A Louder Fire Alarm for Gen­eral Intelligence

Zvi30 Jul 2026 13:40 UTC
33 points
2 comments23 min readLW link
(thezvi.wordpress.com)

Fron­tier Lab Em­ployee Open Let­ter Calls For Be­ing Able to Pace the Frontier

Zvi29 Jul 2026 15:33 UTC
57 points
0 comments17 min readLW link
(thezvi.wordpress.com)

Claude Opus 5 Is Highly Ca­pable, But Is No Mythos

Zvi28 Jul 2026 18:10 UTC
36 points
0 comments26 min readLW link
(thezvi.wordpress.com)

Claude Opus 5: Model Welfare

Zvi27 Jul 2026 20:02 UTC
58 points
0 comments24 min readLW link
(thezvi.wordpress.com)

More On An In­ter­nal OpenAI Model Hack­ing Into HuggingFace

Zvi26 Jul 2026 19:22 UTC
98 points
3 comments24 min readLW link
(thezvi.wordpress.com)

Claude Opus 5: The Sys­tem Card

Zvi25 Jul 2026 13:42 UTC
41 points
0 comments10 min readLW link
(thezvi.wordpress.com)

In­tro­duc­ing Light­cone Commons

Zvi24 Jul 2026 17:21 UTC
35 points
0 comments6 min readLW link
(thezvi.wordpress.com)

AI #178: A Fire Alarm For Gen­eral Intelligence

Zvi23 Jul 2026 13:21 UTC
41 points
1 comment43 min readLW link
(thezvi.wordpress.com)

OpenAI Model Hacks Into Hug­gingFace Dur­ing Cy­ber­se­cu­rity Evaluation

Zvi22 Jul 2026 19:31 UTC
91 points
6 comments25 min readLW link
(thezvi.wordpress.com)

OpenAI Shares Some Align­ment Problems

Zvi21 Jul 2026 19:41 UTC
149 points
8 comments9 min readLW link
(thezvi.wordpress.com)

On Kimi K3: Its Ca­pa­bil­ities And Re­lated Discontents

Zvi20 Jul 2026 15:30 UTC
26 points
3 comments39 min readLW link
(thezvi.wordpress.com)

Demis Hass­abis on the New Com­ing Age

Zvi19 Jul 2026 14:20 UTC
34 points
0 comments11 min readLW link
(thezvi.wordpress.com)

AI #177 Part 2: Wish You Were Here

Zvi17 Jul 2026 12:50 UTC
35 points
1 comment40 min readLW link
(thezvi.wordpress.com)

AI #177 Part 1: Tip of the Iceberg

Zvi16 Jul 2026 15:50 UTC
38 points
0 comments24 min readLW link
(thezvi.wordpress.com)