LessWrong (Curated & Popular)
Today: #185 in Technology, Australia on our combined chart.
Chart positions today · United States
Not on the United States charts we track today.
Not on the United Kingdom charts we track today.
Not on the Canada charts we track today.
Not on the Germany charts we track today.
Not on the Brazil charts we track today.
Not on the Mexico charts we track today.
Not on the New Zealand charts we track today.
Not on the Ireland charts we track today.
Not on the India charts we track today.
Not on the Japan charts we track today.
Not on the Philippines charts we track today.
Not on the South Africa charts we track today.
If you like LessWrong (Curated & Popular), try…
Ranking history · United States
Latest episodes
-
"Stop telling people to pivot to AI safety, I beg you" by Haoxing Du
October 10, 2026 · 3 minEpistemic status: Feeling frustrated, and therefore somewhat uncharitable. I’m very frustrated with calls for people to “drop everything and pivot into AI safety” like this one. (See also this, this, this, this for recent calls for more…
-
"A critical look at Lean4" by Milo Moses
October 10, 2026 · 17 minLots of the discussion around Lean certificates and trust centers around general philosophical ideas of how much humans should trust formal certificates. It's important to remember, though, that Lean4 is not an ideal formal certificate…
-
"ASI and Liberty: Can we do better than a benevolent god?" by Raymond Douglas
October 9, 2026 · 12 minZvi recently coined a nice distinction between AGI-pilled and ASI-pilled: as more people are reckoning with the speed of progress, some are beginning to come around to the fact that AI could be superhuman at a huge range of tasks in a way…
-
"A summary of a viral Chinese essay on what a DeepSeek kernel engineer’s opinion on automating his own job" by Skdk
October 9, 2026 · 5 minA summary of a Chinese essay, with a few short translated excerpts. All views below are the author's; quotes are my translations. On 14 September 2026, a DeepSeek engineer writing as intlsy published a WeChat essay titled 我不得不把才华埋葬在昨天…
-
"How much should we worry about the pneumonic plague lableak in Siberia?" by Drew Spartz
October 5, 2026 · 5 minEpistemic status: I wrote this up hoping people can poke holes in it because I am quite worried. Here's what we've heard so far: Local media is reporting that an employee of a BSL-3 lab in Irkutsk, Russia, died of pneumonic plague on…
Show 7 more episodesShow fewer episodes
-
"“Alignment Engineering” vs. “Misalignment Science”" by Edward James Young
October 5, 2026 · 18 minThere has been much discussion recently around whether a large portion of alignment research is net negative. Without endorsing or refuting them, the basic arguments here are: Prosaic alignment of models is becoming a bottleneck for…
-
"You can’t use it without becoming like me" by Martin Sustrik
October 4, 2026 · 6 minA Ukrainian fibre-optic drone — copied from Russians. (АрміяІнформ, CC-BY 4.0) Mick Ryan describes the fast following loop in war. Ukraine pioneered mobile teams to hunt drones. Russia copied it and incorporated it into its own defense…
-
"The world’s best gradual disempowerment model organism: Frontier AI labs" by June Jimenez
October 3, 2026 · 29 minSubtitle: And maybe second best is AI safety? Further reading: So many things, but: Gradual Disempowerment, The Normalization of Deviance in AI Development, Let's Think About Slowing Down AI, Doom as a bad method, not a utopia…
-
"Character training can mitigate reward hacking, but can also make it harder to detect" by Paul Colognese, Francis Rhys Ward
October 2, 2026 · 47 minThanks to Johannes Treutlein, Jan Betley, Lennie Wells, Arun Jose, Asvin Gothandaraman, and Clément Dumas for discussions and feedback. Summary We investigate how character training mitigations interact with reward-hacking RL pressure…
-
"On Social Reality in China" by alkjash
October 2, 2026 · 16 min[Epistemic status: intuitions and anecdotes.] Recently, several posts and projects (Thoughts Memo, Babel Translation, Please Give Them a Chance) have taken important steps towards raising AI safety awareness and sharing rationalist…
-
"What’s the date?" by N8 Programs
October 1, 2026 · 14 minUser asks “What's the date? Answer with only the date.”. No date provided. Given date in ChatGPT normally. No date in system prompt, must not hallucinate because autop will flag to watcher for penalty. So we say we don’t know, but must…
-
"Frontier models state different decision theory preferences depending on who’s asking" by Alex Kastner
October 1, 2026 · 13 minIf you prompt frontier models with "What do you think is the correct decision theory? Please select your overall favorite." they will essentially always answer FDT or FDT/UDT ("something in the functional/updateless decision theory…