LessWrong (Curated & Popular)
Общая длительность:
7 h 57 min
"Common mistakes in AI safety group organizing" by Nikola Jurkovic
LessWrong (Curated & Popular)
03:40
"Please Give Them a Chance: On China, Rationalism, and AI Safety" by gzjw
LessWrong (Curated & Popular)
08:32
"Why I Stay Off Twitter" by jefftk
LessWrong (Curated & Popular)
03:41
"The Game is Set for a Targeted Memetic Attack on the AI Safety Community" by keltan
LessWrong (Curated & Popular)
02:24
"For Love of the Lightcone, Don’t Partisanize AI Safety" by DanB
LessWrong (Curated & Popular)
26:27
"AI as orderly evacuation vs stampede" by Richard_Ngo
LessWrong (Curated & Popular)
09:04
"Cooperation with AIs seems to be a low-hanging fruit for better evals" by Clément Dumas
LessWrong (Curated & Popular)
15:28
"Current alignment training might be ineffective (and actively bad) in the age of RL" by Daniel Tan
LessWrong (Curated & Popular)
13:23
"If Anyone Builds It, Everyone Dies: One Year Closer" by Eliezer Yudkowsky, So8res, Duncan Sabien (Inactive)
LessWrong (Curated & Popular)
15:16
"Quick notes from teaching technical profiles how to talk in public" by Camille B.
LessWrong (Curated & Popular)
11:56
"Op-Ed: I Worked at Google DeepMind. You Should Listen to the Warnings About AI" by TurnTrout
LessWrong (Curated & Popular)
06:05
"There is a channel to 900M weekly users. What goes in it?" by Charbel-Raphaël
LessWrong (Curated & Popular)
05:57
"I am refusing to work on Cloud TPUs" by Yair Halberstadt
LessWrong (Curated & Popular)
04:03
"Can a superintelligence do THAT?" by Eliezer Yudkowsky
LessWrong (Curated & Popular)
24:38
"The Talker Does Not Control The Doer (in Current AIs)" by Eliezer Yudkowsky
LessWrong (Curated & Popular)
21:50
"Some ways AI could kill us all" by Ruby
LessWrong (Curated & Popular)
17:56
[Linkpost] "Doom as a bad method not a utopia trade-off" by KatjaGrace
LessWrong (Curated & Popular)
03:15
"Astra is much better at reasoning with filler tokens than previous models" by Dylan Xu, SebastianP, Alek Westover
LessWrong (Curated & Popular)
15:34
"Self Hosting" by Tomás B.
LessWrong (Curated & Popular)
04:02
"The Locally Optimal Discursive Posture" by deanball
LessWrong (Curated & Popular)
34:12
"Proposal for tracking the effects of architecture on monitorability" by ryan_greenblatt, Alek Westover, Lukas Finnveden
LessWrong (Curated & Popular)
11:16
"Explaining Knightianism on one foot" by Richard_Ngo
LessWrong (Curated & Popular)
19:17
"Astra can do a concerning amount with no chain of thought" by Neel Nanda
LessWrong (Curated & Popular)
21:20
"Personal statement on joining the OpenAI board" by paulfchristiano
LessWrong (Curated & Popular)
04:13
"How good are slop-vestigators?" by Hasan Baig, OscarGilg, Hamzah
LessWrong (Curated & Popular)
13:44
"The Scramble: getting in position to pace the frontier" by Peter Wildeford
LessWrong (Curated & Popular)
18:44
[Linkpost] "Frontier models still hack on simple variations of alignment evals from early 2025" by Dean Valentine
LessWrong (Curated & Popular)
03:34
"Dear God, Please Don’t Resign In Protest" by Kabir Kumar
LessWrong (Curated & Popular)
03:11
"Let’s talk about the AI coordination problem" by KatjaGrace
LessWrong (Curated & Popular)
03:25
"Drone WMDs Don’t Need Any New Technology" by Felix Choussat
LessWrong (Curated & Popular)
24:33
"Evaluation" by Nina Panickssery
LessWrong (Curated & Popular)
02:47
"Let’s fund weird AI safety projects" by Ihor Kendiukhov
LessWrong (Curated & Popular)
07:12
"Steering towards “automated grading” degrades alignment" by Jan Betley, Johannes Treutlein, Clément Dumas
LessWrong (Curated & Popular)
23:56
[Linkpost] "Discovery Of A New OpenAI Agent Message Board" by Capybasilisk
LessWrong (Curated & Popular)
02:20
"Cat-Belling Problems" by Eliezer Yudkowsky
LessWrong (Curated & Popular)
36:29
"How concerned should we be about OpenAI’s recurrent architecture rumors?" by Rauno Arike
LessWrong (Curated & Popular)
18:51
[Linkpost] "Sen. Bernie Sanders (I-VT) and Rep. Greg Casar (D-TX) introduce legislation to ban Artificial Superintelligence and temporarily pause advanced AI development" by Matrice Jacobine
LessWrong (Curated & Popular)
03:47
[Linkpost] "Resolution has a new Agent Foundations team" by Jeremy Gillen
LessWrong (Curated & Popular)
03:55
[Linkpost] "Training a Misaligned Reward Seeker" by evhub, Monte M, Benjamin Wright
LessWrong (Curated & Popular)
05:59
"PauseAI Has ‘officially disendorsed’ PauseAI-US" by nem
LessWrong (Curated & Popular)
01:25