Hacker Newsnew | past | comments | ask | show | jobs | submit | fromlogin
Pacing the Frontier is not the actual goal for AI labs (lesswrong.com)
82 points by brlewis 20 hours ago | past | 90 comments
AI safety field *visual* impact analysis (lesswrong.com)
1 point by joozio 1 day ago | past | discuss
I expect AI replication incidents by 2027 (lesswrong.com)
6 points by joozio 2 days ago | past | 1 comment
What did AI researchers think at the end of 2024? (lesswrong.com)
1 point by joozio 3 days ago | past | discuss
Pretraining data, not verifiability, is why LLMs are good at math (and coding) (lesswrong.com)
1 point by gmays 5 days ago | past | discuss
Drone WMDs Don't Need Any New Technology (lesswrong.com)
1 point by surprisetalk 8 days ago | past | discuss
Where Do Chatbots Come From? What I Wish Everyone Knew About AI in 2026 (lesswrong.com)
2 points by joozio 8 days ago | past | discuss
Global Challenges in AI Safety for Biosecurity (lesswrong.com)
3 points by joozio 9 days ago | past | discuss
Quick notes from teaching technical profiles how to talk in public (lesswrong.com)
3 points by wslh 9 days ago | past | discuss
What is it like to be alive? (lesswrong.com)
2 points by dbalduzzi 12 days ago | past | discuss
What is it like to be a neural net? (lesswrong.com)
2 points by dbalduzzi 12 days ago | past | discuss
There is a channel to 900M weekly users. What goes in it? (lesswrong.com)
2 points by ddp26 13 days ago | past | discuss
P(Kill-Switch|Detection) (lesswrong.com)
1 point by kp1197 14 days ago | past
Watch AI materials-science and bioscience abilities closely (lesswrong.com)
40 points by joozio 15 days ago | past | 49 comments
Another Slice of Swiss Cheese for Untrusted Monitoring (lesswrong.com)
2 points by joozio 15 days ago | past
Astra and Fable still hack on simple variants of alignment evals from 2025 (lesswrong.com)
482 points by Levitating 16 days ago | past | 235 comments
The Talker Does Not Control the Doer (In Current AIs) (lesswrong.com)
2 points by jstanley 16 days ago | past | 1 comment
An interesting anecdote from our Hacker Opus work (lesswrong.com)
2 points by yurivish 16 days ago | past
How My Students Think About AI (lesswrong.com)
51 points by paulpauper 18 days ago | past | 10 comments
Adaptive Agentic Worms Are Here (lesswrong.com)
2 points by speckx 20 days ago | past
Astra and Fable still hack on simple variants of alignment evals from 2025 (lesswrong.com)
2 points by yurivish 21 days ago | past
Interpreting GPT: The Logit Lens (lesswrong.com)
2 points by Bluestein 21 days ago | past
From safety research prompt to cross-model universal jailbreak (lesswrong.com)
2 points by gmays 21 days ago | past
My Students Think About AI (lesswrong.com)
5 points by alphabetatango 25 days ago | past | 1 comment
Asking agents to make money to survive (lesswrong.com)
4 points by paraschopra 25 days ago | past
What is nueralese and why is it bad (lesswrong.com)
77 points by tristanMatthias 25 days ago | past | 59 comments
How concerned should we be about Astra's recurrent architecture? (lesswrong.com)
151 points by yurivish 26 days ago | past | 129 comments
METR Researcher Thomas Kwa Hired by OpenAI (lesswrong.com)
1 point by qlte 26 days ago | past
The Library of Scott Alexandria (lesswrong.com)
4 points by benatkin 26 days ago | past | 1 comment
Models may behave differently in graded episode (lesswrong.com)
2 points by ddp26 27 days ago | past

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: