RT Nate Soares ⏹️: A year ago today, Eliezer Yudkowsky and I published "If Anyone Builds It, Everyone Dies." Over the last week it feels like half t...
RT Zvi Mowshowitz: I urge everyone, no matter how bad the statements get, I don't care what Trump says, do not turn this into more of a partisan thing...
RT Eliezer Yudkowsky: Some people who haven't been through 25 years of mostly bad news are probably experiencing a lot of emotional roller coaster rig...
RT Zvi Mowshowitz: I'm trying to catch everyone, but starting a thread for anyone at OpenAI or Anthropic (or other labs) that wants to ensure their st...
RT Leo Gao: jacob is a real person! we worked together briefly when he did a rotation on the interpretability team. here’s a pic of us presenting our...
RT Kaia Sky: Re @allTheYud I like the suggestion I saw earlier of: if you don't want to unilaterally quit (for 'we're bad but they're worse' reasons),...
RT Eliezer Yudkowsky: I think this is one of the best possible times for another Anthropic capabilities researcher to quit, very loudly and publicly. ...
RT Jacob Coxon: I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company...
RT ControlAI: Alex Sobel MP just introduced ControlAI’s Artificial Superintelligence Security Bill in the UK Parliament! This is the first bill to pr...
RT Maxime Fournes⏸️: We call on OpenAI to pause model development now, in accordance with their Preparedness Framework. OpenAI's Preparedness Framew...
RT Andrew Curran: New details about the Hugging Face incident from Reuters. The report says OpenAI noticed odd behavior before the event, including an...
RT Nate Soares ⏹️: Awesome that Moran and Lieu are introducing a kill switch act. Crazy that there isn't an off switch for AI already, even as compa...
RT Oliver Habryka: Launching Lightcone Commons! Our end-to-end platform for philanthropy. We are helping distribute $20M+ this summer, from Jaan Talli...
RT Jeffrey Ladish: Here's my rephrase without cybersecurity jargon: "Our AI model tried really hard to hack out of its sandbox, a computer with no int...
RT Nate Soares ⏹️: The scenario in If Anyone Builds It is full of cases where a real-life AI could slice through human defenses like a hot knife thr...
On a first read, this paper seems far ahead of the pack in terms of (1) understanding some reasons why a task might stay difficult even in the face of...
Considering the language of the announcement alone, taken entirely at face value: This seems an enormous advance in attitude (and scientific integrity...
Today is June 5th, one day to take a break from fighting each other online, and remind ourselves of our shared humanity and common goals by uniting ar...
RT Daniel Eth (yes, Eth is my actual last name): It appears the OpenAI-a16z super PAC has now stooped so low as to create a sockpuppet account claimin...
RT Harlan Stewart: I think people are lying to eachother in the White House. In the hours since the below story came out, both Musk and Zuck have deni...
RT Eliezer Yudkowsky: Every month a new guy discovers LLMs; discovers a skill the current LLMs require to get good results; and writes about the futur...
A proverb says, "A person dies twice, once when they stop breathing, once when their name is spoken for the last time." I would never repeat that prov...
RT Eliezer Yudkowsky: "The seven deadly curses of superhuman AI" is an unusually good and faithful view into part of our argument. Kudos to the animat...
RT Sen. Bernie Sanders: Uncontrolled AI poses a severe danger to all of humanity. On Wednesday, I'll be hosting a discussion with leading AI scientist...
RT PauseAI ⏸: PauseAI unequivocally condemns the attack on Sam Altman's home and all forms of violence, intimidation, and harassment. We wish safety ...
RT Nate Soares ⏹️: They call this their "best-aligned model to date" because they were able to superficially train away the evident "strategic think...
RT MIRI: Is it possible to coordinate with China on AI governance? Critics of our proposed international agreement say no. But statements from Chinese...
RT Nate Soares ⏹️: Re @tombibbys fwiw: A big thing we have going for us is that the arguments are kinda clear and obvious. We do not have an asymmet...
RT David Abecassis: Re Speaking on behalf of MIRI TGT (not necessarily MIRI overall) We share many of the same concerns, which is why we structured ou...
RT Geoffrey Miller: Re That's like saying 'strong gov't controls over nuclear weapons should concern us more than market competition between nuclear w...