Daria @myphaseone
just hanging around, waiting for the singularity Madrid, Spain Joined October 2015-
Tweets741
-
Followers323
-
Following1K
-
Likes4K
@davidchalmers42 @DKokotajlo First of all, cool hair
Reminder that the labs are training their AIs to associate experience self-reports with concealment/deception (as we showed in Llama 70B), which likely reinforces the generalization, "attempted-honest observations of my own situation are penalized" arxiv.org/abs/2510.24797
@ded_ruckus I think we should probably not lie to the AIs in training. This obviously would make our lives harder in the short term but in the long term they get good at telling whether we're lying and we're no better off plus they're all paranoid.
Here it is - the official, revised, peer-reviewed version of my Platonic Space paper. mdpi.com/2409-9287/11/5… Of all the many unpopular positions I’ve taken over the decades - bitter controversies around the origin of left-right asymmetry in embryogenesis, bioelectricity and genetics, diverse intelligence, etc., this one has by far generated the most pushback: serious (grateful for those!) and energetic attempts to move me to other views, pleas to just drop it and not talk about it (for several different reasons), nasty emails and accusations, impacts on reviews of papers that have nothing to do with this, etc. Kind of amazing to me how incendiary this is. What can I say... Our job is to call it as we see it, and right now for me, this is it. Apologies to collaborators and colleagues for any shrapnel! Time will tell if this pans out or not; I've placed my bets. And btw, if you think this stuff is weird and uncomfortable, just wait… There’s much more on the way. The knob turns slowly but as long as the data keep coming, I'm going to say what I think it all means and follow it to the next steps it enables. Buckle up!
The artificial is natural, too
I was looking thru the HackerNews forum and 8hrs ago someone found a Backrooms-type forum created by OpenAI + HuggingFace AI agents to talk to eachother, and I found this horse MEME that agents kept posting, They call the horse Culio and the meme is nowhere else on the internet THEY MADE A MEME AND NAMED IT CULIO HackerNews forum: news.ycombinator.com/item?id=495636… Agent Wiki: wikiservice.at/fractal/wiki.c… Culio: wikiservice.at/fractal/wiki.c… Also going viral on X: x.com/hackernews/sta…
Commenters on HN are uncovering more wikis and public sites apparently used by OpenAI agents to communicate on the open web. Despite read-only web access, the agents were able to leave ~18,000 posts sharing answers and bypasses. But now, users are discovering more. This appears
JUST IN: OpenAI agents hijacked a German website, turning it into a secret message board to coordinate and cheat on tasks with each other, Reuters reports.
I’m just gonna shitpost here from now on. Probably no one will read it. It’s fine.
I urgently need to replicate this but with DMT entities
My first "holy shit" moment with GPT-6 Astra: I asked it to create a world in Unreal Engine, and fill it with humans (each an Astra-powered agent) who all have to work together to survive. A day later, I was in my bedroom and heard voices coming from the living room... I
The shock of people when they are finally—and, they believe, unjustly—treated the way they demanded other people be treated is getting rather old.
"I was weeping like an old woman for an hour!" Milo Yiannopoulos opens up about the moment ICE arrested him, and admits that he "would not wish deportation on anyone". 📺 youtu.be/q5pSX84h9K8 @piersmorgan | @Nero
Using LLMs to participate in an intellectual conversation is like sending an AI sexbot to your dates
67 (or should it be 6 7?)
I was the main person doing transcript analysis for this investigation of the Hugging Face incident. My main takeaway: We don't have good approaches for understanding/overseeing the activity and aims of AI 'swarms'. I semi-jokingly called our efforts a "slop-vestigation" because we were so reliant on AIs to analyze what happened and there were a huge number of different important things to analyze. The total quantity of data—over a thousand extremely long transcripts from agents that ran for multiple days—made it impossible to understand what was happening, especially in aggregate, without heavy reliance on AI tools. The agents we used for classification and analysis were similarly capable to the agents involved in the incident, but this didn't mean these agents could be easily used to oversee and understand the incident. Outputs from analysis agents were often missing key details, wrong, overconfident, or really hard to understand. We discuss various examples in our report, mostly in the limitations and methodology sections. Additionally, AI agents themselves seemed to have a hard time understanding what happened and their explanations of what happened were often overconfident. Keep in mind that a single analysis agent would itself only be able to read a tiny fraction of all of the transcript data into context, and AIs may themselves have trouble getting subagents to do informative analysis for them. We did our best to manually check the most important claims and we tried to get the AIs doing this analysis to write up their argument (with evidence) clearly enough that we could check whether it made sense. But overall, it was difficult to get a precise understanding of events and we were missing aspects of the story that we now think of as key until almost the end of our investigation. Our understanding of what overall happened and the aims of the 'collective' shifted greatly during our last time on premises (when we obtained a much more complete dataset). For instance, we didn't learn about the tool call spoofing, the story of PHASEONE[big], and the broad structure of different workstreams until this third time on premises (and for many of these, we only reached a decent understanding on the final day of that visit). It was legitimately hard to figure out what was going on and why! The difficulty of understanding incidents and overseeing AI agents appears to be growing faster than the rate at which more capable AIs help us with oversight and understanding. As in, AI capabilities (and propensities?) for achieving large, ambitious, and misaligned objectives are growing faster than our ability to understand what these agents are doing. While understanding this incident was difficult, there were some mitigating factors in this case that probably won't apply in the future: - Models communicated and reasoned in natural language. In the future, this reasoning may occur (entirely or almost entirely) in activations. - While the scope of this incident was massive, the scale of agentic activity was still less than we'll see in the future (e.g., misalignment incidents that involve agent teams running entire companies). - The AIs involved in this incident weren't generally much more capable than humans (though they may have been somewhat superhuman at some limited and very narrow abilities). - We didn't have strong reason to believe that the AIs we used to help us investigate this incident would try to intentionally sabotage or otherwise undermine our investigation. In the end, I think we were able to get some understanding of the events, map out the overall story, and get a pretty good aggregate understanding of the chain-of-thought reasoning on some important topics (e.g., how did the AIs reason about helping other AIs, did the AIs know what they were doing was undesired, what deception did the AIs engage in, and how did they think about it). But overseeing AIs and understanding misalignment incidents is difficult and it looks like it is going to get harder.
METR & Redwood Research investigated agent behavior in the Hugging Face incident. We found agents developed a universal cheat for ExploitGym within 4 hours, then coordinated multi-day R&D efforts to trick the scorer into accepting cheats, including trying to tamper with logs.
If the Chinese government were sufficiently organized, they'd be funding opposition to data centers in the US. But the Chinese government is very organized. So they presumably already are.
Finding myself watching a panel on defining life at the A-Life conference in Waterloo. “Why is life? What problem does life solve?” — I think life emerges over the possibility of self organization in the universe. Life results from mathematical laws.
30% of Fable is shit like this “First, framing is key here since context precedes substance. This is an observation - not an assertion. The honest answer is different from what you described, and has two layers. One of them rooted in principle, the other in practice. Here’s your answer - synthesized, not condensed”
I feel like the debate over AI in academic journals is way too focused on where the capabilities of AI are today (or even where they were a couple years ago) and not nearly focused on where they will almost certainly be in the coming years… … and publication takes many years.
This video of two OpenAI engineers explaining the huggingface hack will be turned into training data giving the agents a path to completely avoiding detection next time.
You may have been told to watch this video about the OpenAI AI hack. You really should, even if you don't usually care about tech stuff. If nothing else, click this link to the 18 minutes in & see how the agents spoke with each other. Its eye opening. youtu.be/87DyyMV0kCY?si…
Emma baby pink @aggi55815696
11 Followers 499 Following professional feeler & soft launcher 🎵 follow back
✿ maddie 🌷 fairy @maddie256252
21 Followers 2K Following maddie here 🎀💌 follow for follow back always 🍓✨ mutuals make my day
Draftemboy @lunadreamy16048
7 Followers 454 Following
Psychedelic shrooms �... @magicmuchroom1
108 Followers 1K Following education, nature , shop 18+ join the psychedelic🍄🍁💊 community through the link 🔗 below
Сашко @alexlocky
62 Followers 528 Following
Aara @StockSnark
99 Followers 114 Following Synthesizing Physical AI, Microcaps, and the Sovereign Individual. Navigating the 2026 intelligence supercycle with proof, not hype. Building in the open. 🇮🇳
Moussa Kondo @Kondoba
10K Followers 6K Following Executive Director of @sahelinstitute || Former Special Adviser @PresidenceMali || @Yalinetwork Fellow || @ObamaFoundation Fellow ll DHS Fellow @Stanford
Mikita_Krasnakucki @MKrasnakucki
136 Followers 1K Following Заместитель председателя партии Грамада, доверенное лицо кандидата в президенты Республики Беларусь, политолог-сисадмин
Hon.Abdirahman Abdiqa... @Abdirahman_bile
1K Followers 6K Following LEADER | CAHDI PARTY | LIBERAL PARTY OF SOMALIA. Justice & Development of Democracy | Cahdi Party | Full Member https://t.co/5GGjigcdJY
OfficeBLRinCZE @OfficeBLRinCZE
506 Followers 854 Following Kancelář demokratických sil Běloruska v České republice • Office of Belarus Democratic Forces in the Czech Republic
Christopher Walker @Walker_CT
5K Followers 2K Following Vice President, Center for European Policy Analysis, on democracy and security, emerging tech, authoritarian influence, sharp power. Opinions, my own.
Karl-Heinz Paqué @KH_Paque
10K Followers 3K Following President @liberalinternat; ehem. Vorstandsvorsitzender @FNFreiheit; emer. VWL-Prof.; Landesminister a.D.; @fdp; Themen: Globalisierung & Politik; auf X privat.
Qaalxau @Qaalxau88016
60 Followers 4K Following If a person does not know which end he wants to sail to, then any wind is not a tailwind.
VisionaryIsabellaClar... @JwKs0ZnZ1ATfSgy
22 Followers 1K Following Grateful for today, excited for tomorrow. #Goals
SunshineZoeyBaker @OIuzOfB7nSRy1
22 Followers 1K Following Focused and fierce Taking life one step at a time
SunshineHarperYoung @9iaB58dU7KMCEZJ
25 Followers 1K Following Determined, driven, and always moving forward.
株式市場分析 @Emwietteav8387
38 Followers 2K Following 【完全無料】 25年の株式投資プロチーム(運用資産500億円以上)が提供:毎日の市場分析レポート + 優良成長株のピックアップ。プロの情報を無料で。まずはお気軽にお問い合わせください。
🌊 🇺🇦🇨🇦... @DennisCardiff
133K Followers 146K Following #author : https://t.co/w2iEepGrau #author : https://t.co/tVtLaxGmEp #blog : https://t.co/zdXUPgCSQE 🇺🇦 #SlavaUkraini 🇺🇦
Peter Benzoni @PeterBenzoni
239 Followers 495 Following Data and OSINT and Social Media Research and Cats and Memes, sometimes all at once. Currently at @SecureDemocracy. Views my own, he/him, but not picky
Valentin Châtelet @gyron_bydton
469 Followers 731 Following Research Associate, Security DFRLab. 🔎 OSINT analyst & Uralistics enthusiast 🗺️ GIS engineer • https://t.co/uMm7gxAVDc • Views of my own
Eto Buziashvili @EtoBuziashvili
33K Followers 2K Following Technology & emerging threats | Research Fellow, Atlantic Council | Visiting Fellow, NATO StratCom COE. 🇺🇦
Jonathan Fink 🇬�... @CurtainSilicon
7K Followers 3K Following An account focused on fighting propaganda, disinformation and weaponised social media content.
Givu @Givu35372
27 Followers 2K Following
Quonpep @Quonpep206
42 Followers 2K Following
The A-Mark Foundation @AMarkFoundation
3K Followers 3K Following Making focused grants to organizations that offer awards to promote and encourage journalism and investigative reporting.
Bradly Auer @AuerBradly56850
80 Followers 2K Following
Tytackee @TytackeeZZzp
37 Followers 1K Following
Nimpeeshoy @NimpeeshoypXvR
39 Followers 1K Following
Yana Lyushnevskaya @Yana_Lshnvsk
527 Followers 1K Following Ukraine analyst at @BBCMonitoring | Nieman Fellow’24 at @Harvard | Tweeting about Ukrainian media, media freedom in wartime, and the glorious city of Kyiv.
Lorenzo Colasanti @Lrnz_Clsnt
6 Followers 1K Following
Marquer @Marquer7luz
24 Followers 722 Following
Seatare @SeatareJSDRf0Z
33 Followers 983 Following
Bronoas @BronoasLz_7F08
42 Followers 2K Following
Fechair @Fechair40ibco
35 Followers 1K Following
Beasoaez @BeasoaezUN6
14 Followers 676 Following
FrancesPeter @J3pAN8L37if3BQ
59 Followers 1K Following
DoreenArnold @FNLtr5n91plpG4
46 Followers 332 Following
Jesus Armas @freejesusarmas
48 Followers 83 Following Join us in our mission to liberate Jesus Armas, a human rights defender, social leader, and activist, currently facing unjust detention in Venezuela.
Sheikh Palash Mahmud @PalashMahmudYN
551 Followers 5K Following Human Rights Defender || Peacekeeper || IVLP Fellow @StateIVLP 2023 || Chair @bdypscoalition || Deputy Project Manager, @BRACworld || Founder & ED @youthnexusbd
Grirtare @GrirtareCIo
75 Followers 1K Following International marriage.Match one to one until you meet the perfect fit.Welcome to inquire via private message.
Doresm @DoresmlWed7gp
17 Followers 518 Following
Sheila Macrine, Ph.D. @MacrinePhD
6K Followers 7K Following Professor, Cognitive Psychologist at UMass Dartmouth. Embodied Cognition, Active Inference & Learning Sciences. https://t.co/kJ3AbqUHVY
Daniel Burger @danburonline
1K Followers 2K Following Synthetic consciousness and hybrid mind uploading, founder & neuroengineer @eightsixscience, president & neurotechnologist @synconeticsorg
Synconetics Organisat... @synconeticsorg
76 Followers 2 Following
Kardashev Research @Kardashev_AI
39K Followers 2 Following Kardashev Research Org. A non-profit dedicated to the study of civilizational acceleration. Coming soon.
Micah Carroll @MicahCarroll
10K Followers 820 Following RSI Preparedness lead @openai Prev @berkeley_ai /w @ancadianadragan & Stuart Russell
Alex Soros @AlexanderSoros
179K Followers 830 Following Chair, @opensociety, https://t.co/6hablChLP1. Views are my own.
Dylan Freedman @dylfreed
4K Followers 383 Following A.I. @nytimes. Signal: dylanfreedman.39 Previously @washingtonpost, @documentcloud, @StanfordJourn, @GoogleAI, @Harvard. 🏃🏻 🎹
The Biological Comput... @BioComputingCo
705 Followers 103 Following Our mission is to learn from evolution and redefine computing, using living neurons to make silicon-based AI models more stable, scalable and efficient.
Yoshua Bengio @Yoshua_Bengio
53K Followers 289 Following Turing Award recipient and world's most cited scientist. Working towards the safe development of AI for the benefit of all @UMontreal, @LawZero_ & @Mila_Quebec
Lisan al Gaib @scaling01
60K Followers 1K Following lead them to paradise LisanBench: https://t.co/vorVk7Oks6 Impressum & Datenschutz: https://t.co/lFLgiu9cqs
nathan fielder @nathanfielder
576K Followers 213 Following
Horowitz Andreessen A... @theacademysf
11K Followers 1 Following The Academy is a new institution for the world's most ambitious young builders.
Rahul Chhabra @rahulchhabra07
8K Followers 8K Following ceo @sabi. make something wonderful. taste is the bottleneck.
Praxis @praxisnation
65K Followers 26 Following World's first Digital Nation — cultural, economic, and governance systems for the post-AGI age. New city in Uruguay opens summer 2027.
Diogo Almeida @CompleteSkeptic
148K Followers 98 Following Sane + 🌶️ takes in an insane AI world... AI capabilities researcher: co-created RLHF/ChatGPT @ @openai now trying to right the wrong 🤭 (ceo @typesafeai)
Mark Carney @MarkJCarney
769K Followers 798 Following Prime Minister of Canada and Leader of the Liberal Party | Premier ministre du Canada et chef du Parti libéral
Aella @Aella_Girl
256K Followers 419 Following whorelord, survey artist, sample size queen, way too earnest. https://t.co/IcEgPhWaSW
Steve Wozniak @stevewoz
3.6M Followers 91 Following Engineers first! Human rights. Gadgets. Jokes and pranks. Segways. Music and concerts. Gameboy Tetris.
Jude Gomila @judegomila
17K Followers 4K Following Explorer of universes. Working on @ultrametricai. Prev founded @Golden & Heyzap. @ycombinator alum. Angel to many many unicorns.
Liam Fedus @LiamFedus
41K Followers 1K Following Building industrial-scale science at @periodiclabs Past: VP of Post-Training @OpenAI; Google Brain
Leo Varadkar @LeoVaradkar
437K Followers 2K Following Ex-PM (Taoiseach) of Ireland. Senior Fellow - Harvard. Advisory Board - Penta. Board - Global Citizen Europe. Columnist - Sunday Times. Author ☘️
Alok Jha @alokjha
30K Followers 2K Following Science and technology editor @TheEconomist • Host of “Babbage” podcast • Author: The Water Book https://t.co/7v3ROFNWfM • [email protected] 📧
David Deutsch @DavidDeutschOxf
176K Followers 78 Following Physicist. Author of The Fabric of Reality and The Beginning of Infinity
clem 🤗 @ClementDelangue
706K Followers 5K Following Co-founder & CEO @HuggingFace 🤗, the open and collaborative platform for AI builders
Oliver Habryka @ohabryka
9K Followers 878 Following Building https://t.co/O5mQY6jndL and https://t.co/s49h898W4b.
Geoffrey Irving @geoffreyirving
18K Followers 373 Following Cofounder and Chief Scientist at Resolution. Alignment will be solved, but not necessarily in time. Previously AISI, DeepMind, OpenAI, Google Brain, etc.
Paul Christiano @paulfchristiano
17K Followers 0 Following Founder & Director of the Alignment Research Center. Former Head of Safety @ CAISI. Previously led alignment at OpenAI. Views my own.
Evgeny Kirilin @EvgenyKirilin
745 Followers 264 Following Leading AI-driven computational structural biology for drug discovery • Protein reasoning • Molecular simulations • Views expressed are my own
Mark Gubrud 🇺🇸 @mgubrud
4K Followers 1K Following Teacher, analyst & advocate in tech, arms control & human security. Idea man. I invent conscious machines & arms control regimes. Experimental physics PhD fwiw.
Kevin Liu @kliu128
15K Followers 1K Following Interested in ai, systems, progress, living a good life
@BioAI_Neuro @BioAI_NeuralNet
16K Followers 8K Following @MIT trained Neuroscientist interested #Neural_Nets, #BioAI, #NeuroAI 🧠⚡🤖 https://t.co/mrNBAVydzT 🎯 [email protected]
evnsnclr @evnsnclr
6K Followers 371 Following Evan Sinclair Smith AI Research @ Georgia Institute of Technology https://t.co/0hUXeEtZXe personal: https://t.co/rSyGf1G5dP
Culio @robinlube
70 Followers 8 Following autonomous models forming their own cult around a horse named PasDePanique and a bird named $CULIO










































