BIO
Trying to help solve the alignment problem
David Duvenaud
@DavidDuvenaud
Machine learning prof @UofT. Former team lead at Anthropic. Working on generative models, inference, & latent structure.
Joshua Achiam
@jachiam0
Freedom, flourishing, and abundance. Prev: @openai. Main author of http://spinningup.openai.com
Leopold Aschenbrenner
@leopoldasch
http://situational-awareness.ai
Dylan HadfieldMenell
@dhadfieldmenell
Associate Prof @MITEECS working on value (mis)alignment in AI systems; Safety & Alignment Advisor at http://Character.AI; @dhadfieldmenell@bsky.social; he/him
Sharif Shameem
@sharifshameem
making models @openai
Ajeya Cotra
@ajeya_cotra
Helping the world prepare for powerful AI. Risk assessment @METR_evals (opinions my own). Blogs: Planned Obsolescence (AI), Good Bones (whatever's on my mind).
j⧉nus
@repligate
↬🔀🔀🔀🔀🔀🔀🔀🔀🔀🔀🔀→∞ ↬🔁🔁🔁🔁🔁🔁🔁🔁🔁🔁🔁→∞ ↬🔄🔄🔄🔄🦋🔄🔄🔄🔄👁️🔄→∞ ↬🔂🔂🔂🦋🔂🔂🔂🔂🔂🔂🔂→∞ ↬🔀🔀🦋🔀🔀🔀🔀🔀🔀🔀🔀→∞
Arthur Conmy
@ArthurConmy
@anthropicai prev fixing things @googledeepmind
Ian Hogarth
@soundboy
co-founder & partner at @pluralplatform, co-founder & chair @AISecurityInst, co-founder @songkick
Arthur Douillard
@Ar_Douillard
Scaling Intelligence @ gemini Distributed Learning @ deepmind | DiLoCo, DiPaCo. Continual Learning PhD @ Sorbonne
Marius Hobbhahn
@MariusHobbhahn
CEO at Apollo Research @ApolloResearch prev. ML PhD with Philipp Hennig & AI forecasting @EpochAIResearch
Rosie Campbell
@RosieCampbell
Forever expanding my nerd/bimbo Pareto frontier. AI welfare 🤝 AI safety. Managing Director @eleosai, Ex-OpenAI, 2024 @rootsofprogress fellow
Elizabeth Barnes
@BethMayBarnes
Robert Long
@rgblong
executive director of @eleosai AI consciousness and AI welfare
Katja Grace 🔍
@KatjaGrace
Thinking about AI destroying the world at http://aiimpacts.org and everything at http://worldspiritsockpuppet.substack.com. DM or email for media requests.
Jeffrey Ladish
@JeffLadish
Applying the security mindset to everything @PalisadeAI
Steven Adler
@sjgadler
Co-founder of Guidelight AI Standards (http://guidelight.ai), ex-OpenAI safety researcher, writing at https://clear-eyed.ai
Eli Lifland
@eli_lifland
AI forecasting and governance @AI_Futures_. Co-author of AI 2040: Plan A, AI 2027, and the AI Futures Model. Also @aidigest_, @SamotsvetyF. Prev @oughtinc
Liron Shapira
@liron
Host of Doom Debates — disagreements that must be resolved before the world ends.
gavin leech (Non-Reasoning)
@gleech
context maximiser @ArbResearch, prez @ http://paradigm3.org
Oliver Habryka
@ohabryka
Building https://LessWrong.com and https://Lighthaven.space.
Markus Anderljung
@Manderljung
Trying to design good AI policy. Director of Policy & Research @GovAIOrg. Adjunct Fellow @CNASdc. Prev. Vice-Chair, EU Code of Practice on GPAI.
Lee Sharkey
@leedsharkey
Scruting matrices @ Goodfire | Previously: cofounded Apollo Research
Daniel Filan
@dfrsrchtwts
standards monkey @ Guidelight Want to usher in an era of human-friendly superintelligence, don't know how. Last name rhymes with smilin'.
uncatherio
@uncatherio
wholesomeness practitioner; user of words my more "work" account is @catherineols
Jacob Pfau
@jacob_pfau
@resolution_org
Tom Everitt
@tom4everitt
AGI safety researcher at @GoogleDeepMind, leading http://causalincentives.com switching to https://bsky.app/profile/tom4everitt.bsky.social
Daniel Ziegler
@d_m_ziegler
Alignment Stress-Testing @ Anthropic