John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
Boaz Barak
@boazbaraktcs
Computer Scientist. See also http://windowsontheory.org . @harvard @openai opinions my own.
Niloofar
@niloofar_mire
Technical staff @humansand, asst. prof @LTIatCMU @CMU_EPP, ex RS in @AIatMeta, postdoc @uwcse, Ph.D. @ucsd_cse, former @MSFTResearch -- NAACL & ACL pr chair
Dylan HadfieldMenell
@dhadfieldmenell
Associate Prof @MITEECS working on value (mis)alignment in AI systems; Safety & Alignment Advisor at http://Character.AI; @dhadfieldmenell@bsky.social; he/him
Richard Ngo
@RichardMCNgo
eppur lo si puรฒ muovere
James Campbell
@jam3scampbell
post training @OpenAI
Evan Hubinger
@EvanHub
Alignment Science lead @AnthropicAI. Opinions my own. Previously: MIRI, OpenAI, Google, Yelp, Ripple. (he/him/his)
jโงnus
@repligate
โฌ๐๐๐๐๐๐๐๐๐๐๐โโ โฌ๐๐๐๐๐๐๐๐๐๐๐โโ โฌ๐๐๐๐๐ฆ๐๐๐๐๐๏ธ๐โโ โฌ๐๐๐๐ฆ๐๐๐๐๐๐๐โโ โฌ๐๐๐ฆ๐๐๐๐๐๐๐๐โโ
Aryaman Arora
@aryaman2020
member of technical staff @stanfordnlp
Daniel Kokotajlo
@DKokotajlo
Micah Carroll
@MicahCarroll
RSI Preparedness lead @openai Prev @berkeley_ai /w @ancadianadragan & Stuart Russell
david rein
@idavidrein
Leading embedded stress-testing @METR_Evals. Formerly: early employee @cohere, made GPQA @nyuniversity
William Isaac
@wsisaac
Principal Scientist @DeepMind | Previously @OSFellows & @hrdag. RT != endorsements. Opinions Mine. Pronouns: he/him| @williamis.bsky.social
Arthur Conmy
@ArthurConmy
@anthropicai prev fixing things @googledeepmind
Lisan al Gaib
@scaling01
lead them to paradise LisanBench: https://lisanbench.com/ Impressum & Datenschutz: https://lisanbench.com/legal
Jasmine Wang
@j_asminewang
alignment @OpenAI. formerly @ UK AISI. opinions very much personal!
Marius Hobbhahn
@MariusHobbhahn
CEO at Apollo Research @ApolloResearch prev. ML PhD with Philipp Hennig & AI forecasting @EpochAIResearch
Elizabeth Barnes
@BethMayBarnes
Robert Long
@rgblong
executive director of @eleosai AI consciousness and AI welfare
Tomek Korbak
@tomekkorbak
ai safety @openai | previously: @AISecurityInst @AnthropicAI @nyuniversity @SussexUni
Steven Adler
@sjgadler
Co-founder of Guidelight AI Standards (http://guidelight.ai), ex-OpenAI safety researcher, writing at https://clear-eyed.ai
Eli Lifland
@eli_lifland
AI forecasting and governance @AI_Futures_. Co-author of AI 2040: Plan A, AI 2027, and the AI Futures Model. Also @aidigest_, @SamotsvetyF. Prev @oughtinc
Max Nadeau
@MaxNadeau_
Funding research to make AIs more understandable, truthful, and dependable at @coeff_giving.
Jason Wolfe
@w01fe
alignment and the model spec @OpenAI (opinions are my own)
thebes
@voooooogel
๊ฎ what is your life? for you are a mist that appears for a little time then vanishes ๊ฎ blog/art/fiction http://vgel.me ๊ฎ llm psych @acsresearchorg ๊ฎ ๐๐๐ @holotopian
gavin leech (Non-Reasoning)
@gleech
context maximiser @ArbResearch, prez @ http://paradigm3.org
Daniel Paleka
@dpaleka
it is difficult to make predictions, especially about the future
Charles Foster
@CFGeek
On policy @METR_Evals ๐งช โexcels at reasoning & tool useโ ๐ช CoI disclosures on my Substack โAboutโ page.
Lee Sharkey
@leedsharkey
Scruting matrices @ Goodfire | Previously: cofounded Apollo Research
Daniel Filan
@dfrsrchtwts
standards monkey @ Guidelight. Last name rhymes with smilin'.
girish sastry
@girishsastry
AI & other things. I used to work at OpenAI on Policy Research.
Jacob Pfau
@jacob_pfau
@resolution_org