BIO
Trying to understand LLMs.
John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
David Duvenaud
@DavidDuvenaud
Machine learning prof @UofT. Former team lead at Anthropic. Working on generative models, inference, & latent structure.
Riley Goodside
@goodside
Mostly screenshots of chatbots since 2022. Formerly: Google DeepMind, Scale.
Dylan HadfieldMenell
@dhadfieldmenell
Associate Prof @MITEECS working on value (mis)alignment in AI systems; Safety & Alignment Advisor at http://Character.AI; @dhadfieldmenell@bsky.social; he/him
Owain Evans
@OwainEvans_UK
Director of Truthful AI (non-profit AI safety research group) + Affiliate at UC Berkeley. Work: Emergent misalignment, subliminal learning. Prefer email to DM.
Cas (Stephen Casper)
@StephenLCasper
Computer scientist working on AI safeguards, incidents, & gov research. Assistant professor @Kennedy_School @Harvard. https://stephencasper.com/
Evan Hubinger
@EvanHub
Alignment Science lead @AnthropicAI. Opinions my own. Previously: MIRI, OpenAI, Google, Yelp, Ripple. (he/him/his)
Belinda Li
@belindazli
Modeling models @anthropicAI Incoming prof @dsi_uchicago . Formerly PhD @MIT_CSAIL https://belindal.github.io/
Aryaman Arora
@aryaman2020
member of technical staff @stanfordnlp
Laura Ruis
@LauraRuis
Postdoc with @jacobandreas @MIT_CSAIL. PhD from @ucl_dark with @_rockt and @egrefen. Anon feedback: https://www.admonymous.co/laura-ruis
tom white
@dribnet
creations with code and networks
Cem Anil
@cem__anil
Machine learning / AI Safety at @AnthropicAI and University of Toronto / Vector Institute. Prev. @google (Blueshift Team) and @nvidia.
Arthur Conmy
@ArthurConmy
@anthropicai prev fixing things @googledeepmind
Ian Hogarth
@soundboy
co-founder & partner at @pluralplatform, co-founder & chair @AISecurityInst, co-founder @songkick
Samuel Albanie ðŽð§
@SamuelAlbanie
mid training @GoogleDeepMind
Marius Hobbhahn
@MariusHobbhahn
CEO at Apollo Research @ApolloResearch prev. ML PhD with Philipp Hennig & AI forecasting @EpochAIResearch
Rosie Campbell
@RosieCampbell
Forever expanding my nerd/bimbo Pareto frontier. AI welfare ðĪ AI safety. Managing Director @eleosai, Ex-OpenAI, 2024 @rootsofprogress fellow
Daniel Johnson
@_ddjohnson
Member of Technical Staff at @TransluceAI. Building tools to study AI systems and their behaviors. He/him.
Jeffrey Ladish
@JeffLadish
Applying the security mindset to everything @PalisadeAI
Jack Lindsey
@Jack_W_Lindsey
Neuroscience of AI brains @AnthropicAI. Previously neuroscience of real brains @cu_neurotheory.
Steven Adler
@sjgadler
Co-founder of Guidelight AI Standards (http://guidelight.ai), ex-OpenAI safety researcher, writing at https://clear-eyed.ai
Qinyuan Ye
@qinyuan_ye
âïļ Research Scientist @SFResearch | ðū Teaching machines to be versatile and curious. | Prev @nlp_usc
Eric J. Michaud
@ericjmichaud_
Trying to make deep neural networks among the best understood objects in the universe. ðŧðĪð§ ð―ðð
akbir.
@akbirkhan
ur lost, turn around
thebes
@voooooogel
ęŪ what is your life? for you are a mist that appears for a little time then vanishes ęŪ blog/art/fiction http://vgel.me ęŪ llm psych @acsresearchorg ęŪ ððð @holotopian
Samuel Marks
@saprmarks
Scalable oversight lead at @AnthropicAI. Previously: postdoc with @davidbau, math PhD at @Harvard.
Brian Huang
@brianryhuang
midtraining @GoogleDeepmind | prev math and cs @mit
Brian Christian
@brianchristian
Researcher: @CHAI_Berkeley. PhD: @summerfieldlab, Oxford. Author: The Alignment Problem, Algorithms to Live By (w. @cocosci_lab), and The Most Human Human.
Larissa Schiavo
@lfschiavo
cofounder @grove_research previously @eleosai @OpenAI @mural @USC // married to @Miles_Brundage
Miles Wang
@MilesKWang
prev @OpenAI
Daniel Paleka
@dpaleka
it is difficult to make predictions, especially about the future
Daniel Filan
@dfrsrchtwts
standards monkey @ Guidelight. Last name rhymes with smilin'.
girish sastry
@girishsastry
AI & other things. I used to work at OpenAI on Policy Research.