BIO
AI Control at Anthropic
John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
Tim Rocktäschel
@_rockt
Teaching AI the joy of invention at @Recursive_SI, Professor of AI @AI_UCL, PI @BOLD_Lab_AI, Fellow @ELLISforEurope. Ex @GoogleDeepMind @AIatMeta @CompSciOxford
Ethan Perez
@EthanJPerez
Alignment team lead at Anthropic
Riley Goodside
@goodside
Mostly screenshots of chatbots since 2022. Formerly: Google DeepMind, Scale.
Owain Evans
@OwainEvans_UK
Director of Truthful AI (non-profit AI safety research group) + Affiliate at UC Berkeley. Work: Emergent misalignment, subliminal learning. Prefer email to DM.
Rylan Schaeffer
@RylanSchaeffer
Creating something new. Ex-Meta TBD. On-Leave from Stanford w/ @sanmikoyejo. Prev @ Gemini, MIT, Harvard, Uber, UCL, UC Davis
Evan Hubinger
@EvanHub
Alignment Science lead @AnthropicAI. Opinions my own. Previously: MIRI, OpenAI, Google, Yelp, Ripple. (he/him/his)
Sarah Catanzaro
@sarahcat21
“All methods are sacred if they are internally necessary” (GP @amplifypartners, prev @canvasvc; Head of Data @Mattermark; @palantirtech; @c4ads)
Micah Carroll
@MicahCarroll
RSI Preparedness lead @openai Prev @berkeley_ai /w @ancadianadragan & Stuart Russell
Laura Ruis
@LauraRuis
Postdoc with @jacobandreas @MIT_CSAIL. PhD from @ucl_dark with @_rockt and @egrefen. Anon feedback: https://www.admonymous.co/laura-ruis
david rein
@idavidrein
Leading embedded stress-testing @METR_Evals. Formerly: early employee @cohere, made GPQA @nyuniversity
Cem Anil
@cem__anil
Machine learning / AI Safety at @AnthropicAI and University of Toronto / Vector Institute. Prev. @google (Blueshift Team) and @nvidia.
Arthur Conmy
@ArthurConmy
@anthropicai prev fixing things @googledeepmind
Winnie Xu
@winniethexu
making gemini agents think @GoogleDeepmind, hopefully at scale (: bsc cs @UofT. also capturing, potting (http://xu.haus) and wandering.
Maksym Andriushchenko
@maksym_andr
Principal investigator @ELLISInst_Tue & @MPI_IS, advisor @expsecai, mentor @MATSprogram. Past projects: AgentHarm, Claudini, PostTrainBench, Stolen Thoughts.
Marius Hobbhahn
@MariusHobbhahn
CEO at Apollo Research @ApolloResearch prev. ML PhD with Philipp Hennig & AI forecasting @EpochAIResearch
Nitarshan
@nitarshan
compute @anthropic, PhD @cambridge_cl. prev created @aisecurityinst, AI Safety Summit, UK AI Research Resource, EU AI Code of Practice.
Zac Kenton
@ZacKenton1
Amplified Oversight team lead @GoogleDeepMind | AGI safety & alignment | Enabling accurate human supervision of superhuman AI.
akbir.
@akbirkhan
ur lost, turn around
Max Nadeau
@MaxNadeau_
Funding research to make AIs more understandable, truthful, and dependable at @coeff_giving.
mrinank
@MrinankSharma
poet // researcher may we each follow our threads everything has to do with loving and not loving -rumi
Brian Huang
@brianryhuang
midtraining @GoogleDeepmind | prev math and cs @mit
Jenny Zhang
@jennyzhangzt
Improving @Recursive_SI
Leonard Tang
@leonardtang_
vp of ai research @beaconholdings prev ceo @haizelabs
Daniel Paleka
@dpaleka
it is difficult to make predictions, especially about the future
girish sastry
@girishsastry
AI & other things. I used to work at OpenAI on Policy Research.
Miles Turpin
@milesaturpin
Safety and alignment evals @meta Superintelligence Labs. Previously alignment research @scale_AI, @nyuniversity, early employee @cohere