Sara Price
BIO
Head of Alignment Training at Anthropic
Top followers
Sam Bowman
@sleepinyourhat
AI alignment + LLMs at Anthropic. On leave from NYU. Views not employers'. No relation to @s8mb. Into @givingwhatwecan.
John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
Jan Leike
@janleike
AI research @AnthropicAI. Previously OpenAI & DeepMind. Optimizing for a post-AGI future where humanity flourishes. Opinions aren't my employer's.
Noam Brown
@polynoamial
Researching reasoning @OpenAI | Co-created Libratus/Pluribus superhuman poker AIs, CICERO Diplomacy AI, and OpenAI o-series 🍓 reasoning models
Roger Grosse
@RogerGrosse
Barret Zoph
@barret_zoph
VP Research GoogleDeepmind Past: - VP & GM Enterprise (@openai) - CTO & Co-Founder TML (@thinkymachines) - VP Research (Post-Training) @openai
Jascha Sohl-Dickstein
@jaschasd
Member of the technical staff @ Anthropic. Most (in)famous for inventing diffusion models. AI + physics + neuroscience + dynamics.
Ethan Perez
@EthanJPerez
Alignment team lead at Anthropic
Amanda Askell
@AmandaAskell
Philosopher & ethicist trying to make AI be good @AnthropicAI. Personal account. All opinions come from my training data.
Boaz Barak
@boazbaraktcs
Computer Scientist. See also http://windowsontheory.org . @harvard @openai opinions my own.
Catherine Olsson
@catherineols
Hanging out with Claude, improving its behavior, and building tools to support that @AnthropicAI 😁 prev: @open_phil @googlebrain @openai (@microcovid)
Joshua Achiam
@jachiam0
Freedom, flourishing, and abundance. Prev: @openai. Main author of http://spinningup.openai.com
Dylan HadfieldMenell
@dhadfieldmenell
Associate Prof @MITEECS working on value (mis)alignment in AI systems; Safety & Alignment Advisor at http://Character.AI; @dhadfieldmenell@bsky.social; he/him
Alex Tamkin
@AlexTamkin
machine learning, science & society @AnthropicAI | recently: Clio, Anthropic Economic Index, Claude Artifacts | prev: phd @StanfordAILab, @stanfordnlp
Owain Evans
@OwainEvans_UK
Runs an AI Safety research group in Berkeley (Truthful AI) + Affiliate at UC Berkeley. Past: Oxford Uni, TruthfulQA, Reversal Curse. Prefer email to DM.
Richard Ngo
@RichardMCNgo
eppur lo si può muovere
Stephen McAleer
@McaleerStephen
AI researcher at Anthropic
Jesse Mu
@jayelmnop
computational linguistics
kipply
@kipperrii
"drop the forest nymph act we know how much gdp you generate" - max novendstern
Rylan Schaeffer
@RylanSchaeffer
Creating something new. Ex-Meta TBD. On-Leave from Stanford w/ @sanmikoyejo. Prev @ Gemini, MIT, Harvard, Uber, UCL, UC Davis
James Campbell
@jam3scampbell
post training @OpenAI
Zhiqing Sun
@EdwardSun0909
Lead agent research @Meta MSL TBD Lab. previously posttraining/agent research @OpenAI. CS PhD @LTIatCMU
Chenhao Tan
@ChenhaoTan
Professor @UChicagoCS @UChicago. Directing @ChicagoHAI, also part of @UChicagoCI. Email for Postdoc/PhD opportunities. https://chenhaot.com
Evan Hubinger
@EvanHub
Alignment Stress-Testing lead @AnthropicAI. Opinions my own. Previously: MIRI, OpenAI, Google, Yelp, Ripple. (he/him/his)
Ajeya Cotra
@ajeya_cotra
Helping the world prepare for powerful AI. Risk assessment @METR_evals (opinions my own). Blogs: Planned Obsolescence (AI), Good Bones (whatever's on my mind).
Trenton Bricken
@TrentonBricken
Stochastic parrot
Hailey Schoelkopf
@haileysch__
hillclimbing towards generality @anthropicai | prev @AiEleuther | views my own
Jason Phang
@zhansheng
Foundations at @OpenAI. PhD @NYUDataScience, @AiEleuther, 🇸🇬. Prev: @Google, @Microsoft
Peter Hase
@peterbhase
I work in grantmaking for AI safety and interpretability Currently: Schmidt Sciences, Stanford Previously: Anthropic, AI2, Google, Meta, UNC Chapel Hill
Ted Sanders
@sandersted
Research at OpenAI. Be kind to others, and yourself.
Micah Carroll
@MicahCarroll
RSI Preparedness lead @openai Prev @berkeley_ai /w @ancadianadragan & Stuart Russell
david rein
@idavidrein
Leading embedded stress-testing @METR_Evals. Formerly: early employee @cohere, made GPQA @nyuniversity
Yasmin Razavi
@YasminRazavi
but a single story higher
Laurens van der Maaten
@lvdmaaten
Member of Technical Staff at Anthropic. Ex-Meta. t-SNE. Llama 3. DenseNet. Web-scale weakly supervised vision. CrypTen.
tom white
@dribnet
creations with code and networks
Mitchell Wortsman
@Mitchnw
Ming-Wei Chang
@mchang21
GenAI @GoogleDeepMind. LLM: BERT, REALM, Gemini 1, 2 and 3. Others: Image Editing. Multimodal. Retrieval. Computer Use.
Cem Anil
@cem__anil
Machine learning / AI Safety at @AnthropicAI and University of Toronto / Vector Institute. Prev. @google (Blueshift Team) and @nvidia.
Arthur Conmy
@ArthurConmy
@anthropicai prev fixing things @googledeepmind
Joe Carlsmith
@jkcarlsmith
Philosophy, futurism, AI. Working on Claude's values @AnthropicAI. Formerly @coeff_giving. Opinions my own.
Hugh Zhang
@hughbzhang
Jasmine Wang
@j_asminewang
alignment @OpenAI. formerly @ UK AISI. opinions mine!
David Alvarez Melis
@elmelis
Asst. Prof. @hseas @KempnerInst || Researcher @MSRNE || ML + NLP || Previously: @MIT_CSAIL NYU @IBMResearch @ITAM_mx
Marius Hobbhahn
@MariusHobbhahn
CEO at Apollo Research @ApolloResearch prev. ML PhD with Philipp Hennig & AI forecasting @EpochAIResearch
Faeze Brahman
@faeze_brh
Sr. Research Scientist @allen_ai | Prev. Postdoc @allen_ai @uwnlp | Ph.D. from UCSC | Former Intern @MSFTResearch , @allen_ai | Researcher in #NLProc, #ML #AI
Sanmi Koyejo
@sanmikoyejo
I lead @stai_research at Stanford. Co-founder @VirtueAI_co
Anirudh Goyal
@anirudhg9119
Thinking about thinking. Spent time at @Berkeley_EECS, @MPI_IS, @GoogleDeepMind Gemini ♊.
Allan Dafoe
@AllanDafoe
AGI governance: navigating the transition to beneficial AGI (Google DeepMind)
Nitarshan
@nitarshan
compute @anthropic, PhD @cambridge_cl. prev created @aisecurityinst, AI Safety Summit, UK AI Research Resource, EU AI Code of Practice.
Sonia Joseph
@soniajoseph_
co-founder @ stealth | prev world models / JEPA, interpretability @AIatMeta @Mila_Quebec @Princeton