• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    Rohin Shah

    Rohin Shah

    96Following221Top Followers13.6KTotal Followers

    BIO

    AGI Safety & Alignment @ Google DeepMind

    #767in/Tech
    @rohinmshah
    @rohinmshah

    Vibe and topics

    Based on 180 recent X posts Β· click a vibe for evidence

    Vibe
    Informing
    Informing42.5%
    Announcing36.4%
    Teaching10%
    Hopeful5.6%
    Supportive3.6%
    Other1.9%
    Topics
    Alignment Newsletter
    Alignment Newsletter55.7%
    Deceptive Alignment8.6%
    Frontier Safety Framework7.4%
    Goal Misgeneralization7.1%
    Minecraft Benchmarks6.3%
    Chain-of-Thought Monitoring5.6%
    Mechanistic Interpretability4.8%
    AI Safety Hiring4.6%

    Top followers

    Pieter Abbeel
    #21

    Pieter Abbeel

    @pabbeel

    Berkeley & Amazon

    Sam Bowman
    #29

    Sam Bowman

    @sleepinyourhat

    AI alignment + LLMs at Anthropic. On leave from NYU. Views not employers'. No relation to @s8mb. Into @givingwhatwecan.

    Zachary Lipton
    #30

    Zachary Lipton

    @zacharylipton

    Professor: CMU/@acmi_lab, Cofounder: @AbridgeHQ, Creator: @d2l_ai & http://approximatelycorrect.com, Relapsing 🎷

    John Schulman
    #36

    John Schulman

    @johnschulman2

    @thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music

    Jan Leike
    #39

    Jan Leike

    @janleike

    AI research @AnthropicAI. Previously OpenAI & DeepMind. Optimizing for a post-AGI future where humanity flourishes. Opinions aren't my employer's.

    Eric Jang
    #68

    Eric Jang

    @ericjang11

    Zico Kolter
    #81

    Zico Kolter

    @zicokolter

    Professor and Head of Machine Learning Department at @CarnegieMellon. Board member @OpenAI and @Qualcomm. Chief Scientist @GraySwanAI.

    Roger Grosse
    #89

    Roger Grosse

    @RogerGrosse

    Yisong Yue
    #113

    Yisong Yue

    @yisongyue

    AI Professor @Caltech (@YueLabCaltech)

    Diyi Yang
    #118

    Diyi Yang

    @Diyi_Yang

    Assistant Professor @Stanford CS @StanfordNLP @StanfordAILab Part time @humansand LLMs for Humans

    Jakob Foerster
    #126

    Jakob Foerster

    @j_foerst

    Associate Prof in ML @UniofOxford. Something Something Research Scientist @MetaAI. Something @BOLD_LAB_AI. Always #teamhuman. Opinions belong to the world.

    Eric Wallace
    #135

    Eric Wallace

    @Eric_Wallace_

    research @openai

    Jeff Clune
    #142

    Jeff Clune

    @jeffclune

    Co-founder, Recursive. Professor, CS, U. British Columbia. CIFAR AI Chair, Vector Institute. | ML, AI, deep RL, deep learning, AI-Generating Algorithms (AI-GAs)

    Brandon Amos
    #143

    Brandon Amos

    @brandondamos

    πŸ§™ RL @Reflection_AI past: @MetaAi @GoogleDeepmind @SCSatCMU @Cornell_Tech

    Jascha Sohl-Dickstein
    #147

    Jascha Sohl-Dickstein

    @jaschasd

    Member of the technical staff @ Anthropic. Most (in)famous for inventing diffusion models. AI + physics + neuroscience + dynamics.

    Ethan Perez
    #149

    Ethan Perez

    @EthanJPerez

    Alignment team lead at Anthropic

    Amanda Askell
    #153

    Amanda Askell

    @AmandaAskell

    Philosopher & ethicist trying to make AI be good @AnthropicAI. Personal account. All opinions come from my training data.

    Catherine Olsson
    #167

    Catherine Olsson

    @catherineols

    Hanging out with Claude, improving its behavior, and building tools to support that @AnthropicAI 😁 prev: @open_phil @googlebrain @openai (@microcovid)

    Joshua Achiam
    #173

    Joshua Achiam

    @jachiam0

    Freedom, flourishing, and abundance. Prev: @openai. Main author of http://spinningup.openai.com

    Riley Goodside
    #183

    Riley Goodside

    @goodside

    Mostly screenshots of chatbots. Formerly: Google DeepMind, Scale.

    Anca Dragan
    #185

    Anca Dragan

    @ancadianadragan

    Google DeepMind β€’ AI safety & alignment β€’ model behavior β€’ associate professor @ UC Berkeley EECS

    Stephanie Chan
    #216

    Stephanie Chan

    @scychan_brains

    Staff Research Scientist at the DeepMind Institute. Societal impacts of AI + Science of AI. Views are my own.

    Rishabh Agarwal
    #221

    Rishabh Agarwal

    @agarwl_

    Reinforcement Learner

    rishi
    #233

    rishi

    @RishiBommasani

    Senior Research Scholar @StanfordHAI focused on economics and governance of frontier AI Previous: Stanford CS PhD @percyliang @jurafsky, Cornell CS

    Neel Nanda
    #246

    Neel Nanda

    @NeelNanda5

    Mechanistic Interpretability lead DeepMind. Formerly @AnthropicAI, independent. In this to reduce AI X-risk. Neural networks can be understood, let's go do it!

    Shane Legg
    #247

    Shane Legg

    @ShaneLegg

    Chief AGI Scientist & Co-Founder, Google DeepMind Director and Managing Editor, DeepMind Institute Work: http://www.deepmind.com Personal: http://www.vetta.org

    near
    #270

    near

    @nearcyan

    where

    He He
    #276

    He He

    @hhexiy

    NLP researcher. Assistant Professor at NYU CS & CDS.

    Andrew Lampinen
    #277

    Andrew Lampinen

    @AndrewLampinen

    Interested in cognition and artificial intelligence. MTS at @AnthropicAI. Previously @DeepMind, cognitive science @StanfordPsych. Tweets are mine.

    Chris J. Maddison
    #319

    Chris J. Maddison

    @cjmaddison

    ML faculty @UofT

    Zhou Yu
    #348

    Zhou Yu

    @Zhou_Yu_AI

    Founder of http://Arklex.ai, Associate Professor at Columbia. Making ai agent design and deployment easy, safe, and fast! Forbes 30 under 30.

    Owain Evans
    #349

    Owain Evans

    @OwainEvans_UK

    Director of Truthful AI (non-profit AI safety research group) + Affiliate at UC Berkeley. Work: Emergent misalignment, subliminal learning. Prefer email to DM.

    Daniel Fried
    #356

    Daniel Fried

    @dan_fried

    Assistant prof. @LTIatCMU @SCSatCMU. Working on NLP: LLM agents, language-to-code, applied pragmatics, grounding.

    Geoffrey Irving
    #362

    Geoffrey Irving

    @geoffreyirving

    Cofounder and Chief Scientist at Resolution. Alignment will be solved, but not necessarily in time. Previously AISI, DeepMind, OpenAI, Google Brain, etc.

    Richard Ngo
    #363

    Richard Ngo

    @RichardMCNgo

    eppur lo si puΓ² muovere

    Pavel Izmailov
    #396

    Pavel Izmailov

    @Pavel_Izmailov

    Researcher @AnthropicAI πŸ€– Assistant Professor @nyuniversity πŸ™οΈ Previously @OpenAI #StopWar πŸ‡ΊπŸ‡¦

    Kristian Lum
    #411

    Kristian Lum

    @KLdivergence

    Measuring AI's human and societal impacts | Building tools for AI transparency | Research Scientist at GDM | @FAccTConference OG | Ex UPenn, UChicago faculty

    Stephen McAleer
    #424

    Stephen McAleer

    @McaleerStephen

    AI researcher at Anthropic

    Theo Weber
    #434

    Theo Weber

    @theophaneweber

    Research scientist @ DeepMind; currently working on thinking/reasoning in Gemini.

    Roberta Raileanu
    #456

    Roberta Raileanu

    @robertarail

    Open-Endedness Team Lead and Senior Staff Research Scientist @GoogleDeepMind. Adjunct Faculty @bold_lab_ai. ex @Meta | @NYU | @Princeton | IPhO | IOAA.

    Nat McAleese
    #459

    Nat McAleese

    @__nmca__

    Research @AnthropicAI. Previously @OpenAI, @DeepMind. Views my own.

    Andrew Trask
    #464

    Andrew Trask

    @iamtrask

    most important AI paper i've read (1968): https://tinyurl.com/unn5rh3x @openminedorg @GoogleDeepMind @OxfordUni

    Andrew Carr 🀸
    #473

    Andrew Carr 🀸

    @andrew_n_carr

    co-founder leading science @getcartwheel co-founder advisor @arcade_ai Past: Codex @OpenAI, Brain @GoogleAI, world ranked Tetris player

    Leo Gao
    #486

    Leo Gao

    @nabla_theta

    working on AGI alignment. prev: GPT-Neo, the Pile, LM evals, RL overoptimization, scaling SAEs to GPT-4, interp via circuit sparsity. EleutherAI cofounder.

    noahdgoodman
    #487

    noahdgoodman

    @noahdgoodman

    humans& co-founder. professor of natural and artificial intelligence @Stanford. (@StanfordNLP @StanfordAILab) ex: Google DeepMind, UberAI

    Ruiqi Zhong
    #502

    Ruiqi Zhong

    @ZhongRuiqi

    Member of Technical Staff at Thinking Machines. Human+AI collaboration. Scalable Oversight. Explainability. Prev @AnthropicAI PhD UC Berkeley'25; Columbia'19

    James Campbell
    #503

    James Campbell

    @jam3scampbell

    post training @OpenAI

    Kevin Roose
    #504

    Kevin Roose

    @kevinroose

    Tech journalist, author of The AGI Chronicles, co-host of @MachineGodsPod, p(doom) available upon request.

    Zhiqing Sun
    #507

    Zhiqing Sun

    @EdwardSun0909

    Lead agent research @Meta MSL TBD Lab. previously posttraining/agent research @OpenAI. CS PhD @LTIatCMU

    π™·πš’πš–πšŠ π™»πšŠπš”πš”πšŠπš›πšŠπš“πšž
    #515

    π™·πš’πš–πšŠ π™»πšŠπš”πš”πšŠπš›πšŠπš“πšž

    @hima_lakkaraju

    AI Professor @Harvard; Senior Staff Research Scientist @GoogleAI; @trustworthy_ml #AI #XAI; AI PhD from Stanford; Sloan/Kavli Fellow, MIT TR #35Under35