Richard Zhuang
Current: Anthropic Safeguards Research🛡️, Prev: Stanford MSCS🌲, UC Berkeley CS + Applied Math🐻, Research Intern at Bespoke Labs
(Last Updated: 2026.09)
Welcome to my personal space! A quick intro:
- 🛡️ Now: Member of the Safeguards Research Team at Anthropic.
- 🌲 Stanford: MS in Computer Science, where I was a core contributor of the OpenThoughts-Agent project with Prof. Ludwig Schmidt, working on data recipe for post-training agents.
- 🧪 Bespoke Labs: Research Intern in Spring 2025, working on enhancing tool-use capability of LLM agents through RL (blog).
- 🐻 UC Berkeley: BA in Applied Math and Computer Science. Researched on LLM routing (EmbedLLM) with Jiantao Jiao and Tianhao Wu, as well as LLM + Game (PokerBench) with Akshat Gupta.
I’m broadly interested in understanding and improving the capabilities of Large Language Models (LLMs) in a data-centric way. Specifically, I’m intrigued by how certain data “foster” skills that are essential for LLM agents (e.g. reasoning and planning). I have also had a long-standing passion in Sports Analytics.
Outside the realm of AI, you will usually find me playing basketball🏀, pickleball🥒, or poker♠️, or immersing myself in Chinese Hip-hop music🔥.
News
| Jul 15, 2026 | Joined Anthropic as a member of the Safeguards Research Team! 🎉 |
|---|---|
| Jun 25, 2026 | Announcing OpenThoughts-Agent and OpenThinkerAgent-32B — the strongest Qwen-3-based open-data agentic model for terminal use and coding, reaching 44.8% average accuracy across seven agentic benchmarks. We openly share the full stack: paper, model, data, and code. Read the X thread for the highlights. |