Addison J. Wu

Addison J. Wu

I'm a rising senior at Princeton University studying computer science and math, where I'm advised by Tom Griffiths, Zhuang Liu, and Karthik Narasimhan as a part of Princeton Language and Intelligence. Currently, I'm also leading a project with the Center for AI Safety, where I was a research intern in summer 2026, advised by Mantas Mazeika and Dan Hendrycks. I got my first taste of research at the Multi-Physics Interaction Lab, University of Waterloo, advised by Jean-Pierre Hickey

I want AI systems to interact better with each other, with us, and with the world over time so we can do even greater things we find worth doing

Some projects I've especially enjoyed working on are:

Novel social biases
Large Language Models Develop Novel Social Biases Through Adaptive Exploration
Addison J. Wu*, Ryan Liu*, Xuechunzi Bai, Thomas L. Griffiths
ICML 2026 Oral [arXiv]
Media coverage: The Guardian, Forbes, MIT Technology Review, Hacker News (Y Combinator), The Times of India, Gizmodo, Arizona's Family
Adversarial influence in multi-agent systems
How does Adversarial Influence Scale in Multi-Agent Systems?
Addison J. Wu*, Jasin Cekinmez*, Michel Liao*, Karthik Narasimhan, Thomas L. Griffiths
arXiv preprint, 2026 [arXiv]
Ads in AI chatbots
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
Addison J. Wu*, Ryan Liu*, Shuyue Stella Li, Yulia Tsvetkov, Thomas L. Griffiths
COLM 2026 [arXiv]
Media coverage: Hacks/Hackers, The Indian Express; Cited in Senate testimony
Motives behind communication
Are Large Language Models Sensitive to the Motives Behind Communication?
Addison J. Wu*, Ryan Liu*, Kerem Oktar*, Theodore R. Sumers, Thomas L. Griffiths
NeurIPS 2025; PragLM @ COLM 2025 Oral [arXiv]
Polysemous words in multimodal models
Where did the ambiguity go? Examining how multimodal models interpret polysemous words
Jasin Cekinmez*, Addison J. Wu*, Raja Marjieh, Thomas L. Griffiths
Sci-FM @ COLM 2026 Oral [arXiv]
Compute-supervision tradeoffs in RLVR
Quantifying Empirical Compute-Supervision Tradeoffs in RLVR
Ryo Mitsuhashi*, Patrick Chen*, Isabelle Tseng*, Jasin Cekinmez*, Addison J. Wu*
Workshop on Combining Benchmarks and Theory @ ICML 2026 [arXiv]

I love to travel ✈️
🇨🇦 🇺🇸 🇲🇽 🇵🇷 🇵🇦 🇧🇷 🇯🇵 🇰🇷 🇨🇳 🇭🇰 🇮🇪 🇬🇧 🇫🇷 🇪🇸 🇧🇪 🇨🇭 🇮🇹 🇻🇦 🇸🇲 🇩🇰 🇸🇪 🇳🇴