Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Callum McDougall

Callum McDougall

2 indexed papers

Recent (6 mo)
2
With code
0
Influential cites
0
Benchmarked
0

Publications per year

2
26

Top categories

AI×2ML×1

Frequent co-authors

Joshua Engels1×
Bilal Chughtai1×
Janos Kramar1×
Senthoran Rajamanoharan1×
Cindy Wu1×
Arthur Conmy1×

Research Timeline

2026
Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

The paper demonstrates that sparse autoencoders can successfully extract a large set of interpretable, causally influential features from the production-scale Claude 3 Sonnet language model.

How Transparent is DiffusionGemma?

This paper investigates the transparency of DiffusionGemma, a model with a larger fraction of computation in a continuous latent space, and shows that it can be made more transparent by mapping information through an interpretable token bottleneck.

Highlighted terms show continued research focus across papers

Papers

cs.LGcs.AIEmpiricalRecentJun 18, 2026

How Transparent is DiffusionGemma?

Joshua Engels, Callum McDougall, Bilal Chughtai, Janos Kramar +10 more

This paper investigates the transparency of DiffusionGemma, a model with a larger fraction of computation in a continuous latent space, and shows that it can be made more transparent by mapping inform…

View →
cs.AIRecentMay 28, 2026

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

Adly Templeton, Tom Conerly, Jonathan Marcus, Jack Lindsey +22 more

The paper demonstrates that sparse autoencoders can successfully extract a large set of interpretable, causally influential features from the production-scale Claude 3 Sonnet language model.

View →