Mitchell Wortsman

@Mitchnw

@AnthropicAI | prev @uwcse

Joined October 2011

1KFollowing

2KFollowers

Pinned

Mitchell Wortsman@Mitchnw · Sep 28, 2023

Sharing some highlights from our work on small-scale proxies for large-scale Transformer training instabilities: arxiv.org/abs/2309.14322 With fantastic collaborators @peterjliu, @Locchiu, @_katieeverett, many others (see final tweet!), @hoonkp, @jmgilmer, @skornblith! (1/15)

Mitchnw's tweet image. Sharing some highlights from our work on small-scale proxies for large-scale Transformer training instabilities: arxiv.org/abs/2309.14322

With fantastic collaborators @peterjliu, @Locchiu, @_katieeverett, many others (see final tweet!), @hoonkp, @jmgilmer, @skornblith!

(1/15)

345

206

99.0K

Pinned

Mitchell Wortsman Retweeted

Akari Asai@AkariAsai · Dec 4

🚨 I’m on the job market this year! 🚨 I’m completing my @uwcse Ph.D. (2025), where I identify and tackle key LLM limitations like hallucinations by developing new models—Retrieval-Augmented LMs—to build more reliable real-world AI systems. Learn more in the thread! 🧵

118

823

198

126.0K

Pinned

Mitchell Wortsman Retweeted

Anthropic@AnthropicAI · Oct 22

Introducing an upgraded Claude 3.5 Sonnet, and a new model, Claude 3.5 Haiku. We’re also introducing a new capability in beta: computer use. Developers can now direct Claude to use computers the way people do—by looking at a screen, moving a cursor, clicking, and typing text.

483

2.0K

10.0K

3.0K

3.7M

Mitchell Wortsman Retweeted

Ludwig Schmidt@lschmidt3 · Jun 5

Very excited to finally release our paper for OpenThoughts! After DataComp and DCLM, this is the third large open dataset my group has been building in collaboration with the DataComp community. This time, the focus is on post-training, specifically reasoning data.

212

1.0K

875

167.0K

Mitchell Wortsman Retweeted

Anthropic@AnthropicAI · May 22

Introducing the next generation: Claude Opus 4 and Claude Sonnet 4. Claude Opus 4 is our most powerful model yet, and the world’s best coding model. Claude Sonnet 4 is a significant upgrade from its predecessor, delivering superior coding and reasoning.

963

3.0K

21.0K

4.0K

4.2M

Mitchell Wortsman Retweeted

Cade Gordon@CadeGordonML · May 21

Excited to share that I'll be joining @Anthropic to work on pretraining science! I've chosen to defer my Stanford PhD, where I'm honored to be supported by the Hertz Fellowship. There's something special about the science, this place, and these people. Looking forward to joining…

775

109

58.0K

Mitchell Wortsman Retweeted

Mike A. Merrill@Mike_A_Merrill · May 19

Many agents (Claude Code, Codex CLI) interact with the terminal to do valuable tasks, but do they currently work well enough to deploy en masse? We’re excited to introduce Terminal-Bench: An evaluation environment and benchmark for AI agents on real-world terminal tasks. Tl;dr…

234

101

47.0K

Mitchell Wortsman Retweeted

Alex Li@alexlioralexli · Apr 26

Excited to be presenting at #ICLR2025 at 10am today on how generative classifiers are much more robust to distribution shift. Come by to chat and say hello!

6.0K

Mitchell Wortsman Retweeted

Alex Li@alexlioralexli · Dec 12

I'm presenting our #NeurIPS2024 work on Attention Transfer today! Key finding: Pretrained representations aren't essential - just using attention patterns from pretrained models to guide token interactions is enough for models to learn high-quality features from scratch and…

160

14.0K

Mitchell Wortsman Retweeted

Ofir Press@OfirPress · Dec 4

I'm on the academic job market! I develop autonomous systems for: programming, research-level question answering, finding sec vulnerabilities & other useful+challenging tasks. I do this by building frontier-pushing benchmarks and agents that do well on them. See you at NeurIPS!

230

24.0K

Mitchell Wortsman Retweeted

Ross Wightman@wightmanr · Oct 17

OpenCLIP passed 10K stars on GitHub this week. A big milestone for any open-source project. 🍻 to the many collaborators that made that possible. Coincidentally, I pushed a new release with a port of the largest multi-lingual SigLIP -- a SO400M/16 @ 256x256 that appeared on…

148

26.0K

Mitchell Wortsman Retweeted

Katie Everett@_katieeverett · Jul 23, 2024

Come chat with me and @Locchiu at our ICML poster session 1:30-3pm CEST (Vienna time) today at Hall C 4-9 #2500 and see how our theory lets all parameterizations perform hyperparameter transfer! arxiv.org/abs/2407.05872

65.0K

Mitchell Wortsman Retweeted

Vaishaal Shankar@Vaishaal · Jul 18, 2024

We have released our DCLM models on huggingface! To our knowledge these are by far the best performing truly open-source models (open data, open weight models, open training code) 1/5

288

110

51.0K

Mitchell Wortsman Retweeted

Katie Everett@_katieeverett · Jul 18, 2024

We've gotten some great questions about the notion of alignment in our width-scaling parameterization paper! arxiv.org/abs/2407.05872 A deep dive into the alignment metric and intuition 🧵 [1/16]

14.0K

Mitchell Wortsman Retweeted

Tomer Porian@tomerporian · Jul 2, 2024

🧵1/8 We resolve the discrepancy between the compute optimal scaling laws of Kaplan (exponent 0.88, Figure 14, left) et al. and Hoffmann et al. (“Chinchilla”, exponent 0.5). Paper: arxiv.org/abs/2406.19146 Data + Code: github.com/formll/resolvi…

170

124

36.0K

Mitchell Wortsman Retweeted

Anthropic@AnthropicAI · Jun 20, 2024

We're also launching a preview of Artifacts on claude.ai. You can ask Claude to generate docs, code, mermaid diagrams, vector graphics, or even simple games. Artifacts appear next to your chat, letting you see, iterate, and build on your creations in real-time.

194

2.0K

384

572.0K

Mitchell Wortsman Retweeted

Anthropic@AnthropicAI · Jun 20, 2024

Introducing Claude 3.5 Sonnet—our most intelligent model yet. This is the first release in our 3.5 model family. Sonnet now outperforms competitor models on key evaluations, at twice the speed of Claude 3 Opus and one-fifth the cost. Try it for free: claude.ai

424

2.0K

7.0K

1.0K

2.5M

Mitchell Wortsman Retweeted

Josh Gardner@jpgard · Jun 19, 2024

Thrilled to share our paper “Large-Scale Transfer Learning for Tabular Data via Language Modeling,” introducing TabuLa-8B: a foundation model for prediction on tabular data. (with Juan C Perdomo + @lschmidt3) 📖 arxiv.org/abs/2406.12031 🌐 huggingface.co/collections/ml… [long🧵]

5.0K

Mitchell Wortsman Retweeted

Vaishaal Shankar@Vaishaal · Jun 18, 2024

I am really excited to introduce DataComp for Language Models (DCLM), our new testbed for controlled dataset experiments aimed at improving language models. 1/x

279

130

119.0K

Mitchell Wortsman Retweeted

Peter J. Liu@peterjliu · Jun 5, 2024

We recently open-sourced a relatively minimal implementation example of Transformer language model training in JAX, called NanoDO. If you stick to vanilla JAX components, the code is relatively straightforward to read -- the model file is <150 lines. We found it useful as a…

280

242

57.0K

Mitchell Wortsman Retweeted

Anthropic@AnthropicAI · May 21, 2024

New Anthropic research paper: Scaling Monosemanticity. The first ever detailed look inside a leading large language model. Read the blog post here: anthropic.com/research/mappi…

552

2.0K

1.0K

750.0K