RT 1884877516478763140
Original · 1884440764677251515 · Alexandr Wang @alexandr_wang · · en X

What does DeepSeek R1 & v3 mean for LLM data?

Contrary to some lazy takes I’ve seen, DeepSeek R1 was trained on a shit ton of human-generated data—in fact, the DeepSeek models are setting records for the disclosed amount of post-training data for open-source models:

- 600,000 https://t.co/CFhqtaflQ4

RT @alexandr_wang: What does DeepSeek R1 & v3 mean for LLM data?

Contrary to some lazy takes I’ve seen, DeepSeek R1 was trained on a shit…

Fuente verbatim: corpus/posts/1884877516478763140.md · en X · acto Enero 2025
Escrivivir · Scriptorium Skins · Animus Iocandi · Aleph Cero · F.A.R.O. · Material transmedia para agentes del juego ARG · AIGPL · Repositorio