[models] · · 1 min read
ByteDance Is Training a 10 Trillion-Parameter Model, Leaning Into Scale to Chase Anthropic
The TikTok parent is reportedly pre-training a massive model that could rival Anthropic's Mythos 5, signaling a new phase in the global AI race.
By ByteBulletin Editor · Editor
ByteDance is reportedly in the early stages of training a massive AI model with up to 10 trillion parameters—three times larger than the biggest Chinese model released to date—according to three people with knowledge of the matter. The move positions the TikTok parent as the most ambitious Chinese lab yet in the race to match or exceed the top US models from Anthropic.
The model, being trained by ByteDance's Seed team, is currently in pre-training, a phase that typically lasts three to six months, followed by fine-tuning and potential release. The exact parameter count could change, but the reported scale signals a bet on sheer size as a path to frontier capability.
Industry estimates suggest Anthropic's Mythos 5 has around 8 trillion parameters, and its Fable 5 around 5 trillion. While parameter count sets fundamental capacity limits, capability also depends on data quality and training methods—factors that US labs have long used to maintain their edge.
Get the signal, not the noise.
One short email when it matters. No recaps of recaps.
SOURCES
SHARE
RELATED

[models] ·
Anthropic launches Fable 5.1 and Mythos 5.1 with lower costs and refined safeguards

[research] ·
Hacktron exploits Claude Opus 5 to breach OpenAI

[tooling] ·
Anthropic launches Claude Code Projects for multi-agent workflows

[launches] ·
Anthropic merges Claude chat and Cowork into one interface

[research] ·
Anthropic CEO proposes three-step plan to slow AI development

[research] ·
