ByteDance, the company behind TikTok, is reportedly training an artificial intelligence model that could reach around 10 trillion parameters, according to a Financial Times report citing people familiar with the matter. The details are not independently verified, and ByteDance had not responded to a request for comment. Parameter counts are described as a rough indicator of model scale rather than a direct measure of capability. The report says the project is currently in pre-training, typically lasting three to six months, before moving to fine-tuning and any potential release. The final parameter number is not yet fixed and could change. Industry estimates cited by the FT place Anthropic’s Mythos 5 at about 8 trillion parameters and its guardrailed Fable 5 at about 5 trillion, though Anthropic does not officially disclose parameter counts for these models. Similar uncertainty applies to other attributed figures, including estimates for OpenAI models. At 10 trillion parameters, ByteDance’s model would be more than three times larger than Moonshot AI’s Kimi K3, estimated at 2.8 trillion. The development is framed as part of a broader push by Chinese companies to accelerate frontier AI efforts amid high costs and constraints, including limited access to advanced chips.
ByteDance reportedly trains a ~10 trillion-parameter AI model to rival Anthropic’s Mythos
ByteDance, the company behind TikTok, is reportedly training an artificial intelligence model that could reach around 10 trillion parameters, according to a Financial Times report citing people famili...
- ByteDance is reportedly training an AI model with about 10 trillion parameters, citing people familiar with the matter.
- The work is in pre-training and is expected to last months before possible fine-tuning and release.
- Parameter counts are not confirmed by ByteDance or Anthropic, and are based on industry estimates for Anthropic’s Mythos and Fable models.
- The estimated 10-trillion model would be far larger than Moonshot AI’s Kimi K3 (about 2.8 trillion).
- The report notes parameter count alone does not determine performance; data quality and training methods also matter.
ByteDance is reportedly training an AI model with roughly 10 trillion parameters as it tries to close the gap with leading frontier systems such as Anthropic's Mythos. The model is still in early pre-training, and its eventual performance will depend on more than scale alone, but the project underscores how aggressively Chinese firms are pushing frontier AI despite limits on access to advanced chips. The Next Web reports: The size is itself the statement. At roughly 10 trillion parameters, the model would be more than three times as large as Moonshot's Kimi K3, which sits among the biggest Chinese models today at about 2.8 trillion. [...] Parameter count is not everything, of course. Bigger models are not automatically better, and the industry has learned that data quality, training technique and efficiency often matter as much as raw scale. Even so, committing the compute to train a model this size is a declaration in its own right, a signal that ByteDance wants to compete at the very top rather than ship a capable also-ran. Read more of this story at Slashdot.
3 hours agoByteDance, the Chinese technology group that owns TikTok, is training an artificial intelligence model with as many as 10 trillion parameters, a scale that could approach Anthropic's most advanced Mythos system, the FT reported, citing people familiar with the matter.The report said it could not independently verify the details, and ByteDance had not immediately responded to a request for comment.TikTok Stops Working In US; ByteDance App No Longer Available On Google Play Store, Apple App Store, Claim Reports How the numbers stack up?An AI model's parameter count refers to the numerical settings it learns from training data to recognise patterns, generate responses and carry out tasks, and is commonly used as a rough proxy for scale, though not necessarily for capability.At 10 trillion parameters, ByteDance's model would be more than three times the size of Moonshot AI's Kimi K3, a 2.8 trillion-parameter model that is among the largest released by a Chinese AI lab so far.According to the FT, industry estimates put Anthropic's Mythos 5 at around 8 trillion parameters, with its more widely available, guardrailed sibling Fable 5 estimated at about 5 trillion parameters. Anthropic does not officially disclose parameter counts for Mythos, Fable or any of its models, so these figures remain industry estimates rather than confirmed data, a caveat that also applies to parameter counts attributed to other US models such as OpenAI's GPT-5.5.Still early daysThe ByteDance model is currently in the pre-training phase, a process that typically takes three to six months, after which it would move to fine-tuning before any potential public release, the report said. The exact final parameter count has not been locked in and could change before training concludes.The development comes days after ByteDance founder Zhang Yiming reportedly told employees to avoid relying on AI distillation techniques purely to chase short-term gains, following allegations from a senior US official that Moonshot AI had distilled Anthropic's Fable model while building Kimi K3.The bigger pictureMythos is Anthropic's frontier model class, positioned above its Opus tier, and has drawn attention for strong agentic coding, reasoning and cybersecurity capabilities. Access to Mythos has largely been restricted to trusted partners under Anthropic's Project Glasswing, given concerns about misuse in hacking and vulnerability exploitation, while Fable 5 is the safety-hardened version made more broadly available.ByteDance's push comes as Chinese technology firms accelerate their model release cycles to keep pace with US rivals in an increasingly expensive race to build larger and more capable large language models, while also trying to keep the cost of running them from becoming prohibitive. Several Chinese labs are already working on models in the roughly 5 trillion-parameter range associated with Fable, with ByteDance now positioning itself as one of the most ambitious, aiming closer to Mythos-scale.
10 hours agoCongress considers boosting federal funding for quantum computing
U.S. lawmakers from both political parties are looking to increase federal funding for quantum computing, according to r...
Fenix Flexin Admits AI Was Used on ‘Rubberz’ After Months of Accusations
Fenix Flexin says he used AI tools to create “Rubberz,” ending a period of avoiding or denying allegations tied to the s...
GDIT awarded $1.3 billion contract for Army National Guard network operations and cybersecurity
General Dynamics Information Technology (GDIT) is awarded a $1.3 billion contract to support the U.S. Army National Guar...