{"id":55614,"date":"2026-08-08T15:51:13","date_gmt":"2026-08-08T07:51:13","guid":{"rendered":"https:\/\/www.1ai.net\/?p=55614"},"modified":"2026-08-08T15:51:13","modified_gmt":"2026-08-08T07:51:13","slug":"%e6%b6%88%e6%81%af%e7%a7%b0%e5%ad%97%e8%8a%82%e8%b7%b3%e5%8a%a8%e5%bc%80%e5%a7%8b%e9%a2%84%e8%ae%ad%e7%bb%83-10-%e4%b8%87%e4%ba%bf%e5%8f%82%e6%95%b0%e5%a4%a7%e6%a8%a1%e5%9e%8b%ef%bc%8c%e8%a7%84","status":"publish","type":"post","link":"https:\/\/www.1ai.net\/en\/55614.html","title":{"rendered":"Message bytes beats begin pre-training 10 trillion-billion-billion-parameter large model, size or near the Anthropic flagship"},"content":{"rendered":"<p>On August 8th, according to the Financial Times on August 7th<a href=\"https:\/\/www.1ai.net\/en\/tag\/%e5%ad%97%e8%8a%82%e8%b7%b3%e5%8a%a8\" title=\"[View articles tagged with [bytejump]]\" target=\"_blank\" >ByteDance<\/a>We're training up to 10 trillion<a href=\"https:\/\/www.1ai.net\/en\/tag\/%e5%a4%a7%e6%a8%a1%e5%9e%8b\" title=\"[View articles tagged with [large models]]\" target=\"_blank\" >Large Model<\/a>It is still in the early stages of pre-training. According to three informed sources, this size will be more than three times the size of the largest model published in China, Kimi K3, the dark side of the moon, and close to the volume of Mythos 5, the most advanced model of Anthropic, estimated by industry at around 8 trillion parameters\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55615\" title=\"c40f7132j00tjfx300tcd000r300f9m\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/c40f7132j00tjfxs300tcd000r300f9m.jpg\" alt=\"c40f7132j00tjfx300tcd000r300f9m\" width=\"975\" height=\"549\" \/><\/p>\n<p>It was reported that the project was led by a byte by the head of the Seed Foundation, reflecting the idea that CEOs were promoting original research rather than imitating competitors ' models. The pre-training phase usually takes three to six months, after which, if progress is successful, the fine-tuning and eventual release of the exact parameters will take a later stage to determine\u3002<\/p>\n<p>As previously reported in LateLatePost, byte beats are discussing the training of a model of parameters in excess of 5 trillion, more than Ali Qwen 3.8-Max (2.4 trillion parameters) and the dark side of the moon, Kimi K3 (2.8 trillion parameters), which is the largest training programme of known parameters in the country\u3002<\/p>","protected":false},"excerpt":{"rendered":"<p>On 8 August, according to the Financial Times on 7 August, byte beats are training a large model of up to 10 trillion parameters and are still in the early stages of pre-training. According to three informed sources, this size will be more than three times the size of the largest model published in China, Kimi K3, the dark side of the moon, and close to the volume of Mythos 5, the most advanced model of Anthropic, estimated by industry at around 8 trillion parameters. It was reported that the project was led by a byte by the head of the Seed Foundation, reflecting the idea that CEOs were promoting original research rather than imitating competing models. The pre-training phase usually takes between three and six months, after which the fine-tuning and eventual release of the exact parameters will take place at a later stage<\/p>","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[146],"tags":[216,548],"collection":[],"class_list":["post-55614","post","type-post","status-publish","format-standard","hentry","category-news","tag-216","tag-548"],"acf":[],"_links":{"self":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/55614","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/comments?post=55614"}],"version-history":[{"count":0,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/55614\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/media?parent=55614"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/categories?post=55614"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/tags?post=55614"},{"taxonomy":"collection","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/collection?post=55614"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}