BTC $79,807.49 -1.96%
ETH $2,455.27 -2.08%
BNB $720.15 -0.41%
XRP $1.40 -5.15%
SOL $101.70 -3.26%
TRX $0.3313 -0.11%
DOGE $0.0846 -4.53%
ADA $0.2126 -4.52%
BCH $253.16 -1.76%
LINK $11.70 -1.27%
HYPE $84.59 -1.00%
AAVE $130.31 -2.89%
SUI $0.7588 -3.42%
XLM $0.1792 -3.94%
ZEC $1,031.45 +7.26%
BTC $79,807.49 -1.96%
ETH $2,455.27 -2.08%
BNB $720.15 -0.41%
XRP $1.40 -5.15%
SOL $101.70 -3.26%
TRX $0.3313 -0.11%
DOGE $0.0846 -4.53%
ADA $0.2126 -4.52%
BCH $253.16 -1.76%
LINK $11.70 -1.27%
HYPE $84.59 -1.00%
AAVE $130.31 -2.89%
SUI $0.7588 -3.42%
XLM $0.1792 -3.94%
ZEC $1,031.45 +7.26%

parameter

All
Article
Flash

first_img Tencent Hunyuan releases and open-sources Hy4 preview, with a total of 770B parameters and 49B activated

Tencent Hunyuan has released and open-sourced the next-generation large language model Hy4 preview. This model has a total of 770B parameters and 49B active parameters, with a context length exceeding 1M. It demonstrates strong capabilities in real productivity tasks such as coding, office work, and science, firmly placing it in the top tier of open-source models. Hy4 preview significantly expands in model size, context length, and data scale, and enhances real-world performance through high-quality data co-built with Tencent experts in software engineering, gaming, finance, security, and deep collaboration with products like WorkBuddy.In software engineering, the model enhances understanding, planning, debugging, and verification capabilities for long-range development tasks, enabling the construction of complex front-end projects like a Three.js miniature town from scratch. In game development, it supports generating playable prototypes from a single sentence and can complete a full demo in Unity. In smart office applications, it can handle complex financial audits, filtering, analyzing, and delivering from multiple documents. In scientific research, it achieves acceleration in tasks such as molecular dynamics simulations and has initially formed a recursive self-improvement feedback loop.Hy4 preview can be experienced in Tencent products such as WorkBuddy/CodeBuddy domestic and international versions, Yuanbao, ima, and can also be accessed via API calls through Tencent Cloud Tokenhub and OpenRouter. WorkBuddy/CodeBuddy will launch a limited-time free activity for two weeks. Since the reconstruction of the infrastructure, the Hunyuan large model has iterated a major version approximately every two months, continuously optimizing through a preview-first and formal version-following approach.

first_img ByteDance discusses training a model with over 50 trillion parameters, the Seed model team adjusts the architecture

According to LatePost, ByteDance is discussing a large model with training parameters exceeding 50 trillion, surpassing Alibaba's Qwen 3.8-Max (24 trillion) and Moonlight K3 (28 trillion), making it the largest known plan in the country so far. This plan is still in its early stages and does not guarantee a final release. The new model is intended to be led by Xiang Liang, head of Seed Foundation, in collaboration with Shen Ke, who is responsible for the pre-training data of large language models. Seed is reorganizing, dividing responsibilities, and allocating resources based on this.Two weeks ago, ByteDance founder Zhang Yiming held a company-wide meeting with Seed head Wu Yonghui. Zhang reassured the team that training large models is inherently difficult and that it is acceptable to lag behind for a period of time, hoping to aim for the upper limits of intelligence and join the world's top tier. He acknowledged that programming is a key direction at present, advocating for the integration of Volcano Engine, Feishu, and Doubao resources to build computational power and data advantages, while reminding not to be led by a single hot topic. He praised Seedance's differentiated leadership and clearly opposed distillation, believing it is difficult to truly surpass and that AGI barriers should be built from a more fundamental level, stating that the company will continue to increase investment in AI.In the past six months, Seed's multimodal performance has been outstanding, with Seedance 2.0, Seedream, and others driving Volcano Engine MaaS, but the market response to the language model Seed 2.0 has been limited, and its lagging coding capabilities have affected the revenue structure. ByteDance has hired Guo Daye at a high salary to specialize in coding and has consolidated related resources. In the face of the industry's general trend of increasing model sizes, ByteDance hopes to achieve a leapfrog advantage with a larger scale while promoting the elimination of horse racing and breaking down departmental walls to concentrate efforts on tackling challenges.

The dark side of the moon plans to release the Kimi K3 large model soon, with a parameter scale reaching 2 to 3 trillion, closely following the leading teams in the United States

According to the Financial Times, informed sources reveal that the Chinese AI unicorn company Moonshot AI plans to release a new large language model, Kimi K3, in the near future. This model has between 20 trillion to 30 trillion parameters, making it the largest AI model in China by parameter scale, and its performance is expected to surpass the flagship model Claude Opus 4.8 from Anthropic in mainstream benchmark tests (industry speculation suggests its parameter count is around 15 trillion to 20 trillion).Unlike the currently mainstream closed-source and expensive cutting-edge large models in the United States, Kimi K3 will be available as an open-weight model for users to download and modify for free, which may create competitive pressure for leading American labs like OpenAI and Anthropic. Currently, due to the rising service fees for large models in the U.S. (for example, Anthropic has announced a 50% price increase for Opus 4.8 in September), some overseas companies have begun to shift towards using more cost-effective Chinese open-source models.In terms of the capital market, informed sources indicate that Moonshot AI is preparing for a new round of financing, with the latest valuation expected to reach approximately $31.5 billion. Meanwhile, the valuations of other AI giants in China and the U.S. are also rising; DeepSeek is starting a new round of financing with an estimated valuation of about $71 billion, while Anthropic and OpenAI have reached valuations of $965 billion and $852 billion, respectively, in their latest round of financing. In response to the aforementioned release and financing rumors, Moonshot AI has currently declined to comment.
app_icon
ChainCatcher Building the Web3 world with innovations.