Latest Large Language Models News & Updates
New frontier models, context windows, reasoning breakthroughs and benchmark drama — the core technology of this era, tracked daily.
91 articles
Alibaba announces Qwen3.8 model with 2.4T parameters
Alibaba has unveiled the Qwen3.8 preview model, featuring 2.4 trillion parameters. It is currently available for testing via the Token Plan platform and is described as one of the most powerful models available.
Moonshot AI releases 2.8T parameter Kimi K3 model
Moonshot AI has launched Kimi K3, a massive 2.8-trillion-parameter model. The release aims to capture developer loyalty in a highly competitive market.
OpenAI and StepFun challenge the AI hardware market
OpenAI and StepFun are aggressively entering the AI hardware space, signaling a new wave of competition. This move challenges existing players and highlights the integration of LLMs into physical devices.
Zhipu AI Sees 4.7B HKD Net Inflow via Southbound Funds
Zhipu AI led the list of southbound fund net purchases today, attracting 4.737 billion HKD. This highlights significant capital interest in the AI company within the Hong Kong stock market.
Tencent Launches Hunyuan Hy3 Model
Tencent has officially released the Hunyuan Hy3 model, which offers superior intelligence compared to its preview version. It is now integrated into multiple business applications and available via Tencent Cloud's TokenHub API.
Big tech firms race to automate college admission counseling
Major tech companies like Alibaba, Tencent, and Baidu are deploying AI agents to fill the market gap in college admission counseling, aiming to replace human experts with data-driven algorithms.
Github Copilot integrates open-source model Kimi K2.7
Moonshot AI announced that Github Copilot has integrated the Kimi K2.7 open-source model. This marks the first time an open-source model has been incorporated into the Copilot ecosystem.
Huawei Open-Sources 92B-Parameter openPangu-2.0-Flash Model
Huawei has officially released the openPangu-2.0-Flash model, featuring 92 billion parameters. This release is a strategic move to strengthen the company's AI ecosystem and provide developers with high-performance open-source options.
Huawei Open-Sources openPangu-2.0-Flash Model
Huawei has officially launched the open-source version of its openPangu-2.0-Flash model. Additional components of the Pangu model series are scheduled for release starting June 30.
Huawei open-sources 92B parameter openPangu-2.0-Flash model
Huawei has officially open-sourced the openPangu-2.0-Flash model, featuring 92 billion parameters. This release is part of a broader strategy to support the Ascend AI ecosystem.
Alibaba Launches Qwen-Robot Embodied AI Model Series
Alibaba has introduced the Qwen-Robot series, its first large model specifically designed for embodied AI. The model enables robots to perceive, reason, and navigate in physical environments.
Zhipu AI Open-Sources GLM-5.2 With 1M Token Context
Zhipu AI has released GLM-5.2 under an MIT license, featuring a massive 1 million token context window. This release is positioned as a strategic response to US export restrictions affecting access to Anthropic's models.
Claude Mythos 5 Released: 50M Lines of Code in 1 Day
Anthropic has launched Claude Mythos 5, a new model capable of handling massive-scale coding tasks. It claims to process 50 million lines of code within a single day.
Apple to Unveil Major Siri AI Overhaul
Apple is expected to announce a significant overhaul of Siri at the Worldwide Developers Conference. The update will feature a chatbot-style interface and deeper integration across Apple devices.
DeepSeek Nears $7.4 Billion Funding for AI Ambitions
DeepSeek is finalizing a $7.4 billion funding round, including support from Tencent. This capital is intended to scale their AI capabilities to compete with major US-based models.
Anthropic to Release Mythos-class Opus 4.8 Model
Anthropic announced the upcoming release of its new 'Mythos-class' model, Opus 4.8. The model will be available to all customers within the next few weeks after implementing enhanced safety measures.
xAI's Grok V9-Medium Model Training Completed
Elon Musk announced that the Grok V9-Medium (1.5T) foundation model has finished training. The model incorporates significant Cursor data and is expected to be released in 2-3 weeks following reinforcement learning.
Cohere releases Apache 2.0 open-weight model Command A+
Cohere has launched Command A+, a 218-billion-parameter sparse MoE model released under an Apache 2.0 license. It is optimized for complex reasoning and agentic workflows while maintaining high efficiency through advanced quantization.
Ant Group Bailing Ring-2.6-1T Enhances Agent Capabilities
Ant Group has released the Bailing Ring-2.6-1T open-source model, featuring significantly enhanced agent execution capabilities. The model achieved a score of 95.83 on the AIME 26 benchmark.
Amazon integrates Alexa Plus AI into shopping experience
Amazon is replacing its Rufus AI assistant with 'Alexa for Shopping,' powered by the new Alexa Plus LLM. The assistant is now integrated directly into the Amazon.com search bar to provide conversational answers to user queries.
PrismML unveils viable 1-bit LLMs
PrismML announced 1-bit Bonsai, touted as the first commercially viable 1-bit Large Language Models. This breakthrough promises extreme quantization for efficient inference. Details available via linked announcement.
Capital Migration in China's AI Sector
A deep dive into the funding trends of WAIC exhibitors, highlighting the shift toward large-scale capital concentration in LLMs and embodied AI.
Linus Torvalds Clarifies Linux Stance on AI Integration
Linus Torvalds issued a statement clarifying that the Linux kernel project is not 'anti-AI'. He pushed back against developers opposing the use of AI and LLMs in kernel development processes.
The Structural Crisis of AI Companion Robots
The AI companion robot market is facing a 'moat' crisis as technical barriers disappear due to LLMs, leading to product homogenization. High inference costs and a lack of true emotional intelligence create a structural dilemma for startups.
Tencent Launches WorkBuddy AI Agent for WeChat Integration
Tencent has launched WorkBuddy, a local AI coding agent powered by the Hunyuan Hy3 model. It integrates directly with WeChat to streamline file management, task automation, and execution for Chinese users.
Second-gen Doubao phone enters the market
The second generation of the Doubao phone is preparing for market entry, raising questions about how major smartphone manufacturers will respond to this AI-native hardware competition.
MiniMax and Zhipu AI valuation gap widens significantly
The valuation gap between AI unicorns MiniMax and Zhipu AI has reached 600 billion HKD within six months. This reflects shifting investor focus in the competitive Chinese LLM landscape.
SpaceXAI releases coding-focused Grok 4.5 model
SpaceXAI has launched Grok 4.5, a new flagship model optimized for coding tasks. It is now accessible via Grok Build, Cursor, and the official API console.
Zhipu AI's market volatility and the AI bubble
Zhipu AI experienced extreme market volatility due to IPO rumors and lock-up expirations. The article analyzes the valuation gap between Chinese AI firms and global leaders like OpenAI.
Meituan Officially Open-Sources LongCat-2.0 Model
Meituan has officially open-sourced the LongCat-2.0 model, including weights, inference engines, and technical documentation. Major domestic chip manufacturers have already completed inference adaptation for this model.