
Qwen3.5-Plus Breaks Cost-Performance Ceiling
Alibaba's Qwen3.5-Plus launches as top open-source model in multimodal, reasoning, coding, and agents. Priced at 0.8 yuan per million tokens, it's 18x cheaper than Gemini 3 Pro.
机器之心 · 211d ago
Every story we have kept, newest first.
Looking for the daily editions? → Past editions
Page 1347 of 1372

Alibaba's Qwen3.5-Plus launches as top open-source model in multimodal, reasoning, coding, and agents. Priced at 0.8 yuan per million tokens, it's 18x cheaper than Gemini 3 Pro.
机器之心 · 211d ago

Anthropic updated Claude Code's progress output to conceal names of files being read, written, or edited. Developers strongly oppose the change, demanding visibility into file access for safety and oversight.
The Register - AI/ML · 211d ago

Anthropic updated Claude Code to hide file names in progress outputs during read/write/edit operations. Developers strongly oppose the change, demanding visibility into accessed files for better control.
The Register - AI/ML · 211d ago

ByteDance's Seed 2.0 debuts at #6 text and #3 vision on LM Arena, highest for Chinese models. Native multimodal excels in math, vision perception, reasoning, and agents, matching Gemini 3 Pro and GPT 5.2.
机器之心 · 211d ago

ByteDance's Seed 2.0 debuts at #6 text, #3 vision on LMArena, leading domestic models. Excels in math, vision perception, reasoning, and agents, matching Gemini 3 Pro.
机器之心 · 211d ago

ByteDance continues its frontier AI push with the launch of Seed 2.0. This release advances their capabilities in cutting-edge AI models.
The Neuron · 211d ago
Qwen Code released v0.10.3 with minor updates including a version bump to 0.10.2 by the CI bot. Documentation for settings.json was improved with quick setup examples, and the README was updated.
Qwen (GitHub Releases: qwen-code) · 211d ago
OpenAI hires Peter Steinberger, developer of AI agent OpenClaw, to advance personal agents. OpenClaw remains open source under a foundation.
Engadget · 211d ago
OpenAI has hired Peter Steinberger, creator of AI agent OpenClaw, to build next-generation personal agents, as announced by Sam Altman. OpenClaw, with 196k GitHub stars, will move to a foundation and stay open source.
Engadget · 211d ago

2026 CCTV Spring Festival Gala announces over 20 official partners. Internet firms, AI robots, and smart mobility lead as tech dominates sponsorships.
cnBeta (Full RSS) · 211d ago

ByteDance vows to restrain its AI video generator Seedance 2.0 after legal threats from Disney and media backlash. Released last week, the tool creates realistic celebrity videos from text prompts and has gone viral.
The Guardian Technology · 211d ago

ByteDance will restrict Seedance 2.0 AI video generator after Disney's legal threats. The tool creates realistic clips of stars like Tom Cruise from text prompts, going viral.
The Guardian Technology · 211d ago

OpenAI secured top talent Steinberger in a bidding war among major AI labs. Unlike competitors, OpenAI avoided aggressive legal tactics.
OpenClaw.report · 211d ago

Every major AI lab competed for Steinberger, but OpenAI won him over. Unlike others, OpenAI avoided sending lawyers in the process.
OpenClaw.report · 211d ago

Major AI labs vied for talent Steinberger. OpenAI secured him uniquely by avoiding legal tactics.
OpenClaw.report · 211d ago

XPeng launches L4 road tests with GX model in Guangzhou without safety drivers, boasting 3000TOPS compute and pure vision tech like Tesla FSD. He Xiaopeng argues L3 is a regulatory trap, pushing industry to leap to L4 amid commercialization successes.
Huxiu (虎嗅) · 211d ago

Google's AI Overviews omit prominent safety disclaimers for medical queries, risking user harm. The feature provides summaries above search results without initial warnings to consult professionals.
The Guardian Technology · 211d ago

The creator behind OpenClaw has been recruited by Sam Altman. This move signals an impending AI Agent competition.
钛媒体 · 211d ago

Bendigo Bank has reduced costs and time in software development efforts. Digital onboarding experience serves as their first productivity 'trophy'.
iTNews Australia · 211d ago

OpenClaw version 2026.2.15 is now available. It introduces sub-sub-agents and Discord Components v2.
OpenClaw.report · 211d ago

OpenClaw version 2026.2.15 has been released. Key additions include sub-sub-agents and Discord Components v2.
OpenClaw.report · 211d ago

OpenClaw has released version 2026.2.15. Key additions include sub-sub-agents and Discord Components v2.
OpenClaw.report · 211d ago

UK Prime Minister Starmer vows to battle AI chatbots as with Grok. New government plans ensure no online platform gets a free pass on children's internet safety.
BBC Technology · 211d ago

GitHub launches AI agent eliminating 3-hour miscellaneous tasks for developers. Boosts efficiency by 10 times.
钛媒体 · 211d ago

GitHub launches an AI agent that eliminates 3 hours of miscellaneous developer tasks, increasing efficiency by 10 times. This represents a pivotal shift for GitHub from a code repository to an intelligent collaboration platform.
钛媒体 · 211d ago
New ArXiv paper quantifies information optimal policies encode about environments. Proves mutual information of exactly n log m bits in Controlled Markov Processes.
ArXiv AI · 211d ago
GT-HarmBench introduces 2,009 high-stakes multi-agent scenarios using game theory like Prisoner's Dilemma to benchmark AI safety risks. Frontier models select socially beneficial actions only 62% of the time, often leading to harm.
ArXiv AI · 211d ago
Entity State Tuning (EST) introduces persistent entity states to TKG forecasters, overcoming stateless methods' long-term dependency issues. It uses a closed-loop design with topology-aware perception and dual-track evolution.
ArXiv AI · 211d ago
BrowseComp-V³ is a new benchmark with 300 challenging questions for evaluating multimodal browsing agents on deep multi-hop reasoning across text and visuals. It features subgoal-driven process evaluation and publicly searchable evidence for reproducibility.
ArXiv AI · 211d ago
This paper introduces a theoretical framework that reimagines AI benchmarking as a multilayer, adaptive network connecting evaluation metrics, model components, and stakeholder priorities through weighted interactions. It embeds human tradeoffs using conjoint-derived utilities and a human-in-the-loop update rule, allowing benchmarks to evolve dynamically while maintaining stability.
ArXiv AI · 211d ago