SetupAI

Daily AI briefing

This WeekToolsUpdatesSearch繁
繁

Full archive

Every story we have kept, newest first.

Looking for the daily editions? → Past editions

Page 1330 of 1372

February 19, 2026

SetupAIBusiness
Business

Macron Prioritizes Youth AI Curbs in G-7

French President Emmanuel Macron announced protecting children from harmful effects of social media and AI as a key priority for France's G-7 presidency. He urged countries like India to back these measures.

Bloomberg Technology · 209d ago

SetupAIBusiness
Business

Fractal CEO Discusses AI Fears on IPO

Srikanth Velamakanni, CEO of India's first AI unicorn Fractal Analytics, shared views on AI fears and their impact on the company's recent IPO. Fractal raised $313 million last week.

Bloomberg Technology · 209d ago

ByteDance Music App Eyes NetEase Overtake
Business

ByteDance Music App Eyes NetEase Overtake

ByteDance's 汽水音乐 has surged to 1.4B MAU, nearing NetEase Cloud Music's 1.47B amid ByteDance's AI focus. It leverages 抖音 traffic (82% users) and bundled copyrights for 50M songs, plus AI features like singer 大头针.

虎嗅 · 209d ago

NVIDIA Teases Unprecedented Chips at GTC
Infrastructure

NVIDIA Teases Unprecedented Chips at GTC

NVIDIA CEO Jensen Huang revealed in an interview that the company will launch several globally unprecedented new chips at GTC 2026. The keynote is scheduled for March 15 in San Jose, California, focusing on the new era of AI infrastructure competition.

36氪 · 209d ago

Global Recession: China Must Set AI-Era Agenda
Business

Global Recession: China Must Set AI-Era Agenda

Amid global economic stagnation with G7 growth under 1.2%, rising defaults, and AI's job displacement despite limited productivity gains, China should proactively set international agendas. Warns of prolonged recession, youth unemployment surge to 25% NEET rate, and AI stock bubbles risking crisis.

虎嗅 · 209d ago

SetupAIBusiness
Business

ServiceNow Flags AI Software Shakeup

ServiceNow COO Amit Zavery warns of software industry consolidation during AI transition. Firms failing to transform for AI adoption risk failure, per Bloomberg TV interview.

Bloomberg Technology · 209d ago

X Algo Pushes Conservative Content
Research

X Algo Pushes Conservative Content

A Nature study reveals X's 'For You' feed algorithm systematically prioritizes conservative political content and activists over liberal views and news media. This bias not only alters visible content but shifts users' political leanings toward conservatism over weeks.

cnBeta (Full RSS) · 209d ago

Microsoft Tests Ask Copilot in Windows 11 Taskbar
Applications

Microsoft Tests Ask Copilot in Windows 11 Taskbar

Microsoft is testing new AI features in Windows 11, including an 'Ask Copilot' entry in the taskbar and deep integration of Microsoft 365 Copilot into File Explorer. These enhancements aim to boost productivity without altering user habits.

cnBeta (Full RSS) · 209d ago

Musk Predicts AI Binary Coding by 2026
Coding

Musk Predicts AI Binary Coding by 2026

Elon Musk predicts in a recent video that by the end of 2026, AI will directly write binary code, greatly reducing human reliance on programming languages. This could lead to full automation in the programming industry, potentially eliminating traditional programmers.

cnBeta (Full RSS) · 209d ago

Verifiable Semantics for Agent Communication
Research

Verifiable Semantics for Agent Communication

Proposes a certification protocol using stimulus-meaning model to verify shared term understanding in multi-agent systems via tests on observable events. Core-guarded reasoning limits agents to certified terms, provably bounding disagreement.

ArXiv AI · 209d ago

Science of AI Agent Reliability
Research

Science of AI Agent Reliability

AI agents excel on benchmarks but fail in practice due to single-metric evaluations ignoring consistency, robustness, predictability, and safety. This arXiv paper proposes 12 concrete metrics across these four dimensions, grounded in safety-critical engineering.

ArXiv AI · 209d ago

Proxy State Eval Scales LLM Agent Benchmarks
Research

Proxy State Eval Scales LLM Agent Benchmarks

Proxy State-Based Evaluation introduces an LLM-driven simulation framework for benchmarking multi-turn tool-calling agents, avoiding costly deterministic backends. It uses scenarios to define goals and states, with LLM trackers inferring proxy states from traces for verification.

ArXiv AI · 209d ago

PAHF: Personalized Agents from Human Feedback
Research

PAHF: Personalized Agents from Human Feedback

PAHF is a framework for continual personalization of AI agents, learning online from live human interactions via explicit per-user memory. It uses a three-step loop: pre-action clarification, preference-grounded actions, and post-action feedback for memory updates.

ArXiv AI · 209d ago

Mirror Tops GPT-5 on Endo Board Exam
Research

Mirror Tops GPT-5 on Endo Board Exam

January Mirror, an evidence-grounded clinical reasoning system, scored 87.5% on a 120-question 2025 endocrinology board-style exam, outperforming human experts (62.3%) and frontier LLMs like GPT-5.2 (74.6%). It excelled on the hardest questions (76.7% accuracy) under closed-evidence constraints without web access.

ArXiv AI · 209d ago

In-Context Inference Enables Multi-Agent Cooperation
Research

In-Context Inference Enables Multi-Agent Cooperation

Researchers demonstrate that sequence models' in-context learning induces cooperation in multi-agent RL without hardcoded co-player assumptions or timescale separation. Training against diverse co-players leads to best-response strategies on intra-episode timescales.

ArXiv AI · 209d ago

GPSBench Tests LLM GPS Reasoning
Research

GPSBench Tests LLM GPS Reasoning

Researchers launch GPSBench, a 57,800-sample dataset across 17 tasks to probe LLMs' geospatial reasoning without tools. 14 SOTA LLMs show reliability in geographic knowledge but struggle with geometric computations like distance and bearing.

ArXiv AI · 209d ago

FoT: Dynamic LLM Reasoning Optimizer
Research

FoT: Dynamic LLM Reasoning Optimizer

FoT introduces a general-purpose framework for dynamic reasoning schemes in LLMs, overcoming static structures in Chain of Thought, Tree of Thoughts, and Graph of Thoughts. It features hyperparameter tuning, prompt optimization, parallel execution, and caching for better performance.

ArXiv AI · 209d ago

Corecraft RL Env Trains Generalizable Agents
Research

Corecraft RL Env Trains Generalizable Agents

Surge AI launches Corecraft, the first high-fidelity RL environment in EnterpriseGym, simulating enterprise customer support with 2,500+ entities and 23 tools. Training GLM 4.6 via GRPO improves task pass@1 from 25% to 37% on held-out tasks, with gains transferring to BFCL (+4.5%), τ²-Bench Retail (+7.4%), and Toolathlon (+6.8%).

ArXiv AI · 209d ago

CaR Enables Efficient Neural Routing Constraints
Research

CaR Enables Efficient Neural Routing Constraints

Neural solvers excel in simple routing but falter on complex constraints. CaR introduces the first general framework using explicit learning-based feasibility refinement and joint training to generate diverse solutions for lightweight improvement.

ArXiv AI · 209d ago

CAFE: Causal Multi-Agent AFE Breakthrough
Research

CAFE: Causal Multi-Agent AFE Breakthrough

CAFE reformulates automated feature engineering as a causally-guided sequential decision process using causal discovery for soft priors and multi-agent RL for construction. It outperforms baselines by up to 7% on 15 benchmarks and reduces performance drops 4x under covariate shifts.

ArXiv AI · 209d ago

Boosting LLM Feedback-Driven In-Context Learning
Models

Boosting LLM Feedback-Driven In-Context Learning

Proposes a trainable framework for interactive in-context learning using multi-turn feedback from information asymmetry on verifiable tasks. Trained smaller models nearly match performance of 10x larger models and generalize to coding, puzzles, and mazes.

ArXiv AI · 209d ago

AI Long-Term Memory: Store-First Paradigm
Research

AI Long-Term Memory: Store-First Paradigm

This arXiv paper proposes a 'store then on-demand extract' approach for AI memory to retain raw experiences and avoid information loss from the dominant 'extract then store' method. It also explores deriving deeper insights from probabilistic experiences and improving efficiency via shared storage.

ArXiv AI · 209d ago

Agentic AI Fails Paradoxically on Rare Symptoms
Research

Agentic AI Fails Paradoxically on Rare Symptoms

Autonomous agentic workflows exhibit optimization instability, where iterative self-improvement degrades classifier performance, especially for low-prevalence clinical symptoms like Long COVID brain fog (3%). Using the open-source Pythia framework, validation sensitivity oscillated wildly between 1.0 and 0.0.

ArXiv AI · 209d ago

Agent Skills Boost SLMs for Industry
Research

Agent Skills Boost SLMs for Industry

ArXiv paper defines Agent Skill framework mathematically and evaluates its benefits for small language models (SLMs) in industrial settings with data security constraints. Moderate SLMs (12B-30B params) show substantial gains in accuracy and reduced hallucinations; 80B code-specialized variants match proprietary models with better GPU efficiency.

ArXiv AI · 209d ago

Chunwan Sparks China Humanoid Robot Boom
Business

Chunwan Sparks China Humanoid Robot Boom

Morgan Stanley identifies 2026 as a pivotal inflection point for China's humanoid robot market, mirroring the 2019-2020 NEV surge. IDC predicts application scenarios will triple by 2026.

cnBeta (Full RSS) · 209d ago

Tesla FSD Supervised Hits 8B Miles
Research

Tesla FSD Supervised Hits 8B Miles

Tesla announced FSD Supervised cumulative mileage exceeds 8 billion miles, up from 7 billion in December 2024. This data accelerates training for unsupervised Full Self-Driving.

IT之家 · 209d ago

Ant's Trillion-Param Open Model Excels in EQ & Agents
Models

Ant's Trillion-Param Open Model Excels in EQ & Agents

Ant Group launches a trillion-parameter open-source model superior in human understanding and execution. It excels in emotional intelligence and agent combat power.

量子位 · 209d ago

AI & Robots Flock to One Spot After Chunwan
Applications

AI & Robots Flock to One Spot After Chunwan

After China's Spring Festival Gala (Chunwan), various AI systems and robots have converged to a single location. This place has captured the 'stray' elements from the gala, hinted at with a doge meme.

量子位 · 209d ago

AI Robots Rush to Bilibili Post-Chunwan
Business

AI Robots Rush to Bilibili Post-Chunwan

After 2026 Chunwan, AI products like Tencent Yuanbao and robots from Songyan Dynamics and Unitree shifted to Bilibili for live streams and interactions to sustain hype. Bilibili's interest rooms and AI community enabled peaks of 1M+ viewers.

虎嗅 · 209d ago

Giants Splash 45B on 2026 CNY AI Payments
Business

Giants Splash 45B on 2026 CNY AI Payments

Chinese giants invest over 45B RMB in AI voice payment subsidies during 2026 Spring Festival to seize user intents via agents. ByteDance's Doubao AI phone was blocked by rivals' ecosystems, while Alibaba's Qwen thrives in closed-loop integration across its services.

虎嗅 · 209d ago

11329133013311372
Page 1330 of 1372
Back to home
SetupAI

A bilingual daily AI briefing — ten stories a day, each with a deep insight.

Takedown / opt-out: copyright@setupai.uk

© 2026 SetupAI

This WeekToolsUpdatesAboutPrivacyTermsRSS