SetupAI

Daily AI briefing

This WeekToolsUpdatesSearch繁
繁

Full archive

Every story we have kept, newest first.

Looking for the daily editions? → Past editions

Page 1338 of 1372

February 18, 2026

R2U-Net Hits 0.900 DSC in Brain Tumor Segmentation
Research

R2U-Net Hits 0.900 DSC in Brain Tumor Segmentation

Introduces Attention-Gated Recurrent Residual U-Net (R2U-Net) Triplanar model for glioma segmentation, achieving 0.900 Dice Score on BraTS2021 Whole Tumor. Integrates residual, recurrent, and attention mechanisms for efficiency.

ArXiv AI · 210d ago

Panini: Continual Learning via GSW Memory
Research

Panini: Continual Learning via GSW Memory

Panini proposes a non-parametric continual learning framework for LLMs using Generative Semantic Workspaces (GSW), an entity- and event-aware QA network that consolidates experiences without updating the base model. It outperforms RAG baselines by 5-7% on six QA benchmarks while using 2-30x fewer tokens and reducing unsupported answers.

ArXiv AI · 210d ago

Novel da Costian-Tarskian Ontology Heterogeneity Approach
Research

Novel da Costian-Tarskian Ontology Heterogeneity Approach

This arXiv paper proposes da Costian-Tarskianism, blending Carnapian-Goguenism with da Costa's tolerance principle and Tarski's consequence operators for ontological heterogeneity. It introduces extended consequence systems augmented with ontological axioms and extended development graphs for relating ontologies via morphisms, fibring, and splitting.

ArXiv AI · 210d ago

New Model Quantifies LLM Benchmark Validity
Models

New Model Quantifies LLM Benchmark Validity

Presents structured capabilities model to extract interpretable LLM capabilities from benchmarks, addressing construct validity. Outperforms latent factor models on fit and scaling laws on prediction using OpenLLM Leaderboard data.

ArXiv AI · 210d ago

Memory & Planning Excel in Dynamic Navigation
Research

Memory & Planning Excel in Dynamic Navigation

This arXiv paper explores memory strategies for spatial navigation in non-stationary environments with uncertain sensing in a foraging task. It compares simple to sophisticated agents, finding hybrid architectures with episodic memories and on-the-fly planning most efficient for exploration, search, and path optimization.

ArXiv AI · 210d ago

Hybrid Abstention Boosts LLM Reliability
Models

Hybrid Abstention Boosts LLM Reliability

This arXiv paper introduces an adaptive abstention system for LLMs that dynamically adjusts safety thresholds using contextual signals like domain and user history. It features a multi-dimensional detection architecture with five parallel detectors in a hierarchical cascade, reducing latency and false positives.

ArXiv AI · 210d ago

EduEVAL-DB Dataset for AI Tutor Evaluation
Research

EduEVAL-DB Dataset for AI Tutor Evaluation

EduEVAL-DB introduces a dataset of 854 explanations for 139 ScienceQA questions across K-12 subjects, with one human-teacher and six LLM-simulated teacher explanations. It features a pedagogical risk rubric covering factual correctness, depth, focus, appropriateness, and bias, annotated via semi-automatic expert review.

ArXiv AI · 210d ago

EAA Automates Microscopy with VLM Agents
Research

EAA Automates Microscopy with VLM Agents

Experiment Automation Agents (EAA) is a vision-language-model-driven system that automates complex microscopy workflows in materials characterization. It combines multimodal reasoning, tool actions, and long-term memory for autonomous or user-guided experiments.

ArXiv AI · 210d ago

Common Belief Defies KD4: New Axioms
Research

Common Belief Defies KD4: New Axioms

Contrary to common belief, common belief is not KD4 under KD45 individual beliefs, retaining only D and 4 properties plus shift-reflexivity C(Cφ → φ). The paper proves KD4 extended with this axiom is incomplete, requiring an additional agent-number-dependent axiom.

ArXiv AI · 210d ago

AI Predicts Invoice Dilution with Leakage-Free XGBoost & KAN
Research

AI Predicts Invoice Dilution with Leakage-Free XGBoost & KAN

This ArXiv paper proposes an AI/ML framework to predict invoice dilution in supply chain finance, mitigating non-credit risks and margin losses. It employs leakage-free two-stage XGBoost, Kolmogorov-Arnold Networks (KAN), and ensemble models trained on production data across nine transaction fields.

ArXiv AI · 210d ago

AgriWorld: LLM Agents for Verifiable Agri Reasoning
Research

AgriWorld: LLM Agents for Verifiable Agri Reasoning

Researchers introduce AgriWorld, a Python execution environment with unified tools for geospatial queries, remote-sensing analytics, crop simulations, and agri predictors. Agro-Reflective LLM agent uses an execute-observe-refine loop for multi-turn reasoning over agricultural data.

ArXiv AI · 210d ago

Steinberger's OpenClaw Vision as Constitution
Research

Steinberger's OpenClaw Vision as Constitution

Peter Steinberger left a VISION.md document before joining OpenAI, framing OpenClaw's future less as a roadmap and more as a constitution. The article delivers a line-by-line examination of its contents.

OpenClaw.report · 210d ago

OpenClaw v2026.2.17: 1M Context + Sonnet 4.6
Infrastructure

OpenClaw v2026.2.17: 1M Context + Sonnet 4.6

OpenClaw v2026.2.17 release enables Anthropic's 1M token context window for Opus and Sonnet. It introduces Sonnet 4.6 support alongside extensive updates to iOS, Slack, Telegram, Discord, and cron systems.

OpenClaw.report · 210d ago

Palo Alto CEO: AI Lags in Enterprise
Business

Palo Alto CEO: AI Lags in Enterprise

Palo Alto Networks CEO Nikesh Arora reports minimal enterprise AI adoption, limited mainly to coding assistants. Business use trails consumer adoption by at least two years.

The Register - AI/ML · 210d ago

Qwen-Code v0.10.4: Fixes & Region Support
Coding

Qwen-Code v0.10.4: Fixes & Region Support

Qwen-Code released v0.10.4 with a news banner announcing Qwen3.5-Plus launch, fixes for sandbox user permissions in integration tests, and new support for Coding Plan Global/Intl regions. It also bumps the version from 0.10.3, with full changelog available.

Qwen (GitHub Releases: qwen-code) · 210d ago

Spain Probes X, Meta, TikTok on AI CSAM
Image

Spain Probes X, Meta, TikTok on AI CSAM

Spain's government demands investigation into X, Meta, and TikTok for allegedly using AI to create and spread child sexual abuse material. PM Sanchez accuses platforms of harming children's rights and vows to end their impunity.

IT之家 · 210d ago

YouTube Recovers from Recommendation Outage
Infrastructure

YouTube Recovers from Recommendation Outage

YouTube resolved a brief global outage caused by a recommendation system failure that prevented videos from appearing. The issue affected all platforms including YouTube.com, apps, Music, Kids, and TV.

36氪 · 210d ago

GitHub Unveils Copilot CLI Command Cheat Sheet
Coding

GitHub Unveils Copilot CLI Command Cheat Sheet

GitHub has compiled and explained slash commands for GitHub Copilot CLI in an official blog post. Developers can execute quick, repeatable actions in the terminal without switching to editors or web UI.

ITmedia AI+ (日本) · 210d ago

Gartner: Under 20 Humanoids in Production by 2028
Research

Gartner: Under 20 Humanoids in Production by 2028

Gartner predicts fewer than 20 companies will deploy humanoid robots in full production by 2028. The forecast focuses on manufacturing and supply chain sectors amid physical AI hype.

ITmedia AI+ (日本) · 210d ago

AI Adopted by Billions in One Chinese Spring Festival
Business

AI Adopted by Billions in One Chinese Spring Festival

Chinese AI apps Qianwen, Doubao, and Yuanbao exploded during 2026 Spring Festival via red envelope campaigns, logging billions of interactions and onboarding over 130 million new users, including elderly and lower-tier city residents. This achieved unprecedented adoption speed, faster than smartphones (5 years) or mobile payments (3 years).

虎嗅 · 210d ago

SetupAIInfrastructure
Infrastructure

Snapdragon Chipsets Show 71-93% INT8 Accuracy Variance

Same INT8 ONNX model tested on 5 Snapdragon chipsets yields accuracy from 93% (8 Gen 3) to 71% (4 Gen 2), vs 94% cloud. Causes: NPU INT8 rounding differences, operator fusion variations, CPU fallbacks on low-end chips.

Reddit r/MachineLearning · 210d ago

Galaxy Star Brain Enables Real Robot Deployment
Applications

Galaxy Star Brain Enables Real Robot Deployment

Galaxy Universal transitions robots from stage performances to practical on-the-job use via its end-to-end large model, Galaxy Star Brain. A capable working robot debuted at this year's Spring Festival Gala.

量子位 · 210d ago

Tesla Avoids CA Sales Ban on FSD Marketing
Applications

Tesla Avoids CA Sales Ban on FSD Marketing

California DMV confirms Tesla complied with marketing rules for Autopilot and Full Self-Driving, avoiding a 30-day sales ban. This follows a December judge's ruling on exaggerated claims, with Tesla given 90 days for corrections.

36氪 · 210d ago

California Probes xAI Grok Explicit Images
Image

California Probes xAI Grok Explicit Images

California AG Rob Bonta is launching an AI accountability program while investigating xAI's Grok for generating explicit pornographic images without consent, including potentially underage content. The office issued a cease-and-desist order last month amid global scrutiny.

36氪 · 210d ago

Qwen SDK TypeScript v0.1.5-preview.2 Released
Coding

Qwen SDK TypeScript v0.1.5-preview.2 Released

Qwen released SDK TypeScript v0.1.5-preview.2, bundling CLI v0.10.2 with fixes for authentication, logging, and extension issues. New features include experimental skills settings, redesigned CLI UI, and removal of tiktoken dependency.

Qwen (GitHub Releases: qwen-code) · 210d ago

Chrome AI Agent Real-World Test
Applications

Chrome AI Agent Real-World Test

Tests Google's Auto Browse, turning Chrome into an AI-agentic browser. Evaluates performance on shopping, research, and emailing tasks.

ZDNet AI · 210d ago

DeepSeek Tests 1M-Context Model
Models

DeepSeek Tests 1M-Context Model

DeepSeek has begun testing a new long-context model supporting 1 million tokens in its web and app versions since February 13. This has fueled industry speculation about a major Lunar New Year release.

Pandaily · 210d ago

Moonshot AI $700M Raise at $10B+ Valuation
Business

Moonshot AI $700M Raise at $10B+ Valuation

Moonshot AI is closing an oversubscribed $700M+ funding round led by Alibaba, Tencent, and new investor Kaihui Fund. A new round has started at $10-12B valuation with European fund interest.

36氪 · 210d ago

AWS $100M Credits for Federal AI
Infrastructure

AWS $100M Credits for Federal AI

AWS launches accelerator initiatives offering $100M in credits to U.S. federal agencies for cloud and AI services.

GeekWire · 210d ago

Philosophers Guide Gen AI Use
Research

Philosophers Guide Gen AI Use

Senior Google engineer applies Aristotle and Socrates to generative AI. Stresses AI should foster thinking skills, not replace them.

ZDNet AI · 210d ago

11337133813391372
Page 1338 of 1372
Back to home
SetupAI

A bilingual daily AI briefing — ten stories a day, each with a deep insight.

Takedown / opt-out: copyright@setupai.uk

© 2026 SetupAI

This WeekToolsUpdatesAboutPrivacyTermsRSS