SetupAI

Daily AI briefing

This WeekToolsUpdatesSearch繁
繁

Full archive

Every story we have kept, newest first.

Looking for the daily editions? → Past editions

Page 1343 of 1372

February 17, 2026

Kimi Raises $1.2B, Valuation Hits $10B+
Business

Kimi Raises $1.2B, Valuation Hits $10B+

Moonshot AI's Kimi secured over $700M in a new round led by Alibaba, Tencent, and others, just over a month after a $500M raise. Valuation now at $10-12B, doubling previously.

36氪 · 211d ago

EU Parliament Bans AI on Work Devices
data-privacy

EU Parliament Bans AI on Work Devices

European Parliament disabled AI features on official devices via email to all MEPs and staff. Concerns focus on cybersecurity risks from cloud-based processing.

cnBeta (Full RSS) · 211d ago

Chunwan Robots Excel, Nearing Market Launch
Applications

Chunwan Robots Excel, Nearing Market Launch

Humanoid robots dazzled at 2026 Chunwan Gala with jokes, skits, sword dancing, street dance, martial arts, and natural human interactions. Performances far surpassed last year's shaky debut.

cnBeta (Full RSS) · 211d ago

Robots First Replace Humans in Spring Gala
Applications

Robots First Replace Humans in Spring Gala

2026 CCTV Spring Gala exploded with embodied AI and robots, from stage design to visuals and performances. Following 2025 industry boom, it became the most tech-heavy edition ever.

cnBeta (Full RSS) · 211d ago

Unity AI Beta Generates Games via Text Prompts
Coding

Unity AI Beta Generates Games via Text Prompts

Unity plans to launch a Unity AI beta at GDC in March, enabling natural language prompts to create full casual games without coding. The tool leverages top LLMs and Unity's platform context for seamless prototyping to production.

IT之家 · 211d ago

Unity AI Beta Generates Games via Natural Language
Coding

Unity AI Beta Generates Games via Natural Language

Unity plans to launch a Unity AI beta at GDC in March, enabling natural language prompts to create complete casual games without coding. The tool integrates top LLMs and custom models to streamline prototyping to production.

IT之家 · 211d ago

SpaceX Enters Pentagon AI Drone Race
Research

SpaceX Enters Pentagon AI Drone Race

SpaceX and its xAI subsidiary are competing in a classified Pentagon program for voice-controlled autonomous drone swarms. Elon Musk's recent company merger thrusts them into AI-enabled weapons development.

cnBeta (Full RSS) · 211d ago

Micron's $200B Factory Push Breaks AI Memory Bottleneck
Infrastructure

Micron's $200B Factory Push Breaks AI Memory Bottleneck

Micron Technology plans $200 billion investment in new factories to tackle the worst storage chip shortage in 40 years. The expansion targets AI memory constraints, powering data storage for smartphones, autos, laptops, and data centers.

cnBeta (Full RSS) · 211d ago

Unitree Chunwan Robot Video Goes Viral Overseas
Applications

Unitree Chunwan Robot Video Goes Viral Overseas

Unitree humanoid robots performed on the Spring Festival Gala stage. The official video from Unitree reached nearly 100,000 views in under 10 hours overseas.

cnBeta (Full RSS) · 211d ago

Spring Fest Box Office Hits 10B, AI Orders Surge
Applications

Spring Fest Box Office Hits 10B, AI Orders Surge

2026 Spring Festival box office surpassed 10 billion yuan including pre-sales by Feb 17. Qianwen data shows AI ticket buys on Damai jumped 372x in two days.

36氪 · 211d ago

Robotaxi 19% Availability After 8 Months
Applications

Robotaxi 19% Availability After 8 Months

Tesla's Robotaxi service launched in Austin 8 months ago now has just 19% availability. This lags far behind Elon Musk's prior commitments.

cnBeta (Full RSS) · 211d ago

AI Climate Claims Branded Greenwashing
Infrastructure

AI Climate Claims Branded Greenwashing

A report dismisses tech industry claims that AI can fix climate issues as greenwashing. Most references are to traditional machine learning, not energy-intensive generative AI like chatbots and image tools.

The Guardian Technology · 211d ago

X-Blocks: Linguistic Blocks for AV Explanations
Research

X-Blocks: Linguistic Blocks for AV Explanations

X-Blocks introduces a hierarchical framework analyzing natural language explanations for automated vehicles (AVs) at context, syntax, and lexicon levels. RACE, a multi-LLM ensemble with Chain-of-Thought and self-consistency, achieves 91.45% accuracy on Berkeley DeepDrive-X dataset.

ArXiv AI · 211d ago

VeRA: Verified Reasoning Data Augmentation
Research

VeRA: Verified Reasoning Data Augmentation

VeRA converts benchmark problems into executable specifications—templates, generators, and verifiers—to create unlimited verified variants at near-zero cost. VeRA-E generates equivalent problems to detect memorization, while VeRA-H hardens tasks for fresh challenges.

ArXiv AI · 211d ago

VeRA: Scalable Verified Reasoning Data Augmentation
Models

VeRA: Scalable Verified Reasoning Data Augmentation

VeRA is a framework that transforms static benchmark problems into executable specifications for generating unlimited verified variants. It features VeRA-E for equivalent rewrites to detect memorization and VeRA-H for hardened tasks at intelligence frontiers.

ArXiv AI · 211d ago

VaryBalance: Top LLM Text Detector
Models

VaryBalance: Top LLM Text Detector

VaryBalance detects LLM-generated text by exploiting greater variation between human texts and their LLM-rewritten versions versus LLM texts. It quantifies this via mean standard deviation for robust distinction.

ArXiv AI · 211d ago

Trajectory-Dominant Pareto Optimization for Intelligence
Research

Trajectory-Dominant Pareto Optimization for Intelligence

AI systems stagnate in long-horizon adaptability due to trajectory-level Pareto traps, not data or capacity limits. The paper introduces Trajectory-Dominant Pareto Optimization, defining dominance over full trajectories, and Pareto traps as local optima blocking global paths.

ArXiv AI · 211d ago

SSLogic Scales Logic via Agentic Synthesis
Models

SSLogic Scales Logic via Agentic Synthesis

SSLogic is an agentic meta-synthesis framework that scales logical reasoning tasks at the family level using iterative Generate-Validate-Repair loops for Generator-Validator pairs. It features a Multi-Gate Validation Protocol with adversarial blind reviews by independent agents to ensure data reliability.

ArXiv AI · 211d ago

SELFCEST: Learned Parallel Model Clones
Research

SELFCEST: Learned Parallel Model Clones

SELFCEST equips base language models to spawn same-weight clones in parallel contexts via agentic reinforcement learning. It trains end-to-end with global task rewards and shared-parameter rollouts to allocate budgets across branches.

ArXiv AI · 211d ago

PlotChain Benchmark for MLLM Plot Reading
Research

PlotChain Benchmark for MLLM Plot Reading

PlotChain introduces a deterministic benchmark for evaluating multimodal LLMs on extracting quantitative values from engineering plots like Bode and FFT. It features 450 plots across 15 families with ground truth and checkpoint diagnostics for failure analysis.

ArXiv AI · 211d ago

NL2LOGIC: 99% Accurate NL-to-FOL Translation
Research

NL2LOGIC: 99% Accurate NL-to-FOL Translation

NL2LOGIC is a new framework using abstract syntax trees (AST) to translate natural language into first-order logic via large language models. It combines a recursive LLM semantic parser with an AST-guided generator for high syntactic accuracy and semantic faithfulness.

ArXiv AI · 211d ago

MAPLE: Sub-Agent Design for AI Personalization
Research

MAPLE: Sub-Agent Design for AI Personalization

MAPLE decomposes LLM agent limitations by separating memory, learning, and personalization into dedicated sub-agents. Memory manages storage/retrieval, Learning extracts insights asynchronously, and Personalization applies them in real-time.

ArXiv AI · 211d ago

Lang2Act Boosts VLM Visual Reasoning with Emergent Tools
Research

Lang2Act Boosts VLM Visual Reasoning with Emergent Tools

Lang2Act enhances Vision-Language Models (VLMs) via self-emergent linguistic toolchains for fine-grained visual perception in VRAG, avoiding rigid external tools and info loss from image ops. It employs a two-stage RL framework: first to build a reusable action toolbox, second to exploit it for reasoning.

ArXiv AI · 211d ago

Geometric Taxonomy of LLM Hallucinations
Models

Geometric Taxonomy of LLM Hallucinations

Researchers propose a geometric taxonomy classifying LLM hallucinations into three types: unfaithfulness, confabulation, and factual error. Benchmark hallucinations show strong domain-local detection but fail cross-domain, while human-crafted confabulations enable a single global detection direction.

ArXiv AI · 211d ago

Dual-Cycle Framework for Safe Role-Playing LLMs
Research

Dual-Cycle Framework for Safe Role-Playing LLMs

A training-free Dual-Cycle Adversarial Self-Evolution framework addresses jailbreak vulnerabilities in LLM role-playing agents. It couples a Persona-Targeted Attacker cycle for stronger jailbreaks with a Role-Playing Defender cycle that distills failures into a hierarchical safety knowledge base.

ArXiv AI · 211d ago

DPBench Reveals LLM Coordination Failures
Models

DPBench Reveals LLM Coordination Failures

DPBench introduces a benchmark for LLM multi-agent coordination using the Dining Philosophers problem across eight conditions varying timing, group size, and communication. Tests on GPT-5.2, Claude Opus 4.5, and Grok 4.1 show strong sequential performance but >95% deadlock in simultaneous settings due to convergent reasoning.

ArXiv AI · 211d ago

BotzoneBench: Scalable LLM Game Eval
Models

BotzoneBench: Scalable LLM Game Eval

BotzoneBench offers a scalable framework for evaluating LLMs' strategic reasoning in games using fixed hierarchies of skill-calibrated game AIs. It avoids quadratic costs and instability of LLM-vs-LLM tournaments by providing absolute, linear-time measurements.

ArXiv AI · 211d ago

BotzoneBench: Scalable LLM Game Eval Benchmark
Models

BotzoneBench: Scalable LLM Game Eval Benchmark

BotzoneBench introduces a scalable framework for evaluating LLMs' strategic reasoning in interactive games using fixed hierarchies of skill-calibrated game AIs. It assesses five flagship models across eight diverse games via 177,047 state-action pairs, revealing performance gaps and behaviors comparable to mid-tier game AIs.

ArXiv AI · 211d ago

AST-PAC Enhances Code MIA with AST Guidance
Research

AST-PAC Enhances Code MIA with AST Guidance

Researchers introduce AST-PAC, a syntax-aware adaptation of PAC for membership inference attacks on code LLMs. It uses AST-based perturbations to create valid calibration samples, outperforming baselines on larger files but facing limits on small or alphanumeric-rich code.

ArXiv AI · 211d ago

AMOR: Entropy-Gated SSM-Attention Hybrid
Research

AMOR: Entropy-Gated SSM-Attention Hybrid

AMOR is a hybrid model inspired by dual-process cognition theories, dynamically activating sparse attention only when SSM predictions show high entropy uncertainty. It projects Ghost KV from SSM states for O(n) efficiency, outperforming SSM-only and Transformer baselines on retrieval tasks with perfect accuracy using just 22% attention positions.

ArXiv AI · 211d ago

11342134313441372
Page 1343 of 1372
Back to home
SetupAI

A bilingual daily AI briefing — ten stories a day, each with a deep insight.

Takedown / opt-out: copyright@setupai.uk

© 2026 SetupAI

This WeekToolsUpdatesAboutPrivacyTermsRSS