SetupAI

Daily AI briefing

This WeekToolsUpdatesSearch繁
繁

Full archive

Every story we have kept, newest first.

Looking for the daily editions? → Past editions

Page 1365 of 1372

February 12, 2026

Transformer for Experimental NMR Structure Elucidation
Research

Transformer for Experimental NMR Structure Elucidation

NMRTrans uses set transformers on experimental NMR spectra for molecular structure elucidation, trained on NMRSpec corpus from literature. It models spectra as unordered peak sets aligning with NMR physics.

ArXiv AI · 215d ago

Topology Meets NNs Under Uncertainty
Applications

Topology Meets NNs Under Uncertainty

Integrates neural networks, topological data analysis, and Bayesian methods for AI in military domains. Covers image, time-series, graph applications like fraud detection.

ArXiv AI · 215d ago

Tokens Enable Emergent Resource Rationality
Models

Tokens Enable Emergent Resource Rationality

Inference-time scaling in language models leads to adaptive resource rationality without explicit cost rewards. Models shift from brute-force to analytic strategies as task complexity rises.

ArXiv AI · 215d ago

TokaMark Launches Fusion Plasma Benchmark
Infrastructure

TokaMark Launches Fusion Plasma Benchmark

TokaMark standardizes AI evaluation on MAST tokamak data with unified multi-modal access and 14 tasks. Harmonizes formats, metadata, and protocols for reproducible comparisons.

ArXiv AI · 215d ago

Text Boosts Multimodal Anomaly Detection
Video

Text Boosts Multimodal Anomaly Detection

Text-guided framework enhances weakly supervised multimodal video anomaly detection. Employs in-context learning for anomaly text augmentation and multi-scale bottleneck Transformer for fusion.

ArXiv AI · 215d ago

δ_TCB Measures LLM Prediction Stability
Models

δ_TCB Measures LLM Prediction Stability

Introduces δ_TCB metric to quantify LLM internal state robustness against perturbations, beyond traditional accuracy. Linked to output embedding geometry, it reveals prediction instabilities missed by perplexity.

ArXiv AI · 215d ago

Synthetic Underspecification for Agents
Applications

Synthetic Underspecification for Agents

LHAW generates controllable underspecified long-horizon tasks by removing info across goals, constraints, inputs, context. Validates via agent trials, classifying ambiguity impacts.

ArXiv AI · 215d ago

SynergyKGC Handles KG Heterogeneity
Research

SynergyKGC Handles KG Heterogeneity

SynergyKGC fuses entity semantics with heterogeneous topologies via cross-modal synergy. Uses density-dependent anchoring and double-tower consistency.

ArXiv AI · 215d ago

Step 3.5 Flash: Efficient Frontier AI
Models

Step 3.5 Flash: Efficient Frontier AI

Step 3.5 Flash is a 196B MoE model with 11B active params for agentic tasks. Optimized with sliding-window attention and MTP-3 for low-latency inference.

ArXiv AI · 215d ago

Stats Test Spots LLM Degradations
Models

Stats Test Spots LLM Degradations

McNemar's test framework detects post-optimization LLM degradations via per-sample comparisons. Aggregates across benchmarks with controlled false positives.

ArXiv AI · 215d ago

Silence Boosts Collective Taste Judgment
Research

Silence Boosts Collective Taste Judgment

Introduces Silence Routing framework for collective intelligence in taste domains using music preferences. Specifies when contributors should speak, report, or stay silent.

ArXiv AI · 215d ago

SigLIP Boosts Multi-Label ECG Classification
Image

SigLIP Boosts Multi-Label ECG Classification

Adapts SigLIP contrastive learning with a Jaccard-based sigmoid loss for multi-label ECG classification using real-world data. Incorporates medical knowledge and techniques like higher embedding dimensions and random cropping.

ArXiv AI · 215d ago

Semantic Labels Enhance TPRA Retrieval
Applications

Semantic Labels Enhance TPRA Retrieval

Explores semantic labeling for TPRA questionnaires using LLMs and hybrid SSSL. Compares direct labeling vs.

ArXiv AI · 215d ago

Self-Supervised SR Quality Assessor
Image

Self-Supervised SR Quality Assessor

Proposes no-reference IQA for real-world super-resolved images using content-free SSL. Pretrains multi-SR model representations via contrastive learning.

ArXiv AI · 215d ago

ScratchWorld Tests GUI Agents
Applications

ScratchWorld Tests GUI Agents

Introduces ScratchWorld benchmark with 83 tasks for multimodal GUI agents in Scratch. Uses primitive/composite modes and execution-based evaluation.

ArXiv AI · 215d ago

Safety Alignment for Omni-Modal LLMs
Models

Safety Alignment for Omni-Modal LLMs

OmniSteer addresses cross-modality vulnerabilities in OLLMs using AdvBench-Omni dataset and modality-semantics decoupling. Uncovers mid-layer dissolution and extracts golden refusal vector via SVD.

ArXiv AI · 215d ago

SAF Improves Parkinson's ECoG Prediction
Applications

SAF Improves Parkinson's ECoG Prediction

Introduces first reproducible ECoG dataset from rat models for Parkinson's disease prediction. Swap-Adversarial Framework (SAF) uses channel swapping and domain-adversarial training to tackle inter-subject variability and HDLSS issues.

ArXiv AI · 215d ago

RSHallu: Hallucination Eval for RS MLLMs
Image

RSHallu: Hallucination Eval for RS MLLMs

RSHallu studies hallucinations in remote-sensing MLLMs with a new taxonomy, benchmark, and dual-mode checker. Provides datasets for mitigation via training and plug-and-play strategies.

ArXiv AI · 215d ago

Robust Policy Optimization for Recommendations
Applications

Robust Policy Optimization for Recommendations

DRPO tackles model collapse in off-policy generative recommendation via optimistic distributionally robust optimization. Proves hard filtering recovers high-quality data from noisy logs.

ArXiv AI · 215d ago

RLCER Evolves CoT Rubrics
Models

RLCER Evolves CoT Rubrics

RLCER reinforces chain-of-thought via self-evolving rubrics without human labels. Outperforms outcome-centric RLVR on reasoning tasks.

ArXiv AI · 215d ago

Rewiring Sparsifies Efficient GNNs
Research

Rewiring Sparsifies Efficient GNNs

Explores adaptive rewiring and sparsification for scalable GNNs using Erdős-Rényi models. Tested on power grid N-1 analysis with GCN/GIN.

ArXiv AI · 215d ago

RealHD Dataset Detects AI Fake Images
Image

RealHD Dataset Detects AI Fake Images

RealHD offers 730k high-quality real and AI-generated images from advanced methods like text-to-image and inpainting. Addresses prior dataset flaws with diverse prompts, metadata, and masks.

ArXiv AI · 215d ago

Quantum ICO Merges Sensing and Computation
Research

Quantum ICO Merges Sensing and Computation

Proposes quantum scheme using indefinite causal order (ICO) for integrated sensing and computation on one state. Agent superposes observation-then-compute and compute-then-observation orders.

ArXiv AI · 215d ago

Quadrupeds Cooperate for Super Jumps
Applications

Quadrupeds Cooperate for Super Jumps

Co-jump enables two quadrupeds to synchronize jumps up to 1.5m via MAPPO and curriculum, without communication. Achieves 144% height gain over solo robots using proprioception.

ArXiv AI · 215d ago

μpscaling Optimizes Model Warm Starts
Infrastructure

μpscaling Optimizes Model Warm Starts

Proposes principled upscaling for model widths inspired by μP, with theory guaranteeing equivalence to widened versions. Extends μTransfer for hyperparameter scaling, avoiding costly retuning at larger sizes.

ArXiv AI · 215d ago

ProtoGLAD Enables Interpretable Graph Anomalies
Research

ProtoGLAD Enables Interpretable Graph Anomalies

ProtoGLAD detects graph-level anomalies by contrasting with nearest normal prototype graphs discovered via point-set kernels. It iteratively clusters normal graphs for unsupervised detection.

ArXiv AI · 215d ago

Privacy Shield for Mobile GUI Agents
Research

Privacy Shield for Mobile GUI Agents

Framework anonymizes sensitive UI data with type-preserving placeholders for cloud-based GUI agents. Detects PII across screenshots, XML, and instructions via layered architecture.

ArXiv AI · 215d ago

Privacy-Aware XR Collaboration Framework
Applications

Privacy-Aware XR Collaboration Framework

PRISM-XR integrates multimodal LLMs for XR collaboration while filtering sensitive data from XR headset frames on edge servers. It features lightweight registration and customizable content-sharing for efficient synchronization.

ArXiv AI · 215d ago

Policy-Car Swerving Absorbs Traffic Jams
Applications

Policy-Car Swerving Absorbs Traffic Jams

Proposes SVDD-JAD strategy mimicking police swerving to suppress stop-and-go waves via slow-in/fast-out maneuvers. Analyzes five key parameters measurable with roadside detectors.

ArXiv AI · 215d ago

PiT-PO Boosts Equation Discovery with RL
Models

PiT-PO Boosts Equation Discovery with RL

PiT-PO uses reinforcement learning to evolve LLMs for symbolic regression, enforcing physical validity and parsimony. It treats LLMs as adaptive generators updated by search feedback.

ArXiv AI · 215d ago

11364136513661372
Page 1365 of 1372
Back to home
SetupAI

A bilingual daily AI briefing — ten stories a day, each with a deep insight.

Takedown / opt-out: copyright@setupai.uk

© 2026 SetupAI

This WeekToolsUpdatesAboutPrivacyTermsRSS