All Updates

Page 1753 of 1953

March 4, 2026

🛡️
Cloudflare Blog176d ago

Nametag Partnership Defeats Deepfakes

Cloudflare One partners with Nametag to combat laptop farms and AI-enhanced identity fraud. Identity verification is required during employee onboarding. Continuous authentication prevents insider threats.

#deepfake#identity-fraud#onboarding
🛡️
Cloudflare Blog176d ago

Cloudflare One Adds User Risk Scoring

Cloudflare One now incorporates dynamic User Risk Scores into Access policies for automated, adaptive security responses. This moves teams beyond binary allow/deny rules by evaluating continuous behavior signals from internal and third-party sources.

#risk-scoring#zero-trust#behavioral-analytics
📱
Ifanr (爱范儿)176d ago

AI Circle Shares Qwen Farewell Tweet Overnight

Overnight, the global AI community has been widely sharing a farewell tweet related to Qwen. Qwen Station has reached a new crossroads, sparking buzz across the AI world. The news was highlighted by Ifanr.

#farewell-tweet#ai-community#platform-shift
🔥
36氪176d ago

Micro LED CPO power at 5% of copper cables

TrendForce reports generative AI boom drives data center demand for high-speed interconnects. Copper cables struggle with density and energy; Micro LED CPO cuts power to 5% of copper. This positions it as a key alternative for intra-rack transmission.

#data-center#optical-interconnect#energy-efficiency
🐯
虎嗅176d ago

Tokens Outvalue Human Labor

AI tools like Seedance produce viral videos for pennies, slashing human production costs and funneling wealth to compute providers. Humans risk becoming 'meat plugins' for AI via RentAHuman. Warns of 2028 white-collar crisis and systemic collapse.

#ai-economics#job-automation#content-tools
🇬🇧
BBC Technology176d ago

Robot Recruiters for Care Workers?

AI tools are screening care workers for suitability. The article questions if robots can truly assess qualities needed for caring roles. BBC explores limitations in automated hiring.

#robot-recruiter#ai-hiring#care-workers
🐯
虎嗅177d ago

Qwen Tech Lead Lin Junyang Resigns

Alibaba's Qwen technical lead Lin Junyang abruptly announced his departure on X after leading the model to global open-source dominance. Colleagues mourn the loss amid recent Qwen3.5 and Qwen3-Max releases, with two other key engineers also leaving. Potential successors include Alibaba's Zhou Jingren or DeepMind's Hao Zhou.

#leadership-change#llm-strategy#alibaba-ai
🇨🇳
cnBeta (Full RSS)177d ago

Phison Mandates Prepayments in Supply Crunch

Phison Electronics is requiring customers to prepay or shorten payment terms for SSD controller orders. Previously offered 3-month or longer credit periods are discontinued. This stems from upstream suppliers imposing prepayments on Phison amid rising supply chain financial pressures.

#supply-chain#ssd#payment-terms
🇨🇳
cnBeta (Full RSS)177d ago

Windows 12: Modular, AI-First This Year?

Microsoft is reportedly preparing Windows 12 for release this year, featuring a fully modular design. AI will be the core experience, aligning with the company's AI-first operational shift.

#os-update#modular-design#ai-integration
📄
ArXiv AI177d ago

SWE-Hub Unifies Scalable SWE Task Production

SWE-Hub is an end-to-end production system addressing data scarcity for software-engineering AI agents by automating environments, synthesizing bugs at scale, and generating diverse tasks. Key components include Env Agent for reproducible multi-language setups, SWE-Scale for high-throughput bug fixes, Bug Agent for system-level regressions, and SWE-Architect for repo-scale creation from natural language. It enables continuous delivery of executable tasks across the full SWE lifecycle.

#swe-agents#data-factory#bug-synthesis
📄
ArXiv AI177d ago

NFR Patterns for Agentic AI Reliability

Revisits goals-to-aspects methodology for agentic AI, introducing 12 reusable patterns across security, reliability, observability, and cost management. Maps i* goal models to Rust AOP implementations, with agent-specific patterns like prompt injection detection. Validates via case study on open-source agent framework.

#agentic-ai#aop#nfr-patterns
📄
ArXiv AI177d ago

MicroVerse Launches Micro-World Simulations

Introduces MicroWorldBench, a benchmark with 459 expert criteria for microscale simulations across organ, cellular, and molecular levels. Reveals SOTA video models' failures in physics, consistency, and fidelity. Releases MicroSim-10K dataset and trains MicroVerse for accurate microscale reproductions.

#video-gen#biomedical-benchmark
📄
ArXiv AI177d ago

M-JudgeBench Boosts Multimodal Judge Reliability

Introduces M-JudgeBench, a 10-dimensional benchmark assessing MLLM judges across pairwise CoT, length bias, and error detection. Proposes Judge-MCTS for generating diverse reasoning data to train superior M-Judger models. Experiments show M-Judger outperforming priors on benchmarks.

#multimodal-judge#mcts-data#cot-benchmark
📄
ArXiv AI177d ago

LOGIGEN: Logic-Driven Agent Task Generator

LOGIGEN is a framework that synthesizes verifiable training data for agentic LLMs using logic-driven methods and triple-agent orchestration. It generates 20,000 complex tasks across 8 domains with guaranteed validity via state equivalence checks. Models trained with SFT and RL achieve 79.5% success on τ²-Bench, far surpassing baselines.

#agentic-ai#data-synthesis#rl-training
📄
ArXiv AI177d ago

LifeEval: Egocentric AI Assistance Benchmark

LifeEval introduces a multimodal benchmark for real-time, task-oriented human-AI collaboration in egocentric daily life. It features 4,075 QA pairs across 6 capability dimensions, emphasizing holistic evaluation, real-time perception from first-person streams, and natural dialogues. Evaluations of 26 MLLMs reveal major challenges in adaptive interaction.

#multimodal-benchmark#egocentric-ai
📄
ArXiv AI177d ago

IRIS: UMLLM Fairness Benchmark Launch

IRIS Benchmark is the first to synchronously evaluate fairness in UMLLMs' understanding and generation tasks. Powered by ARES classifier and four datasets, it aggregates 60 metrics into a high-dimensional 'fairness space' across IRIS dimensions. It uncovers biases like 'generation gap' and 'personality splits' in leading models.

#fairness#bias#multimodal
📄
ArXiv AI177d ago

HealHGNN Masters Heterophilic Hypergraphs

HealHGNN introduces heterophily-agnostic message passing for hypergraph neural networks using Riemannian geometry. It mitigates oversquashing via adaptive local exchangers based on manifold heat flow, capturing long-range dependencies with Robin conditions and source terms. The model achieves SOTA performance on both homophilic and heterophilic datasets with linear complexity.

#riemannian-geometry#heterophily
📄
ArXiv AI177d ago

EMPA: Persona-Aligned Empathy Evaluation Framework

EMPA is a process-oriented framework for evaluating persona-aligned empathy in LLM-based dialogue agents, addressing challenges like latent user states and sparse feedback. It distills real interactions into controllable scenarios and uses a multi-agent sandbox to test strategic adaptation. Trajectories are scored on directional alignment, cumulative impact, and stability in psychological space.

#empathy-evaluation#persona-alignment#multi-agent
📄
ArXiv AI177d ago

Draft-Thinking Cuts CoT Costs Dramatically

Draft-Thinking trains LLMs to use concise draft-style reasoning, retaining only critical steps via progressive curriculum learning. It reduces reasoning budget by up to 82.6% on MATH500 with just 2.6% performance drop. Adaptive prompting allows flexible reasoning depth.

#chain-of-thought#curriculum-learning#efficient-reasoning
📄
ArXiv AI177d ago

DenoiseFlow: Uncertainty-Aware LLM Agent Denoising

DenoiseFlow tackles accumulated semantic ambiguity in long-horizon LLM agent workflows by modeling them as a Noisy MDP. It uses three stages—sensing uncertainty, adaptive regulation of computation, and targeted correction—for progressive denoising. Achieves highest accuracy (83.3% avg) on six benchmarks while reducing costs by 40-56%.

#agentic-workflows#noisy-mdp
Page 1753 of 1953