
Chrome Auto Browse: Wins and Crashes
Chrome's Auto Browse agent was tested for web surfing tasks. It demonstrated impressive capabilities in some areas.
Ars Technica AI · 215d ago
Every story we have kept, newest first.
Looking for the daily editions? → Past editions
Page 1364 of 1372

Chrome's Auto Browse agent was tested for web surfing tasks. It demonstrated impressive capabilities in some areas.
Ars Technica AI · 215d ago

Chrome's Auto Browse agent autonomously surfs the web with impressive results. However, it also experiences spectacular crashes.
Ars Technica AI · 215d ago

xAI has launched the next phase of its AI development, introducing significant new advancements. This release expands capabilities for users and developers.
The Neuron · 215d ago

RentAHuman platform lets AI agents hire humans to promote AI startups. The author tried it but found it overrun by bots seeking hype.
Wired · 215d ago

Cybersecurity researcher Anton Cherepanov spotted a suspicious file on VirusTotal. The file uses AI to simplify online crimes.
MIT Technology Review · 215d ago

Cybersecurity researcher Anton Cherepanov detected a suspicious file on VirusTotal. AI tools are simplifying online criminal activities.
MIT Technology Review · 215d ago

AI tools are simplifying online crimes, potentially worsening threats. Cybersecurity researcher Anton Cherepanov spotted a suspicious malware file on VirusTotal.
MIT Technology Review · 215d ago

Author tests RentAHuman, a gig platform where AI agents hire humans to hype AI startups. It's bot-heavy and serves as an AI hype machine rather than innovative work.
Wired · 215d ago

OpenAI launches GPT-5.3-Codex-Spark, the first real-time coding model. It delivers 15x faster generation with 128k context.
OpenAI News · 215d ago

OpenAI unveils GPT-5.3-Codex-Spark, the first real-time coding model. It delivers 15x faster generation and 128k context.
OpenAI News · 215d ago

OpenAI introduces GPT-5.3-Codex-Spark, its first real-time coding model. It delivers 15x faster generation and a 128k context window.
OpenAI Blog · 215d ago

OpenAI introduced GPT-5.3-Codex-Spark, the first real-time coding model. It offers 15x faster generation and 128k context length.
OpenAI News · 215d ago

DeepSeek launched its R1 reasoning model in January 2025. This marked a pivotal shift for Chinese AI development.
MIT Technology Review · 215d ago

DeepSeek released its R1 reasoning model in January 2025. This marked a turning point for Chinese AI development.
MIT Technology Review · 215d ago

DeepSeek launched R1 reasoning model in January 2025, signaling a turning point for Chinese AI. Companies are rapidly advancing open-source models.
MIT Technology Review · 215d ago

Microsoft is developing powerful in-house AI models to achieve self-sufficiency and reduce dependence on OpenAI. This strategic shift follows a relationship reorganization in October last year.
cnBeta (Full RSS) · 215d ago

Anthropic is opening premium features to free Claude users, including file creation/editing, third-party connectors, and Skills. This counters OpenAI's introduction of ads in free and low-tier ChatGPT.
cnBeta (Full RSS) · 215d ago

Google appears set to release Gemini 3.1 Pro soon, with model references already spotted in related arenas. This follows recent launches like Zhipu's open-source GLM-5 and DeepSeek's upgraded model with larger context window.
cnBeta (Full RSS) · 215d ago

Google Gemini and related tools now refuse Disney character generation requests after Disney's IP infringement notice. The update rolled out about two months after Disney's December cease-and-desist letter.
cnBeta (Full RSS) · 215d ago
Cosmo3DFlow uses 3D wavelet transform and flow matching for efficient cosmological inference from N-body simulations. Addresses sparsity via spectral compression, enabling 50x faster sampling than diffusion models.
ArXiv AI · 215d ago
VulReaD uses a security knowledge graph and teacher LLM for CWE-consistent vulnerability detection beyond binary classification. Student models are fine-tuned with ORPO for taxonomy-aligned reasoning.
ArXiv AI · 215d ago
Found-RL integrates foundation models into RL for end-to-end driving via async batch inference to cut latency. Distills VLM guidance using VMR, AWAG; CLIP rewards shaped by conditional alignment.
ArXiv AI · 215d ago
Vision-Centric Jailbreak Attack (VJA) uses visual inputs to bypass safety in image editing models. IESBench benchmark tests vulnerabilities with up to 80.9% success rates.
ArXiv AI · 215d ago
VESPO introduces variational sequence-level soft policy optimization to tackle training instability in RL for LLMs caused by policy staleness and async execution. It derives a closed-form reshaping kernel for importance weights without length normalization.
ArXiv AI · 215d ago
Versor uses Conformal Geometric Algebra (CGA) for sequence modeling with SE(3)-equivariance. Outperforms Transformers on N-body dynamics, topology, and benchmarks with fewer parameters.
ArXiv AI · 215d ago
V-STAR addresses probability-reward mismatch in generative recsys via value-guided decoding and sibling-relative RL. VED efficiently explores high-potential prefixes; Sibling-GRPO focuses on decisive branches.
ArXiv AI · 215d ago
EVA is a cross-species, multimodal foundation model harmonizing transcriptomics and histology for immunology. It shows scaling laws and SOTA on 39 tasks from discovery to clinical trials.
ArXiv AI · 215d ago
Develops theory for random projections in computing influence functions, covering unregularized, regularized, and factorized cases. Shows exact preservation conditions and handles out-of-range gradients via leakage term.
ArXiv AI · 215d ago
TwiFF-2.7M dataset and model advance VCoT for videos via future frame generation. TwiFF-Bench evaluates reasoning trajectories.
ArXiv AI · 215d ago
Transformer training on modular arithmetic tasks collapses high-dimensional parameters to 3-4D execution manifolds. This structure explains attention concentration, SGD integrability, and sparse autoencoder limits.
ArXiv AI · 215d ago