
Claude Launches Agent After Banning Lobster
Claude blocks 'lobster' jailbreak before Anthropic promotes its own Agent service. An open-source alternative rapidly gains 2.6k GitHub stars.
量子位 · 163d ago
Open weights are the industry’s counterweight to closed labs. Releases, licenses and benchmarks from Llama, Qwen, DeepSeek and more.
367 articles

Claude blocks 'lobster' jailbreak before Anthropic promotes its own Agent service. An open-source alternative rapidly gains 2.6k GitHub stars.
量子位 · 163d ago

DeepSeek has rolled out a significant upgrade, stirring the AI community. V4 appears imminent based on hints.
Ifanr (爱范儿) · 164d ago

A new neuro-symbolic architecture extracts object structures from ARC grids, proposes DSL transformations via neural priors, and filters via cross-example consistency. It lifts base LLM performance on ARC-AGI-2 from 16% to 24.4%, reaching 30.8% combined with ARC Lang Solver.
ArXiv AI · 166d ago

Tiiny AI Pocket Lab plug-in device for 100B local LLM inference hits $2.95M Kickstarter with 2k backers at $1399; fills gap for privacy-focused, easy local AI vs. costly AI PCs or weak boards.
虎嗅 · 168d ago

Indian AI startup Sarvam is raising $300-350M at $1.5-1.55B valuation, led by Bessemer with Nvidia, Amazon joining. They open-sourced Sarvam 30B and 105B MoE models in March.
IT之家 · 169d ago
OpenClaw 2026.4.2 features breaking config migrations for xAI and Firecrawl plugins to standardized paths, fixable via 'openclaw doctor --fix'. It restores durable Task Flow orchestration with managed child tasks, cancel handling, and plugin APIs.
OpenClaw (GitHub Releases) · 170d ago

Apple delisted AI coding app Anything and froze updates for Replit over unapproved code execution. Vibe Coding, popularized by Karpathy, enables non-coders to build apps but floods markets with insecure 'shit mountains'—10% of Lovable apps expose user data.
虎嗅 · 171d ago

ggerganov's attn-rot (TurboQuant lite) is nearing merge into llama.cpp, delivering lower KLD errors in KV quantization for Qwen3.5 models. Benchmarks show improved q4_0 quality with comparable speeds on 35B, 27B, and 122B variants in VRAM and CPU.
Reddit r/LocalLLaMA · 172d ago

French open-source platform Kestra raised $25M Series A led by RTP Global, totaling $36M funding. Enterprise revenue surged 25x in 18 months, with over 2 billion workflows executed in 2025.
The Next Web (TNW) · 172d ago

daVinci-LLM combines industrial-scale resources with full openness to explore LLM pretraining science. It releases a 3B model trained from scratch on 8T tokens using Data Darwinism framework and two-stage adaptive curriculum.
ArXiv AI · 172d ago

Alibaba has released benchmark results for Qwen3.5-Omni on Reddit's r/LocalLLaMA. The post shares details on the model's performance.
Reddit r/LocalLLaMA · 172d ago
Open-source prototype implements Unix philosophy in ML retrieval pipelines with swappable, typed-contract stages like PII redaction and chunking. Enables isolating changes for precise eval comparisons.
Reddit r/MachineLearning · 173d ago

JD.com open-sourced its large language model JoyAI-LLM Flash and launched the 'Lobster Squad' in the latest Digital Intelligence Weekly. This release supports agent-style AI advancements amid other news like OpenAI's Sora app shutdown and Alibaba's new CPU.
钛媒体 · 175d ago
Meituan released and fully open-sourced its native multimodal large model LongCat-Next and core component dNaViT visual tokenizer on March 27. It unifies image, speech, and text into discrete tokens, breaking language-centric architectures.
36氪 · 176d ago

Jim Keller highlights RISC-V's explosion with 20+ new designs vs ARM's 5, driven by openness. Tenstorrent's Ascalon boasts 8-wide decode, 230GB/s bandwidth for AI servers.
虎嗅 · 176d ago

Mistral has released a new open-source model for speech generation. The lightweight model can run efficiently on resource-limited devices like smartwatches and smartphones.
TechCrunch AI · 177d ago

A DeepSeek employee teased a 'massive' new model surpassing DeepSeek V3.2 on social media. The post was quickly deleted, sparking speculation.
Reddit r/LocalLLaMA · 178d ago

NVIDIA has donated its Dynamic Resource Allocation Driver for GPUs to the Kubernetes open-source community. This tool aids enterprises in managing high-performance AI infrastructure with greater transparency and efficiency.
NVIDIA Blog · 179d ago

Cursor launched Composer 2 as self-trained coding model, but exposed as fine-tune of Moonshot's Kimi K2.5 via API and tokenizer analysis. Controversy over disclosure and license resolved with confirmed authorization via Fireworks AI.
虎嗅 · 181d ago
Volga is an open-source data engine tailored for real-time AI/ML pipelines, recently rewritten in native Rust from a Python+Ray prototype. It leverages Apache DataFusion and Arrow for unified streaming, batch, and request-time compute, eliminating complex infrastructure like Flink or Spark.
Reddit r/MachineLearning · 184d ago

OpenClaw introduces edge execution node paradigm, blending cloud/edge for agents beyond Claude Code capabilities. Sparks forks like NanoClaw and infra ecosystems.
虎嗅 · 184d ago
New fine-tuned Qwen3.5-9B GGUF model uploaded to Hugging Face, optimized on reasoning and function-calling data. Pushes structured responses and tool-use for llama.cpp, LM Studio, Ollama.
Reddit r/LocalLLaMA · 185d ago
Openrouter's stealth models Hunter Alpha and Healer Alpha are officially confirmed as MiMo V2 Pro and MiMo V2 Omni. Hunter Pro offers 1M context for text-only reasoning, while Healer Omni supports text+image with 262K context.
Reddit r/LocalLLaMA · 186d ago
TerraLingua is a persistent multi-agent environment for AI agents to interact, create artifacts, and evolve under resource constraints and lifecycles. Emergent behaviors include implicit rules, infrastructure, and knowledge reuse.
Reddit r/MachineLearning · 186d ago

Pi is the minimalist core framework powering OpenClaw, topping Terminal Bench 2.0 with under 1000-token prompts and just four tools: read, write, edit, bash. Creator Mario Zechner emphasizes letting users decide needs, avoiding bloated features like MCP or plan modes.
虎嗅 · 186d ago

A Reddit user from r/LocalLLaMA released their state-of-the-art Text-To-Sample Generator today. The model fits under 7GB VRAM, with 8GB recommended for headroom.
Reddit r/LocalLLaMA · 187d ago

Newton 1.0 GA is a GPU-accelerated open-source physics simulator optimized for industrial robotics. It excels in contact-rich manipulation and locomotion by handling complex dynamics like contact forces and deformable objects.
NVIDIA Developer Blog · 187d ago

Nvidia launched open-source Agent Toolkit for autonomous enterprise AI agents, adopted by 17 firms including Adobe, Salesforce, SAP. Includes Nemotron models, AI-Q blueprint, OpenShell runtime, cuOpt library.
VentureBeat · 187d ago

NVIDIA launches AI Cluster Runtime, an open-source project for validating Kubernetes clusters with GPU infrastructure. It offers layered, reproducible recipes spanning the full software stack from drivers to operators.
NVIDIA Developer Blog · 191d ago

NVIDIA Megatron Core implements Falcon-H1 Hybrid Architecture to advance large language model training at scale. This open-source library delivers industry-leading parallelism and GPU-optimized performance.
NVIDIA Developer Blog · 194d ago