
Chinese Youth Redefine AI Memory
A group of young Chinese, including 19-year-old Ivy League dropouts, has rebuilt AI memory capabilities. Their system is the only one with native coreference resolution support.
量子位 · 184d ago
Every story we have kept, newest first.
Looking for the daily editions? → Past editions
Page 973 of 1375

A group of young Chinese, including 19-year-old Ivy League dropouts, has rebuilt AI memory capabilities. Their system is the only one with native coreference resolution support.
量子位 · 184d ago

Lenovo has redefined its 'Lobster' product line. This update addresses key computing demand challenges.
量子位 · 184d ago

Developer optimized Kokoro TTS for iOS with CPU-only pipeline, hitting 20x realtime without thermal issues by splitting the model and using Apple's Accelerate framework. Avoids Metal for background audio support.
Reddit r/LocalLLaMA · 184d ago

PrismML, a Caltech AI startup, released Bonasi 8B, a 1-bit large language model competitive with other 8B models. It is 14x smaller and 5x more energy efficient, aiming to enable efficient AI on mobile devices and reduce cloud dependency.
The Register - AI/ML · 184d ago

Users worry about unreleased Qwen 3.6 397B, but small benchmark gaps between Qwen 3.5 and 3.6 suggest quantization like Q2_K_XL on RTX 6000 would negate advantages. Discussion anticipates smaller Qwen models competing with Gemma 4.
Reddit r/LocalLLaMA · 184d ago
China Nonferrous Metals Industry Association Silicon Division surveyed Baotou's silicon sector. Firms report acute supply-demand imbalance, prices below costs, and chain-wide losses.
36氪 · 184d ago

Tesla's global Supercharger network exceeds 80,000 stalls, with over 12,000 in mainland China across 2,500+ stations. It covers 100% of provincial capitals and opens 950+ stations to non-Tesla EVs.
IT之家 · 184d ago
Zivariable Robotics hosted the inaugural global embodied intelligence developer conference in Shenzhen, with 20 post-00s teams hacking for 72 hours on real robotic arms backed by 100+ PFLOPs compute and open models like WALL-OSS. A/B leaderboards tested generalization in fixed vs.
36氪 · 184d ago

Ukraine's use of unmanned ground vehicles has surged exponentially since spring 2024, reshaping the war with Russia into a tech battle. These battery-powered robots vary in design, from tracked milk-float-like units to wheeled models and anti-tank mine carriers.
The Guardian Technology · 184d ago
User shares detailed recipe for quantizing GGUF models like Gemma-4-26B-A4B, requiring 500GB storage and architecture-specific configs. Thanks quantizers like unsloth, bartowski; links full REPRODUCE.md on Hugging Face.
Reddit r/LocalLLaMA · 184d ago
Internet giants are pivoting from lightweight software to massive AI infrastructure investments, with Amazon planning $200B capex in 2026 for power and data centers. Big Tech's combined $650B annual spend rivals global semiconductor revenue, while Chinese firms like ByteDance, Alibaba, and Tencent allocate hundreds of billions RMB to AI chips and facilities.
虎嗅 · 184d ago

Django founder warns that AI will render the skills of 30-year-old programmers worthless. He states his former superpower of rapid prototyping is now achievable by anyone using AI tools.
量子位 · 184d ago

Tiiny AI Pocket Lab plug-in device for 100B local LLM inference hits $2.95M Kickstarter with 2k backers at $1399; fills gap for privacy-focused, easy local AI vs. costly AI PCs or weak boards.
虎嗅 · 184d ago
Reddit thread seeks insights from ML/AI experts with 10+ years experience on public misconceptions. Highlights gaps between public perceptions and frontier research realities.
Reddit r/MachineLearning · 184d ago

Alibaba's Qwen3.6-Plus hit 1.4 trillion tokens on OpenRouter in one day post-launch, shattering the platform's single-model record. It tops China and ranks #2 globally in programming benchmarks.
IT之家 · 184d ago
Researcher debates submitting NeurIPS paper on novel agentic system with formal convergence proof and real-world application. Limited to few examples due to unsuitable benchmarks.
Reddit r/MachineLearning · 184d ago
Qwen AI ride-hailing launched on March 23 and saw orders surge over 1500% week-over-week on April 4 during Qingming holiday. User scale rapidly expanded in under two weeks.
36氪 · 184d ago

YC-Bench benchmark simulates LLMs running a startup for a year with delayed feedback and adversarial clients. GLM-5 achieves $1.21M avg funds, close to Claude Opus's $1.27M but at 11x lower API cost.
Reddit r/LocalLLaMA · 184d ago

Next-gen consoles slated for late 2027 or 2028 despite persistent memory shortages. Sony is alerting developers to gear up for PS6 and a new portable PS handheld.
cnBeta (Full RSS) · 184d ago
Elon Musk reportedly requires companies eyeing SpaceX IPO participation to purchase Grok. This bundles access to the rocket firm's public offering with adoption of xAI's AI chatbot.
36氪 · 184d ago
AI and social media have flooded content supply, diminishing traditional PR impact despite efficient production. Success now hinges on clear core messaging and translating internal narratives for external clarity.
虎嗅 · 184d ago
Dedicated Reddit thread for ACL 2026 decision updates and discussions. Decisions expected to publish within 24 hours.
Reddit r/MachineLearning · 184d ago

Viral GitHub project 'colleague-skill' lets users feed ex-colleagues' messages, docs, and emails into AI to create mimicking 'skills' for work tasks. It sparks debate on distilling human expertise into AI, job losses, and 'cyber immortality' memes.
虎嗅 · 184d ago
Qwen3.6-Plus, released just one day ago on April 4th, has surged to the top of OpenRouter's daily global model API calls leaderboard. It achieved over 1.4 trillion tokens in daily usage, shattering the platform's single-model daily record.
36氪 · 184d ago

Anthropic has blocked OpenClaw access. The project's creator, 'Lobster Father', failed in persuasion efforts.
Ifanr (爱范儿) · 184d ago
Meta's Superintelligence Lab is reportedly assembling a hardware team. The news comes from Cailianpress via 36Kr.
36氪 · 184d ago

Author tested VRAM increases on a Windows 11 PC to assess performance gains. Virtual RAM boosts PC speed when physical resources are limited.
ZDNet AI · 184d ago

Yang Zhilin, founder of Beijing-based Moonshot AI behind Kimi models, was a surprise speaker at Nvidia’s GTC conference. This occurs amid US-China AI rivalry rhetoric compared to an arms race.
SCMP Technology · 184d ago
llama.cpp update fixes Gemma 4's KV cache issue, eliminating excessive VRAM usage previously reaching petabytes. This enables efficient local inference without hardware limitations.
Reddit r/LocalLLaMA · 184d ago

The secondary market for private shares is at peak activity. Anthropic has become the hottest trade, surpassing OpenAI which is losing ground.
TechCrunch AI · 184d ago