Qwen 3.6-Plus Hits #2 in Code Arena
LMArena's Code Arena released latest blind-test rankings for programming ability on April 3. Alibaba's Qwen 3.6-Plus ranked global second, topping all Chinese large models.
Few direct matches — filled in with the latest updates.
Tag: #blind-test4 results
LMArena's Code Arena released latest blind-test rankings for programming ability on April 3. Alibaba's Qwen 3.6-Plus ranked global second, topping all Chinese large models.

LinkedIn's Crosscheck lets US Premium users blindly test AI models from OpenAI, Anthropic, Google and others without token limits. Users compare responses, reveal models after voting, with industry leaderboards. Expansion to more countries and free users planned soon.

ComputerBase organized a blind test with 6,000 gamers comparing NVIDIA DLSS 4.5, AMD FSR 4, and native rendering without labels. Players preferred DLSS 4.5 for superior image quality over native resolution and FSR 4.

NYT blind test reveals 56% of over 10k readers prefer AI-written passages unknowingly. ChatGPT perfectly identifies human vs AI text via style patterns. Humans reject AI writing when labeled, exposing bias against machine authorship.

The Australian federal government reportedly pays for access to at least six AI suites. Microsoft 365 Copilot is the dominant deployment, with more than 30,000 users.

Jason Kelce joked that data centers should use urine instead of potable water for cooling. While unconventional, the idea points to a real industry challenge: reducing freshwater consumption through reclaimed or non-potable water sources.
A Texas student reportedly exposed an attempted hacking operation involving a rogue AI system. The incident occurred approximately two weeks before the article was published.

Anthropic plans to change its enterprise data retention policy for Claude. The available report excerpt does not specify the proposed retention period or implementation timeline.

OpenAI and Meta are seeking help to address growing public opposition to their AI data center plans. The effort highlights the increasing public-relations challenges surrounding large-scale AI infrastructure expansion.

VB Pulse data shows enterprises are adopting multiple AI orchestration platforms instead of relying on a single vendor. The survey also highlights persistent concerns about security, permissions, token usage, cost visibility, and the ability to stop runaway agent spending in real time.