Search

10 results on this page

Rethinking AI Agents on Kubernetes

Rethinking AI Agents on Kubernetes

The article proposes treating Kubernetes Pods as workers rather than complete AI agents. It examines how this shift could provide a more flexible deployment unit for AI agent systems.

InfoQ中国Media13h ago#ai-agents
🤖

Qwen KV Precision Shows Real Quality Gaps

A community test on an AMD R9700 with ROCm reports noticeable quality and long-context retention differences between FP16 and q8_0 KV cache for Qwen3.8-27B. FP16 reportedly produces more careful structured output and maintains performance beyond 120k tokens, challenging the assumption that both formats are equivalent.

Reddit r/LocalLLaMACommunity19h ago#kv-cache#long-context#quantization
A Smaller Qwen3.8 Built by Pruning Layers

A Smaller Qwen3.8 Built by Pruning Layers

A community developer created Qwen3.8-23B-Mini-Me by strategically removing layers from Qwen3.8-27B, reducing the model to approximately 22.7B parameters without severe reasoning degradation. The model is reported to work well for coding, agentic tasks, and multi-turn chats, but it has not yet been benchmarked and struggles more with edge cases and underspecified prompts.

Reddit r/LocalLLaMACommunity1d ago#model-pruning#model-compression#apple-silicon
💼

AI Is Taking Jobs—Where Can Workers Go?

In an interview, Llama Ventures partner and Yidao founder Zhou Hang discusses how AI will reshape careers, entrepreneurship, and traditional industries. He argues that workers should develop commercial judgment and seek AI opportunities in overlooked sectors rather than simply compete for shrinking technical roles.

Page 2