Search

Few direct matches — filled in with the latest updates.

Tag: #tiny-models4 results

Qwen 0.8B Runs on Old S10E at 12 t/s

Qwen 0.8B Runs on Old S10E at 12 t/s

Qwen released its new 0.8B model, which runs locally on a 7-year-old Samsung S10E at 12 tokens per second using llama.cpp and Termux. After resolving missing C libraries, it handles conversations and complex tasks effectively. This demo highlights powerful tiny LLMs on edge devices.

Reddit r/LocalLLaMACommunityMar 2#on-device-inference#mobile-llm#tiny-models
🔬

Recursive Mamba Loops for Tiny Model Reasoning

Experimenter tests recursive looping on a 150M Mamba model to boost reasoning via hidden state feedback. Dynamic depth scaling mimics deeper models but hits 'Cognitive Static' at high loops, degrading language. Seeks community advice on SSM latent space issues.

Reddit r/LocalLLaMACommunityMar 16#ssm#recursion#reasoning