Selective Reasoning Disable in llama.cpp?
User asks if llama.cpp with llama-server allows disabling reasoning for specific requests while keeping it enabled by default. Running unsloth/gemma-4-26B-A4B-it-GGUF model. Aimed at faster responses for chatbots.






