Different models need different tuning for natural conversation: Qwen gets high frequency penalty to fight repetition loops, Llama gets warmer temp to reduce terseness, Grok/Mistral/DeepSeek/Kimi get slightly warmer than Sonnet defaults. Bumps base httpx timeout from 10s to 30s and fallback per-call timeout from 8s to 20s. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>