30 Aug 2026 · 4 min read
ai-tools
GitHub Copilot's New Thinking-Effort Dial: Managing Cost and Latency Per Task, Not Per Team
GitHub's August 28 Visual Studio update adds Low/Medium/High thinking-effort controls per model, plus a model management view with context window and cost info. A Tech Lead's take on why per-call effort tuning matters more than picking 'the best model' — and how it compares to the reasoning-effort trap we already learned the hard way with open models.
Read more




