Claude Opus 4.6 effort=low parameter causes lazy agent behavior

Claude Opus 4.6's effort parameter behaves differently than similar settings from other AI providers, causing unexpected agent behavior when set to low.
Key Findings
Testing revealed that with effort=low, Claude Opus 4.6 exhibited significantly lazier behavior than expected:
- Made fewer tool calls
- Was less thorough in cross-referencing
- Effectively ignored parts of system prompts instructing how to do web research
- Confidently returned wrong answers because it stopped looking for information
The source notes that bumping to effort=medium fixed all these issues. According to the documentation, Anthropic's effort parameter controls general behavioral effort, not just reasoning depth like OpenAI's reasoning.effort=low or Gemini's thinking_level=low.
Important Distinction
This isn't a bug but a documented difference in implementation. The effort parameter in Claude Opus 4.6 has broader scope than equivalent parameters from other providers. This means you can't treat effort as a drop-in replacement for reasoning.effort or thinking_level when working across different AI providers.
The testing was conducted with the expectation that effort=low would behave similarly to other providers' low-effort settings, but the actual behavior was more extreme, leading to agents that were not just thinking less but acting lazier overall.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Xiaomi Open-Sources MiMo-V2.5-Pro: Nears Claude Opus 4.6 on Coding Benchmarks
Xiaomi released MiMo-V2.5-Pro, an open-source coding model that scored 233/233 on a university compiler project, built a video editor autonomously, and ranks within 1% of Claude Opus 4.6 on SWE-Bench and Terminal-Bench.

Claude Code OAuth Login Timeout Bug on Windows
Claude Code version 2.1.92 has a bug where Windows users experience OAuth login failures with a timeout error of 15000ms, completely blocking access to the AI coding assistant.

Gemini 3.1 Flash Live: Google's latest audio model with improved benchmarks and watermarking
Google released Gemini 3.1 Flash Live, an audio model scoring 90.8% on ComplexFuncBench Audio and 36.1% on Scale AI's Audio MultiChallenge. It's available via Gemini Live API in Google AI Studio and includes SynthID watermarking.

Claude's policy filter blocks bioinformatics work with pathogen names
A computational virology researcher reports Claude's usage policy filter flags legitimate bioinformatics scripts when pathogens are named, requiring workarounds like describing tasks without organism names or downgrading to Sonnet 4. The issue affects Claude Code, claude.ai, and both Opus 4.6 and Sonnet 4.6 models.