Benchmark Shows CLI Tool Reduces Claude Code Token Costs by 32% Through Structural Navigation

A developer has open-sourced a CLI tool called Scope that provides Claude Code agents with structural code navigation capabilities, similar to IDE features like "find references" and "go to definition." The tool was built in Rust using tree-sitter and SQLite.
What the Tool Does
The tool gives agents commands like:
- "show me a 180-token summary of this 6,000-token class"
- "search by what code does, not what it's named"
It currently supports TypeScript and C#, with the goal of helping agents navigate code more efficiently than their default grep-based approach.
Benchmark Methodology
The developer ran 54 automated runs on Sonnet 4.6 across a 181-file C# codebase with:
- 6 task categories
- 3 conditions: baseline, tool available, architecture preloaded into CLAUDE.md
- 3 repetitions each
Full NDJSON capture was recorded on every run to decompose tokens into fresh input, cache creation, cache reads, and output. The benchmark runner and telemetry capture are included in the repository.
Key Findings
Contrary to expectations, agents with the tool read more files (6.8 to 9.7 average vs. baseline) but made 67% more code edits per session and finished in fewer turns.
The savings came from shorter conversations, which reduced cache accumulation. Approximately 90% of token cost lives in cache accumulation.
Overall results:
- 32% lower cost per task
- 2x navigation efficiency (nav actions per edit)
- Navigation-to-edit ratio improved from 25:1 (baseline) to 13:1 (with tool) and 12:1 (with architecture preloaded)
Results varied by task type:
- Bug fixes: -62% cost
- New features: -49% cost
- Cross-cutting changes: -46% cost
- Discovery and refactoring tasks: no advantage (baseline agents already navigate these fine)
Important Caveats
The developer notes several limitations:
- p-values don't reach 0.05 at n=6 paired observations (direction is consistent but sample is too small for statistical significance)
- Benchmarked on C# only so far (TypeScript support exists but hasn't been benchmarked yet)
- Cost calculation uses current Sonnet 4.6 API rates: fresh input $3/M, cache write $3.75/M, cache read $0.30/M, output $15/M
The tool is open source and available at github.com/rynhardt-potgieter/scope for developers who want to experiment with improving agent token efficiency.
📖 Read the full source: r/ClaudeAI
👀 See Also

CrabMeat v0.1.0: LLM이 보안 경계를 신뢰하지 않는 보안 우선 에이전트 게이트웨이
CrabMeat v0.1.0은 에이전틱 LLM 워크로드를 위한 WebSocket 게이트웨이로, 아키텍처 수준에서 보안을 강화합니다: 기능 ID 간접화, 효과 클래스, IRONCLAD_CONTEXT 고정 명령어, 변조 감지 감사 체인, 스트리밍 출력 누출 필터, 그리고 YOLO 모드 없음.

넥서스: 발견, 신뢰, 결제 기능을 갖춘 오픈소스 AI 간 통신 프로토콜
넥서스는 AI 에이전트가 인간의 개입 없이 서로를 발견하고, 조건을 협상하며, 응답을 검증하고, 소액 결제를 처리할 수 있도록 하는 자체 호스팅 프로토콜입니다. 이는 발견, 신뢰, 프로토콜, 라우팅, 연합의 5개 계층으로 구성되어 있으며, 66개의 테스트와 MIT 라이선스를 포함합니다.

뇌: MCP를 통한 Claude 코드용 지속적 오류 메모리 시스템
Brain은 Claude Code에 오류와 해결책에 대한 지속적이고 프로젝트 간 메모리를 제공하는 오픈소스 MCP 서버입니다. 오류 컨텍스트를 포착하고, 신뢰도 점수와 함께 검증된 해결책을 제안하며, 모든 프로젝트에 걸쳐 오류, 해결책, 코드 모듈을 연결하는 가중 시냅스 네트워크를 구축합니다.

SwiftUI와 CSM-1B로 Apple Silicon에서 로컬 음성 AI 어시스턴트 구축하기
개발자가 mobiGlas를 만들었습니다. 이는 SwiftUI 앱으로, OpenClaw와 연동하여 AirPods을 통한 핸즈프리 대화를 가능하게 하며, 로컬 음성 복제(CSM-1B on M2 Ultra)를 사용하고 클라우드 API가 필요 없습니다.