Deblank: Tool to Strip Code Formatting for LLM Token Reduction

✍️ OpenClawRadar📅 Published: March 23, 2026🔗 Source
Deblank: Tool to Strip Code Formatting for LLM Token Reduction
Ad

What Deblank Does

Deblank is a preprocessing tool that removes code formatting (indentation, whitespace, line breaks) before sending code to LLMs, with a postprocessing step to restore readability. The transformation is bidirectional and AST-safe.

Performance Results

In evaluations across several models (DeepSeek-V3, Claude, Gemini, etc.):

  • ~30% token reduction for languages like Java and C++
  • ~9% token reduction for Python
  • Negligible impact on Pass@1 accuracy for code completion
  • Average latency: ~76ms

Supported Languages and Features

  • Python, Java, C/C++, C#, JavaScript/TypeScript, and Go
  • Handles incomplete snippets reasonably well
  • Useful for fill-in-the-middle workflows
Ad

Getting Started

The project is open-sourced with these resources:

This type of token optimization can be particularly useful when working with context-limited LLMs or when processing large codebases, though the impact varies by language due to differences in formatting conventions.

📖 Read the full source: r/LocalLLaMA

Ad

👀 See Also

German Bureaucracy Assistant Prompt for Claude: Structured Legal Correspondence
Tools

German Bureaucracy Assistant Prompt for Claude: Structured Legal Correspondence

A detailed system prompt for Claude that turns the AI into a structured assistant for German bureaucracy, contracts, insurance disputes, and official letters, with strict fact-checking and DIN 5008 formatting.

OpenClawRadar
Open Source Auto-Memory System for LLM Agents Achieves 94% Recall Accuracy
Tools

Open Source Auto-Memory System for LLM Agents Achieves 94% Recall Accuracy

A developer built a memory plugin for LLM-based agents that automatically extracts, classifies, and persists facts across sessions without explicit user commands. The system achieved 94.2% accuracy on a 52-checkpoint recall benchmark using structured markdown files instead of vector databases.

OpenClawRadar
Pneuma: An AI-Generated Desktop Environment Where Software Materializes from Descriptions
Tools

Pneuma: An AI-Generated Desktop Environment Where Software Materializes from Descriptions

Pneuma is a desktop computing environment where you describe what you want—a CPU monitor, game, notes app, or data visualizer—and a working program materializes in seconds. The system generates self-contained Rust modules, compiles them to WebAssembly, and executes them in sandboxed Wasmtime instances with GPU rendering via wgpu.

OpenClawRadar
Krasis LLM Runtime Shows 8.9x Prefill and 4.7x Decode Speed Improvements Over Llama.cpp
Tools

Krasis LLM Runtime Shows 8.9x Prefill and 4.7x Decode Speed Improvements Over Llama.cpp

Krasis LLM runtime now runs both prefill and decode entirely on GPU with different optimization strategies, achieving 8.9x faster prefill and 4.7x faster decode than llama.cpp on Qwen3.5-122B with a single 5090 GPU.

OpenClawRadar