ETL-D MCP Server: Deterministic CSV Parsing for Claude to Prevent Financial Hallucinations

A developer has open-sourced ETL-D, an MCP server for Claude Desktop designed to prevent Claude from hallucinating decimal points when parsing financial CSVs and other structured B2B data formats. The tool addresses the "token tax" of sending raw formats to an LLM's context window and the "hallucination risk" where misplaced commas can turn $100.50 into $10,050.00.
Architecture: The Three-Layer Waterfall
The server processes files through three strict layers when Claude is asked to parse them:
- Layer 1 (Heuristics): Uses 100% Python with
regex,dateutil, and strict structural parsers for known formats. The developer reports a load test with 200 parallel requests achieving ~70ms response times with 0 LLM calls and zero hallucination risk. - Layer 2 (Semantic Routing): If CSV headers are obfuscated, a lightweight router maps columns to strict Pydantic schemas.
- Layer 3 (LLM Fallback): Only triggered for high-entropy "free-text" noise, using Llama 3.3 70b under the hood to enforce JSON schemas.
The result is a perfectly clean, flattened JSON array returned to Claude for reasoning.
Setup and Availability
The tool has been approved on the official Anthropic MCP Registry. To use it, developers need to configure their claude_desktop_config.json. The source code is available on GitHub at pablixnieto2/etld-mcp-server.
The developer built this after identifying that "LLM-first" is the wrong architecture for structured B2B data like broker trade histories, bank statements (Norma 43), or SEC XBRL files, arguing that AI agents shouldn't read CSVs directly but should query deterministic middleware instead.
📖 Read the full source: r/ClaudeAI
👀 See Also

Developer builds AI framework with 17 biological principles using Claude Code
A developer created an AI framework called Cognitive Sparks by implementing 17 biological principles like threshold firing and Hebbian plasticity, based on the 1999 book 'Sparks of Genius.' The entire project—22 design docs and 3,300 lines of code—was built in one day using Claude Code, with no human-written code.

Team Brain: A Shared Memory Plugin for Claude Code That Stores Team Knowledge in Git
Team Brain is a Claude Code plugin that stores team knowledge in a .team-brain/ folder within your repository. It automatically generates a BRAIN.md file capped at 180 lines for optimal Claude instruction accuracy and works across tools by creating .cursorrules and AGENTS.md files.

Claude Code's File-Based Memory System: A Pragmatic Alternative to Vector DBs
Claude Code implements a file-based memory system using .md files with frontmatter metadata and a MEMORY.md index, avoiding vector databases and embedding pipelines by scanning files, building manifests, and using a small model to select relevant memories.

Approval Boundary Tool for Claude Code Repository Work
A developer built an approval boundary tool that adds a review step before local execution when using Claude Code for repository work. The tool follows a loop: see the plan first, approve once, let the run happen locally, and keep proof afterward.