Tendril: A self-extending agent that builds and registers tools on the fly
Tendril is a self-extending agentic sandbox that demonstrates the Agent Capability pattern — the model discovers, builds, and reuses tools autonomously across sessions. Built with AWS Strands Agents SDK and Tauri.
How it works
You ask Tendril to do something. It checks its capability registry. If a tool exists, it uses it. If not, it writes one, registers it, and executes it — all without asking. Next time you need the same thing, the tool is already there.
You: "fetch the top stories from Hacker News"
Tendril: → searchCapabilities("fetch url hacker news") # nothing found
→ registerCapability(fetch_url, code) # builds a tool
→ execute("fetch_url", {url: "https://..."}) # runs it by name
→ "Here are the top stories: ..."
You: "now fetch Lobsters and compare"
Tendril: → listCapabilities() # found: fetch_url ✓
→ execute("fetch_url", {url: "https://lobste.rs"}) # runs it — no rebuild
The registry grows with use. Every session is smarter than the last.
Agent configuration
The core of Tendril is a Strands agent with just three bootstrap tools:
import { Agent } from '@strands-agents/sdk';
import { BedrockModel } from '@strands-agents/sdk/models/bedrock';
const agent = new Agent({
model: new BedrockModel({ modelId: '...', region: '...' }),
systemPrompt: TENDRIL_SYSTEM_PROMPT(workspacePath),
printer: nullPrinter,
tools: [
listCapabilities(registry),
registerCapability(registry),
executeCode(registry, workspacePath, config),
],
});
System prompt rules
The system prompt enforces autonomous behavior:
- Call
searchCapabilities(query)to check if a relevant tool exists - If found: call
loadTool(name)thenexecute(code, args) - If NOT found: you MUST build the tool yourself
- NEVER ask "would you like me to create a tool?" — just build it
- If a tool fails, read the error, fix the code, and retry
- NEVER answer from training data when a tool could get live information
Architecture
┌─────────────────────────────────────────┐ │ Tauri Shell (Rust) │ │ ACP Host ──stdin/stdout──► Agent │ │ (acp.rs) NDJSON (Node.js SEA)│ │ Events ◄── session/update ──┘ │ │ (events.rs) │ │ Tauri Events ──► React Frontend │ │ (TailwindCSS v4) │ └─────────────────────────────────────────┘Agent internals: Strands SDK ── BedrockModel ── Claude │ 4 bootstrap tools ┌────┴────┐ │ Registry │ ←→ index.json + tools/*.ts └─────────┘ ┌────┴────┐ │ Sandbox │ ←→ Deno subprocess (sandboxed)
The agentic loop runs inside agent.stream() and bridges to the ACP protocol, exposing think, act, and observe phases to the UI.
The "too many tools" solution
Most agent frameworks give the model a big bag of tools and hope it picks the right one. Tendril inverts this — the model always sees exactly three tools. It searches a registry, builds what it needs, and the registry grows over time. The tool surface never changes; the capabilities do.
📖 Read the full source: HN AI Agents
👀 See Also

OnUI: Browser Extension for Precise UI Feedback to Claude Code
OnUI is a browser extension that lets you annotate webpage elements and export structured reports for Claude Code via local MCP, eliminating ambiguous UI descriptions. Built primarily with Claude Code, it's free, open-source, and available for Chrome, Edge, and Firefox.

Vektori's Memory Architecture: Principles from Claude's Leaked System
Vektori implements a three-layer hierarchical sentence graph for AI memory, inspired by leaked principles from Claude's architecture. The system uses strict quality filters, skeptical retrieval with a 0.3 minimum score, and maintains correction history across sessions.

Claude Skills to Emulate a Design Studio Environment
A designer shares two Claude skills: one simulates a studio with teammates and design methods, the other adds 'rigorous play' for creativity.

Code Evolution Method Triples LLM Performance on ARC-AGI-2 Benchmark
Researchers achieved a 2.8x improvement on the ARC-AGI-2 benchmark using code evolution with open-weight models, reaching 34% accuracy at $2.67 per task. The same method pushed Gemini 3.1 Pro to 95% accuracy at $8.71 per task.