Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
Mnemonic: a local memory daemon that makes AI coding agents stop forgetting
Learn about Mnemonic, a local daemon providing AI coding agents persistent memory by capturing git commits and conversations, ensuring agents stop forgetting.
Mnemonic is an open-source background daemon that gives AI coding agents persistent memory. It passively captures my git commits and my conversations with Claude Code, stores them locally in SQLite with vector, full-text, and knowledge-graph indexes, and serves the relevant slice back to the agent over MCP, so the agent stops re-asking what we already decided.
In the demo I will show it live, end to end: the daemon capturing a real session, hybrid retrieval (BM25 + HNSW vectors + a graph hop, fused with Reciprocal Rank Fusion) answering a query, the recall@5 / recall@20 / MRR eval harness I use to check whether retrieval actually improved, and the macOS menu-bar widget reading the same local HTTP API. Everything runs on-device: no cloud, no API keys.
Local-first SQLite memory engine using HNSW vectors and MCP.
- RustRust is a high-performance systems programming language that guarantees memory and thread safety via its compile-time ownership model.Rust is a statically-typed systems language engineered for performance and reliability, directly challenging C/C++ in speed. Its core innovation is the ownership model and 'borrow checker,' which enforces strict memory and thread safety at compile-time, eliminating data races and null pointer dereferences without a conventional garbage collector. Rust achieves near-native speed through 'zero-cost abstractions,' allowing high-level features to compile into highly optimized code. Major industry players, including Microsoft and Cloudflare, leverage Rust for critical infrastructure, and it is now officially supported for development in the Linux kernel.
- SQLiteSQLite is a C-language library: a self-contained, serverless, zero-configuration SQL database engine embedded directly into the application process.SQLite is the world's most deployed database engine, functioning as a compact, C-language library (under 900KiB with all features) that eliminates the need for a separate server process. It operates as a serverless, zero-configuration system, storing the entire database (up to 281 terabytes) in a single, cross-platform file. This architecture makes it ideal for countless applications: it is built into all major mobile phones, web browsers, and desktop operating systems. The engine guarantees high reliability, supporting full ACID transactions, and its source code is freely available in the public domain for any use.
- HNSWHNSW (Hierarchical Navigable Small World) is a state-of-the-art graph-based algorithm: it executes Approximate Nearest Neighbor (ANN) search on high-dimensional vectors with logarithmic complexity (O(log n)), ensuring lightning-fast similarity retrieval.Hierarchical Navigable Small World (HNSW) is the dominant Approximate Nearest Neighbor (ANN) search algorithm, delivering superior speed and recall for vector databases. It constructs a multi-layer proximity graph: higher layers contain long-range connections for rapid traversal, while lower layers provide fine-grained accuracy for finding the true nearest neighbors. This hierarchical structure, detailed in the 2016 paper by Malkov and Yashunin, achieves logarithmic complexity scaling, making it highly efficient. Use it to power critical applications like large-scale image retrieval, real-time product recommendation engines, and modern Retrieval-Augmented Generation (RAG) systems.
- ONNX all-MiniLM-L6-v2An ultra-lean, 384-dimensional sentence embedding model optimized in ONNX format for lightning-fast CPU and edge-device semantic search.This technology packages the highly popular sentence-transformers/all-MiniLM-L6-v2 model into the portable ONNX runtime format, shrinking the deployment footprint to a mere 80 megabytes. It maps sentences and paragraphs into a dense 384-dimensional vector space, making it a go-to choice for clustering, duplicate detection, and retrieval-augmented generation (RAG) pipelines. By bypassing heavy Python dependencies, developers can run local, high-throughput semantic similarity searches directly in C++, Java, or browser-based JavaScript environments with minimal CPU overhead.
- MCPMCP is the open-source standard for securely connecting AI agents (like LLMs) to external tools, data, and enterprise workflows.The Model Context Protocol (MCP) functions as a standardized integration layer: think of it as a USB-C port for AI applications. Developed and open-sourced by Anthropic, this protocol allows large language models (LLMs) to access real-time context and execute actions via external tools like GitHub, Jira, or proprietary databases . It uses a simple JSON-RPC interface to define tools, schemas, and endpoints, which enables AI agents to perform complex, state-changing tasks—such as creating a GitHub issue or running a test script—rather than just generating text . MCP is essential for building agentic AI systems that can autonomously pursue goals and operate within defined safety and permission boundaries .
Related talks
More from the community
Agent Memory Is the Softest Attack Surface. Let Me Show You.
Minneapolis Saint Paul
Discover how agent memory, crucial for AI, creates vulnerabilities. This talk demonstrates live attacks on multi-tenant memory and…
Persistent Semantic Memory for AI Agents — SQLite-vec Instead of a Vector DB
Zürich
Discover how to give AI agents persistent semantic memory using SQLite-vec, avoiding separate vector databases. This talk covers…
ArgosBrain pushed Opus 4.7 from 87.6% to 95% on SWE-bench Verified
Paris
Learn how to improve AI coding agent efficiency by amplifying grep and Read with structural facts, reducing redundant…
Moka: A Personal Agent OS
Toronto
Discover Moka, a personal AI agent OS running on a private server via Telegram. See how its filesystem…
Building Memory That Agents Can Act On
Seattle
See how to build AI agent memory beyond chat history. This talk shows a system using typed objects…
The Second Brain That Runs My Companies: Live Demo of an Open Source AI Operating Layer
Bogotá
See a live demo of an open-source AI operating layer that manages companies, writing, and life through a…
Compose Email
Loading recent emails...