<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Langhuan Blog</title><description>Author narratives on RAG, MCP, knowledge bases, and traceable retrieval.</description><link>https://langhuan.dev/</link><language>en</language><item><title>How a Chinese Question Sentence Zeroed Out Our Full-Text Search</title><link>https://langhuan.dev/en/blog/chinese-fts-zero-recall/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/chinese-fts-zero-recall/</guid><description>An everyday question like &quot;埃及有哪些民族？&quot; returns literally nothing from SQLite FTS5 Chinese full-text search — and hybrid retrieval quietly degrades to vector-only, with no errors, no alerts, and normal-looking metrics. The mechanism, the query-side fix, and how to check whether your FTS channel is already spinning in neutral.</description><pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate></item><item><title>We Benchmarked Our Own RAG Retrieval — and the First Smoke Run Caught Two Production Bugs</title><link>https://langhuan.dev/en/blog/retrieval-eval-benchmark/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/retrieval-eval-benchmark/</guid><description>The metric that matters most for a knowledge base product is retrieval quality, and all we had were functional tests. How Langhuan built a retrieval eval with 200 human-annotated real queries, a dual-track channel matrix, and bit-identical reproducibility — plus the production bugs it caught on day one and the hybrid-search hypothesis it finally confirmed with data.</description><pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate></item><item><title>How to build a local RAG knowledge base: prove retrieval first, add infrastructure later</title><link>https://langhuan.dev/en/blog/local-rag-knowledge-base/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/local-rag-knowledge-base/</guid><description>A local RAG knowledge base does not need a full database and queue stack on day one. Start with real documents, Chinese-aware retrieval, and evidence; move to production components when the workload earns them.</description><pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate></item><item><title>An MCP knowledge base is not a folder for agents. It is retrieval made shared.</title><link>https://langhuan.dev/en/blog/mcp-knowledge-base-for-agents/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/mcp-knowledge-base-for-agents/</guid><description>The second agent changes the question. Instead of uploading the same documents again, decide who owns versioning, retrieval, access, and source evidence. That is where an MCP knowledge base belongs.</description><pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate></item><item><title>What Happens When Your Project Knowledge Base Hits Context Limits</title><link>https://langhuan.dev/en/blog/project-knowledge-context-limits/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/project-knowledge-context-limits/</guid><description>Stuffing every document into a Claude or ChatGPT project knowledge base eventually hits context limits — answers degrade first, uploads fail later. Symptoms, causes, and moving knowledge out of the context window.</description><pubDate>Mon, 17 Aug 2026 00:00:00 GMT</pubDate></item><item><title>Chinese hybrid search with pgvector + PostgreSQL full-text</title><link>https://langhuan.dev/en/blog/pgvector-chinese-hybrid-search/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/pgvector-chinese-hybrid-search/</guid><description>Store vectors in pgvector and tokenize Chinese with zhparser for full-text search, then fuse with RRF on a single SQL stack without a separate vector database.</description><pubDate>Thu, 13 Aug 2026 00:00:00 GMT</pubDate></item><item><title>Hybrid search explained: vector + full-text + RRF fusion</title><link>https://langhuan.dev/en/blog/hybrid-search-explained/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/hybrid-search-explained/</guid><description>Why pure vector search misses exact keywords, what full-text search adds, how RRF deterministically fuses two ranked lists, and what it costs.</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>MCP over HTTP vs stdio: which transport should your agent use</title><link>https://langhuan.dev/en/blog/mcp-over-http-vs-stdio/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/mcp-over-http-vs-stdio/</guid><description>Comparing stdio, SSE, and streamable HTTP transports for MCP: when to use each, and the deployment and security trade-offs.</description><pubDate>Tue, 11 Aug 2026 00:00:00 GMT</pubDate></item><item><title>What is MCP? A protocol for calling external capabilities from agents</title><link>https://langhuan.dev/en/blog/what-is-mcp/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/what-is-mcp/</guid><description>What the Model Context Protocol solves, what servers, clients, and tools are, and how a knowledge base becomes an MCP-callable service.</description><pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate></item><item><title>By the third enterprise agent, the knowledge base can&apos;t be rebuilt every time</title><link>https://langhuan.dev/en/blog/hello-langhuan/</link><guid isPermaLink="true">https://langhuan.dev/en/blog/hello-langhuan/</guid><description>When building several enterprise agents, I found the thing most often rebuilt isn&apos;t model capability — it&apos;s the ingestion, access control, retrieval, and evidence chain for knowledge.</description><pubDate>Wed, 05 Aug 2026 00:00:00 GMT</pubDate></item></channel></rss>