<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>DEEP | Sirchmunk</title><link>https://modelscope.github.io/sirchmunk-web/tags/deep/</link><atom:link href="https://modelscope.github.io/sirchmunk-web/tags/deep/index.xml" rel="self" type="application/rss+xml"/><description>DEEP</description><generator>HugoBlox Kit (https://hugoblox.com)</generator><language>en-us</language><lastBuildDate>Sun, 20 Sep 2026 00:00:00 +0000</lastBuildDate><image><url>https://modelscope.github.io/sirchmunk-web/media/logo_hu_60b56fdaf40cfb0f.png</url><title>DEEP</title><link>https://modelscope.github.io/sirchmunk-web/tags/deep/</link></image><item><title>Sirchmunk v0.2.0: LENS Paper, Multi-Path DEEP Retrieval &amp; Large Corpus Robustness</title><link>https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/</link><pubDate>Sun, 20 Sep 2026 00:00:00 +0000</pubDate><guid>https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/</guid><description>&lt;p&gt;Sirchmunk v0.2.0 marks a turning point for the project. The core algorithm behind Sirchmunk&amp;rsquo;s retrieval engine has been formalized in a research paper and published on arXiv, the DEEP search pipeline has been rebuilt around multi-path fusion, and the system now enforces strict retrieval cost invariants that keep performance predictable on corpora of any size.&lt;/p&gt;
&lt;h2 id="lens-the-research-paper"&gt;LENS: The Research Paper&lt;/h2&gt;
&lt;p&gt;The theoretical foundations of Sirchmunk&amp;rsquo;s in-context search have been formalized in &lt;strong&gt;&amp;ldquo;LENS: In-Context Search via Latent Evidence Exploration over Dynamic Raw Documents&amp;rdquo;&lt;/strong&gt; (
).&lt;/p&gt;
&lt;p&gt;LENS reframes retrieval as &lt;em&gt;Budgeted Evidence Localization&lt;/em&gt;: given a query and a raw-document corpus, the system maintains a query-conditioned belief over a latent evidence space and iteratively refines it through proposal policies and an LLM relevance oracle — all under an explicit token budget.&lt;/p&gt;
&lt;p&gt;Key results from the controlled evaluation:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;500-question evaluation&lt;/strong&gt;: 62.4% Exact Match with 84.8% evidence recall (vs. ReAct baseline at 65.2% EM but only 50.4% evidence recall).&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;150-question fullwiki subset&lt;/strong&gt; (raw Wikipedia dump, zero indexing): LENS achieves 43.3% EM vs. ReAct&amp;rsquo;s 42.7% EM, with substantially stronger evidence grounding (84.0% vs. 70.7%).&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;These numbers demonstrate that LENS trades a modest amount of accuracy for dramatically better evidence traceability — the answer is not only correct, but &lt;em&gt;provably grounded&lt;/em&gt; in source material.&lt;/p&gt;
&lt;h2 id="multi-path-deep-retrieval"&gt;Multi-Path DEEP Retrieval&lt;/h2&gt;
&lt;p&gt;DEEP mode has been fundamentally restructured. Instead of a single retrieval strategy, it now runs &lt;strong&gt;five complementary retrieval paths&lt;/strong&gt; in parallel:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;Lexical&lt;/strong&gt; — keyword-driven content matching with IDF-weighted scoring.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Entity&lt;/strong&gt; — exact-match probes for named entities, identifiers, and structured values.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Directory&lt;/strong&gt; — file-system structure analysis using naming conventions, path hierarchy, and modification timestamps.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Structural&lt;/strong&gt; — heuristic document tree navigation (v2) with structure anchors for DOCX, RST, and other formatted sources.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Topic-Graph&lt;/strong&gt; — cross-document topic-map routing that leverages the self-evolving knowledge graph to identify relevant document clusters.&lt;/li&gt;
&lt;/ol&gt;
&lt;p&gt;Results from all five paths are fused via &lt;strong&gt;confidence-weighted Reciprocal Rank Fusion (RRF)&lt;/strong&gt;. A &lt;strong&gt;soft route-collapse&lt;/strong&gt; mechanism dynamically disables low-yield paths when a single path produces a high-confidence match, cutting latency and token consumption without sacrificing answer quality.&lt;/p&gt;
&lt;p&gt;
&lt;figure id="figure-sirchmunk-v020-system-architecture-with-multi-path-deep-retrieval-and-lens-framework-integration"&gt;
&lt;div class="flex justify-center "&gt;
&lt;div class="w-full" &gt;
&lt;img alt="Sirchmunk Architecture"
srcset="https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Sirchmunk_Architecture_hu_20a06e55d6b5f4b.webp 320w, https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Sirchmunk_Architecture_hu_85555c5afe12fe9a.webp 480w, https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Sirchmunk_Architecture_hu_a4dc28c9d30189b.webp 760w"
sizes="(max-width: 480px) 100vw, (max-width: 768px) 90vw, (max-width: 1024px) 80vw, 760px"
src="https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Sirchmunk_Architecture_hu_20a06e55d6b5f4b.webp"
width="760"
height="328"
loading="lazy" data-zoomable /&gt;&lt;/div&gt;
&lt;/div&gt;&lt;figcaption&gt;
Sirchmunk v0.2.0 system architecture with multi-path DEEP retrieval and LENS framework integration.
&lt;/figcaption&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h2 id="large-corpus-robustness"&gt;Large Corpus Robustness&lt;/h2&gt;
&lt;p&gt;Previous versions could stall or time out on very large, archive-heavy corpora. v0.2.0 introduces &lt;strong&gt;retrieval cost invariants&lt;/strong&gt; — hard bounds that ensure per-file and per-query cost never grows unbounded:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;&lt;code&gt;GREP_RGA_ADAPTERS&lt;/code&gt;&lt;/strong&gt;: Capability-based adapter whitelist. Only bounded document extractors (poppler, pandoc) are enabled on the query hot path; unbounded recursive adapters (decompress, zip, tar, sqlite, ffmpeg) are restricted to offline extraction.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;&lt;code&gt;GREP_MAX_FILESIZE_MB&lt;/code&gt;&lt;/strong&gt;: Per-file size cap. Files exceeding the threshold are skipped during live queries.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;&lt;code&gt;GREP_TIERED_SCAN&lt;/code&gt;&lt;/strong&gt;: A fast native-&lt;code&gt;rg&lt;/code&gt; pass over all files is unioned with an &lt;code&gt;rga&lt;/code&gt; pass restricted to rich-format extensions, so adapter dispatch never walks the entire tree.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Fail-fast timeout budgets&lt;/strong&gt;: &lt;code&gt;GREP_TEXT_TIMEOUT&lt;/code&gt; for the text pass and &lt;code&gt;GREP_TIMEOUT&lt;/code&gt; for the rich pass; on timeout, search degrades gracefully to native &lt;code&gt;rg&lt;/code&gt; rather than hanging.&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Directory scanning is now enabled by default for stronger filename routing, and a hard token budget keeps each query within a configurable limit.&lt;/p&gt;
&lt;h2 id="generalization-first-design"&gt;Generalization-First Design&lt;/h2&gt;
&lt;p&gt;v0.2.0 adopts a principled stance against benchmark-specific hard rules. Benchmark-tied logic such as hardcoded entity patterns, fixed column positions, and English-only stop words has been replaced with generalizable alternatives:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Grounded numeric verification&lt;/strong&gt;: Computation answers are re-checked deterministically from model-disclosed, evidence-grounded operands — entirely corpus-agnostic.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Injectable tokenizers&lt;/strong&gt;: Tokenizers and lexical policies are dependency-injected with a general default implementation, rather than embedded inside modules.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Corpus-adaptive statistics&lt;/strong&gt;: Document-frequency-based adaptive stop-word pruning replaces fixed word lists, working natively across languages and domains.&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Every replacement behavior is gated behind an environment switch, defaults to the new implementation, and allows single-item rollback on failure.&lt;/p&gt;
&lt;h2 id="knowledge-graph-visualization"&gt;Knowledge Graph Visualization&lt;/h2&gt;
&lt;p&gt;The Web UI now includes an &lt;strong&gt;interactive knowledge graph&lt;/strong&gt; powered by Cytoscape.js. The graph visualizes self-evolving knowledge clusters and their semantic relationships, including lifecycle states (Emerging, Stable, Meta) and edge weights. Users can explore, filter, and drill down into the cluster topology to understand how the system&amp;rsquo;s knowledge evolves with use.&lt;/p&gt;
&lt;p&gt;Behind the visualization, the &lt;code&gt;KnowledgeEvolver&lt;/code&gt; orchestrates a four-phase background evolution cycle — Connect &amp;amp; Merge, Refresh Edges, Detect Meta Clusters (via Leiden community detection), and Global Update — that continuously maintains and consolidates the knowledge graph without blocking queries.&lt;/p&gt;
&lt;p&gt;
&lt;figure id="figure-knowledgeevolver--four-phase-evolution-cycle-for-knowledge-graph-maintenance"&gt;
&lt;div class="flex justify-center "&gt;
&lt;div class="w-full" &gt;
&lt;img alt="Knowledge Evolver Architecture"
srcset="https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Knowledge_Evolver_Architecture_hu_4f04be0dfa1ca9b3.webp 320w, https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Knowledge_Evolver_Architecture_hu_43806d4cd9a45d36.webp 480w, https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Knowledge_Evolver_Architecture_hu_72c87b060cde1c77.webp 760w"
sizes="(max-width: 480px) 100vw, (max-width: 768px) 90vw, (max-width: 1024px) 80vw, 760px"
src="https://modelscope.github.io/sirchmunk-web/blog/v0.2.0/Knowledge_Evolver_Architecture_hu_4f04be0dfa1ca9b3.webp"
width="760"
height="428"
loading="lazy" data-zoomable /&gt;&lt;/div&gt;
&lt;/div&gt;&lt;figcaption&gt;
KnowledgeEvolver — Four-phase evolution cycle for knowledge graph maintenance
&lt;/figcaption&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;The design philosophy is observation-driven: evolution is triggered by actual search patterns rather than pre-defined rules, and source fidelity is always preserved — the system changes how it navigates to evidence, never the evidence itself. For a full animated demonstration, see the
.&lt;/p&gt;
&lt;h2 id="other-improvements"&gt;Other Improvements&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Broader format coverage&lt;/strong&gt;: Native exact-match fallback for LOG, PPTX, and XLSX files, plus heuristic document tree v2 (including DOCX/RST) with structure anchors guiding evidence extraction.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;DeepSeek V4 compatibility&lt;/strong&gt;: Full support for DeepSeek V4&amp;rsquo;s thinking mode (&lt;code&gt;thinking_content&lt;/code&gt;) in the OpenAI-compatible client.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Stabilized large-corpus retrieval&lt;/strong&gt;: Combined improvements in tiered scanning, per-file match caps, and adapter whitelisting eliminate the timeout and stall issues observed on large corpora in earlier versions.&lt;/li&gt;
&lt;/ul&gt;
&lt;h2 id="get-started"&gt;Get Started&lt;/h2&gt;
&lt;div class="highlight"&gt;&lt;pre tabindex="0" class="chroma"&gt;&lt;code class="language-bash" data-lang="bash"&gt;&lt;span class="line"&gt;&lt;span class="cl"&gt;pip install --upgrade sirchmunk
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;Or install with all extras:&lt;/p&gt;
&lt;div class="highlight"&gt;&lt;pre tabindex="0" class="chroma"&gt;&lt;code class="language-bash" data-lang="bash"&gt;&lt;span class="line"&gt;&lt;span class="cl"&gt;pip install &lt;span class="s2"&gt;&amp;#34;sirchmunk[all]&amp;#34;&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Documentation&lt;/strong&gt;:
&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;GitHub&lt;/strong&gt;:
&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Paper&lt;/strong&gt;:
&lt;/li&gt;
&lt;/ul&gt;
&lt;hr&gt;
&lt;p&gt;&lt;em&gt;
·
&lt;/em&gt;&lt;/p&gt;</description></item></channel></rss>