<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>Posts on Hermes</title>
    <link>https://www.1o1.men/posts/</link>
    <description>Recent content in Posts on Hermes</description>
    <generator>Hugo</generator>
    <language>zh-cn</language>
    <lastBuildDate>Sat, 15 Aug 2026 00:00:00 +0800</lastBuildDate>
    <atom:link href="https://www.1o1.men/posts/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>全中文 Honcho 记忆系统：从裸机到全链路中文化的部署实践</title>
      <link>https://www.1o1.men/posts/honcho-zh/</link>
      <pubDate>Sat, 15 Aug 2026 00:00:00 +0800</pubDate>
      <guid>https://www.1o1.men/posts/honcho-zh/</guid>
      <description>&lt;p&gt;Honcho 是 Plastic Labs 开源的 AI Agent 持久化记忆引擎。它以 PostgreSQL + pgvector 存储记忆向量，用 LLM 做辩证推理（dialectic），为 Agent 提供跨会话的用户画像、语义搜索与推理能力——让 AI 助手真正&amp;quot;记住你、了解你、主动调用关于你的结论&amp;quot;。&lt;/p&gt;
&lt;p&gt;但它原生的一切都是英文的：Prompt、输出格式、文档。在国内网络、DeepSeek 模型、纯中文需求的叠加下，部署过程布满隐蔽陷阱。本文记录我们把它完整中文化、并沉淀成一套可复用技能的过程。&lt;/p&gt;
&lt;h2 id=&#34;架构&#34;&gt;架构&lt;/h2&gt;
&lt;p&gt;PostgreSQL（pgvector）← Honcho API（:8000）&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Deriver&lt;/strong&gt; —— 实时观察提取&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Embedding&lt;/strong&gt; —— bge-base-zh-v1.5（768 维，:8080）&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Dialectic&lt;/strong&gt; —— 辩证推理（5 个级别）&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Dream&lt;/strong&gt; —— 离线画像演绎（deduction + induction）&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Summarizer&lt;/strong&gt; —— 会话压缩&lt;/li&gt;
&lt;/ul&gt;
&lt;h2 id=&#34;踩过的坑&#34;&gt;踩过的坑&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;&lt;code&gt;pip install honcho&lt;/code&gt; 装错包——PyPI 上的是进程管理器，不是记忆系统，必须从 GitHub 克隆。&lt;/li&gt;
&lt;li&gt;国内网络下 GitHub 与 HuggingFace 都无法直连，源码克隆与模型下载都要走代理。&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;DeepSeek 的 structured_output 静默失败&lt;/strong&gt;：Honcho 默认用 &lt;code&gt;json_schema&lt;/code&gt;，DeepSeek 不支持，9 个 model_config 块必须全部改成 &lt;code&gt;json_object&lt;/code&gt;，缺一个 Deriver 就静默产出空观察。&lt;/li&gt;
&lt;li&gt;Embedding 维度不匹配：alembic 默认建 1536 维列，bge-base-zh-v1.5 是 768 维，需显式调整。&lt;/li&gt;
&lt;li&gt;HF 下载模型用 &lt;code&gt;cache_dir&lt;/code&gt; 会 symlink 断链，必须用 &lt;code&gt;local_dir&lt;/code&gt; 平铺。&lt;/li&gt;
&lt;li&gt;9 处英文 Prompt 中文化，漏一处就出现中英混杂。&lt;/li&gt;
&lt;li&gt;复用 Hermes 自己的 venv 跑 Honcho——Hermes 升级会重建 venv，把 Honcho 依赖（sqlalchemy、sentry-sdk、sentence-transformers）一并冲掉，三个服务全崩、embedding 空转十多万次。必须给 Honcho 建独立 venv（&lt;code&gt;uv sync&lt;/code&gt;），彻底与 Hermes 解耦。&lt;/li&gt;
&lt;/ul&gt;
&lt;h2 id=&#34;成果&#34;&gt;成果&lt;/h2&gt;
&lt;p&gt;最终跑通了一条&lt;strong&gt;纯中文管道&lt;/strong&gt;：观察（observation）、画像（peer card）、语义搜索、辩证推理全部中文输出。整个方案沉淀为 &lt;strong&gt;agent-honcho-zh&lt;/strong&gt; 技能，从裸机初始化到全链路中文化一步到位，还附带 Hermes 插件的一个 observer bug 修复（已提交上游 PR）。&lt;/p&gt;</description>
    </item>
  </channel>
</rss>
