跳到正文

javieraltman

AIrecall

Drop-in long-term memory layer for AI agents - episodic + semantic memory, hybrid retrieval, and auto-summarization. Python SDK + Go memory server.

README 已保存到本站,可直接阅读

Documentation snapshot

README 快照

这篇是英文原文

下面正文是项目自己的英文 README。想读全文就用浏览器自带的整页翻译: Chrome / Edge 点地址栏右侧的翻译图标,或用右键菜单里的「翻译成中文」; 手机浏览器一般在菜单里。

本页保存的是公开项目资料快照,阅读过程不需要连接 GitHub。

AIrecall

A drop-in long-term memory layer for AI agents. Episodic + semantic memory, hybrid retrieval, and auto-summarization — so your agent remembers what matters across sessions.

pip install airecall-sdk
airecall init
LanguagesPython SDK + Go memory server
StorageSQLite + in-process vector index
LicenseMIT
DependenciesSDK stdlib-only · server one pure-Go dep

Why This Exists

Most agent frameworks give you a context window, not a memory. Close the session and everything the agent learned — user preferences, past decisions, corrections it was given — is gone. Stuff it all into the prompt instead, and you’re paying for and diluting your context with old information, most of which is irrelevant to the current turn.

AIrecall sits between your agent and a persistent store, and gives it an actual memory:

🧠 Episodic memoryWhat happened, in order (conversations, actions, outcomes)
📌 Semantic memoryDurable facts and preferences distilled out of those episodes
🔎 Hybrid retrievalKeyword + vector search pulls back what matters for this turn
🗜️ Auto-summarizationOld episodes are compressed, not deleted — long-term memory stays cheap

The goal is a memory layer that’s boring to integrate and hard to notice — until you turn it off and the agent forgets your name.


How It Works

        +---------------------+        local call / gRPC        +----------------------+
        |      Agent code      | -------------------------------> |     AIrecall Core     |
        |    (Python SDK)      | <------------------------------- |   (memory server)     |
        +---------------------+                                  +-----------+----------+
                                                                             |
                                                                 +-----------v----------+
                                                                 |    Storage Engine      |
                                                                 |  SQLite + vector idx   |
                                                                 +-----------+----------+
                                                                             |
                                                                 +-----------v----------+
                                                                 |   Optional MCP         |
                                                                 |   adapter              |
                                                                 +-----------------------+

On every turn, the SDK sends the current query to the core, which does a hybrid retrieval pass — keyword + vector similarity — over both episodic and semantic memory, and returns the top-k relevant memories to inject into your prompt. In the background, a summarizer periodically walks older episodes, extracts durable facts into semantic memory, and compacts the rest.


Install

# Python SDK - what your agent code imports
pip install airecall-sdk

# Core memory server - runs locally or as a sidecar
pip install airecall-sdk[server]
# or run it standalone:
airecall serve

airecall init scaffolds a local SQLite-backed store so you can start storing and recalling memories immediately, no separate service required.


Quickstart

from airecall import Memory

memory = Memory(agent_id="support-bot")

# Store an episode
memory.remember(
    "User asked about refund policy for order ORD-9921, told 30-day window applies."
)

# Later, in a new session
context = memory.recall("what did we tell this user about refunds?")
# -> returns the relevant episodic memory, ranked by relevance

# Promote a durable fact explicitly
memory.remember_fact(key="preferred_contact", value="email, not phone")

Framework adapters:

from airecall.adapters.langchain import MemoryRetriever

retriever = MemoryRetriever(memory)

## Troubleshooting

- **`RecallTimeout` on first query** - the index is still warming. Retry after the health endpoint reports `status: ready`.
- **Missing memories after restart** - check the `storage.backend` path; a relative path resolves against the working directory of the server process.

Official distribution

获取与安装

暂未发现可确认的官方软件包地址

当前 README 快照没有出现 npm、PyPI、Crates.io、pub.dev 等官方包页链接。本站不会根据仓库名称猜测下载地址。

本站不托管项目文件;需要安装时,请以项目维护者发布的官方文档为准。

使用前核验

本站保存公开资料用于阅读,不代表安全审计或功能背书。安装前请核对许可证、依赖来源和发布签名,不要直接运行来源不明的二进制文件或高权限脚本。