Give any local LLM a billion-token memory. Open-source, local-first context + memory engine for Ollama and AI agents — runs on your own machine, free. 🧠⚡

26 stars 1 forks 26 watchers Python Apache License 2.0
ai-agents context-engineering infinite-context llama llm llm-memory local-first local-llm local-rag long-context memory mistral ollama ollama-rag persistent-memory python qwen rag retrieval-augmented-generation vector-search
5 Open Issues Need Help Last updated: Aug 13, 2026

Open Issues Need Help

View All on GitHub

Give any local LLM a billion-token memory. Open-source, local-first context + memory engine for Ollama and AI agents — runs on your own machine, free. 🧠⚡

Python
#ai-agents#context-engineering#infinite-context#llama#llm#llm-memory#local-first#local-llm#local-rag#long-context#memory#mistral#ollama#ollama-rag#persistent-memory#python#qwen#rag#retrieval-augmented-generation#vector-search

Give any local LLM a billion-token memory. Open-source, local-first context + memory engine for Ollama and AI agents — runs on your own machine, free. 🧠⚡

Python
#ai-agents#context-engineering#infinite-context#llama#llm#llm-memory#local-first#local-llm#local-rag#long-context#memory#mistral#ollama#ollama-rag#persistent-memory#python#qwen#rag#retrieval-augmented-generation#vector-search
documentation good first issue

Give any local LLM a billion-token memory. Open-source, local-first context + memory engine for Ollama and AI agents — runs on your own machine, free. 🧠⚡

Python
#ai-agents#context-engineering#infinite-context#llama#llm#llm-memory#local-first#local-llm#local-rag#long-context#memory#mistral#ollama#ollama-rag#persistent-memory#python#qwen#rag#retrieval-augmented-generation#vector-search

Give any local LLM a billion-token memory. Open-source, local-first context + memory engine for Ollama and AI agents — runs on your own machine, free. 🧠⚡

Python
#ai-agents#context-engineering#infinite-context#llama#llm#llm-memory#local-first#local-llm#local-rag#long-context#memory#mistral#ollama#ollama-rag#persistent-memory#python#qwen#rag#retrieval-augmented-generation#vector-search

Give any local LLM a billion-token memory. Open-source, local-first context + memory engine for Ollama and AI agents — runs on your own machine, free. 🧠⚡

Python
#ai-agents#context-engineering#infinite-context#llama#llm#llm-memory#local-first#local-llm#local-rag#long-context#memory#mistral#ollama#ollama-rag#persistent-memory#python#qwen#rag#retrieval-augmented-generation#vector-search