Fast, Accurate, Production-Ready RAG Pipelines Turn your unstructured data into perfectly optimized vector search indexes, purpose-built for retrieval augmented generation.
Per-user Memory: Each user has a persistent context that retains preferences, history, and decisions separately.
Cross-session Persistence: Memory survives session boundaries, allowing agents to continue interactions weeks later without losing context.
Fast Memory Recall: The system can retrieve relevant memories in under 100 milliseconds through parallel search.
Model-agnostic: The memory layer works with any language model, enabling users to switch models without losing learned information.
Self-installation: Hindsight can be added to agents with a single command, automatically configuring memory tools without boilerplate code.
Persistent User Context: Each user receives a unique memory that retains their preferences, history, and decisions across sessions.
Learning from Mistakes: Agents improve by learning from past errors and user corrections, enhancing their performance over time.
Cross-Agent Knowledge Sharing: Memory allows different agents to share learned user preferences, enabling seamless collaboration and improved service.
Fast Memory Recall: The system can retrieve relevant memories in under 100 milliseconds, ensuring quick responses to user inquiries.
Self-Installation Capability: Hindsight can be integrated into existing systems with a single command, simplifying deployment for developers.
Provides persistent memory for each user, allowing agents to retain context, preferences, and history across sessions.
Enables fast memory recall, retrieving relevant memories in under 100 milliseconds, enhancing user experience.
Learns from mistakes and user corrections, improving agent performance over time by building judgment and understanding patterns.
Offers a model-agnostic memory layer, allowing integration with any large language model (LLM) without losing learned information.
Facilitates automatic installation and setup, simplifying the deployment process for developers.
Self-hosted: Free to run on your own infrastructure with a single Docker command, includes all four memory networks and community support via GitHub.
Hindsight Cloud: Pay-as-you-go model with no fixed monthly fee; billed on token usage, starting free with managed infrastructure and automatic scaling.
Token Costs:
Retain (store memories): $10.00 per million tokens
Recall (retrieve memories): $0.75 per million tokens
Reflect (synthesize patterns): $0.05 per call
Iris Extract (structured extraction): $7.50 per million tokens
Mental Model Retrieve: $0.25 per million tokens
Mental Model Refresh: $0.05 per call
Storage: $0.25 per million tokens stored per month after the first 30 days.
Enterprise Plan: Custom deployment and dedicated support; contact sales for pricing and details.
Free Credits: Available to start using Hindsight Cloud, allowing users to build and test before incurring charges.