Models and memory
Choose local models or explicit hosted and private endpoints. Add documents through a bounded, resumable queue, then build auditable knowledge with sharded search, graph memory, and graceful lexical and graph recall. A visible context meter and rolling compaction keep long conversations manageable while preserving recent complete turns.
Explore models and knowledge →