Memory & long-context systems
Context windows keep growing and systems keep forgetting the things that matter. We work on what to keep, what to compress, what to retrieve on demand, and how a system should decide between them at runtime.
- Hierarchical and episodic memory architectures
- Summarisation and compaction that survives many turns
- Long-horizon state for agents and assistants
- Context-window economics and cache strategy
- Personalisation without leaking across tenants