Lumi-node/infinite-context
0
Infinite Context - Live Demo
Give any LLM unlimited memory with sub-millisecond retrieval.
What This Demo Shows
This is a live demonstration of HAT (Hierarchical Attention Tree) - a retrieval system that:
- 100% accuracy finding relevant conversations
- < 1ms search time across hundreds of thousands of tokens
- 1,400x context extension for small models
How to Use
- Click Initialize to create a simulated conversation history
- Ask natural questions like:
- "What did we do to fix the React error?"
- "How much did we speed up the Python script?"
- "What was causing the Kubernetes pods to crash?"
- See HAT retrieve the exact relevant conversations in milliseconds
Performance
Links
- GitHub
- Docker Hub
- ArXiv Paper (coming soon)
License
MIT
