Phoenix Documentation ===================== **Phoenix** is a high-performance KV cache system designed for efficient memory management in large language model inference. It provides intelligent caching strategies, memory optimization, and seamless integration with AI frameworks. Welcome to Phoenix documentation. .. toctree:: :maxdepth: 2 :caption: Contents: README architecture install adapters kernel-module libphoenix vendor-porting-guide troubleshooting roadmap Key Features ------------ - **Intelligent KV Cache Management**: Optimized memory allocation for transformer models - **High Performance**: Low-latency cache operations for real-time inference - **Multi-Backend Support**: Works with various GPU architectures - **Framework Integration**: Easy integration with popular AI frameworks Links ----- - **GitHub Repository**: https://github.com/xpu-io/Phoenix - **License**: See LICENSE file in the repository Indices and tables ================== * :ref:`genindex` * :ref:`search`