High-throughput, memory-efficient serving engine for LLMs
Stars
93,122
Last commit
7 hours ago
License
Apache-2.0
Inference and serving engine for large language models, built for speed and hardware efficiency with an OpenAI-compatible API and support for a wide range of open models.
Run AI models locally or connect to cloud, privately
Stars
44,776
Last commit
2 days ago
License
Apache-2.0
Jan runs open-source AI models on your own hardware or connects to cloud providers like OpenAI, Anthropic, and Google, keeping your conversations off third-party servers.