High-throughput, memory-efficient serving engine for LLMs
Stars
93,053
Last commit
11 hours ago
License
Apache-2.0
Inference and serving engine for large language models, built for speed and hardware efficiency with an OpenAI-compatible API and support for a wide range of open models.
Run AI models locally or connect to cloud, privately
Stars
44,753
Last commit
12 hours ago
License
Apache-2.0
Jan runs open-source AI models on your own hardware or connects to cloud providers like OpenAI, Anthropic, and Google, keeping your conversations off third-party servers.