Pulse
0
Velocity
Stars
0
Local LLM inference server for fast AI model deployment.
A lightweight, high-performance inference server for local AI model deployment.
Run any LLM model locally at high speed with this inference server.
Why Trending
Enables fast, efficient local inference for various AI models.
Target Audience
Developers running LLMs locally for applications.
Similar Projects