Pulse
0
Velocity
Stars
0
A local inference server or UI for running Llama-family and other LLMs.
Provides a user-friendly interface or server for local execution of large language models.
Run powerful Llama-based LLMs locally with an easy-to-use interface or API.
Why Trending
Simplifies running powerful LLMs locally, reducing cloud dependency and costs.
Target Audience
Developers, researchers running LLMs locally for privacy/cost.
Similar Projects