Pulse
0
Velocity
Stars
0
vLLM for Apple Metal GPUs, offering fast LLM inference.
An optimized vLLM inference engine leveraging Apple Metal for local LLM acceleration.
Accelerate local LLM inference on Apple Silicon using Metal GPU power.
Why Trending
Brings high-performance LLM inference to Apple hardware via Metal.
Target Audience
Developers running LLMs on Apple Silicon Macs.
Similar Projects