Back to RepoInsider
Signal dossier

vllm-project/vllm-omni

AI Tools

Pulse

0

Velocity

Stars

0

High-throughput LLM inference and serving engine.

vLLM Omni is a powerful engine for fast and efficient LLM inference.

Serve LLMs at unmatched speed with vLLM's optimized inference engine.

Why Trending

Achieves significantly higher throughput for LLM inference.

Target Audience

Developers building LLM applications and services.

Similar Projects

NVIDIA Triton Inference ServerOpenVINO
RepoInsider · Spot breakout GitHub repos early