Neural Magic
Neural Magic (Acquired by Red Hat) empowers developers to optimize & deploy LLMs at scale. Our model compression & acceleration enable top performance with vLLM
Pinned Loading
Repositories
Showing 10 of 103 repositories
- vllm Public Forked from vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
- nyann-bench Public
- eval-hub-contrib Public Forked from eval-hub/eval-hub-contrib
Community-contributed evaluation framework adapters for eval-hub
- speculators Public Forked from vllm-project/speculators
A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM
- vllm-openshift-recipes Public
- model-validation-configs Public
Top languages
Loading…
Most used topics
Loading…