Pinned Loading
-
vllm-project/vllm
vllm-project/vllm PublicA high-throughput and memory-efficient inference and serving engine for LLMs
-
vllm-project/llm-compressor
vllm-project/llm-compressor PublicTransformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM
-
-
ECCO-DarwinDiff
ECCO-DarwinDiff PublicA differentiable PyTorch reimplementation of ECCO-Darwin ocean biogeochemistry
Python
-
Quantization-and-Benchmarks
Quantization-and-Benchmarks PublicA comprehensive evaluation of quantization methods for LLM models, comparing BF16, FP8, and FP4 precision formats.
Python
-
nano-LMcache
nano-LMcache PublicA minimal, readable prefix cache for LLM serving (a nano-LMCache): chunk hashing + LRU CPU KV store + model-aware KV geometry. Runs on CPU.
Python 1
If the problem persists, check the GitHub status page or contact support.



