Skip to content
View 2imi9's full-sized avatar
:shipit:
:shipit:

Block or report 2imi9

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. vllm-project/vllm vllm-project/vllm Public

    A high-throughput and memory-efficient inference and serving engine for LLMs

    Python 86.7k 19.6k

  2. vllm-project/llm-compressor vllm-project/llm-compressor Public

    Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM

    Python 3.6k 580

  3. OlmoEarth-Agent OlmoEarth-Agent Public

    Agent for OlmoEarth Studio

    Python 3

  4. ECCO-DarwinDiff ECCO-DarwinDiff Public

    A differentiable PyTorch reimplementation of ECCO-Darwin ocean biogeochemistry

    Python

  5. Quantization-and-Benchmarks Quantization-and-Benchmarks Public

    A comprehensive evaluation of quantization methods for LLM models, comparing BF16, FP8, and FP4 precision formats.

    Python

  6. nano-LMcache nano-LMcache Public

    A minimal, readable prefix cache for LLM serving (a nano-LMCache): chunk hashing + LRU CPU KV store + model-aware KV geometry. Runs on CPU.

    Python 1