skip to content
open source project

vllm-project/llm-compressor

Credential-building open issues here — a merged pull request is verifiable proof of work for your résumé.

1
winnable issues
0%
merge velocity signal
3,576
166 contributors

What it is

Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM

Why get involved

Real code, on a real project, reviewed by a real maintainer — a merged PR here is credential-building proof of work you can point to.

Open issues

Check against my GitHub

Computed from your public GitHub only — nothing is stored.

Connect GitHub to check →

[Performance] Speed up subgraph tracing for large models

vllm-project/llm-compressor · issue #2981 · opened Jul 2026 · no PRs referenced when indexed

good first issueenhancementtracing

LLM Compressor uses Sequential Onloading to avoid loading the entire model into GPU memory. This relies on trace_subgraphs which performs a virtual pass…

How to jump in

Click Claim →on any issue above to open its claim page, where you'll get a terminalhire claim <token> command and a one-tap deep link into your terminal.

terminalhire contribute

Why we picked it

Winnable volume0.13
Merge velocity0.00
Freshness1.00
Skill coverage1.00
Popularity0.60