defilantech/llmkube
Credential-building open issues here — a merged pull request is verifiable proof of work for your résumé.
What it is
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Why get involved
Real code, on a real project, reviewed by a real maintainer — a merged PR here is credential-building proof of work you can point to.
Open issues
Check against my GitHub
Computed from your public GitHub only — nothing is stored.
Connect GitHub to check →We can't currently confirm which issues are still open here — check back shortly.
How to jump in
Click Claim →on any issue above to open its claim page, where you'll get a terminalhire claim <token> command and a one-tap deep link into your terminal.
terminalhire contribute