GitStar
⌘K
GitStar

A ranking dashboard for GitHub momentum, durable repository leaders, package adoption, and editorial context.

Data source · GitHub API and package ecosystem snapshots

Home

HomeTrendingMomentumPulseExplore

Discover

CategoriesLanguagesOrganizationsAI / MLMCP

Workflow

Repo CompareSurprise Me

Knowledge

InsightGuideMethodology

Support

FAQAbout GitStarContactPrivacyTerms

© 2026 GitStar. All rights reserved.

Data sourced from GitHub API

TrendingMomentumPulseExplore
  1. Home
  2. Organizations
  3. kvcache.ai
kvcache.aiOrganization

kvcache.ai

@kvcache-ai • KVCache.AI is an open source orgnization between MADSys and top industry collaborators, focusing on efficient Agent/LLM serving.. Use this route to separate flagship concentration from portfolio breadth before you treat a publisher as broadly strong.

Portfolio concentration

99%

Top three share

Shows whether the organization is driven by one breakout repo or several visible projects.

Breadth

20 repos

Visible snapshot

12 repositories updated in the last 90 days.

Leading language

Python

Portfolio mix

Python (8), Unknown (6), Cuda (2)

Average size

1.5K

Stars per repository

Useful for distinguishing one flagship-heavy publisher from a repeatable portfolio.

Back to organizationsCompare repositories
Updated: 2026-08-07(64d ago)GitHub API fallback20 repositories

Portfolio Shape

99%

of the visible star count comes from this organization's top three repositories.

Average Repository Size

1.5K

stars per repository in this same snapshot.

Current Mix

Python

is the most common language here, with 12 repositories updated in the last 90 days.

Why this rank

This organization stands out because one flagship repo drives 65% of its visible star count.

Flagship share 65%Breakout repo: ktransformers

Organization pages work best when you separate portfolio breadth from flagship concentration. In kvcache.ai's case, the visible top three repositories account for about 99% of total stars in this snapshot, which helps explain whether the organization is known for one breakout project or for a broader repeatable portfolio.

The dominant language mix here is Python (8), Unknown (6), Cuda (2). That makes this page useful not just for popularity checks, but also for seeing what technical shape an organization's public ecosystem actually has.

Source: GitHub API fallback. This is the same cache-first snapshot used by the organization ranking list, so the summary view and the detail view should stay aligned.

Top Repositories

#RepositoryLanguageStars🍴 ForksUpdated
1kvcache-ai/ktransformers

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python19.6K1.6KToday
2kvcache-ai/Mooncake

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++6.8K1.3KToday
3kvcache-ai/AgentENV

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust3.6K330Today
4kvcache-ai/TrEnv-XGo9791 years ago
5kvcache-ai/kvcache-blogJavaScript29242 weeks ago
6kvcache-ai/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

Python1941 months ago
7kvcache-ai/sglang

SGLang is a fast serving framework for large language models and vision language models.

Python14341 weeks ago
8kvcache-ai/custom_flashinfer

FlashInfer: Kernel Library for LLM Serving

Cuda941 years ago
9kvcache-ai/AFD-Ledger

AFD-Ledger is an analytical toolkit for provisioning and evaluating attention–FFN disaggregated LLM inference deployments.

Python622 months ago
10kvcache-ai/DeepEP_fault_tolerance

DeepEP: an efficient expert-parallel communication library that supports fault tolerance

Cuda409 months ago
11kvcache-ai/accelerate

🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support

Python31Today
12kvcache-ai/linux

Linux kernel source tree for PVM

202 months ago
13kvcache-ai/transformers

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python221 weeks ago
14kvcache-ai/sglang_awq

SGLang is a fast serving framework for large language models and vision language models.

Python205 months ago
15kvcache-ai/Model-Optimizer

A unified library of SOTA model optimization techniques like quantization, pruning, distillation, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

103 months ago
16kvcache-ai/gpustack

GPU cluster manager for optimized AI model deployment

1010 months ago
17kvcache-ai/overlaybd

Overlaybd: a block based remote image format. The storage backend of containerd/accelerated-container-image.

002 months ago
18kvcache-ai/kvcache-blog-private002 months ago
19kvcache-ai/evalscope

A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.

Python005 months ago
20kvcache-ai/sglang-npu

SGLang is a fast serving framework for large language models and vision language models.

001 years ago

Next step after the organization read

Open a flagship repository, compare a couple of portfolio leaders, or return to the organization map when you want a broader concentration read.
Open flagship repoCompare repositoriesBack to organizations

Learn and methodology

Keep trust-building context reachable, but behind the first data read instead of ahead of it.
GuideMethodologyArticlesWeekly Digest

How to read this organization snapshot

Total stars are useful as a discovery signal, but they do not tell you whether a team maintains every repository equally. Pair this page with release cadence, maintainer activity, and the flagship concentration shown above before making adoption decisions.

For broader background on GitStar's ranking logic and editorial guidance, see Methodology & Editorial Standards.