Control center / Discover
Discover Browse by category instead of drowning in search results. Every category opens a submenu of subcategories with live counts, and checking one narrows the grid to exactly that slice.
◑ Verified against GitHub · scores pending
Sort by: Best match Skill installs Stars Downloads Last updated Forks Name ↓ DescType: Filter by entry type All types MCP Server Claude Skill Agent Repository Tool Trust: Filter by trust evidence Any trust state Crown eligible Trust checked Audit conflict Maturity: Filter by maturity Any maturity Stable Active Experimental Slowing 11 shown
All (11) ◎ Marketing (1)↗ GTM & Growth (0)✎ Design (0)⛨ Security (0)▶ Video & YouTube (0)⌘ Development (0)⌕ Data & Research (0)⑂ Automation (0)⚡ Token & Cost Optimization (8)◆ AI Agents (0)◐ LLM Ops & Observability (1)◇ RAG & Knowledge (1)
◎ Marketing1 Browse subcategories → blader ★ 33k Claude Skill active Portable writing skill that removes common AI-generated patterns while preserving meaning and voice.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 33k
Downloads —
Updated 2026-07-22
Skill installs 3.4k Audit conflict Momentum not matched
$ npx skills add blader/humanizer --globalCopy Content
⚡ Token & Cost Optimization8 Browse subcategories → sgl-project ★ 31k Repository experimental Fast serving framework for LLMs with radix attention based prefix caching for lower cost.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 31k
Downloads 100357k
Updated 2026-08-03
Skill installs — Trust pending Momentum not matched
$ pip install sglangCopy Prompt Caching
vllm-project ★ 88k Repository experimental High throughput and memory efficient inference and serving engine for large language models.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 88k
Downloads 5581k
Updated 2026-08-03
Skill installs — Trust pending Momentum not matched
$ pip install vllmCopy Model Routing Observability & Cost
predibase ★ 3.8k Repository active Framework for serving thousands of fine tuned LLM adapters on shared GPU infrastructure.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 3.8k
Downloads —
Updated 2026-05-28
Skill installs — Trust pending Momentum not matched
Model Routing
bentoml ★ 12k Tool experimental Platform for running and deploying open source LLMs cost effectively in production.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 12k
Downloads 1.8k
Updated 2026-07-27
Skill installs — Trust pending Momentum not matched
$ pip install openllmCopy Cheaper-Model Agents
lm-sys ★ 40k Repository experimental Platform for training, serving and evaluating multiple open LLMs with routing capabilities.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 40k
Downloads 45k
Updated 2026-05-01
Skill installs — Trust pending Momentum not matched
$ pip install fschatCopy Model Routing
huggingface ★ 11k Tool active Production ready inference server for LLMs optimized for throughput and cost efficiency.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 11k
Downloads —
Updated 2026-03-21
Skill installs — Trust pending Momentum not matched
Observability & Cost
KV cache management layer that speeds up LLM serving by reusing cached context.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 11k
Downloads 68k
Updated 2026-08-02
Skill installs — Trust pending Momentum not matched
$ pip install lmcacheCopy Prompt Caching
Framework for serving and routing queries between strong and weak LLMs to slash costs by 85%.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 5.3k
Downloads 9.3k
Updated 2024-08-10
Skill installs — Trust pending Momentum not matched
$ pip install routellmCopy Model Routing
◐ LLM Ops & Observability1 Browse subcategories → InternLM ★ 8.0k Tool experimental A toolkit for compressing deploying and serving large language models locally with high throughput.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 8.0k
Downloads —
Updated 2026-08-01
Skill installs — Trust pending Momentum not matched
Local Inference
◇ RAG & Knowledge1 Browse subcategories → huggingface ★ 5.0k Tool stable Fast inference server for serving text embedding models in production.
Best match because its name, category, capabilities, or owner matches “serving”. Adoption and trust break close ties.
Stars 5.0k
Downloads —
Updated 2026-07-24
Skill installs — Trust pending Momentum not matched
Embeddings