Control center / Discover

Discover

Browse by category instead of drowning in search results. Every category opens a submenu of subcategories with live counts, and checking one narrows the grid to exactly that slice.

Verified against GitHub · scores pending
25 results for “cost” across 1 categories
Sort by:Type:Trust:Maturity:25 shown

Token & Cost Optimization

25

LiteLLM

BerriAI55kToolstable

Call 100+ LLM APIs using OpenAI format with proxy cost tracking and budget caps.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
55k
Downloads
661278k
Updated
2026-08-03
Skill installs
Trust pendingMomentum not matched

SGLang

sgl-project31kRepositoryexperimental

Fast serving framework for LLMs with radix attention based prefix caching for lower cost.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
31k
Downloads
100357k
Updated
2026-08-03
Skill installs
Trust pendingMomentum not matched

Langfuse

langfuse32kToolstable

Open-source LLM engineering platform for tracing, prompt management, and per-user cost tracking.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
32k
Downloads
7102k
Updated
2026-08-01
Skill installs
Trust pendingMomentum not matched

vLLM

vllm-project88kRepositoryexperimental

High throughput and memory efficient inference and serving engine for large language models.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
88k
Downloads
5581k
Updated
2026-08-03
Skill installs
Trust pendingMomentum not matched

SkyPilot

skypilot-org10kToolexperimental

Framework for running LLM workloads across clouds to minimize compute cost and maximize availability.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
10k
Downloads
1705k
Updated
2026-08-03
Skill installs
Trust pendingMomentum not matched

llama-cpp-python

abetlen11kRepositoryexperimental

Python bindings for llama.cpp enabling low cost local large language model inference.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
11k
Downloads
698k
Updated
2026-08-02
Skill installs
Trust pendingMomentum not matched

AgentOps

AgentOps-AI5.7kToolexperimental

Observability platform for tracking cost, performance and behavior of AI agents in production.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
5.7k
Downloads
287k
Updated
2026-06-25
Skill installs
Trust pendingMomentum not matched

Ollama

ollama178kToolexperimental

Run large language models locally to reduce inference costs and latency.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
178k
Downloads
Updated
2026-07-31
Skill installs
Trust pendingMomentum not matched

OpenMeter

openmeterio2.2kToolstable

Usage metering and billing infrastructure for tracking LLM API consumption and cost.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
2.2k
Downloads
74k
Updated
2026-08-02
Skill installs
Trust pendingMomentum not matched

Vercel AI Gateway

vercel26kToolstable

Unified LLM gateway with provider fallbacks, prompt routing, and spend observability.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
26k
Downloads
Updated
2026-08-02
Skill installs
45k
Trust pendingMomentum not matched

llamafile

Mozilla-Ocho25kToolexperimental

Single file executables that run LLMs locally without heavy infrastructure or API cost.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
25k
Downloads
Updated
2026-07-31
Skill installs
Trust pendingMomentum not matched

OptiLLM

codelion4.2kToolexperimental

Optimizing inference proxy that applies techniques to improve accuracy and reduce LLM cost.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
4.2k
Downloads
Updated
2026-07-18
Skill installs
Trust pendingMomentum not matched

LoRAX

predibase3.8kRepositoryactive

Framework for serving thousands of fine tuned LLM adapters on shared GPU infrastructure.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
3.8k
Downloads
Updated
2026-05-28
Skill installs
Trust pendingMomentum not matched

LangWatch

langwatch3.5kToolstable

Monitoring and analytics platform for tracking LLM usage quality and cost.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
3.5k
Downloads
Updated
2026-08-03
Skill installs
Trust pendingMomentum not matched

OpenLLM

bentoml12kToolexperimental

Platform for running and deploying open source LLMs cost effectively in production.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
12k
Downloads
1.8k
Updated
2026-07-27
Skill installs
Trust pendingMomentum not matched

LLMstudio

TensorOpsAI388Toolexperimental

Gateway and management tool for calling, caching and monitoring multiple LLM providers.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
388
Downloads
Updated
2026-07-29
Skill installs
Trust pendingMomentum not matched

Text Generation Inference

huggingface11kToolactive

Production ready inference server for LLMs optimized for throughput and cost efficiency.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
11k
Downloads
Updated
2026-03-21
Skill installs
Trust pendingMomentum not matched

RouteLLM

lm-sys5.3kToolslowing

Framework for serving and routing queries between strong and weak LLMs to slash costs by 85%.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
5.3k
Downloads
9.3k
Updated
2024-08-10
Skill installs
Trust pendingMomentum not matched

TensorZero

tensorzero12kToolstable

Open source LLM gateway and optimization framework unifying inference, observability and evals.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
12k
Downloads
Updated
2026-06-11
Skill installs
Trust pendingMomentum not matched

Helicone

Helicone6.0kToolstable

Open-source LLM observability platform with real-time token tracking, caching, and cost analytics.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
6.0k
Downloads
3.9k
Updated
2026-07-25
Skill installs
Trust pendingMomentum not matched

Aider

Aider-AI48kAgentexperimental

AI pair-programming CLI with model routing and cost-effective multi-model editing strategies.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
48k
Downloads
813k
Updated
2026-05-22
Skill installs
Trust pendingMomentum not matched

Anthropic Prompt Caching

anthropics51kClaude Skillexperimental

Cache long system prompts and context blocks for up to 90% cost reduction.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
51k
Downloads
Updated
2026-07-23
Skill installs
Trust pendingMomentum not matched

Bifrost

maximhq7.0kToolactive

An AI gateway offering built in semantic and exact prompt caching for cost optimization.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
7.0k
Downloads
Updated
2026-08-03
Skill installs
Trust pendingMomentum not matched

ExLlamaV2

turboderp4.6kRepositoryexperimental

Fast inference library for running quantized LLMs on consumer GPUs at lower cost.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
4.6k
Downloads
Updated
2026-03-04
Skill installs
Trust pendingMomentum not matched

OpenRouter

openrouter-aiToolexperimental

Unified LLM gateway with dynamic price routing, prompt caching, and cost analytics.

Best match because its name, category, capabilities, or owner matches “cost”. Adoption and trust break close ties.

Stars
Downloads
Updated
Skill installs
Trust pendingMomentum not matched
Discover · SkillPilot