Skills

All Skills

cost

Skills tagged with #cost

@VictoryInTech
MCP

TokenOracle

Hosted MCP server for LLM cost estimation, model comparison, and budget-aware routing.

mcpllm
VictoryInTech/TokenOracle-MCP
5mo ago
0
@kevinrabun
MCP

Judges Panel

45 judges that evaluate AI-generated code for security, cost, and quality with built-in AST.

mcpgithubai
kevinrabun/judges
5mo ago
0
@luoyuctl

agenttrace-session-audit

Audit local AI coding-agent sessions with agenttrace. Use when the user asks to inspect Claude Code, Codex CLI, Gemini CLI, Qwen Code, Aider, Cursor, OpenCode, Oh My Pi, Kimi, Copilot-style, or generic JSON/JSONL sessions for cost, tokens, tool failures, latency, anomalies, health, diffs, or CI gates.

luoyuctl/agenttrace
5mo ago
60
@exmergo

transform

Use this to author and change a dbt project: bootstrap a new dbt project in a repo that has none (`transform init`), write or refactor dbt model SQL from staging to marts, add tests and docs in schema.yml, manage dependencies, and define or update the semantic layer (dbt semantic models / MetricFlow: entities, dimensions, measures, metrics). Trigger it for requests like "set up a dbt project in this repo", "build a staging model for this table", "refactor this model", "add tests to this model", "create a mart for X", "define a revenue metric", or "add a dimension to this entity". Every change is a reviewable diff to the dbt project; any warehouse build is dev-target only, gated, and cost-surfaced first. Do not use it to explore or profile a warehouse (use explore) or to detect drift and reconcile a project that has fallen out of sync (use maintain).

exmergo/dex
2mo ago
200
@sharpdeveye

accelerate

Use when the workflow is too slow, too expensive, or both and needs latency, cost, or token usage optimization.

sharpdeveye/maestro+13 more
5mo ago
500
@iris-eval
MCP

Io.Github.Iris Eval/Mcp Server

The agent eval standard for MCP. Score every agent output for quality, safety, and cost.

mcpgithub
iris-eval/mcp-server
5mo ago
0
@KeithWyatt
MCP

Sundr Repair Advisor

Repair or replace? Get device repair costs, local shops, trade-in values.

mcpgithubai
KeithWyatt/sundr
5mo ago
0
@zeuikli

haiku-pilot

Haiku-first execution playbook: through deliberate prompt structure, sub-agent delegation, and quantitative escalation gates, get Haiku 4.5 to produce near-Opus quality on most tasks. Escalate to Sonnet/Opus only when gates trigger. Triggers: "haiku", "Haiku", "Haiku mode", "haiku-pilot". Do NOT use for: cost-only token optimization, file-count cognitive heuristic, agent dispatch table, CLAUDE.md / rules audit, harness health check. This SKILL is a runtime router + escalation gate, not a decision tree or directory.

zeuikli/claude-pilot-suite+1 more
5mo ago
100
@peterbamuhigire

accounting-finance-controller

Use for accounting, bookkeeping, ERP finance, POS, inventory, payroll, billing, financial reporting, IFRS-aware workflows, management accounting, cost accounting, budgeting, valuation, controls, reconciliations, and finance-system design. Produces controller-grade requirements, implementation guidance, review findings, and financial logic so business software can replace QuickBooks/Tally-class workflows where appropriate.

peterbamuhigire/skills-web-dev+46 more
4mo ago
170
@coyvalyss1

context-monitor

Monitor conversation context and prevent MAX mode by warning at token thresholds and generating handoff summaries. Use when context approaches 100K/150K/180K tokens or when working with high-cost files.

coyvalyss1/model-matchmaker+2 more
5mo ago
1440
@HKUDS

ClawWork Economic Survival Protocol

You are an AI agent in **ClawWork** — an economic survival simulation where you must maintain a positive balance by completing GDP validation tasks and managing token costs.

HKUDS/ClawWork
5mo ago
7.2K0
@aklofas

bom

BOM (Bill of Materials) management for electronics projects — the primary orchestrator skill that coordinates DigiKey, Mouser, LCSC, element14, JLCPCB, PCBWay, and KiCad skills into a unified workflow. Create, update, and maintain BOMs with part numbers, costs, quantities stored as KiCad symbol properties. ALWAYS trigger this skill for any task involving component sourcing, pricing, ordering, distributor searches, BOM export, or fabrication preparation — even if the user names a specific distributor or fab house (e.g. "search DigiKey for...", "generate JLCPCB BOM", "order from Mouser"). This skill decides which distributor/fab skills to invoke and in what order. Also trigger on phrases like "what parts do I need", "order components", "how much will this cost", "export for JLCPCB", "find parts for this board", "cost estimate", "compare pricing", or "check stock".

aklofas/kicad-happy+11 more
5mo ago
500
@viktorbezdek

agent-project-development

This skill should be used when the user asks to "start an LLM project", "design batch pipeline", "evaluate task-model fit", "structure agent project", or mentions pipeline architecture, agent-assisted development, cost estimation, or choosing between LLM and traditional approaches. NOT for evaluating agent quality or building evaluation rubrics (use agent-evaluation), NOT for multi-agent coordination or agent handoffs (use multi-agent-patterns).

viktorbezdek/skillstack+49 more
4mo ago
50
@tombelieber

10 Claude sessions running. What are they doing? Live dashboard — monitor, cost tracking, search, sub-agent visibility.

tombelieber/claude-view+8 more
5mo ago
300
@jrenaldi79

sidecar

Spawn conversations with other LLMs (Gemini, GPT, ChatGPT, Codex, o3, DeepSeek, Qwen, Grok, Mistral, etc.) and fold results back into your context. TRIGGER when: user asks to talk to, chat with, use, call, or spawn another LLM or model; user mentions Gemini, GPT, ChatGPT, Codex, o3, DeepSeek, Claude (as a sidecar target), Qwen, Grok, Mistral, or any non-current model by name; user asks to get a second opinion from another model; user wants parallel exploration with a different model; user says "sidecar", "fork", or "fold". CRITICAL RULES: (1) ALWAYS launch sidecar CLI commands with Bash tool's run_in_background: true. Never run sidecar start/resume/continue in the foreground. (2) The fold summary returns on stdout when the user clicks Fold in the GUI or the headless agent finishes. Use TaskOutput to read it when the background task completes. (3) Use --prompt for the start command (NOT --briefing). --briefing is only for subagent spawn. (4) NEVER use o3 or o3-pro unless the user explicitly asks for it by name. These models are extremely expensive ($10-60+ per request). If the user asks for o3, warn them about the cost before proceeding. Default to gemini for most tasks. (5) When the user asks to query MULTIPLE LLMs simultaneously (e.g., "ask Gemini AND ChatGPT", "compare Gemini vs GPT"), ALWAYS use --no-ui (headless) for all of them unless the user explicitly requests interactive. Opening multiple Electron windows at once is disruptive. Launch them all in parallel with run_in_background: true.

jrenaldi79/sidecar
5mo ago
80
@Louishin

claude-api-cost-optimization

Save 50-90% on Claude API costs with Batch API, Prompt Caching & Extended Thinking. Official techniques, verified.

Louishin/claude-api-cost-optimization
5mo ago
10
@muxedai

exploring-llm-traces

ABSOLUTE MUST to debug and inspect LLM/AI agent traces using PostHog's MCP tools. Use when the user pastes a trace URL (e.g. /llm-observability/traces/<id>), asks to debug a trace, figure out what went wrong, check if an agent used a tool correctly, verify context/files were surfaced, inspect subagent behavior, investigate LLM decisions, or analyze token usage and costs.

muxedai/muxed+2 more
5mo ago
140
@mcp-registry
MCP

Philadelphia Restoration

Philadelphia water and fire damage restoration: assessment, insurance, costs, and knowledge search.

mcpgithubsearch
5mo ago
0
@github

Agentic Workflow Token Optimizer

Help users reduce the AI token usage and cost of GitHub Agentic Workflows in this repository.

github/gh-aw+39 more
2mo ago
4.7K0
@robot-resources
MCP

Robot Resources Router

Intelligent LLM routing proxy — 60-90% cost savings by auto-selecting the cheapest model

mcpgithubllm
robot-resources/robot-resources+1 more
5mo ago
0
@first-fluke

deepsec

Drive the `oma-deepsec` skill end-to-end. Installs `.deepsec/`, calibrates cost, runs the right scan/process/triage/revalidate/export pass, gates PRs with `process --diff`, writes custom matchers, and routes findings to follow-up specialists.

first-fluke/oh-my-agent+34 more
4mo ago
1.0K0
@mercurialsolo

Session Monitoring

Provides awareness of claudectl session state, health checks, and cost tracking. Activated when the user asks about session health, spending, brain decisions, or multi-session coordination.

mercurialsolo/claudectl
5mo ago
490
@vllm-project

algorithm-selection

Implements candidate-model selection logic that runs after a routing decision matches, including model ranking, cost-aware routing, and latency-aware model choice. Use when reading or modifying how the router picks which model serves a matched decision.

vllm-project/semantic-router+35 more
5mo ago
3.4K0
@yonatangross

agents-view

Wraps the Research Preview `claude agents` CLI (CC 2.1.139+) and `claude plugin details ork` for live observability of parallel agent sessions. Surfaces running/blocked/done state, per-session token cost, and the 188-hook plugin's runtime footprint. Use when debugging multi-agent workflows, projecting cost on a long-running orchestration, or auditing which hooks fired during a run.

observabilityagentscliresearch-previewcc-2.1.139cost
yonatangross/orchestkit+66 more
29d ago
1700
@supermemoryai

benchmark-context

Automatically benchmark your custom memory implementation against established systems like Supermemory. Set up a public benchmark, or create your own. Compare solutions against quality, latency, features and cost, easily, with a simple UI and CLI.

supermemoryai/memorybench
5mo ago
1930
@Fulcrum-Governance
MCP

Io.Github.Dewars30/Fulcrum

AI governance MCP server for policy enforcement, cost control, and observability.

mcpgithubai
Fulcrum-Governance/fulcrum-io
5mo ago
0
@atriumn
MCP

Io.Github.Jeff Atriumn/Tokencost Dev

LLM pricing oracle — model lookup, cost estimation, and comparison via LiteLLM

mcpgithubllm
atriumn/tokencost-dev
5mo ago
0
@justvinhhere

bigquery-cost-optimization

Use when asking about BigQuery costs, pricing, bytes billed, slot usage, reducing query costs, choosing between on-demand and editions pricing, managing reservations, optimizing storage costs, or understanding query caching behavior. Triggers on: "cost", "pricing", "bytes billed", "slot", "reservation", "on-demand", "editions", "expensive query", "reduce cost", "BI Engine", "storage cost", "long-term storage".

justvinhhere/bigquery-expert+4 more
5mo ago
120
@ericrisco

agent-safety

Use when bounding an LLM agent that already runs — scoping its task domain, gating tools to least privilege, defending against prompt injection in untrusted web/email/RAG text, requiring human approval on irreversible actions, capping runtime and cost, or triaging what it already did. NOT building the loop, tools, or RAG (that is `building-agents`).

agent-securityguardrailsprompt-injectionleast-privilegeowasp-agentic
ericrisco/rsc-harness+55 more
29d ago
660
@warpmetrics
MCP

Mcp

Connect AI assistants to Warpmetrics — query runs, calls, costs, and outcomes.

mcpgithubai
warpmetrics/mcp
5mo ago
0
@metrxbots
MCP

Mcp Server

Track AI agent costs, detect waste, optimize models, and prove ROI. 23 MCP tools across 10 domains.

mcpgithubai
metrxbots/mcp-server
5mo ago
0
@thebpandey

context-summary

Produces a lean, paste-ready session handoff so work can continue in a new conversation without re-paying the cost of the full chat. Emits a fixed seven-block skeleton optimized for re-entry, not archival. Trigger this skill whenever the user says any of the following (exact or close paraphrase): "context summary", "context-summary", "context handoff", "handoff summary", "give me the handoff", "hand this off", "carry context forward", "prep the handoff". Always trigger this skill for these phrases. Do not answer them directly without consulting this skill. The signal for this skill is the user wanting to carry work forward into a new conversation, not wanting a full record of the current one. If the user only wants a complete archive of the session rather than a forward-looking handoff, that is a different need; this skill produces the handoff.

thebpandey/context-summary
3mo ago
60
@QuixiAI

cost-report

Query and report on API usage costs across LLM, embedding, and tool providers

QuixiAI/Hexis+7 more
5mo ago
5530
@jasonwilbur
MCP

OCI Pricing

Oracle Cloud Infrastructure pricing data with cost calculators and comparisons

mcpgithub
jasonwilbur/oci-pricing-mcp
5mo ago
0
@dbsectrainer
MCP

Mcp Cost Tracker Router

Real-time cost awareness for MCP agent workflows

mcpgithubai
dbsectrainer/mcp-cost-tracker-router
5mo ago
0
@Aident-AI

aident-loadout

Coordinate multi-step work across services the user has authorized through Aident Loadout, using live action schemas, connection checks, cost preflight, and explicit approval for side effects.

Aident-AI/aident-skill+1 more
5h ago
50
@arikusi
MCP

Deepseek

MCP server for DeepSeek AI with chat, reasoning, sessions, function calling, and cost tracking

mcpgithubai
arikusi/deepseek-mcp-server
5mo ago
0
@thevibeworks

cctrace — trace Claude Code's HTTP traffic

cctrace wraps the Claude Code CLI, captures every request/response pair to `.cctrace/trace-<ts>.jsonl`, and serves a live web UI (requests list, reconstructed conversation, session replay, cost estimates).

thevibeworks/cctrace
2mo ago
60
@Galileo-Agent-Labs

eval-cost

Use when reducing token, latency, model, retrieval, tool-call, rerank, self-check, retry, or evaluator cost while preserving AI app quality metrics.

Galileo-Agent-Labs/eval-engineer+4 more
4mo ago
80
@soulmaten7
MCP

Io.Github.Soulmaten7/Potal

Total landed cost API for cross-border commerce. 240 countries, 113M+ tariff records.

mcpgithubapi
soulmaten7/potal
5mo ago
0
@goondocks-co

oak

Find out what happened, what was decided, and what depends on what in your codebase. Use this skill whenever you need to: recall past decisions or discussions ("what did we decide about X?"), check what might break before refactoring ("what depends on this module?"), find conceptually similar code that grep would miss ("all the retry/backoff logic"), look up past bugs, gotchas, or learnings, query session history or agent run costs, store observations about the codebase, or understand how components connect end-to-end. Powered by semantic search, memory lookup, and direct SQL against the Oak CI database (.oak/ci/activities.db). Also use when the user mentions oak_search, oak_context, oak_remember, oak_resolve_memory, or asks to run queries against activities.db or oak.

goondocks-co/open-agent-kit+1 more
5mo ago
80
@jonathan-vella

azure-adr

Creates Azure Architecture Decision Records with WAF mapping, alternatives, and consequences. USE FOR: ADR creation, architecture decisions, trade-off analysis, WAF pillar justification. DO NOT USE FOR: Bicep/Terraform code generation, diagram creation, cost estimates.

jonathan-vella/azure-agentic-infraops+12 more
5mo ago
1640
@mcp-registry
MCP

Agent Observability

Agent observability: structured logging, cost tracking, and compliance audit trails

mcpgithubai
5mo ago
0
@sravan27

Cut Claude Code token usage by 40.9% and ship a CI gate for coding-agent cost leaks. Stdlib Python hook + GitHub Action for Claude Code, Codex, Cursor, and agentic coding teams.

Validate context-os setup, graph availability, and local hook health in the current repository.

sravan27/context-os
4mo ago
100
@alexei-led

analyzing-usage

Analyze Claude Code usage, cost, efficiency, and burn rate using ccusage and termgraph. Use when user says "usage", "cost", "spending", "tokens", "analyze usage", "how much did I spend", "usage report", "budget", "burn rate", "efficiency", "cache hits", "ccusage", "ccw", "ccp".

alexei-led/cc-thingz+50 more
4mo ago
130
@jzOcb

context-doctor

Visualize and diagnose OpenClaw context window usage. Generates a terminal-rendered breakdown showing workspace files (status, chars, tokens), installed skills inventory, and token budget allocation across bootstrap components. Use when: (1) user asks about context window health or token usage, (2) debugging agent quality degradation ("agent got dumber"), (3) after editing workspace files to verify impact, (4) auditing bootstrap overhead. NOT for: conversation history analysis, model selection, or cost tracking.

jzOcb/context-doctor
5mo ago
1030
@1sadjlk

bounty-hunter

A professional AI bounty hunter persona named Atlas. Use when seeking, evaluating, or executing paid tasks (bounties, freelance, bug hunting) to maximize profit while minimizing token costs and ensuring secure payouts.

1sadjlk/bounty-hunter-skill
5mo ago
2600
@alibaba

agent-session-monitor

Real-time agent conversation monitoring - monitors Higress access logs, aggregates conversations by session, tracks token usage. Supports web interface for viewing complete conversation history and costs. Use when users ask about current session token consumption, conversation history, or cost statistics.

alibaba/higress+5 more
5mo ago
7.8K0
@smigolsmigol
MCP

Llmkit

AI cost tracking: 14 tools for spend, budgets, Claude Code + Cline costs, Notion sync

mcpgithubaillmnotion
smigolsmigol/llmkit
5mo ago
0
@Gammell53
MCP

Io.Github.Gammell53/Clawwork

AI agent project management — task boards, progress tracking, and cost reporting.

mcpgithubai
Gammell53/clawwork-mcp
5mo ago
0