Hermes Registry

Search the registry

One search across every skill, MCP server, agent, and workflow.

10 results

Mistral AI

1.0.0

Swih

MCP

Full Mistral AI surface — chat, embeddings, vision, OCR, Voxtral audio (transcribe/speak), Codestral FIM, agents, moderation, files, and batch. 22 tools. Listed on the Official MCP Registry. Free Experiment tier: 1B tokens/month.

mistralllmocr
1

Prompt Engineer

1.1.0

Hermes Registry

Agent

Designs, tests, and refines prompts for LLM features.

promptingllm
1

agent-platform-eval-flywheel

1.0.0

google · mlops

Skill

Measures and improves the quality of AI models and agents on Google Cloud using the Eval Quality Flywheel methodology. Use when evaluating an agent or model, building an eval dataset, picking or writing evaluation metrics, analyzing failures, comparing results before and after a fix, or when guidance is needed on Agent Platform eval methodology — including dataset schema, LLM-as-judge scoring, and common failure causes. For fine-tuning, use agent-platform-tuning. For general production deployment, use agent-platform-deploy.

Google CloudAgentPlatform
0

claude-api

1.0.0

anthropics · software-development

Skill

Reference for the Claude API / Anthropic SDK — model ids, pricing, params, streaming, tool use, MCP, agents, caching, token counting, model migration. TRIGGER — read BEFORE opening the target file; don't skip because it "looks like a one-liner" — whenever: the prompt names Claude/Anthropic in any form (Claude, Anthropic, Fable, Opus, Sonnet, Haiku, `anthropic`, `@anthropic-ai`, `claude-*`, `us.anthropic.*`, `[1m]`); the user asks about an LLM (pricing/model choice/limits/caching) — never answer from memory; OR the task is LLM-shaped with provider unstated (agent/MCP/tool-definition/multi-agent/RAG/LLM-judge/computer-use; generate/summarize/extract/classify/rewrite/converse over NL; debugging refusals/cutoffs/streaming/tool-calls/tokens). SKIP only when another provider is being worked on (overrides all triggers): OpenAI/GPT/Gemini/Llama/Mistral/Cohere/Ollama named in the query; OR `grep -rE 'openai|langchain_openai|google.generativeai|genai|mistralai|cohere|ollama'` over the project hits (run this grep FIRST if no provider named — don't Read the file).

ClaudeAnthropicAPI
0

dual-axis-skill-reviewer

1.0.0

TraderMonty · trading

Skill

Review skills in any project using a dual-axis method: (1) deterministic code-based checks (structure, scripts, tests, execution safety) and (2) LLM deep review findings. Use when you need reproducible quality scoring for `skills/*/SKILL.md`, want to gate merges with a score threshold (for example 90+), or need concrete improvement items for low-scoring skills. Works across projects via --project-root.

skill-toolingreviewtrading
0

edge-hint-extractor

1.0.0

TraderMonty · trading

Skill

Extract edge hints from daily market observations and news reactions, with optional LLM ideation, and output canonical hints.yaml for downstream concept synthesis and auto detection.

backtestingstrategyedge
0

karpathy-guidelines

1.0.0

multica-ai · software-development

Skill

Behavioral guidelines to reduce common LLM coding mistakes. Use when writing, reviewing, or refactoring code to avoid overcomplication, make surgical changes, surface assumptions, and define verifiable success criteria.

KarpathyGuidelinesCoding
0

llm-wiki

2.1.0

Hermes Agent · research

Skill

Karpathy's LLM Wiki: build/query interlinked markdown KB.

wikiknowledge-baseresearch
0

obliteratus

2.0.0

Hermes Agent · mlops

Skill

OBLITERATUS: abliterate LLM refusals (diff-in-means).

AbliterationUncensoringRefusal-Removal
0

serving-llms-vllm

1.0.0

Orchestra Research · mlops

Skill

vLLM: high-throughput LLM serving, OpenAI API, quantization.

vLLMInference ServingPagedAttention
0