Hermes Registry

Skills

Self-contained task procedures the Hermes agent can follow — each a single SKILL.md following the agentskills.io open standard.

11 results

Skill

black-forest-labs-flux

1.0.0

Black Forest Labs · media

Use when generating images with FLUX models. Official first-party FLUX image generation skills from Black Forest Labs — the creators of FLUX.1.

image-genfluxblack-forest-labs
0

gemini-api

1.0.0

google · mlops

Use when the user asks about using Gemini in an enterprise environment or explicitly mentions Vertex AI, Google Cloud, or Agent Platform. Guides the usage of the Gemini API on Agent Platform with the Google Gen AI SDK. Covers SDK usage (Python, JS/TS, Go, Java, C#), capabilities like multimodal inputs, tools, media generation, caching, batch prediction, and Live API.

Google CloudGeminiApi
0

gif-search

1.1.0

Hermes Agent · media

Search/download GIFs from Tenor via curl + jq.

GIFMediaSearch
0

heartmula

1.0.0

· media

HeartMuLa: Suno-like song generation from lyrics + tags.

musicaudiogeneration
0

ima-sdk-basics

1.0.0

google · domain

Use this skill for Interactive Media Ads (IMA) SDK client-side ad insertion when you are requesting video ads client-side into websites, apps, TVs or other platforms with VAST or VMAP. Do not use for Dynamic Ad Insertion (DAI), SSAI, or SGAI (use the `ima-sdk-dai-basics` skill instead).

Google AdsImaSdk
0

songsee

1.0.0

community · media

Audio spectrograms/features (mel, chroma, MFCC) via CLI.

AudioVisualizationSpectrogram
0

speech

1.0.0

openai · media

Use when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI (`scripts/text_to_speech.py`) with built-in voices and require `OPENAI_API_KEY` for live calls. Custom voice creation is out of scope.

SpeechTTSAudio
0

spotify

1.0.0

Hermes Agent · media

Spotify: play, search, queue, manage playlists and devices.

spotifymusicplayback
0

transcribe

1.0.0

openai · media

Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.

TranscriptionSTTAudio
0

xurl

1.1.1

xdevplatform + openclaw + Hermes Agent · social-media

X/Twitter via xurl CLI: post, search, DM, media, v2 API.

twitterxsocial-media
0

youtube-content

0.0.0

· media

YouTube transcripts to summaries, threads, blogs.

0