Hermes Registry

Search the registry

One search across every skill, MCP server, agent, and workflow.

13 results

DaVinci Resolve

1.0.0

samuelgursky

MCP

Control DaVinci Resolve Studio via its scripting API — timeline edits, media pool, render setup, color, Fusion, Fairlight. Requires Resolve open with external scripting set to Local.

creatorvideoediting
0

Higgsfield

1.0.0

Higgsfield

MCP

30+ image and video models (Sora, Veo, Kling, Seedance) through one endpoint — research, prompt-optimize, and generate creative assets. Use the official connector, not session-scraping forks.

creatormediavideo
0

black-forest-labs-flux

1.0.0

Black Forest Labs · media

Skill

Use when generating images with FLUX models. Official first-party FLUX image generation skills from Black Forest Labs — the creators of FLUX.1.

image-genfluxblack-forest-labs
0

gemini-api

1.0.0

google · mlops

Skill

Use when the user asks about using Gemini in an enterprise environment or explicitly mentions Vertex AI, Google Cloud, or Agent Platform. Guides the usage of the Gemini API on Agent Platform with the Google Gen AI SDK. Covers SDK usage (Python, JS/TS, Go, Java, C#), capabilities like multimodal inputs, tools, media generation, caching, batch prediction, and Live API.

Google CloudGeminiApi
0

gif-search

1.1.0

Hermes Agent · media

Skill

Search/download GIFs from Tenor via curl + jq.

GIFMediaSearch
0

heartmula

1.0.0

· media

Skill

HeartMuLa: Suno-like song generation from lyrics + tags.

musicaudiogeneration
0

ima-sdk-basics

1.0.0

google · domain

Skill

Use this skill for Interactive Media Ads (IMA) SDK client-side ad insertion when you are requesting video ads client-side into websites, apps, TVs or other platforms with VAST or VMAP. Do not use for Dynamic Ad Insertion (DAI), SSAI, or SGAI (use the `ima-sdk-dai-basics` skill instead).

Google AdsImaSdk
0

songsee

1.0.0

community · media

Skill

Audio spectrograms/features (mel, chroma, MFCC) via CLI.

AudioVisualizationSpectrogram
0

speech

1.0.0

openai · media

Skill

Use when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI (`scripts/text_to_speech.py`) with built-in voices and require `OPENAI_API_KEY` for live calls. Custom voice creation is out of scope.

SpeechTTSAudio
0

spotify

1.0.0

Hermes Agent · media

Skill

Spotify: play, search, queue, manage playlists and devices.

spotifymusicplayback
0

transcribe

1.0.0

openai · media

Skill

Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.

TranscriptionSTTAudio
0

xurl

1.1.1

xdevplatform + openclaw + Hermes Agent · social-media

Skill

X/Twitter via xurl CLI: post, search, DM, media, v2 API.

twitterxsocial-media
0

youtube-content

0.0.0

· media

Skill

YouTube transcripts to summaries, threads, blogs.

0