agent-activity
Streams what the agent is doing into the room, as rows the desktop client renders in an **events drawer** above the composer (collapsed: avatar, pulsing dots,…
Render text to mp3 via OpenAI's tts-1-hd. Use for video narration, demo voiceovers, audio notes.
$ npx -y skills add sonichi/sutando --skill openai-tts --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/openai-ttsContext preview
The summary Claude sees to decide when to auto-load this skill.
Render text to mp3 via OpenAI's tts-1-hd. Use for video narration, demo voiceovers, audio notes.
name: openai-tts description: "Render text to mp3 via OpenAI's tts-1-hd. Use for video narration, demo voiceovers, audio notes." user-invocable: true
Synthesize speech via OpenAI `tts-1-hd`. Reads `OPENAI_API_KEY` from `.env`.
This is offline synthesis — distinct from voice-agent's bidirectional Gemini Live audio.
**Usage**: `/openai-tts [text]`
ARGUMENTS: $ARGUMENTS
`alloy`, `ash`, `coral` (default), `echo`, `fable`, `nova`, `onyx`, `sage`, `shimmer`.
bash "$SKILL_DIR/scripts/synthesize.sh" -- "Hello, this is Sutando." bash "$SKILL_DIR/scripts/synthesize.sh" --voice ash --out /tmp/intro.mp3 -- "Hi."
Default output path: `results/openai-tts-{epoch}.mp3`. Cost: ~$0.02 per 60s of narration.
If ARGUMENTS is empty, ask the user for the text. Otherwise:
bash "$SKILL_DIR/scripts/synthesize.sh" -- "$ARGUMENTS"
My AI Stand — Realtime by Day, Rewriting Itself by Night. Summon my AI superpower. Voice, vision, screen, meetings, calls when I'm engaged. Learns my patterns, ships its own code when I'm not. Runs across my Macs, interacts with people & their Stands.
Repo: sonichi/sutando
Streams what the agent is doing into the room, as rows the desktop client renders in an **events drawer** above the composer (collapsed: avatar, pulsing dots,…
Local Agent Registry — a standalone, dependency-free service that tracks running Claude Code (and other) agent instances. Agents self-register on startup and…
**Prefer the `ag2-space` MCP tools when they are connected and the room exposes them** — availability is per-room and per-actor, so check…
Deterministic final-answer normalizer — a last-step pass for any task that ends in a *precise* answer (a number, a short string, a comma-list). Applies the…
Transcribes audio files and voice notes to text via Gemini 2.5-flash. Integrates with Slack, Discord, and Telegram bridges so voice clips surface as readable…
Act back on the owner's Bee wearable — the TOOL half of the Bee integration (channels-vs-tools split). The Bee CHANNEL (ag2-sparrow's `sources/bee.py` watcher)…