Multi-modal voice agent for your AI agents. Talk with Claude Code, Codex, and others by voice: answer their prompts, give instructions, ask what they did. Or just nod.
# Add to your Claude Code skills
git clone https://github.com/spaceamoeba-t/tapqLast scanned: 10/3/2026
{
"issues": [],
"status": "PASSED",
"scannedAt": "2026-10-03T09:41:21.030Z",
"npmAuditRan": true,
"pipAuditRan": true,
"promptInjectionRan": true
}See how tapq compares with popular alternatives.
tapq is an open-source ai agents skill for AI coding assistants such as Claude Code, Codex CLI, and ChatGPT, built by spaceamoeba-t. Multi-modal voice agent for your AI agents. Talk with Claude Code, Codex, and others by voice: answer their prompts, give instructions, ask what they did. Or just nod. It has 493 GitHub stars.
Yes. tapq passed SkillsLLM's automated security scan — a dependency vulnerability audit plus prompt-injection heuristics — with no high-severity issues. You can read the full report in the Security Report section on this page.
Clone the repository with "git clone https://github.com/spaceamoeba-t/tapq" and add it to your Claude Code skills directory (see the Installation section above).
tapq is primarily written in Swift. It is open-source under spaceamoeba-t on GitHub, so you can review or fork the full source.
Yes. SkillsLLM lists many other AI Agents skills you can browse and compare side by side. Open the AI Agents category from the badge at the top of this page, or use the Related Skills and comparison links further down to weigh tapq against similar tools.
No comments yet. Be the first to share your thoughts!
⚠️ Third-Party Software Notice
This skill is third-party open-source software developed and hosted independently on GitHub. SkillsLLM is an informational directory and does not control or maintain the underlying repository.
Any security checks, ratings, or warnings displayed by SkillsLLM are automated and limited in scope. They do not constitute a security certification or guarantee that the software is safe, error-free, or free from malicious code, vulnerabilities, compromised dependencies, or prompt-injection risks.
Review the source code, permissions, dependencies, and configuration before installing or running any third-party skill. Use is at your own risk. To the maximum extent permitted by applicable law, SkillsLLM is not liable for losses arising from third-party software.
AI scales. Your attention doesn't. Agents work in parallel, and keeping them moving still pulls you back to a screen — to check, to answer, to approve. Checking fragments your focus: which agent finished, which one is stuck. Every switch costs context: open the tool, read the thread, find what needs you. Stepping away stops the work: a question or an approval sits unanswered until you return. You've delegated the work; staying in control still ties you to a screen.
Agents are the first software you don't operate but supervise, and supervision doesn't need a screen — it needs a way to reach you. TapQ is that way. Works today with Claude Code, Codex, Cursor, and OpenCode, on AirPods and macOS.
TapQ is a voice agent that runs in the background all day, connected to your coding agents through their hooks. It listens on-device, speaks only when an agent needs you or you speak to it, and hears your answer through the earbuds. Four things:
Claude Code: Run swift test. Approve? You answer "yes", double-nod, or tilt and tap through the options.
Prompts from every open session arrive as one spoken queue.An afternoon with two agents:
Claude Code: Run swift test. Approve?
"yes"
Codex: Run python migrate.py. Approve?
"no, tell Codex to keep the old table names"
Queued for Codex.
"who's waiting?"
Nothing is waiting. Claude Code finished the migration.
"when Claude Code finishes the changelog, rerun the tests"
After Claude Code finishes: rerun the tests — noted.
Answering needs no wake word, nothing to open, nothing to look at. Anything TapQ cannot answer — a multi-select prompt, a missed gesture, a failed voice pipe — falls back to the agent's on-screen prompt exactly as if TapQ weren't installed. Everything TapQ does unprompted is spoken and attributed, and none of it can approve anything.
TapQ is a runtime on your Mac with an adapter for each agent and each device.
tapq integration <agent> install adds a hook (Claude Code,
Codex, Cursor) or a plugin (OpenCode) to the agent. When the agent stops for a
permission, a question, or a choice, or finishes a turn, the hook sends that event to
the runtime and waits for the answer.--voice-backend openai-realtime, your
speech is sent to OpenAI's realtime API only while a window is open, and the model
turns what you said into one of a fixed set of actions: approve, deny, select an
option, queue an instruction for an agent, answer a question about status or an
agent's transcript, set a follow-up, or start a task. It can speak only what TapQ
passes it; it cannot answer a prompt on its own. The default on-device backend
matches a fixed vocabulary instead.--wearer-gate, speech that does not coincide with it is ignored, so a
colleague or a video cannot answer for you.--attention wake, an on-device recognizer listens for
"hey tapq" whenever nothing else is listening and opens a window with the same rules;
with no session running, a sentence starts one. What you and TapQ say to each other
is appended to a local file, and a recent slice of it is given to the model each turn
so it can resolve "the thing I asked about earlier".TapQ sits between you and tools that can run commands, so the boundaries are part of the design, not a settings page. What ships today:
--voice-backend openai-realtime, audio goes to OpenAI's realtime API only while a
response window is open; wake-word listening between windows stays on-device.wearer-conversation.jsonl in the runtime directory. It keeps 30 days or a couple of
megabytes, whichever comes first, and tapq memory clear wipes it. Tool inputs,
working directories, and permission modes are never spoken and never recorded.The local broker boundary describes what the hooks and the runtime can and cannot do to each other. Privacy principles sets the rules for planned capture modes (explicit start, clear state, consent first) that shipped functionality will be held to.
Nodding to answer an earbud is no longer exotic. Siri Interactions on AirPods let you nod yes or shake no to Siri, and every coding agent now offers some way to approve a prompt from a phone. TapQ is for the gap between them:
| Siri Interactions on AirPods | Approve from the agent's phone app | TapQ | |
|---|---|---|---|
| What a nod answers | Siri's own prompts | Nothing; you tap on a screen | Claude Code, Codex, Cursor, and OpenCode prompts |
| Where the prompt reaches you | Siri | A notification you read | Spoken in your ear, named with its agent |
| Beyond yes and no | Siri's features | The agent's own UI | Options, instructions, q |