by AGIHunt
Show it. Say it. Your AI gets it. Record your screen and talk — your coding agent turns it into bugs, ideas and to-dos, with the exact frames marked. 口喷鸡:边看边喷,AI 全懂。Agent Skill + macOS menu-bar app for Claude Code / Codex.
# Add to your Claude Code skills
git clone https://github.com/AGIHunt/blurtSee how blurt compares with popular alternatives.
blurt is an open-source ai agents skill for AI coding assistants such as Claude Code, Codex CLI, and ChatGPT, built by AGIHunt. Show it. Say it. Your AI gets it. Record your screen and talk — your coding agent turns it into bugs, ideas and to-dos, with the exact frames marked. 口喷鸡:边看边喷,AI 全懂。Agent Skill + macOS menu-bar app for Claude Code / Codex. It has 50 GitHub stars.
blurt's catalog security scan is still queued. You can run an instant dependency and prompt-injection check now with the "Scan for vulnerabilities" button above.
Clone the repository with "git clone https://github.com/AGIHunt/blurt" and add it to your Claude Code skills directory (see the Installation section above).
blurt is primarily written in Python. It is open-source under AGIHunt on GitHub, so you can review or fork the full source.
Yes. SkillsLLM lists many other AI Agents skills you can browse and compare side by side. Open the AI Agents category from the badge at the top of this page, or use the Related Skills and comparison links further down to weigh blurt against similar tools.
No comments yet. Be the first to share your thoughts!
⚠️ Third-Party Software Notice
This skill is third-party open-source software developed and hosted independently on GitHub. SkillsLLM is an informational directory and does not control or maintain the underlying repository.
Any security checks, ratings, or warnings displayed by SkillsLLM are automated and limited in scope. They do not constitute a security certification or guarantee that the software is safe, error-free, or free from malicious code, vulnerabilities, compromised dependencies, or prompt-injection risks.
Review the source code, permissions, dependencies, and configuration before installing or running any third-party skill. Use is at your own risk. To the maximum extent permitted by applicable law, SkillsLLM is not liable for losses arising from third-party software.
The deep catalog scan for this skill is still queued. Run an instant dependency check now instead.
https://github.com/user-attachments/assets/334087d3-9b56-48fa-808a-2fb229ab666a
Voice alone is blind. Screenshots plus typing is slow. The most natural way to tell an AI what you mean is the way you'd tell a colleague sitting next to you: point at the screen and talk. blurt records exactly that, and your agent does the rest:
What used to take two days of screenshots, red boxes and spreadsheet rows is now a 30-minute walkthrough.
| 🐞 Polish a vibe-coded product | An agent built it overnight; now you walk through every page and rant. You get a clean bug list with code pointers, ready for the agent to fix. This is the fastest way to push AI to the finish line. |
| 💡 Capture ideas while browsing | "I like how this site does onboarding… and this pricing page…" You get an idea board: each idea, why you had it, where it came from, the next step, plus a one-page digest. |
| 📐 Browse, then get the spec | Building a landing page or a store? Tour a few dozen reference sites, circle what you like and say why. Your agent turns it into a requirements doc and a design doc, each requirement next to the frame you circled, ready to build from. |
| 🤝 Hand off from anyone | PMs, designers, ops, clients: anyone can record with the Blurt app, no repo needed, and send the video. The developer's agent processes it with the code at hand. |
| 🔎 Research and walkthroughs | Competitor tours, UX research, "how this works": you get notes and findings with the frames to prove them. |
| 🖥️ Not just web apps | Terminals, TUIs and desktop apps record and get boxed the same way. For phones, use the built-in screen recorder with the mic on; for hardware, film it with your phone. Hand the video to your agent: "process this video". |
One recording can mix all of these. The agent decides what each item is, and teams can add their own
lenses (e.g. ux-research, sales-call, sop).
npx skills add AGIHunt/blurt
Claude Code: /plugin marketplace add AGIHunt/blurt · or just paste this repo's URL to your agent and ask it to install the skill.
Then, in any project, tell your agent "start blurt" / 「开始口喷」. The first run picks a speech model for your machine and installs the Blurt menu-bar app.
⌥⇧P pause, ⌥⇧S finish.
When you say "this", hold ⌃ and drag to circle or underline it (macOS). The ink fades after 3 seconds, and
your agent sees exactly what you meant.A keep · X drop · J/K next/prev · Z undo · G list · V overview. Then export,
or say "fix them".Always on: the Blurt app lives in your menu bar. ⌥⇧R starts a recording from anywhere and ⌥⇧R again
finishes it; ⌥⇧B opens the menu. Recordings go to a workspace: ~/Blurt by default, or a project you bind, so
your agent can process them with the code. Turn on After recording → Claude Code / Codex and every recording gets
processed in the background, with the review page popping up when it's ready.
⌥⇧R. Anyone records with the app, then uses Recent recordings → Copy video and pastes it into
Slack/Feishu. Or bind a shared project folder. Items keep the recorder's name, and exports land in the team's
existing tables with your column names.Does it burn a lot of tokens? The video is never fed to the model. Speech is transcribed locally (free), the agent reads the text, and it looks at a few frames only for the moments that become items. Cost follows how many things you talk about, not how long you record; silence and clicking around cost nothing. Quality follows the model: use one with vision, the stronger the better. Two real sessions (Claude Code, Opus 5.5, a production web app):
| recording | speech | items | transcription (local) | recording → review page | new input / output tokens | at API prices |
|---|---|---|---|---|---|---|
| 7.0 min | 2.8 min | 9 | 4 s | ~4 min | ~142k / ~19k | ~$2 |
| 8.6 min | 5.0 min | 20 | 5 s | ~5 min | ~150k / ~20k | ~$2 |
Plus cache reads of the ongoing conversation (~4–5M tokens, 1/20 of the input price, included above), which depend on how long your chat already is.
I don't do frontend. Is it for me? Yes, if you can see the problem on screen or film it with a phone. See Not just web apps above.
Browser capture (console errors and network failures lined up with the video) · the pen on Windows · Windows tray app · a hosted speech API · more lenses and exporters · toward a personal assistant that watches, listens and keeps your projects moving. See TODO.md.