by IRISX-AI
π» Desktop AI assistant built for real productivity. Voice, automation, memory, vision, web search, and workflow tools in one experience.
# Add to your Claude Code skills
git clone https://github.com/IRISX-AI/IRIS-AILast scanned: 6/7/2026
{
"issues": [],
"status": "PASSED",
"scannedAt": "2026-06-07T07:56:30.107Z",
"npmAuditRan": false,
"pipAuditRan": true
}IRIS-AI is an open-source ai agents skill for AI coding assistants such as Claude Code, Codex CLI, and ChatGPT, built by IRISX-AI. π» Desktop AI assistant built for real productivity. Voice, automation, memory, vision, web search, and workflow tools in one experience. It has 182 GitHub stars.
Yes. IRIS-AI passed SkillsLLM's automated security scan β a dependency vulnerability audit plus prompt-injection heuristics β with no high-severity issues. You can read the full report in the Security Report section on this page.
Clone the repository with "git clone https://github.com/IRISX-AI/IRIS-AI" and add it to your Claude Code skills directory (see the Installation section above).
IRIS-AI is primarily written in TypeScript. It is open-source under IRISX-AI on GitHub, so you can review or fork the full source.
Yes. SkillsLLM lists many other AI Agents skills you can browse and compare side by side. Open the AI Agents category from the badge at the top of this page, or use the Related Skills and comparison links further down to weigh IRIS-AI against similar tools.
No comments yet. Be the first to share your thoughts!
β οΈ Third-Party Software Notice
This skill is third-party open-source software developed and hosted independently on GitHub. SkillsLLM is an informational directory and does not control or maintain the underlying repository.
Any security checks, ratings, or warnings displayed by SkillsLLM are automated and limited in scope. They do not constitute a security certification or guarantee that the software is safe, error-free, or free from malicious code, vulnerabilities, compromised dependencies, or prompt-injection risks.
Review the source code, permissions, dependencies, and configuration before installing or running any third-party skill. Use is at your own risk. To the maximum extent permitted by applicable law, SkillsLLM is not liable for losses arising from third-party software.

Build Faster. Automate Workflows. Control your Desktop with Voice Commands.
Speak your command. IRIS executes it.
A voice-first neural execution system powered by Gemini 3.1 Live API with real-time WebRTC audio, biometric security, and autonomous system control.
IRIS is not a chatbot.
It is a Voice-First Desktop AI Assistant that executes real-world actions across your system, applications, and devicesβpowered by Gemini 3.1 Live API with real-time bidirectional audio processing.
Speak naturally. IRIS understands intent. Real actions execute instantly.
β
Voice-First Design β Optimized for natural speech input with real-time WebRTC audio streaming
β
Proprietary Agent Logic β Heavily protected, production-grade agentic orchestration
β
Production-Ready Security β V8 bytecode + ASAR integrity validation + window isolation
β
No Code Exposure β Core agent and tools are completely hidden from public source
β
Autonomous Execution β LangGraph-powered state machine with dynamic tool orchestration
The public repository includes:
The following production components are private:
GitHub Sponsors receive access to additional documentation, implementation examples, architecture breakdowns, and development resources depending on tier.
Sponsorship does not include access to the complete private source code.
Traditional AI assistants are text-first: you type β they respond β you read.
IRIS is voice-first: you speak β they listen & execute β actions happen in real-time.
Your Voice
β (WebRTC Stream)
Gemini 3.1 Live API (Real-time)
β (Intent Recognition)
LangGraph Agent Orchestration
β (Tool Selection)
Protected Tool Execution
β (System Actions)
Results Streamed Back to You
No local-only limitations. IRIS connects to cloud AI, search engines, and APIs for maximum intelligence.
Autonomous voice activation hooks, advanced screen character peeling, and phantom inline input overlays.
Complete native file system and directory access with app process lifecycle controls.
Semantic ingestion using local Vector databases and direct multimodal vision APIs.
Globally accessible NPM package with tunneling and secure CLI execution.
AI-driven coordinate cursor control, scroll tracking, and screen peeler OCR.
Persistent identity tracking, note management, and remote inbox integrations.