by jegly
The most advanced, fully offline client-side AI suite on Android today.
# Add to your Claude Code skills
git clone https://github.com/jegly/BoxGuides for using mcp servers skills like Box.
Last scanned: 5/19/2026
{
"issues": [],
"status": "PASSED",
"scannedAt": "2026-05-19T07:46:47.861Z",
"semgrepRan": false,
"npmAuditRan": true,
"pipAuditRan": true
}Box is an open-source mcp servers skill for AI coding assistants such as Claude Code, Codex CLI, and ChatGPT, built by jegly. The most advanced, fully offline client-side AI suite on Android today. It has 669 GitHub stars.
Yes. Box passed SkillsLLM's automated security scan — a dependency vulnerability audit plus prompt-injection heuristics — with no high-severity issues. You can read the full report in the Security Report section on this page.
Clone the repository with "git clone https://github.com/jegly/Box" and add it to your Claude Code skills directory (see the Installation section above).
Box is primarily written in Kotlin. It is open-source under jegly on GitHub, so you can review or fork the full source.
Yes. SkillsLLM lists many other MCP Servers skills you can browse and compare side by side. Open the MCP Servers category from the badge at the top of this page, or use the Related Skills and comparison links further down to weigh Box against similar tools.
No comments yet. Be the first to share your thoughts!
Top skills in this category by stars
If this project helped you, please ⭐️ star it to help others find it.
Note: If you're using a custom ROM (LineageOS, GrapheneOS, CalyxOS), download the
custom-rom-supportAPK from the latest release instead.
https://github.com/jegly/BoxRecommended for most users: Main version
| Version | For |
|---|---|
| Main | Stock Android (Pixel, Samsung, etc.) |
| Custom ROM | GrapheneOS, LineageOS, CalyxOS — no Google services |
The in-app updater is also available in Settings
Maincustom-rom-supportNote: As of v2.0.0, the in-app App version matches the Box release version (2.0.0) — the earlier mismatch with the upstream Google AI Edge Gallery build number (which showed 1.0.15) is fixed (#67). Box releases are tracked via GitHub tags. Use Settings → Check for updates to see if a newer Box release is available.
Box is a security-hardened, feature rich fork of Google AI Edge Gallery — with on-device image generation (FLUX.2 klein & Z-Image Turbo diffusion), Box Assist (spoken camera assistance for blind and low-vision users), AI image upscaling, face recognition, photo erase/inpainting, music & sound generation, voice mode (speech-to-speech AI chat), voice input, multilingual text-to-speech, document analysis and Q&A, vision AI, full GPU and Snapdragon/Tensor/MediaTek NPU acceleration, a hardened security posture (biometric lock, encrypted chat history, tap-jacking protection), llama.cpp support, and GGUF model import — and more
[!IMPORTANT]
Disclaimer
Box began as a fork of Google AI Edge Gallery and is not affiliated with or endorsed by Google LLC. Google branding has been replaced throughout. Box has since diverged substantially from upstream — active merging with upstream stopped some time ago, and upstream has itself since adopted features that originated in Box. Box now carries roughly 50+ features not present in upstream Google AI Edge Gallery. Credit for the original underlying platform goes to Google and the original contributors.
| Version | Feature | Details |
|---|---|---|
| v3.3.0 | 🦯 Box Assist — a camera that talks (NEW) | Spoken camera assistance for blind and low-vision users, under the Core tab. Live mode calls out people, obstacles and objects with how close they are; Reading mode reads mail, labels and menus aloud; Describe mode describes the scene, spoken as it thinks; voice questions — double-tap, ask out loud, and Box answers against what the camera sees. One download bundles everything (vision models + the Describe brain + speech recognition). Continuous autofocus with pre-capture focus sweeps, automatic flashlight in the dark, a blur check on Reading, physical volume-button controls, hold-to-repeat, TalkBack coexistence, screen never times out, and a launcher long-press shortcut straight into it. Fully offline. |
| v3.3.0 | ⚡ GGUF engine rebuilt — real GPU acceleration | The llama.cpp engine got a ground-up overhaul: full Vulkan GPU offload via the CPU/GPU chip in any GGUF chat, a massively faster CPU mode (a flaw routed CPU prompt processing through the GPU — 0.7 → 21 tok/s on a Pixel 6a), instant replies (weights read up front, reopened chats replay their history during the loading screen), a tokens/sec stat under every GGUF reply, a new Settings → GGUF Models panel (context size, CPU threads, GPU layers, mmap, mlock, Q8 KV cache), sturdier imports with byte-verification, and automatic GPU→CPU retry. llama.cpp updated to a current build. |
| v3.3.0 | 🎨 On-device image generation — FLUX.2 klein & Z-Image Turbo | Two full text-to-image diffusion models running 100% on-device via LiteRT: FLUX.2 klein (4B) — photorealistic images in 4 steps (~7.4 GB download) — and Z-Image Turbo (9 steps), which shares nearly a gigabyte of files with klein so Box doesn't download them twice. Multi-gigabyte downloads now resume without refetching finished files, progress bars show honest totals, and a model only shows "downloaded" when every file is actually present. |
| v3.3.0 | **🔍 Five new vision model |