by jegly
The most advanced, fully offline client-side AI suite on Android today.
# Add to your Claude Code skills
git clone https://github.com/jegly/BoxGuides for using mcp servers skills like Box.
Last scanned: 5/19/2026
{
"issues": [],
"status": "PASSED",
"scannedAt": "2026-05-19T07:46:47.861Z",
"semgrepRan": false,
"npmAuditRan": true,
"pipAuditRan": true
}Box is an open-source mcp servers skill for AI coding assistants such as Claude Code, Codex CLI, and ChatGPT, built by jegly. The most advanced, fully offline client-side AI suite on Android today. It has 710 GitHub stars.
Yes. Box passed SkillsLLM's automated security scan — a dependency vulnerability audit plus prompt-injection heuristics — with no high-severity issues. You can read the full report in the Security Report section on this page.
Clone the repository with "git clone https://github.com/jegly/Box" and add it to your Claude Code skills directory (see the Installation section above).
Box is primarily written in Kotlin. It is open-source under jegly on GitHub, so you can review or fork the full source.
Yes. SkillsLLM lists many other MCP Servers skills you can browse and compare side by side. Open the MCP Servers category from the badge at the top of this page, or use the Related Skills and comparison links further down to weigh Box against similar tools.
No comments yet. Be the first to share your thoughts!
Top skills in this category by stars
If this project helped you, please ⭐️ star it to help others find it. We have hit 20K downloads ! Thank you to everyone for supporting Box.
Note: If you're using a custom ROM (LineageOS, GrapheneOS, CalyxOS), download the
custom-rom-supportAPK from the latest release instead.
https://github.com/jegly/BoxRecommended for most users: Main version
| Version | For |
|---|---|
| Main | Stock Android (Pixel, Samsung, etc.) |
| Custom ROM | GrapheneOS, LineageOS, CalyxOS — no Google services |
The in-app updater is also available in Settings
Maincustom-rom-supportNote: As of v2.0.0, the in-app App version matches the Box release version (2.0.0) — the earlier mismatch with the upstream Google AI Edge Gallery build number (which showed 1.0.15) is fixed (#67). Box releases are tracked via GitHub tags. Use Settings → Check for updates to see if a newer Box release is available.
Box is a security-hardened, feature rich fork of Google AI Edge Gallery — with on-device image generation (FLUX.2 klein & Z-Image Turbo diffusion), Box Assist (spoken camera assistance for blind and low-vision users), AI image upscaling, face recognition, photo erase/inpainting, music & sound generation, voice mode (speech-to-speech AI chat), voice input, multilingual text-to-speech, document analysis and Q&A, vision AI, full GPU and Snapdragon/Tensor/MediaTek NPU acceleration, a hardened security posture (biometric lock, encrypted chat history, tap-jacking protection), llama.cpp support, and GGUF model import — and more
[!IMPORTANT]
Disclaimer
Box began as a fork of Google AI Edge Gallery and is not affiliated with or endorsed by Google LLC. Google branding has been replaced throughout. Box has since diverged substantially from upstream — active merging with upstream stopped some time ago, and upstream has itself since adopted features that originated in Box. Box now carries roughly 50+ features not present in upstream Google AI Edge Gallery. Credit for the original underlying platform goes to Google and the original contributors.
| Version | Feature | Details |
|---|---|---|
| v3.3.2 | Downloads fixed | Model downloads are reliable again after 3.3.1 — no more failing mid-download or stalling at 100%. A previously stuck model downloads normally on the first try. |
| v3.3.2 | GGUF GPU crash fix (really this time) | The Snapdragon GPU crash fix from 3.3.1 now actually ships in the build. |
| v3.3.2 | Biometric lock + database encryption | The biometric app lock works alongside database encryption again — the two are independent, and the app re-locks when reopened. |
| v3.3.1 | Live Translator (NEW, Sound tab) | Two people, two languages — tap your button, speak, and the other person reads and hears it in their language. Runs on your installed Gemma audio model (E2B/E4B), each phrase translated on its own for flat latency. 24 languages, fully offline. |
| v3.3.1 | 4 new models | Granite 4.0 350M (IBM's tiny fast tier, 468 MB), MiniCPM5-1B in int8 and int4 builds, and experimental Gemma 4 26B (A4B) — Google's mixture-of-experts Gemma for 16 GB+ RAM devices. |
| v3.3.1 | Fixes | GGUF models no longer crash on GPU on some Snapdragon devices (Adreno driver quirk). Rotating or folding the phone no longer unloads the model. Custom-ROM: TPU/GPU chat works again on de-Googled devices (GrapheneOS). |
| v3.3.0 | 🦯 Box Assist — a camera that talks (NEW) | Spoken camera assistance for blind and low-vision users, under the Core tab. Live mode calls out people, obstacles and objects with how close they are; Reading mode reads mail, labels and menus aloud; Describe mode describes the scene, spoken as it thinks; voice questions — double-tap, ask out loud, and Box answers against what the camera sees. One download bundles everything (vision models + the Describe brain + speech recognition). Continuous autofocus with pre-capture focus sweeps, automatic flashlight in the dark, a blur check on Reading, physical volume-button controls, hold-to-repeat, TalkBack coexistence, |