A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.
Categories & Topics
Repository Stats
Related Tools
agentic-rag-financial-parser
Enterprise RAG ecosystem managing 32000+ semantic chunks. Features hybrid parsing (LlamaParse/PyMuPDF) and 256-dim MRL embeddings for 512MB RAM environments
ds4
DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm
LiteRT-LM
LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge devices.
Lumina
On-device AI agent runtime with a C/C++ core and Swift host app. Tool execution, permissions, and resumable sessions.
needle
Foundation model for tiny devices; 14mb, 26m params, 1-6k toks/sec on mobiles, wearables smart home and robots.
thunderbolt
AI You Control: Choose your models. Own your data. Eliminate vendor lock-in.