AI/ML Sr Manager @ Accenture | GenAI, Agentic Systems, Enterprise AI | Cloud & Platforms | Speaker | OSS
Pinned Loading
-
turboquant-mlx
turboquant-mlx PublicExtreme weight + KV cache compression for LLMs on Apple Silicon (MLX implementation of Google's TurboQuant)
-
esp32-tinyllm
esp32-tinyllm PublicInteractive on-device storytelling with a 28.9M-parameter LLM on an $8 ESP32-S3, using Gemma-style Per-Layer Embeddings to keep 25M parameters in flash.
C 4
-
esp32-gpio-llm
esp32-gpio-llm PublicA 312K-parameter language model that turns English into GPIO commands, running entirely on an ESP32-S3. No WiFi, no cloud, no API key.
Python 6
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


