A curated list of awesome embedded programming.
-
Updated
Aug 12, 2026
A curated list of awesome embedded programming.
Embedded and mobile deep learning research resources
Embedded agents your customers use to automate work, build views, and connect their tools.
Full face stack that runs entirely in the browser. Detection, 576-point 3D mesh, recognition, anti-spoof, smile — all WebAssembly, zero server. Apache 2.0.
benchmark for embededded-ai deep learning inference engines, such as NCNN / TNN / MNN / TensorFlow Lite etc.
Noodle provides primitive modular functions for convolution layer, dense layer, pooling and activations. It allows streaming the intermediate activations, weights, and biases from/to SD/FFat/SD_MMC filesystems to overcome RAM limitations.
Legend of Elya — N64 game with a real 6.36M-parameter ternary transformer on the VR4300 MIPS III CPU. Zelda-style dungeon, AI NPCs, byte-level inference at 1.23 tok/s scalar / 2.19 tok/s on the RSP overlay (measured under ares, never on silicon). Built with libdragon.
Read chapters directly from this repo - do not use GitHub Pages link.
414 KB WASM runtime for Needle a 14M-parameter tool-calling transformer. Runs in browser, Cloudflare Workers, and Node.js. No backend required.
POC visual search with smart glasses and Qdrant Edge.
World's First NMS-Free YOLOv26n on ESP32-P4. Features end-to-end Int8 QAT and custom C++ optimizations achieving 30% faster inference than the official ESP-DL YOLOv11n (1.7s vs 2.4s).
RPI (Resonant Permutation Inference) — Zero-multiply text generation. 18K tok/s. 868 KB models. Standalone or as speculative draft engine for LLMs. Runs on N64, POWER8, x86, ARM.
Open-source Hardware AI agent. Single Rust binary for cameras, sensors, robots, and IoT fleets — orchestrated by AI agents with memory and real-time telemetry. Runs on Jetson, Raspberry Pi, any Linux.
Culturally-compliant video storage. Embeds searchable text chunks into pixelated media for lightning-fast semantic search. Zero-database, maximum compliance.
Speech Recognition using STM32 and Machine Learning
Auditable offline edge intelligence for low-cost edge devices, with benchmark evidence and public board proof on ESP32-C3.
Run a 28.9M-parameter TinyLM on ESP32-S3 with an RP2040 OLED display node for fully local embedded AI inference.
Ultra-lightweight C++ inference engine for BitMamba-2 (1.58-bit SSM). Runs 1B models on consumer CPUs at 50+ tok/s using <700MB RAM. No heavy dependencies.
AcousticsLab is a cross-platform framework for sound and vibration analysis.
Efficient semantic segmentation model for traversable ground estimation in outdoor robotics, optimized for ESP32 microcontrollers with Rust embeddings and retraining extensions.
Add a description, image, and links to the embedded-ai topic page so that developers can more easily learn about it.
To associate your repository with the embedded-ai topic, visit your repo's landing page and select "manage topics."