High-Performance server for NATS.io, the cloud and edge native messaging system.
-
Updated
Oct 2, 2026 - Go
Edge Computing and Artificial Intelligence of Things (AIoT) involve performing localized data processing directly on edge devices rather than relying entirely on centralized cloud servers. This architectural approach combines Internet of Things (IoT) hardware with optimized, lightweight machine learning models to enable real-time decision-making, drastically reduce network latency, and improve overall data privacy and bandwidth efficiency.
High-Performance server for NATS.io, the cloud and edge native messaging system.
Automation foundation model for tiny devices: 2-bit, 8-29 MB, tool calls, structured extraction and embeddings on phones, wearables, smart homes, robots, cars and microcontrollers.
PyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge devices.
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
Open-Source Personal Cloud OS for Always-On Agents
FEDML - The unified and scalable ML library for large-scale distributed training, model serving, and federated learning. FEDML Launch, a cross-cloud scheduler, further enables running any AI jobs on any GPU cloud or on-premise cluster. Built on this library, TensorOpera AI (https://TensorOpera.ai) is your generative AI platform at scale.
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text, Stable Diffusion, tool calling, and local-network servers. Runs on your CPU, GPU, or NPU. No account, no API key, zero data leaves your device.
[ICLR 2020] Once for All: Train One Network and Specialize it for Efficient Deployment
Build AI agents that run 100% on-device. Sub-100ms latency on Qualcomm NPU. Zero cloud dependency.
Deep learning gateway on Raspberry Pi and other edge devices
Production-grade C++ edge AI engine for video analytics and on-device VLM across Sophon, Rockchip RKNN, and x86, with visual orchestration, real-time OSD, events, and reproducible benchmarks.
jevos is an open-source alternative to Jev for yes/no decisions that runs on your laptop.
Microsoft AI for Good Lab — Biodiversity research hub. Open-source AI models, edge devices, and tools for biodiversity monitoring and conservation. Your source for MegaDetector, SPARROW, PytorchWildlife, Bioacoustics, and more.
DefraDB is a Peer-to-Peer Edge-First Database. It's the core data storage system for the Source Ecosystem.
Android 17 local LLM prototype with Jetpack Compose and ONNX Runtime for offline AI inference experiments.
LibreYOLO is a MIT licensed open source computer vision library