Use your locally running AI models to assist you in your web browsing
-
Updated
Sep 13, 2026 - TypeScript
Use your locally running AI models to assist you in your web browsing
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.
Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc
StenographAI is the secure privacy-first AI notepad & notetaker for confidential conversations in government & defence sectors. On Windows & MacOS.
A generalized information-seeking agent system with Large Language Models (LLMs).
[ICML 2024] SqueezeLLM: Dense-and-Sparse Quantization
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
A flexible, AI powered C2 framework built with operators in mind
Open-source AI IDE powered by local & cloud LLMs. A privacy-first alternative to Cursor.
Unified management and routing for llama.cpp, MLX and vLLM models with web dashboard.
Fenix Ai Trading Bot with LangGraph and ollama and multipe providers
Run Open Source/Open Weight LLMs locally with OpenAI compatible APIs
A Cli, a webUI, and a MCP server for the Z-Image-Turbo text-to-image generation model (Tongyi-MAI/Z-Image-Turbo base model as well as quantized models)
MVP of an idea using multiple local LLM models to simulate and play D&D
The PyVisionAI Official Repo
Run multiple resource-heavy Large Models (LM) on the same machine with limited amount of VRAM/other resources by exposing them on different ports and loading/unloading them on demand
To associate your repository with the localllm topic, visit your repo's landing page and select "manage topics."