Awesome Ai Agents 2026
Section: Local LLM Runners · Desktop app for local LLMs. Beautiful UI. All platforms.
Entry
Appears in 11 awesome lists
LM Studio offers a platform for running various local LLMs like LLaMa, Falcon, MPT, and others offline, featuring a Chat UI, OpenAI-compatible server, and model downloads from Hugging Face, with support for Mac, Windows, and Linux, emphasizing privacy and no data collection, free for personal use…
Section: Local LLM Runners · Desktop app for local LLMs. Beautiful UI. All platforms.
Section: Local LLM Tools · Desktop app for discovering, downloading, and running local LLMs with a built-in chat UI and API server.
Section: Running LLMs Locally · Discover, download, and run local LLMs
Section: 推理 Inference · Discover, download, and run local LLMs.
Section: Inference platforms · discover, download and run local LLMs
Section: Local AI · Discover, download, and run local LLMs with a user-friendly interface.
Section: Repositories · LM Studio offers a platform for running various local LLMs like LLaMa, Falcon, MPT, and others offline, featuring a Chat UI, OpenAI-compatible server, and model downloads from Hugging Face, with support for Mac, Windows, and Linux, emphasizing privacy and no data collection, free for personal use…
Section: LLMs · is a tool to Discover, download, and run local LLMs.
Section: 링크 · 로컬 컴퓨터에서 대규모 언어 모델을 테스트하고 실행할 수 있는 데스크톱 애플리케이션
Section: Local LLM Deployment · Download and run local LLMs on your computer.
Section: AI in Planning Tools and Platforms · This free for personal use software enables users to download large language models and run them locally within a desktop chat interface.
robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video,…
Ollama is a tool for running large language models locally, offering easy setup for macOS, Windows, Linux, and Docker, along with a library of models and quickstart guides for customization and integration github | github profile
State-of-the-art serving engine with PagedAttention and continuous batching. Currently the fastest production-grade LLM server.
is an ecosystem of open-source chatbots trained on a massive collections of clean assistant data including code, stories and dialogue based on LLaMa.
Pure C/C++ inference engine with GGUF format support. The gold standard for CPU/GPU/Apple Silicon on-device running. Includes llama-server for OpenAI-compatible API. Now at 100K+ stars.
Get up and running with Llama 3.3, DeepSeek-R1, Phi-4, Gemma 3, and other large language models. (Source Code) MIT Docker/Python