Awesome Ai Agents 2026
Section: Local LLM Runners · Run LLMs locally. 162k+ stars. Dead simple CLI.
Entry
Appears in 12 awesome lists
Ollama is a tool for running large language models locally, offering easy setup for macOS, Windows, Linux, and Docker, along with a library of models and quickstart guides for customization and integration github | github profile
Section: Local LLM Runners · Run LLMs locally. 162k+ stars. Dead simple CLI.
Section: Langchain · Get up and running with OpenAI gpt-oss, DeepSeek-R1, Gemma 3 and other models.
Section: LLM Providers · Get up and running with large language models locally.
Section: AI · Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Section: 推理 Inference · Get up and running with Llama 3, Mistral, Gemma, and other large language models.
Section: Inference engines · get up and running with LLMs
Section: 3. Inference Engines & Serving · Dead-simple local LLM runner with a one-line install, model registry, and OpenAI-compatible API.
Section: Industry Strength Natural Language Processing · Get up and running with large language models, locally.
Section: Tools & Libraries · Run LLMs locally — desktop app, multimodal, structured outputs
Section: Repositories · Ollama is a tool for running large language models locally, offering easy setup for macOS, Windows, Linux, and Docker, along with a library of models and quickstart guides for customization and integration github | github profile
Section: Local LLM Deployment · Get up and running with large language models locally.
Section: LLM and Inference · Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Langchain integrates various providers like Anthropic, AWS, and OpenAI, and offers tools for components such as LLMs, chat models, and data analysis, supporting functionalities from Alpha Vantage to YouTube github | docs
Unified proxy and SDK that routes to 100+ LLM providers behind a single OpenAI-compatible interface, with a Router handling retry/fallback across deployments, per-project cost and rate-limit tracking, and OTEL callback integrations. The right infrastructure layer when your harness needs provider…
(MIT) provides modules for structured outputs at different levels of abstraction, including output parsers for text completion endpoints, Pydantic programs for mapping prompts to structured outputs using function calling or output parsing, and pre-defined Pydantic programs for specific output types.
robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video,…
State-of-the-art serving engine with PagedAttention and continuous batching. Currently the fastest production-grade LLM server.
Flowise simplifies the creation of applications leveraging large language models (LLMs) by providing a drag-and-drop interface for customizing AI workflows, offering easy installation, Docker support, development tools, and documentation for integrating various functionalities such as…
Open WebUI is an extensible, feature-rich, and user-friendly self-hosted AI platform designed to operate entirely offline. It supports various LLM runners like Ollama and OpenAI-compatible APIs, with built-in inference engine for RAG, making it a powerful AI deployment solution.
LM Studio offers a platform for running various local LLMs like LLaMa, Falcon, MPT, and others offline, featuring a Chat UI, OpenAI-compatible server, and model downloads from Hugging Face, with support for Mac, Windows, and Linux, emphasizing privacy and no data collection, free for personal use…