Open-source repositories gaining traction right now.
94 repositories
#llm
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D taking face running locally across platforms

LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar

Star google / adk-python An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.

Sponsor Star Mintplex-Labs / anything-llm The all-in-one Desktop & Docker AI application with built-in RAG, AI agents, No-code agent builder, MCP compatibility, and more.

Composio powers 1000+ toolkits, tool search, context management, authentication, and a sandboxed workbench to help you build AI agents that turn intent into action.

Star BoundaryML / baml The AI framework that adds the engineering to prompt engineering (Python/TS/Ruby/Java/C#/Rust/Go compatible)

Star memvid / memvid Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

Star SemiAnalysisAI / InferenceX Open Source Continuous Inference Benchmarking Qwen3.5, DeepSeek, GPTOSS - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 vs H100 & soon™ TPUv6e/v7/Trainium2/3

Star NVIDIA-NeMo / Automodel Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Star open-webui / open-webui User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

Star google / langextract A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.

Star SillyTavern / SillyTavern LLM Frontend for Power Users.

Star screenpipe / screenpipe screenpipe turns your computer into a personal AI that knows everything you've done. record. search. automate. all local, all private, all yours.


Star poloclub / transformer-explainer Transformer Explained Visually: Learn How LLM Transformer Models Work with Interactive Visualization

Star maximhq / bifrost Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ models support & <100 µs overhead at 5k RPS.

Star OpenBMB / UltraRAG A Low-Code MCP Framework for Building Complex and Innovative RAG Pipelines

🦞 OpenClaw (Clawdbot/Moltbot) 汉化版 - 开源个人 AI 助手中文版 | Claude/ChatGPT LLM 接入 | WhatsApp/Telegram/Discord 多平台 | 每小时自动同步 | CLI + Dashboard 全中文 | 全流程搭建教程,以及排错指南!

A personal AI assistant built in Rust. Single binary, multi-provider LLMs, long-term memory, sandboxed execution, voice, MCP tools, and multi-channel access (web, Telegram, API).

Sponsor Star icereed / paperless-gpt Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI

Open source implementation and extension of Google Research’s PaperBanana for automated academic figures, diagrams, and research visuals, expanded to new domains like slide generation.

FireRedASR2S is a SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switching, and singing lyrics recognition. FireRedVAD supports speech/singing/music in 100+ langs. FireRedLID supports 100+ langs and 20+ zh dialects.

An open-source AI assistant framework like openclaw

42 AI agents. 9-layer consensus. One prompt. | Multi-agent orchestrator with ONNX neural routing, semantic convergence, and cross-LLM validation. https://nothumanallowed.com

Reduce Claude Desktop consumption by 10x - Integrate Google's Gemini or Z.ai's GLM-5 (744B params) with Claude via MCP for intelligent task delegation

Custom OpenClaw skills for AI agent capabilities - memory systems, prompt injection defense, intent routing, and platform integrations

molecule.ai

Persistent memory for AI coding agents

Context window compression and management utilities

End-to-end RAG pipeline with built-in evaluation metrics

Multi-provider LLM request router with fallback and cost tracking

Type-safe structured output extraction from LLMs
Paste a github.com URL. Submissions are reviewed before they are tracked.