Open Source
Explore the latest AI open-source projects from GitHub and HuggingFace.
Explore the latest AI open-source projects from GitHub and HuggingFace.
Explore the latest AI open-source projects from GitHub and HuggingFace. Discover projects across categories like LLM, Vision, Audio, and more.
760 projects
NVlabs
NVIDIA's open vision-language model family using data-centric strategies, spanning Eagle, Eagle 2, Eagle 2.5 with 128K context, and the new LocateAnything generalist grounding model.
OpenMOSS
Open-source speech and sound generation model family from OpenMOSS with multilingual TTS, dialogue, real-time voice agents, voice design, and sound effects under a unified audio tokenizer.
GPUStack
Open-source GPU cluster manager that turns heterogeneous accelerators into a self-hosted, OpenAI-compatible model-as-a-service platform powered by vLLM, SGLang, and llama.cpp.
kvcache-ai (Moonshot AI)
KV-cache-centric LLM serving platform open-sourced by Moonshot AI, featuring disaggregated prefill and decode, RDMA-based KV transfer, and adapters for vLLM and SGLang.
LMCache
Open-source KV cache layer that lets LLM serving systems reuse previously computed tokens across replicas, dramatically reducing time-to-first-token for RAG, long-context, and multi-turn workloads.
vLLM Project
Kubernetes-native control plane from the vLLM team for GenAI inference, providing autoscaling, cache-aware routing, distributed KV cache, and high-density LoRA adapter serving.
turboderp
MIT-licensed quantization and inference library for running large LLMs on single consumer-class NVIDIA GPUs, with the new EXL3 format and Marlin-inspired memory-bound GEMM kernels.
ModelTC
Pure-Python LLM inference and serving framework from ModelTC with tri-process asynchronous architecture, token-level KV cache management, and Nopad attention for high GPU utilization.
fastgs
CVPR 2026 Highlight: official implementation of 'FastGS' — a general framework that trains 3D Gaussian Splatting scenes in ~100 seconds at PSNR parity with the Inria reference. MIT licensed, 1,100+ stars.
zai-org
Z.ai's open-source GLM-4.1V/4.5V/4.6V vision-language family with explicit 'thinking' reasoning mode and RLCS training, released under Apache 2.0 with 2,300+ GitHub stars.
speaches-ai
MIT-licensed self-hosted server that speaks the OpenAI Audio API, with faster-whisper for streaming transcription and translation plus piper and Kokoro for TTS. 3,300+ stars and 398 forks.
Rohit Ghumare
Apache 2.0 persistent memory layer for AI coding agents. Runs as a single local server on port 3111, talks to Claude Code, Cursor, Codex, Gemini CLI, OpenClaw, Hermes, pi, and OpenCode via 53 MCP tools and 12 auto-hooks. Benchmarked at 95.2 percent recall and 92 percent fewer tokens.