category
Open Source Machine Learning Infrastructure Tools
Open source machine learning infrastructure tools
All Machine Learning Infrastructure tools
Run open-source LLMs locally on your own machine
Run LLMs locally with minimal setup, maximum hardware support
A high-throughput and memory-efficient inference and serving engine for LLMs
Run open-source AI models privately on your own device
LlamaIndex is the document processing platform for AI
Self-hosted AI runtime for text, voice, vision, and agents
📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG
Open source LLM engineering platform for AI-powered applications
Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory across sessions with a self-hosted knowledge graph engi
Turns Data and AI algorithms into production-ready web applications in no time.
An orchestration platform for the development, production, and observation of data assets.
A lightweight, lightning-fast, in-process vector database
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and v
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
Code search MCP for Claude Code. Make entire codebase the context for any coding agent.
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
Always know what to expect from your data.
Open-source LLM tracing & evaluation for AI optimization
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
One Postgres for your application data, full-text search, vector retrieval, and aggregations. Home of the pg_search extension.
Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.
MariaDB server is a community developed fork of MySQL server. Started by core members of the original MySQL team, MariaDB actively works with outside developers
Monitor LLM performance with open-source observability
A framework for efficient model inference with omni-modality models
Open-source framework for building agentic apps in JavaScript, Go, Dart, and Python, built and used in production by Google
Build reliable AI apps with comprehensive observability
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
MineContext is your proactive context-aware AI partner(Context-Engineering+ChatGPT Pulse)
Open-Source Personal Cloud OS for Always-On Agents
Simulation-based testing and evaluation for AI agents
Build, automate, and improve AI agents with your team
The AI-native database built for LLM applications, providing incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text.
Full observability for AI agents in production
Local persistent memory store for LLM applications including claude desktop, github copilot, codex, antigravity, etc.
Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs
High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle
Knowhere extracts, parses, and outputs structured chunks ready for AI Agents and RAG.
AI-powered platform for engineering LLM products
High-performance inference framework for large language models, focusing on efficiency, flexibility, and availability.
The Context Layer for unstructured data: typed, versioned datasets over S3, GCS, Azure
Monitor, debug, and scale LLM applications with ease
OpenLake is a high performance storage engine for efficient LLM inference and GPU Training
Build applications that make decisions (chatbots, agents, simulations, etc...). Monitor, trace, persist, and execute on your own infrastructure.
GPU orchestration across clouds, Kubernetes, and on-prem
🏕️ Reproducible development environment for humans and agents
Serverless GPU compute with sub-second cold starts
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™
🔥 Real-time NVIDIA GPU dashboard
An open source DevOps tool from the CNCF for packaging and versioning AI/ML models, datasets, code, and configuration into an OCI Artifact.
SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.
LightlyStudio - The Unified Data Platform for Multimodal ML
Serverless LLM Serving for Everyone.
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
Experiment tracking platform for ML teams
LLM observability: cost, latency, and traces in one place
Trace, evaluate, and protect every LLM call in one SDK
Frequently asked questions
How many open source Machine Learning Infrastructure tools are there?
This registry tracks 57 open source Machine Learning Infrastructure projects, with 955,231 combined GitHub stars. The list is ranked by stars and refreshed nightly.
What is the most popular open source Machine Learning Infrastructure project?
Ollama leads this category with 181,161 GitHub stars, followed by llama.cpp.
Are these Machine Learning Infrastructure tools free?
Yes — every project listed here is open source. Some also offer paid hosted versions alongside the free self-hosted option; the licence for each project is shown on its card.