Skip to content
89

Awesome Open Source AI

Curated list of the best truly open-source AI projects, models, tools, and infrastructure. Daily updated.

4.8k stars672 forks1,062 entriesLast push Sep 30, 2026 (today)License CC0-1.0

This page lists names, links and short descriptions. The original list on GitHub is the source and belongs to its authors.

1. Core Frameworks & Libraries

PyTorch

Dynamic computation graphs, Pythonic API, dominant in research and production. The current standard for most frontier AI work.

In 16 listsDetails

TensorFlow

End-to-end platform with excellent production deployment, TPU support, and large-scale serving tools.

In 23 listsDetails

JAX

High-performance numerical computing with composable transformations (JIT, vmap, grad). Rising favorite for research and scientific ML.

In 8 listsDetails

Flax

Neural network library for JAX, designed for flexibility. Apache-2.0 licensed.

In 5 listsDetails

dm-haiku

JAX-based neural network library from Google DeepMind. Elegant functional API with state management, widely used in DeepMind's research. Apache 2.0 licensed.

Equinox

Elegant easy-to-use neural networks and scientific computing in JAX. Callable PyTrees with filtered transformations, seamless interoperability with the JAX ecosystem. Apache 2.0 licensed.

In 2 lists

Diffrax

Numerical differential equation solvers in JAX. Autodifferentiable and GPU-capable ODE/SDE/CDE solvers for scientific machine learning and neural differential equations. Apache 2.0 licensed.

In 3 lists

vit-pytorch

Comprehensive Vision Transformer (ViT) implementations in PyTorch. Reference implementations of all major vision transformer variants including ViT, DeiT, Swin, and more. MIT licensed.

In 2 lists

NumPyro

Probabilistic programming with NumPy powered by JAX for autograd and JIT compilation. Bayesian modeling and inference at scale.

In 3 lists

Keras

High-level, beginner-friendly API that now runs on multiple backends (TensorFlow, JAX, PyTorch). Perfect for rapid experimentation.

In 9 listsDetails

tinygrad

Minimalist deep learning framework with tiny code footprint. The "you like PyTorch? you like micrograd? you love tinygrad!" philosophy - simple yet powerful.

In 3 lists

PaddlePaddle

Industrial deep learning platform from Baidu serving 23+ million developers and 760,000+ companies. China's first independent R&D framework with advanced distributed training and deployment capabilities.

In 6 listsDetails

PyTorch Geometric

Library for deep learning on irregular input data such as graphs, point clouds, and manifolds. Part of the PyTorch ecosystem.

In 7 listsDetails

timm (PyTorch Image Models)

The largest collection of PyTorch image encoders and backbones. 900+ pretrained models including ResNet, EfficientNet, Vision Transformer, ConvNeXt, and more with training and inference scripts. Apache 2.0 licensed.

In 2 lists

Triton

Language and compiler for writing highly efficient custom deep-learning primitives. Powers kernel optimizations in PyTorch, JAX, and other frameworks. MIT licensed.

In 2 lists

GGML

Tensor library for machine learning. The foundational C/C++ library powering llama.cpp and many on-device inference engines. MIT licensed.

In 3 lists

MLX

Array framework for machine learning on Apple silicon. Efficient unified memory design with NumPy-like API, automatic differentiation, and multi-device support. MIT licensed.

In 4 listsDetails

notorch

Neural network framework in pure C with reverse-mode automatic differentiation, model training, and GGUF inference, without a Python runtime.

In 2 lists

oneDNN

oneAPI Deep Neural Network Library. Cross-platform performance library of basic building blocks for deep learning, optimized for Intel CPUs, GPUs, and Arm architectures. Apache 2.0 licensed.

In 2 lists

ONNX

Open standard for machine learning interoperability. Open Neural Network Exchange provides an open ecosystem that empowers AI developers to choose the right tools as their project evolves. Apache 2.0 licensed.

In 6 listsDetails

IREE

Retargetable MLIR-based machine learning compiler and runtime toolkit. Lowers ML models to unified IR that scales from datacenter to mobile and edge deployments. Apache 2.0 licensed.

In 2 lists

Modular Platform

Open-source AI compute and programming platform built around the MAX Engine and Mojo programming language.

Burn

Next-generation deep learning framework in Rust. Backend-agnostic with CPU, GPU, WebAssembly support.

In 4 listsDetails

Candle (Hugging Face)

Minimalist ML framework for Rust. PyTorch-like API with focus on performance and simplicity.

In 5 listsDetails

linfa

Comprehensive Rust ML toolkit with classical algorithms. scikit-learn equivalent for Rust with clustering, regression, and preprocessing.

In 3 lists

Flux.jl

100% pure-Julia ML stack with lightweight abstractions on top of native GPU and AD support. Elegant, hackable, and fully integrated with Julia's scientific computing ecosystem.

In 2 lists

MLJ.jl

Comprehensive Julia machine learning framework providing a unified interface to 200+ models with meta-algorithms for selection, tuning, and evaluation. MIT licensed.

ModelingToolkit.jl

High-performance symbolic-numeric modeling framework for scientific machine learning. Automatically generates fast functions for model components like Jacobians and Hessians with automatic sparsification and parallelization. MIT licensed.

In 2 lists

spaCy (Explosion AI)

Industrial-strength natural language processing with 75+ languages, transformer pipelines, and production-grade NER, parsing, and text classification.

In 7 listsDetails

Transformers (Hugging Face)

The de facto standard library for pretrained NLP models. 1M+ models, 250,000+ downloads/day. BERT, GPT, Llama, Qwen, and hundreds more.

In 14 listsDetails

sentence-transformers

Classic library for sentence and image embeddings.

In 5 listsDetails

tokenizers (Hugging Face)

Fast state-of-the-art tokenizers for training and inference.

In 6 listsDetails

fairseq2

FAIR Sequence Modeling Toolkit 2. Complete rewrite of fairseq with modern PyTorch APIs, native support for LLM training (70B+ models), vLLM integration, and first-party recipes for instruction finetuning and preference optimization. MIT licensed.

LibreTranslate

Self-hosted machine translation API powered by the Argos Translate engine. AGPL-3.0 licensed.

In 3 lists

Pandas

The gold standard for data analysis and manipulation in Python.

In 9 listsDetails

Polars

Blazing-fast DataFrame library (Rust backend) - modern alternative to Pandas for large-scale workloads.

In 10 listsDetails

cuDF

GPU DataFrame library from RAPIDS. Accelerates Pandas workflows on NVIDIA GPUs with zero code changes using cuDF.pandas accelerator mode.

In 5 listsDetails

Dask

Parallel computing for big data - scales Pandas/NumPy/scikit-learn to clusters.

In 11 listsDetails

DataFlow

LLM-ready data preparation system for turning raw PDFs, conversations, code, databases, and other sources into SFT, QA, and RAG-ready datasets.

In 3 lists

NumPy

Fundamental array computing library that powers almost every AI stack.

In 6 listsDetails

SciPy

Scientific computing algorithms (optimization, linear algebra, statistics, signal processing).

In 6 listsDetails

CuPy

NumPy and SciPy-compatible array library for GPU-accelerated computing in Python.

In 8 listsDetails

NetworkX

Creation, manipulation, and study of complex networks. The foundational graph analysis library for Python data science.

In 6 listsDetails

cuGraph

GPU graph analytics library with NetworkX-compatible API. 10-100x faster than CPU for large-scale graph algorithms. Apache 2.0 licensed.

In 2 lists

Vaex

Out-of-Core hybrid Apache Arrow/NumPy DataFrame for Python. Visualize and explore billion-row datasets at millions of rows per second. MIT licensed.

In 9 listsDetails

Datashader

High-performance large data visualization. Renders billions of points interactively without aggregation artifacts. BSD-3-Clause licensed.

In 2 lists

Zarr

Chunked, compressed, N-dimensional array storage. Scalable tensor data format optimized for cloud and parallel computing. MIT licensed.

In 4 listsDetails

NVIDIA DALI

GPU-accelerated data loading and augmentation library with highly optimized building blocks for deep learning applications. Apache 2.0 licensed.

In 2 lists

Narwhals

Lightweight compatibility layer between DataFrame libraries. Write Polars-like code that works seamlessly across Pandas, Polars, cuDF, Modin, and more. MIT licensed.

In 3 lists

Ibis

Portable Python dataframe library with 20+ backends. Write Pandas-like code that runs locally with DuckDB or scales to production databases (BigQuery, Snowflake, PostgreSQL) by changing one line. Apache 2.0 licensed.

In 4 listsDetails

skrub

Machine learning with dataframes for dirty categorical data. Preprocessing and feature engineering for heterogeneous data with seamless Pandas/Polars integration. BSD-3-Clause licensed.

In 3 lists

Oxen

Lightning fast data version control for machine learning. Optimized for large datasets with efficient diffing, branching, and collaboration. Apache 2.0 licensed.

Pandera

Statistical data testing and validation for dataframes. Pydantic-like API for Pandas, Polars, and other dataframe libraries with type hints and lazy validation. MIT licensed.

In 4 listsDetails

Snorkel

System for quickly generating training data with weak supervision. Programmatically label, build, and manage training data using labeling functions and probabilistic consensus models. Powers Snorkel Flow and used by Google, Apple, and Intel. Apache 2.0 licensed.

In 6 listsDetails

DuckDB

High-performance analytical in-process SQL database system. Fast, reliable, portable, and easy to use with rich SQL dialect support. Perfect for data processing and analytics workloads. MIT licensed.

In 8 listsDetails

FiftyOne

Visual AI development toolkit for visualizing, labeling, and evaluating visual datasets and models. Supercharges computer vision workflows with dataset exploration and model analysis. Apache 2.0 licensed.

In 3 lists

Label Studio

Multi-type data labeling and annotation tool with standardized output format. Configurable interface for images, text, audio, video, and time series with ML-assisted labeling. Apache 2.0 licensed.

In 3 lists

Delta Lake

Open-source storage framework enabling Lakehouse architecture with ACID transactions, scalable metadata handling, and unified batch/streaming processing. Apache 2.0 licensed.

In 6 listsDetails

Apache Iceberg

High-performance open table format for huge analytic tables. Brings SQL table reliability to big data with time travel, hidden partitioning, and schema evolution. Works with Spark, Trino, Flink, Presto, Hive and Impala. Apache 2.0 licensed.

In 4 listsDetails

Apache Hudi

Open data lakehouse platform for ingesting, indexing, storing, serving, transforming and managing data across cloud environments. Supports upserts, deletes and incremental processing on big data with built-in ingestion tools for Spark and Flink. Apache 2.0 licensed.

In 4 listsDetails

lakeFS

Data version control for your data lake that transforms object storage into Git-like repositories. Enables atomic, versioned data lake operations with branching, committing, and merging for data pipelines. Apache 2.0 licensed.

In 10 listsDetails

Apache Airflow

Platform to programmatically author, schedule, and monitor workflows. Industry-standard orchestration for data pipelines and ML workflows with 500+ integrations. Apache 2.0 licensed.

In 13 listsDetails

Apache Spark

Unified analytics engine for large-scale data processing. In-memory cluster computing with high-level APIs in Python, Scala, Java, and R. Powers MLlib for distributed machine learning and Structured Streaming for real-time data. Apache 2.0 licensed.

In 8 listsDetails

Apache Flink

Stream processing framework with powerful batch and streaming capabilities. High-throughput, low-latency runtime with exactly-once processing guarantees. Ideal for real-time AI inference pipelines and event-driven ML applications. Apache 2.0 licensed.

In 7 listsDetails

Apache Beam

Unified programming model for batch and streaming data processing. Write pipelines once, run anywhere on Flink, Spark, or Google Cloud Dataflow. Portable, extensible, and enterprise-ready for AI data pipelines. Apache 2.0 licensed.

In 6 listsDetails

Scrapy

Fast, high-level web crawling and scraping framework for Python. Extract structured data from websites at scale with built-in support for handling common challenges like pagination, cookies, and concurrent requests. BSD-3-Clause licensed.

In 6 listsDetails

Temporal

Durable execution platform for reliable workflow orchestration. Build resilient data pipelines and ML workflows that survive failures and continue execution exactly where they left off. MIT licensed.

In 6 listsDetails

Luigi

Python module for building complex pipelines of batch jobs. Handles dependency resolution, workflow management, visualization, and Hadoop integration. Built at Spotify and battle-tested in production. Apache 2.0 licensed.

In 15 listsDetails

Mage.ai

Modern open-source data pipeline tool for integrating and transforming data. AI-native ETL/ELT platform with 100+ integrations, real-time monitoring, and collaborative features. Apache 2.0 licensed.

In 2 lists

Hamilton

Declarative dataflow framework for building testable, modular, self-documenting data pipelines. Encode lineage and metadata directly in Python functions. Originally from Stitch Fix, now Apache incubating. Apache 2.0 licensed.

In 2 lists

D-Tale

Visualizer for Pandas data structures with a Flask back-end and React front-end. Interactive data exploration with charting, filtering, and code export. LGPL-2.1 licensed.

In 5 listsDetails

Sweetviz

Beautiful, high-density visualizations for exploratory data analysis in two lines of code. Self-contained HTML reports for dataset comparison and target analysis. MIT licensed.

In 4 listsDetails

TextAttack

Python framework for adversarial attacks, data augmentation, and model training in NLP. Augment datasets to increase model robustness and generate adversarial examples. MIT licensed.

In 6 listsDetails

uv

An extremely fast Python package and project manager, written in Rust. 10-100x faster than pip with built-in virtual environment management, dependency resolution, and lockfiles. Essential for modern AI/ML development workflows. Apache 2.0 and MIT dual-licensed.

In 5 listsDetails

Vector

A high-performance observability data pipeline for collecting, transforming, and routing logs and metrics. Real-time data processing with 50+ sources and sinks including Kafka, S3, and Elasticsearch. Ideal for AI/ML log processing and data ingestion. MPL 2.0 licensed.

In 5 listsDetails

scikit-learn

Industry-standard library for traditional machine learning (classification, regression, clustering, pipelines).

In 10 listsDetails

XGBoost

Scalable, high-performance gradient boosting library. Still dominates Kaggle and tabular competitions.

In 11 listsDetails

LightGBM

Microsoft's ultra-fast gradient boosting framework, optimized for speed and memory.

In 8 listsDetails

CatBoost

Gradient boosting that handles categorical features natively with great out-of-the-box performance.

In 10 listsDetails

sktime

Unified framework for machine learning with time series. scikit-learn compatible API for forecasting, classification, clustering, and anomaly detection.

In 2 lists

StatsForecast

Lightning-fast statistical forecasting with ARIMA, ETS, CES, and Theta models. Optimized for high-performance time series workloads.

In 3 lists

MLForecast

Scalable machine learning for time series forecasting. Train any sklearn-compatible model on millions of time series with efficient feature engineering. Apache 2.0 licensed.

In 3 lists

cuML

GPU-accelerated machine learning algorithms with scikit-learn compatible API. 10-50x faster than CPU implementations for large datasets. Apache 2.0 licensed.

In 7 listsDetails

SynapseML

Distributed machine learning on Apache Spark. Scalable, composable APIs for text analytics, vision, anomaly detection with seamless Python/Scala/R/.NET integration. MIT licensed.

In 2 lists

Darts

User-friendly forecasting and anomaly detection for time series. Unifies classical statistical models (ARIMA, ETS) with modern neural networks (N-BEATS, TFT, DeepAR) in a single scikit-learn compatible API. Apache 2.0 licensed.

In 6 listsDetails

PyTorch Forecasting

Time series forecasting with PyTorch. Multiple neural architectures (N-BEATS, TFT, DeepAR) with in-built interpretation capabilities, built on PyTorch Lightning for distributed training. MIT licensed.

In 3 lists

DataHub

The #1 open-source metadata platform for data and AI. Data discovery, governance, and observability with 80+ connectors, column-level lineage, and AI assistant integration. Originally built at LinkedIn. Apache 2.0 licensed.

In 5 listsDetails

OpenMetadata

Unified metadata platform for data discovery, observability, and governance. Column-level lineage, semantic search, and team collaboration with 70+ data service connectors. Apache 2.0 licensed.

In 2 lists

Amundsen

Data discovery and metadata engine from Lyft. PageRank-style search for data resources with usage-based ranking. LF AI & Data Foundation project. Apache 2.0 licensed.

In 3 lists

Apache Ossie

Vendor-neutral specification to standardize semantic models across analytics, BI, and AI agent platforms. Apache 2.0 licensed.

dbt-core

Transform data using software engineering best practices. The industry-standard framework for analytics engineering with 15M+ monthly downloads. Enables version control, testing, and documentation for SQL transformations. Apache 2.0 licensed.

In 7 listsDetails

SQLMesh

Scalable and efficient data transformation framework with dbt compatibility. Features automatic data lineage, time travel, and virtual data environments for testing. Optimized for large-scale data warehouses. Apache 2.0 licensed.

In 2 lists

SLayer

Semantic layer for AI-powered data analytics. Allows AI agents to describe data models and query the data using an expressive format with measures, dimensions, and filters, without writing raw SQL. MCP, CLI, API, and Python clients. Embeddable as a Python library. MIT licensed.

WrenAI

Open-source Generative BI engine and context layer for AI agents to query databases and produce trusted SQL and dashboards. Apache 2.0 licensed.

In 4 listsDetails

Deequ

Library built on top of Apache Spark for defining "unit tests for data". Measures data quality in large datasets with constraint verification, anomaly detection, and incremental validation. Used at Amazon for production data quality. Apache 2.0 licensed.

In 3 lists

Great Expectations

Always know what to expect from your data. Data validation, profiling, and documentation for data pipelines. Apache 2.0 licensed.

In 5 listsDetails

ydata-profiling

One line of code for comprehensive data quality profiling and exploratory data analysis. Generates detailed reports for Pandas and Spark DataFrames including statistics, correlations, missing values, and data quality alerts. MIT licensed.

In 4 listsDetails

Soda Core

Data contracts engine for the modern data stack. Define data quality checks in YAML and automatically validate schema and data across your pipelines. Supports 20+ data sources including Snowflake, BigQuery, and PostgreSQL. Apache 2.0 licensed.

TFX (TensorFlow Extended)

End-to-end platform for deploying production ML pipelines. Data validation, transformation, model training, and serving with TensorFlow. Powers Google's production ML infrastructure. Apache 2.0 licensed.

In 2 lists

Doccano

Open-source text annotation tool for machine learning practitioners. Features text classification, sequence labeling, and sequence-to-sequence tasks for sentiment analysis, NER, and summarization. MIT licensed.

In 3 lists

OpenRefine

Free, open-source power tool for working with messy data. Clean, transform, and extend data with web services. Formerly Google Refine. BSD-3-Clause licensed.

In 4 listsDetails

Optuna

Modern, define-by-run hyperparameter optimization with pruning and visualizations. Extremely popular in 2026.

In 8 listsDetails

AutoGluon

AWS AutoML toolkit for tabular, image, text, and multimodal data - state-of-the-art with almost zero code.

In 3 lists

FLAML

Microsoft's fast & lightweight AutoML focused on efficiency and low compute.

In 4 listsDetails

Katib (Kubeflow)

Kubernetes-native AutoML for hyperparameter tuning, early stopping, and neural architecture search. Framework-agnostic with support for TensorFlow, PyTorch, XGBoost, and custom training operators. Apache 2.0 licensed.

In 5 listsDetails

Streamlit

The fastest way to build and share data apps. Transform Python scripts into beautiful web applications with minimal code. Widely used for ML model demos, data visualization, and internal tools.

In 10 listsDetails

Gradio

Build and share delightful machine learning apps, all in Python. The de facto standard for creating interactive ML demos with automatic UI generation from function signatures. Powers thousands of Hugging Face Spaces.

In 11 listsDetails

Marimo

A reactive notebook for Python — run reproducible experiments, query with SQL, execute as a script, deploy as an app, and version with git. Stored as pure Python. All in a modern, AI-native editor.

In 5 listsDetails

Hugging Face Accelerate

Simple API to make training scripts run on any hardware (multi-GPU, TPU, mixed precision) with minimal code changes.

In 6 listsDetails

DeepSpeed

Microsoft's deep learning optimization library for extreme-scale training (ZeRO, offloading, MoE).

In 7 listsDetails

FlashAttention

Fast exact attention kernels that reduce memory usage and accelerate transformer training and inference.

In 2 lists

xFormers

Optimized transformer building blocks and attention operators for PyTorch.

In 2 lists

PyTorch Lightning

High-level wrapper for PyTorch that removes boilerplate and adds best practices.

In 3 lists

fastai

Deep learning library providing practitioners with high-level components for state-of-the-art results. Built on PyTorch with a focus on usability and transfer learning. Apache 2.0 licensed.

In 6 listsDetails

PyTorch Ignite

High-level library for training and evaluating neural networks in PyTorch with an engine, events & handlers system for maximum flexibility. BSD-3-Clause licensed.

In 7 listsDetails

ONNX Runtime

High-performance inference and training for ONNX models across hardware.

In 3 lists

einops

Flexible, powerful tensor operations for readable and reliable code. Supports PyTorch, JAX, TensorFlow, NumPy, MLX.

In 3 lists

safetensors

Simple, safe way to store and distribute tensors. Fast, secure alternative to pickle for model serialization.

torchmetrics

Machine learning metrics for distributed, scalable PyTorch applications. 80+ metrics with built-in distributed synchronization.

torchao

PyTorch native quantization and sparsity for training and inference. Drop-in optimizations for production deployment.

In 2 lists

SHAP

Game theoretic approach to explain the output of any machine learning model. Industry standard for model interpretability.

In 3 lists

skorch

scikit-learn compatible neural network library that wraps PyTorch. Seamlessly integrate PyTorch models with scikit-learn pipelines, grid search, and cross-validation.

In 5 listsDetails

Composer

Supercharge your model training. MosaicML's PyTorch training library with built-in algorithms for efficient training (FSDP, gradient compression, progressive resizing) and seamless distributed training on large-scale clusters. Apache 2.0 licensed.

In 2 lists

NVIDIA Apex

PyTorch extension for mixed precision training and distributed training optimizations. Powers many production deep learning workloads with tools for automatic mixed precision (AMP), distributed data parallel, and fused optimizers. BSD-3-Clause licensed.

In 2 lists

2. Model Codebases & Model Families

RWKV

Attention-free language model architecture with linear-time inference, training code, inference examples, and an active open-source ecosystem.

In 3 lists

MiniCPM

Compact open model family with practical code, deployment notes, and active edge/on-device focus.

In 3 lists

GPT-OSS

OpenAI open-weight model repository with inference examples, recipes, and deployment guidance.

In 4 listsDetails

Mamba

State Space Model implementation with pretrained checkpoints, architecture code, and research tooling for efficient long-sequence modeling.

In 2 lists

GPT-NeoX

Large-scale language model training codebase from EleutherAI with distributed training support and historical open-model importance.

In 3 lists

GLM-5

Open-source mixture-of-experts language model family optimized for long-horizon planning, agentic tasks, and coding. Apache 2.0 licensed.

In 2 lists

OpenCLIP

Open implementation of CLIP with training code, pretrained models, and zero-shot evaluation tooling.

In 2 lists

OmniParser

Vision-based GUI parsing model and tooling for computer-use agents.

In 2 lists

MiniCPM-V

Compact vision-language model family with edge-focused deployment examples and strong OCR-oriented use cases.

In 4 listsDetails

Eagle

NVIDIA multimodal model codebase with open checkpoints and reusable research materials for vision-language and video-language tasks.

Moondream

Small vision-language model with practical inference examples for edge and real-time image understanding.

NVIDIA Cosmos

Open platform of world models, tokenizers, and post-training tools designed for physical AI, robotics, and autonomous systems.

In 3 lists

Whisper

Canonical open speech-to-text model codebase with widespread ecosystem support and many downstream implementations.

In 11 listsDetails

FunASR

Speech recognition toolkit with pretrained models, streaming support, diarization, VAD, and production-oriented examples.

In 9 listsDetails

NVIDIA NeMo

Scalable framework and model codebase for speech, language, and multimodal AI with recipes and deployment guidance.

In 2 lists

Sherpa-ONNX

Complete speech toolkit with ASR, TTS, diarization, source separation, and VAD across embedded and edge environments via ONNX Runtime.

In 4 listsDetails

MOSS-TTS

Open speech and sound generation family focused on expressive, long-form text-to-speech with streaming and multi-speaker support.

VoxCPM

Open-sourced tokenizer-free multilingual speech synthesis model with high-quality TTS and style transfer workflows.

In 4 listsDetails

VibeVoice

Open Frontier Voice AI toolkit spanning speech understanding, generation, and multilingual TTS workflows, with active research and deployment tooling.

In 5 listsDetails

SpeechBrain

PyTorch speech toolkit with recipes for ASR, TTS, speaker recognition, and speech enhancement.

In 5 listsDetails

Pocket TTS

Lightweight text-to-speech engine optimized for CPU inference with low latency and streaming support. MIT licensed.

transcribe.cpp

C/C++ speech-to-text inference library running 16+ model families on the ggml runtime with GPU acceleration.

In 2 lists

Moonshine

Open-source on-device voice AI toolkit for low-latency speech-to-text, intent recognition, and text-to-speech.

In 3 lists

3. Inference Engines & Serving

llama.cpp

Pure C/C++ inference engine with GGUF format support. The gold standard for CPU/GPU/Apple Silicon on-device running. Includes llama-server for OpenAI-compatible API. Now at 100K+ stars.

In 10 listsDetails

Ollama

Dead-simple local LLM runner with a one-line install, model registry, and OpenAI-compatible API.

In 12 listsDetails

Foundry Local

Open-source on-device AI platform covering discovery, model running, sandboxed execution, and evaluation of open models.

Potato OS

Linux distribution for fully local AI inference on Raspberry Pi 5 and 4, optimized for running open models at the edge.

MLC-LLM

Deployment engine that compiles and runs LLMs across browsers, mobile devices, and local hardware.

In 4 listsDetails

WebLLM

High-performance in-browser LLM inference engine. Runs models directly in the browser with WebGPU acceleration.

In 2 lists

llama-cpp-python

Official Python bindings for llama.cpp.

In 4 listsDetails

KoboldCpp

User-friendly llama.cpp fork focused on role-playing and creative writing.

In 5 listsDetails

RamaLama

Container-centric tool for simplifying local AI model serving. Automatically detects GPUs, pulls optimized container images, and runs models securely in rootless containers with enterprise-grade isolation.

In 2 lists

LiteRT

Google's production-ready on-device ML and GenAI deployment framework. Supports Android, iOS, Web, Desktop, and IoT targets with GPU/NPU acceleration via a unified edge-first runtime. Apache 2.0 licensed.

In 4 listsDetails

LiteRT-LM

Production-ready runtime for deploying LLMs on edge devices with low-latency inference and optimized hardware paths for mobile and embedded platforms.

In 4 listsDetails

exo

Run frontier AI locally by connecting all your devices into an AI cluster. Features automatic device discovery, RDMA over Thunderbolt for 99% latency reduction, topology-aware auto parallel, and tensor parallelism. Uses MLX backend for distributed inference across Apple Silicon devices. Apache 2.0…

In 4 listsDetails

ds4

Native inference engine optimized for DeepSeek V4 and GLM models with Metal, CUDA, and ROCm support.

In 2 lists

qwen3.8-flash-next-in-c

Native C implementation of Qwen3.8-Flash-Next for local CPU inference, with terminal chat, a resident OpenAI-compatible API, and reproducible benchmarks.

omlx

Apple-centric inference server for local-first AI workflows with model management, GPU orchestration, and OpenAI-compatible APIs for self-hosted deployment. Apache 2.0 licensed.

In 3 lists

llmfit

Terminal tool and TUI that right-sizes LLM models to hardware specs and scores local compatibility across GPU, CPU, and RAM. MIT licensed.

In 4 listsDetails

Needle

Compact 45M-parameter foundation model and 14MB inference engine for tool calling and structured extraction on tiny devices. MIT licensed.

In 2 lists

Colibri

Zero-dependency C inference engine that runs large Mixture-of-Experts models locally by streaming experts across disk, RAM, and VRAM.

In 3 lists

Claude Code Local

MLX-native server that speaks the Anthropic Messages API so the unmodified Claude Code CLI runs against local models on Apple Silicon, with parsing for local models' tool-call formats. MIT licensed.

nemotron-omni-mlx

Pure MLX runtime for the vision and audio towers of NVIDIA Nemotron 3 Nano Omni on Apple Silicon, tested for parity against NVIDIA's PyTorch reference. MIT licensed.

Magnitude

Hardware-aware local inference engine that profiles host hardware, recommends suitable open models, and tunes execution across Apple Silicon, NVIDIA, AMD, and CPU. Apache 2.0 licensed.

llm-vram-calculator

Dependency-free TypeScript core that estimates inference memory (weights, KV cache, overhead) of any Hugging Face model from its config.json and safetensors metadata, covering MoE, MLA, sliding-window and linear-attention layers; runs in the browser at modelvram.com. MIT licensed.

llm-d

Kubernetes-native distributed LLM inference framework. Donated to CNCF by RedHat, Google, and IBM. Intelligent scheduling, KV-cache optimization, and state-of-the-art performance across accelerators.

LMDeploy

Toolkit for compressing, deploying, and serving LLMs from OpenMMLab. 4-bit inference with 2.4x higher performance than FP16, distributed multi-model serving across machines.

In 3 lists

vLLM

State-of-the-art serving engine with PagedAttention and continuous batching. Currently the fastest production-grade LLM server.

In 11 listsDetails

vLLM-Omni

Multi-modal inference stack extending vLLM for image, audio, and video generation workloads with a unified serving interface. Apache 2.0 licensed.

LMCache

Supercharge LLM inference with the fastest KV Cache layer. 3-10x delay savings and GPU cycle reduction for multi-round QA and RAG. Integrates seamlessly with vLLM for distributed, high-throughput deployments. Apache 2.0 licensed.

In 4 listsDetails

vLLM Production Stack

Kubernetes-native production stack for vLLM inference. Automated deployment, autoscaling, and monitoring for enterprise-grade LLM serving. Built by the vLLM team for seamless integration.

In 2 lists

Open Model Engine (OME)

Kubernetes operator for LLM serving with GPU scheduling and model lifecycle management across vLLM, SGLang, and TensorRT-LLM.

nano-vLLM

Minimalist vLLM implementation in ~1,200 lines of Python. Educational yet performant with prefix caching, tensor parallelism, and CUDA graph acceleration. Comparable inference speeds to full vLLM. MIT licensed.

In 5 listsDetails

SGLang

Next-gen serving framework with RadixAttention. Powers xAI's production workloads at 100K+ GPUs scale.

In 9 listsDetails

TensorRT-LLM

NVIDIA's official high-performance inference backend.

In 7 listsDetails

Aphrodite Engine

vLLM fork optimized for role-play and creative writing. Supports extensive quantization methods (AQLM, AWQ, GPTQ, GGUF, FP8) and modern samplers. Active development with multi-LoRA and speculative decoding support.

AIBrix

Cost-efficient and pluggable infrastructure components for GenAI inference. Kubernetes-native control plane for vLLM with distributed KV cache, heterogeneous GPU serving, and intelligent routing. Apache 2.0 licensed.

In 2 lists

AISIX

Self-hosted Apache-2.0-licensed AI gateway written in Rust, with OpenAI-compatible endpoints, an Anthropic-compatible Messages endpoint, model routing, traffic policies, Prometheus metrics, and OTLP trace export.

In 2 lists

Triton Inference Server

NVIDIA's production-grade open-source inference serving software. Supports multiple frameworks (TensorRT, PyTorch, ONNX) with optimized cloud and edge deployment.

In 5 listsDetails

mistral.rs

Fast, flexible Rust-native LLM inference engine built on Candle. Supports text, vision, audio, image generation, and embeddings with hardware-aware auto-tuning.

In 3 lists

KTransformers

Flexible framework for heterogeneous CPU-GPU LLM inference and fine-tuning. Enables running large MoE models by offloading experts to CPU with BF16/FP8 precision support.

In 4 listsDetails

llamafile

Mozilla's single-file distributable LLM solution. Bundle model weights, inference engine, and runtime into one portable executable that runs on six OSes without installation.

In 2 lists

Xinference

Unified, production-ready inference API for LLMs, speech, and multimodal models. Drop-in GPT replacement with single-line code changes. Supports thousands of models with auto-batching and distributed inference.

In 5 listsDetails

RTP-LLM (Alibaba)

Alibaba's high-performance LLM inference acceleration engine. Powers production LLM services across Taobao, Tmall, and Alibaba's international AI platform. Supports PagedAttention, FlashAttention, FlashDecoding, INT8/INT4 quantization, and heterogeneous hardware (GPU/ARM CPU/Intel). Apache 2.0…

LitServe (Lightning AI)

Minimal Python framework for building custom AI inference servers with full control over logic, batching, and scaling. 2x faster than FastAPI with built-in batching, streaming, and multi-GPU autoscaling. Apache 2.0 licensed.

In 2 lists

LightLLM

Pure Python-based LLM inference and serving framework with lightweight design, easy extensibility, and high-speed performance. Integrates optimizations from FasterTransformer, TGI, vLLM, and SGLang.

In 2 lists

TabbyAPI

FastAPI-based API server for ExLlamaV2/V3 backends. OpenAI-compatible API with support for model loading/unloading, embeddings, speculative decoding, multi-LoRA, and streaming.

GPUStack

GPU cluster manager that orchestrates inference engines like vLLM and SGLang. Automated engine selection, parameter optimization, and distributed multi-GPU deployment for high-performance AI workloads.

In 5 listsDetails

OpenLLM (BentoML)

Production-grade platform for running any open-source LLMs as OpenAI-compatible API endpoints. Supports 50+ models with built-in streaming, batching, and auto-acceleration. Apache 2.0 licensed.

In 10 listsDetails

Higress (Alibaba)

AI-native API gateway born from Alibaba's internal infrastructure with 2+ years of production validation. Provides unified LLM API and MCP (Model Context Protocol) management with enterprise-grade 99.99% availability. Apache 2.0 licensed.

In 4 listsDetails

NVIDIA Dynamo

Datacenter-scale distributed inference serving framework from NVIDIA. Orchestration layer above vLLM/SGLang/TensorRT-LLM with disaggregated serving, KV-aware routing, and automatic scaling. Built in Rust with Python extensibility. Apache 2.0 licensed.

In 3 lists

Microsoft BitNet

Official inference framework for 1-bit LLMs (BitNet b1.58). Enables running large models on CPU with minimal memory footprint. Features custom kernels for ternary weight quantization and efficient matmul operations. MIT licensed.

In 5 listsDetails

FreeLLMAPI

OpenAI-compatible proxy gateway that stacks the free tiers of multiple LLM providers behind a single endpoint with automatic failover and rate tracking. MIT licensed.

OmniRoute

Unified AI gateway and proxy supporting over 230 providers with token compression, automatic failover, and routing strategies. MIT licensed.

In 3 lists

Switchyard

Rust proxy and library for routing, protocol translation, and operational metrics across LLM backends and coding agents. Apache 2.0 licensed.

In 2 lists

Bifrost

LLM gateway with a unified OpenAI-compatible API across providers, routing, load balancing, fallbacks, guardrails, and observability. Apache 2.0 licensed.

In 8 listsDetails

DeepEP

Efficient expert-parallel communication library for large MoE models, improving throughput in distributed inference and training.

In 2 lists

DeepGEMM

CUDA FP8/FMA GEMM kernels for efficient LLM inference and training at reduced precision.

AirLLM

Single-GPU 70B inference stack with strong memory/performance optimizations for local deployment on commodity hardware.

In 3 lists

ThunderKittens

High-performance GPU kernel primitives for fast attention and matmul workflows used by LLM stacks.

In 2 lists

Mirage Persistent Kernel

Compiler that fuses model execution into a single mega-kernel for tighter performance.

tt-metal

Operator and kernel toolkit for efficient LLM inference and low-level optimization on Tenstorrent hardware.

vLLM-Ascend

Hardware plugin for running vLLM on Huawei Ascend accelerators.

CTranslate2

Fast inference engine for Transformer models supporting OpenNMT and Hugging Face models. Optimized for CPU and GPU with batching, quantization (INT8/FP16), and dynamic memory management. Powers faster-whisper and other production deployments. MIT licensed.

In 3 lists

llama-swap

Intelligent model swapping proxy for llama.cpp. Enables seamless hot-swapping between different GGUF models without restarting the server, with automatic model loading/unloading and OpenAI-compatible API. MIT licensed.

In 5 listsDetails

optillm

Optimizing inference proxy for LLMs with load balancing, failover, and request routing across multiple providers and models. Improves reliability and performance for production deployments. Apache 2.0 licensed.

In 3 lists

Fugusashi

Open-source intelligent model router with CMA-ES evolved routing weights, federated learning for privacy-preserving collaborative routing, and human-interpretable explanations for every decision. OpenAI-compatible API with web dashboard. MIT licensed.

OpenRoutiQ

Explainable, policy-controlled Python router for selecting models, providers, deployments, and reasoning levels, with optional outcome learning and an OpenAI-compatible proxy. MIT licensed.

mllm

Fast and lightweight multimodal LLM inference engine for mobile and edge devices. Optimized for running vision-language models on resource-constrained hardware with efficient memory management. MIT licensed.

shimmy

Python-free Rust inference server with OpenAI API compatibility. Supports GGUF and SafeTensors formats with hot model swap, auto-discovery, and single binary deployment for zero-dependency inference. Apache 2.0 licensed.

In 7 listsDetails

PowerInfer

High-speed LLM inference for local deployment on consumer GPUs. Achieves up to 11x speedup over llama.cpp on RTX 4090 by exploiting power-law neuron activation patterns. MIT licensed.

In 2 lists

distributed-llama

Distributed LLM inference connecting home devices into a powerful cluster. More devices means faster inference via tensor parallelism over Ethernet. Supports Linux, macOS, Windows, ARM, and x86_64 AVX2 CPUs. MIT licensed.

In 3 lists

ik_llama.cpp

High-performance llama.cpp fork with better CPU and hybrid GPU/CPU performance, SOTA quantization types, first-class Bitnet support, and improved DeepSeek performance via MLA, FlashMLA, and fused MoE operations. MIT licensed.

In 3 lists

xLLM

High-performance inference engine optimized for Chinese AI accelerators (Cambricon MLU, Hygon DCU, Huawei Ascend). Features service-engine decoupled architecture with elastic scheduling, PD disaggregation, and global KV cache management. Powers JD.com's core retail businesses. Apache 2.0 licensed.

In 2 lists

Mooncake

Production-grade serving platform for Kimi (Moonshot AI). Features distributed KV cache pool with intelligent offloading, prefill/decode disaggregation, and cross-instance KV reuse. Integrated with vLLM, SGLang, and TensorRT-LLM. Apache 2.0 licensed.

In 2 lists

gemma.cpp

Lightweight, standalone C++ inference engine for Google's Gemma models. Optimized for on-device deployment with minimal dependencies and efficient memory usage. Apache 2.0 licensed.

In 3 lists

FlashInfer

Kernel library for LLM serving. High-performance CUDA kernels for attention, sampling, and matrix multiplication. Powers vLLM, SGLang, and other inference engines with optimized GPU kernels. Apache 2.0 licensed.

In 2 lists

FlashKDA

High-performance CUTLASS-based Kimi Delta Attention CUDA kernels for linear attention mechanisms. MIT licensed.

RAFT

CUDA-accelerated algorithms and ANN building blocks for high-performance similarity search, clustering, and matrix learning workloads.

mini-sglang

Compact implementation of SGLang designed to demystify modern LLM serving systems. Educational yet production-quality with RadixAttention, continuous batching, and speculative decoding. MIT licensed.

In 3 lists

bitsandbytes

8-bit and 4-bit optimizers + quantization.

In 4 listsDetails

Optimum

Hardware-specific acceleration and quantization.

In 2 lists

4. Agentic AI & Multi-Agent Systems

AutoGPT

The original autonomous AI agent framework that sparked the agent revolution. Vision of accessible AI for everyone with modular agent architecture, benchmark testing, and forge-based agent building. 183k+ stars.

In 10 listsDetails

LangGraph

Stateful, controllable agent orchestration.

In 8 listsDetails

CrewAI

Role-based agent framework.

In 10 listsDetails

AutoGen (AG2)

Flexible multi-agent conversation framework.

In 14 listsDetails

DSPy

Framework for programming language model pipelines with modules, optimizers, and evaluation loops.

In 11 listsDetails

Semantic Kernel

SDK for building and orchestrating AI agents and workflows across multiple programming languages.

In 11 listsDetails

smolagents

Lightweight agent framework centered on tool use and code-executing workflows.

In 6 listsDetails

LangChain

Foundational library for agents, chains, and memory.

In 20 listsDetails

Neuron AI

PHP Agentic Framework for building production-ready AI driven applications. Connect components (LLMs, vector DBs, memory) to agents that can interact with your data. MIT licensed.

II-Agent (Intelligent Internet)

New open-source framework to build and deploy intelligent agents with support for Claude, Gemini, and OpenAI models. Apache 2.0 licensed.

In 2 lists

Hermes Agent (NousResearch)

The agent that grows with you. Autonomous server-side agent with persistent memory that learns and improves over time.

In 7 listsDetails

Strands Agents

Model-driven approach to building AI agents in just a few lines of code. Multi-agent systems, autonomous agents, and streaming support with built-in MCP. Apache 2.0 licensed.

In 4 listsDetails

Agno

Build, run, and manage agentic software at scale. High-performance framework for multi-agent systems with memory, knowledge, and tools.

In 7 listsDetails

Upsonic

Agent framework for fintech and banking with built-in MCP support, guardrails, and tool server architecture.

In 5 listsDetails

VoltAgent

TypeScript-first AI agent engineering platform with memory, RAG, workflows, MCP integration, and voice support.

In 5 listsDetails

PocketFlow

100-line minimalist LLM framework for building agent workflows. Lightweight, extensible architecture for tool use and autonomous task execution.

In 2 lists

Agent Development Kit (Google)

Code-first Python toolkit for building sophisticated AI agents with multi-agent orchestration, built-in evaluation, and flexible deployment. Model-agnostic with tight Google ecosystem integration. Apache 2.0 licensed.

In 5 listsDetails

PydanticAI

Type-safe AI agent framework from the creators of Pydantic. Model-agnostic with 20+ providers, built-in observability via Logfire, MCP/A2A protocol support, and YAML/JSON agent definitions. MIT licensed.

In 12 listsDetails

Griptape

Modular Python framework for AI agents and workflows with chain-of-thought reasoning, tools, and memory. Enforces structures like sequential pipelines and DAG-based workflows for predictable AI systems. Apache 2.0 licensed.

In 3 lists

Langroid

Harness LLMs with multi-agent programming. Mature tool calling system based on Pydantic, supports hundreds of LLM providers including OpenAI and local servers. Built for robust agent behavior in real-world use cases. MIT licensed.

In 3 lists

Octomind

Model-agnostic AI agent runtime written in Rust with specialist agents, MCP support, multiple provider integrations, and zero-config setup. Apache 2.0 licensed.

In 3 lists

Marvin

Python framework for structured outputs and agentic AI workflows. Simplifies LLM interactions with type-safe interfaces, automatic schema generation, and built-in observability. From the creators of Prefect. Apache 2.0 licensed.

In 6 listsDetails

Burr

Apache incubating framework for building stateful AI applications (chatbots, agents, simulations). Monitor, trace, persist, and execute on your own infrastructure with built-in UI and pluggable memory. Apache 2.0 licensed.

In 2 lists

KaibanJS

JavaScript-native framework for building and managing multi-agent systems with a Kanban-inspired approach. Visual task board for AI agents with real-time collaboration features. MIT licensed.

In 2 lists

Jido

Autonomous agent framework for Elixir. Built for distributed, autonomous behavior and dynamic workflows with actor-model concurrency. Apache 2.0 licensed.

In 2 lists

Flue

Programmable TypeScript harness and sandbox agent framework for building autonomous workflows and agents. Apache 2.0 licensed.

In 3 lists

Agent-Native

TypeScript-first framework for building agent-first applications featuring shared database state, real-time multiplayer editing, and action-driven tools. ISC licensed.

rlm

General plug-and-play inference library for Recursive Language Models (RLMs) that programmatically execute sub-LM calls inside isolated code sandboxes.

Page Agent

JavaScript-native, in-page GUI agent framework for controlling web interfaces with natural language without screenshots or external browser automation. MIT licensed.

In 2 lists

Computer (Cloudflare)

Virtual filesystem inside a Durable Object providing sandboxed execution environments for AI agents. MIT licensed.

In 2 lists

Embabel

Agent framework for the JVM written in Kotlin with dynamic goal-oriented planning and Spring Boot integration. Apache 2.0 licensed.

Ouroboros

Self-hosted general-purpose agent with durable identity and memory, specialist subagent coordination, and reviewed changes to its own implementation. MIT licensed.

In 2 lists

ChatDev

Multi-agent software development framework where AI agents collaborate as programmers, designers, and testers to build software. Apache 2.0 licensed.

In 6 listsDetails

CAMEL

First and best multi-agent framework for building scalable agent systems. Apache 2.0 licensed with extensive tooling for agent communication and task automation.

In 7 listsDetails

DeepAgents

Batteries-included LangChain agent harness for building and running structured multi-agent workflows with reusable runtime patterns.

In 7 listsDetails

EvoFlux

Local-first desktop workspace for coordinating AI agent teams across conversations, files, terminal, browser, Git, memory, plugins, and verification.

In 2 lists

Swarms

Bleeding-edge enterprise multi-agent orchestration.

In 3 lists

Mastra

TypeScript-first agent framework with built-in RAG, workflows, tool integrations, observability and observational memory.

In 10 listsDetails

Nika

Workflow engine for AI where agent work is captured as reviewable .nika.yaml DAG files, statically checked before execution (schema, permits, cost floor) with tamper-evident traces after. Local-first (Ollama, llama.cpp, vLLM), MCP client and server. AGPL-3.0 licensed.

In 6 listsDetails

Deer-Flow (ByteDance)

Open-source long-horizon SuperAgent harness that researches, codes, and creates. Handles tasks from minutes to hours with sandboxes, memories, tools, skills, subagents, and message gateway.

In 7 listsDetails

OpenAI Agents SDK

Production-ready lightweight framework for multi-agent workflows. The evolution of Swarm with enhanced orchestration capabilities and enterprise-grade features.

In 9 listsDetails

Symphony

Turns project work into isolated, autonomous implementation runs. Monitors work boards, spawns agents to handle tasks, and provides proof of work including CI status, PR reviews, and walkthrough videos. Engineering preview for managing work instead of supervising coding agents. Apache 2.0 licensed.

In 3 lists

Paperclip

AI agent company and orchestration framework with 55K+ stars. MIT licensed.

In 3 lists

AgentScope

Alibaba's production-ready multi-agent framework with 23K+ stars. Features built-in MCP and A2A support, message hub for flexible orchestration, and AgentScope Runtime for production deployment.

In 4 listsDetails

Microsoft Agent Framework

Microsoft's official framework combining AutoGen's agent abstractions with Semantic Kernel's enterprise features. Supports Python and .NET with graph-based workflows.

In 4 listsDetails

NarraNexus

A ready-to-run AI agent team workspace whose agents remember, collaborate, and use tools from day one. Apache 2.0 licensed.

In 3 lists

Agency Swarm

Reliable multi-agent orchestration framework built on top of the OpenAI Assistants API with organizational structure modeling.

In 3 lists

elizaOS

Autonomous multi-agent framework for building and deploying AI-powered applications. Features Discord/Telegram/Farcaster connectors, RAG support, and a modern web dashboard.

In 2 lists

OpenAgents

AI Agent Networks for Open Collaboration. Platform for building collaborative multi-agent systems with shared knowledge and distributed task execution. Apache 2.0 licensed.

In 3 lists

Hive (Aden)

Production-grade multi-agent orchestration framework with 10K+ stars. Apache 2.0 licensed.

In 5 listsDetails

Agent Squad (AWS Labs)

Flexible multi-agent orchestration framework with intelligent intent classification and context management. Supports Python and TypeScript with pre-built agents for Bedrock, Lex, and custom integrations. Apache 2.0 licensed.

In 2 lists

DeepResearchAgent

Hierarchical multi-agent system for deep research tasks with automated task decomposition and execution across complex domains.

Composio Agent Orchestrator

Agentic orchestrator for parallel coding agents. Plans tasks, spawns agents, and autonomously handles CI fixes, merge conflicts, and code reviews. MIT licensed.

In 2 lists

Open Multi-Agent

MIT-licensed TypeScript-native orchestration framework that plans multi-agent task DAGs at runtime, with approval gates, tracing, evaluation, checkpoints, and resumable execution in your own environment.

In 4 listsDetails

BeeAI Framework (IBM)

Production-ready multi-agent framework in Python and TypeScript. Features workflow orchestration, ACP/MCP protocol support, and deep watsonx integration. Part of Linux Foundation AI & Data program.

AI Town

Deployable starter kit for building virtual towns where AI characters live, chat and socialize. Inspired by Stanford's Generative Agents research with persistent agent memory and social interactions. MIT licensed.

In 4 listsDetails

Conductor OSS

Event-driven agentic orchestration platform providing durable and resilient execution engine for applications and AI agents. Battle-tested at Netflix, Tesla, LinkedIn, and J.P. Morgan with 30K+ stars. Apache 2.0 licensed.

In 4 listsDetails

A2A Protocol

Agent2Agent (A2A) open protocol enabling communication and interoperability between opaque agentic applications. Donated to Linux Foundation by Google with 50+ technology partners. Apache 2.0 licensed.

In 2 lists

777genius/agent-teams-ai

Open-source desktop app for coordinating autonomous coding-agent teams with inter-agent messaging, Kanban task management, and code review across Claude Code, Codex, OpenCode, Cursor, Grok, GitHub Copilot, Kiro, Z.AI, MiniMax, Kimi, and 75+ LLM providers.

In 3 lists

Panniantong/Agent-Reach

Reusable search and web-ingestion layer for AI agents spanning Reddit, X/Twitter, YouTube, GitHub, Bilibili and more through one CLI.

In 3 lists

xerrors/Yuxi

Self-hosted multi-tenant agent harness combining retrieval, knowledge graph grounding, and workflow orchestration for production teams.

Sim Studio

Open-source AI workspace for building, deploying, and orchestrating AI agents. Visual canvas with 1000+ integrations, multi-framework support (Agno, OpenAI, LangChain, Google ADK), and self-hosted or cloud deployment. Apache 2.0 licensed.

In 4 listsDetails

2FastLabs Agent Squad

Flexible, lightweight open-source framework for orchestrating multiple AI agents to handle complex conversations with parallel execution capabilities. Apache 2.0 licensed.

In 2 lists

SIA

Self-improving framework that orchestrates meta, target, and feedback agents to autonomously optimize the performance of AI models and agents on benchmark tasks. MIT licensed.

Council of High Intelligence

Multi-agent deliberation framework that routes specialized personas across different LLM providers to debate topics and reach consensus.

Gas Town

Multi-agent workspace manager and orchestration system for Claude Code and other coding agents with persistent work tracking, mailboxes, and automated merge queues. MIT licensed.

In 3 lists

Agentlas OS Agent Operation Environment (AOE)

Local-first, model-agnostic environment for building, borrowing, and orchestrating specialist agents and teams across supported LLM hosts, with owner-scoped packages and host-enforced permissions. Apache 2.0 licensed.

In 3 lists

Open Deep Research

Open-source deep research assistant that orchestrates language models and search tools to run multi-step web research and compile reports. MIT licensed.

In 3 lists

MiroFish

Multi-agent swarm intelligence engine that constructs parallel digital environments with autonomous agents to simulate and predict real-world outcomes. AGPL-3.0 licensed.

In 2 lists

Maka

Apache-incubating local-first AI-agent workspace with durable append-only agent execution history, licensed under Apache-2.0.

Ruflo

Multi-agent orchestration and meta-harness for coordinating agent swarms with shared memory, licensed under the MIT license.

In 3 lists

munder-difflin

MIT-licensed desktop multi-agent workspace and harness for coordinating local coding agents across multiple LLM backends.

fractal

Hierarchical coding-agent orchestrator with recursive delegation, per-node Git worktrees, configurable limits, persistent SQLite state, and live terminal monitoring and steering.

In 3 lists

AX (Google)

Declarative, high-throughput agent orchestration runtime for sandboxed agent workloads at cluster scale. Apache 2.0 licensed.

In 2 lists

Strands Agent Harness

Production SDK and harness runtime for building, monitoring, and controlling end-to-end AI agent lifecycles in Python and TypeScript. Apache 2.0 licensed.

In 4 listsDetails

StarNet

Local-first desktop multi-agent harness with persistent workspaces, capability-scoped tools, agent memory, budgets, schedules, and live runtime visualization. MIT licensed.

Raven

Multi-agent harness whose host agent plans complex tasks as DAGs and delegates them to built-in research, coding, design, and on-call agents or to third-party agents over ACP, CLI, or OpenAI-compatible APIs, with cross-session memory and an experimental self-evolution loop that installs only…

In 2 lists

Agent Gateway

Next-generation proxy and routing layer for AI agents and MCP servers, with Kubernetes-native transport, protocol interoperability, and service-mesh-style isolation for reliable agent infrastructure. Apache 2.0 licensed.

In 4 listsDetails

Agent Governance Toolkit

Policy, safety, and execution controls for autonomous AI agents, including governance guardrails, sandboxing, and reliability checks. Apache 2.0 licensed.

In 5 listsDetails

Microsandbox

Fast, local-first microVM runtime and library for isolated execution of AI agents and untrusted workloads. Apache 2.0 licensed.

In 3 lists

DESIGN.md (Google)

A format specification for describing visual identity to coding agents, combining YAML tokens and markdown prose to give agents a structured understanding of design systems. Apache 2.0 licensed.

In 4 listsDetails

Agent Skills

Standardized specification and document format for bundling and progressively loading AI agent capabilities, instructions, scripts, and resources. Apache 2.0 licensed.

In 3 lists

FastMCP

A Python framework for building Model Context Protocol (MCP) servers and clients with automatic schema generation and validation. Apache 2.0 licensed.

In 2 lists

OpenViking

Open-source context database for AI agents that unifies agent memory, knowledge RAG, and skills as a virtual filesystem. AGPL 3.0 licensed.

In 7 listsDetails

LWC (Local Wiki CLI)

Local-first, source-grounded project memory for coding agents with bounded MCP retrieval, citations and provenance, atomic changesets, and optional document and code knowledge graphs. Apache 2.0 licensed.

In 5 listsDetails

Obsidian Agent Skills

Agent skills and open-format tooling for Obsidian vaults, Markdown, Bases, and JSON Canvas compatible with Claude Code, Codex, and OpenCode. MIT licensed.

In 2 lists

claude-obsidian

MIT-licensed Obsidian and Claude Code integration that organizes Markdown vaults with local indexing and links.

In 3 lists

Cangjie Skill

Pipeline to distill books, videos, and podcasts into structured, executable agent skills using verification and testing workflows. MIT licensed.

In 2 lists

Hexis

Git-backed platform for managing and sharing skills, tools, and context across AI agents through a remote MCP server. Apache 2.0 licensed.

In 2 lists

Codegraph

Local pre-indexed code knowledge graph for coding agents to reduce token usage and redundant tool calls across Claude, Codex, and other agents.

In 3 lists

Wenlan

Local source-backed AI knowledge base and LLM wiki for coding agents, with Markdown pages, hybrid retrieval, knowledge graphs, CLI workflows, and MCP integrations. Apache 2.0 licensed.

Gortex

Local-first code knowledge graph and code intelligence engine built in Golang with multi-repository support and real-time graph actualisation. Built for AI coding agents aiming to expose only the needed information, reducing token usage by up to 50x. Works natively with Claude, Codex, Hermes, and…

Graphify

AI coding assistant skill that indexes codebases, databases, and documents into a queryable knowledge graph for coding agents. MIT licensed.

In 3 lists

Workspai

Open-source workspace intelligence CLI that models polyglot workspaces, builds proof-backed knowledge graphs, runs health and release gates, and generates bounded context for AI coding agents. MIT licensed.

Headroom

Context compression proxy for tool outputs, logs, and RAG chunks, reducing token pressure while preserving intent for AI agents.

In 4 listsDetails

llmtrim

Self-hosted Rust proxy, MCP server, CLI, and library that compresses LLM prompts, tool outputs, and replies to cut token usage, quality-gated so it never raises your bill. MPL-2.0 licensed.

In 3 lists

MailFathom

Security-first, self-hosted email knowledge system that synchronizes IMAP mailboxes into user-owned PostgreSQL for lexical and semantic retrieval, cited answers, local-model operation, and MCP access. AGPL-3.0 licensed.

In 2 lists

MemPalace

High-performance, benchmarked AI memory system for persistent recall and retrieval in long-horizon autonomous workflows.

In 3 lists

Supermemory

Memory engine and API designed for long-lived AI agents to store, retrieve, and reuse long-horizon context with low latency.

In 3 lists

Tree Ring Memory

Framework-agnostic, local-first memory lifecycle for AI agents with a Rust CLI, SQLite/FTS recall, redaction, forgetting, audit checks, consolidation, and agent-skill guidance. MIT licensed.

In 3 lists

deja

Memory layer over the session transcripts coding agents already write to disk — Claude Code, Codex, Cursor, opencode, Zed and more — including sessions from before it was installed; local BM25 recall over them with no LLM or embeddings, MCP tools, and credentials redacted at index time. MIT…

In 5 listsDetails

codebase-memory-mcp

High-performance C-based codebase intelligence engine and MCP server that indexes repositories into local type-resolved knowledge graphs. MIT licensed.

In 5 listsDetails

book-to-skill

CLI tool that distills technical books, documents, and reference materials into structured, on-demand agent skills. MIT licensed.

In 3 lists

AI Memory

Rust-native long-term memory server and MCP client for agent coding CLIs, featuring Karpathy-style LLM wiki compilation, FTS5 recall, and cross-agent session handoffs. MIT licensed.

In 2 lists

Busabase

Open-source database and workspace that gives AI agents structured data, durable knowledge, docs, skills, and apps through MCP, OpenAPI, CLI, and coding-agent skills; material writes can remain reviewable ChangeRequests before becoming canonical. MIT licensed.

ThreadShelf

Local-first archive and semantic search for AI conversation histories across multiple providers, with local embeddings, LanceDB storage, complete thread retrieval, HTTP and CLI access, and a stdio MCP server.

BitFun

Open-source coding agent with a Rust runtime, desktop and CLI interfaces, self-hosted remote access, and support for custom tools and skills.

In 4 listsDetails

Free Claude Code

Multi-provider proxy and launcher for Claude Code, Codex, and Pi with a local Admin UI to route coding agents to 31+ cloud and local LLM backends. MIT licensed.

In 3 lists

Background Agents

Open-source background coding agent system inspired by Ramp's Inspect, supporting file and environment snapshots, cron-based automation, and multi-provider models. MIT licensed.

In 2 lists

OpenHands (ex-OpenDevin)

Full-featured open-source AI software engineer.

In 7 listsDetails

HEXStrike AI

MCP-powered coding-focused cybersecurity agent framework for automated pentesting and bug-hunting workflows.

In 3 lists

Darkmoon

Open-source autonomous AI pentest platform and MCP host orchestrating 80+ offensive tools via per-tech offensive sub-agents, with Active Directory and Kubernetes support. GPL-3.0 licensed.

In 10 listsDetails

VulnClaw

Autonomous penetration testing agent utilizing Model Context Protocol (MCP) toolchains, blackboard state space search, and structured reasoning.

Goose

Extensible on-machine AI agent for development tasks.

In 3 lists

OpenShell (NVIDIA)

Safe and private runtime for autonomous AI agents with policy-driven execution boundaries and CLI integration.

In 4 listsDetails

OpenCode

Terminal-native autonomous coding agent.

In 5 listsDetails

ECC

Performance-oriented agent harness for coding agents with skills, memory, and security-aware orchestration across Claude Code, Codex, and more.

In 7 listsDetails

oh-my-pi

Terminal coding agent with hash-anchored edits, subagents, LSP, browser integrations, and terminal-native workflows.

In 2 lists

Pi (earendil-works)

Modular agent toolkit with terminal-first CLI, unified model/provider layer, and runtime integrations for coding workflows and TUI/web UIs.

In 4 listsDetails

Aider

Command-line pair-programming agent.

In 8 listsDetails

Pi (badlogic)

Terminal coding agent with hash-anchored edits, LSP integration, subagents, MCP support, and package ecosystem.

Mistral-Vibe (Mistral)

Minimal CLI coding agent by Mistral. Lightweight, fast, and designed for local development workflows.

In 2 lists

Nanocoder (Nano-Collective)

Beautiful local-first coding agent running in your terminal. Built for privacy and control with support for multiple AI providers via OpenRouter.

In 3 lists

Gemini CLI (Google)

Open-source AI agent that brings Gemini's power directly into your terminal. Supports code generation, shell execution, and file editing with full Apache 2.0 licensing.

In 9 listsDetails

Archon

Workflow engine for deterministic AI coding agents. Define development processes as YAML workflows (planning → implementation → validation → review → PR) with isolated Git worktrees for parallel execution. MIT licensed.

In 3 lists

mini-SWE-agent

Lightweight coding agent for repository and issue-fixing workflows, designed for simple agentic software engineering experiments.

Kilo Code

Open-source agentic coding assistant with IDE workflows, tool use, and support for local or OpenAI-compatible models.

In 4 listsDetails

Open SWE

Asynchronous coding agent from the LangChain ecosystem for background software engineering tasks.

In 4 listsDetails

Letta Code

Memory-first coding harness designed for long-lived agents that learn from experience. Persistent agents with portable memory across models (Claude, GPT, Gemini, GLM, Kimi). CLI and desktop app for macOS, Windows, and Linux. Apache 2.0 licensed.

LoopTroop

Local-first AI coding workspace that orchestrates multi-model planning councils, git worktrees, and task loops. MIT licensed.

In 4 listsDetails

gptme

Your agent in your terminal, equipped with local tools: writes code, uses the terminal, browses the web. Make your own persistent autonomous agent on top. MIT licensed.

In 3 lists

Superpowers

Composable skills framework and software development methodology for coding agents, structuring processes like planning, test-driven development, and code review.

In 7 listsDetails

Agent Skills

Production-grade engineering skills and quality gates for AI coding agents, packaging developer workflows like spec refinement, planning, and testing.

In 3 lists

Harness

Team-architecture factory for C‍laude Code that designs domain-specific agent teams, defines specialized agents, and generates their skills. Apache 2.0 licensed.

In 4 listsDetails

jcode

Performance-oriented, memory-efficient coding agent harness built for multi-session workflows and infinite customizability. MIT licensed.

In 3 lists

firstmate

Agent distro for running a crew of autonomous coding agents in isolated Git worktrees across visible terminal session backends. MIT licensed.

DeepSeek-Reasonix

DeepSeek-native AI coding agent for the terminal, designed around prefix-cache stability. MIT licensed.

In 3 lists

gstack

Multi-specialist agent skills framework for Claude Code and coding agents that structures development workflows into planning, design, QA, and release phases. MIT licensed.

In 3 lists

taste-skill

Anti-slop frontend design skills for coding agents that improve layout, typography, and design system alignment during generation. MIT licensed.

In 5 listsDetails

Agents CLI (Google)

CLI and skills that turn coding assistants into experts at creating, evaluating, and deploying AI agents on Google Cloud. Apache 2.0 licensed.

In 2 lists

Claude Code Skills & Plugins

Modular instruction packages, custom commands, and utility scripts for Claude Code, Gemini CLI, Cursor, and other AI coding agents. MIT licensed.

In 5 listsDetails

AG Kit

Antigravity-first agent engineering kit featuring rules, skills, persistent memory, MCP guidance, and a native safety hook. MIT licensed.

Prime Agent

Self-improving recursive language model (RLM) agent for coding workflows and autonomous tasks. MIT licensed.

Agent Skills (Google)

Official collection of Agent Skills for Google Cloud and Google developer platforms, extending AI coding agents with product and technology workflows. Apache 2.0 licensed.

In 3 lists

Agent Skills (Anthropic)

Official collection of Agent Skills and reference implementations for Claude Code, Claude API, and AI agents. Apache 2.0 licensed.

In 9 listsDetails

Superagent

macOS desktop app giving Claude Code and Codex a real browser to drive, an iOS Simulator to install and screenshot apps in, and a phone companion app for remote monitoring. MIT licensed.

In 3 lists

BrowserSkill (Tencent)

CLI and extension enabling AI coding agents to control authenticated browser sessions in background tabs without interrupting user workflows. MIT licensed.

Ordewell

Terminal CLI and TUI that turns one goal into an ordered, editable plan of coding agent tasks, each carrying its own runner (Claude Code, Codex, OpenCode), model and mode, where a task counts as done only when its completion marker appears in that runner's output. Apache-2.0 licensed.

In 3 lists

Orbi

Self-hosted runner that turns a labelled GitHub Issue into a pull request, has a separate model review the frozen diff before merge, and publishes the tagged GitHub Release. Works with any OpenAI-compatible model or a Codex subscription. AGPL-3.0 licensed.

Outlines

Structured outputs for LLMs. Guarantees valid JSON, regex-compliant text, and Pydantic model outputs during generation. Trusted by NVIDIA, Cohere, Hugging Face, and vLLM. Apache 2.0 licensed.

In 6 listsDetails

LangGPT

Pioneering framework for structured and meta-prompt design. Battle-tested by thousands of users worldwide with 10,000+ stars. The most popular prompt engineering paradigm for creating reusable, maintainable prompt templates. Apache 2.0 licensed.

In 3 lists

Prompt Optimizer

AI prompt optimization tool with multi-round iterative improvements, dual-mode optimization for system and user prompts, and multi-model support. Available as web app, desktop app, Chrome extension, and Docker deployment. AGPL-3.0 licensed.

Guidance

Efficient programming paradigm for steering language models. Control output structure with loops, conditionals, and regex constraints inline. Reduces latency and cost vs conventional prompting. MIT licensed.

In 4 listsDetails

XGrammar

Fast, flexible and portable structured generation engine. Default backend for vLLM, SGLang, TensorRT-LLM, and MLC-LLM with flexible grammar support and zero-overhead mask generation. Apache 2.0 licensed.

LM Format Enforcer

Enforce output format (JSON Schema, Regex, etc) of language models by filtering allowed tokens at each generation step. Compatible with Hugging Face, llama-cpp-python, and vLLM. MIT licensed.

AdalFlow

Library to build and auto-optimize LLM applications with LLM-AutoDiff for fine-tuning-free optimization. End-to-end workflow optimization with tracing and human-in-the-loop capabilities. MIT licensed.

In 2 lists

Composio

Tool integration layer for AI agents with 1000+ toolkits, authentication management, and sandboxed workbench. Powers tool use across major frameworks.

In 3 lists

Langflow

Visual low-code platform for agentic workflows.

In 9 listsDetails

Dify

Production-ready agentic workflow platform.

In 14 listsDetails

OWL (camel-ai/owl)

Advanced multi-agent collaboration system.

In 3 lists

gpt-researcher

Autonomous agent that conducts deep online research on any topic. Generates comprehensive reports with citations by orchestrating web searches, content scraping, and synthesis. Apache 2.0 licensed.

In 9 listsDetails

last30days-skill

AI agent that researches and synthesizes topic trends from Reddit, X, YouTube, Hacker News, and prediction markets. MIT licensed.

In 3 lists

Jev Social

Local social-media research agent that uses TypeSafe Jev to select bounded socai CLI browser operations for Instagram, TikTok, and LinkedIn, streams captured evidence, and generates cited reports. MIT licensed.

In 8 listsDetails

pm-skills

Marketplace of product management plugin skills and workflows for Claude Code, Codex, and Claude Cowork. MIT licensed.

In 3 lists

claude-code-best-practice

Custom subagents, skills, environment configurations, and best practices for the Claude Code developer agent.

PPT Master

AI-driven multi-role agent skill for converting documents into editable PowerPoint presentations via SVG and DrawingML workflows. MIT licensed.

In 3 lists

PraisonAI

24/7 AI employee team for automating complex challenges. Low-code multi-agent framework with handoffs, guardrails, memory, RAG, and 100+ LLM providers.

In 9 listsDetails

Agent-S (Simular AI)

Open agentic framework that uses computers like a human. SOTA on OSWorld benchmark (72.6%) for GUI automation and computer control.

In 6 listsDetails

MobileAgent (Alibaba/X-PLUG)

Powerful GUI agent family for autonomous mobile device control. Multimodal agent framework designed to operate smartphone apps through visual UI perception and reasoning. MIT licensed.

In 4 listsDetails

UI-TARS Desktop (ByteDance)

Open-source multimodal AI agent stack with native GUI agent capabilities. Desktop application bringing GUI agent and vision power to your computer, browser, and terminal. Apache 2.0 licensed.

Browser Use

Makes websites accessible for AI agents. Enables autonomous web automation, data extraction, and task completion with natural language instructions. MIT licensed.

In 10 listsDetails

Steel Browser

Open-source browser API for AI agents and apps. Batteries-included browser sandbox for web automation without infrastructure worries. Apache 2.0 licensed.

In 2 lists

Cua

Open-source sandboxes, SDKs, and benchmarks for computer-use agents to control desktop environments. MIT licensed.

In 3 lists

Webwright

A terminal-style web agent framework that enables coding models to run as browser agents by writing and executing Playwright scripts. MIT licensed.

In 2 lists

invisible_playwright_mcp

Self-hosted browser agent that drives a patched Firefox with real pointer and keyboard input, usable as a local MCP server from Claude Code, Codex, Cursor or Gemini CLI, or from its own web UI. MIT licensed.

In 2 lists

TradingAgents

Multi-agent framework for financial trading. Simulates professional trading firm operations with 6+ specialized agent roles, backtesting, risk management, and portfolio optimization. Built with LangGraph, supports multiple LLM providers.

In 4 listsDetails

Parlant

Conversational control layer for customer-facing AI agents. Enterprise-grade context engineering framework optimized for consistent, compliant, and on-brand B2C and sensitive B2B interactions. Apache 2.0 licensed.

In 3 lists

n8n

Self-hostable workflow automation platform with AI agent nodes, tool integrations, and production automation workflows.

In 9 listsDetails

Activepieces

Open-source automation platform with AI agents, MCP integrations, and self-hosted workflow orchestration.

In 5 listsDetails

Julep

Stateful agent workflow platform with memory, tools, branching, and long-running task execution.

In 2 lists

uAgents (Fetch.ai)

Fast and lightweight framework for creating decentralized agents with ease. Agents automatically join the network by registering on the Almanac smart contract. Supports agent-to-agent communication out of the box. Apache 2.0 licensed.

In 3 lists

Tracecat

Self-hostable security automation platform for building agentic workflows across alerts, cases, and operations.

PentAGI

Fully autonomous AI agent system for conducting ethical hacking and complex penetration testing tasks. MIT licensed.

In 5 listsDetails

ToolJet

Self-hostable internal app builder with AI app and agent workflows for operations teams.

In 6 listsDetails

The Agency

Extensive library of specialized developer and workflow agent personas with installer support for Claude Code, Cursor, Codex, and other coding assistants.

CubeSandbox (Tencent Cloud)

High-performance, secure agent sandbox built on RustVMM and KVM, compatible with the E2B SDK. Apache 2.0 licensed.

In 4 listsDetails

DeepTutor

Multi-agent system for lifelong personalized tutoring with interactive and adaptive learning workflows.

In 3 lists

Corsair

Domain-specific AI agent framework for building and running specialized agents.

DeskcommCRM

Self-hosted sales CRM that uses autonomous AI agents, WhatsApp integration, tenant RAG, and MCP tooling for chat-based sales workflows.

OpenResearch

Local-first workspace and harness for turning coding agents into autonomous research agents for literature review, hypothesis generation, and experiments.

In 2 lists

Letta (ex-MemGPT)

Platform for building stateful agents with advanced memory that learn and self-improve over time.

In 4 listsDetails

Mem0

Universal memory layer for AI agents. Persistent, multi-session memory across models and environments.

In 13 listsDetails

Caura

Open-source memory infrastructure for AI agents, providing persistent memory and context management for building stateful, long-running agents.

Rohitg (agentmemory)

Open-source memory service for agents with benchmarked retrieval, structured entities, and API/SDK support for persistent personalized memory in tooling workflows.

In 4 listsDetails

Forgetful

MCP server for persistent AI agent memory with atomic notes, semantic linking, and SQLite or PostgreSQL storage.

Hindsight

State-of-the-art long-term memory for AI agents by Vectorize. Fully self-hosted, MIT-licensed, with integrations for LangChain, CrewAI, LlamaIndex, Vercel AI SDK, and more.

In 4 listsDetails

TencentDB Agent Memory

Fully local long-term memory layer for AI agents with a four-tier progressive pipeline and zero external dependencies. MIT licensed.

In 3 lists

Cognee

AI memory platform for agents that builds a self-hosted knowledge graph for persistent, long-term memory across sessions. Apache 2.0 licensed.

In 6 listsDetails

LoopX

Lightweight state kernel and local control plane for long-running AI agent loops with goal tracking, auto-wake, and verifiable handoffs. MIT licensed.

In 3 lists

Unlimited Context

Local-first context and memory engine that keeps a billion-token encoded pool on disk and pages relevant slices into a small resident window for Ollama and other local LLMs. Apache 2.0 licensed.

inspeximus

Memory layer that retires a corrected fact by key and can undo the correction later from the key alone. Zero-dependency core with an MCP server.

Mnemoverse

MIT-licensed MCP server for agent memory, backed by a hosted engine that re-ranks recall from reported outcomes rather than similarity alone, with one key or OAuth across Claude Code, Cursor, VS Code, and ChatGPT.

In 2 lists

5. Retrieval-Augmented Generation (RAG) & Knowledge

Chroma

Most popular open-source embedding database.

In 8 listsDetails

Qdrant

High-performance vector search engine in Rust.

In 8 listsDetails

Weaviate

GraphQL-native vector search engine.

In 8 listsDetails

Milvus

Scalable cloud-native vector database.

In 14 listsDetails

NornicDB

Low-latency graph and vector hybrid retrieval database in Go with Neo4j and Qdrant-compatible drivers.

In 2 lists

Faiss

Similarity search and clustering library for dense vectors with CPU and GPU implementations.

In 8 listsDetails

LanceDB

Serverless vector DB optimized for multimodal data.

In 4 listsDetails

Vespa

AI + Data platform with hybrid search (vector + keyword) and real-time indexing at scale. Battle-tested serving billions of queries daily.

In 3 lists

pgvector

PostgreSQL extension for vector similarity search.

In 6 listsDetails

pgvectorscale

PostgreSQL extension for scalable vector search with DiskANN algorithm. Complements pgvector with significantly faster search and higher recall at large scale. PostgreSQL licensed.

VectorChord

Scalable, fast, and disk-friendly vector search in Postgres. Successor to pgvecto.rs with production-grade performance and efficient storage. AGPL-3.0 licensed.

In 2 lists

Quickwit

Cloud-native search engine for observability. Open-source alternative to Datadog, Elasticsearch, Loki, and Tempo with native vector search support.

In 3 lists

Tantivy

Full-text search engine library inspired by Apache Lucene and written in Rust. Powers Quickwit and other production search systems.

In 3 lists

Manticore Search

Easy to use open source fast database for search. Good alternative to Elasticsearch with SQL-like interface and vector search capabilities.

In 4 listsDetails

OpenSearch

Open-source distributed and RESTful search and analytics suite with native vector search. Enterprise-grade fork of Elasticsearch with k-NN plugin for semantic search at scale.

In 2 lists

Marqo

Multimodal vector search for text, image, and structured data. End-to-end indexing and search with built-in embedding models. Apache 2.0 licensed.

In 5 listsDetails

Vald

Highly scalable distributed vector search engine. Cloud-native architecture with automatic indexing, horizontal scaling, and multiple ANN algorithm support. Apache 2.0 licensed.

In 3 lists

hnswlib

Header-only C++ library for fast approximate nearest neighbors with Python bindings. Supports CRUD operations and concurrent read/write - unique among ANN libraries. Powers many production vector databases. Apache 2.0 licensed.

In 4 listsDetails

turbovec

Rust-native high-performance vector index with Python bindings optimized for fast ANN search and SIMD acceleration on modern CPUs. MIT licensed.

In 3 lists

sqlite-vec

A vector search SQLite extension that runs anywhere. Extremely small, "fast enough" vector search written in pure C with no dependencies. Perfect for embedded and edge deployments. MIT/Apache-2.0 dual licensed.

In 3 lists

zvec

Lightweight, lightning-fast, in-process vector database from Alibaba. Built on Proxima (Alibaba's battle-tested vector search engine) for production-grade, low-latency similarity search. Apache 2.0 licensed.

In 7 listsDetails

Meilisearch

Lightning-fast search engine API with AI-powered hybrid search. Features typo-tolerant full-text search combined with HNSW-based vector search for semantic retrieval. MIT licensed.

In 6 listsDetails

Typesense

Open source alternative to Algolia + Pinecone. Fast, typo-tolerant, in-memory fuzzy search engine with native vector search capabilities. GPL-3.0 licensed.

In 4 listsDetails

Elasticsearch

Distributed search and analytics engine with native k-NN vector search, hybrid search, and dense vector indexing. Industry-standard for full-text search now with powerful semantic search capabilities. AGPL-3.0/Elastic-2.0 dual licensed.

In 7 listsDetails

Apache Solr

Mature Lucene-based search platform with dense vector search, filtering, faceting, and hybrid retrieval patterns for production search-heavy RAG systems.

In 3 lists

RediSearch

Full-text, secondary indexing, and vector similarity search for Redis deployments. Useful when retrieval needs low-latency Redis-native search.

ParadeDB

Postgres-native search and analytics engine for full-text, faceted, and hybrid retrieval without moving data out of PostgreSQL.

In 4 listsDetails

Orama

Lightweight search engine with full-text, vector, and hybrid search for browser, server, and edge applications.

In 2 lists

HelixDB

Graph-vector database for retrieval systems that need relationship traversal alongside semantic search.

In 3 lists

USearch

Fast single-file similarity search & clustering engine for vectors. Smaller and faster than FAISS with 20+ language bindings (C++, Python, JavaScript, Rust, Java, Go, etc.) and support for custom metrics. Apache 2.0 licensed.

In 6 listsDetails

Deep Lake

AI Data Runtime for Agents with serverless PostgreSQL and multimodal datalake. Store and search vectors, images, text, videos, and more with LangChain/LlamaIndex integrations. Used by Intel, Bayer, Yale, and Oxford. Apache 2.0 licensed.

In 6 listsDetails

DiskANN (Microsoft)

Graph-structured indices for scalable, fast, fresh and filtered approximate nearest neighbor search. Handles billion-vector datasets on a single node with SSD-based indexing. MIT licensed.

SPTAG (Microsoft)

Distributed approximate nearest neighbor search library with high-quality vector index build and online serving toolkits. Powers Bing's vector search at trillion-vector scale. MIT licensed.

In 3 lists

nanoflann

C++11 header-only library for fast nearest neighbor search with KD-trees. Zero dependencies, single-file integration, and 2-3x faster than FLANN with modern C++. BSD licensed.

In 5 listsDetails

NMSLIB

Non-Metric Space Library for efficient similarity search in generic non-metric spaces. Comprehensive toolkit for evaluating k-NN methods with support for exotic distance functions. Apache 2.0 licensed.

In 4 listsDetails

Vearch

Cloud-native distributed vector database for AI-native applications. Efficient similarity search of embedding vectors with horizontal scaling and real-time indexing. Apache 2.0 licensed.

In 3 lists

JVector (DataStax)

The most advanced embedded vector search engine for Java. DiskANN-based algorithm for billion-scale vector search with efficient memory mapping. Apache 2.0 licensed.

VectorDBBench (Zilliz)

Industry-standard benchmark suite for vector databases. Test and compare performance of Milvus, Zilliz Cloud, and other vector DBs with your own datasets. MIT licensed.

BGE (FlagEmbedding)

BAAI's best-in-class embedding family.

In 4 listsDetails

FastEmbed (Qdrant)

Lightweight, fast Python library for embedding generation with ONNX Runtime. Supports text, sparse (SPLADE), and late-interaction (ColBERT) embeddings without GPU dependencies. Apache 2.0 licensed.

In 2 lists

EmbedAnything

Minimalist, highly performant multimodal embedding pipeline built in Rust. Memory-safe, modular, and production-ready for text, image, and audio embeddings with seamless vector DB integration. Apache 2.0 licensed.

In 3 lists

Text Embeddings Inference (Hugging Face)

Blazing fast inference solution for text embedding models. High-performance extraction with token-based dynamic batching, Flash Attention, and support for FlagEmbedding, E5, GTE, and more. OpenAI-compatible API with Docker deployment. Apache 2.0 licensed.

In 2 lists

MTEB

Massive Text Embedding Benchmark covering 1000+ languages and diverse tasks. The industry standard for evaluating and comparing embedding models.

In 2 lists

EmbedChain

Universal memory layer for AI agents. Simple API to create RAG applications over any dataset with support for multiple vector stores, embedding models, and LLM providers. Apache 2.0 licensed.

In 9 listsDetails

LlamaIndex

Full-featured RAG pipeline with advanced indexing.

In 14 listsDetails

Haystack

End-to-end NLP and RAG framework.

In 13 listsDetails

RAGFlow

Deep-document-understanding RAG engine.

In 7 listsDetails

GraphRAG (Microsoft)

Knowledge-graph-based RAG.

In 5 listsDetails

Docling

Document processing toolkit for turning PDFs and other files into structured data for GenAI workflows.

In 5 listsDetails

Unstructured

Best-in-class document preprocessing.

In 3 lists

MinerU

High-accuracy document parsing for LLM and RAG workflows. Converts PDFs, Word, PPTs, and images into structured Markdown/JSON with VLM+OCR dual engine.

In 5 listsDetails

Marker

Fast, accurate PDF-to-markdown converter with table extraction, equation handling, and optional LLM enhancement for RAG pipelines.

In 4 listsDetails

ColPali / ColQwen

Vision-language models for document retrieval.

LightRAG

Graph-based RAG with dual-level retrieval system. Simple and fast with comprehensive knowledge discovery (EMNLP 2025).

In 5 listsDetails

RAG-Anything

All-in-One Multimodal RAG system for seamless processing of text, images, tables, and equations. Built on LightRAG.

In 3 lists

RAGLite (Superlinear)

Python toolkit for RAG with DuckDB or PostgreSQL. Lightweight, efficient retrieval-augmented generation without heavy dependencies. MPL 2.0 licensed.

In 2 lists

GPT-RAG (Azure)

Enterprise RAG pattern for Azure OpenAI at scale. Secure, production-ready architecture using Azure Cognitive Search and Azure OpenAI LLMs for ChatGPT-style Q&A experiences. MIT licensed.

In 2 lists

LangChain4j

Java library for integrating LLMs into Java applications. Implements RAG, tool calling (including MCP support), and agents with seamless integration into enterprise Java frameworks like Spring Boot. Apache 2.0 licensed.

In 4 listsDetails

Kernel Memory (Microsoft)

Memory solution for users, teams, and applications. RAG pipelines with document ingestion, vector indexing, and natural language querying with citations. Supports multiple LLM providers and vector stores. MIT licensed.

txtai

All-in-one AI framework for semantic search, LLM orchestration and language model workflows. Embeddings database with customizable pipelines.

In 9 listsDetails

FlashRAG

Efficient toolkit for RAG research with 40+ retrieval and reranking models, 20+ benchmark datasets, and optimized evaluation pipelines (WWW 2025 Resource). MIT licensed.

In 2 lists

DocsGPT

Private AI platform for building intelligent agents and assistants with enterprise search. Features Agent Builder, deep research tools, multi-format document analysis, and multi-model support. MIT licensed.

In 10 listsDetails

llmware

Unified framework for building enterprise RAG pipelines with small, specialized models. Optimized for AI PC and local deployment with 300+ models in catalog. Apache 2.0 licensed.

In 4 listsDetails

AutoFlow

Graph RAG-based conversational knowledge base tool built on TiDB Vector and LlamaIndex. Features Perplexity-style search with built-in website crawler. Apache 2.0 licensed.

Chonkie

Lightweight document chunking library for fast, efficient RAG pipelines. Memory-safe with multiple chunking strategies (semantic, token, recursive) and direct vector DB integration. MIT licensed.

PageIndex (VectifyAI)

Vectorless, reasoning-based RAG framework using document index structure. Achieves high accuracy without vector databases through intelligent context engineering and reasoning-based retrieval. MIT licensed.

In 2 lists

Kotaemon (Cinnamon)

Open-source RAG-based tool for chatting with your documents. Hybrid RAG pipeline with full-text and vector retriever, re-ranking, and multi-modal capabilities. Clean Gradio-based UI with support for local and API-based LLMs. Apache 2.0 licensed.

In 3 lists

Reader (Jina AI)

Convert any URL to LLM-friendly input with a simple prefix (r.jina.ai). Free service that extracts article content, removes clutter, and returns clean Markdown for RAG and agentic workflows. Apache 2.0 licensed.

In 2 lists

UltraRAG (OpenBMB)

First lightweight RAG framework based on Model Context Protocol (MCP) architecture. Low-code RAG pipeline builder with comprehensive evaluation system and DeepResearch capabilities. From Tsinghua THUNLP, NEUIR, OpenBMB, and AI9stars. Apache 2.0 licensed.

In 2 lists

Semantic Router

Superfast AI decision-making layer for LLMs and agents. Uses semantic vector space to route requests using semantic meaning rather than waiting for slow LLM generations. Cuts routing time from seconds to milliseconds. MIT licensed.

Pathway

Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG. Features 350+ connectors with always-in-sync data from SharePoint, Google Drive, S3, Kafka, PostgreSQL and more. BSL 1.1 license (becomes Apache 2.0 after 4 years).

In 8 listsDetails

Infinity (AI Database)

AI-native database built for LLM applications with incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text. Powers RAGFlow's document engine. Apache 2.0 licensed.

In 9 listsDetails

PrivateGPT

Private document Q&A project for local and offline RAG workflows where data stays inside the user's environment.

In 6 listsDetails

FastGPT

Knowledge-base platform with RAG retrieval, document processing, visual AI workflows, and self-hosted deployment options.

In 4 listsDetails

MaxKB

Self-hostable knowledge-base and agent platform for document ingestion, RAG pipelines, and enterprise assistant workflows.

In 4 listsDetails

DB-GPT

Self-hosted AI data assistant for private knowledge, database-aware conversations, and data-heavy RAG workflows.

In 7 listsDetails

WeKnora

Enterprise knowledge platform that combines RAG retrieval, autonomous reasoning agents, wiki generation, and multi-source document ingestion.

In 3 lists

localGPT

Local document-chat project for private, on-device Q&A over files without sending data to external APIs.

In 7 listsDetails

SurfSense

Privacy-focused NotebookLM-style workspace for teams to search, organize, and query knowledge with self-hosted RAG.

In 5 listsDetails

PieKBS

Local-first knowledge search engine for agents with FTS5 full-text search and graph-based document expansion via citation/support/wiki links. Pure Go, no embedding required.

Morphik

Open-source multimodal RAG framework for building AI apps over private knowledge. Handles text, images, and documents with built-in embedding generation and vector search. MIT licensed.

Beever Atlas

Open-source LLM knowledge base combining Neo4j knowledge graph with Weaviate vector DB for graph-based RAG. Native MCP server, team chat ingestion (Slack, Discord, Teams, Telegram), and BYO LLM via LiteLLM. Apache 2.0 licensed.

VidXP

Local-first multimodal video indexing and semantic search toolkit with transcript extraction, embeddings, scene-aware search. Can be used as a CLI, API and MCP.

Graphiti

Build real-time temporal knowledge graphs for AI agents. Tracks how facts change over time with provenance to source data. Supports prescribed and learned ontology for evolving real-world data. Apache 2.0 licensed.

In 3 lists

Semantica

Graph-native infrastructure for context management, bi-temporal knowledge graphs, GraphRAG document chunking, and explainable AI reasoning. MIT licensed.

In 3 lists

Code-Graph-RAG

Multi-language codebase RAG framework using Tree-sitter and Memgraph knowledge graphs to query, edit, and optimize codebases with AI. MIT licensed.

In 2 lists

Crawl4AI

LLM-friendly web crawler that turns websites into clean Markdown for RAG and agentic workflows.

In 5 listsDetails

Scrapling

Adaptive web scraping and crawling framework for robust structured extraction from pages to large-scale pipelines.

In 4 listsDetails

Lightpanda

Machine-first headless browser in Zig; rendering-free and ultra-lightweight for AI agent browsing.

In 2 lists

Paperless-AI

Automated document analyzer for Paperless-ngx with RAG-powered semantic search across your document archive.

In 2 lists

Firecrawl

Web Data API for AI - search, scrape, and interact with the web at scale. Clean markdown/JSON output with proxy rotation and JS-blocking handled automatically.

In 3 lists

wigolo

Local-first web intelligence and Model Context Protocol (MCP) server for web search, crawling, and structured data extraction. AGPL-3.0 licensed.

Zoom Search

MCP search and evidence tool for AI agents with query rewriting, source zoom-in, sourced answers, and runtime metrics. MIT licensed.

In 3 lists

invisible-playwright

Playwright wrapper for a stealth-patched Firefox 150 binary. Drop-in Browser object for AI agents that need to ingest web data from sites with anti-bot guardrails (reCAPTCHA, FingerprintPro, Cloudflare). Spoofing happens in C++ source, not via JS overrides. MIT (wrapper) + MPL-2.0 (patches).

In 4 listsDetails

pdf-inspector

Fast Rust library for PDF classification and text extraction that detects scanned vs text-based documents to route RAG pipelines without OCR. MIT licensed.

In 2 lists

OpenDataLoader PDF

Accessibility-aware PDF parser and conversion pipeline for AI-ready markdown and structured data workflows.

In 3 lists

MarkItDown (Microsoft)

Python tool for converting files and office documents to Markdown. Supports PDF, PowerPoint, Word, Excel, images, audio, HTML, and more with OCR and transcription capabilities. MIT licensed.

In 9 listsDetails

LiteParse

Lightweight document parsing toolkit for AI and RAG pipelines with PDF/OCR extraction and clean preprocessing defaults.

In 2 lists

PaddleOCR

Large-scale OCR suite with detection, recognition, and layout analysis, used widely for document digitization and downstream RAG pipelines.

In 5 listsDetails

DocETL (UC Berkeley)

Agentic LLM-powered data processing and ETL system for complex document processing. Query rewriting and evaluation for unstructured data analysis with 80% higher accuracy than baselines. MIT licensed.

aisuite

Simple, unified interface to multiple Generative AI providers. Use OpenAI, Anthropic, Google, and 10+ other providers with a standardized API similar to OpenAI's. Switch between models or providers with a single line of code. MIT licensed.

In 4 listsDetails

Spring AI

Application framework for AI engineering in the Spring ecosystem. Unified API for LLMs, vector stores, and embedding models with seamless integration into Spring Boot applications. Supports RAG, tool calling, and structured outputs. Apache 2.0 licensed.

In 2 lists

Rig

Rust library for building scalable, modular LLM-powered applications. Type-safe agent framework with unified LLM interface, built-in vector store integrations, and ergonomic abstractions for production AI systems. MIT licensed.

In 2 lists

Ax

TypeScript framework for building reliable AI applications. "Official" DSPy-inspired framework for TypeScript with type-safe LLM interactions, chain-of-thought reasoning, and structured output validation. Apache 2.0 licensed.

In 2 lists

Genkit

Open-source framework for building full-stack AI-powered applications in JavaScript, Go, and Python. Built and used in production by Google's Firebase. Unified interface for integrating AI models from multiple providers with built-in RAG, tool calling, structured outputs, and developer tools.…

In 4 listsDetails

ContextGem

Effortless LLM extraction framework for documents. Powerful abstractions for building extraction workflows with automated dynamic prompts, data modeling, validation, and precise reference mapping. Apache 2.0 licensed.

In 2 lists

Eino

The ultimate LLM/AI application development framework in Go. Drawing from LangChain and Google ADK, designed to follow Go conventions with composable components for chains, agents, and workflows. Apache 2.0 licensed.

In 3 lists

ruby_llm

One beautiful Ruby API for OpenAI, Anthropic, Gemini, Bedrock, Azure, OpenRouter, DeepSeek, Ollama, and 15+ providers. Agents, Chat, Vision, Audio, PDF, Images, Embeddings, Tools, Streaming and Rails integration. MIT licensed.

In 4 listsDetails

LangChain.rb

Build LLM-powered applications in Ruby. Idiomatic Ruby library for building AI applications with support for multiple LLM providers, vector stores, and RAG pipelines. MIT licensed.

In 3 lists

6. Generative Media Tools

ComfyUI

Node-based visual workflow editor for Stable Diffusion, FLUX, etc.

In 2 lists

Stable Diffusion WebUI Forge - Neo

Actively maintained Forge-based Stable Diffusion web UI with the familiar extension-driven workflow.

Diffusers

PyTorch library for diffusion pipelines spanning image, video, and audio generation.

In 5 listsDetails

InvokeAI

Full-featured creative studio.

In 3 lists

SD.Next

All-in-one WebUI for AI generative image and video creation with multi-platform support, SDNQ quantization, and balanced CPU/GPU memory offload.

In 2 lists

stable-diffusion.cpp

Production-ready C++ inference runtime for SD, Flux, and related diffusion models, optimized for CPU and GPU deployment.

In 3 lists

Upscayl

Free and open-source AI image upscaler for Linux, macOS, and Windows. Uses Real-ESRGAN and Vulkan architecture to enhance images by reconstructing high-resolution details. Cross-platform desktop app with batch processing. AGPL-3.0 licensed.

In 6 listsDetails

upres-cli

Open-source (MIT) CLI and MCP server for UpRes, an image and video upscaler with 14 public model aliases (faces, text, print, art, motion) and 8K output.

Krita AI Diffusion

Streamlined AI image generation plugin for Krita. Inpaint and outpaint with optional text prompt, no tweaking required. Integrates ComfyUI backend for professional digital painting workflows. GPL-3.0 licensed.

In 2 lists

Deep-Live-Cam

Real-time face swap and one-click video deepfake with only a single image. High-quality face swapping for live video streaming and content creation. AGPL-3.0 licensed.

In 2 lists

Faceswap

Open-source deep learning software for recognizing and swapping faces in pictures and videos. GPL-3.0 licensed.

In 4 listsDetails

EchoMimic (Ant Group)

Lifelike audio-driven portrait animations through editable landmark conditioning. High-quality talking head generation with precise lip synchronization and natural head movements. AAAI 2025. Apache 2.0 licensed.

Hyperframes (HeyGen)

Open-source long video generation platform for cinematic and social video creation with timeline control, diffusion-based motion modules, and multimodal conditioning.

In 2 lists

LTX-2 (Lightricks)

Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.

In 3 lists

Helios (PKU-YuanGroup)

Efficient long-video generation framework with 24GB VRAM support for up to 10,000 frames (5+ minutes) and 1280×768 resolution. Apache 2.0 licensed.

In 2 lists

Pixelle-Video (AIDC-AI)

Text-to-video foundation model optimized for long coherent scenes and controllable generation workflows. Apache 2.0 licensed.

MoneyPrinterTurbo

An end-to-end short-video generation pipeline that automates scripts, footage collection, voiceover, and subtitle synthesis. MIT licensed.

In 4 listsDetails

OpenMontage

Agentic video production platform that orchestrates scriptwriting, asset generation, voice synthesis, editing, and rendering. AGPL-3.0 licensed.

In 2 lists

ViMax

Multi-agent video generation framework that orchestrates scriptwriting, storyboarding, character design, and temporal visual consistency for end-to-end video synthesis. MIT licensed.

In 2 lists

WhisperLive

Nearly-live implementation of OpenAI's Whisper for real-time speech-to-text transcription. Supports faster-whisper, tensorrt, and openvino backends with WebSocket streaming. MIT licensed.

In 2 lists

OpenShorts

Self-hosted app that turns long videos into vertical short clips using faster-whisper transcription, LLM clip selection, face-tracking 9:16 reframing, and animated captions. MIT licensed.

ACE-Step 1.5

Local-first music generation model with broad hardware support across Mac, AMD, Intel, and CUDA devices.

In 2 lists

Magenta RealTime 2

Open-weights live music model for streaming generation and real-time interaction. Apache 2.0 licensed.

YuE

Full-song music generation foundation model with symbolic melody-and-chord planning, acoustic synthesis, zero-shot covers, and song editing. Apache 2.0 licensed.

In 3 lists

Stable Audio Tools

Stability AI's open-source audio and music generative models. Latent diffusion model for generating audio conditioned on metadata and timing, providing faster inference times and creative control for sound effects and music production. MIT licensed.

GPT-SoVITS

Few-shot voice cloning with just 1 minute of voice data. Combines GPT and SoVITS architectures for high-quality TTS with cross-lingual support and emotional expression. MIT licensed.

In 3 lists

Supertonic

Lightning-fast, on-device, multilingual text-to-speech system running natively via ONNX. MIT licensed.

Voicebox

Local-first AI voice studio to clone voices, generate speech in multiple languages, and dictate text locally. MIT licensed.

Voice-Pro

Multilingual AI speech recognition, text-to-speech, and zero-shot voice cloning WebUI. GPL-3.0 licensed.

In 3 lists

VoiceStudio

Local-first voice platform for multi-engine text-to-speech, voice cloning and design, video dubbing, and dictation.

In 2 lists

Modly

Desktop application for image-to-3D mesh generation using local GPU-accelerated AI models. MIT licensed.

In 2 lists

gsplat (3D Gaussian Splatting tools)

High-performance 3D Gaussian Splatting library.

In 2 lists

LichtFeld-Studio

Native application for training, editing, and exporting 3D Gaussian Splatting scenes with MCMC optimization and timelapse generation. GPL-3.0 licensed.

In 4 listsDetails

OpenSplat

Production-grade, portable implementation of 3D Gaussian Splatting with CPU/GPU support for Windows, Mac, and Linux. Creates 3D scenes from camera poses and sparse points. AGPL-3.0 licensed.

TRELLIS.2

3D generative model for high-fidelity image-to-3D generation utilizing compact structured latents. MIT licensed.

In 2 lists

7. Training & Fine-tuning Ecosystem

Oumi

Fully open-source platform for the complete foundation model lifecycle - from data preparation and training to evaluation and deployment. Supports 100+ models with 200+ recipes for fine-tuning gpt-oss, Qwen3, DeepSeek-R1, and more. Apache 2.0 licensed.

In 3 lists

Marin

Open-source framework for foundation-model development with composable data, training, and evaluation pipelines for modern LLM research.

LLaMA-Factory

One-stop unified framework for SFT, DPO, ORPO, KTO with web UI.

In 3 lists

Axolotl

YAML-driven full pipeline for SFT, DPO, GRPO.

In 4 listsDetails

ms-swift

Unified training framework for 600+ LLMs and 300+ MLLMs with CPT/SFT/DPO/GRPO (AAAI 2025).

In 3 lists

Unsloth

2× faster, 70% less memory fine-tuning.

In 8 listsDetails

LitGPT

Clean from-scratch implementations of 20+ LLMs.

In 4 listsDetails

torchtune

PyTorch-native library for post-training, fine-tuning, and experimentation with LLMs.

In 2 lists

kohya_ss

Gradio-based GUI and CLI for training Stable Diffusion models (LoRA, Dreambooth, fine-tuning, SDXL). Provides accessible interface to Kohya's powerful training scripts.

In 2 lists

TRL (Transformers Reinforcement Learning)

Official library for RLHF, SFT, DPO, ORPO.

In 8 listsDetails

verl

Volcano Engine Reinforcement Learning for LLMs with PPO, GRPO, REINFORCE++, DAPO (EuroSys 2025).

In 3 lists

NeMo-RL

Scalable toolkit for efficient model reinforcement with DTensor and Megatron backends.

In 3 lists

OpenRLHF

Easy-to-use, scalable RLHF framework based on Ray. Supports PPO, GRPO, REINFORCE++, DAPO with vLLM integration and async training. Apache 2.0 licensed.

In 4 listsDetails

LMFlow

Extensible toolkit for finetuning and inference of large foundation models. Features RAFT alignment algorithm and comprehensive model support. Apache 2.0 licensed.

In 4 listsDetails

XTuner

A next-generation training engine built for ultra-large MoE models with efficient QLoRA and full-parameter fine-tuning. Apache 2.0 licensed.

In 2 lists

Ludwig

Low-code framework for building custom LLMs and deep neural networks. Declarative YAML configuration for training state-of-the-art models with PEFT/LoRA, 4-bit quantization, distributed training via Hugging Face Accelerate, and native Kubernetes support. Linux Foundation AI project. Apache 2.0…

In 6 listsDetails

TorchTitan (PyTorch)

PyTorch native platform for training generative AI models at scale. Showcases 4D parallelism (FSDP, tensor, pipeline, context) for LLM pretraining with 65%+ speedups over optimized baselines. BSD-3-Clause licensed.

VeOmni (ByteDance)

Versatile framework for both single- and multi-modal pre-training and post-training. Model-centric distributed recipe zoo supporting text, vision, audio, and video models with unified training interface. Apache 2.0 licensed.

In 2 lists

H2O LLM Studio

No-code GUI framework for fine-tuning LLMs. Streamlined interface for SFT, reward modeling, and model deployment. Apache 2.0 licensed.

In 3 lists

PRIME-RL

Agentic RL Training at Scale from Prime Intellect. Framework for large-scale reinforcement learning capable of scaling to 1000+ GPUs with fully asynchronous RL, FSDP2 training, and vLLM inference. Apache 2.0 licensed.

In 3 lists

slime

LLM post-training framework for RL Scaling from THUDM. Supports SFT and RL training with multi-turn compilation feedback, powering projects like TritonForge for automated GPU kernel generation. Apache 2.0 licensed.

In 5 listsDetails

rLLM

Democratizing Reinforcement Learning for LLMs. Framework for training AI agents with RL featuring near-zero code changes, CLI-first workflow, and 50+ built-in benchmarks. Supports GRPO, REINFORCE, RLOO with verl and tinker backends. Apache 2.0 licensed.

EasyR1

Efficient, scalable, multi-modality RL training framework based on veRL. Extends veRL to support vision-language models with GRPO algorithm for efficient RL training. Apache 2.0 licensed.

In 3 lists

LeRobot

Making AI for robotics more accessible with end-to-end learning. State-of-the-art approaches for imitation learning and reinforcement learning with pretrained models, datasets, and simulated environments. Apache 2.0 licensed.

In 6 listsDetails

AI-Toolkit

Ultimate training toolkit for finetuning diffusion models. Easy-to-use all-in-one training suite supporting FLUX.1, FLUX.2, Stable Diffusion, and video models with both GUI and CLI interfaces. Consumer-grade hardware friendly with comprehensive LoRA and full fine-tuning support. MIT licensed.

In 2 lists

OneTrainer

One-stop solution for all your Diffusion training needs. Supports FLUX, Stable Diffusion 1.5/2.x/3.x/SDXL, Würstchen, PixArt, Hunyuan Video and more. Features full fine-tuning, LoRA, embeddings, masked training, automatic backups, and TensorBoard integration. GPL-3.0 licensed.

In 2 lists

FluxGym

Dead simple FLUX LoRA training UI with LOW VRAM support (12GB/16GB/20GB). WebUI forked from AI-Toolkit with backend powered by Kohya Scripts. Combines simplicity of Gradio interface with flexibility of Kohya's powerful training scripts. GPL-3.0 licensed.

MiniMind

Train a 64M-parameter LLM from scratch in just 2 hours for $3. Complete from-scratch implementation covering MoE, data cleaning, pretraining, SFT, LoRA, RLHF (DPO/PPO/GRPO), tool use, and model distillation. All core algorithms implemented in pure PyTorch without high-level abstractions.…

In 4 listsDetails

FastChat

Open platform for training, serving, and evaluating large language model chatbots. Powers Chatbot Arena (lmarena.ai) serving 10M+ requests for 70+ LLMs. Includes training code for Vicuna, MT-Bench evaluation, and distributed multi-model serving with OpenAI-compatible APIs. Apache 2.0 licensed.

In 12 listsDetails

PaddleNLP

Easy-to-use and powerful LLM library built on Baidu's PaddlePaddle framework. Supports 100+ models with efficient training, compression, and high-performance inference on diverse hardware. Features RsLoRA+ algorithm, DeepSeek V3/R1 support with FP8/INT8 quantization, and unified checkpointing.…

In 4 listsDetails

Soup

One-config CLI for LLM post-training: SFT, DPO, GRPO, KTO, ORPO, plus eval gating and export. Layer streaming trains an 8B model on a 4 GB laptop GPU by streaming the frozen base from host RAM one decoder layer at a time. Apache 2.0 licensed.

In 2 lists

PEFT (Parameter-Efficient Fine-Tuning)

Official library with LoRA, QLoRA, DoRA, etc.

In 7 listsDetails

Liger Kernel

Ultra-fast custom kernels for training speedup.

In 3 lists

MergeKit

Advanced model merging tools.

In 2 lists

distilabel

End-to-end pipeline for synthetic instruction data.

In 3 lists

Data-Juicer

High-performance data processing for LLM training.

Argilla

Open-source data labeling + synthetic data platform.

In 3 lists

SDV (Synthetic Data Vault)

High-fidelity tabular and relational synthetic data.

In 2 lists

DataTrove (Hugging Face)

Platform-agnostic data processing pipelines for LLM training at scale. Handles filtering, deduplication, and tokenization on local machines or SLURM clusters.

In 3 lists

Bespoke Curator

Synthetic data curation for post-training and structured data extraction. Makes it easy to build pipelines around LLMs with batching and progress tracking. Apache 2.0 licensed.

In 2 lists

SDG (Harbin Institute)

Specialized framework for generating high-quality structured tabular synthetic data with CTGAN models supporting billion-level data processing. Apache 2.0 licensed.

DeepSpeed

Extreme-scale training optimizations.

In 3 lists

Colossal-AI

Unified system for 100B+ models.

In 14 listsDetails

Megatron-LM

Distributed training framework and reference codebase for large transformer models at scale.

In 4 listsDetails

Ray Train

Scalable distributed training.

In 13 listsDetails

Nanotron (Hugging Face)

Minimalistic 3D-parallelism LLM pretraining with tensor, pipeline, and data parallelism. Designed for simplicity and speed.

In 3 lists

RLinf

Scalable open-source RL infrastructure for post-training foundation models via reinforcement learning. Features M2Flow paradigm for embodied AI and agentic workflows with real-world robotics integrations. Apache 2.0 licensed.

In 3 lists

dstack

Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal. MPL-2.0 licensed.

In 4 listsDetails

Streaming (MosaicML)

High-performance data streaming library for efficient neural network training. Streams training data from cloud storage (S3, GCS, Azure) with local caching and deterministic shuffling. Apache 2.0 licensed.

In 2 lists

Higgsfield

Fault-tolerant GPU orchestration and distributed ML framework for training models at billion- to trillion-parameter scale. Apache 2.0 licensed.

LLM Compressor (vLLM)

Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM. Supports GPTQ, AWQ, SmoothQuant, AutoRound, and FP8/INT8 quantization with seamless Hugging Face integration.

In 2 lists

NVIDIA Model Optimizer

Unified library of SOTA model optimization techniques including quantization, pruning, distillation, and speculative decoding. Compresses deep learning models for deployment with TensorRT-LLM, TensorRT, and vLLM to optimize inference speed across NVIDIA hardware.

In 2 lists

8. MLOps / LLMOps & Production

MLflow

End-to-end open platform for the ML/LLM lifecycle.

In 7 listsDetails

DVC (Data Version Control)

Git-like versioning for data and models.

In 8 listsDetails

ClearML

Open-source platform for experiment tracking, orchestration, data management, and model serving.

In 3 lists

Weights & Biases Weave

Open-source tracing and experiment tracking.

In 2 lists

Aim

Self-hosted ML experiment tracker designed to handle 10,000s of training runs with performant UI and SDK for programmatic access. Apache 2.0 licensed.

In 7 listsDetails

Feast

Open source feature store for ML. Manages offline/online feature storage with point-in-time correctness to prevent data leakage. Apache 2.0 licensed.

In 6 listsDetails

OpenLineage

Open standard for lineage metadata collection designed to instrument jobs as they run. Defines a generic model of run, job, and dataset entities for consistent data lineage tracking. Apache 2.0 licensed.

In 3 lists

Marquez

LF AI & Data Foundation Graduated project for metadata collection, aggregation, and visualization. Maintains provenance of how datasets are consumed and produced with global visibility into job runtime and dataset lifecycle management. Integrates with OpenLineage. Apache 2.0 licensed.

In 5 listsDetails

Civitai

Open-source AI model hub and community platform for sharing and discovering generative AI models, with focus on image generation models. Features model versioning, reviews, and integrated inference. Apache 2.0 licensed.

Hugging Face Hub

Official Python client for the Hugging Face Hub. Download, upload, and manage 1M+ open-source ML models and datasets programmatically. The de facto standard for model sharing and distribution. Apache 2.0 licensed.

In 3 lists

ModelScope

Model-as-a-Service platform bringing together 700+ state-of-the-art ML models from the AI community. Covers NLP, CV, Audio, Multi-modality, and AI for Science with streamlined model inference, fine-tuning and evaluation. Apache 2.0 licensed.

In 2 lists

OpenVINO Open Model Zoo

Pre-trained deep learning models and demos optimized for Intel hardware. 200+ public pre-trained models for vision, speech, and NLP with benchmarking tools and accuracy metrics. Apache 2.0 licensed.

ONNX Model Zoo

Collection of pre-trained, state-of-the-art models in the ONNX format. 80+ models spanning vision, NLP, and audio with validation data and reference implementations. Apache 2.0 licensed.

Transformers.js

State-of-the-art Machine Learning for the web. Run Hugging Face Transformers directly in your browser with no server needed. Supports 1000+ models including BERT, GPT-2, T5, and more via ONNX Runtime Web. Apache 2.0 licensed.

In 2 lists

DJL (Deep Java Library)

Engine-agnostic deep learning framework for Java with built-in model zoo. Load and run PyTorch, TensorFlow, MXNet, and ONNX models with a unified API. Includes 80+ pre-trained models for CV and NLP. Apache 2.0 licensed.

In 3 lists

TorchVision Models

PyTorch's official computer vision library with 50+ pre-trained model architectures including ResNet, EfficientNet, Vision Transformers (ViT), ConvNeXt, and more. The de facto standard model zoo for PyTorch computer vision. BSD-3-Clause licensed.

In 9 listsDetails

TensorFlow Model Garden

Official TensorFlow repository of state-of-the-art (SOTA) models and modeling solutions. Contains reference implementations for BERT, ResNet, Transformer, and many more with pre-trained weights and training scripts. Apache 2.0 licensed.

In 12 listsDetails

PINTO Model Zoo

Repository for storing models inter-converted between various frameworks. Supports TensorFlow, PyTorch, ONNX, OpenVINO, TFJS, TFTRT, TensorFlowLite (Float32/16/INT8), EdgeTPU, and CoreML. 4,100+ stars with extensive model conversion tools for edge deployment. MIT licensed.

Cerebras Model Zoo

Collection of deep learning models and utilities optimized for Cerebras hardware. Includes reference implementations for Llama, Mixtral, DINOv2, and Llava with configuration files, data preprocessing tools, and checkpoint converters. 1,150+ stars. Apache 2.0 licensed.

PaddleClas

Comprehensive image recognition and classification toolkit with rich model zoo. 5,800+ stars featuring 24 series of classification networks, 122 pretrained models, and end-to-end image recognition systems including PP-ShiTuV2. Apache 2.0 licensed.

Cog (Replicate)

Containerize and deploy ML models with production-grade inference servers. Packages models into standardized containers with automatic API generation, GPU support, and one-command deployment. Powers thousands of production AI models on Replicate. Apache 2.0 licensed.

In 3 lists

BentoML

Unified framework to build, ship, and scale AI apps.

In 7 listsDetails

ZenML

Pipeline and orchestration framework for taking ML and LLM systems from development to production.

In 5 listsDetails

Kubeflow

Kubernetes-native ML/LLM platform.

In 7 listsDetails

KServe

Kubernetes-based model serving.

In 6 listsDetails

Metaflow

Netflix's ML platform for building and managing real-world AI systems. Powers thousands of projects at Netflix, Amazon, and DoorDash. Apache 2.0 licensed.

In 6 listsDetails

Flyte

Kubernetes-native workflow orchestration platform for AI/ML pipelines. Dynamic, resilient orchestration with strong type safety and reproducibility. Used by Lyft, Spotify, and Gojek. Apache 2.0 licensed.

In 5 listsDetails

Prefect

Workflow orchestration framework for building resilient data and ML pipelines. Python-native with modern observability and 200+ integrations. Apache 2.0 licensed.

In 11 listsDetails

Dagster

Cloud-native orchestration platform for developing and maintaining data assets including ML models. Declarative programming model with integrated lineage and observability. Apache 2.0 licensed.

In 12 listsDetails

Kubeflow Pipelines

Machine Learning Pipelines for Kubeflow. Platform for building and deploying portable, scalable ML workflows using Kubernetes and Argo. Apache 2.0 licensed.

In 3 lists

Argo Workflows

CNCF graduated container-native workflow engine for orchestrating parallel jobs on Kubernetes. Powers Kubeflow Pipelines and widely used for ML/data processing at scale. Apache 2.0 licensed.

In 10 listsDetails

MLRun

Open-source AI orchestration platform for quickly building and managing continuous ML and generative AI applications across their lifecycle. Automates data preparation, model tuning, and deployment. Apache 2.0 licensed.

In 4 listsDetails

Kestra

Event-driven orchestration and scheduling platform for mission-critical workflows. Infrastructure-as-Code approach with declarative YAML, Git version control integration, and hundreds of plugins for data pipelines and ML workflows. Apache 2.0 licensed.

In 8 listsDetails

KitOps

CNCF open source DevOps tool for packaging, versioning, and securely sharing AI/ML models, datasets, code, and configuration. Packages everything into OCI artifacts stored in existing container registries. Apache 2.0 licensed.

In 3 lists

Polyaxon

MLOps Tools For Managing & Orchestrating The Machine Learning LifeCycle. Reproducible and scalable machine learning workflows on Kubernetes with experiment tracking, model management, and pipeline orchestration. Apache 2.0 licensed.

In 9 listsDetails

Netflix Maestro

Netflix's next-generation workflow orchestrator for data and ML pipelines at massive scale. Highly scalable and flexible scheduler designed to handle millions of workflows across thousands of nodes. Apache 2.0 licensed.

In 3 lists

HAMi

Heterogeneous GPU Sharing on Kubernetes. CNCF sandbox project providing GPU virtualization, slicing, and scheduling for efficient AI workload management across heterogeneous accelerators (GPUs, NPUs, MLUs). Apache 2.0 licensed.

In 2 lists

NVIDIA KAI Scheduler

Kubernetes-native GPU scheduler for AI workloads at large scale. Originally developed by Run:ai, now open-sourced by NVIDIA. Optimizes GPU resource allocation with dynamic allocation and efficient queue management. Apache 2.0 licensed.

NVIDIA DeepOps

Infrastructure automation tools for building GPU clusters with Kubernetes and Slurm. Deploys multi-node GPU clusters with monitoring, logging, and storage for AI/HPC workloads. BSD-3-Clause licensed.

In 2 lists

SkyPilot

Run, manage, and scale AI workloads on any AI infrastructure. Unified interface to access and manage compute across Kubernetes, Slurm, and 20+ cloud providers. Used by Shopify and research institutions for training and inference. Apache 2.0 licensed.

In 4 listsDetails

Volcano

Cloud-native batch scheduling system for compute-intensive workloads. CNCF incubating project with gang scheduling, job dependency management, and topology-aware scheduling for AI/ML and deep learning. Apache 2.0 licensed.

In 4 listsDetails

Apache YuniKorn

Kubernetes resource scheduler for batch, data, and ML workloads. Provides hierarchical resource queues, multi-tenancy fairness, and gang scheduling for big data and machine learning applications. Apache 2.0 licensed.

In 2 lists

Kueue

Kubernetes-native job queueing system for batch, HPC, AI/ML, and similar applications. Cloud-native job queueing with resource flavor fungibility, fair sharing, cohorts, and preemption policies. Integrates with Kubeflow, Ray, and JobSet. Apache 2.0 licensed.

In 3 lists

Featuretools

Open-source Python library for automated feature engineering. Transforms transactional and relational datasets into feature matrices for machine learning using Deep Feature Synthesis with reusable primitives. BSD-3-Clause licensed.

In 6 listsDetails

Kedro

Toolbox for production-ready data science. Uses software engineering best practices to help you create data engineering and data science pipelines that are reproducible, maintainable, and modular. Apache 2.0 licensed.

In 7 listsDetails

Feature-engine

Python library with multiple transformers to engineer and select features for machine learning models. scikit-learn compatible with fit() and transform() methods for encoding, imputation, variable transformation, and feature selection. BSD-3-Clause licensed.

In 7 listsDetails

NVTabular

GPU-accelerated feature engineering and preprocessing library for tabular data. Manipulates terabyte-scale datasets to train deep learning recommender systems. Component of NVIDIA Merlin framework. Apache 2.0 licensed.

OpenMLDB

Open-source machine learning database providing a feature platform for consistent features between training and inference. Real-time relational data feature computation system for online ML applications. Apache 2.0 licensed.

Langfuse

#1 open-source LLM observability platform.

In 10 listsDetails

Phoenix (Arize)

AI observability & evaluation platform.

In 10 listsDetails

Evidently

ML & LLM monitoring framework.

In 8 listsDetails

Opik (Comet)

Production-ready LLM evaluation platform.

In 16 listsDetails

LiteLLM

AI Gateway to call 100+ LLM APIs in OpenAI format with unified cost tracking, guardrails, load balancing, and logging.

In 16 listsDetails

OpenLIT

OpenTelemetry-native LLM observability platform with GPU monitoring, evaluations, prompt management, and guardrails.

In 5 listsDetails

flameox

Runtime-evidence toolkit for agents that coordinates PyTorch Profiler and Nsight Systems captures, preserves native traces, and compares GPU-kernel and inference runs. MIT licensed.

In 2 lists

OpenLLMetry (Traceloop)

Open-source observability for GenAI/LLM applications based on OpenTelemetry with 25+ integration backends.

In 7 listsDetails

Agenta

Open-source LLMOps platform combining prompt playground, prompt management, LLM evaluation, and observability.

In 6 listsDetails

Latitude

Open-source agent engineering platform with prompt management, evaluations, and optimization. Features prompt playground, LLM-as-judge evals, and GEPA prompt optimizer for production LLM features. LGPL-3.0 licensed.

In 3 lists

Helicone

Open-source LLM observability with request logging, caching, rate limiting, and cost analytics.

In 8 listsDetails

Giskard

Open-source evaluation and testing library for LLM agents. Red teaming, vulnerability scanning, RAG evaluation, and safety testing with modular architecture. Apache 2.0 licensed.

In 3 lists

Awesome Agentic Engineering

Zero-dependency AI agent production-readiness toolkit with GitHub Actions score and risk-profile gates, prompt-injection fixtures, machine-readable results, and reproducible LangGraph traces. MIT licensed.

Portkey Gateway

Blazing fast AI Gateway to route 200+ LLMs with unified API. Integrated guardrails, load balancing, fallbacks, and cost tracking. MIT licensed.

In 6 listsDetails

Envoy AI Gateway

Manages unified access to generative AI services built on Envoy Gateway. Kubernetes-native AI gateway for routing, load balancing, and managing LLM traffic with enterprise-grade reliability. Apache 2.0 licensed.

In 2 lists

Unified AI System

Self-hosted Apache-2.0 AI gateway and MCP server: per-key virtual tokens with token budgets, exact and lexical-approximate response caching (semantic-grade matching requires an attached embedding endpoint), an append-only audit chain, deterministic local prompt enhancement that makes no provider…

In 2 lists

Pezzo

Cloud-native LLMOps platform with prompt management, versioning, and observability. Features collaborative prompt editing, A/B testing, and cost analytics. Apache 2.0 licensed.

In 4 listsDetails

Microsoft PromptFlow

Comprehensive suite for LLM-based AI app development from prototyping to production. Includes prompt engineering, evaluation, and deployment tools with VS Code integration. MIT licensed.

In 3 lists

ChainForge

Visual programming environment for battle-testing prompts and evaluating LLM outputs. Features node-based prompt chains, multi-model comparison, and hypothesis testing. MIT licensed.

In 2 lists

Future AGI

Open-source self-hostable end-to-end agent engineering and optimization platform that unifies tracing, evals, simulations, datasets, gateway, and guardrails. Built for shipping self-improving AI agents with one feedback loop from prototype to production. Apache 2.0 licensed.

In 6 listsDetails

KubeStellar Console

AI-powered multi-cluster Kubernetes dashboard with GPU workload monitoring, AI pipeline observability, and CNCF ecosystem integrations. Apache 2.0 licensed.

In 13 listsDetails

OrcaReplay

Records a coding-agent run below the harness and replays it offline with the network off, or forks it from a checkpoint onto another model.

In 11 listsDetails

PurpleLlama (Meta)

Comprehensive set of tools to assess and improve LLM security. Includes Llama Guard safety classifiers, CyberSec Eval benchmarks, and Prompt Guard for prompt injection detection. BSD-3-Clause licensed.

In 3 lists

Garak (NVIDIA)

The LLM vulnerability scanner. Probes models for hallucinations, data leakage, prompt injection, misinformation, toxicity, and jailbreaks. Extensive plugin-based architecture with 100+ vulnerability probes. Apache 2.0 licensed.

In 4 listsDetails

Promptfoo

Open-source LLM evaluation and red teaming framework. Test prompts, agents, and RAGs with automated security vulnerability scanning, side-by-side model comparison, and CI/CD integration. Now part of OpenAI. MIT licensed.

In 11 listsDetails

DeepTeam (Confident AI)

Red teaming framework for LLM systems with 50+ vulnerabilities, 20+ adversarial attacks, and production-ready guardrails. Includes OWASP, NIST, and MITRE ATLAS framework mappings. Apache 2.0 licensed.

In 3 lists

SkillSpector (NVIDIA)

Security scanner for AI agent skills that detects vulnerabilities, malicious patterns, and security risks. Apache 2.0 licensed.

In 5 listsDetails

ADR (Uber)

Enterprise security system for AI agents providing telemetry collection, benchmark evaluations, and threat detection. Apache 2.0 licensed.

In 2 lists

9. Evaluation, Benchmarks & Datasets

LiveBench

Contamination-free LLM benchmark with objective ground-truth scoring. ICLR 2025 spotlight paper featuring frequently-updated questions from recent sources. Tests math, coding, reasoning, language, instruction following, and data analysis.

lm-evaluation-harness (EleutherAI)

De-facto standard for generative model evaluation.

In 7 listsDetails

HELM (Stanford)

Holistic Evaluation of Language Models.

In 2 lists

SWE-bench

Evaluates LLMs on real-world GitHub issues from 15+ Python repositories.

GAIA

Real-world multi-step agentic benchmark.

OpenCompass

Evaluation platform for benchmarking language and multimodal models across large benchmark suites.

In 4 listsDetails

MLPerf Inference

Industry-standard ML inference benchmarks with reference implementations for AI accelerators.

In 2 lists

MLPerf Training

Industry-standard ML training benchmarks from MLCommons. Reference implementations for training AI models at scale across image classification, object detection, NLP, and recommendation tasks. Apache 2.0 licensed.

VLMEvalKit

Open-source evaluation toolkit for large multi-modality models (LMMs). Supports 220+ LMMs and 80+ benchmarks including MMMU, MathVista, and ChartQA. Powers the OpenVLM Leaderboard. Apache 2.0 licensed.

In 4 listsDetails

Vectara Hallucination Leaderboard

Leaderboard comparing LLM performance at producing hallucinations when summarizing short documents. Systematic evaluation of factual consistency across major models. Apache 2.0 licensed.

In 2 lists

SWE-rebench (Nebius)

Continuously updated benchmark with 21,000+ real-world SWE tasks for evaluating agentic LLMs. Decontaminated, mined from GitHub.

MLE-bench (OpenAI)

Benchmark for measuring how well AI agents perform at machine learning engineering. Evaluates agents on 75 Kaggle competitions covering diverse ML tasks. MIT licensed.

In 3 lists

PinchBench

Benchmarking system for evaluating LLM models as OpenClaw coding agents. Built with Rust by the kilo.ai team. MIT licensed.

DeepEval

The "Pytest for LLMs".

In 9 listsDetails

Inspect AI

Framework for large language model evaluations from the UK AI Security Institute.

In 3 lists

Lighteval

Evaluation toolkit for LLMs across multiple backends with reusable tasks, metrics, and result tracking.

In 3 lists

Hugging Face Evaluate

Standardized evaluation metrics.

In 4 listsDetails

OpenAI Evals

Framework for evaluating LLMs and LLM systems with an open-source registry of 100+ community-contributed benchmarks. MIT licensed.

In 8 listsDetails

LMMs-Eval

Unified multimodal evaluation toolkit for text, image, video, and audio tasks with 100+ supported benchmarks.

In 2 lists

BrowserGym

Gym environment for web task automation and agent evaluation. Includes MiniWoB, WebArena, WorkArena, and more. Apache 2.0 licensed.

In 2 lists

TruLens

Evaluation and tracking for LLM experiments and AI agents. Provides feedback functions for measuring quality, relevance, and groundedness with LangChain and LlamaIndex integrations. MIT licensed.

In 2 lists

OpenEvals

Open-source evaluation library for LLM and agent applications. Built by LangChain with pre-built evaluators for common use cases including RAG, agents, and structured output validation. MIT licensed.

AutoRAG

RAG AutoML tool for automatically finding optimal RAG pipelines. Evaluates and optimizes retrieval-augmented generation with AutoML-style automation for your own data and use-case. Apache 2.0 licensed.

In 5 listsDetails

E2B Code Interpreter

Python & JS/TS SDK for running AI-generated code in secure isolated sandboxes. Essential infrastructure for evaluating code-generating LLMs with safe execution environments. Apache 2.0 licensed.

In 2 lists

SimpleEvals (OpenAI)

Lightweight library for evaluating language models with transparent accuracy numbers. Reference implementations for MMLU, GPQA, MATH, HumanEval, MGSM, DROP, and SimpleQA benchmarks. MIT licensed.

EvalScope (ModelScope)

Streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking. One-stop evaluation solution with 80+ benchmarks. Apache 2.0 licensed.

In 3 lists

Harbor

Framework for running agent evaluations and creating/using RL environments. Evaluate arbitrary agents like Claude Code, OpenHands, and Codex CLI. Build and share benchmarks and environments. Apache 2.0 licensed.

In 2 lists

YYLO Benchmark

Longitudinal evaluation for agent runs: executes task prompts in private fresh-repository attempt workspaces and retains attempt evidence, evaluator provenance, and reports for later re-evaluation. MIT licensed.

In 2 lists

Hugging Face Datasets

Largest open repository of datasets.

In 6 listsDetails

OrcaPromptVault

Open corpus of the instructions and tool schemas shipping AI agents are actually sent: 119 artifacts from 43 products, 44 of them recorded off the wire with the command that reproduces each, every file marked captured or reported. AGPL-3.0.

In 4 listsDetails

FineWeb / FineWeb-2 (Hugging Face)

Curated 15T+ token web dataset for pre-training.

In 2 lists

OSWorld

Multimodal agent benchmark dataset.

In 3 lists

10. AI Safety, Alignment & Interpretability

AgentOps

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and evaluation. Integrates with CrewAI, Agno, OpenAI Agents SDK, LangChain, Autogen, AG2, and CamelAI. MIT licensed.

In 5 listsDetails

Bloom

Open-source agentic framework for automated behavioral evaluations of frontier AI models. Generates targeted evaluation suites to probe LLMs for specific behaviors (sycophancy, self-preservation, political bias, etc.) with quantitative elicitation rates. From Anthropic's safety research team. MIT…

AgentShield Benchmark

Apache-2.0 corpus of 537 test cases and a TypeScript runner for benchmarking AI agent guardrail products, maintained by a vendor whose own product it scores.

Alignment Handbook

Complete recipes for full-stack alignment.

interpret (Microsoft)

Fit interpretable models and explain blackbox machine learning with state-of-the-art explainability techniques including Explainable Boosting Machines and SHAP-based explanations.

In 8 listsDetails

TransformerLens

Gold-standard for mechanistic interpretability.

In 2 lists

SAELens

Sparse autoencoders for interpretable features.

nnsight

Library for inspecting, tracing, and intervening on neural network internals at scale.

Captum

PyTorch's official interpretability library.

In 3 lists

EasyEdit

Easy-to-use knowledge editing framework for LLMs. Enables precise modification of model knowledge and behavior to correct hallucinations or outdated information. ACL 2024. MIT licensed.

In 2 lists

AIX360

Comprehensive AI explainability toolkit with interpretability algorithms for data and machine learning models. Includes TED, BRCG, and ProtoNN methods for diverse explanation needs. Apache 2.0 licensed.

In 2 lists

ELI5

Library for debugging/inspecting machine learning classifiers and explaining their predictions. Supports scikit-learn, XGBoost, LightGBM, and more with feature importance and explanation visualizations. MIT licensed.

In 2 lists

Shapash

User-friendly explainability library for transparent ML models. Beautiful visualizations with explicit labels that everyone can understand. Generates web reports and integrates with SHAP/LIME. Apache 2.0 licensed.

In 4 listsDetails

AI Fairness 360

Comprehensive toolkit for detecting, understanding, and mitigating unwanted algorithmic bias in datasets and ML models.

In 4 listsDetails

Fairlearn

Python package to assess and improve fairness of machine learning models. Provides metrics for disparity assessment and algorithms for unfairness mitigation with scikit-learn integration. MIT licensed.

In 3 lists

PyRIT (Microsoft)

Python Risk Identification Tool for generative AI. Microsoft's open-source framework for automated red teaming with multi-modal attack support, crescendo strategies, and 100+ operations experience. MIT licensed.

In 2 lists

Heretic

Open-source system for automatic censorship and robustness suppression removal in language-model outputs.

In 3 lists

Agentic Security

Agentic LLM vulnerability scanner and AI red teaming kit with multi-step attack simulation and automated security probing. Apache 2.0 licensed.

In 3 lists

NeMo Guardrails (NVIDIA)

Programmable guardrails toolkit for LLM-based conversational systems. Uses Colang DSL to define safety rules, dialog flows, and content boundaries. Integrates with LangChain, LangGraph, and LlamaIndex for production deployments. Apache 2.0 licensed.

In 6 listsDetails

Guardrails AI

Input/output validation framework for building reliable AI applications. Detects and mitigates risks through composable validators for PII, toxicity, prompt injection, and structured output validation. Features Guardrails Hub with 50+ pre-built validators. Apache 2.0 licensed.

In 6 listsDetails

Detoxify

Trained models and code to predict toxic comments on all 3 Jigsaw Toxic Comment Challenges. Built using PyTorch Lightning and Transformers for toxicity, severe toxicity, obscene, threat, insult, identity attack, and sexual explicit content detection. Apache 2.0 licensed.

RedAmon

AI-powered agentic red team framework that automates offensive security operations from reconnaissance to exploitation to post-exploitation with zero human intervention. Integrates multiple security tools for comprehensive penetration testing. MIT licensed.

CAI

Cybersecurity AI framework for semi- and fully-automating offensive and defensive security tasks. Purpose-built for cybersecurity use cases with agent-based architecture for vulnerability assessment and security operations. MIT licensed.

In 4 listsDetails

AI-Infra-Guard (Tencent)

Full-stack AI Red Teaming platform securing AI ecosystems via OpenClaw Security Scan, Agent Scan, Skills Scan, MCP scan, AI Infra scan and LLM jailbreak evaluation. Apache 2.0 licensed.

In 4 listsDetails

PentestAgent (GH05TCREW)

AI agent framework for black-box security testing, supporting bug bounty, red-team, and penetration testing workflows. MIT licensed.

In 2 lists

Superagent

Protects AI applications against prompt injections, data leaks, and harmful outputs. Embed safety directly into your app and prove compliance to your customers. MIT licensed.

In 4 listsDetails

Anthropic Sandbox Runtime

Portable sandbox execution environment for AI agents with constrained filesystem and network boundaries, designed to reduce blast radius in tool-use and LLM autonomy tests.

SchemaBrain

An open-source schema intelligence and safety layer (like Cloudflare) between AI agents and databases. Exposes a read-only Model Context Protocol (MCP) interface with PII masking, structured query recovery, and tamper-evident SHA-256 audit logs. Apache 2.0 licensed.

Anthropic Cybersecurity Skills

Structured library of 754 cybersecurity skills mapped to five frameworks including MITRE ATT&CK and NIST CSF 2.0, compatible with Claude Code, Cursor, and agentskills.io. Apache 2.0 licensed.

In 4 listsDetails

Strix

Autonomous AI-powered penetration testing agent that identifies vulnerabilities and validates them with proof-of-concept exploits. Apache 2.0 licensed.

In 3 lists

Responsible AI Toolbox

Suite of tools providing model and data exploration, assessment interfaces and libraries for understanding AI systems. Enables developers to develop and monitor AI more responsibly with better data-driven actions. MIT licensed.

Agent Governance Toolkit

Policy, safety, and execution controls for autonomous AI agents, including governance guardrails, sandboxing, and reliability checks. Apache 2.0 licensed.

In 5 listsDetails

Orca AI Incident Archive

Open database of real-world AI agent security incidents, with each sourced record flagged for confirmed harm, AI involvement and kind, plus JSON/CSV exports. CC BY 4.0 licensed.

Presidio (Microsoft)

SDK for detecting, redacting, masking, and anonymizing sensitive personally identifiable information (PII) across text and images. MIT licensed.

11. Specialized Domains

WeatherNext

Global weather and tropical cyclone forecasting framework from Google DeepMind, including WeatherNext 2, GraphCast, and GenCast models. Apache 2.0 licensed.

In 2 lists

NVIDIA Modulus

Open-source deep learning framework for physics-informed machine learning (Physics-ML). Build, train, and fine-tune models for AI4science and engineering applications using state-of-the-art SciML methods. Apache 2.0 licensed.

TorchGeo

PyTorch domain library for geospatial data. Datasets, samplers, transforms, and pre-trained models for multispectral satellite imagery and remote sensing. First library with pre-trained models for Sentinel-2 bands. MIT licensed.

In 3 lists

Astropy

Core library for astronomy and astrophysics in Python. Comprehensive tools for celestial coordinates, FITS I/O, cosmological calculations, and data analysis for professional astronomy. BSD-3-Clause licensed.

In 4 listsDetails

Boltz

Open-source biomolecular interaction prediction models. Boltz-1 was the first fully open source model to approach AlphaFold3 accuracy; Boltz-2 adds binding affinity prediction for drug discovery. MIT licensed.

In 3 lists

Protenix

High-accuracy open-source biomolecular structure prediction model from ByteDance. First fully open-source model to outperform AlphaFold3 across diverse benchmarks with Apache 2.0 licensing for both academic and commercial use.

In 2 lists

DeepChem

Democratizing deep learning for drug discovery, quantum chemistry, materials science, and biology. High-quality open-source toolchain with 50+ models and extensive tutorials. MIT licensed.

In 4 listsDetails

PyMC

Modern, comprehensive probabilistic programming framework in Python. Bayesian modeling with advanced MCMC sampling, variational inference, and seamless integration with ArviZ for visualization. Apache 2.0 licensed.

In 10 listsDetails

ArviZ

Exploratory analysis of Bayesian models with Python. Comprehensive visualization and diagnostics for probabilistic models, supporting PyMC, Pyro, Stan, and other PPLs. Apache 2.0 licensed.

In 3 lists

Stanza

Stanford NLP Python library for 100+ human languages. State-of-the-art neural pipelines for tokenization, NER, parsing, and sentiment analysis with pre-trained models. Apache 2.0 licensed.

In 5 listsDetails

MONAI

Medical Open Network for AI. End-to-end framework for healthcare imaging with state-of-the-art, production-ready training workflows. Apache 2.0 licensed.

In 2 lists

nnU-Net

Self-configuring deep learning method for medical image segmentation. Automatically adapts to any dataset without manual parameter tuning. Widely adopted as the standard baseline for biomedical segmentation challenges. Apache 2.0 licensed.

In 2 lists

OpenMed

Local-first healthcare AI framework and application for clinical text de-identification, entity extraction, and on-device specialized model serving. Apache 2.0 licensed.

In 3 lists

Alex

Open-source desktop novel-writing engine and AI story scaffold with Tauri, FastAPI, SQLite, and plot-generation workflows. MIT licensed.

AI Novel Writer

Local-first desktop workbench for long-form fiction that organizes story premises, characters, worldbuilding, chapter blueprints, drafting, review, and revision, with Ollama support and Windows/macOS releases. GPL-3.0 licensed.

In 2 lists

Aural

Self-hostable AI interview platform for voice, chat, and video interviews with adaptive follow-ups, automated scoring, and detailed feedback reports. MIT licensed.

Unity ML-Agents

Toolkit for training intelligent agents in games and simulations using deep reinforcement learning. Enables NPC behavior control, automated testing, and game design evaluation. Apache 2.0 licensed.

In 3 lists

Tianshou

An elegant PyTorch deep reinforcement learning library with clean API design and comprehensive algorithm implementations. Supports both single-agent and multi-agent RL with GPU acceleration. MIT licensed.

In 3 lists

RL Baselines3 Zoo

A training framework for Stable Baselines3 reinforcement learning agents with hyperparameter optimization, pre-trained agents, and extensive benchmark environments. MIT licensed.

skrl

Modular reinforcement learning library implemented in PyTorch, JAX, and NVIDIA Warp with support for Gymnasium, NVIDIA Isaac Lab, MuJoCo Playground, and other environments. MIT licensed.

In 4 listsDetails

Finetrainers

Scalable and memory-optimized training of diffusion models from Hugging Face. Supports LoRA and full fine-tuning for video and image generation models. Apache 2.0 licensed.

OpenSpiel

Collection of environments and algorithms for research in general reinforcement learning and search/planning in games from Google DeepMind. Apache 2.0 licensed.

In 2 lists

OpenBB

Financial data platform for analysts, quants and AI agents. Open-source investment research infrastructure with extensive data integrations. AGPL-3.0 licensed.

In 5 listsDetails

FinGPT

Open-source financial large language models. Democratizing financial AI with data-centric training pipeline and multiple model releases for trading, analysis, and robo-advising. MIT licensed.

In 6 listsDetails

FinRL

Financial reinforcement learning framework for quantitative trading. Deep RL library for stock trading, portfolio allocation, and market execution with pre-built environments and benchmarks. MIT licensed.

In 5 listsDetails

Qlib

AI-oriented quantitative investment platform from Microsoft. Supports diverse ML modeling paradigms including supervised learning, market dynamics modeling, and RL. Now equipped with RD-Agent for automated R&D process. MIT licensed.

In 5 listsDetails

FinRobot

Open-source AI agent platform for financial analysis using LLMs. Multi-agent system with specialized agents for trading, analysis, and research. Apache 2.0 licensed.

In 4 listsDetails

Kronos

Foundation model for financial candlesticks (K-lines) pre-trained on multi-dimensional market data across global exchanges. MIT licensed.

In 2 lists

Daily Stock Analysis

LLM-powered multi-market stock analysis system providing automated decision dashboard reports and real-time market insights. MIT licensed.

In 2 lists

stock-analysis

Evidence-first market research CLI and Agent Skill for A-share, Hong Kong, and US stocks, funds, and portfolios, with auditable JSON Evidence Packs and explicit data-quality gaps. MIT licensed.

In 3 lists

Vibe-Trading

Open-source personal trading agent and research autopilot platform featuring multi-agent swarms, quantitative backtesting, and broker connectors. MIT licensed.

In 4 listsDetails

AI Hedge Fund

Multi-agent system simulating an AI-powered hedge fund that utilizes specialized investor agent personalities to make trading decisions. MIT licensed.

In 8 listsDetails

etf-pattern-match-pybind11

High-performance ETF pattern matching with C++20/pybind11 acceleration achieving 43x DTW and 58x pattern-matching speedups over pure Python. MIT licensed.

In 2 lists

OpenCV

World's most widely used computer vision library.

In 5 listsDetails

Ultralytics YOLO

State-of-the-art real-time object detection.

In 6 listsDetails

Detectron2

High-performance object detection library.

In 7 listsDetails

CVAT

Industry-leading data annotation platform for computer vision. Interactive video and image annotation tool used by tens of thousands of teams for machine learning at any scale.

In 3 lists

SAM 2

Promptable image and video segmentation model with released checkpoints and training code.

AI Segmentation by TerraLab

QGIS plugin for point-and-click SAM segmentation of buildings, trees and any object in satellite and drone imagery, with a free CPU-only local mode.

In 2 lists

Kornia

Differentiable computer vision library.

In 4 listsDetails

Roboflow Supervision

Reusable computer-vision utilities for detection, tracking, and segmentation pipelines in Python.

In 4 listsDetails

torchaudio

PyTorch audio processing library. Comprehensive toolkit for audio I/O, transformations, and deep learning with support for speech recognition, TTS, and audio classification. BSD-2-Clause licensed.

In 6 listsDetails

MediaPipe

Cross-platform multimodal pipelines.

In 2 lists

OpenEyes

Hardware-agnostic robot vision framework with world models for predictive intelligence on edge devices.

Open3D

Modern library for 3D data processing with Python and C++ APIs. Core features include 3D data structures, processing algorithms, scene reconstruction, surface alignment, 3D visualization, and GPU acceleration. MIT licensed.

In 4 listsDetails

Point Cloud Library (PCL)

Standalone, large-scale open project for 2D/3D image and point cloud processing. Comprehensive algorithms for filtering, feature estimation, surface reconstruction, registration, model fitting, and segmentation. BSD licensed.

In 5 listsDetails

PyTorch3D

FAIR's library of reusable components for deep learning with 3D data. Provides efficient 3D operators, differentiable rendering, and mesh processing tools integrated with PyTorch. BSD licensed.

In 5 listsDetails

RTAB-Map

Real-Time Appearance-Based Mapping library for RGB-D, Stereo and LiDAR SLAM. Graph-based SLAM approach with incremental appearance-based loop closure detection for large-scale and long-term operation. BSD licensed.

In 2 lists

MoveIt 2

Open source robotics manipulation framework for ROS 2. Motion planning, manipulation, 3D perception, kinematics, control, and navigation for robotic arms. BSD-3-Clause licensed.

LingBot-Map

Feed-forward 3D foundation model for transforming and reconstructing streaming 3D scenes. Apache 2.0 licensed.

In 2 lists

Stable-Baselines3

Production-ready RL algorithms.

In 8 listsDetails

Isaac Lab

GPU-accelerated robot learning framework.

In 2 lists

MuJoCo

General-purpose physics simulator for robotics, biomechanics, and ML research. High-fidelity contact dynamics with native Python and C++ bindings. Apache 2.0 licensed.

In 2 lists

Gymnasium (ex-OpenAI Gym)

Standard RL environment API.

In 6 listsDetails

OpenEnv

End-to-end sandbox execution environment framework for agentic reinforcement learning training built on Gymnasium-compatible APIs.

JaxMARL

Multi-agent reinforcement learning library with JAX-accelerated environments and baselines.

MicroDuck RL

Reinforcement learning environments and training tools for the MicroDuck robotics platform.

Time Series Library (TSLib)

Comprehensive benchmark for time-series models.

In 3 lists

Chronos (Amazon)

Pretrained foundation models for time-series forecasting.

In 2 lists

GluonTS (AWS Labs)

Probabilistic time series modeling with deep learning. Powers Amazon SageMaker forecasting with PyTorch and MXNet backends. Apache 2.0 licensed.

In 4 listsDetails

AutoTS

Automated time series forecasting with broad model selection, ensembling, anomaly detection, and holiday effects. Designed for production deployment with minimal setup.

TimesFM (Google Research)

Pretrained decoder-only foundation model developed by Google Research for time-series forecasting, supporting up to 16k context length, covariate support, and Flax / PyTorch backends. Apache 2.0 licensed.

In 5 listsDetails

ExecuTorch

PyTorch runtime and toolchain for deploying AI models on mobile, embedded, and edge devices.

In 2 lists

OpenVINO

Intel's toolkit for edge deployment.

In 3 lists

Apache TVM

Open Machine Learning Compiler Framework. Universal deployment to bring models into minimum deployable modules that can be embedded and run everywhere from datacenter to edge devices. Apache 2.0 licensed.

In 3 lists

NCNN

High-performance neural network inference framework optimized for mobile platforms. No third-party dependencies, cross-platform, and runs faster than all known open-source frameworks on mobile CPU. Powers Tencent apps including QQ, WeChat, and Pitu. BSD-3-Clause licensed.

In 6 listsDetails

MNN

Blazing-fast, lightweight inference engine battle-tested by Alibaba. Supports inference and training with industry-leading on-device performance. Powers high-performance LLMs and Edge AI with MNN-LLM runtime. Apache 2.0 licensed.

In 3 lists

RuView

Real-time spatial intelligence, vital sign monitoring, and human pose estimation using commodity WiFi signals and neural networks. MIT licensed.

In 2 lists

OpenContracts

Self-hosted document annotation platform for legal AI. Semantic search, contract analysis, version control, and MCP integration for building legal knowledge bases. AGPL-3.0 licensed.

In 2 lists

Harvey LAB

Benchmark dataset and execution harness for evaluating AI agents on complex legal work across 24+ practice areas. MIT licensed.

CARLA

Open-source simulator for autonomous driving research. High-fidelity simulation of urban environments with realistic physics, sensors, and traffic scenarios. Widely used for training and validating self-driving algorithms. MIT licensed.

In 5 listsDetails

Webots

Open-source multi-platform robot simulator providing a complete development environment for modeling, programming, and simulating robots, vehicles, and mechanical systems. Used in education, research, and industry. Apache 2.0 licensed.

In 3 lists

Habitat-Sim

High-performance physics-enabled 3D simulator for embodied AI research. Supports 3D scans of indoor/outdoor spaces, CAD models, and configurable sensors. Powers Meta's embodied AI research. MIT licensed.

In 3 lists

Genesis World

General-purpose simulation platform for embodied AI and robotics research with flexible scenes, sensors, and agent workflows. Apache 2.0 licensed.

In 3 lists

OpenPilot

Operating system for robotics. Currently upgrades driver assistance systems on 300+ supported cars. End-to-end autonomous driving stack with open-source hardware and software. MIT licensed.

In 5 listsDetails

Autoware

World's leading open-source software project for autonomous driving. Complete stack from localization and object detection to route planning and control. Used by 50+ companies globally. Apache 2.0 licensed.

12. User Interfaces & Self-hosted Platforms

OpenClaw

Local-first personal AI assistant with multi-channel integrations and full agentic task execution.

In 8 listsDetails

AstrBot

Multi-channel personal AI agent assistant and integration framework for messaging platforms, LLMs, and plugins. AGPL-3.0 licensed.

In 4 listsDetails

Open WebUI

Most popular self-hosted ChatGPT-style interface.

In 11 listsDetails

text-generation-webui

Web UI for running local LLMs with multiple backends, extensions, and model formats.

In 7 listsDetails

LibreChat

Feature-packed multi-LLM interface.

In 4 listsDetails

HuggingChat (self-hosted)

Official open-source codebase for HuggingChat.

In 4 listsDetails

Khoj

Self-hostable personal AI assistant for search, chat, automation, and workflows over local and web data.

In 6 listsDetails

Newelle

GNOME/Linux desktop virtual assistant with integrated file editor, global hotkeys, and profile manager.

In 4 listsDetails

NextChat

Light and fast AI assistant supporting Web, iOS, macOS, Android, Linux, and Windows. One-click deploy with multi-model support. MIT licensed.

In 2 lists

big-AGI

AI suite for power users with multi-model "Beam" chats, AI personas, voice, text-to-image, code execution, and PDF import. MIT licensed.

In 2 lists

Morphic

AI-powered search engine with a generative UI. Supports multiple AI providers (OpenAI, Anthropic, Google, Ollama) and search providers (Tavily, SearXNG, Brave). Features smart search modes, widgets, and image/video search. Apache 2.0 licensed.

In 2 lists

Leon

Your open-source personal assistant. Built around tools, context, memory, and agentic execution. Self-hosted, privacy-focused, and extensible. MIT licensed.

In 5 listsDetails

Willow

Open source, local, and self-hosted Amazon Echo/Google Home competitive voice assistant alternative with hardware support. Apache-2.0 licensed.

In 2 lists

CoPaw

Your Personal AI Assistant; easy to install, deploy on your own machine or on the cloud; supports multiple chat apps with easily extensible capabilities. Apache-2.0 licensed.

Smart2Brain

Privacy-focused Obsidian plugin for AI-powered second brain functionality. Chat with your notes using local or remote LLMs including Ollama and OpenAI. MIT licensed.

In 2 lists

Casibase

Open-source enterprise-level AI knowledge base and agent management platform. Supports multiple LLM providers, RAG, and team collaboration. Apache-2.0 licensed.

In 4 listsDetails

BionicGPT

On-prem ChatGPT replacement for teams with assistants, RAG, access controls, auditing, and enterprise deployment features.

In 2 lists

OpenHuman

A local-first personal assistant and memory harness with Obsidian integration, model routing, and automatic integration sync. GPL-3.0 licensed.

In 3 lists

Open-LLM-VTuber

Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D talking face running locally across platforms. MIT licensed.

In 2 lists

AIRI

Self-hosted cyber companion and VTuber platform with Live2D character rendering, real-time voice chat, and game integration capabilities. MIT licensed.

In 2 lists

Next AI Draw.io

Next.js web application that integrates AI capabilities with draw.io diagrams to create and edit diagrams through natural language. Apache-2.0 licensed.

lucinate

Terminal-native AI chat client supporting OpenClaw, Hermes, Ollama, and any OpenAI-compatible backend — with routines, multi-agent switching, and local agent skills, all in the TUI. Apache-2.0 licensed.

In 2 lists

nanobot

Lightweight personal AI agent with a built-in WebUI, model routing, MCP, and multi-channel chat integrations. MIT licensed.

In 6 listsDetails

LifeOS

General-purpose personal AI harness and assistant framework with persistent memory, custom skills, and goal tracking. MIT licensed.

ThoughtDAG

Local-first visual LLM workspace where graph edges determine the context sent to the model, with branching, merging, document extraction, and Ollama or OpenAI-compatible endpoint support. MIT licensed.

In 3 lists

Thursday

Voice assistant started with npx thursday-agent that runs on the user's computer: an OpenAI GPT-Live call stays in conversation while background bots on OpenAI, Anthropic, Google or xAI models work with a shell, a browser, Agent Skills and MCP servers, and calls, memory and files stay in…

Arelis

Windows desktop research assistant that runs a local model through Ollama, searches the web, drives its own browser, and asks for approval before writing files or sending messages. AGPL-3.0 licensed.

AnythingLLM

All-in-one RAG + agents platform.

In 8 listsDetails

Flowise

Drag-and-drop LLM app builder.

In 12 listsDetails

LocalAI

Open-source AI engine running LLMs, vision, voice, image, and video models on any hardware. Self-hosted OpenAI-compatible API. MIT licensed.

In 14 listsDetails

Onyx

Full-featured AI platform with Chat, RAG, Agents, and Actions. 40+ document connectors and every LLM support. MIT licensed (Community Edition).

In 5 listsDetails

biniou

Self-hosted webUI for 30+ generative AI models. Generate multimedia content with AI on your own computer, even without dedicated GPU (8GB RAM minimum). Works offline once deployed. GPL-3.0 licensed.

Plane

Open-source Jira, Linear, Monday, and ClickUp alternative. AI-powered project management platform with intelligent task triage, sprint planning, and automated workflows. AGPL-3.0 licensed.

In 4 listsDetails

RAG Web UI

Intelligent dialogue system based on RAG technology. Build intelligent Q&A systems on your own knowledge base with modern web interface. Apache-2.0 licensed.

LibreTranslate

Self-hosted machine translation API powered by the Argos Translate engine. AGPL-3.0 licensed.

In 3 lists

Buzz

Self-hostable workspace and Nostr relay implementation where human team members and AI agents collaborate in shared channels, canvases, and workflows. Apache-2.0 licensed.

In 2 lists

Macro

Unified team workspace combining email, messaging, documents, tasks, CRM, and AI agents with shared memory. AGPL-3.0 licensed.

ODS

Apache-2.0-licensed self-hosted local AI stack for inference, chat, voice, agents, workflows, and RAG.

Self-Hosted AI Stack

MIT-licensed Docker Compose stack for private AI with Ollama, a LiteLLM gateway, chat/RAG, voice, and MCP tooling.

Octop (Tencent Cloud)

Self-hosted multi-user, multi-agent AI assistant platform featuring long-term memory, MCP tool integration, and sandboxed execution. MIT licensed.

In 2 lists

ENZO

Self-hosted AI workspace combining agents, skills, and tools (Gmail, Calendar, file conversion) that runs entirely on your own provider API keys, sealed in the browser so the server stores none of them. Apache-2.0 licensed.

In 3 lists

OpenBot

Self-hosted platform for running a team of persistent AI bots that use tools and MCP servers, keep long-term memory, hand work to each other in shared threads, and can pause to ask a human for approval. MIT licensed.

InvoiceFlowAI

Open-source Windows/macOS desktop app that collects emailed PDF/OFD/XML invoices, uses OCR with human review, and exports Excel summaries; optional model providers support extraction.

In 2 lists

AI Language Partner

Local-first Expo and FastAPI Japanese practice app for Korean speakers that uses Whisper-compatible STT, local TTS, and embedding-based dialogue matching without a runtime LLM.

In 2 lists

Jan

Local-first AI app framework.

In 8 listsDetails

Cherry Studio

AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs. AGPL-3.0 licensed.

In 4 listsDetails

DeepChat

A smart assistant that connects powerful AI to your personal world. Built-in MCP and ACP support, multiple search engines, privacy-focused with local data storage. Apache-2.0 licensed.

In 3 lists

SillyTavern

Highly customizable role-playing frontend.

In 4 listsDetails

ChatALL

Concurrently chat with multiple AI bots to discover the best answers. Desktop app for comparing ChatGPT, Claude, Gemini, and 20+ LLMs side-by-side. Apache 2.0 licensed.

In 3 lists

aiFetchly

Open-source desktop AI agent for business automation: lead generation, knowledge library RAG, outreach, and scheduled workflows on Windows, macOS, and Linux.

Chatbox

Powerful desktop AI client for ChatGPT, Claude, and other LLMs. Cross-platform with modern UI. GPLv3 licensed (Community Edition).

In 3 lists

Maid

Free and open-source Android app for interfacing with llama.cpp models locally and remote APIs (Anthropic, DeepSeek, Mistral, Ollama, OpenAI). MIT licensed.

In 2 lists

Dive

Open-source MCP Host Desktop Application with dual Tauri/Electron architecture. Seamlessly integrates with any LLMs supporting function calling. MIT licensed.

PocketPal AI

Open-source app that brings small language models directly to your phone. Run AI 100% privately on iOS and Android with no cloud required. MIT licensed.

Scowld

Native iOS AI voice companion with an animated VRM character, voice and text chat, local conversation history, optional camera context, and BYOK AI, speech-to-text, and text-to-speech providers. MIT licensed.

Hermes Desktop

Desktop companion application for installing, configuring, and chatting with Hermes Agent. MIT licensed.

DSH Studio

Cross-platform desktop application for installing, running, health-checking, and supervising DeepSeek Harness locally. MIT licensed.

In 6 listsDetails

Craft Agents

Open-source desktop and web interface for agentic workflows, featuring built-in MCP support, customizable API connections, and session sharing. Apache-2.0 licensed.

Meetily

Privacy-first local AI meeting assistant for real-time transcription and summary generation. MIT licensed.

In 2 lists

oats

Open-source local-first macOS meeting-notes app with live transcription, speaker labels, AI summaries, and optional fully offline on-device mode. MIT licensed.

In 2 lists

OpenSuperWhisper

macOS dictation application providing real-time audio transcription and drag-and-drop file transcription using Whisper and Parakeet models. MIT licensed.

In 2 lists

T3 Code

Minimal web GUI and desktop app for interacting with coding agents like Codex, Claude Code, Cursor, and OpenCode. MIT licensed.

In 2 lists

OpenWork

Open-source desktop app for sharing AI workflows, skills, and MCP capabilities across agents. MIT licensed.

In 4 listsDetails

ego lite

Desktop browser designed for running web automation tasks and AI agents in parallel within isolated workspaces. MIT licensed.

In 2 lists

Off Grid AI Desktop

Local-first macOS AI app that runs LLM chat, image generation, voice transcription, and personal memory/RAG fully on-device via llama.cpp - nothing leaves the machine. AGPL-3.0 licensed.

Clips Kitty

Local-first Windows desktop app that turns long videos and livestreams into vertical clips, using faster-whisper for word-level captions and a local LLM through Ollama to pick highlights. AGPL-3.0 licensed.

Voz

macOS dictation app that types transcribed speech at the cursor in any application, with whisper.cpp and the Whisper model bundled so it runs offline with no setup. Also records meetings on-device and generates summaries through a bundled llama.cpp. GPL-3.0 licensed.

Codex-X

Cross-platform desktop management application for OpenAI Codex Desktop and CLI with provider switching, local failover routing, session history management, MCP configuration, and token usage analytics. MIT licensed.

PersonalJarvis

Desktop voice and chat assistant for Windows, macOS, and Linux with a custom wake phrase, local speech-to-text, and a workspace that runs CLI coding agents side by side in terminal panes, working with any single model provider key or local models via Ollama. Apache-2.0 licensed.

Speech to Speech

Low-latency, fully modular voice-agent pipeline integrating voice activity detection, speech-to-text, LLMs, and text-to-speech via an OpenAI Realtime-compatible WebSocket API. Apache 2.0 licensed.

In 2 lists

LiveKit Agents

Framework for building realtime voice AI agents with WebRTC transport, STT-LLM-TTS pipelines, and production-grade orchestration. Used by Salesforce Agentforce and Tesla. Apache-2.0 licensed.

In 5 listsDetails

Pipecat

Open-source framework for voice and multimodal conversational AI. Build real-time voice agents with support for speech-to-text, LLMs, text-to-speech, and live video. BSD-2-Clause licensed.

In 7 listsDetails

Agent Chat UI

Web app for interacting with any LangGraph agent (Python & TypeScript) via a chat interface. Stream messages, handle interruptions, and view agent state. MIT licensed.

13. Developer Tools & Integrations

Claude Pet

Desktop companion for Claude Code that reacts to hook events with animated emotions and sounds, maintains a local SQLite graph memory per project injected as a compact recap at session start, promotes repeated patterns into SKILL.md files, and monitors GitHub repository activity.…

In 2 lists

Ralph

Autonomous AI development loop for Claude Code with intelligent exit detection. Automates iterative coding workflows with self-monitoring capabilities. MIT licensed.

In 3 lists

Nimbalyst

Desktop app for running multiple Codex and Claude Code AI sessions in parallel Git worktrees. Test, compare approaches and manage AI-assisted development workflows in one unified interface. MIT licensed.

In 3 lists

Yardlet

Local-first terminal workbench that runs existing Codex, Claude Code, or custom agent CLIs through a deterministic task queue, routing layer, evaluator, and handoff records. MIT licensed.

In 2 lists

Nezha

Code editor for the AI agents era. Run multiple Claude Code and Codex agents across projects on your machine with an intuitive interface. GPL-3.0 licensed.

In 2 lists

Aider Desk

Platform for AI-powered software engineers. Desktop application that enhances the aider terminal experience with a modern UI. Apache 2.0 licensed.

Zed

High-performance, multiplayer code editor with built-in AI features. From the creators of Atom and Tree-sitter. Native AI agentic editing with support for any LLM provider. GPL licensed.

In 6 listsDetails

Code Server

Run VS Code on any machine anywhere and access it in the browser. Self-hosted cloud IDE with full extension support. MIT licensed.

In 5 listsDetails

Gitpod

Cloud development environment platform with automated prebuilds, ephemeral workspaces, and support for any IDE. Self-hostable with open-source core. AGPL-3.0 licensed.

In 2 lists

Onlook

Open-source AI-first design and React editing environment for visually building and modifying frontend applications.

Daytona

Secure elastic infrastructure for running AI-generated code. Self-hosted alternative to GitHub Codespaces with support for multiple IDEs, prebuilds, and any cloud provider. AGPL-3.0 licensed.

In 2 lists

AI Workdeck

Open-source AI-native IDE workspace for legal and document-heavy workflows — "VS Code for lawyers." Self-hosted with MCP agent orchestration, OCR, due-diligence risk flagging, evidence-chain management, WPS WebOffice integration, smart clipboard. Supports air-gapped deployment with Ollama + local…

Orca

An agentic development environment (ADE) for running and orchestrating coding agents in parallel Git worktrees on desktop and mobile. MIT licensed.

In 6 listsDetails

Terax

Lightweight terminal-first AI-native dev workspace (ADE) featuring multi-tab terminals, a code editor with AI edit diffs, source control, and agentic workflows. Apache 2.0 licensed.

Pi Web

Web UI and local workspace for the pi coding agent with session browsing, file previews, and model configuration. MIT licensed.

Garcon

Browser and mobile workspace for running and steering parallel Claude Code, Codex, Cursor Agent, OpenCode, Amp, Droid, and Pi sessions, with integrated terminal, file editing, diff review, Git/PR workflows, mobile approvals, scheduling, and cross-agent transfers. GPL-3.0 licensed.

In 3 lists

screenshot-to-code

MIT-licensed tool that turns screenshots, mockups, and Figma designs into frontend code.

In 4 listsDetails

Atlas

Local-first agentic development environment and source-control system that connects Git commits with agent sessions, tool calls, reasoning traces, and shared memory.

In 2 lists

Continue

Open-source AI coding autopilot for VS Code & JetBrains.

In 10 listsDetails

Tabby

Self-hosted AI coding assistant.

In 7 listsDetails

Cline

Open-source IDE coding agent that can edit files, run commands, and use tools with user approval.

In 7 listsDetails

Open Interpreter

Lets LLMs run code locally.

In 9 listsDetails

Aider

Terminal-based AI pair programmer. Edit code in your local editor and aider implements the changes. Supports multiple LLMs, voice coding, and automatic Git commits. Top scores on SWE Bench. Apache 2.0 licensed.

In 5 listsDetails

Kimi Code CLI

AI coding agent that runs in your terminal to edit code, execute shell commands, and manage Model Context Protocol (MCP) servers and skills. MIT licensed.

In 3 lists

Qwen Code

Open-source AI agent for the terminal, optimized for Qwen series models. Multi-protocol provider support including OpenAI, Anthropic, Gemini, Alibaba Cloud, OpenRouter. Features agentic workflow with Skills and SubAgents. Apache 2.0 licensed.

In 3 lists

Forge

Open-source terminal AI coding agent (Rust TUI) that unifies agent, code editor, and shell in one keyboard-driven workspace with durable SQLite session journals, MCP support, and approval-aware command execution. MIT licensed.

In 2 lists

DeepCode

Transforms research papers and natural language into production-ready code. AI-powered research-to-code automation tool. MIT licensed.

In 5 listsDetails

OpenSpec

Spec-driven development (SDD) framework and CLI tool for AI coding assistants, providing structured workflows for proposing, implementing, and archiving code changes. MIT licensed.

In 2 lists

Spec Kit

Toolkit and CLI for Spec-Driven Development that integrates with AI coding agents to generate structured code implementations from executable specifications. MIT licensed.

In 3 lists

SpecJudge

CLI that reads a project's specification files and recommends which LLM fits the work, using a local judge model via Ollama and citing the spec fragment behind each demand level. MIT licensed.

Open Code Review

AI-powered code review CLI tool combining deterministic analysis pipelines with LLM agents for precise, line-level feedback. Apache-2.0 licensed.

In 4 listsDetails

Open Notebook

Open-source implementation of Notebook LM with multi-modal content support (PDFs, videos, audio, web pages). Features multi-speaker podcast generation, 18+ AI provider integrations, and full-text + vector search. Self-hosted with complete data sovereignty. MIT licensed.

In 2 lists

Deta Surf

Personal AI notebook for organizing files and webpages with AI-generated notes. Local-first data storage, open data formats, and open model choice including local models. Cross-platform desktop app for research and thinking workflows. Apache 2.0 licensed.

In 2 lists

FilePilot AI

Local-first file intelligence app, CLI, and MCP server for searching, summarizing, tagging, deduplicating, and safely organizing local files for AI agent workflows. MIT licensed.

Quarto

Open-source scientific and technical publishing system built on Pandoc. Create dynamic content with Python, R, Julia, and Observable. MIT licensed.

Drawdata

Draw datasets from within Python notebooks. Interactive data visualization tool for creating and editing datasets directly in Jupyter environments. MIT licensed.

Deepnote

Drop-in replacement for Jupyter with AI-first design, sleek UI, and native data integrations. Use Python, R, and SQL locally, then scale to Deepnote cloud for collaboration and deployable data apps. Apache 2.0 licensed.

In 6 listsDetails

Zasper

High-performance IDE for Jupyter Notebooks built with Go. Up to 5x less CPU and 40x less RAM than JupyterLab. Implements Jupyter's wire protocol with massive concurrency support. AGPL-3.0 licensed.

In 3 lists

Archify

Diagram-generation agent skill for Claude Code, Codex, and OpenCode that converts natural language into interactive system, workflow, sequence, data flow, and lifecycle diagrams. MIT licensed.

In 4 listsDetails

Graphify

AI coding assistant skill that maps a codebase and resource directory into an interactive knowledge graph for local query and navigation. MIT licensed.

In 4 listsDetails

llama.vim

Local LLM-powered code completion plugin for Vim/Neovim using llama.cpp. Fast, privacy-first, no API key needed.

CodeCompanion.nvim

AI-powered coding assistant for Neovim. Inline code generation, chat, actions, and tool use with support for multiple LLM providers.

In 4 listsDetails

ProxyAI

Leading open-source AI copilot for JetBrains IDEs. Connect to any model in any environment with auto-apply, image chat, file references, web search, and customizable personas. Apache 2.0 licensed.

In 2 lists

avante.nvim

Neovim plugin that brings Cursor-like AI IDE features to Vim. Edit code with natural language, generate code from context, and chat with AI about your codebase. Apache 2.0 licensed.

In 3 lists

Serena

Powerful MCP toolkit for coding agents providing semantic retrieval and editing capabilities. Integrates language servers for IDE-level code understanding. MIT licensed.

In 4 listsDetails

windsurf.vim

Free, ultrafast Copilot alternative for Vim and Neovim. AI-powered code completion with low latency and large context window. MIT licensed.

Jupyter AI

Chat and code generation inside notebooks.

In 5 listsDetails

Minuet AI

Neovim plugin offering code completion as-you-type from popular LLMs including OpenAI, Gemini, Claude, Ollama, Llama.cpp, Codestral, and more. GPL-3.0 licensed.

In 3 lists

Peekaboo

macOS CLI & MCP server enabling AI agents to capture screenshots and automate UI interactions. Visual question answering through local or remote AI models. MIT licensed.

In 2 lists

Skills

A collection of custom productivity and engineering skills for Claude Code and other AI coding agents. MIT licensed.

In 2 lists

planning-with-files

Persistent file-based planning skill for AI coding agents to survive context loss and restarts with a deterministic completion gate. MIT licensed.

Understand Anything

Turns subdirectories, codebases, or knowledge bases into an interactive, AI-navigable knowledge graph and dashboard. MIT licensed.

In 2 lists

Knowledge Work Plugins (Anthropic)

Collection of role-specific plugins and sub-agents for Claude Cowork and Claude Code. Apache 2.0 licensed.

In 2 lists

Cursor Plugins

Official Cursor plugins for developer tools, workflows, and SaaS integrations. MIT licensed.

Claude Plugins Marketplace

Official, Anthropic-managed directory of high-quality plugins for Claude Code. Apache 2.0 licensed.

In 3 lists

Agent Toolkit for AWS

Official AWS-supported MCP servers, skills, and plugins to help AI coding agents build, deploy, and manage applications on AWS. Apache 2.0 licensed.

In 2 lists

Caveman

Skill and setup script for Claude Code, Codex, Gemini, Cursor, and other coding agents that reduces token output by instructing agents to communicate in concise, caveman-style prose. MIT licensed.

In 3 lists

Chrome DevTools MCP

Official Model Context Protocol (MCP) server from Google that enables coding agents to control and inspect a live Chrome browser for automation, debugging, and web performance analysis. Apache 2.0 licensed.

In 3 lists

Codex Plugin for Claude Code

Official integration plugin for Claude Code that enables developers to run Codex code reviews, delegate tasks, and manage background coding jobs. Apache 2.0 licensed.

In 3 lists

Compound Engineering

Official AI engineering plugin for Claude Code, Codex, Cursor, and other coding agents that implements iterative brainstorm, plan, work, review, and compound workflows. MIT licensed.

In 3 lists

Unity MCP

Model Context Protocol (MCP) server bridging AI assistants with the Unity Editor to automate asset, scene, and script workflows. MIT licensed.

In 2 lists

.NET Agent Skills

Curated set of skills, custom agents, and plugins to assist AI coding agents with .NET and C# development. MIT licensed.

Claude Video

A skill and plugin that enables AI coding assistants to view, transcribe, and analyze video content. MIT licensed.

In 2 lists

Stitch Design Skills

A library of agent skills and plugins designed to work with the Stitch MCP server, following the Agent Skills open standard. Apache 2.0 licensed.

Impeccable

Design guidance skill, CLI tool, and detector rules to improve user interface design generated by AI coding agents. Apache 2.0 licensed.

In 2 lists

Hallmark

Anti-AI-slop design skill and ruleset for Claude Code, Cursor, and Codex to prevent generic, AI-generated layouts. MIT licensed.

In 2 lists

Code Review Graph

Local-first code intelligence graph that builds a persistent map of a codebase to provide token-optimized context for AI coding tools. MIT licensed.

In 2 lists

CAD Skills

A library of modular agent skills for generating, inspecting, slicing, and exporting CAD and robot-description geometry (STEP, STL, URDF, SDF). MIT licensed.

In 2 lists

ADHD

Tree-of-thought divergent reasoning skill for AI agents that spawns parallel cognitive frames to prune traps and score surviving paths. MIT licensed.

MCP Lens

DeepSeek Harness plugin that exposes configured remote MCP tools through a search interface and an explicit server/tool call interface with exact input schemas. MIT licensed.

In 4 listsDetails

jev-use

Claude Code, Codex and pi plugin that routes agent-loop steps needing no text output to a typed judgment model, returning anything it should not decide to the LLM. MIT licensed.

In 3 lists

Assistant UI

React/TypeScript library for building production-grade AI chat interfaces. Drop-in components for streaming messages, tool calls, and multi-modal inputs.

Deep Chat

Fully customizable AI chatbot component for your website. Supports OpenAI, direct API services, and custom endpoints. MIT licensed.

In 4 listsDetails

Markstream

Multi-framework streaming Markdown renderer for AI chat interfaces, with incomplete Markdown handling, Mermaid, KaTeX, Shiki/Monaco code blocks, SSR, and packages for Vue, React, Svelte, and Angular.

In 9 listsDetails

CopilotKit

Best-in-class SDK for building full-stack agentic applications, Generative UI, and chat applications. Creators of the AG-UI Protocol adopted by Google, LangChain, AWS, and Microsoft. MIT licensed.

In 4 listsDetails

json-render

Generative UI framework for rendering dynamic, type-safe interfaces from structured JSON streamed by LLMs and agents. Apache 2.0 licensed.

In 2 lists

Ruler

Central AI agent rule registry. Manages and distributes rules for AI coding agents across projects. MIT licensed.

In 2 lists

PR-Agent (Qodo)

AI-powered code review agent for GitHub, GitLab, Bitbucket, and Azure DevOps. Automated PR analysis, improvement suggestions, and multi-platform deployment via CLI, GitHub Actions, or webhooks. AGPL-3.0 licensed.

In 3 lists

LLM (Simon Willison)

CLI tool and Python library for interacting with dozens of LLMs via remote APIs or locally. Extensible plugin ecosystem, SQLite logging. Apache 2.0 licensed.

In 3 lists

aicommits

CLI that writes your Git commit messages for you with AI. Never write a commit message again. Supports multiple providers including OpenAI, Groq, xAI, Ollama, and LM Studio. MIT licensed.

In 5 listsDetails

Codex CLI

OpenAI's lightweight coding agent that runs in your terminal. Code generation, file editing, and command execution with approval. Apache 2.0 licensed.

In 9 listsDetails

VibePod

Unified CLI for running Claude Code, Codex, OpenCode, and other coding agents in isolated Docker or Podman containers, with local HTTP traffic and token metrics plus a side-by-side comparison dashboard. MIT licensed.

In 2 lists

Repomix

Powerful tool that packs your entire repository into a single AI-friendly file. Perfect for feeding codebases to LLMs with smart filtering and token counting. MIT licensed.

In 4 listsDetails

DevProjex

Open-source GUI, TUI, CLI, and read-only MCP app for selecting project files, previewing, redacting secrets, and packing token-counted context with Git scopes and syntax-aware compression. Apache-2.0 licensed.

In 2 lists

GitIngest

Replace 'hub' with 'ingest' in any GitHub URL to get a prompt-friendly extract of a codebase. Optimized for Python ecosystem and data science workflows. MIT licensed.

GitDiagram

Converts any GitHub repository into an interactive system architecture diagram using LLMs, with clickable source links and Mermaid export. MIT licensed.

Instructor

Python library for extracting structured, validated data from LLMs using Pydantic models. Handles validation, retries, and error handling with 15+ provider support. MIT licensed.

In 4 listsDetails

Mirascope

Python toolkit for building LLM applications with automatic versioning, tracing, and cost tracking. The "LLM Anti-Framework" for developers who want control. MIT licensed.

In 3 lists

Context7

Up-to-date code documentation for LLMs and AI code editors. Fetches latest docs and code examples directly into LLM context via MCP. Eliminates hallucinated APIs. MIT licensed.

In 6 listsDetails

Claude Squad

Manage multiple AI terminal agents like Claude Code, Codex, OpenCode, and Amp. Terminal multiplexer for AI coding agents with session management and parallel execution. AGPL-3.0 licensed.

In 5 listsDetails

YYLO

Command-line orchestrator that runs AI coding agents (Claude Code, Codex, Gemini CLI) in parallel across isolated git worktrees, tracked on a Kanban board with typed merge and release flows. MIT licensed.

In 6 listsDetails

Herdr

Terminal agent multiplexer and workspace manager with mouse-native split-panes and automatic agent state detection. AGPL-3.0 licensed.

In 3 lists

agent-manager

Terminal UI that runs coding-agent CLIs such as Claude Code, Codex, OpenCode, and Gemini CLI side by side, each in its own persistent tmux session, with live status, optional per-session Git worktrees, and a diff review that sends line comments back to the agent. Apache-2.0 licensed.

In 3 lists

agenttrace

Local-first TUI for observing AI coding agent sessions across Claude Code, Codex CLI, Gemini CLI, Aider, Cursor exports, OpenCode, and more.

In 5 listsDetails

agentsview

Local-first session intelligence and cost analytics dashboard for AI coding agents, supporting Claude Code, Codex, and other tools. MIT licensed.

In 3 lists

Uni-CLI

Self-repairing CLI catalog that exposes web, desktop, Electron, and bridge tools as deterministic commands for AI agents.

OpenChamber Mobile Bridge

OpenCode/OpenChamber helper for exposing devcontainer-based coding sessions to private Tailscale mobile access through a host bridge. MIT licensed.

DesktopCommander MCP

MCP server for Claude providing terminal control, file system search, and diff file editing capabilities. Enables autonomous code editing through Model Context Protocol. MIT licensed.

Mobile MCP

MCP server that lets AI agents automate and inspect iOS and Android simulators, emulators, and real devices through accessibility trees, screenshots, and structured device tools. Apache-2.0 licensed.

In 2 lists

Claude Code Action

GitHub Action for running Claude Code in PR and issue workflows with approval-aware automation and coding assistance.

In 3 lists

OfficeCLI

Office suite purpose-built for AI agents to read, edit, and automate Word, Excel, and PowerPoint files without external Office installations. Apache 2.0 licensed.

CLI-Anything

Framework for converting software applications into agent-native command-line interfaces for AI coding agents. Apache 2.0 licensed.

In 3 lists

MulmoTerminal

Browser grid of live Claude Code and Codex sessions started with one npx command. Each cell is a real PTY with a colour-coded status, tmux-backed persistence, and a git worktree per cell. For Claude Code, needs-you is shown separately from done, read from the CLI's own hooks. MIT licensed.

In 3 lists

Worktrunk

CLI for Git worktree management and workflow automation designed for parallel AI coding-agent workflows.

In 3 lists

Vercel AI SDK

Provider-agnostic TypeScript toolkit for building AI-powered applications and agents. Unified API for OpenAI, Anthropic, Google, and 20+ providers with first-class streaming, tool-calling, and structured output support. Apache 2.0 licensed.

In 8 listsDetails

GitHub Copilot SDK

Multi-platform SDK for integrating GitHub Copilot Agent into apps and services. Production-tested agent runtime with planning, tool invocation, and context management. Build Copilot-style agents without writing your own orchestration. MIT licensed.

In 2 lists

IBM MCP Context Forge

Gateway and registry for MCP/A2A/REST APIs with unified discovery, routing, and guardrails for production agent integrations.

In 2 lists

Fern

Open-source SDK generator for REST APIs. Generate type-safe API clients in TypeScript, Python, Go, Java, and more from OpenAPI specs. Powers SDKs for companies like OpenAI, Anthropic, and Cloudflare. Apache 2.0 licensed.

In 2 lists

Cortex

Generates typed SDKs, API documentation, and MCP servers from OpenAPI, AsyncAPI, GraphQL, Protocol Buffer, OpenRPC, and Markdown sources. MIT licensed.

In 6 listsDetails

no-mistakes

A local Git proxy and validation pipeline that runs AI-driven checks and applies fixes in a temporary worktree before forwarding pushes and opening clean PRs. MIT licensed.

In 2 lists

oai-smoke

MIT-licensed, standard-library-only Go CLI that validates /models and opt-in /chat/completions behavior for OpenAI-compatible APIs and emits bounded diagnostics without credentials or response bodies.

Helicone

Open-source LLM observability with request logging, caching, rate limiting, and cost analytics.

In 8 listsDetails

GEPA

Reflective prompt evolution optimizer using natural language reflection and Pareto frontier learning. Outperforms reinforcement learning for prompt optimization. Integrated with DSPY and MLflow. MIT licensed.

In 3 lists

Entroly

Local-first MCP server for explicit-budget context selection, content-addressed exact recovery, and auditable Context Receipts. Apache 2.0 licensed.

In 5 listsDetails

Vibe-Coding Prompt Template

Staged prompt workflow and CLI that turns a product idea into a PRD, technical design, and AGENTS.md instruction files for AI coding agents. MIT licensed.

In 4 listsDetails

Webcmd

Self-learning browser infrastructure for AI agents that compiles a site's navigation into deterministic per-site CLI commands (Apache-2.0, TypeScript).

In 4 listsDetails

14. Resources & Learning

Papers with Code

Definitive database linking papers to open code and datasets.

In 3 lists

Hugging Face Papers

Daily-updated feed of the latest arXiv papers with open weights.

In 2 lists

Open LLM Leaderboard (Hugging Face)

Real-time ranking of open models.

Hugging Face Discussions

Largest open AI forum.

In 2 lists

AI Engineering from Scratch (rohitg00)

Comprehensive curriculum covering machine learning, deep learning, NLP, computer vision, and agents by implementing them from scratch. MIT licensed.

In 3 lists

AMD Strix Halo Local LLM Guide

Reproducible Ubuntu, Ollama, llama.cpp, Vulkan/RADV, and ROCm setup and benchmark evidence for Ryzen AI MAX+ 395 local-AI systems. MIT licensed.

AI Agents in Depth

Open-source textbook and code repository covering AI agent design principles, architectures, and engineering practice. Apache 2.0 licensed.

In 4 listsDetails

Claude How To

Comprehensive learning path and template guide for Claude Code, covering setup, hooks, custom skills, and MCP server integrations. MIT licensed.

In 3 lists

Maths, CS & AI Compendium

An open, unconventional textbook covering mathematics, computer science, and artificial intelligence from the ground up. Apache 2.0 licensed.

In 2 lists

r/LocalLLaMA

Go-to subreddit for local/open-source LLM topics.

In 3 lists

Hugging Face Course

Free hands-on courses using only open models.

In 3 lists

ML For Beginners (Microsoft)

12-week, 26-lesson, 52-quiz classic machine learning course for beginners. Comprehensive curriculum covering regression, classification, clustering, and NLP with practical projects.

In 3 lists

AI For Beginners (Microsoft)

12-week, 24-lesson curriculum on Artificial Intelligence. Covers symbolic AI, neural networks, computer vision, NLP, and reinforcement learning with hands-on labs.

In 4 listsDetails

Generative AI for Beginners (Microsoft)

21 lessons covering generative AI fundamentals, prompt engineering, RAG applications, fine-tuning, and LLM app deployment with practical exercises.

In 5 listsDetails

LangChain Academy

Free courses on agents and RAG.

In 2 lists

Data Science for Beginners (Microsoft)

10-week, 20-lesson curriculum on data science fundamentals. Covers data preparation, visualization, modeling, and deployment with practical projects.

In 4 listsDetails

The Incredible PyTorch

Curated list of PyTorch tutorials, papers, projects, and communities for deep learning researchers.

In 2 lists

Deep RL Class (Hugging Face)

Free deep reinforcement learning course with hands-on exercises and trained agent publishing to the Hugging Face Hub.

In 2 lists

Practical RL (Yandex Data School)

Comprehensive reinforcement learning course covering RL fundamentals, deep RL, policy gradients, actor-critic methods, and practical applications in the wild. The Unlicense.

In 2 lists

NLP Course (Yandex Data School)

YSDA course in Natural Language Processing with 2025 materials covering text classification, language models, transformers, and modern NLP techniques. MIT licensed.

In 4 listsDetails

Large Language Model Notebooks Course

Practical hands-on course about Large Language Models and their applications. Covers Chatbots, Code Generation, OpenAI API, Hugging Face, Vector databases, LangChain, Fine Tuning, PEFT, LoRA, QLoRA. MIT licensed.

In 2 lists

Transformers Tutorials (Niels Rogge)

Comprehensive tutorials and demos using the Hugging Face Transformers library for NLP, vision, and multimodal tasks.

In 3 lists

AI Engineering Hub

93+ production-ready projects with in-depth tutorials on LLMs, RAG, and real-world AI agent applications. Comprehensive resources for all skill levels from beginner to advanced. MIT licensed.

In 3 lists

Complete Agentic AI Engineering Course

6-week comprehensive course on Agentic AI covering autonomous agents, multi-agent systems, and practical agent development. MIT licensed.

OpenMAIC

MIT-licensed multi-agent platform for interactive learning and course generation.

Awesome LLM Apps

A collection of run-ready artificial intelligence templates and agent customizer scripts, featuring over 100 retrieval-augmented generation (RAG) and intelligent agent blueprints. Apache 2.0 licensed.

In 3 lists

Claude Cookbooks

Official collection of recipes and notebooks demonstrating tool use, prompt caching, structured outputs, and agentic workflows with Claude. MIT licensed.

In 6 listsDetails

Hugging Face Transformers Notebooks

Run Transformers, Datasets, and more in Colab.

In 2 lists

TensorFlow Tutorials

Official guides for beginners to advanced users.

Awesome Machine Learning

The definitive curated list of machine learning frameworks, libraries and software organized by language. Covers Python, C++, Java, JavaScript, and more with comprehensive coverage of the ML ecosystem. CC0-1.0 licensed.

In 15 listsDetails

Awesome Artificial Intelligence

Curated list of artificial intelligence courses, books, video lectures, and papers for developers and researchers. MIT licensed.

In 7 listsDetails

Andrej Karpathy Skills

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls. Principles: Think Before Coding, Simplicity First, Surgical Changes, Goal-Driven Execution. MIT licensed.

Awesome DESIGN.md

Curated collection of DESIGN.md files representing popular design systems to guide AI coding agents in consistent UI generation. MIT licensed.

In 2 listsDetails

Awesome Claude Skills

Curated list of Claude Skills, plugins, resources, and custom commands to extend terminal and API workflows. Apache 2.0 licensed.

In 4 listsDetails
See category
94

Table of Contents

hesreallyhim/awesome-claude-code

A hand-picked collection of the finest of resources for the most awesome of agents, Claude Code, the undisputed champion of coding companions, from the unstoppable team…

Fresh★ 55k202 entriesPushed today
94

Awesome Agent Skills

VoltAgent/awesome-agent-skills

A curated collection of 1000+ agent skills from official dev teams and the community, compatible with Claude Code, Codex, Gemini CLI, Cursor, and more.

Fresh★ 35k839 entriesPushed today
93

Awesome Machine Learning

josephmisiti/awesome-machine-learning

A curated list of awesome Machine Learning frameworks, libraries and software.

Fresh★ 74k1188 entriesPushed 7 days ago
92

Awesome Production Machine Learning

EthicalML/awesome-production-machine-learning

A curated list of awesome open source libraries to deploy, monitor, version and scale your machine learning

Fresh★ 21k519 entriesPushed 3 days ago
92

AWESOME DATA SCIENCE

academic/awesome-datascience

:memo: An awesome Data Science repository to learn and apply for real world problems.

Fresh★ 30k881 entriesPushed today
91

Static Analysis

analysis-tools-dev/static-analysis

⚙️ A curated list of static analysis (SAST) tools and linters for all programming languages, config files, build tools, and more. The focus is on tools which improve…

Fresh★ 15k528 entriesPushed 8 days ago