Skip to content

Entry

SGLang

Appears in 9 awesome lists

(MPL-2.0) allows specifying JSON schemas using regular expressions or Pydantic models for constrained decoding. Its high-performance runtime accelerates JSON decoding.

Open github.comsgl-project/sglang

Found in these lists

awesome-ChatGPT-repositories

Section: Langchain · SGLang is a fast serving framework for large language models and vision language models.

FreshScore 87

Awesome LLM JSON List

Section: Python Libraries · (MPL-2.0) allows specifying JSON schemas using regular expressions or Pydantic models for constrained decoding. Its high-performance runtime accelerates JSON decoding.

StaleScore 55

Awesome LLM Resources

Section: 推理 Inference · SGLang is yet another fast serving framework for large language models and vision language models.

FreshScore 87

Awesome local LLM

Section: Inference engines · a fast serving framework for large language models and vision language models

FreshScore 87

awesome-nlp

Section: Efficient and Small Language Models · structured generation and efficient serving.

FreshScore 90

Awesome Open Source AI

Section: 3. Inference Engines & Serving · Next-gen serving framework with RadixAttention. Powers xAI's production workloads at 100K+ GPUs scale.

FreshScore 89

Awesome Production Machine Learning

Section: Deployment and Serving · SGLang is a fast serving framework for large language models and vision language models.

FreshScore 92

Awesome Python

Section: AI and Agents · A high-performance serving framework for large language models and multimodal models.

FreshScore 94

awesome-python

Section: Other · SGLang is a high-performance serving framework for large language models and multimodal models.

FreshScore 81

LangChain

Langchain integrates various providers like Anthropic, AWS, and OpenAI, and offers tools for components such as LLMs, chat models, and data analysis, supporting functionalities from Alpha Vantage to YouTube github | docs

In 20 listsDetails

LiteLLM

Unified proxy and SDK that routes to 100+ LLM providers behind a single OpenAI-compatible interface, with a Router handling retry/fallback across deployments, per-project cost and rate-limit tracking, and OTEL callback integrations. The right infrastructure layer when your harness needs provider…

In 16 listsDetails

LlamaIndex

(MIT) provides modules for structured outputs at different levels of abstraction, including output parsers for text completion endpoints, Pydantic programs for mapping prompts to structured outputs using function calling or output parsing, and pre-defined Pydantic programs for specific output types.

In 14 listsDetails

LocalAI

robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video,…

In 14 listsDetails

Mem0

Mem0 is an intelligent memory layer for Large Language Models that enhances personalized AI experiences by retaining and utilizing contextual information across various applications. github | website | docs | discord | twitter | github profile | linkedin

In 13 listsDetails

PydanticAI

June 2026 harness-first redesign built around the Capability primitive: a single composable unit bundling instructions, tools, lifecycle hooks, and model settings. The split between a small stable core and a fast-moving pydantic-ai-harness lets capabilities graduate as they prove essential, while…

In 12 listsDetails

Ollama

Ollama is a tool for running large language models locally, offering easy setup for macOS, Windows, Linux, and Docker, along with a library of models and quickstart guides for customization and integration github | github profile

In 12 listsDetails

vLLM

State-of-the-art serving engine with PagedAttention and continuous batching. Currently the fastest production-grade LLM server.

In 11 listsDetails