awesome-ChatGPT-repositories
Section: Langchain · SGLang is a fast serving framework for large language models and vision language models.
Entry
Appears in 9 awesome lists
(MPL-2.0) allows specifying JSON schemas using regular expressions or Pydantic models for constrained decoding. Its high-performance runtime accelerates JSON decoding.
Section: Langchain · SGLang is a fast serving framework for large language models and vision language models.
Section: Python Libraries · (MPL-2.0) allows specifying JSON schemas using regular expressions or Pydantic models for constrained decoding. Its high-performance runtime accelerates JSON decoding.
Section: 推理 Inference · SGLang is yet another fast serving framework for large language models and vision language models.
Section: Inference engines · a fast serving framework for large language models and vision language models
Section: Efficient and Small Language Models · structured generation and efficient serving.
Section: 3. Inference Engines & Serving · Next-gen serving framework with RadixAttention. Powers xAI's production workloads at 100K+ GPUs scale.
Section: Deployment and Serving · SGLang is a fast serving framework for large language models and vision language models.
Section: Other · SGLang is a high-performance serving framework for large language models and multimodal models.
Section: AI and Agents · A high-performance serving framework for large language models and multimodal models.
Langchain integrates various providers like Anthropic, AWS, and OpenAI, and offers tools for components such as LLMs, chat models, and data analysis, supporting functionalities from Alpha Vantage to YouTube github | docs
Unified proxy and SDK that routes to 100+ LLM providers behind a single OpenAI-compatible interface, with a Router handling retry/fallback across deployments, per-project cost and rate-limit tracking, and OTEL callback integrations. The right infrastructure layer when your harness needs provider…
(MIT) provides modules for structured outputs at different levels of abstraction, including output parsers for text completion endpoints, Pydantic programs for mapping prompts to structured outputs using function calling or output parsing, and pre-defined Pydantic programs for specific output types.
robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video,…
Mem0 is an intelligent memory layer for Large Language Models that enhances personalized AI experiences by retaining and utilizing contextual information across various applications. github | website | docs | discord | twitter | github profile | linkedin
June 2026 harness-first redesign built around the Capability primitive: a single composable unit bundling instructions, tools, lifecycle hooks, and model settings. The split between a small stable core and a fast-moving pydantic-ai-harness lets capabilities graduate as they prove essential, while…