Awesome Ai Agents 2026
Section: RAG and Knowledge Bases · Live data RAG. Real-time streaming. 50k+ stars.
Entry
Appears in 8 awesome lists
Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG. Features 350+ connectors with always-in-sync data from SharePoint, Google Drive, S3, Kafka, PostgreSQL and more. BSL 1.1 license (becomes Apache 2.0 after 4 years).
Section: RAG and Knowledge Bases · Live data RAG. Real-time streaming. 50k+ stars.
Section: Stream Processing · Performant open-source Python ETL framework with Rust runtime, supporting 300+ data sources.
Section: Retrieval-Augmented Generation · Python ETL framework for stream processing, real-time analytics, LLM pipelines and RAG
Section: 5. Retrieval-Augmented Generation (RAG) & Knowledge · Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG. Features 350+ connectors with always-in-sync data from SharePoint, Google Drive, S3, Kafka, PostgreSQL and more. BSL 1.1 license (becomes Apache 2.0 after 4 years).
Section: Extract, transform, load (ETL) · Performant open-source Python ETL framework with Rust runtime, supporting 300+ data sources.
Section: Data Ingestion / ETL · Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG.
Section: Data processing · Performant open-source Python ETL framework with Rust runtime, supporting 300+ data sources.
Section: Data Science and Analytics · Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG.
Milvus is a cloud-native, open-source vector database built to manage embedding vectors generated by machine learning models and neural networks.
Mem0 is an intelligent memory layer for Large Language Models that enhances personalized AI experiences by retaining and utilizing contextual information across various applications. github | website | docs | discord | twitter | github profile | linkedin
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search,…
Vector Search Engine and Database for the next generation of AI applications. Also available in the cloud
An open-source embedding database for building AI applications with embeddings and semantic search.
The Pinecone vector database makes it easy to build high-performance vector search applications. Developer-friendly, fully managed, and easily scalable without infrastructure hassles.