Explore
Discover the best AI tools in one place
34 tools found
Langfuse is an open-source LLM engineering and observability platform built for teams developing production-grade generative AI applications and autonomous multi-agent pipelines. It captures granular traces across token usage, prompt versions, latency bottlenecks, and retrieval accuracy, giving developers complete visibility into model behavior at runtime. By integrating seamlessly with major AI frameworks such as LangChain, LlamaIndex, LiteLLM, and the OpenAI SDK, Langfuse eliminates the guesswork from debugging complex agent execution trees. Developers can monitor production cost metrics, identify hallucinated responses, and run rigorous continuous evaluation suites on live traffic.
Qdrant is an open-source, high-performance vector database and similarity search engine engineered in Rust for production AI systems, semantic search engines, and Retrieval-Augmented Generation (RAG) pipelines. It provides lightning-fast nearest-neighbor search with rich payload filtering and custom distance metrics. Unlike traditional databases adapted for vectors, Qdrant was designed from day one to handle high-dimensional neural embeddings at scale. Its Rust engine provides memory-efficient vector quantization (scalar, product, and binary), allowing engineering teams to search billions of vectors on cost-effective cloud hardware.
Unstructured is the leading enterprise ETL (Extract, Transform, Load) platform engineered to prepare messy, unstructured business documents for Retrieval-Augmented Generation (RAG) and LLM fine-tuning. Over 80% of enterprise data lives in complex formats like scanned PDFs, PowerPoint decks, Word files, HTML tables, and email threads that break standard text scrapers. Unstructured utilizes specialized computer vision and vision-language models to segment documents into structural semantic elements (titles, paragraphs, headers, embedded tables, and image captions) while preserving exact spatial and hierarchical context. Available as an open-source Python library and a high-throughput serverless cloud API, Unstructured integrates directly with LangChain, LlamaIndex, and major vector databases to power mission-critical enterprise knowledge retrieval.
Vanna AI is an open-source Python-based RAG (Retrieval-Augmented Generation) framework engineered to generate high-accuracy SQL queries from plain English questions. Unlike general-purpose chatbots that frequently hallucinate non-existent database columns and tables, Vanna trains specifically on your database schema, table DDL, documentation, and historical query logs. When a business user or developer asks a natural language question, Vanna retrieves relevant schema definitions and validated SQL examples to construct an accurate, executable SQL query for PostgreSQL, Snowflake, BigQuery, MySQL, SQLite, or SQL Server. With over 12,000 GitHub stars, Vanna allows organizations to deploy self-hosted text-to-SQL agents inside Slack, Streamlit dashboards, or internal REST APIs without exposing private database records to third parties.
Mem0 (formerly Embedchain) is a universal, persistent memory architecture designed to solve the critical context amnesia problem in modern AI applications. While foundational LLMs forget user preferences and past interactions the moment a session ends, Mem0 maintains a continuous, self-improving memory graph across user sessions, agents, and applications. With Mem0, developers can build personalized AI assistants, customer support agents, and autonomous workflow bots that remember user preferences, past project decisions, and communication styles over months and years. Mem0 operates as both an open-source self-hostable Python/TypeScript library and a managed cloud platform, providing sub-100ms vector search, episodic memory extraction, and automated memory consolidation without manual prompt engineering.
Marimo is a modern, open-source reactive Python notebook that solves the reproducibility and out-of-order execution nightmares of traditional Jupyter notebooks. Designed for data scientists, ML engineers, and researchers, Marimo guarantees that notebook state is always consistent by treating code cells as a Directed Acyclic Graph (DAG). When you modify a variable or code cell in Marimo, all dependent downstream cells automatically re-run instantly, eliminating hidden state bugs. Furthermore, Marimo notebooks are stored as pure, version-control-friendly Python scripts rather than messy JSON files. With built-in AI code generation, interactive UI sliders, instant conversion into shareable web apps, and native SQL query cells, Marimo represents the state-of-the-art computational notebook for 2026.
RAGFlow is an open-source enterprise RAG (Retrieval-Augmented Generation) engine based on deep document understanding and multimodal document parsing. Unlike naive RAG systems that slice documents into arbitrary character chunks—scrambling complex tables, footnotes, and multi-column layouts—RAGFlow preserves the original semantic structure of complex documents. Powered by DeepDoc computer vision models, RAGFlow extracts clean text, recognizes complex tables across multiple pages, identifies corporate hierarchies, and visually highlights exact source citations with bounding boxes inside PDF viewers. With self-hosted Docker deployment, low-code workflow orchestration, and native support for local and cloud LLMs, RAGFlow is widely adopted by enterprise organizations requiring zero hallucination in legal, financial, and technical document analysis.
Mermaid Chart helps you visualize complex data with simple charts and diagrams. It is ideal for developers and technical teams.
Sloped is a data analysis tool that converts APIs into easily searchable data sources. It allows users to query and analyze data from various endpoints without complex setup. This tool is perfect for developers and data analysts who need quick access to structured information.
Weaviate is an open-source vector database designed for AI-native applications. It enables developers to store and search data using vector embeddings at scale.
Vespa is an open-source big data serving engine designed for online AI-powered applications at any scale. It enables real-time data processing, search, and recommendation systems using machine learning models. Organizations use it to build low-latency AI applications over large datasets.
Dstack provides a platform for developing and deploying large language models across multiple cloud environments. It simplifies the infrastructure management required for LLM workflows, enabling teams to focus on model development.
H2O.ai provides a platform that combines predictive AI and generative AI capabilities. It enables enterprises to build, deploy, and manage machine learning models at scale. The platform supports both automated and custom model development workflows.
Runcell is an AI agent that integrates directly into Jupyter notebooks. It allows data scientists to generate code, run cells, and debug using natural language commands. Perfect for accelerating data science workflows and reducing boilerplate coding.
RTutor allows users to interact with their datasets through natural language queries, enabling AI-powered data analysis without coding. It simplifies statistical analysis and visualization for researchers and analysts using R.
Teachable Machine is a web-based tool that allows users to quickly create machine learning models without writing any code. It supports training models for recognizing images, sounds, and poses directly in the browser. The tool is designed to make AI accessible to educators, artists, and hobbyists.
K8sGPT uses AI to diagnose and triage issues in Kubernetes clusters. It analyzes cluster data and provides actionable insights to help operators resolve problems quickly.
Gradio is an open-source Python library that allows you to quickly create customizable web interfaces for machine learning models. It enables developers and researchers to demo, debug, and share their models with a simple API. With just a few lines of code, you can launch a fully functional web app.
Abzu provides AI solutions designed for high-stakes environments where accuracy and reliability are paramount. It helps organizations make data-driven decisions with transparent and explainable models.
GitHub Data Explorer uses AI to generate SQL queries for discovering insights from GitHub data. It helps developers and researchers analyze repository trends and patterns.
Metabot AI is a data analysis assistant that helps users explore and visualize their data through conversational interfaces. It simplifies complex data workflows, making analytics accessible to non-technical users.
Beancount.io is an AI-powered accounting tool that lets you manage your finances using plain text, similar to writing code. It provides AI-driven insights to help you understand your spending patterns and financial health. Ideal for developers and finance enthusiasts who prefer a code-based approach to bookkeeping.
Unsloth AI is an open-source web UI designed for training and running open-source language models. It simplifies the process of fine-tuning and deploying LLMs.
Nebius Token Factory provides enterprise-grade open-source AI inference capabilities designed for unlimited scale. It enables businesses to deploy and run large language models efficiently.
KnowledgeGraph GPT transforms unstructured text into structured knowledge graphs using AI. It extracts entities, relationships, and insights from raw text data. Perfect for researchers and data analysts who need to visualize complex information.
Talat is a meeting notes application that uses on-device AI to transcribe and summarize meetings privately. It processes all audio locally on the user's device, ensuring complete data privacy and security. The tool is ideal for professionals who handle sensitive information and require confidential meeting documentation.
Subquadratic introduces a breakthrough large language model architecture that achieves sub-quadratic complexity, enabling efficient reasoning over contexts up to 12 million tokens. It is designed for enterprise-scale document analysis, legal review, and scientific research. The model reduces computational costs while maintaining high accuracy on long-context tasks.
LakeFS provides branch, commit, merge, and revert operations on data stored in object storage, bringing Git-like version control to data engineering workflows. It enables safe experimentation with data without risking production datasets.
Feast is an open-source feature store for managing and serving machine learning features. It helps teams store, version, and serve features consistently across training and production environments.
DVC (Data Version Control) is an open-source tool for versioning data and ML models, enabling reproducibility and collaboration. It works alongside Git to track large files and datasets.
MLflow is an open-source platform for managing the machine learning lifecycle, including experimentation, reproducibility, and deployment. It provides tools for tracking experiments, packaging code, and deploying models.
KNIME is an open-source data analytics platform that enables users to create visual data workflows for data blending, analysis, and machine learning. It supports a wide range of data sources and formats.
RapidMiner is a data science platform that provides an integrated environment for data preparation, machine learning, and model deployment. It supports both visual workflow design and code-based development.
A collaborative platform for the machine learning community to share and build models, datasets, and demo apps (Spaces). It serves as the primary repository for open-source AI.