Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
| Date | Stars |
|---|---|
| 2026-07-24 | 5658 |
| 2026-07-25 | 5660 |
| 2026-07-28 | 5660 |
| 2026-07-30 | 5660 |
| 2026-08-06 | 5660 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
35.0
growth rate 0.00%/day
|Banner| |CI| |PyPI| |Latest Tag| |Coverage| |Slack| `Website <https://www.daft.ai>`_ • `Docs <https://docs.daft.ai>`_ • `Installation <https://docs.daft.ai/en/stable/install/>`_ • `Daft Quickstart <https://docs.daft.ai/en/stable/quickstart/>`_ • `Community and Support <https://github.com/Eventual-Inc/Daft/discussions>`_ Daft: High-Performance Data Engine for AI and Multimodal Workloads ================================================================== |TrendShift| `Daft <https://www.daft.ai>`_ is a high-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale. * **Native multimodal processing:** Process images, audio, video, and embeddings alongside structured data in a single framework * **Built-in AI operations:** Run LLM prompts, generate embeddings, and classify data at scale using OpenAI, Transformers, or custom models * **Python-native, Rust-powered:** Skip the JVM complexity with Python at its core and Rust under the hood for blazing performance * **Seamless scaling:** Start local, scale to distributed clusters on `Ray <https://docs.daft.ai/en/stable/distributed/ray/>`_, `Kubernetes <https://docs.daft.ai/en/stable/distributed/kubernetes/>`_ * **Universal connectivity:** Access data anywhere (S3, GCS, Iceberg, Delta Lake, Hugging Face, Unity Catalog) * **Out-of-box reliability:** Intelligent memory management and sensible defaults eliminate configuration headaches Getting Started --------------- Installation ^^^^^^^^^^^^ Install Daft with ``pip install daft``. Requires Python 3.10 or higher. For more advanced installations (e.g. installing from source or with extra dependencies such as Ray and AWS utilities), please see our `Installation Guide <https://docs.daft.ai/en/stable/install/>`_ Quickstart ^^^^^^^^^^ Get started in minutes with our `Quickstart <https://docs.daft.ai/en/stable/quickstart/>`_ - load a real-world e-commerce dataset, process product images, and run AI inference at scale. More Resources ^^^^^^^^^^^^^^ * `Examples <https://docs.daft.ai/en/stable/examples/>`_ - see Daft in action with use cases across text, images, audio, and more * `User Guide <https://docs.daft.ai/en/stable/>`_ - take a deep-dive into each topic within Daft * `API Reference <https://docs.daft.ai/en/stable/api/>`_ - API reference for public classes/functions of Daft Benchmarks ---------- |Benchmark Image| To see the full benchmarks, detailed setup, and logs, check out our `benchmarking page. <https://docs.daft.ai/en/stable/benchmarks>`_ Contributing ------------ We ❤️ developers! To start contributing to Daft, please read `CONTRIBUTING.md <https://github.com/Eventual-Inc/Daft/blob/main/CONTRIBUTING.md>`_. This document describes the development lifecycle and toolchain for working on Daft. It also details how to add new functionality to the core engine and expose it through a Python API. Here's a list of `good first issues <https://github.com/Eventual-Inc/Daft/issues?q=is%3Aopen+is%3Aissue+label%3A%22good+first+issue%22>`_ to get yourself warmed up with Daft. Comment in the issue to pick it up, and feel free to ask any questions! Telemetry --------- To help improve Daft, we collect non-identifiable data via Scarf (https://scarf.sh). To disable this behavior, set the environment variable ``DO_NOT_TRACK=true``. The data that we collect is: 1. **Non-identifiable:** No session IDs or user identifiers are collected 2. **Metadata-only:** We do not collect any of our users’ proprietary code or data 3. **For development only:** We do not buy or sell any user data Please see our `documentation <https://docs.daft.ai/en/stable/telemetry/>`_ for more details. .. image:: https://static.scarf.sh/a.png?x-pxid=31f8d5ba-7e09-4d75-8895-5252bbf06cf6 Related Projects ---------------- +---------------------------------------------------+-----------------+---------------+-------------+-----------------+-----------------------------+-------------+ | Engine
Excerpt of 8,386 characters
Read on GitHub800
556
Colin Ho · Eventual-Inc · United States
476
Cory Grinstead
348
240
168
Srinivas Lade · United States
168
R. Conner Howell · United States
152
147
102
YK · Eventual
84
Rohit Kulshreshtha
53
51
47
Andrew Gazelka · @indexable-inc · United States
41
40
jay · Bytedance · China
39
37
Sam Stokes · @Eventual-Inc · United States
37
Jeev B
35
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:27000d801c223d11, topic:data-engineering, topic:etl
matched fp:27000d801c223d11, topic:distributed-computing, topic:ray
matched fp:27000d801c223d11, topic:multimodal, desc:multimodal, readme:multimodal
matched fp:27000d801c223d11, topic:embeddings