import-ai/omnibox
quality grade A, 81 out of 100Collect, organize, use, and share, all in OmniBox.
- stars
- 1.5k
- stars gained this week
- +8
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Released model weights, reference implementations and architecture research.
Signals: large-language-models, llm, foundation-models, transformer, gpt, llama, mistral, qwen
1,456 results
Collect, organize, use, and share, all in OmniBox.
Go with your own intelligence - Write Go applications that directly integrate llama.cpp for local inference using hardware acceleration on Linux, macOS, Windows, & WebAssembly.
A neurosymbolic perspective on LLMs
Integrate cutting-edge LLM technology quickly and easily into your apps
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
Turn Windsurf / Devin Desktop's 100+ AI models (Claude, GPT, Gemini, DeepSeek, Kimi, GLM, SWE) into OpenAI-, Anthropic- & Gemini-compatible APIs. Zero-dependency self-hosted reverse proxy for Claude Code, Cline & Cursor. 把 Windsurf/Devin 云端 100+ 模型变成三套兼容 API。
Distributed AI/LLM for the people. Share compute privately or publicly to power your agents and chat.
Web, Desktop & Mobile client for Codex, Claude Code, OpenCode, Kimi, Augment Code, Qwen, fully end-to-end encrypted
VS Code extension for LLM-assisted code/text completion
A Low-Code MCP Framework for Building Complex and Innovative RAG Pipelines
Fast, flexible LLM inference
MACE - Fast and accurate machine learning interatomic potentials with higher order equivariant message passing.
Building AI agents, atomically
《动手学 Pi》:沿 15 个真实 checkpoint 从零构建 Pi-style Agent
Translate full-length books and documents with Ollama, OpenAI-compatible, Gemini, Mistral, DeepSeek, Poe or OpenRouter. Preserves formatting. Resumes where you left off. No file size limits.
BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.
MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.
TimeCopilot: the GenAI Forecasting Agent. Built on LLMs and Time Series Foundation Models, it lets you forecast, cross-validate, and detect anomalies using multiple foundation models through a single API. From finance and energy to web analytics, TimeCopilot turns natural-language queries into production-ready forecasts.
🚀 Efficient implementations for emerging model architectures
LLM KV cache compression made easy
Build APIs your users love ❤️ with Speakeasy. ✨ Polished and type-safe SDKs. 🌐 Terraform providers, MCP servers, CLIs and Contract Tests for your API. OpenAPI native.
SWE-bench: Can Language Models Resolve Real-world Github Issues?
Adding guardrails to large language models.
TabFM (Tabular Foundation Model) is a pretrained tabular foundation model developed by Google Research for tabular data regression and classification.
24,540 repositories in the index in total.