Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Convert files (PDF, image, Word, PPT, Excel, notebooks, code snippets) to markdown using powerful multimodal LLM
| Date | Stars |
|---|---|
| 2026-07-31 | 344 |
| 2026-08-01 | 344 |
| 2026-08-06 | 344 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# MarkEverythingDown [](https://deepwiki.com/RoffyS/MarkEverythingDown) + **MarkEverythingDown** - 你的全能文档Markdown转换神器!🚀 一键将PDF/Office/图片/代码等文件转换为结构清晰的Markdown,专为LLM优化设计。结合Qwen2.5 VL视觉模型,连扫描件都能智能解析! ## ✨ 优势 ✅ **AI超能力** - 深度集成Qwen2.5 VL模型,完美保留表情符号和图像描述 ✅ **格式全覆盖** - 从微信截图到学术论文统统搞定 ✅ **双模处理** - 本地/云端自由切换,隐私与性能兼得 ✅ **小白友好** - 无需代码,拖拽文件立即转换 ✅ **智能分批** - 优化处理大型PDF文档,自动调整批次大小 **MarkEverythingDown** is a versatile document conversion tool that transforms various file formats into clean, structured markdown. Whether you're working with PDFs, Office documents, images, code files, or notebooks, MarkEverythingDown provides a unified interface to convert them all. The tool is specifically designed to leverage **Qwen2.5 VL** models through OpenAI-compatible APIs, supporting both local inference engines like LMStudio and cloud API providers like DashScope. This design enables high-quality processing of visual content while maintaining flexibility in deployment options. I developed this tool to streamline the conversion of documents into markdown format, which is both LLM-friendly and easy for human to read. The goal is to make document processing as seamless as possible, allowing users to easily convert their files for RAG applications or SFT dataset preparations. ## Roadmap ### Recently Implemented (April 2025) #### Enhanced Processing Options - ✅ **Temperature Control**: Added temperature parameter (0.0-1.0) for controlling the determinism of AI output - ✅ **Max Tokens Setting**: Implemented customizable token limits for generation - ✅ **Multi-Page Processing**: Added support for processing multiple PDF pages in a single API call - ✅ **Dynamic Batch Sizing**: Implemented intelligent adjustment of batch sizes based on page complexity - ✅ **Optimized Token Management**: Added max_tokens_per_batch option to prevent token limit issues #### Improved Document Support - ✅ **Enhanced Table Handling in Word Documents**: Better preservation of table structure and formatting in DOCX files - ✅ **Excel Spreadsheet Support**: Full support for XLSX files with proper table formatting - ✅ **Better Visual Elements Preservation**: Improved handling of emojis and image descriptions #### Interface Improvements - ✅ **Enhanced UI Tooltips**: Clearer explanations of processing options - ✅ **Improved Error Handling**: Better feedback for processing issues - ✅ **Progress Indicators**: Added visual feedback during processing ### Planned Features #### Near-term - 🔜 **CSV and TSV Support**: Native support for tabular data files - 🔜 **Custom Templates**: User-defined output formats for different document types - 🔜 **Batch Processing Improvements**: Enhanced management of large document collections #### Long-term - 🔜 **Multi-model Support**: Integration with additional vision-language models - 🔜 **Advanced Document Analysis**: Improved extraction of complex structures like footnotes and citations - 🔜 **API Mode**: Headless operation for integration with other applications - 🔜 **Collaborative Editing**: Real-time collaborative editing of converted documents ## Features - **Multi-format support**: Convert PDFs, DOCX, PPTX, XLSX, images, code files, notebooks, and markdown variants - **Intelligent processing**: Automatically selects the appropriate processor for each file type - **Vision AI support**: Optimized for Qwen2.5 VL models with OpenAI-compatible interface - **Dual processing options**: Support local inference APIs and cloud APIs - **Batch processing**: Process multiple files at once with a simple interface - **User-friendly UI**: Easy-to-use web UI with Gradio and helpful tooltips - **Command line interface**: Quick conversions from the terminal ## Supported Formats | Category | Formats | |----------|---------| | Documents | PDF, DOCX, PPTX, XLSX | | Images | PNG, JPG, JPEG, BMP | | Code | Python, R, and other programming languages | | Notebooks | Jupyter Notebooks (ipynb) | | Markdo
Excerpt of 30,623 characters
Read on GitHubWould you bet a product on this? Bounded 0–100 and slow moving.
matched fp:81af0f66299afa14, topic:multimodal, desc:multimodal
matched fp:81af0f66299afa14, topic:qwen