Doriandarko/texts-to-transformer
quality grade C, 61 out of 100Train a tiny Transformer from scratch on your iMessage history, entirely on your Mac.
- stars
- 467
- stars gained this week
- +3this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Released model weights, reference implementations and architecture research.
Signals: large-language-models, llm, foundation-models, transformer, gpt, llama, mistral, qwen
1,455 results
Train a tiny Transformer from scratch on your iMessage history, entirely on your Mac.
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
Stratified Transformer for 3D Point Cloud Segmentation (CVPR 2022)
No description
Guideline following Large Language Model for Information Extraction
CTNet: A Convolutional Transformer Network for EEG-Based Motor Imagery Classification
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).
[ACL 2022] LinkBERT: A Knowledgeable Language Model 😎 Pretrained with Document Links
[ECCV 2024] Official PyTorch implementation of RoPE-ViT "Rotary Position Embedding for Vision Transformer"
[ICCV 2023 & TPAMI 2026] SparseBEV: High-Performance Sparse 3D Object Detection from Multi-Camera Videos
一个执着于让CPU\端侧-Model逼近GPU-Model性能的项目,CPU上的实时率(RTF)小于0.1
Training neural network potentials
Implementation of Transformer Model in Tensorflow
ChatGPT-Pro is an advanced application that combines the power of ChatGPT and DALL.E.
Zig INferenCe Engine — Local LLM inference on AMD GPUs and Apple Silicon
ChineseNMT: Translate English to Chinese with PyTorch Implementation of Transformer
MetaFormer Baselines for Vision (TPAMI 2024)
A low-latency & high-throughput serving engine for LLMs
PageTransformer for flutter
A Jest transformer using esbuild
FashionCLIP is a CLIP-like model fine-tuned for the fashion domain.
Mask Transfiner for High-Quality Instance Segmentation, CVPR 2022
[ECCV'2020] STTN: Learning Joint Spatial-Temporal Transformations for Video Inpainting
Mass-editing thousands of facts into a transformer memory (ICLR 2023)
24,540 repositories in the index in total.