Github-Ranking-AI

A list of the most popular AI Topic repositories on GitHub based on the number of stars they have received.| AI相关主题Github仓库排名,每日自动更新。

View on GitHub

Github Ranking

Top 100 Stars in MoE

Ranking Project Name Stars Forks Language Open Issues Description Last Commit
1 vllm 91315 21935 Python 2378 A high-throughput and memory-efficient inference and serving engine for LLMs 2026-09-09T07:51:45Z
2 LlamaFactory 74665 9145 Python 998 Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024) 2026-09-08T08:00:35Z
3 sglang 35675 8684 Python 862 SGLang is a high-performance serving framework for large language models and multimodal models. 2026-09-09T08:03:06Z
4 colibri 27160 2976 C 58 Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦 2026-09-07T22:47:21Z
5 ms-swift 15558 1670 Python 482 Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, …) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, …) (AAAI 2025). 2026-09-09T06:13:06Z
6 TensorRT-LLM 14576 2728 Python 583 TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way. 2026-09-09T07:47:33Z
7 FreeToken 12191 1175 Python 160 FreeToken brings datacenter-scale model serving to your desktop. Run massive models locally, fast and efficiently. 2026-09-09T06:37:23Z
8 kimi-k3-in-c 7222 1180 C 6 A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS, no framework, no GPU. 2026-08-26T07:36:53Z
9 flashinfer 6358 1412 Cuda 328 FlashInfer: Kernel Library for LLM Serving 2026-09-09T05:27:45Z
10 MoeKoeMusic 6266 395 Vue 2 一款开源简洁高颜值的酷狗第三方客户端 An open-source, concise, and aesthetically pleasing third-party client for KuGou that supports Windows / macOS / Linux / Web :electron: 2026-09-03T01:48:52Z
11 Bangumi 5914 167 TypeScript 27 :electron: An unofficial https://bgm.tv ui first app client for Android and iOS, built with React Native. 一个无广告、以爱好为驱动、不以盈利为目的、专门做 ACG 的类似豆瓣的追番记录,bgm.tv 第三方客户端。为移动端重新设计,内置大量加强的网页端难以实现的功能,且提供了相当的自定义选项。 目前已适配 iOS / Android。 2026-09-08T14:08:19Z
12 xtuner 5194 448 Python 242 A Next-Generation Training Engine Built for Ultra-Large MoE Models 2026-09-09T03:47:31Z
13 trace.moe 5029 263 None 0 Timestamp Retrieval for Anime Clips Everywhere 2026-08-21T13:33:40Z
14 fastllm 4991 487 C++ 304 fastllm是后端无依赖的高性能大模型推理库。同时支持张量并行推理稠密模型和混合模式推理MOE模型,任意10G以上显卡即可推理满血DeepSeek。双路9004/9005服务器+单显卡部署DeepSeek满血满精度原版模型,单并发20tps;INT4量化模型单并发30tps,多并发可达60+。 2026-09-09T06:22:38Z
15 GLM-4.5 4423 475 Python 27 GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models 2026-02-01T08:28:10Z
16 Moeditor 4106 263 JavaScript 106 (discontinued) Your all-purpose markdown editor. 2020-07-07T01:08:32Z
17 flash-moe 4090 507 Objective-C 10 Running a big model on a small laptop 2026-03-19T17:21:57Z
18 Moe-Counter 3064 299 JavaScript 3 Moe counter badge with multiple themes! - 多种风格可选的萌萌计数器 2026-04-16T03:39:37Z
19 moemail 2809 2591 TypeScript 50 A cute temporary email service built with NextJS + Cloudflare technology stack 🎉 | 一个基于 NextJS + Cloudflare 技术栈构建的可爱临时邮箱服务🎉 2026-08-29T10:20:42Z
20 MoeGoe 2424 241 Python 28 Executable file for VITS inference 2023-08-22T07:17:37Z
21 MoE-LLaVA 2323 139 Python 65 【TMM 2025🔥】 Mixture-of-Experts for Large Vision-Language Models 2025-07-15T07:59:33Z
22 MoBA 2180 158 Python 12 MoBA: Mixture of Block Attention for Long-Context LLMs 2025-04-03T07:28:06Z
23 ICEdit 2104 110 Python 23 [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ MoE ckpt released! Only 4GB VRAM is enough to run! 2025-12-19T19:08:02Z
24 DeepSeek-MoE 1974 310 Python 18 DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models 2024-01-16T12:18:10Z
25 fastmoe 1861 206 Python 25 A fast MoE impl for PyTorch 2025-02-10T06:04:33Z
26 OpenMoE 1695 86 Python 6 A family of open-sourced Mixture-of-Experts (MoE) Large Language Models 2024-03-08T15:08:26Z
27 paimon-moe 1530 284 JavaScript 315 Your best Genshin Impact companion! Help you plan what to farm with ascension calculator and database. Also track your progress with todo and wish counter. 2026-09-01T11:37:25Z
28 uccl 1510 172 C++ 57 UCCL is an efficient communication library for GPUs, covering collectives, P2P (e.g., KV cache transfer, RL weight transfer), and EP (e.g., GPU-driven) 2026-09-06T02:49:05Z
29 SpikingBrain-7B 1380 189 Python 9 Spiking Brain-inspired Large Models, integrating hybrid efficient attention, MoE modules and spike encoding into its architecture 2026-05-14T09:52:16Z
30 moepush 1366 428 TypeScript 15 一个基于 NextJS + Cloudflare 技术栈构建的可爱消息推送服务, 支持多种消息推送渠道✨ 2025-05-10T11:42:44Z
31 SmartImage 1329 81 C# 4 Reverse image search tool (SauceNao, IQDB, Ascii2D, trace.moe, and more) 2026-09-06T01:55:53Z
32 MOE 1321 139 C++ 170 A global, black box optimization engine for real world metric optimization. 2023-03-24T11:00:32Z
33 diy-llm 1320 134 Jupyter Notebook 0 🎓 系统性大语言模型构建课程|🛠️ 覆盖预训练数据工程、Tokenizer、Transformer、MoE、GPU 编程 (CUDA/Triton)、分布式训练、Scaling Laws、推理优化及对齐 (SFT/RLHF/GRPO)|🚀 6 个渐进式作业 + 代码驱动,建立 LLM 全栈认知体系 2026-09-08T11:53:18Z
34 mixture-of-experts 1252 112 Python 6 PyTorch Re-Implementation of “The Sparsely-Gated Mixture-of-Experts Layer” by Noam Shazeer et al. https://arxiv.org/abs/1701.06538 2024-04-19T08:22:39Z
35 MiniMind-in-Depth 1183 90 None 6 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程 2025-06-16T14:13:15Z
36 MoeMemosAndroid 1172 133 Kotlin 97 An app to help you capture thoughts and ideas 2026-09-04T17:13:22Z
37 Uni-MoE 1116 71 Python 27 Uni-MoE: Lychee’s Large Multimodal Model Family. 2026-08-06T10:50:45Z
38 Aria 1088 91 Jupyter Notebook 32 Codebase for Aria - an Open Multimodal Native MoE 2025-01-22T03:25:37Z
39 Tutel 1018 110 C 55 Tutel MoE: Optimized Mixture-of-Experts Library, Support GptOss/DeepSeek/Kimi-K2/Qwen3 using FP8/NVFP4/MXFP4 2026-09-02T06:08:25Z
40 llama-moe 1003 61 Python 6 ⛷️ LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training (EMNLP 2024) 2024-12-06T04:47:07Z
41 Time-MoE 997 115 Python 14 [ICLR 2025 Spotlight] Official implementation of “Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts” 2026-03-21T16:00:55Z
42 MoeTTS 988 72 None 0 Speech synthesis model /inference GUI repo for galgame characters based on Tacotron2, Hifigan, VITS and Diff-svc 2023-03-03T07:30:05Z
43 moebius 966 54 JavaScript 40 Modern ANSI & ASCII Art Editor 2024-05-02T15:54:35Z
44 cudnn-frontend 932 278 Python 94 cuDNN Frontend is NVIDIA’s modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels. 2026-09-09T05:00:45Z
45 ComfyUI-QwenVL 877 142 Python 2 ComfyUI-QwenVL custom node: Integrates the Qwen-VL series, including Qwen2.5-VL, Qwen3-VL, Qwen3.5-VL, Qwen3.6-VL (MoE), and Qwen3.8-VL, with GGUF support for advanced multimodal AI in text generation, image understanding, and video analysis. 2026-09-04T07:09:02Z
46 MoeMemos 876 78 Swift 81 An app to help you capture thoughts and ideas 2026-08-01T01:32:47Z
47 Adan 822 70 Python 6 Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models 2025-06-08T14:35:41Z
48 Hunyuan-A13B 819 118 Python 35 Tencent Hunyuan A13B (short as Hunyuan-A13B), an innovative and open-source LLM built on a fine-grained MoE architecture. 2025-07-08T08:45:27Z
49 MoePeek 816 55 Swift 8 A lightweight macOS selection translator built with pure Swift 6, featuring on-device Apple Translate for privacy, only 5MB install size and stable ~50MB memory usage. 一款轻量级 macOS 划词翻译工具,纯 Swift 6 开发,设备端 Apple 翻译保护隐私,安装体积仅 5MB,后台运行内存稳定约 50MB 2026-09-04T14:19:30Z
50 DeepSeek-671B-SFT-Guide 812 96 Python 1 An open-source solution for full parameter fine-tuning of DeepSeek-V3/R1 671B, including complete code and scripts from training to inference, as well as some practical experiences and conclusions. (DeepSeek-V3/R1 满血版 671B 全参数微调的开源解决方案,包含从训练到推理的完整代码和脚本,以及实践中积累一些经验和结论。) 2025-03-13T03:51:33Z
51 moe-theme.el 773 68 Emacs Lisp 15 A customizable colorful eye-candy theme for Emacser. Moe, moe, kyun! 2026-08-19T15:54:08Z
52 MixtralKit 770 76 Python 11 A toolkit for inference and evaluation of ‘mixtral-8x7b-32kseqlen’ from Mistral AI 2023-12-15T19:10:55Z
53 sonic-moe 761 104 Python 9 Accelerating MoE with IO and Tile-aware Optimizations 2026-08-29T09:26:22Z
54 SwiftLM 760 53 Swift 1 ⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache compression, MACOS + iOS iPhone app. 2026-09-05T23:27:20Z
55 moe 719 36 Nim 43 A command line based editor inspired by Vim. Written in Nim. 2026-09-08T07:25:32Z
56 YOLO-Master 710 155 Python 33 [CVPR2026]🚀🚀🚀Official code for the paper “YOLO-Master: MOE-Accelerated with Specialized Transformers for Enhanced Real-time Detection.” (YOLO = You Only Look Once) 🔥🔥🔥 2026-09-09T02:33:15Z
57 pegainfer 678 104 Rust 66 Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2 2026-09-07T20:19:48Z
58 Awesome-Mixture-of-Experts-Papers 671 49 None 3 A curated reading list of research in Mixture-of-Experts(MoE). 2024-10-30T07:48:14Z
59 moedict-webkit 652 98 Objective-C 1 萌典網站 2026-08-14T03:36:14Z
60 MoeList 647 23 Kotlin 25 Another unofficial Android MAL client 2026-08-09T08:38:47Z
61 vtbs.moe 639 36 Vue 34 Virtual YouTubers in bilibili 2025-07-31T13:39:09Z
62 satania.moe 619 55 HTML 3 Satania IS the BEST waifu, no really, she is, if you don’t believe me, this website will convince you 2022-10-09T23:19:01Z
63 Chinese-Mixtral 612 43 Python 0 中文Mixtral混合专家大模型(Chinese Mixtral MoE LLMs) 2026-04-19T00:59:54Z
64 moebooru 608 81 Ruby 30 Moebooru, a fork of danbooru1 that has been heavily modified 2026-09-06T06:24:48Z
65 moebius 607 42 Elixir 3 A functional query tool for Elixir 2024-10-23T18:55:45Z
66 mixture-of-kittens 582 75 Python 4 Mixture-of-experts (MoE) training megakernel for NVL72s 2026-08-14T03:44:18Z
67 BitSoulStockSkill 581 51 Python 0 由BitSoul出品的A股市场全能Skill,自带免费历史数据,内置100+行业主流因子,完整的回测框架,基于MOE架构的股票筛选与买卖判断,更提供因子挖矿等趣味接口,欢迎安装试用,也欢共同开发交流! 2026-03-21T08:19:00Z
68 MoeGoe_GUI 569 69 C# 8 GUI for MoeGoe 2023-08-22T07:32:08Z
69 trace.moe-telegram-bot 559 78 TypeScript 0 This Telegram Bot can tell the anime when you send an screenshot to it 2026-08-31T17:07:14Z
70 BigMoeOnEdge 549 57 C++ 17 Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp 2026-09-07T18:45:29Z
71 moerail 545 42 JavaScript 22 铁路车站代码查询 × 动车组交路查询 2025-08-13T12:55:25Z
72 vLLM-Moet 540 50 Sass 11 A vLLM patch + hand‑written SM120 SASS kernels: 2‑bit MoE experts + an FP4 “delta” cache that recovers precision — matching the official (NV)FP4 checkpoint’s quality on consumer Blackwell cards 2026-09-08T15:51:27Z
73 Moebius 536 44 Python 3 [ECCV 2026] Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance 2026-08-12T03:16:36Z
74 LPLB 531 43 Python 1 An early research stage expert-parallel load balancer for MoE models based on linear programming. 2025-11-19T07:20:35Z
75 step_into_llm 479 127 Jupyter Notebook 27 MindSpore online courses: Step into LLM 2025-12-22T11:46:46Z
76 FreeMoe 475 8 None 16 Unlock App Vip 2026-01-30T05:50:54Z
77 apex-quant 465 32 Shell 11 Adaptive Precision for EXpert Models: MoE-aware mixed-precision quantization 2026-08-17T09:15:33Z
78 Lvllm 453 39 Python 0 LvLLM is a special NUMA extension of vllm that makes full use of CPU and memory resources, reduces GPU memory requirements, and features an efficient GPU parallel and NUMA parallel architecture, supporting hybrid inference for MOE large models. 2026-09-08T01:58:29Z
79 MoeSR 451 13 JavaScript 8 An application specialized in image super-resolution for ACGN illustrations and Visual Novel CG. 专注于插画/Galgame CG等ACGN领域的图像超分辨率的应用 2026-08-10T15:16:16Z
80 ARIS-in-AI-Offer 446 16 Python 3 Bilingual (中文+EN) ML / LLM / diffusion / agent interview cheat sheets for AI 秋招 — generated by ARIS /interview-cheatsheet, rendered by /render-html into single-file HTML, reads anywhere — plus a CV→DBLP-fact-checked academic homepage generator and hand-authored long-form blogs 🌱 2026-09-09T05:40:56Z
81 DiT-MoE 439 20 Python 7 Scaling Diffusion Transformers with Mixture of Experts 2024-09-09T02:12:12Z
82 moe-sticker-bot 427 73 Go 39 A Telegram bot that imports LINE/kakao stickers or creates/manages new sticker set. 2024-06-06T15:28:28Z
83 awesome-moe-inference 424 18 None 0 Curated collection of papers in MoE model inference 2026-03-12T01:59:19Z
84 MOE 422 80 Java 18 Make Opensource Easy - tools for synchronizing repositories 2022-06-20T22:41:08Z
85 hydra-moe 415 16 Python 10 None 2023-11-02T22:53:15Z
86 MoeLoaderP 413 28 C# 12 🖼二次元图片下载器 Pics downloader for booru sites,Pixiv.net,Bilibili.com,Konachan.com,Yande.re , behoimi.org, safebooru, danbooru,Gelbooru,SankakuComplex,Kawainyan,MiniTokyo,e-shuushuu,Zerochan,WorldCosplay ,Yuriimg etc. 2025-05-19T13:20:58Z
87 Awesome-Efficient-Arch 408 33 None 0 Speed Always Wins: A Survey on Efficient Architectures for Large Language Models 2025-11-11T09:47:37Z
88 zeraix 408 13 TypeScript 0 Open-source local AI workspace — advancing on-device inference. 2026-09-08T09:01:15Z
89 nmoe 399 33 Python 2 MoE training for Me and You and maybe other people 2026-03-15T22:23:47Z
90 ds4-on-spark 394 27 Shell 5 Entrpi/ds4, a Blackwell CUDA perf fork of antirez/ds4 on NVIDIA DGX Spark: one-command install, ~3x upstream prefill, ~1.5x decode, DSpark, and full continuous batch support 2026-08-27T04:25:21Z
91 st-moe-pytorch 387 35 Python 4 Implementation of ST-Moe, the latest incarnation of MoE after years of research at Brain, in Pytorch 2024-06-17T00:48:47Z
92 WThermostatBeca 372 70 C++ 4 Open Source firmware replacement for Tuya Wifi Thermostate from Beca and Moes with Home Assistant Autodiscovery 2023-08-26T22:10:38Z
93 pixiv.moe 371 41 TypeScript 0 😘 A pinterest-style layout site, shows illusts on pixiv.net order by popularity. 2023-03-08T06:54:34Z
94 MOEAFramework 362 129 Java 0 A Free and Open Source Java Framework for Multiobjective Optimization 2026-01-21T16:26:02Z
95 MoE-Infinity 354 38 Python 6 PyTorch library for cost-effective, fast and easy serving of MoE models. 2026-09-07T08:39:10Z
96 notify.moe 353 46 Go 86 :dancer: Anime tracker, database and community. Moved to https://git.akyoto.dev/web/notify.moe 2022-09-26T07:15:05Z
97 pith-train 348 35 Python 4 Compact and Agent-Native MoE Training System 2026-09-07T19:49:39Z
98 soft-moe-pytorch 348 10 Python 4 Implementation of Soft MoE, proposed by Brain’s Vision team, in Pytorch 2025-04-02T12:47:40Z
99 dialogue.moe 344 10 Python 1 None 2022-12-14T14:50:38Z
100 kimi-k3-mlx 337 40 Python 14 MLX port of moonshotai/Kimi-K3 (2.78T multimodal MoE): streaming converter, REAP expert pruning, and per-language expert-overlap analysis 2026-08-01T20:36:39Z