Models AI Projects
Daily ranking page for Models open-source AI repositories.
Models tracks 1673 repositories with 6129355 total GitHub stars.
- NandhaKishorM/laya - Non-autoregressive System 1 decision engine. Typed choice, score and yes/no decisions over any text in a single forward pass, in 100+ languages, with a... (30582 stars, Python, Models)
- datalab-to/lift - Extract structured data from documents quickly and accurately. (1084 stars, Python, Models)
- cactus-compute/needle - Automation foundation model for tiny devices: 2-bit, 8-29 MB, tool calls, structured extraction and embeddings on phones, wearables, smart homes, robots... (13175 stars, Python, Models)
- TokenRhythm/NeoHorse - NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness. (1567 stars, Python, Models)
- jaredpalmer/kev - Jev-like family of decision models built on top of Qwen3.5/3.8 you can train and run on your own (8406 stars, Python, Models)
- volotat/mini-AGI - Continual learning model trained from scratch on 8GB VRAM laptop with batch-1 stream of data. (1154 stars, Python, Models)
- shiyu-coder/Kronos - Kronos: A Foundation Model for the Language of Financial Markets (39896 stars, Python, Models)
- openai/whisper - Robust Speech Recognition via Large-Scale Weak Supervision (109944 stars, Python, Models)
- multimodal-art-projection/YuE - YuE2: frontier music generation with symbolic planning, zero-shot covers, and agentic music editing. (10797 stars, Python, Models)
- MiniMax-AI/MiniMax-H3 - (9535 stars, Python, Models)
- AI4Finance-Foundation/FinGPT - FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace. (21345 stars, Jupyter Notebook, Models)
- google-deepmind/tapnet - Tracking Any Point (TAP) (2149 stars, Jupyter Notebook, Models)
- facebookresearch/brain2qwerty - Non-invasive decoding of typed sentences from MEG and EEG brain recordings using a convolutional encoder, transformer, and character-level language model. (1008 stars, Python, Models)
- Inkloom-art/inkloom - Specialised AI models for logo design — a brand-analysis model turns a business into constraints, typography and symbol models construct the mark, and a... (688 stars, TypeScript, Models)
- om-ai-lab/VLX-Seek - VLX-Seek is a device-native vision-language model that enables machines to see, understand, and reason about the visual world with high precision. (1370 stars, Python, Models)
- ace-step/ACE-Step-1.5 - The most powerful local music generation model that outperforms almost all commercial alternatives, supporting Mac, AMD, Intel, and CUDA devices. (13023 stars, Python, Models)
- firelex/jeff - Millisecond decisions, any domain: a 0.8B open "System 1" model that picks between your options with calibrated probabilities. One base, swappable LoRA... (1359 stars, Python, Models)
- OpenBMB/VoxCPM - VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning (38299 stars, Python, Models)
- Liuziyu77/Valen - Train a Jev-like multimodal model by yourself. System One Model, now with vision. (599 stars, Python, Models)
- google-research/timesfm - TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting. (34095 stars, Python, Models)
- QwenLM/Qwen-Image-2.1 - Qwen's most powerful open-source image generation model (1710 stars, Python, Models)
- hexgrad/kokoro - https://hf.co/hexgrad/Kokoro-82M (9149 stars, JavaScript, Models)
- Lightricks/LTX-2 - Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model. (9587 stars, Python, Models)
- Wan-Video/Wan2.2 - Wan: Open and Advanced Large-Scale Video Generative Models (17718 stars, Python, Models)
- wfzyx/von - The open-source System One decision model. Sub-15ms, non-autoregressive, local drop-in alternative to TypeSafe Jev. (834 stars, Python, Models)
- fishaudio/fish-speech - SOTA Open Source TTS (32937 stars, Python, Models)
- XHToken/Spark-X2.5 - Spark-x2.5 open model series. Pushing the Limits of Agentic Capabilities in On-Device Models (618 stars, Unknown, Models)
- QwenAudio/CosyVoice - Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability. (23833 stars, Python, Models)
- index-tts/index-tts - An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System (24284 stars, Python, Models)
- k2-fsa/sherpa-onnx - Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Intern... (15099 stars, C++, Models)
- QwenLM/Qwen3-TTS - Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech genera... (13626 stars, Python, Models)
- OpenBMB/MiniCPM - MiniCPM5: SOTA on-device LLMs, small yet powerful. (11360 stars, Jupyter Notebook, Models)
- MeiGen-AI/InfiniteTalk - Unlimited-length talking video generation that supports image-to-video and video-to-video generation (7966 stars, Python, Models)
- limix-ldm-ai/LimiX - LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence https://arxiv.org/abs/2609.17488 (4365 stars, Python, Models)
- QwenLM/Qwen3.8 - Qwen3.8 is the large language model series developed by Qwen team, Alibaba Group. (4230 stars, Unknown, Models)
- Lightricks/LTX-Video - Official repository for LTX-Video (11017 stars, Python, Models)
- PriorLabs/TabPFN - ⚡ TabPFN: Foundation Model for Tabular Data ⚡ (8113 stars, Python, Models)
- amazon-science/chronos-forecasting - Chronos: Pretrained Models for Time Series Forecasting (5973 stars, Python, Models)
- urchade/GLiNER - Generalist and Lightweight Model for Named Entity Recognition (Extract any entity types from texts) (4032 stars, Python, Models)
- GVCLab/PersonaLive - [CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming (3920 stars, Python, Models)
- pnnbao97/VieNeu-TTS - Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 48kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text... (2748 stars, Python, Models)
- TianyuCodings/NanoJev - A nano replica of Jev: parallel decisions, dynamic candidates, and an end-to-end training pipeline. (2487 stars, Python, Models)
- huawei-bayerlab/marigold-v2 - Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation (881 stars, Python, Models)
- AlayaLab/Evoke - Official implementation of EVOKE: Endless Interactive World with Bounded State and Long-Horizon Supervision. A three-step, CFG-free interactive world mo... (793 stars, Python, Models)
- microsoft/BitNet - Official inference framework for 1-bit LLMs (40363 stars, C++, Models)
- Robbyant/lingbot-map - (ECCV 2026 oral & best paper candidate) LingBot-Map: Geometric Context Transformer for Streaming 3D Reconstruction (17213 stars, Python, Models)
- kyutai-labs/pocket-tts - A TTS that fits in your CPU (and pocket) (9773 stars, Python, Models)
- NVlabs/Sana - SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer (9198 stars, Python, Models)
- Tencent/WeMM-Embedding - WeMM-Embedding is a family of universal multimodal embedding models by the WeChat Vision Team at Tencent, supporting multimodal understanding and retrie... (1715 stars, Python, Models)
- xinntao/Real-ESRGAN - Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration. (36976 stars, Python, Models)
- SWivid/F5-TTS - Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching" (15335 stars, Python, Models)
- QwenAudio/SenseVoice - Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection. (9436 stars, C, Models)
- Tencent-Hunyuan/Hunyuan3D-2.1 - From Images to High-Fidelity 3D Assets with Production-Ready PBR Material (4120 stars, Python, Models)
- DAXIAORobotics/kairos - Official code for world model Kairos (3004 stars, Python, Models)
- nv-tlabs/lyra - Project Lyra: Open Generative 3D World Models (2528 stars, Python, Models)
- yuantianyuan01/FastWAM - Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination? (1548 stars, Python, Models)
- openai/CLIP - CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image (34408 stars, Jupyter Notebook, Models)
- salute-developers/GigaAM - Foundational Model for Speech Recognition Tasks (843 stars, Python, Models)
- thu-nics/C2C - [ICLR'26] The official code implementation for "Cache-to-Cache: Direct Semantic Communication Between Large Language Models" (704 stars, Python, Models)
- deepinsight/insightface - State-of-the-art 2D and 3D Face Analysis Project (29891 stars, Python, Models)