Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
GitHub projects from awesome lists
Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.
🦉 OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation
Tongyi Deep Research, the Leading Open-source Deep Research Agent
A TTS model capable of generating ultra-realistic dialogue in one pass.
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services
Janus-Series: Unified Multimodal Understanding and Generation Models
🐫 CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.org
Sample code and notebooks for Generative AI on Google Cloud, with Gemini Enterprise Agent Platform
Agent framework and applications built upon Qwen>=3.0, featuring Function Calling, MCP, Code Interpreter, RAG, Chrome extension, etc.
SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts.
ChatGLM2-6B: An Open Bilingual Chat LLM | 开源双语对话语言模型
GUI for ChatGPT API and many LLMs. Supports agents, file-based QA, GPT finetuning and query with web search. All with a neat UI.
Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
Llama中文社区,实时汇总最新Llama学习资料,构建最好的中文Llama大模型开源生态,完全开源可商用
本项目是一个面向小白开发者的大模型应用开发教程,在线阅读地址:https://datawhalechina.github.io/llm-universe/
ChatGLM3 series: Open Bilingual Chat LLMs | 开源双语对话语言模型
Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
Easy-to-use and powerful LLM and SLM library with awesome model zoo.
No description.
The most powerful local music generation model that outperforms almost all commercial alternatives, supporting Mac, AMD, Intel, and CUDA devices.
Go ahead and axolotl questions
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Official implementation of AnimateDiff.
This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.