hub

CodeZen

CodeZen

一个专注中文区的 GitHub 项目发现

avatar

Open-LLM-VTuber

Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D taking face running locally across platforms

ai ai-companion ai-vtuber ai-waifu chatbots
star4.8k
Python
avatar

rags

Build ChatGPT over your data, all with natural language

agent chatbot chatgpt gpts llamaindex
star6.5k
Python
avatar

Mastering-GitHub-Copilot-for-Paired-Programming

A multi-module course teaching everything you need to know about using GitHub Copilot as an AI Peer Programming resource.

copilot csharp dotnet github github-copilot
star7.1k
Python
avatar

mlx-examples

Examples in the MLX framework

mlx
star8.0k
Python
avatar

EmotiVoice

EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine

ai deep-learning emotion emotivoice multi-speaker
star8.4k
Python
avatar

clone-voice

A sound cloning tool with a web interface, using your voice or any sound to record audio / 一个带web界面的声音克隆工具,使用你的音色或任意声音来录制音频

clonevoice speech-analysis sts tts voice-assistant
star8.8k
Python
avatar

InternVL

[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型

gpt gpt-4o gpt-4v image-classification image-text-retrieval
star9.4k
Python
avatar

Amphion

Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.

audio-generation audio-synthesis audioldm audit emilia
star9.5k
Python
avatar

self-operating-computer

A framework to enable multimodal models to operate a computer.

automation openai pyautogui
star10.0k
Python
avatar

StreamDiffusion

StreamDiffusion: A Pipeline-Level Solution for Real-Time Interactive Generation

star10.5k
Python
avatar

magic-animate

[CVPR 2024] Official repository for "MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model"

star10.9k
Python
avatar

OpenVoice

Instant voice cloning by MIT and MyShell. Audio foundation model.

text-to-speech tts voice-clone zero-shot-tts
star35.4k
Python