VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
-
Updated
Sep 29, 2026 - Python
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
This repository contains the Hugging Face Agents Course.
Non-autoregressive System 1 decision engine. Typed choice, score and yes/no decisions over any text in a single forward pass, in 100+ languages, with a router that picks the right checkpoint per request.
🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing tool calling (including MCP support), agents and RAG easy. It integrates seamlessly with enterprise Java frameworks like Quarkus and Spring Boot.
Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepower. Maintained by Orchestra Research.
A PyTorch-based Speech Toolkit
The open source codebase powering HuggingChat
YuE2: frontier music generation with symbolic planning, zero-shot covers, and agentic music editing.
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
The Enterprise-Grade Multi-Agent Orchestration Framework. Website: https://swarms.ai
A scikit-learn compatible neural network library that wraps PyTorch
Chronos: Pretrained Models for Time Series Forecasting
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
A large-scale 7B pretraining language model developed by BaiChuan-Inc.
Framework agnostic sliced/tiled inference + interactive ui + error analysis plots
A blazing fast inference solution for text embeddings models
Create 🔥 videos with Stable Diffusion by exploring the latent space and morphing between text prompts
🤗 AutoTrain Advanced
To associate your repository with the huggingface topic, visit your repo's landing page and select "manage topics."