Skill-Verzeichnis

Wiederverwendbare Skills für AI Agents entdecken.

Durchsuche reale GitHub-Skills nach Aufgabe und prüfe Stars, Trust, Audit, Kategorie und Installationspfad vor der Verwendung.

Jede Empfehlung bleibt mit ihrem Repository, Audit und Installationspfad nachvollziehbar.

Suchergebnisse: synthesis

Englisches Verzeichnis

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer

8.3K
Stars
86/100
Trust
Kategorie: media-automationAudit

Offline Text To Speech synthesis for python

2.5K
Stars
80/100
Trust
Kategorie: media-automationAudit

[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis

2.2K
Stars
83/100
Trust
Kategorie: robotics-iotAudit

[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis

1.6K
Stars
83/100
Trust
Kategorie: media-automationAudit

Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinement (HTR) from the Arbor paper. Use this whenever someone wants to iteratively optimize something over many experiments without overfitting — e.g. "get my model's eval score up", "improve this agent/harness", "tune this pipeline", "beat the baseline on this benchmark", "run a search over approaches and keep the best", "do an MLE-bench / Kaggle-style optimization", or any long-horizon "make this artifact better and don't just memorize the dev set" task. Trigger it even when the user doesn't say "Arbor" or "hypothesis tree" but describes repeated experiment-and-evaluate loops, branching exploration of competing ideas, or worries about a dev/test gap. Runs Claude itself as the coordinator with subagent executors in isolated git worktrees; for the standalone `arbor` CLI tool see references/arbor-upstream.md.

34K
Stars
77/100
Trust
Kategorie: researchAudit

A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows.

1.5K
Stars
76/100
Trust
Kategorie: media-automationAudit

DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code

4.8K
Stars
72/100
Trust
Kategorie: media-automationAudit

:stuck_out_tongue_closed_eyes: TensorFlowTTS: Real-Time State-of-the-art Speech Synthesis for Tensorflow 2 (supported including English, French, Korean, Chinese, German and Easy to adapt for other languages)

4.0K
Stars
72/100
Trust
Kategorie: media-automationAudit

A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)

3.0K
Stars
72/100
Trust
Kategorie: ml-automationAudit

MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java

2.6K
Stars
70/100
Trust
Kategorie: media-automationAudit

HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

2.4K
Stars
72/100
Trust
Kategorie: media-automationAudit

PyTorch implementation of convolutional neural networks-based text-to-speech synthesis models

2.0K
Stars
68/100
Trust
Kategorie: ml-automationAudit