Transcriptionstream
transcriptionstream
turnkey self-hosted offline transcription and diarization service with llm summary
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
401–416 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
Live data is unavailable. A saved snapshot may be shown; check the source before use.
transcriptionstream
turnkey self-hosted offline transcription and diarization service with llm summary
FireRedTeam
An Open-Sourced LLM-empowered Foundation TTS System
deepgram
Official Python SDK for Deepgram.
judahpaul16
ChatGPT at home! A better alternative to commercial smart home assistants, built on the Raspberry Pi using LiteLLM and LangGraph.
ShandaAI
Generative World Renderer: an AI-native Renderer for Games and Virtual Worlds. 面向游戏与虚拟世界的AI原生渲染引擎
lidge-jun
Minimal CLI + web UI for OpenAI GPT Image 2 generation. Dual auth: API Key (paid) or OAuth via ChatGPT (free). Text-to-image, image-to-image, parallel gen, custom sizes.
Lakonik
Official implementation of AsymFlow, pi-Flow, GMFlow
prouast
Desktop implementation of Remote Photoplethysmography – Measuring heart rate using facial video.
yeyupiaoling
基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型
antgroup
[ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis
SamurAIGPT
Remove Seedance 2.0 watermark (AI生成) from videos automatically. No GPU required. Free open-source tool.
HorizonWind2004
[ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potential in Unified Multimodal Models…
gexgd0419
Make Azure natural TTS voices accessible to any SAPI 5-compatible application.
amanchadha
iSeeBetter: Spatio-Temporal Video Super Resolution using Recurrent-Generative Back-Projection Networks | Python3 | PyTorch | GANs | CNNs | ResNets | RNNs | Published in…
opendilab
High-quality and streaming Speech-to-Speech interactive agent in a single file. 只用一个文件实现的流式全双工语音交互原型智能体!
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.