Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis gemelo-ai.github.io/vocos/
onnxruntime-extensions: A specialized pre- and post- processing library for ONNX Runtime
动画壁纸。Free and open-source software that allows users to set animated desktop wallpapers and screensavers powered by WinUI 3.
[EMNLP 2025 Demo] PDF scientific paper translation with preserved formats - 基于 AI 完整保留排版的 PDF 文档全文双语翻译,支持 Google/DeepL/Ollama/OpenAI 等服务,提供 CLI/GUI/MCP/Docker/Zotero
ASR、TTS、说话者日记化、语音增强、源分离和VAD,基于Kaldi和onnxruntime,无需联网。支持嵌入式、Android、iOS、HarmonyOS、Raspberry Pi、RISC-V、RK NPU/NPU、x86_64、websocket,12种编程语言。https://k2-fsa.github.io/sherpa/onnx/index.html
ImageMagick是一个免费的开源软件套件,用于创建、编辑、转换和显示图像。它支持200多种格式,并提供强大的命令行工具和API,用于跨平台的自动化、脚本编写和集成。
A modern and performant C++20 read/write parser of Photoshop Files (*.psd and *.psb) with fully fledged Python bindings hosted on PyPi.
Voice data <= 10 mins can also be used to train a good VC model!
Port of OpenAI's Whisper model in C/C++。OpenAI的Whisper ASR模型在C/C++语言环境下的迁移实现版。对英文支持较好
将Gemini CLI、Antigravity、ChatGPT Codex、Claude Code、Grok Build包装为与OpenAI/Gemini/Claude/Codex兼容的API服务,让您可以通过API享受免费的Gemini 3.1 Pro、GPT 5.5、Grok 4.3、Claud模型。
Firmware for Generic WiFi & Bluetooth Combo SDK(AC791N)
AI 小助手,you run on your own devices. It answers you on the channels you already use (WhatsApp, Telegram, Slack, Discord, Google Chat, Signal, iMessage, Microsoft Teams, WebChat),
No fortress, purely open ground. OpenManus is Coming. openmanus.github.io/
tiny recursive descent expression parser, compiler, and evaluation engine for math expressions web: codeplea.com/tinyexpr