Showing 26 of 62 projects
Convert folders or RSS feeds into Studio pack zip files for Lunii and compatible audio storytelling devices.
A Blazor class library providing Speech Synthesis API access for client-side and server-side Blazor applications.
A shell wrapper for interacting with multiple AI service providers including OpenAI, LocalAI, Ollama, Gemini, and Anthropic via chat, text, and speech endpoints.
A collection of Node-RED nodes to integrate IBM Watson AI services like speech, language, and conversation into applications.
A desktop application that converts clipboard text into speech, allowing background audio narration of web pages.
A Java library for integrating with AI models like ChatGPT, DALL·E, and Cohere using minimal code.
A minimalist, privacy-first menu bar OCR and screen capture tool for macOS that processes text, QR codes, and barcodes entirely on-device.
A React Native module for Android Text-to-Speech functionality, providing speech synthesis and language support.
Example iOS voice-to-voice chat app using Watson Speech to Text, Conversation, and Text to Speech services.
An open-source mobile app that describes photos using audio for blind and visually-impaired users.
An Angular client for Google's Gemini Pro API with rich media support, text-to-speech, and integration with Google AI Studio and VertexAI.
An MCP server that provides AI assistants with 12 tools for video generation using the Creatify AI platform.
A Neovim plugin that reads selected text aloud using multiple TTS backends including Microsoft Edge, Piper, and OpenAI.
A Ruby library for consuming the AT&T Speech API to convert speech to text and text to speech.
A MicroPython library for interfacing with the YuTone VoiceTX SYN6988 text-to-speech module via UART.
Generate SSML fragments to bypass Alexa's text-to-speech censorship for profane words.
An Adobe AIR Native Extension for converting text to speech and speech to text on Android and iOS, working fully in the background.
A Deno module for Windows automation tasks like text-to-speech, screenshots, notifications, and system control using NirCmd.
An Elixir client library for interacting with Microsoft Azure's translation API, providing text translation, language detection, and text-to-speech.
A Capacitor plugin for cross-platform text-to-speech synthesis with full control over voice, pitch, rate, and volume.
A MicroPython port of the classic Software Automatic Mouth (SAM) text-to-speech synthesizer for embedded systems.
A command-line tool that uses text-to-speech to review GTFS stop name pronunciations and identify stops needing tts_stop_name values.
A macOS app that gives your MacBook a personality by reacting to hardware events like slaps, charging, and AI code completion.
An unofficial Elixir SDK for Microsoft Azure Speech Service, providing speech-to-text and text-to-speech capabilities.
A Node.js tool that validates accessibility-related fields and files in GTFS transit data.
A sound pack for INAV flight modes on OpenTX transmitters, providing synthesized voice alerts.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.