Speakoala is an AI-powered text-to-speech (TTS) reading assistant designed to help users consume digital content through listening. It can read any website, email, and local document (including PDF, DOCX, and EPUB) with natural, lifelike voices. The tool supports over 70...
ProductivityVoice & speechAI agents
SurfSense is a highly customizable AI research agent, connected to external sources such as search engines, Google Drive, Slack, Microsoft Teams, Linear, Jira, ClickUp, Confluence, BookStack, Gmail, Notion, YouTube, GitHub, Discord, Airtable, Google Calendar, Luma,...
Image generationCode assistantsWriting & content
YTVidHub is a professional bulk YouTube subtitle downloader and transcript extractor designed for high-volume data collection. It allows users to extract subtitles from entire playlists and channels in a single click, supporting formats like SRT, VTT, and clean TXT. The...
ProductivityResearch & analysisVoice & speech
Dictato is a private, fast voice-to-text dictation application specifically built for macOS. It allows users to transcribe speech directly into any application—such as Gmail, Slack, or VS Code—using a global hotkey. The app operates 100% on-device, meaning no audio data is...
Writing & contentVoice & speech
trnscrb is a local meeting transcription tool for macOS that lives in the menu bar and automatically detects meetings on platforms like Zoom, Google Meet, Microsoft Teams, Slack, and FaceTime. It utilizes OpenAI's Whisper model to perform on-device transcription via...
ProductivityVoice & speech
Prism is an all-in-one AI video creation platform designed for making short-form content without needing multiple external tools. It allows users to generate image and video assets using various state-of-the-art models like Sora, Kling, and Veo, organize them into projects,...
Image generationVideo generationMarketing
Stage Captions is a professional, browser-based real-time closed captioning software designed for live events, conferences, and broadcasts. It utilizes an advanced AI engine to deliver production-ready live transcription with industry-leading low latency. The platform allows...
Writing & contentVoice & speech
Obi, developed by Cor (Corellian Systems), is a voice AI agent designed for customer onboarding and user activation. It functions like a live video call, using voice and on-screen awareness to guide users through product setups, share best practices, and answer questions in...
Voice & speechAI agentsCustomer support
PopAir is a native macOS AI copilot designed for speed and seamless system-wide integration. Built with SwiftUI to consume significantly less RAM than Electron-based apps, it serves as a unified hub for leading AI models including GPT, Claude, Gemini, and DeepSeek. The...
Image generationWriting & contentProductivity
Beni AI is a multimodal AI companion platform designed for real-time, video-first interactions. Unlike text-only AI, Beni responds with voice, motion, and expressions through video calls. The system features persistent memory that allows the companion to remember past...
Image generationWriting & contentVoice & speech
Medeo Seedance 2.0 is an advanced AI-powered video generation tool designed to streamline end-to-end video creation through a chat-based, multimodal workflow. Users can generate fully composed videos by combining text prompts, images, audio, and short video clips—without...
Video generationVoice & speech
Seedance2 Love is a professional AI video generation platform that specializes in creating high-quality cinematic videos using the Seedance 2.0 model. It allows users to generate videos from text, images, or multi-modal inputs with a focus on native multi-shot storytelling,...
Video generationVoice & speech
HeyVid AI is a comprehensive, all-in-one AI creative platform designed for generating high-quality videos, images, voices, and music. It provides a centralized hub to access over 18 leading AI models, including Sora, Kling, Runway, Midjourney, and Flux. The platform supports...
Image generationVideo generationAudio & music
MagicBoat AI is the ultimate AI filmmaking platform. Turn novels into scripts, generate consistent storyboards, and produce cinematic videos—all in one seamless workflow. Long Description MagicBoat AI is a professional-grade, all-in-one AI video production platform designed...
Image generationVideo generationVoice & speech
Crun AI is a unified AI API platform that provides developers and businesses with a single entry point to access over 100 top-tier AI models for video, image, and audio generation. It features models like Veo 3.1, Sora 2 Pro, Wan 2.6, and Flux, offering an OpenAI-compatible...
Image generationVideo generationAudio & music
Seedance 2.0 AI is a revolutionary multimodal video generation platform that allows users to create cinematic-quality videos up to 2 minutes long. It supports inputs from text descriptions, images, video clips, or audio assets, enabling advanced multi-shot storytelling while...
Video generationVoice & speech
Storyship is an AI-powered platform designed to create professional product demo videos instantly from screen recordings. It eliminates the need for complex editing skills by automatically adding AI voiceover, generating transcriptions, ensuring perfect audio-video sync, and...
Image generationVideo generationVoice & speech
Kling 3.0 is marketed as the ultimate 4K AI video generator released in 2026. It redefines AI storytelling by offering cinematic 4K precision, native audio integration (generating visuals, voice, and sound effects simultaneously), and advanced motion control for precise...
Image generationVideo generationVoice & speech
NoteAI is a Next-Gen AI Knowledge Extractor and all-in-one knowledge assistant designed for blazing-fast learning and creation. It is an AI-powered summarizer that supports various content formats, including YouTube videos, PDFs, documents (Word, PowerPoint, Excel), images,...
ProductivityTranslationVoice & speech
MuseGen | AI Music Generator
MuseGen is an all-in-one AI music generator studio powered by AI, utilizing Suno’s state-of-the-art AI model. It transforms text prompts and creative ideas into full-length, radio-ready songs, including melodies, harmony, expressive lyrics, and lifelike vocals, all generated...
Image generationAudio & musicVoice & speech
Disertus Labs (Milo AI Speech Therapist)
Disertus Labs offers Milo, an AI Speech Therapist accessible 24/7 via iMessage, Web, and WhatsApp. Milo allows users to practice speech therapy anytime, anywhere, receiving instant feedback, personalized exercises, and comprehensive progress tracking. It utilizes real-time...
Voice & speechAI agents
RED is a smart, floating AI hub that deeply integrates into your workflow, combining screen analysis, real-time transcription, and automation capabilities. It helps users think, organize, and act in real-time without needing screenshots or context switching, leveraging Deep...
ProductivityVoice & speechAutomation
Sayline is a native macOS application designed for private, local voice dictation in any text field. It allows users to replace manual typing with voice commands using global hotkeys across various applications like Gmail, Slack, VS Code, or Notes. Utilizing on-device...
Writing & contentProductivityVoice & speech
Kuku is a truly native, local-first markdown editor designed for macOS, built using Tauri for superior performance and minimal resource consumption compared to Electron applications. It adheres to an open format, storing notes as plain .md files with support for wikilinks,...
Writing & contentProductivityVoice & speech
Genspark Speakly is an AI voice dictation application designed to convert spoken language into clear, polished messages, emails, and writings. It is marketed as being 4x faster than typing. The app integrates advanced AI features like Auto-Edits (which remove filler words,...
Writing & contentProductivityVoice & speech
Nafy AI is a powerful online AI music generator that enables users to transform concepts into royalty-free, studio-quality audio tracks instantly. It provides comprehensive tools for creating beats, vocals, and full songs, utilizing Text To Music, Lyrics to Song, AI Song...
Image generationAudio & musicVoice & speech