Speakr
Speakr transforms audio recordings into organized, searchable, AI-enhanced notes with speaker recognition that identifies who said what across your entire recording library. The Python/Flask backend with Vue.js 3 and Tailwind CSS frontend deploys via Docker on port 8899, offering multiple transcription engines through auto-detected connectors: WhisperX for local processing with speaker diarization and voice embeddings, OpenAI Whisper and GPT-4o-transcribe, Mistral Voxtral for cloud diarization, AssemblyAI for multi-hour files, and any custom ASR webservice. Speaker voice profiles use embedding comparison to recognize individuals across different recordings automatically, while custom vocabulary biases the transcriber toward domain-specific jargon. The AI layer goes well beyond transcription: customizable summaries with per-recording, per-tag, and per-folder prompt templates; event extraction surfacing action items and calendar events; per-recording chat with streaming responses; and Inquire Mode for semantic search and natural-language queries across your entire library simultaneously. Smart tags execute custom AI prompts on transcripts for automatic categorization. The REST API with Swagger documentation supports signed webhooks integrating with n8n, Zapier, and Make. Auto-export pushes to Obsidian and Logseq, auto-processing watches directories, and the installable PWA provides mobile-first, offline-capable access with share-target support. 3,600+ stars since May 2025. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. AGPL-3.0 licensed.
OpenMontage
Reaching #1 on GitHub Trending with over 48,000 stars, OpenMontage is the first open-source agentic video production system — transforming AI coding assistants like Claude Code, Cursor, Copilot, Windsurf, and Codex into complete video studios that handle research, scripting, scene planning, asset generation, editing, and final rendering through natural language prompts. Twelve production pipelines cover animated explainers, cinematic trailers, documentary montages, talking heads, screen demos, podcast repurposing, character animation, localization and dubbing, avatar spokesperson videos, hybrid productions, clip factory batch processing, and animation workflows. Over 100 registered Python tools connect to 60+ providers including Kling, Runway Gen-4, Google Veo 3.1, FLUX, Google Imagen 4, ElevenLabs, and Suno AI for cloud generation, plus Piper TTS, WAN 2.1, Hunyuan, and CogVideo for fully local GPU rendering — while free footage from Archive.org, NASA, Wikimedia Commons, Pexels, and Unsplash powers the documentary montage pipeline's CLIP-indexed retrieval system for real-motion video without paid generation APIs. Two composition engines — Remotion for React-based programmatic video and HyperFrames for HTML/GSAP motion graphics — render final output with spring animations, word-level captions, kinetic typography, and SVG character rigs. A seven-dimension scored provider selector, pre-compose validation gates, post-render ffprobe self-review, slideshow risk scoring, configurable budget caps with per-action approval thresholds, and the Backlot live web dashboard for visual production monitoring enforce production-grade quality at every stage. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. AGPL-3.0 licensed.