BookOrbit
Consolidate your entire digital library into a self-hosted reading sanctuary with BookOrbit, an open-source media platform that organizes ebooks, audiobooks, comic archives, and research PDFs within a personal private cloud. Readers can enjoy multi-format media directly through responsive browser readers that support EPUB, CBZ, and M4B formats without external browser extensions. The platform synchronizes reading bookmarks, highlights, and completion statuses across web interfaces, Kobo ereaders, and KOReader devices. An automated metadata enrichment engine queries fourteen upstream bibliographic indexes, downloading cover artwork, chapter breakdowns, author biographies, and series taxonomies. The staging workspace allows librarians to inspect incoming files, review suggested metadata revisions, and embed updated tags directly into physical storage archives before shelving. Custom reading dashboards track annual book goals, calculate daily streaks, and plot reading habit radar charts. Multi-user configurations support distinct per-user collections, granular content permissions, and single sign-on authentication through OpenID Connect providers. Power users can export private OPDS feeds to external mobile reading applications or push books directly to connected Kindle hardware. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. GNU AGPL v3.0 licensed.
PodFetch
Podcast enthusiasts and homelab archivers automate digital episode preservation and stream personal audio feeds using PodFetch, an open-source self-hosted podcast manager that combines scheduled RSS downloads, Podcasting 2.0 transcript indexing, and browser playback into a modern listening platform. Listeners can subscribe to international podcast channels through direct feed URLs, Apple Podcasts directory searches, or bulk OPML subscription imports, configuring background polling timers to download new releases automatically upon publication. The integrated audio player supports variable speed playback, episode bookmarks, listening histories, and fullscreen visualization modes directly inside responsive desktop and mobile browser sessions. Advanced search capabilities leverage Podcasting 2.0 synchronized transcripts and automated speech-to-text generation via OpenAI-compatible Whisper APIs, allowing collectors to run full-text keyword searches across every spoken phrase in their audio archive. Built-in GPodder API synchronization maintains subscription lists and playback positions across popular smartphone apps like AntennaPod and native companion clients without relying on commercial cloud accounts. Server operators can assign multi-user accounts with private favorites lists, export customized RSS aggregation streams, and route storage paths between fast solid-state drives and large secondary mechanical hard disks. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. Apache-2.0 licensed.
BookLore
BookLore centralizes personal electronic book, audiobook, and comic collections into an organized private media server equipped with automated metadata scrapers and e-reader synchronization. Users can drop EPUB, PDF, CBZ, and MOBI files into a watched BookDrop directory to trigger automated format parsing, background cover retrieval, and staged queue imports. Integrated metadata scrapers query Google Books, Open Library, Goodreads, and Amazon to automatically fill synopsis summaries, author records, publishing dates, and review scores. Dynamic Magic Shelves categorize volumes through custom rule-based filters, author collections, and full-text search indexes across your complete literature catalog. The browser-based reader renders documents with adjustable typography, night modes, audio playback controls, text annotations, and persistent bookmarking. Hardware e-readers and mobile devices connect through native Kobo Store emulation APIs, bidirectional KOReader progress synchronization, and OPDS catalog feeds. Multi-tenant permissions provide each household member with private reading statistics, personalized shelves, email book delivery, and direct send-to-Kindle dispatch. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. GNU AGPLv3 licensed.
Mstream
"The easiest music streaming server available" is mStream's own billing, and the claim holds up: a lightweight Node.js app that turns a folder of audio files into a private streaming service in minutes, no external database required. Its filesystem-based design is the clever part - the API mirrors your folder structure, so you can browse and play music immediately, before any library scan finishes, and your organization on disk is your organization in the app. It streams flac, mp3, wav, ogg, opus, aac, and m4a, which matters to the audiophile crowd: FLAC plays uncompressed, bit-perfect, with gapless playback for live albums and continuous mixes. The web player runs anywhere a browser does and packs personality - a Milkdrop-style visualizer (Butterchurn), playlist sharing via links, and drag-and-drop uploads straight through the file explorer. Native iOS and Android apps add the feature streaming subscriptions can't match: sync your collection to your phone for true offline playback of music you own. Multi-user support assigns separate directories and permissions per account. Resource usage is famously light - mStream is tested on multi-terabyte libraries and runs happily on a Raspberry Pi, so a small RepoCloud instance serves a lifetime's collection. GPL-licensed, with zero listening-habit telemetry.
Navidrome
Spotify economics without the subscription or catalog gaps: Navidrome, the reference self-hosted music server, streams your own FLAC, MP3, and ALAC collection from a single Go binary with a React/Material UI web player. Its Subsonic/OpenSubsonic API compatibility is the superpower: 50+ existing clients work out of the box, from Symfonium and DSub on Android to Feishin and Sonixd on desktop, plus Android Auto, CarPlay, and Android TV apps. Transcoding is server-managed and FFmpeg-backed - FLAC direct-plays at home and downsamples to MP3, AAC, or Opus over mobile bandwidth, with the OpenSubsonic transcoding extension letting clients declare capabilities and receive per-track direct-play or transcode decisions automatically. Multi-user support gives every account its own play counts, favorites, ratings, and playlists, and multi-library support scopes different collections to different users. The feature list covers serious listening: Last.fm and ListenBrainz scrobbling, artist bios and images, embedded and external lyrics, audiobook bookmarks, saved play queues that resume on another device, internet radio, jukebox mode, and M3U playlist auto-import kept in sync with your folder. Resource usage is famously low - it runs happily on a Raspberry Pi and scales to six-figure track counts.
Black Candy
With 4,300+ GitHub stars and native mobile apps on three platforms, Black Candy transforms any VPS into a private Spotify-style streaming service for your personal music collection. The Ruby on Rails 7 backend with Hotwire Turbo and Stimulus delivers a responsive single-page-feeling web player supporting album browsing, artist views, playlists, favorites, and queue management without full page reloads. Point it at a media directory containing MP3, FLAC, OGG, AAC, or WAV files and Black Candy indexes metadata, fetches album artwork from Discogs API, and begins streaming immediately with on-the-fly transcoding that adapts bitrate to client bandwidth. Multi-user support gives each account independent playlists, favorites, and listening history while sharing the same music library — ideal for families or shared households. Native iOS, Android, and F-Droid apps maintained as separate repositories provide offline caching, background playback, and server discovery for mobile listening. The admin panel manages user accounts, configures media paths, and sets Discogs API tokens for automatic cover art retrieval. Deployment requires one Docker command — `docker run -p 80:80 ghcr.io/blackcandy-org/blackcandy:latest` — with persistent storage volumes for the SQLite database and media directory. For larger deployments, switch to PostgreSQL via environment variables with dedicated database URLs for ActionCable, SolidQueue, and SolidCache. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.
AzuraCast
AzuraCast is a solution for running internet radio stations, packaging the entire broadcast stack into a single Docker installation that gets you on air in minutes. Built on PHP with a Vue.js frontend, it uses Liquidsoap as the AutoDJ engine to compile media from playlists, live DJ inputs, and scheduled programming into composited output streams, then broadcasts via Icecast-KH, SHOUTcast 2 DNAS, or Rocket Streaming Audio Server to thousands of concurrent listeners. The media library supports drag-and-drop uploads with automatic metadata extraction, while playlist management handles sequential, random, weighted, and time-scheduled rotations. Live DJ accounts let remote broadcasters connect via BUTT, Mixxx, or any Icecast-compatible software, with automatic crossfade transitions between AutoDJ and live input. The built-in Web DJ tool enables broadcasting directly from the browser without additional software. Each station includes public-facing player pages with now-playing metadata, album art, song history, and embeddable widgets for external websites. The listener request system lets audiences queue songs through the public interface. Multi-station administration hosts unlimited stations on a single server with individual user accounts and granular role-based permissions. Comprehensive analytics track listener counts, geographic distribution, unique listeners, and listening duration with exportable reports. HLS streaming delivers adaptive bitrate broadcasts for mobile compatibility. LADSPA audio plugins enable professional equalization and processing in the stream pipeline. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. AGPL-3.0 licensed.
Kokoro FastAPI
Kokoro-FastAPI turns text into natural-sounding speech across eight languages by serving the 82-million-parameter Kokoro-82M model through an OpenAI-compatible REST API, so any existing OpenAI SDK client can generate audio by just changing the base URL. With over 5,300 GitHub stars since December 2024, the fully Dockerized FastAPI server covers American English, British English, Spanish, French, Hindi, Italian, Japanese, Brazilian Portuguese, and Mandarin Chinese with language-specific phoneme processing. Inline voice mixing blends multiple profiles using weighted ratios like af_bella(2)+af_heart(1), automatically normalizing weights and caching combined voicepacks as PyTorch tensor files for reuse. Audio streams in real time over HTTP with configurable chunk sizes, or generates complete files in MP3, WAV, OPUS, FLAC, AAC, or PCM formats with speed control from 0.25x to 4.0x. Per-word timestamped captions with speaker-tagged voice labels enable subtitle generation for podcasts, audiobooks, and accessibility workflows. Pre-built Docker images support NVIDIA GPU acceleration via CUDA, experimental AMD GPU inference via ROCm, and CPU-only deployment on linux/amd64 and linux/arm64 architectures, with Apple Silicon MPS support available through direct UV execution. The integrated web interface at port 8880 provides browser-based speech generation, while the Swagger UI at /docs exposes the full API reference. Debug endpoints report system statistics for monitoring inference load. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. Apache 2.0 licensed.
Kima
Streaming personal music collections with neural discovery and intelligent playlist curation is what Kima accomplishes as a modern self-hosted audio server that replaces Spotify and Apple Music. Listeners organize their FLAC, MP3, and AAC audio files into responsive grid layouts that scale up to eight columns on ultra-wide displays while automatically fetching album covers, artist biographies, and synchronized line-by-line lyrics. The built-in Vibe System utilizes CLAP neural network audio embeddings to map an entire music collection across two-dimensional clusters and three-dimensional interactive galaxy views, enabling users to explore acoustic relationships or queue smooth sonic paths between distant tracks. Music fans can import playlists directly from Spotify, Deezer, and YouTube using deterministic ISRC matching, while automated Made For You mixes generate era-specific playlists, workout stations, and discovery queues based on individual listening habits. Integrated podcast search pulls subscriptions from iTunes with automatic episode tracking across devices, while native Audiobookshelf connectors synchronize spoken-word chapter positions and playback progress directly into the unified web player. OpenSubsonic API endpoints allow mobile users to stream libraries through native clients like Symfonium and DSub using scoped authentication tokens, while administrators manage multi-user accounts with time-based two-factor authentication. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. GPL-3.0 licensed.
Chatterbox TTS
With 26,000 GitHub stars and consistent victories over ElevenLabs in blind evaluations, Chatterbox delivers state-of-the-art text-to-speech with zero-shot voice cloning requiring only 5 seconds of reference audio. The model family spans three architectures: Chatterbox Multilingual V3 (500M parameters, 23+ languages including Arabic, Chinese, Japanese, Korean, Hindi, French, German, Spanish, and Portuguese), Chatterbox-Turbo (350M parameters optimized for voice agents with a single-step distilled decoder achieving ~200ms time-to-first-speech), and Chatterbox-Nano (110M parameters running 3x faster than realtime on 8 CPU cores for edge deployment). Unique among open-source TTS systems, Chatterbox introduces emotion exaggeration control — adjusting intensity from monotone to dramatically expressive via a single parameter — and native paralinguistic tagging where tokens like [laugh], [cough], [chuckle], and [gasp] inject natural vocal reactions inline without post-processing. The alignment-informed inference pipeline eliminates hallucinations and repetition artifacts common in autoregressive TTS. Built-in PerTh neural watermarking embeds imperceptible forensic identifiers in generated audio for provenance tracking. Trained on 500,000 hours of cleaned speech data across all supported languages. Voice conversion scripts enable transforming existing audio into any cloned voice. Deploy via pip install with PyTorch, serve through Gradio interfaces or custom FastAPI endpoints, and expose via HTTP streaming or WebSocket for sub-200ms conversational applications. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.
Castopod
Every podcast you publish through Castopod automatically becomes a Fediverse social account that Mastodon, Pleroma, and Pixelfed users can follow, boost, and comment on from their own timelines. This ActivityPub integration transforms podcast distribution from a one-way broadcast into a two-way conversation where listeners interact with episodes through the social platforms they already use. Beyond federation, the platform manages the complete lifecycle: upload an episode, and auto-generated RSS feeds distribute it to Apple Podcasts, Spotify, Deezer, Podcast Addict, and every compatible directory. IABv2-compliant analytics measure downloads, geographic distribution, and listening apps while maintaining GDPR, CCPA, and LGPD compliance through anonymized data collection. Podcasting 2.0 namespace support unlocks chapters with images, SRT/VTT transcripts, location tags, and person credits that modern podcast apps surface to listeners. Video episodes publish alongside audio, and auto-generated clips feed social media sharing workflows. Monetization tools span Value4Value micropayments, premium subscriptions, funding links, and cookieless advertising. Multi-user roles operate per-podcast: contributors, editors, and admins each work within defined permission boundaries. The PHP application runs behind nginx with MariaDB and Redis on a RepoCloud VPS with dedicated CPU, RAM, and SSD. AGPL-3.0 licensed.
ScribeWizard
Audio lectures become structured, Markdown-formatted notes in about a minute with ScribeWizard (also known as GroqNotes). Upload an MP3, WAV, or M4A file - or paste a YouTube link - and the app runs a three-stage pipeline on Groq's LPU inference hardware: Whisper Large v3 transcribes the audio, a larger Llama model drafts a comprehensive outline of the material, and a faster Llama model fills each section with detailed content. This scaffolded prompting strategy is the core idea: the strong model handles structure where quality matters most, the fast model handles volume, and Groq's 1200+ tokens-per-second inference keeps the whole process near real time. Output renders as clean Markdown with support for tables and code blocks, and finished notes download as text or PDF. Model selection is configurable - swap in other Groq-hosted open models like Mixtral or Gemma to trade speed against quality or work around rate limits. Built as a single Streamlit app by Benjamin Klieger at Groq, it needs only a Groq API key to run, making it one of the simplest self-hosted AI tools to operate.
Substreamer
A free, polished web client for Subsonic-compatible music servers: Substreamer is the browser-based frontend you point at your existing streaming backend to play your own library from anywhere. It speaks the Subsonic API (v1.13 and higher), which makes it compatible with the whole ecosystem that has grown around that protocol: the original Subsonic server, its forks Airsonic and Madsonic, and modern implementations like Navidrome and Ampache. That decoupling is the point - your music files, transcoding, and library indexing live on whichever server you prefer, while Substreamer provides the listening experience: browse by artist, album, and genre, build and manage playlists, search your collection, and stream on demand. This RepoCloud deployment runs the containerized web edition, so the same interface is available from any browser without installing a native app, and it pairs with the Substreamer mobile apps that made the client popular. For anyone assembling a self-hosted Spotify replacement - typically Navidrome for the backend plus a good client - Substreamer fills the client half with a clean, familiar player UI. Because it is a stateless client, the container is lightweight and low-maintenance: connect it to your server's URL and credentials, and your entire collection is streaming in minutes, with no subscription and no catalog that can disappear.