Upload an image file or a PDF document, and the app will read the text contained in it. PDFs are first split into separate page images, then each page is processed to recognize the words, sending t...
Open-source, local-first AI video clipper. Long video in, caption-burned speaker-tracked 9:16 clips out. Fully offline with Whisper + Ollama, or bring your own API key.
PDFCraft is a free, privacy-focused PDF toolkit that runs entirely in your browser. With 90+ professional tools, you can edit, convert, merge, split, and secure your PDF files without ever uploading them to a server.
A self-hosted Vietnamese Text-to-Speech tool that runs entirely on your machine. No subscriptions, no usage quotas, no data sent to external servers. Your text and audio never leave your computer.
[LINUX + NVIDIA ONLY] Real-time interactive world model. Drive an infinite, action-conditioned world rollout at 720p/16fps on a single desktop GPU (~19GB VRAM). https://github.com/amap-cvlab/ABot-World