A Claude Code / Codex skill that automates the entire film pipeline, using Remotion to turn research and narration into deterministic 720p motion graphics with TTS voiceover.
Image · Video · Audio
underdogs breaking outMost short-video SaaS tools watermark your work and charge rent; this local Python pipeline offers a watermark-free OpusClip alternative for creators with existing long-form content.
A ground-up reimplementation of NVIDIA's neural rendering runtime that lets RDNA 3 and 4 cards run DLSS 5 in FSR-enabled DirectX 12 games.
A self-hosted 'fashion' platform whose repository name, topics, and sponsor link suggest it is marketing something far less SFW than wardrobe recommendations.
SolarWM exists to make long-horizon video world models reproducible, open-sourcing the full stack from a unified data contract to staged training recipes and checkpoints across three major backbones.
4DAnyone generates dense multi-view video from a single casual clip, feeding standard 4D Gaussian Splatting pipelines without a real camera rig.
An open-source answer to paywalled AI camera-angle tools, wrapping Qwen-Image-Edit-Plus in a Three.js control rig so anyone can generate multi-angle images without a subscription.
GigaAM exists because Russian call centers, music, and atypical speech deserve a dedicated open-source foundation model instead of hand-me-down multilingual checkpoints.
It gives modern audio models a shared native runtime so you can stop managing Python package conflicts and start generating speech, music, and transcripts locally.
OpenWhispr is the open-source, privacy-first alternative to WisprFlow and Granola that lets you choose between local Whisper/Parakeet models or your own cloud API keys.
JoyAI-Video-Edit applies natural-language instructions to live or uploaded video streams frame by frame, without buffering the whole clip or revisiting future frames.
This ComfyUI node turns the official MiniMax H3 pipeline into a multi-segment editing bay for long-form video with stitched motion and native stereo audio.
It gives generative models a first-person camera and spatial ears, producing synchronized sight and sound as you navigate.
It inventories the shapes, counts, and spatial rhythm of a photograph, then renders only the observable facts as sparse editorial abstractions.
It provides the theoretically correct causal initialization that autoregressive video distillation was missing, enabling one-step to four-step generation without extra training overhead.
AICON wires script parsing, character-aware generation, and video publishing into one node-based canvas so you don't lose your actors between scenes.
It unifies Stable Diffusion, GGUF chat, Whisper, and Kokoro TTS into a single offline desktop GUI so you can skip cloud APIs, subscriptions, and censorship filters.
A Tauri desktop app that auto-detects a dozen local AI backends so you don't have to wrestle with Docker or API keys.
Maestro is a local creative suite that uses a built-in LLM to plan, shoot, and edit multi-clip music videos and short films without sending anything to the cloud.
It bundles the entire AI drama workflow—script, storyboard, voice, and final cut—into a single pipeline you can host yourself.





