Image · Video · Audio

Image · Video · Audio

underdogs · picking up speed
01
basketikun/infinite-canvas
+38% /wk +273 ★/dayaccelerating

A self-hostable infinite canvas that wires AI image generation, reference editing, and chat into one collaborative workspace.

5k TypeScript Creative · Design · explained
02
jegly/Box
+22% /wk +24 ★/dayaccelerating

A privacy-first Android fork that runs LLMs, image generation, and speech AI entirely offline, then locks itself behind your fingerprint.

749 Kotlin Inference · Serving · explained
03
artokun/comfyui-mcp
+13% /wk +11 ★/dayaccelerating

It turns any LLM into a ComfyUI operator that edits live graphs, manages models, and runs workflows instead of just forwarding prompts.

577 TypeScript Agents · explained
04
AgnesAI-Labs/AgnesAI-Models
+14% /wk +53 ★/dayaccelerating

Agnes AI is a hosted multimodal API that speaks OpenAI's protocol, letting you reroute existing clients to its text, image, video, and agent models by changing a base URL.

2.6k Inference · Serving · explained
07
jau123/MeiGen-AI-Design-MCP
+8.9% /wk +22 ★/dayaccelerating

It plugs nine image and video models into your AI coding tools via MCP, letting them generate visuals in parallel without cluttering your chat context.

1.7k TypeScript Coding Assistants · explained
09
VigoZhao/AI-Visual-Prompt-Cookbook
+8.7% /wk +6.9 ★/dayaccelerating

Most AI image prompts are one-off text blobs; this repo distills 96 visual styles into structured JSON templates so you can swap variables without losing style direction.

552 Python Image · Video · Audio · explained
10
ModelTC/LightX2V
+7.3% /wk +28 ★/dayaccelerating

An inference framework that corrals a dozen video and image generation models into a single runtime optimized to outrun Diffusers and FastVideo on both H100s and RTX 4090Ds.

2.7k Python Image · Video · Audio · explained
11
techjarves/Uncensored-Local-Studio
+12% /wk +15 ★/dayaccelerating

It unifies Stable Diffusion, GGUF chat, Whisper, and Kokoro TTS into a single offline desktop GUI so you can skip cloud APIs, subscriptions, and censorship filters.

916 JavaScript Inference · Serving · explained
12
Vincentwei1021/video-shotcraft
+25% /wk +172 ★/dayaccelerating

It gives Claude Code and Codex a motion-design vocabulary—106 shot recipes, 161 previews, and a Remotion template—so they can direct cinematic product videos instead of generic slideshows.

4.8k TypeScript Agents · explained
13
wuyoscar/GPT-Image2-Skill
+7.9% /wk +50 ★/dayaccelerating

It curates GPT Image 2 prompts and packages them as copy-paste examples, an agent skill, and a lightweight CLI.

4.4k Python Coding Assistants · explained
14
IAHispano/Applio
+4.3% /wk +22 ★/dayaccelerating

Applio offers a browser-based voice conversion toolkit that prioritizes stability and plugin extensibility over constant feature churn.

3.6k Python Image · Video · Audio · explained
15
mcmonkeyprojects/SwarmUI
+3.5% /wk +22 ★/dayaccelerating

It wants to be the only browser tab you need for local AI image and video generation, serving both beginners and node-graph tinkerers.

4.5k C# Image · Video · Audio · explained
16
spotify/basic-pitch
+3.2% /wk +24 ★/dayaccelerating

It gives you polyphonic audio-to-MIDI transcription, complete with pitch bends, without the resource bill of heavier research tools.

5.4k Python Domain Apps · explained
17
bytedance/LatentSync
+2.6% /wk +22 ★/dayaccelerating

It generates lip-synced faces by feeding Whisper audio embeddings straight into a latent diffusion U-Net, skipping the usual motion-representation detour.

6k Python Image · Video · Audio · explained
19
vladmandic/sdnext
+2.4% /wk +25 ★/dayaccelerating

SD.Next exists to run Stable Diffusion and related models on virtually any consumer hardware while bundling quantization, captioning, and video tools into one interface.

7.3k Python Image · Video · Audio · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.