Computer Vision

Computer Vision

big names · picking up speed
01
ruvnet/RuView
+743 ★/dayaccelerating

Because commodity WiFi already bounces off your body, RuView uses cheap ESP32 nodes to detect presence, vital signs, and even body pose without cameras or wearables.

86.4k Rust Domain Apps · explained Feature
02
microsoft/OmniParser
+29 ★/dayaccelerating

OmniParser turns raw screenshots into structured, labeled UI elements so vision-language models can finally click what they mean to click.

25.2k Jupyter Notebook Agents · explained
03
TencentARC/GFPGAN
+18 ★/dayaccelerating

A Tencent research project that restores degraded faces by tapping into the rich priors locked inside a pretrained StyleGAN2 model.

37.6k Python Computer Vision · explained
05
graphdeco-inria/gaussian-splatting
+13 ★/dayaccelerating

It renders high-quality novel views of real-world scenes at 30 fps by replacing costly neural radiance fields with optimized 3D Gaussians.

22.8k Python Computer Vision · explained
07
google-ai-edge/mediapipe
+17 ★/dayaccelerating

It exists to let developers run customized vision, text, and audio machine learning across mobile, web, and edge hardware without cloud round-trips.

36.3k C++ Computer Vision · explained
08
ultralytics/yolov5
+7.0 ★/dayaccelerating

YOLOv5 made real-time object detection as easy as `torch.hub.load`, then exported to everything from iOS to edge chips.

57.7k Python Computer Vision · explained
09

To give developers a zero-shot image segmentation model that generates masks from a click or a bounding box, no retraining required.

54.6k Jupyter Notebook Computer Vision · explained
11
serengil/deepface
+4.0 ★/daysteady

DeepFace wraps a zoo of pre-trained face models into a single Python API so you can verify identities, search databases, and analyze attributes without hand-rolling a Keras pipeline.

23.2k Python Computer Vision · explained
12
xinntao/Real-ESRGAN
+11 ★/daysteady

Real-ESRGAN turns the ESRGAN research model into a practical tool for upscaling and restoring real-world images and videos using only synthetic training data.

36.3k Python Image · Video · Audio · explained
13
google-research/google-research
+5.1 ★/daysteady

Google Research releases all its code and datasets in one place, which has grown so large that the README treats a full clone as a hazard.

38.4k Jupyter Notebook Other AI · explained
14
ultralytics/ultralytics
+35 ★/daycooling

Ultralytics wants to stop you from stitching together separate repos for every computer vision task by bundling detection, segmentation, tracking, and pose estimation into one YOLO-backed package.

59.9k Python Computer Vision · explained
15
JaidedAI/EasyOCR
+5.6 ★/daycooling

It glues together CRAFT detection and CRNN recognition so you can pull text out of images without tuning neural networks yourself.

29.8k Python Computer Vision · explained
16
deepinsight/insightface
+6.9 ★/daycooling

It bundles detection, recognition, alignment, and reconstruction into a single research-grade toolbox.

29.3k Python Computer Vision · explained
19
opencv/opencv
+21 ★/daycooling

OpenCV is an open-source C++ computer vision library whose own README acts as a portal rather than a product page.

90.1k C++ Computer Vision · explained
20
blakeblackshear/frigate
+19 ★/daycooling

Frigate performs real-time, local object detection on IP camera streams using OpenCV and TensorFlow, designed to integrate tightly with Home Assistant.

34.5k TypeScript Computer Vision · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.