Language Models

Language Models

underdogs breaking out
01
slvDev/esp32-ai
+202% /wk +392 ★/daysteady

The project squeezes a 28.9-million-parameter language model onto an ESP32-S3 microcontroller by keeping most of its weights in slow flash memory and reading only what each token needs.

1.4k Python Language Models · explained
03
openJiuwen-ai/jiuwenswarm
+32% /wk +80 ★/dayaccelerating

It breaks complex tasks across a team of specialized LLM agents that refine their own skills as they work.

1.8k Python Agents · explained
04
ximeiorg/Xime
+13% /wk +13 ★/dayaccelerating

Xime is a deliberately minimal, Rime-based Android input method that serves as its author's personal testbed for on-device AI experiments in predictive text and speech recognition.

702 Kotlin Language Models · explained
05
google-research/tabfm
+12% /wk +37 ★/daycooling

TabFM exists so you can run classification and regression on messy, mixed-type tables without retraining a model on your data.

2.1k Python Language Models · explained
06
techjarves/Uncensored-Local-Studio
+11% /wk +12 ★/daysteady

It unifies Stable Diffusion, GGUF chat, Whisper, and Kokoro TTS into a single offline desktop GUI so you can skip cloud APIs, subscriptions, and censorship filters.

734 JavaScript Inference · Serving · explained
07
verl-project/verl-omni
+10% /wk +9.6 ★/dayaccelerating

It split off from `verl` to give diffusion, video, and omni-modality models an RL post-training framework that doesn't treat them like chatbots.

661 Python ML Frameworks · explained
08
OpenDCAI/DataFlex
+9.4% /wk +23 ★/dayaccelerating

DataFlex stops LLM training loops from wasting compute on static data mixes by dynamically selecting, mixing, and reweighting samples inside LLaMA-Factory.

1.7k Python Language Models · explained
09
AtomicBot-ai/Atomic-Chat
+9.2% /wk +15 ★/dayaccelerating

It turns your local machine into an OpenAI-compatible inference endpoint so agents and IDEs can run on offline models without reconfiguration.

1.2k TypeScript Inference · Serving · explained
10
MarkPDFdown/markpdfdown
+8.4% /wk +23 ★/dayaccelerating

Uses multimodal LLMs to transcribe PDFs into Markdown, preserving complex layouts that traditional extractors mangle.

1.9k Python Data Tooling · explained
11
ForceInjection/AI-fundamentals
+8.2% /wk +23 ★/dayaccelerating

Curated technical deep-dives covering everything from NVLink signal integrity to Kubernetes GPU scheduling and Huawei NPU porting.

2k HTML Learning · explained
12
voocel/ainovel-cli
+7.6% /wk +17 ★/dayaccelerating

This Go CLI turns a single sentence into a full novel by making Architect, Writer, and Editor LLM agents plan, draft, and review inside a long-loop state machine—no human hand-holding required.

1.5k Go Agents · explained
16
Bytez-com/docs
+5.6% /wk +18 ★/dayaccelerating

Bytez wraps 175,000+ AI models behind a single endpoint so you don't have to host them yourself.

2.3k TypeScript Inference · Serving · explained
17
chrisliu298/awesome-on-policy-distillation
+5.3% /wk +4.3 ★/daysteady

A curated reading list that maps how the post-training world is moving from static SFT to student-rollout distillation with live teacher feedback.

566 Learning · explained
18
sapientinc/HRM-Text
+5.3% /wk +13 ★/dayaccelerating

HRM-Text claims to cut pretraining costs by 130–600× compute and 150–900× data, shipping a full 1B-parameter framework with FSDP2, FlashAttention 3, and a hierarchical recurrent architecture.

1.7k Python Language Models · explained
19
kyegomez/OpenMythos
+5.2% /wk +109 ★/dayaccelerating

OpenMythos is an independent attempt to reconstruct Anthropic’s rumored Claude Mythos architecture as a trainable Recurrent-Depth Transformer with switchable attention and sparse MoE layers.

14.8k Python Language Models · explained
20
Starmel/OpenSuperWhisper
+5.2% /wk +17 ★/daycooling

It’s a macOS dictation app built for people who want real-time Whisper transcription without leaving the keyboard.

2.3k Swift Inference · Serving · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.