Language Models

Language Models

underdogs breaking out
02
sqliteai/waste
+62% /wk +172 ★/daysteady

WASTE exists to find out how far local inference can be pushed when model weights live mostly on fast storage instead of RAM.

1.9k C Inference · Serving · explained Feature
03
drumih/turbo-fieldfare
+32% /wk +249 ★/dayaccelerating

TurboFieldfare streams individual experts from SSD on demand so Gemma 4 26B-A4B fits inside ~2 GB of RAM on any Apple Silicon Mac.

5.5k Swift Inference · Serving · explained
04
slvDev/esp32-ai
+29% /wk +156 ★/daycooling

The project squeezes a 28.9-million-parameter language model onto an ESP32-S3 microcontroller by keeping most of its weights in slow flash memory and reading only what each token needs.

3.8k Python Language Models · explained Feature
06
jegly/Box
+21% /wk +22 ★/dayaccelerating

A privacy-first Android fork that runs LLMs, image generation, and speech AI entirely offline, then locks itself behind your fingerprint.

735 Kotlin Inference · Serving · explained
07
openJiuwen-ai/jiuwenswarm
+19% /wk +67 ★/dayaccelerating

It breaks complex tasks across a team of specialized LLM agents that refine their own skills as they work.

2.5k Python Agents · explained
08
chrisliu298/awesome-on-policy-distillation
+16% /wk +15 ★/dayaccelerating

A curated reading list that maps how the post-training world is moving from static SFT to student-rollout distillation with live teacher feedback.

638 Learning · explained
09
SimonSchubert/Kai
+14% /wk +24 ★/dayaccelerating

An open-source AI assistant built in Kotlin that remembers context across chats, draws interactive screens, and runs a sandboxed Linux distro on your phone.

1.2k Kotlin Chat Assistants · explained
10
datawhalechina/diy-llm
+14% /wk +23 ★/dayaccelerating

A Chinese-language LLM curriculum that rebuilds Stanford CS336 into six hands-on assignments, from writing a tokenizer to distributed training and GRPO.

1.1k Jupyter Notebook Learning · explained
11
AgnesAI-Labs/AgnesAI-Models
+14% /wk +53 ★/dayaccelerating

Agnes AI is a hosted multimodal API that speaks OpenAI's protocol, letting you reroute existing clients to its text, image, video, and agent models by changing a base URL.

2.6k Inference · Serving · explained
12
AutoArk/TinyEngram
+13% /wk +17 ★/dayaccelerating

TinyEngram open-sources experiments showing that DeepSeek's Engram architecture can inject domain knowledge into Qwen and Stable Diffusion more efficiently than LoRA, with less catastrophic forgetting.

918 Python Language Models · explained
13
huggingface/speech-to-speech
+13% /wk +213 ★/daycooling

A modular speech-to-speech pipeline that exposes an OpenAI Realtime-compatible WebSocket API so you can run voice agents on local or open-source models instead of proprietary cloud services.

11.7k Python Agents · explained
14
Datus-ai/Datus-agent
+12% /wk +26 ★/dayaccelerating

Datus is an open-source agent that keeps LLMs from hallucinating SQL by building a living, learning knowledge base around your data stack.

1.5k Python Data Tooling · explained
15
Blaizzy/nativ
+8.8% /wk +15 ★/daycooling

Nativ bundles an embedded mlx-vlm server into a SwiftUI app to turn your Apple Silicon Mac into a local AI workspace with OpenAI-compatible APIs.

1.2k Swift Inference · Serving · explained
16
sapientinc/HRM-Text
+8.4% /wk +22 ★/dayaccelerating

HRM-Text claims to cut pretraining costs by 130–600× compute and 150–900× data, shipping a full 1B-parameter framework with FSDP2, FlashAttention 3, and a hierarchical recurrent architecture.

1.8k Python Language Models · explained
17
verl-project/verl-omni
+8.2% /wk +8.9 ★/daycooling

It split off from `verl` to give diffusion, video, and omni-modality models an RL post-training framework that doesn't treat them like chatbots.

759 Python ML Frameworks · explained
18
basketikun/chatgpt2api
+7.5% /wk +60 ★/dayaccelerating

It turns ChatGPT’s browser-only image generation into a poolable, OpenAI-compatible API so you can self-host programmatic access to GPT-Image-2 and friends.

5.6k Python Inference · Serving · explained
19
techjarves/Uncensored-Local-Studio
+7.3% /wk +8.7 ★/daycooling

It unifies Stable Diffusion, GGUF chat, Whisper, and Kokoro TTS into a single offline desktop GUI so you can skip cloud APIs, subscriptions, and censorship filters.

840 JavaScript Inference · Serving · explained
20
neavo/LinguaGacha
+6.7% /wk +22 ★/dayaccelerating

LinguaGacha uses LLMs to batch-translate novels, subtitles, and game scripts while auto-generating glossaries so character names stay consistent across an entire work.

2.3k TypeScript Language Models · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.