Language Models

Language Models

underdogs · picking up speed
02
openJiuwen-ai/jiuwenswarm
+32% /wk +80 ★/dayaccelerating

It breaks complex tasks across a team of specialized LLM agents that refine their own skills as they work.

1.8k Python Agents · explained
03
ximeiorg/Xime
+13% /wk +13 ★/dayaccelerating

Xime is a deliberately minimal, Rime-based Android input method that serves as its author's personal testbed for on-device AI experiments in predictive text and speech recognition.

702 Kotlin Language Models · explained
04
AtomicBot-ai/Atomic-Chat
+9.2% /wk +15 ★/dayaccelerating

It turns your local machine into an OpenAI-compatible inference endpoint so agents and IDEs can run on offline models without reconfiguration.

1.2k TypeScript Inference · Serving · explained
05
MarkPDFdown/markpdfdown
+8.4% /wk +23 ★/dayaccelerating

Uses multimodal LLMs to transcribe PDFs into Markdown, preserving complex layouts that traditional extractors mangle.

1.9k Python Data Tooling · explained
06
ForceInjection/AI-fundamentals
+8.2% /wk +23 ★/dayaccelerating

Curated technical deep-dives covering everything from NVLink signal integrity to Kubernetes GPU scheduling and Huawei NPU porting.

2k HTML Learning · explained
07
verl-project/verl-omni
+10% /wk +9.6 ★/dayaccelerating

It split off from `verl` to give diffusion, video, and omni-modality models an RL post-training framework that doesn't treat them like chatbots.

661 Python ML Frameworks · explained
08
Bytez-com/docs
+5.6% /wk +18 ★/dayaccelerating

Bytez wraps 175,000+ AI models behind a single endpoint so you don't have to host them yourself.

2.3k TypeScript Inference · Serving · explained
09
sapientinc/HRM-Text
+5.3% /wk +13 ★/dayaccelerating

HRM-Text claims to cut pretraining costs by 130–600× compute and 150–900× data, shipping a full 1B-parameter framework with FSDP2, FlashAttention 3, and a hierarchical recurrent architecture.

1.7k Python Language Models · explained
10
kyegomez/OpenMythos
+5.2% /wk +109 ★/dayaccelerating

OpenMythos is an independent attempt to reconstruct Anthropic’s rumored Claude Mythos architecture as a trainable Recurrent-Depth Transformer with switchable attention and sparse MoE layers.

14.8k Python Language Models · explained
11
huggingface/speech-to-speech
+4.9% /wk +45 ★/dayaccelerating

A modular speech-to-speech pipeline that exposes an OpenAI Realtime-compatible WebSocket API so you can run voice agents on local or open-source models instead of proprietary cloud services.

6.5k Python Agents · explained
12
Arthur-Ficial/apfel
+3.5% /wk +31 ★/dayaccelerating

Apfel surfaces Apple’s built-in FoundationModels as a pipe-friendly UNIX tool and OpenAI-compatible local server, no API keys required.

6.2k Swift Inference · Serving · explained
13
voocel/ainovel-cli
+7.6% /wk +17 ★/dayaccelerating

This Go CLI turns a single sentence into a full novel by making Architect, Writer, and Editor LLM agents plan, draft, and review inside a long-loop state machine—no human hand-holding required.

1.5k Go Agents · explained
14
andrewyng/aisuite
+2.9% /wk +64 ★/dayaccelerating

A thin Python wrapper that lets you swap GPT-4o for Claude or Gemini without touching your code.

15.4k Python Language Models · explained
15
fla-org/flash-linear-attention
+2.6% /wk +20 ★/dayaccelerating

It corrals the latest subquadratic sequence-model research into hardware-efficient, training-ready PyTorch layers verified across NVIDIA, AMD, and Intel GPUs.

5.4k Python ML Frameworks · explained
16
OpenDCAI/DataFlex
+9.4% /wk +23 ★/dayaccelerating

DataFlex stops LLM training loops from wasting compute on static data mixes by dynamically selecting, mixing, and reweighting samples inside LLaMA-Factory.

1.7k Python Language Models · explained
17
qualcomm/GenieX
+1.8% /wk +22 ★/dayaccelerating

NexaSDK is a local inference engine that squeezes frontier LLMs and vision models onto Qualcomm silicon through NPU, GPU, and CPU backends.

8.3k Rust Inference · Serving · explained
18
bigscience-workshop/petals
+1.8% /wk +27 ★/dayaccelerating

Petals lets you run and fine-tune models like Llama 3.1 405B from a desktop by distributing layers across a public swarm of consumer GPUs.

10.4k Python Inference · Serving · explained
20
jingyaogong/minimind-o
+3.5% /wk +11 ★/dayaccelerating

MiniMind-O packs listen-see-speak intelligence into a 0.1B-parameter model you can retrain from the first line of code on a single desktop GPU.

2.2k Python Language Models · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.