Language Models

Language Models

big names on the move
01
jingyaogong/minimind
+328 ★/daycooling

MiniMind is an educational training ground that rebuilds every stage of a modern language model—from tokenizer to RLHF—in raw PyTorch so you can see the gears turning instead of just calling high-level APIs.

60.6k Python Language Models · explained Feature
02
TauricResearch/TradingAgents
+298 ★/dayaccelerating

A research framework that assigns LLMs to trading-floor roles—analyst, researcher, trader, risk manager—to debate and execute simulated stock decisions.

104.5k Python Agents · explained Feature
03
sgl-project/sglang
+243 ★/dayaccelerating

SGLang exists to push low-latency, high-throughput inference for LLMs and multimodal models from a single GPU up to massive clusters.

35.8k Python Inference · Serving · explained
04
nashsu/llm_wiki
+137 ★/dayaccelerating

It turns your document pile into a persistent, interlinked wiki so the LLM doesn't have to re-read everything every time you ask a question.

18.1k TypeScript RAG · Search · explained
05
Tencent/WeKnora
+133 ★/dayaccelerating

WeKnora exists to turn scattered enterprise documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining wiki.

22.1k Go RAG · Search · explained
06
ggml-org/llama.cpp
+115 ★/daycooling

It exists to run large language models on virtually any hardware—from Apple Silicon to RISC-V to your browser—with zero external dependencies and minimal setup.

127.8k C++ Inference · Serving · explained
07
p-e-w/heretic
+114 ★/daycooling

It automates the removal of transformer safety alignment so you don't have to hand-tune abliteration parameters or pay for expensive post-training.

31.1k Python Language Models · explained
08
agentscope-ai/agentscope
+106 ★/daysteady

A Python framework for building production multi-agent systems that leans on LLM reasoning instead of rigid prompt choreography.

31.3k Python Agents · explained
09
ollama/ollama
+74 ★/dayaccelerating

It exists so you can download, run, and chat with open-weight LLMs locally through one CLI and REST API, keeping inference on your own silicon.

180.6k Go Inference · Serving · explained
10
chatanywhere/GPT_API_free
+73 ★/dayaccelerating

A hosted proxy that offers free, rate-limited API access to GPT, DeepSeek, and others for Chinese users who'd rather not tunnel through a VPN.

42.3k Inference · Serving · explained
11
BerriAI/litellm
+70 ★/daycooling

Because swapping from GPT-4o to Claude shouldn't require rewriting your request plumbing.

58.5k Python LLMOps · Eval · explained
12
openai/whisper
+70 ★/dayaccelerating

To give developers a single, general-purpose speech model that handles transcription, translation, and language identification by treating tasks as tokens to predict.

108.9k Python Image · Video · Audio · explained
13
lyogavin/airllm
+64 ★/daycooling

AirLLM slices giant transformers into layer shards so they fit in consumer VRAM without quantization or distillation.

34.1k Jupyter Notebook Inference · Serving · explained Feature
14
rasbt/LLMs-from-scratch
+60 ★/dayaccelerating

It teaches how LLMs work by implementing tokenization, attention, pretraining, and finetuning in pure PyTorch, one notebook at a time.

104.7k Jupyter Notebook Language Models · explained
15
huggingface/transformers
+47 ★/dayaccelerating

It centralizes model definitions so the same architecture works across PyTorch, JAX, vLLM, and llama.cpp without rewrites.

165.1k Python Language Models · explained
17
liyupi/ai-guide
+42 ★/dayaccelerating

Curated tutorials and tool reviews covering vibe coding, DeepSeek, Cursor, and the rest of the generative-AI menagerie, maintained as a free, open-source knowledge base.

19.8k JavaScript Learning · explained
18
huggingface/agents-course
+36 ★/daycooling

To teach developers agent engineering from scratch using smolagents, LangGraph, and LlamaIndex, capped by an automated benchmark.

32.4k MDX Learning · explained
19
shiyu-coder/Kronos
+35 ★/daycooling

Kronos recasts noisy, multi-dimensional candlestick data as hierarchical discrete tokens so an autoregressive Transformer can forecast financial markets like a language model.

38.6k Python Language Models · explained Feature
20
datawhalechina/happy-llm
+29 ★/dayaccelerating

A systematic Chinese tutorial for developers who want to stop treating LLMs as black boxes and hand-build a 215-million-parameter model from the ground up.

33.7k Jupyter Notebook Learning · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.