Data Tooling

Data Tooling

underdogs · picking up speed
01
TickDB/tickdb-unified-realtime-marketdata-api
+23% /wk +25 ★/dayaccelerating

TickDB open-sources its MCP server, Skill definitions, and CLI so AI agents can query unified real-time market data across stocks, forex, and crypto without writing REST clients.

776 Python Data Tooling · explained
02
magicrew/doc7
+21% /wk +33 ★/dayaccelerating

doc7 converts documents into Markdown by showing page images to a local vision model, bypassing OCR pipelines and cloud per-page fees.

1.1k Go Data Tooling · explained
03
HiThink-Tech/Financial-API
+29% /wk +126 ★/dayaccelerating

To give AI agents and quant scripts a legitimate, unified pipe into official Tonghuashun A-share market data.

3.1k TypeScript Data Tooling · explained Feature
04
Rizzo-AI-Academy/rizzo-pii
+14% /wk +19 ★/dayaccelerating

It exists so Italian law firms can use ChatGPT on sensitive documents without ever shipping a codice fiscale off their laptop.

956 Python Data Tooling · explained
05
liliu-z/stashbase
+11% /wk +9.4 ★/dayaccelerating

It builds an agent-ready knowledge base over your local files without dragging them into a proprietary workspace.

605 TypeScript RAG · Search · explained
07
lycohana/BiliSum
+9.0% /wk +7.4 ★/dayaccelerating

BiliSum exists to convert Bilibili, YouTube, and local videos into structured, offline notes and a personal knowledge base that keeps data off the cloud.

577 Python RAG · Search · explained
08
brightdata/brightdata-mcp
+5.7% /wk +21 ★/dayaccelerating

An MCP server that hands LLMs live web access, anti-bot evasion, and 60+ data tools—if you pay past the first 5,000 requests.

2.6k JavaScript Coding Assistants · explained
09
ix-infrastructure/Ix
+6.9% /wk +8.7 ★/dayaccelerating

Ix builds a persistent symbol graph of your repository so both you and your AI assistant can ask structural questions instead of grepping blind.

882 TypeScript Coding Assistants · explained
10
bawadou/ai-data-extractor
+4.9% /wk +3.9 ★/dayaccelerating

Your AI coding conversations are trapped in undocumented SQLite schemas and scattered log files; this extracts them before the next update wipes everything.

555 Python Coding Assistants · explained
11
xerj-org/xerj
+8.2% /wk +22 ★/dayaccelerating

A Rust search engine that auto-indexes folders so AI agents can query code and documents without stuffing their context windows.

1.9k Rust RAG · Search · explained
12
huggingface/huggingface_hub
+4.0% /wk +22 ★/dayaccelerating

A Python client that turns the Hub's model zoo into a local filesystem with caching, search, and inference baked in.

3.9k Python Data Tooling · explained
13
antvis/mcp-server-chart
+3.6% /wk +22 ★/dayaccelerating

An MCP server that turns conversational data requests into AntV visualizations without you picking colors or chart types.

4.4k TypeScript Coding Assistants · explained
14
ScrapingBee/chatgpt-scraper-api
+14% /wk +20 ★/dayaccelerating

Documentation and SDK examples for a paid API that replaces brittle CSS selectors with natural-language prompts and optional live search.

958 RAG · Search · explained
15
OpenSenseNova/SenseNova-Skills
+5.4% /wk +43 ★/dayaccelerating

Modular agent skills that hand LLMs the boring office work—Excel analysis, slide decks, research reports, and infographics—instead of letting them hallucinate about it.

5.5k JavaScript Agents · explained
16
grobidOrg/grobid
+2.9% /wk +22 ★/dayaccelerating

A battle-tested Java toolkit that extracts metadata, references, and full text from academic PDFs using a cascade of ML models.

5.1k Java Domain Apps · explained
17
memgraph/memgraph
+3.2% /wk +21 ★/dayaccelerating

Memgraph wants to be the single database operation your GraphRAG pipeline actually needs.

4.5k C++ RAG · Search · explained
18
Eventual-Inc/Daft
+2.7% /wk +23 ★/dayaccelerating

Daft is a Rust-backed dataframe engine that treats images, audio, and embeddings as native types rather than generic Python objects, letting AI pipelines scale from a laptop to a Ray cluster without JVM baggage.

5.8k Rust Data Tooling · explained
19
huohuoer/wechat-cli
+6.8% /wk +22 ★/dayaccelerating

Scans running WeChat process memory to extract SQLCipher keys, then exposes your chat database as a JSON-first CLI designed for AI agent consumption.

2.3k Coding Assistants · explained
20
Ontos-AI/knowhere
+7.9% /wk +35 ★/dayaccelerating

A pipeline that turns messy PDFs and slides into structured, navigable memory for AI agents instead of flat text shards.

3.1k Python RAG · Search · explained
loading more…

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.