Open-source Claude skills that run SEO and generative-engine optimization on your actual Google data instead of scraped estimates.
Data Tooling
underdogs breaking outSolarWM exists to make long-horizon video world models reproducible, open-sourcing the full stack from a unified data contract to staged training recipes and checkpoints across three major backbones.
To give AI agents and quant scripts a legitimate, unified pipe into official Tonghuashun A-share market data.
EverRoom exists because the hardest part of using AI agents is everything before the prompt: finding relevant material, remembering past decisions, and keeping generated work tied to its sources.
It exists to deliver a point-in-time Postgres diagnosis from a single static binary, no write privileges or external services required.
An open-source pipeline that curates, licenses, and retrieves task-specific instructions so agents don't have to wing it.
TickDB open-sources its MCP server, Skill definitions, and CLI so AI agents can query unified real-time market data across stocks, forex, and crypto without writing REST clients.
doc7 converts documents into Markdown by showing page images to a local vision model, bypassing OCR pipelines and cloud per-page fees.
Utopia ingests your documents into a self-hosted, time-aware knowledge graph so you can ask what was true last quarter—and prove it with citations.
Documentation and SDK examples for a paid API that replaces brittle CSS selectors with natural-language prompts and optional live search.
It exists so Italian law firms can use ChatGPT on sensitive documents without ever shipping a codice fiscale off their laptop.
Jonex unifies multimodal ingestion, ontology compilation, and source-grounded retrieval into a single governed enterprise platform.
It builds an agent-ready knowledge base over your local files without dragging them into a proprietary workspace.
OpenLake wants storage to bypass the host entirely and land straight in GPU memory.
A self-hosted Go API that turns websites into LLM-ready markdown or structured JSON, keeping most scraping local and only burning paid fallbacks when a page actually demands it.
It gives AI agents complete social media research workflows—outlier detection, comment mining, competitor teardowns—that produce actionable business artifacts instead of raw API responses.
BiliSum exists to convert Bilibili, YouTube, and local videos into structured, offline notes and a personal knowledge base that keeps data off the cloud.
A Rust search engine that auto-indexes folders so AI agents can query code and documents without stuffing their context windows.
A pipeline that turns messy PDFs and slides into structured, navigable memory for AI agents instead of flat text shards.
Ix builds a persistent symbol graph of your repository so both you and your AI assistant can ask structural questions instead of grepping blind.




