VulnClaw exists so that a single natural-language sentence can trigger the entire reconnaissance-to-report pipeline without manually orchestrating a dozen separate tools.
LLMOps · Eval
underdogs breaking outAn agent skill that forces you to approve the visual metaphor and static layout before it spends your Gemini credits on video generation.
A single Zig binary with an embedded Svelte dashboard that installs, supervises, and cross-wires local AI agents, workflow engines, and tracing tools so you don't have to juggle separate terminals.
A skill that spawns parallel reasoning processes under distorted cognitive frames, then scores and prunes them with a separate critic pass.
Hallmark is a design skill that stops Claude, Cursor, and Codex from producing the same generic AI slop.
Raven wraps agents in durable memory and self-refining skills so workflows survive past the chat session.
iFixAi runs up to 32 inspections against any LLM or agent and returns a letter-grade scorecard in minutes, using a separate provider as judge so the model isn't grading its own homework.
codex-keysmith exists because copying a Markdown file and editing one TOML key is simple, but existing files, active hooks, and interrupted writes are not.
pilotfish keeps Claude Fable 5 in the boardroom by delegating the grunt work to cheaper Sonnet and Haiku subagents.
Hyperresearch turns Claude Code into a tiered, multi-agent research pipeline that persists every source in a searchable, compounding vault.
It exists to stop your AI gateway from quietly burning through quotas, cash, and expired OAuth tokens without leaving a paper trail.
Token Monitor reads local logs from two dozen AI coding tools to surface live token burn, costs, and limits in one place, synced across all your machines.
Moss exists because calling out to a remote vector database adds 200–500 ms of latency—enough to kill a real-time conversation—so it runs embedding and search inside your process instead.
Agentlas OS compiles plain-language requests into portable, ownable agent packages that run locally across whatever LLM host you already use, while a Hub and owner-scoped Cloud handle sharing and retrieval.
It packages a professor's decade of SIGMOD and NeurIPS experience into structured AI skills, bridging the 'last-mile' gap where generic guides and busy advisors leave grad students stranded.
PhyAgentOS treats robot hardware like pluggable drivers so the same agentic session can run in simulation or on a real arm without rewriting the control stack.
This runtime layer gives Codex a structured, opt-in red-team workflow that stays out of your way until you explicitly enable it.
AIHelms wraps LiteLLM in a Vue management layer so finance can trace every token back to the department that spent it.
Books are too good to leave on the shelf; this system distills them into structured, callable agent skills.
repowise indexes a codebase into five queryable intelligence layers—dependency graphs, git history, docs, architectural decisions, and deterministic health scores—so MCP-compatible agents can answer "why" instead of grepping for "what".



