Repositories list
22 repositories
little-canary
PublicDetects prompt injection by its effect on a sacrificial canary model, not just pattern matching: untrusted input hits a powerless model first, a behavioral chec…lintlang
PublicStatic analysis for AI agent configs, tool descriptions, and system prompts — catches vague tool descriptions, missing stop conditions, and schema gaps before t…plugins
PublicCanonical generated plugin catalog for Claude Code, Codex, GitHub Copilot, and portable Agent Pluginshomebrew-tap
PublicHomebrew formulae for Hermes Labs softwarefidelis
PublicZero-LLM agent memory for Claude Code and AI agents: local-first BM25, dense-vector, and reciprocal-rank-fusion retrieval. Returns original passages verbatim by…hermes-blind
PublicRecovers the original goal of a long Claude Code or Codex session from its first user turn, so you can restate it before continuing — plus a prompt wrapper that…supersearch
PublicDeadline-bounded local search fan-out with source-explicit JSON receipts for agents and engineers.langstate
PublicInspectable context compression for LLM conversations: turn older history into visible scaffold state, keep recent turns verbatim, and check named facts with le…hermes-rubric
PublicEvidence-first LLM-as-judge scoring for AI artifacts — papers, PRs, prompts, cold emails: synthesizes a rubric, collects quoted-evidence citations, scores only …hermeneutic
PublicTurn corrections from AI agent logs into guidance for similar tasks. Local memory with semantic retrieval, prompt-context hooks, and response checks.- Hermes Labs offline CVE JSON checker for CVSS 3 score/vector consistency
hermes-jailbench
PublicZero-LLM deterministic jailbreak regression benchmark: runs a repeatable battery of known-pattern attacks against an LLM endpoint and scores refusal/partial/com…zer0dex
PublicA local dual-layer memory pattern for AI agents: a compact, human-readable markdown index paired with semantic retrieval from a local vector store, queried befo…hermes-publications
PublicResearch and DOI publications hub for Hermes Labs: a Zenodo-canonical, DOI-anchored index of papers on epistemic and hermeneutic failure modes in large language…tool-differentia
PublicArchival repository for the Zenodo technical note Tool Differentia: Relational Static Analysis for AI Agent Tool Descriptions. The check itself ships in hermes-…behavioral-canarying
PublicArchival repository for the Zenodo technical note Behavioral Canarying for Prompt Injection: Powerless Model Probes with Explicit Coverage Semantics..github
PublicAI reliability engineering for production AI agents and LLM applications. Open-source tools and research that find silent failures in prompts, tools, retrieval,…agent-gorgon
PublicRuntime policy guard for AI agent processes: watches process, file, and network activity from user space, applies deterministic policy, and attempts SIGSTOP or …- Archival repository for the Zenodo paper The Asymmetric Burden of Proof.
- Archival repository for the Zenodo paper A Taxonomy of Epistemic Failure Modes in Large Language Models.
the-generative-horizon
PublicApplied hermeneutics, linguistic attractors, and the limits of model self-report in language-model systems.- Measurement validity, construct validity, and unsupported claims derived from AI agent telemetry.
ProTip! When viewing an organization's repositories, you can use the
props. filter to filter by custom property.