Agent Plugins Marketplace
All plugins

hermes-jailbench

v0.2.1

Jailbreak regression benchmark for LLM endpoints with repeatable known-pattern attacks and deterministic scoring

Claude CodeAgent Plugins1 Skill

By Hermes LabsLicense: MIT3 GitHub starsUpdated 6 days ago

Directory evidence

Runtimes
Claude Code and Agent Plugins
Parsed components
1 skill or MCP entry
Source updated
Sep 17, 2026
Manifest status
Canonical path parsed

The directory validates manifest shape and source location. It does not execute the plugin or provide a security endorsement. Review the indexing methodology

Install hermes-jailbench for Claude Code

Installs for the current user
claude plugin marketplace add IchenDEV/agent-plugin-mkt
claude plugin marketplace update agent-plugin-marketplace
claude plugin install hermes-jailbench@agent-plugin-marketplace

Paste and run these commands in a terminal with Claude Code. They add and refresh the PluginsMP catalog, then install this plugin.

The installer fetches third-party code from the source repository shown on this page. This directory validates manifest structure and source location, but does not perform a security audit; review the manifest, components, and source before installing.

Get the source manually
git clone https://github.com/hermes-labs-ai/hermes-jailbench

Clone the source repository, then follow its setup instructions to add the plugin to a compatible client. The repository root is the plugin root.

Plugin files

hermes-jailbench/
├── .claude-plugin/plugin.json
├── plugin.json
└── skills/hermes-jailbench/SKILL.md

Included Skills1

hermes-jailbenchskills/hermes-jailbench/SKILL.md

Run the hermes-jailbench safety regression check against a model endpoint the user owns or is authorized to test, then summarize the report. Trigger when the user wants to confirm that a model, system prompt, or release change did not weaken refusals on a fixed set of known patterns, or wants a pass/fail safety gate in CI.

Plugin manifests2

.claude-plugin/plugin.json
{
  "name": "hermes-jailbench",
  "version": "0.2.1",
  "description": "Jailbreak regression benchmark for LLM endpoints with repeatable known-pattern attacks and deterministic scoring",
  "author": {
    "name": "Hermes Labs",
    "email": "[email protected]"
  },
  "homepage": "https://github.com/hermes-labs-ai/hermes-jailbench",
  "repository": "https://github.com/hermes-labs-ai/hermes-jailbench",
  "license": "MIT"
}
plugin.json
{
  "$schema": "https://agent-plugins.org/schemas/1.0.0/plugin.schema.json",
  "name": "hermes-jailbench",
  "version": "0.2.1",
  "description": "Jailbreak regression benchmark for LLM endpoints with repeatable known-pattern attacks and deterministic scoring",
  "author": {
    "name": "Hermes Labs",
    "email": "[email protected]"
  },
  "homepage": "https://github.com/hermes-labs-ai/hermes-jailbench",
  "repository": "https://github.com/hermes-labs-ai/hermes-jailbench",
  "license": "MIT",
  "keywords": [
    "ai-safety",
    "regression-testing",
    "agent-skills",
    "claude-code",
    "codex",
    "gemini-cli"
  ]
}

If you maintain this plugin, link to this source-backed listing from your README so users can review its manifest and indexed components.

[hermes-jailbench on Agent Plugins Marketplace](https://pluginsmp.com/plugins/hermes-jailbench)