Move the code and terraform audits into the reviews plugin

Copy the standalone code-review and terraform-review skills into
plugins/reviews as audit-code and audit-terraform. The rename separates the
automated, linter-driven audits from the guided review-pr walkthrough that
already lived here.

Resolve bundled script paths through ${SKILL_DIR}, exported in a new step 0.
CLAUDE_PLUGIN_ROOT is not set in the Bash tool environment, so the obvious
substitution would have expanded to nothing and broken every collection
script invocation.

Replace the PLAN and DESIGN docs with READMEs written from the current
SKILL.md and scripts. The old docs had drifted badly: they named semgrep
where the code calls opengrep, scoped five review agents where there are
now eight, and predated Lua, PowerShell, and GitHub Actions support.

Add CONSISTENCY_NORMS to the audit-terraform agent inputs. The collection
script writes consistency_norms.json and the agent prompt declares it, but
SKILL.md never listed it, leaving the variable unsubstituted.

Drop the --ingest-verdicts instruction from both skills. review_stats.py
parses no arguments, so the ref-mode verdict template it told users to feed
back could never be read.

Point audit-terraform's smoke test at README.md and resolve its fixture
paths relative to the test file rather than an absolute home directory.

Tests: 197 passing (audit-code), 106 passing (audit-terraform).
This commit is contained in:
2026-07-21 11:11:05 -05:00
parent 600c1fef86
commit f5934181ec
179 changed files with 20779 additions and 3 deletions
@@ -0,0 +1,84 @@
"""Scrape the CIS AWS Foundations Benchmark page.
The page is a mapping table:
| Control ID and title (FSBP-style) | CIS v5.0.0 | CIS v3.0.0 | CIS v1.4.0 | CIS v1.2.0 |
We emit one CISRow per unified control. The `requirement` field summarises
which CIS versions reference it.
"""
from __future__ import annotations
import re
from dataclasses import dataclass, field
from bs4 import BeautifulSoup
from scripts.control_resource_map import control_id_to_types
CIS_URL = (
"https://docs.aws.amazon.com/securityhub/latest/userguide/"
"cis-aws-foundations-benchmark.html"
)
@dataclass
class CISRow:
control_id: str
title: str
severity: str
requirement: str
source_url: str
resource_types: list[str] = field(default_factory=list)
_TITLE_RE = re.compile(r"^\s*\[(?P<id>[A-Za-z][A-Za-z0-9]*\.\d+)\]\s*(?P<title>.+)$")
_VERSION_RE = re.compile(r"CIS\s+v(?P<ver>[\d.]+)", re.IGNORECASE)
def _extract_versions(headers: list[str]) -> list[str]:
"""From header cells, pull out version strings like '5.0.0' for each
column. Non-version columns yield empty string."""
versions: list[str] = []
for h in headers:
m = _VERSION_RE.search(h)
versions.append(m.group("ver") if m else "")
return versions
def parse_cis_page(html: str, source_url: str) -> list[CISRow]:
soup = BeautifulSoup(html, "html.parser")
rows: list[CISRow] = []
for table in soup.find_all("table"):
header_cells = table.find("tr").find_all(["th", "td"]) if table.find("tr") else []
header_texts = [c.get_text(strip=True) for c in header_cells]
versions = _extract_versions(header_texts)
if not any(versions):
continue
for tr in table.find_all("tr")[1:]:
cells = tr.find_all("td")
if len(cells) < 2:
continue
title_text = cells[0].get_text(strip=True)
m = _TITLE_RE.match(title_text)
if not m:
continue
unified_id = m.group("id")
title = m.group("title").strip()
version_refs: list[str] = []
for ver, cell in zip(versions[1:], cells[1:]):
if not ver:
continue
num = cell.get_text(strip=True)
if num:
version_refs.append(f"v{ver} §{num}")
requirement = "CIS " + ", ".join(version_refs) if version_refs else ""
rows.append(CISRow(
control_id=f"CIS {unified_id}",
title=title,
severity="medium",
requirement=requirement,
source_url=source_url,
resource_types=control_id_to_types(f"CIS {unified_id}"),
))
return rows