Skip to content
Back to skills

Recovery Evaluation

ASecurity

Evaluate soft-mode APT recovery records for source-boundary, metric, and mechanism-faithfulness checks.

  • 247 stars
  • 0 votes
  • 0 copies
  • 1 view
  • Added September 9, 2026
testingpython

Security analysis

A100/100

Scanned September 9, 2026

npx -y skills add VectorSpaceLab/AREX-Skill --skill recovery_evaluation --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Recovery Evaluation?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Recovery Evaluation
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/vectorspacelab-recovery-evaluation/badge)](https://www.skillsdirectory.com/skills/vectorspacelab-recovery-evaluation)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

SKILL.md
---
name: recovery_evaluation
description: Evaluate soft-mode APT recovery records for source-boundary, metric, and mechanism-faithfulness checks.
---

# Recovery Evaluation

Use this skill after a recovery harness has produced `recovery_result.json`. It is especially useful for soft-mode APT runs where the selected target is reduced or proxy rather than a full benchmark reproduction.

## Inputs
- Recovery result JSON with `metrics`, `paper_target`, and `mechanism_checks`.
- Experiment validation JSON from the recovery gate.
- Generated skill invocation log.

## Outputs
- Boolean checks for numeric metric, target consistency, gate pass, optimizer execution, and mechanism coverage.
- A concise recommendation of `accept` or `refine` for the analysis phase.

## Workflow
1. Verify that the recovery gate reports `ok: true`.
2. Confirm the configured metric is present and numeric.
3. Require mechanism flags for proposal correction, atomic loss, sequential update, generated skill invocation, and reduced optimizer execution when proxy mode is used.
4. Return actionable missing-check feedback rather than silently accepting a high scalar metric.

## Validation
Run `python scripts/evaluate_recovery_record.py --self-test`.

## Limitations
This skill complements but does not replace the Distiller recovery validator or final analysis report.

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…