Skip to content
Back to skills

Benchmark

ASecurity

Benchmark commands and HTTP load with hyperfine, oha, and Locust.

  • 10 stars
  • 0 votes
  • 0 copies
  • 0 views
  • Added September 23, 2026
ai-agentsdebugginggitperformance

Security analysis

A100/100

Pro scans all 3 files and shows the line behind each finding

Scanned October 6, 2026

npx -y skills add fmind/dot --skill benchmark --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Benchmark?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Benchmark
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/fmind-benchmark/badge)](https://www.skillsdirectory.com/skills/fmind-benchmark)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
name: benchmark
description: "Benchmark commands and HTTP load with hyperfine, oha, and Locust."
license: MIT
metadata:
  kind: task
  author: Médéric HURIER (Fmind)
  source: github.com/fmind/dot/tree/main/skills/benchmark
  created: "2026-09-02"
  updated: "2026-10-05"
---

# Benchmark

Produce comparable measurements for a stated performance question. Use hyperfine for commands, oha for simple HTTP load, and Locust for realistic concurrent user scenarios; diagnosing why something is slow belongs to [systematic-debugging](../systematic-debugging/SKILL.md).

## Workflow

1. **Fix the question**: one command or endpoint, one metric (mean latency, p99, requests per second), one hypothesis.
1. **Confirm authority**: never load-test a remote service you do not own or a production system without explicit approval; agree a load bound and stop condition appropriate to the target's capacity.
1. **Control the machine**: close heavy processes, run on AC power, and pin versions; record CPU, OS, and tool versions in the report.
1. **Measure with the matching guide**: read only that guide and its required resources; use at least 3 warmup runs and 10 measured runs for commands and at least 30 seconds for endpoints, and compare against a baseline measured the same way in the same session.
1. **Quantify uncertainty**: compare repeated, equivalently controlled runs and report uncertainty in the difference; a run's standard deviation alone does not decide significance. Alternate baseline/candidate measurements when machine drift matters.
1. **Report**: the command lines, the exported table, the relative change, and the conditions. Keep the raw export (`bench.json`, `oha.json`, or Locust CSV) if the number will be tracked over time.

## Task guides

<!-- guides:start -->

- [command-http](references/command-http.md): Compare command latency with hyperfine or measure HTTP throughput with oha.
- [locust](references/locust.md): Concurrent user scenarios, capacity tests, and acceptance thresholds.

<!-- guides:end -->

Files in this skill

  • SKILL.md923 B
  • references/command-http.md3.5 KB
  • references/locust.md2.4 KB

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…