Skip to content
Back to skills

Sample Factory Async Topology

ASecurity

Design and validate Sample Factory-style asynchronous rollout, policy, and learner topology with policy-lag estimates.

  • 247 stars
  • 0 votes
  • 0 copies
  • 1 view
  • Added September 9, 2026
developmentpythonbash

Security analysis

A100/100

Scanned September 9, 2026

npx -y skills add VectorSpaceLab/AREX-Skill --skill sample_factory_async_topology --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Sample Factory Async Topology?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Sample Factory Async Topology
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/vectorspacelab-sample-factory-async-topology/badge)](https://www.skillsdirectory.com/skills/vectorspacelab-sample-factory-async-topology)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

SKILL.md
---
name: sample_factory_async_topology
description: Design and validate Sample Factory-style asynchronous rollout, policy, and learner topology with policy-lag estimates.
---

# Sample Factory Asynchronous Topology

Use this skill when designing or auditing a Sample Factory-style APPO training harness where environment simulation, policy inference, and learning should be separated into asynchronous components.

Do not use it to claim real throughput without a wall-clock benchmark. It provides topology contracts and deterministic estimates.

## Inputs
- Number of rollout workers, policy workers, and learners.
- Number of environments per rollout worker.
- Trajectory length and learner batch size.
- Optional component responsibility map.

## Outputs
- Component roles for rollout workers, policy workers, and learner.
- Queue contracts for observation requests, action replies, completed trajectories, and policy updates.
- Produced samples per rollout iteration.
- Policy-lag pressure estimate.

## Workflow
1. Keep rollout workers environment-only: they step environments and write transition fields.
2. Keep policy workers stateless: they batch observation-buffer indices and return action-buffer indices.
3. Keep the learner as the only component that consumes completed trajectories and updates trainable parameters for a policy.
4. Use compact queue messages that carry buffer indices, not serialized observations or trajectories.
5. Estimate lag pressure as `max(0, produced_samples / learner_batch_size - 1)`.
6. Reject designs where rollout workers compute gradients or policy workers modify rewards/returns.

## Validation
Run:

```bash
python scripts/topology_contract.py --rollout-workers 2 --policy-workers 1 --learners 1 --envs-per-worker 4 --trajectory-length 8 --learner-batch-size 32
python tests/test_topology_contract.py
```

## Limitations
The estimate abstracts away real IPC, GPU scheduling, and environment variance. Use it as mechanism guidance before running real profiling.

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…