Skip to content
Back to skills

Vision Multimodal

ASecurity

"Operate AutoTrain Advanced image classification, image regression,

  • 247 stars
  • 0 votes
  • 0 copies
  • 1 view
  • Added September 8, 2026
toolsapibackend

Works with

  • cli
  • api

Security analysis

A100/100

Pro scans all 5 files and shows the line behind each finding

Scanned September 8, 2026

npx -y skills add VectorSpaceLab/AREX-Skill --skill vision-multimodal --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Vision Multimodal?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Vision Multimodal
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/vectorspacelab-vision-multimodal/badge)](https://www.skillsdirectory.com/skills/vectorspacelab-vision-multimodal)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
name: vision-multimodal
description: "Operate AutoTrain Advanced image classification, image regression,
  object detection, and VLM app/API/config workflows."
disable-model-invocation: true
metadata:
  disco-role: operating
  parent-skill: autotrain-advanced
license: Apache 2.0
---

# AutoTrain vision and multimodal workflows

Use this sub-skill for image classification, image regression/scoring, object detection, and VLM dataset/task flows.

## Supported entry points

- `autotrain image-classification --help`
- `autotrain image-regression --help`
- `autotrain object-detection --help`
- YAML aliases: `image-classification`, `image-regression`, `image-scoring`, `object-detection`, `image-object-detection`, `vlm:captioning`, `vlm:vqa`
- App/API task keys: `image-classification`, `image-regression`, `image-object-detection`, `vlm:captioning`, `vlm:vqa`

Important: there is no top-level `autotrain vlm` command in this checkout. Route VLM through the app/API/config flow.

## Data-layout summary

| Task | Local layout |
| --- | --- |
| image classification | directory with at least two class subfolders; each class folder has at least five jpg/jpeg/png files and no extra files/subfolders |
| image regression | directory with at least five images plus `metadata.jsonl` containing `file_name` and `target` |
| object detection | directory with at least five images plus `metadata.jsonl` containing `file_name` and `objects` |
| VLM | directory with at least five images plus `metadata.jsonl` containing `file_name` and all mapped text/prompt columns |

Use `scripts/validate_vision_data.py` for bounded local layout checks.

## Safe validation sequence

1. Inspect relevant CLI help with the root `inspect_cli.py` helper.
2. Validate YAML with the root `validate_config.py` helper.
3. Validate local image folders/metadata with `scripts/validate_vision_data.py`.
4. If VLM or UI/API parameters are involved, use `../app-backends/` to inspect app/API task params and backend auth.
5. Launch only after data layout, model id, task alias, backend, and Hub credentials are explicit.

## References

- `references/workflows.md` — route patterns and launch/check examples.
- `references/data-formats.md` — local folder and metadata schemas.
- `references/troubleshooting.md` — image file counts, metadata parsing, VLM route, and backend issues.

Files in this skill

  • SKILL.md2.3 KB
  • references/data-formats.md2 KB
  • references/troubleshooting.md2.1 KB
  • references/workflows.md2.5 KB
  • scripts/validate_vision_data.py9 KB

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…