Skip to content
Back to skills

Text And Tabular

ASecurity

"Operate AutoTrain Advanced text, NLP, sentence-transformers,

  • 247 stars
  • 0 votes
  • 0 copies
  • 1 view
  • Added September 8, 2026
toolspythonbashbackend

Works with

  • cli

Security analysis

A100/100

Pro scans all 5 files and shows the line behind each finding

Scanned September 8, 2026

npx -y skills add VectorSpaceLab/AREX-Skill --skill text-and-tabular --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Text And Tabular?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Text And Tabular
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/vectorspacelab-text-and-tabular/badge)](https://www.skillsdirectory.com/skills/vectorspacelab-text-and-tabular)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
name: text-and-tabular
description: "Operate AutoTrain Advanced text, NLP, sentence-transformers,
  extractive QA, and tabular training workflows."
disable-model-invocation: true
metadata:
  disco-role: operating
  parent-skill: autotrain-advanced
license: Apache 2.0
---

# AutoTrain text, NLP, embeddings, and tabular workflows

Use this sub-skill for non-LLM text tasks, sentence-transformers, extractive QA, and tabular workflows.

## Supported entry points

- `autotrain text-classification --help`
- `autotrain text-regression --help`
- `autotrain token-classification --help`
- `autotrain seq2seq --help`
- `autotrain sentence-transformers --help`
- `autotrain tabular --help`
- `autotrain extractive-qa --help`
- YAML aliases such as `text-classification`, `text-regression`, `token-classification`, `seq2seq`, `extractive-qa`, `sentence-transformers:pair`, `st:triplet`, and `tabular`

If the request is LLM-specific, route to `../llm-training/`; the data validator in this sub-skill still supports `--task llm` for local column checks.

## Task families

| Family | Typical data shape | Notes |
| --- | --- | --- |
| text classification/regression | text column + target/label column | Binary and multi-class classification resolve to the same text-classification trainer family. |
| token classification | token list column + tag list column | Values may be Python-list strings that AutoTrain parses with `ast.literal_eval`. |
| seq2seq | source text + target text | Use for summarization/translation-style data. |
| extractive QA | context text + question + answer dict/string | Answer values should be dict-like or parseable into dicts. |
| sentence-transformers | pair/triplet/QA rows | Trainer variants control whether target labels/scores are needed. |
| tabular | id column + one or more target columns | `task` distinguishes classification vs regression. |

## Safe validation sequence

1. Inspect the specific command help with the root `inspect_cli.py` helper.
2. Validate YAML configs with the root `validate_config.py` helper.
3. Validate local CSV/JSONL columns with `scripts/validate_text_data.py`.
4. Launch training only after model, data path, split names, backend, and Hub credentials are confirmed.

## Useful script

```bash
python skills/disco/autotrain-advanced/sub-skills/text-and-tabular/scripts/validate_text_data.py \
  --task text-classification \
  --text-column text \
  --target-column label \
  data.csv
```

The script performs local schema checks only. It does not import trainer code, upload data, or start training.

## References

- `references/workflows.md` — task/alias map and command/config patterns.
- `references/data-formats.md` — column schemas and validator examples.
- `references/troubleshooting.md` — column mapping, parser, Hub, and local dataset recovery.

Files in this skill

  • SKILL.md2.8 KB
  • references/data-formats.md2.5 KB
  • references/troubleshooting.md2.3 KB
  • references/workflows.md2.8 KB
  • scripts/validate_text_data.py9.1 KB

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…