Skip to content
Back to skills

Llm Training

ASecurity

"Operate AutoTrain Advanced LLM finetuning workflows, config

  • 247 stars
  • 0 votes
  • 0 copies
  • 3 views
  • Added September 8, 2026
toolspythonapibackend

Works with

  • cli
  • api

Security analysis

A100/100

Pro scans all 3 files and shows the line behind each finding

Scanned September 8, 2026

npx -y skills add VectorSpaceLab/AREX-Skill --skill llm-training --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Llm Training?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Llm Training
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/vectorspacelab-llm-training/badge)](https://www.skillsdirectory.com/skills/vectorspacelab-llm-training)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
name: llm-training
description: "Operate AutoTrain Advanced LLM finetuning workflows, config
  aliases, data columns, PEFT/quantization knobs, and adapter flow."
disable-model-invocation: true
metadata:
  disco-role: operating
  parent-skill: autotrain-advanced
license: Apache 2.0
---

# AutoTrain LLM training

Use this sub-skill for `autotrain llm`, LLM YAML configs, app/API LLM task keys, PEFT/LoRA, quantization, unsloth, and adapter-related decisions.

## Supported entry points

- CLI: `autotrain llm --help`
- CLI training: `autotrain llm --train ...`
- YAML config aliases: `llm`, `llm-sft`, `llm-dpo`, `llm-orpo`, `llm-reward`, `llm-generic`
- App/API task keys: `llm:sft`, `llm:dpo`, `llm:orpo`, `llm:reward`, `llm:generic`
- Config examples: `configs/llm_finetuning/*.yml`

Do not suggest `--deploy` or `--inference` as working LLM commands; in this checkout those branches raise `NotImplementedError`.

## Required CLI fields for training

For `autotrain llm --train`, make sure the user provides at least:

- `--project-name`
- `--data-path`
- `--model`

If `--push-to-hub` is used, `--username` and `--token` are required. Hosted backends such as `spaces-*` and `ep-*` also require Hub push credentials.

## Common LLM knobs

- Data: `data_path`, `train_split`, `valid_split`, `text_column`, `prompt_text_column`, `rejected_text_column`, `chat_template`.
- Sequence: `block_size` / `block-size`, `model_max_length`, `max_prompt_length`, `max_completion_length`, `padding`, `add_eos_token`.
- Training: `trainer`, `epochs`, `batch_size`, `gradient_accumulation`, `lr`, `scheduler`, `optimizer`, `mixed_precision`, `auto_find_batch_size`.
- PEFT/LoRA: `peft`, `quantization`, `target_modules`, `lora_r`, `lora_alpha`, `lora_dropout`, `merge_adapter`.
- Preference tuning: `model_ref`, `dpo_beta`, `rejected_text_column`, `prompt_text_column`.
- Acceleration: `use_flash_attention_2`, `unsloth`, `distributed_backend`.

## Safe validation sequence

1. Inspect the command: `python ../../scripts/inspect_cli.py llm --help` from this sub-skill directory's parent skill root, or use the absolute root helper path.
2. Validate a YAML file without launching: `python skills/disco/autotrain-advanced/scripts/validate_config.py path/to/llm.yml`.
3. Validate local CSV/JSONL columns with `../text-and-tabular/scripts/validate_text_data.py --task llm ...` when the dataset is local.
4. Route adapter merging to `../model-tools/` rather than duplicating that logic here.

## References

- `references/workflows.md` — command templates, config aliases, trainers, and field groups.
- `references/troubleshooting.md` — auth, backend, PEFT, quantization, and unsupported deploy/inference recovery.

Files in this skill

  • SKILL.md2.6 KB
  • references/troubleshooting.md2.4 KB
  • references/workflows.md2.9 KB

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…