Skip to content
Back to skills

266 Usage Guide 0d4cfc27

ASecurity

Unique Molecular Identifiers (UMIs) are random sequences added to molecules before PCR amplification. They enable distinguishing PCR duplicates from biological duplicates, crucial for accurate quantification in RNA-seq, targeted sequencing, and single-cell applications.

  • 4 stars
  • 0 votes
  • 0 copies
  • 1 view
  • Added May 31, 2026
databashdocumentation

Security analysis

A100/100

Pro scans all 2 files and shows the line behind each finding

Scanned May 31, 2026

npx -y skills add tools-only/X-Skills --skill 266-usage-guide_0d4cfc27 --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of 266 Usage Guide 0d4cfc27?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for 266 Usage Guide 0d4cfc27
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/tools-only-266-usage-guide-0d4cfc27/badge)](https://www.skillsdirectory.com/skills/tools-only-266-usage-guide-0d4cfc27)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
# UMI Processing - Usage Guide

## Overview
Unique Molecular Identifiers (UMIs) are random sequences added to molecules before PCR amplification. They enable distinguishing PCR duplicates from biological duplicates, crucial for accurate quantification in RNA-seq, targeted sequencing, and single-cell applications.

## Prerequisites
```bash
conda install -c bioconda umi_tools
```

## Quick Start
Tell your AI agent what you want to do:
- "Extract UMIs from my reads and deduplicate after alignment"
- "Process my UMI-tagged library for accurate quantification"
- "Remove PCR duplicates using UMI information"

## Example Prompts

### Standard Workflow
> "Extract 8bp UMIs from read 1 and move them to the read header"

> "Deduplicate my aligned BAM file using UMI information"

### Library-Specific
> "Process my 10X Genomics library with cell barcodes and UMIs"

> "Handle my NEBNext library with 8bp UMIs"

### Deduplication Options
> "Use directional deduplication method for my RNA-seq data"

> "Deduplicate with cluster method for high-error-rate UMIs"

## What the Agent Will Do
1. Extract UMIs from reads and add to read headers
2. Pass through alignment with UMI information preserved
3. Deduplicate aligned reads based on UMI + alignment position
4. Generate deduplication statistics
5. Output deduplicated BAM for downstream analysis

## UMI Pattern Syntax

| Pattern | Description |
|---------|-------------|
| `N` | UMI base |
| `C` | Cell barcode base |
| `X` | Base to discard |
| `NNNNNNNN` | 8bp UMI |
| `CCCCCCCCNNNNNNNN` | Cell barcode + UMI |

## Common Library Types

| Library | Pattern |
|---------|---------|
| NEBNext | `NNNNNNNN` (8bp in R1) |
| 10X 3' v3 | `CCCCCCCCCCCCCCCCNNNNNNNNNNNN` (16bp CB + 12bp UMI) |
| TruSeq UMI | `NNNNNNNNN` (9bp in index) |

## Tips
- Extract UMIs before alignment; they must be in read headers for deduplication
- Directional deduplication is recommended for most RNA-seq applications
- High deduplication rates (>70%) may indicate library over-amplification
- Low deduplication rates (<10%) may indicate under-sequencing
- Check UMI diversity in the deduplication stats to assess library complexity

## Resources
- [umi_tools Documentation](https://umi-tools.readthedocs.io/)
- [umi_tools Publication](https://doi.org/10.1101/gr.209601.116)

Files in this skill

  • README.md800 B
  • skill.md2.2 KB

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…