Step 2 · dimensional structure

Embedding

Select highly variable genes, inspect PCA structure, then compute UMAP clustering for downstream DEA, visualization and labeling.

Analysis pending PCA not run UMAP not run t-SNE optional
Phase 1 · HVG and PCA

PCA parameters

Choose a high-variable gene set and compute principal components before neighborhood graph construction.

Not run

Usually 1,000-5,000; 2,000 is a stable starting point for many datasets.

Use Seurat by default; Cell Ranger can be useful for 10x-style inputs.