Docker image for Clair3 in Fred Hutch OCDO's WILDS
915
This directory contains Docker images for Clair3, a germline small variant caller optimized for long-read sequencing data (Oxford Nanopore and PacBio HiFi) that identifies SNPs and indels using a pileup-plus-full-alignment approach with PyTorch-based deep learning models.
latest ( Dockerfile | Vulnerability Report )2.0.0 ( Dockerfile | Vulnerability Report )These Docker images are built from nvidia/cuda:12.6.3-runtime-ubuntu24.04 with Miniforge installed on top, and include:
--use_gpu to enable)The images support both CPU and GPU execution — GPU mode is opt-in via the --use_gpu flag, so the image works in CPU-only environments without any changes. All 14 pre-trained sequencing models are bundled in the image at /opt/models/.
If you use Clair3 in your research, please cite the original authors:
Zheng, Z., Li, S., Su, J., Leung, A.W.S., Lam, T.W., & Luo, R. (2022).
Symphonizing pileup and full-alignment for deep learning-based long-read variant calling.
Nature Computational Science, 2, 797–803.
https://doi.org/10.1038/s43588-022-00387-x
Tool homepage: https://github.com/HKU-BAL/Clair3
# Pull the latest version
docker pull getwilds/clair3:latest
# Or pull a specific version
docker pull getwilds/clair3:2.0.0
# Alternatively, pull from GitHub Container Registry
docker pull ghcr.io/getwilds/clair3:latest
# Pull the latest version
apptainer pull docker://getwilds/clair3:latest
# Or pull a specific version
apptainer pull docker://getwilds/clair3:2.0.0
# Alternatively, pull from GitHub Container Registry
apptainer pull docker://ghcr.io/getwilds/clair3:latest
All 14 pre-trained models are bundled in the image at /opt/models/. Pass the appropriate model path for your sequencing platform via --model_path.
# Run Clair3 on ONT R10.4.1 data (CPU mode)
docker run --rm \
-v /path/to/data:/data \
getwilds/clair3:latest \
run_clair3.sh \
--bam_fn=/data/sample.bam \
--ref_fn=/data/reference.fa \
--threads=4 \
--platform=ont \
--model_path=/opt/models/r1041_e82_400bps_sup_v500 \
--output=/data/clair3_output
# Run Clair3 on ONT R10.4.1 data with GPU acceleration (requires NVIDIA Container Toolkit)
docker run --rm --gpus all \
-v /path/to/data:/data \
getwilds/clair3:latest \
run_clair3.sh \
--bam_fn=/data/sample.bam \
--ref_fn=/data/reference.fa \
--threads=4 \
--platform=ont \
--model_path=/opt/models/r1041_e82_400bps_sup_v500 \
--output=/data/clair3_output \
--use_gpu
# Run Clair3 on PacBio HiFi data (CPU mode)
docker run --rm \
-v /path/to/data:/data \
getwilds/clair3:latest \
run_clair3.sh \
--bam_fn=/data/sample.hifi.bam \
--ref_fn=/data/reference.fa \
--threads=8 \
--platform=hifi \
--model_path=/opt/models/hifi_revio \
--output=/data/clair3_hifi_output
# Run using Apptainer with GPU
apptainer run --nv \
--bind /path/to/data:/data \
docker://getwilds/clair3:latest \
run_clair3.sh \
--bam_fn=/data/sample.bam \
--ref_fn=/data/reference.fa \
--threads=4 \
--platform=ont \
--model_path=/opt/models/r1041_e82_400bps_sup_v500 \
--output=/data/clair3_output \
--use_gpu
GPU mode is opt-in via --use_gpu and requires the NVIDIA Container Toolkit on the host. Pass --gpus all to Docker or --nv to Apptainer. Without these flags the image runs in CPU mode regardless of available hardware. The image uses CUDA 12.6, which is compatible with NVIDIA driver ≥525.60.13.
These images are built for linux/amd64 only.
The Dockerfile follows these main steps:
nvidia/cuda:12.6.3-runtime-ubuntu24.04 as the base image/opt/models/run_clair3.sh --version and a PyTorch import check as smoke testsThese images are regularly scanned for vulnerabilities using Docker Scout. However, due to the nature of bioinformatics software and their dependencies, some Docker images may contain components with known vulnerabilities (CVEs).
Use at your own risk: While we strive to minimize security issues, these images are primarily designed for research and analytical workflows in controlled environments.
For the latest security information about this image, please check the CVEs_*.md files in this directory, which are automatically updated through our GitHub Actions workflow. If a particular vulnerability is of concern, please file an issue in the GitHub repo citing which CVE you would like to be addressed.
These Dockerfiles are maintained in the WILDS Docker Library repository.
Content type
Image
Digest
sha256:246a5a91f…
Size
5.8 GB
Last updated
6 months ago
docker pull getwilds/clair3