The path to the taglign is gs://caper_in/CRC_finemap/tagalign/Crypt_Fibroblasts_4.tsv.gz
Pipeline version
v1.10.0
Pipeline type
atac
Genome
hg38
Aligner
bowtie2
Sequencing endedness
{'rep1': {'paired_end': True}}
Peak caller
macs2
Alignment quality metrics
Annotated genomic region enrichment
rep1
Fraction of Reads in universal DHS regions
0.5445440024553571
Fraction of Reads in blacklist regions
0.0008083839450109955
Fraction of Reads in promoter regions
0.22229981104782803
Fraction of Reads in enhancer regions
0.4079347028476817
Signal to noise can be assessed by considering whether reads are falling into
known open regions (such as DHS regions) or not. A high fraction of reads
should fall into the universal (across cell type) DHS set. A small fraction
should fall into the blacklist regions. A high set (though not all) should
fall into the promoter regions. A high set (though not all) should fall into
the enhancer regions. The promoter regions should not take up all reads, as
it is known that there is a bias for promoters in open chromatin assays.
Replication quality metrics
IDR (Irreproducible Discovery Rate) plots
rep1-pr1_vs_rep1-pr2
Reproducibility QC and peak detection statistics
overlap
idr
Nt
0
0
N1
134758
56951
Np
0
0
N optimal
134758
56951
N conservative
134758
56951
Optimal Set
rep1-pr1_vs_rep1-pr2
rep1-pr1_vs_rep1-pr2
Conservative Set
rep1-pr1_vs_rep1-pr2
rep1-pr1_vs_rep1-pr2
Rescue Ratio
0.0
0.0
Self Consistency Ratio
1.0
1.0
Reproducibility Test
pass
pass
Reproducibility QC
N1: Replicate 1 self-consistent peaks (comparing two pseudoreplicates generated by subsampling Rep1 reads)
N2: Replicate 2 self-consistent peaks (comparing two pseudoreplicates generated by subsampling Rep2 reads)
Ni: Replicate i self-consistent peaks (comparing two pseudoreplicates generated by subsampling RepX reads)
Nt: True Replicate consistent peaks (comparing true replicates Rep1 vs Rep2)
Np: Pooled-pseudoreplicate consistent peaks (comparing two pseudoreplicates generated by subsampling pooled reads from Rep1 and Rep2)
Self-consistency Ratio: max(N1,N2) / min (N1,N2)
Rescue Ratio: max(Np,Nt) / min (Np,Nt)
Reproducibility Test: If Self-consistency Ratio >2 AND Rescue Ratio > 2, then 'Fail' else 'Pass'
Number of raw peaks
rep1
Number of peaks
230553
Top 300000 raw peaks from macs2 with p-val threshold 0.01
Peak calling statistics
Peak region size
rep1
idr_opt
overlap_opt
Min size
150.0
150.0
150.0
25 percentile
199.0
486.0
323.0
50 percentile (median)
306.0
680.0
479.0
75 percentile
537.0
907.0
710.0
Max size
2238.0
2238.0
2238.0
Mean
409.6833005859824
720.9364892626995
551.2454993395569
rep1idr_optoverlap_opt
Peak enrichment
Fraction of reads in peaks (FRiP)
FRiP for macs2 raw peaks
rep1
rep1-pr1
rep1-pr2
Fraction of Reads in Peaks
0.4560967972611622
0.4505213338710789
0.4413612612413144
FRiP for overlap peaks
rep1-pr1_vs_rep1-pr2
Fraction of Reads in Peaks
0.39925525225291514
FRiP for IDR peaks
rep1-pr1_vs_rep1-pr2
Fraction of Reads in Peaks
0.282871584385769
For macs2 raw peaks:
repX: Peak from true replicate X
repX-prY: Peak from Yth pseudoreplicates from replicate X
pooled: Peak from pooled true replicates (pool of rep1, rep2, ...)
pooled-pr1: Peak from 1st pooled pseudo replicate (pool of rep1-pr1, rep2-pr1, ...)
pooled-pr2: Peak from 2nd pooled pseudo replicate (pool of rep1-pr2, rep2-pr2, ...)
For overlap/IDR peaks:
repX_vs_repY: Comparing two peaks from true replicates X and Y
repX-pr1_vs_repX-pr2: Comparing two peaks from both pseudoreplicates from replicate X
pooled-pr1_vs_pooled-pr2: Comparing two peaks from 1st and 2nd pooled pseudo replicates