We have decided to continue providing and supporting DepthOfCoverage and DiagnoseTargets for the foreseeable future. Going forward, we'll try to integrate them and develop their features to address the main needs of the community. To this end we welcome your continuing feedback, so please feel free to contribute comments and ideas in this thread.

To all who took the time to tell us what you find useful about DoC and DT (and what you wish it could do), a big thank you! This is always very useful to us because it helps us identify which features are most valuable to our users.



mengyuankan on 12 Mar 2013


Thanks Geraldine. A newbie question: what's the difference between DepthOfCoverage and DiagnoseTargets, and which one do you recommend to use (or both)?

Geraldine_VdAuwera on 12 Mar 2013


The two are somewhat convergent, but basically DepthOfCoverage allows you to evaluate the depth of coverage in your data at certain sites or over intervals generally, whereas DiagnoseTargets is more specifically designed to evaluate the quality of data covering exome target intervals. The best thing to do is to read their respective documentations pages (links below) and see which sounds like it will suit you needs best. http://www.broadinstitute.org/gatk/gatkdocs/org_broadinstitute_sting_gatk_walkers_coverage_DepthOfCoverage.html http://www.broadinstitute.org/gatk/gatkdocs/org_broadinstitute_sting_gatk_walkers_diagnostics_targets_DiagnoseTargets.html

kevyin on 12 Mar 2013


Hi @Geraldine_VdAuwera and thanks for support on this. The DepthOfCoverage documentation mentions -nt as a parallelism option but I get an error. `##### ERROR MESSAGE: Invalid command line: Argument nt has a bad value: The analysis DepthOfCoverage aggregates results by interval. Due to a current limitation of the GATK, analyses of this type do not currently support parallel execution. Please run your analysis without the -nt option.` My command line: `java -jar GenomeAnalysisTK.jar -T DepthOfCoverage -I ExampleBAM.bam -R exampleFASTA.fasta -o ./test_coverage_out -nt 4` Is this an error in the documentation? Or am I missing something? Thanks.

Geraldine_VdAuwera on 12 Mar 2013


Hi @kevyin, DOC is currently set up to aggregate statistics over intervals by default, which is incompatible with the `-nt` mode. You can disable this behavior by using the `--omitIntervalStatistics` flag, which should make `-nt` work. Let me know if you have any issues with that. We may change the default behavior in future; in any case we will add a note about this to the documentation. Thanks for reporting this!

SerenaRhie on 12 Mar 2013


Hello, I want the concise depth of coverage summary as well as the per-location covering number of bases for my RNA-Seq data. With not splitting the N containing reads, what does DoC exactly do with -U ALLOW_N_CIGAR_READS option? Is it still counting the bases in the CIGAR array represented as matched? I think the bases represented as N in CIGAR should not be counted, just as the deletions (D) are treated. (Maybe an option such as --includeNBases might be appropriate for special cases.) A good explanation will be helpful to decide the DoC or not. And please do something with the last line of the ERROR MESSAGE "Notice however that if you were to choose the latter, an unspecified subset of the analytical outputs of an unspecified subset of the tools will become unpredictable.", which I think is totally nonsense! :smile:

Geraldine_VdAuwera on 12 Mar 2013


@SerenaRhie, Hah, I think the developer who wrote that error message was feeling particularly sarcastic that day. I'll see what we can do to make that a little more informative. Regarding the behavior question, I believe the Ns in the CIGAR are correctly interpreted as not contributing any coverage. That said, you could run DoC after running the SplitNTrim step, which will give you effective depth after processing.

SerenaRhie on 12 Mar 2013


@Geraldine_VdAuwera Thanks!

MUHAMMADSOHAILRAZA on 12 Mar 2013


Hi @Geraldine_VdAuwera i try to use -nt option while running DoC tool. I added -omitIntervalStatistics. But it still prompts an error. My command was: java -Xmx15g -jar /softwares/GTK/GenomeAnalysisTK.jar -T DepthOfCoverage \ -log DepthofCov/depth.log -nt 15 -omitIntervalStatistics \ -R /Ref/human_g1k_v37.fasta \ -I /DepthofCov/CHG000691_2_3bam.list \ -geneList /Ref/humanRefseqROD_Final.refseq \ -o /DepthofCov/HM_trios_Samples Error Message: ##### ERROR ------------------------------------------------------------------------------------------ ##### ERROR A USER ERROR has occurred (version 3.3-0-g37228af): ##### ERROR ##### ERROR This means that one or more arguments or inputs in your command are incorrect. ##### ERROR The error message below tells you what is the problem. ##### ERROR ##### ERROR If the problem is an invalid argument, please check the online documentation guide ##### ERROR (or rerun your command with --help) to view allowable command-line arguments for this tool. ##### ERROR ##### ERROR Visit our website and forum for extensive documentation and answers to ##### ERROR commonly asked questions http://www.broadinstitute.org/gatk ##### ERROR ##### ERROR Please do NOT post this error to the GATK forum unless you have really tried to fix it yourself. ##### ERROR ##### ERROR MESSAGE: Argument with name 'omitIntervalStatistics' isn't defined. ##### ERROR ------------------------------------------------------------------------------------------ Can you help me how to resolve this issue? Thanks..

Geraldine_VdAuwera on 12 Mar 2013


Try using two dashes instead of one (--).




- Recent posts


- Upcoming events

See Events calendar for full list and dates


- Recent events

See Events calendar for full list and dates



- Follow us on Twitter

GATK Dev Team

@gatk_dev

@hdeus To be clear we're not yet using convolutional neural nets (CNNs) in copy number (CNV) analysis. We're using… https://t.co/MYPQVjvzyC
18 May 18
Small correction on this #BioIT18 agenda item: Mark will actually be introducing Lee Lichtenstein from our team to… https://t.co/0UxLmX1uN0
16 May 18
RT @broadinstitute: Geraldine Van der Auwera of @gatk_dev will be presenting on the team’s Best Practices Pipeline tomorrow (5/16) at 12:40…
16 May 18
At Bio-IT World? Check out our abundant lineup of GATK4 talks and demos at #BioIT18 https://t.co/ahkUjny6Cw
15 May 18
Hey that pipeline looks familiar :) https://t.co/68lupDUkxN
15 May 18

- Our favorite tweets from others

Bioinformatics in a nutshell. 😑 #genetics #research #phd #gatk https://t.co/EjqaeFf4YZ
18 May 18
Want to hear the latest on WDL, Cromwell, FireCloud, and GATK #BioIT18 ? See this blog for tomorrow's schedule of t… https://t.co/S7kf58tECu
16 May 18
We caught up with @broadinstitute's Anthony Philippakis and Illumina's Susan Tousi today for perspective on today's… https://t.co/PNaVXbNB0r
15 May 18
#BioIT18 folks - come to booth #410 on 5/15 at 5:00 to learn about our $5 genome analysis pipeline (5 is clearly th… https://t.co/GH0zwrLlea
14 May 18
Excited to be registered for #GCCBOSC! Looking forward to the GATK4.0 and snakemake workshops. https://t.co/fMfcH38OhO
11 May 18

See more of our favorite tweets...