Hi, A very helpful feature to add would be the demultiplexing of the reads as an o

<a class="user-mention notranslate" data-hovercard-type="user" data-hovercard-url="/us

Demultiplexing could be done via cutadapt as documented <a href="https://cutadapt.read

Add demultiplexing step about ampliseq HOT 7 OPEN

nf-core commented on June 4, 2024

Add demultiplexing step

from ampliseq.

Comments (7)

d4straub commented on June 4, 2024 1

What about adding a few optional columns (such as fw_index, rv_index) to the sample sheet. If those columns are present, demultiplexing will run. If that might mess too much with existing routines, a separate input file (e.g. --demultiplex "sheet.tsv") that contains the necessary information (samplesheet & demultiplexsheet have identical IDs) might be an option?
While ampliseq does not require a samplesheet (folder input & fasta input are also allowed), for demultiplexing that would be fine. After all, a samplesheet can handle more info than a folder input. Not all input options need to support all functionality, imho.

from ampliseq.

d4straub commented on June 4, 2024

This might be a helpful feature. As far as I know there is work ongoing for wrapping DADA2 directely in this pipeline instead of QIIME2 using DADA2. Therefore I am unsure how to integrate this feature sustainably with the major changes that are planned to the early workflow. However, PRs are welcome.

from ampliseq.

d4straub commented on June 4, 2024

@DiegoBrambilla is planning to implement dada2 for PacBio analysis and could immediately add that demultiplexing step :)

from ampliseq.

DiegoBrambilla commented on June 4, 2024

We take it into consideration.
For the time being, implementing the R-DADA2 pipeline, taxonomy annotation from several sources and dealing with PacBio reads take priority.

from ampliseq.

d4straub commented on June 4, 2024

Demultiplexing could be done via cutadapt as documented here. I never come across the need for demultiplexing in the pipeline, but if anyone does, please mention it here and I might further look into it.

from ampliseq.

a4000 commented on June 4, 2024

I want to add demultiplexing (with Cutadapt) to Ampliseq. The way I've handled demultiplexing in my own nf-core style pipeline is to ask the user to specify the path to their raw data in the command line --raw_data "/path/to/data/*{R1,R2}*.fastq.gz". Then in the sample sheet the user has to add the columns fw_index, rv_index, fw_primer, and rv_primer (the two rv_ columns can be empty for single-end data). I use the _index columns for demultiplexing and the _primer columns for trimming after demultiplexing. The main issue I see is that Ampliseq doesn't require a sample sheet as input, so I'm wondering if anyone has a suggestion for a better way of adding this feature to Ampliseq? Maybe the sample sheet should be required if the user wants to demultiplex?

from ampliseq.

erikrikarddaniel commented on June 4, 2024

To me, adding columns to the sample sheet sounds best.

from ampliseq.

Recommend Projects

Add demultiplexing step about ampliseq HOT 7 OPEN

Comments (7)

Related Issues (20)

Recommend Projects

React

Vue.js

Typescript

TensorFlow

Django

Laravel

D3

Recommend Topics

javascript

web

server

Machine learning

Visualization

Game

Recommend Org

Facebook

Microsoft

Google

Alibaba

D3

Tencent