# seq lenght after Pyronoise

**URL:** https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329
**Category:** Theory behind mothur
**Created:** [May 1, 2013, 2:58pm UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329 "2013-05-01T14:58:17Z")
**Posts on this page:** 6
**Page:** 1

<div class="post-metadata">

### Author: ![lopesas](https://avatars.discourse-cdn.com/v4/letter/l/f07891/32.png) [@lopesas](https://forum.mothur.org/u/lopesas)
#### Post date: [May 1, 2013, 2:58pm UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329/1 "2013-05-01T14:58:17Z")

</div>

Hello Patrick,  
my sequences coming out of shhh.flows are mostly longer than (1-5 bases) the input ones except a fraction that got filtered out. I was wondering why this happened and if this is an artifact.  
Thanks a lot,  
Adriana.

---

<div class="post-metadata">

### Author: ![pschloss](https://yyz2.discourse-cdn.com/flex036/user_avatar/forum.mothur.org/pschloss/32/4_2.png) [@pschloss](https://forum.mothur.org/u/pschloss)
#### Post date: [May 1, 2013, 3:57pm UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329/2 "2013-05-01T15:57:03Z")

</div>

Sorry but I’m not sure what you’re asking… Also, nothing gets “filtered” out - you see a reduction in the number of sequences because the method effectively dereplicates the flowgrams and sequences. The redundant sequence names are in the shhh.names file.

---

<div class="post-metadata">

### Author: ![lopesas](https://avatars.discourse-cdn.com/v4/letter/l/f07891/32.png) [@lopesas](https://forum.mothur.org/u/lopesas)
#### Post date: [May 2, 2013, 7:30am UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329/3 "2013-05-02T07:30:02Z")

</div>

Sorry for not being clear. When I look the shhh.fasta the sequences have in general 1 - 5 nucleotides more.  
Example:  
fusion35. - 2839seq - max lenght - 284/ min lenght - 221  
fusion35.shhh - 2395seq - max lenght 286/ min lenght - 223  
I did not understand why…  
Thanks a lot for your help,  
Adriana.

---

<div class="post-metadata">

### Author: ![pschloss](https://yyz2.discourse-cdn.com/flex036/user_avatar/forum.mothur.org/pschloss/32/4_2.png) [@pschloss](https://forum.mothur.org/u/pschloss)
#### Post date: [May 2, 2013, 12:21pm UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329/4 "2013-05-02T12:21:45Z")

</div>

Not sure but it could be because the flows (with the flow order A) came in clusters of four bases at a time. The default trimming that’s in the fasta file produced by sffinfo is based on the quality score (not sure how that’s done) and the fasta file produced by shhh.flows is based on complete cycles of four bases at a time.

---

<div class="post-metadata">

### Author: ![lopesas](https://avatars.discourse-cdn.com/v4/letter/l/f07891/32.png) [@lopesas](https://forum.mothur.org/u/lopesas)
#### Post date: [May 3, 2013, 1:04pm UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329/5 "2013-05-03T13:04:42Z")

</div>

Hello Pat,

thanks a lot …i am running again the Pyronoise using the files generated after the trim.flows (i trimmed before using the RDP pipeline) just to see if i get the same result. However i got this message… ☹

> > > > > Processing fusion35\_good.scrap.flow (file 1 of 1) \<\<\<\<\<  
> > > > > Reading flowgrams…  
> > > > > [ERROR]: St9bad\_alloc has occurred in the ShhherCommand class function getFlowDa ta. Please contact Pat Schloss at , and be sure to include the mothur.logFile with your inquiry.  
> > > > > [alopesdossantos@n0 bacteria]$

Adriana.

---

<div class="post-metadata">

### Author: ![pschloss](https://yyz2.discourse-cdn.com/flex036/user_avatar/forum.mothur.org/pschloss/32/4_2.png) [@pschloss](https://forum.mothur.org/u/pschloss)
#### Post date: [May 3, 2013, 5:53pm UTC](https://forum.mothur.org/t/seq-lenght-after-pyronoise/1329/6 "2013-05-03T17:53:28Z")

</div>

So you probably don’t want to be processing fusion35\_good.scrap.flow. Those are the bad sequences. You probably want to use the files option as described here:

[http://www.mothur.org/wiki/454\_SOP](http://www.mothur.org/wiki/454_SOP)

Unless they’ve updated it in the last year, the RDP pipeline does very little to improve the quality of the output data. So you should expect some differences.
