GFF2GTF.py2

时间：2021-04-23 00:10:52 阅读：33 评论：0 收藏：0 [点我收藏+]

import sys

inFile = open(sys.argv[1],‘r‘)

for line in inFile:
  #skip comment lines that start with the ‘#‘ character
  if line[0] != ‘#‘:
    #split line into columns by tab
    data = line.strip().split(‘\t‘)

    ID = ‘‘

    #if the feature is a gene 
    if data[2] == "gene":
      #get the id
      ID = data[-1].split(‘ID=‘)[-1].split(‘;‘)[0]

    #if the feature is anything else
    else:
      # get the parent as the ID
      ID = data[-1].split(‘Parent=‘)[-1].split(‘;‘)[0]
    
    #modify the last column
    data[-1] = ‘gene_id "‘ + ID + ‘"; transcript_id "‘ + ID

    #print out this new GTF line
    print ‘\t‘.join(data)

https://www.jianshu.com/p/c284a6b4e1c6

GFF2GTF.py2

原文：https://www.cnblogs.com/3Dgenome/p/14690619.html

踩

(0)

评论一句话评论（0）

分享档案

更多>

2021年09月23日 (328)
2021年09月24日 (313)
2021年09月17日 (191)
2021年09月15日 (369)
2021年09月16日 (411)
2021年09月13日 (439)
2021年09月11日 (398)
2021年09月12日 (393)
2021年09月10日 (160)
2021年09月08日 (222)