cuffdiff 和 edgeR 对差异表达基因的描述

ASE又走到了关键的一步要生成能决定是否有差异表达的table.

准备借鉴一下cuffdiff和edgeR 的结果

cuffdiff对差异表达基因的描述：

一共十四列：

第一列， test_id

a unique identifer describing the transcript, gene, primary transcript, or CDS being tested.

eg XLOC_000003

第二列，gene_id

eg XLOC_000003

第三列， gene

第四列， locus

genomic coordinates for easy browsing to the genes or transcripts being tested.

eg contig_23646:3511-3922

第五列， sample1

label (or number if no labels provided) of the first sample being tested

eg Sample_E

第六列， sample2

label (or number if no labels provided) of the second sample being tested

eg Sample_FHM

第七列， status

can be one of OK(test successful), NOTEST(not enough alignments for testing), LOWDATA（too many fragments in locus）, or FAIL, when an ill-conditioned covariance matrix or other numerical exception prevents testing

eg OK

第八列 value_1

FPKM of the gene in sample 1

eg 339.567

第九列 value_2

FPKM of the gene in sample 2

eg 465.939

第十列 log2(fold change)

the (base 2 ) log of the fold change 1/2

eg 0.456447

第十一列 test stat

the value of the test statistic used to compute significance of the observed change in FPKM

不懂什么意思估计要去翻统计书的节奏了

eg 0.361712

第十二列 p_value

the uncorrected p-value of the test statistic

eg 0.4849

第十三列 q_value

the FDR-adjusted p-value of the test statistic

eg 0.756741

第十四列 significant

can be either 'yes' or 'no' , depending on whether p is greater than the FDR after Benjamini-Hochberg correction for multiple-testing

eg no

The FPKM value represents the concentration of a transcript in your samples, normalized for observed read counts and gene length. Thus fields 7,8 represent measurements for your samples and field 9 is simply a ratio of the two. You might look up FPKM or RPKM values if you're unsure what they represent. Fields 11 and 12 are p-value and q-value. These are values associated with the measured variation or uncertainty when you make repeated measurements of something. You should look up what a p-value and an "adjusted p-value" are (the adjusted one is important for you to understand if you're going to do any genomic data analysis). The 13th field is simply a flag based on whether the value in field 11 or 12 is less than 0.05 (I forget which one, but you could figure it out by exploring your data).

edge R 结果对差异表达基因的描述：

Differential expression analysis of RNA-seq and digital gene expression profiles with biological replication. Uses empirical Bayes estimation and exact tests based on the negative binomial distribution. Also useful for differential signal analysis with other types of genome-scale count data.（貌似两者采用的分布模型是不一样的哦~~）

by freemao

FAFU

free_mao@qq.com

cuffdiff 和 edgeR 对差异表达基因的描述的更多相关文章

RNA-seq差异表达基因分析之TopHat篇
RNA-seq差异表达基因分析之TopHat篇发表于2012 年 10 月 23 日 TopHat是基于Bowtie的将RNA-Seq数据mapping到参考基因组上,从而鉴定可变剪切(exon-e ...
使用GEO数据库来筛选差异表达基因，KOBAS进行KEGG注释分析
前言本文主要演示GEO数据库的一些工具,使用的数据是2015年在Nature Communications上发表的文章Regulation of autophagy and the ubiquiti ...
使用Trinity拼接以及分析差异表达一个小例子
使用Trinity拼接以及分析差异表达一个小例子 2017-06-12 09:42:47 293 0 0 Trinity 将测序数据分为许多独立的de Brujin grap ...
使用limma、Glimma和edgeR，RNA-seq数据分析易如反掌
使用limma.Glimma和edgeR,RNA-seq数据分析易如反掌 Charity Law1, Monther Alhamdoosh2, Shian Su3, Xueyi Dong3, Luyi ...
Differential expression analysis for paired RNA-seq data 成对RNA-seq数据的差异表达分析
Differential expression analysis for paired RNA-seq data 抽象背景:RNA-Seq技术通过产生序列读数并在不同生物条件下计数其频率来测量转录本丰 ...
RNA-Seq differential expression analysis: An extended review and a software tool RNA-Seq差异表达分析：扩展评论和软件工具
RNA-Seq differential expression analysis: An extended review and a software tool RNA-Seq差异表达分析: 扩展 ...
差异基因分析：fold change(差异倍数), P-value(差异的显著性)
在做基因表达分析时必然会要做差异分析(DE) DE的方法主要有两种: Fold change t-test fold change的意思是样本质检表达量的差异倍数,log2 fold change的意 ...
edgeR使用学习【转载】
转自:http://yangl.net/2016/09/27/edger_usage/ 1.Quick start 2. 利用edgeR分析RNA-seq鉴别差异表达基因: #加载软件包 librar ...
Sensitivity, specificity, and reproducibility of RNA-Seq differential expression calls RNA-Seq差异表达调用的灵敏度特异性重复性
Sensitivity, specificity, and reproducibility of RNA-Seq differential expression calls RNA-Seq差异表达调用 ...

随机推荐

[Js]表格排序
思路:遍历每个li,并把它们存放到数组中去,然后通过sort()方法进行排序,再插入 <body> <input type="button" value=& ...
超简单的NDK单步调试方法
令人兴奋的是,ADTr20已经支持JNI单步调试,再也不需要如上这么麻烦的步骤了你现在需要做的只需以下2步: 1.使用ndk-build编译时,加上如下参数NDK_DEBUG=1,之后生成so文件之 ...
HDU 4405 Aeroplane chess 概率DP 难度:0
http://acm.hdu.edu.cn/showproblem.php?pid=4405 明显,有飞机的时候不需要考虑骰子,一定是乘飞机更优设E[i]为分数为i时还需要走的步数期望,j为某个可能 ...
NOIP2005 篝火晚会解题报告
佳佳刚进高中,在军训的时候,由于佳佳吃苦耐劳,很快得到了教官的赏识,成为了“小教官”.在军训结束的那天晚上,佳佳被命令组织同学们进行篝火晚会.一共有n个同学,编号从1到n.一开始,同学们按照1,2,… ...
wp8.1 Study9：针对不同的屏幕和手机方向调整UI
一.预备知识现在不同屏幕大小WP8.1手机越来越多,那么在设计UI时,这需要我们考虑这个问题.在WP中,比例因子(a scale factor)能很好的解决问题,而且在微软系统的PC/平板/手机都是 ...
C#根据当前日期获取星期和阴历日期
private string GetWeek(int dayOfWeek) { string returnWeek = ""; switch (dayOfWeek) { case ...
Android 系统基础
当系统启动一个组件,它其实就启动了这个程序的进程(如果这个进程还未被启动的话)并实例化这个组件所需要的类. 例如,如果你的程序启动了相机程序里的activity去拍照,这个activity实际上是运行 ...
黑马程序员——C语言基础语法关键字标识符注释数据及数据类型
Java培训.Android培训.iOS培训..Net培训.期待与您交流! (一下内容是对黑马苹果入学视频的个人知识点总结) (一)C语言简单介绍 (1)C语言程序是由函数组成的任何C语言程序都是由一 ...
<select>标签使用方法
前台页面: <form id="form1" runat="server"> <select runat="server" ...
win8系统 host文件无法修改解决之道
host文件,路径为:C:\windows\system32\drivers\etc\hosts 方法/步骤: 方法1:用notepad++打开host文件,修改和保存方法2:(1)首先用管理管权限 ...

cuffdiff 和 edgeR 对差异表达基因的描述

cuffdiff 和 edgeR 对差异表达基因的描述的更多相关文章

随机推荐

热门专题