計量生物学
Online ISSN : 2185-6494
Print ISSN : 0918-4430
ISSN-L : 0918-4430
総説
ゲノム・プロテオミクスデータを用いた予測解析:機械学習による新しい統計的手法
小森 理江口 真透
著者情報
ジャーナル フリー

2011 年 32 巻 1 号 p. 49-73

詳細
抄録

At the present day, it becomes imperative to develop appropriate statistical methods for high-dimensional and small sample data analysis because data formats in the biological or medical fields have been dramatically changed. Especially, it will be common in the near future to analyze clinical data together with genomic data. In this review paper, we introduce several current approaches to the analysis relating to genomic and proteomic data, and describe some limitations or problems in the statistical performance.
In the former part of this paper, we explain a problem of p»n, which is the fundamental challenge in data analysis in bioinformatics. In particular, we consider a typical problem of p»n in prediction of treatment effects using microarray data as feature vectors. Then, we introduce some new boosting methods based on the area under the ROC curve. After showing some applications of the boosting methods, we summarize the present problems and refer to outlook for the future.

著者関連情報
© 2011 日本計量生物学会
前の記事
feedback
Top