Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
GSEA-InContext: identifying novel and common pa...
Search
Y-h. Taguchi
PRO
August 13, 2018
Science
310
1
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
GSEA-InContext: identifying novel and common patterns in expression experiments
ISMB2018読み会
https://atnd.org/events/98383
でのプレゼンです。
Y-h. Taguchi
PRO
August 13, 2018
More Decks by Y-h. Taguchi
See All by Y-h. Taguchi
知能とはなにか -ヒトとAIのあいだ-
tagtag
PRO
0
180
サンプル対応のない複数遺伝子発現プロファイルに対するテンソル分解型統合解析の要約
tagtag
PRO
0
250
AI(人工知能)の過去・現在・未来 ~AIは人類を越えるのか~
tagtag
PRO
0
160
presen_司法書士学員会.pdf
tagtag
PRO
1
100
生成AIと司法書士の未来.pdf
tagtag
PRO
0
170
データ駆動型ゲノム解析で迫る睡眠研究
tagtag
PRO
0
100
適応テンソル分解と主成分分析に基づく教師なし特徴抽出は、従来手法よりも生物学的に妥当な発現量差のある遺伝子を選択する
tagtag
PRO
0
70
知能とはなにか -ヒトとAIのあいだ-
tagtag
PRO
0
120
Genomic Differentiation of Sleep and Anesthesia: The Role of RHO GTPase and Cortical Neurons
tagtag
PRO
0
73
Other Decks in Science
See All in Science
Leitner Inauguration Lecture Chalmers University of Technology
xleitix
0
350
AI for Phage-Host prediction
michielstock
0
120
Bリーグのショットデータを活用した得点期待値モデルの構築 / Construction of expected points model using shot data of B.LEAGUE
konakalab
0
210
AlgorithAlgorihms for Decision Making
mickey_kubo
0
140
データベース04: SQL (1/3) 単純質問 & 集約演算
trycycle
PRO
0
1.7k
O(log n)-Approximation Algorithms for Bipartiteness Ratio
tasusu
0
190
2026 Introduction to University Math 01
kanaya
0
140
機械学習 - 授業概要
trycycle
PRO
0
620
Utiliser Bitcoin sans Internet
rlifchitz
0
370
因果探索の発展と展望
sshimizu2006
2
1.1k
CVPR2026_VGGTとその仲間たち
mickey_0226
0
1.1k
因果推論と機械学習
sshimizu2006
1
1.4k
Featured
See All Featured
Building Flexible Design Systems
yeseniaperezcruz
330
41k
Designing for humans not robots
tammielis
254
26k
First, design no harm
axbom
PRO
2
1.3k
Color Theory Basics | Prateek | Gurzu
gurzu
0
460
The Anti-SEO Checklist Checklist. Pubcon Cyber Week
ryanjones
0
230
Visualizing Your Data: Incorporating Mongo into Loggly Infrastructure
mongodb
49
10k
The Spectacular Lies of Maps
axbom
PRO
1
990
How Software Deployment tools have changed in the past 20 years
geshan
1
34k
Fight the Zombie Pattern Library - RWD Summit 2016
marcelosomers
234
17k
Effective software design: The role of men in debugging patriarchy in IT @ Voxxed Days AMS
baasie
0
520
Collaborative Software Design: How to facilitate domain modelling decisions
baasie
1
320
A designer walks into a library…
pauljervisheath
211
25k
Transcript
ISMB2018読み会 GSEA-InContext: identifying novel and common patterns in expression experiments
Rani K. Powers, Andrew Goodspeed, Harrison Pielke-Lombardo, Aik-Choon Tan and James C. Costello Bioinformatics, 34, 2018, i555–i564 doi: 10.1093/bioinformatics/bty271 報告者: 中央大学理工学部物理学科 田口善弘
論文の目的: 論文の目的: GSEA(Gene Set Enrichment Analysis)は「遺伝子を『何か(例: 発現差の大きさ)の順番』で並べた場合、順番には意味があ る。ある遺伝子セットAが有意に上位に並ぶなら、そのセットに は『何か』の大きさが有意に大きい遺伝子のセットであるとい えるだろう」
というものですが、その場合「有意に上位に並ぶ」の判定をす るときの比較対象(=帰無仮説)が「完全にランダムな並び」 になっている。しかし、遺伝子はお互いに相関しているんだか ら、『何か』と全く無関係じゃない限り、遺伝子セットAはどっち にしろグループで動く(上位に来る)だろう。そうなると「遺伝 子セットAに意味があるか?」という検証にはなっても「順位 付けした『何か』と関係している」と言えなくないか? この問題は解決するには比較対象を完全にランダムな並び じゃなく、いろいろな実験での並びの集合に置き換えないとい けないのでは?
比較対象 比較対象 ・GEOから集めた ・Afymetrix Human Genome U133 Plus 2.0 Array限定
・small molecule test限定 ・遺伝子の順位リストを442個作成 GSEAPreranked: GSEAPreranked:入力がm個の遺伝子場合、m個の遺伝子をラ ンダムに選んで比較、入力が有意に上位のあるかを比較 GSEA-InContext GSEA-InContext: :442個の順位リストをつかい、これらのリスト で上位にあるという重み付けをしてm個の遺伝子を選んで比較、 入力が有意に上位にあるかを比較
B(α、β):β関数 β二項分布: バイアスのあるコインがたくさん入った袋がある。そこから一枚コイ ンを一枚抜き出して、n 回投げた。表の出る回数 k が従う分布は? ただし袋の中のコインの表の出る確率 p はベータ分布に従うこと
とする。 α、βの値は442個の遺伝子ランクをつかって、遺伝子ごとに決定 ある遺伝子がr位になる確率:β二項分布 ∫0 1 p(α−1)(1−p)α dp
単純ランダムより実験に基づくほうがランクの期待値の幅(分散) は大きい →順位が高いものは高く、低いものは低くなりやすい。
試験用遺伝子データセット MSigDB :The Hallmarks collection (50クラス) 442遺伝子順位セット(バックグランウンド) →薬剤の標的蛋白で予めグループ化 標的蛋白 臓 器
確 か に こ う す る と 遺 伝
子 セ ッ ト の 有 意 度 は 低 下 論 文 で は こ れ を ア | テ ィ フ ァ ク ト の 減 少 と 解 釈
GSEA-InContextだけで有意になるものもある (バックグラウンドをうまく選べば)
なんで なんでISMB ISMB2018に採択されたの? 2018に採択されたの? 正直、何が面白いのか皆目わかりません。 コレポンはgoogle scholarの引用数が4000ある(2007年から論文 を書き始めた)。実験系の論文が多く、ファーストやコレポンは少な い。 多分、最初の論文から10年でgoogle
scholarの引用数を4000くら いにするのがISMBに論文通すコツなのでは? (僕にはもう実現できないハードルですが、過去のことなので)