Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
ICML2013読み会 "ELLA: An Efficient Lifelong Learni...
Search
Yuya Unno
July 09, 2013
Research
24
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
ICML2013読み会 "ELLA: An Efficient Lifelong Learning Algorithm"
Yuya Unno
July 09, 2013
More Decks by Yuya Unno
See All by Yuya Unno
深層学習で切り拓くパーソナルロボットの未来 @東京大学 先端技術セミナー 工学最前線
unnonouno
0
31
深層学習時代の自然言語処理ビジネス @DLLAB 言語・音声ナイト
unnonouno
0
56
ベンチャー企業で言葉を扱うロボットの研究開発をする @東京大学 電子情報学特論I
unnonouno
0
53
PFNにおけるセミナー活動 @NLP2018 言語処理研究者・技術者の育成と未来への連携WS
unnonouno
0
21
進化するChainer @JSAI2017
unnonouno
0
33
予測型戦略を知るための機械学習チュートリアル @BigData Conference 2017 Spring
unnonouno
0
32
深層学習フレームワーク Chainerとその進化
unnonouno
0
38
深層学習による機械とのコミュニケーション @DeNA TechCon 2017
unnonouno
0
45
最先端NLP勉強会 “Learning Language Games through Interaction” @第8回最先端NLP勉強会
unnonouno
0
29
Other Decks in Research
See All in Research
VRID: View-Invariant Representation through Dual-Axis Transformation for Cross-iew Pose Estimation
satai
3
110
ros2-perf-multihost: Automated Coordination Framework for Objective Architecture Evaluation in Distributed Systems
takasehideki
0
110
LA-Bench 2025:実験指示から実行可能手順を生成するためのデータセット/LA-Bench 2025: A Dataset for Generating Executable Experimental Procedures from Experimental Instructions
stktu
0
190
Spatial Active Noise Control Based onSound Field Interpolation Incorporating Physical Constraints
skoyamalab
0
190
多様なデータを許容し学習し続ける模倣学習 / Advanced Imitation Learning for VLA
prinlab
0
330
[最先端NLP勉強会2026] Agentic Rubrics as Contextual Verifiers for SWE Agents
rfujii
1
360
Google Cloud Next 2026 DM Recap Agentic Data Cloudを添えて / Google Cloud Next 2026 DM Recap
nnaka2992
0
140
Physical AIでモデリングはどう変わるか/How Physical AI Will Change Modeling
stktu
0
260
「AIとWhyを深堀る」をAIと深堀る
iflection
0
660
某助成金プロジェクト採択に向けて企業研究所のアウトリーチ専任者がやったこと
afroscript
0
190
JPA2026_NetworkTutorial_JunKashihara
junkashihara
0
180
J-STAGEの現況と全文XML登載必須化について
xspa2012
0
270
Featured
See All Featured
WENDY [Excerpt]
tessaabrams
14
39k
Git: the NoSQL Database
bkeepers
PRO
432
67k
More Than Pixels: Becoming A User Experience Designer
marktimemedia
3
530
How to optimise 3,500 product descriptions for ecommerce in one day using ChatGPT
katarinadahlin
PRO
3
3.8k
Bridging the Design Gap: How Collaborative Modelling removes blockers to flow between stakeholders and teams @FastFlow conf
baasie
0
700
HU Berlin: Industrial-Strength Natural Language Processing with spaCy and Prodigy
inesmontani
PRO
0
710
How to Think Like a Performance Engineer
csswizardry
28
2.8k
A Guide to Academic Writing Using Generative AI - A Workshop
ks91
PRO
1
480
Test your architecture with Archunit
thirion
2
2.4k
Speed Design
sergeychernyshev
33
2.1k
Dominate Local Search Results - an insider guide to GBP, reviews, and Local SEO
greggifford
PRO
0
340
The Psychology of Web Performance [Beyond Tellerrand 2023]
tammyeverts
49
3.6k
Transcript
ELLA: An Efficient Lifelong Learning Algorithm 株式会社Preferred Infrastructure 海野 裕也
(@unnonouno) 2013/07/09 ICML2013読み会@東大
⾃自⼰己紹介 l 海野 裕也 (@unnonouno) l プリファードインフラストラクチャー l 情報検索索、レコメンド l 機械学習・データ解析研究開発
l Jubatusチームリーダー l 分散オンライン機械学習フレームワーク l 専⾨門 l ⾃自然⾔言語処理理 l テキストマイニング 2
要旨 l Lifelong learningのためにGO-MTLの精度度をほとんど落落 とさずに、1000倍早くした l ⼿手法の要旨は以下の2点 l テーラー展開して元の最適化の式を簡略略化 l
再計算の必要な項の計算を簡略略化 3
Lifelong learning 4
Lifelong learning l タスクが次々やってくる l Z(1), …, Z(Tmax) l 学習者はタスクの数も順番も知らない
l 各Zは教師有りの問題(分類か回帰) l 各タスクにはn t 個の教師ありデータが与えられる マルチタスクで、タスクが次々やってくるイメージ 5
Lifelong learningのキモチ(ホントか?) l ずっと学習し続ける l データセットはオンラインでやってくる l 過去の学習結果をうまく活かしたい(似たような問題、 組み合わせの問題が多い) 例例えば将来的に、ずっと学習し続けるインフラのようなモ
ノができた時を想定している(のかも) 6
Grouping and Overlap in Multi-Task Learning (GO-MTL) [Kumar&Daume III ’12]
l L: 損失関数 l w = Ls: モデルパラメータ l L: k個の隠れタスクの重み l s: 各タスクをLの線形和で表現する役割 l sは疎にしたいのでL1正則化 7 収束の証明のために ちょっと変えてある
GO-MTLが遅い l GO-MTL⾃自体はマルチタスクのバッチ学習⼿手法なので データが次々やってくるLifelong learningに適⽤用しよう とすると遅い l 2重ループが明らかに遅そう 8
⼯工夫1: 損失関数の部分をテーラー展開 9 θ(t)の周りで2次の テーラー展開
⼯工夫2: 全てのtに対するs(t)の最適化を⾏行行うのは⾮非効 率率率 10 s(t)の最適化を 順次行う
実際の更更新式 l L = A-1b l 実際に計算するときは、Aとbは差分更更新できるような⼯工 夫が⼊入っている 11
実験結果 12 バッチとほとんど同じ精度度で1000倍以上速い!!
あれ、よく⾒見見ると・・・ 13 Single Task Leaning (STL) でもそこそこだし、 当然もっと速い・・・
まとめ l マルチタスクのバッチ学習であるGO-MTLをLifelong learningに適⽤用するために、⾮非効率率率な部分を効率率率化した l ほとんど精度度を下げずに、1000倍以上⾼高速化した l タスクを独⽴立立に解いてもそこそこの精度度が出ていて、実 験設定はもう少し考慮しても良良かったのかも 14