Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
[論文紹介] Probabilistic Matrix Factorization
Search
ysekky
January 27, 2015
Research
680
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
[論文紹介] Probabilistic Matrix Factorization
Gunosy研究会 #82 2015/01/27
ysekky
January 27, 2015
More Decks by ysekky
See All by ysekky
スタートアップの開発サイクルに学ぶ 研究活動の進め方 / research practices inspired by startup business strategy
ysekky
0
2.5k
[論文紹介] A Method to Anonymize Business Metrics to Publishing Implicit Feedback Datasets (Recsys2020) / recsys20-reading-gunosy-datapub
ysekky
3
2.9k
JSAI2020 OS-12 広告とAI オープニング / JSAI2020-OS-12-ads-and-ai-opening
ysekky
0
2.3k
JSAI2020インダストリアルセッション - Gunosyにおける研究開発 / jsai2020-gunosy-rd-examples
ysekky
1
840
ウェブサービス事業者における研究開発インターン[株式会社Gunosy] - テキストアナリティクスシンポジウム2019 / research-intern-case-study-at-gunosy
ysekky
0
3.1k
Gunosyにおけるニュース記事推薦/ news-recommendation-in-gunosy-webdbf2019
ysekky
1
1.6k
DEIM2019技術報告セッション - Gunosyの研究開発 / deim-2019-sponsor-session-gunosy-research
ysekky
0
1.3k
Analysis of Bias in Gathering Information Between User Attributes in News Application (ABCCS 2018)
ysekky
1
2.5k
世代による政治ニュース記事の閲覧傾向の違いの分析 - JSAI2018 / Analysis of differences in viewing behavior of politics news by age
ysekky
0
4.2k
Other Decks in Research
See All in Research
データサイエンティストの就労意識~2015 → 2026 一般(個人)会員アンケートより
datascientistsociety
PRO
0
940
超効率化への挑戦:1bit LLMの現状と展望
yumaichikawa
0
780
Karkada さんの論文 × 2 の紹介: (1) Closed-Form Training Dynamics Reveal Learned Features and Linear Structure in Word2Vec-like Models, (2) Symmetry in language statistics shapes the geometry of model representations
eumesy
PRO
1
800
The story of RefactoringMiner. Slow research, long-term impact
tsantalis
0
180
Source Code Diff Revolution
tsantalis
0
180
実例から見るLLMのマンガ理解:実務VQAタスクによる長期的文脈と視覚情報の定性評価
kzmssk
0
160
20260624 NLP colloquium: 単一のhubテキストがCLIPを壊す:hubnessによる埋め込みの脆弱性特定
de9uch1
2
280
生成AIなんでも展示会vol6 LT登壇資料 NexteraBERT
rikkabotan7
0
160
最先端NLP勉強会2026 論文紹介:Reasoning with Sampling: Your Base Model is Smarter Than You Think (ICLR 2026 paper)
kogoro
4
660
Easy to Guess, Hard to Verify: Lessons from AIMO 3 for Olympiad-Level AI Mathematics
corochann
0
120
[CV勉強会@関東 CVPR2026] PSDesigner: Automated Graphic Design with a Human-Like Creative Workflow / kantocv 67th CVPR 2026
shunk031
0
350
全国町字単位空き家率推定データver1.0データ仕様
microbaseinc
0
270
Featured
See All Featured
SEO Brein meetup: CTRL+C is not how to scale international SEO
lindahogenes
2
2.9k
Docker and Python
trallard
47
4.2k
First, design no harm
axbom
PRO
2
1.3k
Leading Effective Engineering Teams in the AI Era
addyosmani
9
2.7k
Marketing to machines
jonoalderson
1
5.8k
GitHub's CSS Performance
jonrohan
1033
470k
Producing Creativity
orderedlist
PRO
348
41k
Raft: Consensus for Rubyists
vanstee
142
7.7k
Visualizing Your Data: Incorporating Mongo into Loggly Infrastructure
mongodb
50
10k
Lightning Talk: Beautiful Slides for Beginners
inesmontani
PRO
2
710
The Straight Up "How To Draw Better" Workshop
denniskardys
239
140k
CoffeeScript is Beautiful & I Never Want to Write Plain JavaScript Again
sstephenson
162
16k
Transcript
[論文紹介] Probalis,c Matrix Factoriza,on Ruslan Salakhutdinov
and Andriy Mnih (University of Toronto) NIPS2008 Yoshifumi Seki (Gunosy Inc) 2015.01.27 @Gunosy研究会 #82
概要 • 協調フィルタリングのための次元削減手法の 提案 • NeQlix – 超大規模なデータ
• pLSAなどのギブスサンプリングでは遅いし正確性に欠 ける – バランスの悪いデータ(スパース, 偏り有り) • SVDなどでは評価値が少ないユーザの評価が平均的 なユーザに近づいてしまう
Probabilis,c Matrix Factoriza,on • MのアイテムとNのユー ザ • それぞれD次元の次元 を与えることを考える
• V, Uの各要素は平均0 のガウス分布を仮定
Probabilis,c Matrix Factoriza,on I : i, jに評価値があるときに1, それ以外は0 U, Vの対数事後分布を,
下記の条件の元で最大化する
Probabilis,c Matrix Factoriza,on • 値がすべて満たされている場合にはSVDの確率 モデルへの拡張とみなすことができる • ユーザとアイテムの内積をロジスティクス関数に 通す
• 評価値を0-‐1の範囲にMapする
Automa,c Complexity Control • 新しいデータにも適切に反映できるようにした い – 次元数を調整するのはアンバランスなデータの 場合は適切ではない
– これまでにないようなデータが入ってくる可能性 がある. • 分散のパラメータを調整することで対応させ る
Constrained PMF • PMFではデータが少な いユーザの情報が平均 に近づいてしまう • そのユーザが評価した アイテムの情報が評価
されやすいように制約 を与える
Constrained PMF ユーザのベクタは以下のようにして与えられる よって評価値は以下のようなモデルになる PMFと同様にパラメータ推定を行いモデルを生成する
Experimental Result • Dataset – a subset of NeQlix
• 50,000 users • 1,850 movies • 1,082,982 values – 半分以上は10個以下の評価しかない(sparse) • Parameter – Learning rate: 0.005 – Momentum: 0.9 – D: 30 – λ U, λV, λW, λY = 0.002
Experimental Result
None
None