Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
機械学習のための音声信号処理(基礎編)/ Speech Signal Processing
Search
moonlight-aska
October 21, 2018
1.3k
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
機械学習のための音声信号処理(基礎編)/ Speech Signal Processing
2018年10月21日開催の「大江橋Pythonの会#3」の資料です.
moonlight-aska
October 21, 2018
More Decks by moonlight-aska
See All by moonlight-aska
Create Your Own AI with Dify×Gemma3
aska
0
82
Generative AI Prototyping
aska
0
37
【入門】プロンプトの書き方のコツ / Tips for writing prompts
aska
0
250
CHATGPT。はじめの一歩 / ChatGPT. Get Started
aska
0
160
「Kingyo AI Navi」アプリ / Kingyo AI Navi App
aska
0
290
Kingo AI Navi LINEをもっと使い倒せ!!
aska
0
170
Depth画像で物体検知やってみたー。/ Objects Detection with Depth Images
aska
0
880
Kingyo AI Naviアプリ開発 / Kingyo AI Navi App
aska
0
460
AutoML Vision Edgeで金魚分類モデルを学習してみた / Kingyo Classification Model with AutoML Vision Edge
aska
0
600
Featured
See All Featured
From π to Pie charts
rasagy
1
380
Exploring anti-patterns in Rails
aemeredith
4
510
My Coaching Mixtape
mlcsv
0
320
JAMstack: Web Apps at Ludicrous Speed - All Things Open 2022
reverentgeek
1
600
Building Adaptive Systems
keathley
44
3.2k
The State of eCommerce SEO: How to Win in Today's Products SERPs - #SEOweek
aleyda
2
12k
Highjacked: Video Game Concept Design
rkendrick25
PRO
1
460
Digital Ethics as a Driver of Design Innovation
axbom
PRO
1
430
Technical Leadership for Architectural Decision Making
baasie
3
580
I Don’t Have Time: Getting Over the Fear to Launch Your Podcast
jcasabona
35
2.8k
A designer walks into a library…
pauljervisheath
211
25k
The Impact of AI in SEO - AI Overviews June 2024 Edition
aleyda
6
1.2k
Transcript
None
NARA
안녕하세요
None
f P t
None
None
None
None
None
None
Convolutional Recurrent Neural Network
[音情報処理論 音声処理における信号処理1より引用]
None
None
×
×
[音声言語処理特論 第2回音声認識の基礎、DPマッチングの基礎より引用]
None
None
None
None
None
None
https://www.kaggle.com/c/tensorflow-speech-recognition-challenge
None
MFCC12 ΔMFCC12 ΔΔMFCC12
None
None
None
None
--- No. 6795 edc53350_nohash_0.wav (house ) --- * 1位 :
house (0.999820) 2位 : cat (0.000060) 3位 : off (0.000032) 4位 : yes (0.000024) 5位 : down (0.000013) --- No. 6796 e95c70e2_nohash_0.wav (house ) --- * 1位 : house (0.999953) 2位 : off (0.000012) 3位 : cat (0.000005) 4位 : eight (0.000004) 5位 : happy (0.000004) --- No. 6797 258f4559_nohash_0.wav (house ) --- * 1位 : house (0.999980) 2位 : off (0.000007) 3位 : happy (0.000004) 4位 : cat (0.000004) 5位 : eight (0.000002) --- No. 6798 1657c9fa_nohash_0.wav (house ) --- * 1位 : house (0.999972) 2位 : off (0.000011) 3位 : yes (0.000003) 4位 : happy (0.000003) 5位 : cat (0.000003) ---------- Total Accuracy ---------- 1位 : 93.57 % ( 6361 / 6798 ) 2位 : 96.78 % ( 6579 / 6798 ) 3位 : 97.78 % ( 6647 / 6798 ) 4位 : 98.35 % ( 6686 / 6798 ) 5位 : 98.57 % ( 6701 / 6798 )
None
NARA
None
None
None
https://ai.googleblog.com/2017/08/launching-speech-commands-dataset.html