Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
小型ローカルAIで日本語の音声会話botを作った話
Search
Sponsored
·
SiteGround - Reliable hosting with speed, security, and support you can count on.
→
wancoimo
July 26, 2026
Programming
86
1
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
小型ローカルAIで日本語の音声会話botを作った話
LiquidAI社が日本語特化のモデルを発表していたので、それを使って日本語の会話ボットを作ってみた。
wancoimo
July 26, 2026
More Decks by wancoimo
See All by wancoimo
日本語を汚染するAI
route250
1
230
Java嫌いのぼやき
route250
1
32
くそゲーをLLMにやらせてみた
route250
1
17
サーバールームのトラブル〜雪が降るサーバルームの怪〜
route250
1
22
LLMの精度ってどうなの?
route250
1
25
Browser UseをWeb化してみた
route250
1
27
ネット記事からツイートを生成する話
route250
1
22
Other Decks in Programming
See All in Programming
thread_parallel_with_free-threaded_Python_and_NumPy.pdf
riku_sakamoto
0
340
iOSDC Japan 2026 - Swiftで作って学ぼう!データベース自作入門
kaseken
2
170
GraphRAGのKnowledge Graphを 直接!見る/View-GraphRAG's-KnowledgeGraph-directly!
tyumugi1113
1
340
App Intentsのビルドプロセスを支える技術
kntkymt
0
390
Vue Fes Japan 2026 タイムテーブル徹底解説
448jp
1
250
AI に Inclusive UI を書かせよう — Design Rules Skill で Compose UI を作り直す
theoriatec2024
1
500
『寄り添うラジオ』をAIで作る 体験価値から逆算した、会話しないUXと品質設計
theoriatec2024
3
180
What We Talk About When We Talk About XP
m_seki
2
650
ゲームコントローラやキーボードのファームウェアをSwiftで書く
kishikawakatsumi
1
250
Jetpack Compose メカニズム
skydoves
0
440
Augmenting AI with the Power of Jakarta EE
ivargrimstad
0
170
Java 27新機能 / Java 27 new features
kishida
2
150
Featured
See All Featured
Claude Code どこまでも/ Claude Code Everywhere
nwiizo
67
58k
技術選定の審美眼(2025年版) / Understanding the Spiral of Technologies 2025 edition
twada
PRO
120
120k
Lessons Learnt from Crawling 1000+ Websites
charlesmeaden
PRO
1
1.6k
Taking LLMs out of the black box: A practical guide to human-in-the-loop distillation
inesmontani
PRO
3
2.4k
Darren the Foodie - Storyboard
khoart
PRO
4
3.9k
Tell your own story through comics
letsgokoyo
1
1.1k
Are puppies a ranking factor?
jonoalderson
2
3.9k
How to Grow Your eCommerce with AI & Automation
katarinadahlin
PRO
2
280
Performance Is Good for Brains [We Love Speed 2024]
tammyeverts
12
1.8k
Responsive Adventures: Dirty Tricks From The Dark Corners of Front-End
smashingmag
254
22k
職位にかかわらず全員がリーダーシップを発揮するチーム作り / Building a team where everyone can demonstrate leadership regardless of position
madoxten
69
65k
実際に使うSQLの書き方 徹底解説 / pgcon21j-tutorial
soudai
PRO
203
76k
Transcript
AIエージェント開発LT大会 小型ローカルAIで 日本語の音声会話botを 作った LFM2.5-Audio-1.5B-JP × LFM2.5-1.2B-JP-202606 顔認証 / 擬似フルデュプレックス会話
2026.07.25
LIQUID AI / VOICE BOT 02 Liquid AI:会社概要と日本語対応の経緯 会社概要 本社
設立 日本拠点 米国マサチューセッツ州 ケンブリッジ 2023年 日本法人:Liquid AI株式会社 (公式は日付非公開) CTCが2024年に出資を発表 特色 データセンターの外で動く基盤モデル 低遅延・プライバシー・実機ハードウェアの制約を前提に、端末上で動くLFMを開発 日本語対応の経緯 2024.10 2025.07 2026.01 CTCと協業 LFM2 LFM2.5-JP 日本語LLMの共同開発・ 多言語対応の重点言語に 日本語アプリ向けモデルを公開 技術検証を発表 日本語を含める (文化・言語ニュアンスを重視)
LIQUID AI / VOICE BOT 03 今回試した2つの日本語モデル TEXT LLM AUDIO
LLM LFM2.5-1.2B-JP-202606 LFM2.5-Audio-1.5B-JP 日本語に特化した 音声認識と音声合成を扱う 1.2Bパラメータのチャットモデル 1.5Bパラメータの音声モデル 会話・応答生成 ASR / TTS
LIQUID AI / VOICE BOT 04 Audioモデルは、軽量で応答が速い 試用した印象 ASR Whisper
small程度の認識精度と感じた TTS 女性音声は1種類だが、自然に聞こえる 実行感 小型モデルでメモリ使用量が少なく、応答が速い ※ ASR精度・応答速度は、発表者の実装環境における主観評価 音声入力をテキスト化 テキストを音声化 ローカル実行に向く
LIQUID AI / VOICE BOT 05 一体型ではなく、4段パイプラインにした VADで発話区間を切り出すため、分離しても実装量は大きく増えない VAD ASR
LLM TTS 発話区間を検出 Audio-1.5B-JP 1.2B-JP-202606 Audio-1.5B-JP 分離した理由 • VADの処理が必要 • ツール制御をLLM側で実装したい • モデルごとの挙動を個別に調整できる
LIQUID AI / VOICE BOT 06 会話botとしては、LLM側の制御に課題が残った LFM2.5-1.2B-JP-202606 追加した機能 自由会話は自然
顔認証 利用者を識別して会話を開始 役割・出力形式・ツール実行条件を含む 会話bot用プロンプトでは、意図した制御を 毎回同じように得るには調整が必要だった ※ 課題は発表者のプロンプト・実装条件での観察 擬似フルデュプレックス 音声再生中も次の入力を受ける
DEMO 実機デモ 顔認証つき・擬似フルデュプレックス 音声会話bot 画面を切り替えます 07
LIQUID AI / VOICE BOT まとめ:Audioは有望、JP LLMは用途を選ぶ 今回わかったこと 01 LFM2.5-Audio-1.5B-JP
ASR・TTSともに実用的。小型でローカル会話botに使いやすい 02 LFM2.5-1.2B-JP-202606 自然な日本語会話はできるが、bot制御にはプロンプト検証が必要 03 次の展開 LFM2.5-VLを使い、カメラ画像も入力できる会話botへ 08
ご清聴ありがとうございました AIエージェント開発LT大会