Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
ゼロから作るDeepLearning 第5章 誤差逆伝播法による重み更新を追ってみる
Search
dproject21
February 20, 2017
Science
1.3k
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
ゼロから作るDeepLearning 第5章 誤差逆伝播法による重み更新を追ってみる
dproject21
February 20, 2017
More Decks by dproject21
See All by dproject21
ISTQB/JSTQBシラバスから学ぶAgileTesting / A guide of agile testing based on ISTQB syllabus
dproject21
4
4.1k
JSTQB Advanced Level 模擬問題作成方法 / methodology to questions creation for JSTQB advanced level
dproject21
3
1.6k
試験に絶対出ないJSTQB AL TA,TM問題 / Questions that will never be given on the exam of JSTQB advanced level
dproject21
0
1.6k
The official zip code book is terrible. And what should I do with the address you wrote.
dproject21
0
250
TDD applied Data Cleansing
dproject21
0
2k
Data preprocessing for MachineLearning/BI by Golang and MySQL UDF
dproject21
1
1k
高精度名寄せシステムを支える テキスト処理 (の、ほんのさわり)
dproject21
3
2.7k
ゼロから作るDeepLearning 第7章前半ざっくりまとめ
dproject21
0
1.1k
ゼロから作るDeepLearning 第6章ざっくりまとめ
dproject21
2
1.5k
Other Decks in Science
See All in Science
科学で迫る勝敗の法則-スポーツデータ分析の最前線 (刈谷市連携講座.2026年7月) / The principle of victory discovered by science. at Kariya City, 2027.07
konakalab
0
160
因果推論と機械学習
sshimizu2006
1
1.4k
O(log n)-Approximation Algorithms for Bipartiteness Ratio
tasusu
0
190
CVPR2026_VGGTとその仲間たち
mickey_0226
0
1.1k
ゲームと人工知能
miyayou
0
190
Does the Efficient Compute Frontier Represent New Physics?
drqz
0
120
サンプル対応のない複数遺伝子発現プロファイルに対するテンソル分解型統合解析の要約
tagtag
PRO
0
250
知能とはなにか -ヒトとAIのあいだ-
tagtag
PRO
0
190
Leitner Inauguration Lecture Chalmers University of Technology
xleitix
0
350
Utiliser Bitcoin sans Internet
rlifchitz
0
380
Conwayの法則を"ちゃんと"使うために — 原典でConwayは何を言っていたのか
bonotake
10
7.2k
(CVPR2026) Back to Basics: Let Denoising Generative Models Denoise
shumpei777
0
350
Featured
See All Featured
16th Malabo Montpellier Forum Presentation
akademiya2063
PRO
0
380
Save Time (by Creating Custom Rails Generators)
garrettdimon
PRO
32
4.8k
HU Berlin: Industrial-Strength Natural Language Processing with spaCy and Prodigy
inesmontani
PRO
0
700
個人開発の失敗を避けるイケてる考え方 / tips for indie hackers
panda_program
123
22k
Unsuck your backbone
ammeep
672
58k
Practical Tips for Bootstrapping Information Extraction Pipelines
honnibal
25
2.1k
Refactoring Trust on Your Teams (GOTO; Chicago 2020)
rmw
35
3.8k
Easily Structure & Communicate Ideas using Wireframe
afnizarnur
194
17k
Noah Learner - AI + Me: how we built a GSC Bulk Export data pipeline
techseoconnect
PRO
0
430
JAMstack: Web Apps at Ludicrous Speed - All Things Open 2022
reverentgeek
1
600
Visual Storytelling: How to be a Superhuman Communicator
reverentgeek
2
660
Ethics towards AI in product and experience design
skipperchong
2
380
Transcript
「ゼロから作るDeepLearning」 第5章 誤差逆伝播法の流れをまとめてみる 2017.2.20 たのっち @dproject21
前回質問を頂いた内容を改めて確認しま した。 • 「ゼロから作るDeepLearning」斎藤 康毅 著 オライリー・ジャパンより2016年9⽉ 発⾏ https://www.oreilly.co.jp/books/9784873117584/ •
公式サポートページ https://github.com/oreilly-japan/deep-learning-from-scratch • 第5章「誤差逆伝播法」の重み更新部分です。 https://deeplearning-yokohama.connpass.com/
勾配の計算について " # " # 1 ℎ( ) 勾配 :
すべての変数の偏微分をベクト ルでまとめたもの。 ニューラルネットワークでは、損失関 数の値ができるかぎり⼩さくなるベク トルを、勾配降下法を⽤いて求め、重 み付けを更新する。 . = . − . 学習率 の値は0.01など事前に決めて おく。この学習率の値を変更しながら、 正しく学習できているか確認していく。
勾配の計算について 4.4.1 勾配法で出てくる例を解いてみる。 問: 4 , " = 4 #
+ " # の最⼩値を勾配法で求める。( = 0.1 とする) 1回⽬ : 4 = −3.0, " = 4.0に対して、4 # = −6.0, " # = 8 となる。 4 # = −0.6, " # = 0.8となるので、4 = −2.4, " = 3.2に更新する。 2回⽬ : 4 = −2.4, " = 3.2に対して、 4 # = −4.8, " # = 6.4 となる。 4 # = −0.48, " # = 0.64となるので、4 = −1.92, " = 2.56に更新する。 以降、計算を続けていくと、0に集約されていく。
勾配の計算について では、ニューラルネットワークに対する勾配は? 重みは、最初ランダムな値(正規分布からランダムな値)が⽤いられ、 ← − で更新される。 では、 DE DF の値は、どうやって計算されるか。
損失関数を交差エントロピー誤差 = − ∑ . . log . として求めていく。
勾配の計算について 交差エントロピー誤差 = − ∑ . . log . の偏微分は…
の微分 = 1 O . . log . の微分 = −1 . log . の微分 = それぞれ − 1 log . の微分 = −. . の微分 = − PQ RQ ( = log , DR DS = " S より) (以降、詳細な計算は省略。テキストを参照。)
勾配の計算について 同様に、Softmax関数の偏微分を求めると、 . − . となる。
勾配の計算について シグモイド関数の偏微分は、 (1 − ) ReLU関数の偏微分は、 = T 1 (
> 0) 0 ( ≦ 0) となる。
勾配の計算について Affineレイヤの逆伝播は、ReLUレイヤの各ニューロンからの逆伝播の値を受けて、 DE DW が⼊⼒となる。 Affineレイヤの出⼒Y = + に対して、 バイアスの逆伝播はDE
DW 、⼊⼒データと重みの乗算に対する逆伝播はDE DW ⼊⼒データの逆伝播はDE D[ = DE DW \ ] 重みの逆伝播は DE DF = ] \ DE DW
勾配の計算について 重みの更新は、 それぞれの値に対して⾏うので、 DE DF に学習係数を適⽤し、 ← − ← ""
#" _" "# ## _# − "" #" _" "# ## _# となる。次の学習では、ごくわずかな更新をした重みを⽤いて、 = + に 対する⼊⼒データとの誤差を求める。 4.4.1 勾配法と同様のプロセスで、更新量が漸減していく。