Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
オッカムの剃刀と汎化誤差解析
Search
Masanari Kimura
August 31, 2021
Research
5.4k
3
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
オッカムの剃刀と汎化誤差解析
Masanari Kimura
August 31, 2021
More Decks by Masanari Kimura
See All by Masanari Kimura
Equivalence of Geodesics and Importance Weighting from the Perspective of Information Geometry
mkimura
0
380
機械学習における重要度重み付けとその応用
mkimura
3
3.5k
Paper Intro: Human Rademacher Complexity
mkimura
0
250
On the principle of Invariant Risk Minimization
mkimura
0
410
論文紹介:Clustering with Bregman Divergences: an Asymptotic Analysis
mkimura
0
640
Generalization Bounds for Set-to-Set Matching with Negative Sampling
mkimura
0
210
論文紹介:On the Importance of Gradients for Detecting Distributional Shifts in the Wild
mkimura
2
930
論文紹介:Dangers of Bayesian Model Averaging under Covariate Shift
mkimura
0
390
Information Geometry of Dropout Training
mkimura
0
370
Other Decks in Research
See All in Research
GLIM とMegaParticles:正規分布近似の限界とタイトカップリング&パーティクルフィルタの進展 / GLIM and MegaParticles : Progress of the distribution representation in SLAM
koide3
0
750
Research Engineerという仕事 / Research Engineering: Bridging Research and Business
chck
1
280
シングルチャネルマルチトーカー音声認識の進展
ryomasumura
0
240
SAKURAONE:An Open Ethernet-based AI HPC System And Its Observed Workload Dynamicsin a Single-Tenant LLM Development Environment
yuukit
1
550
Data Visualization Tools in the Age of AI
flekschas
0
190
さくらインターネット研究所テックトーク2026春、研究開発Gr.25年度成果26年度方針
kikuzo
0
180
【中間報告】国会議員の立法・政策実務を支える環境を巡る現状と課題
polipoli
0
500
LA-Bench 2025:実験指示から実行可能手順を生成するためのデータセット/LA-Bench 2025: A Dataset for Generating Executable Experimental Procedures from Experimental Instructions
stktu
0
140
敵対生成プロンプト同時探索による内省型プロンプト最適化
kinoue_smarthr
0
390
論文読み会 SNLP2026 Tau2-Bench: Evaluating Conversational Agents in a Dual-Control Environment
s_mizuki_nlp
0
150
RS-Agent: Automating Remote Sensing Tasks through Intelligent Agent
satai
3
500
kintone リサーチ副部/UXリサーチャー 業務紹介
cybozuinsideout
PRO
0
150
Featured
See All Featured
How STYLIGHT went responsive
nonsquared
100
6.2k
Ecommerce SEO: The Keys for Success Now & Beyond - #SERPConf2024
aleyda
1
2.1k
The SEO identity crisis: Don't let AI make you average
varn
0
540
How To Stay Up To Date on Web Technology
chriscoyier
790
250k
Introduction to Domain-Driven Design and Collaborative software design
baasie
1
950
Leading Effective Engineering Teams in the AI Era
addyosmani
9
2.4k
Agile Leadership in an Agile Organization
kimpetersen
PRO
0
210
Raft: Consensus for Rubyists
vanstee
141
7.7k
Designing Experiences People Love
moore
143
24k
Templates, Plugins, & Blocks: Oh My! Creating the theme that thinks of everything
marktimemedia
31
2.9k
Building Flexible Design Systems
yeseniaperezcruz
330
41k
What’s in a name? Adding method to the madness
productmarketing
PRO
24
4.1k
Transcript
Intro Occan Bound Additional Discussions References オッカムの剃刀と汎化誤差解析 Masanari Kimura
[email protected]
August 31, 2021
Intro Occan Bound Additional Discussions References Intro 2/11
Intro Occan Bound Additional Discussions References TL;DR ▶ オッカムの剃刀の概念について説明; ▶
オッカムの剃刀の形式化と汎化誤差解析への応用について説明. 3/11
Intro Occan Bound Additional Discussions References オッカムの剃刀(Occam’s Razor) オッカム [Drouhin,
2006] 必要が無いなら多くのものを定立してはならない.少数の論理でよい場合は多数の論理を 定立してはならない. ▶ ある二つの理論が同程度にデータを説明できているとき,より単純な方が好まれる; ▶ 統計的機械学習において単純さは直感的にだけでなく定量的に測れる; ▶ 以下ではオッカムの剃刀を形式的に記述していく. 4/11
Intro Occan Bound Additional Discussions References Occan Bound 5/11
Intro Occan Bound Additional Discussions References Occam Bound Theorem 独立かつ同一なサンプルサイズ
m のデータセット S = {x, y} とある仮説 h ∈ H について 少なくとも 1 − δ の確率で以下が成り立つ: L(h) ≤ ˆ L(h) + √ (ln 2)|h| + ln 1 δ 2m . (1) ただし,|h| は仮説 h を記述するのに必要な bit 数であり, L(h) := E [ 1[h(x) ̸= y] ] , (2) ˆ L(h) := 1 m m ∑ i=1 1[h(xi) ̸= yi]. (3) 6/11
Intro Occan Bound Additional Discussions References Proof of the Occam
Bound Proof. 定理に矛盾する仮説集合を B とする: B := { L(h) ≥ ˆ L(h) + √ (ln 2)|h| + ln 1 δ 2m ; h ∈ H } (4) このとき, P [ h ∈ B ] ≤ ∑ h∈H exp { −2m (√ (ln 2)|h| + ln 1 δ 2m )2 } (∵ Chernoff bound) (5) = ∑ h∈H δ2−|h| = δ ∑ h∈H 2−|h| ≤ δ (∵ Kraft inequality) (6) 7/11
Intro Occan Bound Additional Discussions References Occam Bound と仮説選択 Occam
bound は期待誤差の上界を与えるので,これを最小化するように仮説選択をする ことが考えられる: ˆ h = arg min h∈H ˆ L(h) + √ (ln 2)|h| + ln 1 δ 2m . (7) ▶ この最適化は,手元へのデータの説明能力(第一項)とモデルのシンプルさ(第二 項)の最小化のトレードオフになっている; ▶ これは,ある h1 , h2 ∈ H がもし同じだけデータを説明できるとき,よりシンプルな方 が未知のデータへの誤差を小さくできる可能性が高いことを意味している; ▶ これはまさしくオッカムの剃刀の形式的な記述になっている. 8/11
Intro Occan Bound Additional Discussions References Additional Discussions 9/11
Intro Occan Bound Additional Discussions References Occam Bound のベイズ的解釈 P
を h に関する確率分布とし,|h|P を以下のように定義する: |h|P := log 2 1 P(h) . (8) このとき,Occam bound は次のように書き換えることができる: L(h) ≤ ˆ L(h) + √ (ln 2)|h|P + ln 1 δ 2m . (9) これはまさしく仮説集合に関する任意の事前分布を考えた場合の Occam bound に相当 する. 10/11
Intro Occan Bound Additional Discussions References References I Nicolas Drouhin.
Pluralitas non est ponenda sine neccesitate. Technical report, GRID Working paper, 2006. 11/11