Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
Learning Lightweight Lane Detection CNNs by Sel...
Search
catla
September 13, 2019
Research
640
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
Learning Lightweight Lane Detection CNNs by Self Attention Distillation(ICCV2019)の紹介
catla
September 13, 2019
More Decks by catla
See All by catla
ベイズ深層学習(6.3)
catla
2
240
ベイズ深層学習(6.2)
catla
3
240
[読み会資料] Federated Learning for Vision-and-Language Grounding Problems
catla
0
340
ベイズ深層学習(5.1~5.2)
catla
0
250
ベイズ深層学習(4.1)
catla
0
470
ベイズ深層学習(3.3~3.4)
catla
19
11k
ベイズ深層学習(2.2~2.4)
catla
6
1.3k
23回アルゴリズムコンテスト 1位解法
catla
6
690
TGS Salt Identification Challenge 12th place solution
catla
3
12k
Other Decks in Research
See All in Research
SLAMはどこまで解決されたのか?
tomonom
0
1.1k
Cross-Media Human-Information Interaction
signer
PRO
0
180
Visual SLAM未来予測 / Future Prediction in Visual SLAM
koide3
1
880
2026年版中小企業白書・小規模企業白書の概要
ozekinote
0
160
Claude Code × autoresearch 実践
mathbullet
0
230
マーケットストリート 社会実験2024 in 秋葉原ジャンク通り 調査報告書
izumiyama_lab
1
110
Apache Gravitinoで実現する Icebergカタログ統合とアクセスの一元化
matsumooon
0
440
[最先端NLP勉強会2026] Checklists Are Better Than Reward Models For Aligning Language Models
nzw0301
1
130
typst の使い方:言語学を研究する学生のために
gitomochang
0
550
ScoreMatchingRiesz for Automatic Debiased Machine Learning and Policy Path Estimation with an Application to Japanese Monetary Policy Evaluation
masakat0
0
320
多様なデータを許容し学習し続ける模倣学習 / Advanced Imitation Learning for VLA
prinlab
0
280
Sleuthcon Keynote - How Cybercriminals (ab)use AI
fr0gger
0
290
Featured
See All Featured
Creating an realtime collaboration tool: Agile Flush - .NET Oxford
marcduiker
35
2.5k
Save Time (by Creating Custom Rails Generators)
garrettdimon
PRO
32
4.4k
DevOps and Value Stream Thinking: Enabling flow, efficiency and business value
helenjbeal
1
340
No one is an island. Learnings from fostering a developers community.
thoeni
21
3.8k
From π to Pie charts
rasagy
0
300
Understanding Cognitive Biases in Performance Measurement
bluesmoon
32
3k
Marketing to machines
jonoalderson
1
5.7k
Exploring the Power of Turbo Streams & Action Cable | RailsConf2023
kevinliebholz
37
6.6k
Accessibility Awareness
sabderemane
1
180
Building a Modern Day E-commerce SEO Strategy
aleyda
45
9.2k
Digital Ethics as a Driver of Design Innovation
axbom
PRO
1
370
Organizational Design Perspectives: An Ontology of Organizational Design Elements
kimpetersen
PRO
1
800
Transcript
https://arxiv.org/abs/1908.00821 桂 尚輝 (Katsura Naoki) #7【画像処理 & 機械学習】論文LT会 2019.09.13
どんなもの? 先行研究と比べてどこがすごい? 技術や手法のキモはどこ? どうやって有効かを検証した? 議論はある? 入力画像(RGB)から車線を検出する.
蒸留に追加データやラベルを必要とせず,先行研究である message passing機能を有するSCNNは順伝播に全体の 35%の時間を占めるのに対し,提案手法の推論時間はベー スモデルと同程度で精度の向上を達成. ベーシックな蒸留は,教師モデルを用いて新たなモデルを 学習させるが,提案手法のSelf Attention Distillation(SAD) は,自分自身の深いレイヤーにおけるアテンションマップを 浅いレイヤーの蒸留ラベルに使用. したがって, 教師モデ ル等から得られる追加ラベルが必要なく, モデル自体も大 きくならない. Lane Detectionにおける3つのベンチマーク(TuSimple, BDD100K, CULane)で実験を行い,先行研究との比較や SADに関するablation studyを行い有効性を検証. SADを導入することで各レイヤーのアテンションマップが 良くなった. 細部にこだわるタスクにおいても有効かもしれ ない.
背景&先行研究 (1) 車線が画像に対してスパースであるので学習が難しい問題. (2) また,前に走ってる車によってレーンが見えない,実際のレーンが曖昧に引かれている(視覚情報が弱い),路 面状況が悪いといった時にうまく予測できない問題もある.
(3) 先行研究でSoTAなSCNNはmessage passing(MP)によって精度向上を達成したがMP部分に推論時間の 35%が占められていて計算コストが高い. (4) アプローチ方法としては,セグメンテーションとして解く,セグメンテーションしたのちに点をサンプリングして 多項式に変形,多項式のパラメータを予測する方法があるっぽい (?). この論文はsemantic segmentationで解く
Self Attention Distillation 一つレイヤーの深いアテンションマップを蒸 留に用いる教師ラベルとして使用する. 蒸留に対する損失関数は, L2 Loss
Self Attention Distillation 次ココ!
AT-GEN Bilinear Upsample Spatial Softmax p=2で実験したもの が一番よかったらし い C ×
H × W H × W H’ × W’ H’ × W’
Result (Attention map)
Result (vs Previous work)
Result (vs Deep supervision)
Message Passing (MP) https://arxiv.org/abs/1712.06080 MRF/CRF SCNN 近くのピクセル情報を関連づける
Message Passing (MP) https://arxiv.org/abs/1712.06080 CNNは、細く連続な物体や大きな物体のsemantic情報 に弱いらしく、SCNNでそれを改善できたそう