Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
異常検知の評価指標って何を使えばいいの? / Metrics for one-class cl...
Search
Kon
October 19, 2018
Science
7.4k
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
異常検知の評価指標って何を使えばいいの? / Metrics for one-class classification
https://netadashi.connpass.com/event/100334/
Kon
October 19, 2018
More Decks by Kon
See All by Kon
Numerai はいいぞ / An encouragement of Numerai
yohrn
0
3.3k
M5 Forecasting 参加報告 / 143rd place solution of M5 Forecasting Accuracy
yohrn
1
1.5k
AutoML はお好きですか? / 8th place solution of AutoWSL 2019
yohrn
1
3.6k
3rd Place Solution of AutoSpeech 2019
yohrn
0
520
自然言語処理初心者が AutoNLP に挑戦した話 / 8th place solution of AutoNLP 2019
yohrn
0
990
AutoML パッケージの開発を円滑に進めたい / How to develop AutoML package
yohrn
1
3.7k
機械学習の再現性 / Enabling Reproducibility in Machine Learning Workshop
yohrn
9
3.1k
35th ICML における異常検知に関する論文紹介 / Deep One-Class Classification
yohrn
0
9.8k
機械学習の公平性と解釈可能性 / Fairness, Interpretability, and Explainability Federation of Workshops
yohrn
5
2.7k
Other Decks in Science
See All in Science
コーヒー豆様核 (Coffee-bean nuclei) における形態学的サブタイピングと精選・焙煎特性の同定
jagupath
PRO
0
140
不動産業界における業界特化のデータ整備とAI活用 ─Vertical DataとVertical AI─
estie
1
890
Physical AIを支えるWeights & Biases
olachinkei
1
510
データベース04: SQL (1/3) 単純質問 & 集約演算
trycycle
PRO
0
1.6k
大黒市で発生した大規模インシデント の ポストモーテムから読み解く、 記憶媒体消去の大切さ
shucho0103
0
220
Endel Tulvingとエピソード記憶
rmaruy
0
170
データベース02: データベースの概念
trycycle
PRO
2
1.3k
プロジェクト「Azayaka」のSARの数式とジオメトリ
syuchimu
0
430
第67回コンピュータビジョン勉強会論文紹介「RoboWheel: A Data Engine from Real-World Human Demonstrations for Cross-Embodiment Robotic Learning」
x_ttyszk
0
150
How a camera trap data standard enabled an ecosystem of interoperable tools
peterdesmet
0
110
チュートリアル:世界モデル
hf149
0
2.1k
LLMs vs Chess
ianozsvald
0
110
Featured
See All Featured
Raft: Consensus for Rubyists
vanstee
141
7.6k
Why Our Code Smells
bkeepers
PRO
340
58k
Automating Front-end Workflow
addyosmani
1369
210k
Digital Ethics as a Driver of Design Innovation
axbom
PRO
1
360
Become a Pro
speakerdeck
PRO
31
6.2k
Agile Actions for Facilitating Distributed Teams - ADO2019
mkilby
0
250
BBQ
matthewcrist
89
10k
RailsConf 2023
tenderlove
30
1.5k
Mobile First: as difficult as doing things right
swwweet
225
10k
Measuring Dark Social's Impact On Conversion and Attribution
stephenakadiri
2
250
The Curious Case for Waylosing
cassininazir
1
450
Product Roadmaps are Hard
iamctodd
55
12k
Transcript
Netadashi Meetup #7 Oct 19, 2018 異常検知の評価指標って何を使えばいいの?
Yu Ohori (a.k.a. Kon) NS Solutions Corporation (Apr 2017 -
) • Researcher • Data Science & Infrastructure Technologies • System Research & Development Center • Technology Bureau @Y_oHr_N @Y-oHr-N #SemiSupervisedLearning #AnomalyDetection #DataOps
学習を終えたらモデルの性能を評価しなければならない Chapman, P., et al., "CRISP-DM 1.0 Step-by-step data mining
guides," 2000. 3
不均衡データの場合,評価指標に F 値を使う事が多い 適合率(precision)と 再現率(recall)の調和平均で表される評価指標 実ラベル Y 混同行列 (confusion matrix)
正常 pos: +1 異常 neg: -1 予測ラベル f(X) 正常 pos: +1 true positive (tp) false positive (fp) 異常 neg: -1 false negative (fn) true negative (tn) 4
新規性検知の評価指標は F 値を使えばいいの? 新規性検知の場合,異常標本を一つも入手できない事がある このとき,F 値(正確に言うと適合率)は算出できない いいえ 5
F 値に似た Lee-Liu metric と呼ばれる評価指標がある 適合率と再現率の幾何平均の二乗の定数倍で表される 評価指標 実ラベル Y 混同行列
(confusion matrix) 正常 pos: +1 異常 neg: -1 予測ラベル f(X) 正常 pos: +1 true positive (tp) false positive (fp) 異常 neg: -1 false negative (fn) true negative (tn) Lee, W. S, and Liu, B., "Learning with positive and unlabeled examples using weighted Logistic Regression," In Proceedings of ICML, pp. 448-455, 2003. 6
新規性検知の評価指標は Lee-Liu metric を使えばいいの? ベイズの定理より式変形することで適合率が消える したがって,明示的に適合率を求めることなく算出できる https://stats.stackexchange.com/questions/192530/metrics-for-one-class-classification はい 7