Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
P 値と有意差/分散分析 / P-value, Significant Difference ...
Search
Kenji Saito
PRO
January 03, 2025
Technology
0
120
P 値と有意差/分散分析 / P-value, Significant Difference and Analysis of Variance
早稲田大学大学院経営管理研究科「企業データ分析」2024 冬の第9-10回で使用したスライドです。
Kenji Saito
PRO
January 03, 2025
Tweet
Share
More Decks by Kenji Saito
See All by Kenji Saito
AI が研究する時代に、人はどう育つのか? — GAMER PAT にみる "シリアスゲームとしての知的訓練" / In an era where AI conducts research, how will humans develop? — "Intellectual Training as a Serious Game" Seen in GAMER PAT
ks91
PRO
0
36
FinTech 5-6 : The World of Apps
ks91
PRO
0
67
生成AI による論文執筆サポート・ワークショップ ─ サーベイ/リサーチクエスチョン編 / Workshop on AI-Assisted Paper Writing Support: Survey/Research Question Edition
ks91
PRO
0
66
ブロックチェーン概論とインストール大会 / Introduction to Blockchain and Installation Workshop
ks91
PRO
0
2
FinTech 3-4 : Internet Technology and Governance
ks91
PRO
0
71
民主主義と博愛(Humanitarianism) / Democracy and Humanitarianism
ks91
PRO
0
6
ブロックチェーン概論 / Introduction to Blockchain
ks91
PRO
0
12
ブロックチェーンと分散ファイナンス概論 / Introduction to Blockchain and Decentralized Finance
ks91
PRO
0
65
Proof of Authenticity of General IoT Information with Tamper-Evident Sensors and Blockchain
ks91
PRO
0
7
Other Decks in Technology
See All in Technology
HR Force における DWH の併用事例 ~ サービス基盤としての BigQuery / 分析基盤としての Snowflake ~@Cross Data Platforms Meetup #2「BigQueryと愉快な仲間たち」
ryo_suzuki
0
210
ガバメントクラウド(AWS)へのデータ移行戦略の立て方【虎の巻】 / 20251011 Mitsutosi Matsuo
shift_evolve
PRO
2
200
スタートアップにおけるこれからの「データ整備」
shomaekawa
2
480
AWS Top Engineer、浮いてませんか? / As an AWS Top Engineer, Are You Out of Place?
yuj1osm
2
210
incident_commander_demaecan__1_.pdf
demaecan
0
130
なぜAWSを活かしきれないのか?技術と組織への処方箋
nrinetcom
PRO
4
880
プレーリーカードを活用しよう❗❗デジタル名刺交換からはじまるイベント会場交流のススメ
tsukaman
0
160
Contract One Engineering Unit 紹介資料
sansan33
PRO
0
8.8k
Introduction to Sansan for Engineers / エンジニア向け会社紹介
sansan33
PRO
5
43k
ガバメントクラウドの概要と自治体事例(名古屋市)
techniczna
2
240
綺麗なデータマートをつくろう_データ整備を前向きに考える会 / Let's create clean data mart
brainpadpr
3
510
Adminaで実現するISMS/SOC2運用の効率化 〜 アカウント管理編 〜
shonansurvivors
4
450
Featured
See All Featured
Helping Users Find Their Own Way: Creating Modern Search Experiences
danielanewman
30
2.9k
CSS Pre-Processors: Stylus, Less & Sass
bermonpainter
359
30k
Creating an realtime collaboration tool: Agile Flush - .NET Oxford
marcduiker
33
2.3k
We Have a Design System, Now What?
morganepeng
53
7.8k
Navigating Team Friction
lara
190
15k
Understanding Cognitive Biases in Performance Measurement
bluesmoon
31
2.7k
YesSQL, Process and Tooling at Scale
rocio
173
14k
How to Think Like a Performance Engineer
csswizardry
27
2k
Rebuilding a faster, lazier Slack
samanthasiow
84
9.2k
Building a Scalable Design System with Sketch
lauravandoore
463
33k
実際に使うSQLの書き方 徹底解説 / pgcon21j-tutorial
soudai
PRO
189
55k
The World Runs on Bad Software
bkeepers
PRO
72
11k
Transcript
Corporate data analysis — generated by Stable Diffusion XL v1.0
2024 9-10 P (WBS) 2024 9-10 P — 2025-01-06 – p.1/33
https://speakerdeck.com/ks91/collections/corporate-data-analysis-2024-winter 2024 9-10 P — 2025-01-06 – p.2/33
( ) 1 12 2 • 2 12 2 (B
A ) • 3 12 9 • 4 12 9 • 5 12 16 • 6 12 16 t • 7 12 23 2 ( ) t • 8 12 23 2 ( ) t • 9 1 6 P • 10 1 6 • 11 1 20 12 1 20 13 1 27 14 1 27 W-IOI 2024 9-10 P — 2025-01-06 – p.3/33
( 20 25 ) 1 (20 ) • 2 R
( 55 ) • 3 (32 ) • 4 (14 ) • 5 ( Git) (22 ) • 6 ( ) (24 ) • 7 (1) (25 ) • 8 (2) (25 ) • 9 R ( ) (1) — Welch (17 ) • 10 R ( ) (2) — (21 ) • 11 R ( ) (1) — (15 ) • 12 R ( ) (2) — (19 ) • 13 GPT-4 (19 ) • 14 GPT-4 (29 ) • 15 ( ) LaTeX Overleaf (40 ) • 8 (12/16 ) / (2 ) OK / 2024 9-10 P — 2025-01-06 – p.4/33
( Student µ 95% ) 7 2 t ( t
) 2 ( ) 2 d ( ) ← [ 3] σd 2 t 8 2 t ( t ) 2 ( ) ( ) ← [ 4] σ 2 t 2024 9-10 P — 2025-01-06 – p.5/33
2 2 t 1 9 P P 10 H0 HA
k, N, ¯ ¯ x σ2 ( )MSwithin ( )MSbetween MStotal F F 2024 9-10 P — 2025-01-06 – p.6/33
2024 9-10 P — 2025-01-06 – p.7/33
4. t (1) 2 t (2) 2 t (3) 2025
1 2 ( ) 23:59 JST ( ) Waseda Moodle (Q & A ) (1)(2) Discord 2024 9-10 P — 2025-01-06 – p.8/33
. . . . . . 17 14 (1/3( )
) ( ) → 14 ( ) ( ) → 6 → 3 ( ) → 5 ( ) ( OK) 2 t . . . . . . / . . . ( ) 2024 9-10 P — 2025-01-06 – p.9/33
t t ⇒ ( ) A A xA 2 B
B xB 2 df . . . ⇒ t σ z0.05 . . . ⇒ ( ) t 2024 9-10 P — 2025-01-06 – p.10/33
N (1/2) 2 t 2 2 “ ” 1. 1
2 2. 3. - 2 (n − 1) 4. ÷ ÷ t 5. t (n − 1) t t ⇒ . . . 0 ( ) 2024 9-10 P — 2025-01-06 – p.11/33
N (2/2) 2 t 2 2 1 2 1 2
2 “ ” 1. 2 1 2 2. 3. ( -2) 4. t 1÷ 2 t 5. t (n1 + n2 − 2) t t ⇒ 2024 9-10 P — 2025-01-06 – p.12/33
M ( ) [ 2 t ] 1Day 1Day 1Day
⇒ 2024 9-10 P — 2025-01-06 – p.13/33
K ⇒ . . . 2024 9-10 P — 2025-01-06
– p.14/33
2 t d : µd 0 ( 2 ) :
(1) d d, sd , n, df (2) |d| sd n |t| (3) t0.05 (df) < |t| ( ) R > t.test(sample2, sample1, paired=T) 2024 9-10 P — 2025-01-06 – p.15/33
2 t ( ) 10 ( ) ( ) (
) ( ) ( ) d ( ) d, ( ) sd , ( ) n, ( ) df ( ) t ( ) t ( ) d ( ) sd ( ) n ( ) t df 5% ( ) ( ) ( ) ( ) ( ) 2024 9-10 P — 2025-01-06 – p.16/33
2 t xA xB : µA − µB 0 (
2 ) : (1) xA − xB , sp , nA nB , df (2) |xA − xB | sp nA nB |t| (3) t0.05 (df) < |t| ( ) R > t.test(sample2, sample1, var.equal=T) 2024 9-10 P — 2025-01-06 – p.17/33
2 t ( ) ( ) ( ) ( )
( ) ( ) ( ( ) A B ( ) ) ( ) xA − xB , A B ( ) ( ) sp , ( ) nA nB , ( ) df = nA + nB − 2 ( ) t ( ) t ( ) xA − xB ( ) sp ( ) nA ,nB ( ) t df 5% ( ) ( ) ( ) ( ) ( ) ( ) 2024 9-10 P — 2025-01-06 – p.18/33
K 2 t ( ) ⇒ 2 2024 9-10 P
— 2025-01-06 – p.19/33
N ⇒ (σ) ( σ √n ) ( ) p.121
(standard error) (p.121) (sampling distribution) (p.120) (p.120) ( : ) 2024 9-10 P — 2025-01-06 – p.20/33
K ⇒ . . . AI ( ) . .
. ^^; ( ) 2024 9-10 P — 2025-01-06 – p.21/33
H t 2 Student t t 1 sin(α + β)
= sinαcosβ + cosαsinβ . . . ⇒ 2024 9-10 P — 2025-01-06 – p.22/33
U R ChatGPT ⇒ AI ( ) 2024 9-10 P
— 2025-01-06 – p.23/33
9 P P 2024 9-10 P — 2025-01-06 – p.24/33
α β P P H0 ( ) P 0.05 (P
= 0.015) (P = 0.361) 2024 9-10 P — 2025-01-06 – p.25/33
10 H0 HA k, N, ¯ ¯ x σ2 (
) MSwithin ( )MSbetween MStotal ( SStotal dftotal ) F F 2024 9-10 P — 2025-01-06 – p.26/33
(1/3) k (1) : (2) : σ2 ( ) N(µ,
σ2) µ1 = µ2 = · · · = µk N ( ) ¯ ¯ x ¯ ¯ x = k j=1 nj i=1 xji N (j i N ) 2024 9-10 P — 2025-01-06 – p.27/33
(2/3) ( )MSwithin σ2 MSwithin = SSwithin dfwithin = k
j=1 nj i=1 (xji − ¯ xj )2 N − k ( N− ) ( )MSbetween σ2 MSbetween = SSbetween dfbetween = k j=1 nj (¯ xj − ¯ ¯ x)2 k − 1 ( −1 ) ( H0 σ2 ) 2024 9-10 P — 2025-01-06 – p.28/33
(3/3) MStotal MStotal = SStotal dftotal = k j=1 nj
i=1 (xji − ¯ ¯ x)2 N − 1 ( N − 1 ) : SStotal = SSbetween + SSwithin, dftotal = dfbetween + dfwithin F F = MSbetween MSwithin F0.05 (dfbetween, dfwithin ) < F ( H0 ) 2024 9-10 P — 2025-01-06 – p.29/33
U ( p.227) 20 4 “ U.R” ( anova() )
pp.226–227 2024 9-10 P — 2025-01-06 – p.30/33
2024 9-10 P — 2025-01-06 – p.31/33
5. (1) ( ) (2) 2025 1 16 ( )
23:59 JST ( ) Waseda Moodle (Q & A ) (1)(2) Discord 2024 9-10 P — 2025-01-06 – p.32/33
2024 9-10 P — 2025-01-06 – p.33/33