Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
P 値と有意差/分散分析 / P-value, Significant Difference ...
Search
Kenji Saito
PRO
January 03, 2025
Technology
0
29
P 値と有意差/分散分析 / P-value, Significant Difference and Analysis of Variance
早稲田大学大学院経営管理研究科「企業データ分析」2024 冬の第9-10回で使用したスライドです。
Kenji Saito
PRO
January 03, 2025
Tweet
Share
More Decks by Kenji Saito
See All by Kenji Saito
関連2群のt検定/独立2群のt検定 / Related 2-group t-test and independent 2-group t-test
ks91
PRO
0
51
A Guide to Paper Writing Support with Generative AI - A Joint Zemi
ks91
PRO
0
9
正規分布と簡単な統計理論/t分布と信頼区間 / Normal distribution, simple statistical theory, t-distribution and confidence intervals
ks91
PRO
0
43
じわじわ迫ってきている自動化社会 (その先にメタ・ネイチャー) / The Slowly Approaching Automated Society (and its beyond: Meta-Nature)
ks91
PRO
0
6
起こりうる誤った推論/平均・分散・標準偏差・自由度 / Possible false inferences, means, variances, standard deviations and degrees of freedom
ks91
PRO
0
59
LaTeX と Overleaf によるショートペーパー作成 / Short paper writing with LaTeX and Overleaf
ks91
PRO
0
23
R を用いた検定(補講) (1) — Welch 検定 / Tests using R (supplementary) (1) - Welch test
ks91
PRO
0
12
R を用いた検定(補講) (2) — カイ二乗検定 / Tests using R (supplementary) (2) - Chi-squared test
ks91
PRO
0
13
R を用いた分析(補講) (1) — 重回帰分析 / Analysis using R (supplementary) (1) - Multiple regression analysis
ks91
PRO
0
13
Other Decks in Technology
See All in Technology
Server-Side Engineer of LINE Sukimani
lycorp_recruit_jp
1
470
DevFest 2024 Incheon / Songdo - Compose UI 조합 심화
wisemuji
0
230
「完全に理解したTalk」完全に理解した
segavvy
1
230
Oracle Cloudの生成AIサービスって実際どこまで使えるの? エンジニア目線で試してみた
minorun365
PRO
5
330
小学3年生夏休みの自由研究「夏休みに Copilot で遊んでみた」
taichinakamura
0
200
大規模言語モデルとそのソフトウェア開発に向けた応用 (2024年版)
kazato
1
240
能動的ドメイン名ライフサイクル管理のすゝめ / Practice on Active Domain Name Lifecycle Management
nttcom
0
300
MasterMemory v3 最速確認会
yucchiy
0
260
なぜCodeceptJSを選んだか
goataka
0
200
プロダクト組織で取り組むアドベントカレンダー/Advent Calendar in Product Teams
mixplace
0
570
開発生産性向上! 育成を「改善」と捉えるエンジニア育成戦略
shoota
2
760
成果を出しながら成長する、アウトプット駆動のキャッチアップ術 / Output-driven catch-up techniques to grow while producing results
aiandrox
0
430
Featured
See All Featured
The Web Performance Landscape in 2024 [PerfNow 2024]
tammyeverts
3
310
Large-scale JavaScript Application Architecture
addyosmani
510
110k
個人開発の失敗を避けるイケてる考え方 / tips for indie hackers
panda_program
97
17k
The Cost Of JavaScript in 2023
addyosmani
46
7k
Fight the Zombie Pattern Library - RWD Summit 2016
marcelosomers
232
17k
Java REST API Framework Comparison - PWX 2021
mraible
28
8.3k
CoffeeScript is Beautiful & I Never Want to Write Plain JavaScript Again
sstephenson
159
15k
Into the Great Unknown - MozCon
thekraken
34
1.6k
RailsConf 2023
tenderlove
29
960
Scaling GitHub
holman
459
140k
Code Reviewing Like a Champion
maltzj
521
39k
The Cult of Friendly URLs
andyhume
78
6.1k
Transcript
Corporate data analysis — generated by Stable Diffusion XL v1.0
2024 9-10 P (WBS) 2024 9-10 P — 2025-01-06 – p.1/33
https://speakerdeck.com/ks91/collections/corporate-data-analysis-2024-winter 2024 9-10 P — 2025-01-06 – p.2/33
( ) 1 12 2 • 2 12 2 (B
A ) • 3 12 9 • 4 12 9 • 5 12 16 • 6 12 16 t • 7 12 23 2 ( ) t • 8 12 23 2 ( ) t • 9 1 6 P • 10 1 6 • 11 1 20 12 1 20 13 1 27 14 1 27 W-IOI 2024 9-10 P — 2025-01-06 – p.3/33
( 20 25 ) 1 (20 ) • 2 R
( 55 ) • 3 (32 ) • 4 (14 ) • 5 ( Git) (22 ) • 6 ( ) (24 ) • 7 (1) (25 ) • 8 (2) (25 ) • 9 R ( ) (1) — Welch (17 ) • 10 R ( ) (2) — (21 ) • 11 R ( ) (1) — (15 ) • 12 R ( ) (2) — (19 ) • 13 GPT-4 (19 ) • 14 GPT-4 (29 ) • 15 ( ) LaTeX Overleaf (40 ) • 8 (12/16 ) / (2 ) OK / 2024 9-10 P — 2025-01-06 – p.4/33
( Student µ 95% ) 7 2 t ( t
) 2 ( ) 2 d ( ) ← [ 3] σd 2 t 8 2 t ( t ) 2 ( ) ( ) ← [ 4] σ 2 t 2024 9-10 P — 2025-01-06 – p.5/33
2 2 t 1 9 P P 10 H0 HA
k, N, ¯ ¯ x σ2 ( )MSwithin ( )MSbetween MStotal F F 2024 9-10 P — 2025-01-06 – p.6/33
2024 9-10 P — 2025-01-06 – p.7/33
4. t (1) 2 t (2) 2 t (3) 2025
1 2 ( ) 23:59 JST ( ) Waseda Moodle (Q & A ) (1)(2) Discord 2024 9-10 P — 2025-01-06 – p.8/33
. . . . . . 17 14 (1/3( )
) ( ) → 14 ( ) ( ) → 6 → 3 ( ) → 5 ( ) ( OK) 2 t . . . . . . / . . . ( ) 2024 9-10 P — 2025-01-06 – p.9/33
t t ⇒ ( ) A A xA 2 B
B xB 2 df . . . ⇒ t σ z0.05 . . . ⇒ ( ) t 2024 9-10 P — 2025-01-06 – p.10/33
N (1/2) 2 t 2 2 “ ” 1. 1
2 2. 3. - 2 (n − 1) 4. ÷ ÷ t 5. t (n − 1) t t ⇒ . . . 0 ( ) 2024 9-10 P — 2025-01-06 – p.11/33
N (2/2) 2 t 2 2 1 2 1 2
2 “ ” 1. 2 1 2 2. 3. ( -2) 4. t 1÷ 2 t 5. t (n1 + n2 − 2) t t ⇒ 2024 9-10 P — 2025-01-06 – p.12/33
M ( ) [ 2 t ] 1Day 1Day 1Day
⇒ 2024 9-10 P — 2025-01-06 – p.13/33
K ⇒ . . . 2024 9-10 P — 2025-01-06
– p.14/33
2 t d : µd 0 ( 2 ) :
(1) d d, sd , n, df (2) |d| sd n |t| (3) t0.05 (df) < |t| ( ) R > t.test(sample2, sample1, paired=T) 2024 9-10 P — 2025-01-06 – p.15/33
2 t ( ) 10 ( ) ( ) (
) ( ) ( ) d ( ) d, ( ) sd , ( ) n, ( ) df ( ) t ( ) t ( ) d ( ) sd ( ) n ( ) t df 5% ( ) ( ) ( ) ( ) ( ) 2024 9-10 P — 2025-01-06 – p.16/33
2 t xA xB : µA − µB 0 (
2 ) : (1) xA − xB , sp , nA nB , df (2) |xA − xB | sp nA nB |t| (3) t0.05 (df) < |t| ( ) R > t.test(sample2, sample1, var.equal=T) 2024 9-10 P — 2025-01-06 – p.17/33
2 t ( ) ( ) ( ) ( )
( ) ( ) ( ( ) A B ( ) ) ( ) xA − xB , A B ( ) ( ) sp , ( ) nA nB , ( ) df = nA + nB − 2 ( ) t ( ) t ( ) xA − xB ( ) sp ( ) nA ,nB ( ) t df 5% ( ) ( ) ( ) ( ) ( ) ( ) 2024 9-10 P — 2025-01-06 – p.18/33
K 2 t ( ) ⇒ 2 2024 9-10 P
— 2025-01-06 – p.19/33
N ⇒ (σ) ( σ √n ) ( ) p.121
(standard error) (p.121) (sampling distribution) (p.120) (p.120) ( : ) 2024 9-10 P — 2025-01-06 – p.20/33
K ⇒ . . . AI ( ) . .
. ^^; ( ) 2024 9-10 P — 2025-01-06 – p.21/33
H t 2 Student t t 1 sin(α + β)
= sinαcosβ + cosαsinβ . . . ⇒ 2024 9-10 P — 2025-01-06 – p.22/33
U R ChatGPT ⇒ AI ( ) 2024 9-10 P
— 2025-01-06 – p.23/33
9 P P 2024 9-10 P — 2025-01-06 – p.24/33
α β P P H0 ( ) P 0.05 (P
= 0.015) (P = 0.361) 2024 9-10 P — 2025-01-06 – p.25/33
10 H0 HA k, N, ¯ ¯ x σ2 (
) MSwithin ( )MSbetween MStotal ( SStotal dftotal ) F F 2024 9-10 P — 2025-01-06 – p.26/33
(1/3) k (1) : (2) : σ2 ( ) N(µ,
σ2) µ1 = µ2 = · · · = µk N ( ) ¯ ¯ x ¯ ¯ x = k j=1 nj i=1 xji N (j i N ) 2024 9-10 P — 2025-01-06 – p.27/33
(2/3) ( )MSwithin σ2 MSwithin = SSwithin dfwithin = k
j=1 nj i=1 (xji − ¯ xj )2 N − k ( N− ) ( )MSbetween σ2 MSbetween = SSbetween dfbetween = k j=1 nj (¯ xj − ¯ ¯ x)2 k − 1 ( −1 ) ( H0 σ2 ) 2024 9-10 P — 2025-01-06 – p.28/33
(3/3) MStotal MStotal = SStotal dftotal = k j=1 nj
i=1 (xji − ¯ ¯ x)2 N − 1 ( N − 1 ) : SStotal = SSbetween + SSwithin, dftotal = dfbetween + dfwithin F F = MSbetween MSwithin F0.05 (dfbetween, dfwithin ) < F ( H0 ) 2024 9-10 P — 2025-01-06 – p.29/33
U ( p.227) 20 4 “ U.R” ( anova() )
pp.226–227 2024 9-10 P — 2025-01-06 – p.30/33
2024 9-10 P — 2025-01-06 – p.31/33
5. (1) ( ) (2) 2025 1 16 ( )
23:59 JST ( ) Waseda Moodle (Q & A ) (1)(2) Discord 2024 9-10 P — 2025-01-06 – p.32/33
2024 9-10 P — 2025-01-06 – p.33/33