Overview

Dataset statistics

Number of variables9
Number of observations10000
Missing cells0
Missing cells (%)0.0%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory839.8 KiB
Average record size in memory86.0 B

Variable types

DateTime3
Numeric4
Categorical2

Dataset

Description충북농업기술원 농가 경영기록장의 지역별 관측지점의 해당 날씨정보가 제공됩니다.
Author충청북도
URLhttps://www.data.go.kr/data/15050256/fileData.do

Alerts

발표시각 has constant value ""Constant
강수형태 is highly imbalanced (73.1%)Imbalance
강수량 has 9155 (91.5%) zerosZeros

Reproduction

Analysis started2023-12-12 10:00:56.397070
Analysis finished2023-12-12 10:00:59.612147
Duration3.22 seconds
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

Distinct105
Distinct (%)1.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum2017-01-01 00:00:00
Maximum2017-04-15 00:00:00
2023-12-12T19:00:59.708801image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:59.911813image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)

발표시각
Date

CONSTANT 

Distinct1
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum2023-12-12 00:00:00
Maximum2023-12-12 00:00:00
2023-12-12T19:01:00.039348image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:01:00.190405image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=1)

예보지점X좌표
Real number (ℝ)

Distinct20
Distinct (%)0.2%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean73.8987
Minimum66
Maximum86
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-12T19:01:00.319829image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum66
5-th percentile68
Q170
median73
Q377
95-th percentile83
Maximum86
Range20
Interquartile range (IQR)7

Descriptive statistics

Standard deviation4.6356863
Coefficient of variation (CV)0.062730282
Kurtosis-0.27930686
Mean73.8987
Median Absolute Deviation (MAD)3
Skewness0.59513312
Sum738987
Variance21.489587
MonotonicityNot monotonic
2023-12-12T19:01:00.476895image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=20)
ValueCountFrequency (%)
68 940
9.4%
74 910
 
9.1%
73 907
 
9.1%
71 878
 
8.8%
72 817
 
8.2%
77 784
 
7.8%
75 781
 
7.8%
69 655
 
6.6%
70 560
 
5.6%
76 474
 
4.7%
Other values (10) 2294
22.9%
ValueCountFrequency (%)
66 96
 
1.0%
67 286
 
2.9%
68 940
9.4%
69 655
6.6%
70 560
5.6%
71 878
8.8%
72 817
8.2%
73 907
9.1%
74 910
9.1%
75 781
7.8%
ValueCountFrequency (%)
86 127
 
1.3%
84 291
 
2.9%
83 261
 
2.6%
82 189
 
1.9%
81 416
4.2%
80 97
 
1.0%
79 93
 
0.9%
78 438
4.4%
77 784
7.8%
76 474
4.7%

예보지점Y좌표
Real number (ℝ)

Distinct26
Distinct (%)0.3%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean108.9788
Minimum93
Maximum119
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-12T19:01:00.674409image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum93
5-th percentile96
Q1104
median111
Q3114
95-th percentile118
Maximum119
Range26
Interquartile range (IQR)10

Descriptive statistics

Standard deviation6.6274688
Coefficient of variation (CV)0.060814294
Kurtosis-0.7378265
Mean108.9788
Median Absolute Deviation (MAD)4
Skewness-0.57998988
Sum1089788
Variance43.923343
MonotonicityNot monotonic
2023-12-12T19:01:00.824339image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=26)
ValueCountFrequency (%)
114 891
 
8.9%
113 852
 
8.5%
111 777
 
7.8%
115 697
 
7.0%
110 608
 
6.1%
117 605
 
6.0%
107 519
 
5.2%
106 448
 
4.5%
116 439
 
4.4%
118 424
 
4.2%
Other values (16) 3740
37.4%
ValueCountFrequency (%)
93 87
 
0.9%
95 216
2.2%
96 200
2.0%
97 166
1.7%
98 375
3.8%
99 273
2.7%
100 186
1.9%
101 178
1.8%
102 287
2.9%
103 311
3.1%
ValueCountFrequency (%)
119 99
 
1.0%
118 424
4.2%
117 605
6.0%
116 439
4.4%
115 697
7.0%
114 891
8.9%
113 852
8.5%
112 393
3.9%
111 777
7.8%
110 608
6.1%

기온
Real number (ℝ)

Distinct369
Distinct (%)3.7%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean7.80108
Minimum-20
Maximum24.8
Zeros67
Zeros (%)0.7%
Negative1239
Negative (%)12.4%
Memory size166.0 KiB
2023-12-12T19:01:01.044269image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum-20
5-th percentile-2.8
Q13.3
median8.1
Q312.6
95-th percentile18
Maximum24.8
Range44.8
Interquartile range (IQR)9.3

Descriptive statistics

Standard deviation6.4282327
Coefficient of variation (CV)0.8240183
Kurtosis-0.18716341
Mean7.80108
Median Absolute Deviation (MAD)4.6
Skewness-0.16744467
Sum78010.8
Variance41.322175
MonotonicityNot monotonic
2023-12-12T19:01:01.190362image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
8.4 77
 
0.8%
7.4 72
 
0.7%
8.8 71
 
0.7%
7.9 70
 
0.7%
7.6 68
 
0.7%
0.0 67
 
0.7%
13.2 66
 
0.7%
10.4 65
 
0.7%
12.3 64
 
0.6%
7.1 64
 
0.6%
Other values (359) 9316
93.2%
ValueCountFrequency (%)
-20.0 1
< 0.1%
-19.4 1
< 0.1%
-19.0 1
< 0.1%
-18.1 1
< 0.1%
-17.9 2
< 0.1%
-16.7 1
< 0.1%
-16.3 1
< 0.1%
-16.1 2
< 0.1%
-15.7 1
< 0.1%
-15.5 1
< 0.1%
ValueCountFrequency (%)
24.8 2
< 0.1%
24.7 2
< 0.1%
24.6 3
< 0.1%
24.5 3
< 0.1%
24.3 1
 
< 0.1%
24.2 1
 
< 0.1%
24.1 4
< 0.1%
23.9 2
< 0.1%
23.8 1
 
< 0.1%
23.7 4
< 0.1%

강수량
Real number (ℝ)

ZEROS 

Distinct63
Distinct (%)0.6%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean0.09327
Minimum0
Maximum18
Zeros9155
Zeros (%)91.5%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-12T19:01:01.342732image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum0
5-th percentile0
Q10
median0
Q30
95-th percentile0.3
Maximum18
Range18
Interquartile range (IQR)0

Descriptive statistics

Standard deviation0.66561702
Coefficient of variation (CV)7.1364535
Kurtosis333.35604
Mean0.09327
Median Absolute Deviation (MAD)0
Skewness15.997833
Sum932.7
Variance0.44304601
MonotonicityNot monotonic
2023-12-12T19:01:01.524327image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
0.0 9155
91.5%
0.1 184
 
1.8%
0.2 133
 
1.3%
0.5 62
 
0.6%
0.3 61
 
0.6%
0.4 43
 
0.4%
1.0 30
 
0.3%
0.9 26
 
0.3%
2.0 24
 
0.2%
0.6 23
 
0.2%
Other values (53) 259
 
2.6%
ValueCountFrequency (%)
0.0 9155
91.5%
0.1 184
 
1.8%
0.2 133
 
1.3%
0.3 61
 
0.6%
0.4 43
 
0.4%
0.5 62
 
0.6%
0.6 23
 
0.2%
0.7 21
 
0.2%
0.8 15
 
0.1%
0.9 26
 
0.3%
ValueCountFrequency (%)
18.0 2
< 0.1%
17.3 1
 
< 0.1%
16.5 1
 
< 0.1%
14.5 1
 
< 0.1%
14.0 3
< 0.1%
13.5 1
 
< 0.1%
12.5 1
 
< 0.1%
12.2 1
 
< 0.1%
11.5 1
 
< 0.1%
9.5 1
 
< 0.1%

하늘상태
Categorical

Distinct4
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
4
4240 
1
4141 
3
863 
2
756 

Length

Max length1
Median length1
Mean length1
Min length1

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row4
2nd row4
3rd row4
4th row2
5th row4

Common Values

ValueCountFrequency (%)
4 4240
42.4%
1 4141
41.4%
3 863
 
8.6%
2 756
 
7.6%

Length

2023-12-12T19:01:01.668046image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T19:01:01.827138image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
4 4240
42.4%
1 4141
41.4%
3 863
 
8.6%
2 756
 
7.6%

강수형태
Categorical

IMBALANCE 

Distinct3
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
0
9147 
1
 
845
2
 
8

Length

Max length1
Median length1
Mean length1
Min length1

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row0
2nd row0
3rd row0
4th row0
5th row1

Common Values

ValueCountFrequency (%)
0 9147
91.5%
1 845
 
8.5%
2 8
 
0.1%

Length

2023-12-12T19:01:01.978039image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T19:01:02.112249image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
0 9147
91.5%
1 845
 
8.5%
2 8
 
0.1%
Distinct899
Distinct (%)9.0%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum2017-03-09 15:23:00
Maximum2017-04-15 09:50:00
2023-12-12T19:01:02.271895image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:01:02.442289image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)

Interactions

2023-12-12T19:00:58.701541image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.208083image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.714204image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:58.209061image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:58.885459image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.362723image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.853445image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:58.340614image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:59.024401image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.495953image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.954102image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:58.445389image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:59.151856image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:57.608328image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:58.086945image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:00:58.559333image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-12T19:01:02.560795image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
예보지점X좌표예보지점Y좌표기온강수량하늘상태강수형태
예보지점X좌표1.0000.7590.1350.0960.0460.078
예보지점Y좌표0.7591.0000.1070.0340.0000.068
기온0.1350.1071.0000.1170.3450.193
강수량0.0960.0340.1171.0000.1280.425
하늘상태0.0460.0000.3450.1281.0000.257
강수형태0.0780.0680.1930.4250.2571.000
2023-12-12T19:01:03.029639image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
하늘상태강수형태
하늘상태1.0000.246
강수형태0.2461.000
2023-12-12T19:01:03.159803image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
예보지점X좌표예보지점Y좌표기온강수량하늘상태강수형태
예보지점X좌표1.0000.371-0.0780.0230.0260.045
예보지점Y좌표0.3711.000-0.047-0.0240.0000.040
기온-0.078-0.0471.0000.0120.2130.117
강수량0.023-0.0240.0121.0000.0760.280
하늘상태0.0260.0000.2130.0761.0000.246
강수형태0.0450.0400.1170.2800.2461.000

Missing values

2023-12-12T19:00:59.352481image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-12T19:00:59.539750image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

발표일자발표시각예보지점X좌표예보지점Y좌표기온강수량하늘상태강수형태등록일시
621412017-04-010:00751174.60.0402017-04-01 23:50
421862017-03-240:0074977.60.0402017-03-24 22:50
491832017-03-270:007010910.20.0402017-03-27 18:50
927482017-04-140:006811317.20.0202017-04-14 09:50
461142017-03-260:00711178.90.1412017-03-26 12:50
307512017-03-200:0077981.00.0102017-03-20 07:50
591612017-03-310:0070998.30.2412017-03-31 18:50
839862017-04-100:00821199.90.0402017-04-10 20:50
470412017-03-260:00711175.50.0402017-03-26 21:50
613872017-04-010:006810910.40.0302017-04-01 16:50
발표일자발표시각예보지점X좌표예보지점Y좌표기온강수량하늘상태강수형태등록일시
382272017-03-230:0075115-0.20.0402017-03-23 07:50
137332017-03-130:00751120.00.0102017-03-13 08:50
383082017-03-230:00731070.60.0402017-03-23 08:50
610452017-04-010:00861177.30.0402017-04-01 13:50
20532017-02-050:00841150.20.0102017-03-09 16:56
400802017-03-240:00751150.40.0102017-03-24 01:50
643752017-04-020:00701097.00.0102017-04-02 21:50
158462017-03-140:0074110-3.80.0102017-03-14 06:50
67642017-03-100:00701005.30.0102017-03-10 12:47
249602017-03-170:00761154.60.0102017-03-17 22:50