Overview

Dataset statistics

Number of variables7
Number of observations10000
Missing cells0
Missing cells (%)0.0%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory644.5 KiB
Average record size in memory66.0 B

Variable types

Numeric2
DateTime2
Boolean1
Text1
Categorical1

Dataset

Description충청북도 농업기술원 농가경영기록장(농가의 소득을 증진시킬 수 있는 회원전용 농가경영 관리 프로그램)의 수입지출관련 이용자 접속기록, 거래, 거래처 등의 관리시스템으로 파일일련번호, 등록일시, 수정일시, 상태, 파일명, 파일크기, Content-Type 등을 제공합니다.
Author충청북도
URLhttps://www.data.go.kr/data/15050315/fileData.do

Alerts

상태 has constant value ""Constant
파일일련번호 is highly overall correlated with Content-TypeHigh correlation
Content-Type is highly overall correlated with 파일일련번호High correlation
파일일련번호 has unique valuesUnique

Reproduction

Analysis started2023-12-12 19:26:38.941934
Analysis finished2023-12-12 19:26:40.364687
Duration1.42 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

파일일련번호
Real number (ℝ)

HIGH CORRELATION  UNIQUE 

Distinct10000
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean37834.507
Minimum4
Maximum76878
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-13T04:26:40.442182image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum4
5-th percentile3810.85
Q118556
median37429.5
Q357008.25
95-th percentile72755.6
Maximum76878
Range76874
Interquartile range (IQR)38452.25

Descriptive statistics

Standard deviation22132.046
Coefficient of variation (CV)0.58496983
Kurtosis-1.204501
Mean37834.507
Median Absolute Deviation (MAD)19248.5
Skewness0.037599813
Sum3.7834508 × 108
Variance4.8982744 × 108
MonotonicityNot monotonic
2023-12-13T04:26:40.593143image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
66090 1
 
< 0.1%
37071 1
 
< 0.1%
64589 1
 
< 0.1%
68641 1
 
< 0.1%
25376 1
 
< 0.1%
26898 1
 
< 0.1%
70199 1
 
< 0.1%
71793 1
 
< 0.1%
6843 1
 
< 0.1%
45013 1
 
< 0.1%
Other values (9990) 9990
99.9%
ValueCountFrequency (%)
4 1
< 0.1%
8 1
< 0.1%
12 1
< 0.1%
13 1
< 0.1%
24 1
< 0.1%
25 1
< 0.1%
36 1
< 0.1%
38 1
< 0.1%
51 1
< 0.1%
81 1
< 0.1%
ValueCountFrequency (%)
76878 1
< 0.1%
76875 1
< 0.1%
76867 1
< 0.1%
76861 1
< 0.1%
76836 1
< 0.1%
76827 1
< 0.1%
76825 1
< 0.1%
76823 1
< 0.1%
76822 1
< 0.1%
76815 1
< 0.1%
Distinct4815
Distinct (%)48.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum1900-01-01 00:00:00
Maximum2019-11-11 00:26:00
2023-12-13T04:26:40.746321image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-13T04:26:40.887188image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
Distinct4803
Distinct (%)48.0%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum2017-03-09 16:49:00
Maximum2019-11-11 00:26:00
2023-12-13T04:26:41.049927image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-13T04:26:41.176449image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)

상태
Boolean

CONSTANT 

Distinct1
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size87.9 KiB
False
10000 
ValueCountFrequency (%)
False 10000
100.0%
2023-12-13T04:26:41.273402image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Distinct8960
Distinct (%)89.6%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
2023-12-13T04:26:41.482124image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length60
Median length56
Mean length25.2439
Min length7

Characters and Unicode

Total characters252439
Distinct characters124
Distinct categories10 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique8946 ?
Unique (%)89.5%

Sample

1st row20190410192847_1.jpg
2nd row20181219142454_1.jpg
3rd row128_20131207211231_1_temp.jpg
4th row2785_66_20140809200834_1_temp.jpg
5th row20181207104330_4.jpg
ValueCountFrequency (%)
temp1.jpg 562
 
5.5%
temp2.jpg 306
 
3.0%
temp3.jpg 159
 
1.6%
smr13245 17
 
0.2%
juhi1231 17
 
0.2%
lee190 11
 
0.1%
jck6218 11
 
0.1%
사진 11
 
0.1%
www.bkc321 8
 
0.1%
nsgsapkh 8
 
0.1%
Other values (9020) 9104
89.1%
2023-12-13T04:26:41.893448image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
1 35770
14.2%
0 33568
13.3%
2 28205
11.2%
_ 23176
 
9.2%
p 15855
 
6.3%
3 12539
 
5.0%
5 10822
 
4.3%
4 10741
 
4.3%
. 10095
 
4.0%
j 10056
 
4.0%
Other values (114) 61612
24.4%

Most occurring categories

ValueCountFrequency (%)
Decimal Number 162900
64.5%
Lowercase Letter 55356
 
21.9%
Connector Punctuation 23176
 
9.2%
Other Punctuation 10170
 
4.0%
Space Separator 598
 
0.2%
Uppercase Letter 112
 
< 0.1%
Other Letter 102
 
< 0.1%
Close Punctuation 9
 
< 0.1%
Open Punctuation 9
 
< 0.1%
Dash Punctuation 7
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
12
 
11.8%
11
 
10.8%
3
 
2.9%
3
 
2.9%
3
 
2.9%
2
 
2.0%
2
 
2.0%
2
 
2.0%
2
 
2.0%
2
 
2.0%
Other values (59) 60
58.8%
Lowercase Letter
ValueCountFrequency (%)
p 15855
28.6%
j 10056
18.2%
g 10045
18.1%
m 6000
 
10.8%
e 5985
 
10.8%
t 5890
 
10.6%
a 161
 
0.3%
o 153
 
0.3%
b 145
 
0.3%
c 136
 
0.2%
Other values (14) 930
 
1.7%
Uppercase Letter
ValueCountFrequency (%)
G 30
26.8%
P 16
14.3%
J 16
14.3%
M 14
12.5%
I 14
12.5%
C 7
 
6.2%
D 4
 
3.6%
H 3
 
2.7%
R 3
 
2.7%
T 2
 
1.8%
Other values (2) 3
 
2.7%
Decimal Number
ValueCountFrequency (%)
1 35770
22.0%
0 33568
20.6%
2 28205
17.3%
3 12539
 
7.7%
5 10822
 
6.6%
4 10741
 
6.6%
6 8392
 
5.2%
8 7901
 
4.9%
9 7896
 
4.8%
7 7066
 
4.3%
Other Punctuation
ValueCountFrequency (%)
. 10095
99.3%
@ 75
 
0.7%
Close Punctuation
ValueCountFrequency (%)
) 8
88.9%
] 1
 
11.1%
Open Punctuation
ValueCountFrequency (%)
( 8
88.9%
[ 1
 
11.1%
Connector Punctuation
ValueCountFrequency (%)
_ 23176
100.0%
Space Separator
ValueCountFrequency (%)
598
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 7
100.0%

Most occurring scripts

ValueCountFrequency (%)
Common 196869
78.0%
Latin 55468
 
22.0%
Hangul 102
 
< 0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
12
 
11.8%
11
 
10.8%
3
 
2.9%
3
 
2.9%
3
 
2.9%
2
 
2.0%
2
 
2.0%
2
 
2.0%
2
 
2.0%
2
 
2.0%
Other values (59) 60
58.8%
Latin
ValueCountFrequency (%)
p 15855
28.6%
j 10056
18.1%
g 10045
18.1%
m 6000
 
10.8%
e 5985
 
10.8%
t 5890
 
10.6%
a 161
 
0.3%
o 153
 
0.3%
b 145
 
0.3%
c 136
 
0.2%
Other values (26) 1042
 
1.9%
Common
ValueCountFrequency (%)
1 35770
18.2%
0 33568
17.1%
2 28205
14.3%
_ 23176
11.8%
3 12539
 
6.4%
5 10822
 
5.5%
4 10741
 
5.5%
. 10095
 
5.1%
6 8392
 
4.3%
8 7901
 
4.0%
Other values (9) 15660
8.0%

Most occurring blocks

ValueCountFrequency (%)
ASCII 252337
> 99.9%
Hangul 102
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
1 35770
14.2%
0 33568
13.3%
2 28205
11.2%
_ 23176
 
9.2%
p 15855
 
6.3%
3 12539
 
5.0%
5 10822
 
4.3%
4 10741
 
4.3%
. 10095
 
4.0%
j 10056
 
4.0%
Other values (45) 61510
24.4%
Hangul
ValueCountFrequency (%)
12
 
11.8%
11
 
10.8%
3
 
2.9%
3
 
2.9%
3
 
2.9%
2
 
2.0%
2
 
2.0%
2
 
2.0%
2
 
2.0%
2
 
2.0%
Other values (59) 60
58.8%

파일크기
Real number (ℝ)

Distinct9798
Distinct (%)98.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean1160590.5
Minimum8454
Maximum12190875
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-13T04:26:42.046838image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum8454
5-th percentile313694.1
Q1528842.25
median752935
Q31656674
95-th percentile3124880.9
Maximum12190875
Range12182421
Interquartile range (IQR)1127831.8

Descriptive statistics

Standard deviation957124.3
Coefficient of variation (CV)0.82468737
Kurtosis5.0586968
Mean1160590.5
Median Absolute Deviation (MAD)284395
Skewness1.788728
Sum1.1605905 × 1010
Variance9.1608693 × 1011
MonotonicityNot monotonic
2023-12-13T04:26:42.217989image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
570744 5
 
0.1%
1800943 4
 
< 0.1%
1433919 4
 
< 0.1%
519865 3
 
< 0.1%
486906 3
 
< 0.1%
1253183 3
 
< 0.1%
1034948 3
 
< 0.1%
2131990 3
 
< 0.1%
2406375 3
 
< 0.1%
1713311 3
 
< 0.1%
Other values (9788) 9966
99.7%
ValueCountFrequency (%)
8454 1
< 0.1%
8719 1
< 0.1%
21012 1
< 0.1%
23393 1
< 0.1%
25859 1
< 0.1%
32611 1
< 0.1%
32654 1
< 0.1%
35402 1
< 0.1%
35452 1
< 0.1%
36731 1
< 0.1%
ValueCountFrequency (%)
12190875 1
< 0.1%
10073585 1
< 0.1%
8131857 1
< 0.1%
7989830 1
< 0.1%
7491652 1
< 0.1%
7393218 1
< 0.1%
7030760 1
< 0.1%
6825378 1
< 0.1%
6552617 1
< 0.1%
6208691 1
< 0.1%

Content-Type
Categorical

HIGH CORRELATION 

Distinct6
Distinct (%)0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
image/jpeg
4974 
application/octet-stream
3856 
jpeg
1167 
image/bmp
 
1
image/svg+xml
 
1

Length

Max length24
Median length15
Mean length14.6989
Min length4

Unique

Unique3 ?
Unique (%)< 0.1%

Sample

1st rowapplication/octet-stream
2nd rowapplication/octet-stream
3rd rowimage/jpeg
4th rowimage/jpeg
5th rowapplication/octet-stream

Common Values

ValueCountFrequency (%)
image/jpeg 4974
49.7%
application/octet-stream 3856
38.6%
jpeg 1167
 
11.7%
image/bmp 1
 
< 0.1%
image/svg+xml 1
 
< 0.1%
application/pdf 1
 
< 0.1%

Length

2023-12-13T04:26:42.366848image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-13T04:26:42.516935image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
image/jpeg 4974
49.7%
application/octet-stream 3856
38.6%
jpeg 1167
 
11.7%
image/bmp 1
 
< 0.1%
image/svg+xml 1
 
< 0.1%
application/pdf 1
 
< 0.1%

Interactions

2023-12-13T04:26:39.963708image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-13T04:26:39.737632image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-13T04:26:40.078029image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-13T04:26:39.850643image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-13T04:26:42.599921image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
파일일련번호파일크기Content-Type
파일일련번호1.0000.5600.762
파일크기0.5601.0000.610
Content-Type0.7620.6101.000
2023-12-13T04:26:42.691981image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
파일일련번호파일크기Content-Type
파일일련번호1.0000.4230.536
파일크기0.4231.0000.356
Content-Type0.5360.3561.000

Missing values

2023-12-13T04:26:40.200718image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-13T04:26:40.310935image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

파일일련번호등록일시수정일시상태파일명파일크기Content-Type
64515660902019-04-10 19:182019-04-10 19:18N20190410192847_1.jpg4099707application/octet-stream
60133615172018-12-19 14:252018-12-19 14:25N20181219142454_1.jpg1676329application/octet-stream
33163332651900-01-01 00:002017-03-09 22:53N128_20131207211231_1_temp.jpg377009image/jpeg
847984801900-01-01 00:002017-03-09 22:23N2785_66_20140809200834_1_temp.jpg550273image/jpeg
59774611472018-12-07 10:382018-12-07 10:38N20181207104330_4.jpg2339854application/octet-stream
41567419892017-06-14 21:522017-06-14 21:52Ntemp1.jpg287901application/octet-stream
43096435982017-07-24 15:192017-07-24 15:19Ntemp2.jpg236540jpeg
23377234151900-01-01 00:002017-03-09 22:41N2292_66_20160616070624_2_temp.jpg432123image/jpeg
41836422672017-06-19 21:592017-06-19 21:59Ntemp2.jpg444455jpeg
43831443702017-08-16 20:172017-08-16 20:17N20170816202635_2.jpg431283jpeg
파일일련번호등록일시수정일시상태파일명파일크기Content-Type
51495524522018-04-22 21:422018-04-22 21:42N20180422214350_2.jpg955700application/octet-stream
17467174761900-01-01 00:002017-03-09 22:34N2458_20_20151117231111_1_temp.jpg619105image/jpeg
846484651900-01-01 00:002017-03-09 22:23N1016_140_20140808160855_1_temp.jpg1072227image/jpeg
30243303081900-01-01 00:002017-03-09 22:50N25_1_20150520220521_1_temp.jpg681503image/jpeg
59746611192018-12-06 15:262018-12-06 15:26N20181206153135_1.jpg1789948application/octet-stream
21642216761900-01-01 00:002017-03-09 22:39N74_81_20160503200520_2_temp.jpg602042image/jpeg
69477711832019-06-28 14:492019-06-28 14:49N20190628144905_1.jpg2185319application/octet-stream
728772881900-01-01 00:002017-03-09 22:22N16_23_20140607110600_1_temp.jpg900915image/jpeg
50465514012018-04-05 17:212018-04-05 17:21N20180405172638_1.jpg2603571application/octet-stream
59287606512018-11-21 20:172018-11-21 20:17N20181121202013_1.jpg1687260application/octet-stream