Overview

Dataset statistics

Number of variables5
Number of observations10000
Missing cells10001
Missing cells (%)20.0%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory478.5 KiB
Average record size in memory49.0 B

Variable types

Numeric1
Text3
Categorical1

Dataset

Description경기도 평택시 지역화폐 가맹점 현황에 대한 데이터로 상호명, 소재지도로명주소, 소재지지번주소 등의 정보를 제공합니다.
Author경기도 평택시
URLhttps://www.data.go.kr/data/3065043/fileData.do

Alerts

기준일자 has constant value ""Constant
소재지지번주소 has 9996 (> 99.9%) missing valuesMissing
연번 has unique valuesUnique

Reproduction

Analysis started2023-12-13 00:32:15.223707
Analysis finished2023-12-13 00:32:16.507408
Duration1.28 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

연번
Real number (ℝ)

UNIQUE 

Distinct10000
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean5898.268
Minimum1
Maximum11801
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-13T09:32:16.560828image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1
5-th percentile601.95
Q12955.75
median5907.5
Q38828.25
95-th percentile11208.05
Maximum11801
Range11800
Interquartile range (IQR)5872.5

Descriptive statistics

Standard deviation3400.6095
Coefficient of variation (CV)0.57654375
Kurtosis-1.1995714
Mean5898.268
Median Absolute Deviation (MAD)2937.5
Skewness0.0012172858
Sum58982680
Variance11564145
MonotonicityNot monotonic
2023-12-13T09:32:16.663420image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
3844 1
 
< 0.1%
9928 1
 
< 0.1%
11631 1
 
< 0.1%
1777 1
 
< 0.1%
1977 1
 
< 0.1%
238 1
 
< 0.1%
1402 1
 
< 0.1%
2443 1
 
< 0.1%
3175 1
 
< 0.1%
9529 1
 
< 0.1%
Other values (9990) 9990
99.9%
ValueCountFrequency (%)
1 1
< 0.1%
2 1
< 0.1%
3 1
< 0.1%
4 1
< 0.1%
5 1
< 0.1%
7 1
< 0.1%
8 1
< 0.1%
9 1
< 0.1%
11 1
< 0.1%
12 1
< 0.1%
ValueCountFrequency (%)
11801 1
< 0.1%
11800 1
< 0.1%
11799 1
< 0.1%
11798 1
< 0.1%
11797 1
< 0.1%
11795 1
< 0.1%
11794 1
< 0.1%
11793 1
< 0.1%
11792 1
< 0.1%
11791 1
< 0.1%
Distinct9811
Distinct (%)98.1%
Missing1
Missing (%)< 0.1%
Memory size156.2 KiB
2023-12-13T09:32:16.934180image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length46
Median length30
Mean length6.7589759
Min length1

Characters and Unicode

Total characters67583
Distinct characters1060
Distinct categories13 ?
Distinct scripts4 ?
Distinct blocks6 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique9646 ?
Unique (%)96.5%

Sample

1st row중앙당구장
2nd row머슬짐
3rd row김치찌개 미화식당
4th row김밥천국용이점
5th row안중 EED PRO
ValueCountFrequency (%)
평택점 86
 
0.7%
세븐일레븐 57
 
0.5%
주식회사 52
 
0.4%
안중점 43
 
0.4%
송탄점 32
 
0.3%
지에스(gs)25 31
 
0.3%
씨유 27
 
0.2%
소사벌점 26
 
0.2%
청북점 23
 
0.2%
평택안중점 23
 
0.2%
Other values (10321) 11625
96.7%
2023-12-13T09:32:17.327277image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
2049
 
3.0%
1973
 
2.9%
1537
 
2.3%
1187
 
1.8%
1167
 
1.7%
1101
 
1.6%
907
 
1.3%
901
 
1.3%
844
 
1.2%
809
 
1.2%
Other values (1050) 55108
81.5%

Most occurring categories

ValueCountFrequency (%)
Other Letter 61286
90.7%
Space Separator 2049
 
3.0%
Uppercase Letter 1360
 
2.0%
Decimal Number 822
 
1.2%
Close Punctuation 617
 
0.9%
Open Punctuation 614
 
0.9%
Lowercase Letter 544
 
0.8%
Other Punctuation 171
 
0.3%
Other Symbol 100
 
0.1%
Dash Punctuation 14
 
< 0.1%
Other values (3) 6
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
1973
 
3.2%
1537
 
2.5%
1187
 
1.9%
1167
 
1.9%
1101
 
1.8%
907
 
1.5%
901
 
1.5%
844
 
1.4%
809
 
1.3%
635
 
1.0%
Other values (971) 50225
82.0%
Uppercase Letter
ValueCountFrequency (%)
S 166
 
12.2%
C 144
 
10.6%
G 126
 
9.3%
E 74
 
5.4%
U 72
 
5.3%
A 72
 
5.3%
B 69
 
5.1%
M 67
 
4.9%
O 62
 
4.6%
P 49
 
3.6%
Other values (16) 459
33.8%
Lowercase Letter
ValueCountFrequency (%)
a 63
11.6%
i 57
10.5%
e 54
 
9.9%
o 51
 
9.4%
n 35
 
6.4%
r 32
 
5.9%
t 32
 
5.9%
s 31
 
5.7%
l 29
 
5.3%
m 23
 
4.2%
Other values (14) 137
25.2%
Decimal Number
ValueCountFrequency (%)
2 229
27.9%
1 150
18.2%
5 141
17.2%
4 77
 
9.4%
0 56
 
6.8%
3 56
 
6.8%
9 43
 
5.2%
8 34
 
4.1%
6 19
 
2.3%
7 17
 
2.1%
Other Punctuation
ValueCountFrequency (%)
& 66
38.6%
. 47
27.5%
, 31
18.1%
' 10
 
5.8%
# 5
 
2.9%
/ 3
 
1.8%
! 3
 
1.8%
: 3
 
1.8%
· 2
 
1.2%
? 1
 
0.6%
Other Symbol
ValueCountFrequency (%)
99
99.0%
1
 
1.0%
Space Separator
ValueCountFrequency (%)
2049
100.0%
Close Punctuation
ValueCountFrequency (%)
) 617
100.0%
Open Punctuation
ValueCountFrequency (%)
( 614
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 14
100.0%
Connector Punctuation
ValueCountFrequency (%)
_ 3
100.0%
Math Symbol
ValueCountFrequency (%)
+ 2
100.0%
Modifier Symbol
ValueCountFrequency (%)
` 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 61382
90.8%
Common 4294
 
6.4%
Latin 1904
 
2.8%
Han 3
 
< 0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
1973
 
3.2%
1537
 
2.5%
1187
 
1.9%
1167
 
1.9%
1101
 
1.8%
907
 
1.5%
901
 
1.5%
844
 
1.4%
809
 
1.3%
635
 
1.0%
Other values (969) 50321
82.0%
Latin
ValueCountFrequency (%)
S 166
 
8.7%
C 144
 
7.6%
G 126
 
6.6%
E 74
 
3.9%
U 72
 
3.8%
A 72
 
3.8%
B 69
 
3.6%
M 67
 
3.5%
a 63
 
3.3%
O 62
 
3.3%
Other values (40) 989
51.9%
Common
ValueCountFrequency (%)
2049
47.7%
) 617
 
14.4%
( 614
 
14.3%
2 229
 
5.3%
1 150
 
3.5%
5 141
 
3.3%
4 77
 
1.8%
& 66
 
1.5%
0 56
 
1.3%
3 56
 
1.3%
Other values (18) 239
 
5.6%
Han
ValueCountFrequency (%)
1
33.3%
1
33.3%
1
33.3%

Most occurring blocks

ValueCountFrequency (%)
Hangul 61283
90.7%
ASCII 6195
 
9.2%
None 101
 
0.1%
CJK 2
 
< 0.1%
CJK Compat Ideographs 1
 
< 0.1%
Misc Symbols 1
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
2049
33.1%
) 617
 
10.0%
( 614
 
9.9%
2 229
 
3.7%
S 166
 
2.7%
1 150
 
2.4%
C 144
 
2.3%
5 141
 
2.3%
G 126
 
2.0%
4 77
 
1.2%
Other values (66) 1882
30.4%
Hangul
ValueCountFrequency (%)
1973
 
3.2%
1537
 
2.5%
1187
 
1.9%
1167
 
1.9%
1101
 
1.8%
907
 
1.5%
901
 
1.5%
844
 
1.4%
809
 
1.3%
635
 
1.0%
Other values (968) 50222
82.0%
None
ValueCountFrequency (%)
99
98.0%
· 2
 
2.0%
CJK Compat Ideographs
ValueCountFrequency (%)
1
100.0%
Misc Symbols
ValueCountFrequency (%)
1
100.0%
CJK
ValueCountFrequency (%)
1
50.0%
1
50.0%
Distinct9257
Distinct (%)92.6%
Missing4
Missing (%)< 0.1%
Memory size156.2 KiB
2023-12-13T09:32:17.571770image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length63
Median length51
Mean length23.367247
Min length13

Characters and Unicode

Total characters233579
Distinct characters487
Distinct categories12 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique8677 ?
Unique (%)86.8%

Sample

1st row경기도 평택시 세교상가1길 8
2nd row경기도 평택시 안중읍 안현로서6길 7, 4층(대건빌딩)
3rd row경기도 평택시 청북읍 청원로 382-1, 1층
4th row경기도 평택시 용죽3로 21, 1층
5th row경기도 평택시 현덕면 송담1로 50, 101호
ValueCountFrequency (%)
평택시 9999
 
19.5%
경기도 9996
 
19.5%
1층 1894
 
3.7%
안중읍 929
 
1.8%
포승읍 411
 
0.8%
팽성읍 374
 
0.7%
청북읍 354
 
0.7%
2층 344
 
0.7%
고덕면 326
 
0.6%
101호 311
 
0.6%
Other values (5653) 26208
51.2%
2023-12-13T09:32:17.933290image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
41297
17.7%
1 13017
 
5.6%
11570
 
5.0%
11196
 
4.8%
10493
 
4.5%
10276
 
4.4%
10258
 
4.4%
10058
 
4.3%
8393
 
3.6%
2 7566
 
3.2%
Other values (477) 99455
42.6%

Most occurring categories

ValueCountFrequency (%)
Other Letter 130713
56.0%
Decimal Number 48204
 
20.6%
Space Separator 41297
 
17.7%
Other Punctuation 5913
 
2.5%
Open Punctuation 2346
 
1.0%
Close Punctuation 2341
 
1.0%
Dash Punctuation 2256
 
1.0%
Uppercase Letter 391
 
0.2%
Math Symbol 91
 
< 0.1%
Lowercase Letter 25
 
< 0.1%
Other values (2) 2
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
11570
 
8.9%
11196
 
8.6%
10493
 
8.0%
10276
 
7.9%
10258
 
7.8%
10058
 
7.7%
8393
 
6.4%
4369
 
3.3%
3128
 
2.4%
3103
 
2.4%
Other values (416) 47869
36.6%
Uppercase Letter
ValueCountFrequency (%)
A 83
21.2%
B 61
15.6%
S 31
 
7.9%
K 25
 
6.4%
C 24
 
6.1%
E 20
 
5.1%
D 14
 
3.6%
I 14
 
3.6%
J 14
 
3.6%
H 12
 
3.1%
Other values (15) 93
23.8%
Lowercase Letter
ValueCountFrequency (%)
a 5
20.0%
e 4
16.0%
k 2
 
8.0%
z 2
 
8.0%
l 2
 
8.0%
s 2
 
8.0%
b 2
 
8.0%
g 1
 
4.0%
n 1
 
4.0%
o 1
 
4.0%
Other values (3) 3
12.0%
Decimal Number
ValueCountFrequency (%)
1 13017
27.0%
2 7566
15.7%
0 5180
 
10.7%
3 5133
 
10.6%
5 3960
 
8.2%
4 3559
 
7.4%
6 2995
 
6.2%
7 2404
 
5.0%
9 2200
 
4.6%
8 2190
 
4.5%
Other Punctuation
ValueCountFrequency (%)
, 5884
99.5%
. 18
 
0.3%
@ 5
 
0.1%
/ 3
 
0.1%
& 2
 
< 0.1%
: 1
 
< 0.1%
Space Separator
ValueCountFrequency (%)
41297
100.0%
Open Punctuation
ValueCountFrequency (%)
( 2346
100.0%
Close Punctuation
ValueCountFrequency (%)
) 2341
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 2256
100.0%
Math Symbol
ValueCountFrequency (%)
~ 91
100.0%
Modifier Symbol
ValueCountFrequency (%)
` 1
100.0%
Letter Number
ValueCountFrequency (%)
1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 130713
56.0%
Common 102449
43.9%
Latin 417
 
0.2%

Most frequent character per script

Hangul
ValueCountFrequency (%)
11570
 
8.9%
11196
 
8.6%
10493
 
8.0%
10276
 
7.9%
10258
 
7.8%
10058
 
7.7%
8393
 
6.4%
4369
 
3.3%
3128
 
2.4%
3103
 
2.4%
Other values (416) 47869
36.6%
Latin
ValueCountFrequency (%)
A 83
19.9%
B 61
14.6%
S 31
 
7.4%
K 25
 
6.0%
C 24
 
5.8%
E 20
 
4.8%
D 14
 
3.4%
I 14
 
3.4%
J 14
 
3.4%
H 12
 
2.9%
Other values (29) 119
28.5%
Common
ValueCountFrequency (%)
41297
40.3%
1 13017
 
12.7%
2 7566
 
7.4%
, 5884
 
5.7%
0 5180
 
5.1%
3 5133
 
5.0%
5 3960
 
3.9%
4 3559
 
3.5%
6 2995
 
2.9%
7 2404
 
2.3%
Other values (12) 11454
 
11.2%

Most occurring blocks

ValueCountFrequency (%)
Hangul 130713
56.0%
ASCII 102865
44.0%
Number Forms 1
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
41297
40.1%
1 13017
 
12.7%
2 7566
 
7.4%
, 5884
 
5.7%
0 5180
 
5.0%
3 5133
 
5.0%
5 3960
 
3.8%
4 3559
 
3.5%
6 2995
 
2.9%
7 2404
 
2.3%
Other values (50) 11870
 
11.5%
Hangul
ValueCountFrequency (%)
11570
 
8.9%
11196
 
8.6%
10493
 
8.0%
10276
 
7.9%
10258
 
7.8%
10058
 
7.7%
8393
 
6.4%
4369
 
3.3%
3128
 
2.4%
3103
 
2.4%
Other values (416) 47869
36.6%
Number Forms
ValueCountFrequency (%)
1
100.0%

소재지지번주소
Text

MISSING 

Distinct4
Distinct (%)100.0%
Missing9996
Missing (%)> 99.9%
Memory size156.2 KiB
2023-12-13T09:32:18.063859image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length21
Median length17.5
Mean length17.75
Min length15

Characters and Unicode

Total characters71
Distinct characters31
Distinct categories4 ?
Distinct scripts2 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique4 ?
Unique (%)100.0%

Sample

1st row경기도 평택시 통복동65-1
2nd row경기도 평택시 안중읍 금곡리 202-6
3rd row경기도 평택시 소사동403-1
4th row경기도 평택시 진위면 야막리 443
ValueCountFrequency (%)
경기도 4
25.0%
평택시 4
25.0%
통복동65-1 1
 
6.2%
안중읍 1
 
6.2%
금곡리 1
 
6.2%
202-6 1
 
6.2%
소사동403-1 1
 
6.2%
진위면 1
 
6.2%
야막리 1
 
6.2%
443 1
 
6.2%
2023-12-13T09:32:18.304371image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
12
16.9%
4
 
5.6%
4
 
5.6%
4
 
5.6%
4
 
5.6%
4
 
5.6%
4
 
5.6%
4 3
 
4.2%
- 3
 
4.2%
1 2
 
2.8%
Other values (21) 27
38.0%

Most occurring categories

ValueCountFrequency (%)
Other Letter 42
59.2%
Decimal Number 14
 
19.7%
Space Separator 12
 
16.9%
Dash Punctuation 3
 
4.2%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
2
 
4.8%
2
 
4.8%
1
 
2.4%
1
 
2.4%
Other values (12) 12
28.6%
Decimal Number
ValueCountFrequency (%)
4 3
21.4%
1 2
14.3%
2 2
14.3%
3 2
14.3%
0 2
14.3%
6 2
14.3%
5 1
 
7.1%
Space Separator
ValueCountFrequency (%)
12
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 3
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 42
59.2%
Common 29
40.8%

Most frequent character per script

Hangul
ValueCountFrequency (%)
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
2
 
4.8%
2
 
4.8%
1
 
2.4%
1
 
2.4%
Other values (12) 12
28.6%
Common
ValueCountFrequency (%)
12
41.4%
4 3
 
10.3%
- 3
 
10.3%
1 2
 
6.9%
2 2
 
6.9%
3 2
 
6.9%
0 2
 
6.9%
6 2
 
6.9%
5 1
 
3.4%

Most occurring blocks

ValueCountFrequency (%)
Hangul 42
59.2%
ASCII 29
40.8%

Most frequent character per block

ASCII
ValueCountFrequency (%)
12
41.4%
4 3
 
10.3%
- 3
 
10.3%
1 2
 
6.9%
2 2
 
6.9%
3 2
 
6.9%
0 2
 
6.9%
6 2
 
6.9%
5 1
 
3.4%
Hangul
ValueCountFrequency (%)
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
4
 
9.5%
2
 
4.8%
2
 
4.8%
1
 
2.4%
1
 
2.4%
Other values (12) 12
28.6%

기준일자
Categorical

CONSTANT 

Distinct1
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
2021-07-07
10000 

Length

Max length10
Median length10
Mean length10
Min length10

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row2021-07-07
2nd row2021-07-07
3rd row2021-07-07
4th row2021-07-07
5th row2021-07-07

Common Values

ValueCountFrequency (%)
2021-07-07 10000
100.0%

Length

2023-12-13T09:32:18.400557image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-13T09:32:18.465067image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
2021-07-07 10000
100.0%

Interactions

2023-12-13T09:32:16.215181image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-13T09:32:18.502473image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
연번소재지지번주소
연번1.0001.000
소재지지번주소1.0001.000

Missing values

2023-12-13T09:32:16.306885image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-13T09:32:16.383416image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.
2023-12-13T09:32:16.464973image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
The correlation heatmap measures nullity correlation: how strongly the presence or absence of one variable affects the presence of another.

Sample

연번상호명소재지도로명주소소재지지번주소기준일자
38433844중앙당구장경기도 평택시 세교상가1길 8<NA>2021-07-07
65496550머슬짐경기도 평택시 안중읍 안현로서6길 7, 4층(대건빌딩)<NA>2021-07-07
1003410035김치찌개 미화식당경기도 평택시 청북읍 청원로 382-1, 1층<NA>2021-07-07
81188119김밥천국용이점경기도 평택시 용죽3로 21, 1층<NA>2021-07-07
1173511736안중 EED PRO경기도 평택시 현덕면 송담1로 50, 101호<NA>2021-07-07
1168411685제이알모터스경기도 평택시 포승읍 내기새싹길 12-11<NA>2021-07-07
11931194상록수뚝배기경기도 평택시 평택5로 228, 상가-107(비전동, 아파트)<NA>2021-07-07
686687나은떡집경기도 평택시 평택2로 108, 1층<NA>2021-07-07
89288929다다공인중개사사무소경기도 평택시 점촌로32번길 55, 1층<NA>2021-07-07
1046810469무지개커피숍경기도 평택시 통복로 33<NA>2021-07-07
연번상호명소재지도로명주소소재지지번주소기준일자
12371238명약국경기도 평택시 비전5로 11, 403호(비전동)<NA>2021-07-07
44244425돈벼락우리가쏠께경기도 평택시 지산로 82, B동(지산동)<NA>2021-07-07
54565457한솥도시락평택여중사거리점경기도 평택시 중앙로 158<NA>2021-07-07
902903생각하는 수학나무 센트럴자이3단지경기도 평택시 상서재로 55, 304동 101호(평택센트럴자이3단지)<NA>2021-07-07
1029610297아바이순대경기도 평택시 통복시장2로9번길 10<NA>2021-07-07
29052906사거리덕성식당경기도 평택시 국제로 38<NA>2021-07-07
1012510126굽네치킨청북신도시점경기도 평택시 청북읍 안청로 312-19, 102호<NA>2021-07-07
25642565천사열쇠도장경기도 평택시 평남로 653<NA>2021-07-07
1175011751풍천장어경기도 평택시 현덕면 서동대로 266<NA>2021-07-07
55415542미샤경기도 평택시 중앙2로 8<NA>2021-07-07