Overview

Dataset statistics

Number of variables5
Number of observations50
Missing cells0
Missing cells (%)0.0%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory2.1 KiB
Average record size in memory43.6 B

Variable types

Numeric1
Categorical1
Text3

Dataset

Description화성시 관광편의시설업에 관한 데이터로 관광펜션업의 연번, 업종, 상호, 소재지지번주소, 소재지도로명주소에 대한 데이터를 포함하고 있습니다.
Author경기도 화성시
URLhttps://www.data.go.kr/data/15076283/fileData.do

Alerts

업종 is highly imbalanced (73.9%)Imbalance
연번 has unique valuesUnique
상호 has unique valuesUnique

Reproduction

Analysis started2023-12-12 15:03:51.274786
Analysis finished2023-12-12 15:03:51.782359
Duration0.51 seconds
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

연번
Real number (ℝ)

UNIQUE 

Distinct50
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean25.5
Minimum1
Maximum50
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size582.0 B
2023-12-13T00:03:51.852800image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1
5-th percentile3.45
Q113.25
median25.5
Q337.75
95-th percentile47.55
Maximum50
Range49
Interquartile range (IQR)24.5

Descriptive statistics

Standard deviation14.57738
Coefficient of variation (CV)0.57166195
Kurtosis-1.2
Mean25.5
Median Absolute Deviation (MAD)12.5
Skewness0
Sum1275
Variance212.5
MonotonicityStrictly increasing
2023-12-13T00:03:52.015316image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
1 1
 
2.0%
39 1
 
2.0%
29 1
 
2.0%
30 1
 
2.0%
31 1
 
2.0%
32 1
 
2.0%
33 1
 
2.0%
34 1
 
2.0%
35 1
 
2.0%
36 1
 
2.0%
Other values (40) 40
80.0%
ValueCountFrequency (%)
1 1
2.0%
2 1
2.0%
3 1
2.0%
4 1
2.0%
5 1
2.0%
6 1
2.0%
7 1
2.0%
8 1
2.0%
9 1
2.0%
10 1
2.0%
ValueCountFrequency (%)
50 1
2.0%
49 1
2.0%
48 1
2.0%
47 1
2.0%
46 1
2.0%
45 1
2.0%
44 1
2.0%
43 1
2.0%
42 1
2.0%
41 1
2.0%

업종
Categorical

IMBALANCE 

Distinct4
Distinct (%)8.0%
Missing0
Missing (%)0.0%
Memory size532.0 B
관광펜션업
46 
한옥체험업(구)
 
2
관광극장식당업
 
1
관광궤도업
 
1

Length

Max length8
Median length5
Mean length5.16
Min length5

Unique

Unique2 ?
Unique (%)4.0%

Sample

1st row관광극장식당업
2nd row관광펜션업
3rd row관광펜션업
4th row관광펜션업
5th row관광펜션업

Common Values

ValueCountFrequency (%)
관광펜션업 46
92.0%
한옥체험업(구) 2
 
4.0%
관광극장식당업 1
 
2.0%
관광궤도업 1
 
2.0%

Length

2023-12-13T00:03:52.143283image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-13T00:03:52.238282image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
관광펜션업 46
92.0%
한옥체험업(구 2
 
4.0%
관광극장식당업 1
 
2.0%
관광궤도업 1
 
2.0%

상호
Text

UNIQUE 

Distinct50
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Memory size532.0 B
2023-12-13T00:03:52.462019image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length14
Median length12
Mean length5.88
Min length3

Characters and Unicode

Total characters294
Distinct characters125
Distinct categories6 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique50 ?
Unique (%)100.0%

Sample

1st rowS관광나이트클럽
2nd row바다뜰이야기펜션
3rd row다소니펜션
4th row수에뇨펜션
5th row미고펜션
ValueCountFrequency (%)
솔잎펜션 2
 
3.4%
작은섬 2
 
3.4%
2 2
 
3.4%
1 2
 
3.4%
여행스케치 2
 
3.4%
일마레 1
 
1.7%
양지팬션 1
 
1.7%
양지리조텔 1
 
1.7%
토닥토닥 1
 
1.7%
나그랑펜션 1
 
1.7%
Other values (44) 44
74.6%
2023-12-13T00:03:52.870541image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
33
 
11.2%
32
 
10.9%
9
 
3.1%
7
 
2.4%
7
 
2.4%
6
 
2.0%
5
 
1.7%
5
 
1.7%
4
 
1.4%
4
 
1.4%
Other values (115) 182
61.9%

Most occurring categories

ValueCountFrequency (%)
Other Letter 268
91.2%
Space Separator 9
 
3.1%
Decimal Number 9
 
3.1%
Open Punctuation 3
 
1.0%
Close Punctuation 3
 
1.0%
Uppercase Letter 2
 
0.7%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
33
 
12.3%
32
 
11.9%
7
 
2.6%
7
 
2.6%
6
 
2.2%
5
 
1.9%
5
 
1.9%
4
 
1.5%
4
 
1.5%
4
 
1.5%
Other values (106) 161
60.1%
Decimal Number
ValueCountFrequency (%)
2 4
44.4%
1 3
33.3%
8 1
 
11.1%
0 1
 
11.1%
Uppercase Letter
ValueCountFrequency (%)
B 1
50.0%
S 1
50.0%
Space Separator
ValueCountFrequency (%)
9
100.0%
Open Punctuation
ValueCountFrequency (%)
( 3
100.0%
Close Punctuation
ValueCountFrequency (%)
) 3
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 268
91.2%
Common 24
 
8.2%
Latin 2
 
0.7%

Most frequent character per script

Hangul
ValueCountFrequency (%)
33
 
12.3%
32
 
11.9%
7
 
2.6%
7
 
2.6%
6
 
2.2%
5
 
1.9%
5
 
1.9%
4
 
1.5%
4
 
1.5%
4
 
1.5%
Other values (106) 161
60.1%
Common
ValueCountFrequency (%)
9
37.5%
2 4
16.7%
( 3
 
12.5%
1 3
 
12.5%
) 3
 
12.5%
8 1
 
4.2%
0 1
 
4.2%
Latin
ValueCountFrequency (%)
B 1
50.0%
S 1
50.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 268
91.2%
ASCII 26
 
8.8%

Most frequent character per block

Hangul
ValueCountFrequency (%)
33
 
12.3%
32
 
11.9%
7
 
2.6%
7
 
2.6%
6
 
2.2%
5
 
1.9%
5
 
1.9%
4
 
1.5%
4
 
1.5%
4
 
1.5%
Other values (106) 161
60.1%
ASCII
ValueCountFrequency (%)
9
34.6%
2 4
15.4%
( 3
 
11.5%
1 3
 
11.5%
) 3
 
11.5%
B 1
 
3.8%
8 1
 
3.8%
S 1
 
3.8%
0 1
 
3.8%
Distinct48
Distinct (%)96.0%
Missing0
Missing (%)0.0%
Memory size532.0 B
2023-12-13T00:03:53.141397image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length37
Median length30
Mean length22.72
Min length19

Characters and Unicode

Total characters1136
Distinct characters46
Distinct categories8 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique46 ?
Unique (%)92.0%

Sample

1st row경기도 화성시 반송동 93-6 센트럴S타운 9,10층
2nd row경기도 화성시 서신면 제부리 40-1
3rd row경기도 화성시 서신면 제부리 10-20
4th row경기도 화성시 서신면 제부리 270-8
5th row경기도 화성시 서신면 제부리 103
ValueCountFrequency (%)
경기도 50
19.4%
화성시 50
19.4%
서신면 49
19.0%
제부리 48
18.6%
190-81 2
 
0.8%
290-1 2
 
0.8%
217-2 1
 
0.4%
센트럴s타운 1
 
0.4%
93-6 1
 
0.4%
반송동 1
 
0.4%
Other values (53) 53
20.5%
2023-12-13T00:03:53.519340image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
258
22.7%
1 52
 
4.6%
51
 
4.5%
50
 
4.4%
50
 
4.4%
50
 
4.4%
50
 
4.4%
50
 
4.4%
50
 
4.4%
49
 
4.3%
Other values (36) 426
37.5%

Most occurring categories

ValueCountFrequency (%)
Other Letter 613
54.0%
Space Separator 258
22.7%
Decimal Number 214
 
18.8%
Dash Punctuation 45
 
4.0%
Close Punctuation 2
 
0.2%
Open Punctuation 2
 
0.2%
Other Punctuation 1
 
0.1%
Uppercase Letter 1
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
51
8.3%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
49
8.0%
49
8.0%
49
8.0%
Other values (20) 115
18.8%
Decimal Number
ValueCountFrequency (%)
1 52
24.3%
0 35
16.4%
9 28
13.1%
2 27
12.6%
8 22
10.3%
4 14
 
6.5%
7 12
 
5.6%
5 11
 
5.1%
6 9
 
4.2%
3 4
 
1.9%
Space Separator
ValueCountFrequency (%)
258
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 45
100.0%
Close Punctuation
ValueCountFrequency (%)
) 2
100.0%
Open Punctuation
ValueCountFrequency (%)
( 2
100.0%
Other Punctuation
ValueCountFrequency (%)
, 1
100.0%
Uppercase Letter
ValueCountFrequency (%)
S 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 613
54.0%
Common 522
46.0%
Latin 1
 
0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
51
8.3%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
49
8.0%
49
8.0%
49
8.0%
Other values (20) 115
18.8%
Common
ValueCountFrequency (%)
258
49.4%
1 52
 
10.0%
- 45
 
8.6%
0 35
 
6.7%
9 28
 
5.4%
2 27
 
5.2%
8 22
 
4.2%
4 14
 
2.7%
7 12
 
2.3%
5 11
 
2.1%
Other values (5) 18
 
3.4%
Latin
ValueCountFrequency (%)
S 1
100.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 613
54.0%
ASCII 523
46.0%

Most frequent character per block

ASCII
ValueCountFrequency (%)
258
49.3%
1 52
 
9.9%
- 45
 
8.6%
0 35
 
6.7%
9 28
 
5.4%
2 27
 
5.2%
8 22
 
4.2%
4 14
 
2.7%
7 12
 
2.3%
5 11
 
2.1%
Other values (6) 19
 
3.6%
Hangul
ValueCountFrequency (%)
51
8.3%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
50
8.2%
49
8.0%
49
8.0%
49
8.0%
Other values (20) 115
18.8%
Distinct48
Distinct (%)96.0%
Missing0
Missing (%)0.0%
Memory size532.0 B
2023-12-13T00:03:53.752396image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length52
Median length40
Mean length24.3
Min length18

Characters and Unicode

Total characters1215
Distinct characters56
Distinct categories8 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique46 ?
Unique (%)92.0%

Sample

1st row경기도 화성시 메타폴리스로 47-25 (반송동)
2nd row경기도 화성시 서신면 해안길 96-19
3rd row경기도 화성시 서신면 해안길442번길 62-6
4th row경기도 화성시 서신면 해안길 377
5th row경기도 화성시 서신면 해안길178번길 48-7
ValueCountFrequency (%)
경기도 50
17.9%
화성시 50
17.9%
서신면 49
17.6%
해안길 30
10.8%
해안길178번길 13
 
4.7%
해안길442번길 4
 
1.4%
1동 3
 
1.1%
1층 3
 
1.1%
1 3
 
1.1%
57 2
 
0.7%
Other values (69) 72
25.8%
2023-12-13T00:03:54.170851image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
229
18.8%
70
 
5.8%
1 62
 
5.1%
50
 
4.1%
50
 
4.1%
50
 
4.1%
50
 
4.1%
50
 
4.1%
50
 
4.1%
49
 
4.0%
Other values (46) 505
41.6%

Most occurring categories

ValueCountFrequency (%)
Other Letter 676
55.6%
Decimal Number 251
 
20.7%
Space Separator 229
 
18.8%
Dash Punctuation 29
 
2.4%
Other Punctuation 18
 
1.5%
Close Punctuation 5
 
0.4%
Open Punctuation 5
 
0.4%
Uppercase Letter 2
 
0.2%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
70
10.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
49
 
7.2%
49
 
7.2%
49
 
7.2%
Other values (30) 159
23.5%
Decimal Number
ValueCountFrequency (%)
1 62
24.7%
2 34
13.5%
8 27
10.8%
4 26
10.4%
7 26
10.4%
3 22
 
8.8%
6 19
 
7.6%
0 16
 
6.4%
5 12
 
4.8%
9 7
 
2.8%
Space Separator
ValueCountFrequency (%)
229
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 29
100.0%
Other Punctuation
ValueCountFrequency (%)
, 18
100.0%
Close Punctuation
ValueCountFrequency (%)
) 5
100.0%
Open Punctuation
ValueCountFrequency (%)
( 5
100.0%
Uppercase Letter
ValueCountFrequency (%)
B 2
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 676
55.6%
Common 537
44.2%
Latin 2
 
0.2%

Most frequent character per script

Hangul
ValueCountFrequency (%)
70
10.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
49
 
7.2%
49
 
7.2%
49
 
7.2%
Other values (30) 159
23.5%
Common
ValueCountFrequency (%)
229
42.6%
1 62
 
11.5%
2 34
 
6.3%
- 29
 
5.4%
8 27
 
5.0%
4 26
 
4.8%
7 26
 
4.8%
3 22
 
4.1%
6 19
 
3.5%
, 18
 
3.4%
Other values (5) 45
 
8.4%
Latin
ValueCountFrequency (%)
B 2
100.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 676
55.6%
ASCII 539
44.4%

Most frequent character per block

ASCII
ValueCountFrequency (%)
229
42.5%
1 62
 
11.5%
2 34
 
6.3%
- 29
 
5.4%
8 27
 
5.0%
4 26
 
4.8%
7 26
 
4.8%
3 22
 
4.1%
6 19
 
3.5%
, 18
 
3.3%
Other values (6) 47
 
8.7%
Hangul
ValueCountFrequency (%)
70
10.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
50
 
7.4%
49
 
7.2%
49
 
7.2%
49
 
7.2%
Other values (30) 159
23.5%

Interactions

2023-12-13T00:03:51.555859image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-13T00:03:54.276114image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
연번업종상호소재지 지번주소소재지 도로명주소
연번1.0000.4661.0000.9690.969
업종0.4661.0001.0001.0001.000
상호1.0001.0001.0001.0001.000
소재지 지번주소0.9691.0001.0001.0001.000
소재지 도로명주소0.9691.0001.0001.0001.000
2023-12-13T00:03:54.376176image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
연번업종
연번1.0000.270
업종0.2701.000

Missing values

2023-12-13T00:03:51.652180image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-13T00:03:51.742656image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

연번업종상호소재지 지번주소소재지 도로명주소
01관광극장식당업S관광나이트클럽경기도 화성시 반송동 93-6 센트럴S타운 9,10층경기도 화성시 메타폴리스로 47-25 (반송동)
12관광펜션업바다뜰이야기펜션경기도 화성시 서신면 제부리 40-1경기도 화성시 서신면 해안길 96-19
23관광펜션업다소니펜션경기도 화성시 서신면 제부리 10-20경기도 화성시 서신면 해안길442번길 62-6
34관광펜션업수에뇨펜션경기도 화성시 서신면 제부리 270-8경기도 화성시 서신면 해안길 377
45관광펜션업미고펜션경기도 화성시 서신면 제부리 103경기도 화성시 서신면 해안길178번길 48-7
56관광펜션업초원펜션경기도 화성시 서신면 제부리 40-5경기도 화성시 서신면 해안길 96-1
67관광펜션업리베펜션경기도 화성시 서신면 제부리 187 (및 187-2번지)경기도 화성시 서신면 해안길178번길 72 (및 해안길 178번길 74)
78관광펜션업해송펜션경기도 화성시 서신면 제부리 41-1경기도 화성시 서신면 해안길 116-7
89관광펜션업동미산펜션경기도 화성시 서신면 제부리 76경기도 화성시 서신면 제부말길 165
910관광펜션업테라스의 아침펜션경기도 화성시 서신면 제부리 288-58 외 제부리 288-100경기도 화성시 서신면 해안길 412-13 (외 1 (해안길 412-19))
연번업종상호소재지 지번주소소재지 도로명주소
4041관광펜션업아띠펜션경기도 화성시 서신면 제부리 190-102경기도 화성시 서신면 해안길 246-4
4142관광펜션업리멤버8펜션경기도 화성시 서신면 제부리 79-1 (제부리 79-5)경기도 화성시 서신면 해안길178번길 8-11, 1동, (해안길178번길 8-9, 1동,2동)
4243관광펜션업바다마을펜션경기도 화성시 서신면 제부리 41-17경기도 화성시 서신면 해안길 128
4344관광펜션업여행스케치 B경기도 화성시 서신면 제부리 190-29경기도 화성시 서신면 해안길 308-5, 여행스케치B 3층
4445관광펜션업바닷길펜션경기도 화성시 서신면 제부리 290-9경기도 화성시 서신면 해안길 330
4546관광펜션업나드리 펜션경기도 화성시 서신면 제부리 41-4경기도 화성시 서신면 해안길 122
4647관광펜션업제부 포레스트경기도 화성시 서신면 제부리 8-3경기도 화성시 서신면 해안길442번길 25
4748관광궤도업제부도해상케이블카 주식회사경기도 화성시 서신면 장외리 618-12경기도 화성시 서신면 전곡항로 1-10
4849한옥체험업(구)옥란재경기도 화성시 서신면 용두리 95-1경기도 화성시 서신면 영종이길 120-8
4950한옥체험업(구)백미응서재경기도 화성시 서신면 백미리 564경기도 화성시 서신면 밸미길 84