Overview

Dataset statistics

Number of variables6
Number of observations649
Missing cells0
Missing cells (%)0.0%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory31.2 KiB
Average record size in memory49.2 B

Variable types

Numeric1
Categorical2
Text3

Dataset

Description도 내에 자리잡고 있는 모범음식점 현황으로, 시도, 시군, 주메뉴, 업소명, 주소와 같은 항목에 대한 정보들을 제공합니다.
Author경상남도
URLhttps://www.data.go.kr/data/3079051/fileData.do

Alerts

시도 has constant value ""Constant
연번 is highly overall correlated with 시군High correlation
시군 is highly overall correlated with 연번High correlation
연번 has unique valuesUnique

Reproduction

Analysis started2024-03-14 16:17:07.254968
Analysis finished2024-03-14 16:17:08.806853
Duration1.55 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

연번
Real number (ℝ)

HIGH CORRELATION  UNIQUE 

Distinct649
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean325
Minimum1
Maximum649
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size5.8 KiB
2024-03-15T01:17:08.945134image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1
5-th percentile33.4
Q1163
median325
Q3487
95-th percentile616.6
Maximum649
Range648
Interquartile range (IQR)324

Descriptive statistics

Standard deviation187.49444
Coefficient of variation (CV)0.57690598
Kurtosis-1.2
Mean325
Median Absolute Deviation (MAD)162
Skewness0
Sum210925
Variance35154.167
MonotonicityStrictly increasing
2024-03-15T01:17:09.304374image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
1 1
 
0.2%
447 1
 
0.2%
429 1
 
0.2%
430 1
 
0.2%
431 1
 
0.2%
432 1
 
0.2%
433 1
 
0.2%
434 1
 
0.2%
435 1
 
0.2%
436 1
 
0.2%
Other values (639) 639
98.5%
ValueCountFrequency (%)
1 1
0.2%
2 1
0.2%
3 1
0.2%
4 1
0.2%
5 1
0.2%
6 1
0.2%
7 1
0.2%
8 1
0.2%
9 1
0.2%
10 1
0.2%
ValueCountFrequency (%)
649 1
0.2%
648 1
0.2%
647 1
0.2%
646 1
0.2%
645 1
0.2%
644 1
0.2%
643 1
0.2%
642 1
0.2%
641 1
0.2%
640 1
0.2%

시도
Categorical

CONSTANT 

Distinct1
Distinct (%)0.2%
Missing0
Missing (%)0.0%
Memory size5.2 KiB
경상남도
649 

Length

Max length4
Median length4
Mean length4
Min length4

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row경상남도
2nd row경상남도
3rd row경상남도
4th row경상남도
5th row경상남도

Common Values

ValueCountFrequency (%)
경상남도 649
100.0%

Length

2024-03-15T01:17:09.754622image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2024-03-15T01:17:09.926478image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
경상남도 649
100.0%

시군
Categorical

HIGH CORRELATION 

Distinct21
Distinct (%)3.2%
Missing0
Missing (%)0.0%
Memory size5.2 KiB
거제시
50 
창원시 성산구
46 
남해군
44 
창원시 의창구
43 
거창군
 
42
Other values (16)
424 

Length

Max length9
Median length3
Mean length4.3374422
Min length3

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row창원시 마산합포구
2nd row창원시 마산합포구
3rd row창원시 마산합포구
4th row창원시 마산합포구
5th row창원시 마산합포구

Common Values

ValueCountFrequency (%)
거제시 50
 
7.7%
창원시 성산구 46
 
7.1%
남해군 44
 
6.8%
창원시 의창구 43
 
6.6%
거창군 42
 
6.5%
사천시 40
 
6.2%
창원시 마산합포구 39
 
6.0%
통영시 37
 
5.7%
창원시 진해구 32
 
4.9%
합천군 31
 
4.8%
Other values (11) 245
37.8%

Length

2024-03-15T01:17:10.252407image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category
ValueCountFrequency (%)
창원시 185
22.2%
거제시 50
 
6.0%
성산구 46
 
5.5%
남해군 44
 
5.3%
의창구 43
 
5.2%
거창군 42
 
5.0%
사천시 40
 
4.8%
마산합포구 39
 
4.7%
통영시 37
 
4.4%
진해구 32
 
3.8%
Other values (12) 276
33.1%
Distinct394
Distinct (%)60.7%
Missing0
Missing (%)0.0%
Memory size5.2 KiB
2024-03-15T01:17:11.450890image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length30
Median length26
Mean length5.972265
Min length1

Characters and Unicode

Total characters3876
Distinct characters240
Distinct categories5 ?
Distinct scripts2 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique311 ?
Unique (%)47.9%

Sample

1st row닭요리
2nd row삼계탕
3rd row순대전골
4th row곱창전골
5th row생선회
ValueCountFrequency (%)
생선회 49
 
5.5%
돼지갈비 25
 
2.8%
매운탕 23
 
2.6%
삼겹살 17
 
1.9%
삼계탕 16
 
1.8%
오리불고기 14
 
1.6%
갈비 13
 
1.5%
정식 12
 
1.4%
한정식 11
 
1.2%
갈비탕 11
 
1.2%
Other values (360) 697
78.5%
2024-03-15T01:17:13.030006image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
, 301
 
7.8%
240
 
6.2%
164
 
4.2%
114
 
2.9%
114
 
2.9%
108
 
2.8%
106
 
2.7%
100
 
2.6%
92
 
2.4%
87
 
2.2%
Other values (230) 2450
63.2%

Most occurring categories

ValueCountFrequency (%)
Other Letter 3320
85.7%
Other Punctuation 309
 
8.0%
Space Separator 240
 
6.2%
Open Punctuation 4
 
0.1%
Close Punctuation 3
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
164
 
4.9%
114
 
3.4%
114
 
3.4%
108
 
3.3%
106
 
3.2%
100
 
3.0%
92
 
2.8%
87
 
2.6%
78
 
2.3%
76
 
2.3%
Other values (225) 2281
68.7%
Other Punctuation
ValueCountFrequency (%)
, 301
97.4%
. 8
 
2.6%
Space Separator
ValueCountFrequency (%)
240
100.0%
Open Punctuation
ValueCountFrequency (%)
( 4
100.0%
Close Punctuation
ValueCountFrequency (%)
) 3
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 3320
85.7%
Common 556
 
14.3%

Most frequent character per script

Hangul
ValueCountFrequency (%)
164
 
4.9%
114
 
3.4%
114
 
3.4%
108
 
3.3%
106
 
3.2%
100
 
3.0%
92
 
2.8%
87
 
2.6%
78
 
2.3%
76
 
2.3%
Other values (225) 2281
68.7%
Common
ValueCountFrequency (%)
, 301
54.1%
240
43.2%
. 8
 
1.4%
( 4
 
0.7%
) 3
 
0.5%

Most occurring blocks

ValueCountFrequency (%)
Hangul 3320
85.7%
ASCII 556
 
14.3%

Most frequent character per block

ASCII
ValueCountFrequency (%)
, 301
54.1%
240
43.2%
. 8
 
1.4%
( 4
 
0.7%
) 3
 
0.5%
Hangul
ValueCountFrequency (%)
164
 
4.9%
114
 
3.4%
114
 
3.4%
108
 
3.3%
106
 
3.2%
100
 
3.0%
92
 
2.8%
87
 
2.6%
78
 
2.3%
76
 
2.3%
Other values (225) 2281
68.7%
Distinct637
Distinct (%)98.2%
Missing0
Missing (%)0.0%
Memory size5.2 KiB
2024-03-15T01:17:14.029598image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length18
Median length16
Mean length5.5315871
Min length1

Characters and Unicode

Total characters3590
Distinct characters410
Distinct categories7 ?
Distinct scripts2 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique627 ?
Unique (%)96.6%

Sample

1st row아자방
2nd row백제령삼계탕
3rd row개성순대 주식회사
4th row인혜돌곱창
5th row청송해변집
ValueCountFrequency (%)
양평해장국 3
 
0.4%
푸주옥 3
 
0.4%
주남오리궁 2
 
0.3%
주식회사 2
 
0.3%
2
 
0.3%
양덕점 2
 
0.3%
남해 2
 
0.3%
뷔페 2
 
0.3%
한우 2
 
0.3%
싱싱게장 2
 
0.3%
Other values (685) 695
96.9%
2024-03-15T01:17:15.257813image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
101
 
2.8%
94
 
2.6%
76
 
2.1%
70
 
1.9%
70
 
1.9%
67
 
1.9%
55
 
1.5%
53
 
1.5%
52
 
1.4%
51
 
1.4%
Other values (400) 2901
80.8%

Most occurring categories

ValueCountFrequency (%)
Other Letter 3464
96.5%
Space Separator 70
 
1.9%
Decimal Number 16
 
0.4%
Other Punctuation 14
 
0.4%
Open Punctuation 12
 
0.3%
Close Punctuation 12
 
0.3%
Dash Punctuation 2
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
101
 
2.9%
94
 
2.7%
76
 
2.2%
70
 
2.0%
67
 
1.9%
55
 
1.6%
53
 
1.5%
52
 
1.5%
51
 
1.5%
50
 
1.4%
Other values (383) 2795
80.7%
Decimal Number
ValueCountFrequency (%)
2 4
25.0%
3 3
18.8%
1 2
12.5%
0 2
12.5%
4 2
12.5%
9 1
 
6.2%
6 1
 
6.2%
5 1
 
6.2%
Other Punctuation
ValueCountFrequency (%)
. 5
35.7%
· 4
28.6%
& 3
21.4%
/ 1
 
7.1%
, 1
 
7.1%
Space Separator
ValueCountFrequency (%)
70
100.0%
Open Punctuation
ValueCountFrequency (%)
( 12
100.0%
Close Punctuation
ValueCountFrequency (%)
) 12
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 2
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 3464
96.5%
Common 126
 
3.5%

Most frequent character per script

Hangul
ValueCountFrequency (%)
101
 
2.9%
94
 
2.7%
76
 
2.2%
70
 
2.0%
67
 
1.9%
55
 
1.6%
53
 
1.5%
52
 
1.5%
51
 
1.5%
50
 
1.4%
Other values (383) 2795
80.7%
Common
ValueCountFrequency (%)
70
55.6%
( 12
 
9.5%
) 12
 
9.5%
. 5
 
4.0%
· 4
 
3.2%
2 4
 
3.2%
& 3
 
2.4%
3 3
 
2.4%
- 2
 
1.6%
1 2
 
1.6%
Other values (7) 9
 
7.1%

Most occurring blocks

ValueCountFrequency (%)
Hangul 3464
96.5%
ASCII 122
 
3.4%
None 4
 
0.1%

Most frequent character per block

Hangul
ValueCountFrequency (%)
101
 
2.9%
94
 
2.7%
76
 
2.2%
70
 
2.0%
67
 
1.9%
55
 
1.6%
53
 
1.5%
52
 
1.5%
51
 
1.5%
50
 
1.4%
Other values (383) 2795
80.7%
ASCII
ValueCountFrequency (%)
70
57.4%
( 12
 
9.8%
) 12
 
9.8%
. 5
 
4.1%
2 4
 
3.3%
& 3
 
2.5%
3 3
 
2.5%
- 2
 
1.6%
1 2
 
1.6%
0 2
 
1.6%
Other values (6) 7
 
5.7%
None
ValueCountFrequency (%)
· 4
100.0%

주소
Text

Distinct646
Distinct (%)99.5%
Missing0
Missing (%)0.0%
Memory size5.2 KiB
2024-03-15T01:17:16.695064image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length55
Median length46
Mean length22.594761
Min length11

Characters and Unicode

Total characters14664
Distinct characters297
Distinct categories10 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique643 ?
Unique (%)99.1%

Sample

1st row 마산합포구 3·15대로 383-1 (중성동,(1.2층))
2nd row 마산합포구 3·15대로 383-4 (중성동)
3rd row 마산합포구 가포로 686 (덕동동,(지상1층))
4th row 마산합포구 교방천남길 320 (오동동)
5th row 마산합포구 구산면 해양관광로 1285-57
ValueCountFrequency (%)
경상남도 262
 
8.6%
1층 59
 
1.9%
거제시 50
 
1.6%
성산구 46
 
1.5%
남해군 44
 
1.4%
의창구 43
 
1.4%
거창군 42
 
1.4%
사천시 40
 
1.3%
마산합포구 39
 
1.3%
거창읍 33
 
1.1%
Other values (1231) 2393
78.4%
2024-03-15T01:17:18.713196image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
2655
 
18.1%
1 685
 
4.7%
459
 
3.1%
450
 
3.1%
2 423
 
2.9%
406
 
2.8%
356
 
2.4%
) 352
 
2.4%
( 352
 
2.4%
335
 
2.3%
Other values (287) 8191
55.9%

Most occurring categories

ValueCountFrequency (%)
Other Letter 8154
55.6%
Decimal Number 2689
 
18.3%
Space Separator 2655
 
18.1%
Close Punctuation 352
 
2.4%
Open Punctuation 352
 
2.4%
Other Punctuation 270
 
1.8%
Dash Punctuation 179
 
1.2%
Uppercase Letter 6
 
< 0.1%
Math Symbol 5
 
< 0.1%
Lowercase Letter 2
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
459
 
5.6%
450
 
5.5%
406
 
5.0%
356
 
4.4%
335
 
4.1%
280
 
3.4%
275
 
3.4%
266
 
3.3%
212
 
2.6%
192
 
2.4%
Other values (263) 4923
60.4%
Decimal Number
ValueCountFrequency (%)
1 685
25.5%
2 423
15.7%
3 293
10.9%
5 217
 
8.1%
4 199
 
7.4%
0 194
 
7.2%
8 192
 
7.1%
6 177
 
6.6%
7 170
 
6.3%
9 139
 
5.2%
Other Punctuation
ValueCountFrequency (%)
, 257
95.2%
. 9
 
3.3%
· 3
 
1.1%
/ 1
 
0.4%
Uppercase Letter
ValueCountFrequency (%)
A 2
33.3%
G 2
33.3%
H 1
16.7%
B 1
16.7%
Space Separator
ValueCountFrequency (%)
2655
100.0%
Close Punctuation
ValueCountFrequency (%)
) 352
100.0%
Open Punctuation
ValueCountFrequency (%)
( 352
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 179
100.0%
Math Symbol
ValueCountFrequency (%)
~ 5
100.0%
Lowercase Letter
ValueCountFrequency (%)
l 2
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 8154
55.6%
Common 6502
44.3%
Latin 8
 
0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
459
 
5.6%
450
 
5.5%
406
 
5.0%
356
 
4.4%
335
 
4.1%
280
 
3.4%
275
 
3.4%
266
 
3.3%
212
 
2.6%
192
 
2.4%
Other values (263) 4923
60.4%
Common
ValueCountFrequency (%)
2655
40.8%
1 685
 
10.5%
2 423
 
6.5%
) 352
 
5.4%
( 352
 
5.4%
3 293
 
4.5%
, 257
 
4.0%
5 217
 
3.3%
4 199
 
3.1%
0 194
 
3.0%
Other values (9) 875
 
13.5%
Latin
ValueCountFrequency (%)
l 2
25.0%
A 2
25.0%
G 2
25.0%
H 1
12.5%
B 1
12.5%

Most occurring blocks

ValueCountFrequency (%)
Hangul 8154
55.6%
ASCII 6507
44.4%
None 3
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
2655
40.8%
1 685
 
10.5%
2 423
 
6.5%
) 352
 
5.4%
( 352
 
5.4%
3 293
 
4.5%
, 257
 
3.9%
5 217
 
3.3%
4 199
 
3.1%
0 194
 
3.0%
Other values (13) 880
 
13.5%
Hangul
ValueCountFrequency (%)
459
 
5.6%
450
 
5.5%
406
 
5.0%
356
 
4.4%
335
 
4.1%
280
 
3.4%
275
 
3.4%
266
 
3.3%
212
 
2.6%
192
 
2.4%
Other values (263) 4923
60.4%
None
ValueCountFrequency (%)
· 3
100.0%

Interactions

2024-03-15T01:17:08.086072image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2024-03-15T01:17:19.135153image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
연번시군
연번1.0000.984
시군0.9841.000
2024-03-15T01:17:19.272789image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
연번시군
연번1.0000.895
시군0.8951.000

Missing values

2024-03-15T01:17:08.516805image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2024-03-15T01:17:08.718140image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

연번시도시군주메뉴업소명주소
01경상남도창원시 마산합포구닭요리아자방마산합포구 3·15대로 383-1 (중성동,(1.2층))
12경상남도창원시 마산합포구삼계탕백제령삼계탕마산합포구 3·15대로 383-4 (중성동)
23경상남도창원시 마산합포구순대전골개성순대 주식회사마산합포구 가포로 686 (덕동동,(지상1층))
34경상남도창원시 마산합포구곱창전골인혜돌곱창마산합포구 교방천남길 320 (오동동)
45경상남도창원시 마산합포구생선회청송해변집마산합포구 구산면 해양관광로 1285-57
56경상남도창원시 마산합포구추어탕황토추어탕마산합포구 동서동2길 41 (신포동2가,기전씨티텔 101호)
67경상남도창원시 마산합포구해물탕백일섭의 전복예찬 마산점마산합포구 동서동로 18 (신포동2가)
78경상남도창원시 마산합포구모듬쌈밥해송마산합포구 동서북14길 16 (동성동)
89경상남도창원시 마산합포구비빔밥, 수육홍화마산합포구 동서북9길 8-17 (남성동)
910경상남도창원시 마산합포구한정식시골밥상마산합포구 동서북9길 8-26 (남성동)
연번시도시군주메뉴업소명주소
639640경상남도합천군갈비찜,청국장청마루경상남도 합천군 대양면 동부로 21
640641경상남도합천군한식,추어탕,된장찌개청솔음식마을경상남도 합천군 청덕면 동부로 2488-6
641642경상남도합천군청국장,보리밥토속식당경상남도 합천군 대병면 회양관광단지길 36-1
642643경상남도합천군오리요리합천가든경상남도 합천군 합천읍 옥산로 26
643644경상남도합천군돼지국밥합천돼지국밥경상남도 합천군 합천읍 문화로 9
644645경상남도합천군돼지갈비찜합천명품돼지경상남도 합천군 합천읍 남정길 82
645646경상남도합천군돼지갈비,돼지구이합천명품토종돼지경상남도 합천군 묘산면 묘산로 215
646647경상남도합천군돼지고기합천명품토종흑돼지경상남도 합천군 가야면 가야시장로 106
647648경상남도합천군산채한정식해인사 맛집 감로식당경상남도 합천군 가야면 치인1길 8-1
648649경상남도합천군돼지갈비, 한우모듬황금나무숯불갈비경상남도 합천군 합천읍 옥산로 15