Overview

Dataset statistics

Number of variables10
Number of observations266
Missing cells61
Missing cells (%)2.3%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory21.2 KiB
Average record size in memory81.5 B

Variable types

Numeric1
Text7
Categorical2

Dataset

Description중랑구립면목정보 도서관 신착 자료 2014년 12월 분
Author중랑구시설관리공단
URLhttps://www.data.go.kr/data/15044312/fileData.do

Alerts

[2014년 12월] 중랑구립면목정보도서관 신착자료 목록 is highly overall correlated with Unnamed: 9High correlation
Unnamed: 6 is highly overall correlated with Unnamed: 9High correlation
Unnamed: 9 is highly overall correlated with [2014년 12월] 중랑구립면목정보도서관 신착자료 목록 and 1 other fieldsHigh correlation
Unnamed: 9 is highly imbalanced (50.1%)Imbalance
Unnamed: 8 has 53 (19.9%) missing valuesMissing

Reproduction

Analysis started2023-12-12 04:48:56.637238
Analysis finished2023-12-12 04:48:58.515960
Duration1.88 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

Distinct264
Distinct (%)100.0%
Missing2
Missing (%)0.8%
Infinite0
Infinite (%)0.0%
Mean132.5
Minimum1
Maximum264
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size2.5 KiB
2023-12-12T13:48:58.618695image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1
5-th percentile14.15
Q166.75
median132.5
Q3198.25
95-th percentile250.85
Maximum264
Range263
Interquartile range (IQR)131.5

Descriptive statistics

Standard deviation76.354437
Coefficient of variation (CV)0.5762599
Kurtosis-1.2
Mean132.5
Median Absolute Deviation (MAD)66
Skewness0
Sum34980
Variance5830
MonotonicityStrictly increasing
2023-12-12T13:48:58.804499image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
183 1
 
0.4%
169 1
 
0.4%
170 1
 
0.4%
171 1
 
0.4%
172 1
 
0.4%
173 1
 
0.4%
174 1
 
0.4%
175 1
 
0.4%
176 1
 
0.4%
177 1
 
0.4%
Other values (254) 254
95.5%
(Missing) 2
 
0.8%
ValueCountFrequency (%)
1 1
0.4%
2 1
0.4%
3 1
0.4%
4 1
0.4%
5 1
0.4%
6 1
0.4%
7 1
0.4%
8 1
0.4%
9 1
0.4%
10 1
0.4%
ValueCountFrequency (%)
264 1
0.4%
263 1
0.4%
262 1
0.4%
261 1
0.4%
260 1
0.4%
259 1
0.4%
258 1
0.4%
257 1
0.4%
256 1
0.4%
255 1
0.4%
Distinct265
Distinct (%)100.0%
Missing1
Missing (%)0.4%
Memory size2.2 KiB
2023-12-12T13:48:59.070030image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length12
Median length12
Mean length11.969811
Min length4

Characters and Unicode

Total characters3172
Distinct characters17
Distinct categories3 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique265 ?
Unique (%)100.0%

Sample

1st row등록번호
2nd rowDM0000085457
3rd rowDM0000085458
4th rowDM0000085454
5th rowDM0000085455
ValueCountFrequency (%)
dm0000085468 1
 
0.4%
dz0100084750 1
 
0.4%
dz0100084759 1
 
0.4%
dm0000084818 1
 
0.4%
dm0000084819 1
 
0.4%
dm0000084820 1
 
0.4%
dm0000084821 1
 
0.4%
dm0000084822 1
 
0.4%
dm0000084823 1
 
0.4%
dm0000084824 1
 
0.4%
Other values (255) 255
96.2%
2023-12-12T13:48:59.567769image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
0 1314
41.4%
8 353
 
11.1%
D 264
 
8.3%
4 245
 
7.7%
M 212
 
6.7%
5 148
 
4.7%
7 143
 
4.5%
6 132
 
4.2%
2 116
 
3.7%
1 106
 
3.3%
Other values (7) 139
 
4.4%

Most occurring categories

ValueCountFrequency (%)
Decimal Number 2640
83.2%
Uppercase Letter 528
 
16.6%
Other Letter 4
 
0.1%

Most frequent character per category

Decimal Number
ValueCountFrequency (%)
0 1314
49.8%
8 353
 
13.4%
4 245
 
9.3%
5 148
 
5.6%
7 143
 
5.4%
6 132
 
5.0%
2 116
 
4.4%
1 106
 
4.0%
9 45
 
1.7%
3 38
 
1.4%
Other Letter
ValueCountFrequency (%)
1
25.0%
1
25.0%
1
25.0%
1
25.0%
Uppercase Letter
ValueCountFrequency (%)
D 264
50.0%
M 212
40.2%
Z 52
 
9.8%

Most occurring scripts

ValueCountFrequency (%)
Common 2640
83.2%
Latin 528
 
16.6%
Hangul 4
 
0.1%

Most frequent character per script

Common
ValueCountFrequency (%)
0 1314
49.8%
8 353
 
13.4%
4 245
 
9.3%
5 148
 
5.6%
7 143
 
5.4%
6 132
 
5.0%
2 116
 
4.4%
1 106
 
4.0%
9 45
 
1.7%
3 38
 
1.4%
Hangul
ValueCountFrequency (%)
1
25.0%
1
25.0%
1
25.0%
1
25.0%
Latin
ValueCountFrequency (%)
D 264
50.0%
M 212
40.2%
Z 52
 
9.8%

Most occurring blocks

ValueCountFrequency (%)
ASCII 3168
99.9%
Hangul 4
 
0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
0 1314
41.5%
8 353
 
11.1%
D 264
 
8.3%
4 245
 
7.7%
M 212
 
6.7%
5 148
 
4.7%
7 143
 
4.5%
6 132
 
4.2%
2 116
 
3.7%
1 106
 
3.3%
Other values (3) 135
 
4.3%
Hangul
ValueCountFrequency (%)
1
25.0%
1
25.0%
1
25.0%
1
25.0%
Distinct213
Distinct (%)80.4%
Missing1
Missing (%)0.4%
Memory size2.2 KiB
2023-12-12T13:48:59.885182image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length22
Median length20
Mean length14.381132
Min length4

Characters and Unicode

Total characters3811
Distinct characters75
Distinct categories8 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique166 ?
Unique (%)62.6%

Sample

1st row청구기호
2nd rowNB 320.942-ㅂ158ㅂ
3rd rowNB 340.9-ㅎ899ㅈ
4th rowNB 834-ㅁ176ㅇ
5th rowNB 834-ㅁ176ㄴ
ValueCountFrequency (%)
jb 103
19.5%
nb 92
 
17.4%
bb 66
 
12.5%
747-r282w-v.5 3
 
0.6%
747-r282w-v.6 3
 
0.6%
747-r282w-v.4 3
 
0.6%
747-r282w-v.2 3
 
0.6%
747-r282w-v.1 3
 
0.6%
747-g573e-v.8 2
 
0.4%
747-b236a-v.1 2
 
0.4%
Other values (208) 249
47.1%
2023-12-12T13:49:00.376203image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
7 412
 
10.8%
B 362
 
9.5%
- 360
 
9.4%
4 292
 
7.7%
264
 
6.9%
2 193
 
5.1%
3 190
 
5.0%
8 187
 
4.9%
. 184
 
4.8%
5 155
 
4.1%
Other values (65) 1212
31.8%

Most occurring categories

ValueCountFrequency (%)
Decimal Number 1838
48.2%
Uppercase Letter 760
19.9%
Dash Punctuation 360
 
9.4%
Space Separator 264
 
6.9%
Other Letter 258
 
6.8%
Other Punctuation 184
 
4.8%
Lowercase Letter 146
 
3.8%
Math Symbol 1
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
44
17.1%
42
16.3%
24
9.3%
24
9.3%
20
7.8%
18
7.0%
17
 
6.6%
11
 
4.3%
11
 
4.3%
8
 
3.1%
Other values (20) 39
15.1%
Uppercase Letter
ValueCountFrequency (%)
B 362
47.6%
J 108
 
14.2%
N 93
 
12.2%
W 31
 
4.1%
E 24
 
3.2%
T 24
 
3.2%
G 24
 
3.2%
R 18
 
2.4%
S 17
 
2.2%
H 9
 
1.2%
Other values (13) 50
 
6.6%
Decimal Number
ValueCountFrequency (%)
7 412
22.4%
4 292
15.9%
2 193
10.5%
3 190
10.3%
8 187
10.2%
5 155
 
8.4%
1 129
 
7.0%
6 121
 
6.6%
9 81
 
4.4%
0 78
 
4.2%
Lowercase Letter
ValueCountFrequency (%)
v 96
65.8%
e 23
 
15.8%
a 10
 
6.8%
d 6
 
4.1%
b 4
 
2.7%
t 4
 
2.7%
l 2
 
1.4%
m 1
 
0.7%
Dash Punctuation
ValueCountFrequency (%)
- 360
100.0%
Space Separator
ValueCountFrequency (%)
264
100.0%
Other Punctuation
ValueCountFrequency (%)
. 184
100.0%
Math Symbol
ValueCountFrequency (%)
= 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Common 2647
69.5%
Latin 906
 
23.8%
Hangul 258
 
6.8%

Most frequent character per script

Latin
ValueCountFrequency (%)
B 362
40.0%
J 108
 
11.9%
v 96
 
10.6%
N 93
 
10.3%
W 31
 
3.4%
E 24
 
2.6%
T 24
 
2.6%
G 24
 
2.6%
e 23
 
2.5%
R 18
 
2.0%
Other values (21) 103
 
11.4%
Hangul
ValueCountFrequency (%)
44
17.1%
42
16.3%
24
9.3%
24
9.3%
20
7.8%
18
7.0%
17
 
6.6%
11
 
4.3%
11
 
4.3%
8
 
3.1%
Other values (20) 39
15.1%
Common
ValueCountFrequency (%)
7 412
15.6%
- 360
13.6%
4 292
11.0%
264
10.0%
2 193
7.3%
3 190
7.2%
8 187
7.1%
. 184
7.0%
5 155
 
5.9%
1 129
 
4.9%
Other values (4) 281
10.6%

Most occurring blocks

ValueCountFrequency (%)
ASCII 3553
93.2%
Compat Jamo 239
 
6.3%
Hangul 19
 
0.5%

Most frequent character per block

ASCII
ValueCountFrequency (%)
7 412
11.6%
B 362
 
10.2%
- 360
 
10.1%
4 292
 
8.2%
264
 
7.4%
2 193
 
5.4%
3 190
 
5.3%
8 187
 
5.3%
. 184
 
5.2%
5 155
 
4.4%
Other values (35) 954
26.9%
Compat Jamo
ValueCountFrequency (%)
44
18.4%
42
17.6%
24
10.0%
24
10.0%
20
8.4%
18
7.5%
17
 
7.1%
11
 
4.6%
11
 
4.6%
8
 
3.3%
Other values (5) 20
8.4%
Hangul
ValueCountFrequency (%)
4
21.1%
2
 
10.5%
1
 
5.3%
1
 
5.3%
1
 
5.3%
1
 
5.3%
1
 
5.3%
1
 
5.3%
1
 
5.3%
1
 
5.3%
Other values (5) 5
26.3%
Distinct260
Distinct (%)98.1%
Missing1
Missing (%)0.4%
Memory size2.2 KiB
2023-12-12T13:49:00.793869image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length131
Median length57
Mean length32.909434
Min length2

Characters and Unicode

Total characters8721
Distinct characters522
Distinct categories10 ?
Distinct scripts4 ?
Distinct blocks4 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique255 ?
Unique (%)96.2%

Sample

1st row서명
2nd row불평등 민주주의 : 자유에 가려진 진실
3rd row정치 질서의 기원 : 불안정성을 극복할 정치적 힘은 어디서 오는가
4th row여자라는 생물 : 마스다 미리 에세이
5th row나는 사랑을 하고 있어 : 마스다 미리 에세이
ValueCountFrequency (%)
109
 
6.0%
the 96
 
5.3%
a 29
 
1.6%
and 26
 
1.4%
little 22
 
1.2%
my 21
 
1.2%
of 19
 
1.0%
to 13
 
0.7%
영어 13
 
0.7%
터지는 12
 
0.7%
Other values (1014) 1454
80.2%
2023-12-12T13:49:01.390193image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
1550
 
17.8%
e 483
 
5.5%
a 283
 
3.2%
t 282
 
3.2%
o 279
 
3.2%
i 263
 
3.0%
r 243
 
2.8%
n 212
 
2.4%
s 204
 
2.3%
h 189
 
2.2%
Other values (512) 4733
54.3%

Most occurring categories

ValueCountFrequency (%)
Lowercase Letter 3448
39.5%
Other Letter 2497
28.6%
Space Separator 1550
17.8%
Uppercase Letter 481
 
5.5%
Other Punctuation 216
 
2.5%
Close Punctuation 199
 
2.3%
Open Punctuation 199
 
2.3%
Decimal Number 98
 
1.1%
Math Symbol 27
 
0.3%
Dash Punctuation 6
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
80
 
3.2%
79
 
3.2%
67
 
2.7%
67
 
2.7%
60
 
2.4%
55
 
2.2%
54
 
2.2%
47
 
1.9%
43
 
1.7%
37
 
1.5%
Other values (434) 1908
76.4%
Lowercase Letter
ValueCountFrequency (%)
e 483
14.0%
a 283
 
8.2%
t 282
 
8.2%
o 279
 
8.1%
i 263
 
7.6%
r 243
 
7.0%
n 212
 
6.1%
s 204
 
5.9%
h 189
 
5.5%
l 161
 
4.7%
Other values (15) 849
24.6%
Uppercase Letter
ValueCountFrequency (%)
T 88
18.3%
M 48
10.0%
D 39
8.1%
A 38
7.9%
S 35
 
7.3%
L 31
 
6.4%
P 28
 
5.8%
F 28
 
5.8%
H 28
 
5.8%
G 24
 
5.0%
Other values (14) 94
19.5%
Decimal Number
ValueCountFrequency (%)
0 31
31.6%
1 24
24.5%
2 12
 
12.2%
3 9
 
9.2%
5 8
 
8.2%
7 5
 
5.1%
9 3
 
3.1%
8 3
 
3.1%
6 2
 
2.0%
4 1
 
1.0%
Other Punctuation
ValueCountFrequency (%)
: 89
41.2%
, 44
20.4%
' 26
 
12.0%
. 25
 
11.6%
! 20
 
9.3%
? 7
 
3.2%
· 2
 
0.9%
" 2
 
0.9%
& 1
 
0.5%
Math Symbol
ValueCountFrequency (%)
= 19
70.4%
+ 6
 
22.2%
> 1
 
3.7%
< 1
 
3.7%
Close Punctuation
ValueCountFrequency (%)
) 147
73.9%
] 52
 
26.1%
Open Punctuation
ValueCountFrequency (%)
( 147
73.9%
[ 52
 
26.1%
Space Separator
ValueCountFrequency (%)
1550
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 6
100.0%

Most occurring scripts

ValueCountFrequency (%)
Latin 3929
45.1%
Hangul 2496
28.6%
Common 2295
26.3%
Han 1
 
< 0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
80
 
3.2%
79
 
3.2%
67
 
2.7%
67
 
2.7%
60
 
2.4%
55
 
2.2%
54
 
2.2%
47
 
1.9%
43
 
1.7%
37
 
1.5%
Other values (433) 1907
76.4%
Latin
ValueCountFrequency (%)
e 483
 
12.3%
a 283
 
7.2%
t 282
 
7.2%
o 279
 
7.1%
i 263
 
6.7%
r 243
 
6.2%
n 212
 
5.4%
s 204
 
5.2%
h 189
 
4.8%
l 161
 
4.1%
Other values (39) 1330
33.9%
Common
ValueCountFrequency (%)
1550
67.5%
) 147
 
6.4%
( 147
 
6.4%
: 89
 
3.9%
[ 52
 
2.3%
] 52
 
2.3%
, 44
 
1.9%
0 31
 
1.4%
' 26
 
1.1%
. 25
 
1.1%
Other values (19) 132
 
5.8%
Han
ValueCountFrequency (%)
1
100.0%

Most occurring blocks

ValueCountFrequency (%)
ASCII 6222
71.3%
Hangul 2496
28.6%
None 2
 
< 0.1%
CJK 1
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
1550
24.9%
e 483
 
7.8%
a 283
 
4.5%
t 282
 
4.5%
o 279
 
4.5%
i 263
 
4.2%
r 243
 
3.9%
n 212
 
3.4%
s 204
 
3.3%
h 189
 
3.0%
Other values (67) 2234
35.9%
Hangul
ValueCountFrequency (%)
80
 
3.2%
79
 
3.2%
67
 
2.7%
67
 
2.7%
60
 
2.4%
55
 
2.2%
54
 
2.2%
47
 
1.9%
43
 
1.7%
37
 
1.5%
Other values (433) 1907
76.4%
None
ValueCountFrequency (%)
· 2
100.0%
CJK
ValueCountFrequency (%)
1
100.0%
Distinct195
Distinct (%)73.6%
Missing1
Missing (%)0.4%
Memory size2.2 KiB
2023-12-12T13:49:01.760417image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length128
Median length68
Mean length31.630189
Min length3

Characters and Unicode

Total characters8382
Distinct characters314
Distinct categories8 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique146 ?
Unique (%)55.1%

Sample

1st row저작자
2nd row래리 M. 바텔스 지음 ; 위선주 옮김
3rd row프랜시스 후쿠야마 지음 ; 함규진 옮김
4th row마스다 미리 지음 ; 권남희 옮김
5th row마스다 미리 지음 ; 박정임 옮김
ValueCountFrequency (%)
by 241
 
14.3%
207
 
12.3%
지음 101
 
6.0%
illustrated 97
 
5.7%
옮김 48
 
2.8%
written 29
 
1.7%
그림 24
 
1.4%
park 21
 
1.2%
adapted 16
 
0.9%
joyce 15
 
0.9%
Other values (450) 889
52.7%
2023-12-12T13:49:02.291013image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
1425
17.0%
e 507
 
6.0%
a 472
 
5.6%
r 444
 
5.3%
t 424
 
5.1%
i 365
 
4.4%
y 337
 
4.0%
l 327
 
3.9%
n 322
 
3.8%
b 271
 
3.2%
Other values (304) 3488
41.6%

Most occurring categories

ValueCountFrequency (%)
Lowercase Letter 4824
57.6%
Space Separator 1425
 
17.0%
Other Letter 1265
 
15.1%
Uppercase Letter 589
 
7.0%
Other Punctuation 241
 
2.9%
Open Punctuation 15
 
0.2%
Close Punctuation 15
 
0.2%
Dash Punctuation 8
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
106
 
8.4%
103
 
8.1%
75
 
5.9%
48
 
3.8%
46
 
3.6%
36
 
2.8%
34
 
2.7%
28
 
2.2%
24
 
1.9%
24
 
1.9%
Other values (245) 741
58.6%
Lowercase Letter
ValueCountFrequency (%)
e 507
10.5%
a 472
9.8%
r 444
 
9.2%
t 424
 
8.8%
i 365
 
7.6%
y 337
 
7.0%
l 327
 
6.8%
n 322
 
6.7%
b 271
 
5.6%
s 219
 
4.5%
Other values (14) 1136
23.5%
Uppercase Letter
ValueCountFrequency (%)
C 55
 
9.3%
B 54
 
9.2%
A 52
 
8.8%
J 50
 
8.5%
W 43
 
7.3%
S 41
 
7.0%
M 36
 
6.1%
D 30
 
5.1%
L 25
 
4.2%
P 25
 
4.2%
Other values (14) 178
30.2%
Other Punctuation
ValueCountFrequency (%)
; 207
85.9%
. 13
 
5.4%
· 11
 
4.6%
, 8
 
3.3%
2
 
0.8%
Open Punctuation
ValueCountFrequency (%)
[ 14
93.3%
1
 
6.7%
Close Punctuation
ValueCountFrequency (%)
] 14
93.3%
1
 
6.7%
Space Separator
ValueCountFrequency (%)
1425
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 8
100.0%

Most occurring scripts

ValueCountFrequency (%)
Latin 5413
64.6%
Common 1704
 
20.3%
Hangul 1265
 
15.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
106
 
8.4%
103
 
8.1%
75
 
5.9%
48
 
3.8%
46
 
3.6%
36
 
2.8%
34
 
2.7%
28
 
2.2%
24
 
1.9%
24
 
1.9%
Other values (245) 741
58.6%
Latin
ValueCountFrequency (%)
e 507
 
9.4%
a 472
 
8.7%
r 444
 
8.2%
t 424
 
7.8%
i 365
 
6.7%
y 337
 
6.2%
l 327
 
6.0%
n 322
 
5.9%
b 271
 
5.0%
s 219
 
4.0%
Other values (38) 1725
31.9%
Common
ValueCountFrequency (%)
1425
83.6%
; 207
 
12.1%
[ 14
 
0.8%
] 14
 
0.8%
. 13
 
0.8%
· 11
 
0.6%
- 8
 
0.5%
, 8
 
0.5%
2
 
0.1%
1
 
0.1%

Most occurring blocks

ValueCountFrequency (%)
ASCII 7102
84.7%
Hangul 1265
 
15.1%
None 15
 
0.2%

Most frequent character per block

ASCII
ValueCountFrequency (%)
1425
20.1%
e 507
 
7.1%
a 472
 
6.6%
r 444
 
6.3%
t 424
 
6.0%
i 365
 
5.1%
y 337
 
4.7%
l 327
 
4.6%
n 322
 
4.5%
b 271
 
3.8%
Other values (45) 2208
31.1%
Hangul
ValueCountFrequency (%)
106
 
8.4%
103
 
8.1%
75
 
5.9%
48
 
3.8%
46
 
3.6%
36
 
2.8%
34
 
2.7%
28
 
2.2%
24
 
1.9%
24
 
1.9%
Other values (245) 741
58.6%
None
ValueCountFrequency (%)
· 11
73.3%
2
 
13.3%
1
 
6.7%
1
 
6.7%
Distinct137
Distinct (%)51.7%
Missing1
Missing (%)0.4%
Memory size2.2 KiB
2023-12-12T13:49:02.579669image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length32
Median length29
Mean length7.4150943
Min length1

Characters and Unicode

Total characters1965
Distinct characters239
Distinct categories6 ?
Distinct scripts4 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique101 ?
Unique (%)38.1%

Sample

1st row발행자
2nd row21세기북스
3rd row웅진지식하우스
4th row이봄
5th row이봄
ValueCountFrequency (%)
eplis 42
 
12.5%
books 16
 
4.8%
wonder&learn 15
 
4.5%
press 15
 
4.5%
노란우산 11
 
3.3%
walker 10
 
3.0%
dkbooks 8
 
2.4%
university 7
 
2.1%
oxford 7
 
2.1%
disney 7
 
2.1%
Other values (135) 197
58.8%
2023-12-12T13:49:03.017364image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
s 133
 
6.8%
i 107
 
5.4%
o 107
 
5.4%
e 105
 
5.3%
r 105
 
5.3%
l 98
 
5.0%
n 74
 
3.8%
70
 
3.6%
p 62
 
3.2%
a 51
 
2.6%
Other values (229) 1053
53.6%

Most occurring categories

ValueCountFrequency (%)
Lowercase Letter 1078
54.9%
Other Letter 573
29.2%
Uppercase Letter 217
 
11.0%
Space Separator 70
 
3.6%
Other Punctuation 21
 
1.1%
Decimal Number 6
 
0.3%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
29
 
5.1%
18
 
3.1%
18
 
3.1%
18
 
3.1%
16
 
2.8%
16
 
2.8%
15
 
2.6%
12
 
2.1%
12
 
2.1%
10
 
1.7%
Other values (184) 409
71.4%
Lowercase Letter
ValueCountFrequency (%)
s 133
12.3%
i 107
9.9%
o 107
9.9%
e 105
9.7%
r 105
9.7%
l 98
9.1%
n 74
 
6.9%
p 62
 
5.8%
a 51
 
4.7%
k 50
 
4.6%
Other values (14) 186
17.3%
Uppercase Letter
ValueCountFrequency (%)
E 42
19.4%
B 28
12.9%
W 26
12.0%
L 21
9.7%
D 21
9.7%
P 18
8.3%
K 14
 
6.5%
C 13
 
6.0%
O 9
 
4.1%
H 6
 
2.8%
Other values (6) 19
8.8%
Other Punctuation
ValueCountFrequency (%)
& 15
71.4%
' 6
 
28.6%
Decimal Number
ValueCountFrequency (%)
1 3
50.0%
2 3
50.0%
Space Separator
ValueCountFrequency (%)
70
100.0%

Most occurring scripts

ValueCountFrequency (%)
Latin 1295
65.9%
Hangul 570
29.0%
Common 97
 
4.9%
Han 3
 
0.2%

Most frequent character per script

Hangul
ValueCountFrequency (%)
29
 
5.1%
18
 
3.2%
18
 
3.2%
18
 
3.2%
16
 
2.8%
16
 
2.8%
15
 
2.6%
12
 
2.1%
12
 
2.1%
10
 
1.8%
Other values (181) 406
71.2%
Latin
ValueCountFrequency (%)
s 133
 
10.3%
i 107
 
8.3%
o 107
 
8.3%
e 105
 
8.1%
r 105
 
8.1%
l 98
 
7.6%
n 74
 
5.7%
p 62
 
4.8%
a 51
 
3.9%
k 50
 
3.9%
Other values (30) 403
31.1%
Common
ValueCountFrequency (%)
70
72.2%
& 15
 
15.5%
' 6
 
6.2%
1 3
 
3.1%
2 3
 
3.1%
Han
ValueCountFrequency (%)
1
33.3%
1
33.3%
1
33.3%

Most occurring blocks

ValueCountFrequency (%)
ASCII 1392
70.8%
Hangul 570
29.0%
CJK 3
 
0.2%

Most frequent character per block

ASCII
ValueCountFrequency (%)
s 133
 
9.6%
i 107
 
7.7%
o 107
 
7.7%
e 105
 
7.5%
r 105
 
7.5%
l 98
 
7.0%
n 74
 
5.3%
70
 
5.0%
p 62
 
4.5%
a 51
 
3.7%
Other values (35) 480
34.5%
Hangul
ValueCountFrequency (%)
29
 
5.1%
18
 
3.2%
18
 
3.2%
18
 
3.2%
16
 
2.8%
16
 
2.8%
15
 
2.6%
12
 
2.1%
12
 
2.1%
10
 
1.8%
Other values (181) 406
71.2%
CJK
ValueCountFrequency (%)
1
33.3%
1
33.3%
1
33.3%

Unnamed: 6
Categorical

HIGH CORRELATION 

Distinct16
Distinct (%)6.0%
Missing0
Missing (%)0.0%
Memory size2.2 KiB
2014
122 
2010
46 
2011
25 
2013
21 
2012
 
11
Other values (11)
41 

Length

Max length4
Median length4
Mean length3.9962406
Min length3

Unique

Unique5 ?
Unique (%)1.9%

Sample

1st row<NA>
2nd row발행년
3rd row2012
4th row2012
5th row2014

Common Values

ValueCountFrequency (%)
2014 122
45.9%
2010 46
 
17.3%
2011 25
 
9.4%
2013 21
 
7.9%
2012 11
 
4.1%
2007 10
 
3.8%
2009 9
 
3.4%
2008 9
 
3.4%
2006 4
 
1.5%
2005 2
 
0.8%
Other values (6) 7
 
2.6%

Length

2023-12-12T13:49:03.141884image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category
ValueCountFrequency (%)
2014 122
45.9%
2010 46
 
17.3%
2011 25
 
9.4%
2013 21
 
7.9%
2012 11
 
4.1%
2007 10
 
3.8%
2009 9
 
3.4%
2008 9
 
3.4%
2006 4
 
1.5%
2005 2
 
0.8%
Other values (6) 7
 
2.6%
Distinct66
Distinct (%)24.9%
Missing1
Missing (%)0.4%
Memory size2.2 KiB
2023-12-12T13:49:03.324533image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length5
Median length5
Mean length4.0792453
Min length1

Characters and Unicode

Total characters1081
Distinct characters12
Distinct categories2 ?
Distinct scripts2 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique35 ?
Unique (%)13.2%

Sample

1st row가격
2nd row25000
3rd row30000
4th row12500
5th row12500
ValueCountFrequency (%)
0 52
19.6%
11000 23
 
8.7%
13000 22
 
8.3%
15000 21
 
7.9%
9000 12
 
4.5%
9800 11
 
4.2%
14000 9
 
3.4%
12000 8
 
3.0%
10000 8
 
3.0%
15800 6
 
2.3%
Other values (56) 93
35.1%
2023-12-12T13:49:03.657832image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
0 604
55.9%
1 194
 
17.9%
5 54
 
5.0%
3 45
 
4.2%
9 45
 
4.2%
8 43
 
4.0%
2 40
 
3.7%
4 29
 
2.7%
6 18
 
1.7%
7 7
 
0.6%
Other values (2) 2
 
0.2%

Most occurring categories

ValueCountFrequency (%)
Decimal Number 1079
99.8%
Other Letter 2
 
0.2%

Most frequent character per category

Decimal Number
ValueCountFrequency (%)
0 604
56.0%
1 194
 
18.0%
5 54
 
5.0%
3 45
 
4.2%
9 45
 
4.2%
8 43
 
4.0%
2 40
 
3.7%
4 29
 
2.7%
6 18
 
1.7%
7 7
 
0.6%
Other Letter
ValueCountFrequency (%)
1
50.0%
1
50.0%

Most occurring scripts

ValueCountFrequency (%)
Common 1079
99.8%
Hangul 2
 
0.2%

Most frequent character per script

Common
ValueCountFrequency (%)
0 604
56.0%
1 194
 
18.0%
5 54
 
5.0%
3 45
 
4.2%
9 45
 
4.2%
8 43
 
4.0%
2 40
 
3.7%
4 29
 
2.7%
6 18
 
1.7%
7 7
 
0.6%
Hangul
ValueCountFrequency (%)
1
50.0%
1
50.0%

Most occurring blocks

ValueCountFrequency (%)
ASCII 1079
99.8%
Hangul 2
 
0.2%

Most frequent character per block

ASCII
ValueCountFrequency (%)
0 604
56.0%
1 194
 
18.0%
5 54
 
5.0%
3 45
 
4.2%
9 45
 
4.2%
8 43
 
4.0%
2 40
 
3.7%
4 29
 
2.7%
6 18
 
1.7%
7 7
 
0.6%
Hangul
ValueCountFrequency (%)
1
50.0%
1
50.0%

Unnamed: 8
Text

MISSING 

Distinct162
Distinct (%)76.1%
Missing53
Missing (%)19.9%
Memory size2.2 KiB
2023-12-12T13:49:03.906214image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length42
Median length33
Mean length17.84507
Min length4

Characters and Unicode

Total characters3801
Distinct characters55
Distinct categories9 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique146 ?
Unique (%)68.5%

Sample

1st row형태사항
2nd row491p.:도표;23cm
3rd row597p.:삽도;24cm
4th row207p.:삽도;20cm
5th row207p.:삽도;20cm
ValueCountFrequency (%)
1매 61
 
15.2%
1v.:col 40
 
10.0%
1 23
 
5.7%
v.:col 23
 
5.7%
ill.;23cm+오디오cd 20
 
5.0%
1책:삽도;22x22cm+오디오cd 10
 
2.5%
ill.;28cm 6
 
1.5%
ill.;29cm+오디오cd 6
 
1.5%
ill 5
 
1.2%
score;29cm+dvd 5
 
1.2%
Other values (170) 203
50.5%
2023-12-12T13:49:04.293550image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
. 355
 
9.3%
2 336
 
8.8%
c 299
 
7.9%
l 250
 
6.6%
1 236
 
6.2%
; 212
 
5.6%
m 212
 
5.6%
189
 
5.0%
: 168
 
4.4%
3 138
 
3.6%
Other values (45) 1406
37.0%

Most occurring categories

ValueCountFrequency (%)
Lowercase Letter 1181
31.1%
Decimal Number 1002
26.4%
Other Punctuation 758
19.9%
Other Letter 448
 
11.8%
Space Separator 189
 
5.0%
Uppercase Letter 141
 
3.7%
Math Symbol 56
 
1.5%
Open Punctuation 13
 
0.3%
Close Punctuation 13
 
0.3%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
98
21.9%
72
16.1%
68
15.2%
61
13.6%
49
10.9%
23
 
5.1%
22
 
4.9%
18
 
4.0%
10
 
2.2%
8
 
1.8%
Other values (7) 19
 
4.2%
Lowercase Letter
ValueCountFrequency (%)
c 299
25.3%
l 250
21.2%
m 212
18.0%
p 133
11.3%
o 89
 
7.5%
i 84
 
7.1%
v 63
 
5.3%
x 30
 
2.5%
r 6
 
0.5%
e 5
 
0.4%
Other values (5) 10
 
0.8%
Decimal Number
ValueCountFrequency (%)
2 336
33.5%
1 236
23.6%
3 138
13.8%
5 54
 
5.4%
4 48
 
4.8%
9 42
 
4.2%
6 41
 
4.1%
0 41
 
4.1%
8 34
 
3.4%
7 32
 
3.2%
Uppercase Letter
ValueCountFrequency (%)
D 66
46.8%
C 56
39.7%
M 7
 
5.0%
P 7
 
5.0%
V 5
 
3.5%
Other Punctuation
ValueCountFrequency (%)
. 355
46.8%
; 212
28.0%
: 168
22.2%
, 23
 
3.0%
Space Separator
ValueCountFrequency (%)
189
100.0%
Math Symbol
ValueCountFrequency (%)
+ 56
100.0%
Open Punctuation
ValueCountFrequency (%)
[ 13
100.0%
Close Punctuation
ValueCountFrequency (%)
] 13
100.0%

Most occurring scripts

ValueCountFrequency (%)
Common 2031
53.4%
Latin 1322
34.8%
Hangul 448
 
11.8%

Most frequent character per script

Latin
ValueCountFrequency (%)
c 299
22.6%
l 250
18.9%
m 212
16.0%
p 133
10.1%
o 89
 
6.7%
i 84
 
6.4%
D 66
 
5.0%
v 63
 
4.8%
C 56
 
4.2%
x 30
 
2.3%
Other values (10) 40
 
3.0%
Common
ValueCountFrequency (%)
. 355
17.5%
2 336
16.5%
1 236
11.6%
; 212
10.4%
189
9.3%
: 168
8.3%
3 138
 
6.8%
+ 56
 
2.8%
5 54
 
2.7%
4 48
 
2.4%
Other values (8) 239
11.8%
Hangul
ValueCountFrequency (%)
98
21.9%
72
16.1%
68
15.2%
61
13.6%
49
10.9%
23
 
5.1%
22
 
4.9%
18
 
4.0%
10
 
2.2%
8
 
1.8%
Other values (7) 19
 
4.2%

Most occurring blocks

ValueCountFrequency (%)
ASCII 3353
88.2%
Hangul 448
 
11.8%

Most frequent character per block

ASCII
ValueCountFrequency (%)
. 355
 
10.6%
2 336
 
10.0%
c 299
 
8.9%
l 250
 
7.5%
1 236
 
7.0%
; 212
 
6.3%
m 212
 
6.3%
189
 
5.6%
: 168
 
5.0%
3 138
 
4.1%
Other values (28) 958
28.6%
Hangul
ValueCountFrequency (%)
98
21.9%
72
16.1%
68
15.2%
61
13.6%
49
10.9%
23
 
5.1%
22
 
4.9%
18
 
4.0%
10
 
2.2%
8
 
1.8%
Other values (7) 19
 
4.2%

Unnamed: 9
Categorical

HIGH CORRELATION  IMBALANCE 

Distinct4
Distinct (%)1.5%
Missing0
Missing (%)0.0%
Memory size2.2 KiB
유아어린이자료실
172 
종합자료실
92 
<NA>
 
1
자료실명
 
1

Length

Max length8
Median length8
Mean length6.9323308
Min length4

Unique

Unique2 ?
Unique (%)0.8%

Sample

1st row<NA>
2nd row자료실명
3rd row종합자료실
4th row종합자료실
5th row종합자료실

Common Values

ValueCountFrequency (%)
유아어린이자료실 172
64.7%
종합자료실 92
34.6%
<NA> 1
 
0.4%
자료실명 1
 
0.4%

Length

2023-12-12T13:49:04.434633image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T13:49:04.560437image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
유아어린이자료실 172
64.7%
종합자료실 92
34.6%
na 1
 
0.4%
자료실명 1
 
0.4%

Interactions

2023-12-12T13:48:57.551350image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-12T13:49:04.639488image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
[2014년 12월] 중랑구립면목정보도서관 신착자료 목록Unnamed: 6Unnamed: 7Unnamed: 9
[2014년 12월] 중랑구립면목정보도서관 신착자료 목록1.0000.5890.8680.967
Unnamed: 60.5891.0000.9400.977
Unnamed: 70.8680.9401.0000.986
Unnamed: 90.9670.9770.9861.000
2023-12-12T13:49:04.751081image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Unnamed: 9Unnamed: 6
Unnamed: 91.0000.805
Unnamed: 60.8051.000
2023-12-12T13:49:04.853898image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
[2014년 12월] 중랑구립면목정보도서관 신착자료 목록Unnamed: 6Unnamed: 9
[2014년 12월] 중랑구립면목정보도서관 신착자료 목록1.0000.2810.831
Unnamed: 60.2811.0000.805
Unnamed: 90.8310.8051.000

Missing values

2023-12-12T13:48:57.735840image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-12T13:48:57.907072image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.
2023-12-12T13:48:58.382223image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
The correlation heatmap measures nullity correlation: how strongly the presence or absence of one variable affects the presence of another.

Sample

[2014년 12월] 중랑구립면목정보도서관 신착자료 목록Unnamed: 1Unnamed: 2Unnamed: 3Unnamed: 4Unnamed: 5Unnamed: 6Unnamed: 7Unnamed: 8Unnamed: 9
0<NA><NA><NA><NA><NA><NA><NA><NA><NA><NA>
1<NA>등록번호청구기호서명저작자발행자발행년가격형태사항자료실명
21DM0000085457NB 320.942-ㅂ158ㅂ불평등 민주주의 : 자유에 가려진 진실래리 M. 바텔스 지음 ; 위선주 옮김21세기북스201225000491p.:도표;23cm종합자료실
32DM0000085458NB 340.9-ㅎ899ㅈ정치 질서의 기원 : 불안정성을 극복할 정치적 힘은 어디서 오는가프랜시스 후쿠야마 지음 ; 함규진 옮김웅진지식하우스201230000597p.:삽도;24cm종합자료실
43DM0000085454NB 834-ㅁ176ㅇ여자라는 생물 : 마스다 미리 에세이마스다 미리 지음 ; 권남희 옮김이봄201412500207p.:삽도;20cm종합자료실
54DM0000085455NB 834-ㅁ176ㄴ나는 사랑을 하고 있어 : 마스다 미리 에세이마스다 미리 지음 ; 박정임 옮김이봄201412500207p.:삽도;20cm종합자료실
65DM0000085460NB 304-ㅇ296ㄷ단속사회 : 쉴 새 없이 접속하고 끊임없이 차단한다엄기호 지음창비201415000306p.;22cm종합자료실
76DM0000085461NB 004.583-ㄱ318ㅇ(Do it!) HTML + CSS3 웹 표준의 정석 : 기초부터 반응형 웹까지! HTML 권위자에게 정석으로 배워라!고경희 지음이지스퍼블리싱201432000700p.:천연색삽도;27cm종합자료실
87DM0000085462NB 189-ㅍ28ㄲ끌리는 얼굴은 무엇이 다른가 : 과학으로 밝혀낸 매력의 비밀데이비드 페렛 지음 ; 박여진 옮김엘도라도201416800487p.:삽도, 사진;23cm종합자료실
98DM0000085463NB 234.9-ㅈ784ㅇ엄마 마음 내려놓기 : 빵점 엄마 주견자 사모의 맡기는 교육주견자 지음두란노서원201410000244p.;21cm종합자료실
[2014년 12월] 중랑구립면목정보도서관 신착자료 목록Unnamed: 1Unnamed: 2Unnamed: 3Unnamed: 4Unnamed: 5Unnamed: 6Unnamed: 7Unnamed: 8Unnamed: 9
256255DM0000086264NB 181.845-ㄱ635ㅅ습관의 재발견 : 기적 같은 변화를 불러오는 작은 습관의 힘스티븐 기즈 지음 ; 구세희 옮김비즈니스북스201413000239p.:삽도;21cm종합자료실
257256DM0000086265BB 808.9-ㅇ175우-v.88머나먼 여행에런 베커 글·그림웅진주니어201412000[49]p.:삽도;25x28cm유아어린이자료실
258257DM0000086266NB 150-ㅅ898ㄷ동양철학 인생과 맞짱 뜨다 : 삶의 지혜를 넘어 도전의 철학으로신정근 지음21세기북스201417000418p.:삽도;22cm종합자료실
259258DM0000086267NB 322.01-ㅎ145ㅈ자본의 17가지 모순 : 이 시대 자본주의의 위기와 대안데이비드 하비 지음 ; 황성원 옮김동녘201419800464p.;23cm종합자료실
260259DM0000086268NB 182.2-ㅈ526ㅈ-v.2사람 VS 사람정혜신 지음개마고원200910000319p.:삽도;23cm종합자료실
261260DM0000086269NB 189-ㄱ619ㅁ미움받을 용기 : 자유롭고 행복한 삶을 위한 아들러의 가르침기시미 이치로, 고가 후미타케 [같이] 지음 ; 전경아 옮김인플루엔셜201414900331p.;21cm종합자료실
262261DM0000086270NB 813.6-ㅇ754ㅎ한복 입은 남자 = A Man in Korean Costume : 이상훈 장편소설이상훈 지음박하201413000536p.;20cm종합자료실
263262DM0000086271NB 558.92-ㅎ435ㅇ우주비행사의 지구생활 안내서 : 나는 우주정거장에서 인생을 배웠다크리스 해드필드 지음 ; 노태복 옮김더퀘스트201414500336p.:삽도;23cm종합자료실
264263DM0000086272BB 808.9-ㅍ84ㅋ-v.153그림자가 사는 마을마이클 바틀로스 글·그림 ; 김영미 옮김키즈엠201410000[30]p.:삽도;26cm유아어린이자료실
265264DM0000086273JB 808.9-ㅂ596ㅁ-v.33탄탄동 사거리 만복전파사김려령 지음 ; 조승연 그림문학동네201410000121p.:삽도;23cm유아어린이자료실