Overview

Dataset statistics

Number of variables9
Number of observations3176
Missing cells2152
Missing cells (%)7.5%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory226.5 KiB
Average record size in memory73.0 B

Variable types

Text7
Numeric1
Categorical1

Dataset

Description한국방송광고진흥공사 광고박물관이 소장하고 있는 분야별 광고소재의 디지털 아카이브 데이터로 라디오광고 관련 항목들을 제공합니다.
Author한국방송광고진흥공사
URLhttps://www.data.go.kr/data/15044289/fileData.do

Alerts

년도 has 1343 (42.3%) missing valuesMissing
대행사 has 803 (25.3%) missing valuesMissing
파일명 has unique valuesUnique

Reproduction

Analysis started2023-12-11 22:53:07.179971
Analysis finished2023-12-11 22:53:09.066868
Duration1.89 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

파일명
Text

UNIQUE 

Distinct3176
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Memory size24.9 KiB
2023-12-12T07:53:09.268647image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length17
Median length12
Mean length12.349496
Min length12

Characters and Unicode

Total characters39222
Distinct characters29
Distinct categories5 ?
Distinct scripts2 ?
Distinct blocks1 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique3176 ?
Unique (%)100.0%

Sample

1st rowA1A00001.wma
2nd rowA1A00004.wma
3rd rowA1A00005.wma
4th rowA1A00006.wma
5th rowA1A00007.wma
ValueCountFrequency (%)
a1a00001.wma 1
 
< 0.1%
a9a00062.wma 1
 
< 0.1%
radio2000_125.mp3 1
 
< 0.1%
a9a00077.wma 1
 
< 0.1%
a9a00078.wma 1
 
< 0.1%
a9a00086.wma 1
 
< 0.1%
a9a00090.wma 1
 
< 0.1%
a9a00091.wma 1
 
< 0.1%
a9a00092.wma 1
 
< 0.1%
a9a00096.wma 1
 
< 0.1%
Other values (3166) 3166
99.7%
2023-12-12T07:53:09.703782image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
0 9625
24.5%
A 4875
12.4%
. 3176
 
8.1%
m 3176
 
8.1%
a 3176
 
8.1%
w 2954
 
7.5%
2 1853
 
4.7%
1 1531
 
3.9%
3 1518
 
3.9%
6 1494
 
3.8%
Other values (19) 5844
14.9%

Most occurring categories

ValueCountFrequency (%)
Decimal Number 19754
50.4%
Lowercase Letter 10416
26.6%
Uppercase Letter 5654
 
14.4%
Other Punctuation 3176
 
8.1%
Connector Punctuation 222
 
0.6%

Most frequent character per category

Decimal Number
ValueCountFrequency (%)
0 9625
48.7%
2 1853
 
9.4%
1 1531
 
7.8%
3 1518
 
7.7%
6 1494
 
7.6%
4 938
 
4.7%
5 923
 
4.7%
9 646
 
3.3%
8 639
 
3.2%
7 587
 
3.0%
Uppercase Letter
ValueCountFrequency (%)
A 4875
86.2%
F 393
 
7.0%
K 119
 
2.1%
B 116
 
2.1%
H 96
 
1.7%
J 29
 
0.5%
G 14
 
0.2%
C 7
 
0.1%
I 5
 
0.1%
Lowercase Letter
ValueCountFrequency (%)
m 3176
30.5%
a 3176
30.5%
w 2954
28.4%
p 222
 
2.1%
o 222
 
2.1%
i 222
 
2.1%
d 222
 
2.1%
r 222
 
2.1%
Other Punctuation
ValueCountFrequency (%)
. 3176
100.0%
Connector Punctuation
ValueCountFrequency (%)
_ 222
100.0%

Most occurring scripts

ValueCountFrequency (%)
Common 23152
59.0%
Latin 16070
41.0%

Most frequent character per script

Latin
ValueCountFrequency (%)
A 4875
30.3%
m 3176
19.8%
a 3176
19.8%
w 2954
18.4%
F 393
 
2.4%
p 222
 
1.4%
o 222
 
1.4%
i 222
 
1.4%
d 222
 
1.4%
r 222
 
1.4%
Other values (7) 386
 
2.4%
Common
ValueCountFrequency (%)
0 9625
41.6%
. 3176
 
13.7%
2 1853
 
8.0%
1 1531
 
6.6%
3 1518
 
6.6%
6 1494
 
6.5%
4 938
 
4.1%
5 923
 
4.0%
9 646
 
2.8%
8 639
 
2.8%
Other values (2) 809
 
3.5%

Most occurring blocks

ValueCountFrequency (%)
ASCII 39222
100.0%

Most frequent character per block

ASCII
ValueCountFrequency (%)
0 9625
24.5%
A 4875
12.4%
. 3176
 
8.1%
m 3176
 
8.1%
a 3176
 
8.1%
w 2954
 
7.5%
2 1853
 
4.7%
1 1531
 
3.9%
3 1518
 
3.9%
6 1494
 
3.8%
Other values (19) 5844
14.9%

년도
Real number (ℝ)

MISSING 

Distinct29
Distinct (%)1.6%
Missing1343
Missing (%)42.3%
Infinite0
Infinite (%)0.0%
Mean1991.934
Minimum1969
Maximum1998
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size28.0 KiB
2023-12-12T07:53:09.856499image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1969
5-th percentile1980
Q11992
median1993
Q31995
95-th percentile1997
Maximum1998
Range29
Interquartile range (IQR)3

Descriptive statistics

Standard deviation4.9561953
Coefficient of variation (CV)0.0024881323
Kurtosis3.1322813
Mean1991.934
Median Absolute Deviation (MAD)2
Skewness-1.8247131
Sum3651215
Variance24.563871
MonotonicityNot monotonic
2023-12-12T07:53:09.960825image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=29)
ValueCountFrequency (%)
1993 295
 
9.3%
1994 291
 
9.2%
1995 236
 
7.4%
1996 233
 
7.3%
1992 231
 
7.3%
1997 113
 
3.6%
1991 87
 
2.7%
1985 57
 
1.8%
1986 37
 
1.2%
1982 30
 
0.9%
Other values (19) 223
 
7.0%
(Missing) 1343
42.3%
ValueCountFrequency (%)
1969 1
 
< 0.1%
1971 4
 
0.1%
1972 12
0.4%
1973 1
 
< 0.1%
1974 2
 
0.1%
1975 1
 
< 0.1%
1976 6
 
0.2%
1977 2
 
0.1%
1978 13
0.4%
1979 29
0.9%
ValueCountFrequency (%)
1998 1
 
< 0.1%
1997 113
 
3.6%
1996 233
7.3%
1995 236
7.4%
1994 291
9.2%
1993 295
9.3%
1992 231
7.3%
1991 87
 
2.7%
1990 20
 
0.6%
1989 14
 
0.4%
Distinct847
Distinct (%)26.7%
Missing0
Missing (%)0.0%
Memory size24.9 KiB
2023-12-12T07:53:10.204960image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length21
Median length15
Mean length6.6665617
Min length1

Characters and Unicode

Total characters21173
Distinct characters505
Distinct categories9 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique509 ?
Unique (%)16.0%

Sample

1st row엘프
2nd row케스트롤뉴트론X
3rd row엘지칼텍스정유(주)여
4th row엘지칼텍스정유(주)여
5th row엘지칼텍스정유(주)여
ValueCountFrequency (%)
진흥기업(주)부산백화점 105
 
3.3%
주)뉴코아 105
 
3.3%
주)대우 75
 
2.3%
남영나이론 68
 
2.1%
삼성물산(주 63
 
2.0%
주)엘칸토 61
 
1.9%
동서식품(주 55
 
1.7%
제일모직(주 51
 
1.6%
제일제당 48
 
1.5%
주)희망백화점 48
 
1.5%
Other values (846) 2534
78.9%
2023-12-12T07:53:10.695781image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
2252
 
10.6%
( 2168
 
10.2%
) 2168
 
10.2%
417
 
2.0%
408
 
1.9%
309
 
1.5%
304
 
1.4%
293
 
1.4%
272
 
1.3%
270
 
1.3%
Other values (495) 12312
58.1%

Most occurring categories

ValueCountFrequency (%)
Other Letter 16702
78.9%
Open Punctuation 2168
 
10.2%
Close Punctuation 2168
 
10.2%
Uppercase Letter 65
 
0.3%
Space Separator 44
 
0.2%
Other Punctuation 12
 
0.1%
Other Symbol 10
 
< 0.1%
Dash Punctuation 2
 
< 0.1%
Lowercase Letter 2
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
2252
 
13.5%
417
 
2.5%
408
 
2.4%
309
 
1.9%
304
 
1.8%
293
 
1.8%
272
 
1.6%
270
 
1.6%
257
 
1.5%
257
 
1.5%
Other values (467) 11663
69.8%
Uppercase Letter
ValueCountFrequency (%)
S 12
18.5%
K 12
18.5%
A 6
9.2%
I 6
9.2%
V 5
7.7%
O 3
 
4.6%
C 3
 
4.6%
T 3
 
4.6%
F 3
 
4.6%
L 2
 
3.1%
Other values (8) 10
15.4%
Other Punctuation
ValueCountFrequency (%)
. 7
58.3%
& 4
33.3%
, 1
 
8.3%
Lowercase Letter
ValueCountFrequency (%)
c 1
50.0%
i 1
50.0%
Open Punctuation
ValueCountFrequency (%)
( 2168
100.0%
Close Punctuation
ValueCountFrequency (%)
) 2168
100.0%
Space Separator
ValueCountFrequency (%)
44
100.0%
Other Symbol
ValueCountFrequency (%)
10
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 2
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 16712
78.9%
Common 4394
 
20.8%
Latin 67
 
0.3%

Most frequent character per script

Hangul
ValueCountFrequency (%)
2252
 
13.5%
417
 
2.5%
408
 
2.4%
309
 
1.8%
304
 
1.8%
293
 
1.8%
272
 
1.6%
270
 
1.6%
257
 
1.5%
257
 
1.5%
Other values (468) 11673
69.8%
Latin
ValueCountFrequency (%)
S 12
17.9%
K 12
17.9%
A 6
9.0%
I 6
9.0%
V 5
7.5%
O 3
 
4.5%
C 3
 
4.5%
T 3
 
4.5%
F 3
 
4.5%
L 2
 
3.0%
Other values (10) 12
17.9%
Common
ValueCountFrequency (%)
( 2168
49.3%
) 2168
49.3%
44
 
1.0%
. 7
 
0.2%
& 4
 
0.1%
- 2
 
< 0.1%
, 1
 
< 0.1%

Most occurring blocks

ValueCountFrequency (%)
Hangul 16702
78.9%
ASCII 4461
 
21.1%
None 10
 
< 0.1%

Most frequent character per block

Hangul
ValueCountFrequency (%)
2252
 
13.5%
417
 
2.5%
408
 
2.4%
309
 
1.9%
304
 
1.8%
293
 
1.8%
272
 
1.6%
270
 
1.6%
257
 
1.5%
257
 
1.5%
Other values (467) 11663
69.8%
ASCII
ValueCountFrequency (%)
( 2168
48.6%
) 2168
48.6%
44
 
1.0%
S 12
 
0.3%
K 12
 
0.3%
. 7
 
0.2%
A 6
 
0.1%
I 6
 
0.1%
V 5
 
0.1%
& 4
 
0.1%
Other values (17) 29
 
0.7%
None
ValueCountFrequency (%)
10
100.0%

대분류
Categorical

Distinct23
Distinct (%)0.7%
Missing0
Missing (%)0.0%
Memory size24.9 KiB
패션
790 
제약 및 의료
374 
유통
370 
그룹 및 기업광고
272 
음료 및 기호식품
216 
Other values (18)
1154 

Length

Max length12
Median length11
Mean length4.934194
Min length2

Unique

Unique1 ?
Unique (%)< 0.1%

Sample

1st row기초재
2nd row기초재
3rd row기초재
4th row기초재
5th row기초재

Common Values

ValueCountFrequency (%)
패션 790
24.9%
제약 및 의료 374
11.8%
유통 370
11.6%
그룹 및 기업광고 272
 
8.6%
음료 및 기호식품 216
 
6.8%
식품 201
 
6.3%
화장품 및 보건용품 195
 
6.1%
출판 194
 
6.1%
가정용품 127
 
4.0%
서비스 105
 
3.3%
Other values (13) 332
10.5%

Length

2023-12-12T07:53:10.834999image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category
ValueCountFrequency (%)
1266
21.8%
패션 790
13.6%
제약 376
 
6.5%
의료 374
 
6.5%
유통 370
 
6.4%
그룹 272
 
4.7%
기업광고 272
 
4.7%
기호식품 216
 
3.7%
음료 216
 
3.7%
식품 201
 
3.5%
Other values (27) 1442
24.9%
Distinct125
Distinct (%)3.9%
Missing0
Missing (%)0.0%
Memory size24.9 KiB
2023-12-12T07:53:11.164456image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length15
Median length13
Mean length5.0050378
Min length2

Characters and Unicode

Total characters15896
Distinct characters172
Distinct categories4 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique16 ?
Unique (%)0.5%

Sample

1st row석탄, 석유 및 가스
2nd row석탄, 석유 및 가스
3rd row석탄, 석유 및 가스
4th row석탄, 석유 및 가스
5th row석탄, 석유 및 가스
ValueCountFrequency (%)
375
 
8.3%
대형유통 342
 
7.6%
그룹광고 272
 
6.0%
기타 245
 
5.4%
패션 203
 
4.5%
캐주얼의류 178
 
4.0%
정장의류 134
 
3.0%
신발류 116
 
2.6%
대사성의약 113
 
2.5%
비알콜음료 93
 
2.1%
Other values (149) 2430
54.0%
2023-12-12T07:53:11.553964image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
1328
 
8.4%
682
 
4.3%
623
 
3.9%
514
 
3.2%
477
 
3.0%
466
 
2.9%
455
 
2.9%
375
 
2.4%
370
 
2.3%
362
 
2.3%
Other values (162) 10244
64.4%

Most occurring categories

ValueCountFrequency (%)
Other Letter 14241
89.6%
Space Separator 1328
 
8.4%
Other Punctuation 171
 
1.1%
Uppercase Letter 156
 
1.0%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
682
 
4.8%
623
 
4.4%
514
 
3.6%
477
 
3.3%
466
 
3.3%
455
 
3.2%
375
 
2.6%
370
 
2.6%
362
 
2.5%
360
 
2.5%
Other values (157) 9557
67.1%
Other Punctuation
ValueCountFrequency (%)
, 93
54.4%
/ 78
45.6%
Uppercase Letter
ValueCountFrequency (%)
S 78
50.0%
W 78
50.0%
Space Separator
ValueCountFrequency (%)
1328
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 14241
89.6%
Common 1499
 
9.4%
Latin 156
 
1.0%

Most frequent character per script

Hangul
ValueCountFrequency (%)
682
 
4.8%
623
 
4.4%
514
 
3.6%
477
 
3.3%
466
 
3.3%
455
 
3.2%
375
 
2.6%
370
 
2.6%
362
 
2.5%
360
 
2.5%
Other values (157) 9557
67.1%
Common
ValueCountFrequency (%)
1328
88.6%
, 93
 
6.2%
/ 78
 
5.2%
Latin
ValueCountFrequency (%)
S 78
50.0%
W 78
50.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 14241
89.6%
ASCII 1655
 
10.4%

Most frequent character per block

ASCII
ValueCountFrequency (%)
1328
80.2%
, 93
 
5.6%
S 78
 
4.7%
W 78
 
4.7%
/ 78
 
4.7%
Hangul
ValueCountFrequency (%)
682
 
4.8%
623
 
4.4%
514
 
3.6%
477
 
3.3%
466
 
3.3%
455
 
3.2%
375
 
2.6%
370
 
2.6%
362
 
2.5%
360
 
2.5%
Other values (157) 9557
67.1%
Distinct375
Distinct (%)11.8%
Missing0
Missing (%)0.0%
Memory size24.9 KiB
2023-12-12T07:53:11.863282image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length17
Median length15
Mean length5.2267003
Min length1

Characters and Unicode

Total characters16600
Distinct characters355
Distinct categories5 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique111 ?
Unique (%)3.5%

Sample

1st row윤활유
2nd row연료유
3rd row연료유
4th row연료유
5th row연료유
ValueCountFrequency (%)
백화점 329
 
7.2%
274
 
6.0%
그룹pr 272
 
6.0%
기업pr 218
 
4.8%
패션 182
 
4.0%
캐주얼웨어 159
 
3.5%
기타 148
 
3.2%
신사정장 90
 
2.0%
s/w 66
 
1.4%
신사숙녀화 65
 
1.4%
Other values (437) 2766
60.5%
2023-12-12T07:53:12.270011image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
1393
 
8.4%
667
 
4.0%
602
 
3.6%
P 485
 
2.9%
R 458
 
2.8%
385
 
2.3%
359
 
2.2%
338
 
2.0%
314
 
1.9%
274
 
1.7%
Other values (345) 11325
68.2%

Most occurring categories

ValueCountFrequency (%)
Other Letter 13871
83.6%
Space Separator 1393
 
8.4%
Uppercase Letter 1143
 
6.9%
Other Punctuation 125
 
0.8%
Lowercase Letter 68
 
0.4%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
667
 
4.8%
602
 
4.3%
385
 
2.8%
359
 
2.6%
338
 
2.4%
314
 
2.3%
274
 
2.0%
272
 
2.0%
272
 
2.0%
271
 
2.0%
Other values (333) 10117
72.9%
Uppercase Letter
ValueCountFrequency (%)
P 485
42.4%
R 458
40.1%
W 77
 
6.7%
S 77
 
6.7%
C 38
 
3.3%
T 4
 
0.3%
V 4
 
0.3%
Other Punctuation
ValueCountFrequency (%)
/ 77
61.6%
, 48
38.4%
Lowercase Letter
ValueCountFrequency (%)
r 34
50.0%
p 34
50.0%
Space Separator
ValueCountFrequency (%)
1393
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 13871
83.6%
Common 1518
 
9.1%
Latin 1211
 
7.3%

Most frequent character per script

Hangul
ValueCountFrequency (%)
667
 
4.8%
602
 
4.3%
385
 
2.8%
359
 
2.6%
338
 
2.4%
314
 
2.3%
274
 
2.0%
272
 
2.0%
272
 
2.0%
271
 
2.0%
Other values (333) 10117
72.9%
Latin
ValueCountFrequency (%)
P 485
40.0%
R 458
37.8%
W 77
 
6.4%
S 77
 
6.4%
C 38
 
3.1%
r 34
 
2.8%
p 34
 
2.8%
T 4
 
0.3%
V 4
 
0.3%
Common
ValueCountFrequency (%)
1393
91.8%
/ 77
 
5.1%
, 48
 
3.2%

Most occurring blocks

ValueCountFrequency (%)
Hangul 13871
83.6%
ASCII 2729
 
16.4%

Most frequent character per block

ASCII
ValueCountFrequency (%)
1393
51.0%
P 485
 
17.8%
R 458
 
16.8%
W 77
 
2.8%
/ 77
 
2.8%
S 77
 
2.8%
, 48
 
1.8%
C 38
 
1.4%
r 34
 
1.2%
p 34
 
1.2%
Other values (2) 8
 
0.3%
Hangul
ValueCountFrequency (%)
667
 
4.8%
602
 
4.3%
385
 
2.8%
359
 
2.6%
338
 
2.4%
314
 
2.3%
274
 
2.0%
272
 
2.0%
272
 
2.0%
271
 
2.0%
Other values (333) 10117
72.9%
Distinct1789
Distinct (%)56.3%
Missing0
Missing (%)0.0%
Memory size24.9 KiB
2023-12-12T07:53:12.510093image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length18
Median length16
Mean length4.8743703
Min length1

Characters and Unicode

Total characters15481
Distinct characters750
Distinct categories9 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique1378 ?
Unique (%)43.4%

Sample

1st row엘프
2nd row케스트롤뉴트론X
3rd row테크론
4th row테크론
5th row테크론
ValueCountFrequency (%)
부산백화점 73
 
2.2%
대우그룹캠페인 67
 
2.1%
인천희망백화점 47
 
1.4%
엘칸토 40
 
1.2%
하티스트 21
 
0.6%
옴파로스 21
 
0.6%
기업pr 21
 
0.6%
모닝글로리 20
 
0.6%
유니온베이 20
 
0.6%
프로월드컵 19
 
0.6%
Other values (1831) 2905
89.3%
2023-12-12T07:53:12.854252image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
561
 
3.6%
423
 
2.7%
278
 
1.8%
268
 
1.7%
267
 
1.7%
261
 
1.7%
245
 
1.6%
205
 
1.3%
196
 
1.3%
192
 
1.2%
Other values (740) 12585
81.3%

Most occurring categories

ValueCountFrequency (%)
Other Letter 14676
94.8%
Uppercase Letter 442
 
2.9%
Decimal Number 148
 
1.0%
Space Separator 79
 
0.5%
Close Punctuation 40
 
0.3%
Open Punctuation 40
 
0.3%
Lowercase Letter 24
 
0.2%
Other Punctuation 21
 
0.1%
Dash Punctuation 11
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
561
 
3.8%
423
 
2.9%
278
 
1.9%
268
 
1.8%
267
 
1.8%
261
 
1.8%
245
 
1.7%
205
 
1.4%
196
 
1.3%
192
 
1.3%
Other values (684) 11780
80.3%
Uppercase Letter
ValueCountFrequency (%)
P 67
15.2%
R 63
14.3%
S 28
 
6.3%
F 27
 
6.1%
C 25
 
5.7%
O 23
 
5.2%
L 22
 
5.0%
B 19
 
4.3%
I 19
 
4.3%
G 18
 
4.1%
Other values (15) 131
29.6%
Lowercase Letter
ValueCountFrequency (%)
e 4
16.7%
r 3
12.5%
n 3
12.5%
o 3
12.5%
u 2
8.3%
i 2
8.3%
p 2
8.3%
y 1
 
4.2%
f 1
 
4.2%
t 1
 
4.2%
Other values (2) 2
8.3%
Decimal Number
ValueCountFrequency (%)
0 41
27.7%
1 34
23.0%
2 16
 
10.8%
5 14
 
9.5%
3 13
 
8.8%
7 9
 
6.1%
9 8
 
5.4%
4 7
 
4.7%
8 4
 
2.7%
6 2
 
1.4%
Other Punctuation
ValueCountFrequency (%)
& 8
38.1%
, 6
28.6%
. 5
23.8%
/ 1
 
4.8%
' 1
 
4.8%
Space Separator
ValueCountFrequency (%)
79
100.0%
Close Punctuation
ValueCountFrequency (%)
) 40
100.0%
Open Punctuation
ValueCountFrequency (%)
( 40
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 11
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 14676
94.8%
Latin 466
 
3.0%
Common 339
 
2.2%

Most frequent character per script

Hangul
ValueCountFrequency (%)
561
 
3.8%
423
 
2.9%
278
 
1.9%
268
 
1.8%
267
 
1.8%
261
 
1.8%
245
 
1.7%
205
 
1.4%
196
 
1.3%
192
 
1.3%
Other values (684) 11780
80.3%
Latin
ValueCountFrequency (%)
P 67
14.4%
R 63
13.5%
S 28
 
6.0%
F 27
 
5.8%
C 25
 
5.4%
O 23
 
4.9%
L 22
 
4.7%
B 19
 
4.1%
I 19
 
4.1%
G 18
 
3.9%
Other values (27) 155
33.3%
Common
ValueCountFrequency (%)
79
23.3%
0 41
12.1%
) 40
11.8%
( 40
11.8%
1 34
10.0%
2 16
 
4.7%
5 14
 
4.1%
3 13
 
3.8%
- 11
 
3.2%
7 9
 
2.7%
Other values (9) 42
12.4%

Most occurring blocks

ValueCountFrequency (%)
Hangul 14676
94.8%
ASCII 805
 
5.2%

Most frequent character per block

Hangul
ValueCountFrequency (%)
561
 
3.8%
423
 
2.9%
278
 
1.9%
268
 
1.8%
267
 
1.8%
261
 
1.8%
245
 
1.7%
205
 
1.4%
196
 
1.3%
192
 
1.3%
Other values (684) 11780
80.3%
ASCII
ValueCountFrequency (%)
79
 
9.8%
P 67
 
8.3%
R 63
 
7.8%
0 41
 
5.1%
) 40
 
5.0%
( 40
 
5.0%
1 34
 
4.2%
S 28
 
3.5%
F 27
 
3.4%
C 25
 
3.1%
Other values (46) 361
44.8%

제목
Text

Distinct2921
Distinct (%)92.1%
Missing6
Missing (%)0.2%
Memory size24.9 KiB
2023-12-12T07:53:13.138832image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length36
Median length28
Mean length6.9646688
Min length1

Characters and Unicode

Total characters22078
Distinct characters930
Distinct categories10 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique2803 ?
Unique (%)88.4%

Sample

1st row윤활류
2nd row한국의명차들과함께
3rd row어서오세요
4th rowLG정유라고해주세요
5th row장애인돕기에참여하세요
ValueCountFrequency (%)
젊음에게 66
 
1.6%
38
 
0.9%
하이웨이 18
 
0.4%
전지점 16
 
0.4%
순천점 16
 
0.4%
16
 
0.4%
가격인하 13
 
0.3%
인터넷 10
 
0.2%
세일 9
 
0.2%
봄정기바겐세일 8
 
0.2%
Other values (3584) 3926
94.9%
2023-12-12T07:53:13.562118image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
968
 
4.4%
455
 
2.1%
374
 
1.7%
362
 
1.6%
359
 
1.6%
338
 
1.5%
309
 
1.4%
298
 
1.3%
288
 
1.3%
277
 
1.3%
Other values (920) 18050
81.8%

Most occurring categories

ValueCountFrequency (%)
Other Letter 20403
92.4%
Space Separator 968
 
4.4%
Decimal Number 344
 
1.6%
Uppercase Letter 94
 
0.4%
Other Punctuation 78
 
0.4%
Open Punctuation 71
 
0.3%
Close Punctuation 71
 
0.3%
Dash Punctuation 29
 
0.1%
Lowercase Letter 18
 
0.1%
Math Symbol 2
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
455
 
2.2%
374
 
1.8%
362
 
1.8%
359
 
1.8%
338
 
1.7%
309
 
1.5%
298
 
1.5%
288
 
1.4%
277
 
1.4%
236
 
1.2%
Other values (866) 17107
83.8%
Uppercase Letter
ValueCountFrequency (%)
P 12
12.8%
R 10
 
10.6%
S 8
 
8.5%
A 6
 
6.4%
G 6
 
6.4%
N 5
 
5.3%
C 5
 
5.3%
T 5
 
5.3%
I 5
 
5.3%
K 4
 
4.3%
Other values (13) 28
29.8%
Lowercase Letter
ValueCountFrequency (%)
w 4
22.2%
k 2
11.1%
s 2
11.1%
p 2
11.1%
r 2
11.1%
c 1
 
5.6%
o 1
 
5.6%
m 1
 
5.6%
v 1
 
5.6%
i 1
 
5.6%
Decimal Number
ValueCountFrequency (%)
1 67
19.5%
0 62
18.0%
2 50
14.5%
3 40
11.6%
5 30
8.7%
9 28
8.1%
4 23
 
6.7%
8 18
 
5.2%
7 17
 
4.9%
6 9
 
2.6%
Other Punctuation
ValueCountFrequency (%)
, 57
73.1%
% 13
 
16.7%
. 5
 
6.4%
2
 
2.6%
& 1
 
1.3%
Space Separator
ValueCountFrequency (%)
968
100.0%
Open Punctuation
ValueCountFrequency (%)
( 71
100.0%
Close Punctuation
ValueCountFrequency (%)
) 71
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 29
100.0%
Math Symbol
ValueCountFrequency (%)
~ 2
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 20403
92.4%
Common 1563
 
7.1%
Latin 112
 
0.5%

Most frequent character per script

Hangul
ValueCountFrequency (%)
455
 
2.2%
374
 
1.8%
362
 
1.8%
359
 
1.8%
338
 
1.7%
309
 
1.5%
298
 
1.5%
288
 
1.4%
277
 
1.4%
236
 
1.2%
Other values (866) 17107
83.8%
Latin
ValueCountFrequency (%)
P 12
 
10.7%
R 10
 
8.9%
S 8
 
7.1%
A 6
 
5.4%
G 6
 
5.4%
N 5
 
4.5%
C 5
 
4.5%
T 5
 
4.5%
I 5
 
4.5%
K 4
 
3.6%
Other values (24) 46
41.1%
Common
ValueCountFrequency (%)
968
61.9%
( 71
 
4.5%
) 71
 
4.5%
1 67
 
4.3%
0 62
 
4.0%
, 57
 
3.6%
2 50
 
3.2%
3 40
 
2.6%
5 30
 
1.9%
- 29
 
1.9%
Other values (10) 118
 
7.5%

Most occurring blocks

ValueCountFrequency (%)
Hangul 20403
92.4%
ASCII 1673
 
7.6%
Punctuation 2
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
968
57.9%
( 71
 
4.2%
) 71
 
4.2%
1 67
 
4.0%
0 62
 
3.7%
, 57
 
3.4%
2 50
 
3.0%
3 40
 
2.4%
5 30
 
1.8%
- 29
 
1.7%
Other values (43) 228
 
13.6%
Hangul
ValueCountFrequency (%)
455
 
2.2%
374
 
1.8%
362
 
1.8%
359
 
1.8%
338
 
1.7%
309
 
1.5%
298
 
1.5%
288
 
1.4%
277
 
1.4%
236
 
1.2%
Other values (866) 17107
83.8%
Punctuation
ValueCountFrequency (%)
2
100.0%

대행사
Text

MISSING 

Distinct120
Distinct (%)5.1%
Missing803
Missing (%)25.3%
Memory size24.9 KiB
2023-12-12T07:53:13.749016image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length14
Median length7
Mean length6.4378424
Min length2

Characters and Unicode

Total characters15277
Distinct characters177
Distinct categories7 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique48 ?
Unique (%)2.0%

Sample

1st row(주)엘지애드
2nd row(주)엘지애드
3rd row(주)엘지애드
4th row(주)엘지애드
5th row(주)엘지애드
ValueCountFrequency (%)
주)제일기획 830
34.8%
주)오리콤 436
18.3%
주)대홍기획 150
 
6.3%
주)엘지애드 129
 
5.4%
삼희기획 127
 
5.3%
뉴코아종합기획 109
 
4.6%
주)mbc애드컴 71
 
3.0%
나라기획 52
 
2.2%
주)대방기획 46
 
1.9%
동방기획 26
 
1.1%
Other values (113) 409
17.1%
2023-12-12T07:53:14.077001image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
1824
11.9%
( 1821
11.9%
) 1821
11.9%
1402
 
9.2%
1402
 
9.2%
849
 
5.6%
844
 
5.5%
456
 
3.0%
443
 
2.9%
443
 
2.9%
Other values (167) 3972
26.0%

Most occurring categories

ValueCountFrequency (%)
Other Letter 11212
73.4%
Open Punctuation 1821
 
11.9%
Close Punctuation 1821
 
11.9%
Uppercase Letter 386
 
2.5%
Other Punctuation 24
 
0.2%
Space Separator 12
 
0.1%
Dash Punctuation 1
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
1824
16.3%
1402
12.5%
1402
12.5%
849
 
7.6%
844
 
7.5%
456
 
4.1%
443
 
4.0%
443
 
4.0%
252
 
2.2%
228
 
2.0%
Other values (144) 3069
27.4%
Uppercase Letter
ValueCountFrequency (%)
C 104
26.9%
M 94
24.4%
B 72
18.7%
S 21
 
5.4%
Q 16
 
4.1%
D 15
 
3.9%
R 15
 
3.9%
P 14
 
3.6%
Y 11
 
2.8%
K 6
 
1.6%
Other values (7) 18
 
4.7%
Other Punctuation
ValueCountFrequency (%)
& 16
66.7%
. 8
33.3%
Open Punctuation
ValueCountFrequency (%)
( 1821
100.0%
Close Punctuation
ValueCountFrequency (%)
) 1821
100.0%
Space Separator
ValueCountFrequency (%)
12
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 11212
73.4%
Common 3679
 
24.1%
Latin 386
 
2.5%

Most frequent character per script

Hangul
ValueCountFrequency (%)
1824
16.3%
1402
12.5%
1402
12.5%
849
 
7.6%
844
 
7.5%
456
 
4.1%
443
 
4.0%
443
 
4.0%
252
 
2.2%
228
 
2.0%
Other values (144) 3069
27.4%
Latin
ValueCountFrequency (%)
C 104
26.9%
M 94
24.4%
B 72
18.7%
S 21
 
5.4%
Q 16
 
4.1%
D 15
 
3.9%
R 15
 
3.9%
P 14
 
3.6%
Y 11
 
2.8%
K 6
 
1.6%
Other values (7) 18
 
4.7%
Common
ValueCountFrequency (%)
( 1821
49.5%
) 1821
49.5%
& 16
 
0.4%
12
 
0.3%
. 8
 
0.2%
- 1
 
< 0.1%

Most occurring blocks

ValueCountFrequency (%)
Hangul 11212
73.4%
ASCII 4065
 
26.6%

Most frequent character per block

Hangul
ValueCountFrequency (%)
1824
16.3%
1402
12.5%
1402
12.5%
849
 
7.6%
844
 
7.5%
456
 
4.1%
443
 
4.0%
443
 
4.0%
252
 
2.2%
228
 
2.0%
Other values (144) 3069
27.4%
ASCII
ValueCountFrequency (%)
( 1821
44.8%
) 1821
44.8%
C 104
 
2.6%
M 94
 
2.3%
B 72
 
1.8%
S 21
 
0.5%
Q 16
 
0.4%
& 16
 
0.4%
D 15
 
0.4%
R 15
 
0.4%
Other values (13) 70
 
1.7%

Interactions

2023-12-12T07:53:08.578312image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-12T07:53:14.156578image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
년도대분류
년도1.0000.525
대분류0.5251.000
2023-12-12T07:53:14.225948image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
년도대분류
년도1.0000.222
대분류0.2221.000

Missing values

2023-12-12T07:53:08.733018image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-12T07:53:08.864403image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.
2023-12-12T07:53:09.000850image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
The correlation heatmap measures nullity correlation: how strongly the presence or absence of one variable affects the presence of another.

Sample

파일명년도광고주대분류중분류소분류제품명제목대행사
0A1A00001.wma1991엘프기초재석탄, 석유 및 가스윤활유엘프윤활류<NA>
1A1A00004.wma1996케스트롤뉴트론X기초재석탄, 석유 및 가스연료유케스트롤뉴트론X한국의명차들과함께<NA>
2A1A00005.wma1997엘지칼텍스정유(주)여기초재석탄, 석유 및 가스연료유테크론어서오세요(주)엘지애드
3A1A00006.wma1996엘지칼텍스정유(주)여기초재석탄, 석유 및 가스연료유테크론LG정유라고해주세요(주)엘지애드
4A1A00007.wma1997엘지칼텍스정유(주)여기초재석탄, 석유 및 가스연료유테크론장애인돕기에참여하세요(주)엘지애드
5A1A00008.wma<NA>엘지칼텍스정유(주)여기초재석탄, 석유 및 가스연료유테크론말못하는차도스트레스(주)엘지애드
6A1A00009.wma<NA>엘지칼텍스정유(주)여기초재석탄, 석유 및 가스연료유테크론추석맞이(주)엘지애드
7A1A00010.wma1993쌍용정유기초재기초재 기타기초재 기업PR쌍용정유변함없는친절함(주)제일기획
8A1A00011.wma1993쌍용정유기초재기초재 기타기초재 기업PR쌍용정유10부제편(주)제일기획
9A1A00012.wma<NA>쌍용정유기초재기초재 기타기초재 기업PR쌍용정유도와주세요(주)제일기획
파일명년도광고주대분류중분류소분류제품명제목대행사
3166radio2000_013.mp3<NA>(주)롯데삼강그룹 및 기업광고그룹광고그룹PR롯데삼강기술과 전통으로 앞서가는 롯데삼강(주)대홍기획
3167radio2000_114.mp3<NA>한국통신엠닷컴(주)그룹 및 기업광고그룹광고그룹PR한국통신 기업 PR한국통신이 만들어갑니다(주)휘닉스커뮤니케이션즈
3168radio2000_120.mp3<NA>(주)한글과컴퓨터그룹 및 기업광고그룹광고그룹PR한글과 컴퓨터 기업PR인터넷 국민기업티비더블유에이코리아(주)
3169radio2000_145.mp3<NA>유한킴벌리(주)그룹 및 기업광고그룹광고그룹PR유한킴벌리(가을)우리강산 푸르게 푸르게(주)오리콤
3170radio2000_146.mp3<NA>유한킴벌리(주)그룹 및 기업광고그룹광고그룹PR유한킴벌리(비)우리강산 푸르게 푸르게(주)오리콤
3171radio2000_147.mp3<NA>유한킴벌리(주)그룹 및 기업광고그룹광고그룹PR유한킴벌리(별)우리강산 푸르게 푸르게(주)오리콤
3172radio2000_179.mp3<NA>(주)에스엠엔터테인먼트그룹 및 기업광고그룹광고그룹PRHOT그들은 우리와 다르지 않다(주)아나기획
3173radio2000_212.mp3<NA>(주)경농그룹 및 기업광고그룹광고그룹PR경농기업PR농작물 튼튼 자연은 소중하게<NA>
3174radio2000_219.mp3<NA>포항종합제철(주)그룹 및 기업광고그룹광고그룹PR포스코철이 없다면…소리없이 세상을 움직입니다(주)하쿠호도제일
3175radio2000_221.mp3<NA>SK글로벌(주)그룹 및 기업광고그룹광고그룹PRSK글로벌세상보다 먼저 변화는 기업의 강력한 네트워크(주)제일기획