Overview

Dataset statistics

Number of variables10
Number of observations10000
Missing cells1855
Missing cells (%)1.9%
Duplicate rows1
Duplicate rows (%)< 0.1%
Total size in memory869.1 KiB
Average record size in memory89.0 B

Variable types

Numeric1
Categorical6
Text2
Boolean1

Dataset

Description한국산업기술진흥원 통합사업관리시스템에 등재되는 정보입니다. NTIS 지식재산권 정보로 성과발생년도 지식재산권, 출원방법, 출원국, 출원구분 및 공개엽부, 사업화활용, 발명 명칭 등을 제공합니다.
Author한국산업기술진흥원
URLhttps://www.data.go.kr/data/15071136/fileData.do

Alerts

Dataset has 1 (< 0.1%) duplicate rowsDuplicates
출원방법 is highly overall correlated with 출원국High correlation
출원국 is highly overall correlated with 출원방법High correlation
지식재산권종류구분 is highly imbalanced (52.4%)Imbalance
출원국 is highly imbalanced (87.6%)Imbalance
지식재산권검색여부 has 1842 (18.4%) missing valuesMissing

Reproduction

Analysis started2023-12-12 12:11:12.020952
Analysis finished2023-12-12 12:11:15.242962
Duration3.22 seconds
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

성과발생년도
Real number (ℝ)

Distinct16
Distinct (%)0.2%
Missing2
Missing (%)< 0.1%
Infinite0
Infinite (%)0.0%
Mean2015.7596
Minimum2006
Maximum2021
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2023-12-12T21:11:15.309722image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum2006
5-th percentile2011
Q12013
median2016
Q32018
95-th percentile2021
Maximum2021
Range15
Interquartile range (IQR)5

Descriptive statistics

Standard deviation3.0192269
Coefficient of variation (CV)0.001497811
Kurtosis-0.79388916
Mean2015.7596
Median Absolute Deviation (MAD)2
Skewness-0.069326926
Sum20153564
Variance9.115731
MonotonicityNot monotonic
2023-12-12T21:11:15.463271image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=16)
ValueCountFrequency (%)
2014 1232
12.3%
2017 1222
12.2%
2013 1186
11.9%
2018 1042
10.4%
2015 1032
10.3%
2016 791
7.9%
2020 751
7.5%
2019 740
7.4%
2012 665
6.7%
2021 606
6.1%
Other values (6) 731
7.3%
ValueCountFrequency (%)
2006 1
 
< 0.1%
2007 2
 
< 0.1%
2008 6
 
0.1%
2009 159
 
1.6%
2010 246
 
2.5%
2011 317
 
3.2%
2012 665
6.7%
2013 1186
11.9%
2014 1232
12.3%
2015 1032
10.3%
ValueCountFrequency (%)
2021 606
6.1%
2020 751
7.5%
2019 740
7.4%
2018 1042
10.4%
2017 1222
12.2%
2016 791
7.9%
2015 1032
10.3%
2014 1232
12.3%
2013 1186
11.9%
2012 665
6.7%

지식재산권종류구분
Categorical

IMBALANCE 

Distinct7
Distinct (%)0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
특허
7283 
<NA>
1512 
디자인
 
516
상표
 
423
SW등록
 
144
Other values (2)
 
122

Length

Max length4
Median length2
Mean length2.394
Min length2

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row<NA>
2nd row특허
3rd row<NA>
4th row특허
5th row특허

Common Values

ValueCountFrequency (%)
특허 7283
72.8%
<NA> 1512
 
15.1%
디자인 516
 
5.2%
상표 423
 
4.2%
SW등록 144
 
1.4%
기타 66
 
0.7%
실용신안 56
 
0.6%

Length

2023-12-12T21:11:15.689665image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T21:11:15.849338image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
특허 7283
72.8%
na 1512
 
15.1%
디자인 516
 
5.2%
상표 423
 
4.2%
sw등록 144
 
1.4%
기타 66
 
0.7%
실용신안 56
 
0.6%

출원방법
Categorical

HIGH CORRELATION 

Distinct4
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
국내 출원
7775 
<NA>
1514 
일반 해외 출원
 
399
PCT 해외 출원
 
312

Length

Max length9
Median length5
Mean length5.0931
Min length4

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row<NA>
2nd row국내 출원
3rd row<NA>
4th row국내 출원
5th row국내 출원

Common Values

ValueCountFrequency (%)
국내 출원 7775
77.8%
<NA> 1514
 
15.1%
일반 해외 출원 399
 
4.0%
PCT 해외 출원 312
 
3.1%

Length

2023-12-12T21:11:16.043277image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T21:11:16.171960image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
출원 8486
44.2%
국내 7775
40.5%
na 1514
 
7.9%
해외 711
 
3.7%
일반 399
 
2.1%
pct 312
 
1.6%

출원국
Categorical

HIGH CORRELATION  IMBALANCE 

Distinct38
Distinct (%)0.4%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
대한민국
9230 
미국
 
237
중국
 
138
일본
 
95
가나
 
70
Other values (33)
 
230

Length

Max length8
Median length4
Mean length3.8793
Min length1

Unique

Unique12 ?
Unique (%)0.1%

Sample

1st row대한민국
2nd row대한민국
3rd row대한민국
4th row대한민국
5th row대한민국

Common Values

ValueCountFrequency (%)
대한민국 9230
92.3%
미국 237
 
2.4%
중국 138
 
1.4%
일본 95
 
0.9%
가나 70
 
0.7%
국제 60
 
0.6%
유럽연합 38
 
0.4%
베트남 14
 
0.1%
크리스마스도 12
 
0.1%
인도 11
 
0.1%
Other values (28) 95
 
0.9%

Length

2023-12-12T21:11:16.316250image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category
ValueCountFrequency (%)
대한민국 9230
92.3%
미국 237
 
2.4%
중국 138
 
1.4%
일본 95
 
0.9%
가나 70
 
0.7%
국제 60
 
0.6%
유럽연합 38
 
0.4%
베트남 14
 
0.1%
크리스마스도 12
 
0.1%
기타지역 11
 
0.1%
Other values (28) 95
 
0.9%

출원구분
Categorical

Distinct3
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
특허출원
6952 
특허등록
2991 
<NA>
 
57

Length

Max length4
Median length4
Mean length4
Min length4

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row특허출원
2nd row특허출원
3rd row특허등록
4th row특허출원
5th row특허출원

Common Values

ValueCountFrequency (%)
특허출원 6952
69.5%
특허등록 2991
29.9%
<NA> 57
 
0.6%

Length

2023-12-12T21:11:16.469894image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T21:11:16.587517image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
특허출원 6952
69.5%
특허등록 2991
29.9%
na 57
 
0.6%
Distinct5344
Distinct (%)53.5%
Missing2
Missing (%)< 0.1%
Memory size156.2 KiB
2023-12-12T21:11:16.854866image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length142
Median length118
Mean length8.9981996
Min length1

Characters and Unicode

Total characters89964
Distinct characters754
Distinct categories12 ?
Distinct scripts5 ?
Distinct blocks6 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique3762 ?
Unique (%)37.6%

Sample

1st row한화케미칼 주식회사
2nd row(주)세일섬유
3rd row유노특허법률사무소
4th row오스템
5th row주은테크, 덕양산업
ValueCountFrequency (%)
주식회사 1597
 
11.9%
산학협력단 749
 
5.6%
특허청 143
 
1.1%
재단법인 115
 
0.9%
97
 
0.7%
한국기계연구원 86
 
0.6%
한국광기술원 83
 
0.6%
영남대학교 71
 
0.5%
전자부품연구원 65
 
0.5%
52
 
0.4%
Other values (5321) 10312
77.1%
2023-12-12T21:11:17.334228image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
5270
 
5.9%
3732
 
4.1%
3372
 
3.7%
) 3208
 
3.6%
( 3202
 
3.6%
2401
 
2.7%
2358
 
2.6%
2261
 
2.5%
2250
 
2.5%
2180
 
2.4%
Other values (744) 59730
66.4%

Most occurring categories

ValueCountFrequency (%)
Other Letter 76677
85.2%
Space Separator 3372
 
3.7%
Close Punctuation 3208
 
3.6%
Open Punctuation 3202
 
3.6%
Uppercase Letter 1314
 
1.5%
Other Punctuation 706
 
0.8%
Other Symbol 693
 
0.8%
Lowercase Letter 556
 
0.6%
Decimal Number 115
 
0.1%
Math Symbol 100
 
0.1%
Other values (2) 21
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
5270
 
6.9%
3732
 
4.9%
2401
 
3.1%
2358
 
3.1%
2261
 
2.9%
2250
 
2.9%
2180
 
2.8%
2166
 
2.8%
1967
 
2.6%
1858
 
2.4%
Other values (673) 50234
65.5%
Uppercase Letter
ValueCountFrequency (%)
C 134
 
10.2%
I 109
 
8.3%
O 104
 
7.9%
T 99
 
7.5%
S 94
 
7.2%
N 91
 
6.9%
A 89
 
6.8%
E 87
 
6.6%
L 72
 
5.5%
D 63
 
4.8%
Other values (14) 372
28.3%
Lowercase Letter
ValueCountFrequency (%)
e 61
11.0%
n 61
11.0%
o 59
10.6%
t 52
9.4%
i 49
8.8%
c 41
 
7.4%
r 38
 
6.8%
a 37
 
6.7%
s 28
 
5.0%
l 21
 
3.8%
Other values (14) 109
19.6%
Decimal Number
ValueCountFrequency (%)
1 46
40.0%
0 23
20.0%
2 19
16.5%
3 8
 
7.0%
4 6
 
5.2%
9 4
 
3.5%
5 4
 
3.5%
8 2
 
1.7%
7 2
 
1.7%
6 1
 
0.9%
Other Punctuation
ValueCountFrequency (%)
; 359
50.8%
, 196
27.8%
. 93
 
13.2%
& 31
 
4.4%
/ 26
 
3.7%
: 1
 
0.1%
Space Separator
ValueCountFrequency (%)
3372
100.0%
Close Punctuation
ValueCountFrequency (%)
) 3208
100.0%
Open Punctuation
ValueCountFrequency (%)
( 3202
100.0%
Other Symbol
ValueCountFrequency (%)
693
100.0%
Math Symbol
ValueCountFrequency (%)
| 100
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 20
100.0%
Connector Punctuation
ValueCountFrequency (%)
_ 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 77351
86.0%
Common 10724
 
11.9%
Latin 1870
 
2.1%
Katakana 17
 
< 0.1%
Han 2
 
< 0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
5270
 
6.8%
3732
 
4.8%
2401
 
3.1%
2358
 
3.0%
2261
 
2.9%
2250
 
2.9%
2180
 
2.8%
2166
 
2.8%
1967
 
2.5%
1858
 
2.4%
Other values (659) 50908
65.8%
Latin
ValueCountFrequency (%)
C 134
 
7.2%
I 109
 
5.8%
O 104
 
5.6%
T 99
 
5.3%
S 94
 
5.0%
N 91
 
4.9%
A 89
 
4.8%
E 87
 
4.7%
L 72
 
3.9%
D 63
 
3.4%
Other values (38) 928
49.6%
Common
ValueCountFrequency (%)
3372
31.4%
) 3208
29.9%
( 3202
29.9%
; 359
 
3.3%
, 196
 
1.8%
| 100
 
0.9%
. 93
 
0.9%
1 46
 
0.4%
& 31
 
0.3%
/ 26
 
0.2%
Other values (12) 91
 
0.8%
Katakana
ValueCountFrequency (%)
2
11.8%
2
11.8%
2
11.8%
2
11.8%
1
 
5.9%
1
 
5.9%
1
 
5.9%
1
 
5.9%
1
 
5.9%
1
 
5.9%
Other values (3) 3
17.6%
Han
ValueCountFrequency (%)
1
50.0%
1
50.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 76657
85.2%
ASCII 12594
 
14.0%
None 693
 
0.8%
Katakana 17
 
< 0.1%
CJK 2
 
< 0.1%
Compat Jamo 1
 
< 0.1%

Most frequent character per block

Hangul
ValueCountFrequency (%)
5270
 
6.9%
3732
 
4.9%
2401
 
3.1%
2358
 
3.1%
2261
 
2.9%
2250
 
2.9%
2180
 
2.8%
2166
 
2.8%
1967
 
2.6%
1858
 
2.4%
Other values (657) 50214
65.5%
ASCII
ValueCountFrequency (%)
3372
26.8%
) 3208
25.5%
( 3202
25.4%
; 359
 
2.9%
, 196
 
1.6%
C 134
 
1.1%
I 109
 
0.9%
O 104
 
0.8%
| 100
 
0.8%
T 99
 
0.8%
Other values (60) 1711
13.6%
None
ValueCountFrequency (%)
693
100.0%
Katakana
ValueCountFrequency (%)
2
11.8%
2
11.8%
2
11.8%
2
11.8%
1
 
5.9%
1
 
5.9%
1
 
5.9%
1
 
5.9%
1
 
5.9%
1
 
5.9%
Other values (3) 3
17.6%
CJK
ValueCountFrequency (%)
1
50.0%
1
50.0%
Compat Jamo
ValueCountFrequency (%)
1
100.0%

공개여부
Categorical

Distinct3
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
공개
4477 
비공개
3886 
<NA>
1637 

Length

Max length4
Median length3
Mean length2.716
Min length2

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row<NA>
2nd row공개
3rd row<NA>
4th row공개
5th row비공개

Common Values

ValueCountFrequency (%)
공개 4477
44.8%
비공개 3886
38.9%
<NA> 1637
 
16.4%

Length

2023-12-12T21:11:17.472788image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T21:11:17.585882image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
공개 4477
44.8%
비공개 3886
38.9%
na 1637
 
16.4%
Distinct6
Distinct (%)0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
<NA>
4652 
상업화
2836 
기타
1369 
해당없음
737 
후속특허
 
233

Length

Max length4
Median length4
Mean length3.4426
Min length2

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row<NA>
2nd row기타
3rd row<NA>
4th row후속특허
5th row기타

Common Values

ValueCountFrequency (%)
<NA> 4652
46.5%
상업화 2836
28.4%
기타 1369
 
13.7%
해당없음 737
 
7.4%
후속특허 233
 
2.3%
기술이전 173
 
1.7%

Length

2023-12-12T21:11:17.692505image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T21:11:17.819544image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
na 4652
46.5%
상업화 2836
28.4%
기타 1369
 
13.7%
해당없음 737
 
7.4%
후속특허 233
 
2.3%
기술이전 173
 
1.7%
Distinct2
Distinct (%)< 0.1%
Missing1842
Missing (%)18.4%
Memory size97.7 KiB
False
5116 
True
3042 
(Missing)
1842 
ValueCountFrequency (%)
False 5116
51.2%
True 3042
30.4%
(Missing) 1842
 
18.4%
2023-12-12T21:11:17.927557image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Distinct9226
Distinct (%)92.3%
Missing9
Missing (%)0.1%
Memory size156.2 KiB
2023-12-12T21:11:18.278760image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length175
Median length116
Mean length25.104994
Min length1

Characters and Unicode

Total characters250824
Distinct characters1325
Distinct categories14 ?
Distinct scripts7 ?
Distinct blocks10 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique8662 ?
Unique (%)86.7%

Sample

1st row폴리실리콘 제조 장치
2nd row열차단성능이우수한차량내장재용 코팅제 및 이를 활용한 열 차단 성능이 우수한 원단의 코팅방법
3rd row풍력 발전기용 기어박스
4th row차량시트용리클라이너
5th row엠보 패턴 다양화 구현형 사출금형
ValueCountFrequency (%)
3521
 
5.8%
방법 1473
 
2.4%
제조방법 1342
 
2.2%
이용한 1215
 
2.0%
장치 949
 
1.6%
조성물 901
 
1.5%
시스템 861
 
1.4%
690
 
1.1%
포함하는 685
 
1.1%
이를 674
 
1.1%
Other values (17978) 48002
79.6%
2023-12-12T21:11:18.870874image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
50321
 
20.1%
6029
 
2.4%
4487
 
1.8%
4343
 
1.7%
4142
 
1.7%
3601
 
1.4%
3582
 
1.4%
3406
 
1.4%
3400
 
1.4%
3370
 
1.3%
Other values (1315) 164143
65.4%

Most occurring categories

ValueCountFrequency (%)
Other Letter 175231
69.9%
Space Separator 50322
 
20.1%
Uppercase Letter 12480
 
5.0%
Lowercase Letter 9642
 
3.8%
Decimal Number 1167
 
0.5%
Other Punctuation 803
 
0.3%
Dash Punctuation 497
 
0.2%
Open Punctuation 322
 
0.1%
Close Punctuation 312
 
0.1%
Connector Punctuation 32
 
< 0.1%
Other values (4) 16
 
< 0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
6029
 
3.4%
4487
 
2.6%
4343
 
2.5%
4142
 
2.4%
3601
 
2.1%
3582
 
2.0%
3406
 
1.9%
3400
 
1.9%
3370
 
1.9%
3000
 
1.7%
Other values (1164) 135871
77.5%
Uppercase Letter
ValueCountFrequency (%)
E 1097
 
8.8%
A 951
 
7.6%
I 904
 
7.2%
O 825
 
6.6%
T 818
 
6.6%
N 793
 
6.4%
S 773
 
6.2%
R 765
 
6.1%
D 713
 
5.7%
C 700
 
5.6%
Other values (42) 4141
33.2%
Lowercase Letter
ValueCountFrequency (%)
e 1017
 
10.5%
i 897
 
9.3%
o 886
 
9.2%
a 807
 
8.4%
n 763
 
7.9%
t 755
 
7.8%
r 650
 
6.7%
s 442
 
4.6%
c 434
 
4.5%
l 411
 
4.3%
Other values (39) 2580
26.8%
Decimal Number
ValueCountFrequency (%)
3 239
20.5%
2 227
19.5%
1 202
17.3%
0 171
14.7%
5 67
 
5.7%
4 66
 
5.7%
6 56
 
4.8%
9 51
 
4.4%
7 40
 
3.4%
8 27
 
2.3%
Other values (8) 21
 
1.8%
Other Punctuation
ValueCountFrequency (%)
, 647
80.6%
/ 70
 
8.7%
. 54
 
6.7%
· 7
 
0.9%
& 7
 
0.9%
: 7
 
0.9%
' 6
 
0.7%
3
 
0.4%
2
 
0.2%
Open Punctuation
ValueCountFrequency (%)
( 253
78.6%
[ 65
 
20.2%
{ 3
 
0.9%
1
 
0.3%
Close Punctuation
ValueCountFrequency (%)
) 245
78.5%
] 64
 
20.5%
} 2
 
0.6%
1
 
0.3%
Math Symbol
ValueCountFrequency (%)
+ 4
40.0%
3
30.0%
~ 2
20.0%
= 1
 
10.0%
Dash Punctuation
ValueCountFrequency (%)
- 481
96.8%
9
 
1.8%
7
 
1.4%
Space Separator
ValueCountFrequency (%)
50321
> 99.9%
  1
 
< 0.1%
Other Number
ValueCountFrequency (%)
2
66.7%
1
33.3%
Letter Number
ValueCountFrequency (%)
1
50.0%
1
50.0%
Connector Punctuation
ValueCountFrequency (%)
_ 32
100.0%
Format
ValueCountFrequency (%)
­ 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 174743
69.7%
Common 53469
 
21.3%
Latin 22113
 
8.8%
Han 289
 
0.1%
Katakana 103
 
< 0.1%
Hiragana 96
 
< 0.1%
Greek 11
 
< 0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
6029
 
3.5%
4487
 
2.6%
4343
 
2.5%
4142
 
2.4%
3601
 
2.1%
3582
 
2.0%
3406
 
1.9%
3400
 
1.9%
3370
 
1.9%
3000
 
1.7%
Other values (959) 135383
77.5%
Han
ValueCountFrequency (%)
12
 
4.2%
12
 
4.2%
10
 
3.5%
10
 
3.5%
9
 
3.1%
9
 
3.1%
6
 
2.1%
6
 
2.1%
5
 
1.7%
5
 
1.7%
Other values (132) 205
70.9%
Latin
ValueCountFrequency (%)
E 1097
 
5.0%
e 1017
 
4.6%
A 951
 
4.3%
I 904
 
4.1%
i 897
 
4.1%
o 886
 
4.0%
O 825
 
3.7%
T 818
 
3.7%
a 807
 
3.6%
N 793
 
3.6%
Other values (88) 13118
59.3%
Common
ValueCountFrequency (%)
50321
94.1%
, 647
 
1.2%
- 481
 
0.9%
( 253
 
0.5%
) 245
 
0.5%
3 239
 
0.4%
2 227
 
0.4%
1 202
 
0.4%
0 171
 
0.3%
/ 70
 
0.1%
Other values (38) 613
 
1.1%
Katakana
ValueCountFrequency (%)
11
 
10.7%
7
 
6.8%
7
 
6.8%
6
 
5.8%
5
 
4.9%
4
 
3.9%
4
 
3.9%
4
 
3.9%
4
 
3.9%
4
 
3.9%
Other values (28) 47
45.6%
Hiragana
ValueCountFrequency (%)
12
12.5%
10
10.4%
9
 
9.4%
7
 
7.3%
7
 
7.3%
6
 
6.2%
6
 
6.2%
5
 
5.2%
5
 
5.2%
5
 
5.2%
Other values (15) 24
25.0%
Greek
ValueCountFrequency (%)
β 4
36.4%
α 3
27.3%
Ν 2
18.2%
π 1
 
9.1%
Ι 1
 
9.1%

Most occurring blocks

ValueCountFrequency (%)
Hangul 174735
69.7%
ASCII 75084
29.9%
None 495
 
0.2%
CJK 289
 
0.1%
Katakana 103
 
< 0.1%
Hiragana 96
 
< 0.1%
Punctuation 9
 
< 0.1%
Compat Jamo 8
 
< 0.1%
Arrows 3
 
< 0.1%
Number Forms 2
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
50321
67.0%
E 1097
 
1.5%
e 1017
 
1.4%
A 951
 
1.3%
I 904
 
1.2%
i 897
 
1.2%
o 886
 
1.2%
O 825
 
1.1%
T 818
 
1.1%
a 807
 
1.1%
Other values (70) 16561
 
22.1%
Hangul
ValueCountFrequency (%)
6029
 
3.5%
4487
 
2.6%
4343
 
2.5%
4142
 
2.4%
3601
 
2.1%
3582
 
2.0%
3406
 
1.9%
3400
 
1.9%
3370
 
1.9%
3000
 
1.7%
Other values (953) 135375
77.5%
None
ValueCountFrequency (%)
41
 
8.3%
27
 
5.5%
25
 
5.1%
25
 
5.1%
21
 
4.2%
20
 
4.0%
20
 
4.0%
20
 
4.0%
19
 
3.8%
17
 
3.4%
Other values (57) 260
52.5%
Hiragana
ValueCountFrequency (%)
12
12.5%
10
10.4%
9
 
9.4%
7
 
7.3%
7
 
7.3%
6
 
6.2%
6
 
6.2%
5
 
5.2%
5
 
5.2%
5
 
5.2%
Other values (15) 24
25.0%
CJK
ValueCountFrequency (%)
12
 
4.2%
12
 
4.2%
10
 
3.5%
10
 
3.5%
9
 
3.1%
9
 
3.1%
6
 
2.1%
6
 
2.1%
5
 
1.7%
5
 
1.7%
Other values (132) 205
70.9%
Katakana
ValueCountFrequency (%)
11
 
10.7%
7
 
6.8%
7
 
6.8%
6
 
5.8%
5
 
4.9%
4
 
3.9%
4
 
3.9%
4
 
3.9%
4
 
3.9%
4
 
3.9%
Other values (28) 47
45.6%
Punctuation
ValueCountFrequency (%)
9
100.0%
Arrows
ValueCountFrequency (%)
3
100.0%
Compat Jamo
ValueCountFrequency (%)
2
25.0%
2
25.0%
1
12.5%
1
12.5%
1
12.5%
1
12.5%
Number Forms
ValueCountFrequency (%)
1
50.0%
1
50.0%

Interactions

2023-12-12T21:11:14.496714image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-12T21:11:18.979001image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
성과발생년도지식재산권종류구분출원방법출원국출원구분공개여부사업화활용여부지식재산권검색여부
성과발생년도1.0000.1380.0810.1100.1380.4980.1730.527
지식재산권종류구분0.1381.0000.1940.1680.1980.1020.1140.132
출원방법0.0810.1941.0000.8520.0390.0000.0660.009
출원국0.1100.1680.8521.0000.0840.0880.1010.155
출원구분0.1380.1980.0390.0841.0000.2350.0420.361
공개여부0.4980.1020.0000.0880.2351.0000.0560.077
사업화활용여부0.1730.1140.0660.1010.0420.0561.0000.047
지식재산권검색여부0.5270.1320.0090.1550.3610.0770.0471.000
2023-12-12T21:11:19.133513image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
사업화활용여부공개여부출원방법지식재산권종류구분출원구분출원국지식재산권검색여부
사업화활용여부1.0000.0690.0490.0770.0520.0480.058
공개여부0.0691.0000.0000.0730.1510.0700.049
출원방법0.0490.0001.0000.0820.0650.6610.015
지식재산권종류구분0.0770.0730.0821.0000.1420.0740.095
출원구분0.0520.1510.0650.1421.0000.0700.235
출원국0.0480.0700.6610.0740.0701.0000.131
지식재산권검색여부0.0580.0490.0150.0950.2350.1311.000
2023-12-12T21:11:19.257201image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
성과발생년도지식재산권종류구분출원방법출원국출원구분공개여부사업화활용여부지식재산권검색여부
성과발생년도1.0000.0730.0460.0370.1060.3830.1110.406
지식재산권종류구분0.0731.0000.0820.0740.1420.0730.0770.095
출원방법0.0460.0821.0000.6610.0650.0000.0490.015
출원국0.0370.0740.6611.0000.0700.0700.0480.131
출원구분0.1060.1420.0650.0701.0000.1510.0520.235
공개여부0.3830.0730.0000.0700.1511.0000.0690.049
사업화활용여부0.1110.0770.0490.0480.0520.0691.0000.058
지식재산권검색여부0.4060.0950.0150.1310.2350.0490.0581.000

Missing values

2023-12-12T21:11:14.637782image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-12T21:11:14.873406image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.
2023-12-12T21:11:15.083837image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
The correlation heatmap measures nullity correlation: how strongly the presence or absence of one variable affects the presence of another.

Sample

성과발생년도지식재산권종류구분출원방법출원국출원구분출원등록기관공개여부사업화활용여부지식재산권검색여부발명의명칭
103182013<NA><NA>대한민국특허출원한화케미칼 주식회사<NA><NA>N폴리실리콘 제조 장치
299042017특허국내 출원대한민국특허출원(주)세일섬유공개기타N열차단성능이우수한차량내장재용 코팅제 및 이를 활용한 열 차단 성능이 우수한 원단의 코팅방법
152122014<NA><NA>대한민국특허등록유노특허법률사무소<NA><NA>N풍력 발전기용 기어박스
185272015특허국내 출원대한민국특허출원오스템공개후속특허N차량시트용리클라이너
355692018특허국내 출원대한민국특허출원주은테크, 덕양산업비공개기타<NA>엠보 패턴 다양화 구현형 사출금형
453832021특허국내 출원대한민국특허등록주식회사 다이아덴트비공개상업화N치과용 근관 충전 시멘트 조성물 및 그의 제좡법
181162014특허국내 출원대한민국특허출원(주)에스엘라이팅비공개<NA>Y차량용 램프(BF 램프에서 Ref&H_Sink 결합구조)
385912019SW등록국내 출원대한민국특허출원주식회사 이티엑스웍스공개상업화<NA>3차원 Depth 기반 객체 검출 엔진
207222015특허국내 출원대한민국특허출원주식회사 케이비공개상업화Y증기터빈용 티타늄 합금 블레이드의 제조방법
287742017특허국내 출원대한민국특허등록유한회사한풍제약비공개해당없음Y몰약, 당귀 및 오미자를 유효성분으로 함유하는 기억력 개선을 위한 조성물
성과발생년도지식재산권종류구분출원방법출원국출원구분출원등록기관공개여부사업화활용여부지식재산권검색여부발명의명칭
309482017특허국내 출원대한민국특허출원(주)바이오에프디엔씨비공개상업화N세포투과형 혈관내피세포성장인자 융합 단백질을 함유하는 피부 생리활성 피부외용제 조성물
165322014<NA><NA>대한민국특허등록(주)태진기술<NA><NA>NDC-DC 컨버터
311072017특허국내 출원대한민국특허등록(주)한국티이에스비공개상업화Y슈퍼커패시터와 재충전배터리를 이용한 발광장치
16532010특허국내 출원대한민국특허등록강릉원주대학교산학협력단공개<NA>N해삼의 건조분말 또는 추출물을 유효성분으로 함유하는 당뇨병의 예방 및 치료용 조성물
358022018특허국내 출원대한민국특허출원더블유스코프코리아 주식회사공개기타N다공성 분리막의 제조방법
453252021특허국내 출원대한민국특허등록(주)군장조선공개상업화N스텝 헐과 센터 킬 구조를 적용한 고속단정
17922010특허국내 출원대한민국특허출원㈜액트공개<NA>YRF형 방열회로기판 및 그 의 제조방법
37692012특허국내 출원대한민국<NA>포스코에너지 주식회사공개<NA>Y연료전지를 이용한 발전 시스템
119132013특허국내 출원대한민국특허출원(주)티오피에스공개<NA>Y워터젯 식각장치
273432017특허국내 출원대한민국특허출원(주)비젼사이언스비공개상업화N컨택트렌즈용 전자차폐 안료

Duplicate rows

Most frequently occurring

성과발생년도지식재산권종류구분출원방법출원국출원구분출원등록기관공개여부사업화활용여부지식재산권검색여부발명의명칭# duplicates
0<NA><NA><NA><NA><NA><NA><NA><NA><NA><NA>2