Overview

Dataset statistics

Number of variables4
Number of observations83
Missing cells0
Missing cells (%)0.0%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory2.8 KiB
Average record size in memory34.6 B

Variable types

Numeric1
Text3

Dataset

Description방위사업청 정책실명제 중전관리 대상사업 목록 정보를 제공하며, 정책실명제등록번호,사업명,사업요지,담당부서로 구성되어 있음
Author방위사업청
URLhttps://www.data.go.kr/data/3073015/fileData.do

Alerts

정책실명제등록번호 has unique valuesUnique

Reproduction

Analysis started2023-12-12 22:51:46.896598
Analysis finished2023-12-12 22:51:47.918761
Duration1.02 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

정책실명제등록번호
Real number (ℝ)

UNIQUE 

Distinct83
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean201516.43
Minimum201301
Maximum201915
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size879.0 B
2023-12-13T07:51:48.003804image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum201301
5-th percentile201305.1
Q1201321.5
median201401
Q3201703.5
95-th percentile201910.9
Maximum201915
Range614
Interquartile range (IQR)382

Descriptive statistics

Standard deviation234.71442
Coefficient of variation (CV)0.0011647408
Kurtosis-1.1711893
Mean201516.43
Median Absolute Deviation (MAD)96
Skewness0.69631011
Sum16725864
Variance55090.858
MonotonicityStrictly increasing
2023-12-13T07:51:48.179767image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
201301 1
 
1.2%
201602 1
 
1.2%
201703 1
 
1.2%
201702 1
 
1.2%
201701 1
 
1.2%
201700 1
 
1.2%
201606 1
 
1.2%
201605 1
 
1.2%
201604 1
 
1.2%
201603 1
 
1.2%
Other values (73) 73
88.0%
ValueCountFrequency (%)
201301 1
1.2%
201302 1
1.2%
201303 1
1.2%
201304 1
1.2%
201305 1
1.2%
201306 1
1.2%
201307 1
1.2%
201308 1
1.2%
201309 1
1.2%
201310 1
1.2%
ValueCountFrequency (%)
201915 1
1.2%
201914 1
1.2%
201913 1
1.2%
201912 1
1.2%
201911 1
1.2%
201910 1
1.2%
201909 1
1.2%
201908 1
1.2%
201907 1
1.2%
201906 1
1.2%
Distinct82
Distinct (%)98.8%
Missing0
Missing (%)0.0%
Memory size796.0 B
2023-12-13T07:51:48.455281image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length27
Median length18
Mean length11.036145
Min length5

Characters and Unicode

Total characters916
Distinct characters176
Distinct categories10 ?
Distinct scripts3 ?
Distinct blocks4 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique81 ?
Unique (%)97.6%

Sample

1st row군 위성통신장비
2nd row한국형 기동헬기 초도양산
3rd rowFA-50 개조개발 (R&D)
4th rowKHP 사업(R&D)
5th rowK-21보병전투차량
ValueCountFrequency (%)
성능개량 9
 
5.8%
2차 5
 
3.2%
울산급 4
 
2.6%
사업 3
 
1.9%
차기 3
 
1.9%
batchⅰ 3
 
1.9%
유도무기 3
 
1.9%
2
 
1.3%
장거리 2
 
1.3%
fa-50 2
 
1.3%
Other values (111) 119
76.8%
2023-12-13T07:51:48.865366image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
72
 
7.9%
- 37
 
4.0%
32
 
3.5%
I 30
 
3.3%
25
 
2.7%
) 19
 
2.1%
( 19
 
2.1%
17
 
1.9%
1 15
 
1.6%
15
 
1.6%
Other values (166) 635
69.3%

Most occurring categories

ValueCountFrequency (%)
Other Letter 527
57.5%
Uppercase Letter 124
 
13.5%
Space Separator 72
 
7.9%
Decimal Number 53
 
5.8%
Lowercase Letter 46
 
5.0%
Dash Punctuation 37
 
4.0%
Close Punctuation 19
 
2.1%
Open Punctuation 19
 
2.1%
Other Punctuation 10
 
1.1%
Letter Number 9
 
1.0%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
32
 
6.1%
25
 
4.7%
17
 
3.2%
15
 
2.8%
15
 
2.8%
15
 
2.8%
13
 
2.5%
13
 
2.5%
13
 
2.5%
12
 
2.3%
Other values (116) 357
67.7%
Uppercase Letter
ValueCountFrequency (%)
I 30
24.2%
K 12
 
9.7%
A 11
 
8.9%
B 10
 
8.1%
R 9
 
7.3%
D 7
 
5.6%
F 6
 
4.8%
U 5
 
4.0%
X 5
 
4.0%
V 4
 
3.2%
Other values (10) 25
20.2%
Lowercase Letter
ValueCountFrequency (%)
c 9
19.6%
h 8
17.4%
t 8
17.4%
a 8
17.4%
m 4
8.7%
k 2
 
4.3%
l 2
 
4.3%
o 1
 
2.2%
b 1
 
2.2%
s 1
 
2.2%
Other values (2) 2
 
4.3%
Decimal Number
ValueCountFrequency (%)
1 15
28.3%
2 11
20.8%
5 10
18.9%
0 9
17.0%
6 3
 
5.7%
3 3
 
5.7%
4 2
 
3.8%
Other Punctuation
ValueCountFrequency (%)
& 7
70.0%
, 1
 
10.0%
/ 1
 
10.0%
· 1
 
10.0%
Letter Number
ValueCountFrequency (%)
4
44.4%
3
33.3%
2
22.2%
Space Separator
ValueCountFrequency (%)
72
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 37
100.0%
Close Punctuation
ValueCountFrequency (%)
) 19
100.0%
Open Punctuation
ValueCountFrequency (%)
( 19
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 527
57.5%
Common 210
 
22.9%
Latin 179
 
19.5%

Most frequent character per script

Hangul
ValueCountFrequency (%)
32
 
6.1%
25
 
4.7%
17
 
3.2%
15
 
2.8%
15
 
2.8%
15
 
2.8%
13
 
2.5%
13
 
2.5%
13
 
2.5%
12
 
2.3%
Other values (116) 357
67.7%
Latin
ValueCountFrequency (%)
I 30
16.8%
K 12
 
6.7%
A 11
 
6.1%
B 10
 
5.6%
R 9
 
5.0%
c 9
 
5.0%
h 8
 
4.5%
t 8
 
4.5%
a 8
 
4.5%
D 7
 
3.9%
Other values (25) 67
37.4%
Common
ValueCountFrequency (%)
72
34.3%
- 37
17.6%
) 19
 
9.0%
( 19
 
9.0%
1 15
 
7.1%
2 11
 
5.2%
5 10
 
4.8%
0 9
 
4.3%
& 7
 
3.3%
6 3
 
1.4%
Other values (5) 8
 
3.8%

Most occurring blocks

ValueCountFrequency (%)
Hangul 527
57.5%
ASCII 379
41.4%
Number Forms 9
 
1.0%
None 1
 
0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
72
19.0%
- 37
 
9.8%
I 30
 
7.9%
) 19
 
5.0%
( 19
 
5.0%
1 15
 
4.0%
K 12
 
3.2%
2 11
 
2.9%
A 11
 
2.9%
5 10
 
2.6%
Other values (36) 143
37.7%
Hangul
ValueCountFrequency (%)
32
 
6.1%
25
 
4.7%
17
 
3.2%
15
 
2.8%
15
 
2.8%
15
 
2.8%
13
 
2.5%
13
 
2.5%
13
 
2.5%
12
 
2.3%
Other values (116) 357
67.7%
Number Forms
ValueCountFrequency (%)
4
44.4%
3
33.3%
2
22.2%
None
ValueCountFrequency (%)
· 1
100.0%
Distinct82
Distinct (%)98.8%
Missing0
Missing (%)0.0%
Memory size796.0 B
2023-12-13T07:51:49.124974image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length43
Median length30
Mean length21.590361
Min length10

Characters and Unicode

Total characters1792
Distinct characters203
Distinct categories11 ?
Distinct scripts4 ?
Distinct blocks4 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique81 ?
Unique (%)97.6%

Sample

1st row군 위성통신장비를 연구 개발로 확보하는 사업
2nd row한국형 기동헬기를 초도 양산하는 사업
3rd rowTA-50 항공기를 경공격기로 개조 개발하는 사업
4th row한국형 기동헬기를 연구 개발하는 사업
5th rowK-21 보병전투차량을 양산 하는 사업
ValueCountFrequency (%)
사업 82
22.3%
확보하는 28
 
7.6%
양산하는 22
 
6.0%
국외구매로 9
 
2.4%
국내건조로 7
 
1.9%
연구개발로 6
 
1.6%
차기 5
 
1.4%
성능개량 4
 
1.1%
하는 3
 
0.8%
3
 
0.8%
Other values (174) 199
54.1%
2023-12-13T07:51:49.508001image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
286
 
16.0%
84
 
4.7%
82
 
4.6%
73
 
4.1%
71
 
4.0%
46
 
2.6%
44
 
2.5%
36
 
2.0%
31
 
1.7%
29
 
1.6%
Other values (193) 1010
56.4%

Most occurring categories

ValueCountFrequency (%)
Other Letter 1325
73.9%
Space Separator 286
 
16.0%
Uppercase Letter 73
 
4.1%
Decimal Number 47
 
2.6%
Dash Punctuation 23
 
1.3%
Lowercase Letter 18
 
1.0%
Open Punctuation 7
 
0.4%
Close Punctuation 7
 
0.4%
Other Punctuation 4
 
0.2%
Letter Number 1
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
84
 
6.3%
82
 
6.2%
73
 
5.5%
71
 
5.4%
46
 
3.5%
44
 
3.3%
36
 
2.7%
31
 
2.3%
29
 
2.2%
28
 
2.1%
Other values (152) 801
60.5%
Uppercase Letter
ValueCountFrequency (%)
I 23
31.5%
K 12
16.4%
A 7
 
9.6%
F 4
 
5.5%
S 4
 
5.5%
P 3
 
4.1%
M 2
 
2.7%
L 2
 
2.7%
G 2
 
2.7%
B 2
 
2.7%
Other values (7) 12
16.4%
Lowercase Letter
ValueCountFrequency (%)
m 4
22.2%
k 2
11.1%
n 2
11.1%
i 2
11.1%
h 2
11.1%
c 2
11.1%
t 2
11.1%
a 2
11.1%
Decimal Number
ValueCountFrequency (%)
1 15
31.9%
5 10
21.3%
0 8
17.0%
3 4
 
8.5%
6 4
 
8.5%
2 4
 
8.5%
4 2
 
4.3%
Other Punctuation
ValueCountFrequency (%)
, 2
50.0%
/ 1
25.0%
& 1
25.0%
Space Separator
ValueCountFrequency (%)
286
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 23
100.0%
Open Punctuation
ValueCountFrequency (%)
( 7
100.0%
Close Punctuation
ValueCountFrequency (%)
) 7
100.0%
Letter Number
ValueCountFrequency (%)
1
100.0%
Math Symbol
ValueCountFrequency (%)
+ 1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 1324
73.9%
Common 375
 
20.9%
Latin 92
 
5.1%
Han 1
 
0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
84
 
6.3%
82
 
6.2%
73
 
5.5%
71
 
5.4%
46
 
3.5%
44
 
3.3%
36
 
2.7%
31
 
2.3%
29
 
2.2%
28
 
2.1%
Other values (151) 800
60.4%
Latin
ValueCountFrequency (%)
I 23
25.0%
K 12
13.0%
A 7
 
7.6%
F 4
 
4.3%
S 4
 
4.3%
m 4
 
4.3%
P 3
 
3.3%
M 2
 
2.2%
k 2
 
2.2%
n 2
 
2.2%
Other values (16) 29
31.5%
Common
ValueCountFrequency (%)
286
76.3%
- 23
 
6.1%
1 15
 
4.0%
5 10
 
2.7%
0 8
 
2.1%
( 7
 
1.9%
) 7
 
1.9%
3 4
 
1.1%
6 4
 
1.1%
2 4
 
1.1%
Other values (5) 7
 
1.9%
Han
ValueCountFrequency (%)
1
100.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 1324
73.9%
ASCII 466
 
26.0%
Number Forms 1
 
0.1%
CJK 1
 
0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
286
61.4%
I 23
 
4.9%
- 23
 
4.9%
1 15
 
3.2%
K 12
 
2.6%
5 10
 
2.1%
0 8
 
1.7%
A 7
 
1.5%
( 7
 
1.5%
) 7
 
1.5%
Other values (30) 68
 
14.6%
Hangul
ValueCountFrequency (%)
84
 
6.3%
82
 
6.2%
73
 
5.5%
71
 
5.4%
46
 
3.5%
44
 
3.3%
36
 
2.7%
31
 
2.3%
29
 
2.2%
28
 
2.1%
Other values (151) 800
60.4%
Number Forms
ValueCountFrequency (%)
1
100.0%
CJK
ValueCountFrequency (%)
1
100.0%
Distinct56
Distinct (%)67.5%
Missing0
Missing (%)0.0%
Memory size796.0 B
2023-12-13T07:51:49.756274image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length12
Median length10
Mean length7.6987952
Min length5

Characters and Unicode

Total characters639
Distinct characters93
Distinct categories6 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique35 ?
Unique (%)42.2%

Sample

1st row위성사업팀
2nd rowKUH사업팀
3rd row경공격기사업팀
4th rowKUH사업팀
5th row장갑차사업팀
ValueCountFrequency (%)
중고도유도무기사업팀 5
 
5.6%
사업팀 4
 
4.5%
포병사업팀 4
 
4.5%
전차사업팀 4
 
4.5%
무인기사업팀 4
 
4.5%
화생방사업팀 3
 
3.4%
전투기사업팀 3
 
3.4%
해상유도무기사업팀 3
 
3.4%
전투함사업팀 3
 
3.4%
상륙함사업팀 3
 
3.4%
Other values (38) 53
59.6%
2023-12-13T07:51:50.125849image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
83
 
13.0%
82
 
12.8%
82
 
12.8%
43
 
6.7%
39
 
6.1%
21
 
3.3%
20
 
3.1%
17
 
2.7%
15
 
2.3%
13
 
2.0%
Other values (83) 224
35.1%

Most occurring categories

ValueCountFrequency (%)
Other Letter 580
90.8%
Space Separator 43
 
6.7%
Uppercase Letter 12
 
1.9%
Other Punctuation 2
 
0.3%
Dash Punctuation 1
 
0.2%
Letter Number 1
 
0.2%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
83
14.3%
82
14.1%
82
14.1%
39
 
6.7%
21
 
3.6%
20
 
3.4%
17
 
2.9%
15
 
2.6%
13
 
2.2%
10
 
1.7%
Other values (74) 198
34.1%
Uppercase Letter
ValueCountFrequency (%)
K 4
33.3%
H 3
25.0%
U 3
25.0%
D 1
 
8.3%
X 1
 
8.3%
Space Separator
ValueCountFrequency (%)
43
100.0%
Other Punctuation
ValueCountFrequency (%)
. 2
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 1
100.0%
Letter Number
ValueCountFrequency (%)
1
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 580
90.8%
Common 46
 
7.2%
Latin 13
 
2.0%

Most frequent character per script

Hangul
ValueCountFrequency (%)
83
14.3%
82
14.1%
82
14.1%
39
 
6.7%
21
 
3.6%
20
 
3.4%
17
 
2.9%
15
 
2.6%
13
 
2.2%
10
 
1.7%
Other values (74) 198
34.1%
Latin
ValueCountFrequency (%)
K 4
30.8%
H 3
23.1%
U 3
23.1%
1
 
7.7%
D 1
 
7.7%
X 1
 
7.7%
Common
ValueCountFrequency (%)
43
93.5%
. 2
 
4.3%
- 1
 
2.2%

Most occurring blocks

ValueCountFrequency (%)
Hangul 580
90.8%
ASCII 58
 
9.1%
Number Forms 1
 
0.2%

Most frequent character per block

Hangul
ValueCountFrequency (%)
83
14.3%
82
14.1%
82
14.1%
39
 
6.7%
21
 
3.6%
20
 
3.4%
17
 
2.9%
15
 
2.6%
13
 
2.2%
10
 
1.7%
Other values (74) 198
34.1%
ASCII
ValueCountFrequency (%)
43
74.1%
K 4
 
6.9%
H 3
 
5.2%
U 3
 
5.2%
. 2
 
3.4%
- 1
 
1.7%
D 1
 
1.7%
X 1
 
1.7%
Number Forms
ValueCountFrequency (%)
1
100.0%

Interactions

2023-12-13T07:51:47.631184image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-13T07:51:50.231481image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
정책실명제등록번호사업명사업요지담당부서
정책실명제등록번호1.0000.7881.0000.895
사업명0.7881.0000.9990.976
사업요지1.0000.9991.0001.000
담당부서0.8950.9761.0001.000

Missing values

2023-12-13T07:51:47.790958image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-13T07:51:47.879957image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

정책실명제등록번호사업명사업요지담당부서
0201301군 위성통신장비군 위성통신장비를 연구 개발로 확보하는 사업위성사업팀
1201302한국형 기동헬기 초도양산한국형 기동헬기를 초도 양산하는 사업KUH사업팀
2201303FA-50 개조개발 (R&D)TA-50 항공기를 경공격기로 개조 개발하는 사업경공격기사업팀
3201304KHP 사업(R&D)한국형 기동헬기를 연구 개발하는 사업KUH사업팀
4201305K-21보병전투차량K-21 보병전투차량을 양산 하는 사업장갑차사업팀
5201306K-1 중구난차량K-1 중구난 차량을 추가로 양산하는 사업전차사업팀
6201307K-2 전차K-2 전차를 양산하는 사업전차사업팀
7201308경구난차량(K-21)경구난차량을 양산하는 사업장갑차사업팀
8201309K-11 복합형소총K-11 복합형소총을 양산하는 사업기동장비사업팀
9201310육군 과학화전투 훈련단 부대개편과학화전투 훈련 규모와 부대를 확장하는 사업부대개편사업팀
정책실명제등록번호사업명사업요지담당부서
73201906장거리레이더장거리레이더를 연구개발하는 사업레이더사업팀
74201907GPS유도폭탄(2,000lbs급)GPS유도폭탄을 구매하는 사업항공유도무기사업팀
75201908Link-16 성능개량Link-16 성능개량 사업연합전술데이터링크사업팀
76201909K1E1전차 성능개량K1E1전차를 성능개량하는 사업전차사업팀
77201910화생방정찰차-II(차량형)화생방정찰차-II(차량형)을 양산하는 사업화생방사업팀
78201911대함유도탄방어유도탄대함유도탄방어유도탄을 확보하는 사업해상유도무기사업팀
79201912중어뢰-II중어뢰-II를 확보하는 사업해상유도무기사업팀
8020191330mm 차륜형대공포30mm차륜형대공포를 확보하는 사업방공유도무기 사업팀
81201914울산급 Batch-III호위함을 국내건조하는 사업호위함사업팀
82201915대형기동헬기-II대형기동헬기를 국외구매하는 사업특수헬기사업팀