Overview

Dataset statistics

Number of variables4
Number of observations167
Missing cells0
Missing cells (%)0.0%
Duplicate rows2
Duplicate rows (%)1.2%
Total size in memory5.5 KiB
Average record size in memory33.8 B

Variable types

Categorical1
Text1
DateTime1
Numeric1

Dataset

Description방위사업청의 최근 SW 및 ICT(컴퓨팅, 네트워크)장비 구매현황에 대한 데이터로 분류, 제품명, 계약일자, 금액 정보가 제공됩니다.
Author방위사업청
URLhttps://www.data.go.kr/data/15104354/fileData.do

Alerts

Dataset has 2 (1.2%) duplicate rowsDuplicates

Reproduction

Analysis started2023-12-12 10:39:07.042905
Analysis finished2023-12-12 10:39:07.535364
Duration0.49 seconds
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

분류
Categorical

Distinct3
Distinct (%)1.8%
Missing0
Missing (%)0.0%
Memory size1.4 KiB
SW
126 
HW > 컴퓨팅장비
24 
HW > 네트워크 장비
17 

Length

Max length12
Median length2
Mean length4.1676647
Min length2

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st rowSW
2nd rowSW
3rd rowSW
4th rowSW
5th rowSW

Common Values

ValueCountFrequency (%)
SW 126
75.4%
HW > 컴퓨팅장비 24
 
14.4%
HW > 네트워크 장비 17
 
10.2%

Length

2023-12-12T19:39:07.612900image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2023-12-12T19:39:07.750622image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
sw 126
47.4%
hw 41
 
15.4%
41
 
15.4%
컴퓨팅장비 24
 
9.0%
네트워크 17
 
6.4%
장비 17
 
6.4%
Distinct159
Distinct (%)95.2%
Missing0
Missing (%)0.0%
Memory size1.4 KiB
2023-12-12T19:39:08.034056image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length67
Median length36
Mean length18.922156
Min length8

Characters and Unicode

Total characters3160
Distinct characters246
Distinct categories11 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique152 ?
Unique (%)91.0%

Sample

1st row2021_국방망 데스크탑 가상화 사용자용 방화벽
2nd row2021_CTI DB용 MSSQL 2019로 교체
3rd row2021_CTI, CTI DB, CTI REC, CTI IVR, 영상회의, APM 서버용 OS 윈도우 서버2019로 교체
4th row2021_VOIP 교환용 인사정보 반영 교체
5th row2021_전자조달 모바일 간편인증
ValueCountFrequency (%)
서버 13
 
3.1%
sw 13
 
3.1%
증설 10
 
2.4%
2021_국방망 8
 
1.9%
교체 8
 
1.9%
2020_통합관제 8
 
1.9%
메모리 7
 
1.7%
2021_인터넷망 6
 
1.4%
v2.0 5
 
1.2%
데스크탑 5
 
1.2%
Other values (248) 331
80.0%
2023-12-12T19:39:08.511875image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
2 353
 
11.2%
0 287
 
9.1%
247
 
7.8%
_ 167
 
5.3%
1 88
 
2.8%
S 69
 
2.2%
e 55
 
1.7%
50
 
1.6%
( 41
 
1.3%
) 41
 
1.3%
Other values (236) 1762
55.8%

Most occurring categories

ValueCountFrequency (%)
Other Letter 1115
35.3%
Decimal Number 758
24.0%
Uppercase Letter 378
 
12.0%
Lowercase Letter 362
 
11.5%
Space Separator 247
 
7.8%
Connector Punctuation 167
 
5.3%
Other Punctuation 43
 
1.4%
Open Punctuation 41
 
1.3%
Close Punctuation 41
 
1.3%
Dash Punctuation 4
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
50
 
4.5%
38
 
3.4%
37
 
3.3%
28
 
2.5%
27
 
2.4%
23
 
2.1%
22
 
2.0%
22
 
2.0%
22
 
2.0%
22
 
2.0%
Other values (168) 824
73.9%
Uppercase Letter
ValueCountFrequency (%)
S 69
18.3%
A 41
10.8%
W 34
9.0%
B 28
 
7.4%
M 27
 
7.1%
D 24
 
6.3%
C 23
 
6.1%
I 18
 
4.8%
P 18
 
4.8%
N 16
 
4.2%
Other values (14) 80
21.2%
Lowercase Letter
ValueCountFrequency (%)
e 55
15.2%
n 40
11.0%
o 33
 
9.1%
t 30
 
8.3%
a 25
 
6.9%
r 24
 
6.6%
g 21
 
5.8%
i 18
 
5.0%
s 17
 
4.7%
v 13
 
3.6%
Other values (14) 86
23.8%
Decimal Number
ValueCountFrequency (%)
2 353
46.6%
0 287
37.9%
1 88
 
11.6%
4 9
 
1.2%
6 7
 
0.9%
5 6
 
0.8%
3 4
 
0.5%
9 3
 
0.4%
7 1
 
0.1%
Other Punctuation
ValueCountFrequency (%)
, 14
32.6%
/ 12
27.9%
. 12
27.9%
# 5
 
11.6%
Math Symbol
ValueCountFrequency (%)
+ 3
75.0%
~ 1
 
25.0%
Space Separator
ValueCountFrequency (%)
247
100.0%
Connector Punctuation
ValueCountFrequency (%)
_ 167
100.0%
Open Punctuation
ValueCountFrequency (%)
( 41
100.0%
Close Punctuation
ValueCountFrequency (%)
) 41
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 4
100.0%

Most occurring scripts

ValueCountFrequency (%)
Common 1305
41.3%
Hangul 1115
35.3%
Latin 740
23.4%

Most frequent character per script

Hangul
ValueCountFrequency (%)
50
 
4.5%
38
 
3.4%
37
 
3.3%
28
 
2.5%
27
 
2.4%
23
 
2.1%
22
 
2.0%
22
 
2.0%
22
 
2.0%
22
 
2.0%
Other values (168) 824
73.9%
Latin
ValueCountFrequency (%)
S 69
 
9.3%
e 55
 
7.4%
A 41
 
5.5%
n 40
 
5.4%
W 34
 
4.6%
o 33
 
4.5%
t 30
 
4.1%
B 28
 
3.8%
M 27
 
3.6%
a 25
 
3.4%
Other values (38) 358
48.4%
Common
ValueCountFrequency (%)
2 353
27.0%
0 287
22.0%
247
18.9%
_ 167
12.8%
1 88
 
6.7%
( 41
 
3.1%
) 41
 
3.1%
, 14
 
1.1%
/ 12
 
0.9%
. 12
 
0.9%
Other values (10) 43
 
3.3%

Most occurring blocks

ValueCountFrequency (%)
ASCII 2045
64.7%
Hangul 1115
35.3%

Most frequent character per block

ASCII
ValueCountFrequency (%)
2 353
17.3%
0 287
14.0%
247
 
12.1%
_ 167
 
8.2%
1 88
 
4.3%
S 69
 
3.4%
e 55
 
2.7%
( 41
 
2.0%
) 41
 
2.0%
A 41
 
2.0%
Other values (58) 656
32.1%
Hangul
ValueCountFrequency (%)
50
 
4.5%
38
 
3.4%
37
 
3.3%
28
 
2.5%
27
 
2.4%
23
 
2.1%
22
 
2.0%
22
 
2.0%
22
 
2.0%
22
 
2.0%
Other values (168) 824
73.9%
Distinct28
Distinct (%)16.8%
Missing0
Missing (%)0.0%
Memory size1.4 KiB
Minimum2020-01-01 00:00:00
Maximum2021-11-09 00:00:00
2023-12-12T19:39:08.706028image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2023-12-12T19:39:08.917876image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=28)

금액(천원)
Real number (ℝ)

Distinct144
Distinct (%)86.2%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean61590.922
Minimum540
Maximum807750
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size1.6 KiB
2023-12-12T19:39:09.127108image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum540
5-th percentile1958
Q111286
median30000
Q367680
95-th percentile244226.4
Maximum807750
Range807210
Interquartile range (IQR)56394

Descriptive statistics

Standard deviation103456.82
Coefficient of variation (CV)1.6797414
Kurtosis21.870979
Mean61590.922
Median Absolute Deviation (MAD)21875
Skewness4.1115619
Sum10285684
Variance1.0703314 × 1010
MonotonicityNot monotonic
2023-12-12T19:39:09.339239image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
33000 3
 
1.8%
30000 3
 
1.8%
49500 2
 
1.2%
660 2
 
1.2%
135200 2
 
1.2%
22500 2
 
1.2%
11900 2
 
1.2%
16600 2
 
1.2%
39800 2
 
1.2%
32340 2
 
1.2%
Other values (134) 145
86.8%
ValueCountFrequency (%)
540 1
0.6%
660 2
1.2%
792 1
0.6%
880 1
0.6%
1044 1
0.6%
1048 1
0.6%
1518 1
0.6%
1940 1
0.6%
2000 1
0.6%
2088 1
0.6%
ValueCountFrequency (%)
807750 1
0.6%
622160 1
0.6%
432000 1
0.6%
387420 1
0.6%
360000 1
0.6%
279887 1
0.6%
272250 1
0.6%
261800 1
0.6%
252252 1
0.6%
225500 1
0.6%

Interactions

2023-12-12T19:39:07.252494image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2023-12-12T19:39:09.452395image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
분류계약일자금액(천원)
분류1.0000.9520.182
계약일자0.9521.0000.758
금액(천원)0.1820.7581.000
2023-12-12T19:39:09.557121image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
금액(천원)분류
금액(천원)1.0000.116
분류0.1161.000

Missing values

2023-12-12T19:39:07.390081image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2023-12-12T19:39:07.492400image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

분류제품명계약일자금액(천원)
0SW2021_국방망 데스크탑 가상화 사용자용 방화벽2021-10-085500
1SW2021_CTI DB용 MSSQL 2019로 교체2021-10-0811066
2SW2021_CTI, CTI DB, CTI REC, CTI IVR, 영상회의, APM 서버용 OS 윈도우 서버2019로 교체2021-10-0810626
3SW2021_VOIP 교환용 인사정보 반영 교체2021-10-0833000
4SW2021_전자조달 모바일 간편인증2021-10-0829700
5SW2021_제안서평가용 ORACLE DBMS 교체2021-10-0848400
6SW2021_DB 모니터링2021-10-08432000
7SW2021_DB접근제어시스템2021-10-08215535
8SW2021_WAS 모니터링2021-10-08166320
9SW2021_데스크탑가상화 SW2021-10-0854000
분류제품명계약일자금액(천원)
157HW > 네트워크 장비2020_무정전전원장치(UPS)2020-09-1880000
158HW > 네트워크 장비2020_방화벽2020-09-1885000
159HW > 네트워크 장비2020_백업용 SAN스위치2020-09-1858000
160HW > 네트워크 장비2020_시스템용 SAN스위치2020-09-18116000
161HW > 네트워크 장비2020_통신센터 서버백업2020-09-1835160
162HW > 네트워크 장비2020_통합서버2020-09-1868000
163HW > 네트워크 장비2020_항온항습기2020-09-1820000
164HW > 네트워크 장비2020_무정전전원장치(UPS)용 축전지2020-09-1813800
165HW > 네트워크 장비2020_서버메모리 증설2020-09-183894
166HW > 네트워크 장비2020_스토리지 증설2020-09-18132000

Duplicate rows

Most frequently occurring

분류제품명계약일자금액(천원)# duplicates
0SW2020_DBMS2020-03-23166652
1SW2020_DB품질진단툴2020-03-23434502