Overview

Dataset statistics

Number of variables8
Number of observations10000
Missing cells0
Missing cells (%)0.0%
Duplicate rows5
Duplicate rows (%)0.1%
Total size in memory722.7 KiB
Average record size in memory74.0 B

Variable types

Numeric2
Categorical2
Text2
DateTime2

Alerts

Dataset has 5 (0.1%) duplicate rowsDuplicates

Reproduction

Analysis started2024-03-12 23:33:21.743672
Analysis finished2024-03-12 23:33:22.912716
Duration1.17 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

사업개시년도
Real number (ℝ)

Distinct7
Distinct (%)0.1%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean2020.3035
Minimum2017
Maximum2023
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2024-03-13T08:33:22.956264image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum2017
5-th percentile2017
Q12019
median2020
Q32022
95-th percentile2023
Maximum2023
Range6
Interquartile range (IQR)3

Descriptive statistics

Standard deviation1.9740764
Coefficient of variation (CV)0.00097711871
Kurtosis-1.1917815
Mean2020.3035
Median Absolute Deviation (MAD)2
Skewness-0.17053224
Sum20203035
Variance3.8969774
MonotonicityNot monotonic
2024-03-13T08:33:23.057618image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=7)
ValueCountFrequency (%)
2023 1794
17.9%
2022 1603
16.0%
2020 1503
15.0%
2021 1453
14.5%
2019 1364
13.6%
2018 1207
12.1%
2017 1076
10.8%
ValueCountFrequency (%)
2017 1076
10.8%
2018 1207
12.1%
2019 1364
13.6%
2020 1503
15.0%
2021 1453
14.5%
2022 1603
16.0%
2023 1794
17.9%
ValueCountFrequency (%)
2023 1794
17.9%
2022 1603
16.0%
2021 1453
14.5%
2020 1503
15.0%
2019 1364
13.6%
2018 1207
12.1%
2017 1076
10.8%

시군명
Categorical

Distinct31
Distinct (%)0.3%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
고양시
937 
성남시
791 
수원시
739 
부천시
674 
용인시
 
425
Other values (26)
6434 

Length

Max length4
Median length3
Mean length3.0888
Min length3

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row의왕시
2nd row양주시
3rd row안산시
4th row파주시
5th row수원시

Common Values

ValueCountFrequency (%)
고양시 937
 
9.4%
성남시 791
 
7.9%
수원시 739
 
7.4%
부천시 674
 
6.7%
용인시 425
 
4.2%
화성시 424
 
4.2%
안산시 420
 
4.2%
의정부시 401
 
4.0%
시흥시 399
 
4.0%
남양주시 377
 
3.8%
Other values (21) 4413
44.1%

Length

2024-03-13T08:33:23.177086image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category
ValueCountFrequency (%)
고양시 937
 
9.4%
성남시 791
 
7.9%
수원시 739
 
7.4%
부천시 674
 
6.7%
용인시 425
 
4.2%
화성시 424
 
4.2%
안산시 420
 
4.2%
의정부시 401
 
4.0%
시흥시 399
 
4.0%
남양주시 377
 
3.8%
Other values (21) 4413
44.1%
Distinct351
Distinct (%)3.5%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
2024-03-13T08:33:23.346570image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length25
Median length22
Mean length9.6335
Min length3

Characters and Unicode

Total characters96335
Distinct characters213
Distinct categories6 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique9 ?
Unique (%)0.1%

Sample

1st row의왕시니어클럽
2nd row양주시 회천노인복지관 양주실버인력뱅크
3rd row대한노인회 경기 안산상록구지회
4th row파주시니어클럽
5th rowSK청솔노인복지관
ValueCountFrequency (%)
경기 715
 
5.7%
대한노인회 566
 
4.5%
화성시니어클럽 236
 
1.9%
부천시니어클럽 215
 
1.7%
안산시니어클럽 214
 
1.7%
실버인력뱅크 207
 
1.7%
고양시니어클럽 188
 
1.5%
수원시니어클럽 167
 
1.3%
시흥시니어클럽 157
 
1.3%
군포시니어클럽 157
 
1.3%
Other values (336) 9647
77.4%
2024-03-13T08:33:23.618282image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
6683
 
6.9%
6446
 
6.7%
5770
 
6.0%
4253
 
4.4%
4215
 
4.4%
3882
 
4.0%
2893
 
3.0%
2786
 
2.9%
2785
 
2.9%
2785
 
2.9%
Other values (203) 53837
55.9%

Most occurring categories

ValueCountFrequency (%)
Other Letter 93074
96.6%
Space Separator 2474
 
2.6%
Close Punctuation 305
 
0.3%
Open Punctuation 270
 
0.3%
Uppercase Letter 160
 
0.2%
Decimal Number 52
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
6683
 
7.2%
6446
 
6.9%
5770
 
6.2%
4253
 
4.6%
4215
 
4.5%
3882
 
4.2%
2893
 
3.1%
2786
 
3.0%
2785
 
3.0%
2785
 
3.0%
Other values (193) 50576
54.3%
Uppercase Letter
ValueCountFrequency (%)
S 58
36.2%
K 58
36.2%
Y 11
 
6.9%
A 11
 
6.9%
C 11
 
6.9%
M 11
 
6.9%
Space Separator
ValueCountFrequency (%)
2474
100.0%
Close Punctuation
ValueCountFrequency (%)
) 305
100.0%
Open Punctuation
ValueCountFrequency (%)
( 270
100.0%
Decimal Number
ValueCountFrequency (%)
1 52
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 93074
96.6%
Common 3101
 
3.2%
Latin 160
 
0.2%

Most frequent character per script

Hangul
ValueCountFrequency (%)
6683
 
7.2%
6446
 
6.9%
5770
 
6.2%
4253
 
4.6%
4215
 
4.5%
3882
 
4.2%
2893
 
3.1%
2786
 
3.0%
2785
 
3.0%
2785
 
3.0%
Other values (193) 50576
54.3%
Latin
ValueCountFrequency (%)
S 58
36.2%
K 58
36.2%
Y 11
 
6.9%
A 11
 
6.9%
C 11
 
6.9%
M 11
 
6.9%
Common
ValueCountFrequency (%)
2474
79.8%
) 305
 
9.8%
( 270
 
8.7%
1 52
 
1.7%

Most occurring blocks

ValueCountFrequency (%)
Hangul 93074
96.6%
ASCII 3261
 
3.4%

Most frequent character per block

Hangul
ValueCountFrequency (%)
6683
 
7.2%
6446
 
6.9%
5770
 
6.2%
4253
 
4.6%
4215
 
4.5%
3882
 
4.2%
2893
 
3.1%
2786
 
3.0%
2785
 
3.0%
2785
 
3.0%
Other values (193) 50576
54.3%
ASCII
ValueCountFrequency (%)
2474
75.9%
) 305
 
9.4%
( 270
 
8.3%
S 58
 
1.8%
K 58
 
1.8%
1 52
 
1.6%
Y 11
 
0.3%
A 11
 
0.3%
C 11
 
0.3%
M 11
 
0.3%

사업유형명
Categorical

Distinct4
Distinct (%)< 0.1%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
공익활동형
6605 
시장형
2093 
사회서비스형
1269 
취업알선형
 
33

Length

Max length6
Median length5
Mean length4.7083
Min length3

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row공익활동형
2nd row시장형
3rd row공익활동형
4th row시장형
5th row사회서비스형

Common Values

ValueCountFrequency (%)
공익활동형 6605
66.0%
시장형 2093
 
20.9%
사회서비스형 1269
 
12.7%
취업알선형 33
 
0.3%

Length

2024-03-13T08:33:23.725803image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2024-03-13T08:33:23.807513image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
공익활동형 6605
66.0%
시장형 2093
 
20.9%
사회서비스형 1269
 
12.7%
취업알선형 33
 
0.3%
Distinct4450
Distinct (%)44.5%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
2024-03-13T08:33:24.014049image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length48
Median length39
Mean length9.9861
Min length2

Characters and Unicode

Total characters99861
Distinct characters665
Distinct categories14 ?
Distinct scripts4 ?
Distinct blocks7 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique2577 ?
Unique (%)25.8%

Sample

1st row환경지킴이 1 사업단
2nd row반찬 및 도시락 사업단
3rd row경로당 도우미(우리서로도우미)
4th row청춘찬찬찬
5th row사회서비스형 보육시설지원
ValueCountFrequency (%)
노노케어 304
 
2.1%
195
 
1.3%
사업 168
 
1.1%
지원 161
 
1.1%
노노카페 129
 
0.9%
시니어 121
 
0.8%
사업단 100
 
0.7%
도우미 98
 
0.7%
경로당 92
 
0.6%
봉사 84
 
0.6%
Other values (4403) 13171
90.1%
2024-03-13T08:33:24.352363image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
4633
 
4.6%
4592
 
4.6%
4353
 
4.4%
2672
 
2.7%
2640
 
2.6%
2605
 
2.6%
2083
 
2.1%
1915
 
1.9%
1906
 
1.9%
( 1763
 
1.8%
Other values (655) 70699
70.8%

Most occurring categories

ValueCountFrequency (%)
Other Letter 88415
88.5%
Space Separator 4633
 
4.6%
Open Punctuation 1783
 
1.8%
Close Punctuation 1781
 
1.8%
Decimal Number 1440
 
1.4%
Other Punctuation 746
 
0.7%
Uppercase Letter 377
 
0.4%
Dash Punctuation 332
 
0.3%
Lowercase Letter 126
 
0.1%
Connector Punctuation 109
 
0.1%
Other values (4) 119
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
4592
 
5.2%
4353
 
4.9%
2672
 
3.0%
2640
 
3.0%
2605
 
2.9%
2083
 
2.4%
1915
 
2.2%
1906
 
2.2%
1653
 
1.9%
1649
 
1.9%
Other values (577) 62347
70.5%
Lowercase Letter
ValueCountFrequency (%)
o 14
11.1%
m 14
11.1%
e 13
10.3%
h 10
 
7.9%
a 9
 
7.1%
s 9
 
7.1%
i 9
 
7.1%
c 9
 
7.1%
t 6
 
4.8%
u 6
 
4.8%
Other values (11) 27
21.4%
Uppercase Letter
ValueCountFrequency (%)
E 47
12.5%
C 43
11.4%
T 36
9.5%
A 33
8.8%
G 31
8.2%
M 29
7.7%
I 28
7.4%
S 23
 
6.1%
V 21
 
5.6%
F 19
 
5.0%
Other values (9) 67
17.8%
Other Punctuation
ValueCountFrequency (%)
' 330
44.2%
" 200
26.8%
& 66
 
8.8%
, 48
 
6.4%
: 35
 
4.7%
. 27
 
3.6%
! 18
 
2.4%
· 15
 
2.0%
? 4
 
0.5%
% 2
 
0.3%
Decimal Number
ValueCountFrequency (%)
2 378
26.2%
1 319
22.2%
3 190
13.2%
0 188
13.1%
9 160
11.1%
6 77
 
5.3%
5 65
 
4.5%
8 26
 
1.8%
7 19
 
1.3%
4 18
 
1.2%
Open Punctuation
ValueCountFrequency (%)
( 1763
98.9%
[ 19
 
1.1%
1
 
0.1%
Close Punctuation
ValueCountFrequency (%)
) 1761
98.9%
] 19
 
1.1%
1
 
0.1%
Letter Number
ValueCountFrequency (%)
13
44.8%
12
41.4%
4
 
13.8%
Final Punctuation
ValueCountFrequency (%)
28
63.6%
16
36.4%
Initial Punctuation
ValueCountFrequency (%)
28
63.6%
16
36.4%
Space Separator
ValueCountFrequency (%)
4633
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 332
100.0%
Connector Punctuation
ValueCountFrequency (%)
_ 109
100.0%
Modifier Symbol
ValueCountFrequency (%)
` 2
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 88381
88.5%
Common 10914
 
10.9%
Latin 532
 
0.5%
Han 34
 
< 0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
4592
 
5.2%
4353
 
4.9%
2672
 
3.0%
2640
 
3.0%
2605
 
2.9%
2083
 
2.4%
1915
 
2.2%
1906
 
2.2%
1653
 
1.9%
1649
 
1.9%
Other values (567) 62313
70.5%
Latin
ValueCountFrequency (%)
E 47
 
8.8%
C 43
 
8.1%
T 36
 
6.8%
A 33
 
6.2%
G 31
 
5.8%
M 29
 
5.5%
I 28
 
5.3%
S 23
 
4.3%
V 21
 
3.9%
F 19
 
3.6%
Other values (33) 222
41.7%
Common
ValueCountFrequency (%)
4633
42.5%
( 1763
 
16.2%
) 1761
 
16.1%
2 378
 
3.5%
- 332
 
3.0%
' 330
 
3.0%
1 319
 
2.9%
" 200
 
1.8%
3 190
 
1.7%
0 188
 
1.7%
Other values (25) 820
 
7.5%
Han
ValueCountFrequency (%)
12
35.3%
4
 
11.8%
4
 
11.8%
3
 
8.8%
2
 
5.9%
2
 
5.9%
2
 
5.9%
2
 
5.9%
2
 
5.9%
1
 
2.9%

Most occurring blocks

ValueCountFrequency (%)
Hangul 88381
88.5%
ASCII 11312
 
11.3%
Punctuation 88
 
0.1%
Number Forms 29
 
< 0.1%
CJK 22
 
< 0.1%
None 17
 
< 0.1%
CJK Compat Ideographs 12
 
< 0.1%

Most frequent character per block

ASCII
ValueCountFrequency (%)
4633
41.0%
( 1763
 
15.6%
) 1761
 
15.6%
2 378
 
3.3%
- 332
 
2.9%
' 330
 
2.9%
1 319
 
2.8%
" 200
 
1.8%
3 190
 
1.7%
0 188
 
1.7%
Other values (58) 1218
 
10.8%
Hangul
ValueCountFrequency (%)
4592
 
5.2%
4353
 
4.9%
2672
 
3.0%
2640
 
3.0%
2605
 
2.9%
2083
 
2.4%
1915
 
2.2%
1906
 
2.2%
1653
 
1.9%
1649
 
1.9%
Other values (567) 62313
70.5%
Punctuation
ValueCountFrequency (%)
28
31.8%
28
31.8%
16
18.2%
16
18.2%
None
ValueCountFrequency (%)
· 15
88.2%
1
 
5.9%
1
 
5.9%
Number Forms
ValueCountFrequency (%)
13
44.8%
12
41.4%
4
 
13.8%
CJK Compat Ideographs
ValueCountFrequency (%)
12
100.0%
CJK
ValueCountFrequency (%)
4
18.2%
4
18.2%
3
13.6%
2
9.1%
2
9.1%
2
9.1%
2
9.1%
2
9.1%
1
 
4.5%

사업량(명)
Real number (ℝ)

Distinct327
Distinct (%)3.3%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean53.1972
Minimum1
Maximum1051
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size166.0 KiB
2024-03-13T08:33:24.497775image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1
5-th percentile6
Q115
median30
Q368
95-th percentile165
Maximum1051
Range1050
Interquartile range (IQR)53

Descriptive statistics

Standard deviation65.743554
Coefficient of variation (CV)1.2358461
Kurtosis30.747505
Mean53.1972
Median Absolute Deviation (MAD)20
Skewness4.1469449
Sum531972
Variance4322.2149
MonotonicityNot monotonic
2024-03-13T08:33:24.638942image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
10 673
 
6.7%
20 665
 
6.7%
30 549
 
5.5%
40 415
 
4.2%
12 328
 
3.3%
50 310
 
3.1%
8 278
 
2.8%
15 271
 
2.7%
60 266
 
2.7%
100 232
 
2.3%
Other values (317) 6013
60.1%
ValueCountFrequency (%)
1 14
 
0.1%
2 84
 
0.8%
3 54
 
0.5%
4 136
 
1.4%
5 111
 
1.1%
6 184
 
1.8%
7 75
 
0.8%
8 278
2.8%
9 103
 
1.0%
10 673
6.7%
ValueCountFrequency (%)
1051 1
 
< 0.1%
1014 1
 
< 0.1%
800 2
< 0.1%
711 1
 
< 0.1%
700 3
< 0.1%
690 1
 
< 0.1%
663 1
 
< 0.1%
653 1
 
< 0.1%
620 1
 
< 0.1%
619 1
 
< 0.1%
Distinct226
Distinct (%)2.3%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum2017-01-01 00:00:00
Maximum2023-07-01 00:00:00
2024-03-13T08:33:24.742190image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2024-03-13T08:33:24.840937image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
Distinct81
Distinct (%)0.8%
Missing0
Missing (%)0.0%
Memory size156.2 KiB
Minimum2017-02-28 00:00:00
Maximum2023-12-31 00:00:00
2024-03-13T08:33:24.944906image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2024-03-13T08:33:25.067109image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)

Interactions

2024-03-13T08:33:22.544774image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2024-03-13T08:33:22.381157image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2024-03-13T08:33:22.639827image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
2024-03-13T08:33:22.458611image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2024-03-13T08:33:25.133399image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
사업개시년도시군명사업유형명사업량(명)사업종료일자
사업개시년도1.0000.0000.2650.0601.000
시군명0.0001.0000.2050.2940.579
사업유형명0.2650.2051.0000.2070.643
사업량(명)0.0600.2940.2071.0000.224
사업종료일자1.0000.5790.6430.2241.000
2024-03-13T08:33:25.208090image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
사업유형명시군명
사업유형명1.0000.107
시군명0.1071.000
2024-03-13T08:33:25.286662image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
사업개시년도사업량(명)시군명사업유형명
사업개시년도1.0000.0820.0000.189
사업량(명)0.0821.0000.1140.133
시군명0.0000.1141.0000.107
사업유형명0.1890.1330.1071.000

Missing values

2024-03-13T08:33:22.745424image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2024-03-13T08:33:22.857220image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

사업개시년도시군명기관명사업유형명사업내용사업량(명)사업시작일자사업종료일자
14072023의왕시의왕시니어클럽공익활동형환경지킴이 1 사업단652023-01-012023-11-30
58092020양주시양주시 회천노인복지관 양주실버인력뱅크시장형반찬 및 도시락 사업단242020-03-022020-12-31
83902018안산시대한노인회 경기 안산상록구지회공익활동형경로당 도우미(우리서로도우미)2212018-01-022018-12-31
15752023파주시파주시니어클럽시장형청춘찬찬찬182023-01-012023-12-31
40502021수원시SK청솔노인복지관사회서비스형사회서비스형 보육시설지원102021-01-012021-11-30
81682018성남시수정중앙노인종합복지관공익활동형노인돌봄지원사업(9개월)192018-03-012018-12-31
28692022양평군양평군종합사회복지관사회서비스형행복사랑 노인지킴이72022-01-102022-11-30
70222019수원시팔달노인복지관공익활동형행궁전통지킴이662019-02-012019-10-31
4392023남양주시대한노인회 경기 남양주시지회공익활동형환경개선사업1902023-02-012023-12-31
75032019의정부시의정부시니어클럽시장형시니어재활용사업단102019-01-012019-12-31
사업개시년도시군명기관명사업유형명사업내용사업량(명)사업시작일자사업종료일자
13032023오산시오산시니어클럽시장형소리울카페162023-01-092023-12-31
13582023용인시뉴딜사회적협동조합사회서비스형노지혜서 꿈을 키운다252023-01-012023-11-30
55542020수원시수원시광교노인복지관공익활동형노노케어602020-01-012020-12-31
84612018안양시안양시노인종합복지관공익활동형EM환경지킴이402018-02-012018-11-30
53882020성남시수정중앙노인종합복지관공익활동형장난감 클린 도우미 사업102020-01-012020-11-30
52242020남양주시남양주시니어클럽사회서비스형노인맞춤돌봄서비스지원사업302020-03-012020-12-31
3462023김포시김포시니어클럽공익활동형보육시설도우미(공익형)902023-01-022023-11-30
79212018구리시구리시니어클럽공익활동형무지개지킴이사업단302018-03-012018-12-28
52472020남양주시남양주실버인력뱅크사회서비스형노인시설지원사업702020-02-012020-11-30
52292020남양주시남양주시니어클럽시장형실버카페운영사업(행복일번지분식카페)152020-01-012020-12-31

Duplicate rows

Most frequently occurring

사업개시년도시군명기관명사업유형명사업내용사업량(명)사업시작일자사업종료일자# duplicates
02017평택시대한노인회평택지회공익활동형노노케어(연중)502017-01-092017-12-312
12018안양시경기실버포럼공익활동형노노케어(경기실버포럼)482018-03-012018-11-302
22020화성시경기 화성시지회시장형짚풀수공예제작162020-01-012020-12-312
32020화성시화성시남부노인복지관사회서비스형시니어선생님182020-03-022020-12-312
42023하남시하남시니어클럽시장형이음누리재봉152023-01-012023-12-312