Overview

Dataset statistics

Number of variables7
Number of observations150
Missing cells85
Missing cells (%)8.1%
Duplicate rows0
Duplicate rows (%)0.0%
Total size in memory8.5 KiB
Average record size in memory57.9 B

Variable types

Numeric1
Categorical3
Text3

Dataset

Description경기도 용인시 구인, 구직의 신청을 받아 구직자, 구인자를 탐색하거나 모집하여 구인자와 구직자 간에 고용계약이 성립되도록 알선하는 직업소개소현황입니다. 유무료구분, 법인명, 사업소주소 등의 데이터를 제공합니다.※ 데이터기준일자 : 2024-01-31
Author경기도 용인시
URLhttps://www.data.go.kr/data/3078096/fileData.do

Alerts

데이터기준일자 has constant value ""Constant
유무료구분 is highly overall correlated with 법인개인구분High correlation
법인개인구분 is highly overall correlated with 유무료구분High correlation
유무료구분 is highly imbalanced (64.7%)Imbalance
사업소전화번호 has 85 (56.7%) missing valuesMissing
순번 has unique valuesUnique
사업소주소 has unique valuesUnique

Reproduction

Analysis started2024-05-04 08:15:09.347858
Analysis finished2024-05-04 08:15:11.061810
Duration1.71 second
Software versionydata-profiling vv4.5.1
Download configurationconfig.json

Variables

순번
Real number (ℝ)

UNIQUE 

Distinct150
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Infinite0
Infinite (%)0.0%
Mean75.5
Minimum1
Maximum150
Zeros0
Zeros (%)0.0%
Negative0
Negative (%)0.0%
Memory size1.4 KiB
2024-05-04T08:15:11.279965image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Quantile statistics

Minimum1
5-th percentile8.45
Q138.25
median75.5
Q3112.75
95-th percentile142.55
Maximum150
Range149
Interquartile range (IQR)74.5

Descriptive statistics

Standard deviation43.445368
Coefficient of variation (CV)0.57543534
Kurtosis-1.2
Mean75.5
Median Absolute Deviation (MAD)37.5
Skewness0
Sum11325
Variance1887.5
MonotonicityStrictly increasing
2024-05-04T08:15:12.061717image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram with fixed size bins (bins=50)
ValueCountFrequency (%)
1 1
 
0.7%
96 1
 
0.7%
98 1
 
0.7%
99 1
 
0.7%
100 1
 
0.7%
101 1
 
0.7%
102 1
 
0.7%
103 1
 
0.7%
104 1
 
0.7%
105 1
 
0.7%
Other values (140) 140
93.3%
ValueCountFrequency (%)
1 1
0.7%
2 1
0.7%
3 1
0.7%
4 1
0.7%
5 1
0.7%
6 1
0.7%
7 1
0.7%
8 1
0.7%
9 1
0.7%
10 1
0.7%
ValueCountFrequency (%)
150 1
0.7%
149 1
0.7%
148 1
0.7%
147 1
0.7%
146 1
0.7%
145 1
0.7%
144 1
0.7%
143 1
0.7%
142 1
0.7%
141 1
0.7%

유무료구분
Categorical

HIGH CORRELATION  IMBALANCE 

Distinct2
Distinct (%)1.3%
Missing0
Missing (%)0.0%
Memory size1.3 KiB
유료
140 
무료
 
10

Length

Max length2
Median length2
Mean length2
Min length2

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row유료
2nd row유료
3rd row유료
4th row유료
5th row유료

Common Values

ValueCountFrequency (%)
유료 140
93.3%
무료 10
 
6.7%

Length

2024-05-04T08:15:12.694075image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2024-05-04T08:15:13.048138image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
유료 140
93.3%
무료 10
 
6.7%
Distinct148
Distinct (%)98.7%
Missing0
Missing (%)0.0%
Memory size1.3 KiB
2024-05-04T08:15:13.666226image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length20
Median length15
Mean length7.0866667
Min length2

Characters and Unicode

Total characters1063
Distinct characters229
Distinct categories7 ?
Distinct scripts3 ?
Distinct blocks3 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique146 ?
Unique (%)97.3%

Sample

1st row조흥인력
2nd row떼부자 인력사무소
3rd row현대인력사무소
4th row드림인력 용인점
5th rowTS인력
ValueCountFrequency (%)
주식회사 6
 
3.2%
용인지사 3
 
1.6%
대영인력 2
 
1.1%
우리인력 2
 
1.1%
2
 
1.1%
일가자 2
 
1.1%
모두인력 2
 
1.1%
미래인력개발 1
 
0.5%
조흥인력 1
 
0.5%
신갈인력공사 1
 
0.5%
Other values (165) 165
88.2%
2024-05-04T08:15:15.069104image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
117
 
11.0%
82
 
7.7%
37
 
3.5%
30
 
2.8%
24
 
2.3%
21
 
2.0%
19
 
1.8%
19
 
1.8%
18
 
1.7%
17
 
1.6%
Other values (219) 679
63.9%

Most occurring categories

ValueCountFrequency (%)
Other Letter 976
91.8%
Space Separator 37
 
3.5%
Open Punctuation 16
 
1.5%
Close Punctuation 16
 
1.5%
Uppercase Letter 12
 
1.1%
Other Punctuation 4
 
0.4%
Lowercase Letter 2
 
0.2%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
117
 
12.0%
82
 
8.4%
30
 
3.1%
24
 
2.5%
21
 
2.2%
19
 
1.9%
19
 
1.9%
18
 
1.8%
17
 
1.7%
17
 
1.7%
Other values (205) 612
62.7%
Uppercase Letter
ValueCountFrequency (%)
S 3
25.0%
T 2
16.7%
K 2
16.7%
M 2
16.7%
G 1
 
8.3%
H 1
 
8.3%
Y 1
 
8.3%
Other Punctuation
ValueCountFrequency (%)
. 2
50.0%
· 2
50.0%
Lowercase Letter
ValueCountFrequency (%)
o 1
50.0%
a 1
50.0%
Space Separator
ValueCountFrequency (%)
37
100.0%
Open Punctuation
ValueCountFrequency (%)
( 16
100.0%
Close Punctuation
ValueCountFrequency (%)
) 16
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 976
91.8%
Common 73
 
6.9%
Latin 14
 
1.3%

Most frequent character per script

Hangul
ValueCountFrequency (%)
117
 
12.0%
82
 
8.4%
30
 
3.1%
24
 
2.5%
21
 
2.2%
19
 
1.9%
19
 
1.9%
18
 
1.8%
17
 
1.7%
17
 
1.7%
Other values (205) 612
62.7%
Latin
ValueCountFrequency (%)
S 3
21.4%
T 2
14.3%
K 2
14.3%
M 2
14.3%
o 1
 
7.1%
a 1
 
7.1%
G 1
 
7.1%
H 1
 
7.1%
Y 1
 
7.1%
Common
ValueCountFrequency (%)
37
50.7%
( 16
21.9%
) 16
21.9%
. 2
 
2.7%
· 2
 
2.7%

Most occurring blocks

ValueCountFrequency (%)
Hangul 976
91.8%
ASCII 85
 
8.0%
None 2
 
0.2%

Most frequent character per block

Hangul
ValueCountFrequency (%)
117
 
12.0%
82
 
8.4%
30
 
3.1%
24
 
2.5%
21
 
2.2%
19
 
1.9%
19
 
1.9%
18
 
1.8%
17
 
1.7%
17
 
1.7%
Other values (205) 612
62.7%
ASCII
ValueCountFrequency (%)
37
43.5%
( 16
18.8%
) 16
18.8%
S 3
 
3.5%
. 2
 
2.4%
T 2
 
2.4%
K 2
 
2.4%
M 2
 
2.4%
o 1
 
1.2%
a 1
 
1.2%
Other values (3) 3
 
3.5%
None
ValueCountFrequency (%)
· 2
100.0%

법인개인구분
Categorical

HIGH CORRELATION 

Distinct3
Distinct (%)2.0%
Missing0
Missing (%)0.0%
Memory size1.3 KiB
<NA>
73 
개인
63 
법인
14 

Length

Max length4
Median length2
Mean length2.9733333
Min length2

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row개인
2nd row개인
3rd row개인
4th row개인
5th row개인

Common Values

ValueCountFrequency (%)
<NA> 73
48.7%
개인 63
42.0%
법인 14
 
9.3%

Length

2024-05-04T08:15:15.527699image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2024-05-04T08:15:15.951372image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
na 73
48.7%
개인 63
42.0%
법인 14
 
9.3%

사업소전화번호
Text

MISSING 

Distinct65
Distinct (%)100.0%
Missing85
Missing (%)56.7%
Memory size1.3 KiB
2024-05-04T08:15:16.795476image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length13
Median length12
Mean length11.984615
Min length9

Characters and Unicode

Total characters779
Distinct characters11
Distinct categories2 ?
Distinct scripts1 ?
Distinct blocks1 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique65 ?
Unique (%)100.0%

Sample

1st row031-322-3141
2nd row031-332-8898
3rd row031-244-3997
4th row031-336-9230
5th row031-764-2402
ValueCountFrequency (%)
031-332-5840 1
 
1.5%
031-322-0829 1
 
1.5%
031-287-7711 1
 
1.5%
031-287-0664 1
 
1.5%
031-283-0099 1
 
1.5%
031-273-6821 1
 
1.5%
031-284-9965 1
 
1.5%
031-895-3250 1
 
1.5%
070-4145-0166 1
 
1.5%
031-286-0392 1
 
1.5%
Other values (55) 55
84.6%
2024-05-04T08:15:18.149025image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
3 156
20.0%
- 129
16.6%
0 105
13.5%
1 102
13.1%
2 70
9.0%
8 56
 
7.2%
6 36
 
4.6%
7 34
 
4.4%
4 33
 
4.2%
5 30
 
3.9%

Most occurring categories

ValueCountFrequency (%)
Decimal Number 650
83.4%
Dash Punctuation 129
 
16.6%

Most frequent character per category

Decimal Number
ValueCountFrequency (%)
3 156
24.0%
0 105
16.2%
1 102
15.7%
2 70
10.8%
8 56
 
8.6%
6 36
 
5.5%
7 34
 
5.2%
4 33
 
5.1%
5 30
 
4.6%
9 28
 
4.3%
Dash Punctuation
ValueCountFrequency (%)
- 129
100.0%

Most occurring scripts

ValueCountFrequency (%)
Common 779
100.0%

Most frequent character per script

Common
ValueCountFrequency (%)
3 156
20.0%
- 129
16.6%
0 105
13.5%
1 102
13.1%
2 70
9.0%
8 56
 
7.2%
6 36
 
4.6%
7 34
 
4.4%
4 33
 
4.2%
5 30
 
3.9%

Most occurring blocks

ValueCountFrequency (%)
ASCII 779
100.0%

Most frequent character per block

ASCII
ValueCountFrequency (%)
3 156
20.0%
- 129
16.6%
0 105
13.5%
1 102
13.1%
2 70
9.0%
8 56
 
7.2%
6 36
 
4.6%
7 34
 
4.4%
4 33
 
4.2%
5 30
 
3.9%

사업소주소
Text

UNIQUE 

Distinct150
Distinct (%)100.0%
Missing0
Missing (%)0.0%
Memory size1.3 KiB
2024-05-04T08:15:18.892878image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Length

Max length57
Median length44
Mean length33.613333
Min length22

Characters and Unicode

Total characters5042
Distinct characters196
Distinct categories8 ?
Distinct scripts3 ?
Distinct blocks2 ?
The Unicode Standard assigns character properties to each code point, which can be used to analyse textual variables.

Unique

Unique150 ?
Unique (%)100.0%

Sample

1st row경기도 용인시 처인구 모현읍 능원로 97
2nd row경기도 용인시 처인구 원삼면 고당로15번길 6-1, 3층
3rd row경기도 용인시 처인구 이동읍 백옥대로 687-12, 1층 01호
4th row경기도 용인시 처인구 중부대로 1425, 일삼빌딩 2층 201호 (김량장동)
5th row경기도 용인시 처인구 원삼면 문촌로 100-4, 2호
ValueCountFrequency (%)
경기도 150
 
13.8%
용인시 150
 
13.8%
처인구 77
 
7.1%
기흥구 45
 
4.1%
수지구 28
 
2.6%
김량장동 22
 
2.0%
2층 20
 
1.8%
신갈동 16
 
1.5%
중부대로 16
 
1.5%
3층 14
 
1.3%
Other values (357) 547
50.4%
2024-05-04T08:15:20.172099image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Most occurring characters

ValueCountFrequency (%)
935
 
18.5%
232
 
4.6%
1 201
 
4.0%
197
 
3.9%
170
 
3.4%
158
 
3.1%
157
 
3.1%
153
 
3.0%
150
 
3.0%
149
 
3.0%
Other values (186) 2540
50.4%

Most occurring categories

ValueCountFrequency (%)
Other Letter 2927
58.1%
Space Separator 935
 
18.5%
Decimal Number 794
 
15.7%
Other Punctuation 121
 
2.4%
Close Punctuation 114
 
2.3%
Open Punctuation 114
 
2.3%
Dash Punctuation 32
 
0.6%
Uppercase Letter 5
 
0.1%

Most frequent character per category

Other Letter
ValueCountFrequency (%)
232
 
7.9%
197
 
6.7%
170
 
5.8%
158
 
5.4%
157
 
5.4%
153
 
5.2%
150
 
5.1%
149
 
5.1%
136
 
4.6%
78
 
2.7%
Other values (167) 1347
46.0%
Decimal Number
ValueCountFrequency (%)
1 201
25.3%
2 118
14.9%
3 94
11.8%
0 88
11.1%
4 63
 
7.9%
5 60
 
7.6%
6 50
 
6.3%
8 43
 
5.4%
7 42
 
5.3%
9 35
 
4.4%
Uppercase Letter
ValueCountFrequency (%)
B 3
60.0%
A 1
 
20.0%
R 1
 
20.0%
Other Punctuation
ValueCountFrequency (%)
, 120
99.2%
/ 1
 
0.8%
Space Separator
ValueCountFrequency (%)
935
100.0%
Close Punctuation
ValueCountFrequency (%)
) 114
100.0%
Open Punctuation
ValueCountFrequency (%)
( 114
100.0%
Dash Punctuation
ValueCountFrequency (%)
- 32
100.0%

Most occurring scripts

ValueCountFrequency (%)
Hangul 2927
58.1%
Common 2110
41.8%
Latin 5
 
0.1%

Most frequent character per script

Hangul
ValueCountFrequency (%)
232
 
7.9%
197
 
6.7%
170
 
5.8%
158
 
5.4%
157
 
5.4%
153
 
5.2%
150
 
5.1%
149
 
5.1%
136
 
4.6%
78
 
2.7%
Other values (167) 1347
46.0%
Common
ValueCountFrequency (%)
935
44.3%
1 201
 
9.5%
, 120
 
5.7%
2 118
 
5.6%
) 114
 
5.4%
( 114
 
5.4%
3 94
 
4.5%
0 88
 
4.2%
4 63
 
3.0%
5 60
 
2.8%
Other values (6) 203
 
9.6%
Latin
ValueCountFrequency (%)
B 3
60.0%
A 1
 
20.0%
R 1
 
20.0%

Most occurring blocks

ValueCountFrequency (%)
Hangul 2927
58.1%
ASCII 2115
41.9%

Most frequent character per block

ASCII
ValueCountFrequency (%)
935
44.2%
1 201
 
9.5%
, 120
 
5.7%
2 118
 
5.6%
) 114
 
5.4%
( 114
 
5.4%
3 94
 
4.4%
0 88
 
4.2%
4 63
 
3.0%
5 60
 
2.8%
Other values (9) 208
 
9.8%
Hangul
ValueCountFrequency (%)
232
 
7.9%
197
 
6.7%
170
 
5.8%
158
 
5.4%
157
 
5.4%
153
 
5.2%
150
 
5.1%
149
 
5.1%
136
 
4.6%
78
 
2.7%
Other values (167) 1347
46.0%

데이터기준일자
Categorical

CONSTANT 

Distinct1
Distinct (%)0.7%
Missing0
Missing (%)0.0%
Memory size1.3 KiB
2024-01-31
150 

Length

Max length10
Median length10
Mean length10
Min length10

Unique

Unique0 ?
Unique (%)0.0%

Sample

1st row2024-01-31
2nd row2024-01-31
3rd row2024-01-31
4th row2024-01-31
5th row2024-01-31

Common Values

ValueCountFrequency (%)
2024-01-31 150
100.0%

Length

2024-05-04T08:15:20.768780image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Histogram of lengths of the category

Common Values (Plot)

2024-05-04T08:15:21.340220image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
ValueCountFrequency (%)
2024-01-31 150
100.0%

Interactions

2024-05-04T08:15:10.054772image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/

Correlations

2024-05-04T08:15:21.556076image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
순번유무료구분법인개인구분사업소전화번호
순번1.0000.0000.0001.000
유무료구분0.0001.0000.8141.000
법인개인구분0.0000.8141.0001.000
사업소전화번호1.0001.0001.0001.000
2024-05-04T08:15:21.839394image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
법인개인구분유무료구분
법인개인구분1.0000.605
유무료구분0.6051.000
2024-05-04T08:15:22.071363image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
순번유무료구분법인개인구분
순번1.0000.0000.000
유무료구분0.0001.0000.605
법인개인구분0.0000.6051.000

Missing values

2024-05-04T08:15:10.439511image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
A simple visualization of nullity by column.
2024-05-04T08:15:10.891937image/svg+xmlMatplotlib v3.7.2, https://matplotlib.org/
Nullity matrix is a data-dense display which lets you quickly visually pick out patterns in data completion.

Sample

순번유무료구분상호명법인개인구분사업소전화번호사업소주소데이터기준일자
01유료조흥인력개인031-322-3141경기도 용인시 처인구 모현읍 능원로 972024-01-31
12유료떼부자 인력사무소개인031-332-8898경기도 용인시 처인구 원삼면 고당로15번길 6-1, 3층2024-01-31
23유료현대인력사무소개인<NA>경기도 용인시 처인구 이동읍 백옥대로 687-12, 1층 01호2024-01-31
34유료드림인력 용인점개인<NA>경기도 용인시 처인구 중부대로 1425, 일삼빌딩 2층 201호 (김량장동)2024-01-31
45유료TS인력개인<NA>경기도 용인시 처인구 원삼면 문촌로 100-4, 2호2024-01-31
56유료태산인력개인<NA>경기도 용인시 처인구 모현읍 초부로 153, 3호2024-01-31
67무료(사)한국후계농업경영인용인특례시연합회법인<NA>경기도 용인시 처인구 원삼면 농촌파크로 80-1, 농촌테마파크 농경문화전시관 1층2024-01-31
78유료다모아 건설인력개인<NA>경기도 용인시 처인구 백암면 근창로35번길 522024-01-31
89유료국제인력개인<NA>경기도 용인시 처인구 금령로 144, 3층 (마평동)2024-01-31
910유료한일건설인력개인<NA>경기도 용인시 처인구 원삼면 원양로 81, 한덕개발2024-01-31
순번유무료구분상호명법인개인구분사업소전화번호사업소주소데이터기준일자
140141유료직업소개·파출 광장<NA><NA>경기도 용인시 수지구 광교중앙로 301, 701-1호 (상현동)2024-01-31
141142무료지구촌사회복지재단 수지장애인종합복지관<NA><NA>경기도 용인시 수지구 포은대로 435, 3층 (풍덕천동, 수지복지센터)2024-01-31
142143유료에스크나우이티오<NA><NA>경기도 용인시 수지구 광교중앙로 301, 드림타워 706-2호 (상현동)2024-01-31
143144유료동진인력<NA><NA>경기도 용인시 수지구 풍덕천로 183, 401호 (풍덕천동)2024-01-31
144145유료대우파출부<NA><NA>경기도 용인시 수지구 문정로 46, 403호 (풍덕천동,한길타운)2024-01-31
145146유료성우인력개발<NA><NA>경기도 용인시 수지구 정든로 18 (죽전동)2024-01-31
146147유료고.휴먼엔지니어링<NA><NA>경기도 용인시 수지구 광교중앙로296번길 4, 골드리치안오피스텔 1018호 (상현동)2024-01-31
147148유료폴라리스 써어치 앤 컴퍼니<NA><NA>경기도 용인시 수지구 광교중앙로295번길 13, B동 314호 (상현동, 광교2차푸르지오시티)2024-01-31
148149유료수지파출부<NA><NA>경기도 용인시 수지구 수지로 64, 102동 지하층 2호 (상현동, 대진아파트)2024-01-31
149150유료창덕취업정보<NA><NA>경기도 용인시 수지구 풍덕천로 146, 2층 (풍덕천동)2024-01-31