Employer Industry Classification Using Job Postings

2017 
In the recruitment domain, knowing the employer industry of jobs is important to get an insight about the demand in each industry. The existing system at CareerBuilder uses an employer name normalization system and an employer knowledge base to infer the employer industry of a job. However, errors may occur during the computation of the job employer and in the construction of the employer knowledge base with the industry attributes. Since the knowledge base is huge, it is not possible to manually detect the errors. Therefore, in this paper we use Machine Learning techniques to automatically detect the errors. With the observation that the main jobs posted by an employer often relate to the employer industry, e.g., truck driver jobs often correspond to employers belonging to the transportation industry, we develop a system that classifies the industry of an employer using job posting data. We aggregate job postings from an employer and use job titles and employer names as features for predicting the industry of the employer. We used two models for classification: (1) Support Vector Machine, and (2) Gradient Boosted Decision Trees, and observed that while both the models perform similarly in classifying job employers that were correctly computed, GBDT is more effective than SVM in identifying job employers that were wrongly computed. We also show the utility of our system in detecting normalization errors and knowledge base errors.
    • Correction
    • Source
    • Cite
    • Save
    • Machine Reading By IdeaReader
    12
    References
    5
    Citations
    NaN
    KQI
    []