Conference Agenda
Overview and details of the sessions of this conference. Please select a date or location to show only sessions at that day or location. Please select a single session for detailed view (with abstracts and downloads if available).
Please note that all times are shown in the time zone of the conference. The current conference time is: 24th Aug 2026, 04:43:56am America, Santiago
|
Daily Overview |
| Session | |
|
51B: Computer Science Location: Room 03: Manquehue | |
| Presentation 4 | |
9:06am - 9:18am
Automated Classification of Tax Service Requests Using Machine Learning and Natural Language Processing Universidad Andrés Bello - (CL), Chile Classifying citizen service requests in public tax offices is harder than it looks. Citizens describe their problems in their own words, select the wrong service category more often than not, and force staff to reclassify every submission before routing can begin. This paper examines whether that reclassification burden can be automated using NLP and supervised machine learning, drawing on 4,021 real requests submitted to Chile's Taxpayer Ombudsman Office (DEDECON) in Spanish free text. The feature engineering strategy combines TF-IDF text representation with two domain-grounded signals: a proxy for the user's knowledge level (scored 0 to 2) and binary indicators of class-specific keyword presence. Six classifiers were evaluated under identical conditions: Logistic Regression, Support Vector Machines, Naive Bayes, Decision Trees, Random Forest, and K-Nearest Neighbors. Random Forest reached 92.18% accuracy and an F1-macro score of 0.921, competitive with transformer-based approaches at a fraction of the computational cost. The main source of residual error is semantic overlap between adjacent service categories, a problem that text features alone are unlikely to resolve. Beyond DEDECON, the pipeline is directly applicable to other Spanish-language public service institutions facing the same intake classification problem. | |
