View metadata
| dc.title | Automatic Product Classification in International Trade: Machine Learning and Large Language Models |
| dc.contributor.author | Marra de Artiñano, Ignacio |
| dc.contributor.author | Riottini Depetris, Franco |
| dc.contributor.author | Volpe Martincus, Christian |
| dc.contributor.orgunit | Productivity, Trade and Innovation Sector |
| dc.contributor.orgunit | Trade and Investment Division |
| dc.coverage | Latin America |
| dc.date.available | 2023-07-21T00:07:00 |
| dc.date.issue | 2023-07-21T00:07:00 |
| dc.description.abstract | Accurately classifying products is essential in international trade. Virtually all countries categorize products into tariff lines using the Harmonized System (HS) nomenclature for both statistical and duty collection purposes. In this paper, we apply and assess several different algorithms to automatically classify products based on text descriptions. To do so, we use agricultural product descriptions from several public agencies, including customs authorities and the United States Department of Agriculture (USDA). We find that while traditional machine learning (ML) models tend to perform well within the dataset in which they were trained, their precision drops dramatically when implemented outside of it. In contrast, large language models (LLMs) such as GPT 3.5 show a consistently good performance across all datasets, with accuracy rates ranging between 60% and 90% depending on HS aggregation levels. Our analysis highlights the valuable role that artificial intelligence (AI) can play in facilitating product classification at scale and, more generally, in enhancing the categorization of unstructured data. |
| dc.format.extent | 37 |
| dc.identifier.doi | http://dx.doi.org/10.18235/0005012 |
| dc.identifier.url | https://publications.iadb.org/publications/english/document/Automatic-Product-Classification-in-International-Trade-Machine-Learning-and-Large-Language-Models.pdf |
| dc.language.iso | en |
| dc.publisher | Inter-American Development Bank |
| dc.subject | Export of Goods |
| dc.subject | Customs Administration |
| dc.subject | International Trade |
| dc.subject | Machine Learning |
| dc.subject | Small Business |
| dc.subject | Artificial Intelligence |
| dc.subject | Integration and Trade |
| dc.subject | Rating |
| dc.subject | Tariff System |
| dc.subject.jelcode | F10 - Trade: General |
| dc.subject.jelcode | C55 - Large Data Sets: Modeling and Analysis |
| dc.subject.jelcode | C81 - Methodology for Collecting, Estimating, and Organizing Microeconomic Data • Data Access |
| dc.subject.jelcode | C88 - Other Computer Software |
| dc.subject.keywords | Product Classification;machine learning;Large Language Models;Trade |
| dc.type | Working Papers |
| idb.identifier.pubnumber | IDB-WP-01494 |
| idb.operation | RG-E1716 |