Document detail
ID

oai:arXiv.org:2404.05632

Topic
Computer Science - Computation and...
Author
Hammami, Haitham Baligand, Louis Petrovski, Bojan
Category

Computer Science

Year

2024

listing date

4/17/2024

Keywords
data parsing address
Metrics

Abstract

In the financial industry, identifying the location of parties involved in payments is a major challenge in the context of various regulatory requirements.

For this purpose address parsing entails extracting fields such as street, postal code, or country from free text message attributes.

While payment processing platforms are updating their standards with more structured formats such as SWIFT with ISO 20022, address parsing remains essential for a considerable volume of messages.

With the emergence of Transformers and Generative Large Language Models (LLM), we explore the performance of state-of-the-art solutions given the constraint of processing a vast amount of daily data.

This paper also aims to show the need for training robust models capable of dealing with real-world noisy transactional data.

Our results suggest that a well fine-tuned Transformer model using early-stopping significantly outperforms other approaches.

Nevertheless, generative LLMs demonstrate strong zero-shot performance and warrant further investigations.

Hammami, Haitham,Baligand, Louis,Petrovski, Bojan, 2024, Fighting crime with Transformers: Empirical analysis of address parsing methods in payment data

Document

Open

Share

Source

Articles recommended by ES/IODE AI

Should we consider Systemic Inflammatory Response Index (SIRI) as a new diagnostic marker for rectal cancer?
inflammation rectal surgery overall survival complication significantly diagnostic value cancer rectal 38 siri