Paper Reading AI Learner

READ: Improving Relation Extraction from an ADversarial Perspective

2024-04-02 16:42:44
Dawei Li, William Hogan, Jingbo Shang

Abstract

Recent works in relation extraction (RE) have achieved promising benchmark accuracy; however, our adversarial attack experiments show that these works excessively rely on entities, making their generalization capability questionable. To address this issue, we propose an adversarial training method specifically designed for RE. Our approach introduces both sequence- and token-level perturbations to the sample and uses a separate perturbation vocabulary to improve the search for entity and context perturbations. Furthermore, we introduce a probabilistic strategy for leaving clean tokens in the context during adversarial training. This strategy enables a larger attack budget for entities and coaxes the model to leverage relational patterns embedded in the context. Extensive experiments show that compared to various adversarial training methods, our method significantly improves both the accuracy and robustness of the model. Additionally, experiments on different data availability settings highlight the effectiveness of our method in low-resource scenarios. We also perform in-depth analyses of our proposed method and provide further hints. We will release our code at this https URL.

Abstract (translated)

近年来,在关系抽取(RE)领域取得了一些的有希望的基准准确度;然而,我们的攻击实验表明,这些工作过度依赖实体,导致其泛化能力值得怀疑。为了解决这个问题,我们提出了一个专门针对RE的攻击训练方法。我们的方法引入了样本级和标记级扰动,并使用了一个单独的扰动词汇表来提高对实体和上下文扰动的搜索。此外,我们还引入了一种概率策略,让其在上下文训练过程中留下干净的标记。这种策略使得实体和上下文中的关系模式得到更充分的利用。大量实验证明,与各种攻击训练方法相比,我们的方法显著提高了模型的准确性和鲁棒性。此外,在不同数据可用性设置的实验中,我们的方法在低资源场景中表现出有效的效果。我们还对我们所提出的方法进行了深入的分析,并提供了进一步的提示。我们将发布我们的代码在這個 URL 上。

URL

https://arxiv.org/abs/2404.02931

PDF

https://arxiv.org/pdf/2404.02931.pdf


Tags
3D Action Action_Localization Action_Recognition Activity Adversarial Agent Attention Autonomous Bert Boundary_Detection Caption Chat Classification CNN Compressive_Sensing Contour Contrastive_Learning Deep_Learning Denoising Detection Dialog Diffusion Drone Dynamic_Memory_Network Edge_Detection Embedding Embodied Emotion Enhancement Face Face_Detection Face_Recognition Facial_Landmark Few-Shot Gait_Recognition GAN Gaze_Estimation Gesture Gradient_Descent Handwriting Human_Parsing Image_Caption Image_Classification Image_Compression Image_Enhancement Image_Generation Image_Matting Image_Retrieval Inference Inpainting Intelligent_Chip Knowledge Knowledge_Graph Language_Model LLM Matching Medical Memory_Networks Multi_Modal Multi_Task NAS NMT Object_Detection Object_Tracking OCR Ontology Optical_Character Optical_Flow Optimization Person_Re-identification Point_Cloud Portrait_Generation Pose Pose_Estimation Prediction QA Quantitative Quantitative_Finance Quantization Re-identification Recognition Recommendation Reconstruction Regularization Reinforcement_Learning Relation Relation_Extraction Represenation Represenation_Learning Restoration Review RNN Robot Salient Scene_Classification Scene_Generation Scene_Parsing Scene_Text Segmentation Self-Supervised Semantic_Instance_Segmentation Semantic_Segmentation Semi_Global Semi_Supervised Sence_graph Sentiment Sentiment_Classification Sketch SLAM Sparse Speech Speech_Recognition Style_Transfer Summarization Super_Resolution Surveillance Survey Text_Classification Text_Generation Tracking Transfer_Learning Transformer Unsupervised Video_Caption Video_Classification Video_Indexing Video_Prediction Video_Retrieval Visual_Relation VQA Weakly_Supervised Zero-Shot