Improving Cross-lingual Text Classification with Zero-shot Instance-Weighting

被引:0
|
作者
Li, Irene [1 ]
Sen, Prithviraj [2 ]
Zhu, Huaiyu [2 ]
Li, Yunyao [2 ]
Radev, Dragomir [1 ]
机构
[1] Yale Univ, New Haven, CT 06520 USA
[2] IBM Res Almaden, San Jose, CA USA
关键词
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Cross-lingual text classification (CLTC) is a challenging task made even harder still due to the lack of labeled data in low-resource languages. In this paper, we propose zero-shot instance-weighting, a general model-agnostic zero-shot learning framework for improving CLTC by leveraging source instance weighting. It adds a module on top of pre-trained language models for similarity computation of instance weights, thus aligning each source instance to the target language. During training, the framework utilizes gradient descent that is weighted by instance weights to update parameters. We evaluate this framework over seven target languages on three fundamental tasks and show its effectiveness and extensibility, by improving on F1 score up to 4% in single-source transfer and 8% in multi-source transfer. To the best of our knowledge, our method is the first to apply instance weighting in zeroshot CLTC. It is simple yet effective and easily extensible into multi-source transfer.
引用
收藏
页码:1 / 7
页数:7
相关论文
共 50 条
  • [31] Rumour Detection via Zero-Shot Cross-Lingual Transfer Learning
    Tian, Lin
    Zhang, Xiuzhen
    Lau, Jey Han
    MACHINE LEARNING AND KNOWLEDGE DISCOVERY IN DATABASES, 2021, 12975 : 603 - 618
  • [32] Deep Exploration of Cross-Lingual Zero-Shot Generalization in Instruction Tuning
    Han, Janghoon
    Lee, Changho
    Shin, Joongbo
    Choi, Stanley Jungkyu
    Lee, Honglak
    Bae, Kyunghoon
    FINDINGS OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS: ACL 2024, 2024, : 15436 - 15452
  • [33] Zero-shot Cross-lingual Dialogue Systems with Transferable Latent Variables
    Liu, Zihan
    Shin, Jamin
    Xu, Yan
    Winata, Genta Indra
    Xu, Peng
    Madotto, Andrea
    Fung, Pascale
    2019 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING AND THE 9TH INTERNATIONAL JOINT CONFERENCE ON NATURAL LANGUAGE PROCESSING (EMNLP-IJCNLP 2019): PROCEEDINGS OF THE CONFERENCE, 2019, : 1297 - 1303
  • [34] Zero-shot Cross-lingual Transfer is Under-specified Optimization
    Wu, Shijie
    Van Durme, Benjamin
    Dredze, Mark
    PROCEEDINGS OF THE 7TH WORKSHOP ON REPRESENTATION LEARNING FOR NLP, 2022, : 236 - 248
  • [35] Zero-shot Cross-Lingual Phonetic Recognition with External Language Embedding
    Gao, Heting
    Ni, Junrui
    Zhang, Yang
    Qian, Kaizhi
    Chang, Shiyu
    Hasegawa-Johnson, Mark
    INTERSPEECH 2021, 2021, : 1304 - 1308
  • [36] Exposing the limits of Zero-shot Cross-lingual Hate Speech Detection
    Nozza, Debora
    ACL-IJCNLP 2021: THE 59TH ANNUAL MEETING OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS AND THE 11TH INTERNATIONAL JOINT CONFERENCE ON NATURAL LANGUAGE PROCESSING, VOL 2, 2021, : 907 - 914
  • [37] Zero-shot learning based cross-lingual sentiment analysis for sanskrit text with insufficient labeled data
    Kumar, Puneet
    Pathania, Kshitij
    Raman, Balasubramanian
    APPLIED INTELLIGENCE, 2023, 53 (09) : 10096 - 10113
  • [38] Zero-shot learning based cross-lingual sentiment analysis for sanskrit text with insufficient labeled data
    Puneet Kumar
    Kshitij Pathania
    Balasubramanian Raman
    Applied Intelligence, 2023, 53 : 10096 - 10113
  • [39] Zero-Shot Cross-Lingual Transfer in Legal Domain Using Transformer Models
    Shaheen, Zein
    Wohlgenannt, Gerhard
    Mouromtsev, Dmitry
    2021 INTERNATIONAL CONFERENCE ON COMPUTATIONAL SCIENCE AND COMPUTATIONAL INTELLIGENCE (CSCI 2021), 2021, : 450 - 456
  • [40] Label modification and bootstrapping for zero-shot cross-lingual hate speech detection
    Bigoulaeva, Irina
    Hangya, Viktor
    Gurevych, Iryna
    Fraser, Alexander
    LANGUAGE RESOURCES AND EVALUATION, 2023, 57 (04) : 1515 - 1546