Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching

Kunbo Ding; Weijie Liu; Yuejian Fang; Zhe Zhao; Qi Ju; Xuefeng Yang; Rong Tian; Tong Zhu; Haoyan Liu; Guohong Han; Xiaohong Bai; Weiquan Mao; Yudong Li; Wei Guo; Taiqiang Wu; Ning Sun

doi:10.60692/nz2ww-kmd10

JOURNAL ARTICLE

Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching

Kunbo Ding Weijie Liu Yuejian Fang Zhe Zhao Qi Ju Xuefeng Yang Rong Tian Tong Zhu Haoyan Liu Guohong Han Xiaohong Bai Weiquan Mao Yudong Li Wei Guo Taiqiang Wu Ning Sun

Year: 2022 Journal: Greater South Information System

DOI: 10.60692/nz2ww-kmd10

Get Full-Text PDF Get Analytical Report

Abstract

Previous studies have proved that cross-lingual knowledge distillation can significantly improve the performance of pre-trained models for cross-lingual similarity matching tasks.However, the student model needs to be large in this operation.Otherwise, its performance will drop sharply, thus making it impractical to be deployed to memory-limited devices.To address this issue, we delve into cross-lingual knowledge distillation and propose a multistage distillation framework for constructing a small-size but high-performance cross-lingual model.In our framework, contrastive learning, bottleneck, and parameter recurrent strategies are combined to prevent performance from being compromised during the compression process.The experimental results demonstrate that our method can compress the size of XLM-R and MiniLM by more than 50%, while the performance is only reduced by about 1%.

Keywords:

Distillation Matching (statistics) Similarity (geometry) Compression (physics) Key (lock) Semantic similarity

Metrics

Cited By

0.00

FWCI (Field Weighted Citation Impact)

Refs

0.23

Citation Normalized Percentile

Is in top 1%

Is in top 10%

Topics

Multimodal Machine Learning Applications

Physical Sciences → Computer Science → Computer Vision and Pattern Recognition

Topic Modeling

Physical Sciences → Computer Science → Artificial Intelligence

Domain Adaptation and Few-Shot Learning

Physical Sciences → Computer Science → Artificial Intelligence

Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching

Abstract

Metrics

Topics

Related Documents

Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching

Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching

The Study on Three-Stage Matching Algorithm Framework of Semantic Similarity

Ensemble transformer for cross-lingual semantic textual similarity

Measuring cross-lingual semantic similarity across European languages