Jointly Learning to Label Sentences and Tokens

Marek Rei, Anders Søgaard

2018-11-14Sentence Classification Grammatical Error Detection

Abstract

Learning to construct text representations in end-to-end systems can be difficult, as natural languages are highly compositional and task-specific annotated datasets are often limited in size. Methods for directly supervising language composition can allow us to guide the models based on existing knowledge, regularizing them towards more robust and interpretable representations. In this paper, we investigate how objectives at different granularities can be used to learn better language representations and we propose an architecture for jointly learning to label sentences and tokens. The predictions at each level are combined together using an attention mechanism, with token-level labels also acting as explicit supervision for composing sentence-level representations. Our experiments show that by learning to perform these tasks jointly on multiple levels, the model achieves substantial improvements for both sentence classification and sequence labeling.

Results

Task	Dataset	Metric	Value	Model
Grammatical Error Correction	CoNLL-2014 A1	F0.5	22.14	BiLSTM-JOINT (trained on FCE)
Grammatical Error Correction	CoNLL-2014 A2	F0.5	29.65	BiLSTM-JOINT (trained on FCE)
Grammatical Error Correction	JFLEG	F0.5	52.52	BiLSTM-JOINT (trained on FCE)
Grammatical Error Correction	FCE	F0.5	52.07	BiLSTM-JOINT

Related Papers

IMPARA-GED: Grammatical Error Detection is Boosting Reference-free Grammatical Error Quality Estimator2025-06-03 A Personalized Conversational Benchmark: Towards Simulating Personalized Conversations2025-05-20 Detecting Spelling and Grammatical Anomalies in Russian Poetry Texts2025-05-07 ARWI: Arabic Write and Improve2025-04-16 Tougher Text, Smarter Models: Raising the Bar for Adversarial Defence Benchmarks2025-01-05 Consolidating and Developing Benchmarking Datasets for the Nepali Natural Language Understanding Tasks2024-11-28 Cyber-Attack Technique Classification Using Two-Stage Trained Large Language Models2024-11-27 Multi-label Sequential Sentence Classification via Large Language Model2024-11-23