
doi: 10.5281/zenodo.21552632 , 10.5281/zenodo.21541894 , 10.5281/zenodo.21566312 , 10.5281/zenodo.21256947 , 10.5281/zenodo.21542298 , 10.5281/zenodo.21544788 , 10.5281/zenodo.21391857 , 10.5281/zenodo.21322707 , 10.5281/zenodo.21479110 , 10.5281/zenodo.21566311 , 10.5281/zenodo.21544787 , 10.5281/zenodo.21322708 , 10.5281/zenodo.21552633 , 10.5281/zenodo.21562864 , 10.5281/zenodo.21256948 , 10.5281/zenodo.21541895 , 10.5281/zenodo.21542297 , 10.5281/zenodo.21391856 , 10.5281/zenodo.21479109 , 10.5281/zenodo.21562863
doi: 10.5281/zenodo.21552632 , 10.5281/zenodo.21541894 , 10.5281/zenodo.21566312 , 10.5281/zenodo.21256947 , 10.5281/zenodo.21542298 , 10.5281/zenodo.21544788 , 10.5281/zenodo.21391857 , 10.5281/zenodo.21322707 , 10.5281/zenodo.21479110 , 10.5281/zenodo.21566311 , 10.5281/zenodo.21544787 , 10.5281/zenodo.21322708 , 10.5281/zenodo.21552633 , 10.5281/zenodo.21562864 , 10.5281/zenodo.21256948 , 10.5281/zenodo.21541895 , 10.5281/zenodo.21542297 , 10.5281/zenodo.21391856 , 10.5281/zenodo.21479109 , 10.5281/zenodo.21562863
Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuning again on the target task---often improves model performance substantially on language understanding tasks in monolingual English settings. We investigate whether English intermediate-task training is still helpful on non-English target tasks. Using nine intermediate language-understanding tasks, we evaluate intermediate-task transfer in a zero-shot cross-lingual setting on the XTREME benchmark. We see large improvements from intermediate training on the BUCC and Tatoeba sentence retrieval tas Research goal: Does multimodal intermediate-task training (e.g., image-captioning) improve zero-shot cross-lingual performance on XTREME compared to text-only intermediate tasks, as measured by accuracy on downstream tasks? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 9.3/10.
This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 9.3/10.
enhance, task, further, addition, models, zero-shot, image-text, audio, intermediate, combining, image, tasks, training, like, effect, datasets, multimodal, CLIP, alignment, intermediate-task, improve, data, impact, reasoning, cross-lingual, visual, modality, text, image-captioning, performance
enhance, task, further, addition, models, zero-shot, image-text, audio, intermediate, combining, image, tasks, training, like, effect, datasets, multimodal, CLIP, alignment, intermediate-task, improve, data, impact, reasoning, cross-lingual, visual, modality, text, image-captioning, performance
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
