Ali Azmoudeh /

Publication · 2026

VisAffect at MWE-2026 AdMIRe 2: IMMCAN Idiom Multimodal Cross-Attention Network

Barış Bilen, Ali Azmoudeh, Hazım Kemal Ekenel, Hatice Köse

Proceedings of the 22nd Workshop on Multiword Expressions (MWE 2026)
Association for Computational Linguistics · pp. 149–153 · Equal contribution with Barış Bilen

Research summary

Studies multilingual image ranking using frozen language and vision–language encoders with a trainable cross-attention module. Visual grounding improves zero-shot transfer, while text-only modeling performs better on the smaller in-domain test.

DOI: 10.18653/v1/2026.mwe-1.19

Citation

Barış Bilen, Ali Azmoudeh, Hazım Kemal Ekenel, Hatice Köse. 2026. VisAffect at MWE-2026 AdMIRe 2: IMMCAN Idiom Multimodal Cross-Attention Network. Proceedings of the 22nd Workshop on Multiword Expressions (MWE 2026).

View BibTeX
@inproceedings{bilen2026visaffect,
  title = {{VisAffect at MWE-2026 AdMIRe 2: IMMCAN Idiom Multimodal Cross-Attention Network}},
  author = {Barış Bilen and Ali Azmoudeh and Hazım Kemal Ekenel and Hatice Köse},
  year = {2026},
  booktitle = {Proceedings of the 22nd Workshop on Multiword Expressions (MWE 2026)},
  doi = {10.18653/v1/2026.mwe-1.19},
  url = {https://aclanthology.org/2026.mwe-1.19/},
  pages = {149--153},
  publisher = {Association for Computational Linguistics}
}