A

Improving weakly supervised phrase grounding via visual representation contextualization with contrastive learning

Applied Intelligence, vol. 53, pp. 14690–14702

Authors 4

  1. Leiden University · Xi'an Jiaotong University

    Affiliation as printed

    Faculty of Electronic and Information Engineering, Xi’an Jiaotong University, Xi’an, 710049, China

    Leiden Institute of Advanced Computer Science, Leiden University, Street, Leiden, 2333 CA, Leiden, The Netherlands

    Faculty of Electronic and Information Engineering, Xi'an Jiaotong University, Xi'an, 710049, China

  2. Youtian Du corresponding

    Xi'an Jiaotong University

    Affiliation as printed

    Faculty of Electronic and Information Engineering, Xi’an Jiaotong University, Xi’an, 710049, China

    Faculty of Electronic and Information Engineering, Xi'an Jiaotong University, Xi'an, 710049, China

  3. Leiden University

    Affiliation as printed

    Leiden Institute of Advanced Computer Science, Leiden University, Street, Leiden, 2333 CA, Leiden, The Netherlands

  4. Leiden University

    Affiliation as printed

    Leiden Institute of Advanced Computer Science, Leiden University, Street, Leiden, 2333 CA, Leiden, The Netherlands

Cited by 0 stored of 0

No patents citing this paper on Lens.org (checked 2026-10-11).

References 42