基本信息

张亚萍  硕士生导师

中国科学院自动化研究所
电子邮件: yaping.zhang@nlpr.ia.ac.cn
通信地址: 北京市海淀区中关村东路95号




研究领域

自然语言处理,多模态理解与生成

招生信息

招收致力于在多模态文字图像领域、科技文献理解领域理论研究和关键技术突破的研究生,让我们一起让大模型具有更好的文字阅读能力。


招生专业
081104-模式识别与智能系统
招生方向
多模态理解与生成,文档图像翻译,科学文献理解

工作经历

张亚萍,中国科学院自动化研究所副研究员,硕士生导师,长期从事鲁棒性多语言多模态序列理解和生成研究。视觉大数据专委会、文档图像分析与识别专委会以及中文信息学会青年工作委员会委员。相关工作发表CCFA/B论文20余篇,含IEEE TPAMI、IEEE TIP、CVPR等领域顶级期刊和会议,曾获ICASSP最佳学生论文奖(10/1406) 和CCMT最佳论文奖。近年来,作为项目负责人主持国家自然科学基金面上、青年基金十四五装备预研等项目,作为技术骨干参与国家重点研发计划、重点基金以及多项特定领域应用项目。

个人主页:https://aprilyapingzhang.github.io/

工作简历
2023-07~现在, 中国科学院自动化研究所, 副研究员
2020-07~2023-07,中国科学院自动化研究所, 助理研究员

出版信息

论文发表
  • Zirui Zhang, Yaping Zhang, Lu Xiang, Yang Zhao, Feifei Zhai, Yu Zhou, Chengqing Zong. PromptDLA: A Domain-aware Prompt Document Layout Analysis Framework with Descriptive Knowledge as a Cue. IEEE Transactions on Multimedia (TMM), 2026. (CCF A, 通讯作者)
  • Zhiyang Zhang, Yaping Zhang, Yupu Liang, Cong Ma, Lu Xiang, Yang Zhao, Yu Zhou, and Chengqing Zong. Understand Layout and Translate Text: Unified Feature-Conductive End-to-End Document Image Translation. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2025. (CCF A, 通讯作者)
  • Yupu Liang, Yaping Zhang, Zhiyang Zhang, Yang Zhao, Lu Xiang, Chengqing Zong, Yu Zhou "Single-to-mix modality alignment with multimodal large language model for document image machine translation." Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025). (CCF A, 通讯作者)
  • Zhiyang Zhang, Yaping Zhang, Yupu Liang, Cong Ma, Lu Xiang, Yang Zhao, Yu Zhou, Chengqing Zong. "A Query-Response Framework for Whole-Page Complex-Layout Document Image Translation with Relevant Regional Concentration." Findings of the Association for Computational Linguistics: ACL 2025. (CCF A, 通讯作者)
  • Yaping Zhang, Shuai Nie, Wenju Liu, Xing Xu, Dongxiang Zhang, Heng Tao Shen. Sequence-to-sequence domain adaptation network for robust text image recognition. Proceedings of the IEEE/CVF conference on CVPR. 2019. (CCF A)
  • Yaping Zhang, Shuai Nie, Shan Liang, Wenju Liu. Robust text image recognition via adversarial sequence-to-sequence domain adaptation. IEEE trans. on Image Processing. 2021. (CCF A)
  • Gengluo Li, Chengquan Zhang, Yupu Liang, Huawen Shen, Yaping Zhang, Pengyuan Lyu, Weinong Wang, Xingyu Wan, Gangyan Zeng, Han Hu, Can Ma, Yu Zhou. "MMTIT-Bench: A Multilingual and Multi-Scenario Benchmark with Cognition-Perception-Reasoning Guided Text-Image Machine Translation."  CVPR, 2026. (CCF A, accepted)
  • Lu Xiang, Yang Zhao, Yaping Zhang, Zixuan Ren, Xingquan Zhang, Zhiyuan Chen, Chengqing Zong. "An Empirical Study on Reasoning and Generalization in Large Language Models." Computational Linguistics (2026): 1-51. (JCR Q1)
  • Yupu Liang, Yaping Zhang, Zhiyang Zhang, Zhiyuan Chen, Yang Zhao, Lu Xiang, Chengqing Zong, Yu Zhou. "Improving MLLM’s Document Image Machine Translation via Synchronously Self-reviewing Its OCR Proficiency." Findings of the Association for Computational Linguistics: ACL 2025. (CCF A)
  • Zhiyang Zhang, Yaping Zhang, Yupu Liang, Cong Ma, Lu Xiang, Yang Zhao, Yu Zhou, Chengqing Zong. "Reading When Translating: Multi-Modal Document Image Machine Translation With Reading Flow Prediction." IEEE Transactions on Audio, Speech and Language Processing (IEEE TASLP), 2025. (CCF B)
  • Yang Zhao, Yaping Zhang, Lu Xiang, Jiajun Zhang, Yu Zhou, Chengqing Zong. Imitate, Reward and Introspect: Enhancing Neural Machine Translation by Utilizing Large Language Models as Environments[J]. IEEE Transactions on Audio, Speech and Language Processing (IEEE TASLP), 2025, 33: 4736-4747. (CCF B)
  • Jing Ye, Lu Xiang, Yaping Zhang, Chengqing Zong. SweetieChat: A Strategy-Enhanced Role-playing Framework for Diverse Scenarios Handling Emotional Support Agent. Proceedings of the 31st International Conference on Computational Linguistics(COLING-2025). Abu Dhabi, UAE. pages 4646–4669.(CCF B)
  • Zhiyang Zhang, Yaping Zhang, Yupu Liang, Lu Xiang, Yang Zhao, Yu Zhou, Chengqing Zong. From Chaotic OCR Words to Coherent Document: A Fine-to-Coarse Zoom-Out Network for Complex-Layout Document Image Translation. Proceedings of the 31st International Conference on Computational Linguistics(COLING-2025). Abu Dhabi, UAE. pages 10877–10890. (CCF B)
  • Yupu Liang, Yaping Zhang, Cong Ma, Zhiyang Zhang, Yang Zhao, Lu Xiang, Chengqing Zong, Yu Zhou. Document Image Machine Translation with Dynamic Multi-pre-trained Models Assembling. In The 2024 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2024). Mexico City, Mexico. June 16-21, 2024.(CCF B)
  • Cong Ma, Yaping Zhang, Zhiyang Zhang, Yupu Liang, Yang Zhao, Yu Zhou, Chengqing Zong. Born a BabNet with Hierarchical Parental Supervision for End-to-End Text Image Machine Translation. In The 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024). Torino, Italia. May 20-25, 2024.(CCF B)
  • Cong Ma, Yaping Zhang, Yang Zhao, Yu Zhou, Chengqing Zong. Vector Quantization Knowledge Transfer for End-to-End Text Image Machine Translation. In The 49th IEEE International Conference on Acoustics, Speech, & Signal Processing (ICASSP 2024). COEX, Seoul, Korea. April 14-19, 2024. IEEE Xplore Version.(CCF B)
  • Zhiyang Zhang, Yaping Zhang, Lu Xiang, Yang Zhao, Yu Zhou, Chengqing Zong. LayoutDIT: Layout-Aware End-to-End Document Image Translation with Multi-Step Conductive Decoder. In Findings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023), Singapore. December 6-10, 2023. pp. 4959–4965. ACL_Anthology_version. (CCF B)
  • Cong Ma, Yaping Zhang, Mei Tu, Yang Zhao, Yu Zhou, Chengqing Zong. CCIM: Cross-Modal Cross-Lingual Interactive Image Translation. In Findings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023), Singapore. December 6-10, 2023. pp. 4959–4965. ACL_Anthology_version. (CCF B)
  • Yaping Zhang, Shan Liang, Shuai Nie, Wenju Liu, Shouye Peng. Robust offline handwritten character recognition through exploring writer-independent features under the guidance of printed data. Pattern Recognition Letters. 2018. (CCF B)
  • Bin Liu, Shuai Nie, Yaping Zhang, Dengfeng Ke, Shan Liang, Wenju Liu. Boosting noise robustness of acoustic model via deep adversarial training. Proceedings of the ICASSP. 2018. (CCF B, 最佳学生论文)
  • Zhiyang Zhang, Yaping Zhang, Yupu Liang, Lu Xiang, Yang Zhao, Yu Zhou, Chengqing Zong. A Novel Dataset and Benchmark Analysis on Document Image Translation. In Proceddings of the 2023 China Conference on Machine Translation (CCMT 2023, 最佳论文).

科研活动

   
学术活动

IEEE TMM、 NeuNet、CVMJ、 ICML、ICLR、NeuralPS、CVPR、ECCV、AAAI、 ACL、 EMNLP、 COLING 等期刊会议审稿人