


default search action
Yizhi Song
This is just a disambiguation page, and is not intended to be the bibliography of an actual person. The links to all actual bibliographies of persons of the same or a similar name can be found below. Any publication listed on this page has not been assigned to an actual author yet. If you know the true author of one of the publications listed below, you are welcome to contact us.
Person information
Other persons with the same name
- Yizhi Song 0001
— Hong Kong University of Science and Technology (Guangzhou), China
Refine list

refinements active!
zoomed in on ?? of ?? records
view refined list in
2020 – today
- 2026
[j5]Lesong Zheng
, Yunbo Guo
, Ying Liang, Lirong Wang, Siyu Meng, Yiwen Xu, Lei Liu, Yizhi Song
, Yuguo Tang:
GMDM-MoE: A biologically-inspired growth-to-morphology and dual-magnification mixture-of-experts for bacterial detection. Biomed. Signal Process. Control. 112: 108639 (2026)
[j4]Ying Liang, Lesong Zheng, Zihao Li, Siyu Meng, Yizhi Song
, Yiwen Xu, Lirong Wang
:
Bacterial perception-enhanced detection transformer in time-lapse images. Biomed. Signal Process. Control. 113: 109160 (2026)
[j3]Zhexiao Xiong, Wei Xiong, Jing Shi, He Zhang, Yizhi Song, Nathan Jacobs:
GroundingBooth: Grounding Text-to-Image Customization. Trans. Mach. Learn. Res. 2026 (2026)
[c9]Yolo Yunlong Tang, Jing Bi, Chao Huang, Susan Liang, Daiki Shimada, Hang Hua, Yunzhong Xiao, Yizhi Song, Pinxin Liu, Mingqian Feng, Junjia Guo, Zhuo Liu, Luchuan Song, Ali Vosoughi, Jinxi He, Liu He, Zeliang Zhang, Jiebo Luo, Chenliang Xu:
Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting. AAAI 2026: 41697-41699
[c8]Liu He, Xiao Zeng, Yizhi Song, Albert Y. C. Chen, Lu Xia, Shashwat Verma, Sankalp Dayal, Min Sun, Cheng-Hao Kuo, Daniel G. Aliaga:
Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation. WACV 2026: 5886-5897
[i15]Hengjia Li, Liming Jiang, Qing Yan, Yizhi Song, Hao Kang, Zichuan Liu, Xin Lu, Boxi Wu, Deng Cai:
ThinkRL-Edit: Thinking in Reinforcement Learning for Reasoning-Centric Image Editing. CoRR abs/2601.03467 (2026)
[i14]Zhexiao Xiong, Yizhi Song, Liu He, Wei Xiong, Yu Yuan, Feng Qiao, Nathan Jacobs:
PhysAlign: Physics-Coherent Image-to-Video Generation through Feature and 3D Representation Alignment. CoRR abs/2603.13770 (2026)
[i13]Zhexiao Xiong, Yizhi Song, Hao Kang, Qing Yan, Liming Jiang, Jenson Yang, Zhoujie Fu, Stathi Fotiadis, Angtian Wang, Zichuan Liu, Bo Liu, Yiding Yang, Xin Lu, Nathan Jacobs:
ActWorld: From Explorable to Interactive World Model via Action-Aware Memory. CoRR abs/2606.17730 (2026)- 2025
[j2]Lesong Zheng
, Yunbo Guo
, Ying Liang
, Lirong Wang
, Siyu Meng
, Yizhi Song
, Yuguo Tang
:
TrCL-AGS: A Universal Sequential Triple-Stage Contrastive Learning Framework for Bacterial Detection With Across-Growth-Stage Information. IEEE Internet Things J. 12(10): 14886-14896 (2025)
[c7]Yunlong Tang, Junjia Guo, Pinxin Liu, Zhiyuan Wang, Hang Hua, Jia-Xing Zhong, Yunzhong Xiao, Chao Huang, Luchuan Song, Susan Liang, Yizhi Song, Liu He, Jing Bi, Mingqian Feng, Xinyang Li, Zeliang Zhang, Chenliang Xu:
Generative AI for Cel-Animation: A Survey. ICCVW 2025: 3837-3850
[c6]Yizhi Song, Liu He, Zhifei Zhang, Soo Ye Kim, He Zhang, Wei Xiong, Zhe Lin, Brian L. Price, Scott Cohen, Jianming Zhang, Daniel G. Aliaga:
Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment. ICLR 2025
[c5]Hang Hua, Ziyun Zeng, Yizhi Song, Yunlong Tang, Liu He, Daniel G. Aliaga, Wei Xiong, Jiebo Luo:
MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models. NeurIPS 2025
[i12]Yunlong Tang, Junjia Guo, Pinxin Liu, Zhiyuan Wang, Hang Hua, Jia-Xing Zhong, Yunzhong Xiao, Chao Huang, Luchuan Song, Susan Liang, Yizhi Song, Liu He, Jing Bi, Mingqian Feng, Xinyang Li, Zeliang Zhang, Chenliang Xu:
Generative AI for Cel-Animation: A Survey. CoRR abs/2501.06250 (2025)
[i11]Yunlong Tang, Jing Bi, Chao Huang, Susan Liang, Daiki Shimada, Hang Hua, Yunzhong Xiao, Yizhi Song, Pinxin Liu, Mingqian Feng, Junjia Guo, Zhuo Liu, Luchuan Song, Ali Vosoughi, Jinxi He, Liu He, Zeliang Zhang, Jiebo Luo
, Chenliang Xu:
Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting. CoRR abs/2504.05541 (2025)
[i10]Hang Hua, Ziyun Zeng, Yizhi Song, Yunlong Tang, Liu He, Daniel G. Aliaga, Wei Xiong, Jiebo Luo
:
MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models. CoRR abs/2505.19415 (2025)
[i9]Liu He, Xiao Zeng, Yizhi Song, Albert Y. C. Chen, Lu Xia, Shashwat Verma, Sankalp Dayal, Min Sun, Cheng-Hao Kuo, Daniel G. Aliaga:
Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation. CoRR abs/2507.08513 (2025)
[i8]Yolo Yunlong Tang, Jing Bi, Pinxin Liu, Zhenyu Pan, Zhangyun Tan, Qianxiang Shen, Jiani Liu, Hang Hua, Junjia Guo, Yunzhong Xiao, Chao Huang, Zhiyuan Wang, Susan Liang, Xinyi Liu, Yizhi Song, Junhua Huang, Jia-Xing Zhong, Bozheng Li, Daiqing Qi, Ziyun Zeng, Ali Vosoughi, Luchuan Song, Zeliang Zhang, Daiki Shimada
, Han Liu, Jiebo Luo
, Chenliang Xu:
Video-LMM Post-Training: A Deep Dive into Video Reasoning with Large Multimodal Models. CoRR abs/2510.05034 (2025)- 2024
[c4]Yizhi Song, Zhifei Zhang, Zhe Lin, Scott Cohen, Brian L. Price, Jianming Zhang, Soo Ye Kim, He Zhang, Wei Xiong, Daniel G. Aliaga:
IMPRINT: Generative Object Compositing by Learning Identity-Preserving Representation. CVPR 2024: 8048-8058
[c3]Gemma Canet Tarrés
, Zhe Lin
, Zhifei Zhang
, Jianming Zhang
, Yizhi Song, Dan Ruta
, Andrew Gilbert
, John P. Collomosse
, Soo Ye Kim
:
Thinking Outside the BBox: Unconstrained Generative Object Compositing. ECCV (62) 2024: 476-495
[c2]Dongyu Lv, Yizhi Song
, Chao Xu
:
Nested and Interleaved Ticketing for Multiple Travelers. IJTCS-FAW 2024: 94-105
[i7]Yizhi Song, Zhifei Zhang, Zhe Lin, Scott Cohen, Brian L. Price, Jianming Zhang, Soo Ye Kim, He Zhang, Wei Xiong, Daniel G. Aliaga:
IMPRINT: Generative Object Compositing by Learning Identity-Preserving Representation. CoRR abs/2403.10701 (2024)
[i6]Liu He, Yizhi Song, Hejun Huang, Daniel G. Aliaga, Xin Zhou:
Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation. CoRR abs/2408.10453 (2024)
[i5]Gemma Canet Tarrés, Zhe Lin, Zhifei Zhang, Jianming Zhang, Yizhi Song, Dan Ruta, Andrew Gilbert, John P. Collomosse, Soo Ye Kim:
Thinking Outside the BBox: Unconstrained Generative Object Compositing. CoRR abs/2409.04559 (2024)
[i4]Zhexiao Xiong, Wei Xiong, Jing Shi, He Zhang, Yizhi Song, Nathan Jacobs:
GroundingBooth: Grounding Text-to-Image Customization. CoRR abs/2409.08520 (2024)
[i3]Yizhi Song, Liu He, Zhifei Zhang, Soo Ye Kim, He Zhang, Wei Xiong, Zhe Lin, Brian L. Price, Scott Cohen, Jianming Zhang, Daniel G. Aliaga:
Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment. CoRR abs/2412.00306 (2024)- 2023
[c1]Yizhi Song, Zhifei Zhang, Zhe Lin, Scott Cohen, Brian L. Price, Jianming Zhang, Soo Ye Kim, Daniel G. Aliaga:
ObjectStitch: Object Compositing with Diffusion Model. CVPR 2023: 18310-18319- 2022
[i2]Yizhi Song, Zhifei Zhang, Zhe Lin, Scott Cohen, Brian L. Price, Jianming Zhang, Soo Ye Kim, Daniel G. Aliaga:
ObjectStitch: Generative Object Compositing. CoRR abs/2212.00932 (2022)
2010 – 2019
- 2019
[j1]Yizhi Song, Ruochen Fan, Sharon X. Huang, Zhe Zhu, Ruofeng Tong:
A three-stage real-time detector for traffic signs in large panoramas. Comput. Vis. Media 5(4): 403-416 (2019)- 2017
[i1]Yizhi Song, Cheng Xu, Daoxin Ding, Hang Zhou, Tingwei Quan, Shiwei Li:
Properties on n-dimensional convolution for image deconvolution. CoRR abs/1711.11224 (2017)
Coauthor Index

manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.
Unpaywalled article links
Add open access links from
to the list of external document links (if available).
Privacy notice: By enabling the option above, your browser will contact the API of unpaywall.org to load hyperlinks to open access articles. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Unpaywall privacy policy.
Archived links via Wayback Machine
For web page which are no longer available, try to retrieve content from the
of the Internet Archive (if available).
Privacy notice: By enabling the option above, your browser will contact the API of archive.org to check for archived content of web pages that are no longer available. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Internet Archive privacy policy.
Reference lists
Add a list of references from
,
, and
to record detail pages.
load references from crossref.org and opencitations.net
Privacy notice: By enabling the option above, your browser will contact the APIs of crossref.org, opencitations.net, and semanticscholar.org to load article reference information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Crossref privacy policy and the OpenCitations privacy policy, as well as the AI2 Privacy Policy covering Semantic Scholar.
Citation data
Add a list of citing articles from
and
to record detail pages.
load citations from opencitations.net
Privacy notice: By enabling the option above, your browser will contact the API of opencitations.net and semanticscholar.org to load citation information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the OpenCitations privacy policy as well as the AI2 Privacy Policy covering Semantic Scholar.
OpenAlex data
Load additional information about publications from
.
Privacy notice: By enabling the option above, your browser will contact the API of openalex.org to load additional information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the information given by OpenAlex.
last updated on 2026-07-29 03:32 CEST by the dblp team
all metadata released as open data under CC0 1.0 license
see also: Terms of Use | Privacy Policy | Imprint


Google
Google Scholar
Semantic Scholar
Internet Archive Scholar
CiteSeerX
ORCID






