Skip to main navigation Skip to search Skip to main content

Fontanimate: High Quality Few-Shot Font Generation Via Animating Font Transfer Process

  • Bin Fu
  • , Zixuan Wang
  • , Kainan Yan
  • , Shitian Zhao
  • , Qi Qin
  • , Jie Wen*
  • , Junjun He
  • , Peng Gao*
  • *Corresponding author for this work
  • Shenzhen Institute of Advanced Technology
  • Shanghai Artificial Intelligence Laboratory

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Few-shot font generation (FFG) aims to create new font images by imitating the style from a limited set of reference images, while maintaining the content from the source images. Although this task has achieved significant progress, most existing methods still suffer from incorrect generation of complicated character structure and detailed font style. To address the above issues, in this paper, we regard font generation as a font transfer process from the source font to the target font, and construct a video generation framework to model this process. Moreover, a test-time condition alignment mechanism is further developed to enhance the consistency between the generated samples and the provided condition samples. Specifically, we first construct a diffusion-based image-to-image font generation framework for the few-shot font generation task. This framework is expanded into an image-to-video font generation framework by integrating temporal components and frame-index information, enabling the production of high-quality font videos that transition from the source font to the target font. Based on this framework, we develop a noise inversion mechanism in the generative process to perform content and style alignment between the generated samples and the provided condition samples, enhancing style consistency and structural accuracy. The experimental results show that our model achieves superior performance on FFG tasks, demonstrating the effectiveness of our method. The code is available at: https://github.com/fubinfb/FontAnimate.

Original languageEnglish
Title of host publicationProceedings - 2025 IEEE/CVF International Conference on Computer Vision, ICCV 2025
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages16015-16025
Number of pages11
ISBN (Electronic)9798331587758
DOIs
StatePublished - 2025
Event2025 IEEE/CVF International Conference on Computer Vision, ICCV 2025 - Honolulu, United States
Duration: 19 Oct 202523 Oct 2025

Publication series

NameProceedings of the IEEE International Conference on Computer Vision
ISSN (Print)1550-5499
ISSN (Electronic)2380-7504

Conference

Conference2025 IEEE/CVF International Conference on Computer Vision, ICCV 2025
Country/TerritoryUnited States
CityHonolulu
Period19/10/2523/10/25

Fingerprint

Dive into the research topics of 'Fontanimate: High Quality Few-Shot Font Generation Via Animating Font Transfer Process'. Together they form a unique fingerprint.

Cite this