Pytorch implementation of the paper Improving Text-to-Image Synthesis Using Contrastive Learning

Last update: Dec 31, 2022

Related tags

Deep Learning T2I_CL

Overview

T2I_CL

This is the official Pytorch implementation of the paper Improving Text-to-Image Synthesis Using Contrastive Learning

Requirements

Linux
Python ≥ 3.6
PyTorch ≥ 1.4.0

Prepare Data

Download the preprocessed datasets from AttnGAN

Alternatively, another site is from DM-GAN

Training

Pretrain DAMSM+CL:
- For bird dataset: python pretrain_DAMSM.py --cfg cfg/DAMSM/bird.yml --gpu 0
- For coco dataset: python pretrain_DAMSM.py --cfg cfg/DAMSM/coco.yml --gpu 0
Train AttnGAN+CL:
- For bird dataset: python main.py --cfg cfg/bird_attn2.yml --gpu 0
- For coco dataset: python main.py --cfg cfg/coco_attn2.yml --gpu 0
Train DM-GAN+CL:
- For bird dataset: python main.py --cfg cfg/bird_DMGAN.yml --gpu 0
- For coco dataset: python main.py --cfg cfg/coco_DMGAN.yml --gpu 0

Pretrained Models

DAMSM+CL for bird. Download and save it to DAMSMencoders/
DAMSM+CL for coco. Download and save it to DAMSMencoders/
AttnGAN+CL for bird. Download and save it to models/
AttnGAN+CL for coco. Download and save it to models/
DM-GAN+CL for bird. Download and save it to models/
DM-GAN+CL for coco. Download and save it to models/

Evaluation

Sampling and get the R-precision:
- python main.py --cfg cfg/eval_bird.yml --gpu 0
- python main.py --cfg cfg/eval_coco.yml --gpu 0
Inception score:
- python inception_score_bird.py --image_folder fake_images_bird
- python inception_score_coco.py fake_images_coco
FID:
- python fid_score.py --gpu 0 --batch-size 50 --path1 real_images_bird --path2 fake_images_bird
- python fid_score.py --gpu 0 --batch-size 50 --path1 real_images_coco --path2 fake_images_coco

Citation

If you find this work useful in your research, please consider citing:

@article{ye2021improving,
  title={Improving Text-to-Image Synthesis Using Contrastive Learning},
  author={Ye, Hui and Yang, Xiulong and Takac, Martin and Sunderraman, Rajshekhar and Ji, Shihao},
  journal={arXiv preprint arXiv:2107.02423},
  year={2021}
}

Acknowledge

Our work is based on the following works:

Pytorch implementation of the paper Improving Text-to-Image Synthesis Using Contrastive Learning

Related tags

Overview

T2I_CL

Requirements

Prepare Data

Training

Pretrained Models

Evaluation

Citation

Acknowledge

Owner

DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models

UNet model with VGG11 encoder pre-trained on Kaggle Carvana dataset

Retinal vessel segmentation based on GT-UNet

Unsupervised 3D Human Mesh Recovery from Noisy Point Clouds

A modular framework for vision & language multimodal research from Facebook AI Research (FAIR)

Analysis of Smiles through reservoir sampling & RDkit

SIEM Logstash parsing for more than hundred technologies

Clean and readable code for Decision Transformer: Reinforcement Learning via Sequence Modeling

Ludwig is a toolbox that allows to train and evaluate deep learning models without the need to write code.

fklearn: Functional Machine Learning

Cobalt Strike teamserver detection.

Fully Convlutional Neural Networks for state-of-the-art time series classification

Cognition-aware Cognate Detection

网络协议2天集训

High-resolution networks and Segmentation Transformer for Semantic Segmentation

Self-Supervised Pre-Training for Transformer-Based Person Re-Identification

Convnet transfer - Code for paper How transferable are features in deep neural networks?

[ICCV 2021 Oral] Mining Latent Classes for Few-shot Segmentation

AnimationKit: AI Upscaling & Interpolation using Real-ESRGAN+RIFE

Embracing Single Stride 3D Object Detector with Sparse Transformer