chaerinkong/paper-reviews

(very personal) Deep learning literature review

25

199 commits

updated Dec 21, 2023

See the code

README

Paper Reviews

Machine Learning paper reviews for personal archiving purpose.
Refer to [Issues] tab. (issues without labels are TBD)



Paper List

NoTitleVenueYearLink
1Learning representation from backpropagating errors1986[pdf]
2Unpaired Image to Image Translation using cycle-consistent advarsarial netICCV2017[pdf]
3Generative Adversarial NetsNIPS2014[pdf]
4Understanding Deep Learning requires rethinking GeneralizationCVPR2017[pdf]
5Show and Tell - A neural Image Caption GeneratorCVPR2015[pdf]
6Unsupervised Representation Learning with Deep Convolutional Generative Adversarial NetworksICLR2016[pdf]
7Explainable Artificial Intelligence: Understanding, Visualizing and Interpreting Deep Learning Models2017[pdf]
8Conditional Generative Adversarial Nets2014[pdf]
9CartoonGAN: Generative Adversarial Networks for Photo CartoonizationCVPR2018[pdf]
10Individualized Indicator for All: Stock-wise Technical Indicator Optimization with Stock EmbeddingKDD2019[pdf]
11Jukebox: A Generative Model for Music2020[pdf]
12Revisiting Self-supervised Visual Representation LearningCVPR2019[pdf]
13Reposing Humans by Warping 3D FeaturesCVPRW2020[pdf]
14Improving Language Understanding by Generative Pre-Training2018[pdf]
15Language Models are Unsupervised Multitask Learners2019[pdf]
16Implicit Maximum Likelihood Estimation2018[pdf]
17Pose Guided Person Image GenerationNIPS2017[pdf]
18Progressive Growing of GANs for Improved Quality, Stability and VariationICLR2018[pdf]
19Arbitrary Style Transfer in Real-time with Adaptive Instance NormalizationICCV2017[pdf]
20Style Transfer for Anime Sketches with Enhanced Residual U-net and Auxiliary Classifier GANACPR2017[pdf]
21Continuous Control with Deep Reinforcement Learning2015[pdf]
22Every Model Learned by Gradient Descent is Apporximately a Kernel Machine2020[pdf]
23Closed-Form Factorization of Latent Semantics in GANsCVPR2021[pdf]
24Large Scale GAN Training for High Fidelity Natural Image SynthesisICLR2019[pdf]
25Im2Pencil: Controllable Pencil Illustration from PhotographsCVPR2019[pdf]
26End-to-End Time-Lapse Video Synthesis from a Single Outdoor ImageCVPR2019[pdf]
27Improved Precision and Recall Metric for Assessing Generative ModelsNIPS2019[pdf]
28Freeze the Discriminator: a Simple Baseline for Fine-Tuning GANsCVPRW2020[pdf]
29Photo-Realistic Single Image Super-Resolution Using a Generative Adversarial NetworkCVPR2017[pdf]
30Sharpness-Aware Minimization for Efficiently Improving GeneralizationICLR2020[pdf]
31Differentiable Augmentation for Data-Efficient GAN TrainingNIPS2020[pdf]
32ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustnessICLR2019[pdf]
33U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationICLR2020[pdf]
34In-Domain GAN Inversion for Real Image EditingECCV2020[pdf]
35Feature Quantization Improves GAN TrainingICML2020[pdf]
36Improved Techniques for Training GANsNIPS2016[pdf]
37Generating Diverse High-Fidelity Images with VQ-VAE-2NIPS2019[pdf]
38Wide-Context Semantic Image ExtrapolationCVPR2019[pdf]
39Do 2D GANs Know 3D Shape? Unsupervised 3D Shape Reconstruction from 2D Image GANsICLR2021[pdf]
40Swapping Autoencoder for Deep Image ManipulationNIPS2020[pdf]
41Training GANs with Stronger Augmentations via Contrastive DiscriminatorICLR2021[pdf]
42Using Latent Space Regression to Analyze and Leverage Compositionality in GANs2021[pdf]
43On Self-supervised Image Representations For GAN EvaluationICLR2021[pdf]
44A Good Image Generator is What You Need for High-Resolution Video SynthesisICLR2021[pdf]
45Adversarial Latent Autoencoders2020[pdf]
46The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural NetworksICLR2019[pdf]
47How Neural Networks Extrapolate: From Feedforward to Graph Neural NetworksICLR2021[pdf]
48Density Estimation Using Real NVPICLR2017[pdf]
49Unsupervised Learning of Probably Symmetric Deformable 3D Object from Images in the WildCVPR2020[pdf]
50NeRF: Representing Scenes as Neural Radiance Fields for View SynthesisCVPR2020[pdf]
51Dataset Condensation with Gradient MatchingICLR2021[pdf]
52Momentum Contrast for Unsupervised Visual Representation LearningCVPR2020[pdf]
53Interpreting the Latent Space of GANs for Semantic Face EditingCVPR2020[pdf]
54Learning Continuous Image Representation with Local Implicit Image FunctionCVPR2021[pdf]
55Bootstrap your own latent: A new approach to self-supervised LearningNIPS2020[pdf]
56PULSE: Self-Supervised Photo Upsampling via Latent Space Exploration of Generative ModelsCVPR2020[pdf]
57Denoising Diffusion Probabilistic ModelsNIPS2020[pdf]
58Analyzing and Improving the Image Quality of StyleGANCVPR2020[pdf]
59Are Convolutional Neural Networks or Transformers more like human vision?2021[pdf]
60Pay Attention to MLPs2021[pdf]
61Rethinking and Improving the Robustness of Image Style TransferCVPR2021[pdf]
62Style-Aware Normalized Loss for Improving Arbitrary Style TransferCVPR2021[pdf]
63GIRAFFE: Representing Scenes as Compositional Generative Neural Feature FieldsCVPR2021[pdf]
64Densely connected multidilated convolutional networks for dense prediction tasksCVPR2021[pdf]
65ArtFlow: Unbiased Image Style Transfer via Reversible Neural FlowsCVPR2021[pdf]
66Dual Contradistinctive Generative AutoencoderCVPR2021[pdf]
67Exponential Moving Average Normalization for Self-supervised and Semi-supervised LearningCVPR2021[pdf]
68AdCo: Adversarial Contrast for Efficient Learning of Unsupervised Representations from Self-Trained Negative AdversariesCVPR2021[pdf]
69Training Networks in Null Space of Feature Covariance for Continual LearningCVPR2021[pdf]
70NeRF in the Wild: Neural Radiance Fields for Unconstrained Photo CollectionsCVPR2021[pdf]
71Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsNIPS2020[pdf]
72What Uncertainties Do We Need in Bayesian Deep Learning for Computer Vision?NIPS2017[pdf]
73GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisNIPS2020[pdf]
74Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene RepresentationsNIPS2019[pdf]
75Neural Volumes: Learning Dynamic Renderable Volumes from ImagesSIGGRAPH2019[pdf]
76pixelNeRF: Neural Radiance Fields from One or Few ImagesCVPR2021[pdf]
77PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationICCV2019[pdf]
78pi-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image SynthesisCVPR2021[pdf]
79Edge Guided Progressively Generative Image OutpaintingCVPRW2021[pdf]
803D Shape Generation with Grid-based Implicit FunctionsCVPR2021[pdf]
81Deep Image PriorCVPR2018[pdf]
82Training BatchNorm and Only BatchNorm: On the Expressive Power of Random Features in CNNsICLR2021[pdf]
83Taming Transformers for High-Resolution Image SynthesisCVPR2021[pdf]
84Neural Geometric Level of Detail: Real-time Rendering with Implicit 3D ShapesCVPR2021[pdf]
85Neural Sparse Voxel FieldsNIPS2020[pdf]
86Triplet is All You Need with Random Mappings for Unsupervised Visual Representation Learning2021[pdf]
87Few-shot Image Generation via Cross-domain CorrespondenceCVPR2021[pdf]
88Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?ICCV2019[pdf]
89NeX: Real-time View Synthesis with neural Basis ExpansionCVPR2021[pdf]
90Few-shot Image Generation with Elastic Weight ConsolidationNIPS2020[pdf]
91Towards Unsupervised Learning of Generative Models for 3D Controllable Image SynthesisCVPR2020[pdf]
92On the Effectiveness of Weight-Encoded Neural Implicit 3D ShapesICML2021[pdf]
93AdaCoF: Adaptive Collaboration of Flows for Video Frame InterpolationCVPR2020[pdf]
94Softmax Splatting for Video Frame InterpolationCVPR2020[pdf]
95Time Lens: Event-based Video Frame InterpolationCVPR2021[pdf]
96Revisiting Adaptive Convolutions for Video Frame InterpolationWACV2021[pdf]
97Depth-Aware Video Frame InterpolationCVPR2019[pdf]
98Channel Attention is All You Need for Video Frame InterpolationAAAI2020[pdf]
99PWC-Net: CNNs for Optical Flow Using Pyramid,Warping, and Cost VolumeCVPR2018[pdf]
100RAFT: Recurrent All-Pairs Field Transforms for Optical FlowECCV2020[pdf]
101What Matters in Unsupervised Optical FlowECCV2020[pdf]
102Occlusion Aware Unsupervised Learning of Optical FlowCVPR2018[pdf]
103UnFlow: Unsupervised Learning of Optical Flow with a Bidirectional Census LossAAAI2018[pdf]
104DDFlow: Learning Optical Flow with Unlabeled Data DistillationAAAI2019[pdf]
105SelFlow: Self-Supervised Learning of Optical FlowCVPR2019[pdf]
106GeoNet: Unsupervised Learning of Dense Depth, Optical Flow and Camera PoseCVPR2018[pdf]
107Zero Shot Text to Image GenerationOpenAI2021[pdf]
108LayoutTransformer: Scene Layout Generation With Conceptual and Spatial DiversityCVPR2021[pdf]
109StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial NetworksICCV2017[pdf]
110AttnGAN: Fine-Grained Text to Image Generation with Attentional Generative Adversarial NetworksCVPR2018[pdf]
111Image Synthesis From Reconfigurable Layout and StyleICCV2019[pdf]
112Image Generation from LayoutCVPR2019[pdf]
113Context-aware Layout to Image Generation with Enhanced Object AppearanceCVPR2021[pdf]
114Attribute-Guided Image Generation From LayoutBMVC2020[pdf]
115Object-Centric Image Generation from LayoutsAAAI2021[pdf]
116Learning Layout and Style Reconfigurable GANs for Controllable Image SynthesisTPAMI2021[pdf]
117BachGAN: High-Resolution Image Synthesis from Salient Object LayoutCVPR2020[pdf]
118Region-aware Adaptive Instance Normalization for Image HarmonizationCVPR2021[pdf]
119StEP: Style-based Encoder Pre-training for Multi-modal Image SynthesisCVPR2021[pdf]
120Optimizing the Latent Space of Generative NetworksICML2018[pdf]
121StyleSpace Analysis: Disentangled Controls for StyleGAN Image GenerationCVPR2021[pdf]
122Autoencoder Image Interpolation by Shaping the Latent SpaceICML2021[pdf]
123Autoencoding Under Normalization ConstraintsICML2021[pdf]
124Cross-Modal Contrastive Contrastive Learning for Text-to-Image GenerationCVPR2021[pdf]
125Action-Conditioned 3D Human Motion Synthesis with Transformer VAEICCV2021[pdf]
126Every Pixel Matters: Center-aware Feature Alignment for Domain Adaptive Object DetectorECCV2020[pdf]
127A Unified 3D Human Motion Synthesis Model via Conditional Variational Auto-EncoderICCV2021[pdf]
128NUWA: Visual Synthesis Pre-training for Neural visUal World creAtion--[pdf]
129Pix2seq: A Language Modeling Framework for Object Detection--[pdf]
130Masked Autoencoders Are Scalable Vision Learners-2021[pdf]
131Vokenization: Improving Language Understanding with Contextualized, Visual-Grounded SupervisionEMNLP2020[pdf]
132Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksNIPS2020[pdf]
133VidLanKD: Improving Language Understanding via Video-Distilled Knowledge TransferNIPS2021[pdf]
134You Only Learn One Representation: Unified Network for Multiple Tasks--[pdf]
135StyleNeRF: A Style-based 3D-Aware Generator for High-resolution Image SynthesisICLR2022[pdf]
136Co2L: Contrastive Continual LearningICCV2021[pdf]
137Adversarial Generation of Continuous ImagesCVPR2021[pdf]
138MetaFormer is Actually What You Need for Vision--[pdf]
139Rehearsal revealed: The limits and merits of revisiting samples in continual learningICCV2021[pdf]
140GDumb: A Simple Approach that Questions Our Progress in Continual LearningECCV2020[pdf]
141GAN Memory with No ForgettingNIPS2020[pdf]
142Few-Shot and Continual Learning With Attentive Independent MechanismsICCV2021[pdf]
143Self-Supervised GANs with Label AugmentationNIPS2021[pdf]
144Overcoming Catastrophic Forgetting in Incremental Few-Shot Learning by Finding Flat MinimaNIPS2021[pdf]
145On Memorization in Probabilistic Deep Generative ModelsNIPS2021[pdf]
146Low-Rank Subspaces in GANsNIPS2021[pdf]
147Multimodal Few-Shot Learning with Frozen Language ModelsNIPS2021[pdf]
148Encoding in Style: a StyleGAN Encoder for Image-to-Image TranslationNIPS2021[pdf]
149Do Vision Transformers See Like Convolutional Neural Networks?NIPS2021[pdf]
150Generating Videos with Dynamics-aware Implicit Generative Adversarial NetworksICLR2022[pdf]
151On the Measure of Intelligence-2019[pdf]
152Collapse by Conditioning: Training Class-conditional GANs with Limited DataICLR2022[pdf]
153Regularizing Generative Adversarial Networks under Limited DataCVPR2021[pdf]
154Denoising Diffusion Probabilistic ModelsNIPS2020[pdf]
155Improved Denoising Diffusion Probabilistic ModelsICML2021[pdf]
156Denoising Diffusion Implicit ModelsICLR2021[pdf]
157TRGP: Trust Region Gradient Projection for Continual LearningICLR2022[pdf]
158Rethinking the Representational Continuity: Towards Unsupervised Continual LearningICLR2022[pdf]
159Diffusion Models Beat GANs on Image SynthesisNIPS2021[pdf]
160Classifier-free Diffusion GuidanceNIPS-W2021[pdf]
161GLIDE: Towards Photorealistic Image Generation and Editing with Text-guided Diffusion Models-2021[pdf]
162VITON: An Image-based Virtual Try-on NetworkCVPR2018[pdf]
163Toward Characteristic-Preserving Image-based Virtual Try-On NetworkECCV2018[pdf]
164VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware NormalizationCVPR2021[pdf]
165SharpContour: A Contour-based Boundary Refinement Approach for Efficient and Accurate Instance SegmentationCVPR2022[pdf]
166Deblur-NeRF: Neural Radiance Fields from Blurry ImagesCVPR2022[pdf]
167Fashion Attribute-to-Image Synthesis Using Attention-based Generative Adversarial NetworkWACV2019[pdf]
168Attribute Manipulaiton Generative Adversarial Networks for Fashion ImagesICCV2019[pdf]
169FlexIT: Towards Flexible Semantic Image TranslationCVPR2022[pdf]
170Point-NeRF: Point-based Neural Radiance FieldsCVPR2022[pdf]
171CLIP-Event: Connecting Text and Images with Event StructuresCVPR2022[pdf]
172InfoNeRF: Ray Entropy Minimization for Few-Shot Neural Volume RenderingCVPR2022[pdf]
173StyleT2I: Toward Compositional and High-Fidelity Text-to-Image SynthesisCVPR2022[pdf]
174CAM-GAN: Continual Adaptation Modules for Generative Adversarial NetworksNIPS2021[pdf]
175Blended Diffusion for Text-driven Editing of Natural ImagesCVPR2022[pdf]
176Multimodal Contrastive Learning with LIMoE: the Language-Image Mixture of Experts-2022[pdf]
177BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationICML2022[pdf]
178Vision-Language Pre-Training with Triple Contrastive LearningCVPR2022[pdf]
179Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks-2022[pdf]
180Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding-2022[pdf]
181Prompt-to-Prompt Image Editing with Cross Attention Control-2022[pdf]
182An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion-2022[pdf]
183Video Diffusion Models-2022[pdf]
184Diffusion Probabilistic Modeling for Video Generation-2022[pdf]
185High-resolution Image Synthesis with Latent Diffusion ModelsCVPR2022[pdf]
186Visual Prompting via Image InpaintingNeurIPS2022[pdf]
187On Distillation of Guided Diffusion Models-2022[pdf]
188Unifying Diffusion Models' Latent Space, with Applications to CycleDiffusion and Guidance-2022[pdf]
189eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers-2022[pdf]
190Any-resolution Training for High-resolution Image SynthesisECCV2022[pdf]
191Variational Diffusion ModelsNeurIPS2021[pdf]
192Elucidating the Design Space of Diffusion-Based Generative ModelsNeurIPS2022[pdf]
193Visual Prompt Tuning for Generative Transfer Learning-2022[pdf]
194The Euclidean Space is Evil: Hyperbolic Attribute Editing for Few-shot Image Generation-2022[pdf]
195Language Models with Image Descriptors are Strong Few-Shot Video-Language LearnersNeurIPS2022[pdf]
196DreamFusion: Text-to-3D Using 2D DiffusionICLR2023[pdf]
197Sketch-Guided Text-to-Image Diffusion Models-2023[pdf]
198Dataset Distillation by Matching Training TrajectoriesCVPR2022[pdf]
199Is synthetic data from generative models ready for image recognition?ICLR2023[pdf]
200Generative Models as a Data Source for Multiview Representation LearningICLR2022[pdf]
201Scalable Diffusion Models with Transformers-2022[pdf]
202CAFE: Learning to Condense Dataset by Aligning FeaturesCVPR2022[pdf]
203Learning to Learn with Generative Models of Neural Network Checkpoints-2022[pdf]
204Synthesizing Informative Training Samples with GANNeurIPSW2022[pdf]
205Sliced Score Matching: A Scalable Approach to Density and Score EstimationUAI2019[pdf]
206Generative Modeling by Estimating Gradients of the Data DistributionNeurIPS2019[pdf]
207Dataset Distillation via FactorizationNeurIPS2022[pdf]
208Visual Classification via Description from Large Language ModelsICLR2023[pdf]
209Training Language Models to Follow Instructions from Human Feedback-2022[pdf]
210Adding Conditional Control to Text-to-Image Diffusion Models-2023[pdf]
211LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention-2023[pdf]
212BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models-2023[pdf]
213InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning-2023[pdf]
214Diffusion Models already have a Semantic Latent SpaceICLR2023[pdf]
215Train Short, Test Long: Attention with Linear Biases Enables Input Length ExtrapolationICLR2022[pdf]
216Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter TransferNeurIPS2021[pdf]
217Any-to-Any Generation via Composable Diffusion-2023[pdf]
218Too Large; Data Reduction for Vision-Language Pre-Training-2023[pdf]
219Controllable Text-to-Image Generation with GPT-4-2023[pdf]
220VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners-2023[pdf]
221Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video CaptioningCVPR2023[pdf]
222Visual Instruction Tuning-2023[pdf]
223PaLM-E: An Embodied Multimodal Language Model-2023[pdf]
224Flamingo: a Visual Language Model for Few-Shot LearningNeurIPS2022[pdf]
225Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsNeurIPS2022[pdf]
226SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking-2023[pdf]
227Object-centric Learning with Cyclic Walks between Parts and WholeNeurIPS2023[pdf]
228Characterizing the Impacts of Semi-supervised Learning for Weak SupervisionNeurIPS2023[pdf]
229Brain Decoding: Toward Real-time Reconstruction of Visual Perception-2023[pdf]
230One-step Diffusion with Distribution Matching Distillation-2023[pdf]
231Analyzing and Improving the Training Dynamics of Diffusion Models-2023[pdf]
232Common Diffusion Noise Schedules and Sample Steps are Flawed-2023[pdf]
233On the Importance of Noise Scheduling for Diffusion Models-2023[pdf]
234StyleGAN-T: Unlocking the Power of GANs for Fast Large-Scale Text-to-Image Synthesis-2023[pdf]
deep-learning
machine-learning

Contributors

chaerinkong

199 commits

chaerinkong/paper-reviews

(very personal) Deep learning literature review

25

199 commits

updated Dec 21, 2023

See the code

README

Paper Reviews

Machine Learning paper reviews for personal archiving purpose.
Refer to [Issues] tab. (issues without labels are TBD)



Paper List

NoTitleVenueYearLink
1Learning representation from backpropagating errors1986[pdf]
2Unpaired Image to Image Translation using cycle-consistent advarsarial netICCV2017[pdf]
3Generative Adversarial NetsNIPS2014[pdf]
4Understanding Deep Learning requires rethinking GeneralizationCVPR2017[pdf]
5Show and Tell - A neural Image Caption GeneratorCVPR2015[pdf]
6Unsupervised Representation Learning with Deep Convolutional Generative Adversarial NetworksICLR2016[pdf]
7Explainable Artificial Intelligence: Understanding, Visualizing and Interpreting Deep Learning Models2017[pdf]
8Conditional Generative Adversarial Nets2014[pdf]
9CartoonGAN: Generative Adversarial Networks for Photo CartoonizationCVPR2018[pdf]
10Individualized Indicator for All: Stock-wise Technical Indicator Optimization with Stock EmbeddingKDD2019[pdf]
11Jukebox: A Generative Model for Music2020[pdf]
12Revisiting Self-supervised Visual Representation LearningCVPR2019[pdf]
13Reposing Humans by Warping 3D FeaturesCVPRW2020[pdf]
14Improving Language Understanding by Generative Pre-Training2018[pdf]
15Language Models are Unsupervised Multitask Learners2019[pdf]
16Implicit Maximum Likelihood Estimation2018[pdf]
17Pose Guided Person Image GenerationNIPS2017[pdf]
18Progressive Growing of GANs for Improved Quality, Stability and VariationICLR2018[pdf]
19Arbitrary Style Transfer in Real-time with Adaptive Instance NormalizationICCV2017[pdf]
20Style Transfer for Anime Sketches with Enhanced Residual U-net and Auxiliary Classifier GANACPR2017[pdf]
21Continuous Control with Deep Reinforcement Learning2015[pdf]
22Every Model Learned by Gradient Descent is Apporximately a Kernel Machine2020[pdf]
23Closed-Form Factorization of Latent Semantics in GANsCVPR2021[pdf]
24Large Scale GAN Training for High Fidelity Natural Image SynthesisICLR2019[pdf]
25Im2Pencil: Controllable Pencil Illustration from PhotographsCVPR2019[pdf]
26End-to-End Time-Lapse Video Synthesis from a Single Outdoor ImageCVPR2019[pdf]
27Improved Precision and Recall Metric for Assessing Generative ModelsNIPS2019[pdf]
28Freeze the Discriminator: a Simple Baseline for Fine-Tuning GANsCVPRW2020[pdf]
29Photo-Realistic Single Image Super-Resolution Using a Generative Adversarial NetworkCVPR2017[pdf]
30Sharpness-Aware Minimization for Efficiently Improving GeneralizationICLR2020[pdf]
31Differentiable Augmentation for Data-Efficient GAN TrainingNIPS2020[pdf]
32ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustnessICLR2019[pdf]
33U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationICLR2020[pdf]
34In-Domain GAN Inversion for Real Image EditingECCV2020[pdf]
35Feature Quantization Improves GAN TrainingICML2020[pdf]
36Improved Techniques for Training GANsNIPS2016[pdf]
37Generating Diverse High-Fidelity Images with VQ-VAE-2NIPS2019[pdf]
38Wide-Context Semantic Image ExtrapolationCVPR2019[pdf]
39Do 2D GANs Know 3D Shape? Unsupervised 3D Shape Reconstruction from 2D Image GANsICLR2021[pdf]
40Swapping Autoencoder for Deep Image ManipulationNIPS2020[pdf]
41Training GANs with Stronger Augmentations via Contrastive DiscriminatorICLR2021[pdf]
42Using Latent Space Regression to Analyze and Leverage Compositionality in GANs2021[pdf]
43On Self-supervised Image Representations For GAN EvaluationICLR2021[pdf]
44A Good Image Generator is What You Need for High-Resolution Video SynthesisICLR2021[pdf]
45Adversarial Latent Autoencoders2020[pdf]
46The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural NetworksICLR2019[pdf]
47How Neural Networks Extrapolate: From Feedforward to Graph Neural NetworksICLR2021[pdf]
48Density Estimation Using Real NVPICLR2017[pdf]
49Unsupervised Learning of Probably Symmetric Deformable 3D Object from Images in the WildCVPR2020[pdf]
50NeRF: Representing Scenes as Neural Radiance Fields for View SynthesisCVPR2020[pdf]
51Dataset Condensation with Gradient MatchingICLR2021[pdf]
52Momentum Contrast for Unsupervised Visual Representation LearningCVPR2020[pdf]
53Interpreting the Latent Space of GANs for Semantic Face EditingCVPR2020[pdf]
54Learning Continuous Image Representation with Local Implicit Image FunctionCVPR2021[pdf]
55Bootstrap your own latent: A new approach to self-supervised LearningNIPS2020[pdf]
56PULSE: Self-Supervised Photo Upsampling via Latent Space Exploration of Generative ModelsCVPR2020[pdf]
57Denoising Diffusion Probabilistic ModelsNIPS2020[pdf]
58Analyzing and Improving the Image Quality of StyleGANCVPR2020[pdf]
59Are Convolutional Neural Networks or Transformers more like human vision?2021[pdf]
60Pay Attention to MLPs2021[pdf]
61Rethinking and Improving the Robustness of Image Style TransferCVPR2021[pdf]
62Style-Aware Normalized Loss for Improving Arbitrary Style TransferCVPR2021[pdf]
63GIRAFFE: Representing Scenes as Compositional Generative Neural Feature FieldsCVPR2021[pdf]
64Densely connected multidilated convolutional networks for dense prediction tasksCVPR2021[pdf]
65ArtFlow: Unbiased Image Style Transfer via Reversible Neural FlowsCVPR2021[pdf]
66Dual Contradistinctive Generative AutoencoderCVPR2021[pdf]
67Exponential Moving Average Normalization for Self-supervised and Semi-supervised LearningCVPR2021[pdf]
68AdCo: Adversarial Contrast for Efficient Learning of Unsupervised Representations from Self-Trained Negative AdversariesCVPR2021[pdf]
69Training Networks in Null Space of Feature Covariance for Continual LearningCVPR2021[pdf]
70NeRF in the Wild: Neural Radiance Fields for Unconstrained Photo CollectionsCVPR2021[pdf]
71Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsNIPS2020[pdf]
72What Uncertainties Do We Need in Bayesian Deep Learning for Computer Vision?NIPS2017[pdf]
73GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisNIPS2020[pdf]
74Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene RepresentationsNIPS2019[pdf]
75Neural Volumes: Learning Dynamic Renderable Volumes from ImagesSIGGRAPH2019[pdf]
76pixelNeRF: Neural Radiance Fields from One or Few ImagesCVPR2021[pdf]
77PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationICCV2019[pdf]
78pi-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image SynthesisCVPR2021[pdf]
79Edge Guided Progressively Generative Image OutpaintingCVPRW2021[pdf]
803D Shape Generation with Grid-based Implicit FunctionsCVPR2021[pdf]
81Deep Image PriorCVPR2018[pdf]
82Training BatchNorm and Only BatchNorm: On the Expressive Power of Random Features in CNNsICLR2021[pdf]
83Taming Transformers for High-Resolution Image SynthesisCVPR2021[pdf]
84Neural Geometric Level of Detail: Real-time Rendering with Implicit 3D ShapesCVPR2021[pdf]
85Neural Sparse Voxel FieldsNIPS2020[pdf]
86Triplet is All You Need with Random Mappings for Unsupervised Visual Representation Learning2021[pdf]
87Few-shot Image Generation via Cross-domain CorrespondenceCVPR2021[pdf]
88Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?ICCV2019[pdf]
89NeX: Real-time View Synthesis with neural Basis ExpansionCVPR2021[pdf]
90Few-shot Image Generation with Elastic Weight ConsolidationNIPS2020[pdf]
91Towards Unsupervised Learning of Generative Models for 3D Controllable Image SynthesisCVPR2020[pdf]
92On the Effectiveness of Weight-Encoded Neural Implicit 3D ShapesICML2021[pdf]
93AdaCoF: Adaptive Collaboration of Flows for Video Frame InterpolationCVPR2020[pdf]
94Softmax Splatting for Video Frame InterpolationCVPR2020[pdf]
95Time Lens: Event-based Video Frame InterpolationCVPR2021[pdf]
96Revisiting Adaptive Convolutions for Video Frame InterpolationWACV2021[pdf]
97Depth-Aware Video Frame InterpolationCVPR2019[pdf]
98Channel Attention is All You Need for Video Frame InterpolationAAAI2020[pdf]
99PWC-Net: CNNs for Optical Flow Using Pyramid,Warping, and Cost VolumeCVPR2018[pdf]
100RAFT: Recurrent All-Pairs Field Transforms for Optical FlowECCV2020[pdf]
101What Matters in Unsupervised Optical FlowECCV2020[pdf]
102Occlusion Aware Unsupervised Learning of Optical FlowCVPR2018[pdf]
103UnFlow: Unsupervised Learning of Optical Flow with a Bidirectional Census LossAAAI2018[pdf]
104DDFlow: Learning Optical Flow with Unlabeled Data DistillationAAAI2019[pdf]
105SelFlow: Self-Supervised Learning of Optical FlowCVPR2019[pdf]
106GeoNet: Unsupervised Learning of Dense Depth, Optical Flow and Camera PoseCVPR2018[pdf]
107Zero Shot Text to Image GenerationOpenAI2021[pdf]
108LayoutTransformer: Scene Layout Generation With Conceptual and Spatial DiversityCVPR2021[pdf]
109StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial NetworksICCV2017[pdf]
110AttnGAN: Fine-Grained Text to Image Generation with Attentional Generative Adversarial NetworksCVPR2018[pdf]
111Image Synthesis From Reconfigurable Layout and StyleICCV2019[pdf]
112Image Generation from LayoutCVPR2019[pdf]
113Context-aware Layout to Image Generation with Enhanced Object AppearanceCVPR2021[pdf]
114Attribute-Guided Image Generation From LayoutBMVC2020[pdf]
115Object-Centric Image Generation from LayoutsAAAI2021[pdf]
116Learning Layout and Style Reconfigurable GANs for Controllable Image SynthesisTPAMI2021[pdf]
117BachGAN: High-Resolution Image Synthesis from Salient Object LayoutCVPR2020[pdf]
118Region-aware Adaptive Instance Normalization for Image HarmonizationCVPR2021[pdf]
119StEP: Style-based Encoder Pre-training for Multi-modal Image SynthesisCVPR2021[pdf]
120Optimizing the Latent Space of Generative NetworksICML2018[pdf]
121StyleSpace Analysis: Disentangled Controls for StyleGAN Image GenerationCVPR2021[pdf]
122Autoencoder Image Interpolation by Shaping the Latent SpaceICML2021[pdf]
123Autoencoding Under Normalization ConstraintsICML2021[pdf]
124Cross-Modal Contrastive Contrastive Learning for Text-to-Image GenerationCVPR2021[pdf]
125Action-Conditioned 3D Human Motion Synthesis with Transformer VAEICCV2021[pdf]
126Every Pixel Matters: Center-aware Feature Alignment for Domain Adaptive Object DetectorECCV2020[pdf]
127A Unified 3D Human Motion Synthesis Model via Conditional Variational Auto-EncoderICCV2021[pdf]
128NUWA: Visual Synthesis Pre-training for Neural visUal World creAtion--[pdf]
129Pix2seq: A Language Modeling Framework for Object Detection--[pdf]
130Masked Autoencoders Are Scalable Vision Learners-2021[pdf]
131Vokenization: Improving Language Understanding with Contextualized, Visual-Grounded SupervisionEMNLP2020[pdf]
132Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksNIPS2020[pdf]
133VidLanKD: Improving Language Understanding via Video-Distilled Knowledge TransferNIPS2021[pdf]
134You Only Learn One Representation: Unified Network for Multiple Tasks--[pdf]
135StyleNeRF: A Style-based 3D-Aware Generator for High-resolution Image SynthesisICLR2022[pdf]
136Co2L: Contrastive Continual LearningICCV2021[pdf]
137Adversarial Generation of Continuous ImagesCVPR2021[pdf]
138MetaFormer is Actually What You Need for Vision--[pdf]
139Rehearsal revealed: The limits and merits of revisiting samples in continual learningICCV2021[pdf]
140GDumb: A Simple Approach that Questions Our Progress in Continual LearningECCV2020[pdf]
141GAN Memory with No ForgettingNIPS2020[pdf]
142Few-Shot and Continual Learning With Attentive Independent MechanismsICCV2021[pdf]
143Self-Supervised GANs with Label AugmentationNIPS2021[pdf]
144Overcoming Catastrophic Forgetting in Incremental Few-Shot Learning by Finding Flat MinimaNIPS2021[pdf]
145On Memorization in Probabilistic Deep Generative ModelsNIPS2021[pdf]
146Low-Rank Subspaces in GANsNIPS2021[pdf]
147Multimodal Few-Shot Learning with Frozen Language ModelsNIPS2021[pdf]
148Encoding in Style: a StyleGAN Encoder for Image-to-Image TranslationNIPS2021[pdf]
149Do Vision Transformers See Like Convolutional Neural Networks?NIPS2021[pdf]
150Generating Videos with Dynamics-aware Implicit Generative Adversarial NetworksICLR2022[pdf]
151On the Measure of Intelligence-2019[pdf]
152Collapse by Conditioning: Training Class-conditional GANs with Limited DataICLR2022[pdf]
153Regularizing Generative Adversarial Networks under Limited DataCVPR2021[pdf]
154Denoising Diffusion Probabilistic ModelsNIPS2020[pdf]
155Improved Denoising Diffusion Probabilistic ModelsICML2021[pdf]
156Denoising Diffusion Implicit ModelsICLR2021[pdf]
157TRGP: Trust Region Gradient Projection for Continual LearningICLR2022[pdf]
158Rethinking the Representational Continuity: Towards Unsupervised Continual LearningICLR2022[pdf]
159Diffusion Models Beat GANs on Image SynthesisNIPS2021[pdf]
160Classifier-free Diffusion GuidanceNIPS-W2021[pdf]
161GLIDE: Towards Photorealistic Image Generation and Editing with Text-guided Diffusion Models-2021[pdf]
162VITON: An Image-based Virtual Try-on NetworkCVPR2018[pdf]
163Toward Characteristic-Preserving Image-based Virtual Try-On NetworkECCV2018[pdf]
164VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware NormalizationCVPR2021[pdf]
165SharpContour: A Contour-based Boundary Refinement Approach for Efficient and Accurate Instance SegmentationCVPR2022[pdf]
166Deblur-NeRF: Neural Radiance Fields from Blurry ImagesCVPR2022[pdf]
167Fashion Attribute-to-Image Synthesis Using Attention-based Generative Adversarial NetworkWACV2019[pdf]
168Attribute Manipulaiton Generative Adversarial Networks for Fashion ImagesICCV2019[pdf]
169FlexIT: Towards Flexible Semantic Image TranslationCVPR2022[pdf]
170Point-NeRF: Point-based Neural Radiance FieldsCVPR2022[pdf]
171CLIP-Event: Connecting Text and Images with Event StructuresCVPR2022[pdf]
172InfoNeRF: Ray Entropy Minimization for Few-Shot Neural Volume RenderingCVPR2022[pdf]
173StyleT2I: Toward Compositional and High-Fidelity Text-to-Image SynthesisCVPR2022[pdf]
174CAM-GAN: Continual Adaptation Modules for Generative Adversarial NetworksNIPS2021[pdf]
175Blended Diffusion for Text-driven Editing of Natural ImagesCVPR2022[pdf]
176Multimodal Contrastive Learning with LIMoE: the Language-Image Mixture of Experts-2022[pdf]
177BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationICML2022[pdf]
178Vision-Language Pre-Training with Triple Contrastive LearningCVPR2022[pdf]
179Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks-2022[pdf]
180Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding-2022[pdf]
181Prompt-to-Prompt Image Editing with Cross Attention Control-2022[pdf]
182An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion-2022[pdf]
183Video Diffusion Models-2022[pdf]
184Diffusion Probabilistic Modeling for Video Generation-2022[pdf]
185High-resolution Image Synthesis with Latent Diffusion ModelsCVPR2022[pdf]
186Visual Prompting via Image InpaintingNeurIPS2022[pdf]
187On Distillation of Guided Diffusion Models-2022[pdf]
188Unifying Diffusion Models' Latent Space, with Applications to CycleDiffusion and Guidance-2022[pdf]
189eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers-2022[pdf]
190Any-resolution Training for High-resolution Image SynthesisECCV2022[pdf]
191Variational Diffusion ModelsNeurIPS2021[pdf]
192Elucidating the Design Space of Diffusion-Based Generative ModelsNeurIPS2022[pdf]
193Visual Prompt Tuning for Generative Transfer Learning-2022[pdf]
194The Euclidean Space is Evil: Hyperbolic Attribute Editing for Few-shot Image Generation-2022[pdf]
195Language Models with Image Descriptors are Strong Few-Shot Video-Language LearnersNeurIPS2022[pdf]
196DreamFusion: Text-to-3D Using 2D DiffusionICLR2023[pdf]
197Sketch-Guided Text-to-Image Diffusion Models-2023[pdf]
198Dataset Distillation by Matching Training TrajectoriesCVPR2022[pdf]
199Is synthetic data from generative models ready for image recognition?ICLR2023[pdf]
200Generative Models as a Data Source for Multiview Representation LearningICLR2022[pdf]
201Scalable Diffusion Models with Transformers-2022[pdf]
202CAFE: Learning to Condense Dataset by Aligning FeaturesCVPR2022[pdf]
203Learning to Learn with Generative Models of Neural Network Checkpoints-2022[pdf]
204Synthesizing Informative Training Samples with GANNeurIPSW2022[pdf]
205Sliced Score Matching: A Scalable Approach to Density and Score EstimationUAI2019[pdf]
206Generative Modeling by Estimating Gradients of the Data DistributionNeurIPS2019[pdf]
207Dataset Distillation via FactorizationNeurIPS2022[pdf]
208Visual Classification via Description from Large Language ModelsICLR2023[pdf]
209Training Language Models to Follow Instructions from Human Feedback-2022[pdf]
210Adding Conditional Control to Text-to-Image Diffusion Models-2023[pdf]
211LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention-2023[pdf]
212BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models-2023[pdf]
213InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning-2023[pdf]
214Diffusion Models already have a Semantic Latent SpaceICLR2023[pdf]
215Train Short, Test Long: Attention with Linear Biases Enables Input Length ExtrapolationICLR2022[pdf]
216Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter TransferNeurIPS2021[pdf]
217Any-to-Any Generation via Composable Diffusion-2023[pdf]
218Too Large; Data Reduction for Vision-Language Pre-Training-2023[pdf]
219Controllable Text-to-Image Generation with GPT-4-2023[pdf]
220VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners-2023[pdf]
221Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video CaptioningCVPR2023[pdf]
222Visual Instruction Tuning-2023[pdf]
223PaLM-E: An Embodied Multimodal Language Model-2023[pdf]
224Flamingo: a Visual Language Model for Few-Shot LearningNeurIPS2022[pdf]
225Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsNeurIPS2022[pdf]
226SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking-2023[pdf]
227Object-centric Learning with Cyclic Walks between Parts and WholeNeurIPS2023[pdf]
228Characterizing the Impacts of Semi-supervised Learning for Weak SupervisionNeurIPS2023[pdf]
229Brain Decoding: Toward Real-time Reconstruction of Visual Perception-2023[pdf]
230One-step Diffusion with Distribution Matching Distillation-2023[pdf]
231Analyzing and Improving the Training Dynamics of Diffusion Models-2023[pdf]
232Common Diffusion Noise Schedules and Sample Steps are Flawed-2023[pdf]
233On the Importance of Noise Scheduling for Diffusion Models-2023[pdf]
234StyleGAN-T: Unlocking the Power of GANs for Fast Large-Scale Text-to-Image Synthesis-2023[pdf]
deep-learning
machine-learning

Contributors

chaerinkong

199 commits