Category Research Paper

Research Paper

173. BiSeNet

Background Most of the previous semantic segmentation model’s architecture can be categorized into 2 types. Encoder-Decoder Backbone: (Ex. FCN, UNet) This architecture requires all information to flow through the deep encoding-decoding structure leading to high latency, also suffering in restoring…

Kyosuke
August 3, 2022

Computer Vision, Research Paper

162. Residual Blocks

Why are residual blocks called “residual” blocks? The reason why I was confused was that the equation in the diagram explaining the residual blocks on the research paper was f(x) + x. So I thought, “Where is the residual..?” When…

Kyosuke
July 23, 2022

AI, Research Paper

161. ESRGAN

Abstract Even though SR-GAN was able to make a huge improvement, there was still a gap between the generated image and the ground truth image. The proposed ESR-GAN further enhances the performance. Three Key Modification Components Network Remove all batch…

Kyosuke
July 22, 2022

Research Paper

160. DeepPose

DeepPose DeepPose is a research done by Google for human pose estimation. Pose Vector First, the paper encodes all “k” body joints into a pose vector. To avoid using absolute coordinates for the body joints like right now, the paper…

Kyosuke
July 21, 2022

Research Paper

159. M-RNN

M-RNN Multi-Modal Recurrent Neural Network is a research done by The University of California and the Baidu Research Team which generates captions for images. In this research,Deep Recurrent Neural Network is used for sentences, and Deep Convolutional Neural Network is…

Kyosuke
July 20, 2022

Research Paper

158: SR-GAN

SR-GAN Today I learned about SR(Super Resolution)-GAN, so I’d like to share it here. In previous research, super-resolution tasks(Enhancing resolution) struggled when recovering finer text details at large upscale factors. SR-GAN is the first to be able to infer images…

Kyosuke
July 19, 2022

Research Paper

157. CycleGAN

CycleGAN Today I’ve learned about CycleGAN, so I’d like it here. Before this research, Image-to-Image translation tasks(Learning how to map an input image to a different style image) required “PAIR” data sets for training. Unfortunately, in most cases, you don’t…

Kyosuke
July 18, 2022

Research Paper

52. Implementing InceptionNet From Thesis

■ InceptionNet I’ve been using this model to run inference on jetson for a while, but I didn’t know what was actually going on inside, so for this time I’ve decided to create the network from scratch using Pytorch. ■…

Kyosuke
March 4, 2022