Knowledge Transfer Graph for Deep Collaborative Learning

Soma  Minami (Chubu university)*; Tsubasa Hirakawa (Chubu University); Takayoshi  Yamashita (Chubu University); Hironobu Fujiyoshi (Chubu University)

Knowledge Transfer Graph for Deep Collaborative Learning

Soma Minami (Chubu university)*, Tsubasa Hirakawa (Chubu University), Takayoshi Yamashita (Chubu University), Hironobu Fujiyoshi (Chubu University)

Keywords: Deep Learning for Computer Vision

Abstract: Knowledge transfer among multiple networks using their outputs or intermediate activations have evolved through manual design from a simple teacher-student approach to a bidirectional cohort one. The major components of such knowledge transfer framework involve the network size, the number of networks, the transfer direction, and the design of the loss function. However, because these factors are enormous when combined and become intricately entangled, the methods of conventional knowledge transfer have explored only limited combinations. In this paper, we propose a novel graph representation called knowledge transfer graph that provides a unified view of the knowledge transfer and has the potential to represent diverse knowledge transfer patterns. We also propose four gate functions that control the gradient and can deliver diverse combinations of knowledge transfer. Searching the graph structure enables us to discover more effective knowledge transfer methods than a manually designed one. Experimental results show that the proposed method achieved performance improvements.

Knowledge Transfer Graph for Deep Collaborative Learning

Soma Minami (Chubu university)*, Tsubasa Hirakawa (Chubu University), Takayoshi Yamashita (Chubu University), Hironobu Fujiyoshi (Chubu University)

SlidesLive

Similar Papers

Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network

Yuxuan Wu (The University of Tokyo)*, Hideki Nakayama (The University of Tokyo)

Learn more, forget less: Cues from human brain

Arijit Patra (University of Oxford), Tapabrata Chakraborti (University of Oxford)*

Large-Scale Cross-Domain Few-Shot Learning

Jiechao Guan (Renmin University of China), Manli Zhang (Renmin University of China), Zhiwu Lu (Renmin University of China)*