Generation is Required for Data-Efficient Perception
arxiv.org·1d
🧠Learned Codecs
Preview
Report Post

Title:Generation is Required for Data-Efficient Perception

View PDF HTML (experimental)

Abstract:It has been hypothesized that human-level visual perception requires a generative approach in which internal representations result from inverting a decoder. Yet today’s most successful vision models are non-generative, relying on an encoder that maps images to representations without decoder inversion. This raises the question of whether generation is, in fact, necessary for machines to achieve human-level visual perception. To address this, we study whether generative and non-generative methods can achieve compositional generalization, a hallmark of human perception. Under a compositional data generating p…

Similar Posts

Loading similar posts...