Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

>As an example (28:57), he described how the human brain does not have any innate convolutional structure – but it doesn’t need to, because as an effective unsupervised learner, the brain can learn the same low-level image features (e.g. oriented edge detectors) as a ConvNet, even without the convolutional weight-sharing constraint.

I think the first part of the sentence should be that the brain doesn't have an innate weight sharing (like is stated in the end of the sentence), not that it is not convolutional. I believe the convolutional structure is actually copied from the visual cortex (but with no weight sharing as far as we know)



Convolutions are mathematically defined by application of the same kernel at each possible position. If that kernel has finite support, you also get locality. The visual cortex has locality, but without weight-sharing between functionally identical neurons, it's not convolutional.


The visual cortex neurons have not only locality, but also they learn/recognize similar features. This was the original inspiration for the conv nets: it was a way to learn efficient local features and apply them to the whole image. To me it seems that cortex neurons and the initial layers of conv nets do similar work, but with constraints of their implementations: biological neurons cannot share weights, and for artificial neurons it is more efficient to learn and compute dense convolution.


Disclaimer: I'm learning deep learning mainly from HN comments and just want to provoke more insights. I have no idea what weight sharing is or how kernels are represented in networks, but I do know that e.g. a blur filter or edge filter is represented as convolution matrix.

It is dangerously confusing to reapply the neural-net-terminology to neuronal-nets isn't it? The weight of a kernel of biological neurons, what is that supposed to mean?

If you haven't stopped reading yet, please consider: in case, as I have to assume, you mean there is a specific ensemble of neurons that represents a kernel of given weights corresponding to exactly one area of retina, then isn't sharing between "pixels" achieved simply by the eye's jittering?

For better or worse, assume I'm the adversary in a GAN and ignore me if it doesn't make sense.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: