Extract activation from lower audio layers #90

seunggookim · 2022-06-15T14:15:12Z

Hi, I was wondering how I can extract activations from the lower audio layers. I guess "embeddings" are the same as "MaxPool_3"? and if that's correct, then "MaxPool", "MaxPool_1", and "MaxPool_2" corresponds to the first, second, and third max-pooling layers in the Audio ConvNet as explained in Arandjelovic and Zisserman 2018 (https://arxiv.org/abs/1712.06651)?

auroracramer · 2022-07-01T20:52:53Z

That sounds right! You should be able to construct a new model stopping at the appropriate level and selectively load the model weights up to that layer from the existing ones. If you wanted, you could also swap out the output max pooling layer if you wanted to play around with further downsampling the embedding.

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Extract activation from lower audio layers #90

Extract activation from lower audio layers #90

seunggookim commented Jun 15, 2022

auroracramer commented Jul 1, 2022

Extract activation from lower audio layers #90

Extract activation from lower audio layers #90

Comments

seunggookim commented Jun 15, 2022

auroracramer commented Jul 1, 2022