/home/tony/Work/glockenspiel/s3prl/s3prl/upstream/byol_s/byol_a/common.py:20: UserWarning: torchaudio._backend.set_audio_backend has been deprecated. With dispatcher enabled, this function is no-op. You can remove the function call. torchaudio.set_audio_backend("sox_io") /home/tony/Work/glockenspiel/s3prl/s3prl/run_downstream.py:157: UserWarning: torchaudio._backend.set_audio_backend has been deprecated. With dispatcher enabled, this function is no-op. You can remove the function call. torchaudio.set_audio_backend('sox_io') 2023-10-14 22:17:49 | INFO | s3prl.upstream.hubert.hubconf | Converting a fairseq checkpoint: /home/tony/Data/MERT/hubert_d2v2_200k.pt 2023-10-14 22:17:49 | INFO | s3prl.upstream.hubert.hubconf | To: /home/tony/Data/MERT/hubert_d2v2_200k.converted.pt 2023-10-14 22:17:56 | INFO | fairseq.models.hubert.hubert | HubertModel Config: HubertConfig(_name='hubert', label_rate=50.0, extractor_mode='default', encoder_layers=12, encoder_embed_dim=768, encoder_ffn_embed_dim=3072, encoder_attention_heads=12, activation_fn='gelu', layer_type='transformer', dropout=0.1, attention_dropout=0.1, activation_dropout=0.0, encoder_layerdrop=0.05, dropout_input=0.1, dropout_features=0.1, final_dim=256, untie_final_proj=True, layer_norm_first=False, conv_feature_layers='[(512,10,5)] + [(512,3,2)] * 4 + [(512,2,2)] * 3', conv_bias=False, logit_temp=0.1, target_glu=False, feature_grad_mult=0.1, mask_length=10, mask_prob=0.8, mask_selection='static', mask_other=0.0, no_mask_overlap=False, mask_min_space=1, mask_channel_length=10, mask_channel_prob=0.0, mask_channel_selection='static', mask_channel_other=0.0, no_mask_channel_overlap=False, mask_channel_min_space=1, conv_pos=128, conv_pos_groups=16, conv_pos_batch_norm=False, latent_temp=[2.0, 0.5, 0.999995], skip_masked=False, skip_nomask=False, checkpoint_activations=False, required_seq_len_multiple=2, depthwise_conv_kernel_size=31, attn_type='', pos_enc_type='abs', fp16=False) /home/tony/anaconda3/envs/suno_env/lib/python3.10/site-packages/torch/nn/utils/weight_norm.py:30: UserWarning: torch.nn.utils.weight_norm is deprecated in favor of torch.nn.utils.parametrizations.weight_norm. warnings.warn("torch.nn.utils.weight_norm is deprecated in favor of torch.nn.utils.parametrizations.weight_norm.") 2023-10-14 22:17:58 | INFO | fairseq.models.hubert.hubert | cannot find dictionary. assume will be used for fine-tuning [Featurizer] - Take a list of 13 features and weighted sum them. [Featurizer] - The selected feature hidden_states's downsample rate is 640 [Runner] - Start a new experiment [(512, 10, 5), (512, 3, 2), (512, 3, 2), (512, 3, 2), (512, 3, 2), (512, 2, 2), (512, 2, 2), (512, 2, 2)] 640 overall: 0%| | 0/4000 [00:00