MCPcopy Create free account

hub / github.com/AIGC-Audio/AudioGPT / functions

Functions2,752 in github.com/AIGC-Audio/AudioGPT

↓ 1 callersFunctionssim
(img1, img2, window_size=11, size_average=True)
NeuralSeq/modules/commons/ssim.py:383
↓ 1 callersMethodstem
(self, x)
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/model.py:223
↓ 1 callersMethodstepwise_process
Postprocessing after the whole step-by-step autoregressive decoding
audio_to_text/captioning/models/base_model.py:206
↓ 1 callersMethodstepwise_process_step
Postprocessing (save output values) after each timestep t
audio_to_text/captioning/models/base_model.py:198
↓ 1 callersFunctionstring2symbols
(chinese_string, system)
NeuralSeq/utils/text_norm.py:245
↓ 1 callersMethodsum_with_attention
(self, embedding, top_k, selected_embeddings)
audio_detection/target_sound_detection/src/models.py:1170
↓ 1 callersFunctionsummary_stats
(args)
audio_detection/audio_infer/utils/plot_statistics.py:1614
↓ 1 callersFunctiontable_values
(args)
audio_detection/audio_infer/utils/plot_statistics.py:1260
↓ 1 callersFunctiontensors_to_scalars
(metrics)
NeuralSeq/utils/__init__.py:17
↓ 1 callersMethodtest
(self, model)
NeuralSeq/utils/pl_utils.py:584
↓ 1 callersMethodtest_end
(self, outputs)
NeuralSeq/tasks/tts/tts.py:115
↓ 1 callersMethodtest_start
(self)
NeuralSeq/tasks/tts/tts.py:107
↓ 1 callersMethodtoJson
(self)
NeuralSeq/data_gen/tts/data_gen_utils.py:264
↓ 1 callersMethodto_rgb
(self, x)
text_to_audio/Make_An_Audio/ldm/models/diffusion/ddpm_audio_inpaint.py:1074
↓ 1 callersFunctiontrain
(args)
audio_detection/audio_infer/pytorch/finetune_template.py:65
↓ 1 callersFunctiontrain
Train AudioSet tagging model. Args: dataset_dir: str workspace: str data_type: 'balanced_train' | 'full_train' window_si
audio_detection/audio_infer/pytorch/main.py:50
↓ 1 callersMethodtrain_forward
(self, input_dict)
audio_to_text/captioning/models/base_model.py:127
↓ 1 callersFunctiontrain_inception_scorer
(cfg)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/train_melception.py:36
↓ 1 callersMethodtrain_process
(self, output, input_dict)
audio_to_text/captioning/models/base_model.py:138
↓ 1 callersMethodtraining_end
(self, *args, **kwargs)
NeuralSeq/tasks/base_task.py:274
↓ 1 callersMethodtraining_forward
Handle forward for each training case (distributed, single gpu, etc...) :param batch: :param batch_idx: :return:
NeuralSeq/utils/pl_utils.py:1564
↓ 1 callersMethodtranslate_english
(self, audio_path)
audio-chatgpt.py:574
↓ 1 callersFunctiontrim_long_silences
Ensures that segments without voice in the waveform remain no longer than a threshold determined by the VAD parameters in params.py. :p
NeuralSeq/data_gen/tts/emotion/audio.py:58
↓ 1 callersFunctiontuneThresholdfromScore
(scores, labels, target_fa, target_fr=None)
NeuralSeq/data_gen/tts/emotion/test_emotion.py:32
↓ 1 callersMethodtxt2audio
(self, text, seed = 55, scale = 1.5, ddim_steps = 100, n_samples = 3, W = 624, H = 80)
audio-chatgpt.py:158
↓ 1 callersFunctionunify_energy_torch
(*args)
sound_extraction/utils/create_mixtures.py:69
↓ 1 callersFunctionvariance_scaling_
(tensor, scale=1.0, mode='fan_in', distribution='normal')
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/htsat.py:223
↓ 1 callersMethodvggishish16
(self, pretrained: bool = True)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/lpaps.py:127
↓ 1 callersFunctionwhitespace_clean
(text)
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/tokenizer.py:62
↓ 1 callersFunctionwindow_reverse
Args: windows: (num_windows*B, window_size, window_size, C) window_size (int): Window size H (int): Height of image
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/htsat.py:263
↓ 1 callersFunctionwindow_sumsquare
# from librosa 0.6 Compute the sum-square envelope of a window function at a given hop length. This is used to estimate modulation effect
sound_extraction/utils/stft.py:10
↓ 1 callersMethodwrite_logs
(self, loss, logits, targets)
text_to_audio/Make_An_Audio/ldm/models/diffusion/classifier.py:162
↓ 1 callersFunctionzero_module
Zero out the parameters of a module and return it.
text_to_audio/Make_An_Audio/ldm/modules/attention.py:67
FunctionRoberta_embeddings
(text)
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/bert.py:17
Method__call__
(self, text)
NeuralSeq/data_gen/tts/txt_processors/en.py:15
Method__call__
(self, audio_list)
audio_to_text/inference_waveform.py:98
Method__call__
(self, word)
audio_to_text/captioning/utils/build_vocab.py:23
Method__call__
(self, word)
audio_to_text/captioning/utils/build_vocab_ltp.py:22
Method__call__
(self, word)
audio_to_text/captioning/utils/build_vocab_spacy.py:22
Method__call__
(self, x)
audio_to_text/captioning/utils/train_util.py:109
Method__call__
(self, *args, **kwargs)
audio_detection/audio_infer/utils/crash.py:5
Method__call__
(self, n, **kwargs)
text_to_audio/Make_An_Audio/ldm/lr_scheduler.py:32
Method__call__
(self, n, **kwargs)
text_to_audio/Make_An_Audio/ldm/lr_scheduler.py:77
Method__call__
(self, outputs, targets, to_weight=True)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/loss.py:12
Method__call__
(self, item)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/transforms.py:25
Method__call__
(self, item)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/transforms.py:69
Method__call__
(self, item)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/transforms.py:89
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:28
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:45
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:56
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:67
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:78
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:89
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:99
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:111
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:122
Method__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:133
Method__call__
(self, wav)
text_to_audio/Make_An_Audio/vocoder/bigvgan/models.py:413
Method__del__
(self)
NeuralSeq/utils/indexed_datasets.py:21
Method__enter__
(self)
NeuralSeq/utils/__init__.py:231
Method__exit__
(self, exc_type, exc_val, exc_tb)
NeuralSeq/utils/__init__.py:234
Method__getitem__
(self, index)
NeuralSeq/modules/GenerSpeech/task/dataset.py:96
Method__getitem__
Get ndarray for a given key.
NeuralSeq/modules/parallel_wavegan/utils/utils.py:151
Method__getitem__
(self, i)
NeuralSeq/utils/indexed_datasets.py:25
Method__getitem__
(self, index)
NeuralSeq/tasks/base_task.py:42
Method__getitem__
(self, index)
NeuralSeq/tasks/svs/diffsinger_task.py:103
Method__getitem__
(self, index)
NeuralSeq/tasks/svs/diffsinger_task.py:255
Method__getitem__
(self, index)
NeuralSeq/tasks/tts/fs2_utils.py:60
Method__getitem__
(self, index)
NeuralSeq/tasks/tts/dataset_utils.py:137
Method__getitem__
(self, index)
NeuralSeq/tasks/tts/dataset_utils.py:239
Method__getitem__
(self, index)
NeuralSeq/tasks/tts/pe.py:48
Method__getitem__
(self, index)
NeuralSeq/tasks/vocoder/dataset_utils.py:80
Method__getitem__
(self, word_id)
audio_to_text/captioning/utils/build_vocab.py:28
Method__getitem__
Load waveform and target of an audio clip. Args: meta: { 'hdf5_path': str, 'index_in_hdf5': int}
audio_detection/audio_infer/utils/data_generator.py:28
Method__getitem__
(self, idx)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/dataset.py:47
Method__init__
(self, device)
audio-chatgpt.py:105
Method__init__
(self, device)
audio-chatgpt.py:127
Method__init__
(self, device)
audio-chatgpt.py:141
Method__init__
(self, device)
audio-chatgpt.py:215
Method__init__
(self, device=None)
audio-chatgpt.py:276
Method__init__
(self, device= None)
audio-chatgpt.py:299
Method__init__
(self, device=None)
audio-chatgpt.py:342
Method__init__
(self, device)
audio-chatgpt.py:384
Method__init__
(self, device)
audio-chatgpt.py:419
Method__init__
(self, device)
audio-chatgpt.py:561
Method__init__
(self, device)
audio-chatgpt.py:579
Method__init__
(self, device=None)
audio-chatgpt.py:590
Method__init__
(self, device)
audio-chatgpt.py:613
Method__init__
(self, device)
audio-chatgpt.py:676
Method__init__
(self, device)
audio-chatgpt.py:714
Method__init__
(self, device)
audio-chatgpt.py:776
Method__init__
(self, device="cuda", model_name="espnet/Wangyou_Zhang_chime4_enh_train_enh_conv_tasnet_raw")
audio-chatgpt.py:963
Method__init__
(self, device="cuda", model_name="lichenda/wsj0_2mix_skim_noncausal")
audio-chatgpt.py:1010
Method__init__
(self)
audio-chatgpt.py:1052
Method__init__
(self, sampling_rate=48000)
mono2binaural/src/warping.py:93
Method__init__
(self, model_name="network", use_cuda=True)
mono2binaural/src/utils.py:16
Method__init__
(self, sampling_rate=48000)
mono2binaural/src/models.py:12
Method__init__
(self, view_dim=7, warpnet_layers=4, warpnet_channels=64,
mono2binaural/src/models.py:87
Method__init__
(self)
NeuralSeq/modules/GenerSpeech/task/generspeech.py:26
Method__init__
(self, prefix, shuffle=False, test_items=None, test_sizes=None, data_dir=None)
NeuralSeq/modules/GenerSpeech/task/dataset.py:31
← previousnext →1,101–1,200 of 2,752, ranked by callers