Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/AIGC-Audio/AudioGPT
/ functions
Functions
2,752 in github.com/AIGC-Audio/AudioGPT
⨍
Functions
2,752
◇
Types & classes
604
↓ 1 callers
Function
ssim
(img1, img2, window_size=11, size_average=True)
NeuralSeq/modules/commons/ssim.py:383
↓ 1 callers
Method
stem
(self, x)
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/model.py:223
↓ 1 callers
Method
stepwise_process
Postprocessing after the whole step-by-step autoregressive decoding
audio_to_text/captioning/models/base_model.py:206
↓ 1 callers
Method
stepwise_process_step
Postprocessing (save output values) after each timestep t
audio_to_text/captioning/models/base_model.py:198
↓ 1 callers
Function
string2symbols
(chinese_string, system)
NeuralSeq/utils/text_norm.py:245
↓ 1 callers
Method
sum_with_attention
(self, embedding, top_k, selected_embeddings)
audio_detection/target_sound_detection/src/models.py:1170
↓ 1 callers
Function
summary_stats
(args)
audio_detection/audio_infer/utils/plot_statistics.py:1614
↓ 1 callers
Function
table_values
(args)
audio_detection/audio_infer/utils/plot_statistics.py:1260
↓ 1 callers
Function
tensors_to_scalars
(metrics)
NeuralSeq/utils/__init__.py:17
↓ 1 callers
Method
test
(self, model)
NeuralSeq/utils/pl_utils.py:584
↓ 1 callers
Method
test_end
(self, outputs)
NeuralSeq/tasks/tts/tts.py:115
↓ 1 callers
Method
test_start
(self)
NeuralSeq/tasks/tts/tts.py:107
↓ 1 callers
Method
toJson
(self)
NeuralSeq/data_gen/tts/data_gen_utils.py:264
↓ 1 callers
Method
to_rgb
(self, x)
text_to_audio/Make_An_Audio/ldm/models/diffusion/ddpm_audio_inpaint.py:1074
↓ 1 callers
Function
train
(args)
audio_detection/audio_infer/pytorch/finetune_template.py:65
↓ 1 callers
Function
train
Train AudioSet tagging model. Args: dataset_dir: str workspace: str data_type: 'balanced_train' | 'full_train' window_si
audio_detection/audio_infer/pytorch/main.py:50
↓ 1 callers
Method
train_forward
(self, input_dict)
audio_to_text/captioning/models/base_model.py:127
↓ 1 callers
Function
train_inception_scorer
(cfg)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/train_melception.py:36
↓ 1 callers
Method
train_process
(self, output, input_dict)
audio_to_text/captioning/models/base_model.py:138
↓ 1 callers
Method
training_end
(self, *args, **kwargs)
NeuralSeq/tasks/base_task.py:274
↓ 1 callers
Method
training_forward
Handle forward for each training case (distributed, single gpu, etc...) :param batch: :param batch_idx: :return:
NeuralSeq/utils/pl_utils.py:1564
↓ 1 callers
Method
translate_english
(self, audio_path)
audio-chatgpt.py:574
↓ 1 callers
Function
trim_long_silences
Ensures that segments without voice in the waveform remain no longer than a threshold determined by the VAD parameters in params.py. :p
NeuralSeq/data_gen/tts/emotion/audio.py:58
↓ 1 callers
Function
tuneThresholdfromScore
(scores, labels, target_fa, target_fr=None)
NeuralSeq/data_gen/tts/emotion/test_emotion.py:32
↓ 1 callers
Method
txt2audio
(self, text, seed = 55, scale = 1.5, ddim_steps = 100, n_samples = 3, W = 624, H = 80)
audio-chatgpt.py:158
↓ 1 callers
Function
unify_energy_torch
(*args)
sound_extraction/utils/create_mixtures.py:69
↓ 1 callers
Function
variance_scaling_
(tensor, scale=1.0, mode='fan_in', distribution='normal')
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/htsat.py:223
↓ 1 callers
Method
vggishish16
(self, pretrained: bool = True)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/lpaps.py:127
↓ 1 callers
Function
whitespace_clean
(text)
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/tokenizer.py:62
↓ 1 callers
Function
window_reverse
Args: windows: (num_windows*B, window_size, window_size, C) window_size (int): Window size H (int): Height of image
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/htsat.py:263
↓ 1 callers
Function
window_sumsquare
# from librosa 0.6 Compute the sum-square envelope of a window function at a given hop length. This is used to estimate modulation effect
sound_extraction/utils/stft.py:10
↓ 1 callers
Method
write_logs
(self, loss, logits, targets)
text_to_audio/Make_An_Audio/ldm/models/diffusion/classifier.py:162
↓ 1 callers
Function
zero_module
Zero out the parameters of a module and return it.
text_to_audio/Make_An_Audio/ldm/modules/attention.py:67
Function
Roberta_embeddings
(text)
text_to_audio/Make_An_Audio/ldm/modules/encoders/open_clap/bert.py:17
Method
__call__
(self, text)
NeuralSeq/data_gen/tts/txt_processors/en.py:15
Method
__call__
(self, audio_list)
audio_to_text/inference_waveform.py:98
Method
__call__
(self, word)
audio_to_text/captioning/utils/build_vocab.py:23
Method
__call__
(self, word)
audio_to_text/captioning/utils/build_vocab_ltp.py:22
Method
__call__
(self, word)
audio_to_text/captioning/utils/build_vocab_spacy.py:22
Method
__call__
(self, x)
audio_to_text/captioning/utils/train_util.py:109
Method
__call__
(self, *args, **kwargs)
audio_detection/audio_infer/utils/crash.py:5
Method
__call__
(self, n, **kwargs)
text_to_audio/Make_An_Audio/ldm/lr_scheduler.py:32
Method
__call__
(self, n, **kwargs)
text_to_audio/Make_An_Audio/ldm/lr_scheduler.py:77
Method
__call__
(self, outputs, targets, to_weight=True)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/loss.py:12
Method
__call__
(self, item)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/transforms.py:25
Method
__call__
(self, item)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/transforms.py:69
Method
__call__
(self, item)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/transforms.py:89
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:28
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:45
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:56
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:67
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:78
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:89
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:99
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:111
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:122
Method
__call__
(self, x)
text_to_audio/Make_An_Audio/ldm/data/extract_mel_spectrogram.py:133
Method
__call__
(self, wav)
text_to_audio/Make_An_Audio/vocoder/bigvgan/models.py:413
Method
__del__
(self)
NeuralSeq/utils/indexed_datasets.py:21
Method
__enter__
(self)
NeuralSeq/utils/__init__.py:231
Method
__exit__
(self, exc_type, exc_val, exc_tb)
NeuralSeq/utils/__init__.py:234
Method
__getitem__
(self, index)
NeuralSeq/modules/GenerSpeech/task/dataset.py:96
Method
__getitem__
Get ndarray for a given key.
NeuralSeq/modules/parallel_wavegan/utils/utils.py:151
Method
__getitem__
(self, i)
NeuralSeq/utils/indexed_datasets.py:25
Method
__getitem__
(self, index)
NeuralSeq/tasks/base_task.py:42
Method
__getitem__
(self, index)
NeuralSeq/tasks/svs/diffsinger_task.py:103
Method
__getitem__
(self, index)
NeuralSeq/tasks/svs/diffsinger_task.py:255
Method
__getitem__
(self, index)
NeuralSeq/tasks/tts/fs2_utils.py:60
Method
__getitem__
(self, index)
NeuralSeq/tasks/tts/dataset_utils.py:137
Method
__getitem__
(self, index)
NeuralSeq/tasks/tts/dataset_utils.py:239
Method
__getitem__
(self, index)
NeuralSeq/tasks/tts/pe.py:48
Method
__getitem__
(self, index)
NeuralSeq/tasks/vocoder/dataset_utils.py:80
Method
__getitem__
(self, word_id)
audio_to_text/captioning/utils/build_vocab.py:28
Method
__getitem__
Load waveform and target of an audio clip. Args: meta: { 'hdf5_path': str, 'index_in_hdf5': int}
audio_detection/audio_infer/utils/data_generator.py:28
Method
__getitem__
(self, idx)
text_to_audio/Make_An_Audio/ldm/modules/losses_audio/vggishish/dataset.py:47
Method
__init__
(self, device)
audio-chatgpt.py:105
Method
__init__
(self, device)
audio-chatgpt.py:127
Method
__init__
(self, device)
audio-chatgpt.py:141
Method
__init__
(self, device)
audio-chatgpt.py:215
Method
__init__
(self, device=None)
audio-chatgpt.py:276
Method
__init__
(self, device= None)
audio-chatgpt.py:299
Method
__init__
(self, device=None)
audio-chatgpt.py:342
Method
__init__
(self, device)
audio-chatgpt.py:384
Method
__init__
(self, device)
audio-chatgpt.py:419
Method
__init__
(self, device)
audio-chatgpt.py:561
Method
__init__
(self, device)
audio-chatgpt.py:579
Method
__init__
(self, device=None)
audio-chatgpt.py:590
Method
__init__
(self, device)
audio-chatgpt.py:613
Method
__init__
(self, device)
audio-chatgpt.py:676
Method
__init__
(self, device)
audio-chatgpt.py:714
Method
__init__
(self, device)
audio-chatgpt.py:776
Method
__init__
(self, device="cuda", model_name="espnet/Wangyou_Zhang_chime4_enh_train_enh_conv_tasnet_raw")
audio-chatgpt.py:963
Method
__init__
(self, device="cuda", model_name="lichenda/wsj0_2mix_skim_noncausal")
audio-chatgpt.py:1010
Method
__init__
(self)
audio-chatgpt.py:1052
Method
__init__
(self, sampling_rate=48000)
mono2binaural/src/warping.py:93
Method
__init__
(self, model_name="network", use_cuda=True)
mono2binaural/src/utils.py:16
Method
__init__
(self, sampling_rate=48000)
mono2binaural/src/models.py:12
Method
__init__
(self, view_dim=7, warpnet_layers=4, warpnet_channels=64,
mono2binaural/src/models.py:87
Method
__init__
(self)
NeuralSeq/modules/GenerSpeech/task/generspeech.py:26
Method
__init__
(self, prefix, shuffle=False, test_items=None, test_sizes=None, data_dir=None)
NeuralSeq/modules/GenerSpeech/task/dataset.py:31
← previous
next →
1,101–1,200 of 2,752, ranked by callers