MCPcopy Create free account

hub / github.com/bytedance/LatentSync / types & classes

Types & classes103 in github.com/bytedance/LatentSync

↓ 37 callersClassConv2d
latentsync/models/wav2lip_syncnet.py:71
↓ 7 callersClassInflatedConv3d
latentsync/models/resnet.py:10
↓ 6 callersClassLinear
latentsync/whisper/whisper/model.py:34
↓ 6 callersClassResnetBlock3D
latentsync/models/resnet.py:104
↓ 5 callersClassLayerNorm
latentsync/whisper/whisper/model.py:29
↓ 5 callersClassPretrainVisionTransformer
Vision Transformer with support for patch or hybrid CNN input stage
latentsync/trepa/third_party/VideoMAEv2/videomaev2_pretrain.py:238
↓ 5 callersClassTargetFC
Fully connection operations for target net Note: Weights & biases are different for different images in a batch, thus here w
eval/hyper_iqa.py:158
↓ 4 callersClassImageProcessor
latentsync/utils/image_processor.py:34
↓ 3 callersClassAttention
latentsync/models/attention.py:202
↓ 3 callersClassAudio2Feature
latentsync/whisper/audio2feature.py:10
↓ 3 callersClassBlock
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:261
↓ 3 callersClassInflatedGroupNorm
latentsync/models/resnet.py:21
↓ 3 callersClassL2Norm
eval/detectors/s3fd/nets.py:8
↓ 3 callersClassStableSyncNet
latentsync/models/stable_syncnet.py:28
↓ 3 callersClassSyncNetDataset
latentsync/data/syncnet_dataset.py:29
↓ 3 callersClassSyncNetDetector
eval/syncnet_detect.py:20
↓ 3 callersClassSyncNetEval
eval/syncnet/syncnet_eval.py:40
↓ 3 callersClassTransformer3DModel
latentsync/models/attention.py:23
↓ 2 callersClassConv1d
latentsync/whisper/whisper/model.py:41
↓ 2 callersClassDecodingResult
latentsync/whisper/whisper/decoding.py:104
↓ 2 callersClassDownEncoder2D
latentsync/models/stable_syncnet.py:172
↓ 2 callersClassDownsample3D
latentsync/models/resnet.py:78
↓ 2 callersClassLipsyncPipeline
latentsync/pipelines/lipsync_pipeline.py:43
↓ 2 callersClassMultiHeadAttention
latentsync/whisper/whisper/model.py:57
↓ 2 callersClassPatchEmbed
Image to Patch Embedding
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:323
↓ 2 callersClassResidualAttentionBlock
latentsync/whisper/whisper/model.py:103
↓ 2 callersClassTREPALoss
latentsync/trepa/loss.py:22
↓ 2 callersClassUpsample3D
latentsync/models/resnet.py:32
↓ 2 callersClassVideoProcessor
latentsync/utils/image_processor.py:103
↓ 1 callersClassAVReader
Individual audio video reader with convenient indexing function. Parameters ---------- uri: str Path of file. ctx: decord.Con
latentsync/utils/av_reader.py:13
↓ 1 callersClassAlignRestore
latentsync/utils/affine_transform.py:10
↓ 1 callersClassApplyTimestampRules
latentsync/whisper/whisper/decoding.py:405
↓ 1 callersClassAttention
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:210
↓ 1 callersClassAttentionBlock2D
latentsync/models/stable_syncnet.py:136
↓ 1 callersClassAudioEncoder
latentsync/whisper/whisper/model.py:131
↓ 1 callersClassBasicTransformerBlock
latentsync/models/attention.py:127
↓ 1 callersClassBeamSearchDecoder
latentsync/whisper/whisper/decoding.py:281
↓ 1 callersClassChart
eval/draw_syncnet_lines.py:19
↓ 1 callersClassCosAttention
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:154
↓ 1 callersClassCrossAttnDownBlock3D
latentsync/models/unet_blocks.py:263
↓ 1 callersClassCrossAttnUpBlock3D
latentsync/models/unet_blocks.py:519
↓ 1 callersClassDecodingOptions
latentsync/whisper/whisper/decoding.py:72
↓ 1 callersClassDecodingTask
latentsync/whisper/whisper/decoding.py:444
↓ 1 callersClassDetect
eval/detectors/s3fd/box_utils.py:133
↓ 1 callersClassDownBlock3D
latentsync/models/unet_blocks.py:410
↓ 1 callersClassDropPath
Drop paths (Stochastic Depth) per sample (when applied in main path of residual blocks).
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:119
↓ 1 callersClassEnglishNumberNormalizer
Convert any spelled-out numbers into arabic numbers, while handling: - remove any commas - keep the suffixes such as: `1960s`, `274th`,
latentsync/whisper/whisper/normalizers/english.py:12
↓ 1 callersClassEnglishSpellingNormalizer
Applies British-American spelling mappings as listed in [1]. [1] https://www.tysto.com/uk-us-spelling-list.html
latentsync/whisper/whisper/normalizers/english.py:443
↓ 1 callersClassFVD
eval/eval_fvd.py:25
↓ 1 callersClassFaceDetector
latentsync/utils/face_detector.py:8
↓ 1 callersClassFaceDetector
preprocess/filter_high_resolution.py:37
↓ 1 callersClassFaceDetector
preprocess/remove_incorrect_affined.py:22
↓ 1 callersClassFeatureStats
Class to store statistics of features, including all features and mean/covariance. Args: capture_all: Whether to store all the featu
latentsync/trepa/utils/metric_utils.py:18
↓ 1 callersClassFileslistWriter
tools/write_fileslist.py:19
↓ 1 callersClassGreedyDecoder
latentsync/whisper/whisper/decoding.py:253
↓ 1 callersClassHyperNet
Hyper network for learning perceptual rules. Args: lda_out_channels: local distortion aware module output size. hyper_in_cha
eval/hyper_iqa.py:19
↓ 1 callersClassMaximumLikelihoodRanker
Select the sample with the highest log probabilities, penalized using either a simple length normalization or Google NMT paper's length penal
latentsync/whisper/whisper/decoding.py:173
↓ 1 callersClassMish
latentsync/models/resnet.py:226
↓ 1 callersClassMlp
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:133
↓ 1 callersClassModelDimensions
latentsync/whisper/whisper/model.py:16
↓ 1 callersClassPositionalEncoding
latentsync/models/motion_module.py:221
↓ 1 callersClassPretrainVisionTransformerDecoder
Vision Transformer with support for patch or hybrid CNN input stage
latentsync/trepa/third_party/VideoMAEv2/videomaev2_pretrain.py:146
↓ 1 callersClassPretrainVisionTransformerEncoder
Vision Transformer with support for patch or hybrid CNN input stage
latentsync/trepa/third_party/VideoMAEv2/videomaev2_pretrain.py:27
↓ 1 callersClassPriorBox
eval/detectors/s3fd/box_utils.py:180
↓ 1 callersClassPyTorchInference
latentsync/whisper/whisper/decoding.py:132
↓ 1 callersClassResNetBackbone
eval/hyper_iqa.py:220
↓ 1 callersClassResize
latentsync/trepa/third_party/VideoMAEv2/utils.py:31
↓ 1 callersClassResnetBlock2D
latentsync/models/stable_syncnet.py:65
↓ 1 callersClassS
eval/syncnet/syncnet.py:18
↓ 1 callersClassS3FD
eval/detectors/s3fd/__init__.py:14
↓ 1 callersClassS3FDNet
eval/detectors/s3fd/nets.py:28
↓ 1 callersClassSuppressBlank
latentsync/whisper/whisper/decoding.py:387
↓ 1 callersClassSuppressTokens
latentsync/whisper/whisper/decoding.py:397
↓ 1 callersClassTargetNet
Target network for quality prediction.
eval/hyper_iqa.py:123
↓ 1 callersClassTemporalTransformer3DModel
latentsync/models/motion_module.py:76
↓ 1 callersClassTemporalTransformerBlock
latentsync/models/motion_module.py:154
↓ 1 callersClassTextDecoder
latentsync/whisper/whisper/model.py:174
↓ 1 callersClassToFloatTensorInZeroOne
latentsync/trepa/third_party/VideoMAEv2/utils.py:26
↓ 1 callersClassTokenizer
A thin wrapper around `GPT2TokenizerFast` providing quick access to special tokens
latentsync/whisper/whisper/tokenizer.py:130
↓ 1 callersClassTransformer3DModelOutput
latentsync/models/attention.py:19
↓ 1 callersClassUNet3DConditionOutput
latentsync/models/unet.py:35
↓ 1 callersClassUNetDataset
latentsync/data/unet_dataset.py:29
↓ 1 callersClassUNetMidBlock3DCrossAttn
latentsync/models/unet_blocks.py:153
↓ 1 callersClassUpBlock3D
latentsync/models/unet_blocks.py:669
↓ 1 callersClassVanillaTemporalModule
latentsync/models/motion_module.py:39
↓ 1 callersClassVersatileAttention
latentsync/models/motion_module.py:237
↓ 1 callersClassVideoData
Class to create dataloaders for video datasets Args: data_path: Path to the folder with video frames or videos. image_folder: I
latentsync/trepa/utils/data_utils.py:77
↓ 1 callersClassVisionTransformer
Vision Transformer with support for patch or hybrid CNN input stage
latentsync/trepa/third_party/VideoMAEv2/videomaev2_finetune.py:371
↓ 1 callersClassWhisper
latentsync/whisper/whisper/model.py:220
↓ 1 callersClassdummy_context
latentsync/utils/util.py:284
ClassBasicTextNormalizer
latentsync/whisper/whisper/normalizers/basic.py:55
ClassBottleneck
eval/hyper_iqa.py:181
ClassEnglishTextNormalizer
latentsync/whisper/whisper/normalizers/english.py:458
ClassFrameDataset
Generic dataset for videos stored as images. The loading will iterates over all the folders and subfolders in the provided directory. Ea
latentsync/trepa/utils/data_utils.py:211
ClassInference
latentsync/whisper/whisper/decoding.py:118
ClassLogitFilter
latentsync/whisper/whisper/decoding.py:371
ClassPredictor
predict.py:21
ClassSequenceRanker
latentsync/whisper/whisper/decoding.py:164
ClassTemporalTransformer3DModelOutput
latentsync/models/motion_module.py:25
ClassTokenDecoder
latentsync/whisper/whisper/decoding.py:199
next →1–100 of 103, ranked by callers