Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/CompVis/zigma
/ types & classes
Types & classes
98 in github.com/CompVis/zigma
⨍
Functions
471
◇
Types & classes
98
↓ 15 callers
Class
ZigMa
A DiT-styled Mamba model with ZigZag scan.
model_zigma.py:544
↓ 11 callers
Class
Unit3D
video_metrics/fvd/videogpt/pytorch_i3d.py:37
↓ 10 callers
Class
WebDataModuleFromConfig
datasets/wds_dataloader.py:46
↓ 9 callers
Class
InceptionModule
video_metrics/fvd/videogpt/pytorch_i3d.py:107
↓ 5 callers
Class
MaxPool3dSamePadding
video_metrics/fvd/videogpt/pytorch_i3d.py:7
↓ 4 callers
Class
FrechetVideoDistance
r"""Calculate Fréchet inception distance (FID_) which is used to access the quality of generated images. .. math:: FID = \|\mu - \mu_w\|^2
utils/torchmetric_fvd.py:211
↓ 4 callers
Class
MyMetric
my_metrics.py:13
↓ 4 callers
Class
NoTrainInceptionV3
Module that never leaves evaluation mode.
utils/torchmetric_sfid.py:63
↓ 2 callers
Class
DropPath
Drop paths (Stochastic Depth) per sample (when applied in main path of residual blocks).
model_zigma.py:162
↓ 2 callers
Class
FrechetDinovDistance
r"""Calculate Fréchet inception distance (FID_) which is used to access the quality of generated images. .. math:: FID = \|\mu - \mu_w\|^
utils/torchmetric_fdd.py:131
↓ 2 callers
Class
InferenceParams
Inference parameters that are passed to the main model in order to efficienly calculate and store the context during inference.
dis_mamba/mamba_ssm/utils/generation.py:18
↓ 2 callers
Class
Sampler
Sampler class for the transport model
transport/transport.py:236
↓ 2 callers
Class
ode
ODE solver class
transport/integrators.py:83
↓ 2 callers
Class
sFrechetInceptionDistance
r"""Calculate Fréchet inception distance (FID_) which is used to access the quality of generated images. .. math:: FID = \|\mu - \mu_w\|^
utils/torchmetric_sfid.py:202
↓ 1 callers
Class
Block
model_zigma.py:340
↓ 1 callers
Class
Block
dis_mamba/mamba_ssm/modules/mamba_simple.py:611
↓ 1 callers
Class
CrossAttention
model_zigma.py:95
↓ 1 callers
Class
DINOv2Encoder
utils/torchmetric_fdd.py:82
↓ 1 callers
Class
DecodingCGCache
dis_mamba/mamba_ssm/utils/generation.py:243
↓ 1 callers
Class
FinalLayer
The final layer of DiT.
model_zigma.py:313
↓ 1 callers
Class
InceptionI3d
Inception-v1 I3D architecture. The model is introduced in: Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset Joa
video_metrics/fvd/videogpt/pytorch_i3d.py:135
↓ 1 callers
Class
InceptionScore
r"""Calculate the Inception Score (IS) which is used to access how realistic generated images are. .. math:: IS = exp(\mathbb{E}_x KL(p(y
utils/torchmetric_inception.py:34
↓ 1 callers
Class
KernelInceptionDistance
r""" Calculates Kernel Inception Distance (KID) which is used to access the quality of generated images. Given by .. math:: KID = MMD
utils/torchmetric_kid.py:67
↓ 1 callers
Class
LabelEmbedder
Embeds class labels into vector representations. Also handles label dropout for classifier-free guidance.
model_zigma.py:278
↓ 1 callers
Class
Mamba
dis_mamba/mamba_ssm/modules/mamba_simple.py:64
↓ 1 callers
Class
MixerModel
dis_mamba/mamba_ssm/models/mixer_seq_simple.py:83
↓ 1 callers
Class
PRDC
Example: >>> import torch >>> _ = torch.manual_seed(123) >>> from torchmetric_prdc import PRDC >>> prdc = PRDC(ne
utils/torchmetric_prdc.py:74
↓ 1 callers
Class
PatchEmbed_Video
2D Image to Patch Embedding
model_zigma.py:66
↓ 1 callers
Class
RandomHorizontalFlipVideo
Flip the video clip along the horizontal direction with a given probability Args: p (float): probability of the clip being flipped. D
datasets/video_utils.py:425
↓ 1 callers
Class
TemporalRandomCrop
Temporally crop the given frame indices at a random location. Args: size (int): Desired length of frames will be seen in the model.
datasets/video_utils.py:453
↓ 1 callers
Class
TimestepEmbedder
Embeds scalar timesteps into vector representations.
model_zigma.py:232
↓ 1 callers
Class
ToTensorVideo
Convert tensor data type from uint8 to float, divide value by 255.0 and permute the dimensions of clip tensor
datasets/video_utils.py:403
↓ 1 callers
Class
TrainState
utils/train_state_utils.py:21
↓ 1 callers
Class
Transport
transport/transport.py:43
↓ 1 callers
Class
UCFCenterCropVideo
First scale to the specified size in equal proportion to the short edge, then center cropping
datasets/video_utils.py:279
↓ 1 callers
Class
VideoDetector
utils/torchmetric_fvd.py:169
↓ 1 callers
Class
sde
SDE solver class
transport/integrators.py:9
Class
AbstractEncoder
datasets/clip.py:5
Class
Allreduce
dis_causal_conv1d/csrc/causal_conv1d_common.h:47
Class
Allreduce<2>
dis_causal_conv1d/csrc/causal_conv1d_common.h:58
Class
BiMambaInnerFn
dis_mamba/mamba_ssm/ops/selective_scan_interface.py:437
Class
BytesToType
dis_causal_conv1d/csrc/causal_conv1d_common.h:12
Class
BytesToType
dis_mamba/csrc/selective_scan/selective_scan_common.h:29
Class
BytesToType<16>
dis_causal_conv1d/csrc/causal_conv1d_common.h:14
Class
BytesToType<16>
dis_mamba/csrc/selective_scan/selective_scan_common.h:31
Class
BytesToType<1>
dis_causal_conv1d/csrc/causal_conv1d_common.h:34
Class
BytesToType<1>
dis_mamba/csrc/selective_scan/selective_scan_common.h:51
Class
BytesToType<2>
dis_causal_conv1d/csrc/causal_conv1d_common.h:29
Class
BytesToType<2>
dis_mamba/csrc/selective_scan/selective_scan_common.h:46
Class
BytesToType<4>
dis_causal_conv1d/csrc/causal_conv1d_common.h:24
Class
BytesToType<4>
dis_mamba/csrc/selective_scan/selective_scan_common.h:41
Class
BytesToType<8>
dis_causal_conv1d/csrc/causal_conv1d_common.h:19
Class
BytesToType<8>
dis_mamba/csrc/selective_scan/selective_scan_common.h:36
Class
CachedWheelsCommand
The CachedWheelsCommand plugs into the default bdist wheel, which is ran by pip when it cannot find an existing wheel (which is currently the
dis_causal_conv1d/setup.py:191
Class
CachedWheelsCommand
The CachedWheelsCommand plugs into the default bdist wheel, which is ran by pip when it cannot find an existing wheel (which is currently the
dis_mamba/setup.py:199
Class
CaptionEmbedder
Embeds class labels into vector representations. Also handles label dropout for classifier-free guidance.
model_zigma.py:177
Class
CausalConv1dFn
dis_causal_conv1d/causal_conv1d/causal_conv1d_interface.py:10
Class
CenterCropResizeVideo
First use the short side for cropping length, center crop video, then resize to the specified size
datasets/video_utils.py:237
Class
CenterCropVideo
datasets/video_utils.py:346
Class
ConvParamsBase
dis_causal_conv1d/csrc/causal_conv1d.h:9
Class
ConvParamsBwd
dis_causal_conv1d/csrc/causal_conv1d.h:37
Class
Converter
dis_mamba/csrc/selective_scan/selective_scan_common.h:59
Class
Converter<at::BFloat16, N>
dis_mamba/csrc/selective_scan/selective_scan_common.h:79
Class
Converter<at::Half, N>
dis_mamba/csrc/selective_scan/selective_scan_common.h:67
Class
DatasetFromCSV
load video according to the csv file. Args: target_video_len (int): the number of video frames will be load. align_transform (cal
datasets/video_utils.py:470
Class
EasyDict
transport/utils.py:3
Class
Encoder
utils/torchmetric_fdd.py:63
Class
FrozenCLIPEmbedder
Uses the CLIP transformer encoder for text (from Hugging Face)
datasets/clip.py:13
Class
GVPCPlan
transport/path.py:174
Class
GenerationMixin
dis_mamba/mamba_ssm/utils/generation.py:203
Class
ICPlan
Linear Coupling Plan
transport/path.py:18
Class
ImageFolder_FakeWrapper
datasets/dataset_wrapper.py:6
Class
KineticsRandomCropResizeVideo
Slide along the long edge, with the short edge as crop size. And resie to the desired size.
datasets/video_utils.py:319
Class
LayerNormFn
dis_mamba/mamba_ssm/ops/triton/layernorm.py:380
Class
LayerNormLinearFn
dis_mamba/mamba_ssm/ops/triton/layernorm.py:506
Class
MambaEvalWrapper
dis_mamba/evals/lm_harness_eval.py:15
Class
MambaInnerFn
dis_mamba/mamba_ssm/ops/selective_scan_interface.py:292
Class
MambaInnerFnNoOutProj
dis_mamba/mamba_ssm/ops/selective_scan_interface.py:155
Class
MambaLMHeadModel
dis_mamba/mamba_ssm/models/mixer_seq_simple.py:173
Class
ModelType
Which type of output the model predicts.
transport/transport.py:13
Class
NormalizeVideo
Normalize the video clip by mean subtraction and division by standard deviation Args: mean (3-tuple): pixel RGB mean std (3-t
datasets/video_utils.py:378
Class
PathType
Which type of path to use.
transport/transport.py:23
Class
RMSNorm
dis_mamba/mamba_ssm/ops/triton/layernorm.py:481
Class
RandomCropVideo
datasets/video_utils.py:198
Class
SSMParamsBase
dis_mamba/csrc/selective_scan/selective_scan.h:26
Class
SSMParamsBwd
dis_mamba/csrc/selective_scan/selective_scan.h:71
Class
SSMScanOp
dis_mamba/csrc/selective_scan/selective_scan_common.h:108
Class
SSMScanOp<complex_t>
dis_mamba/csrc/selective_scan/selective_scan_common.h:118
Class
SSMScanOp<float>
dis_mamba/csrc/selective_scan/selective_scan_common.h:111
Class
SSMScanParamsBase
dis_mamba/csrc/selective_scan/selective_scan.h:9
Class
SSMScanPrefixCallbackOp
dis_mamba/csrc/selective_scan/selective_scan_common.h:132
Class
SelectiveScanFn
dis_mamba/mamba_ssm/ops/selective_scan_interface.py:14
Class
SumOp
dis_causal_conv1d/csrc/causal_conv1d_common.h:42
Class
VPCPlan
class for VP path flow matching
transport/path.py:139
Class
WeightType
Which type of weighting to use.
transport/transport.py:33
Class
_FeatureExtractorInceptionV3
utils/torchmetric_sfid.py:42
Class
_FeatureExtractorInceptionV3
utils/torchmetric_fdd.py:42
Class
_FeatureExtractorInceptionV3
utils/torchmetric_fvd.py:42