MCPcopy Create free account

hub / github.com/alipay/Ant-Multi-Modal-Framework / functions

Functions4,004 in github.com/alipay/Ant-Multi-Modal-Framework

↓ 9 callersMethodget
(self, item)
antmmf/datasets/database/features_database.py:122
↓ 9 callersMethodget_cross_output
(self, cap_embed, visual_embed, cap_mask, visual_mask, n_clips)
prj/snps3_vtp/roi_univl/univl/model/univl_video_base.py:224
↓ 9 callersMethodget_cross_output
(self, cap_embed, visual_embed, cap_mask, visual_mask, n_clips)
prj/dmae_vtp/roi_univl/univl/model/univl_video_base.py:224
↓ 9 callersMethodget_cross_output
(self, cap_embed, visual_embed, cap_mask, visual_mask, n_clips)
prj/base_vtp/roi_univl/univl/model/univl_video_base.py:224
↓ 9 callersMethodget_size
(self)
antmmf/utils/vocab.py:153
↓ 8 callersMethod__init__
(self, config)
prj/M2_omni/models/qwen2_vit.py:291
↓ 8 callersMethod__init__
(self, num_features, num_layers, hidden, num_classes)
antmmf/modules/graph.py:178
↓ 8 callersMethod__init__
( self, embed_dim: int, # vision image_resolution: int, vision_layers:
antmmf/modules/vision/backbone/clip/model.py:340
↓ 8 callersMethod_calc_text_embeddings
( self, text_tokens, position_2d_embeddings=None, position_1d_embeddings=None,
antmmf/modules/embeddings/univl_layout_embedding.py:86
↓ 8 callersFunctiondraw_multiple_line_text
refer to: python PIL draw multiline text on image: https://stackoverflow.com/a/7698300/395857
antmmf/utils/visual_utils/visualization_utils.py:52
↓ 8 callersMethodfrom_config
(cls, nn_config: Union[str, dict])
prj/M2_Encoder/m2_encoder.py:10
↓ 8 callersMethodget
(self, key, default_val=None)
antmmf/common/configuration.py:232
↓ 8 callersMethodget_loss_metric
TODO: add document here.
antmmf/modules/transformers/heads/itm.py:48
↓ 8 callersMethodinit_weights
Initialize the weights in backbone. Args: pretrained (str, optional): Path to pre-trained weights. Defaults to No
antmmf/modules/vision/backbone/cctt.py:963
↓ 8 callersFunctionmake_block_temporal
(stage, this_segment)
antmmf/modules/vision/temporal_shift.py:117
↓ 8 callersMethodmedian
Return the median value over the data in given window_size.
antmmf/common/meter.py:58
↓ 8 callersMethodprepare_batch
(self, batch)
antmmf/tasks/base_task.py:161
↓ 8 callersFunctiontrunc_normal_
(tensor, mean=0.0, std=1.0)
prj/M2_Encoder/vlmo/modules/modeling_utils.py:17
↓ 7 callersMethod__init__
( self, in_features, hidden_features=None, out_features=None, act_laye
antmmf/modules/vision/backbone/video_swin.py:70
↓ 7 callersFunction_download
(url: str, root: str)
antmmf/modules/vision/backbone/clip/cn_model.py:319
↓ 7 callersMethod_loose_similarity_row
implemented ["wti", "ti"] interactions Args: sequence_output: repeat by batch or row, only sequence_output[0]
prj/dmae_vtp/roi_univl/univl/model/dmae_utils.py:465
↓ 7 callersFunctionavailable_models
Returns the names of available CLIP models
antmmf/modules/vision/backbone/clip/model.py:665
↓ 7 callersMethodget_cross_output
(self, cap_embed, visual_embed, cap_mask, visual_mask, n_clips)
prj/cnvid_vtp/roi_univl/univl/model/univl_video_base.py:224
↓ 7 callersFunctionget_same_padding_conv2d
Chooses static padding if you have specified an image size, and dynamic padding otherwise. Static padding is necessary for ONNX exporting of mo
antmmf/modules/layers/padding.py:215
↓ 7 callersMethodnum_position_embeddings
(self)
prj/M2_Encoder/vlmo/torchscale/component/embedding.py:61
↓ 7 callersMethodreset
(self)
antmmf/utils/timer.py:15
↓ 6 callersMethod__init__
(self, in_features, hidden_features=None, out_features=None, act_layer=nn.GELU, drop=0.)
prj/Pink/pink/model/eva_vit.py:53
↓ 6 callersMethod__init__
( self, d_model=512, nhead=8, num_encoder_layers=6, num_decoder_layers
antmmf/modules/transformers/base.py:30
↓ 6 callersMethod__init__
(self, config: Configuration, *args, **kwargs)
antmmf/modules/encoders/visual_encoder.py:43
↓ 6 callersMethod__init__
( self, in_features, hidden_features=None, out_features=None, act_laye
antmmf/modules/vision/backbone/pvt.py:273
↓ 6 callersMethod__init__
( self, num_classes=512, gating=True, space_to_depth=False, with_text_
antmmf/models/s3dg.py:249
↓ 6 callersMethod_get_partial_loss
(self, sim_matrix, sim_matrix_bar)
prj/dmae_vtp/roi_univl/univl/model/dmae_utils.py:379
↓ 6 callersMethod_update_meter
(self, report, meter, sync=True)
antmmf/trainers/base_trainer.py:692
↓ 6 callersFunctionapply_rotary_pos_emb_vision
(tensor: torch.Tensor, freqs: torch.Tensor)
prj/M2_omni/models/qwen2_vit.py:105
↓ 6 callersMethodforward_head
:param encoder_output: bsz, seq_length, hidden :param decoder_output: :return:
antmmf/modules/transformers/heads/itm.py:34
↓ 6 callersMethodget_input_embeddings
(self)
prj/M2_omni/models/modeling_llama_3d.py:873
↓ 6 callersMethodget_node_info
(self, node)
antmmf/modules/utils.py:187
↓ 6 callersMethodget_output_embeddings
(self)
prj/M2_omni/models/modeling_llama_3d.py:1142
↓ 6 callersFunctionget_package_version
Args: package_name(str): The name of package. Returns: Package's version(str) if package was found, otherwise return None.
antmmf/utils/general.py:557
↓ 6 callersMethodget_parser
(self)
antmmf/utils/flags.py:10
↓ 6 callersMethodget_time_since_start
(self, format=None)
antmmf/utils/timer.py:18
↓ 6 callersMethodinit_weights
(self)
antmmf/models/vilbert.py:960
↓ 6 callersFunctionis_master
()
antmmf/utils/distributed.py:40
↓ 6 callersFunctionone_hot
(indices: torch.Tensor, num_classes: int, unsqueeze_indices=False)
prj/M2_Encoder/vlmo/torchscale/component/xmoe/routing.py:232
↓ 6 callersMethodpreprocessor
(self)
antmmf/datasets/processors/image_processors.py:204
↓ 6 callersFunctionreduce_dict
Args: dictionary (dict): all the values will be reduced Reduce the values in the dictionary from all processes so that all processes
antmmf/utils/distributed_utils.py:200
↓ 6 callersFunctionrepeat_kv
This is the equivalent of torch.repeat_interleave(x, dim=1, repeats=n_rep). The hidden states go from (batch, num_key_value_heads, seqlen, he
prj/M2_omni/models/modeling_llama_3d.py:256
↓ 6 callersMethodreset
(self)
antmmf/modules/metrics/f1.py:26
↓ 6 callersMethodtransform
(self, x)
antmmf/datasets/processors/image_processors.py:439
↓ 6 callersMethodtranspose_for_scores
(self, x)
antmmf/models/vilbert.py:318
↓ 6 callersMethodwith_pos_embed
(self, tensor, pos: Optional[Tensor])
antmmf/modules/transformers/base.py:400
↓ 5 callersMethod__init__
(self, config)
prj/snps3_vtp/roi_univl/roi/model.py:428
↓ 5 callersMethod__init__
(self, config)
prj/dmae_vtp/roi_univl/roi/model.py:428
↓ 5 callersMethod__init__
(self, )
prj/dmae_vtp/roi_univl/univl/model/dmae_utils.py:540
↓ 5 callersMethod__init__
(self, config)
prj/base_vtp/roi_univl/roi/model.py:428
↓ 5 callersMethod__init__
(self, config)
prj/cnvid_vtp/roi_univl/roi/model.py:428
↓ 5 callersMethod__init__
(self, config: Configuration, *args, **kwargs)
antmmf/modules/encoders/text_encoder.py:23
↓ 5 callersMethod__init__
(self, image_feat_dim, ques_emb_dim, **kwargs)
antmmf/modules/layers/modal_combine_layer.py:51
↓ 5 callersMethod__init__
(self, config)
antmmf/models/layoutlm.py:66
↓ 5 callersMethod_box_mode_keeper
In some inplace operations, we may change the box_mode of Boxes, this mode keeper context will help us to keep the box_mode unchanged
antmmf/structures/boxes.py:78
↓ 5 callersMethod_len_of_loader_list
(self, loader_list)
antmmf/trainers/base_trainer.py:166
↓ 5 callersFunctionbuild_processors
(config, *args, **kwargs)
antmmf/datasets/build.py:53
↓ 5 callersMethodcheck_input
(self, input_img)
antmmf/modules/encoders/visual_encoder.py:52
↓ 5 callersFunctionckpt_name_from_core_args
(config)
antmmf/utils/general.py:71
↓ 5 callersMethodfrom_pretrained
create an efficientnet model according to name. Args: model_name (str): Name for efficientnet. weights_path (None or
antmmf/modules/vision/backbone/efficientnet.py:330
↓ 5 callersFunctionget_low_clip
(img)
antmmf/utils/dataset_utils.py:215
↓ 5 callersMethodget_time_hhmmss
Calculates time since `start` and formats as a string.
antmmf/utils/timer.py:21
↓ 5 callersFunctionget_world_size
()
antmmf/utils/distributed.py:44
↓ 5 callersFunctionkeep_till_eos
(item)
antmmf/utils/text_utils.py:315
↓ 5 callersMethodload_state_dict
(self, name, download_root=None)
prj/dmae_vtp/roi_univl/univl/model/clip_text_encoder.py:194
↓ 5 callersMethodmkdirs
(path: str)
antmmf/utils/file_io.py:97
↓ 5 callersMethodpairwise_iou
Implementation from https://github.com/kuangliu/torchcv/blob/master/torchcv/utils/box.py with slight modifications. Assume t
antmmf/structures/boxes.py:302
↓ 5 callersFunctionplain_run
(args: Namespace)
antmmf/run.py:40
↓ 5 callersMethodprepare_cross_text
(self, input_ids, input_mask)
prj/snps3_vtp/roi_univl/univl/model/univl_video_base.py:168
↓ 5 callersMethodprepare_cross_text
(self, input_ids, input_mask)
prj/dmae_vtp/roi_univl/univl/model/univl_video_base.py:168
↓ 5 callersMethodprepare_cross_text
(self, input_ids, input_mask)
prj/base_vtp/roi_univl/univl/model/univl_video_base.py:168
↓ 5 callersMethodreset
(self)
antmmf/modules/metrics/accuracy.py:22
↓ 5 callersFunctionround_filters
Calculate and round number of filters based on width multiplier. Use width_coefficient, depth_divisor and min_depth of global_params. Args
antmmf/modules/vision/backbone/efficientnet.py:763
↓ 5 callersMethodstep
Performs a single optimization step.
antmmf/optimizer/adan.py:98
↓ 5 callersFunctionsynchronize
()
antmmf/utils/distributed_utils.py:21
↓ 5 callersMethodtokenize
(self, text)
antmmf/modules/vision/backbone/clip/cn_tokenizer.py:190
↓ 5 callersMethodzero_grad
(self, set_to_none: Optional[bool] = ...)
antmmf/optimizer/combine_optimizers.py:129
↓ 4 callersMethod__init__
( self, in_features, hidden_features=None, out_features=None, act_laye
prj/M2_Encoder/vlmo/modules/multiway_transformer.py:32
↓ 4 callersMethod__init__
(self, image_dim, question_dim, **kwargs)
antmmf/modules/attention.py:9
↓ 4 callersMethod__init__
(self, *args, **kwargs)
antmmf/modules/metrics/f1.py:20
↓ 4 callersMethod__init__
(self, modalities, topk, *args, **kwargs)
antmmf/modules/metrics/mm_retrieval_recall.py:65
↓ 4 callersMethod__init__
(self, name: str = "accuracy")
antmmf/modules/metrics/accuracy.py:18
↓ 4 callersMethod__init__
(self, block, n_segment)
antmmf/modules/vision/non_local.py:171
↓ 4 callersMethod__init__
(self, *args, **params)
antmmf/utils/vocab.py:14
↓ 4 callersMethod__init__
(self, config)
antmmf/models/spkResNet.py:153
↓ 4 callersMethod_calc_img_embeddings
:param visual_tokens: B, seq, h :param position_2d_embeddings: B, seq, h :param position_1d_embeddings: B, seq, h :p
antmmf/modules/embeddings/univl_layout_embedding.py:52
↓ 4 callersMethod_calculate
(self, scores, expected)
antmmf/modules/metrics/f1.py:30
↓ 4 callersMethod_forward_pass
(self, batch, enable_amp=False)
antmmf/trainers/base_trainer.py:609
↓ 4 callersMethod_make_layer
(self, planes, blocks, stride=1)
antmmf/modules/vision/backbone/clip/model.py:182
↓ 4 callersMethod_make_layer
(self, block, planes, num_blocks, stride)
antmmf/models/spkResNet.py:112
↓ 4 callersMethodarea
Computes the area of all the boxes. Returns: torch.Tensor: a vector with areas of each box.
antmmf/structures/boxes.py:167
↓ 4 callersMethodattention
(self, x: torch.Tensor)
antmmf/modules/vision/backbone/clip/model.py:245
↓ 4 callersMethodbackward
(self, samples, optimizer, loss)
antmmf/models/mm_adversarial.py:152
↓ 4 callersMethodbuild
(self)
antmmf/models/cnn.py:38
↓ 4 callersFunctionbuild_embedder
(config, *args, **kwargs)
antmmf/modules/build.py:80
← previousnext →101–200 of 4,004, ranked by callers