MCPcopy Create free account

hub / github.com/VITA-MLLM/VITA / functions

Functions1,650 in github.com/VITA-MLLM/VITA

↓ 1 callersFunctioncan_infer_option
(answer, choices)
VLMEvalKit/vlmeval/utils/matching_util.py:7
↓ 1 callersFunctioncan_infer_text
(answer, choices)
VLMEvalKit/vlmeval/utils/matching_util.py:51
↓ 1 callersFunctionchange_file
(file_path, mm_vision_tower)
VLMEvalKit/vlmeval/vlm/video_llm/llama_vid.py:23
↓ 1 callersMethodchat_inner
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/idefics.py:235
↓ 1 callersMethodchat_inner
(self, inputs, **kwargs)
VLMEvalKit/vlmeval/api/base.py:141
↓ 1 callersMethodchat_inner_v2
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/internvl_chat.py:401
↓ 1 callersFunctionchat_mt
(model, messages, dataset_name)
VLMEvalKit/vlmeval/inference_mt.py:20
↓ 1 callersFunctioncheck_denotation
Return True if the predicted denotation is correct. Args: target_values (list[Value]) predicted_values (list[Value]) Returns:
VLMEvalKit/vlmeval/dataset/utils/tablevqabench.py:456
↓ 1 callersMethodcheck_install
(self)
VLMEvalKit/vlmeval/vlm/deepseek_vl.py:13
↓ 1 callersFunctioncheck_video_file
(video_name, set_id)
data_tools/check_video_mmco.py:42
↓ 1 callersFunctioncollect_error_files
(log_file_path)
data_tools/mmco_collect.py:3
↓ 1 callersFunctioncollect_results
(llm_output)
web_demo/server.py:385
↓ 1 callersFunctioncompleted
(m, d, suf)
VLMEvalKit/vlmeval/tools.py:101
↓ 1 callersMethodcompute_scores
(self)
VLMEvalKit/vlmeval/dataset/image_caption.py:20
↓ 1 callersFunctionconcat_item
(items)
data_tools/concat_data_frameCat.py:69
↓ 1 callersFunctionconcat_item
(items)
data_tools/concat_data_patch_nemo.py:69
↓ 1 callersFunctionconcat_item
(items)
data_tools/concat_data_patch.py:69
↓ 1 callersFunctionconcat_item
(items)
data_tools/concat_data_patch_qwen.py:70
↓ 1 callersFunctionconcat_item
(items)
data_tools/concat_data.py:91
↓ 1 callersFunctionconvert_webm_to_mp4
(input_file, output_file)
web_demo/web_ability_demo.py:150
↓ 1 callersFunctiondefault_rating
(data_file)
VLMEvalKit/vlmeval/dataset/utils/yorn.py:144
↓ 1 callersFunctiondetermine_dataset
(eval_file)
VLMEvalKit/scripts/mmb_eval_gradio.py:45
↓ 1 callersFunctiondisable_torch_init
Disable the redundant torch default initialization to accelerate model creation.
VLMEvalKit/vlmeval/vlm/yi_vl.py:43
↓ 1 callersFunctiondo_setup
()
VLMEvalKit/setup.py:87
↓ 1 callersMethoddtype
(self)
vita/model/multimodal_encoder/clip/clip_encoder.py:58
↓ 1 callersMethoddump_image
(self, line)
VLMEvalKit/vlmeval/dataset/image_base.py:98
↓ 1 callersMethoddump_image
Dump the image(s) of the input line to the corresponding dataset folder. Args: line (line of pd.DataFrame): The raw input line.
VLMEvalKit/vlmeval/api/sensechat_vision.py:31
↓ 1 callersFunctiondynamic_preprocess
( image, min_num=2, max_num=12, image_size=448, use_thumbnail=False, img_mean=0 )
data_tools/concat_data_frameCat.py:36
↓ 1 callersFunctiondynamic_preprocess
(image, min_num=1, max_num=12, image_size=448, use_thumbnail=True)
data_tools/concat_data_patch_nemo.py:35
↓ 1 callersFunctiondynamic_preprocess
( image, min_num=2, max_num=12, image_size=448, use_thumbnail=False, img_mean=0 )
data_tools/statistics_token_num_frameCat.py:47
↓ 1 callersFunctiondynamic_preprocess
(image, min_num=1, max_num=12, image_size=448, use_thumbnail=True)
data_tools/concat_data_patch_qwen.py:36
↓ 1 callersFunctiondynamic_preprocess
(image, min_num=1, max_num=12, image_size=448, use_thumbnail=True)
data_tools/statistics_token_num_patch.py:40
↓ 1 callersFunctiondynamic_preprocess
(image, min_num=1, max_num=12, image_size=448, use_thumbnail=True)
data_tools/statistics_token_num_patch_video.py:50
↓ 1 callersFunctiondynamic_preprocess
( image, min_num=1, max_num=6, image_size=448, use_thumbnail=False )
VLMEvalKit/vlmeval/vlm/mmalaya.py:101
↓ 1 callersFunctiondynamic_preprocess
(image, min_num=1, max_num=6, image_size=448, use_thumbnail=False)
VLMEvalKit/vlmeval/vlm/internvl_chat.py:49
↓ 1 callersMethodembed_gst
(self, x)
vita/model/vita_tts/decoder/ticodec/models.py:703
↓ 1 callersMethodencode_jwt_token
(self, ak, sk)
VLMEvalKit/vlmeval/api/sensechat_vision.py:71
↓ 1 callersFunctioneval_score
(gt, pred, answer_type)
VLMEvalKit/vlmeval/dataset/mmlongbench.py:296
↓ 1 callersFunctionevaluate_fintabnet
(data, score_keys)
VLMEvalKit/vlmeval/dataset/utils/tablevqabench.py:129
↓ 1 callersFunctionevaluate_tabfact
(data, score_keys)
VLMEvalKit/vlmeval/dataset/utils/tablevqabench.py:56
↓ 1 callersFunctionevaluate_wtq
(data, score_keys)
VLMEvalKit/vlmeval/dataset/utils/tablevqabench.py:94
↓ 1 callersFunctionexpand2even
(pil_img, new_target_width, new_target_height, background_color)
vita/util/data_utils_video_audio_neg_frameCat.py:1316
↓ 1 callersFunctionexpand2square
(pil_img, background_color)
video_audio_demo.py:85
↓ 1 callersFunctionexpand2square
(pil_img, background_color)
videomme/yt_video_inference_qa.py:160
↓ 1 callersFunctionexpand2square
(pil_img, background_color)
videomme/yt_video_inference_qa_imgs.py:130
↓ 1 callersMethodexpand2square
(self, pil_img, background_color)
vita/model/language_model/vita_fo_qwen2.py:188
↓ 1 callersMethodexpand2square
(self, pil_img, background_color)
vita/model/language_model/vita_nemo.py:246
↓ 1 callersMethodexpand2square
(self, pil_img, background_color)
vita/model/language_model/vita_qwen2.py:266
↓ 1 callersMethodexpand2square
(self, pil_img, background_color)
vita/model/language_model/vita_mixtral.py:385
↓ 1 callersFunctionexpand_question_into_multimodal
( question_text, image_token_len, im_st_token, im_ed_token, im_patch_token )
VLMEvalKit/vlmeval/vlm/omnilmm.py:55
↓ 1 callersFunctionextract_json_objects
(text, decoder=JSONDecoder())
VLMEvalKit/vlmeval/smp/misc.py:206
↓ 1 callersFunctionfind_all_linear_names
(model)
vita/train/train.py:157
↓ 1 callersFunctionfind_best_match
Finds the best matching n-gram in the haystack for the given needle. Parameters: needle (str): The string to find. hay (str): The te
VLMEvalKit/vlmeval/dataset/vcr.py:86
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
vita/util/data_utils_video_audio_patch.py:1340
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
vita/util/data_utils_video_audio_neg_patch_fo.py:1491
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
vita/util/data_utils_video_patch_audio.py:1390
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
vita/util/data_utils_video_audio_neg_frameCat.py:1219
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
vita/util/data_utils_video_audio_patch_sf.py:1358
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
VLMEvalKit/vlmeval/vlm/mmalaya.py:85
↓ 1 callersFunctionfind_closest_aspect_ratio
(aspect_ratio, target_ratios, width, height, image_size)
VLMEvalKit/vlmeval/vlm/internvl_chat.py:33
↓ 1 callersMethodfind_min_sum_index
Find the index with the minimum sum of a sliding window in the given audio segment and perform operations based on this index.
vita/model/vita_tts/decoder/llm2tts.py:70
↓ 1 callersMethodfix_init_weight
(self)
vita/model/multimodal_encoder/eva_clip/eva_vit.py:717
↓ 1 callersFunctionfloat_cvt
(s)
VLMEvalKit/vlmeval/dataset/utils/mmvet.py:46
↓ 1 callersFunctionfloat_to_int16
(audio: np.ndarray)
web_demo/web_ability_demo.py:34
↓ 1 callersMethodfor_one_step
(self, xin, idx)
vita/model/vita_tts/decoder/ticodec/models.py:583
↓ 1 callersMethodfor_one_step_gst
(self, xin)
vita/model/vita_tts/decoder/ticodec/models.py:617
↓ 1 callersFunctionframe2img
(img_path_list, font, save_path=None, idx_start=0)
VLMEvalKit/vlmeval/dataset/mmlongbench.py:206
↓ 1 callersFunctiongen_packages_items
()
VLMEvalKit/setup.py:65
↓ 1 callersFunctiongen_table
(models, datasets)
VLMEvalKit/scripts/summarize.py:72
↓ 1 callersMethodgenerate
The main function to generate the answer. Will call `generate_inner` with the preprocessed input messages. Args: message: raw inp
VLMEvalKit/vlmeval/api/base.py:195
↓ 1 callersMethodgenerate_brief
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2_4KHD.py:142
↓ 1 callersMethodgenerate_brief
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2.py:125
↓ 1 callersMethodgenerate_brief
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2d5.py:198
↓ 1 callersMethodgenerate_inner
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/mmalaya.py:40
↓ 1 callersMethodgenerate_inner
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/base.py:46
↓ 1 callersMethodgenerate_inner
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/deepseek_vl.py:62
↓ 1 callersMethodgenerate_inner_image
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/llava/llava.py:420
↓ 1 callersMethodgenerate_inner_video
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/llava/llava.py:460
↓ 1 callersMethodgenerate_list
(self, full_inputs, offset=0, **kwargs)
VLMEvalKit/vlmeval/api/hf_chat_model.py:165
↓ 1 callersMethodgenerate_mme
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2_4KHD.py:117
↓ 1 callersMethodgenerate_mme
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2.py:96
↓ 1 callersMethodgenerate_mme
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2d5.py:170
↓ 1 callersMethodgenerate_multichoice
(self, image_path, prompt)
VLMEvalKit/vlmeval/vlm/monkey.py:49
↓ 1 callersMethodgenerate_multichoice
(self, image_path, prompt)
VLMEvalKit/vlmeval/vlm/monkey.py:131
↓ 1 callersMethodgenerate_multichoice
(self, image_path, text, dataset)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2_4KHD.py:124
↓ 1 callersMethodgenerate_multichoice
(self, image_path, text, dataset)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2.py:103
↓ 1 callersMethodgenerate_multichoice
(self, image_path, text, dataset)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2d5.py:177
↓ 1 callersFunctiongenerate_self_signed_cert
Generate a self-signed SSL certificate and its corresponding private key. Parameters: - cert_file (str): The file path where the generat
web_demo/vita_html/web/pem.py:14
↓ 1 callersMethodgenerate_str
(self, input, **kwargs)
VLMEvalKit/vlmeval/api/hf_chat_model.py:127
↓ 1 callersMethodgenerate_v1_2
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/internvl_chat.py:272
↓ 1 callersMethodgenerate_v1_5
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/internvl_chat.py:285
↓ 1 callersMethodgenerate_v2
(self, message, dataset=None)
VLMEvalKit/vlmeval/vlm/internvl_chat.py:312
↓ 1 callersMethodgenerate_vqa
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2_4KHD.py:134
↓ 1 callersMethodgenerate_vqa
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2.py:113
↓ 1 callersMethodgenerate_vqa
(self, image_path, text)
VLMEvalKit/vlmeval/vlm/xcomposer/xcomposer2d5.py:188
↓ 1 callersFunctionget_args
()
web_demo/server.py:28
↓ 1 callersMethodget_audio_encoder
(self)
vita/model/vita_arch.py:143
↓ 1 callersFunctionget_audio_feature_size
(audio: torch.Tensor)
web_demo/vllm_tools/vllm_file/qwen2.py:272
↓ 1 callersMethodget_chunk_size
(self)
web_demo/wakeup_and_vad/wakeup_and_vad.py:125
↓ 1 callersMethodget_context_emb
(self, conv, model, img_list, answer_prompt=None, print_res=False)
VLMEvalKit/vlmeval/vlm/video_llm/videochat2.py:244
← previousnext →401–500 of 1,650, ranked by callers