MCPcopy Create free account

hub / github.com/AIS-SNU/Smart-Infinity / functions

Functions4,989 in github.com/AIS-SNU/Smart-Infinity

↓ 4 callersMethodfork
Fork the cuda rng state, perform operations, and exit with the original state.
DeepSpeedExample/megatron/core/tensor_parallel/random.py:149
↓ 4 callersMethodfp_tensor_constructor
(self, fn: Callable, target_fp_dtype: torch.dtype)
deepspeed/utils/init_on_device.py:42
↓ 4 callersMethodfrom_meta
(cls, meta, local_part, group, device=get_accelerator().device_name())
deepspeed/runtime/utils.py:635
↓ 4 callersMethodfull
(self, device=None)
deepspeed/runtime/utils.py:668
↓ 4 callersFunctionget_attr_wrapped_model
Get an attribute from a wrapped model
DeepSpeedExample/megatron/core/utils.py:29
↓ 4 callersMethodget_axis_comm_lists
Construct lists suitable for a communicator group along axis ``axis``. Example: >>> topo = Topo(axes=['pipe', 'data', 'model'],
deepspeed/runtime/pipe/topology.py:127
↓ 4 callersMethodget_axis_names
Return a list of the axis names in the ordering of the topology.
deepspeed/runtime/pipe/topology.py:65
↓ 4 callersFunctionget_compression_config
(param_dict)
deepspeed/compression/config.py:11
↓ 4 callersFunctionget_cuda_rng_tracker
Get cuda rng tracker.
deepspeed/runtime/activation_checkpointing/checkpointing.py:193
↓ 4 callersMethodget_dp_process_group
Return the communication group with all data-parallel ranks
deepspeed/runtime/zero/partition_parameters.py:1527
↓ 4 callersFunctionget_file_size
deepspeed/ops/csrc/aio/common/deepspeed_aio_utils.cpp:98
↓ 4 callersFunctionget_files
(dir)
deepspeed/checkpoint/reshape_utils.py:33
↓ 4 callersMethodget_full_hp_param
(self, param, optim_state_key=None)
deepspeed/runtime/zero/stage3.py:2181
↓ 4 callersFunctionget_indexed_dataset_
(data_prefix, data_impl, skip_warmup)
DeepSpeedExample/megatron/data/dataset_utils.py:455
↓ 4 callersFunctionget_lazy_path
Gets directory path where lazy files are stored.
DeepSpeedExample/megatron/deprecated_data_utils/lazy_loader.py:26
↓ 4 callersMethodget_mask
(self, pruning_type='sparse')
deepspeed/compression/basic_layer.py:528
↓ 4 callersFunctionget_model
Build the model.
DeepSpeedExample/megatron/training.py:135
↓ 4 callersFunctionget_model_type
(model)
DeepSpeedExample/megatron/core/utils.py:48
↓ 4 callersMethodget_module
(self, sd)
deepspeed/runtime/state_dict_factory.py:149
↓ 4 callersFunctionget_module_name
get the associated module name from the model based on the key_word provided by users
deepspeed/compression/compress.py:30
↓ 4 callersFunctionget_pipeline_model_parallel_group
Get the pipeline model parallel group the caller rank belongs to.
DeepSpeedExample/megatron/core/parallel_state.py:355
↓ 4 callersMethodget_sd_loader
(ckpt_list, checkpoint_engine, sd_type='Megatron', version=None)
deepspeed/runtime/state_dict_factory.py:41
↓ 4 callersFunctionget_sequence_parallel_world_size
Return world size for the sequence parallel group.
DeepSpeedExample/megatron/core/parallel_state.py:448
↓ 4 callersMethodget_states
Get rng states. Copy the dictionary so we have direct pointers to the states, not just a pointer to the dictionary.
DeepSpeedExample/megatron/mpu/random.py:147
↓ 4 callersFunctionget_weight_norm
Get norm of an iterable of parameters. This is adapted from torch.nn.utils.clip_grad.clip_grad_norm_ and added functionality to handle model
deepspeed/runtime/utils.py:472
↓ 4 callersMethodhas_all_gather_into_tensor
(self)
deepspeed/comm/torch.py:78
↓ 4 callersMethodhas_gradients
(self)
deepspeed/runtime/swap_tensor/optimizer_utils.py:62
↓ 4 callersMethodinclude_paths
Returns list of include paths, relative to root of deepspeed package (i.e., DeepSpeed/deepspeed)
deepspeed/ops/op_builder/builder.py:205
↓ 4 callersFunctioninitialize_megatron
Set global variables, initialize distributed, and set autoresume and random seeds. `allow_no_cuda` should not be set unless using megatron for
DeepSpeedExample/megatron/initialize.py:34
↓ 4 callersMethodinsert_tensor
(self, tensor, swap_path, aligned_numel)
deepspeed/runtime/swap_tensor/utils.py:50
↓ 4 callersFunctionis_activation_to_checkpoint
Is an activation to be checkpointed
deepspeed/runtime/activation_checkpointing/checkpointing.py:358
↓ 4 callersMethodis_compatible
Check if all non-python dependencies are satisfied to build this op
deepspeed/ops/op_builder/builder.py:223
↓ 4 callersMethodis_initialized
(self)
deepspeed/comm/torch.py:227
↓ 4 callersMethodis_moe_group
(self, group)
deepspeed/runtime/zero/stage_1_and_2.py:535
↓ 4 callersFunctionload
(module, state_dict, prefix, mp_group=None)
deepspeed/module_inject/replace_module.py:874
↓ 4 callersFunctionload_checkpoint
Load a model checkpoint and return the iteration.
DeepSpeedExample/megatron/checkpointing.py:175
↓ 4 callersMethodlog_timers
(self, timer_names)
deepspeed/runtime/zero/stage3.py:1641
↓ 4 callersFunctionmaybe_copy_qkv
(module, sd, weight_quantizer, mp_replace, dst_name, src_names, split_qkv=False)
deepspeed/module_inject/policy.py:175
↓ 4 callersFunctionmerge_state
(state_a, state_b, key_list=[])
deepspeed/checkpoint/reshape_utils.py:83
↓ 4 callersMethodmy_igather
(self, rank, size, comm, sendbuf, recbuf, root)
deepspeed/runtime/comm/mpi.py:24
↓ 4 callersMethodnormalize
(self, v)
deepspeed/runtime/eigenvalue.py:46
↓ 4 callersMethodprint_data
(self, tag)
deepspeed/checkpoint/reshape_meg_2d.py:43
↓ 4 callersFunctionprint_object
(obj, name, exclude_list=[])
deepspeed/runtime/swap_tensor/utils.py:29
↓ 4 callersMethodquantize_data
(self, data, quantize_bits, groups, key=None)
deepspeed/runtime/weight_quantizer.py:21
↓ 4 callersFunctionread_longs
(f, n)
DeepSpeedExample/megatron/data/indexed_dataset.py:82
↓ 4 callersFunctionread_longs
(f, n)
deepspeed/runtime/data_pipeline/data_sampling/indexed_dataset.py:91
↓ 4 callersFunctionreplace_dict
Replaces values in dict d with values in dict u. Args: d (dict): the target dict to overwrite u (dict): the dict containing the v
deepspeed/autotuning/utils.py:110
↓ 4 callersFunctionreport_file_error
deepspeed/ops/csrc/aio/common/deepspeed_aio_common.cpp:258
↓ 4 callersMethodreset
Set to the initial state (no tracker).
DeepSpeedExample/megatron/mpu/random.py:142
↓ 4 callersMethodreset_activation_shape
Reset the buffers when the shape of activation and gradient change. For example, for curriculum learning that changes the seqlen of each
deepspeed/runtime/pipe/engine.py:276
↓ 4 callersMethodschedule_experiments
(self, exp_paths)
deepspeed/autotuning/scheduler.py:59
↓ 4 callersMethodset_command_tokens
(self, command_tokens)
DeepSpeedExample/megatron/deprecated_data_utils/tokenization.py:74
↓ 4 callersMethodset_state
(self, state)
deepspeed/runtime/data_pipeline/curriculum_scheduler.py:119
↓ 4 callersMethodset_tensor_parallel_config
(self, mp_size, mp_group)
deepspeed/module_inject/containers/base.py:153
↓ 4 callersMethodsources
Returns list of source files for your op, relative to root of deepspeed package (i.e., DeepSpeed/deepspeed)
deepspeed/ops/op_builder/builder.py:122
↓ 4 callersFunctionsplit_index
(start_idx, end_idx, num_partitions)
deepspeed/runtime/data_pipeline/data_sampling/utils.py:34
↓ 4 callersMethodstart_timers
(self, timer_names)
deepspeed/runtime/zero/stage3.py:1647
↓ 4 callersMethodstep
If no closure is supplied, :attr:`step` should be called after ``fp16_optimizer_obj.backward(loss)``. :attr:`step` updates th
DeepSpeedExample/megatron/fp16/fp16.py:420
↓ 4 callersMethodstop_timers
(self, timer_names)
deepspeed/runtime/zero/stage3.py:1654
↓ 4 callersMethodswappable_tensor
(self, param=None, numel=None)
deepspeed/runtime/swap_tensor/optimizer_utils.py:193
↓ 4 callersFunctiontrim_mean
Compute the trimmed mean of a list of numbers. Args: data (list): List of numbers. trim_percent (float): Percentage of data to tr
deepspeed/utils/timer.py:247
↓ 4 callersMethodtune_space
(self, tuning_space, prev_max_mbs=0, prev_best_mbs=0, prev_best_metric_val=0)
deepspeed/autotuning/autotuner.py:523
↓ 4 callersMethodupdate_difficulty
(self, global_steps)
deepspeed/runtime/data_pipeline/curriculum_scheduler.py:155
↓ 4 callersMethodversion_dependent_macros
(self)
deepspeed/ops/op_builder/builder.py:587
↓ 4 callersMethodwait
(self)
deepspeed/runtime/zero/partition_parameters.py:51
↓ 4 callersFunctionwrite_longs
(f, a)
DeepSpeedExample/megatron/data/indexed_dataset.py:88
↓ 4 callersFunctionwrite_longs
(f, a)
deepspeed/runtime/data_pipeline/data_sampling/indexed_dataset.py:97
↓ 4 callersMethodwriter
(cls, path, dtype)
deepspeed/runtime/data_pipeline/data_sampling/indexed_dataset.py:375
↓ 4 callersMethodzero_overlap_comm
(self)
deepspeed/runtime/engine.py:680
↓ 4 callersFunctionzero_wrapper_for_fp_tensor_constructor
(fn: Callable, target_fp_dtype: torch.dtype)
deepspeed/runtime/zero/partition_parameters.py:224
↓ 3 callersMethodEvent
(self)
deepspeed/accelerator/cpu_accelerator.py:99
↓ 3 callersMethodHasDropout
deepspeed/ops/csrc/includes/dropout.h:63
↓ 3 callersMethodIncrementStep_
deepspeed/ops/csrc/includes/cpu_adam.h:195
↓ 3 callersMethodSetMask
deepspeed/ops/csrc/includes/dropout.h:67
↓ 3 callersMethodSoftmax
deepspeed/ops/csrc/includes/softmax.h:38
↓ 3 callersMethod__check_params
(model: Module, dtype: torch.dtype)
deepspeed/runtime/engine.py:1017
↓ 3 callersMethod__init__
(self, input_size: int, output_size: int, *, config: ModelParallelConfig, in
DeepSpeedExample/megatron/core/tensor_parallel/layers.py:685
↓ 3 callersMethod__init__
A context manager to partition the model parameters during the model construction with MiCS partition strategy. Model states are partitioned
deepspeed/runtime/zero/mics.py:56
↓ 3 callersMethod_aligned_size
(self, param)
deepspeed/runtime/zero/partition_parameters.py:1033
↓ 3 callersMethod_apply_injection_policy
(self, config, client_module=None)
deepspeed/inference/engine.py:413
↓ 3 callersFunction_bias_dropout_add_func
(x, bias, residual, prob, training)
DeepSpeedExample/megatron/core/fusions/fused_bias_dropout.py:6
↓ 3 callersMethod_combine_output_splits
Join the splits of the output into a single result. Args: outputs (List[Any]): The reduced outputs for each output split.
deepspeed/runtime/zero/tiling.py:195
↓ 3 callersMethod_create_checkpoint_file
(self, save_dir, tag, zero_checkpoint)
deepspeed/runtime/engine.py:3051
↓ 3 callersMethod_decode
(self, x, return_dict=True)
deepspeed/model_implementations/diffusers/vae.py:33
↓ 3 callersFunction_drop_tokens
Divide a tensor among the tensor parallel ranks
deepspeed/moe/mappings.py:46
↓ 3 callersMethod_encode
(self, x, return_dict=True)
deepspeed/model_implementations/diffusers/vae.py:76
↓ 3 callersMethod_ensure_availability_of_partitioned_params
(self, params)
deepspeed/runtime/zero/partition_parameters.py:1043
↓ 3 callersFunction_ensure_var_is_initialized
Make sure the input variable is not None.
DeepSpeedExample/megatron/global_vars.py:145
↓ 3 callersMethod_forward
(self, sample, timestamp, encoder_hidden_states, return_dict=True, cross_attention_kwargs=None)
deepspeed/model_implementations/diffusers/unet.py:65
↓ 3 callersMethod_forward
(self, sample, timestamp, encoder_hidden_states, return_dict=True)
deepspeed/model_implementations/diffusers/vae.py:149
↓ 3 callersMethod_fuse_lora
(self, params, lora_params)
deepspeed/runtime/hybrid_engine.py:130
↓ 3 callersFunction_gather
Gather tensors and concatinate along the last dimension.
DeepSpeedExample/megatron/mpu/mappings.py:54
↓ 3 callersFunction_gather_along_last_dim
Gather tensors and concatinate along the last dimension.
DeepSpeedExample/megatron/core/tensor_parallel/mappings.py:68
↓ 3 callersFunction_gather_tokens
Gather tensors and concatenate them along a dimension
deepspeed/moe/mappings.py:28
↓ 3 callersMethod_generate
(self, *inputs, **kwargs)
deepspeed/inference/engine.py:620
↓ 3 callersMethod_get_buffer
(self, index)
deepspeed/runtime/swap_tensor/async_swapper.py:157
↓ 3 callersMethod_get_ckpt_name
(self, checkpoints_path, tag, mp_placeholder=None)
deepspeed/runtime/engine.py:2509
↓ 3 callersFunction_get_data_parallel_group
Get the data parallel group the caller rank belongs to.
deepspeed/utils/groups.py:319
↓ 3 callersMethod_get_expert_ckpt_name
(checkpoints_path, layer_id, expert_id, tag, mpu=None)
deepspeed/runtime/engine.py:2538
↓ 3 callersFunction_get_expert_data_parallel_group
Get the expert data parallel group the caller rank belongs to.
deepspeed/utils/groups.py:292
← previousnext →401–500 of 4,989, ranked by callers