Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/AIS-SNU/Smart-Infinity
/ functions
Functions
4,989 in github.com/AIS-SNU/Smart-Infinity
⨍
Functions
4,989
◇
Types & classes
624
↓ 15 callers
Method
state_dict
(self)
deepspeed/runtime/lr_schedules.py:619
↓ 15 callers
Method
update_records
(self, space_name, exp, metric_val, num_exps)
deepspeed/autotuning/autotuner.py:707
↓ 15 callers
Method
write
(self, sizes, doc_idx)
deepspeed/runtime/data_pipeline/data_sampling/indexed_dataset.py:400
↓ 14 callers
Method
GetWorkSpace
deepspeed/ops/csrc/transformer/inference/includes/inference_context.h:183
↓ 14 callers
Method
clear
Clear experiment queues, does not reset self.experiment_count
deepspeed/autotuning/scheduler.py:247
↓ 14 callers
Method
exists
(path)
DeepSpeedExample/megatron/data/indexed_dataset.py:201
↓ 14 callers
Method
flush
(self)
deepspeed/monitor/tensorboard.py:54
↓ 14 callers
Function
recursive_getattr
Recursively get the attribute of a module. Args: model (`torch.nn.Module`) The model to get the attribute from. m
deepspeed/compression/helper.py:17
↓ 14 callers
Function
set_random_seed
Set random seed for reproducability.
DeepSpeedExample/megatron/mpu/tests/commons.py:34
↓ 14 callers
Method
write
given a generator of metrics for each of the data points X_i, write the metrics, text, and labels to a csv file
DeepSpeedExample/megatron/deprecated_data_utils/datasets.py:299
↓ 13 callers
Method
GetMaxTokenLength
deepspeed/ops/csrc/transformer/inference/includes/inference_context.h:178
↓ 13 callers
Method
free
(self, buffers)
deepspeed/runtime/swap_tensor/utils.py:216
↓ 13 callers
Function
get_model_parallel_group
Get the model parallel group the caller rank belongs to.
DeepSpeedExample/megatron/mpu/initialize.py:97
↓ 13 callers
Method
get_param_id
(self, param)
deepspeed/runtime/zero/stage3.py:1083
↓ 13 callers
Method
get_param_id
(self, param)
deepspeed/runtime/zero/stage_1_and_2.py:817
↓ 13 callers
Function
get_tensor_model_parallel_rank
Return my rank for the tensor model parallel group.
DeepSpeedExample/megatron/core/parallel_state.py:505
↓ 13 callers
Method
is_last_stage
True if this process is in the last stage in the pipeline.
deepspeed/runtime/pipe/engine.py:462
↓ 13 callers
Method
print_allocation
(self, resolution=200)
deepspeed/runtime/zero/contiguous_memory_allocator.py:121
↓ 13 callers
Function
print_separator
(message)
DeepSpeedExample/megatron/mpu/tests/commons.py:76
↓ 13 callers
Method
wait
(self)
deepspeed/runtime/swap_tensor/pipelined_optimizer_swapper.py:33
↓ 13 callers
Method
write
(self, sizes, doc_idx)
DeepSpeedExample/megatron/data/indexed_dataset.py:363
↓ 12 callers
Method
Stream
(self)
deepspeed/accelerator/cpu_accelerator.py:85
↓ 12 callers
Function
close_mmap_dataset_builder
(builder, fname)
deepspeed/runtime/data_pipeline/data_sampling/utils.py:52
↓ 12 callers
Function
create_mmap_dataset_builder
(fname, dtype)
deepspeed/runtime/data_pipeline/data_sampling/utils.py:47
↓ 12 callers
Method
cupy2torch
(self, cupy_tensor)
deepspeed/runtime/compression/cupy.py:19
↓ 12 callers
Function
get_model_parallel_rank
Return my rank for the model parallel group.
DeepSpeedExample/megatron/mpu/initialize.py:131
↓ 12 callers
Method
gradient_accumulation_steps
(self)
deepspeed/runtime/engine.py:777
↓ 12 callers
Method
mp_size
(self)
deepspeed/autotuning/autotuner.py:247
↓ 12 callers
Method
quantize
(self, inputs, qkv=True, count=1, parallel_dim=0)
deepspeed/module_inject/replace_module.py:152
↓ 12 callers
Method
release_tensor
(self, tensor)
deepspeed/runtime/zero/contiguous_memory_allocator.py:97
↓ 12 callers
Method
write
deepspeed/ops/csrc/aio/py_lib/deepspeed_py_aio_handle.cpp:102
↓ 11 callers
Method
_io_aligned_numel
(self, numel)
deepspeed/runtime/swap_tensor/optimizer_utils.py:680
↓ 11 callers
Method
decode
(self, *inputs, **kwargs)
deepspeed/model_implementations/diffusers/vae.py:55
↓ 11 callers
Method
get_dim
Return the number of processes along the given axis. For example: >>> X = ProcessTopology(axes=['x', 'y'], dims=[2,3])
deepspeed/runtime/pipe/topology.py:98
↓ 11 callers
Function
get_model_parallel_world_size
Return world size for the model parallel group.
DeepSpeedExample/megatron/mpu/initialize.py:117
↓ 11 callers
Method
get_partition_dp_group
(self, param)
deepspeed/runtime/zero/partition_parameters.py:1515
↓ 11 callers
Method
is_gradient_accumulation_boundary
Query whether the current micro-batch is at the boundary of gradient accumulation, and thus will trigger gradient reductions and
deepspeed/runtime/engine.py:1908
↓ 11 callers
Function
is_moe_param
(param: torch.Tensor)
deepspeed/moe/utils.py:23
↓ 11 callers
Function
is_zero_param
(parameter)
deepspeed/runtime/zero/partition_parameters.py:101
↓ 11 callers
Function
iter_params
(module: Module, recurse=False)
deepspeed/runtime/zero/partitioned_param_coordinator.py:30
↓ 11 callers
Function
memory_to_string
(n, postfix="", units=None, precision=2)
deepspeed/autotuning/utils.py:416
↓ 11 callers
Method
pin_memory
(self, tensor)
deepspeed/accelerator/cpu_accelerator.py:216
↓ 11 callers
Function
print_rank_0
(message, debug=False, force=False)
deepspeed/runtime/zero/linear.py:30
↓ 11 callers
Function
quantize_weight
(weight)
deepspeed/module_inject/module_quantize.py:25
↓ 11 callers
Method
set_rng_state
(self, new_state, device_index=None)
deepspeed/accelerator/cpu_accelerator.py:63
↓ 11 callers
Method
torch2cupy
(self, tensor)
deepspeed/runtime/compression/cupy.py:16
↓ 11 callers
Method
zero_optimization_partition_weights
(self)
deepspeed/runtime/engine.py:717
↓ 10 callers
Method
_swappable_optimizer_subgroup
(self, sub_group_id)
deepspeed/runtime/zero/stage3.py:880
↓ 10 callers
Method
add_export
(self, key, var)
deepspeed/launcher/multinode_runner.py:36
↓ 10 callers
Method
cumsum
(sequence)
DeepSpeedExample/megatron/deprecated_data_utils/datasets.py:49
↓ 10 callers
Function
divide
Ensure that numerator is divisible by the denominator and return the division value.
DeepSpeedExample/megatron/core/utils.py:23
↓ 10 callers
Method
eigenvalue_enabled
(self)
deepspeed/runtime/engine.py:493
↓ 10 callers
Function
einsum
(rule, a, b)
deepspeed/moe/sharded_moe.py:116
↓ 10 callers
Function
get_timers
Return timers.
DeepSpeedExample/megatron/global_vars.py:58
↓ 10 callers
Method
metric
(self)
deepspeed/autotuning/autotuner.py:238
↓ 10 callers
Method
random
(self)
deepspeed/accelerator/cpu_accelerator.py:60
↓ 10 callers
Method
tokenize
(self, *text)
DeepSpeedExample/tools/preprocess_data.py:52
↓ 10 callers
Method
zero_optimization_stage
(self)
deepspeed/runtime/engine.py:702
↓ 9 callers
Function
_communicate
Communicate tensors between stages. Used as helper method in other communication methods that are used in megatron/schedules.py. Arguments:
DeepSpeedExample/megatron/core/pipeline_parallel/p2p_communication.py:223
↓ 9 callers
Method
add
Track the rng state.
DeepSpeedExample/megatron/mpu/random.py:160
↓ 9 callers
Method
add
(self, b)
deepspeed/runtime/sparse_tensor.py:54
↓ 9 callers
Method
all_gather_into_tensor
(self, output_tensor, input_tensor, group=None, async_op=False)
deepspeed/comm/torch.py:123
↓ 9 callers
Method
allocate
(self, num_elems, count, dtype)
deepspeed/runtime/swap_tensor/utils.py:197
↓ 9 callers
Method
backward
(ctx, *args)
DeepSpeedExample/megatron/mpu/random.py:279
↓ 9 callers
Method
eval
r
deepspeed/runtime/engine.py:1656
↓ 9 callers
Function
get_files_with_prefix
(all_files, prefix)
deepspeed/checkpoint/reshape_utils.py:17
↓ 9 callers
Function
get_inactive_params
(param_list)
deepspeed/runtime/utils.py:972
↓ 9 callers
Function
get_pipeline_model_parallel_rank
Return my rank for the pipeline model parallel group.
DeepSpeedExample/megatron/core/parallel_state.py:513
↓ 9 callers
Function
is_model_parallel_parameter
(p)
deepspeed/runtime/utils.py:78
↓ 9 callers
Method
load_state_dict
(self, sd)
deepspeed/runtime/lr_schedules.py:622
↓ 9 callers
Method
memory_breakdown
(self)
deepspeed/runtime/engine.py:597
↓ 9 callers
Method
steps_per_print
(self)
deepspeed/runtime/engine.py:806
↓ 9 callers
Method
strided_copy
(self, dst: Optional[torch.Tensor], src: Optional[torch.Tensor],
deepspeed/module_inject/replace_module.py:48
↓ 9 callers
Method
train_batch_size
(self)
deepspeed/runtime/engine.py:632
↓ 9 callers
Method
transpose_impl
(self, data)
deepspeed/module_inject/containers/base.py:278
↓ 9 callers
Method
write_events
(self, event_list)
deepspeed/monitor/monitor.py:20
↓ 8 callers
Method
ByteTensor
(self)
deepspeed/accelerator/cpu_accelerator.py:193
↓ 8 callers
Method
EncodeAsIds
encode text using text tokenizer and shift Id values for command tokens
DeepSpeedExample/megatron/deprecated_data_utils/tokenization.py:319
↓ 8 callers
Method
UseMean
deepspeed/ops/csrc/includes/normalize_layer.h:188
↓ 8 callers
Method
__getattr__
Pass through attributes defined in the model if they are not overridden by ds-engine.
deepspeed/runtime/engine.py:449
↓ 8 callers
Function
_compare
(arg_name)
DeepSpeedExample/megatron/checkpointing.py:47
↓ 8 callers
Method
_valid_stage
(self, stage_id)
deepspeed/runtime/pipe/schedule.py:83
↓ 8 callers
Function
add_to_logging
(name)
DeepSpeedExample/megatron/training.py:428
↓ 8 callers
Function
bwc_tensor_model_parallel_rank
Backwards-compatible way of querying the tensor model parallel rank from an ``mpu`` object. *Tensor* model parallelism means that tensors are
deepspeed/runtime/utils.py:88
↓ 8 callers
Function
clean_text
Remove new lines and multiple spaces and adjust end of sentence dot.
DeepSpeedExample/tasks/data_utils.py:22
↓ 8 callers
Method
data
(self)
deepspeed/runtime/utils.py:711
↓ 8 callers
Method
device_count
(self)
deepspeed/accelerator/cpu_accelerator.py:40
↓ 8 callers
Method
encode
(self, json_line)
DeepSpeedExample/tools/preprocess_data.py:78
↓ 8 callers
Method
get_coord
Return the coordinate owned by a process rank. The axes of the returned namedtuple can be directly accessed as members. For example:
deepspeed/runtime/pipe/topology.py:110
↓ 8 callers
Function
get_pipeline_model_parallel_next_rank
Return the global rank that follows the caller in the pipeline
DeepSpeedExample/megatron/core/parallel_state.py:688
↓ 8 callers
Function
get_pipeline_model_parallel_prev_rank
Return the global rank that preceeds the caller in the pipeline
DeepSpeedExample/megatron/core/parallel_state.py:696
↓ 8 callers
Function
get_pipeline_model_parallel_world_size
Return world size for the pipeline model parallel group.
DeepSpeedExample/megatron/core/parallel_state.py:462
↓ 8 callers
Method
gradient_clipping
(self)
deepspeed/runtime/engine.py:818
↓ 8 callers
Method
gradient_predivide_factor
(self)
deepspeed/runtime/engine.py:803
↓ 8 callers
Function
instrument_w_nvtx
decorator that causes an NVTX range to be recorded for the duration of the function call.
deepspeed/utils/nvtx.py:9
↓ 8 callers
Method
is_available
(self)
deepspeed/accelerator/cpu_accelerator.py:160
↓ 8 callers
Method
load_state_dict
(self, state_dict, strict=True)
DeepSpeedExample/megatron/fp16/fp16.py:84
↓ 8 callers
Method
log
Log a group of timers.
DeepSpeedExample/megatron/global_vars.py:221
↓ 8 callers
Method
match
(self, module)
deepspeed/module_inject/containers/vae.py:24
↓ 8 callers
Method
partition
(param_list=None, hierarchy=0, has_been_updated=False)
deepspeed/runtime/zero/partition_parameters.py:948
← previous
next →
101–200 of 4,989, ranked by callers