Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/ZJU-REAL/SkillZero
/ functions
Functions
3,531 in github.com/ZJU-REAL/SkillZero
⨍
Functions
3,531
◇
Types & classes
522
↳
Endpoints
90
↓ 5 callers
Function
new_PseudoActionEffect
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/memory.c:677
↓ 5 callers
Function
new_ef
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/relax.c:1290
↓ 5 callers
Function
parallel_put
Puts a list of data into the Ray object store in parallel using a thread pool. Args: data_list (List[Any]): A list of Python objects
verl/utils/ray_utils.py:49
↓ 5 callers
Function
print_NormOperator
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/output.c:607
↓ 5 callers
Function
print_Operator
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/output.c:534
↓ 5 callers
Function
print_model_size
(model: nn.Module, name: str = None)
verl/utils/model.py:154
↓ 5 callers
Function
print_type
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/output.c:1220
↓ 5 callers
Function
read_html_template
(path)
agent_system/environments/env_package/webshop/webshop/web_agent_site/engine/engine.py:111
↓ 5 callers
Function
replace_var_with_const_in_exp
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/inst_pre.c:2132
↓ 5 callers
Method
reset
(self, idx=None)
agent_system/environments/env_package/webshop/webshop/baseline_models/env.py:211
↓ 5 callers
Function
reset_fixpoint
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/relax.c:1346
↓ 5 callers
Method
run
(self, config)
recipe/spin/main_spin.py:44
↓ 5 callers
Method
save
(self)
agent_system/environments/env_package/webshop/webshop/baseline_models/agent.py:158
↓ 5 callers
Method
save_checkpoint
(self, local_path, hdfs_path=None, global_step=0, max_ckpt_to_keep=None)
recipe/prime/prime_fsdp_workers.py:341
↓ 5 callers
Function
slice_input_tensor
(x: Tensor, dim: int, padding: bool = True, group: ProcessGroup = None)
verl/utils/ulysses.py:117
↓ 5 callers
Function
source_to_dest
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/search.c:2351
↓ 5 callers
Method
step
(self, text_actions: List[str])
agent_system/environments/env_manager.py:142
↓ 5 callers
Method
step
(self, action_str)
agent_system/environments/env_package/alfworld/alfworld/agents/controller/mrcnn.py:376
↓ 5 callers
Method
store
Stores a new batch of records into memory.
agent_system/memory/base.py:43
↓ 5 callers
Method
submit_chat_completions
Submit a chat completion request to chat scheduler and wait until it is done. To submit multiple requests in parallel, please use `generate_se
verl/workers/rollout/async_server.py:308
↓ 5 callers
Function
sum_hand
(hand)
agent_system/environments/env_package/gym_cards/gym-cards/gym_cards/envs/blackjack.py:59
↓ 5 callers
Function
torch_to_numpy
(tensor, is_object=False)
agent_system/multi_turn_rollout/utils.py:39
↓ 5 callers
Method
train
Tell the agent that it's training phase.
agent_system/environments/env_package/alfworld/alfworld/agents/agent/base_agent.py:233
↓ 5 callers
Method
update
(self, **kwargs)
agent_system/environments/env_package/alfworld/alfworld/agents/detector/utils.py:152
↓ 5 callers
Method
update_actor
(self, data: DataProto)
verl/workers/fsdp_workers.py:601
↓ 4 callers
Function
GetGenerationQValue
Shape: - y_pred: batch x time x vocab - y_true: batch x time - mask: batch x time
agent_system/environments/env_package/alfworld/alfworld/agents/modules/layers.py:49
↓ 4 callers
Function
NegativeLogLoss
Shape: - y_pred: batch x time x vocab - y_true: batch x time - mask: batch x time
agent_system/environments/env_package/alfworld/alfworld/agents/modules/layers.py:28
↓ 4 callers
Method
__init__
(self)
verl/single_controller/ray/base.py:711
↓ 4 callers
Method
__init__
(self, dim, max_position_embeddings=2048, base=10000, device=None)
verl/models/llama/megatron/layers/parallel_attention.py:39
↓ 4 callers
Method
_balance_batch
Reorder the data on single controller such that each dp rank gets similar total tokens
recipe/spin/spin_trainer.py:902
↓ 4 callers
Function
_broadcast_tp_shard_tensor
broadcast tensor in tp shards across mp_group
verl/models/qwen2/megatron/checkpoint_utils/qwen2_saver.py:158
↓ 4 callers
Function
_broadcast_tp_shard_tensor
broadcast tensor in tp shards across mp_group
verl/models/mcore/saver.py:168
↓ 4 callers
Function
_broadcast_tp_shard_tensor
broadcast tensor in tp shards across mp_group
verl/models/llama/megatron/checkpoint_utils/llama_saver.py:158
↓ 4 callers
Method
_build_rollout
(self, trust_remote_code=False)
verl/workers/fsdp_workers.py:396
↓ 4 callers
Function
_check_dispatch_mode
(dispatch_mode)
verl/single_controller/base/decorator.py:478
↓ 4 callers
Method
_compute_loss_and_backward
Compute loss with optional sequence parallelism and remove padding features
verl/trainer/fsdp_sft_trainer.py:303
↓ 4 callers
Function
_compute_response_info
Computes information about prompts and responses from a batch. This is an internal helper function that extracts masks and lengths for p
verl/trainer/ppo/metric_utils.py:49
↓ 4 callers
Method
_convert_single
Convert a single trajectory text to an image with optimized packing. Args: trajectory_text: Trajectory text stri
agentocr/ocrtool.py:394
↓ 4 callers
Function
_get_cached_font
Get or create a cached font object to avoid repeated font loading. This provides significant speedup for repeated text rendering.
agentocr/utils.py:77
↓ 4 callers
Method
_get_config
Get configuration dictionary, merging defaults with overrides. Args: **override_kwargs: Parameters to override
agentocr/ocrtool.py:479
↓ 4 callers
Function
_get_cpu_tensor
(tensor: torch.Tensor)
verl/models/qwen2/megatron/checkpoint_utils/qwen2_saver.py:112
↓ 4 callers
Function
_get_cpu_tensor
(tensor: torch.Tensor)
verl/models/mcore/saver.py:122
↓ 4 callers
Function
_get_cpu_tensor
(tensor: torch.Tensor)
verl/models/llama/megatron/checkpoint_utils/llama_saver.py:112
↓ 4 callers
Function
_get_gpt_model
(model)
verl/models/qwen2/megatron/checkpoint_utils/qwen2_loader_depracated.py:60
↓ 4 callers
Function
_get_gpt_model
(model)
verl/models/qwen2/megatron/checkpoint_utils/qwen2_loader.py:60
↓ 4 callers
Function
_get_gpt_model
(model)
verl/models/mcore/loader.py:63
↓ 4 callers
Function
_get_gpt_model
(model)
verl/models/llama/megatron/checkpoint_utils/llama_loader.py:62
↓ 4 callers
Function
_get_gpt_model
(model)
verl/models/llama/megatron/checkpoint_utils/llama_loader_depracated.py:62
↓ 4 callers
Function
_is_non_local
(path: str)
verl/utils/hdfs_io.py:148
↓ 4 callers
Function
_iter_opts
(opt)
verl/utils/megatron_utils.py:312
↓ 4 callers
Method
_load_checkpoint
(self)
recipe/spin/spin_trainer.py:849
↓ 4 callers
Method
_parse_html
Returns web request result wrapped in BeautifulSoup object Arguments: url (`str`): If no url or html is provided, use the cu
agent_system/environments/env_package/webshop/webshop/web_agent_site/envs/web_agent_text_env.py:181
↓ 4 callers
Function
_repeat_interleave
(value: Union[torch.Tensor, np.ndarray], repeats: int)
verl/workers/rollout/vllm_rollout/vllm_rollout_spmd.py:87
↓ 4 callers
Method
add_assistant_message
Currently, we only support chatml format.
verl/workers/rollout/schemas.py:115
↓ 4 callers
Function
add_to_ehc_space
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/search.c:321
↓ 4 callers
Function
apply_kl_penalty
Apply KL penalty to the token-level rewards. This function computes the KL divergence between the reference policy and current policy, then a
verl/trainer/ppo/ray_trainer.py:152
↓ 4 callers
Function
bootstrap_metric
Performs bootstrap resampling to estimate statistics of metrics. This function uses bootstrap resampling to estimate the mean and standard d
verl/trainer/ppo/metric_utils.py:297
↓ 4 callers
Function
box_displacement_score
Calculates the sum of all Manhattan distances, between the boxes and their origin box targets. :param box_mapping: :return:
agent_system/environments/env_package/sokoban/sokoban/room_utils.py:551
↓ 4 callers
Method
build_text_obs
This function builds the text observation for the agent.
agent_system/environments/env_manager.py:606
↓ 4 callers
Method
choose_random_action
Select an action randomly.
agent_system/environments/env_package/alfworld/alfworld/agents/agent/text_dqn_agent.py:29
↓ 4 callers
Function
compute_data_metrics
Computes various metrics from a batch of data for PPO training. This function calculates metrics related to scores, rewards, advantages, ret
verl/trainer/ppo/metric_utils.py:79
↓ 4 callers
Function
compute_reward
We compute dense reward here so that we can directly train RL without SFT
tests/e2e/envs/digit_completion/task.py:135
↓ 4 callers
Method
compute_rm_score
(self, data: DataProto)
recipe/spin/fsdp_workers.py:508
↓ 4 callers
Function
compute_timing_metrics
Computes timing metrics for different processing stages in PPO training. This function calculates both raw timing metrics (in seconds) a
verl/trainer/ppo/metric_utils.py:222
↓ 4 callers
Method
compute_values
(self, data: DataProto)
verl/workers/fsdp_workers.py:1047
↓ 4 callers
Function
create_colocated_worker_cls
This function should return a class instance that delegates the calls to every cls in cls_dict
verl/single_controller/ray/base.py:692
↓ 4 callers
Function
create_sft_dataset
Create a dataset.
verl/trainer/fsdp_sft_trainer.py:586
↓ 4 callers
Function
default_compute_score
Compute the score for a given solution based on the data source. Args: data_source (str): The source dataset identifier which determines
verl/utils/reward_score/__init__.py:19
↓ 4 callers
Method
encode
(self, query_list: List[str], is_query=True)
examples/search/retriever/retrieval_server.py:70
↓ 4 callers
Method
execute_rank_zero_async
Execute a method on rank zero worker asynchronously. Args: method_name: Name of the method to execute *args: Position
verl/single_controller/ray/base.py:522
↓ 4 callers
Method
forward
( self, hidden_states: torch.FloatTensor, vocab_weights: torch.FloatTensor, in
verl/utils/experimental/torch_functional.py:203
↓ 4 callers
Function
free_FactList
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/memory.c:922
↓ 4 callers
Function
free_TokenList
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/memory.c:906
↓ 4 callers
Function
fsdp2_load_full_state_dict
Loads the full state dict (could be only on rank 0) into the sharded model. This is done by broadcasting the parameters from rank 0 to all ot
verl/utils/fsdp_utils.py:394
↓ 4 callers
Function
get_1P_and_H
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/relax.c:242
↓ 4 callers
Method
get_any
(self)
agent_system/environments/env_package/alfworld/alfworld/gen/utils/py_util.py:44
↓ 4 callers
Function
get_constant_schedule_with_warmup
Create a constant LR schedule with a linear warmup phase. Args: optimizer (Optimizer): Wrapped optimizer. num_warmup_steps (
verl/utils/torch_functional.py:505
↓ 4 callers
Function
get_cosine_schedule_with_warmup
Create a schedule with a learning rate that decreases following the values of the cosine function between the initial lr set in the optimizer
verl/utils/torch_functional.py:461
↓ 4 callers
Method
get_current_plan
(self, force_update=False)
agent_system/environments/env_package/alfworld/alfworld/gen/game_states/task_game_state.py:44
↓ 4 callers
Method
get_page_name
Determine which page (i.e. item_page, search_results) the given URL is pointing at
agent_system/environments/env_package/webshop/webshop/web_agent_site/envs/web_agent_text_env.py:604
↓ 4 callers
Function
get_predefined_dispatch_fn
(dispatch_mode)
verl/single_controller/base/decorator.py:443
↓ 4 callers
Method
get_resource_pool
Get the resource pool of the worker_cls
verl/trainer/ppo/ray_trainer.py:120
↓ 4 callers
Function
init_execution_pool
(num_workers: int, enable_global_rate_limit=True, rate_limit=10, mode: PoolMode = PoolMode.ThreadMode)
verl/tools/sandbox_fusion_tools.py:88
↓ 4 callers
Function
init_mcore_model
Initialize a Mcore model. Args: tfconfig: The transformer config. hf_config: The HuggingFace config. pre_process: Op
verl/models/mcore/registry.py:197
↓ 4 callers
Method
init_model
(self)
recipe/spin/fsdp_workers.py:369
↓ 4 callers
Method
init_workers
Init resource pool and worker group
recipe/spin/spin_trainer.py:730
↓ 4 callers
Method
initialize
(self, **kwargs)
verl/models/mcore/model_initializer.py:141
↓ 4 callers
Function
instantiate_exp
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/inst_hard.c:814
↓ 4 callers
Function
is_artificial_fluent
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/expressions.c:2201
↓ 4 callers
Function
is_digit
(s)
verl/utils/reward_score/prime_math/grader.py:110
↓ 4 callers
Function
is_dnf
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/inst_pre.c:3252
↓ 4 callers
Function
is_wff
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/parse.c:1051
↓ 4 callers
Function
load_fsdp_optimizer
(optimizer, device_id)
verl/utils/fsdp_utils.py:200
↓ 4 callers
Function
load_mcore_dist_weights
(parallel_model, dist_weight_path, is_value_model=False)
verl/utils/model.py:398
↓ 4 callers
Function
load_megatron_gptmodel_weights
Load weights for mcore GPT model.
verl/utils/model.py:348
↓ 4 callers
Function
make_Fluent
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/inst_pre.c:719
↓ 4 callers
Method
make_iterator
r"""Make an iterator from the DataProto. This is built upon that TensorDict can be used as a normal Pytorch dataset. See https://pytorch.org/t
verl/protocol.py:685
↓ 4 callers
Function
new_Fact
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/memory.c:270
↓ 4 callers
Function
new_Literal
agent_system/environments/env_package/alfworld/alfworld/gen/ff_planner/memory.c:396
↓ 4 callers
Function
normalize_answer
(s)
verl/utils/reward_score/search_r1_like_qa_em.py:23
← previous
next →
301–400 of 3,531, ranked by callers