MCPcopy Create free account
hub / github.com/codefuse-ai/codefuse-devops-eval / load_model

Method load_model

src/models/internlm_model.py:80–88  ·  view source on GitHub ↗

加载模型

(self, model_path, peft_path=None, trust_remote_code=True, tensor_parallel_size=1, gpu_memory_utilization=0.25)

Source from the content-addressed store, hash-verified

78 return params
79
80 def load_model(self, model_path, peft_path=None, trust_remote_code=True, tensor_parallel_size=1, gpu_memory_utilization=0.25):
81 '''加载模型'''
82 print(model_path, peft_path, trust_remote_code)
83 self.tokenizer = AutoTokenizer.from_pretrained(model_path, trust_remote_code=trust_remote_code)
84 self.model = AutoModelForCausalLM.from_pretrained(model_path, device_map="auto", trust_remote_code=trust_remote_code).eval()
85 if peft_path:
86 self.model = PeftModel.from_pretrained(self.model, peft_path)
87
88 # self.model = LLM(model=model_path, trust_remote_code=trust_remote_code, tensor_parallel_size=tensor_parallel_size, gpu_memory_utilization=gpu_memory_utilization)

Callers 1

__init__Method · 0.95

Calls

no outgoing calls

Tested by

no test coverage detected