MCPcopy Create free account
hub / github.com/FlashSampling/FlashSampling / SyntheticInputs

Class SyntheticInputs

src/fused_mm_sampling/testing.py:15–24  ·  view source on GitHub ↗

In the tensor-parallel case V is sharded among ranks

Source from the content-addressed store, hash-verified

13from .tp_info import TP1, TPInfo
14
15
16@dataclass
17class SyntheticInputs:
18 """In the tensor-parallel case V is sharded among ranks"""
19
20 weights: torch.Tensor # [V, D], bfloat16
21 hidden_states: torch.Tensor # [n_hidden_states, D], bfloat16
22 logits: (
23 torch.Tensor
24 ) # [n_hidden_states, V], float32 (the exact logits before bf16 quantization)
25 vocab_size: int
26 hidden_size: int
27

Callers 1

make_synthetic_inputsFunction · 0.85

Calls

no outgoing calls

Tested by 1

make_synthetic_inputsFunction · 0.68