MCPcopy Create free account
hub / github.com/bitsandbytes-foundation/bitsandbytes / _

Function _

bitsandbytes/backends/cuda/ops.py:79–81  ·  view source on GitHub ↗
(A: torch.Tensor, B: torch.Tensor)

Source from the content-addressed store, hash-verified

77
78@register_kernel("bitsandbytes::int8_linear_matmul", "cuda")
79def _(A: torch.Tensor, B: torch.Tensor):
80 out = torch.empty((*A.shape[:-1], B.shape[0]), device=A.device, dtype=torch.int32)
81 return _int8_linear_matmul_impl(A, B, out)
82
83
84@register_kernel("bitsandbytes::int8_linear_matmul.out", "cuda")

Callers

nothing calls this directly

Calls 9

_cuda_device_ofFunction · 0.90
_get_col_absmaxFunction · 0.85
_dequant_linear_fallbackFunction · 0.85
_gemm_4bit_kernel_implFunction · 0.85
_int8_linear_matmul_implFunction · 0.70
_dequantize_4bit_implFunction · 0.70
_gemv_4bit_implFunction · 0.70
toMethod · 0.45

Tested by

no test coverage detected