MCPcopy Create free account

hub / github.com/bitsandbytes-foundation/bitsandbytes / functions

Functions720 in github.com/bitsandbytes-foundation/bitsandbytes

↓ 1 callersFunctionfmt_mb
(nbytes)
examples/xpu/benchmark_paged_memory.py:142
↓ 1 callersFunctionformat_alpaca
(example)
examples/xpu/paged_xpu_training.py:59
↓ 1 callersFunctionformat_alpaca
(example)
examples/cpu/cpu_training.py:63
↓ 1 callersFunctionformat_detail
Full detail view of an issue including body and comments.
agents/query_issues.py:201
↓ 1 callersFunctionformat_list_line
Compact one-line summary for list view, with date and key metadata.
agents/query_issues.py:182
↓ 1 callersFunctionformat_with_label
(label: str, value: Any)
tests/helpers.py:71
↓ 1 callersMethodforward
Forward pass to dequantize the parameter. Args: quantized_param (`torch.Tensor`): The quantized parameter tensor (from .
bitsandbytes/nn/parametrize.py:29
↓ 1 callersMethodfrom_prequantized
( cls, data: torch.Tensor, quantized_stats: dict[str, Any], requires_grad: boo
bitsandbytes/nn/modules.py:356
↓ 1 callersFunctiongemm_4bit_inference_naive_bf16
csrc/pythonInterface.cpp:46
↓ 1 callersFunctiongemm_4bit_inference_naive_fp16
csrc/pythonInterface.cpp:39
↓ 1 callersFunctiongemm_4bit_inference_naive_fp32
csrc/pythonInterface.cpp:53
↓ 1 callersFunctiongemv_4bit_inference_bf16
csrc/pythonInterface.cpp:312
↓ 1 callersFunctiongemv_4bit_inference_fp16
csrc/pythonInterface.cpp:305
↓ 1 callersFunctiongemv_4bit_inference_fp32
csrc/pythonInterface.cpp:321
↓ 1 callersFunctionget_4bit_config
()
tests/test_generation.py:12
↓ 1 callersFunctionget_args
()
examples/xpu/benchmark_paged_memory.py:22
↓ 1 callersFunctionget_args
()
examples/xpu/paged_xpu_training.py:21
↓ 1 callersFunctionget_args
()
examples/cpu/cpu_training.py:25
↓ 1 callersFunctionget_available_cuda_binary_versions
Get formatted CUDA/ROCm versions from existing library files.
bitsandbytes/cextension.py:152
↓ 1 callersFunctionget_compute_capabilities
()
bitsandbytes/cuda_specs.py:23
↓ 1 callersFunctionget_cuda_version_string
Get CUDA/HIP version as a string.
bitsandbytes/cuda_specs.py:46
↓ 1 callersFunctionget_gaudi_sw_version
Returns the installed version of Gaudi SW.
bitsandbytes/backends/utils.py:78
↓ 1 callersFunctionget_global_lut
csrc/cpu_ops.cpp:563
↓ 1 callersFunctionget_inputs
(tokenizer)
benchmarking/xpu/inference_benchmark.py:34
↓ 1 callersMethodget_lut
csrc/cpu_ops.cpp:539
↓ 1 callersFunctionget_max_threads
csrc/cpu_ops.h:62
↓ 1 callersFunctionget_model_and_tokenizer
(config)
tests/test_generation.py:24
↓ 1 callersFunctionget_native_library
Load CUDA library XOR CPU, as the latter contains a subset of symbols of the former.
bitsandbytes/cextension.py:348
↓ 1 callersMethodget_outliers
(self, weight)
bitsandbytes/utils.py:69
↓ 1 callersFunctionget_package_version
(name: str)
bitsandbytes/diagnostics/main.py:42
↓ 1 callersFunctionget_potentially_lib_path_containing_env_vars
()
bitsandbytes/diagnostics/cuda.py:85
↓ 1 callersFunctionget_rocm_gpu_arch
Get ROCm GPU architecture.
bitsandbytes/cuda_specs.py:82
↓ 1 callersFunctionget_streamer
(tokenizer)
benchmarking/xpu/inference_benchmark.py:45
↓ 1 callersFunctionget_test_dims
(min: int, max: int, *, n: int)
tests/helpers.py:67
↓ 1 callersFunctionget_torch_dtype
(name)
examples/xpu/benchmark_paged_memory.py:37
↓ 1 callersFunctionget_xpu_bnb_library_path
Get the path to the XPU native library matching the oneAPI toolchain. Prefers the versioned library (e.g. libbitsandbytes_xpu2026) matching the f
bitsandbytes/cextension.py:334
↓ 1 callersFunctiongh_graphql
Execute a GraphQL query via the gh CLI, passing the full payload as JSON on stdin.
agents/fetch_issues.py:69
↓ 1 callersFunctionigemmlt_32
csrc/pythonInterface.cpp:223
↓ 1 callersFunctionigemmlt_8
csrc/pythonInterface.cpp:230
↓ 1 callersFunctionigemmlt_8_rowscale
csrc/pythonInterface.cpp:237
↓ 1 callersMethodinit_8bit_state
(self)
bitsandbytes/nn/modules.py:1159
↓ 1 callersMethodinit_state
(self, group, p, gindex, pindex)
bitsandbytes/optim/optimizer.py:368
↓ 1 callersMethodinitialize
(self)
bitsandbytes/optim/optimizer.py:36
↓ 1 callersMethodinitialize
(self)
bitsandbytes/autograd/_functions.py:31
↓ 1 callersFunctionis_relevant_candidate_env_var
(env_var: str, value: str)
bitsandbytes/diagnostics/cuda.py:72
↓ 1 callersFunctionlinear8bit
(request)
tests/test_linear8bitlt.py:176
↓ 1 callersFunctionload_data
(path: str)
agents/query_issues.py:162
↓ 1 callersFunctionmain
()
bitsandbytes/diagnostics/main.py:70
↓ 1 callersFunctionmain
()
scripts/stale.py:30
↓ 1 callersFunctionmain
()
tests/fsdp_state_dict_save.py:63
↓ 1 callersFunctionmain
()
examples/xpu/benchmark_paged_memory.py:150
↓ 1 callersFunctionmain
()
examples/xpu/paged_xpu_training.py:290
↓ 1 callersFunctionmain
()
examples/cpu/cpu_training.py:306
↓ 1 callersFunctionmain
()
agents/fetch_issues.py:223
↓ 1 callersFunctionmain
()
agents/query_issues.py:537
↓ 1 callersFunctionmake_batch
Create a random batch of input_ids and labels.
examples/xpu/benchmark_paged_memory.py:67
↓ 1 callersFunctionmeasure_training
Run a few training steps and return peak GPU memory in bytes.
examples/xpu/benchmark_paged_memory.py:82
↓ 1 callersFunctionneon_dequant_4bit_16values
Vectorized 4-bit dequantization: process 8 packed bytes = 16 output values Each byte contains two 4-bit values: high nibble first, low nibble second
csrc/cpu_ops.cpp:89
↓ 1 callersFunctionneon_fp4_lut
ARM NEON FP4 lookup table
csrc/cpu_ops.cpp:72
↓ 1 callersFunctionneon_nf4_lut
ARM NEON NF4 lookup table (16 float values indexed by 4-bit code)
csrc/cpu_ops.cpp:44
↓ 1 callersFunctionneon_norm_to_lut_index_x4
NEON-optimized norm_to_lut_index for 4 float values at a time Maps [-1, 1] → [0, 65535]
csrc/cpu_ops.cpp:221
↓ 1 callersFunctionnorm_to_lut_index
Convert a normalized value in [-1, 1] to LUT index [0, 65535]
csrc/cpu_ops.cpp:569
↓ 1 callersFunctionone_step
(max_unorm)
tests/test_optim.py:703
↓ 1 callersFunctionoptimizer_update_8bit_blockwise_impl
( optimizer_name: str, g: torch.Tensor, p: torch.Tensor, state1: torch.Tensor, state2: Opt
bitsandbytes/backends/triton/kernels_optim.py:1099
↓ 1 callersFunctionpack_dict_to_tensor
Pack a dictionary into a torch tensor for storing quant_state items in state_dict. Parameters: - source_dict: The dictionary to be packe
bitsandbytes/utils.py:166
↓ 1 callersFunctionparallel_2d
csrc/cpu_ops.h:78
↓ 1 callersFunctionparse_args
()
benchmarking/inference_benchmark.py:83
↓ 1 callersFunctionparse_arguments
()
benchmarking/xpu/inference_benchmark.py:82
↓ 1 callersFunctionparse_cuda_version
Convert a raw version code string (e.g. '118', '713') to a dotted version (e.g. '11.8', '7.13').
bitsandbytes/cextension.py:159
↓ 1 callersMethodprefetch_state
(self, p)
bitsandbytes/optim/optimizer.py:392
↓ 1 callersFunctionprint_diagnostics
(cuda_specs: CUDASpecs)
bitsandbytes/diagnostics/cuda.py:165
↓ 1 callersFunctionprint_header
(txt: str, width: int = HEADER_WIDTH, filler: str = "=")
bitsandbytes/diagnostics/utils.py:6
↓ 1 callersMethodprint_report
(self)
benchmarking/xpu/inference_benchmark.py:68
↓ 1 callersFunctionquantizeBlockwise_bf16
csrc/pythonInterface.cpp:139
↓ 1 callersFunctionquantizeBlockwise_bf16_fp4
csrc/pythonInterface.cpp:145
↓ 1 callersFunctionquantizeBlockwise_bf16_nf4
csrc/pythonInterface.cpp:151
↓ 1 callersFunctionquantizeBlockwise_fp16
csrc/pythonInterface.cpp:127
↓ 1 callersFunctionquantizeBlockwise_fp16_fp4
csrc/pythonInterface.cpp:131
↓ 1 callersFunctionquantizeBlockwise_fp16_nf4
csrc/pythonInterface.cpp:135
↓ 1 callersFunctionquantizeBlockwise_fp32
csrc/pythonInterface.cpp:157
↓ 1 callersFunctionquantizeBlockwise_fp32_fp4
csrc/pythonInterface.cpp:161
↓ 1 callersFunctionquantizeBlockwise_fp32_nf4
csrc/pythonInterface.cpp:165
↓ 1 callersFunctionquantize_blockwise_triton
(A, code, blocksize, absmax=None, out=None)
bitsandbytes/backends/triton/kernels_8bit_quant.py:107
↓ 1 callersFunctionquantize_cpu
csrc/cpu_ops.cpp:667
↓ 1 callersFunctionquantize_cpu_bf16
csrc/cpu_ops.cpp:671
↓ 1 callersFunctionquantize_cpu_fp16
csrc/cpu_ops.cpp:675
↓ 1 callersMethodquantize_weight
(self, w, outlier_idx)
bitsandbytes/nn/modules.py:1206
↓ 1 callersMethodreset_parameters
(self)
bitsandbytes/nn/modules.py:101
↓ 1 callersFunctionrun_benchmark
(args, config, batch_size)
benchmarking/inference_benchmark.py:120
↓ 1 callersFunctionrun_compare
Compare paged_adamw vs adamw numerically.
examples/xpu/paged_xpu_training.py:250
↓ 1 callersFunctionrun_compare
Compare bnb AdamW vs torch AdamW on CPU to verify correctness.
examples/cpu/cpu_training.py:193
↓ 1 callersFunctionrun_single
Train with one optimizer and report results.
examples/xpu/paged_xpu_training.py:146
↓ 1 callersFunctionrun_single
(args)
examples/cpu/cpu_training.py:152
↓ 1 callersFunctionrun_with_trainer
Train using HuggingFace Trainer with a bnb optimizer.
examples/xpu/paged_xpu_training.py:184
↓ 1 callersFunctionrun_with_trainer
Train using HuggingFace Trainer with a bnb optimizer on CPU.
examples/cpu/cpu_training.py:239
↓ 1 callersFunctionsanity_check
()
bitsandbytes/diagnostics/main.py:27
↓ 1 callersMethodset_compute_type
(self, x)
bitsandbytes/nn/modules.py:575
↓ 1 callersFunctionset_fp4_lut
csrc/cpu_ops.cpp:271
↓ 1 callersFunctionset_nf4_lut
csrc/cpu_ops.cpp:263
↓ 1 callersFunctionshow_environment
Simple utility to print out environment information.
bitsandbytes/diagnostics/main.py:50
← previousnext →201–300 of 720, ranked by callers