Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/antirez/llama.cpp-deepseek-v4-flash
/ functions
Functions
13,473 in github.com/antirez/llama.cpp-deepseek-v4-flash
⨍
Functions
13,473
◇
Types & classes
3,460
↳
Endpoints
3
↓ 8 callers
Function
ggml_node_get_use_count
ggml/src/ggml-impl.h:629
↓ 8 callers
Function
ggml_op_is_empty
ggml/src/ggml-impl.h:90
↓ 8 callers
Function
ggml_opt_result_reset
ggml/src/ggml-opt.cpp:647
↓ 8 callers
Function
ggml_rope_impl
ggml/src/ggml.c:4125
↓ 8 callers
Function
ggml_status_to_string
ggml/src/ggml.c:432
↓ 8 callers
Function
ggml_swiglu_split
ggml/src/ggml.c:3027
↓ 8 callers
Function
ggml_ue4m3_to_fp32
UE4M3: unsigned, 4 exp bits (bias=7), 3 mantissa bits Returns value * 0.5 to match kvalues_mxfp4 convention (kvalues = 2 * E2M1_float)
ggml/src/ggml-impl.h:502
↓ 8 callers
Function
ggml_vec_dot_f32
ggml/src/ggml-cpu/vec.cpp:11
↓ 8 callers
Function
ggml_vec_set_f32
ggml/src/ggml-cpu/vec.h:89
↓ 8 callers
Function
ggml_vk_cpy_to_contiguous
ggml/src/ggml-vulkan/ggml-vulkan.cpp:7412
↓ 8 callers
Function
ggml_vqtbl1q_u8
NOTE: not tested
ggml/src/ggml-cpu/ggml-cpu-impl.h:266
↓ 8 callers
Function
gguf_get_key
ggml/src/gguf.cpp:904
↓ 8 callers
Function
gguf_get_meta_size
ggml/src/gguf.cpp:1545
↓ 8 callers
Function
gguf_type_name
ggml/src/gguf.cpp:867
↓ 8 callers
Function
hex_l2fetch
ggml/src/ggml-hexagon/htp/hex-utils.h:77
↓ 8 callers
Function
hvx_vec_exp_f32
ggml/src/ggml-hexagon/htp/hvx-exp.h:23
↓ 8 callers
Method
init
src/llama-sampler.cpp:543
↓ 8 callers
Method
insert
common/json-schema-to-grammar.cpp:573
↓ 8 callers
Method
is_running
tools/server/server-models.h:73
↓ 8 callers
Method
is_static
ggml/src/ggml-openvino/openvino/node_context.h:95
↓ 8 callers
Function
llama_chat_apply_template
src/llama.cpp:447
↓ 8 callers
Function
llama_get_logits
src/llama-context.cpp:3123
↓ 8 callers
Function
llama_grammar_init_impl
src/llama-grammar.cpp:1126
↓ 8 callers
Function
llama_memory_seq_pos_max
src/llama-context.cpp:3335
↓ 8 callers
Function
llama_model_chat_template
src/llama-model.cpp:9618
↓ 8 callers
Function
llama_sampler_init_greedy
src/llama-sampler.cpp:1011
↓ 8 callers
Function
llama_sampler_init_top_p
src/llama-sampler.cpp:1513
↓ 8 callers
Function
llama_time_us
src/llama.cpp:110
↓ 8 callers
Function
lsx_packs_w
ggml/src/ggml-cpu/arch/loongarch/quants.c:28
↓ 8 callers
Function
lsx_shuffle_b
ggml/src/ggml-cpu/arch/loongarch/quants.c:68
↓ 8 callers
Method
n_gqa
src/llama-hparams.cpp:54
↓ 8 callers
Function
parse_bool
* @brief Verify whether the environment variable is a valid value. */
ggml/src/ggml-cann/ggml-cann.cpp:116
↓ 8 callers
Method
pos_get
note: call only if the cell is not empty
src/llama-kv-cells.h:363
↓ 8 callers
Method
post_task
tools/server/server-queue.cpp:343
↓ 8 callers
Function
quantize_row_q8_1_ref
reference implementation for deterministic creation of model files
ggml/src/ggml-quants.c:260
↓ 8 callers
Function
random_string
tools/server/server-common.cpp:64
↓ 8 callers
Method
read_raw
src/llama-mmap.cpp:419
↓ 8 callers
Function
replace_all
src/llama-impl.cpp:69
↓ 8 callers
Method
reset
src/llama-kv-cells.h:34
↓ 8 callers
Function
serialize_tensor
ggml/src/ggml-rpc/ggml-rpc.cpp:424
↓ 8 callers
Method
set_input_k_idxs
src/llama-kv-cache.cpp:1354
↓ 8 callers
Method
set_input_kq_mask
src/llama-kv-cache.cpp:1608
↓ 8 callers
Function
stripLegacyContextMarkers
(content: string)
tools/server/webui/tests/unit/agentic-strip.test.ts:11
↓ 8 callers
Function
summarize
Print a tensor in llama.cpp debug style. Supports: - 2D tensors (seq, hidden) - 3D tensors (batch, seq, hidden) - 4D tensors (ba
examples/model-conversion/scripts/utils/common.py:30
↓ 8 callers
Method
takeUnless
examples/llama.android/lib/src/main/java/com/arm/aichat/internal/gguf/GgufMetadataReaderImpl.kt:479
↓ 8 callers
Method
text_to_token
src/llama-vocab.cpp:3625
↓ 8 callers
Function
tokenizer
(message)
tools/server/bench/script.js:33
↓ 8 callers
Method
tool
Low-level tag methods
common/chat-peg-parser.h:79
↓ 8 callers
Method
tool_args
common/chat-peg-parser.h:84
↓ 8 callers
Method
tool_close
common/chat-peg-parser.h:81
↓ 8 callers
Function
trim
============================================================== Utility: trim ==============================================================
ggml/src/ggml-webgpu/pre_wgsl.hpp:27
↓ 8 callers
Function
trim_leading_whitespace
common/chat-auto-parser-helpers.cpp:33
↓ 8 callers
Function
unpack_mxfp4_quants
ggml/src/ggml-hexagon/ggml-hexagon.cpp:1038
↓ 8 callers
Function
unpack_q4_0_quants
ggml/src/ggml-hexagon/ggml-hexagon.cpp:351
↓ 8 callers
Function
unpack_q8_0_quants
ggml/src/ggml-hexagon/ggml-hexagon.cpp:693
↓ 8 callers
Function
until_common_prefix
Returns the prefix of `full` up until the first occurrence of the common prefix of `left` and `right`
common/chat-auto-parser-helpers.cpp:210
↓ 8 callers
Method
updateSession
(conversationId: string, update: Partial<AgenticSession>)
tools/server/webui/src/lib/stores/agentic.svelte.ts:144
↓ 8 callers
Method
write_tensor
src/llama-context.cpp:2357
↓ 8 callers
Method
write_tensors_to_file
(self, *, progress: bool = False)
gguf-py/gguf/gguf_writer.py:438
↓ 7 callers
Function
SHA1Update
examples/gguf-hash/deps/sha1/sha1.c:198
↓ 7 callers
Function
XXH64_avalanche
! @copydoc XXH32_avalanche */
examples/gguf-hash/deps/xxhash/xxhash.h:3435
↓ 7 callers
Function
XXH_mult32to64
! * @brief Calculates a 32-bit to 64-bit long multiply. * * Implemented as a macro. * * Wraps `__emulu` on MSVC x86 because it tends to call `__a
examples/gguf-hash/deps/xxhash/xxhash.h:4328
↓ 7 callers
Function
XXH_swap32
examples/gguf-hash/deps/xxhash/xxhash.h:2776
↓ 7 callers
Method
_try_set_pooling_type
(self)
convert_hf_to_gguf.py:1862
↓ 7 callers
Function
aclnn_muls
* @brief Multiplies elements of a tensor by a scalar value, optionally * in-place. * * This function multiplies each element of the source tensor `
ggml/src/ggml-cann/aclnn_ops.cpp:265
↓ 7 callers
Method
add_chat_template
(self, value: str | Sequence[Mapping[str, str]])
gguf-py/gguf/gguf_writer.py:1105
↓ 7 callers
Method
add_expert_gating_func
(self, value: ExpertGatingFuncType)
gguf-py/gguf/gguf_writer.py:847
↓ 7 callers
Method
add_mask
src/models/deepseek4.cpp:59
↓ 7 callers
Method
add_ssm_conv_kernel
(self, value: int)
gguf-py/gguf/gguf_writer.py:1027
↓ 7 callers
Method
add_token_merges
(self, merges: Sequence[str] | Sequence[bytes] | Sequence[bytearray])
gguf-py/gguf/gguf_writer.py:1057
↓ 7 callers
Function
apir_decode_ggml_buffer
ggml/src/ggml-virtgpu/backend/shared/apir_cs_ggml.h:100
↓ 7 callers
Function
apir_encode_bool_t
ggml/src/ggml-virtgpu/backend/shared/apir_cs.h:340
↓ 7 callers
Function
apir_load_library_error
ggml/src/ggml-virtgpu/backend/shared/api_remoting.h:59
↓ 7 callers
Method
apply
src/llama-adapter.cpp:94
↓ 7 callers
Function
atomic_fetch_add_explicit
ggml/src/ggml-cpu/ggml-cpu.c:136
↓ 7 callers
Function
atomic_store
ggml/src/ggml-cpu/ggml-cpu.c:119
↓ 7 callers
Function
atomic_store_explicit
ggml/src/ggml-cpu/ggml-cpu.c:122
↓ 7 callers
Function
byte_level_permute
ggml/src/ggml-sycl/dpct/helper.hpp:2968
↓ 7 callers
Function
check_max_size
tests/test-alloc.cpp:195
↓ 7 callers
Function
common_log_set_verbosity_thold
common/log.cpp:30
↓ 7 callers
Function
common_ngram_cache_update
common/ngram-cache.cpp:12
↓ 7 callers
Function
common_perf_print
common/sampling.cpp:475
↓ 7 callers
Function
cosine_similarity
(a, b=None)
examples/model-conversion/scripts/utils/semantic_check.py:14
↓ 7 callers
Method
dev_layer
src/llama-model.cpp:8495
↓ 7 callers
Function
div_up
ggml/src/ggml-cpu/amx/common.h:34
↓ 7 callers
Method
drain
(self)
tools/server/tests/unit/test_kv_keep_only_active.py:12
↓ 7 callers
Function
dsv4_apply_rope_tail
src/models/deepseek4.cpp:297
↓ 7 callers
Function
dsv4_cache_offset
src/llama-memory-hybrid-iswa.cpp:74
↓ 7 callers
Function
dsv4_make_state_layout
src/models/deepseek4.cpp:199
↓ 7 callers
Method
empty
Check if a grammar is set
common/common.h:192
↓ 7 callers
Method
encode
(self, text: str)
tests/test-tokenizer-random.py:121
↓ 7 callers
Function
extractJsonRpcMethods
(body: BodyInit | null | undefined)
tools/server/webui/src/lib/utils/request-helpers.ts:93
↓ 7 callers
Function
filterByLeafNodeId
( messages: readonly DatabaseMessage[], leafNodeId: string, includeRoot: boolean = false )
tools/server/webui/src/lib/utils/branching.ts:40
↓ 7 callers
Function
format_codepoint
Helper to format a code point as a readable string
tests/test-chat.cpp:130
↓ 7 callers
Function
fp32_from_bits
ggml/src/ggml-impl.h:366
↓ 7 callers
Method
getMessageByIdWithRole
( messageId: string, expectedRole?: MessageRole )
tools/server/webui/src/lib/stores/chat.svelte.ts:304
↓ 7 callers
Method
getServerDefaults
* Helper method to get server defaults with null safety * Centralizes the pattern of getting and extracting server defaults
tools/server/webui/src/lib/stores/settings.svelte.ts:76
↓ 7 callers
Method
getServers
()
tools/server/webui/src/lib/stores/mcp.svelte.ts:375
↓ 7 callers
Function
get_gguf_split_info
common/download.cpp:509
↓ 7 callers
Function
get_kv_str
tools/export-lora/export-lora.cpp:22
← previous
next →
1,001–1,100 of 13,473, ranked by callers