Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/Maratyszcza/NNPACK
/ functions
Functions
873 in github.com/Maratyszcza/NNPACK
⨍
Functions
873
◇
Types & classes
270
↓ 900 callers
Method
errorLimit
test/testers/relu.h:70
↓ 452 callers
Method
testInference
test/testers/convolution.h:372
↓ 441 callers
Function
cimagf
include/nnpack/complex.h:16
↓ 441 callers
Function
crealf
include/nnpack/complex.h:12
↓ 321 callers
Method
inputSize
test/testers/pooling.h:118
↓ 321 callers
Method
iterations
test/testers/relu.h:61
↓ 288 callers
Method
batchSize
test/testers/relu.h:94
↓ 258 callers
Function
butterflyfc
src/ref/fft/complex.h:18
↓ 246 callers
Function
vst1q_f32_aligned
include/nnpack/arm_neon.h:20
↓ 244 callers
Function
vld1q_f32_aligned
include/nnpack/arm_neon.h:16
↓ 228 callers
Function
vmuladdq_f32
include/nnpack/arm_neon.h:65
↓ 226 callers
Method
testOutput
test/testers/relu.h:121
↓ 198 callers
Function
scalar_butterfly
src/scalar/butterfly.h:11
↓ 101 callers
Method
kernelSize
test/testers/convolution.h:149
↓ 98 callers
Function
psimd_butterfly_f32
src/psimd/butterfly.h:6
↓ 98 callers
Function
vmuladd_f32
include/nnpack/arm_neon.h:113
↓ 96 callers
Function
conv3
* VGG model A conv3 layer: * input channels = 128 * output channels = 256 * input size = 56x56 * implicit padding = 1 * k
test/models/vgg-a.h:80
↓ 96 callers
Function
conv4
* VGG model A conv4 layer: * input channels = 256 * output channels = 256 * input size = 56x56 * implicit padding = 1 * k
test/models/vgg-a.h:110
↓ 96 callers
Function
conv5
* VGG model A conv5 layer: * input channels = 256 * output channels = 512 * input size = 28x28 * implicit padding = 1 * k
test/models/vgg-a.h:128
↓ 89 callers
Method
testInputGradient
test/testers/relu.h:180
↓ 84 callers
Method
simdWidth
test/testers/fourier.h:41
↓ 81 callers
Function
min
include/nnpack/utils.h:21
↓ 81 callers
Method
outputChannels
test/testers/convolution.h:121
↓ 79 callers
Method
fftSize
test/testers/fourier.h:32
↓ 74 callers
Method
accumulateC
test/testers/gemm-ukernel.h:87
↓ 72 callers
Function
conv2
* VGG model A conv2 layer: * input channels = 64 * output channels = 128 * input size = 112x112 * implicit padding = 1 *
test/models/vgg-a.h:50
↓ 72 callers
Method
inputChannels
test/testers/convolution.h:112
↓ 72 callers
Function
scalar_swap
src/scalar/butterfly.h:4
↓ 72 callers
Function
vld1q_f32_f16
include/nnpack/arm_neon.h:26
↓ 69 callers
Method
testKernelGradient
test/testers/convolution.h:319
↓ 64 callers
Method
kernelHeight
test/testers/convolution.h:153
↓ 64 callers
Method
kernelWidth
test/testers/convolution.h:157
↓ 63 callers
Function
vst1q_f16_f32
include/nnpack/arm_neon.h:35
↓ 59 callers
Function
butterfly
(a, b, negate_a=False, negate_b=False, scale_a=None, scale_b=None, negate_out_b=False, writeback=True)
src/x86_64-fma/common.py:65
↓ 58 callers
Method
multithreading
test/testers/relu.h:85
↓ 54 callers
Method
testSGEMM
test/testers/gemm-ukernel.h:290
↓ 51 callers
Method
inputPadding
test/testers/pooling.h:185
↓ 48 callers
Method
channels
test/testers/relu.h:103
↓ 42 callers
Function
scalar_butterfly_with_negated_b
src/scalar/butterfly.h:25
↓ 42 callers
Function
swapfc
src/ref/fft/complex.h:25
↓ 40 callers
Function
conv1
* VGG model A conv1 layer: * input channels = 3 * output channels = 64 * input size = 224x224 * implicit padding = 1 * ke
test/models/vgg-a.h:20
↓ 40 callers
Method
mr
test/testers/gemm-ukernel.h:37
↓ 40 callers
Method
nr
test/testers/gemm-ukernel.h:46
↓ 40 callers
Method
poolingSize
test/testers/pooling.h:136
↓ 40 callers
Method
poolingStride
test/testers/pooling.h:154
↓ 39 callers
Function
nnp_initialize
src/init.c:601
↓ 37 callers
Method
kc
test/testers/gemm-ukernel.h:55
↓ 36 callers
Function
vmuladd_lane0_f32
include/nnpack/arm_neon.h:129
↓ 36 callers
Function
vmuladd_lane1_f32
include/nnpack/arm_neon.h:137
↓ 32 callers
Function
conv6
* VGG model A conv6 layer: * input channels = 512 * output channels = 512 * input size = 28x28 * implicit padding = 1 * k
test/models/vgg-a.h:158
↓ 32 callers
Function
conv8
* VGG model A conv8 layer: * input channels = 512 * output channels = 512 * input size = 14x14 * implicit padding = 1 * k
test/models/vgg-a.h:176
↓ 31 callers
Function
max
include/nnpack/utils.h:17
↓ 31 callers
Function
vmuladdq_lane0_f32
include/nnpack/arm_neon.h:81
↓ 30 callers
Function
doz
include/nnpack/utils.h:13
↓ 30 callers
Method
outputSubsampling
test/testers/convolution.h:182
↓ 29 callers
Function
median
bench/median.c:43
↓ 29 callers
Function
neon_reluq_f32
include/nnpack/activations.h:27
↓ 28 callers
Method
testOutputInplace
test/testers/relu.h:151
↓ 27 callers
Function
neon_relu_f32
include/nnpack/activations.h:37
↓ 27 callers
Function
relu
include/nnpack/activations.h:6
↓ 27 callers
Function
round_down
include/nnpack/utils.h:33
↓ 27 callers
Function
vmuladdq_lane1_f32
include/nnpack/arm_neon.h:89
↓ 26 callers
Function
divide_round_up
include/nnpack/utils.h:37
↓ 26 callers
Function
vmulsubq_f32
include/nnpack/arm_neon.h:73
↓ 24 callers
Method
outputSize
test/testers/pooling.h:158
↓ 24 callers
Function
vmulsub_f32
include/nnpack/arm_neon.h:121
↓ 23 callers
Method
testOptimizedComplex
* Validates that optimized complex 1D FFT produces the same output as reference implementation. */
test/testers/fourier.h:379
↓ 23 callers
Function
winograd_f6k3_output_transformq
src/neon/winograd/f6x6k3x3.h:222
↓ 20 callers
Function
psimd_exp_f32
src/psimd/exp.h:6
↓ 16 callers
Function
store_ymm_result
(variable, result)
src/x86_64-fma/fft16x16.py:26
↓ 15 callers
Function
_MM_SHUFFLE
(z, y, x, w)
src/x86_64-fma/common.py:51
↓ 14 callers
Function
winograd_f6k3_output_transform
src/neon/winograd/f6x6k3x3.h:160
↓ 13 callers
Function
psimd_ifft16_real_f32
src/psimd/fft/real.h:161
↓ 13 callers
Function
transpose2x2x2x64
(ymm_a, ymm_b, use_blend=True)
src/x86_64-fma/common.py:178
↓ 12 callers
Function
nnp_convolution_inference
include/nnpack.h:635
↓ 12 callers
Function
psimd_transpose4x4_f32
src/psimd/transpose.h:6
↓ 12 callers
Method
testHXGEMM
test/testers/gemm-ukernel.h:194
↓ 12 callers
Method
testOptimizedDualReal
* Validates that optimized dual-sequence real 1D FFT produces the same output as reference implementation. */
test/testers/fourier.h:393
↓ 12 callers
Method
testOptimizedReal
* Validates that optimized real 1D FFT produces the same output as reference implementation. */
test/testers/fourier.h:386
↓ 12 callers
Function
transpose2x2x128
(ymm_a, ymm_b, use_blend=True)
src/x86_64-fma/common.py:152
↓ 11 callers
Function
TEST
test/sgemm/neon.cc:8
↓ 11 callers
Function
extract_time
(line, prefix)
benchmark.py:6
↓ 10 callers
Function
psimd_butterfly_with_negated_b_f32
src/psimd/butterfly.h:26
↓ 10 callers
Function
read_timer
bench/perf_counter.h:74
↓ 10 callers
Method
testSoa
* Validates that complex 1D FFT with structure-of-arrays layout produces the same output * as FFT with array-of-structures layout. This function wor
test/testers/fourier.h:117
↓ 10 callers
Function
winograd_f6k3_kernel_transform
src/neon/winograd/f6x6k3x3.h:85
↓ 9 callers
Function
conv1_relu
* VGG model A conv1 ReLU layer: * channels = 64 * image size = 224x224 */
test/models/vgg-a.h:35
↓ 9 callers
Function
conv2_relu
* VGG model A conv2 ReLU layer: * channels = 128 * image size = 224x224 */
test/models/vgg-a.h:65
↓ 9 callers
Function
conv3_relu
* VGG model A conv3 ReLU layer: * channels = 256 * image size = 56x56 */
test/models/vgg-a.h:95
↓ 9 callers
Function
fc6
* VGG model A fc6 layer: * input channels = 25088 * output channels = 4096 */
test/models/vgg-a.h:203
↓ 9 callers
Function
fc6_relu
* VGG model A fc6 ReLU layer: * channels = 4096 */
test/models/vgg-a.h:214
↓ 9 callers
Function
fc7
* VGG model A fc7 layer: * input channels = 4096 * output channels = 4096 */
test/models/vgg-a.h:225
↓ 9 callers
Function
fc8
* VGG model A fc8 layer: * input channels = 4096 * output channels = 1000 */
test/models/vgg-a.h:237
↓ 9 callers
Function
fc8_relu
* VGG model A fc8 ReLU layer: * channels = 1000 */
test/models/vgg-a.h:248
↓ 9 callers
Function
grad_relu
include/nnpack/activations.h:10
↓ 9 callers
Method
testInferenceF16F32
test/testers/fully-connected.h:181
↓ 9 callers
Method
testInferenceF32
test/testers/fully-connected.h:147
↓ 8 callers
Function
psimd_cmul_soa_f32
src/psimd/fft/soa.h:8
↓ 8 callers
Function
psimd_cmulc_soa_f32
src/psimd/fft/soa.h:23
↓ 8 callers
Function
psimd_relu_f32
include/nnpack/activations.h:15
next →
1–100 of 873, ranked by callers