MCPcopy Create free account

hub / github.com/alibaba/BladeDISC / functions

Functions6,088 in github.com/alibaba/BladeDISC

↓ 1 callersFunctioncreateDiscMathApproximationPass
tao_compiler/mlir/disc/transforms/disc_math_approximation.cc:46
↓ 1 callersFunctioncreateDiscMemRefCSEPass
tao_compiler/mlir/disc/transforms/disc_memref_cse.cc:113
↓ 1 callersFunctioncreateDiscMemrefCanonicalizerPass
tao_compiler/mlir/disc/transforms/disc_memref_canonicalizer.cc:80
↓ 1 callersFunctioncreateDiscMemrefCopyToLinalgPass
tao_compiler/mlir/disc/tools/disc-transform/transforms/memref_copy_to_linalg.cc:65
↓ 1 callersFunctioncreateDiscMhloDecompositionRewriterPass
tao_compiler/mlir/disc/transforms/mhlo_decomp_rewriters.cc:152
↓ 1 callersFunctioncreateDiscOutlineCpuKernelPass
tao_compiler/mlir/disc/transforms/disc_outline_cpu_kernel.cc:438
↓ 1 callersFunctioncreateDiscParallelLoopCollapsingPass
tao_compiler/mlir/disc/transforms/parallel_loop_collapsing.cc:69
↓ 1 callersFunctioncreateDiscQuantizedConvRewriter
tao_compiler/mlir/disc/transforms/conv_rewriter.cc:485
↓ 1 callersFunctioncreateDiscQuantizedDotMergePass
tao_compiler/mlir/disc/transforms/disc_quantized_dot_merge.cc:511
↓ 1 callersFunctioncreateDiscQuantizedDotRewriter
tao_compiler/mlir/disc/transforms/quantized_dot_rewriter.cc:208
↓ 1 callersFunctioncreateDiscReductionRewriterPass
tao_compiler/mlir/disc/transforms/reduction_rewriters.cc:133
↓ 1 callersFunctioncreateDiscRemoveDeadBufferPass
tao_compiler/mlir/disc/transforms/disc_remove_dead_buffer.cc:82
↓ 1 callersFunctioncreateDiscRemoveShapeConstraintsPass
tao_compiler/mlir/disc/transforms/remove_shape_constraints.cc:81
↓ 1 callersFunctioncreateDiscRewritePayloadIRForRALPass
tao_compiler/mlir/disc/tools/disc-transform/transforms/rewrite_payload_ir_for_ral.cc:134
↓ 1 callersFunctioncreateDiscSparseGemmTransposeSimplifierPass
tao_compiler/mlir/disc/transforms/disc_dense_to_sparse.cc:556
↓ 1 callersFunctioncreateDiscSplitLargeOpsPass
tao_compiler/mlir/disc/transforms/split_large_ops.cc:116
↓ 1 callersFunctioncreateDiscStdBufferizePass
tao_compiler/mlir/disc/transforms/disc_std_bufferize.cc:156
↓ 1 callersFunctioncreateDiscStitchFusionPass
tao_compiler/mlir/disc/transforms/disc_stitch_fusion.cc:114
↓ 1 callersFunctioncreateDiscToLLVMPass
tao_compiler/mlir/disc/transforms/disc_to_llvm.cc:1355
↓ 1 callersFunctioncreateDiscTransformDialectEraseSchedulePass
tao_compiler/mlir/disc/tools/disc-transform/transforms/transform_dialect_interpreter.cc:188
↓ 1 callersFunctioncreateDiscTransformDialectInterpreterPass
tao_compiler/mlir/disc/tools/disc-transform/transforms/transform_dialect_interpreter.cc:181
↓ 1 callersFunctioncreateDiscTransformLegalizeToLoopPass
tao_compiler/mlir/disc/transforms/disc_transform_legalize_to_loop.cc:406
↓ 1 callersFunctioncreateDiscUnhandledAtomicRMWConverterPass
tao_compiler/mlir/disc/transforms/disc_unhandled_atomic_rmw_converter.cc:91
↓ 1 callersFunctioncreateForLoopUnrollInterleavePass
tao_compiler/mlir/disc/transforms/disc_for_loop_unroll_interleave.cc:158
↓ 1 callersFunctioncreateFunctionDeadArgumentEliminationPass
tao_compiler/mlir/disc/transforms/disc_function_dead_argument_elimination.cc:101
↓ 1 callersFunctioncreateInitTensor
Helper to create a tensor filled with the given scalar. Scalar would be converted the to the element type of the given tensor type.
pytorch_blade/pytorch_blade/torch-mlir/lib/Dialect/TorchConversion/Transforms/DiscDecomposeComplexOps.cpp:41
↓ 1 callersFunctioncreateLLVMInsertValueSimplifierPass
tao_compiler/mlir/disc/transforms/disc_llvm_insert_value_simplifier.cc:111
↓ 1 callersFunctioncreateLhloFusionInlinerPass
tao_compiler/mlir/disc/transforms/lhlo_fusion_inliner.cc:53
↓ 1 callersFunctioncreateLoadOpsArray
tao_compiler/mlir/disc/transforms/revise_kernel_outlining.cc:97
↓ 1 callersFunctioncreateMemRef1DReinterpretCastWithStaticShape
tao_compiler/mlir/disc/transforms/lhlo_elemental_utils.cc:1221
↓ 1 callersFunctioncreateOffsetLoad
tao_compiler/mlir/disc/transforms/lhlo_elemental_utils.cc:1282
↓ 1 callersFunctioncreateOffsetStore
tao_compiler/mlir/disc/transforms/lhlo_elemental_utils.cc:1276
↓ 1 callersFunctioncreateParallelLoopTilingPass
tao_compiler/mlir/disc/transforms/parallel_loop_tiling.cc:88
↓ 1 callersFunctioncreatePlacerPass
tao_compiler/mlir/disc/transforms/mhlo_placer.cc:668
↓ 1 callersFunctioncreatePrintFusionParams
Print the params of a fusion with tensor shape at runtime. For debugging.
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:3848
↓ 1 callersFunctioncreateRalInjectExecutionContextPass
tao_compiler/mlir/disc/transforms/ral_inject_execution_context.cc:134
↓ 1 callersFunctioncreateRank0Tensor
Helper to create a rank 0 tensor filled with the given `scalar`. `scalar` would be converted to the element type of the given `inputType`.
pytorch_blade/pytorch_blade/torch-mlir/lib/Dialect/TorchConversion/Transforms/DiscDecomposeComplexOps.cpp:61
↓ 1 callersFunctioncreateReviseArgsForStaticRankPass
tao_compiler/mlir/disc/transforms/revise_args_for_static_rank.cc:145
↓ 1 callersFunctioncreateReviseGpuKernelOutliningPass
tao_compiler/mlir/disc/transforms/revise_kernel_outlining.cc:444
↓ 1 callersFunctioncreateSideEffectLoopInvariantCodeMotionPass
tao_compiler/mlir/disc/transforms/disc_side_effect_loop_invariant_code_motion.cc:113
↓ 1 callersFunctioncreateStripShapeConstraintOpsPass
tao_compiler/mlir/disc/transforms/disc_strip_shape_constraint_ops.cc:51
↓ 1 callersFunctioncreateTransposeSimplifierPass
tao_compiler/mlir/disc/transforms/disc_transpose_simplifier.cc:1081
↓ 1 callersMethodcreate_input
()
pytorch_blade/tests/test_tools.py:62
↓ 1 callersFunctioncreate_list_construct
(graph, vals, list_type)
pytorch_blade/torch_blade/utils.py:230
↓ 1 callersFunctioncreate_prim_constant_with_val
pytorch_blade/pytorch_blade/compiler/jit/tool_funcs.cpp:183
↓ 1 callersFunctioncu_prof_start
()
examples/PyTorch/Train/Dynamo/Bert/test_bert.py:38
↓ 1 callersFunctioncu_prof_start
()
examples/PyTorch/Inference/CUDA/T5/main.py:28
↓ 1 callersFunctioncu_prof_start
()
examples/PyTorch/Inference/CUDA/BERT/main.py:34
↓ 1 callersFunctioncu_prof_start
()
examples/PyTorch/Inference/CUDA/S2T/main.py:28
↓ 1 callersFunctioncu_prof_stop
()
examples/PyTorch/Train/Dynamo/Bert/test_bert.py:44
↓ 1 callersFunctioncu_prof_stop
()
examples/PyTorch/Inference/CUDA/T5/main.py:34
↓ 1 callersFunctioncu_prof_stop
()
examples/PyTorch/Inference/CUDA/BERT/main.py:40
↓ 1 callersFunctioncu_prof_stop
()
examples/PyTorch/Inference/CUDA/S2T/main.py:34
↓ 1 callersFunctioncudaAvailable
pytorch_blade/pytorch_blade/compiler/jit/torch/shape_analysis_test.cpp:22
↓ 1 callersMethodd2d
tao_compiler/mlir/xla/ral/device/gpu/gpu_driver.cc:154
↓ 1 callersFunctiondata
pytorch_blade/pytorch_blade/ltc/include/torch/csrc/lazy/ts_backend/ts_backend_impl.h:48
↓ 1 callersFunctiondata_processing
(num_samples, batch_size)
examples/PyTorch/Train/Lazy/Bert/bert_train_ltc.py:85
↓ 1 callersFunctiondata_processing
(num_samples, batch_size)
examples/PyTorch/Train/Dynamo/Bert/test_bert.py:95
↓ 1 callersFunctiondeduce_cuda_info
Deduce cuda major and minor version and cuda directory.
pytorch_blade/common_setup.py:429
↓ 1 callersFunctiondependsOnInBlock
Check whether a depends on b.
tao_compiler/mlir/disc/transforms/codegen_utils.cc:381
↓ 1 callersFunctiondetect_cuda_version
return a tuple with major and minor version
scripts/python/tao_common.py:78
↓ 1 callersFunctiondetect_host_tf_version
()
tao/setup.py:34
↓ 1 callersFunctiondict_map
(fn, dic, leaf_type)
examples/PyTorch/Inference/CUDA/AlphaFold/FoldAcc/foldacc/model/modules/utils.py:85
↓ 1 callersFunctiondisc_optimize
(model, inputs, out_file: str)
examples/PyTorch/Inference/CUDA/T5/main.py:66
↓ 1 callersFunctiondisc_optimize
(model, inputs, out_file: str)
examples/PyTorch/Inference/CUDA/BERT/main.py:130
↓ 1 callersFunctiondisc_optimize
(model, inputs, out_file: str)
examples/PyTorch/Inference/CUDA/S2T/main.py:83
↓ 1 callersMethoddispatch
tao_compiler/mlir/disc/transforms/disc_transform_schedule.cc:1552
↓ 1 callersMethoddoCodeGeneration
Do the first level (tile level) codegen for stitch fusion pattern. An example is: Input IR ``` lmhlo.fusion() { lmhlo.exp(%in, %exp_out) : (memref<?x?
tao_compiler/mlir/disc/transforms/fusion_utils.cc:3342
↓ 1 callersFunctiondoTfTopK
tao_compiler/mlir/xla/ral/context/dynamic_sort_impl.cc:38
↓ 1 callersFunctiondo_optimize
(model, inputs, backend)
examples/PyTorch/Inference/CPU/albert/main.py:52
↓ 1 callersFunctiondumpFusionPattern
tao_compiler/mlir/disc/transforms/fusion_utils.cc:48
↓ 1 callersFunctiondumpModule
tao_compiler/mlir/disc/transforms/disc_transpose_simplifier.cc:1027
↓ 1 callersFunctiondumpTilePlan
tao_compiler/mlir/disc/transforms/fusion_utils.cc:1692
↓ 1 callersFunctionemitArrayAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:793
↓ 1 callersFunctionemitBcastOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:204
↓ 1 callersFunctionemitBoolAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:726
↓ 1 callersFunctionemitBytes
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:687
↓ 1 callersFunctionemitCUDASourceAndAttach
tao_compiler/mlir/disc/tools/disc-source-emitter/disc-cuda-emitter.cc:53
↓ 1 callersFunctionemitConstOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:192
↓ 1 callersFunctionemitDenseElementsAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:750
↓ 1 callersFunctionemitDictAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:700
↓ 1 callersFunctionemitDotGeneralOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:87
↓ 1 callersFunctionemitFirstRoundShuffle
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:1306
↓ 1 callersFunctionemitFirstRoundShuffleStitch
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:2345
↓ 1 callersFunctionemitFloatAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:742
↓ 1 callersFunctionemitInitLoops
we can emit them in one kernel based on the fact that all the fused col reduction should be in the same shape.
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:848
↓ 1 callersFunctionemitIntegerAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:734
↓ 1 callersFunctionemitLmhloFusionOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:386
↓ 1 callersFunctionemitLmhloOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:325
↓ 1 callersFunctionemitOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:397
↓ 1 callersFunctionemitReturnOp
tao_compiler/mlir/disc/tools/disc-transform/transforms/legalize_lmhlo_fusion_to_linalg.cc:79
↓ 1 callersFunctionemitRowReduceThreadBlock
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:2586
↓ 1 callersFunctionemitRowReduceThreadBlockV2
* The generated kernel for row-reduce is as following. The code is augly and * will be removed. * * template <typename T> * __global__ void row_re
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:3307
↓ 1 callersFunctionemitSecondRoundShuffle
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:1393
↓ 1 callersFunctionemitSecondRoundShuffleStitch
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:2447
↓ 1 callersFunctionemitStrAttr
tao_compiler/mlir/disc/transforms/disc_lower_to_library_call.cc:718
↓ 1 callersFunctionemitSwitchOperandIdx
For concat op having many operands, we use following schedule: parallel.for %idx in range(num_operands_of_concat) { if (%idx == 0) { copy operand #0
tao_compiler/mlir/disc/transforms/lhlo_legalize_roots_to_loops.cc:4092
↓ 1 callersFunctionenableEagerTransposeFusion
tao_compiler/mlir/disc/transforms/fusion_utils.cc:64
↓ 1 callersFunctionenableInOutLayoutTuning
tao_compiler/mlir/xla/ral/context/common_context_impl_mkldnn.cc:157
↓ 1 callersFunctionenableTransposeLibraryCall
tao_compiler/mlir/disc/transforms/fusion_utils.cc:68
← previousnext →2,001–2,100 of 6,088, ranked by callers