317 Commits (tmp-test)

Author SHA1 Message Date
  Megvii Engine Team b04ad06f84 refactor(megdnn): refactor matmul algo in conv backward filter 4 years ago
  Megvii Engine Team 25089e520e refactor(megdnn): refactor matmul algo in conv backward data 4 years ago
  Megvii Engine Team 0d720653ac refactor(megdnn): add default algo for convolution forward 4 years ago
  Megvii Engine Team 659217acd2 refactor(megdnn): refactor bfloat16 convbias to recursive inteface 4 years ago
  Megvii Engine Team 4a1d52c9c6 refactor(megdnn): refactor bfloat16 matmul to recursive inteface 4 years ago
  Megvii Engine Team b8febaf91f refactor(megdnn): refactor bfloat16 convolutionbackwardfilter to recursive inteface 4 years ago
  Megvii Engine Team f14e0c17e7 feat(mgb): add recursive for fastrun and megdnn test 4 years ago
  Megvii Engine Team 0e8b81c20e fix(dnn/opencl): fix elemwise negative stride support 4 years ago
  Megvii Engine Team 364afec033 chore(mge): update copyright years 4 years ago
  Megvii Engine Team ae8b38f634 fix(cmake/whl): reduce wheel size 4 years ago
  Megvii Engine Team 3bda334798 fix(dnn/fallback): fix segmentfault caused by im2col/conv1x1 using 4 years ago
  Megvii Engine Team 409a877267 feat(dnn): add algo interface for rocm&fallback matmul and batched matrix mul 4 years ago
  Megvii Engine Team a85531dd0f feat(mgb/opr): add tqt opr 4 years ago
  Megvii Engine Team c3a4b2225d feat(dnn/cuda): add cutlass impls for fused convolution reformat operation 4 years ago
  Megvii Engine Team 5f44203d7b feat(dnn/cuda): add a cutlass impl for fusing convolution and dimshuffle 4 years ago
  Megvii Engine Team 61f917fb8e feat(dnn/cuda): add impl for fusing warp perspective and dimshuffle 4 years ago
  Megvii Engine Team fc0fcd2f7f chore(winograd): remove winograd transform code 4 years ago
  Megvii Engine Team 7e2b2dbffc fix(dnn/test): delete large size in ARM_COMMON.FP32_GEVM 4 years ago
  Megvii Engine Team 3bf73ff16f feat(dnn): add cuda preprocess fusion 4 years ago
  Megvii Engine Team 86cf7490ec feat(dnn/aarch64): add quantizeds4 matmul int4x4x16_k8x8x8 4 years ago
  Megvii Engine Team 142f31a875 perf(dnn/cuda): change conv_bias heu, prefer dnn chanwise impl, dislike dnn batch gemm conv1x1 4 years ago
  Megvii Engine Team 98a74e4a7b refactor(dnn): refactor opr proxy in test 4 years ago
  Megvii Engine Team 7066ad5ba6 feat(dnn): add uint16 support 4 years ago
  Megvii Engine Team a1877ee0fa refactor(dnn): refactor algo interface, use algoinfo instead of global algorithm 4 years ago
  Megvii Engine Team 6f5d0febf1 perf(dnn/cuda): enhance performance for pooling forward 4 years ago
  Megvii Engine Team 0560a218af chore(dnn/test): refactor megdnn arm_common test 4 years ago
  Megvii Engine Team 6856ce9ce2 feat(dnn): support conv bias activation for nchw4 input tensor format and nchw output tensor format 4 years ago
  Megvii Engine Team 60c6d59fc9 feat(mbg/core): support bias preprocess in conv_bias 4 years ago
  Megvii Engine Team 1e71e0afe0 refactor(dnn): refactor deconv algo 4 years ago
  Megvii Engine Team 89ad33aeb3 feat(dnn/cuda): support weight preprocessing for cutlass algorithms 4 years ago
  Megvii Engine Team c03249c059 feat(dnn/opr): add megdnn fake quant opr 4 years ago
  Megvii Engine Team 739f927c4c feat(dnn/cuda): opt dp4a conv for small channel base on cutlass 4 years ago
  Megvii Engine Team 4aa277a203 refactor(dnn/cuda): misc 4 years ago
  Megvii Engine Team e39f938662 refactor(dnn): remove ProfileCache and matmul algo in x86 4 years ago
  Megvii Engine Team 89303cd829 feat(megdnn/rocm): add bn for rocm backend 4 years ago
  Megvii Engine Team aea829c9fa feat(megdnn/rocm): add average inclusive mode for pooling 4 years ago
  Megvii Engine Team ba66e1d039 feat(dnn): add nchw_fp32 nchw44_qint8 cuda dct 4 years ago
  Megvii Engine Team 8764a6c8ff feat(dnn/cuda): add volta dp4a int8 sass kernel 4 years ago
  Megvii Engine Team 92b12685db feat(dnn/aarch64): add aarch64 int8X8X16_mk4_k8x8x8 matmul, performance is better 4 years ago
  Megvii Engine Team edb32495c6 feat(dnn/opr): add megdnn adaptive pooling opr 4 years ago
  Megvii Engine Team 310c805f20 fix(dnn/cuda): use kernel parameter instead of user constant memory 4 years ago
  Megvii Engine Team 3a03fa7a50 fix(dnn/cuda): disable pascal sass conv2d 4 years ago
  Megvii Engine Team a5fad7d07c feat(dnn): add compile for riscv64 4 years ago
  Megvii Engine Team 76fa71573b feat(dnn/cuda): add cutlass nchw4 convolution 4 years ago
  Megvii Engine Team 5b6ebeb563 fix(mgb): append json file for dump and ready for midout open source 4 years ago
  Megvii Engine Team 16324e3076 feat(dnn/cuda): add remap backward 5 years ago
  Megvii Engine Team bd73dabbe2 fix(dnn/build): add CUDNN_INCLUDE_DIR to the megdnn_test target 4 years ago
  Megvii Engine Team 343335932a fix(dnn/arm): fix read invalid data in arm kernel 4 years ago
  Megvii Engine Team 6e882c1a86 feat(whl/imperative): compat for build python whl imperative and legacy runtime 4 years ago
  Megvii Engine Team 7f857bd471 feat(mgb/rocm): add cmake for rocm and fix compile errors and bn 4 years ago