InfiniTensor

History

constroy Li f60767a770 impl distributed launch with NCCL (#106 ) * add cmake bits about NCCL * move example to examples/NNmodel * impl NCCL communicator * add comm related function to Runtime * export runtime interface * add launch.py * use unique name to distingush the the NCCL ID file * add timeout to communicator init * expose communicator obj from runtime obj, add unit test for nccl communicator * reformat files * Add allReduce operator and cuda nccl allReduce kernel * impl model parallel for resnet * add allGather nccl kernel and operator * Add allreduce allgather operator tests, change allgather kernel to output list of tensor, fix shape infer, handle nullptr output * fix format of onnx.py * use concat following AllGather * get tensor parallel for resnet * fix format of graph_handler.cc * change BUILD_DIST default to OFF * polish code of communicator * update .gitignore * Add broadcast operator and cuda kernel * Add comments for operators * remove const of class member * move communicator to CudaRuntimeObj * Add an empty line at EOF. --------- Co-authored-by: panzezhong <panzezhong@qiyuanlab.com> Co-authored-by: Haojie Wang <haojie0429@gmail.com>		2023-09-05 09:47:35 +08:00
..
core	Issue 107: Add copyin Numpy and covertion to Numpy (#126 )	2023-09-01 11:20:26 +08:00
cuda	impl distributed launch with NCCL (#106 )	2023-09-05 09:47:35 +08:00
kernels	impl distributed launch with NCCL (#106 )	2023-09-05 09:47:35 +08:00
nnet	NNET supports TVM backend and kernels (#78 )	2023-04-18 00:26:36 +08:00
operators	impl distributed launch with NCCL (#106 )	2023-09-05 09:47:35 +08:00
script	build: 实现格式化 git added c/c++ 源码的脚本 (#98 )	2023-07-21 12:29:50 +08:00