Commits · f3d53c3d5f2820c78cbc0d52e6708d62f1b316db · tsoc / superbenchmark

30 Aug, 2021 2 commits

Benchmarks: Add Benchmark - Add gemm flops microbenchmark for amd (#152) · f3d53c3d

Yuting Jiang authored Aug 30, 2021

**Description**
Add gemm flops microbenchmark for amd.

**Major Revision**
- Add gemm flops microbenchmark for amd.
- Add related example and test file.

f3d53c3d

Benchmarks: Code Revision - Extract base class for gemm flops microbenchmark (#165) · b0df66f7

Yuting Jiang authored Aug 30, 2021

**Description**
Extract base class for gemm flops microbenchmark.

**Major Revision**
- extract base class for gemm flops microbenchmark and add related test.
- revise gemm_flops_performance for cuda.

b0df66f7

27 Aug, 2021 4 commits

Benchmarks: Code Revision - Rename kernel_launch_overhead metrics (#171) · 35114bae

guoshzhao authored Aug 28, 2021

**Description**
Rename `kernel_launch_overhead_event` to `event_overhead`, `kernel_launch_overhead_wall` to `wall_overhead`.

35114bae

Benchmarks: Add Benchmark - Add memory bus bandwidth performance microbenchmark for amd (#153) · 666e3a94

Yuting Jiang authored Aug 27, 2021

**Description**
Add memory bus bandwidth performance microbenchmark for amd.

**Major Revision**
- Add memory bus bandwidth performance microbenchmark for amd.
- Add related example and test file.

666e3a94

Benchmarks: Add Benchmark - Add GPU SM copy benchmark (#162) · 2880f71e
Ziyue Yang authored Aug 27, 2021
```
**Description**
This commit adds the benchmark program for GPU-initiated data transfer benchmark.
```
2880f71e

Benchmarks: Fix Bug - fix bug of microbenmark building cublas and cudnn for... · 958ebc0e

Yuting Jiang authored Aug 27, 2021

Benchmarks: Fix Bug - fix bug of microbenmark building cublas and cudnn for amd in build pipeline (#166)

**Description**
Fix bug of microbenmark building cublas and cudnn for amd

**Major Revision**
- remove cuda LANGUAGES in project()
- check CUDAToolkit quiet and then build if found

958ebc0e

26 Aug, 2021 1 commit

Benchmarks: Code Revision - Rename computation_communication_overlap microbenchmark metric (#167) · 34cd2e8c

Yuting Jiang authored Aug 26, 2021

**Description**
Rename computation_communication_overlap microbenchmark metric .

**Major Revision**
- remove rank info in metric.
- simplify and rename metric.

34cd2e8c

25 Aug, 2021 1 commit

Benchmarks: Code Revision - Extract base class for memory bandwidth microbenchmark (#159) · e5e84a2e

Yuting Jiang authored Aug 26, 2021

**Description**
extract base class for memory bandwidth microbenchmark.

**Major Revision**
- revise and optimize cuda_memory_bandwidth_performance
- extract base class for memory bandwidth microbenchmark
- add test for base class

e5e84a2e

22 Aug, 2021 1 commit
- Benchmarks: Revise Benchmark - Add readwrite I/O pattern (#161) · 6774d7b7
  Ziyue Yang authored Aug 22, 2021
```
**Description**
This commit adds readwrite I/O pattern for FIO benchmark. Read/write ratio is fixed at 4:1.
```
  6774d7b7
06 Aug, 2021 2 commits
- Benchmarks: Add Feature - Set reduce type for current benchmarks' metrics. (#149) · acf365a8
  guoshzhao authored Aug 06, 2021
```
**Description**
Set reduce type for current benchmarks' metrics, including model benchmarks and ShardingMatmul.
```
  acf365a8
- Benchmarks: Code Revision - Calculate average value by using statistics module. (#148) · bc1a61b9
  guoshzhao authored Aug 06, 2021
```
**Description**
Replace `sum(results) / len(results)` with `statistics.mean(results)`
```
  bc1a61b9
05 Aug, 2021 1 commit

Benchmarks: Add Feature - Add reduce function support for output summary. (#147) · e41b1f62

guoshzhao authored Aug 05, 2021

**Description**
Add reduce function support for output summary.

**Major Revision**
- Add reducer class to maintain all reduce functions.
- Save reduce type of each metric into `BenchmarkResult`
- Fix UT.

e41b1f62

30 Jul, 2021 1 commit
- Benchmarks: Add Benchmark - Revise and add rccl microbenchmark for rocm (#143) · 157b4e2d
  Yuting Jiang authored Jul 30, 2021
```
**Description**
Add rccl bandwidth microbenchmark for rocm.

**Major Revision**
- Register rccl-bw benchmark.
```
  157b4e2d
27 Jul, 2021 2 commits

Benchmarks: Add Benchmark - Add the source code of rocm kernel launch overhead benchmark. (#136) · 1ee8f7dc

Yuting Jiang authored Jul 27, 2021

**Description**
Add the source code of rocm kernel launch overhead benchmark. 

**Major Revision**
- Revise cmake build logic to support both cuda and rocm

1ee8f7dc

Benchmarks: Build Pipeline - Support rocm cmake build (#137) · fdc33f40

Yuting Jiang authored Jul 27, 2021

**Description**
Support rocm cmake build. 

**Major Revision**
- Add  some envs in rocm_common.cmake to support rocm cmake build.

fdc33f40

26 Jul, 2021 1 commit

Benchmarks: Add Benchmark - Add NCCL performance benchmark (#113) · e083a598

Yuting Jiang authored Jul 26, 2021

**Description**
Add NCCL performance microbenchmark.

**Major Revision**
- Add microbenchmark, example, test, config for NCCL

e083a598

23 Jul, 2021 2 commits

Benchmarks: Add Benchmark - Add IB Loopback performance benchmark. (#112) · b0c5addc

Yuting Jiang authored Jul 24, 2021

**Description**
Add RDMA Loopback performance microbenchmark.

**Major Revision**
- Add microbenchmark, example, test, config for RDMA Loopback

b0c5addc

Benchmarks: Add Benchmark - Add disk performance benchmark (#132) · db297fb4

Ziyue Yang authored Jul 23, 2021

**Description**
Add disk performance microbenchmark.

**Major Revision**
- Add microbenchmark, example, test, config for disk performance.

**Minor Revision**
- Fix bugs in executor unit test related to default enabled tests.

db297fb4

13 Jul, 2021 1 commit

Benchmarks: Add Benchmark - Add memory bandwidth benchmark for cuda. (#114) · f9550bd6

Yuting Jiang authored Jul 13, 2021

Add microbenchmark, example, test, config for cuda memory performance and Add cuda-samples(tag with cuda version) as git submodule and update related makefile

f9550bd6

30 Jun, 2021 1 commit
- Benchmarks: Fix Bug - Fix typo in gemm-flops benchmark. (#109) · 1e96c27e
  guoshzhao authored Jun 30, 2021
  
  1e96c27e
29 Jun, 2021 1 commit
- Benchmarks: Fix Bug - Fix gemm kernel bug for nvidia v100. (#105) · 8ffaddfa
  guoshzhao authored Jun 29, 2021
```
* fix bug for nvidia v100
* hard code the supported dict for different arch.
```
  8ffaddfa
21 Jun, 2021 1 commit
- Benchmarks: Add Feature - Add DistributedImpl and DistributedBackend arguments... · 216c5b5c
  guoshzhao authored Jun 21, 2021
```
Benchmarks: Add Feature - Add DistributedImpl and DistributedBackend arguments for micro benchmark. (#100)
```
  216c5b5c
20 Jun, 2021 1 commit
- Bug bash - Rename bin name and metric name of cublas and cudnn microbenchmark (#99) · 3d72c078
  Yuting Jiang authored Jun 20, 2021
```
rename bin name and result metric of cublas and cudnn microbenchmark
```
  3d72c078
02 Jun, 2021 2 commits
- Benchmarks: Code Revision - Change default shape of sharding-matmul. (#92) · 44c5103b
  guoshzhao authored Jun 02, 2021
```
* Change default shape of sharding-matmul.
```
  44c5103b
- Benchmarks: Add Benchmark - Add FLOPs performance benchmark for cuda. (#87) · 6c6f5269
  guoshzhao authored Jun 02, 2021
```
* add cuda flops performance benchmark.
```
  6c6f5269
01 Jun, 2021 3 commits
- Benchmarks: Add benchmark - add micro benchmark for cudnn test (#89) · 83235433
  Yuting Jiang authored Jun 01, 2021
```
* add python related cudnn microbenchmark
```
  83235433
- Benchmarks: Code Revision - add error return code for cublas microbenchmark (#90) · 08317481
  Yuting Jiang authored Jun 01, 2021
```
* add error return code for cublas micro benchmark
```
  08317481
- Benchmarks: Add benchmark - add source code of cudnn function micro benchmark (#78) · 61c258fe
  Yuting Jiang authored Jun 01, 2021
```
* Benchmarks: Add benchmark - add source code of cudnn function micro benchmark
```
  61c258fe
31 May, 2021 1 commit

Benchmarks: Add benchmark - add micro benchmark for cublas test (#80) · 18398fba

Yuting Jiang authored May 31, 2021



* add benchmark for cublas test

* format

* revise error handling and test

* add interface to read json file, revise json file path and include .json in packaging

* add random_seed in arguments

* revise preprocess of cublas benchmark

* fix lint error and note error in source code

* update according comments

* revise input arguments from json file to custom str and convert json file to built-in dict list

* restore package config

* fit lint issue

* update platform and comments

* rename files to match source code dir and fix comments error
Co-authored-by: root <root@sb-validation-000001.51z1chmys5fuzfqyo4niepozre.bx.internal.cloudapp.net>

18398fba

27 May, 2021 1 commit

Benchmarks: Add benchmark - add source code of cublas function micro benchmark (#77) · 87f6b371

Yuting Jiang authored May 27, 2021



* Superbenchmark: Add benchmarks - add cublas function micro benchmark

* format

* add python benchmark for cublas functions, example and test file

* detele python related and rename some files

* revise cmd_helper and move json package to cmake

* resolve conflict

* revise error handing to try-catch and update some code style

* revise cmd_helper.h, cublas_helper.h, cublas_helper.cpp

* revise structure of the cublas function

* add some comments and move cuda_init and cuda_free

* add comments for class member

* add ramdom seed, revise input from file to json string, simplify cmake

* delete json file in source code of cublas

* update according comments

* limit batchcount=1 in initialization of cublas function which do not use batch count

* revise and fix some errors of annotations

* update according comments and revise construction of CublasFunction
Co-authored-by: root <root@sb-validation-000001.51z1chmys5fuzfqyo4niepozre.bx.internal.cloudapp.net>

87f6b371

26 May, 2021 1 commit
- Benchmarks: Build Pipeline - Revise path of installing cmake projects (#83) · e9965162
  Yuting Jiang authored May 26, 2021
```
* Unify SB_MICRO_PATH and SB_MICRO_LIB

* fix bug of lib path
```
  e9965162
19 May, 2021 1 commit
- Benchmarks: Add Benchmark - Add kernel launch overhead benchmark. (#74) · e977bbc1
  guoshzhao authored May 19, 2021
```
* add kernel launch overhead benchmark.
```
  e977bbc1
18 May, 2021 1 commit

Benchmarks: Add Benchmark - Add the source code of cuda kernel launch overhead benchmark. (#71) · 7cfe7c16

guoshzhao authored May 18, 2021

* add cuda kernel launch overhead benchmark - source part.
* can customize the nvcc_archs_support.
* set SB_MICRO_PATH for azure pipeline tests.

7cfe7c16

13 May, 2021 1 commit

Benchmarks: Code Revision - Revise MicroBenchmark class to be more flexible. (#66) · 729e04ab

guoshzhao authored May 13, 2021

* Revise MicroBenchmark class to be more flexible.
* use command index not the command as the parameter.
* changes according to discussion.

729e04ab

14 Apr, 2021 1 commit

Benchmarks: Add Benchmark - Add computation and communication overlap micro benchmark (#39) · 435b2d5e

Yuting Jiang authored Apr 14, 2021



* Benchmarks: Add Benchmark - add computation and communication overlap micro benchmark

* Benchmarks: Add benchmark - fix some format issues and typo

* Benchmarks: Add Benchmark - update according comments and add test

* revise tests

* skip multi gpu test due to no multi gpu
Co-authored-by: v-yujiang <v-yujiang@microsoft.com>

435b2d5e

12 Apr, 2021 4 commits
- add _post_process() implementation in sharding_matmul.py to clean up distributed resource. (#46) · 7f6deabb
  guoshzhao authored Apr 12, 2021
```
Co-authored-by: Guoshuai Zhao <guzhao@microsoft.com>
```
  7f6deabb
- rename metric name of sharding-matmul (#48) · 7bd41649
  guoshzhao authored Apr 12, 2021
```
Co-authored-by: Guoshuai Zhao <guzhao@microsoft.com>
```
  7bd41649
- remove unused code. (#47) · 020cefbd
  guoshzhao authored Apr 12, 2021
```
Co-authored-by: Guoshuai Zhao <guzhao@microsoft.com>
```
  020cefbd
- change the condition when execute self.__matmul_nosharding() (#49) · 82e267b9
  guoshzhao authored Apr 12, 2021
```
Co-authored-by: Guoshuai Zhao <guzhao@microsoft.com>
```
  82e267b9
09 Apr, 2021 1 commit

Benchmarks: Add Benchmark - Add op-sharding microbenchmark, including matmul... · f0f65a71

guoshzhao authored Apr 09, 2021


Benchmarks: Add Benchmark - Add op-sharding microbenchmark, including matmul and sharding_matmul. (#36)

* add microbenchmark - sharding matmul.
* address comments.
Co-authored-by: Guoshuai Zhao <guzhao@microsoft.com>

f0f65a71