mirror of
https://github.com/NVIDIA/cuda-samples.git
synced 2025-04-19 06:20:50 +01:00
37 lines
1.4 KiB
Markdown
37 lines
1.4 KiB
Markdown
# simpleAWBarrier - Simple Arrive Wait Barrier
|
|
|
|
## Description
|
|
|
|
A simple demonstration of arrive wait barriers.
|
|
|
|
## Key Concepts
|
|
|
|
Arrive Wait Barrier
|
|
|
|
## Supported SM Architectures
|
|
|
|
[SM 7.0 ](https://developer.nvidia.com/cuda-gpus) [SM 7.2 ](https://developer.nvidia.com/cuda-gpus) [SM 7.5 ](https://developer.nvidia.com/cuda-gpus) [SM 8.0 ](https://developer.nvidia.com/cuda-gpus) [SM 8.6 ](https://developer.nvidia.com/cuda-gpus) [SM 8.7 ](https://developer.nvidia.com/cuda-gpus) [SM 8.9 ](https://developer.nvidia.com/cuda-gpus) [SM 9.0 ](https://developer.nvidia.com/cuda-gpus)
|
|
|
|
## Supported OSes
|
|
|
|
Linux, Windows, QNX
|
|
|
|
## Supported CPU Architecture
|
|
|
|
x86_64, armv7l, aarch64
|
|
|
|
## CUDA APIs involved
|
|
|
|
### [CUDA Runtime API](http://docs.nvidia.com/cuda/cuda-runtime-api/index.html)
|
|
cudaStreamCreateWithFlags, cudaFree, cudaDeviceGetAttribute, cudaMallocHost, cudaFreeHost, cudaStreamSynchronize, cudaLaunchCooperativeKernel, cudaMalloc, cudaOccupancyMaxActiveBlocksPerMultiprocessor, cudaMemcpyAsync, cudaOccupancyMaxPotentialBlockSize
|
|
|
|
## Dependencies needed to build/run
|
|
[CPP11](../../../README.md#cpp11), [MBCG](../../../README.md#mbcg)
|
|
|
|
## Prerequisites
|
|
|
|
Download and install the [CUDA Toolkit 12.5](https://developer.nvidia.com/cuda-downloads) for your corresponding platform.
|
|
Make sure the dependencies mentioned in [Dependencies]() section above are installed.
|
|
|
|
## References (for more details)
|