使用 OpenACC 進行 GPU 程式設計培訓
OpenACC 是一個異構程式設計的開放標準,讓程式碼能在不同的平台和裝置上執行,例如多核 CPU、GPU、FPGA 等。
此由導師主導的培訓課程(線上或線下)專為初級至中級的開發人員設計,讓他們學習如何使用 OpenACC 為異構裝置撰寫程式並發揮其平行運算能力。
完成本培訓後,參與者將能夠:
- 建置 OpenACC 開發環境。
- 編寫並執行基本的 OpenACC 程式。
- 以 OpenACC 指示詞與子句注解程式碼。
- 使用 OpenACC API 與函式庫。
- 為 OpenACC 程式進行效能分析、除錯與最佳化。
課程形式
- 互動式講授與討論。
- 大量的練習與實作。
- 在即時實驗室環境中進行實作演練。
客製化選項
- 如需針對此課程要求客製化培訓,請聯繫我們以安排。
感謝您提交詢問!我們的一位團隊成員將在短時間內與您聯繫。
感謝您提交預訂!我們的一位團隊成員將在短時間內與您聯繫。
課程簡介
簡介
- 什麼是 OpenACC?
- OpenACC 與 OpenCL、CUDA、SYCL 的比較
- OpenACC 功能與架構概覽
- 設定開發環境
入門
- 在 Visual Studio Code 中建立 OpenACC 專案
- 探索專案結構與檔案
- 編譯並執行程式
- 使用 printf 與 fprintf 顯示輸出結果
OpenACC 指示詞與子句
- 了解 OpenACC 指示詞與子句
- 使用平行指示詞建立平行區域
- 使用 kernels 指示詞以進行編譯器管理的平行化
- 使用 loop 指示詞平行化迴圈
- 以 data 指示詞管理資料移動
- 以 update 指示詞同步化資料
- 以 cache 指示詞提升資料重用率
- 以 routine 指示詞建立裝置函式
- 以 wait 指示詞同步化事件
OpenACC API
- 了解 OpenACC API 的角色
- 查詢裝置資訊與功能
- 設定裝置號碼與類型
- 處理錯誤與例外狀況
- 建立並同步化事件
OpenACC 函式庫與互操作性
- 了解 OpenACC 函式庫與互操作性
- 使用數學、隨機數與複雜數列函式庫
- 與其他模型整合(CUDA、OpenMP、MPI)
- 與 GPU 函式庫整合(cuBLAS、cuFFT)
OpenACC 工具
- 了解 OpenACC 在開發中的應用
- 為 OpenACC 程式進行效能分析與除錯
- 使用 PGI 編譯器、NVIDIA Nsight Systems、Allinea Forge 進行效能分析
最佳化
- 影響 OpenACC 程式效能的因素
- 最佳化資料局部性並減少傳輸
- 最佳化迴圈平行化與合併
- 最佳化核心平行化與合併
- 最佳化向量化與自動調參
總結與下一步
最低要求
- 了解 C/C++ 或 Fortran 語言以及平行程式設計的概念
- 具備電腦架構與記憶體階層的基礎知識
- 熟悉指令列工具與程式碼編輯器
受眾
- 希望學習如何使用 OpenACC 為異構裝置撰寫程式並發揮其平行運算能力的開發人員
- 希望編寫可在不同平台與裝置上執行的可攜式且具擴充性的程式碼的開發人員
- 希望探索異構程式設計高階層面,並提升程式生產力的程式設計師
28 小時
公開培訓課程需要5名以上參與者。
使用 OpenACC 進行 GPU 程式設計培訓 - 訂單
使用 OpenACC 進行 GPU 程式設計培訓 - 詢問
使用 OpenACC 進行 GPU 程式設計 - 咨詢詢問
即將到來的課程
相關課程
使用華為 Ascend 與 CANN 開發 AI 應用程式
21 小時
此由講師主導的 台灣 培訓課程,指導中級 AI 工程師使用華為 Ascend 平台與 CANN 工具包建構並最佳化神經網路模型。學員將配置環境、利用 MindSpore 開發應用程式,並部署至邊緣或雲端環境。
更多...
使用CANN與Ascend AI處理器部署AI模型
14 小時
這門台灣即時培訓課程,指引中階AI開發人員使用CANN工具包在Ascend處理器上部署模型。學習如何轉換PyTorch和TensorFlow等框架,優化效能,並除錯問題,以實現高效的邊緣端和雲端推理場景。
更多...
使用 CloudMatrix 進行 AI 推論與部署
21 小時
This instructor-led training in 台灣 introduces CloudMatrix for scalable AI inference. Learn to deploy, optimize, and monitor models using CANN and MindSpore. Hands-on exercises cover packaging, conversion, serving, and performance tuning for real-time and batch workloads.
更多...
Biren AI加速器的GPU編程
21 小時
This live training in 台灣 equips developers with the skills to program and optimize applications on Biren AI accelerators. Participants will learn the GPU architecture, set up the SDK, and translate CUDA code to Biren. It focuses on performance tuning and debugging techniques.
更多...
使用 BANGPy 和 Neuware 進行寒武紀 MLU 開發
21 小時
This instructor-led live training in 台灣 equips developers with the skills to build and deploy AI models using BANGPy and Neuware on Cambricon MLUs. Participants will configure environments, develop optimized models, and integrate MLU acceleration into edge and data center applications.
更多...
CANN for AI Framework Developers 入門
7 小時
This live training in 台灣 introduces the CANN toolkit for AI framework developers. Learn to set up environments, convert models, and deploy applications on Ascend hardware using MindSpore, TensorFlow, or PyTorch, covering the full workflow from training to inference.
更多...
CANN 邊緣 AI 部署
14 小時
This instructor-led, live training in 台灣 covers the core concepts and hands-on fundamentals of deploying AI models on Ascend edge devices using the CANN toolkit, helping participants build practical skills for compiling, optimizing, and managing constrained environments.
更多...
理解華為的 AI 計算堆疊:從 CANN 到 MindSpore
14 小時
This instructor-led live training in 台灣 explores Huawei's AI stack, from the CANN SDK to the MindSpore framework. It helps beginners and intermediate professionals understand how these components integrate on Ascend hardware for lifecycle management and deployment.
更多...
利用 CANN SDK 優化神經網路效能
14 小時
Optimize neural network inference performance on Ascend AI processors with this advanced, instructor-led training in 台灣. Explore CANN's runtime architecture, leveraging the Graph Engine, TIK, and TVM for profiling, custom operator development, and memory bottleneck resolution.
更多...
CANN SDK 用於電腦視覺與自然語言處理管線
14 小時
This instructor-led training in 台灣 covers deploying and optimizing CV and NLP models using the CANN SDK for Ascend hardware. Participants will learn to convert models, integrate them into live pipelines, and enhance inference performance for real-time detection and analysis.
更多...
使用 CANN TIK 和 TVM 構建自訂 AI 運算元
14 小時
This instructor-led, live training in 台灣 equips advanced developers with the skills to build, deploy, and tune custom AI operators. Participants will master CANN TIK and Apache TVM integration, enabling advanced optimization and scheduling on Huawei Ascend hardware for real-world performance.
更多...
將 CUDA 應用程式遷移至中國 GPU 架構
21 小時
Migrate CUDA applications to Chinese GPU architectures like Huawei Ascend and Biren in 台灣. This instructor-led course guides advanced programmers through code translation and performance optimization, covering hands-on labs for porting CUDA codebases to new SDKs.
更多...
Ascend、Biren與寒武紀的性能優化
21 小時
Optimize AI workloads on Ascend, Biren, and Cambricon with this hands-on training in 台灣. Learn to benchmark models, identify bottlenecks, and apply graph, kernel, and operator-level optimizations. Tune deployment pipelines to enhance throughput and latency across these leading platforms.
更多...