이미지 검색을 사용해 보세요
검색창 이전화면 이전화면
최근 검색어
인기 검색어

소득공제 수입
직수입양서 Introduction to GPU Programming
Paperback, 1st Edition
Kuldeep Singh Kaswan, Balamurugan Balusamy, Rekha R. Nair, Jagjit Singh Dhatterwal, Kiran Malik
Chapman and Hall/CRC 2027.03.05.
가격
129,220
10 116,290
YES포인트?
5,820원 (5%) 마니아추가적립
5만원 이상 구매 시 2천원 추가 적립
결제혜택
최대 2,000원 즉시할인

책소개

목차

Chapter
1. Evolution of Parallel Computing.
1.1 The Need for High-Performance Computing.
1.2 From CPUs to GPUs: Architectural Shifts.
1.3 Applications Driving GPU Programming. Chapter
2. Fundamentals of GPU Architecture.
2.1 GPU vs CPU: Key Differences.
2.2 SIMD, SIMT, and Thread-Level Parallelism.
2.3 Memory Hierarchy in GPUs.
2.4 Execution Model and Warp Scheduling. Chapter
3. Getting Started with GPU Programming.
3.1 Introduction to CUDA and OpenCL.
3.2 GPU Programming Workflow.
3.3 Environment and Installing.
3.4 Hello GPU: First GPU Program. Chapter
4. CUDA Programming Essentials.
4.1 Kernel Functions and Thread Hierarchy.
4.2 Grids, Blocks, and Threads.
4.3 Synchronization and Barriers.
4.4 Error Handling in CUDA. Chapter
5. Memory Management in GPUs.
5.1 Types of GPU Memory: Global, Shared, Local, Constant.
5.2 Memory Allocation and Transfer.
5.3 Coalesced Memory Access.
5.4 Optimization Strategies.
5.5 GPU Cache Architecture. Chapter
6. Performance Optimization Techniques in GPU.
6.1 Occupancy and Thread Divergence.
6.2 Shared Memory Utilization.
6.3 Latency Hiding and Instruction Pipelining.
6.4 Profiling GPU Programs.
6.5 Example 1: Using Shared Memory for Faster Access. Chapter
7. Parallel Algorithms on GPUs.
7.1 Introduction Parallel Reduction.
7.2 Scan (Prefix Sum) Algorithms.
7.3 Sorting on GPUs.
7.4 Matrix Operations and Linear Algebra Kernels.
7.5 Example 1: Parallel Reduction (Sum). Chapter
8. OpenCL Programming Model.
8.1 Introduction OpenCL Architecture and Execution Model.
8.2 Kernels and Work-Items.
8.3 Memory Objects and Buffers.
8.4 Comparing CUDA and OpenCL.
8.5 Example 1: Simple OpenCL Kernel. Chapter
9. GPU Libraries and Frameworks.
9.1 cuBLAS and cuFFT.
9.2 Thrust Library for Parallel Algorithms.
9.3 TensorRT and GPU-Accelerated AI Libraries.
9.4 Cross-Vendor Libraries and Portable Frameworks.
9.5 Integration with Python (PyCUDA, Numba). Chapter
10. GPU Programming for Data Science.
10.1 GPU Acceleration in Data Analytics.
10.2 RAPIDS Framework.
10.3 GPU-Accelerated Machine Learning.
10.4 Deep Learning with GPUs.
10.5 Bias Detection and Fairness in AI Lending Decisions derived from NLP Sentiment. Chapter
11. Advancements in GPU Programming.
11.1 Multi-GPU Programming.
11.2 Unified Memory and Heterogeneous Computing.
11.3 Dynamic Parallelism.
11.4 GPU Virtualization and Cloud GPUs. Chapter
12. OpenMP Offloading Architecture and Execution Model.
12.1 How the OpenMP Target Model Maps Work to GPU Devices.
12.2 Overview of Teams, Threads, and SIMD Constructs.
12.3 Interaction with Vendor Backends (NVIDIA, AMD, Intel, ARM).
12.4 Compilation Pipeline: Clang/LLVM, GCC, Intel oneAPI Tooling.
12.5 Example 1: OpenMP Target Offload to GPU. Chapter
13. Case Studies and Applications.
13.1 Scientific Simulations on GPUs.
13.2 Image and Video Processing.
13.3 Cryptography and Blockchain Acceleration.
13.4 Real-Time Systems and Gaming Engines.
13.5 Example 1: GPU Image Smoothing (3×3 Filter). Chapter
14. Debugging and Profiling Tools.
14.1 NVIDIA Nsight and Visual Profiler.
14.2 Performance Counters and Tracing.
14.3 Debugging CUDA Applications.
14.4 Benchmarking Best Practices.
14.5 Example 1: Adding NVTX Annotations for Profiling. Chapter
15. Challenges and Future of GPU Programming.
15.1 Power Consumption and Energy Efficiency.
15.2 Portability Across GPU Architectures.
15.3 GPUs vs TPUs and Other Accelerators.
15.4 Future Trends in GPU Computing.
15.5 Example 1: Measuring GPU Power Using NVML. Chapter
16. Multi-GPU, Multi-Node, and Cloud-Native GPU Execution.
16.1 NCCL, RCCL, and Distributed Communication Topologies.
16.2 Pipeline Parallelism, ZeRO, and Sharded Training Strategies.
16.3 Kubernetes GPU Orchestration, MIG, MPS, and Virtualized GPUs.
16.4 Cloud GPU Computing Workflows on AWS/GCP/Azure.
16.5 Performance Engineering in Large-Scale Distributed Environments.
16.6 NCCL All-Reduce. Chapter
17. GPU Acceleration in Emerging Computational Domains.
17.1 GPU Acceleration in Computer Vision and Imaging Pipelines.
17.2 AI/ML Pipelines Accelerated by GPUs.
17.3 Blockchain and Cryptography on GPUs.
17.4 Quantum Computing Simulation on GPUs.

품목정보

발행 예정일
미정
쪽수, 무게, 크기
360쪽 | 160*230mm
ISBN13
9781041306795

리뷰/한줄평0

리뷰

첫번째 리뷰어가 되어주세요.

한줄평

첫번째 한줄평을 남겨주세요.