2026.09.04 · FRI SEOUL VOL.02 / NO.09
Search Letter
delay″
Unbound Visions, Inspired Living.
Beauty Health Technology Gastronomy Culture Dispatches
2026.09.04
SearchLetter
delay″
Beauty Health Technology Gastronomy Culture Dispatches
← Technology 2024.04.05 5 min read Irang 글 · 편집부
Index
Technology · The Maker

엔비디아, CUDA 6 출시…병렬 프로그래밍 간소화하고 성능 최대 8배 향상

무슨 발표인가

  • CPU 라이브러리 교체로 애플리케이션 최대 8배 가속
  • Unified Memory로 CPU/GPU 메모리 관리 자동화
  • 8개 GPU까지 확장 시 9테라플롭스 이상

원문 (영어)

SANTA CLARA, CA -- NVIDIA today announced NVIDIA CUDA 6, the latest version of the world's most pervasive parallel computing platform and programming model. The CUDA 6 platform makes parallel programming easier than ever, enabling software developers to dramatically decrease the time and effort required to accelerate their scientific, engineering, enterprise and other applications with GPUs.

It offers new performance enhancements that enable developers to instantly accelerate applications up to 8X by simply replacing existing CPU-based libraries. Key features of CUDA 6 include: Unified Memory -- Simplifies programming by enabling applications to access CPU and GPU memory without the need to manually copy data from one to the other, and makes it easier to add support for GPU acceleration in a wide range of programming languages.

Drop-in Libraries -- Automatically accelerates applications' BLAS and FFTW calculations by up to 8X by simply replacing the existing CPU libraries with the GPU-accelerated equivalents. Multi-GPU Scaling -- Re-designed BLAS and FFT GPU libraries automatically scale performance across up to eight GPUs in a single node, delivering over nine teraflops of double precision performance per node, and supporting larger workloads than ever before (up to 512GB).

Multi-GPU scaling can also be used with the new BLAS drop-in library. "By automatically handling data management, Unified Memory enables us to quickly prototype kernels running on the GPU and reduces code complexity, cutting development time by up to 50 percent," said Rob Hoekstra, manager of Scalable Algorithms Department at Sandia National Laboratories.

"Having this capability will be very useful as we determine future programming model choices and port more sophisticated, larger codes to GPUs."

원문: NVIDIA News — "NVIDIA Dramatically Simplifies Parallel Programming With CUDA 6" (2024-04-05) 공식 원문: https://nvidianews.nvidia.com/news/nvidia-dramatically-simplifies-parallel-programming-with-cuda-6

NVIDIA News
delayseconds · 2024.04.05
Read next
Technology
보나주, AWS와 AI 음성 에이전트 통합 협력 발표
Technology
Meta, 미 연방정부에 Llama AI 무료 제공
Culture
완판 신화를 스스로 단종시킨 펜티뷰티의 승부수
d″
이 기사가 좋았다면, 월 1회 레터로 받아보세요.
Subscribe
← Technology

엔비디아, CUDA 6 출시…병렬 프로그래밍 간소화하고 성능 최대 8배 향상

2024.04.05 · 5 min · Irang

무슨 발표인가

원문 (영어)

SANTA CLARA, CA -- NVIDIA today announced NVIDIA CUDA 6, the latest version of the world's most pervasive parallel computing platform and programming model. The CUDA 6 platform makes parallel programming easier than ever, enabling software developers to dramatically decrease the time and effort required to accelerate their scientific, engineering, enterprise and other applications with GPUs.

It offers new performance enhancements that enable developers to instantly accelerate applications up to 8X by simply replacing existing CPU-based libraries. Key features of CUDA 6 include: Unified Memory -- Simplifies programming by enabling applications to access CPU and GPU memory without the need to manually copy data from one to the other, and makes it easier to add support for GPU acceleration in a wide range of programming languages.

Drop-in Libraries -- Automatically accelerates applications' BLAS and FFTW calculations by up to 8X by simply replacing the existing CPU libraries with the GPU-accelerated equivalents. Multi-GPU Scaling -- Re-designed BLAS and FFT GPU libraries automatically scale performance across up to eight GPUs in a single node, delivering over nine teraflops of double precision performance per node, and supporting larger workloads than ever before (up to 512GB).

Multi-GPU scaling can also be used with the new BLAS drop-in library. "By automatically handling data management, Unified Memory enables us to quickly prototype kernels running on the GPU and reduces code complexity, cutting development time by up to 50 percent," said Rob Hoekstra, manager of Scalable Algorithms Department at Sandia National Laboratories.

"Having this capability will be very useful as we determine future programming model choices and port more sophisticated, larger codes to GPUs."

원문: NVIDIA News — "NVIDIA Dramatically Simplifies Parallel Programming With CUDA 6" (2024-04-05) 공식 원문: https://nvidianews.nvidia.com/news/nvidia-dramatically-simplifies-parallel-programming-with-cuda-6

NVIDIA News
delayseconds · 2024.04.05
Read next
Technology
보나주, AWS와 AI 음성 에이전트 통합 협력 발표
Technology
Meta, 미 연방정부에 Llama AI 무료 제공
Culture
완판 신화를 스스로 단종시킨 펜티뷰티의 승부수
delayseconds
Unbound Visions, Inspired Living..
얽매이지 않는 시선, 영감을 주는 삶
Beauty Health Technology
Gastronomy Culture Dispatches
LetterSearchAbout
© 2026 delayseconds The mark uses the double prime ″ (U+2033)