copy from http://developer.amd.com/community/blog/2014/10/31/opencl-2-0-pipes/

OpenCL™ 2.0 – Pipes

In the previous post, we saw one of the important features of OpenCL™ 2.0, Shared Virtual Memory (SVM). In this blog, we will see another feature of OpenCL 2.0 called “pipes”.

To get the most from our discussions, we recommend the same approach as in the previous post:

Review the code snippets in each blog post along with the explanatory text.

Download the AMD OpenCL 2.0 Driver located here. This page also has a complete list of supported platforms.

Download the sample code from our OpenCL 2.0 Samples page here.

Write your own OpenCL 2.0 samples and share your results with the OpenCL community.

The code will run on a variety of AMD platforms, such as the Radeon HD8000 series. The driver page has a complete list of supported product families.

Overview

OpenCL 2.0 introduces a new mechanism for passing data between kernels, called “pipes.” A pipe is essentially a structured buffer containing some space for a set of “packets”—kernel-specified type objects. As the name suggests, these packets of data are ordered in the pipe. There is a write end of the pipe into which the data is written and a read end of the pipe from which the data is read. A pipe is essentially an addition to buffer objects, such as buffers and images. Pipes can only be accessed using the built-in functions provided by the kernel and cannot be accessed from the host.

Special built-in functions, read_pipe and write_pipe, provide access to pipes from the kernel. A given kernel can either read from or write to a pipe, but not both. Pipes are only coherent at the standard synchronization points; the result of concurrent accesses to the same pipe by multiple kernels (even hardware permitting) is undefined. The host side cannot access a pipe.

It is easy to create pipes. On the host, invoke clCreatePipe, and you are done.

You can use the pipes for a variety of functions. You can pass the pipes between kernels. Even better, combine pipes with the device-size enqueue feature in OpenCL 2.0 to dynamically construct computational data flow graphs.

Pipes come in two types: a read pipe, from which a number of packets can be read, and a write pipe, to which a number of packets can be written.

Note: You cannot write to a pipe specified as read-only, nor can you read from a pipe specified as write-only. You cannot both read from and write to a pipe at the same time.

Functions for Accessing Pipes

OpenCL 2.0 adds a new host API function to create a pipe.

cl_mem clCreatePipe ( cl_context context, cl_mem_flags flags,
cl_uint pipe_packet_size, cl_uint pipe_max_packets,
const cl_pipe_properties * properties,
cl_int *errcode_ret)

The memory allocated in this function can pass to kernels as read-only or write-only pipes. The pipe objects can only be passed as kernel arguments or kernel functions and cannot be declared inside a kernel or as program-scoped objects.

Also, the OpenCL 2.0 spec adds a set of built-in functions for operating on the pipes. The important ones are:

  • read_pipe (pipe p, gentype *ptr): for reading packet from pipe p into ptr.
  • write_pipe (pipe p, gentype *ptr): for writing packet pointed to by ptr to pipe p.

To ensure you have enough space in the pipe structure for reading and writing (before you actually do it), you can use built-in functions to “reserve” enough space. For example, you could reserve room by calling reserve_read_pipe or reserve_write_pipe. These functions return a reservation ID, which can be used when the actual operations are performed. Similarly, the standard has built-in functions for workgroup level reservations, such as work_group_reserve_read_pipe and work_group_reserve_write_pipe and for the workgroup order (in the program). These workgroup built-in functions operate at the workgroup level. Ordering across workgroups is undefined. Calls to commit_read_pipe and commit_write_pipe, as the names suggest, commit the actual operations (read/write).

Using Pipes—A Simple Example

Let’s look at a typical usage of pipes in the example code (SimplePipe). The code contains two kernels: producer_kernel, which writes to the pipe, and consumer_kernel, which reads from the same pipe. In the example, the producer writes a sequence of random numbers; the consumer reads them and creates a histogram.

The host creates the pipe, which both kernels will use, as follows:

rngPipe = clCreatePipe(context,
CL_MEM_READ_WRITE,
szPipePkt,
szPipe,
NULL,
&status);

This code makes a pipe that the program kernels can access (read/write). The host creates two kernels, producer_kernel and consumer_kernel. The producer kernel first reserves enough space for the write pipe:

//reserve space in pipe for writing random numbers.
reserve_id_t rid = work_group_reserve_write_pipe(rng_pipe, szgr);

Next, the kernel writes and commits to the pipe by invoking the following functions:

write_pipe(rng_pipe,rid,lid, &gfrn);
work_group_commit_write_pipe(rng_pipe, rid);

Similarly, the consumer kernel reads from the pipe:

//reserve pipe for reading
reserve_id_t rid = work_group_reserve_read_pipe(rng_pipe, szgr);
if(is_valid_reserve_id(rid)) {
//read random number from the pipe.
read_pipe(rng_pipe,rid,lid, &rn);
work_group_commit_read_pipe(rng_pipe, rid);
}

The consumer_kernel then uses this set of random number and constructs the histogram. The CPU creates the same histogram and verifies whether the histogram created by the kernel is correct. Here, lid is the local id of the work item, obtained by get_local_id(0).

The example code demonstrates how you can use a pipe as a convenient data structure that allows two kernels to communicate. It’s really pretty simple.

In OpenCL 1.2, this kind of communication typically involves the host – although kernels can communicate without returning control to the host. Pipes, however, ease programming by reducing the amount of code that some applications require. There are additional examples of pipes used in conjunction with device enqueue, which we will explore in later blogs in this series.

To conclude, using pipes in OpenCL 2.0 can make your code simpler and more readable. Don’t believe us? Write your own OpenCL 2.0 programs and tell us about the difference!

Sample code and readme

The sample code demonstrates the use of the pipes feature of OpenCL.2.0 using the SimplePipe (producer/consumer kernels) sample:

  • The sample code and readme are provided here.
  • Install AMD OpenCL 2.0 Driver located here. This page also has a complete list of supported platforms.
  • Build the sample using the instructions in the readme.
  • Let us know your feedback and comments on the Developer Central OpenCL forum.

– Prakash Raghavendra

Dr. Prakash Raghavendra is a technical lead at AMD. He has several years of experience in developing compilers and run-time. His postings are his own opinions and may not represent AMD’s positions, strategies or opinions. Links to third party sites, and references to third party trademarks, are provided for convenience and illustrative purposes only. Unless explicitly stated, AMD is not responsible for the contents of such links, and no third party endorsement of AMD or any of its products is implied.

OpenCL™ 2.0 – Pipes的更多相关文章

  1. OpenCL介绍

    OpenCL(全称Open Computing Language,开放运算语言)是第一个面向异构系统通用目的并行编程的开放式.免费标准,也是一个统一的编程环境,便于软件开发人员为高性能计算服务器.桌面 ...

  2. OpenCL Kernel设计优化

    使用Intel® FPGA SDK for OpenCL™ 离线编译器,不需要调整kernel代码便可以将其最佳的适应于固定的硬件设备,而是离线编译器会根据kernel的要求自适应调整硬件的结构. 通 ...

  3. 基于SoCkit的opencl实验1-基础例程

    基于SoCkit的opencl实验1-基础例程 准备软硬件 Arrow SoCkit Board 4GB or larger microSD Card Quartus II v14.1 SoCEDS ...

  4. 面向OPENCL的ALTERA SDK

    面向OPENCL的ALTERA SDK 使用面向开放计算语言 (OpenCL™) 的 Altera® SDK,用户可以抽象出传统的硬件 FPGA 开发流程,采用更快.更高层面的软件开发流程.在基于 x ...

  5. OpenCL中三种内存创建image的效率对比

    第一种:使用ION: cl_mem_ion_host_ptr ion_host_ptr1; ion_host_ptr1.ext_host_ptr.allocation_type = CL_MEM_IO ...

  6. OpenCL科普及在ubuntu 16.04 LTS上的安装

    OpenCL(Open Computing Language,开放计算语言)是一个为异构平台编写程序的框架,此异构平台可由CPU.GPU.DSP.FPGA或其他类型的处理器與硬體加速器所组成.Open ...

  7. OpenCL 查询平台和设备

    ▶ 查询平台和设备的代码以结果,放在这里方便以后逐渐扩充和查询(没有营养) #include <stdio.h> #include <stdlib.h> #include &l ...

  8. OpenCL Hello World

    ▶ OpenCL 的环境配置与第一个程序 ● CUDA 中自带 OpenCL 需要的头文件和库,直接拉近项目里边去就行:AMD 需要下载 AMD APP SDK(https://community.a ...

  9. OpenCL双边滤波实现美颜功能

    OpenCL是一个并行异构计算的框架,包括intel,AMD,英伟达等等许多厂家都有对它的支持,不过英伟达只到1.2版本,主要发展自己的CUDA去了.虽然没有用过CUDA,但个人感觉CUDA比Open ...

随机推荐

  1. SSL证书是“盾牌“还是”鸡肋“?

    德国联邦安全与IT办公室(BSI,职能相当于美国的国家安全与信息技术局)近日发布公告警告:网络攻击者冒充其发布了“关于Meltdown与Spectre攻击信息”的垃圾邮件,该邮件中包含指向修复补丁的页 ...

  2. devstack环境中不能创建cinder volume

    刚安装好的devstack环境中无法成功创建cinder volume,创建的volume的status为error:在cinder scheduler中看到失败log:2015-10-15 14:1 ...

  3. Sqlserver 查询 临时字段

    临时字段格式   字段名=N'字段值' 例子如下: select cEmp_C, cEmp_N, oper_id=N'001', log_pw=N'123', sSex, cDept_C, cDept ...

  4. centos下使用fdisk扩展分区容量大小

    硬盘空间为20G,VMware增加磁盘大小,需要再增加10G空间 扩展完后,重启系统,再次使用fdisk -l查看,会发现硬盘空间变大了: 重新创建分区,调整分区信息 本次实验主要对/dev/sda4 ...

  5. Django进阶Template篇002 - 模板包含和继承

    包含 {% include %} 允许在模板中包含其他模板的内容. {% include "foo/bar.html" %} {% include template_name %} ...

  6. Microsoft Edge Certified with EBS 12.1 and 12.2

    I am very pleased to announce that Microsoft Edge is certified as a new browser for Oracle E-Busines ...

  7. Prism5.0新内容 What's New in Prism Library 5.0 for WPF(英汉对照版)

    Prism 5.0 includes guidance in several new areas, resulting in new code in the Prism Library for WPF ...

  8. 内存保护机制及绕过方案——利用未启用SafeSEH模块绕过SafeSEH

    前言:之前关于safeSEH保护机制的原理等信息,可在之前的博文(内存保护机制及绕过方案中查看). 利用未启用SafeSEH模块绕过SafeSEH ⑴.  原理分析: 一个不是仅包含中间语言(1L)且 ...

  9. 《修炼Java开发技术 在架构中体验设计模式和算法之美》 - 书摘精要

    (P7) 建议直接加入到软件公司中去,这样会学到很多实际的东西: 程序员最主要的发展方向是资深技术专家,无论是 Java..Net 还是数据库领域,都要首先成为专家,然后才可能继续发展为架构师: 增强 ...

  10. 剑指offer--26.顺时针打印矩阵

    1,2,3,45,6,7,88,10,11,1213,14,15,16 每次输出第一行,然后删除第一行,逆时针旋转剩下的矩阵. ------------------------------------ ...