- Java 双重比较
- java - 比较器与 Apache BeanComparator
- Objective-C 完成 block 导致额外的方法调用?
- database - RESTful URI 是否应该公开数据库主键?
我想知道 .cu
编译单元中隐式包含了哪些 header 。例如,让我们举一些基本的例子:
#include <iostream>
__global__ void kernel(int a, int b, int *c)
{
*c = a * b;
}
int main()
{
int c, *dev_c;
cudaMalloc((void **) &dev_c, sizeof(*dev_c));
kernel<<<1,1>>>(5, 3, dev_c);
cudaMemcpy(&c, dev_c, sizeof(*dev_c), cudaMemcpyDeviceToHost);
cudaFree(dev_c);
std::cout << c << std::endl;
}
根据 C++ 标准,所有函数都必须在使用前声明(在 C99 之后的 C 中也是如此),所以我猜想这些 CUDA 头文件包含在“幕后”。
这个问题很重要,因为我想在 Visual C++ 中为 CUDA API 调用获取 IntelliSense,但我还没有找到任何关于此的“官方”引用。
最佳答案
我在 Linux 上为一个空文件运行了以下命令:
nvcc --generate-dependencies empty.cu
这导致了以下列表:
empty.o : empty.cu \
/usr/include/stdc-predef.h \
/opt/cuda/bin/..//include/cuda_runtime.h \
/opt/cuda/bin/..//include/host_config.h \
/usr/include/features.h \
/usr/include/x86_64-linux-gnu/sys/cdefs.h \
/usr/include/x86_64-linux-gnu/bits/wordsize.h \
/usr/include/x86_64-linux-gnu/gnu/stubs.h \
/usr/include/x86_64-linux-gnu/gnu/stubs-64.h \
/opt/cuda/bin/..//include/builtin_types.h \
/opt/cuda/bin/..//include/device_types.h \
/opt/cuda/bin/..//include/host_defines.h \
/opt/cuda/bin/..//include/driver_types.h \
/usr/lib/gcc/x86_64-linux-gnu/4.8/include-fixed/limits.h \
/usr/lib/gcc/x86_64-linux-gnu/4.8/include-fixed/syslimits.h \
/usr/include/limits.h \
/usr/include/x86_64-linux-gnu/bits/posix1_lim.h \
/usr/include/x86_64-linux-gnu/bits/local_lim.h \
/usr/include/linux/limits.h \
/usr/include/x86_64-linux-gnu/bits/posix2_lim.h \
/usr/include/x86_64-linux-gnu/bits/xopen_lim.h \
/usr/include/x86_64-linux-gnu/bits/stdio_lim.h \
/usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h \
/opt/cuda/bin/..//include/surface_types.h \
/opt/cuda/bin/..//include/texture_types.h \
/opt/cuda/bin/..//include/vector_types.h \
/opt/cuda/bin/..//include/channel_descriptor.h \
/opt/cuda/bin/..//include/cuda_runtime_api.h \
/opt/cuda/bin/..//include/cuda_device_runtime_api.h \
/opt/cuda/bin/..//include/driver_functions.h \
/opt/cuda/bin/..//include/vector_functions.h \
/opt/cuda/bin/..//include/vector_functions.hpp \
/opt/cuda/bin/..//include/common_functions.h \
/usr/include/string.h \
/usr/include/xlocale.h \
/usr/include/time.h \
/usr/include/x86_64-linux-gnu/bits/time.h \
/usr/include/x86_64-linux-gnu/bits/types.h \
/usr/include/x86_64-linux-gnu/bits/typesizes.h \
/usr/include/x86_64-linux-gnu/bits/timex.h \
/usr/include/c++/4.8/new \
/usr/include/x86_64-linux-gnu/c++/4.8/bits/c++config.h \
/usr/include/x86_64-linux-gnu/c++/4.8/bits/os_defines.h \
/usr/include/x86_64-linux-gnu/c++/4.8/bits/cpu_defines.h \
/usr/include/c++/4.8/exception \
/usr/include/c++/4.8/bits/atomic_lockfree_defines.h \
/usr/include/stdio.h \
/usr/include/libio.h \
/usr/include/_G_config.h \
/usr/include/wchar.h \
/usr/lib/gcc/x86_64-linux-gnu/4.8/include/stdarg.h \
/usr/include/x86_64-linux-gnu/bits/sys_errlist.h \
/usr/include/stdlib.h \
/usr/include/x86_64-linux-gnu/bits/waitflags.h \
/usr/include/x86_64-linux-gnu/bits/waitstatus.h \
/usr/include/endian.h \
/usr/include/x86_64-linux-gnu/bits/endian.h \
/usr/include/x86_64-linux-gnu/bits/byteswap.h \
/usr/include/x86_64-linux-gnu/bits/byteswap-16.h \
/usr/include/x86_64-linux-gnu/sys/types.h \
/usr/include/x86_64-linux-gnu/sys/select.h \
/usr/include/x86_64-linux-gnu/bits/select.h \
/usr/include/x86_64-linux-gnu/bits/sigset.h \
/usr/include/x86_64-linux-gnu/sys/sysmacros.h \
/usr/include/x86_64-linux-gnu/bits/pthreadtypes.h \
/usr/include/alloca.h \
/usr/include/x86_64-linux-gnu/bits/stdlib-float.h \
/usr/include/assert.h \
/opt/cuda/bin/..//include/math_functions.h \
/usr/include/math.h \
/usr/include/x86_64-linux-gnu/bits/huge_val.h \
/usr/include/x86_64-linux-gnu/bits/huge_valf.h \
/usr/include/x86_64-linux-gnu/bits/huge_vall.h \
/usr/include/x86_64-linux-gnu/bits/inf.h \
/usr/include/x86_64-linux-gnu/bits/nan.h \
/usr/include/x86_64-linux-gnu/bits/mathdef.h \
/usr/include/x86_64-linux-gnu/bits/mathcalls.h \
/usr/include/c++/4.8/cmath \
/usr/include/c++/4.8/bits/cpp_type_traits.h \
/usr/include/c++/4.8/ext/type_traits.h \
/usr/include/c++/4.8/cstdlib \
/opt/cuda/bin/..//include/math_functions.hpp \
/opt/cuda/bin/..//include/math_functions_dbl_ptx3.h \
/opt/cuda/bin/..//include/math_functions_dbl_ptx3.hpp \
/opt/cuda/bin/..//include/cuda_surface_types.h \
/opt/cuda/bin/..//include/cuda_texture_types.h \
/opt/cuda/bin/..//include/device_functions.h \
/opt/cuda/bin/..//include/device_functions.hpp \
/opt/cuda/bin/..//include/device_atomic_functions.h \
/opt/cuda/bin/..//include/device_atomic_functions.hpp \
/opt/cuda/bin/..//include/device_double_functions.h \
/opt/cuda/bin/..//include/device_double_functions.hpp \
/opt/cuda/bin/..//include/sm_20_atomic_functions.h \
/opt/cuda/bin/..//include/sm_20_atomic_functions.hpp \
/opt/cuda/bin/..//include/sm_32_atomic_functions.h \
/opt/cuda/bin/..//include/sm_32_atomic_functions.hpp \
/opt/cuda/bin/..//include/sm_35_atomic_functions.h \
/opt/cuda/bin/..//include/sm_20_intrinsics.h \
/opt/cuda/bin/..//include/sm_20_intrinsics.hpp \
/opt/cuda/bin/..//include/sm_30_intrinsics.h \
/opt/cuda/bin/..//include/sm_30_intrinsics.hpp \
/opt/cuda/bin/..//include/sm_32_intrinsics.h \
/opt/cuda/bin/..//include/sm_32_intrinsics.hpp \
/opt/cuda/bin/..//include/sm_35_intrinsics.h \
/opt/cuda/bin/..//include/surface_functions.h \
/opt/cuda/bin/..//include/surface_functions.hpp \
/opt/cuda/bin/..//include/texture_fetch_functions.h \
/opt/cuda/bin/..//include/texture_fetch_functions.hpp \
/opt/cuda/bin/..//include/texture_indirect_functions.h \
/opt/cuda/bin/..//include/texture_indirect_functions.hpp \
/opt/cuda/bin/..//include/surface_indirect_functions.h \
/opt/cuda/bin/..//include/surface_indirect_functions.hpp \
/opt/cuda/bin/..//include/device_launch_parameters.h
所以默认情况下 nvcc 已经包含很多 CUDA header (在我的例子中是 /opt/cuda/...
)。
然后我创建了一个包含以下内容的文件 test.cpp
并运行 g++ -H -D__CUDACC__
(伪造 nvcc):
#include "/opt/cuda/bin/..//include/cuda_runtime.h"
这导致了以下列表:
. /opt/cuda/bin/..//include/cuda_runtime.h
.. /opt/cuda/bin/..//include/host_config.h
... /usr/include/features.h
.... /usr/include/x86_64-linux-gnu/sys/cdefs.h
..... /usr/include/x86_64-linux-gnu/bits/wordsize.h
.... /usr/include/x86_64-linux-gnu/gnu/stubs.h
..... /usr/include/x86_64-linux-gnu/gnu/stubs-64.h
.. /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/device_types.h
.... /opt/cuda/bin/..//include/host_defines.h
... /opt/cuda/bin/..//include/driver_types.h
.... /usr/lib/gcc/x86_64-linux-gnu/4.8/include-fixed/limits.h
..... /usr/lib/gcc/x86_64-linux-gnu/4.8/include-fixed/syslimits.h
...... /usr/lib/gcc/x86_64-linux-gnu/4.8/include-fixed/limits.h
....... /usr/include/limits.h
........ /usr/include/x86_64-linux-gnu/bits/posix1_lim.h
......... /usr/include/x86_64-linux-gnu/bits/local_lim.h
.......... /usr/include/linux/limits.h
........ /usr/include/x86_64-linux-gnu/bits/posix2_lim.h
........ /usr/include/x86_64-linux-gnu/bits/xopen_lim.h
......... /usr/include/x86_64-linux-gnu/bits/stdio_lim.h
.... /usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h
... /opt/cuda/bin/..//include/surface_types.h
... /opt/cuda/bin/..//include/texture_types.h
... /opt/cuda/bin/..//include/vector_types.h
.... /opt/cuda/bin/..//include/builtin_types.h
..... /opt/cuda/bin/..//include/vector_types.h
.. /opt/cuda/bin/..//include/channel_descriptor.h
... /opt/cuda/bin/..//include/cuda_runtime_api.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/cuda_device_runtime_api.h
.. /opt/cuda/bin/..//include/driver_functions.h
... /opt/cuda/bin/..//include/builtin_types.h
.. /opt/cuda/bin/..//include/vector_functions.h
... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/vector_functions.hpp
.... /opt/cuda/bin/..//include/builtin_types.h
.. /opt/cuda/bin/..//include/common_functions.h
... /opt/cuda/bin/..//include/builtin_types.h
... /usr/include/string.h
.... /usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h
.... /usr/include/xlocale.h
... /usr/include/time.h
.... /usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h
.... /usr/include/x86_64-linux-gnu/bits/time.h
..... /usr/include/x86_64-linux-gnu/bits/types.h
...... /usr/include/x86_64-linux-gnu/bits/wordsize.h
...... /usr/include/x86_64-linux-gnu/bits/typesizes.h
..... /usr/include/x86_64-linux-gnu/bits/timex.h
... /opt/cuda/bin/..//include/math_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /usr/include/math.h
..... /usr/include/x86_64-linux-gnu/bits/huge_val.h
..... /usr/include/x86_64-linux-gnu/bits/huge_valf.h
..... /usr/include/x86_64-linux-gnu/bits/huge_vall.h
..... /usr/include/x86_64-linux-gnu/bits/inf.h
..... /usr/include/x86_64-linux-gnu/bits/nan.h
..... /usr/include/x86_64-linux-gnu/bits/mathdef.h
..... /usr/include/x86_64-linux-gnu/bits/mathcalls.h
..... /usr/include/x86_64-linux-gnu/bits/mathcalls.h
..... /usr/include/x86_64-linux-gnu/bits/mathcalls.h
.... /usr/include/stdlib.h
..... /usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h
..... /usr/include/x86_64-linux-gnu/bits/waitflags.h
..... /usr/include/x86_64-linux-gnu/bits/waitstatus.h
...... /usr/include/endian.h
....... /usr/include/x86_64-linux-gnu/bits/endian.h
....... /usr/include/x86_64-linux-gnu/bits/byteswap.h
........ /usr/include/x86_64-linux-gnu/bits/wordsize.h
........ /usr/include/x86_64-linux-gnu/bits/byteswap-16.h
..... /usr/include/x86_64-linux-gnu/sys/types.h
...... /usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h
...... /usr/include/x86_64-linux-gnu/sys/select.h
....... /usr/include/x86_64-linux-gnu/bits/select.h
........ /usr/include/x86_64-linux-gnu/bits/wordsize.h
....... /usr/include/x86_64-linux-gnu/bits/sigset.h
....... /usr/include/x86_64-linux-gnu/bits/time.h
...... /usr/include/x86_64-linux-gnu/sys/sysmacros.h
...... /usr/include/x86_64-linux-gnu/bits/pthreadtypes.h
....... /usr/include/x86_64-linux-gnu/bits/wordsize.h
..... /usr/include/alloca.h
...... /usr/lib/gcc/x86_64-linux-gnu/4.8/include/stddef.h
..... /usr/include/x86_64-linux-gnu/bits/stdlib-float.h
.... /usr/include/c++/4.8/cmath
..... /usr/include/x86_64-linux-gnu/c++/4.8/bits/c++config.h
...... /usr/include/x86_64-linux-gnu/c++/4.8/bits/os_defines.h
...... /usr/include/x86_64-linux-gnu/c++/4.8/bits/cpu_defines.h
..... /usr/include/c++/4.8/bits/cpp_type_traits.h
..... /usr/include/c++/4.8/ext/type_traits.h
.... /usr/include/c++/4.8/cstdlib
.... /opt/cuda/bin/..//include/math_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/math_functions_dbl_ptx3.h
..... /opt/cuda/bin/..//include/math_functions_dbl_ptx3.hpp
.. /opt/cuda/bin/..//include/cuda_surface_types.h
... /opt/cuda/bin/..//include/builtin_types.h
.. /opt/cuda/bin/..//include/cuda_texture_types.h
... /opt/cuda/bin/..//include/builtin_types.h
.. /opt/cuda/bin/..//include/device_functions.h
... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/device_functions.hpp
.... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/device_atomic_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/device_atomic_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/device_double_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/device_double_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/sm_20_atomic_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/sm_20_atomic_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/sm_32_atomic_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/sm_32_atomic_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/sm_35_atomic_functions.h
... /opt/cuda/bin/..//include/sm_20_intrinsics.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/sm_20_intrinsics.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/sm_30_intrinsics.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/sm_30_intrinsics.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/sm_32_intrinsics.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/sm_32_intrinsics.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/sm_35_intrinsics.h
... /opt/cuda/bin/..//include/surface_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/surface_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/texture_fetch_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/texture_fetch_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/texture_indirect_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/texture_indirect_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
... /opt/cuda/bin/..//include/surface_indirect_functions.h
.... /opt/cuda/bin/..//include/builtin_types.h
.... /opt/cuda/bin/..//include/surface_indirect_functions.hpp
..... /opt/cuda/bin/..//include/builtin_types.h
.. /opt/cuda/bin/..//include/device_launch_parameters.h
看起来 cuda_runtime.h
包含所有其他与 CUDA 相关的 header 。
关于c++ - .cu 源文件中默认包含哪些 header ?,我们在Stack Overflow上找到一个类似的问题: https://stackoverflow.com/questions/29706730/
我有一个类似于以下的结构。 class A { string title; List bItem; } class B { int pric
本地流 和 远程流 两者都是“媒体流列表 ”。 本地流 包含“本地媒体流 ” 对象 但是,远程流 包含“媒体流 ” 对象 为什么差别这么大? 当我使用“本地流 “- 这个对我有用: localVide
我正在尝试将 8 列虚拟变量转换为 8 级排名的一列。 我试图用这个公式来做到这一点: =IF(OR(A1="1");"1";IF(OR(B1="1");"2";IF(OR(C1="1");"3";I
我正在使用面向对象编程在 Python 中创建一个有点复杂的棋盘游戏的实现。 我的问题是,许多这些对象应该能够与其他对象交互,即使它们不包含在其中。 例如Game是一个对象,其中包含PointTrac
有没有办法获取与 contains 语句匹配的最深元素? 基本上,如果我有嵌套的 div,我想要最后一个元素而不是父元素: Needle $("div:contains('Needle')")
出于某种原因,我无法在 Google 上找到答案!但是使用 SQL contains 函数我怎么能告诉它从字符串的开头开始,即我正在寻找等同于的全文 喜欢 'some_term%'。 我知道我可以使用
我正在尝试创建一个正则表达式来匹配具有 3 个或更多元音的字符串。 我试过这个: [aeiou]{3,} 但它仅在元音按顺序排列时才有效。有什么建议吗? 例如: 塞缪尔 -> 有效 琼 -> 无效 S
嘿所以我遇到了这样的情况,我从数据库中拉回一个客户,并通过包含的方式包含所有案例研究 return (from c in db.Clients.Include("CaseStudies")
如果关键字是子字符串,我无法弄清楚为什么这个函数不返回结果。 const string = 'cake'; const substring = 'cak'; console.log(string.in
我正在尝试将包含特定文本字符串的任何元素更改为红色。在我的示例中,我可以将子元素变为蓝色,但是我编写“替换我”行的方式有些不正确;红色不会发生变化。我注意到“contains”方法通常写为 :cont
我想问一下我是否可以要求/包含一个语法错误的文件,如果不能,则require/include返回一个值,这样我就知道所需/包含的文件存在语法错误并且不能被要求/包含? file.php语法错误 inc
我想为所有包含youtube链接的链接添加一个rel。 这就是我正在使用的东西-但它没有用。有任何想法吗? $('a [href:contains(“youtube.com”)]')。attr('re
我正在尝试在 Elasticsearch 中查询。除搜索中出现“/”外,此功能均正常运行。查询如下所示 GET styling_rules/product_line_filters/_search {
我正在开发名为eBookRepository的ASP.NET MVC应用程序,其中包含在线图书。 电子书具有自己的标题,作者等。因此,现在我正在尝试实现搜索机制。我必须使用Elasticsearch作
我已阅读Firebase Documentation并且不明白什么是 .contains()。 以下是文档中 Firebase 数据库的示例规则: { "rules": { "rooms"
我的问题是我可以给出条件[ 'BookTitleMaster.id' => $xtitid, ] 如下所示 $bbookinfs = $this->BookStockin->BookIssue->fi
我需要能够使用 | 检查模式在他们中。例如,对于像“dtest|test”这样的字符串,像 d*|*t 这样的表达式应该返回 true。 我不是正则表达式英雄,所以我只是尝试了一些事情,例如: Reg
我想创建一个正则表达式来不匹配某些单词... 我的字符:var test = "é123rr;and;ià456;or;456543" 我的正则表达式:test.match(\((?!and)(?!o
我在 XSLT 中有一个名为 variable_name 的变量,如果相关产品具有名称为 A 或 B 或两者均为 A & 的属性,我将尝试将其设置为 1 B.
您好,我想让接待员和经理能够查看工作类型和费率并随后进行更新。但是技术人员只能查看不能更新。该图是否有效? 我读到扩展用例是由发起基本用例的参与者发起的。我应该如何区分技术人员只能启动基本案例而不能启
我是一名优秀的程序员,十分优秀!