- 新着としてマーク
- ブックマーク
- 購読
- ミュート
- RSS フィードを購読する
- ハイライト
- 印刷
- 不適切なコンテンツを報告
OS: Ubuntu 24.10
Linux Kernel: tested on both 6.11.0-generic and 6.12.0-rc3
CPU: intel ultra7 258v
GPU: 64a0
How should I run OpenCL applications using this GPU on Linux?
Compile models using OpenVINO:
RuntimeError: Exception from src/inference/src/cpp/core.cpp:104:
Exception from src/inference/src/dev/plugin.cpp:53:
Check 'false' failed at src/plugins/intel_gpu/src/plugin/program_builder.cpp:185:
[GPU] ProgramBuilder build failed!
Exception from src/plugins/intel_gpu/src/runtime/ocl/ocl_stream.cpp:384:
[GPU] clFinish, error code: -5
Run `dmesg | grep xe`:
[ 4.248318] xe 0000:00:02.0: vgaarb: deactivate vga console
[ 4.248451] xe 0000:00:02.0: [drm] Found LUNARLAKE (device ID 64a0) display version 20.00 stepping B0
[ 4.250010] xe 0000:00:02.0: [drm] Using GuC firmware from xe/lnl_guc_70.bin version 70.29.2
[ 4.264005] xe 0000:00:02.0: [drm] Using GuC firmware from xe/lnl_guc_70.bin version 70.29.2
[ 4.268181] xe 0000:00:02.0: [drm] Using HuC firmware from xe/lnl_huc.bin version 9.4.13
[ 4.270292] xe 0000:00:02.0: [drm] Using GSC firmware from xe/lnl_gsc_1.bin version 104.0.0.1161
[ 4.305868] xe 0000:00:02.0: vgaarb: VGA decodes changed: olddecodes=io+mem,decodes=none:owns=io+mem
[ 4.320340] xe 0000:00:02.0: [drm] Finished loading DMC firmware i915/xe2lpd_dmc.bin (v2.21)
[ 4.573311] xe 0000:00:02.0: [drm] ccs1 fused off
[ 4.573318] xe 0000:00:02.0: [drm] ccs2 fused off
[ 4.573319] xe 0000:00:02.0: [drm] ccs3 fused off
[ 4.585725] xe 0000:00:02.0: [drm] vcs1 fused off
[ 4.585727] xe 0000:00:02.0: [drm] vcs2 fused off
[ 4.585728] xe 0000:00:02.0: [drm] vcs3 fused off
[ 4.585728] xe 0000:00:02.0: [drm] vcs4 fused off
[ 4.585729] xe 0000:00:02.0: [drm] vcs5 fused off
[ 4.585729] xe 0000:00:02.0: [drm] vcs6 fused off
[ 4.585730] xe 0000:00:02.0: [drm] vcs7 fused off
[ 4.585731] xe 0000:00:02.0: [drm] vecs1 fused off
[ 4.585731] xe 0000:00:02.0: [drm] vecs2 fused off
[ 4.585732] xe 0000:00:02.0: [drm] vecs3 fused off
[ 4.657505] [drm] Initialized xe 1.1.0 for 0000:00:02.0 on minor 0
[ 4.755404] xe 0000:00:02.0: [drm] GT1: found GSC cv104.1.0
[ 6.535955] xe 0000:00:02.0: [drm] Reducing the compressed framebuffer size. This may lead to less power savings than a non-reduced-size. Try to increase stolen memory size if available in BIOS.
[ 6.638552] xe 0000:00:02.0: [drm] fb0: xedrmfb frame buffer device
[ 8.467021] Modules linked in: cpuid snd_soc_cs35l56_sdw snd_soc_cs35l56 snd_soc_wm_adsp cs42l43_sdw snd_soc_cs35l56_shared regmap_sdw snd_soc_cs_amp_lib snd_hda_codec_hdmi cs42l43 cs_dsp snd_soc_dmic binfmt_misc nls_iso8859_1 snd_sof_pci_intel_lnl snd_sof_pci_intel_mtl intel_uncore_frequency snd_sof_intel_hda_generic intel_uncore_frequency_common soundwire_intel x86_pkg_temp_thermal intel_powerclamp soundwire_cadence iwlmvm snd_sof_intel_hda_common snd_soc_hdac_hda coretemp snd_sof_intel_hda_mlink xe snd_sof_intel_hda snd_sof_pci snd_sof_xtensa_dsp snd_sof snd_sof_utils snd_hda_ext_core snd_soc_acpi_intel_match soundwire_generic_allocation snd_soc_acpi soundwire_bus kvm_intel mac80211 snd_soc_core snd_compress ac97_bus snd_pcm_dmaengine snd_hda_intel snd_intel_dspcfg snd_intel_sdw_acpi snd_hda_codec kvm cmdlinepart spi_nor snd_hda_core mtd snd_hwdep mei_gsc_proxy snd_pcm intel_rapl_msr libarc4 crct10dif_pclmul snd_seq_midi polyval_clmulni snd_seq_midi_event polyval_generic uvcvideo snd_rawmidi iwlwifi
[ 8.709305] Modules linked in: cpuid snd_soc_cs35l56_sdw snd_soc_cs35l56 snd_soc_wm_adsp cs42l43_sdw snd_soc_cs35l56_shared regmap_sdw snd_soc_cs_amp_lib snd_hda_codec_hdmi cs42l43 cs_dsp snd_soc_dmic binfmt_misc nls_iso8859_1 snd_sof_pci_intel_lnl snd_sof_pci_intel_mtl intel_uncore_frequency snd_sof_intel_hda_generic intel_uncore_frequency_common soundwire_intel x86_pkg_temp_thermal intel_powerclamp soundwire_cadence iwlmvm snd_sof_intel_hda_common snd_soc_hdac_hda coretemp snd_sof_intel_hda_mlink xe snd_sof_intel_hda snd_sof_pci snd_sof_xtensa_dsp snd_sof snd_sof_utils snd_hda_ext_core snd_soc_acpi_intel_match soundwire_generic_allocation snd_soc_acpi soundwire_bus kvm_intel mac80211 snd_soc_core snd_compress ac97_bus snd_pcm_dmaengine snd_hda_intel snd_intel_dspcfg snd_intel_sdw_acpi snd_hda_codec kvm cmdlinepart spi_nor snd_hda_core mtd snd_hwdep mei_gsc_proxy snd_pcm intel_rapl_msr libarc4 crct10dif_pclmul snd_seq_midi polyval_clmulni snd_seq_midi_event polyval_generic uvcvideo snd_rawmidi iwlwifi
[ 8.927816] Modules linked in: cpuid snd_soc_cs35l56_sdw snd_soc_cs35l56 snd_soc_wm_adsp cs42l43_sdw snd_soc_cs35l56_shared regmap_sdw snd_soc_cs_amp_lib snd_hda_codec_hdmi cs42l43 cs_dsp snd_soc_dmic binfmt_misc nls_iso8859_1 snd_sof_pci_intel_lnl snd_sof_pci_intel_mtl intel_uncore_frequency snd_sof_intel_hda_generic intel_uncore_frequency_common soundwire_intel x86_pkg_temp_thermal intel_powerclamp soundwire_cadence iwlmvm snd_sof_intel_hda_common snd_soc_hdac_hda coretemp snd_sof_intel_hda_mlink xe snd_sof_intel_hda snd_sof_pci snd_sof_xtensa_dsp snd_sof snd_sof_utils snd_hda_ext_core snd_soc_acpi_intel_match soundwire_generic_allocation snd_soc_acpi soundwire_bus kvm_intel mac80211 snd_soc_core snd_compress ac97_bus snd_pcm_dmaengine snd_hda_intel snd_intel_dspcfg snd_intel_sdw_acpi snd_hda_codec kvm cmdlinepart spi_nor snd_hda_core mtd snd_hwdep mei_gsc_proxy snd_pcm intel_rapl_msr libarc4 crct10dif_pclmul snd_seq_midi polyval_clmulni snd_seq_midi_event polyval_generic uvcvideo snd_rawmidi iwlwifi
[ 65.105259] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 65.105296] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967170, lrc_seqno=4294967170, guc_id=17, flags=0x4 in no process [-1]
[ 104.135116] xe 0000:00:02.0: [drm] GT0: Engine reset: engine_class=ccs, logical_mask: 0x1, guc_id=18
[ 110.099636] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 110.099654] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967170, lrc_seqno=4294967170, guc_id=17, flags=0x4 in no process [-1]
[ 110.732908] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 110.732937] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967170, lrc_seqno=4294967170, guc_id=17, flags=0x4 in no process [-1]
[ 111.099513] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 111.329237] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 111.329259] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967170, lrc_seqno=4294967170, guc_id=17, flags=0x4 in no process [-1]
[ 111.529902] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 1319.353766] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 1319.353790] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967170, lrc_seqno=4294967170, guc_id=17, flags=0x4 in no process [-1]
[ 1509.202668] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967169, lrc_seqno=4294967169, guc_id=17, flags=0x4 in no process [-1]
[ 1509.202691] xe 0000:00:02.0: [drm] GT0: Timedout job: seqno=4294967170, lrc_seqno=4294967170, guc_id=17, flags=0x4 in no process [-1]
[ 1544.114323] xe 0000:00:02.0: [drm] GT0: Engine reset: engine_class=ccs, logical_mask: 0x1, guc_id=18
Run clinfo:
```
Number of platforms 1
Platform Name Intel(R) OpenCL Graphics
Platform Vendor Intel(R) Corporation
Platform Version OpenCL 3.0
Platform Profile FULL_PROFILE
Platform Extensions cl_khr_byte_addressable_store cl_khr_device_uuid cl_khr_fp16 cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_icd cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_intel_command_queue_families cl_intel_subgroups cl_intel_required_subgroup_size cl_intel_subgroups_short cl_khr_spir cl_intel_accelerator cl_intel_driver_diagnostics cl_khr_priority_hints cl_khr_throttle_hints cl_khr_create_command_queue cl_intel_subgroups_char cl_intel_subgroups_long cl_khr_il_program cl_intel_mem_force_host_memory cl_khr_subgroup_extended_types cl_khr_subgroup_non_uniform_vote cl_khr_subgroup_ballot cl_khr_subgroup_non_uniform_arithmetic cl_khr_subgroup_shuffle cl_khr_subgroup_shuffle_relative cl_khr_subgroup_clustered_reduce cl_intel_device_attribute_query cl_khr_extended_bit_ops cl_khr_suggested_local_work_size cl_intel_split_work_group_barrier cl_khr_fp64 cl_khr_subgroups cl_intel_spirv_subgroups cl_khr_spirv_linkonce_odr cl_khr_spirv_no_integer_wrap_decoration cl_intel_unified_shared_memory cl_khr_mipmap_image cl_khr_mipmap_image_writes cl_ext_float_atomics cl_khr_external_memory cl_intel_planar_yuv cl_intel_packed_yuv cl_khr_int64_base_atomics cl_khr_int64_extended_atomics cl_khr_image2d_from_buffer cl_khr_depth_images cl_khr_3d_image_writes cl_intel_bfloat16_conversions cl_intel_create_buffer_with_properties cl_intel_subgroup_local_block_io cl_intel_subgroup_matrix_multiply_accumulate cl_intel_subgroup_matrix_multiply_accumulate_tf32 cl_khr_subgroup_named_barrier cl_intel_subgroup_extended_block_read cl_intel_subgroup_2d_block_io cl_intel_subgroup_buffer_prefetch cl_khr_integer_dot_product cl_khr_gl_sharing cl_khr_gl_depth_images cl_khr_gl_event cl_khr_gl_msaa_sharing cl_intel_sharing_format_query cl_khr_pci_bus_info
Platform Extensions with Version cl_khr_byte_addressable_store 0x400000 (1.0.0)
cl_khr_device_uuid 0x400000 (1.0.0)
cl_khr_fp16 0x400000 (1.0.0)
cl_khr_global_int32_base_atomics 0x400000 (1.0.0)
cl_khr_global_int32_extended_atomics 0x400000 (1.0.0)
cl_khr_icd 0x400000 (1.0.0)
cl_khr_local_int32_base_atomics 0x400000 (1.0.0)
cl_khr_local_int32_extended_atomics 0x400000 (1.0.0)
cl_intel_command_queue_families 0x400000 (1.0.0)
cl_intel_subgroups 0x400000 (1.0.0)
cl_intel_required_subgroup_size 0x400000 (1.0.0)
cl_intel_subgroups_short 0x400000 (1.0.0)
cl_khr_spir 0x400000 (1.0.0)
cl_intel_accelerator 0x400000 (1.0.0)
cl_intel_driver_diagnostics 0x400000 (1.0.0)
cl_khr_priority_hints 0x400000 (1.0.0)
cl_khr_throttle_hints 0x400000 (1.0.0)
cl_khr_create_command_queue 0x400000 (1.0.0)
cl_intel_subgroups_char 0x400000 (1.0.0)
cl_intel_subgroups_long 0x400000 (1.0.0)
cl_khr_il_program 0x400000 (1.0.0)
cl_intel_mem_force_host_memory 0x400000 (1.0.0)
cl_khr_subgroup_extended_types 0x400000 (1.0.0)
cl_khr_subgroup_non_uniform_vote 0x400000 (1.0.0)
cl_khr_subgroup_ballot 0x400000 (1.0.0)
cl_khr_subgroup_non_uniform_arithmetic 0x400000 (1.0.0)
cl_khr_subgroup_shuffle 0x400000 (1.0.0)
cl_khr_subgroup_shuffle_relative 0x400000 (1.0.0)
cl_khr_subgroup_clustered_reduce 0x400000 (1.0.0)
cl_intel_device_attribute_query 0x400000 (1.0.0)
cl_khr_extended_bit_ops 0x400000 (1.0.0)
cl_khr_suggested_local_work_size 0x400000 (1.0.0)
cl_intel_split_work_group_barrier 0x400000 (1.0.0)
cl_khr_fp64 0x400000 (1.0.0)
cl_khr_subgroups 0x400000 (1.0.0)
cl_intel_spirv_subgroups 0x400000 (1.0.0)
cl_khr_spirv_linkonce_odr 0x400000 (1.0.0)
cl_khr_spirv_no_integer_wrap_decoration 0x400000 (1.0.0)
cl_intel_unified_shared_memory 0x400000 (1.0.0)
cl_khr_mipmap_image 0x400000 (1.0.0)
cl_khr_mipmap_image_writes 0x400000 (1.0.0)
cl_ext_float_atomics 0x400000 (1.0.0)
cl_khr_external_memory 0x9001 (0.9.1)
cl_intel_planar_yuv 0x400000 (1.0.0)
cl_intel_packed_yuv 0x400000 (1.0.0)
cl_khr_int64_base_atomics 0x400000 (1.0.0)
cl_khr_int64_extended_atomics 0x400000 (1.0.0)
cl_khr_image2d_from_buffer 0x400000 (1.0.0)
cl_khr_depth_images 0x400000 (1.0.0)
cl_khr_3d_image_writes 0x400000 (1.0.0)
cl_intel_bfloat16_conversions 0x400000 (1.0.0)
cl_intel_create_buffer_with_properties 0x400000 (1.0.0)
cl_intel_subgroup_local_block_io 0x400000 (1.0.0)
cl_intel_subgroup_matrix_multiply_accumulate 0x400000 (1.0.0)
cl_intel_subgroup_matrix_multiply_accumulate_tf32 0x400000 (1.0.0)
cl_khr_subgroup_named_barrier 0x400000 (1.0.0)
cl_intel_subgroup_extended_block_read 0x400000 (1.0.0)
cl_intel_subgroup_2d_block_io 0x400000 (1.0.0)
cl_intel_subgroup_buffer_prefetch 0x400000 (1.0.0)
cl_khr_integer_dot_product 0x800000 (2.0.0)
cl_khr_gl_sharing 0x400000 (1.0.0)
cl_khr_gl_depth_images 0x400000 (1.0.0)
cl_khr_gl_event 0x400000 (1.0.0)
cl_khr_gl_msaa_sharing 0x400000 (1.0.0)
cl_intel_sharing_format_query 0x400000 (1.0.0)
cl_khr_pci_bus_info 0x400000 (1.0.0)
Platform Numeric Version 0xc00000 (3.0.0)
Platform Extensions function suffix INTEL
Platform Host timer resolution 1ns
Platform External memory handle types DMA buffer
Platform Name Intel(R) OpenCL Graphics
Number of devices 1
Device Name Intel(R) Graphics [0x64a0]
Device Vendor Intel(R) Corporation
Device Vendor ID 0x8086
Device Version OpenCL 3.0 NEO
Device UUID 8680a064-0400-0000-0002-000000000000
Driver UUID 32342e33-352e-3330-3837-320000000000
Valid Device LUID No
Device LUID d0b8-5aaffd7f0000
Device Node Mask 0
Device Numeric Version 0xc00000 (3.0.0)
Driver Version 24.35.30872
Device OpenCL C Version OpenCL C 1.2
Device OpenCL C all versions OpenCL C 0x400000 (1.0.0)
OpenCL C 0x401000 (1.1.0)
OpenCL C 0x402000 (1.2.0)
OpenCL C 0xc00000 (3.0.0)
Device OpenCL C features __opencl_c_int64 0xc00000 (3.0.0)
__opencl_c_3d_image_writes 0xc00000 (3.0.0)
__opencl_c_images 0xc00000 (3.0.0)
__opencl_c_read_write_images 0xc00000 (3.0.0)
__opencl_c_atomic_order_acq_rel 0xc00000 (3.0.0)
__opencl_c_atomic_order_seq_cst 0xc00000 (3.0.0)
__opencl_c_atomic_scope_all_devices 0xc00000 (3.0.0)
__opencl_c_atomic_scope_device 0xc00000 (3.0.0)
__opencl_c_generic_address_space 0xc00000 (3.0.0)
__opencl_c_program_scope_global_variables 0xc00000 (3.0.0)
__opencl_c_work_group_collective_functions 0xc00000 (3.0.0)
__opencl_c_subgroups 0xc00000 (3.0.0)
__opencl_c_ext_fp32_global_atomic_add 0xc00000 (3.0.0)
__opencl_c_ext_fp32_local_atomic_add 0xc00000 (3.0.0)
__opencl_c_ext_fp32_global_atomic_min_max 0xc00000 (3.0.0)
__opencl_c_ext_fp32_local_atomic_min_max 0xc00000 (3.0.0)
__opencl_c_ext_fp16_global_atomic_load_store 0xc00000 (3.0.0)
__opencl_c_ext_fp16_local_atomic_load_store 0xc00000 (3.0.0)
__opencl_c_ext_fp16_global_atomic_min_max 0xc00000 (3.0.0)
__opencl_c_ext_fp16_local_atomic_min_max 0xc00000 (3.0.0)
__opencl_c_fp64 0xc00000 (3.0.0)
__opencl_c_ext_fp64_global_atomic_add 0xc00000 (3.0.0)
__opencl_c_ext_fp64_local_atomic_add 0xc00000 (3.0.0)
__opencl_c_ext_fp64_global_atomic_min_max 0xc00000 (3.0.0)
__opencl_c_ext_fp64_local_atomic_min_max 0xc00000 (3.0.0)
__opencl_c_integer_dot_product_input_4x8bit 0xc00000 (3.0.0)
__opencl_c_integer_dot_product_input_4x8bit_packed 0xc00000 (3.0.0)
Latest conformance test passed v2024-02-27-00
Device Type GPU
Device PCI bus info (KHR) PCI-E, 0000:00:02.0
Device Profile FULL_PROFILE
Device Available Yes
Compiler Available Yes
Linker Available Yes
Max compute units 64
Max clock frequency 1950MHz
Device IP (Intel) 0x5010004 (20.16.4)
Device ID (Intel) 25760
Slices (Intel) 1
Sub-slices per slice (Intel) 8
EUs per sub-slice (Intel) 8
Threads per EU (Intel) 8
Feature capabilities (Intel) DP4A, DPAS
Device Partition (core)
Max number of sub-devices 0
Supported partition types None
Supported affinity domains (n/a)
Max work item dimensions 3
Max work item sizes 1024x1024x1024
Max work group size 1024
Preferred work group size multiple (device) 32
Preferred work group size multiple (kernel) 32
Max sub-groups per work group 64
Max named sub-group barriers <printDeviceInfo:71: get CL_DEVICE_MAX_NAMED_BARRIER_COUNT_KHR : error -30>
Sub-group sizes (Intel) 16, 32
Preferred / native vector sizes
char 16 / 16
short 8 / 8
int 4 / 4
long 1 / 1
half 8 / 8 (cl_khr_fp16)
float 1 / 1
double 1 / 1 (cl_khr_fp64)
Half-precision Floating-point support (cl_khr_fp16)
Denormals Yes
Infinity and NANs Yes
Round to nearest Yes
Round to zero Yes
Round to infinity Yes
IEEE754-2008 fused multiply-add Yes
Support is emulated in software No
Single-precision Floating-point support (core)
Denormals Yes
Infinity and NANs Yes
Round to nearest Yes
Round to zero Yes
Round to infinity Yes
IEEE754-2008 fused multiply-add Yes
Support is emulated in software No
Correctly-rounded divide and sqrt operations Yes
Double-precision Floating-point support (cl_khr_fp64)
Denormals Yes
Infinity and NANs Yes
Round to nearest Yes
Round to zero Yes
Round to infinity Yes
IEEE754-2008 fused multiply-add Yes
Support is emulated in software No
Address bits 64, Little-Endian
External memory handle types DMA buffer
Global memory size 14832754688 (13.81GiB)
Error Correction support No
Max memory allocation 14832754688 (13.81GiB)
Unified memory for Host and Device Yes
Shared Virtual Memory (SVM) capabilities (core)
Coarse-grained buffer sharing Yes
Fine-grained buffer sharing No
Fine-grained system sharing No
Atomics No
Unified Shared Memory (USM) (cl_intel_unified_shared_memory)
Host USM capabilities (Intel) USM access, USM atomic access
Device USM capabilities (Intel) USM access, USM atomic access
Single-Device USM caps (Intel) USM access, USM atomic access
Cross-Device USM caps (Intel) USM access, USM atomic access
Shared System USM caps (Intel) (n/a)
Minimum alignment for any data type 128 bytes
Alignment of base address 1024 bits (128 bytes)
Preferred alignment for atomics
SVM 64 bytes
Global 64 bytes
Local 64 bytes
Atomic memory capabilities relaxed, acquire/release, sequentially-consistent, work-group scope, device scope, all-devices scope
Atomic fence capabilities relaxed, acquire/release, sequentially-consistent, work-item scope, work-group scope, device scope, all-devices scope
Max size for global variable 65536 (64KiB)
Preferred total size of global vars 14832754688 (13.81GiB)
Global Memory cache type Read/Write
Global Memory cache size 8388608 (8MiB)
Global Memory cache line size 256 bytes
Image support Yes
Max number of samplers per kernel 16
Max size for 1D images from buffer 927047168 pixels
Max 1D or 2D image array size 2048 images
Base address alignment for 2D image buffers 4 bytes
Pitch alignment for 2D image buffers 4 pixels
Max 2D image size 16384x16384 pixels
Max planar YUV image size 16384x16128 pixels
Max 3D image size 16384x16384x2048 pixels
Max number of read image args 128
Max number of write image args 128
Max number of read/write image args 128
Pipe support No
Max number of pipe args 0
Max active pipe reservations 0
Max pipe packet size 0
Local memory type Local
Local memory size 131072 (128KiB)
Max number of constant args 8
Max constant buffer size 14832754688 (13.81GiB)
Generic address space support Yes
Max size of kernel argument 2048 (2KiB)
Queue properties (on host)
Out-of-order execution Yes
Profiling Yes
Device enqueue capabilities (n/a)
Queue properties (on device)
Out-of-order execution No
Profiling No
Preferred size 0
Max size 0
Max queues on device 0
Max events on device 0
Device queue families ccs (1)
Queue properties Out-of-order execution, Profiling
Capabilities create single-queue events, create cross-queue events
cccs (1)
Queue properties Out-of-order execution, Profiling
Capabilities create single-queue events, create cross-queue events
bcs (1)
Queue properties Out-of-order execution, Profiling
Capabilities create single-queue events, create cross-queue events
Prefer user sync for interop Yes
Profiling timer resolution 52ns
Execution capabilities
Run OpenCL kernels Yes
Run native kernels No
Non-uniform work-groups Yes
Work-group collective functions Yes
Sub-group independent forward progress Yes
IL version SPIR-V_1.3 SPIR-V_1.2 SPIR-V_1.1 SPIR-V_1.0
ILs with version SPIR-V 0x403000 (1.3.0)
SPIR-V 0x402000 (1.2.0)
SPIR-V 0x401000 (1.1.0)
SPIR-V 0x400000 (1.0.0)
SPIR versions 1.2
printf() buffer size 4194304 (4MiB)
Built-in kernels (n/a)
Built-in kernels with version (n/a)
Device Extensions cl_khr_byte_addressable_store cl_khr_device_uuid cl_khr_fp16 cl_khr_global_int32_base_atomics cl_khr_global_int32_extended_atomics cl_khr_icd cl_khr_local_int32_base_atomics cl_khr_local_int32_extended_atomics cl_intel_command_queue_families cl_intel_subgroups cl_intel_required_subgroup_size cl_intel_subgroups_short cl_khr_spir cl_intel_accelerator cl_intel_driver_diagnostics cl_khr_priority_hints cl_khr_throttle_hints cl_khr_create_command_queue cl_intel_subgroups_char cl_intel_subgroups_long cl_khr_il_program cl_intel_mem_force_host_memory cl_khr_subgroup_extended_types cl_khr_subgroup_non_uniform_vote cl_khr_subgroup_ballot cl_khr_subgroup_non_uniform_arithmetic cl_khr_subgroup_shuffle cl_khr_subgroup_shuffle_relative cl_khr_subgroup_clustered_reduce cl_intel_device_attribute_query cl_khr_extended_bit_ops cl_khr_suggested_local_work_size cl_intel_split_work_group_barrier cl_khr_fp64 cl_khr_subgroups cl_intel_spirv_subgroups cl_khr_spirv_linkonce_odr cl_khr_spirv_no_integer_wrap_decoration cl_intel_unified_shared_memory cl_khr_mipmap_image cl_khr_mipmap_image_writes cl_ext_float_atomics cl_khr_external_memory cl_intel_planar_yuv cl_intel_packed_yuv cl_khr_int64_base_atomics cl_khr_int64_extended_atomics cl_khr_image2d_from_buffer cl_khr_depth_images cl_khr_3d_image_writes cl_intel_bfloat16_conversions cl_intel_create_buffer_with_properties cl_intel_subgroup_local_block_io cl_intel_subgroup_matrix_multiply_accumulate cl_intel_subgroup_matrix_multiply_accumulate_tf32 cl_khr_subgroup_named_barrier cl_intel_subgroup_extended_block_read cl_intel_subgroup_2d_block_io cl_intel_subgroup_buffer_prefetch cl_khr_integer_dot_product cl_khr_gl_sharing cl_khr_gl_depth_images cl_khr_gl_event cl_khr_gl_msaa_sharing cl_intel_sharing_format_query cl_khr_pci_bus_info
Device Extensions with Version cl_khr_byte_addressable_store 0x400000 (1.0.0)
cl_khr_device_uuid 0x400000 (1.0.0)
cl_khr_fp16 0x400000 (1.0.0)
cl_khr_global_int32_base_atomics 0x400000 (1.0.0)
cl_khr_global_int32_extended_atomics 0x400000 (1.0.0)
cl_khr_icd 0x400000 (1.0.0)
cl_khr_local_int32_base_atomics 0x400000 (1.0.0)
cl_khr_local_int32_extended_atomics 0x400000 (1.0.0)
cl_intel_command_queue_families 0x400000 (1.0.0)
cl_intel_subgroups 0x400000 (1.0.0)
cl_intel_required_subgroup_size 0x400000 (1.0.0)
cl_intel_subgroups_short 0x400000 (1.0.0)
cl_khr_spir 0x400000 (1.0.0)
cl_intel_accelerator 0x400000 (1.0.0)
cl_intel_driver_diagnostics 0x400000 (1.0.0)
cl_khr_priority_hints 0x400000 (1.0.0)
cl_khr_throttle_hints 0x400000 (1.0.0)
cl_khr_create_command_queue 0x400000 (1.0.0)
cl_intel_subgroups_char 0x400000 (1.0.0)
cl_intel_subgroups_long 0x400000 (1.0.0)
cl_khr_il_program 0x400000 (1.0.0)
cl_intel_mem_force_host_memory 0x400000 (1.0.0)
cl_khr_subgroup_extended_types 0x400000 (1.0.0)
cl_khr_subgroup_non_uniform_vote 0x400000 (1.0.0)
cl_khr_subgroup_ballot 0x400000 (1.0.0)
cl_khr_subgroup_non_uniform_arithmetic 0x400000 (1.0.0)
cl_khr_subgroup_shuffle 0x400000 (1.0.0)
cl_khr_subgroup_shuffle_relative 0x400000 (1.0.0)
cl_khr_subgroup_clustered_reduce 0x400000 (1.0.0)
cl_intel_device_attribute_query 0x400000 (1.0.0)
cl_khr_extended_bit_ops 0x400000 (1.0.0)
cl_khr_suggested_local_work_size 0x400000 (1.0.0)
cl_intel_split_work_group_barrier 0x400000 (1.0.0)
cl_khr_fp64 0x400000 (1.0.0)
cl_khr_subgroups 0x400000 (1.0.0)
cl_intel_spirv_subgroups 0x400000 (1.0.0)
cl_khr_spirv_linkonce_odr 0x400000 (1.0.0)
cl_khr_spirv_no_integer_wrap_decoration 0x400000 (1.0.0)
cl_intel_unified_shared_memory 0x400000 (1.0.0)
cl_khr_mipmap_image 0x400000 (1.0.0)
cl_khr_mipmap_image_writes 0x400000 (1.0.0)
cl_ext_float_atomics 0x400000 (1.0.0)
cl_khr_external_memory 0x9001 (0.9.1)
cl_intel_planar_yuv 0x400000 (1.0.0)
cl_intel_packed_yuv 0x400000 (1.0.0)
cl_khr_int64_base_atomics 0x400000 (1.0.0)
cl_khr_int64_extended_atomics 0x400000 (1.0.0)
cl_khr_image2d_from_buffer 0x400000 (1.0.0)
cl_khr_depth_images 0x400000 (1.0.0)
cl_khr_3d_image_writes 0x400000 (1.0.0)
cl_intel_bfloat16_conversions 0x400000 (1.0.0)
cl_intel_create_buffer_with_properties 0x400000 (1.0.0)
cl_intel_subgroup_local_block_io 0x400000 (1.0.0)
cl_intel_subgroup_matrix_multiply_accumulate 0x400000 (1.0.0)
cl_intel_subgroup_matrix_multiply_accumulate_tf32 0x400000 (1.0.0)
cl_khr_subgroup_named_barrier 0x400000 (1.0.0)
cl_intel_subgroup_extended_block_read 0x400000 (1.0.0)
cl_intel_subgroup_2d_block_io 0x400000 (1.0.0)
cl_intel_subgroup_buffer_prefetch 0x400000 (1.0.0)
cl_khr_integer_dot_product 0x800000 (2.0.0)
cl_khr_gl_sharing 0x400000 (1.0.0)
cl_khr_gl_depth_images 0x400000 (1.0.0)
cl_khr_gl_event 0x400000 (1.0.0)
cl_khr_gl_msaa_sharing 0x400000 (1.0.0)
cl_intel_sharing_format_query 0x400000 (1.0.0)
cl_khr_pci_bus_info 0x400000 (1.0.0)
NULL platform behavior
clGetPlatformInfo(NULL, CL_PLATFORM_NAME, ...) Intel(R) OpenCL Graphics
clGetDeviceIDs(NULL, CL_DEVICE_TYPE_ALL, ...) Success [INTEL]
clCreateContext(NULL, ...) [default] Success [INTEL]
clCreateContextFromType(NULL, CL_DEVICE_TYPE_DEFAULT) Success (1)
Platform Name Intel(R) OpenCL Graphics
Device Name Intel(R) Graphics [0x64a0]
clCreateContextFromType(NULL, CL_DEVICE_TYPE_CPU) No devices found in platform
clCreateContextFromType(NULL, CL_DEVICE_TYPE_GPU) Success (1)
Platform Name Intel(R) OpenCL Graphics
Device Name Intel(R) Graphics [0x64a0]
clCreateContextFromType(NULL, CL_DEVICE_TYPE_ACCELERATOR) No devices found in platform
clCreateContextFromType(NULL, CL_DEVICE_TYPE_CUSTOM) No devices found in platform
clCreateContextFromType(NULL, CL_DEVICE_TYPE_ALL) Success (1)
Platform Name Intel(R) OpenCL Graphics
Device Name Intel(R) Graphics [0x64a0]
ICD loader properties
ICD loader Name OpenCL ICD Loader
ICD loader Vendor OCL Icd free software
ICD loader Version 2.3.2
ICD loader Profile OpenCL 3.0
```
コピーされたリンク
0 返答(返信)
