Changelog for mesa-tools-native: 26.1.6 -> 26.2.0 Source: docs/relnotes/25.3.0.rst, docs/relnotes/26.0.5.rst, docs/relnotes/26.0.6.rst, docs/relnotes/26.0.7.rst, docs/relnotes/26.0.8.rst, docs/relnotes/26.1.5.rst, docs/relnotes/26.1.6.rst, docs/relnotes/26.2.0.rst Mesa 25.3.0 Release Notes / 2025-11-14 ====================================== Mesa 25.3.0 is a new development release. People who are concerned with stability and reliability should stick with a previous release or wait for Mesa 25.3.1. Mesa 25.3.0 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 25.3.0 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: SHA256: 0fd54fea7dbbddb154df05ac752b18621f26d97e27863db3be951417c6abe8ae mesa-25.3.0.tar.xz SHA512: 46df9e5e27f9a36cf893a68ad4a465fcc6efe1bcb46ad8d4b015699ad1a11e582b8d41f4157326556af603fe454b2ff34ecc17a0c742b5fd9ce5f0097106fec5 mesa-25.3.0.tar.xz Note for distros: ----------------- Starting with Mesa 25.3.0, we no longer support the Vulkan API on Linux kernel versions older than 6.0 due to the deprecation of implicit synchronization in favor of the dma-buf sync file import/export ioctls. More information can be found in MR `!36783 `__. New features ------------ - EGL_EXT_create_context_robustness support on Panfrost V10+ - GL_ARB_robust_buffer_access_behavior, GL_KHR_robust_buffer_access_behavior and GL_KHR_robustness support on Panfrost - VK_EXT_mutable_descriptor_type on panvk/v9+ - GL_KHR_robustness on v3d - VK_ARM_shader_core_builtins on panvk - VK_KHR_shader_untyped_pointers on anv - cl_ext_immutable_memory_objects - VK_KHR_video_encode_intra_refresh on radv - VK_KHR_video_encode_quantization_map on radv - GL_ATI_meminfo and GL_NVX_gpu_memory_info on r300 - VK_KHR_shader_untyped_pointers on anv and RADV - VK_KHR_maintenance8 on NVK - VK_KHR_maintenance9 on NVK - cl_khr_semaphore on radeonsi and zink - cl_khr_external_semaphore on radeonsi and zink - cl_khr_external_semaphore_sync_fd on radeonsi and zink - GL_NV_shader_atomic_int64 on radeonsi and Panfrost V9+ - VK_KHR_maintenance7 on panvk/v10+ - VK_KHR_maintenance8 on panvk/v10+ - VK_KHR_maintenance9 on panvk - VK_AMD_buffer_marker on NVK - VK_EXT_ycbcr_2plane_444_formats on radv - Removed VDPAU frontend - GL_NV_representative_fragment_test on zink - VK_KHR_maintenance9 on HoneyKrisp - sparseBinding on panvk/v10+ - sparseResidencyBuffer on panvk/v10+ - Vulkan 1.2 on pvr - VK_KHR_create_renderpass2 on pvr - VK_KHR_dedicated_allocation on pvr - VK_KHR_depth_stencil_resolve on pvr - VK_KHR_descriptor_update_template on pvr - VK_KHR_imageless_framebuffer on pvr - VK_KHR_line_rasterization on pvr - VK_KHR_maintenance1 on pvr - VK_KHR_maintenance2 on pvr - VK_KHR_maintenance3 on pvr - VK_KHR_multiview on pvr - VK_KHR_robustness2 on pvr - VK_KHR_separate_depth_stencil_layouts on pvr - VK_KHR_shader_draw_parameters on pvr - VK_KHR_shader_float_controls on pvr - VK_KHR_shader_subgroup_extended_types on pvr - VK_KHR_spirv_1_4 on pvr - VK_KHR_shader_terminate_invocation on pvr - VK_KHR_swapchain_mutable_format on pvr - VK_KHR_vertex_attribute_divisor on pvr - VK_EXT_border_color_swizzle on pvr - VK_EXT_color_write_enable on pvr - VK_EXT_custom_border_color on pvr - VK_EXT_depth_clamp_zero_one on pvr - VK_EXT_depth_clip_enable on pvr - VK_EXT_extended_dynamic_state on pvr - VK_EXT_extended_dynamic_state2 on pvr - VK_EXT_extended_dynamic_state3 on pvr - VK_EXT_image_2d_view_of_3d on pvr - VK_EXT_line_rasterization on pvr - VK_EXT_physical_device_drm on pvr - VK_EXT_provoking_vertex on pvr - VK_EXT_robustness2 on pvr - VK_EXT_queue_family_foreign on pvr - VK_EXT_separate_stencil_usage on pvr - VK_EXT_shader_demote_to_helper_invocation on pvr - VK_EXT_vertex_attribute_divisor on pvr - imageCubeArray on pvr - independentBlend on pvr - sampleRateShading on pvr - logicOp on pvr - drawIndirectFirstInstance on pvr - alphaToOne on pvr - samplerAnisotropy on pvr - shaderStorageImageExtendedFormats on pvr - shaderStorageImageReadWithoutFormat on pvr - shaderStorageImageWriteWithoutFormat on pvr - shaderClipDistance on pvr - shaderCullDistance on pvr - VK_EXT_zero_initialize_device_memory on pvr - VK_KHR_sampler_mirror_clamp_to_edge on pvr - VK_KHR_shader_non_semantic_info on pvr - VK_KHR_shader_relaxed_extended_instruction on pvr - VK_EXT_shader_replicated_composites on pvr - VK_KHR_device_group_creation on pvr - VK_KHR_map_memory2 on pvr - VK_EXT_map_memory_placed on pvr - VK_KHR_device_group on pvr - VK_KHR_buffer_device_address on pvr - GL_EXT_mesh_shader on zink - VK_KHR_wayland_surface on pvr - VK_NVX_image_view_handle on NVK Bug fixes --------- - amdgpu: ring gfx_0.0.0 timeout, in vr when opening apps - zink/radv: new cts fails on rdna3 - Penumbra: Overture OpenGL game has graphical glitch for ice - mesa: regression caused by hash_table sizing - RustiCL: fence fd leak on CL-GL interop - Uniform variable not updated correctly with shared contexts - [radv] Borderlands 4 triggers a consistent GPU page fault on RDNA2 - radv: RE4 Separate Ways DLC hangs RDNA2 GPU - ACO: fix a hazard when the number of attributes loaded/consumed don't match with VS prologs - ACO: loading 64-bit attributes can override the fetch index in VS prologs - [RADV][bisected][regression] - Doom: The Dark Ages (3017860) - Square flickering artifacts around Hebeth - nvk, nak: Broken icons in ENDLESS Legend 2 on a RTX 4080 - LLVMPipe's \`VkPhysicalDeviceAccelerationStructurePropertiesKHR::maxPrimitiveCount` is lower than Vulkan requires. - asahi: DMABuf import of multi-plane YCbCr (NV12 from ISP) not renderer correctly - brw: Gfx9 sampler messages violate r127 rule - radv: No Man's Sky XESS page fault GPU reset - r600/sfn: Assertion \`cir.alu_vec.empty()` failed - radv: Hit assert when over maxFragmentDualSrcAttachments but vkCmdSetColorBlendEnableEXT is set to false - [ANV][PTL][DG2] Flickering textures in Assassin's Creed Valhalla benchmark - ADL, ANV: Wuthering Waves leads to gpu reset on Alder Lake iGPU - RADV: ANGLE deqp regression - [ANV][EXT_debug_utils] descriptor set object_name leak when not calling vkFreeDescriptorSets - nvk: CTS failures in sample_locations_ext.verify_interpolation.samples_1 - [regression] [bisected] RuneLite GPU Experimental - GPU crash - Missing definition of __builtin_ia32_clflush since "util/cache_ops: Add some cache flush helpers" - LLVM instruction selection compilation error - v3d: green screen when rpivid hevc decoder is used - [radv] Stuttering with latest mesa git (21 sept) on radv/6900 XT - BFN with UW sources gets munged by lower regioning - zink: chromium flickers in youtube when fullscreening videos - r600: Attribute stride updates may be skipped - [ANV][TGL]: test_buffer_feedback_instructions_sm51 on vkd3d-proton crashes - some video file are not shown in mpv when using vaapi hardware decoding on amd apu - [ANV][PTL] Indiana Jones and the Great Circle - GPU Hang - [ANV] [PTL] Hades 2 game freeze on start of gameplay - [anv][ptl] GPU hang in Dying Light dx12 - radv: Only look at statically used descriptors. - RADV: Consider always using the global bo list - anv: Age of Wonders 4 corruption on a Arc b580 - nvk: Incorrect rendering in Baldur's Gate 3 shadows starting with e6dae6ef5fc134f9ed5dd93b1a462084bc3aadfd - nvk commets cause problems with kepler - anv: Assert in brew when descriptor indexing with modulo - tu: VK_EXT_zero_initialize_device_memory - ResourceTracker.cpp:40:10: fatal error: perfetto/tracing.h: No such file or directory - A bunch of CTS tests are failing on Gfx12.0 trying to use the blitter with TILE_X - radv: meta pipeline cache appears to be broken - mesa:amd+compiler / aco_tests assembler.mubuf/gfx11 failure with llvm-21.1.2 - [ANV] Bunch of tests in dEQP-VK.pipeline.*.render_to_image.*3d.*2d_compatible failing on gen9/11 - elk: segfault in lower_txd_cb - bisected: Regression in EXT_shader_framebuffer_fetch_non_coherent test after !37527 - VK_QUERY_RESULT_WAIT_BIT does not work for VK_QUERY_TYPE_VIDEO_ENCODE_FEEDBACK_KHR - a618-traces often times out - bisected build failure in clc_helpers.ccp with llvm 22 - anv: GL mesh tests crash/fail on zink with shader object - 25.2.1 fails to build on risc-v with llvm 21 - RISC-V builds with llvmpipe against LLVM 21 fail due to API changes - Confidential issue #14013 - implicit-function-declaration error when compiling mesa 25.2.0 devel - vl_stubs.c:105:1: error: conflicting types for 'vl_mpg12_bs_decode' - [ANV][LNL] - FINAL FANTASY XVI (2515020) - Title crashes to Desktop immediately following the splash card. - Segfault in init_source at ../src/gallium/auxiliary/vl/vl_idct.c:597 when trying to play DVD on r600 - nvk: Failure in vkd3d-proton ibfe tests - nvk, nak: NAK panic in Call of the Wild: The Angler on RTX 4080 - Simple External Semaphore test hangs in vk_sync_wait - nir_builtin_builder.h:108:43: error: 'M_LOG2E' undeclared - regression: windows: msys2 - undeclared M_PI and M_LOG2E probably since !37289 21b8e7604ba51f90682adeff650fc866c71c57f2 - dEQP-VK.spirv_assembly.instruction.compute.float_controls.fp32.input_args.reflect_denorm_flush_to_zero regression on nvk - mesa-25.2.3/src/gallium/drivers/radeonsi/radeon_uvd.c:658: array index used before check ? - lp_test_arit.c:200:14: error: static declaration of ‘rsqrtf’ follows non-static declaration - build failure with glibc 2.42 - [bisected] 44aaf884254 regressing FSR vulkan cts tests on PTL - [bisected] f416a529 "egl: refine dma buf export to support multi plane" results in piglit crash - Crash on game Elite Dangerous at 0% planetary generation, on Tigerlake+ Iris Xe and Arc GPUs. - regression;bisected;amd: 0a266f0256025d271945adb3478fc2c1291d4c79 leads pgadmin4-qt to crashes - segfault with mesa >= 24.1.0 on nvidia - segfault through lavapipe - Confidential issue #13807 - [bisected] 25b97a mesa/st: mark internal texture map calls as UNSYNCHRONIZED breaks r600 - Gallium: Segfault while trying to compile a shader with differing UBO contents in fragment and vertex stage - With reproduction case - aco: generate wrong code when gl_DrawID is used by primitive indices in mesh shader - Regression since mesa 25.2.0: applications waiting for dGPU to start - ci: libX11 upgrade tracker - anv: Regression in dEQP-VK.graphicsfuzz.cov-nested-loops-set-struct-data-verify-in-function - brw: regression crash on dEQP-VK.graphicsfuzz.cov-dfdx-dfdy-after-nested-loops - a618-traces often times out - ci: crosvm dumping log spam from host gl when the job fails - panfrost: assertion fail in pan_image_get_wsi_row_pitch - virgl: guest memory leak with qemu + virtio-gpu-gl - [ANV][LNL] - Horizon Forbidden West™ Complete Edition (2420110) - Orbicular artifacts near heads of machines (wildlife). - iris: Assertion failures in piglit tests on all platforms - [radv] [Regression) Shadow of the Tomb Raider - flickering/missing textures - Minecraft 1.12.2 visual artifacts when running on zink/radv - [RADV][VEGA 64][bisected] Cyberpunk 2077 - Massive performance regression due to https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/37025/diffs?commit_id=d7f401c2bbadd192dbbcaaeede2805bad71f6193 - [PTL] hitting assert when starting Xorg - GZDoom 4.11/Raze 1.7 exhibit very high memory usage during shader compilation under OpenGL - anv: Assertion failure replaying q2rtx fossil - [ANV] [REGRESSION] PCSX2/Midnight Club 3 crashes with VK_ERROR_DEVICE_LOST on Mesa 25.0.7 - Hollow Knight Silksong segfaults with zink on radv - zink: crash in KHR-GL46.framebuffer_blit.scissor_blit - Request: RADV support for VK_EXT_ycbcr_2plane_444_formats.. - mesa: state parameters duplicated - ARB_vertex_program and ARB_fragment_program are broken - d3d10umd fails to create basic pixel shader, outputs "unknown TGSI opcode: RET" - turnip: FDM failures with forcebin - a7xx_state_location - [ANV] Assertion with VVL GPU-AV around robust UBO - radv: avoid advertising unsupported global queue priorities for the client - crocus: SIGSEGV crash at pbo compressed teximage - nvk: test_conditional_rendering fails on vkd3d-proton - Segfault in x11_xlib_display_is_thread_safe - [ANV][BMG] Witcher 3 ray tracing freeze on a Arc b580 - anv: optimize utrace overhead from bo memset - radv: watching vp9 encoded video with vulkan hwdecode result in artifacts - ci: zink-venus-lavapipe errors - [ANV][DG2][BMG] 3DMark Solar Bay Misrendering - lavapipe defaults to memfd when multiple export types are requested - anv: Simple vulkan compute shader causes Intel GPU hang due to excessive loop unrolling - INTEL_DEBUG=spill_fs regression - NIR validation failed after nir_lower_io in DOOM The Dark Ages - radv: NIR validation failed after nir_shrink_vec_array_vars in ../src/amd/vulkan/radv_shader.c:171 (bisected) - Wayland EGL missing pbuffer surface support - missing sparse synchronization in zink - ACO validation failed in DOOM: The Dark Ages - Undef operand to \`p_parallelcopy` - ACO validation failure in DOOM: The Dark Ages Shader - Dragon Age Veilguard / Ability Wheel Targeting Visual Distortion - [AMD] glTexturePageCommitmentEXT triggers an error if level is higher than 0 - deqp-egl multithread link flakes - Ratchet and Clank "[gfxhub] page fault" Mesa25.3 - [RADV] Support for VK_KHR_video_encode_quantization_map - System Crashes when starting VR on rx 9070 (xt) - [Build][32bit] Meson does not find libdisplay-info in 32-bit builds - freedreno,decode: Lua environment picks up the A6XX register offset instead of A7XX - Confidential issue #13351 - d3d10umd: Build regression on 25.2.0-rc1 - turnip: LRZ bug with TU_DEBUG=gmem,forcebin - nvk/nak regression: memory_model.message_passing fails on KeplerA - [ANV] dEQP-VK.api.copy_and_blit.copy_commands2.image_to_buffer_transfer_queue.2d_images.whole_r32g32b32_uint_linear and possibly others asserts on LNL - nir: validation failed after nir_remove_dead_variables in 3DMark Solar Bay - Build dependency on intel_wa.h missing in Intel vulkan driver - Missing polygons/vertices in CS2 on BMG - \`vn_ring`: use-after-release crash after \`vn_ring_destroy` on Virtio-Vulkan - venus: vkmark --winsys headless segfault (regression) - Vulkan headless WSI crashes when initializing swapchain on Asahi Linux running Apple M1 Max - lavapipe: Crashes on simple Descriptor Buffer test - make zink-radv-navi31-valve a pre-merge job - [RADV] Graphical glitches in Ghost of Tsushima on Polaris - radv: RT regressions - macOS: use of undeclared identifier 'free_zombie_glx_drawable' - macOS: Undefined symbols "_pipe_loader_drm_probe_fd", referenced from: _dri_get_drm_device_info in libdri.a[7](dri_util.c.o) - Segfault when activating DPMS on i915 hardware - RADV caps reported sparse address size at 4 GiB - nvk Blackwell support - hk: framerate limited/locked to 23 in RDR2 ingame menus (Vulkan) - i915: multiple dEQP tests asserts at nir_opt_group_loads.c:75: get_load_resource: Assertion \`!"" "tex instr should have a resource"' failed. - tu: VK_KHR_fragment_shading_rate broken when HelperInvocation is used - radv: regression: commit a7291074c800 break lighting in Like a Dragon: Infinite Wealth - [bisected regression] Latest mesa-git keeps /dev/dri/renderD129 always open with new applications even though they don't use it at all - spec\@arb_shader_storage_buffer_object\@max-ssbo-size\@fs stack overflow since cb558b2b88c2 - anv: enable compression on ASTC LDR emulation surface - High GPU usage when using Zink for eglgears_x11 (on X11) - Segfault in X11 image acquire code with timeout=0 - v3dv: regression in vkAllocateMemory importing gbm bo - Crash from iris_set_sampler_views in chromium/chrome with accelerated video decoding - rusticl: aco: LLVM outperforms ACO in clpeak for \`short` benchmarks on hawaii - rusticl: aco: Performance regression in clpeak for char benchmarks on hawaii - nir: Deprecate NIR_PASS_V - zink on tu assertion failing doing shader-db runs. - Race condition with timeline semaphores - nir_algebraic silently ignores operand conditions in some cases - lavapipe: valgrind triggers errors with CTS unit tests when creating a vulkan device - radv: gfx12 RGP captures don't support instruction timings for graphics pipelines - xe2: DMA Buffer exported modifier is incorrect - cleanup CI kernel patches - radv: more glcts fails KHR-GL46.shading_language_420pack.initializer_list_initializer* - radv: regression in KHR-GL46.gpu_shader5_gl.float_encoding - !36097 breaks Xwayland (& others) - GPU process crash via WebGPU shader - heap-buffer-overflow in Mesa build_interference_graph - radeonsi: Broken VAAPI video color conversion - Gallium HUD broken since !34054 Changes ------- Aaron Ruby (13): - gfxstream: Rename platform/linux to platform/drm - gfxstream: init vk_queues in CreateDevice() based on queueCreateInfo - gfxstream: Remove all "Yoda conditions" in gfxstream_vk_device.cpp - gfxstream: Downgrade some debug prints to traces - gfxstream: Remove duplicate/unnecessary frees in destroyDevice - gfxstream: Modify deviceName, driverVersion, driverName, driverInfo ... - gfxstream: Use the Mesa common tss_* TLS helper functions - gfxstream: Remove on_vkGetDeviceQueue* impls entirely - gfxstream: Pre-fetch the VkQueue objects from the host - gfxstream: Add init+tracking for the host-equivalent queue_family information - vulkan/wsi: No commandPool allocation required for WSI_SWAPCHAIN_NO_BLIT - gfxstream: Prune all guest-side KHR entrypoints that are provided with VK_VERSION_1_1 - gfxstream: address-space graphics requires kParamResourceBlob and kParamHostVisible Agate, Jesse (1): - amd/vpelib: Use Ceil Division Macro Ahmed Hesham (1): - rusticl: Fix negative CTS device tests Aitor Camacho (7): - nir: Set cursor in lower_sampler_lod_bias - meson: static link spirv-tools for darwin - wsi/metal: Cleanup unused members in wsi_metal_swapchain - wsi/metal: Fix wsi_metal_surface_get_formats2 - wsi/metal: Disable reference counting - wsi/metal: Fix size query and present result - wsi/metal: Backend addition for drivers built on top of Metal Aksel Hjerpbakk (5): - panvk: avoid cs jump block with no allocator - panvk: implement cs_extract64 & cs_extract_tuple - panvk: Use a single FBD for IR - panvk: pool large TLS allocations - panvk: clear big_bos on cmd pool reset with release bit Alejandro Piñeiro (4): - broadcom/compiler: update compact arrays comment - docs: GL_ARB_compute_shader is not a ES extension - v3d: use directly MESA_TRACE_SCOPE for additional context - v3d: expose GL_KHR_shader_subgroup for v71+ Aleksi Sapon (11): - meson: add missing x11 dependency on libloader_x11 - util: SWAP macro implementation for older MSVC versions - wsi/metal: current extents might not be known until swapchain is created - draw: fix missing line viewport transformation - draw: don't set the clipped window coordinate to NaN in debug - nir: Fix gnu-empty-initializer warning - nir: Fix nir.h MSVC compilation for C++ source files - wsi/metal: move VkFormat -> MTLPixelFormat conversion to wsi_common_metal_layer.m - wsi/metal: add support for color spaces - wsi/metal: fix cleanup on swapchain image creation failure - vk: Fix MSVC warning C4189 Alessio Belle (4): - pvr: Fix error value returned by pvr_rt_datas_init - pvr: Replace check on Mlist size with assert - pvr: Pass the PM/FW protect flag to the Mlist allocation - pvr: add device info for BXM-4-64 (36.56.104.183) Alexandros Frantzis (1): - egl/wayland: Support pbuffer surfaces Ali, Nawwar (1): - amd/vpelib: add FL capabilitie and lut container size Alyssa Rosenzweig (145): - nir/opt_preamble: add sampler class - nir: add bindless_sampler_agx intrinsic - hk: dedupe hk_buffer_view_descriptor - hk: push descriptor set addresses - hk: embed texture desc in set - hk: stop pushing image heap - hk: stop reserving uniform for image heap - hk: drop image heap - asahi: drop image heap decode - agx: report sampler state count - hk: plumb sampler state counts - hk,agx: promote bindless samplers - hk: optimize desc set addr push - hk: only pass sampler heap if needed - nir: add nir_mov_scalar helper - treewide: use nir_mov_scalar - util: crib SWAP macro from freedreno - nir: mark exact fmul in ldexp lowering - nir: introduce "inexact associative" property - nir: restrict associativity to binary operations - nir: unmark 24b multiply as associative - agx: fix dead phis - agx: simplify block image store offset - agx: optimize txl LOD - agx: optimize imgwblk uniform - agx: add immediate load ts/ss encodings - agx: use immediate load ts/ss forms - hk: use amul instead of imul - hk: always lower bindless samplers - hk: readvertise required bgra4 format - nir: introduce ergonomic tex builder - nir/lower_drawpixels: use tex builder - nir/lower_bitmap: use more effective NIR - vulkan/nir_convert_ycbcr: use more effective nir - radv: remove redundant nir->info.internal = true - tu: use more effective NIR in meta shaders - freedreno: use tex builder - asahi: use tex builders - dzn: drop redundant internal = true writes - nir: add vbo_stride_agx - hk: support static vertex input state - util: make SWAP safe for MSVC - nir: add nir_alu_src_rewrite_scalar helper - nir: add ALU reassocation pass - agx: make sure denorm flushing really happens - agx: run more opt passes - agx: reassociate ALU - vulkan: fix shader linking with common pipelines - glsl,nir: factor out nir_opt_varyings_bulk - nir: handle frag_coord_z/w intrinsics - nir/opt_vectorize_io: allow i/o semantics w/o component - nir/divergence_analysis: handle more AGX - agx/nir_lower_gs: handle XFB corner - hk: optimize varyings - dzn: use common SWAP - treewide: use SWAP macro - nir/lower_system_values: simplify load_helper_invocation lowering - nir: drop load_sample_id_no_per_sample - nir: add nir_def_as_* helpers - nir: add nir_def_block helper - treewide: use nir_def_as_* - treewide: simplify nir_def_rewrite_uses_after - treewide: use nir_def_block - asahi: clang-format - clc: force exact! across libclc - asahi: drop sink/move in GS code - agx: try to rematerialize to improve occupancy - asahi: use native colour masking - hk: kill psiz writes via topology, not feature - hk: only enable image view min LOD for dx12 - asahi: optimize pass type with depth-only passes - asahi,hk: optimize no-op FS - asahi: rename compressed 1 to just compressed - agx: add foreach_reg_{src,dest} - agx: track block divergence - agx: fix reg cache printing - agx: fix export instructions in the IR - agx: fix simd reduce forcing no cache bit - agx: fix cache bit packing - agx: plumb is_alu query for reg cache opt - agx: lower export even later - agx: set register cache hints - agx: handle 16-bit coordinates - asahi: use 16-bit coordinates for bg program - libagx: factor out query_report - libagx: port reset query helper to libagx - hk: use new reset query kernel - people: add John Anthony - nir: add nir_inline_sysval pass - brw: replace lower_fs_msaa with nir_inline_sysval - pan/bi: replace specialize_idvs with nir_inline_sysval - lvp: replace lower_ray_tracing_stack_base with inline_sysval - panfrost: don't use nir_lower_printf_buffer - nir,agx: pull lower_printf_buffer into backend - nir: gather info in opt_varyings_bulk - nir: gather interpolation qualifiers - nir/opt_varyings: link interpolation qualifiers - asahi: use NIR gathered interpolation - asahi: inline UVS indices - asahi: enable virtgpu support - panvk: rewrite pan_nir_lower_static_noperspective - agx: gate scratch opt on internal shaders - asahi: clang-format - asahi: reduce ppp alignment - hk: fix todo - hk: clarify command pool types - hk: fix pathological RAM use for tess emulation - hk: drop unused - hk: reduce storage desc - nir/lower_subgroups: add lower_fp64 option - nir: plumb ballot options - glsl: lower fp64 subgroup ops - agx: lower fmin/fmax scans - asahi: implement KHR_shader_subgroup - agx: drop bounds check optimize pass - people: update Alyssa's email - mailmap: add Alyssa's Intel e-mail address - hk: assume largePoints always set - asahi: fix drm-shim - util: add util_bit_swap macro - util: add boolean lookup table helpers - util: add unit tests for util/lut.h - agx: use util_lut2 - nir/lower_flatshade: clean up - brw: drop unused brw_kernel code - brw: drop indirection on compiler options - brw: hoist shared options out of the stage loop - brw: cleanup int64 option set - anv,hasvk: do not use unify_interfaces - brw: drop printf info plumbing - intel: drop clamp_fragment_color handling - intel: drop legacy flatshade handling - util/shader_stats: allow "hidden" stats - brw,anv: use XML-based stats - util: add BITSET_CALLOC helper - treewide: use BITSET_CALLOC - brw/nir_lower_alpha_to_coverage: eliminate goto - brw/nir_lower_fs_barycentrics: avoid nir_def_rewrite_uses_after - brw/nir_lower_sample_index_in_coord: use helpers - brw/nir_lower_shader_calls: use helpers - brw/nir_lower_storage_image: use helper - intel/nir_blockify_uniform_loads: use helpers - treewide: don't check before free - anv: use D3D-compatible texturing for Proton - asahi,ail: fix multi-plane imports Alyssa Ross (4): - gfxstream: guest: don't use transitional LFS64 API - docs: update GitLab option name - meson.build: remove dead code - meson.build: set with_clc for asahi tools Anna Maniscalco (4): - tu: Add support for realtime vk priority - mailmap: Update my name - freedreno/registers: add CP_ALWAYS_ON_CONTEXT - freedreno/afuc: Add x1e fw-id Ansari, Muhammad (1): - amd/vpelib: VPE Events Antonio Ospite (32): - ci/android: update comment about ANDROID_CTS_MODULES - ci/android: fix exit code from android-cts-runner.sh and android-deqp-runner.sh - zink: fix assigning _Bool to _Bool* - nir: fix returning _Bool instead of pointer - crocus: fix returning _Bool instead of pointer - zink: fix returning _Bool instead of pointer - anv: fix returning _Bool instead of pointer - nak: fix returning _Bool instead of pointer - radv: fix returning _Bool instead of pointer - dril: fix returning _Bool instead of pointer - microsoft/compiler: fix returning _Bool instead of pointer - asahi: fix returning _Bool instead of pointer - etnaviv: fix returning _Bool instead of pointer - lima: fix returning _Bool instead of pointer - broadcom/compiler: prevent FALLTHROUGH error with C23 - glsl: rename state name to avoid conflicts with future changes - build: stop calling unreachable() without arguments - build: avoid redefining unreachable() which is standard in C23 - util: avoid calling UNREACHABLE(str) macro without arguments - libcl: avoid calling UNREACHABLE(str) macro without arguments - nak/nouveau: silence errors about never used methods - compiler/rust: fix errors about hiding elided lifetime - ci/android: add rust compiler to create-android-cross-file.sh - ci/android: add comment about updating tags to create-android-cross-file.sh - nvk: silence error when cross-building for Android - subprojects: fix ignore exception for files under packagefiles/ - meson: handle dep_libdrm before the driver specific libdrm modules - ci: bump DEBIAN_BUILD_TAG to include all the android/rust changes for nvk and panvk - ci/android: enable cross-building nvk and panvk for Android - radv: don't include amdgpu.h directly - radv: fix building with libdrm as a submodule - device-select: fix build errors on some stricter build configurations Arkadiusz Hiler (1): - wsi/display: Avoid connector reprobes in wsi_GetRandROutputDisplayEXT Arseny Kapoulkine (1): - ac/rgp: Warn when RGP capture can't be saved without libelf Asahi Lina (1): - asahi: Ensure shared BOs have a prime_fd Ashish Chauhan (10): - pvr: temporary spm tweaks - pvr: Add support for gpu multicore MC1 configurations - pvr: Implement WA BRN_72168 - pvr: Implement WA BRN_72463 - pvr: Enable PBE_FILTERABLE_F16 - pvr: Feature support TPU_PARALLEL_INSTANCES - pvr: Enable PDS_DDMADT - pvr: Enable shaderStorageImageExtendedFormats - pvr: Drop broken driver environment variable check for BXS-4-64 - pvr: Drop '-experimental' suffix from the 'imagination' build option Ashley Smith (4): - mesa: Fix support for GL_EXT_shader_clock - panfrost: Enable shader_atomic_int64 for gallium - panfrost,mesa: Fix versions for EXT_shader_realtime_clock - panfrost,mesa: Fix versions for EXT_shader_clock Assadian, Navid (3): - amd/vpelib: Exit when VPE not support in debug - amd/vpelib: Add necessary pointer casting - amd/vpelib: Add new colors to visual confirm Autumn Ashton (4): - radv: Implement VK_KHR_video_encode_quantization_map - radv: Support VK_IMAGE_TILING_OPTIMAL for quantization maps - radv: Allow MUTABLE_FORMAT and EXTENDED_USAGE for qp_map images - nvk: Implement VK_NVX_image_view_handle Bas Nieuwenhuizen (2): - device-select: Fix error check. - radv: use vk_drm_syncobj_copy_payloads Benjamin Cheng (11): - vulkan/video: Add vk_video_is_profile_supported() - radv/video: Fix video profile reporting - radv/video: Report extra image usages - vulkan/query_pool: Store video encode feedback - radv: Output requested encode query results only - radv/video: Fill maxCodedExtent caps first - radv/video_enc: Cleanup slice count assert - radv/video: Override H265 SPS block size parameters - radv/video: Override H265 SPS unaligned resolutions - vulkan/video: NULL check codec-specific chain - radv/video: Fix dummy DPB addresses Benjamin Otte (1): - device_select: Allow shortcut names for device types Bo Hu (2): - gfxstream: update codegen for event save and load - gfxstream: [vulkan snapshot]: update code gen for vkUpdateDescriptorSet change Bohan Yu (1): - gallium: Fix LLVMpipe function parameter of Vector type call load mismatch Boris Brezillon (53): - panfrost: Add get_device_reset_status() to the CSF backend - panfrost: Add a GPU fault injection mechanism - panfrost: Log when an unusable group caused a context re-initialization - util/format: Auto-generate the enum pipe_format definition - util/format: Use more descriptive names for YUV formats - util/format: Add subsampling info to our YUV-as-RGB format names - util/format: Auto-generate a bunch of YUV helpers - pan/mod: Add a pan_mod_get_handler() implementation when PAN_ARCH is defined - pan/mod: Replace ::supports_format() by ::test_props() - pan/image: Provide two helpers to check image viability - panvk: Use pan_image_test_props() to do our modifier check - panfrost: Don't check for MTK_TILED when walking the native modifiers list - dri: Don't pretend we can lower NV15/NV20 when we can't - panfrost: Use pan_image_test_modifier_with_format() to do our modifier check - panvk: Remove leftovers from CPU-side min/max index calculation - panvk: Fix disjoint image memory binding - panvk: Fix panvk_image_can_use_afbc() for GetPhysicalDeviceImageFormatProperties2() - panvk: Pass a correct aspect to panvk_plane_index() - panvk/jm: Preload the FB even if we have no draws queued - panvk/jm: Automatically open a batch in dispatch_precomp() - panvk/jm: Add a JM barrier on clear AFBC jobs - panfrost: Fix panfrost_batch_to_fb_info() for stencil-only attachments - pan/mod: Allow testing if a modifier is optimal - pan/format: Fix the mapping for Z32_FLOAT on v7+ - panfrost: Explicitly reject AFBC(Z32) - pan/afbc: Add missing S8 and Z32 cases to pan_afbc_format() - panvk: Hook-up optimal modifier selection - util/format: Autogen type conversion helpers - pan/afbc: Cache the pan_afbc_mode selection - panfrost: Explictly filter out AFBC(SNORM) - pan/desc: Upgrade writeback format to RAW32 on v9+ when AFBC(RAW24) - pan/afbc: Allow AFBC on UINT/SINT/SNORM types on v9+ - panvk: Don't allow AFBC if the format format is mutable on v7- - panvk: Make panvk_meta.h per-gen - panvk: Consolidate image copy format selection - panvk: Disallow AFBC(D24S8) if separateDepthStencilLayouts=true - panvk: Make AFBC an opt-out - util/format: Add a Z24_UNORM_PACKED format - pan/lib: Hook-up Z24_UNORM_PACKED support - panvk: Initialize panvk_image::plane_count early - panvk: Pass an image to panvk_plane_count() - panvk: Stop using panvk_image_can_use_afbc() in panvk_image_can_use_mod() - panvk: Add planar Z24S8 support - drm-uapi: Sync panfrost_drm.h - pan/kmod: query and cache available context priorities from KMD - panfrost: Support JM context creation and destruction - panfrost: Support debugging JM context priorities with env vars - panvk: Fix ordering in prepare_draw() - panvk: Don't expose low/high priority queues on Bifrost - vk/meta: Support DS <-> color copies - panvk: Fix panvk_interleaved_copy() formatting - panvk: Fix host copies on planar DS resources - panvk: Only use Z24_UNORM_PACKED for AFBC images Boyuan Zhang (5): - pipe: add gaps_in_frame for h264 - frontends/va: get gaps_in_frame for h264 dec - radeon/vcn: add gaps_in_frame flag to h264 sps - ci/fluster: remove 3 pass cases resulted by gaps_in_frame - radeonsi/vcn: adjust subsample size alignment Brais Solla (2): - r300: Added support for GL_ATI_meminfo and GL_NVX_gpu_memory_info - r300: move r300_query_memory_info to r300_screen.c Caio Oliveira (93): - brw: Fix cmat conversion between bfloat16 and non-float32 - brw: Move insert/remove code to the block - brw: Add more specific brw_builder helpers - brw: Use a more specific builder helper in combine constants - brw: Use a builder to track position in lower_simd - brw: Make brw_builder() shader constructor use CFG if available - intel/decoder/tests: Sort gentest.xml file - intel/genxml: Add support for dword/bits in fields to gen_sort_tags.py script - intel/genxml: Add support for dword/bits in fields to rest of the code - intel/genxml: Convert field format from start/end to dword/bits - intel/genxml: Remove support for start/end atttributes - spirv: Load block descriptors as soon as we hit them - spirv: Implement SPV_KHR_untyped_pointers - brw: Use ralloc helpers for string handling in brw_eu_validate - brw: Remove extra iteration on instructions from brw_opt_address_reg_load - spirv: Update headers and metadata from latest Khronos commit - vulkan: Update enum_to_str conversion to handle ARM enum names - vulkan: Update headers/xml for 1.4.325 - anv: Advertise VK_KHR_shader_untyped_pointers - brw: Define order for fixes in 3-src operand fix - brw: Make sure copied instruction don't copy the list pointers - brw: Move resize_sources() earlier when lowering FIND_LIVE_CHANNELS - brw: Only access valid sources in lower_btd_logical_send() - brw: If the instruction is already a SEND, no need to resize sources - brw: Avoid invalid access when compacting out-of-bounds JIP/UIP - brw: Add disabled test for MAD constant folding - brw: Fix folding case for MAD instruction with all immediates - brw: Fix checking sources of wrong instruction in opt_address_reg_load - brw: Add brw_shader_params - brw: Pass per_primitive_offset in brw_shader_params - anv: Allocate prog_data->param array when making internal kernels - intel/brw: Remove brw_shader::import_uniforms() - intel/brw: Simplify tracking of dispatch_width_limit in brw_compile_fs - intel/brw: Simplify variant tracking in brw_compile_fs - intel/brw: Take shader in the brw_generator::generate_code() parameters - brw: Run validation as soon as we have the CFG around - brw: Fix printing of blocks in disassembly when BRW is available - util: Avoid invalid access in ralloc_print_info() - brw: Add \`FILE \*\` parameter to dump_assembly - brw: Add and use more brw_validate.cpp macros - brw: Use uint16_t for size_written - brw: Centralize brw_inst allocation - brw: Allocate brw_inst::src with ralloc - brw: Remove builtin sources from brw_inst - brw: Bundle the allocation of brw_inst and its sources - brw: Let the builder fill the sources of brw_inst - brw: Allow emit instruction with only number of sources - brw: Pass brw_shader in fold_instruction - brw: Add and use brw_transform_inst() - brw: Add brw_builder::SEND() helper - brw: Add brw_builder::URB_READ and URB_WRITE helpers - brw: Remove the extra function call when lowering samplers - brw: Add initial support for different instruction kinds - brw: Add brw_send_inst - brw: Add brw_tex_inst - brw: Add brw_mem_inst - brw: Add brw_dpas_inst - brw: Add brw_load_payload_inst - brw: Add brw_urb_inst - brw: Add brw_fb_write_inst - brw: Add a generic LOGICAL instruction kind - brw: Allocate only brw_inst for BASE instructions - brw: Repack brw_inst fields - brw: Don't use individual rallocs for each instruction - brw: Fix encoding of 3-src dst in Xe2+ - egl: Set atexit() handler during initialization - egl: Don't maintain a list of AtExit functions - intel/mda: Add code to produce mesa debug archives - brw: Use debug archive file with INTEL_DEBUG=mda - brw: Include some NIR states in the debug archive - brw: Also include the final disassembly in the debug archive - anv: Refactor anv_shader_compile result handling - anv: Create archive file when using INTEL_DEBUG=mda - iris: Create archive file when using INTEL_DEBUG=mda - intel/mda: Add tool to inspect mesa debug archives - intel/mda: Add search/searchall commands - intel/mda: Add -U and -Y diff options - intel/mda: Handle non-contiguous object versions in mda.tar files - intel/mda: Add pager support - intel/mda: Add MDA_OUTPUT_DIR and MDA_PREFIX environment variable support - intel/mda: If MDA_PREFIX=timestamp use the actual timestamp as a prefix - intel/mda: Allow more toplevel directory names inside mda.tar files - intel/mda: Use archive filename as directory name instead of hardcoded "mda/" - intel/mda: Add MDA_FILTER to select which archives to generate - brw: Identify if/break/endif special case before emission - intel/executor: Destroy syncobjs after using them - intel/executor: Expose extra command line arguments to script - intel/executor: Drop check_ver and check_verx10 functions - intel/executor: Expose a devinfo table - intel/executor: Add script directory to \`package.path` - intel/executor: Add DPAS examples for HF/F, UB/UD and BF/F - intel/executor: Add a matrix multiplication example - brw: Add variable for opcode in the brw_set_* high-level helpers Calder Young (13): - nir/builder: Add helper for building uvec8 immediates - brw,anv: Reduce UBO robustness size alignment to 16 bytes - isl: Add support for creating layered surfaces for video encode/decode - anv: Add support for creating layered surfaces for video encode/decode - anv: Add support for using layered surfaces in H.264 and H.265 video coding - anv: Add support for using layered surfaces in AV1 video decoding - anv: Add support for using layered surfaces in VP9 video decoding - anv: Report disjoint images as unsupported for video usage - anv: Update video test expectations for layered_dpb - anv: Advertise only OUTPUT_COINCIDE_BIT for AV1 video decoding - anv: Add support for AV1 film grain sythesis on Xe2+ - anv: Fix tiling for AV1 IntraBC surface on Gfx125+ - isl: Fix noncoherent framebuffer fetch when base_level != 0 Caleb Callaway (6): - spirv: Fix RT raygen hit attribute validation error - compiler: use PATH_MAX for SPIR-V capture filename - compiler: BLAKE3 ID for SPIR-V capture - compiler: auto-stage file ext for SPIR-V capture - compiler: SPIR-V shader replacement - compiler: document SPIR-V capture + replace Caterina Shablia (17): - vulkan/runtime: add vk_image_subresource_slice_count - panvk/csf: change get_cs_deps to be add_cs_deps - panvk: add a meta command for transitioning image layout - panvk: call cmd_transition_image_layout for each image memory barrier - panvk: do not zero AFBC when an image is being bound - panvk/csf: plop the stage and access masks into panvk_sync_scope - panvk: adjust formatting in csf/panvk_queue.h - pan/kmod,panvk: use uint64_t and not size_t for device sizes - pan/kmod: introduce pan_kmod_vm::pgsize_bitmap - panvk: introduce panvk_get_gpu_page_size - pan/kmod,panvk: rewrite how alignment for an allocation is chosen - panvk: add blackhole bo - panvk: add PANVK_DEBUG=force_blackhole - panvk: implement sparse resources - panvk: add bind queue - panvk: report support for sparse{Binding,ResidencyBuffer} - docs/features: add sparse{Binding,ResidencyBuffer} on panvk/v10+ Chan, Roy (2): - amd/vpelib: fix memory corruption - amd/vpelib: check stream_count as well before accessing streams Chang, Tomson (2): - amd/vpelib: Add missing swizzle and dcc info - amd/vpelib: Update register header and definitions macros Charles Giessen (1): - docs: Use correct ICD path in install.rst Chia-I Wu (2): - panvk: require gpu_can_query_timestamp for calibrated timestamps - panvk: use common calibrated timestamp support Christian Gmeiner (63): - v3dv: Make use of hash table helpers - freedreno/rddecompiler: Make use of hash table helpers - etnaviv: Update headers from rnndb - etnaviv: Handle 64-bit pixel formats in texture sampler TS setup - etnaviv: Fix vertex format normalization for signed integer formats - etnaviv: Fix negative LOD value encoding in texture descriptors - etnaviv: Emulate rasterizer_discard - etnaviv: hwdb: Add MSAA_FRAGMENT_OPERATION feature - etnaviv: Only emit VIVS_PS_MSAA_CONFIG if GPU support it - etnaviv: Update headers from rnndb - etnaviv: Emit alpha-to-coverage dither - etnaviv: Add support for alpha_to_coverage - etnaviv: blt: Add r8_unorm format support - etnaviv: blt: Add r8g8_unorm format support - etnaviv: blt: Clear only requested color buffers - etnaviv: rs: Clear only requested color buffers - etnaviv: Optimize sampler view iteration with u_foreach_bit(..) - etnaviv: blt: Extend translate_blt_format(..) - etnaviv: blt: Add hardware based mipmap generation - etnaviv: Enable texture_multisample for deqp testing - etnaviv: isa: Add tg4 instruction - etnaviv: nir: Add nir_texop_tg4 offset lowering - etnaviv: Add support for ARB_texture_gather - etnaviv: Do not update derived states during non-draw force flush - etnaviv: re-format using clang-format - etnaviv: Replace unsupported blit debug message with detailed dump and assertion - r300: re-format using clang-format - radv: re-format using clang-format - nak: Move dataflow to compiler crate - etnaviv: hwdb: Add S8 feature - etnaviv: Update headers from rnndb - etnaviv: rs: Support 8bpp for clears - etnaviv: Support PIPE_FORMAT_S8_UINT stencil format - imagination: Re-format using clang-format - clang-format: Add src/imagination to .clang-format-include - nir/opt_algebraic: optimize f2i32(fround_even(x)) to f2i32_rtne(x) - etnaviv: blt: Enable scissored clear - etnaviv: Update headers from rnndb - etnaviv: hwdb: Add HWTFB cap - etnaviv: Support hw based rasterizer_discard - etnaviv: Pass context to acc sample provider supports(..) function - etnaviv: Support PIPE_QUERY_PRIMITIVES_EMITTED - etnaviv: Implement stream output target management - etnaviv: Implement hardware based streamout support - etnaviv: Fix util_blitter_save_so_targets(..) call - docs/features: Mark GL_EXT_transform_feedback as done for etnaviv/HWTFB - etnaviv: Update headers from rnndb - etnaviv: Support ARB_stencil_texturing - etnaviv: Expose faked xfb support when DEQP debug flag is enabled - pvr, pco: Set has_f2i32_rtne to true - etnaviv/ci: Add per-gpu GLES2 extension lists - etnaviv: Allow 128-bit formats when DEQP debug flag is enabled - etnaviv: Add 128bit emulated formats - etnaviv: Add 128 bit format helper - etnaviv: Add 128-bit format tilling - etnaviv: Support 128 bit formats transfers - etnaviv: 128 bit format needs to be CPU tiled - etnaviv: Do not use TS for emulated 128 bit formats - etnaviv: Implement 128-bit format emulation using dual 64-bit layout - etnaviv: blt: Support 128 bit clear operations - etnaviv: blt: Support 128 bit blit operations - anv: Fix needs_temp_copy() incorrectly matching depth/stencil formats - meson: require sysprof-capture-4 >= 4.49.0 Christian Meissl (1): - panfrost: take reference from pool used for allocation Christoph Neuhauser (3): - egl: Fix DRI utility function compilation on macOS - iris: Increase max_shader_buffer_size to max_buffer_size - egl: Fix invalid device UUID returned by EGL_EXT_device_persistent_id Christoph Pillmayer (25): - panvk: hide utrace behind more generic interface - panvk: Make panvk_utrace_record_ts wait mask configurable - panvk: Make ts in panvk_instr_begin_work synchronous - panvk: Make most end work instrumentation synchronous - panvk: Support VK_DESCRIPTOR_TYPE_MUTABLE_EXT on v9+ - panvk: Support DESCRIPTOR_POOL_CREATE_HOST_ONLY_BIT - panvk: Advertise VK_EXT_mutable_descriptor_type on v9+ - vk/sync: Pass dependencyFlags in vk_common_CmdPipelineBarrier - panvk: Fix preserved metadata in lower_input_attachment_load - panvk/utrace: Alloc utrace copy buf from userspace heap - panvk/utrace: Remove dynamic alloc from utrace clone builder - panvk/perfetto: Handle re-submittable command buffers - panvk/perfetto: Drop zero duration events - panvk: Add support for moving constants to the FAU - pan/bi: Move some constants into FAU entries - pan/va: Pull out constant swizzle handling - pan/bi: Prioritize consts moved to the FAU - nir/opt_algebraic: Convert a + b + a to b + 2a - pan: Add gpu variant to compile inputs - panfrost: Wire up gpu_variant to pan_compile_inputs - panvk: Wire up gpu_variant to pan_compile_inputs - pan/clc: Wire up gpu_variant to pan_compile_inputs - pan: Lift pan_get_model into its own lib - pan/bi: Normalize with pan_model.rates - pan/va: Remove redundant MOVs from va_lower_split_64bit Collabora's Gfx CI Team (11): - Uprev ANGLE to 6a04a50f98cac71b25464d10289ce7a013841caf - Uprev Piglit to 0980079dcfb5adbad873d88e00181268f55cb8ef - Uprev Piglit to c3a3e29d59e0972650a6d30d20de930c87739c14 - Uprev ANGLE to 995c4c4d89ed6a5c28b210e9c0f83eb4f8b6e2f5 - Uprev Piglit to 28d1349844eacda869f0f82f551bcd4ac0c4edfe - Uprev ANGLE to 1df3b59f8730b56b4770595d4d69f36d5283333f - Uprev Piglit to 517270ccca11a795d2f29bd723c362eb6ef9ce8f - Uprev Piglit to a70c33045c59310f972dbbdb33f322eb209971bc - Uprev ANGLE to 538129c6b3c17dc864101c7a4af4b74b00706f82 - Uprev ANGLE to 8ed16003f27125f27cbb87578368e447043420d3 - Uprev Piglit to 4147e9d7aeb8ba26ffc25a90fc237588bcb3bb11 Connor Abbott (62): - tu: Don't keep track of acceleration structure sizes - freedreno: Add bin scaling registers - freedreno: Document GRAS_SC_BIN_CNTL::FORCE_LRZ_DIS - freedreno: Add HW bin scaling feature - tu: Add documentation for VK_EXT_fragment_density_map - tu: Use GRAS bin offset registers - tu: Enable LRZ with FDM - ir3: Simplify and rationalize shading rate LUT - freedreno: Add common VRS helpers - ir3: Use common shading rate lookup table - tu, freedreno: Document GRAS shading rate LUT - vulkan/queue: Fix VkTimelineSemaphoreSubmitInfo sanitization - tu: Refactor BO deletion - freedreno/drm: Import new UABI for VM_BIND - tu: Align BO size to page size - tu: Fix CmdBindTransformFeedbackBuffersEXT size handling - tu/drm: Enable VM_BIND - tu/knl: Add an API for sparse binding - tu/drm: Add support for sparse binding - tu/kgsl: Add support for sparse binding - tu: Initial support for sparse binding - tu: Support sparseResidencyAliased - freedreno/ci: Add sparse-related a618 skips - freedreno/ci: Skip dEQP-VK.memory.mapping.*.full.variable.* - freedreno/ci: Update kernel with VM_BIND fixes - freedreno/ci: Update a750 expectations - zink: Make sparse always wait on pending gfx commands - tu: Don't decrement implicit_sync_bo_count with VM_BIND - freedreno/fdl: Expose fdl6_is_r8g8_layout() publicly - freedreno/fdl: Refactor and expose bank swizzling logic - freedreno/fdl: Handle cpp=32 and cpp=64 when getting macrotile size - freedreno/fdl: Handle layout differences for r8g8 images - freedreno/fdl: Add sparse layout support - tu: Support sparse residency for images - ir3: Assemble and disassemble rck modifier - ir3: Implement sparse residency check - tu: Expose shaderResourceResidency - ir3: Assemble and disassemble .clp modifier - ir3: Support min_lod tex source - tu: Advertise shaderResourceMinLod - freedreno/ci: Add a750 sparse skips - tu: Lower ViewIndex to 0 when multiview is disabled - freedreno: Add blit_wfi_quirk and use in turnip - tu/drm: Split out iova allocation and BO allocation - tu: Add support for a "lazy" sparse VMA - tu: Make tu_image point to tu_device_memory instead of tu_bo - tu: Implement transient attachments and lazily allocated memory - freedreno: Don't program non-context reg with CRB - tu: Fix 3d load and clear when FDM bin offsets are in use - tu/fdm: Use better bounds for LRZ overallocation with FDM offset - tu: Expose VK_EXT_dynamic_rendering_unused_attachments - tu: Reset \*_BIN_FOVEAT when not using FDM - freedreno: Don't stomp VSC registers - tu: Pass tu_queue to kernel create/destroy functions - tu/drm: Emulate combined gfx/sparse queues - tu: Support sparse binds on the gfx queue - tu: Fix RT count with remapped color attachments - tu: Don't patch GMEM for input attachments never in GMEM - tu: Fix 3d load path with D24S8 on a7xx - tu: Also disable stencil load for attachments not in GMEM - tu: Rename tu_render_pass_attachment::clear_views to used_views - tu: Fix attachment stores with subpasses with partial views Corentin Noël (8): - virgl: Stop using deprecated util_framebuffer_init - ci/piglit: Allow traces content-type to be binary/octet-stream - docs/features: Add missing llvmpipe extensions - docs/features: Add missing virgl extensions - tgsi: Drop TGSI_SEMANTIC_TESS_DEFAULT_OUTER/INNER_LEVEL - tgsi: Remove return type from tgsi_instruction_texture - android: Only include libdrm_intel for i915 as iris do not depend on it - virgl: Skip resource destruction only when there are actually needed references Daivik Bhatia (7): - v3d: remove unused functions from v3d_bufmgr.h - v3d: use Texture Data Formats enum in Texture Shader State struct - v3d: move format helpers to v3dx_format_table.h - v3d: replace raw integers with enum types in helper functions - broadcom/common: Optimize CSD super-group packing - broadcom/common: Add subgroup support to CSD super-group packing - broadcom/compiler: support arithmetic subgroup operations Dallas Strouse (1): - rusticl/device: skip loading devices in cfg(test) Daniel Almeida (2): - nouveau/headers: Import the video class headers from NVIDIA - nouveau: Handle video decode in nv_push_print() Daniel Schürmann (74): - util/time: add os_time_nanosleep_until() function - vulkan: implement VK_AMD_anti_lag as implicit vulkan layer - aco/tests: Fix p_startpgm definitions to registers - aco/ra: generalize register affinities - aco/ra: collect register affinities for all precolored operands. - aco/ra: don't optimize encodings on precolor affinity mismatch - aco/ra: propagate precolor affinities through phis - aco/ra: propagate precolor affinities through parallelcopies and tied definitions - aco/scheduler: improve scheduling heuristic - nir/opt_load_store_vectorize: only attempt to vectorize shared2 after exhausting other possibilities - nir/opt_load_store_vectorize: don't vectorize large shared2_amd loads - radv: only vectorize shared2 instructions during late optimizations - aco/isel: allow for large 8-bit vectors in extract_8_16_bit_sgpr_element() - ac/nir: use HW-requirements on alignment for vectorizing LDS - ac/nir_lower_mem_access_bit_sizes: Split unsupported shared memory instructions - aco/isel: rename emit_readfirstlane() -> emit_vector_as_uniform() - aco/isel: refactor load_shared() by directly matching NIR intrinsics to ACO opcodes - radv: unconditionally call ac_nir_lower_mem_access_bit_sizes() - aco/isel: refactor store_shared() by directly matching NIR intrinsics to ACO opcodes - aco/scheduler: check dependencies of entire clause upfront - aco/scheduler: Stop downwards scheduling after encountering the first clause - aco/scheduler: split downwards_move_clause() from downwards_move() - aco/scheduler: remove DownwardsCursor::insert_demand_clause - aco/scheduler: remove DownwardsCursor::clause_demand - aco/scheduler: short-cut downwards_move_clause() when no movement is done - aco/scheduler: ignore potential SMEM stalls when forming clauses - aco/scheduler: move clauses as batch - aco/scheduler: schedule VMEM store clauses during the regular forward pass - aco/scheduler: small refactor of schedule_VMEM() - aco/ra: don't clear lateKill operands in get_reg_create_vector() - aco/ra: add vector_info::index to indicate the Operand's index into the vector - aco/ra: don't set precolor affinities for already assigned temporaries - aco/ra: consider precolor affinities in get_reg_vector() - aco/ra: coalesce vector affinities with tied definitions - radv/rt: use ACCESS_CAN_REORDER when loading SBT entries - nir/algebraic: add pattern for (a << #b) * #c => a * (#c << #b) - nir/load_store_vectorize: also parse offsets through u2u64 if additions don't wrap around - nir/load_store_vectorize: hoist base addr instead of subtracting - nir/opt_offsets: allow for unsigned wraps when folding load/store_shared2_amd offsets - radv: allow for unsigned wraps for shared memory intrinsics in nir_opt_offsets - radeonsi: allow for unsigned wraps for shared memory intrinsics in nir_opt_offsets - aco/optimizer: remove DS offset optimization - aco: remove excess offset handling for load/store_shared - amd: don't allow unsigned wraps for shared memory offsets on GFX6 - nir/opt_offsets: call allow_offset_wrap() for try_fold_shared2() - nir/load_store_vectorize: Fix parsing offsets through u2u64 - radv: delay lowering global access - radv: delay lowering int64 - nir/divergence_analysis: check ACCESS_SMEM_AMD - ac/nir_lower_global_access: require no_unsigned wrap when extracting from 32-bit additions - ac/nir_lower_global_access: don't assume pack_64_2x32 is the same as u2u64 - radv: delay nir_opt_shrink_vectors - radeonsi: delay nir_lower_global_access - radv,radeonsi: call ac_nir_lower_global_access and nir_lower_int64 for gs copy shaders - ac/nir: switch load_smem_amd to use load_global - nir/divergence: don't assume that load_sample_positions_amd is always uniform - radv: use load_global instead of load_global_amd for load_sample_positions_amd - amd/lower_mem_access_bit_sizes: lower all SMEM instructions to supported sizes - amd/lower_mem_access_bit_sizes: also use SMEM for subdword loads - amd/common: merge radv_nir_opt_access_speculate() into ac_nir_flag_smem_for_loads() - radv: delay ac_nir_lower_mem_access_bit_sizes - ac/nir_flag_smem_for_loads: call divergence analysis internally - radv/rt: fix LDS size calculation with LLVM for inlined stages - radv: fix max_waves calculation for tesselation - radv: use lds_alloc_granularity alignment for stats - amd: change ac_shader_config::lds_size to bytes - radv: calculate LDS allocation requirements independently from the compiler - radeonsi: pass calculated LDS size to ACO - amd: add and use utility functions for LDS size encoding - amd/common: remove radeon_info::lds_alloc_granularity and radeon_info::lds_encode_granularity - aco: remove DeviceInfo::lds_encoding_granule and DeviceInfo::lds_alloc_granule - amd: keep ac_shader_config::lds_size unaligned - amd: change radeon_info::lds_size_per_workgroup for GFX10+ to 64KB - radv/null_device: set more options which affect compilation Daniel Stone (2): - ci/panfrost: Add wider EGL/multithread flakes - ci/freedreno: Skip overly-slow trace Danylo Piliaiev (30): - tu: Use safe-const binning VS when safe-const full VS is used - util/u_trace: Add scripts for perf analysis based on u_trace results - tu: Fix nullptr dereference in cmd_buffer tracepoint - util: Add function os_get_option_secure - util/disk_cache: Use os independent functions instead of getenv - util/disk_cache: Fallback to ftruncate if posix_fallocate not supported - util/disk_cache: Allow disk cache on Android if explicitly enabled - tu: Fix unaligned image_to_buffer on close to (1 << 14) width - tu/a6xx: Fix unaligned buffer_to_image on close to (1 << 14) width - ir3: Add EOLM and EOGM a7xx flags to NOP - tu: Use approx square tiles when FDM is enabled - freedreno/a750: Fix typo in recent magic regs change - tu: Fix the lack of IB size sanitization in several cases in tu_cs - tu/a7xx: Don't disable LRZ for empty FS when FDM is used - tu: Reset rp_trace on tu_reset_cmd_buffer - tu: Prevent dangling start_sysmem_clear_all tracepoint - egl: Bring back util_cpu_trace_init - tu: Reset BIN_FOVEAT regs for tiling with and without HW binning - freedreno/decode: Fix preamble decoding - tu/a7xx: Update reg stomping info to fix GPU crashes when stomping - tu: Destroy all mutexes used for device - tu/perfetto: Don't check sync_gpu_ts when emitting renderstage - tu/perfetto: Track GPU timestamps per-device - tu/perfetto: Make GPU clock sequence-scoped - tu/perfetto: Init perfetto datasources once - tu/perfetto: Use a separate track for VK_EXT_debug_utils labels - tu: Prevent GPU hang with occlusion query + certain depth state - tu: Synchronize access to copy_timestamp_cs_pool - vulkan: Always fill DS state for EXT_dynamic_rendering_unused_attachments - tu: Use cmd->rp_trace u_trace for draw calls Dave Airlie (11): - nak: disable imma 8x8x16 on Blackwell+ - nvk: add sm120 latencies via csv files. - spirv: move cmat store barrier after the store. - nouveau: Handle subchannels better in nv_push_print() - nir: add coop mat flexible dimensions lowering. - radv: add support for coopmat2 flexible dimensions - radv: consolidate cooperative matrix array sizes enumeration - nir: add nir_intrinsic_cmat_load_shared_nv - gallivm: handle u8/u16 const loads properly on big-endian. - nir/coopmat: fix non square load/store lowering for flexible dimensions - c11/threads: fix build on c23 David Rosca (129): - radeonsi/vcn: Correctly handle tile swizzle - radv/video: Fix encode when using layered source image - ac/surface: Add ac_modifier_supports_video - radeonsi/video: Use ac_modifier_supports_video - radv/video: Support DRM format modifier tiling - radeonsi/uvd: Set H264 gaps_in_frame_num_value_allowed_flag - radv/video: Don't allow DRM format modifier tiling on GFX < 9 - radv/ci: Add dEQP-VK.video.formats.* fails for navi10 and vega10 - radv/video: Add bit depth and profile check for AV1 encode - radv/video: Add bit depth and profile check for VP9 decode - radv/video: Set encodeInputPictureGranularity for AV1 encode - radv/video: Add radv_video_is_profile_supported - radv/video: Rework GetPhysicalDeviceVideoFormatPropertiesKHR - radv/video: Remove 10 to 8bit dithering support - radv: Reject linear modifier for video decode DPB - radv/ci: Update navi10 and vega10 expected failures - radv/video: Remove disabled slice header code for field encoding - radv/video: Set H264 encode cabac_init_idc and Cb/Cr QP offsets - radv/video: Always send the latency command - radv/video: Send slice control, spec misc and deblocking params every frame - radv/video: Add more encode session params overrides - radv/video: Fix encode bitstream buffer offset and alignment - radv/video: Fix setting H265 encode cu_qp_delta on VCN2 - radv/video: Fix session_init and rc_per_pic on VCN2 - radv/video: Disable rate control modes for H265 encode on VCN1 - radv/video: Use the new defines for H264 SPS info flags - frontends/va: Add H264 encode more_rbsp_data PPS flag - radeonsi/vcn: Use more_rbsp_data flag for H264 PPS encode - radeonsi: Add missing DEBUG_NAMED_VALUE_END to radeonsi_shader_debug_options - radeonsi/vcn: Always enable decode tier2 when supported - vulkan/video: Fix h265 level values - radeonsi: Move multimedia debug options to its own flags - radeonsi: Add debug option to disable tiling for video - radeonsi: Add debug options to disable video decode/encode tiers - wsi/display: Report supported formats based on plane formats - wsi/display: Add RGBA16, RGBA16F and A2RGB10(SRGB) formats - radv: Add timeout to video encode query - radv/video: Don't init vp9 probs table in message buffer - radv/video: Simplify vp9 q params - radv/video: Remove unused enum - ac/vcn_dec: Add RDECODE_IT_SCALING_TABLE_SIZE - radv/video: Use more common defines - radv: Fix alignment for linear video decode dst images - rusticl/ptr: Fix hidden lifetime warning - ac/vcn_dec: Add av1_intrabc_workaround - radeonsi/vcn: Enable AV1 decode workaround for gfx1153 - radv/video: Enable AV1 decode workaround for gfx1153 - vulkan/video: Add intra refresh support - radv/video: Add support for VK_KHR_video_encode_intra_refresh - auxiliary/vl: Map X6R10/X6R10X6G10 formats to R16/R16G16 - radeonsi: Map X6R10/X6R10X6G10 formats to R16/R16G16 - frontends/va: Cleanup CreateContext - frontends/va: Refactor vlVaVidEngineBlit - frontends/va: Change vlVaPostProcCompositor to take pipe_vpp_desc arg - frontends/va: Remove EFC support - frontends/va: Add support for decode/encode processing - radeonsi/vcn: Support EFC with encode processing - radeonsi/vcn: Support VPE with decode processing - radeonsi: Remove now unused si_vid_is_target_buffer_supported - pipe: Remove now unused is_video_target_buffer_supported - subprojects: Remove libdisplay-info wrap file - radeonsi/vcn: Disable H264 encode 8x8 transform when CABAC is disabled - radv/video: Disable H264 encode 8x8 transform when CABAC is disabled - radeonsi/vcn: Disable H264/5 constrained intra pred with rate control - radeonsi/vcn: Fix compatibility with old FW for encode - radeonsi/vcn: Fix HEVC encode cu_qp_delta with old FW - radeonsi/vcn: Fix HEVC encode transform_skip with old FW - ci: Add missing rust subprojects to meson/build.sh - radeonsi/vcn: Correctly set chroma location with EFC - radv: Use extra context for video encode queue with multiple VCN instances - radv/video: Fix VP9 loop filter and segmentation params - util/format: Add RGB lowering for single plane YUV formats - ac/vcn: Add RADEON_VCN_IB_COMMON_OP_RESOLVEINPUTPARAMLAYOUT - radv/video: Set rate control to default on reset - radv/video: Support quantization map on VCN5 - util/format: Add VK_EXT_ycbcr_2plane_444_formats formats - vulkan/format: Map VK_EXT_ycbcr_2plane_444_formats to pipe format - radv: Enable VK_EXT_ycbcr_2plane_444_formats - ci: Stop building VDPAU driver - mesa: Remove NV_vdpau_interop - Remove VDPAU - gallium/vl: Remove now unused filters - radeonsi/video: Remove support for interlaced buffers - pipe: Remove PIPE_VIDEO_CAP_PREFERS/SUPPORTS_INTERLACED - radeonsi/vcn: Fix calculating QP map region dimensions - radeonsi/vcn: Get rid of PIPE_ALIGN_IN_BLOCK_SIZE - radv/video: Always use OBU_FRAME in AV1 encode - radeonsi/uvd: Swap order of comparison to avoid warning - r600: Remove mpeg12 shader decoder support - r300: Remove mpeg12 shader decoder support - nouveau: Remove mpeg12 shader decoder support - gallium/vl: Remove mpeg12 shader decoder - gallium/vl: Fix building vl_stubs - r600: Implement resource_get_param - d3d12: Implement resource_get_param - frontends/va: Use resource_get_param instead of resource_get_info - pipe: Remove resource_get_info - radv: Change radv_vcn_write_event to a write memory func - radv/video: Check FW version before using WRITE_MEMORY - radv/video: Fix waiting on encode feedback query - radeonsi/vpe: Fix transfer function mapping to vpelib - frontends/va: Fix parsing VP9 frame header - frontends/va: Add VP9 use_prev_frame_mvs and segmentation_update_data flags - radeonsi/vcn: Use VP9 use_prev_frame_mvs and segmentation_update_data - ac/gfx10_format_table: Use new names for 422 subsampled formats - gallium/vl: Add new function to get RGB YUV conversion matrix - frontends/va: Set color properties when not using explicit color standard - frontends/va: Use new RGB YUV conversion matrix - gallium/vl: Remove vl_csc_get_matrix - frontends/va: Always advertise explicit color standard support - radeonsi/vcn: Stop using vpp colors standard - radeonsi/vpe: Stop using vpp colors standard - frontends/va: Stop using vpp colors standard - vl,frontends/va: Implement YUV->YUV matrix coeff conversion - vl,frontends/va: Implement gamma and primaries conversion - gallium/vl: Remove luma key support - gallium/vl: Remove vl_compositor_set_csc_matrix - pipe: Remove PIPE_VIDEO_CAP_VPP_SUPPORT_HDR_INPUT/OUTPUT - pipe: Remove pipe_video_vpp_color_standard_type - radeonsi/vcn: Support BT2020 matrix with EFC - ac/surface: Limit video modifiers to 64K_S also for VCN 2.2 - radv/video: Introduce two levels of write_memory support - radv/video: Only use write_memory for encode feedback with full support - radeonsi/vcn: Fix AV1 bidir compound encode with order_hint disabled - radv/video: Don't require encode FW version >= interface version - radv/video: Fix AV1 bidir compound encode with order_hint disabled - vulkan/video: Avoid NULL pointers in session parameters - radv/video: Correctly handle no feedback query for encode - radv/video: Add NULL checks for picture parameters Deborah Brouwer (1): - android: fall back to SwiftShader’s LLVM Derek Foreman (2): - dril: Skip some pipe formats to avoid breaking X - zink: Don't use VK_PRESENT_MODE_IMMEDIATE_KHR on wayland Dhruv Mark Collins (1): - tu/util: Allow setting all TU_DEBUG options from envvar and file Dmitry Baryshkov (2): - glx: provide glx.pc - ci: drop google-freedreno remnants Dmitry Osipenko (1): - virtio/vdrm: Fix varying offsets of struct vdrm_device members Dylan Baker (31): - meson: set the \`legacy-x11` option as deprecated - anv: avoid potential integer overflow in video address calculation - intel/brw: Fix implementaiton of \|= operator for enum - isl: prevent potential overflow before widen - blorp: Fix potential read of uninitaized elk fields in debug paths - anv: add assertion that tes and tcs data is non-null - anv: remove dead code - mailmap: Update for Dylan Baker - calendar: Update release dates and change 25.3 to Dylan - meson: use the wayland module - anv: don't attempt to memcpy if allocation fails - iris: Fix potential null deref in debug archiver - VERSION: bump for 25.3.0-rc1 - .pick_status.json: Update to 3b2f7ed918a5ad78c1d3756e9823a1616c1f21d7 - .pick_status.json: Update to ad421cdf2e68a1ccef80cb810c012c8469579cb6 - .pick_status.json: Mark c20e2733bf8f9bb595f1bcc68ebb3d0686ef28e4 as denominated - .pick_status.json: Update to 28fbc6addbda2ce3e264b41b6ad91a7a0d8eb788 - .pick_status.json: Update to e38491eb1850ab8b0082716b00f514f75e2a0e1a - VERSION: bump for rc2 - .pick_status.json: Update to fd55e874ed09a04447ebd4dae25c98df2621ef7d - .pick_status.json: Update to 45a762727cf8708392b6de38616909543c799923 - intel/compiler/brw: Add assert that we don't have a negative value - .pick_status.json: Update to 32b646c5976f64152a004d4c83962ca14c46154f - VERSION: bump for rc3 - .pick_status.json: Update to 33342848451ca06deb054fad94de3cea3a9efe63 - .pick_status.json: Update to e44a776f4751d665efc447d8fe8e6c01d25a60c5 - .pick_status.json: Update to 27d9e4ec2a13a957f416a234a93bf2f0c2c9c56c - VERSION: bump for 25.3.0-rc4 - .pick_status.json: Update to 04a0d512fa68a48bc2a2632a0a4ff2c3ac10c6ca - .pick_status.json: Update to 294e72e2b517bc744f909fbce9e154efa698dd10 - .pick_status.json: Update to 8f13905c5e38ac3921c4804b19fc0f50531b0317 Ella Stanforth (22): - util/list: Fix next instruction removal usecase for non safe iterators - util/list: Add iterator debug to more routines. - util/tests: Add list iterator tests - pvr: Use demote - nir: assert when we do not have a sample count when not using intrinsic - pco: Switch to common alpha_to_coverage intrinsic - pco: Switch to common alpha to coverage lowering - pco: Cleanup meson.build files - pco: Switch back to util/list - v3d: rename msaa resolve - v3d: Always lower frag color - v3d: Fallback to software blend support for formats that do not support blend. - v3d/compiler: Add unpacking instructions for normalised 16bit formats. - v3d/compiler: Lower load_output after logic operations - nir: add v3d specific intrinsic normalised to float conversion - v3d/compiler: implement normalised to float conversions - v3d/compiler: Implement 16bit normalised render targets. - v3d: Add support for 16bit normalised formats - v3dv: Take format plane when packing hw clear color - v3dv: Add normalisation flags to the format table - v3dv: Add support for 16bit normalised formats - pvr: implement buffer device address Emma Anholt (49): - wsi/display: Add some comments about what's going on in the code. - wsi/display: Add error messages to some shouldn't-be-hit paths. - wsi/display: Pull DRM format translation up a level. - wsi/display: Do connector setup before swapchain init. - ir3: Rename per_samp to sample_shading. - tu: Rename per_samp to sample_shading to match ir3. - freedreno: Drop min_samples handling code. - tu: Implement sampleShadingEnable by flagging uses_sample_shading. - nir: Move ST's force-persample-shading NIR pass to shared code. - nir/lower_sample_shading: Set the sample qualifier on in vars. - zink: Lower sample shading before we add_derefs(). - ci/radeonsi: Add a flake on mendocino that appeared yesterday. - nir,agx: Move AGX's loop (generalized) to shared NIR code. - tu: Use nir_opt_reassociate. - ci/tu: Generalize the subgroupclustered pre-merge skips. - ci/tu: Do more generalization of the tess flakes. - i915: Avoid calling drm_intel_get_aperture_sizes(). - Revert "tu: Use nir_opt_reassociate." - vk/runtime: Set GPU_MULTI_WAIT on the drm syncobj type. - tu: Use the common syncobj sync type for the layered timelines. - tu: Fix the comment about DRM_CAP_SYNCOBJ_TIMELINE support. - ci/tu: Generalize the FDM flakes and link an issue. - ci/tu: Drop highp.scalar xfail. - ci/tu: generalize the multisample_resolve tess/gs flakes. - tu: Disable LRZ writes after most stencil-write operations. - vulkan/wsi: Add comments about the WSI's syncing, and KHR_display stuff. - vulkan/wsi: Add a test for kernel 6.0 sync file import/export ioctls. - wsi/drm: Do the dma_buf_semaphore setup at swapchain creation time. - wsi/drm: Don't request implicit sync if we're doing implicit sync ourselves. - tu: Move the BO implicit sync flag handling to a BO allocation flag. - ir3: Don't try to use indirect access in the alias table. - util/u_queue: Fix data race on num_threads during finish. - ir3: Enable nir_opt_shrink_stores. - ir3: Enable nir_opt_shrink_shrink_vec_array_vars. - ir3: Use a bitset for the defs-seen table. - ir3: Use a linear allocation context for ir3_registers. - ir3: Use a linear allocation context for ir3_instructions. - d3d10umd: Add missing dependency on u_formats codegen. - treewide: Make exported DRM FDs read-write. - ir3: Avoid O(n^2) behavior in rpt validation. - nir: Add a shader bisect tool. - radv: Restore marking WSI image's mem->buffer as uncached. - radv: Allocate BOs as implicit sync even if the WSI is doing implicit sync. - ir3: Move the big block of C support code out of the parser .y file. - ir3/parser: Make sure relative accesses have a size set. - ir3: Use bitset range operations. - wsi: Fix the flagging of dma_buf_sync_file for the amdgpu workaround. - nir/shrink_stores: Don't shrink stores to an invalid num_components. - v3dv: Fix assertion failure for not-found primary_fd during enumeration. Eric Engestrom (247): - VERSION: bump to 25.3 - docs: reset new_features.txt - docs/releasing: add missing "track remote staging branch" command in instructions - docs: update calendar for 25.2.0-rc1 - docs: update calendar for 25.1.6 - docs: add release notes for 25.1.6 - docs: add sha sum for 25.1.6 - gfxstream: move variables into the #ifdef that uses them - docs/linkcheck: drop cgit exception as nothing links to it anymore - docs/linkcheck: ignore sourceforge subdomains as well - docs/linkcheck: ignore vulkan.org failures as it also blocks non-browsers - freedreno/ci: disable defunct baremetal jobs - wsi/display: setup the connector earlier - wsi/display: also select a plane when selecting a crtc - ci: fix rustfmt job rules - radv/ci: lower timeouts for newly added gfx1201 jobs - radv/ci: lower timeouts for vkd3d jobs - ci: fix rustfmt job rules (one more case) - radv/ci: sort navi21 flakes - broadcom/ci: sort rpi4 flakes - zink+radv/ci: sort cezanne flakes - radeonsi/ci: document recent flakes - radv/ci: document recent flakes - broadcom/ci: document recent flakes - zink+radv/ci: document recent flakes - lavapipe/ci: document recent flakes - docs: update calendar for 25.2.0-rc2 - ci/lava: fix heredoc-in-yaml syntax - wsi/display: pass the image's DRM modifiers to the kernel - wsi/display: pass the plane's modifiers to the image - docs: update calendar for 25.2.0-rc3 - docs: update calendar for 25.1.7 - docs: add release notes for 25.1.7 - docs: add sha sum for 25.1.7 - ci-tron: set pipefail to show the correct error message when failing to download the install tarball - ci-tron: drop unnecessary \`HWCI_TEST_SCRIPT: deqp-runner.sh` re-defines - ci-tron: cleanup redundancy in artifacts exclude variable - ci-tron: set SCRIPTS_DIR where its path is defined - radv/ci: deduplicate \`DEQP_SUITE: radv-valve` in ci-tron jobs - radv/ci: deduplicate GPU_VERSION in ci-tron jobs - turnip/ci: drop redundant GPU_VERSION - broadcom/ci: drop redundant \`script:` already set by .broadcom-test - broadcom/ci: drop redundant HWCI_TEST_SCRIPT already set by .broadcom-test - anv/ci: drop already included skip list - iris/ci: drop already included skip list - nouveau/ci: drop already included \*-skips.tx - llvmpipe/ci: set DRIVER_NAME to not have to manually add llvmpipe-skips.txt in asan job variant - ci/deqp-runner: fix path to install folder - ci/prepare-artifacts: move git version dump out of static file copy block - ci/prepare-artifacts: drop redundant copy - ci/prepare-artifacts: turn file copies into a loop - meson: fix VkLayer_MESA_device_select in the devenv - meson: include VkLayer_MESA_screenshot in the devenv - meson: include VkLayer_MESA_vram_report_limit in the devenv - meson: include VkLayer_MESA_anti_lag in the devenv - radv/ci: add missing GPU_VERSION for navi10 in kws farm - ci: fix PYTHONPATH variable - turnip/ci: document new vkd3d crash - ci/vkd3d: fix "unexpected results" check - ci: uprev vkd3d to fix some nvk tests - ci: cleanup weston invocations - llvmpipe/ci: use weston's Xwayland instead of broken Xvfb - llvmpipe/ci: document two regressions - llvmpipe/ci: document flakes seen during stress-testing - ci: dedupe weston setup - ci: document image tag to bump for rust build changes - docs/llvmpipe: fix links to defunct drdobbs.com website - docs/linkcheck: ignore crates.io links as it also blocks non-browsers - zink+nvk/ci: fix flakes - ci: drop unnecessary rename of \*.log into \*.log.txt - freedreno/ci: run a618-gl job on xwayland instead of xorg - intel/ci: run iris-{apl,glk,amly}-egl jobs on xwayland instead of xorg - ci: drop xorg + weston workaround now that no user is left - zink+nvk/ci: sort ad106 fails - zink+nvk/ci: give piglit tests a display to use - ci-tron: keep \*.qpa in job artifacts - ci-tron: move vkcts shader cache out of $CI_PROJECT_DIR - ci-tron: move vkd3d shader cache out of $CI_PROJECT_DIR - ci: mark igalia farm as offline - broadcom/ci: skip two more slow CL tests - radv/ci: mark all of dEQP-VK.ray_tracing_pipeline.pipeline_library.configurations.* as flaky - radeonsi/ci: document recent flakes - radv/ci: document recent flakes - broadcom/ci: document recent flakes - zink+radv/ci: document recent flakes - lavapipe/ci: document recent flakes - docs: update calendar for 25.2.0 - docs: add release notes for 25.2.0 - docs: add sha sum for 25.2.0 - docs: add 25.2.x release dates - Revert "ci: mark igalia farm as offline" - radeonsi/ci: document fixes test - r300/ci: document fixes tests and one regression in c64c6a0c...bf8ebb6a - turnip/ci: document regression in 0a12ff6f...8fe0a347 - broadcom/ci: fix another slow & flaky CL test on rpi4 - radeonsi/ci: document recent flakes - radv/ci: document recent flakes - zink+radv/ci: document recent flakes - llvmpipe/ci: document fixed test - llvmpipe/ci: document recent flakes - lavapipe/ci: document recent flakes - ci: track changes to new src/x11/ folder - ci: uprev vkd3d - ci/init-stage2: drop no-op "copy python path into python path" - ci: move setting python path for structured_logger.py to where it's actually used - docs: update calendar for 25.1.8 - docs: add release notes for 25.1.8 - docs: add sha sum for 25.1.8 - freedreno/ci: consistently use x11- prefix for deqp-egl-x11 - iris/ci: consistently use x11- prefix for deqp-egl-x11 - llvmpipe/ci: consistently use x11- prefix for deqp-egl-x11 - softpipe/ci: document fixed tests - ci: set DRIVER_NAME in jobs that are implicitly inheriting skip lists - ci/deqp-runner: drop implicit skips of \`GALLIUM_DRIVER` or \`VK_DRIVER` - ci/deqp-runner: simplify handling the various \*-skips.txt files - ci/deqp-runner: add support for all the prefixes for \*-flakes.txt files - ci/deqp-runner: remove duplicate values to avoiding read the same file multiple times - ci/deqp-runner: add support for all the prefixes for \*-fails.txt files - lavapipe/ci: drop asan fails that are already tracked as normal fails - softpipe/ci: drop asan fails that are already tracked as normal fails - zink+radv/ci: set DRIVER_NAME=zink-radv to allow using common expectation files - zink+radv/ci: deduplicate zink-radv-\*-skips.txt lists - zink+radv/ci: deduplicate zink-radv-\*-fails.txt files - zink+radv/ci: fix typo in skips comment - zink+radv/ci: add common fails for the next commits - zink+radv/ci: give polaris10 piglit tests a display to use - zink+radv/ci: give navi10 piglit tests a display to use - zink+radv/ci: give navi31 piglit tests a display to use - zink+radv/ci: give vangogh piglit tests a display to use - zink+radv/ci: give gfx1201 piglit tests a display to use - panfrost/meson: drop invalid C++ arg - zink+turnip/ci: document regression in b22806705c...cac3b4f404 - zink+turnip/ci: document fixed tests - r300/ci: document flake - etnaviv/ci: document some flakes - turnip/ci: document a flake - nvk/ci: document some flakes - meson: add spirv-tools option to disable the optional dependency - docs: stub pipe_format & pipe_video_chroma_format - docs: update calendar for 25.2.1 - docs: add release notes for 25.2.1 - docs: add sha sum for 25.2.1 - meson: fixup b_sanitize checks - ci-tron: drop meaningless timestamp in initial section message - virgl/ci: drop invalid but overridden empty caching proxy - vmware/ci: fix caching proxy url - ci/piglit: automatically use LAVA proxy - ci/piglit: automatically use baremetal proxy - broadcom/ci: drop unnecessary variables redefinitions - ci-tron: move s3_jwt token file to the project dir - ci-tron: avoid uploading downloaded traces - piglit/ci: configure ci-tron to download traces and upload renders - broadcom/ci: add ci-tron variant of the piglit traces job - docs/ci: drop redundant/dead fork rule - docs/ci: drop unnecessary comment - docs/ci: always build the docs - docs: update calendar for 25.1.9 - docs: add release notes for 25.1.9 - docs: add sha sum for 25.1.9 - ci: document what scope the ci_run_n_monitor token needs - zink+radv/ci: add traces job on vangogh - zink+radv/ci: add traces job on gfx1201 - broadcom/ci: document recent flakes - radeonsi/ci: document recent flakes - radv/ci: document recent flakes - zink+radv/ci: document recent flakes - zink+lavapipe/ci: document recent flakes - docs: update calendar for 25.2.2 - docs: add release notes for 25.2.2 - docs: add sha sum for 25.2.2 - bin/ci: let filter_dag() caller define job filter once (instead of 3 times) - ci/gitlab_gql: keep track of job tags - ci_run_n_monitor: add --job-tags filter - radv/ci: deduplicate navi10 GPU_VERSION - radv/ci: document whether ci-tron jobs runs on an APU or a dGPU - etnaviv/ci: document fixed tests - r300/ci: document fixed tests - nvk/ci: document fixed tests - zink+nvk/ci: document fixed tests - zink+turnip/ci: document fixed tests - venus/ci: document fixed tests - zink+radv/ci: comment out the two checksums - ci/update_traces_checksum: fix decoding of log lines - ci/update_traces_checksum: fix regex detecting PIGLIT_REPLAY_DEVICE_NAME in job logs - intel/perf: fix enum type for eu stall props - zink+radv/ci: sort vangogh flakes - zink+radv/ci: document recent flakes - radv/ci: document recent flakes - broadcom/ci: document recent flakes - zink+lvp/ci: document recent flakes - broadcom/ci: update test expectations - etnaviv/ci: update test expectations - turnip/ci: update test expectations - zink+turnip/ci: update test expectations - zink+nvk/ci: update test expectations - doc/features.txt: add missing supported anv extensions - doc/features.txt: add missing supported tu extensions - doc/features.txt: add missing supported lvp extensions - doc/features.txt: add missing supported v3dv extensions - doc/features.txt: add missing supported nvk extensions - docs/release-calendar: add 25.2.x dates, and 25.3 branchpoint and release candidates - docs: update calendar for 25.2.3 - docs: add release notes for 25.2.3 - docs: add sha sum for 25.2.3 - doc/features.txt: add missing supported dzn extensions - radv: make sure fp16 is enabled consistently on gfx8 - radv: add comment explaining why fp16 is disabled by default on gfx8 - meson: require glslang >= 12.2 for bvh preample - meson: only require glslang >= 12.2 when anv/radv/turnip are built - ci/fedora: manage rust version ourselves - ci/alpine: install and manage rust version ourselves - ci/rust: install components with the initial install command - ci: use MSRV for build-for-tests jobs and recent version in build-only jobs and CI components - ci/build-rust: strip rust libs and binaries - zink+nvk/ci: fix test expectations - zink/ci: drop gbm override now that debian has a usable xorg - util/meson: make sure shader_stats.h is generated in time for anything that depends on mesautil - egl/meson: generate wayland presentation-time header before it gets included - panvk/meson: generate git_sha1.h before compiling panvk_vX_physical_device.c - gfxstream/meson: generate git_sha1.h before compiling ResourceTracker.cpp - intel/meson: generate spirv_info.h before compiling brw_spirv.c - etnaviv/meson: generate enums.h before compiling assembler.c - freedreno/meson: generate xml headers before compiling gmemtool - i915/meson: generate intel_device_info_gen.h before compiling i915_drm_winsys.c - meson: use vcs_tag() instead of custom script - llvmpipe/ci: document fixed tests - docs: update calendar for 25.2.4 - docs: add release notes for 25.2.4 - docs: add sha sum for 25.2.4 - iris/meson: generate git_sha1.h before compiling iris_program.c - docs: finish converting the docs job into a meson build job - ci/alpine: install the real \`ninja` package - ci: check for missing meson dependencies - Revert "meson: use vcs_tag() instead of custom script" - ci-tron: bump job template commit to get cached job templates - docs: update khronos wiki url - nvk/ci: document some flakes - nvk/ci: document fixed tests - broadcom/ci: document fixed tests - docs: update calendar for 25.2.5 - docs: add release notes for 25.2.5 - docs: add sha sum for 25.2.5 - asahi/virtio: fix memleak - util/meson: don't build libmesa_util_clflushopt unless needed - util/meson: don't build libmesa_util_clflush unless needed - ci: track src/c11/ changes - ci: track src/android_stub/ changes Eric R. Smith (9): - panvk: use minimum attachment size for frame buffer size - panvk: fix a NULL pointer dereference in occlusion queries - mesa: fix off by one in MSRTT handling - panfrost: add some sanity checks for nr_samples - panvk: revised occlusion query pointer fix - panfrost: fix typo in register allocation - panfrost: fix debug print of spilled registers - panfrost: align spills to reduce TLS memory usage - glcpp: prevent accidental token pasting Erico Nunes (10): - lima: fix array limit in texture mipmap descriptor - lima: ppir: fix check for discard_block in optimization - lima: ppir: fix store_output optimization for modifiers - ci: lima farm maintenance - Revert "ci: lima farm maintenance" - kmsro: enable with zink - pvr: add VK_EXT_physical_device_drm support - v3dv: rename primary_fd to display_fd - v3dv: use v3d primary node for VK_EXT_physical_device_drm - pvr: enable KHR_wayland_surface Erik Faye-Lund (89): - panfrost: enable robust_buffer_access_behavior - docs: document new panfrost extensions - docs: add GL_KHR_robustness to panfrost - r300/ci: update expected failures - mesa/st: do not check single-sampled for max_samples - Revert "lima: make fp16 render-targets opt-in with driconf" - Revert "upanfrost: make 128-bit opt-in with driconf on v4" - panfrost: add new skips - panvk/ci: try to remove all previously slow tests - pan/ci: remove non-existent flag from PAN_MESA_DEBUG - docs/features: add missing panvk extension - panvk: fix EXT_texture_compression_astc_support - crocus: use os_get_total_physical_memory instead of open-coding - iris: use os_get_total_physical_memory instead of open-coding - panfrost: use os_get_page_size() - winsys/radeon: use os_get_page_size and error-check - winsys/radeon: use util_get_cpu_caps()-helper - prefer _SC_PAGESIZE over _SC_PAGE_SIZE - meson/util: properly detect sysconf - nvk: drop some needless definitions and deps - docs/features: sort drivers - docs/panfrost: update exposed vulkan version - pan/util: use nir_component_mask instead of BITFIELD_MASK - pan: use translate_s_format for stencil - pan/lib: do not duplicate enum mali_pixel_kill - panvk: avoid implicit cast-warning on Clang - pan/midgard: avoid implicit cast-warning on Clang - pan/bi: plug leak - pan/bi: bail from optimizing on oom - pan/bi: use ralloc - pan/midgard: r1w should be set - pan/midgard: initialize last_next_tag to TAG_BREAK - pan/decode: detect error on fseek - pan/clc: handle seek-error - pan/bi: use os_read_file-helper - pan/midgard: fix check for negative texture offset - pan/va: check branch_offset for overflow - panvk: properly handle errors from utrace_context_init - pan/lib: clamp format size to 4 - pan/lib: clean up tilebuffer size helpers - panvk: enable KHR_maintenance7 - doc/features: update VK_KHR_maintenance8 - panvk: enable KHR_maintenance8 - panvk: respect VK_QUERY_POOL_CREATE_RESET_BIT_KHR-flag - panvk: enable KHR_maintenance9 - panvk: fix up vk1.4 properties - panvk: clean up feature-bits - panvk: clean up limits and properties - panvk: explicitly list unsupported features - panvk: expose missed vulkan 1.4 properties - zink: update profile schema - zink: add missing gpl requirement - zink: use polygonModePointSize instead of open-coding - aux/pp: fixup sampler-view release - pan/lib: set afbc mode based on plane-format, not view - panfrost: add per-gpu GLES2 extension lists - panvk: do not export needless symbols - pvr: use vulkan_icd_link_args - pvr: report vulkan 1.4 to the loader - pvr: wire up version-overriding - pvr: remove unused enum - pvr: drop pointless PVR_FROM_HANDLE macro - pvr: move event/sampler cast defs to correct header - pvr: remove bogus forward-declaration - pvr: include pvr_common.h instead of pvr_private.h - pvr: use pvr_memlayout instead of uint32_t - pvr: remove stale comment about pvr_pds_upload - pvr: move pvr_pds_upload to pvr_common.h - pvr: break out queue to separate header - pvr: break out instance/device to separate header - pvr: break out image to separate header - pvr: break out buffer to separate header - pvr: break out render-pass to separate headers - pvr: break out cmd-buffer to separate header - pvr: break out queries to separate header - pvr: break out pipelines to separate header - pvr: break out descriptor sets to separate header - pvr: break out wsi to separate header - pvr: break out macros to separate header - pvr: avoid including pvr_private.h from headers - pvr: kill off pvr_private.h - pvr: include pvr_csb.h first in implementation - pvr: kill rogue_hwdefs.h - pvr: split out rogue hw-defs to separate folder - v3dv: use ld_args_build_id - docs/pvr: update conformance status - docs/pvr: update vulkan version - aux/pp: release correct sampler-views - gallium/aux: unconditionally write buffer Ernst Persson (3): - meson: Raise minimum Python version to 3.9 - vulkan/util: Use str.removeprefix() from Python 3.9 - amd/vulkan: Use str.removesuffix() from Python 3.9 Fafa Kitten (1): - meson: detect \`memfd_create()` and \`getrandom()` from headers, not system libraries Faith Ekstrand (205): - nak: Wire up the mma predicate on Hopper+ - nir/instr_set: Rework tex instr hash/compare - nil: Add a ViewAccess enum and plumb it through from NVK - nil: Use an extent in samples for MSAA storage images - nir,nak: Add a nir_texop_sample_pos_nv and plumb it through - nak/lower_tex: Don't use remap_sampler_dim() for images - nak/lower_tex: Add texture query helpers - nak/lower_tex: Handle NULL image queries pre-Volta - nvk: Drop the pre-Volta texture query workaround - nak: Lower MSAA image load/store/atomic/size - nvk: Delete the old MSAA image workarounds and trust NIL and NAK - nouveau/headers: Skip duplicate enumerants in rust enums and switches - nouveau,nvk: Import the Blackwell and Hopper DMA class headers - nvk: Move KHR_timeline_semaphore to the right spot in the list - nvk: Bump the conformance version to 1.4.3 - nvk: Add an nvk_is_conformant() helper - vulkan/meta: Supply image view usage in vk_meta_clear_*_image() - loader: Ignore NOUVEAU_USE_ZINK on Hopper+ - vulkan: Rename a bunch of vk_sync_timeline helpers - vulkan: Hold a reference to pending vk_sync_timeline_points - nak/lower_tex: Re-order arguments to put can_speculate at the end - vulkan/wsi/x11: Handle VK_NOT_READY in AcquireNextImage() - spirv: Assert !ptr_as_array for blocks and acceleration structures - spirv: Drop block_index/offset pointers - spirv: Simplify pointer_to/from_ssa a bit - spirv: Assert that vtn_pointer_to_deref() doesn't return NULL - compiler/rust: Add a CFG::loop_depth() method - nak: Take loops into account in static cycle estimates - nvk: Blackwell is now Vulkan 1.4 conformant - nvk: Handle empty pushes in nvk_queue_push() - nouveau/class_parser: Strip unnecessary parens - nouveau/headers: Import video encode/decode headers from NVIDIA - nouveau/push: Map b0 classes to subchannel 4 - nouveau/winsys: Allow subchan_dealloc() on zeroed subchans - nouveau/winsys: Refactor nouveau_ws_context_create() - nvk: Advertise KHR_shader_untyped_pointers - vulkan/video: Switch vk_video_session_parameters to create/destroy - vulkan: Add handle casts for vk_video_session[_parameters] - vulkan: Add common VideoSessionParametersKHR entrypoints - anv: Delete anv_video_session_params - radv: Delete radv_video_session_params - vulkan: Add a vk_video_session_finish() helper - nvk: Allow kepler in nvk_is_conformant() - anv: Set the Shader capability when compiling the FP64 shader - anv/i915: Require HAS_EXEC_ASYNC - anv/i915: Require HAS_EXEC_CAPTURE - anv/i915: Require HAS_EXEC_TIMELINE_FENCES - intel/gem: Add an intel_gem_supports_dma_buf_sync_file() helper - anv: Require Linux 6.0 for dma-buf sync file import/export - anv/wsi: Stop requesting signal_*_with_memory - anv: Dead code anv_bo_sync - hasvk: Require HAS_EXEC_ASYNC - hasvk: Require HAS_EXEC_CAPTURE - hasvk: Require HAS_EXEC_TIMELINE_FENCES - hasvk: Require Linux 6.0 for dma-buf sync file import/export - hasvk/wsi: Stop requesting signal_*_with_memory - hasvk: Dead code anv_bo_sync - dozen: Drop dzn_create_sync_for_memory() - vulkan/wsi: Drop signal_fence/semaphore_with_memory - vulkan/wsi: Stop setting wsi_memory_signal_submit_info - vulkan: Drop implicit sync support - vulkan/wsi: Style nits - vulkan/wsi: Sanitize the result of wsi_drm_check_dma_buf_sync_file_import_export() - vulkan/wsi: Only test for dma-buf sync file support once - subprojects: Stop calling add_languages() in paste-1-rs/meson.build - meson: Add a rust_2024_lint_args helper - meson: Disable unsafe_op_in_unsafe_fn in bindgen for now - meson: Disable unsafe_attr_outside_unsafe for now - nil/copy: Wrap all unsafe code in unsafe blocks - nil/copy: Use saturating_sub() instead of doing it manually - nil: Fix a couple of clippy lints - nak: Use .as_ref().unwrap() instead ofv &* - nak/hw_runner: Wrap all unsafe code in unsafe blocks - nak: Use +use<> to avoid unnecessary lifetime captures - nouveau: Use rust_2024_lint_args - nouveau/class_parser: Stop shifting by zero - nouveau/class_parser: Add a helper for address expression filtering - nouveau/struct_parser: Stop generationg i * 1 - nouveau/bitview: Drop an unneeded lifetime - compiler/rust: Use .as_ref().unwrap() instead of &* - compiler/rust: Stop using NonNull in the NIR bindings - meson: Add --wrap-unsafe-ops to bindgen - compiler/rust: Add Rust 2024 lints - compiler/rust/nir: Drop a bunch of explicit lifetimes - compiler/rust: Don't use assert_eq!() with booleans - compiler/rust: Add a bunch of clippy lints - compiler/rust: Stop using try_into() for u8 -> usize - compiler/rust/bitset: Don't use a vector for expected sets in tests - compiler/rust/cfg: Use slices instead of &Vec - vulkan/sync: Return early in vk_sync_timeline_wait() if wait_value == 0 - vulkan/drm_syncobj: Use SWAP() in vk_drm_syncobj_move() - vulkan/sync: Make the can_wait_many() check faster - vulkan/sync: Add vk_sync_signal/reset_many() - vulkan/drm_syncobj: Implement signal/reset_many - vulkan: Add a vk_sync_wait_unwrap() helper - vulkan/queue: Move timeline point allocation to vk_queue_submit_final() - vulkan: Add a vk_sync_signal_unwrap() helper - vulkan: Add a vk_device_copy_semaphore_payloads() helper - vulkan/drm_syncobj: Add a vk_drm_syncobj_copy_payloads helper - anv,hasvk: Use vk_drm_syncobj_copy_payloads - nvk: Use vk_drm_syncobj_copy_payloads - panvk: Use vk_drm_syncobj_copy_payloads - anv: Stop picking our own blit queue - vulkan/wsi: Switch to vkQueueSubmit2() - vulkan,anv,hasvk: Drop vk_queue_wait_before_present() - vulkan/wsi: Take a vk_queue in wsi_common_queue_present() - vulkan/wsi: Make get_blit_queue return a struct vk_queue * - vulkan/wsi: Add a QueueSubmit2() wrapper - vulkan/wsi: Gather per-swapchain results in an array in queue_present() - vulkan/wsi: Handle throttling in a separate loop - vulkan/wsi: Consolodate vkQueueSubmit2() calls across swapchains - vulkan/wsi: Skip the vkQueueSubmit() entirely if we aren't blitting - vulkan/wsi: Always use VK_PIPELINE_STAGE_2_TRANSFER_BIT for semaphore ops - nak: Lower away ldcx when NAK_DEBUG=no_ugpr is set - nvk/nvkmd: Stop setting WAIT_FOR_SUBMIT for sync - nvk/nvkmd: Track all memory objects by default - nvk,nvkmd: Move push dumping to NVKMD - nouveau/push: Handle more recent versions of 6F - nak: Add a nak_qmd_size_B() query - nak/hw_runner: Allow for variable sized QMDs - nvk: Allow for larger QMDs - nak/qmd: QMD versions 4.0 and 5.0 are both 384B - nouveau/headers: Add a MAX_BIT for structs - nak: Assert that QMDs are big enough - nak: NAK_MAX_QMD_SIZE_B should be 384 - nak: Increase Imma latencies on Blackwell by 4 - compiler/rust: Fix the DFS loop detection algorithm - lavapipe: Always use dma-buf for external memory when we can - vulkan/wsi: Move a couple of dma-buf sync checks - vulkan/wsi: Don't dma-buf sync import/export on success - nir: Add an option to make lower_phis_to_regs_block() less clever - nak,nir: Use a simpler version of phis_to_regs_block in lower_cf - nil: Delete some useless image alignment code - turnip: Use vk_drm_syncobj_copy_payloads - nouveau/push: Fix SET_OBJECT handling - nvk: Use the image format for depth views - vulkan/meta: Always set VK_IMAGE_VIEW_CREATE_DRIVER_INTERNAL_BIT_MESA - vulkan: Handle VK_IMAGE_VIEW_CREATE_DRIVER_INTERNAL_BIT_MESA automatically - nvk: Use VK_IMAGE_VIEW_CREATE_DRIVER_INTERNAL_BIT_MESA - radv: Use VK_IMAGE_VIEW_CREATE_DRIVER_INTERNAL_BIT_MESA - v3dv: Use VK_IMAGE_VIEW_CREATE_DRIVER_INTERNAL_BIT_MESA - vulkan: Drop the driver_internal from vk_image_view_init/create() - nvk: Stop adding Vulkan image usage flags - nvk: Use Vulkan formats for SET_ZT_FORMAT instead of NIL - mesa: Use mesa_log_if_debug() for no context errors - util/log: Add a MESA_LOG_LEVEL environment variable - vulkan/wsi/x11: Use mesa_logX() instead of fprintf() - vulkan/queue: Move shared binary semaphores to temps - spirv: Add support for OpBitcast in OpSpecConstantOp - nvk: Actually reserve 1/2 for FALCON - compiler/rust: Add a DepthFirstSearch trait - compiler/rust/cfg: Use DepthFirstSearch for rev_post_order_sort() - compiler/rust/cfg: Use DepthFirstSearch for calc_dominance() - compiler/rust/cfg: Use DepthFirstSearch for find_back_edges() - compiler/rust/cfg: Use DepthFirstSearch for finding reaches sets - compiler/rust: Implement dfs() non-recursively - nil: Add a GOB_TYPE_MODIFIER_INFOS table - nil: Add GOBType::TegraColor - util/cache_ops: Add some cache flush helpers - util/cache_ops/x86: Call util_get_cpu_caps() less - hasvk: Switch to util/cache_ops.h - anv: Switch to util/cache_ops.h - intel/sanitize-gpu: Use util_flush_inval_range() - crocus: Use util_flush_inval_range() - intel: Drop intel_mem.c/h - turnip: Use the util cache helpers - nouveau/winsys: Add a NOUVEAU_WS_BO_COHERENT flag - nvk/nvkmd: Add an NVKMD_MEM_COHERENT flag - nvk/nvkmd: Add map sync to/from GPU helpers - nvk: Implement Flush/InvalidateMappedMemoryRanges() - nvk: Flush pushbufs in EndCommandBuffer() - nvk/nvkmd: Invalidate maps before dumping pushbufs - nvk: Use a coherent map for the event heap - nvk: Flush descriptor tables and heap maps on submit - nvk/mem_stream: Flush maps in nvk_mem_stream_flush() - nvk: Flush after zeroing memory - nvk: Flush the zero page - nvk: Flush/invalidate around host image copies - nvk: Use _B suffixes in descriptor sets - nvk: Use a pool offset instead of an address in nvk_descriptor_set - nvk: Add an nvk_descriptor_writer - nvk: Route more descriptor types through write_desc() - nvk: Flush descriptor set maps - nvk: Flush indirect execution set maps - nvk/query: Rework offset helpers - nvk/query: Pass an IS_TIMESTAMP flag explicitly to the CL kernel - nvk/query: Add a vk_query_pool_report_count() helper - nvk/query: Add an interleaved query layout - nvk/query: Rework query waits - nvk/query: Handle non-coherent query pool memory - nvk: Expose cached and coherent as separate types on Tegra - panvk: Fix integer dot product properties - util: Don't advertise cache ops on x86 without SSE2 - util: Build util/cache_ops_x86.c with -msse2 - nvk: Include the chipset in the pipeline/binary cache UUID - nvk: Disable sampleLocationsSampleCounts for 1x MSAA - nvk: Emit inactive vertex attributes - nvk: Look at the right pointer in GetDescriptorInfo for SSBOs - nvk: Capture/replay buffer addresses for EDB capture/replay - panvk/shader: [de]serialize desc_info.max_varying_loads - panvk/shader: Use the right copy size for deserializing dynamic UBOs/SSBOs - nvk: Don't re-initialize the descriptor writer if the set matches - drm-uapi: Import the new NVIDIA modifiers - nil: Add support for Blackwell 8 and 16-bit modifiers - nir: Add a couple panfrost sysvals to divergence analysis Francisco Jerez (16): - intel/brw/xe3+: Handle SENDG in instruction scheduler. - intel/brw: Fix behavior of scheduler around flag register writes. - intel/brw/xe3+: Define BRW_SCHEDULE_PRE_LATENCY scheduling mode. - util/ra: Allow driver to override class P value. - intel/brw/xe3+: Override P value of GRF register classes to increase thread parallelism. - intel/brw/xe3+: Model trade-off between parallelism and GRF use in performance analysis. - intel/brw/xehp+: Adjust performance model weights of LSC atomic ops. - intel/brw/xe3+: Adjust weights of discard control flow for non-EU-fused platforms. - intel/brw/xe3+: Tweak render target write timings in performance modeling pass. - intel/brw: Allow using performance analysis pass pre-register allocation. - intel/brw: Make sure we don't use stale analysis after inst. order restore in brw_allocate_registers(). - intel/brw/xe3+: Select scheduler heuristic with best trade-off between register pressure and latency. - intel/brw: Apply 7e1362e9c070ad037 to pre-xe3 codepath of brw_compile_fs(). - intel/brw/xe3+: Re-enable static analysis-based SIMD32 FS heuristic for the moment. - intel/brw: Fix regression in brw_allocate_registers() compiling large shaders with throughput==0. - intel/brw/gfx12.0+: Sync on all pending send messages after halt target. Frank Binns (30): - pvr: correctly return core count for pvrsrvkm - pvr: update conformance version - pvr: only share scratch buffers when they're the required size - pvr: apply PBE stride alignment when setting up image physical extents - pvr: implement VK_(EXT|KHR)_vertex_attribute_divisor - pvr: advertise VK_EXT_queue_family_foreign - pvr: implement VK_EXT_depth_clip_enable - pvr: Implement VK_KHR_descriptor_update_template - pvr: add support for VK_FORMAT_D32_SFLOAT_S8_UINT - pvr: setup tpu_tag_cdm_ctrl when present (pvrsrvkm) - pvr: support VK_FORMAT_R8G8_SSCALED for vertex attribs - pvr: add some more pixel formats needed by Zink - pvr: implement KHR_shader_float_controls - pvr: disable gs_rta_support for BXS-4-64 to workaround some conformance failures - pvr: enable KHR_create_renderpass2 - pvr: advertise KHR_shader_subgroup_extended_types - pvr: advertise KHR_spirv_1_4 - pvr: setup Vulkan 1.1 & 1.2 features, properties, version - docs: add pvr VK 1.0, extensions and optional features to new_features.txt - pvr: advertise VK_EXT_zero_initialize_device_memory - docs/features: claim vk 1.2 for pvr - pvr: add device info for BXE-4-32 (36.50.54.182) - pvr: add device info for GX6250 (4.45.2.58) - pvr: add device info for G6110 (5.9.1.46) - pvr: add device info for GX6650 (4.46.6.62) - pvr: add device info for BXM-4-64 (36.52.104.182) - pvr: add device info for BXE-2-32 (36.29.52.182) - pvr: add device info for GE8300 (22.102.54.38) - pvr: add device info for GE8300 (22.68.54.30) - pvr: support VK_KHR_device_group GKraats (1): - crocus: fix SIGSEGV crash at pbo compressed teximage Georg Lehmann (175): - ac/nir/lower_mem_access_bit_sizes: make 8/16bit access 32bit if possible - nir/lower_int64: lower 64bit bitfield_select - aco/isel: don't create literal operands for SALU bitfield_select - aco: supported 64bit or vectorized bitfield_select - ac/nir: don't lower 8/16bit bitfield_select - nir/opt_generate_bfi: create vector and non 32bit bitfield_select - nir/opt_algebraic: create non 32bit bitfield_select - radv: vectorize 8/16bit bitfield_select - lavapipe: use NIR_PASS(_, ...) instead of NIR_PASS_V - gallium/draw: use NIR_PASS(_, ...) instead of NIR_PASS_V - gallivm: use NIR_PASS(_, ...) instead of NIR_PASS_V - nir/schedule: return progress and fix metadata - broadcom/compiler: use NIR_PASS for nir_schedule - llvmpipe: use NIR_PASS(_, ...) for nir_lower_fragcolor - svga: use NIR_PASS(_, ...) for gl_nir_lower_images - nir/opt_remove_phis: skip unreachable phis - pvr/rogue: return progress in rogue_nir_pfo - pvr/rogue: replace NIR_PASS_V with NIR_PASS(_, ...) - lima: rework lima_nir_duplicate_modifiers - lima: rework lima_nir_duplicate_intrinsic - lima: rework lima_nir_duplicate_load_consts - lima: fix metadata in lima_nir_split_loads - lima: replace NIR_PASS_V with NIR_PASS(_, ...) - aco: optimize get_alu_src with constant source and size > 1 - nir: remove NIR_PASS_V - aco/statistics: add latency to WMMA - aco/statistics: update GFX12 WMMA cost - aco: insert VALU s_delay_alu for WMMA - aco/select_alu: avoid vector get_alu_src for instructions with scalar operands - aco/isel: refactor shared vgpr usage - aco/gfx10: optimize subgroupRotate(x, 32) and subgroupShuffleXor(x, 32) - nir/search: support swizzles on expressions in replacement patterns - radv/nir/lower_cmat: load gfx11 8bit ACC using the B layout to get aligned loads - nir/opt_algebraic: remove 8bit roundtrip when vectorizing i2i16(unpack_4x8(a).zw) - aco/print_asm: use real true16 instr on gfx11+ - aco/ra: convert bitwise instruction to gfx11+ 16bit on demand - nir/opt_algebraic: optimize fsat(fmax(a, b)) where b is not positive - nir/opt_algebraic: push fsat into bcsel with constant - nir/opt_algebraic: use range analysis to detect no-op fmin/fmax - nir/range_analysis: look through f2f - nir/range_analysis: look through vec2 - nir/opt_algebraic: make fmin/fmax(a, #b) 16bit if only used by f2f16 - nir/opt_algebraic: remove fneg around fmin/fmax - nir/opt_algebraic: create 16bit fmin/fmax if only used by pack_half_2x16_rtz_split - nir/opt_algebraic: optimize pack_half_rtz of bcsel with constant - nir/opt_algebraic: optimize pack_half_rtz of b2f - nir/opt_tex_skip_helpers: don't skip helpers for terminate_if source - nir/opt_tex_skip_helpers: never require helpers for stores/atomics - nir: print skip_helpers for tex instrs - nir: rename to nir_opt_load_skip_helpers and add options struct - nir: add ACCESS_SKIP_HELPERS - nir: add access for scratch loads - nir/opt_load_skip_helpers: optionally handle intrinsics - aco/insert_exec: remove p_jump_to_epilog from needs exact - aco: add a post-RA pass to disable wqm - aco/insert_exec: new way to handle instructions that need wqm disabled - aco: use new disable_wqm for mubuf/mtbuf - aco: use new disable_wqm for flatlike - aco: use new disable_wqm for mimg - aco/builder: support new disable_wqm - aco: use new disable_wqm for exp - aco: use new disable_wqm for p_dual_src_export_gfx11 - aco/insert_exec: remove per instruction wqm/exact exec handling - aco: use a smaller wqm section for strict_wqm sampling - aco: don't restrict vmem load scheduling by inserting p_end_wqm early - aco: disable wqm for tex loads when not needed - aco: disable wqm for sampled buffer loads when not needed - aco/disable_wqm: optimize local mask creation - amd: replace ACCESS_TYPE_SMEM with ACCESS_SMEM_AMD - amd: stop using custom gl_access_qualifier for access type - amd/ci: update checksums for restricted traces - nir/uub: guard against division by 0 - aco/isel: fix vectorized i2i16 with 8bit vec8 source - nir/uub: fix exclusive scans - nir/uub: decrease default max subgroup size to 128 - nir/uub: handle more reduction ops - nir/uub: handle bit_count - nir/shrink_vec_array_vars: allow nir_var_mem_shared - radv: shrink shared arrays - nir/shrink_vec_array_vars: use range analysis for non constant indices - aco: fix ra validation for flat/global/scratch/ds load sbyte_d16 - aco/optimizer: don't apply packed clamp to v_fma_mix - aco/optimizer: don't create undef copies from p_create_vector - nir: constant fold txd with 0 ddx/ddy to txl - nir/shrink_vec_array_vars: update constant initializer after shrinking - nir/shrink_vec_array_vars: detect zero init shared memory using constant initializer - radv/nir/lower_cmat: split up larger nested switches - radv: reorder cmat properties according to performance - ac/nir: do not assume mesh cull flag is 1bit - nir/lower_io: fix boolean output stores - nir/peephole_select: allows more lowered io - nir/opt_algebraic: optimize some post peephole select patterns - radv: set ACCESS_CAN_SPECULATE for smem buffer loads with known good descriptors - aco/isel: add init_disable_wqm helper - aco: implement skip_helpers for image loads - aco: implement skip_helpers for load_ssbo/ubo/constant - aco: implement skip_helpers for load_scratch - aco: implement skip_helpers for load_global_amd - aco: never end wqm early for vmem - nir: make inverse_ballot 1bit only - nir/builder: add nir_inverse_ballot_imm - nir: make ballot_bitfield_extract 1bit only - spirv: handle ballot bit_extract separately - nir: make ballot find_lsb/msb/bit_count 32bit only - spirv: ensure ballot find_lsb/find_msb/bit_count have 32bit result - nir/lower_subgroups: don't use get_max_subgroup_size for lowering boolean rotates - nir/lower_subgroups: change filter to intrinsic callback - nir/lower_subgroups: recursively lower ballot scans - mesa: clamp fog scale to -FLT_MAX instead of FLT_MIN - intel/ci: update restricted trace checksums - radv/nir/lower_cmat: add shuffle_xor_imm helper - radv/nir/lower_cmat: clean up gfx12 transpose - radv/nir/lower_cmat: clean up GFX11 ACC->B convert - nir/lower_subgroup: optimize reduce/scans with unknown subgroup size - mesa/st: make double subgroup lowering more precise - nir: remove subgroup size related nir_shader_compiler_options members - nir/lower_subgroups: remove lower_fp64 option - nir: remove unused shader_info param in nir_create_shader - nir: define new subgroup size info - vulkan: set nir subgroup size shader info - mesa,glsl,spirv: set new subgroup size info - intel: switch to new subgroup size info - radeonsi: switch to new subgroup size info - rusticl: switch to new subgroup size info - microsoft: switch to new subgroup size info - shader_info: remove gl_subgroup_size enum - radv: add varying subgroup size to shader stage key - ac/llvm: remove unused ballot size - radv: remove unused ballot_bit_size from shader info - ac/nir: set subgroup size for gs copy shader - radv: determine subgroup/wave size early - radv: remove uses_rt from radv_shader_info - nir: remove has_ddx_intrinsics option - aco/isel: fix output args init stack buffer overflow - nir/uub: remove vertex input handling - nir/uub: use shader_info subgroup size - nir/uub: remove max_workgroup_size from config - nir: remove unsigned upper bound config - radv: allow application required fragment shader subgroup size - radv: use rt wave size in fragment shaders with ray queries - radv,aco: don't end monolithic ray tracing with unconditional terminate - aco: remove existing dealloc_vgprs use - aco: dealloc vgprs if there is a pending non scratch store and no pending export - aco: don't insert s_sendmsg dealloc_vgprs with little vgprs allocated - util: add util_round_down_npot - aco: use maximum RT vgpr_limit that doesn't reduce wave count - aco/lower_branches: update branch hints after changing jump targets - radv: call nir_opt_undef late too - nir/opt_undef: prefer 0 over NaN for pack_half_2x16_rtz_split - aco/optimizer: fix incorrect operand order assumption for neg(mul) opt - aco/insert_waitcnt: don't merge waitcnts for LDS clauses - nir: add atomic isub - ac/llvm: support nir_atomic_op_isub - aco/isel: support nir_op_atomic_isub - nir: optimize atomic isub if supported - aco: fix global_atomic_swap offset overflow check - nir: fix nir_get_io_offset_src for global_atomic_swap_amd - aco/gfx10+: only work around split execution of uniform LDS in WGP mode - nir/opt_uniform_atomics: optimize xchg with uniform address and data - nir/opt_intrinsics: don't pass nir options around - nir/opt_intrinsics: optimize atomics to atomic load/store - ac/nir: enable nir atomic load/store opts - aco/tests: allow even more literals - aco/optimizer: add a new dce helper - aco/optimizer: add alu_opt_info helpers - aco/optimizer: use new helpers to apply literals - aco/optimizer: use new helpers to propagate constants/neg/abs - aco/optimizer: rework packed fneg opt - aco/optimizer: apply sgprs/extract with new helpers - aco/optimizer: delete apply_extract - aco/optimizer: remove can_apply_extract - aco/optimizer: apply f2f16 conversion with the new helpers - aco/optimizer: unify constant labels - radv: do not report wave32 in gl_SubgroupSize for Doom Dark Ages - aco/gfx10_3: work around NSA hazard Gert Wollny (95): - r600/sfn: lower bany/ball \*(n)equal in nir - r600/sfn: lower ineg in nir - r600/sfn: remove some dead code - r600/sfn: remove obsolete index and address register handling - r600/sfn: remove code used for vectorized ALU ops - r60/sfn: Update .clang-format - r600/sfn: Move RA helper class declaration into implementation file - r600/sfn: lower b2f64 in nir - r600/sfn: Allow f2f64 to use vec2 - r600/sfn: remove first call to r600_split_64bit_alu_and_phi - r600/sfn: lower u2f64 and i2f64 in nir - r600/sfn: check number of fsat64 source uses properly - r600/sfn: rename free_slots and improve updating it - r600/sfn: Simplify test code when scheduling a vec instr into trans - r600/sfn: unify and fix naming of group readport reserver - r600/sfn: reuse readport for already loaded registers - r600/sfn: Fix update readports method - r600/sfn: update readports before trying to schedule group instrutions - r600: Update GPR count when adding a GDS instruction - r600/sfn: allow skipping RA for shader ID ranges - r600/sfn: factor out adding an input in GS - r600/sfn: Handle indirect access to GS input arrays - r00/sfn: Fix copy propagation into buffer load address - r600/sfn: resolve constant indices into local arrays better - r600/sfn: Lower all GS indirect input loads after lowering IO - r600/sfn: cleanup GS shader emission - r600/sfn: When splitting an ALU CF update possible start of next CF - r600/sfn: Fix AR use tracking off-by-one error - r600/sfn: remove extra slot of AR use - r600/sfn: remove early emmission of ALU last op - r600/sfn: Take allowed dest mask into account in copy-prop - r600/sfn: Only map ssa index to register index if pinning is not free - r600/sfn: Fix test when allocating registers more freely - r600/sfn: Take slot count into account when pinning registers - r600/sfn: Fix the mods when splitting ALU op - r600/sfn: replace hard-coded multislot dot handling - r600/sfn: Handle more ops in desk mask evaluation - r600/sfn: op1v_flt64_to_flt32 as multi-slot instruction - r600/sfn: give more liberty to the channel selection in simple two-slot ops - r600/sfn: Emit thread position as two-slot op - r600/sfn: pass group into AluInstr::split instead of creating it - R600/sfn: split one-dest multi-slot ops late when scheduling - r600/sfn: stop early when looking for ALU vec ready ops - r600/sfn: remove some useless boolean parameters - r600/sfn: add an unreachable if the creation of a fp64 group fails - r600/sfn: rework testing readport config for more than one source - r600/sfn: factor out common code for readport validation - r600/sfn: preloading sources for fp64 ops with common code path - r600/sfn/tests: Update source pinning when loading from string - r600/sfn: Pin registers to channel only after scheduling - r600/sfn: try all possible configurations when splitting multi-slot instructions - r600: remove hack to force a new CF if TEX grad is set - r600/sfn: Increase limit for lowering local arrays to scratch - r600/sfn: remove superfluous semicolon - egl,glx,X11: Handle case when PlatformDisplay is EGL_DEFAULT_DISPLAY - r600/sfn: make pin_dest_to_chan a virtual function - r600/sfn: Simplify scheduling - r600/sfn: preselect fetch by using TC and VC in scheduler - r600/sfn: Prepare scheduler to handle WaitAck instructions - r600/sfn: Emit and schedule WaitACK as a separate instruction - r600/sfn: Add more CF instruction types - r600/sfn: Add a CF block start member and handle it in the tests - r600/sfn: chain group barrier and predicate instructions - r600/sfn: Add method to query whether an ALU group sets the predicate - r600/sfn: Add method to emit ALU_PUSH_BEFORE in assembler - r600/sfn: Drop test for address register in assembler IF predicate - r600/sfn: Add method to query whether ALU block will need ALU_EXTENDED - r600/sfn: extract handling of ALU_PUSH_BEFORE in assembler code - r600/sfn: make sure that kill and update pred are not in the same group - r600/sfn: handle the IF predicate in the scheduler - r600/sfn: start scheduling memory writes earlier - r600/sfn: Don't fall through if a WaitACK was scheduled - r600/sfn: fix op2_pred_sete_64 opcode - r600/sfn: Pass chan and dest_clamp to alu op if no dest register is given - r600/sfn: Add handling of channels for dest-less ALU ops - r600/sfn: don't use dummy regs in alu ops when no dest register is needed - r600/sfn: optimize comparison results - r600/sfn: emit 64 bit predicates like normal ALU ops - r600/sfn: relax restrictions when optimizing predicate evaluation with a register - r600/sfh: Handle 64 bit comparisons in predicate optimization - r600/sfn: Optimize pred(not X != 0) to pred(X == 0) - r600/sfn: Filter lowering of b2f32(comp(x,y)) for 64 bit sources - r600/sfn: Propagate pred and exec update flags when splitting ops - r600/sfn: Add omod to AluInstr and assembler - r600/sfn: Wire up some omod optimizations - nir+r600: add option to avoid contracting fabs into ffma - r600/sfn: replace hand coded comparison opts with opt_algebraic - r600/sfn: clear PIPE_MAP_UNSYNCRONIZED for partial DS texture writes - r600: Fix comparison of strides array when emitting vertex buffers - r600/sfn: extract function to update group after instr insert - r600/sfn: move some common code into try_readport - r600/sfn: Track whether a ALU group has a exec flag update - r600/sfn: make sure kill and update_exec don't happen in one group - r600/sfn: AR loads are not dependend on the future and other code blocks - r600/sfn: Don't start a new ALU-CF if LDS pipeline loads are pending Guilherme Gallo (12): - ci/bare-metal: Fix exit code variable - ci/panfrost: Disable DUTs under maintenance - Revert "ci/panfrost: Disable DUTs under maintenance" - ci: Fix for GitLab 18.2.2 upgrade - ci: Disable vmware farm - ci/radeonsi: Document a new flake - ci/baremetal: Use find_s3_project_artifact on baremetal_build.sh - ci/android: Use find_s3_project_artifact in build script - ci/android: Use curl-with-retry in build scripts - ci/baremetal: Use curl-with-retry in build scripts - ci/zink: Document bypassed failures - ci: Bump image tags to force recreation of s3 artifacts Gurchetan Singh (13): - gfxstream: null-check in vulkan-mapper - gfxstream: vulkan-mapper: special case Nvidia - gfxstream: correct Android API level check - mesa: define peripheral support for src/util/rust - util: rust: make stubs simpler - gfxstream: ANDROID --> VK_USE_PLATFORM_ANDROID_KHR - vulkan: #if DETECT_OS_ANDROID --> #if defined(VK_USE_PLATFORM_ANDROID_KHR) - util: rust: fix some warnings - mesa3d: util: rust: add proper stubs - util: rust: spelling and whitespace fixes - gfxstream: determine page size based on guest properties too - virtio: virtgpu_kumquat: clippy fixes - gfxstream: delete magma-over-gfxstream Hans-Kristian Arntzen (10): - anti-lag: Only consider timestamps from queues which have presented. - anti-lag: Submit timestamps early in a frame. - ac/nir: Avoid 0/0 when computing texel buffer size on Polaris. - nvk: Return 0 for opaque memory capture replay. - nvk: Avoid passing garbage data in descriptor buffers for UBOs. - anti-lag: Fix stype for submit2 semaphores. - anti-lag: Don't force enable every supported feature on device creation. - radv/sqtt: Ensure that present fence gets signalled. - anti-lag: Do not enable layer by default. - radv: Actually fail custom border color sampler creation. Hsieh, Mike (3): - amd/vpelib: add format, colorspace check function - amd/vpelib: bug fix: remove unnecessary free - amd/vpelib: add max/min input output capability Hyunjun Ko (18): - vulkan/video: fix to write a h264 slice header for CAVLC mode - vulkan/video: fix to set ref_pic_list_modification_flag_l1 correctly - anv/video: Fix to set high profile to PPS if high profile provided - anv/video: implement GetPhysicalDeviceVideoEncodeQualityLevelPropertiesKHR - vulkan/video: align with spec correctly for h265 slice header. - anv/video: fix to set some attributes for HCP_PIC_STATE. - anv/genxml: the type of POC delta changes correctly - anv/video: set short term ref list1 even if P frames provided - anv/video: don't set the MVDL1Zero for encoding - anv/video: create Motion Vector buffers for encoding too - anv/video: add VK_VIDEO_ENCODE_H265_CTB_SIZE_32_BIT_KHR for minimum ctb sizes - vulkan/video: fix h265 decoding with LT enabled. - vulkan/video: fix h265 encoding with LT enabled. - vulkan/video: fix misuse of CLAMP in h265 slice parsing. - anv/video: fix to set slice block size correctly for h265 decoding. - anv/video: Make the query result for video profiles and formats more precisely. - anv/video: remove support for VK_IMAGE_TILING_DRM_FORMAT_MODIFIER_EXT - anv/ci: added video tests failures on tgl/jsl Iago Toral Quiroga (2): - nir/serialize: make alu src deserialization consistent for unused swizzles - panfrost: fix swapped stats for varing and position shaders Ian Romanick (40): - brw/reg_allocate: Don't access out of bounds in non-debug builds - brw: Split virtual GRFs again at the end of optimizations - nir/print: Don't segfault checking has_debug_info - brw: Add and use brw_reg_is_arf to test for a specific ARF - brw: Implement Wa_22012725308 for flags via SWSB too - brw: Allow additional flags registers on Xe2+ - brw: Do cmod prop again after brw_lower_subgroup_ops - brw: Don't emit redundant flags initialization for subgroup op lowering - brw: Strategically place flags initialization to help cmod prop - brw: Use nir_opt_sink and more nir_opt_move - elk: Use nir_opt_sink and more nir_opt_move - iris: Limit max_shader_buffer_size to INT32_MAX - brw: Increase the size of some structure fields in combine_constants - elk: Increase the size of some structure fields in combine_constants - brw/nir: nir_intrinsic_load_reloc_const_intel may not be scalar [v3] - elk: Set lower_txd_data to devinfo - nir: Add saturating float to integer conversion opcodes - brw: Enable saturating float to integer conversion opcodes - elk: Enable saturating float to integer conversion opcodes - nir/algebraic: Elide range clamping of f2u sources - nir/algebraic: Remove useless ftrunc inside f2i/f2u - nir/algebraic: Don't introduce undefined behavior in f2u conversion - nir/algebraic: Optimize f2u of negative value to zero - nir/algebraic: Prefer bfi over bitfield_select for bitfield_insert - nir/range_analysis: Handle bfi and bitfield_select in get_alu_uub - brw/disasm: Fix BFN disassembly of src1 and src2 - brw/disasm: Pretty print the BFN equation as an annotation - brw: Basic validation for BFN - brw: BFN does not support source modifiers - brw: Constant propagation and constant combining support for BFN - brw/builder: Add BFN - brw/cmod: Enable limited cmod propagation for BFN - brw: Use BFN to implement nir_opt_bitfield_select - nir/algebraic: Optimize bfi with odd-valued mask to bitfield_select - brw: elk: Fix name of function in comment - brw: Mark src3 of BFN as is_control_source - brw: Don't do non-obvious things with BFN parameter ordering - brw: Apply Gfx9 vgrf127 workaround in more cases - elk: Apply vgrf127 workaround in more cases - brw: Correctly generate conditional modifier for BFN Icenowy Zheng (4): - pvr: fix for GCC - pvr: implement samplerAnisotropy - gallivm: orcjit: put object cache under the protect of lookup_mutex - gallivm: orcjit: remember Context in addition to ThreadSafeContext Igor Naigovzin (1): - zink: fix clamping gl_Layer output to 0 when framebuffer is not layered Iliyan Dinev (3): - pvr: fix pvr_CmdResetQueryPool barriers - pvr: add support for VK_FORMAT_X8_D24_UNORM_PACK32 - pvr: re-emit ppp state update when ds depth bits are set Iván Briano (15): - intel: Re-disable ray tracing on 32 bits - anv: check for pending_db_mode when dirtying descriptor mode - anv: dirty descriptor state on CmdSetDescriptorBufferOffets - anv: fix capture/replay of sparse images with descriptor buffer - anv, hasvk: allow using a 3D image as a resolve target - anv: pass only isl_format to helper functions - anv: drop EXT from host_image_copy stuff - anv: handle multiple aspects in vkCopyImageToImage - anv: drop height_pitch parameter from anv_copy_image_memory - anv: intermediate RGB <-> RGBX copy for HIC - anv: fix FS output <-> attachment map building - anv: use the color_map if present for calculating color_mask - anv: handle compiling of mesh shader separately from task shader - brw/mesh: drop brw_tue_map::per_task_data_start_dw - anv: report maint5::earlyFragment*SampleCounting correctly James Fitzpatrick (2): - pvr: update WClamp value to 1.0e-13f - pvr: add support for (EXT|KHR)_line_rasterization Janne Grunau (1): - hk: Report the correct plane count in VkDrmFormatModifierProperties2?EXT Jarred Davies (3): - pvr: Disable PBE resolve on cores without gs_rta_support - pvr: Reduce number of stencil dependency barriers needed - pvr: Mark barrier load subcmd as not empty Jason Macnak (4): - gfxstream: Add gfxstream TLS connection manager reset - gfxstream: add a vkTraceAsyncGOOGLE - gfxstream: hide vkTraceAsyncGOOGLE behind new capset flag - gfxstream: Address some Werror errors from ag/35389434 Jeffrey Zhuang (1): - zink: remove ALWAYS_INLINE from zink_batch_usage_unflushed_wait Jeongik Cha (1): - gfxstream: Generate goldfish dispatch code for AHB extension Jesse Natalie (19): - gallium/aux: nir_lower_pstipple_fs progress and metadata - microsoft/compiler: Use NIR_PASS instead of NIR_PASS_V - microsoft/clc: Use NIR_PASS instead of NIR_PASS_V - dozen: Use NIR_PASS instead of NIR_PASS_V - d3d12: Use NIR_PASS instead of NIR_PASS_V - winsys/d3d12: Use DComp swapchains to support transparency - nir: Add missing #include for c99_alloca.h - util: Disable inline asm for arm64 for MSVC - d3d12: Stop using util_framebuffer_init - d3d12: Support more logic op formats - d3d12: Move logicop emulation resource from surface to resource - d3d12: Move logicop descriptor initialization to after all blits - d3d12: Flush command queue when destroying or resizing - wgl: Always revalidate framebuffer when front is requested - d3d12: Only use DComp swapchains when alpha is present in the framebuffer - wgl: Fix zink depth buffers - dlist: Flush the context during EndList if it's part of a share group and uploaded during recording - microsoft/compiler: Use lower_mem_access_bit_sizes for scratch/shared - microsoft/compiler: Respect write masks when lowering unaligned loads and stores Jianxun Zhang (7): - anv: No compression on host memory allocation (xe2) - anv: Fix PAT entry in importing (xe2) - iris: Disable compression on sharing without modifier - iris: Ensure type of bo's heap is consistent with modifier - iris: Assert no disabling aux in first query (xe2) - isl: Reuse Xe2 modifers on newer platforms - iris: Enable Xe2 modifiers on all newer platforms Job Noorman (75): - ir3/cp: disable cat3 hw bug workaround on a6xx+ - freedreno: remove ir3_cmdline - ir3/legalize: add asserts to prevent OOB array access - ir3/postsched/legalize: ignore prefetch sam dummy src - ir3: use dummy dst for descriptor prefetches - ir3/shared_ra: don't reuse src of different halfness - tu: add constlen shader stat - ir3/a750: don't allocate const space for primitive_param/map - ir3: treat consts_ubo as normal UBO - tu: remove consts_ubo upload code - freedreno/a7xx: disable consts_ubo upload - tu: disable VK_EXT_post_depth_coverage - tu: enable fragmentShadingRateWithShaderSampleMask - ir3/legalize: prevent infinite loop when inserting (ss)nop - ir3/ra: fix file start wraparound - ir3: add pointer from ir3_shader_variant to ir3_shader - ir3: add shader bisect debug tool - v3d/drm-shim: add support for multisync - nir/opt_uniform_subgroup: use ballot_bit_count - ir3: allow 2 const srcs in scalar cat2 - ir3: align alias sequences to work around hardware bug - ir3: don't add array stores to block keeps - ir3: allow shared srcs for ldc - ir3: use isam for txf with LOD 0 - ir3/array_to_ssa: fix updating/removing phis - ir3/array_to_ssa: remove trivial all-undef phis - ir3: allow shared srcs for ldc.k - ir3: use ir3_get_predicate for demote/kill - ir3: use shared srcs for demote/kill condition - ir3/legalize: don't special-case early-preamble a1 reads - ir3: make backend aware of scalar predicates - ir3/isa: add encoding for scalar predicates - ir3/opt_predicates: move some helpers up - ir3: enable scalar predicates - tu: pass SSBO/UBO min alignment to SPIR-V frontend - nir: add nir_src_is_deref helper - nir: add offset_shift intrinsic index - nir: add some helpers for dealing with offset_shift - nir,ir3: add offset_shift index to SSBO access intrinsics - nir/lower_atomics: add support for offset_shift - nir/lower_io_to_scalar: add support for offset_shift - nir/lower_wrmasks: don't adjust BASE - nir/lower_wrmasks: add support for offset_shift - nir/opt_shrink_vectors: add support for offset_shift - nir/lower_mem_access_bit_sizes: add partial support for offset_shift - nir/opt_load_store_vectorize: allow per-instruction offset scaling - nir/opt_load_store_vectorize: add support for offset_shift - nir/opt_load_store_vectorize: fix wrap check for scaled offsets - nir/lower_explicit_io: make offset calculation reusable - nir/lower_explicit_io: add helper to build address - nir/lower_explicit_io: use nir_io_offset to pass around addresses - nir/lower_explicit_io: add alignment parameters to address builder - nir/lower_explicit_io: add support for offset_shift - ir3: use offset_shift for SSBO intrinsics - ir3: don't vectorize nir_op_sdot_4x8_iadd[_sat] - ir3: emit descriptor prefetch in block dominated by its sources - freedreno/drm-shim: disable VM_BIND - ir3: use shared masks for cov when scalar ALU is supported - freedreno/computerator: fix cs builder conversion errors - nir/opt_offsets: rename max_offset_data to cb_data - nir/opt_offsets: add callback to set need_nuw per intrinsic - ir3/cf: don't swap signedness of (sat) instructions - ir3: use nir_lower_bit_size for 8-bit bit_count - bin/rb: update Alyssa's email address in test case - ir3/spill: initialize base reg as late as possible - ir3/ra: make main shader reg select independent of preamble - ir3: don't create merge sets for subreg moves - ir3/parser: don't use instr as ralloc context - freedreno/computerator: disable disk cache - nir: add nir_shr builder - nir/lower_alu: use Knuth's Algorithm M for [iu]mul_high - nir,ir3: rename umul_low to umul_16x16 - nir: mark fneg distribution through fadd/ffma as nsz - ir3/ra: fix assert during file start reset - spirv: don't set in_bounds for structs John Anthony (4): - nir,agx: unvendor core_id_agx - nir,spirv: Add support for SPV_ARM_core_builtins - pan/va: Add support for SPV_ARM_core_builtins - panvk: Enable VK_ARM_shader_core_builtins Jonathan Marek (1): - wsi/display: use atomic mode setting Jordan Justen (6): - intel/dev: Add WCL platform enum - intel/dev/mesa_defs.json: Add WCL WA entries - intel/dev: Add WCL device info - intel/dev: Add WCL PCI IDs - intel/dev: Add BMG 0xe209 PCI ID - anv: Use image view base-layer in can_fast_clear_color_att() Jose Maria Casanova Crespo (13): - v3dv: Move V3D_TFU_READAHEAD_SIZE to src/broadcom/common - v3d: Add V3D_TFU_READAHEAD padding for allocated resources - v3dv: limit V3D_TFU_READAHEAD to buffers/images with USAGE_TRANSFER_SRC flag - v3d: glMemoryBarriers only flush jobs with tmu_dirty_rcl - v3d: Mark DIRTY_ZSA if disable_ez is changed from FS. - v3d: Reduce CLE submission of CLIP_WINDOW packets - v3d: Add V3D_TFU_READAHEAD padding for renderonly resources - vc4/simulator: pass and return sim_file on vc4_simulator init/destroy - vc4/simulator: avoid free simulator memory on destroy - v3dv: Fix stencil clear values for only stencil clears - v3d: Don't enable Early-z with discards when stencil updates are enabled - v3d: use helpers util_writes_depth/stencil - v3d: mark FRAG_RESULT_COLOR as output_written on SAND blits FS Josh Simmons (2): - util: Fix \`BITSET_EXTRACT` out-of-bounds read - radv: Fix crash in sqtt due to uninitalized value Joshua Ashton (5): - wsi/common: Track VkColorSpaceKHR with wsi swapchain - wsi/display: Implement VK_EXT_hdr_metadata on KHR_display swapchain - wsi/display: Clean up DRM hdr/color state on swapchain destruction - build: Add dependency on libdisplay-info - wsi/display: Expose HDR10 colorspace based on EDID Joshua Simmons (1): - vtn: Fix OpCopyLogical destination type José Roberto de Souza (23): - intel/brw: Nuke unused brw_message_desc_header_present() - intel/brw: Add comment to reg_unit() - intel/brw: Remove duplicated implementation of brw_imm_uq/brw_imm_u64() - gallium/llvmpipe/test: Rename rsqrtf() to _rsqrtf() - intel/decode: Add support to new version of Xe KMD devcoredump with canonical addresses - intel/brw: Use ASR over SHR for SHADER_OPCODE_ISUB_SAT - intel/brw: Move brw_s0() to brw_reg.h - anv/allocator: Move definition of ANV_FREE_LIST_EMPTY to anv_allocator - anv/allocator: Drop uncessary function - anv/allocator: Change some parameters and variables from 32bit to 64bits - anv/allocator: Don't call anv_block_pool_map() with an offset that includes start_offset - anv/allocator: Subtract start_offset in chunk_offset - anv: Add comment to anv_state->offset - anv: Define bt_block only in the block that uses it in anv_cmd_buffer_alloc_binding_table() - anv: Replace duplicated code set shader relocs by a function - anv: Drop shader relocs from anv_shader_bin_create() - anv: Simply anv_shader_set_relocs() parameters - anv: Rename anv_shader_bin to anv_shader_internal - intel/brw: Share mode code in lower_lsc_varying_pull_constant_logical_send() - intel/brw: Add comment to first_non_payload_grf - intel/brw: Fix LSC fence scope and flush type - intel/brw: Call lower_hdc_memory_fence_and_interlock() with brw_send_inst - intel/brw: Store and set sfid in memory fences Juan A. Suarez Romero (20): - broadcom/ci: disable baremetal jobs for ci-tron - v3d/ci: unlock rusticl citron jobs - broadcom: remove obvious comment - drm-uapi: update v3d_drm.h for reset counters - broadcom: check for GPU reset counters support - broadcom/simulator: add support for GPU reset counters - v3d: implement get device reset status - v3d: handle QUNIFORM_GET_UBO_SIZE - v3d: implement robust buffer access - broadcom/ci: disable baremetal rusticl jobs for ci-tron - meson: check for no_sanitize function attributes - util: add DECLARE_LINEAR_ZALLOC with no sanitize - glsl: disable UBSan vptr check for ir_instruction - broadcom/ci: comment some of the failures - broadcom/ci: unlock CI-Tron jobs for arm32 - v3d/ci: update expected results - ci: uprev VKCTS to 1.4.3.3 - glsl: use array element type to validate assignment - vc4/ci: disable asan job - v3d/v3dv/ci: switch to asan rpi5 Julia Zhang (2): - virgl: Small fix of converting format - pps: init driver in OnSetup Julian Orth (2): - ci: build and install native libwayland - kms-swrast: export dmabufs with DRM_RDWR Juston Li (3): - anv/android: refactor anb resolve to fix align assertion - anv: fix uninitialized mutex lock in anv_slab_bo_deinit() - android/gralloc0: add CROS_GRALLOC_DRM_GET_BUFFER_COLOR_INFO K900 (1): - gfxstream: fix build on 32-bit Karmjit Mahil (10): - freedreno/registers: Fix SP_READ_SEL_LOCATION - pvr: fix spm-related renderpass hwr - pvr: Remove shareds_dest_offset from load_op - pvr: Move renderpass load op setup into a separate function - nir: Add more matches for \`fmulz` - nir, ir3: Add \`lower_fmulz_with_abs_min` backend option - freedreno/registers: Fix typo - tu: Add VK_EXT_zero_initialize_device_memory - ci,crnm: Fix f-string print error - freedreno/decode: Add 2d_to_json lua script Karol Herbst (125): - vtn/opencl: set exact on all ffmas and mads - zink: disallow intensity buffer images - zink: disable shader images for intensity formats - rusticl/mem: set swizzle for intensity images - rusticl/mesa: add return status to PipeFence::wait - rusticl/queue: offload waiting on fences to another thread - rusticl/mem: relax flags validation for clGetSupportedImageFormats - rusticl/queue: do not return event status errors on flush/finish - rusticl/kernel: fix clippy lint needless-question-mark - zink: properly unbind sampler views with imported 2D resource - rusticl/mesa: use pipe_sampler_view_reference - rusticl/queue: clear shader images when destroying queues - rusticl/queue: pass a mut reference to QueueContext around - rusticl/queue: commit lifetime crimes - rusticl/queue: remove RefCell - rusticl/kernel: stop clearing sampler views on kernel launches - rusticl/queue: cache samplers - rusticl/kernel: unbind trailing shader images - nak: fix wrong argument order in calls to build_txq_size - nak: optimize load_subgroup_id - nv50: fully migrate away from util_framebuffer_init - nak: use MemScope::CTA for shared memory scoped SCOPE_WORKGROUP barriers - nak: copy late_algebraic iadd3 rules without the constant restriction - rusticl: fix impl_trait_overcaptures lint errors - rusticl: fix unsafe_attr_outside_unsafe lint errors - rusticl: add lints relevant for edition 2024 migration - rusticl: use pipe_sampler_view_release - rusticl/mesa: wire up fence_server - rusticl/gl: store the mesa_glinterop_export_in - st/interup: flushing objects is a no-op when no context is bound - rusticl/gl: only flush objects on import if we get a valid fd - rusticl/gl: flush and wait on gl objects inside clEnqueueAcquireGLObjects - vulkan: use p_atomic_read on vk_descriptor_set_layout::ref_cnt - zink: fix data race in descriptor_util_pool_key_get - rusticl: silence warnings in generated sources - rusticl: silence new warnings from rustc versions above our rustc target - anv: do not map from_host_ptr bos in image_bind_address - zink: set zink_bo is_user_ptr on creation - anv/i915: print bo->map when dumping exec buffers bos - nak: set max_gpr to multiple of 8s - nak: add more helpers for predicates - nak: relayout opt_uniform_instrs - nak: support bra.u with a upred source on Ampere and newer - rusticl/mesa: add ResourceType::Immutable - rusticl/kernel: create shader constants as immutable - rusticl/mem: split out mem_flags validation for creation operations - rusticl/mem: turn bool argument into enum in validate_mem_flags - rusticl: implement cl_ext_immutable_memory_objects - rusticl: fix a bunch of warnings - rusticl/util: add read_and_advance methods for pointers - rusticl/util: use read_and_advance in Properties - rusticl/util: drop uneccesary Arc in event_list_from_cl - rusticl/icd: qualify CLResult inside impl_cl_type_trait_base macro - rusticl/icd: sort extension functions by extension name - rusticl: handle failures when importing fences - rusticl/mesa: port PipeFence to use ThreadSafeCPtr - rusticl: specify FD type when importing fences - nak: run nir_opt_move nir_move_load_ubo - nak: run nir_opt_move nir_move_comparisons - rusticl: add SPDX tags - aux/trace: move fence_server calls outside the locked area - nak: rework scale argument of compute_mat and rename it - nak: protect static cycle counting against overflows - nak: use logarithmic scaling in estimate_block_weight - nak: extract nir_intrinsic_cmat_load lowering into a function - nak/hw_runner: support shared memory - nak/hw_runner: add ldsm tests - nak: use ldsm - rusticl/mesa: rename PipeResource to PipeResourceOwned - rusticl/mesa: add borrow/to_owned semantics to our pipe_resource wrapper - rusticl/kernel: reduce CPU overhead of set_global_binding - rusticl/kernel: move add_pointer into KernelExecBuilder - rusticl/kernel: move add_global into KernelExecBuilder - rusticl/kernel: move add_sysval into KernelExecBuilder - rusticl/kernel: add KernelExecBuilder::add_values - rusticl/kernel: add KernelExecBuilder::add_zero_padding - rusticl/kernel: add KernelExecBuilder::get_resources_and_globals - rusticl/kernel: move workgroup id offset handling into KernelExecBuilder - rusticl/kernel: add KernelExecBuilder::input - rusticl/kernel: allocate the full input buffer at creation time - rusticl/kernel: rework KernelExecBuilder::get_resources_and_globals to reduce allocations - rusticl/device: add DeviceCaps::has_create_fence_fd and use it - docs/gallium: Clarify ordering requiremenets on fence_server_signal and fence_server_sync - rusticl/event: fix create_and_queue for deps in error states - rusticl/util: add MultiValProperties - gallium/noop: add fence_server_signal - gallium: add pipe_screen::semaphore_create - rusticl/mesa: wire up semaphores - zink: factor out fence creation function - zink: implement pipe_screen::semaphore_create - radeonsi: implement pipe_screen::semaphore_create - rusticl: add stubs for semaphores and external_memory - rusticl: implement cl_khr_semaphore - rusticl: implement cl_khr_external_semaphore - util: move typed_memcpy into macros.h - nvk: prepare for higher shared memory sizes - nouveau/winsys: add shared memory size tables - nak/qmd: base shared mem size allocation on hardware limits - nvk: use hardware limits for maxComputeSharedMemorySize - nak/qmd: properly set target shared mem size - rusticl: drop unneeded dependency to generated sources - rusticl: drop global allow statements - rusticl: specify allowed lints for tests in lib.rs - rusticl: add a bunch of trivial tests - rusticl/mem: fix Image::read for 1Darray images - rusticl/mesa: fix NULL pointer access in set_constant_buffer_stream - ac/llvm: fix get_global_address for global atomics - rusticl: reference resource in sampler and image view wrappers - ci: document what version to specify in RUST_VERSION - rusticl/util: make ThreadSafeCPtr Copy, Clone and transparent - rusticl/mesa: add PipeScreen::pipe - rusticl/mesa: rework Context creation - rusticl/mesa: make PipeScreen transparent - rusticl/mesa: make PipeScreen refcounted - libagx: fix heap argument type in libagx_draw_robust_index - clc: Fix createDiagnostics for LLVM-22 - nak: extract cmat load/store element offset calculation - nak: ensure deref has a ptr_stride in cmat load/store lowering - nak: fix MMA latencies on Ampere - st/interop: fix fence leak - rusticl/queue: fix error code for invalid queue properties part 1 - rusticl/queue: fix error code for invalid queue properties part 2 - rusticl/queue: fix error code for invalid sampler kernel arg - rusticl/kernel: take no kernel_info reference inside the launch closure - rusticl/spirv: preserve signed zeroes by default Kenneth Graunke (45): - brw: Refactor copy propagation checks for EOT send restrictions - brw: Fix units in copy propagation EOT restriction size calculation - brw: Update copy propagation into EOT sends handling for Xe2 units - crocus: Drop 16X MSAA code remnants - crocus: Fix a comment about supporting 16x MSAA - intel: Disable 16x MSAA support on Xe3 - brw: Use BAD_FILE instead of ARF null for second send payload - brw: Assert that EOT is always SHADER_OPCODE_SEND on pre-Xe3 - brw: Stop checking inst->is_send_from_grf() for g127 register hack - brw: Stop using is_send_from_grf() in CSE pass - brw: Drop inst->mlen check from is_send() - brw: Rename is_send_from_grf to is_send, replace other is_send() helper - brw: Properly resolve non-sendable sources in a few logical opcodes - brw: Enumerate SHADER_OPCODE_SEND sources and standardize how many - brw: Drop INTERPOLATE_AT_* opcodes from is_send() - brw: Drop interlock and memory fence logical opcodes from is_send() - brw: Drop uniform pull constant load virtual opcode from is_send() - brw: Drop INTERPOLATE_AT_* opcodes from is_payload() - brw: Drop interlock and memory fence logical opcodes from is_payload() - brw: Validate that send payloads can't be imms or have source mods - brw: Remove brw_inst::no_dd_check/no_dd_clear - nir: Add load_simd_width_intel to divergence analysis - intel/nir: Make ffma peephole optimization preserve fp_fast_math flags - brw: Move "SSA form" printing to after divergence analysis is run - brw: Lower certain subgroup size modes in brw_preprocess_nir - brw: Split brw_postprocess_nir() into two pieces - brw: Do most of NIR postprocessing before cloning for SIMD variants - brw: Add a quick NIR-based register pressure estimate pass - brw: Skip compilation of larger SIMDs when pressure is too high - iris/ci: Update trace checksums - brw: Only skip SIMD widths based on pressure if an smaller one compiled - elk: Delete ELK_SHADER_RELOC_DESCRIPTORS_ADDR_HIGH - brw: Rename brw_shader_reloc to intel_shader_reloc - intel: Move intel_shader_reloc to common code and drop elk_shader_reloc - brw: Drop ir_expression_operation_h from build system - brw: Rename brw_nir_trig build target to brw_nir_workarounds - intel: Make a libintel_compiler_nir internal static library - intel: Re-unify brw_prim.h and elk_prim.h - brw: Drop compiler/ from brw includes - brw: Move into a new src/intel/compiler/brw subdirectory - brw: Stop using type_size_dvec4 for fragment shader outputs - brw: Replace type_size_xvec4 with glsl_count_attribute_slots - brw: Refactor clip/cull distance mask setting into a helper - brw: Use BITFIELD_{MASK,RANGE} in clip/cull distance mask handling code - brw: Fix mesh shader asserts in clip/cull distance setting Konstantin Seurer (63): - radv: Optimize ray tracing position fetch - radv: Disable pointer flags and the GFX12 WA for emulated RT - radv: Implement watertightness for emulated RT - radv/rt: Optimize emulated ray-triangle tests - radv/rt: Use inv_dir for software ray-triangle tests - radv/rt: Implement null acceleration structure in shader code - radv/rra: Only write used BLAS - radv/rra: Increase rra_validation_context::location - radv/rra/gfx12: Handle box nodes without children - radv/rra/gfx12: Add validation - gallivm: Silence a warning - gallium/util: Fix an assert in util_resource_copy_region - lavapipe: Adjust imageGranularity for block formats - lavapipe/ci: Add context to some vkd3d-proton test fails - lavapipe: Set image_array for input attachment loads - gallivm: Implement txs with divergent explicit lod - gallivm: Implement arrayed non-arrayed descriptor compatibility - util: Fix sparse tile size when dimensions=1 - lavapipe/rt: Fix watertightness for real this time - lavapipe/rt: Set push_constant_size - lavapipe/rt: Do not use vk_acceleration_structure::size - radv: Add and use RADV_OFFSET_UNUSED - radv: Only write leaf node offsets when required - radv/bvh: Fix flush in bit_writer_skip_to - radv/bvh: Use a fixed indices midpoint on GFX12 - radv: Initialize base IDs when doing a BVH update with src!=dst - radv/bvh: Update leaf nodes before refitting - radv/bvh: Specialize the update shader for geometryCount==1 - vulkan/cmd_queue: Do not free if driver_free_cb is provided - vulkan/cmd_queue: Improve struct free code indentation - vulkan/cmd_queue: Recursively free struct members - vulkan/cmd_queue: Clean up generating copies - vulkan/cmd_queue: Reorder memcpy in get_struct_copy - radv: Use vk_acceleration_struct_vtx_format_supported - lavapipe: Use vk_acceleration_struct_vtx_format_supported - radv/rra/gfx12: Handle compressed primitive nodes - radv: Emit compressed primitive nodes on GFX12 - vulkan: Add MESA_VK_SHADER_STAGE_ALL - lavapipe: Mask invalid shader stage flags - radv: Rename radv_printf files to radv_debug_nir - radv: Add RADV_DEBUG=validatevas for address validation in nir - radv: Store parent node IDs inside nodes on GFX12 - radv/bvh: Copy parent_id during updates on GFX12 - nir: Use nir_def_as_* in more places - nir: Use nir_def_block in more places - radv/bvh: Do not write pointer flag related data on GFX103 - vulkan: Use a struct for debug markers - vulkan: Add more detail to encode debug markers - radv: Use vk_barrier_compute_w_to_compute_r more - radv,vulkan: Avoid a useless barrier in radv_update_bind_pipeline - nir/opt_ray_queries: Cleanup and return if functions is not singular - vulkan/bvh: Enable glsl extensions in meson - vulkan/cmd_queue: Remove unused variable - vulkan/cmd_queue: Handle internal structs - vulkan/cmd_queue: Handle struct arrays with pNext - Revert "lavapipe/ci: Disable stack-use-after-return detection for ASan" - vulkan/vk_cmd_queue: Clone VkSampleLocationsInfoEXT extending VkRenderingInfo - aco: Fixup out_launch_size_y in the RT prolog for 1D dispatch - lavapipe: Bump maxPrimitiveCount - lavapipe: Zero image null descriptors - lavapipe: Bump MAX_DESCRIPTOR_UNIFORM_BLOCK_SIZE - gallivm/nir/soa: Use the sign of src1 for imod - llvmpipe: Always recompute 1/w Kovac, Krunoslav (2): - amd/vpelib: Fix Possible dereferencing null - amd/vpelib: Minor Refactor Lars-Ivar Hesselberg Simonsen (20): - u_trace: Indirect capture fixes - panvk: Fix instrumentation on v12+ - panvk: Fix IUB decode - panvk/utrace: Pass async_op instead of mask - panvk/utrace: Make indirect capture wait optional - panvk/utrace: Add support for storing registers - panvk/utrace: Add sync32/64_wait support - panvk/utrace: Add sync32/64_add support - panvk/utrace: Add flush_cache support - panvk: Add utrace tracepoints in queue_submit - vulkan: Stop combining subpass dependencies - vulkan: Find first_subpass when creating renderpass - vulkan: Add transition_view_mask calculation - vulkan: Optimize implicit begin_subpass barrier - vulkan: Optimize implicit end_subpass barrier - panvk/ci: Add uncovered CTS issue to flakes - radv/ci: Add uncovered CTS issue to gfx1201 fails - panvk: Fix IUB decode - pan/format: Fix mapping for I16F - pan/format: Disable PAN_BIND_STORAGE_IMAGE for RGBA4/BGRA4 Leder, Brendan Steve (Brendan) (1): - amd/vpelib: General cleanup / optimization tasks Lewis Cooper (2): - pvr: Implement VK_KHR_maintenance3 - pvr: Implement VK_KHR_dedicated_allocation LingMan (7): - ci/rust: Drop date from Rust release channel selection - docs/rusticl: Update documented version requirements for meson and bindgen - mesa: Bump required Rust version to 1.82 - rusticl: Use \`is_aligned` from std - rusticl: Drop include paths for \`size_of`, \`size_of_val`, and \`align_of` - rusticl: Use std::mem::offset_of!() - nak: Drop include paths for \`size_of` and \`size_of_val` Lionel Landwerlin (148): - anv: reuse runtime descriptor set layout base object - anv: remove unused helper arguments - brw: fix NIR metadata invalidation with closest-hit shaders - brw: fixup source depth enabling with coarse pixel shading - brw: fixup coarse_z computation - brw: consider LOAD_PAYLOAD fully defined - brw: always ensure coarse pixel is disabled on Gfx9 - anv: fix wsi image aliasing - compiler: add gl_shader_stage_is_graphics - brw: make more passes printable through NIR_DEBUG - anv: move over to common descriptor set & pipeline layouts - anv: expose helper function outside of anv_pipeline.c - anv: rename vertex input emission helper - anv: reuse runtime flags field for descriptor set layout - anv: make anv_pipeline_sets_layout looks more like vk_pipeline_layout - anv: stop using anv_pipeline_sets_layout - anv: extract embedded samplers from pipeline_cache - anv: break ANV_CMD_DIRTY_PIPELINE into each stage - anv: avoid storing L3 config on the pipeline - intel: move deref_block_size to intel_urb_config - intel: reuse intel_urb_config for mesh - anv: store layout_type on the bind_map for convenience - anv: move URB programming to dynamic emission path - anv: avoid looking at the pipeline to flush push descriptors - anv: constify some helpers - anv: store gfx/compute bound shaders on command buffer state - meson: remove intel-clc options - brw: implement ACCESS_COHERENT on Gfx12.5+ - anv: fix source hash utrace prints - anv/brw: store min_sample_shading on wm_prog_data - anv/brw: move sample_shading_enable to wm_prog_data - anv: move primitive_replication emission to dynamic path - anv: move 3DSTATE_SF dynamic emission path - anv: simplify SBE emission - anv: move SBE emission to dynamic path - anv: move 3DSTATE_CLIP emission to dynamic path - anv: move 3DSTATE_VFG emission to dynamic path - anv: move 3DSTATE_TE::TessellationDistributionMode to dynamic path - anv: pass active stages to push descriptor flushing - anv: remove pipeline_stage unused field - anv: use a local variable for batch - anv: actually use the COMPUTE_WALKER_BODY prepacked field - anv: rework gfx state emission (again) - anv: subclass vk_pipeline - brw: compute consistent clip/cull distance masks with VUE - anv: Do not consider task as prerasterization - anv: fix missing meson dep - vulkan/runtime: add a few more shader properties - vulkan/runtime: add ray tracing pipeline support - brw: reorder reloc enums to leave embedded samplers at the end - anv: stop using descriptor layouts for descriptor buffers push sizes - brw: move URB channel mask shifting to the lowering pass - anv: fix R64* vertex buffer format support - vulkan/runtime: use a pipeline flag for unaligned dispatches - brw: enable register allocation to deal with multiple EOTs - brw: enable opt_register_coalesce to work with multiple EOT blocks - brw: workaround broken indirect RT messages on Gfx11 - brw: fix analysis dirtying with pulled constants - brw: make assign_curb_setup visible in optimizer debug - anv: fix uninitialized return value - brw: remove uniform from opt_offsets - brw: use a scalar builder for the load_payload on transpose loads - brw: fix INTEL_DEBUG=spill_fs - brw: fix broadcast opcode - anv: move input coverage mask setup to runtime flush - anv: temporary disable KHR_maintenance8 - Revert "anv: enable non uniform texture offset lowering" - Revert "brw: move texture offset packing to NIR" - intel: update code owners - anv: fix pipeline barriers with pre-rasterization stages - anv/utrace: avoid memseting timestamp buffers by using tracepoint flags - anv: fix partial queries - nir: add a new intrinsic for load dynamic tessellation config - brw: add ability to compute VUE map for separate tcs/tes - anv/brw/iris: move VS VUE computation to backend - brw: add support for separate tessellation shader compilation - anv: prep work for separate tessellation shaders - compiler: add stage_is_graphics() helper - anv: add infrastructure for common vk_pipeline - anv: move internal RT shaders around - anv: add runtime shader statistic support - anv: add shader instruction emission - anv: store a few default instructions - anv: switch over to runtime pipelines - anv: remove unused gfx/compute pipeline code - anv: expose VK_EXT_shader_object - anv: add an undocumented HW workaround for Gfx12.5 - anv: fixup robust_ubo_range mask - vulkan: remove incorrect assert - anv: remove divergence requirement - brw: don't use brw_null_reg() for unused SEND sources - anv: run nir_opt_acquire_release_barriers - brw: remove unused RT write code - brw: improve eot_reg computation in register allocate - anv: fixup 3DSTATE_COARSE_PIXEL emission - anv: avoid unnecessary 3DSTATE_PS_EXTRA emissions - brw: lower non coherent FS load_output in NIR - brw/blorp: lower MCS fetching in NIR - brw: lower shader opcode into tex_instr - brw: simplify texture surface/sampler handle sources - brw: fix split_sends with txf combining - brw: layout patch in VUE in position independent way - anv: fix streamout config comparison - anv: fix crash in ESO tests - brw: fix type conversion in tex operation params - nir/lower_tex: add an callback to lower txd ops - brw: use the new lower_txd_cb - elk: remove txd bindless sampler lowering - elk: use the new lower_txd_cb - nir/lower_tex: remove unused options - brw: fix render target indexing in FS output reads - vulkan/render_pass: fixup renderpasses barriers for 2D views of 3D images - nir: add pass to propagate image format to intrinsics - anv: run image/intrinsic update pass - iris: run image/intrinsic update pass - brw: avoid looking at variables to get image formats - u_trace: use os_get_option instead of getenv - intel/ds: lump all the draw under the same toggle - intel/ds: disable draw/blorp tracepoints by default on android - brw: prevent LOAD_REG modifications on MOV_INDIRECT/BROADCAST - anv: fix companion usage for emulated image - nir/divergence: add a new mode to cover fused threads on Intel HW - nir/lower_io: add get_io_index_src_number support for image intrinsics - compiler: add an access flag for intel EU fusion - brw: serialize messages on Gfx12.x if required - brw: add serialize send stats - anv: fix query copy with shaders - intel/ci: remove old comments - brw: fix invalid sparse bitfield offset computation - Revert "wsi: Implements scaling controls for DRI3 presentation." - anv: fix image-to-image copies of TileW images - brw: constant fold u2u16 conversion on MCS messages - brw: only consider cross lane access on non scalar VGRFs - brw: fix ballot() type operations in shaders with HALT instructions - nir/divergence: fix handling of intel uniform block load - anv: rename structure holding 3DSTATE_WM_DEPTH_STENCIL state - brw: handle GLSL/GLSL tessellation parameters - nir/lower_io: add missing levels intrinsics to get_io_index_src_number - anv/brw: fix output tcs vertices - anv: destroy sets when destroying pool - vulkan/render_pass: Add a missing sType - u_trace: reserve chunk space before emitting copies - anv: avoid null pointer access in utrace copies on CCS - brw: avoid invalid URB messages - anv: avoid invalid timestamp generation due to skipped commands - vulkan/runtime: simplify robustness state hashing - anv/blorp/iris: rework Wa_14025112257 - anv: disable software detiling on Xe2+ for image atomics 64bits Lorenzo Rossi (3): - nak: Fix pre-volta iadd3 panic during compilation - nak/kepler: Refine instruction scheduling - nvk: Fix QMD buffer length on upload Luc Ma (1): - dri: use XCB_PRESENT_EVENT_* enum instead of macros for consistency Lucas Fryzek (14): - lp: Don't allocate sampler functions if count is 0 - anv: Enable compression on astc emulation plane - vulkan/util: update pd feature codegen to use platform guards - anv: Remove special CROS_GRALLOC path from format logic - hasvk: Remove special CROS_GRALLOC path from format logic - anv: Update viewport/scissor state when count changes - vulkan/runtime: Error if ahb has more than one layer - anv: Assert that we only import ahb image with one layer - anv: Enable R10X6 & R10X6G10X6 unorm formats - anv: Modify anv feature (dis)enable code to match other drivers - vulkan/android: Add rp_attachment_has_external_format helper - vulkan/runtime: Add logic to set external format resolve mode - anv: Add external format resolve operation using blorp - anv: Enable VK_ANDROID_external_format_resolve Lucas Stach (6): - etnaviv: Update headers from rnndb - etnaviv: stop touching code steering bits while updating uniforms - etnaviv: update code steering bit when writing shader instructions - etnaviv: don't emit start/end PC states when unified instmem is present - etnaviv: use new shader range registers when icache is present - etnaviv: fix YUV tiler blits Ludvig Lindau (1): - panfrost: Make instrs_equal check res table/index Luigi Santivetti (22): - pvr: rename pvr tex format description variables for clarity - pvr: rename pvr_{create,generate} to appear at the end - pvr: split out missing output register write handling into separate function - pvr: determine rt layers based on rta support - pvr: fix logic for setting vdm instance count present - pvr: don't csb emit multi-layer clear attachments without rta support - pvr: reset the pds info map entries pointer to avoid double free - pvr: align texture stride for spm as the PBE requires - pvr: take zonlyrender into account when setting up ZLS control - pvr: add support for VK_KHR_maintenance1 - pvr: add support for VK_KHR_maintenance2 - pvr: unify the creation of load_op objects and shaders - pvr: rename job field holding pds PR background objects - pvr: rename {init,setup} command buffer helpers - pvr: drop unused argument from pvr_load_op_shader_generate() - pvr: add support for U16U16U16 texture state format - pvr: restrict signed A2-10 bits per component formats to vertex only - Revert "pvr: treat VK_IMAGE_CREATE_MUTABLE_FORMAT_BIT as not supported" - pvr: add initial driver support for VK_KHR_multiview - pvr: improve unemitted resolve attachments readability - pvr: restrict the scope of copy_{buffer,image}_to_{image,buffer} - pvr: propagate image samples when doing a blit from DS surface Marek Olšák (168): - gallium: make pipe_screen::finalize_nir return void - gallium: replace get_compiler_options with pipe_screen::nir_options - st/mesa: don't expect pipe_screen::nir_options to be NULL for supported shaders - mesa: use pipe_screen::nir_options instead of NirOptions - glsl: use pipe_screen::nir_options instead of NirOptions - ac/surface/gfx12: add addr_from_coord for sparse MSAA textures - ac/surface/gfx12: select 64K tiling for sparse MSAA textures - radeonsi/gfx12: enable sparse textures - ac/nir: don't vectorize to 96-bit and 128-bit LDS loads (it's slower) - ac/nir: mark all input loads as reorderable and speculatable (for LICM) - ac/llvm: rewrite global & shared stores to share code - ac/llvm: rewrite global & shared loads to share code - ac/llvm: always use opaque pointers - ac/llvm: fix readlane with vectors - radeonsi: disallow the compute copy for Z/S - radeonsi: add a workaround for gfx10.3-11 corruption with R9G9B9E5_FLOAT - radeonsi: recompute FS output IO bases to prevent an LLVM crash - radeonsi: get si_shader_info::input::usage_mask from NIR - radeonsi: flatten struct si_vs_tcs_input_info - radv,radeonsi: mark VS input loads and poly stipple load speculatable - radv: don't sink VS input loads and move them to the top - nir: add nir_instr_can_speculate helper (for LICM) - nir: add nir_tex_instr::can_speculate - nir: add access to load_smem_amd (for ACCESS_CAN_SPECULATE) - nir/divergence_analysis: simplify nir_vertex_divergence_analysis - nir/opt_move_to_top: check can_reorder & can_speculate - nir: silence a warning in nir_opt_shrink_vectors - nir: handle store_buffer_amd in nir_intrinsic_writes_external_memory - radeonsi/ci: import piglit & cts build scripts - radeonsi/ci: don't build GLES CTS separately - radeonsi/ci: update gfx12 and other failures - nir/group_loads: handle more loads - nir/group_loads: allow moving loads across instructions without defs - nir/group_loads: split is_barrier into is_barrier + is_terminate - nir/group_loads: group any reorderable intrinsics regardless of barriers - nir/group_loads: invert the return value of can_move to reflect its true meaning - nir/group_loads: remove mostly duplicated function is_memory_load - nir/group_loads: make is_grouped_load use get_load_resource - nir/group_loads: use nir_instr_next/prev - nir/group_loads: store our custom instr->index in an array - nir/group_loads: don't use pass_flags to store the indirection level - nir/group_loads: rename to nir_opt_group_loads - nir: mark inverse_ballot & is_subgroup_invocation_lt_amd as CAN_REORDER - nir: change how can_mov_out_of_loop is set for intrinsics in nir_can_move_instr - nir: handle can_reorder robustly in nir_can_move_instr - nir: renumber nir_move_options - nir: split nir_move_load_frag_coord from nir_move_load_input - nir: handle load_input_vertex in nir_can_move_instr - nir: add more nir_move_options - nir: add nir_move_only_convergent/divergent - glsl: fork exec_node/list -> ir_exec_node/list as private GLSL IR utility - intel: fork exec_node/list -> brw_exec_node/list as a private Intel utility - nir: move list.h outside the glsl directory - nir: remove C++ stuff from list.h - nir: remove unused stuff from list.h - glsl: remove unused stuff from ir_list.h - glsl: remove unused symbol_table_entry::get_interface - glsl: remove reparent_ir - nir/opt_group_loads: support tex instructions without resource srcs for i915 - glsl/tests: fix memory leaks - ralloc/linalloc: allow adding custom code to LINEAR_ALLOC new operator - glsl: add support for linear_ctx into ir_instruction - glsl: switch ir_instruction to linear_ctx to eliminate malloc overhead - glsl: switch ir_variable_refcount to linear_ctx - mesa: switch symbol_table to linear_ctx - dri: fail creating DRI images that exceed hw limits - nir: don't allocate nir_constant::elements if there are none - nir: add nir_variable_{set,append,steal}_name{f}() to modify nir_variable names - nir: eliminate most ralloc/malloc for nir_variable names - nir/clone: don't call ralloc_strdup with a NULL pointer for intrinsic names - nir: don't use variables as ralloc parents, use the shader instead - nir: add nir_variable_create_zeroed helper - nir: use gc_ctx for nir_variable to reduce ralloc/malloc overhead - meson: reinstate LLVM requirement for r300 and enforce it for i915 too - meson: remove unused -DLLVM_AVAILABLE - mesa: move src/mapi to src/mesa/glapi - docs,ci: update mapi relocation - mesa: remove inc_mapi - mesa: stop using inc_mesa in most places that have nothing to do with GL - glsl: use pipe caps in opt_shader - glsl: replace LowerBuiltinVariablesXfb with pipe caps - glsl: replace LowerPrecisionFP16/Int16 with pipe caps - glsl: replace LowerPrecisionDerivatives with pipe caps - glsl: replace LowerPrecisionFloat16Uniforms with pipe caps - glsl: replace LowerPrecision16BitLoadDst with pipe caps - glsl: replace LowerPrecisionConstants with pipe caps - st/mesa: replace EmitNoIndirect* with pipe caps - glsl: move PositionAlwaysInvariant/Precise options to gl_constants - glsl: remove gl_shader_compiler_options - ac/nir/meta: allow compute blits with R5G6B5 & R5G5B5A1 formats on GFX9+ - radeonsi/gfx12: print swizzle modes for AMD_TEST=imagecopy - ac/nir: clarify the behavior of ac_nir_lower_ngg_options::can_cull - ac/llvm: inline ac_array_in_const*_addr_space - ac/nir: inline ac_get_ptr_arg - ac/nir: remove unused ac_get_ptr_arg & ac_arg_type_to_pointee_type - ac: simplify AC_ARG_CONST_*PTR enums - ac/llvm: make ac_get_arg non-inline - radeonsi: bitcast shader args to float in LLVM IR manually - ac/llvm: make AC_ARG_FLOAT equal to AC_ARG_INT - ac: merge AC_ARG_INT & AC_ARG_FLOAT into single AC_ARG_VALUE - egl,glx: allow OpenGL with old libx11, but disable glthread if it's unsafe - util/set: improve support for usage without "set" structure allocation - radv,zink,st/mesa: use _mesa_set_fini instead of ralloc_free - util/set: start with 16 entries to reduce reallocations when growing the set - util/set: don't allocate the smallest table, declare it in the struct - util/set: set _mesa_set_init return type to void - util/set: add _mesa_set_copy, a cloning helper without allocation - util/hash_table: start with 16 entries to reduce reallocations - util/hash_table: improve support for usage without "hash_table" allocation - util/hash_table: don't allocate the smallest table, declare it in the struct - util/hash_table: set _mesa_hash_table_init return type to void - util/hash_table: don't allocate hash_table_u64::table, declare it statically - util/hash_table: add _mesa_hash_table_copy, a cloning helper without allocation - nir/dominance: don't allocate 0-sized dom_children - nir/dominance: eliminate ralloc overhead for allocating dom_children - nir: make nir_block::predecessors & dom_frontier sets non-malloc'd - nir/lower_vars_to_ssa: don't ralloc sets - nir/instr_set: don't ralloc the set - nir/remove_dead_variables: don't ralloc the set - nir/opt_vectorize: don't ralloc the set - nir/gather_info: don't ralloc the set - nir/search: don't ralloc the hash table - nir/opt_copy_prop_vars: don't allocate vars_written::derefs hash table - nir/opt_copy_prop_vars: don't allocate vars_written_map hash table - nir/opt_copy_prop_vars: don't allocate copies::ht hash table - nir/lower_vars_to_ssa: don't ralloc the hash table - nir/opt_find_array_copies: don't allocate the hash tables - nir/split_vars: don't allocate the hash tables - nir/serialize: don't allocate the hash tables - nir/opt_load_store_vectorize: don't allocate 0-sized offset_defs - nir: convert nir_instr_worklist to init/fini semantics w/out allocation - nir/opt_dead_write_vars: don't use ralloc context, share dynarray among blocks - nir/gather_info: don't allocate the ralloc context - glsl/opt_function_inlining: don't ralloc the hash table - glsl/ir_constant_expression: don't ralloc the hash table - glsl/ir_variable_refcount: don't ralloc the hash table - glsl_to_nir: don't allocate 0-sized num_params & subroutine_types - glsl_to_nir: don't allocate 0-sized arrays for Uniform/ShaderStorageBlocks - nir/opt_call: handle load_global(_amd) with SPECULATE as rematerializable - nir/opt_sink: handle load_global_amd - nir/opt_move_to_top: handle load_global_amd with ACCESS_SMEM_AMD - aco: check that global addresses are 64bit, apply_nuw_to_ssa to global_amd/smem - ac/llvm: fix handling COHERENT and VOLATILE flags for global access - ac/llvm: port load_smem_amd behavior to load_global_amd - aco,radeonsi: expand 32-bit shader arg pointers to 64 bits for ACO - ac/nir: switch nir_load_smem_amd uses to ac_nir_load_smem wrapper - radv: fix load_smem alignment - radeonsi: always set TC_L2 for CP DMA on GFX12 - radeonsi: inline si_upload_const_buffer - radeonsi: if rebinding the same constbuf, don't update refcount with atomics - radeonsi: remove recursion from si_set_constant_buffer - radeonsi: don't ref and unref an index buffer uploaded from a user buffer - radeonsi: switch VBO descriptor uploads from u_upload_alloc_ref to u_upload_alloc - radeonsi/ci: primitive_counter failures are no longer reproducible on gfx12 - radeonsi: compute blake3 hashes of internal shaders if they are not set - gallium/u_threaded: remove refcounting for draw indirect buffers - gallium/u_threaded: remove refcounting for dispatch compute indirect buffers - gallium/u_threaded: remove refcounting for clear_buffer - gallium/u_threaded: remove refcounting for draw mesh indirect buffers - gallium/u_threaded: remove refcounting for get_query_result_resource - gallium/u_threaded: remove refcounting for buffer_unmap - gallium/u_threaded: remove refcounting for buffer_subdata - nir: remove load_smem_amd - r300: fix DXTC blits - winsys/radeon: fix completely broken tessellation for gfx6-7 - zink: fix mesh and task shader pipeline statistics - Revert ABI breakage "amd: Add user queue HQD count to hw_ip info" - gallium/noop: don't unref buffers passed to set_vertex_buffers to fix crashes Marek Vasut (4): - etnaviv: hwdb: update gc_feature_database from ST - etnaviv: Turn ETNA_CORE\_ into ETNA_FEATURE_CORE\_ - pvr: fix features pointer on GX6650 (4.46.6.62) - pvr: fix device info for GX6250 (4.45.2.58) Mario Kleiner (6): - asahi: Fix lseek failure error handling in agx_bo_import(). - asahi: Set PIPE_BIND_SCANOUT in agx_resource_from_handle(). - wsi/display: Accept 0 nits for HDR light level properties for "undefined" - wsi/display: Initially set default HDR metadata from EDID for HDR modes - wsi/display: Allow atomic modeset for change of Colorspace or HDR poperties - wsi/wayland: Zero min_luminance, max_luminance HDR light levels are valid. Mark Collins (1): - freedreno/drm: Only initialize memory data source when Perfetto is active Martin Krastev (1): - Revert "ci: Disable vmware farm" Martin Roukala (né Peres) (24): - radv/ci: add post-merge jobs for gfx1201 - zink/ci: add post-merge jobs for gfx1201 - zink/ci: update the nvk expectations - nvk/ci: document a new fail and flakes - radv/ci: document new flakes - freedreno/ci: document new flakes - radv/ci: disable hang detection in navi31-vkcts - ci: disable the valve-kws farm - Revert "ci: disable the valve-kws farm" - ci/ci-tron: uprev the job submission template - freedreno/ci: uprev the kernel for the a750 - nvk/ci: document some vk3d fails - ci-tron: uprev b2c to v0.9.17 - radv/ci: switch to default kernel to b2c's default kernel - nvk/ci: switch to default kernel to b2c's default kernel - zink/ci: raise the job timeout from 5 to 8 minutes - turnip/ci: document more flakes - zink/ci: document more flakes in the a750 job - turnip/ci: switch vkcts testing to the KWS farm - ci,crnm: remove unsupported arguments by console.print - ci,crnm: remove unused imports - turnip/ci: enable a750_vk in marge pipelines - turnip/ci: squeeze a750-vk into 4 jobs - zink/ci: run the a750 job in pre-merge Mary Guillemard (85): - panvk: Fix nullDescriptor for dynamic descriptors - panvk: Wire robustness2 buffer info down to pan/bi - panvk: Exposes robustBufferAccess2 on v11+ - pan/genxml: Add missing parenthesis on pan_cast_and_pack macros - pan/genxml: Make resource table optional on RUN_COMPUTE{_INDIRECT} - panvk: Add basic infrastructure for shader variants - pan/bi: Fuse FCMP/ICMP on Valhall - pan/bi: Properly handle SWZ.v4i8 lowering on v11+ - panvk: Always use varying_count in emit_varying_attrs - panvk: track oq write jobs in JM - panvk: Directly use index buffer tracked value in JM - libcl: Add stdatomic.h - panfrost: Allow to pass job dependencies in grid for precomp JM - libpan: Add draw indexed and indirect helper for Bifrost - panvk: Prepare draw_emit_attrib_buf and draw_emit_attrib for indirect - panvk: Move JM draw preparation logic to prepare_draw - panvk: Prepare panvk_draw_prepare_varyings for JM indirect - panvk: Prepare tiler and vertex dcd for JM indirect - panvk: Implement indirect draw for Bifrost on JM - panvk: Use indirect path for indexed draw on JM - panvk: Make indexed draw use indirect indexed draw - panvk: Parallelize min max index search on JM - panvk: Call nir_opt_access - pan/bi: Switch to nir_lower_alu_width - pan/bi: Vectorize UBOs load/store - pan/bi: Handle needless conversions in nir_lower_bool_to_bitsize - pan/bi: Revamp bi_optimize_nir - pan/bi: Move pan_lower_sample_pos to next block - pan/bi: Stop exposing bifrost_nir_lower_load_output - panvk: Remove unused color_output_var function in fb_preload - panvk: Lower sampler and texture index in case of offset - panfrost: Split compilers preprocess_nir - panfrost: Move nir_lower_io outside of postprocess - panfrost: Split texture lowering passes - pan/bi: Split bi_optimize_nir and run bi_optimize_loop_nir in preprocess - pan/bi: remove dead variables in preprocess - pan/bi: Run opt_sink and opt_move in preprocess - nouveau/headers: Properly parse DMA classes for Turing and Ampere A - nouveau/headers: Mark SET_POINT_SIZE as using float - nouveau/headers: Handle Ampere A GPFIFO in dumper - nouveau/headers: Add missing M2MF parsing and set it for subchan 2 - nouveau/headers: Fix nv_push rust push_inline_data implementation - nouveau/headers: Add raw INC methods in nv_push rust impl - nvk: Force GART for command buffers - nvk: Use MEM_LOCAL for nvk_cmd_mem_create - nak: add Ldsm - hk: Return 0 for opaque memory capture replay - pan/bi: Ensure to merge adjacent ifs after bifrost_nir_lower_shader_output - pan/bi: Reintroduce bi_fuse_small_int_to_f32 on v11+ - pan/bi: Make va_optimize_forward run until there is no progress - pan/bi: Propagate MKVEC.v2i8 and V2X8_TO_V2X16 for replicate swizzle - panvk: Do not clamp blend constants in command buffer - panvk: Enable SNORM rendering - panvk/ci: Update waivered tests - pan/decode: Fix SYNC_SET32 double dots - panvk: Fix wrong type for sb_mask in CmdSetEvent2 - panvk: Take VK_DEPENDENCY_ASYMMETRIC_EVENT_BIT_KHR into account - docs/features: Mark VK_KHR_maintenance9 as done for ANV - hk: Move query pool creation/destruction - hk: Add support for VK_QUERY_POOL_CREATE_RESET_BIT_KHR - hk: Rework queue creation logic - hk: Advertise VK_KHR_maintenance9 - nir/print: Fix load_converted_output_pan and load_readonly_output_pan - panvk: Follow nir_lower_io for subpass lowering - panvk: Properly set shader binary properties - nouveau/headers: Autogenerate push method dumpers - nouveau/headers: Handle all compute classes in vk_push_print - nouveau/headers: Handle all DMA classes in vk_push_print - nouveau/headers: Handle all 3D classes in vk_push_print - nouveau/headers: Handle more gpfifo classes in vk_push_print - nouveau/headers: Include class headers instead of redefining class ids - nouveau/headers: Add Blackwell support to nv_push_dump - nouveau/headers: Properly set subchannel 3 to 2D engine in vk_push_print - nouveau/headers: Import Blackwell host class headers - nouveau/headers: Handle unbound sub channels in vk_push_print - panvk, vk/meta: Move D/S sanitizing to panvk - asahi: Add base expectation on VKCTS main - nouveau/headers: Define fake devices in a table for nv_push_dump - nouveau/headers: Add missing Kepler, Maxwell and Pascal defs to nv_push_dump - nouveau/headers: Properly reformat nv_push_dump - hk: Fix maxVariableDescriptorCount with inline uniform block - hk: Disable 1x in sampleLocationsSampleCounts - hk: Remove unused allocation in queue_submit - hk: Make width and height per block in HIC - hk: Allocate the temp tile buffer in copy_image_to_image_cpu Matt Coster (6): - pvr: Fill in missing {u,s}norm equivalents for tex formats - pvr: Add missing format adjustment for e5b9g9r9 - pvr: Add macros to iterate all supported tex formats - pvr: Cleanup compressed border colour support - pvr: Use 2D texstate for buffer views to allow for >8k sizes - pvr: Add support for custom border colors Matt Turner (4): - meson: Allow controlling perfetto fallback - meson: Allow configuring with Android-internal perfetto - brw/algebraic: Protect SHUFFLE from OOB indices - elk/algebraic: Protect SHUFFLE from OOB indices Mauro Rossi (4): - intel/mda: Fix gnu-empty-initializer warning - amd: require LLVM when amd-use-llvm is enabled - android: fix building rules for i915, r300 - util: Fix gnu-empty-initializer error Max R (2): - d3d10umd: De-bufferize OutputMerger - d3d10umd: Flush on present Maíra Canal (3): - vulkan: create a wrapper struct for vk_sync_timeline - vulkan: don't destroy vk_sync_timeline if a point is still pending - broadcom/ci: remove synchronization-related flakes and skips Mel Henning (68): - nouveau/headers: Update g_nv_name_released.h - nak/mark_lcssa_invariants: Invalidate divergence - loader: Don't load nouveau GL on nvidia kmd - meson,nvk: Require rustc-hash 2.0 or later - nvk: Call cmd_buffer_begin_* based on queue flags - nvk: Factor out nvk_queue_engines_from_queue_flags - nvk: Check subchannels are valid in nv_push - nvk: Disable non-graphics timestamp queries - zink: Fix a few profile errors - zink: Convert profile tabs to spaces - zink: Add zink_check_requirements - loader: Don't fall back to nouveau GL without zink - nvk: Split out NVC0_FIFO_SUBC_FROM_PKHDR helper - nvK: Add nvk_cmd_buffer_last_subchannel - nvk: Reduce subc switches in cmd_invalidate_deps - nvk/copy: Split out nvk_remap_insert_aspect - nvk/copy: Split out nvk_remap_extract_aspect - nvk/copy: Split out nvk_remap_copy_aspect - nvk/copy: Implement CopyImage2 between R and D/S - nvk: Expose VK_KHR_maintenance8 - nvk: Clear cond_render_gart_* in reset_cmd_buffer - nak/hw_runner: Make a few more items public - nak: Add a test to check how RENDER_ENABLE works - nvk/cmd_pool: NVK_DEBUG=trash_mem for alloc_mem - nvk: Clear second SET_RENDER_ENABLE operand - nvk: Remove gart from the name of cond_render_mem - nvk: Move cond rendering memory out of gart - nvk: Reuse the same cond render temp in a cmd_buf - nvk: Don't re-initialize cond rendering operand B - nvk: Only copy 32-bits for cond render operand A - nir: Don't require nir_metadata_control_flow - nir/phi_builder: Adjust valid_metadata assert - util: Add range_minimum_query - nir: Add a faster lowest common ancestor algorithm - treewide: Spell indices correctly - nak: Remove Option<> from SSARef::file() return - nak: impl HasRegFile for SSARef and &[SSAValue] - nak/assign_regs: Make src_ssa_ref return a slice - nak: Make BindlessSSA store [SSAValue; 2] - compiler/rust: impl AsSlice for Box - nak: Special case Box in derive_from_variants - nak: impl SM*Op for Op - nak: Place most Op structs in Box<> - nak: Don't copy-prop adds that flush to zero - nak: Fix divergence test for redux availability - util/macros: Add ATTRIBUTE_COLD - nouveau/headers: Mark vk_push_print as cold - nouveau/headers: Split out "cases" in template - nouveau/headers: Deduplicate push dump impls - nouveau/headers: Use previous method for default - nak: Add OpSgxt - nak: Implement bitfield_extract with OpSgxt - nvk: Only run one INVALIDATE_SHADER_CACHES - nvk: Combine BARRIER_{COMPUTE,RENDER}_WFI - nvk: Fix execution deps in pipeline barriers - nvk/cmd_buffer: Remove redundant tests for access - vulkan: Drop vk_pipeline_stage_flags2_has_*_shader - nvk: INVALIDATE_SHADER_CACHES on most recent subc - nvk: WFI on the most recent subc - nvk/cmd_copy: Use PIPELINED for user transfers - nvk/cmd_copy: Pipeline user copy_rect operations - nvk: Reduce subc switches with events - nvk: Call INVALIDATE_RASTER_CACHE for shading rate - nvk: FLUSH_PENDING_WRITES in gr semaphore release - nvk: Fix maxVariableDescriptorCount with iub - nvk: Really fix maxVariableDescriptorCount w/ iub - nvk: VK_DEPENDENCY_ASYMMETRIC_EVENT_BIT_KHR - nak/opt_lop: Don't handle modifiers in dedup_srcs Michal Krol (3): - gallium: Do not flush subnormals during tessellation. - lavapipe: Bump maxTransformFeedbackBufferDataStride to 2048. - llvmpipe: Add support for 8x MSAA. Michel Dänzer (2): - egl/dri: Name struct dri2_egl_buffer - egl/gbm: Destroy excess BOs Mike Blumenkrantz (217): - gallium/hud: set the framebuffer texture when drawing - ci: bump VVL to 1.4.322ish - zink: fix valid contents check for adding new bind - lavapipe: call nir_lower_int64 - lavapipe: maintenance9 - lavapipe: VK_KHR_unified_image_layouts - zink: use maint9 implicit query resets when available - zink: flag dmabuf exports on usage set, not synchronization - zink: simplify sampler bufferview change for non-db path - egl/x11: don't leak device_name when choosing zink - zink: account for generated tcs when pruning programs - zink: remove extra gfx prog unref during separable replacement - anv: fix format compatibility check typo - ci: add venus-lavapipe flake - ci: disable xwm decorations in weston - zink: create a dummy image for shaderdb runs - zink: drop primitiveTopologyPatchListRestart from profile - zink: just check multiview availability to advertise extensions - crocus: silence perf_debug -Waddress warnings - iris: silence perf_debug -Waddress warnings - vulkan: silence typed_memcpy -Waddress warnings - zink: skip all glx piglit tests on anv-adl - zink: verify that no generated tcs is ever in zink_context::gfx_stages - kopper: fix initial swapinterval setting - zink: also add access stage sync when rebinding buffers - zink: check for multi-context image/buffer rebinds during dispatch - zink: fix tc buffer replacement rebind condition - zink: trigger multi-context buffer invalidate on internal buffer invalidate - mesa/fbobject: tweak attachment validation - crocus: stop using util_framebuffer_init - i915: stop using util_framebuffer_init - zink: add cezanne skip for a device loss flake - mesa: fix and advertise GL_EXT_sRGB - zink: zero dynamic rendering resolve views on rp end - tc: also inline depth resolves - zink: add ZINK_DEBUG=rploads to mimic tiler behavior - zink: fix assert for unsynchronized non-GENERAL image barriers - tc: don't clobber CSO info when renderpass has ended - zink: don't access ctx in submit_queue - zink: stop always syncing threaded flushes - perfetto: unify init - mesa: make _mesa_bufferobj_release_buffer static - mesa: add a ctx param to _mesa_bufferobj_release_buffer - mesa/st: check for tc on context create - util/tc: don't print END_BATCH in debug - tc: break out buffer list busy check - tc: add a function to check the internal buffer lists - freedreno: stop using util_set_vertex_buffers - r300: stop using util_set_vertex_buffers - r600: stop using util_set_vertex_buffers - zink: destroy u_uploaders earlier in context destroy - gallium: set prefer_real_buffer_in_constbuf0 for all drivers using tc - gallium: always upload cbuf0 when cap is set - mesa/st: rework thread scheduler handling + add dispatch tracking - tc: remove user cbuf uploads - zink: optimize a GENERAL layout case in pre-draw/dispatch barriers - zink: fix image sync deferral - zink: remove UNSYNCHRONIZED map flag during unmap flush for non-subdata calls - zink: improve deferred buffer barrier heuristics - glthread: mark internal bufferobjs for the ctx they belong to - st/program: stop calling st_finalize_nir() unnecessarily for variants - kopper: don't sync glthread from swapbuffers - glx/egl/kopper: explicitly pass __DRI2_FLUSH_CONTEXT when appropriate - glx/kopper: don't call glFlush from swapbuffers - zink: sprinkle in a bunch of MESA_TRACE_FUNC - zink: inline zink_resource_access_is_write() - zink: ALWAYS_INLINE resource inlines - zink: break out unflushed batch waiting into separate function/mechanism - zink: pass ctx to sparse bind functions - zink: when sparse unbinding, always wait on main timeline semaphore - zink: trigger fb unbind barrier on resolve images too - zink: fix sizing on resolve resource array - zink: update resized swapchain depth buffer layout while blitting - zink: unify/fix clear flushing - zink: fixes for flushing clears - zink: also set msrtss stencil - zink: always flush clears when doing single-aspect blit to avoid data loss - zink: enable single-aspected blitting of mixed z/s formats - zink: fix some weird indentation in update_binds_for_samplerviews() - zink: flag resources for layout eval in update_binds_for_samplerviews() - zink: unset validate_all_dirty_states - zink: set can_bind_const_buffer_as_vertex - radv: ALWAYS_INLINE radv_upload_graphics_shader_descriptors and relateds - zink: add a util function for appending a batch state - zink: split out batch state finding - zink: null out zink_batch_state::next when reusing a batch state - zink: defer batch state resets more competently - zink: check ctx batch states first when finding a usable one - zink: stop using atomics to check fence submit/complete - zink: stop trying to oom prune batch states - zink: rename zink_batch_state::unref_resources -> unref_resource_objs - zink: move buffer hashlist clear to normal batch state reset - zink: stop deferring resource object unrefs - zink: once there are many outstanding submits, check for timeline updates - zink: zero db offset on batch reset - zink: don't init non-db batch stuff in db mode - zink: reset batch descriptor states again before use on recycle - zink: don't increase db scale when resizing a db up to the current scale - zink: add some cml flakes - mesa: tag a couple framebuffer commands for MESA_VERBOSE=api - mesa: add MESA_DEBUG=fallback_tex - kopper: unwrap screen before checking cpu flag - tc: don't unset resolve resource in set_framebuffer_state - mesa/varray: inline a bunch of functions - zink: reeneable OVR_multiview2 - mesa: add task/mesh to _mesa_shader_stage_to_subroutine_prefix() - aux/trace: dump more mesh draw info - zink: remove rebar requirement for descriptor buffer support - zink: add another flag to determine whether linked program compile is done - zink: toggle ctx->has_swapchain when flushing clears - zink: flag pipeline_changed when updating shader modules - zink: clamp subgroup op return types to required int/uint types - zink: fix edgeflags check on program creation - zink: correctly handle batch_id==0 in check_last_finished() - zink: only set compute module info on dispatch (after compile fence) - zink: set current compute prog after comparing against current compute prog - zink: do bindless init when binding a bindless shader, not on create - zink: just reference compute progs to batch on delete - zink: ensure transient surface is created when doing msaa expand - gallium: add pipe_context::resource_release to eliminate buffer refcounting - zink: eliminate buffer refcounting to improve performance - zink: flag vertex element state for rebind after vstate draws - zink: don't init batch descriptors for copy contexts - zink: simplify state iterating in find_completed_batch_state() - zink: make find_completed_batch_state() only return state for COPY_ONLY ctx - zink: update gfx pipeline less frequently - zink: use implicit offsets for function temp variables in ntv - zink: more vvl exceptions - cso: unbind vertex buffers when unbinding context - tc: eliminate refcounting for set_shader_buffers - ci: bump vvl to another random version - zink: store last index buffer - zink: always use vkCmdBindVertexBuffers2 - zink: simplify index type access to normal array - zink: move draw state flag resets into their blocks - zink: add some pre-checks before calling query update/suspend/resume - zink: add another tu flake - mesa: support GL_NV_representative_fragment test - zink: support NV_representative_fragment_test - zink: add a fastpath for nooping vertex and draw buffer barriers - zink: ALWAYS_INLINE zink_set_vertex_buffers_internal - zink: split update_res_bind_count - zink: use velems buffer count in blitter instead of gfx mask - zink: move zink_bind_vertex_elements_state() to zink_context.c - zink: move vbo unbind to bind_vertex_state - zink: rescope some zink_set_vertex_buffers_internal variables - zink: use memcpy for vbo bind - zink: delete some function decls that no longer exist - zink: only remove buffer deferred sync on release - zink: eliminate even more calls to sync functions - util/vbuf: stop nooping set_vertex_buffers calls - Revert "util/vbuf: stop nooping set_vertex_buffers calls" - zink: mark dirty_gfx_stages using util function - zink: delete weird prog->pipelines sizing - zink: make zink_descriptor_util_push_layouts_get() static - zink: unify ntv code for storing shared/scratch memory - zink: unify ntv code for loading shared/scratch memory - zink: add enum zink_pipeline_idx to distinguish between types of pipelines - zink: break out setting draw-time dynamic state into separate function - zink: some minor tweaks to descriptor template code - zink: use a better array loop sizing for gfx descriptor program init - zink: stop unsetting zink_gfx_pipeline::modules on shader unbind - zink: don't use screen ralloc context for screen::pipeline_libs - zink: imagelessFramebuffer is no longer required/used - tc: don't sync on internal UNSYNCHRONIZED texture_map calls - mesa/st: add a flags param to st_texture_create() - mesa/st: mark internal texture map calls as UNSYNCHRONIZED - mesa/st: mark internal buffer map call as UNSYNCHRONIZED - zink: make zink-anv-adl jobs use descriptor buffer - zink: hook up VK_EXT_mesh_shader - zink: implement compiler-side handling for mesh shaders - zink: split out descriptor invalidation to be more explicit - zink: use pipeline_idx for descriptor invalidation - zink: implement mesh shaders - zink: wait on queues during screen destroy - zink: account for kopper dt not having a swapchain when pruning batch usage - zink: prune active queries in reset_batch_state_ctx() - zink: call post_submit directly from submit_queue - zink: check for zink_batch_state::ctx before using during descriptor state reset - zink: null out zink_batch_state::ctx when adding to the screen list - zink: reset batch states on destroy - zink: flag gfx pipeline_changed if switching from a shader object draw - zink: flag mesh pipeline_changed if switching from a shader object draw - zink: only try update descriptors on draw/dispatch when necessary - zink: fix descriptor array indexing for mesh pipeline - zink: set OutputPoints for mesh point output - zink: various cleanups for mesh+multiview - zink: stop creating GPL inputs for mesh - zink: disable single-aspected blits for now - tu: don't deref end info in tu_CmdEndRendering2EXT - zink: add ZINK_DEBUG=nogeneral to disable unified image layouts - mesa: don't assert when finding a renderbuffer miplevel fails - zink: fix u_blitting when clears are pending - hud: delete buffer refcounting - zink: convert task_payload offset to array index in prepass - vulkan: update spec to 1.4.328 - lavapipe: move copy_depth_box to lvp_image.c - lavapipe: handle aspected depth/stencil memory->image HIC transfers - lavapipe: VK_KHR_copy_memory_indirect - mesa: delete task and mesh programs on context destroy - zink: fix disabling multiview mesh with shader objects - zink: various fixes for custom sample locations - zink: stop using vk lazy allocations / transient attachments - zink: strip dmabuf bind flags when creating transient image - zink: always add mutable to transient surface creation when needed - zink: only add mutable bind for transient surfaces when necessary - zink: disable msrtss handling when blitting - glsl: fix gl_ViewID_OVR type to uint - mesa: copy NumSamples in reuse_framebuffer_texture_attachment - zink: enable GL_EXT_mesh_shader - zink: enable srgb-mutable for dmabufs when possible - zink: defer swapchain updates for interval changes if acquired image is active - zink: consistently set/unset msrtss in begin_rendering - zink: disable primitiveFragmentShadingRateMeshShader feature - zink: collapse gfx pipeline fetching and binding conditionals - zink: collapse mesh pipeline fetching and binding conditionals - zink: don't destroy old push layout when enabling fbfetch descriptor Mohamed Ahmed (12): - nvk: Dynamically allocate queues - nak: Fix 64-bit bit_count, ufind_msb, ifind_msb, find_lsb - nak: Enable lowering for bitfield manipulation at <32bit sizes - nvk: Ensure we have nvkmd before shader upload - nvk: Ensure we have nvkmd before sampler descriptor upload - nvk: Skip creating a nvkmd device if we don't have to - nvk: Add support for VK_QUERY_POOL_CREATE_RESET_BIT_KHR - nvk: Advertise VK_KHR_maintenance9 - nil: Add missing compressible PTE kinds - nouveau/headers: Add AMPERE_B compute subchannel definition - nouveau/mme: Add unit tests for sharing between compute and 3D scratch registers - nvk: Use the compute MME for compute dispatch Myrrh Periwinkle (1): - gallium: Properly handle non-contiguous used sampler view indexes Nagulendran, Iswara (3): - amd/vpelib: Fix Issues with Background Color insertions - amd/vpelib: Fix cost profiling support - amd/vpelib: Handle Destination Rect with zero dimensions Nanley Chery (18): - anv: Disable CCS if image bound to wrong heap on Xe2+ - anv: Disable fast-clears on linear surfaces - iris: Disable fast-clears on linear surfaces - iris: Add PIPE_BIND_SCANOUT when exporting textures - iris: Fix image reallocation for sharing - intel/isl: Only set CMF on renderable views on Xe2+ - intel: Enable CCS_E on linear surfaces on Xe2+ - iris: Drop iris_resource_image_is_pat_compressible - anv,hasvk: Take trace submission ID out of lock - anv: Rework locking for sparse binding with TR-TT - intel/isl: Define initial state of non-zeroed CCS on gfx9-11 - anv: Query ISL for the aux-state of undefined layouts - intel: Delete the has_illegal_ccs_values bool - intel/isl: Update the initial HiZ state for Xe2+ - intel/isl: Update the aux-state of zeroed HiZ - iris: Don't zero the CCS in an already zeroed BO - iris: Initialize HiZ to the CLEAR state on BDW-ICL - iris: Drop iris_resource_level_has_hiz() Natalie Vock (18): - radv/winsys: Support vm_always_valid in the NULL winsys - radv: Only expose indirect raytracing on gfx7+ - aco: Add RegisterDemand::operator!= - aco: Add function call attributes - aco: Add ABI and Pseudo CALL format - aco: Add call-related program/block properties - aco: Add call info - aco/lower_to_hw_instr: Lower calls - aco/live_var_analysis: Handle calls - aco/sched: Handle calls - aco/validate: Validate call instructions - aco/vn: Don't combine expressions across calls - aco/opt: Work around GCC compiler issue - aco/scheduler: Bail early on unreorderable instructions - vulkan/bvh: Mark instances with NAN AABBs as inactive - radv/bvh: Encode empty AS bounds as NaN - nir/lower_shader_calls: Repair SSA after wrap_instrs - radv: Fix PSO history with RT pipelines Nataraj Deshpande (1): - anv: add feature flags for linearly tiled ASTC images Okenczyc, Andrzej (1): - amd/vpelib: Move predication size calculation to bufs_req Olivia Lee (16): - panvk: stop CPU mapping all index buffers on JM - perfetto: allow specifying clock domain for cpu timestamps - panvk/perfetto: improve clock synchronization using CLOCK_MONOTONIC_RAW - editorconfig: move OpenCL configuration to root - vulkan: move internal vulkan pseudo-extensions to a common file - vulkan/util: add vk_topology_to_mesa helper function - hk: replace vk_conv_topology with vk_topology_to_mesa from vulkan/util - lavapipe: replace vk_conv_topology with vk_topology_to_mesa from vulkan/util - v3dv: replace vk_to_mesa_prim with vk_topology_to_mesa from vulkan/util - panvk: pass correct variant shader/compile inputs to panvk_lower_nir - pan/va: fix bi_is_imm_desc_handle early return - panvk: fix FS driver set layout when LD_VAR_BUF is disabled - vtn_bindgen2: use anonymous namespace to avoid name collisions - util/macros: coerce likely/unlikely to bool even without __builtin_expect - panfrost: fix cl_local_size for precompiled shaders - hk: fix data race when initializing poly_heap Paolo Bonzini (2): - meson: rename Rust subprojects to NAME-SEMVER-rs - docs: document naming convention for Rust subprojects Patrick Lerda (23): - dri: fix image_loader_extensions array - dri: complete the support for ARGB4444 - r600: refactor r600_is_buffer_format_supported() for the next update - r600: fix remaining pbo issues - r600: fix arb_shader_image_load_store incomplete - r600: refactor step 1 - r600_texture cast is replaced by a function - r600: refactor step 2 - r600_resource cast is replaced by a function - r600: refactor step 3 - split r600_framebuffer - r600: refactor step 4 - clean up r600_surface width0 and height0 elements - r600: refactor step 5 - evergreen clean up an incompatible mechanism - r600: refactor step 6 - pre-evergreen clean up - r600: refactor step 7 - split r600_surface - r600: refactor step 8 - pre-evergreen operations - r600: refactor step 9 - remove util_framebuffer_init - r600: refactor step 10 - drop create_surface - r600: refactor step 11 - change r600_aligned_buffer_create() return type - r600: fix evergreen gds atomic_counter_comp_swap - r600: fix r600_resource_copy_region behavior for some formats - r600: update multi_draw_indirect_params drm version requirement - r600: fix emit_ssbo_atomic_op when ssbo_image_offset is non-zero - r600: fix r600_draw_rectangle refcnt imbalance - r600: update nplanes support - r600: limit pre-evergreen predicate ready size Paul Gofman (1): - driconf: add a workaround for Investigation Stories : gunsound Paulo Zanoni (32): - brw: remove unnecessary inclusions - brw: store 'volatile' GLSL/SPIR-V access in MEMORY_LOGICAL_FLAGS - brw: consider 'volatile' memory access when doing CSE - brw: mark 'volatile' sends as uncached on LSC messages - brw: adjust comment pasted from a commit message - brw: remove unnecessary casts to unsigned after calling LSC_CACHE() - brw: null-tile sends don't need to skip L3 on Xe2 and newer - anv/sparse: don't claim Xe2's non-standard MSAA shapes as unsupported - anv/sparse: declare sparse MSAA block shapes as standard before Xe2 - anv/sparse: allow multiple sample bits in anv_sparse_image_check_support - anv/sparse: don't support depth/stencil with sparse - anv/sparse: we can support R64 and other atomics emulated formats - anv/sparse: call sparse_image_check_support from get_image_format_properties - zink: new expected failures for sparse depth buffers - intel: rework the way sparse forces CCS/MCS/HIZ to be disabled - isl: allow sparse with CCS on Xe2 and newer - isl: allow sparse with STC_CCS on DG2 - iris: fix indentation during command submission - iris/xe: move error checking to inside the devinfo->no_hw case - iris: devinfo->no_hw is unlikely - anv/i915: bring info->no_hw handling to anv_gem_execbuffer() - anv/xe: extract xe_exec_ioctl() - anv/xe: rework set_lost handling in xe_exec_ioctl() - anv/i915: rework set_lost handling in anv_gem_execbuffer() - anv/xe: set the queue as lost instead of the device on execbuf failure - anv: we never set I915_EXEC_FENCE_OUT - intel/i915: add i915_gem_execbuf_ioctl() - intel/i915: sleep a little bit between retries of the execbuf ioctl - intel/i915: give up the execbuf ioctl after ~16s of ENOMEMs - intel/i915: warn the user about repeated execbuf ENOMEM after ~2s - intel/xe: unify behavior with i915.ko regarding ENOMEM on DRM_IOCTL_XE_EXEC - intel: unify parameters for the exec ioctl retries Pavel Asyutchenko (1): - radv: report full sparse address space size Pavel Ondračka (5): - r300/ci: check gles2 extensions - r300/ci: add one recent flake - r300/ci: add RS740 piglit and dEQP testing - r300/ci: remove emulated swtcl testing - i915/ci: update CI expectations Peter Quayle (2): - pvr: various multiview fixes - pvr: add view index support for vertex shaders Philipp Zabel (1): - rusticl: Fix hidden lifetime warnings Pierre-Eric Pelloux-Prayer (31): - bufferobj: init the return value for GetParam functions - radeonsi/tests: enable vk interop testing - radeonsi: fix refcount with memobj - radeonsi/gfx12: dont use HTILE for imported textures - nir/lower_io: make sure range is not 0 - mesa/st: always use base_serialized_nir for draw - nir/opt_varyings: fix build with PRINT_RELOCATE_SLOT - mesa/st: check buf before dereferencing it - radeonsi/tests: update rasterpos results - radeonsi: sync harder on finish - radeonsi/sqtt: retry a frame capture after reiszing the buffer - radeonsi/sqtt: update the shader after scratch config - mesa: clear TransformFeedback.NumVarying on error - mesa: add u_overflow.h - util, vulkan: use u_overflow.h - nir/opcodes: use u_overflow to fix incorrect checks - nir/opcodes: remove invalid comment - glthread, tc: Fix buffer release with glthread and tc - st: add early to st_prune_releasebufs - tc: prevent flush of incomplete batches - tc: add debug code for tc_set_vertex_elements_for_call_pending - util: mimic KCMP_FILE via epoll when KCMP is missing - util: use F_DUPFD_QUERY on Linux - radeonsi/tests: use black to fix style issues - radeonsi/tests: allow to test radv - radeonsi/tests: add gfx11_5 to the list - radeonsi/tests: rename --no-xxx arguments - radeonsi/tests: rename glcts_path -> vk_gl_cts_path - radeonsi/tests: add an argument to specify a folder with the must pass files - radeonsi/tests: add a flag to specify a folder with the cts binaries - radeonsi: propagate shader updates for merged shaders Pohsiang (John) Hsu (11): - mediafoundation: change frame preanalysis rc from ifdef to runtime control - d3d12: Fix mediafoundation build - mediafoundation: fix deadlock when user call shutdown and endGetEvent concurrently - gallium/pipebuffer: fix multithread issue on pb_slab_manager_create_buffer - mediafoundation: periodic clang-format, no code changes - mediafoundation: update doc to remove gallium-vdpau from build setup - mediafoundation: return adjusted LTR frame (need to remove one for short term) - mediafoundation: create sample allocator for SW input sample on demand to save video memory - mediafoundation: periodic clang format - no code changes - mediafoundation: remove extra ';' - mediafoundation: update version to 1.07 Qiang Yu (103): - all: rename PIPE_SHADER_VERTEX to MESA_SHADER_VERTEX - all: rename PIPE_SHADER_TESS_CTRL to MESA_SHADER_TESS_CTRL - all: rename PIPE_SHADER_TESS_EVAL to MESA_SHADER_TESS_EVAL - all: rename PIPE_SHADER_GEOMETRY to MESA_SHADER_GEOMETRY - all: rename PIPE_SHADER_FRAGMENT to MESA_SHADER_FRAGMENT - all: rename PIPE_SHADER_COMPUTE to MESA_SHADER_COMPUTE - all: rename PIPE_SHADER_TASK to MESA_SHADER_TASK - all: rename PIPE_SHADER_MESH to MESA_SHADER_MESH - all: rename PIPE_SHADER_TYPES to MESA_SHADER_STAGES - all: rename PIPE_SHADER_MESH_TYPES to MESA_SHADER_MESH_STAGES - glsl: remove miss declaration of struct gl_shader_stage - all: rename gl_shader_stage to mesa_shader_stage - all: rename pipe_shader_type to mesa_shader_stage - mesa,gallium: remove pipe_shader_type_from_mesa - all: rename gl_shader_stage_is_compute to mesa_shader_stage_is_compute - all: rename gl_shader_stage_is_mesh to mesa_shader_stage_is_mesh - compiler: remove gl_shader_stage_is_graphics - all: rename gl_shader_stage_uses_workgroup to mesa_shader_stage_uses_workgroup - compiler: rename gl_shader_stage_is_callable to mesa_shader_stage_is_callable - all: rename gl_shader_stage_is_rt to mesa_shader_stage_is_rt - all: rename gl_shader_stage_can_set_fragment_shading_rate - all: rename gl_shader_stage_name to mesa_shader_stage_name - compiler,gallium: remove PIPE_SHADER_* and adjust some macro usage - gallium: add mesh shader caps - mesa,gallium: remove tgsi_processor_to_shader_stage - mesa/st: use shader_caps.max_instructions to check shader present - compiler: adjust comments for mesa_shader_stage - radeonsi: do not init nir_options for mesh shader - gallium/dd: enlarge shader string for mesh shader - mesa: enlarge the shader resourse limits for mesh shader - mesa: init program constants for mesh shader - glsl,gallium,mesa: replace MESA_SHADER_STAGES with MESA_SHADER_MESH_STAGES - mesa: set a more accurate value for combined limits - mesa: count mesh shader when init limits - mesa: add mesh shader extension state - nir/opt_varying: remove assert for mesh shader crash - nir: lower io support task and mesh shader - nir: compute io base for fragment shader inputs which maybe per primitive - Update OpenGL headers for GL_EXT_mesh_shader - mesa,mapi: add EXT_mesh_shader extension - mesa: implement EXT_mesh_shader glGet* values - mesa: implement EXT_mesh_shader glGetProgrameiv values - mesa: implement EXT_mesh_shader glGetActive* values - mesa,glsl: add mesh shader subrotine handling - mesa: implement mesh shader queries - mesa: support mesh shader when glCreateShader - mesa: remove mtype.h include from st_atom.h - mesa: fix glTexPageCommitmentARB and glTexturePageCommitmentEXT level check - mesa: use bitset for driver states tracker - gallium: cso context support mesh shader - mesa: add mesh shader states - mesa: handle mesh shader in state management - mesa: implement mesh shader draw calls - mesa,gallium: handle mesh shader create and delete - gallium: threaded context support mesh shader - gallium/u_blitter: save mesh shader - gallium/ddebug: support mesh shader - mesa: allow NULL for vertex shader when mesh pipeline - gallium/trace: dump mesh shader queries - mesa/st: convert mesh shader to gl stages - mesa: not fail the assert when detach mesh shader - mesa: program pipeline support mesh shader - gallium/noop: add mesh shader callbacks - panfrost: fix image plane array copy - panfrost: fix lowered multi plane resource offset/stride param get - ac/surface: refine supported modifier list for multi block size - ac/surface: add radeonsi exported modifiers to supported list - ac/surface: add ac_compute_surface_modifier - gallium: add PIPE_RESOURCE_PARAM_DISJOINT_PLANES - egl: refine dma buf export to support multi plane - radeonsi: really support eglExportDMABUFImageQueryMESA - mesa: fix draw mesh shader indirect buffer size check - radeonsi: fix use aco/llvm debug options - radeonsi: hide real modifier export behind AMD_DEBUG - glsl: prepare parse state for mesh shader - glsl: handle taskPayloadSharedEXT variables - glsl: handle PerPrimitiveEXT qualifier - glsl: allow shared variables in task and mesh shader - glsl: handle mesh shader primitive type layout qualifier - glsl: handle max_vertices/primitives for mesh shader - glsl: handle work group in layout for mesh shader - glsl: add input builtin variables for mesh shader - glsl: add mesh shader builtin outputs - glsl: assign mesh shader output variable array size - glsl: handle mesh shader output block - glsl: add mesh shader builtin functions - glsl: nir_build_program_resource_list support mesh shader - glsl: gl_nir_link_glsl handle mesh shader - glsl: validate MS/FS interstage in/out block - glsl: handle per primitive varying when link - glsl: validate MS/FS interstage in/out variable type - glsl: disable mesh shader output remove when separate shader - glsl: pack vertex pipeline varying linkage into a function - glsl: pack varying limit check code into functions - glsl: add mesh pipeline varying linkage - glsl: handle mesh shader when optimize varying - glsl: handle explicit location for mesh shader - glsl: lower shared and task playload for mesh shader - glsl: no xfb buffer qualifier for mesh shader - glsl: flat qualifier is not needed for per primitive IO - glsl: translate mesa stage for mesh shader - glsl: allow barrier builtin functions for mesh shader - gallium: fix eglExportDMABUFImageQueryMESA crash for r600 Quentin Schulz (3): - nvk: remove unused relative_dir variable - meson: replace global_source_root/global_build_root with project_* - meson: fix libcl assert() reproducibility Renato Pereyra (1): - anv: Enable anv_emulate_read_without_format for Android 15+ Rhys Perry (107): - aco/lower_phis: add bld_before_logical_end helper - nir/divergence: ignore boolean phis for ignore_undef_if_phi_srcs - aco: optimize s_and(s_cselect, exec) - aco: stop labeling first def of and(uniform_bool/uniform_bitwise, exec) - aco: don't both flip s_cselect and label uniform_bool - aco/opt: add some comments - aco: optimize uniform s_not - aco/isel: optimize uniform vote - nir/cf: have nir_remove_after_cf_node remove phis at the start too - nir/search: check variable requirements even if it's already seen - nir/uub: fix 8/16-bit overflow - nir/opt_access: support RT/callable shaders - nir/load_store_vectorize: check for interfering shared2 before vectorizing - nir/load_store_vectorize: set is_store for shared append/consume - nir/load_store_vectorize: always set num_components correctly - glsl_to_nir,vtn: insert barriers around begin/end invocation interlock - ac/nir/lower_ps: remove barrier for end_invocation_interlock - aco/gfx12: fix printing of temporal hints - aco: align scratch size after isel - aco: fix possible scratch offset overflow - vtn: fix placement of barriers for MakeAvailable/MakeVisible - nir: don't move accesses across make visible/available barriers - vtn: remove acquire/release around make visible/available barriers - nir/lower_memory_model: remove empty lowered barriers - aco/ra: set late-kill for operands of temporary p_create_vector - nir: add global_amd to nir_get_io_offset_src/nir_get_io_index_src - nir/opt_load_skip_helpers: move divergence check earlier - nir/opt_load_skip_helpers: always require helpers for handles - nir/search: add nir_search_state - nir/search: don't clear empty hash tables - nir/search: reorder match_value to check constants first - nir: add nir_def_num_lsb_zero - nir/algebraic: improve is_unsigned_multiple_of_4 and use it more - nir/algebraic: allow non-const for iand(iadd()) -> iadd(iand()) - nir/load_store_vectorize: use nir_def_num_lsb_zero in check_for_robustness - nir/load_store_vectorize: use nir_def_num_lsb_zero in calc_alignment - device-select: clang-format - device-select: move get_default_device to it's own file - device-select: simplify adding/removing instances - device-select: do all getenv during instance creation - device-select: use debug_get_bool_option for FORCE_DEFAULT_DEVICE - device-select: refactor device_select_get_default - nir/divergence: make smem load_global_amd uniform - drm-shim: use atomics for inited - drm-shim: fix with asan - aco: fix signed integer overflow - radv: fix shift overflow in radv_pipeline_init_dynamic_state - vtn: use vtn_has_decoration more - nir/load_store_vectorize: refactor offset parsing - nir/load_store_vectorize: refactor entry key creation - nir/load_store_vectorize: call nir_def_num_lsb_zero less - nir/load_store_vectorize: optimize accesses with u2u64(ishl.nuw(iadd)) - nir/opt_offsets: report progress if NUW is set - nir/opt_offsets: fix progress determination with offsets that add to zero - nir/opt_offsets: improve shared2 optimization - nir/load_store_vectorize: remove offset check in try_vectorize_shared2 - aco: reduce cost of using values defined in predecessors - aco: add is_atomic_or_control_instr helper - aco: don't move release barriers after interlock end - aco: don't move acquire barriers before interlock begin - aco: refactor waitcnt pass to use barrier_info - aco: add a separate barrier_info for release/acquire barriers - aco: delay barrier waitcnt until they are needed - aco: remove waitcnt code for SMEM stores - aco: remove waitcnt code for POPS - aco: update waitcnt events for exports - aco: use a separate event for sendmsg_rtn - aco: fix workgroup-scope barrier between vmem and lds - aco/gfx10: skip waitcnts or use vm_vsrc(0) for workgroup vmem barriers - aco/gfx10: skip waitcnts or use vm_vsrc(0) for workgroup lds barriers - aco/tests: add barrier-to-waitcnt tests - aco: avoid wraparound for smem global loads with both offsets - aco: avoid unaligned offsets when selecting load_global_amd - zink/ntv: fix coherent image load/store - vtn: skip make-available/visible for shared - zink/ntv: use MakePointerAvailable/Visible for shared load/store - nir/lower_atomics_to_ssbo: set ACCESS_COHERENT for loads - nir/lower_atomics: set ACCESS_COHERENT - aco: workaround load tearing for load_shared2_amd - aco: fix SGPR 8-bit nir_op_vec with mixed constant and non-constant - ac/nir: fix progress reporting in ac_nir_lower_tex - nir: fix progress reporting in nir_io_add_const_offset_to_base - radv: fix progress reporting in lower_rt_derefs - nir/opt_if: fix progress reporting with multiple function impls - nir/opt_if: rewrite progress reporting and metadata invalidation - nir: fix NIR_DEBUG=extended_validation - nir: add NIR_DEBUG=progress_validation - rusticl: support NIR_DEBUG=invalidate_metadata/extended_validation - rusticl: support NIR_DEBUG=progress_validation - aco: remove buffer_load_lds instructions - nir: add ACCESS_ATOMIC - vtn: set ACCESS_ATOMIC - zink/ntv: use ACCESS_ATOMIC - nir,vtn: add shader_info::assume_no_data_races - nir: assume non-atomic loads don't tear - aco: only workaround load tearing for atomic loads - aco: set atomic semantic for atomic load/store - aco: remove barrier acquire/release workaround - aco: use MTBUF for 64-bit atomic load/store - radv: move nir_opt_algebraic loop for NGG culling earlier - radv: only call radv_should_use_wgp_mode() once - radv: use CU mode when LDS is used - radv: allow WGP mode with task/mesh - amd/lower_mem_access_bit_sizes: don't create subdword UBO loads with LLVM - amd/lower_mem_access_bit_sizes: improve subdword/unaligned SMEM lowering - amd/lower_mem_access_bit_sizes: be more careful with 8/16-bit scratch load - amd/lower_mem_access_bit_sizes: fix shared access when bytes - freedreno/drm-shim: Handle GET/SET_METADATA - freedreno/registers: Add a way to disable deprecated warnings - freedreno/registers: Generate variant builder always - freedreno/a6xx: Convert to variant reg packers - freedreno/computerator: Convert to variant reg packers - freedreno/registers: Fix variant ranges - freedreno/registers: Add implicit reg32 for empty arrays - freedreno/registers: De-open-code some offsets - freedreno/registers: Cleanup the bin_cntl's - freedreno/registers: Move descriptor related enums - freedreno/registers: Prep for upcoming things - freedreno/registers: Make TPL1_BICUBIC_WEIGHTS_TABLE an array - freedreno: Name a few events - freedreno/a6xx: Drop VPC table magic - freedreno/a6xx: Require write support for images - freedreno/a6xx: Disallow impossible image swizzles - freedreno/a6xx: Mark tex and samp descriptors for dumping - freedreno/a6xx: Format table fixes - nir/lower-amul: Fix crash with unused SSBO - nir/lower-amul: Comment fix - freedreno/registers: Add A7XX_CX_DBGC - freedreno/registers: Re-enable validation for gen_header.py - freedreno/registers: Remove license/etc from generated headers - freedreno/registers: remove python 3.9 dependency for compiling msm - freedreno/registers: Generate _HI/LO builders for reg64 - freedreno/registers: Update GMU register xml - freedreno/a6xx: Fallback to original blit in the snorm_copy path - freedreno/blitter: Don't ignore blit swizzle - freedreno/a6xx: Add missing format - freedreno/a6xx: Fix snorm rounding - freedreno/devices: Update chicken bits - freedreno/decode: Add test to check for conflicting regs - freedreno/registers: Remove conflicting RBBM regs - freedreno/registers: Fix x_CONTEXT_SWITCH_GFX_PREEMPTION_SAFE_MODE - freedreno/decode: checkreg handling for bitsize/stride - freedreno/decode/scripts: Add license comments - freedreno/fdl: Set pitch for buffers - freedreno/a6xx: Drop arbitrary import restrictions - freedreno: Handle buffer import - freedreno: Always use aux-ctx for export blits - freedreno: Allow TC async fences to have an fd - freedreno: Disable explicit sync heuristic for Xwayland - freedreno/a6xx: Move reg to static-non-context - freedreno/decode/crashdec: Limit snapshot BO size - freedreno/afuc: Add missing varset check - freedreno/registers: More register prep - freedreno/registers: Rename some unknowns - freedreno/registers: x_ADDR_MODE_CNTL is a6xx and earlier - freedreno/registers: Fix a couple reg names - freedreno/registers: Extract out bitset for roq_avail - freedreno/decode: Add gen8 support - freedreno/decode: Move enum lookup out of snapshot - freedreno/registers: Common-ize PIPE definitions - freedreno/registers: Add gen8 regs - freedreno/registers: Add gen8 descriptor layout - freedreno/registers: pm4 updates for gen8 - freedreno/a6xx: Slight re-org of sampler descriptor building - freedreno/layout: Convert fd6_view to c++ - freedreno/layout: gen8 descriptor support Rob Hughes (1): - llvmpipe: Work around WSL 1 missing support for memfd_create() Robert Mader (8): - anv: Enable G8_B8_R8_3PLANE_422 and G8_B8_R8_3PLANE_444 formats - gallium: Set and count all extra samplers - mesa: Add support for NV61, NV24 and NV42 pixel formats - panfrost: Add lowerings for the NV61, NV24 and NV42 pixel formats - nir: Fixup 10/12 bit SW decoder YCbCr formats - sw_winsys: Add winsys_handle to displaytarget_create_mapped - kms-dri-sw: Implement create_mapped() - kms-dri-sw: Report linear modifiers in get_handle() Rohan Garg (1): - intel/compiler: use the WA framework when emitting WA 14014595444 Rohit Athavale (6): - mediafoundation: Add guids for the newly added Input Delta QP & Absolute QP APIs - mediafoundation: Add IsSupported() & GetValue() for CODECAPI_AVEncVideoInputDeltaQPBlockSettings - d3d12: Make delta QP min and max to be bit-depth dependent for HEVC - pipe: Add pipe_enc_qpmap_input_info to contain GPU & CPU QP Maps - d3d12: Update d3d12 back to use pipe_enc_qpmap_input_info - mediafoundation: Lock QP Map Buffer when in use, unlock after Roland Scheidegger (13): - llvmpipe: minor cleanup - llvmpipe: Fix array mismatch when accessing shader images - llvmpipe: Fix attribute interpolation setup when rendering lines with msaa - llvmpipe: Fix wrong pixel shader invocation count with discard - llvmpipe: Fix wrong GS invocation count when using instanced GS - llvmpipe: add bitcasts around fptrunc/fpext operations - docs: fix up old comment about fake msaa for llvmpipe - lavapipe: don't leak the temporary msaa resource - llvmpipe: fix incorrect scissor planes - lavapipe: expose support for msaa 8x - gallium,mesa/st: reverse logic for y flip for programmable sample locations - llvmpipe: implement GL_ARB_sample_locations - lavapipe: implement VK_EXT_sample_locations Romaric Jodin (11): - pan/bi: use only 1 MKVEC.v2i8 to generate v4i8 when possible - pan/va: improve lowering of SWZ_V4I8 - pan/bi: add pass to simplify control flow - pan/bi: schedule simple iterators to avoid extra move - panfrost/perfetto: Use Android-internal perfetto - meson: remove '--outdir' argument in script - meson: add vk_enum_defines.h to idep_vulkan_util_headers - meson: add depend_files for gl_enums.py - meson: update xml files list in mesa/glapi - meson: sort xml files in mesa/glapi - glapi: static_data: do not use __file__ to get gl symbols file Ruijing Dong (2): - radeonsi/vcn: vcn5 av1 decoding context buffer fix - radeonsi/vcn: Correct a typo condition for jpeg decoding Ryan Houdek (1): - freedreno/fdl: Fix typo in tiled_to_linear_2cpp Sagar Ghuge (24): - intel/genxml: Update CS_CHICKEN1 register field - anv: Use thread group preemption granularity - vulkan/radix_sort: Fix subgroup invocation id - anv: Use vk_get_bvh_build_pipeline_spv helper - vulkan/runtime: Add VK_SHADER_CREATE_UNALIGNED_DISPATCH_BIT_MESA flag - anv: Mask off excessive invocations - intel/genxml: Drop all unused struct/fields - intel/compiler: Fix ray geometry index - anv: Add missing ACCELERATION_STRUCTURE_READ in barrier handling - anv: Enable CS stall for ACCELERATION_STRUCTURE_COPY stage - anv: Add missing L3 flushes - anv: Apply pipe flushes for outstanding PC bits - anv: Emit state cache invalidation after every compute dispatch - blorp: Emit state cache invalidation after every compute dispatch - iris: Emit state cache invalidation after every compute dispatch - isl: Respect driconf option for EnableSamplerRoutetoLSC - Revert "intel: Always set Cube Face Enables for all surfaces." - anv: Call brw_nir_lower_rt_intrinsics_pre_trace lowering pass - brw/rt: Move nir_build_vec3_mat_mult_col_major helper to header - brw/rt: fix ray_object_(direction|origin) for closest-hit shaders - vulkan/runtime: Fix typo in stack size calculation - anv: Use correct engine class for companion RCS - anv: Drop unwanted untyped flush for AS query - intel/common: Consider 0 threads while setting TG Samuel Pitoiset (352): - Revert "ci: Disable Valve keywords farm" - radv: adjust conservative rasterization configuration on GFX12 - radv: use vk_optimize_depth_stencil_state() for optimal settings - radv: add RADV_DEBUG=novideo to disable all video extensions - radv: fix SQTT shaders relocation on GFX12 - radv: simplify emitting SQTT shaders relocation for GFX6-GFX11.5 - radv: fix reporting instance/vertex_count for direct draws with RGP on GFX12 - radv: reject 1D block-compresed formats with mips on GFX6 - zink/ci: update list of expected failures for NAVI31 - zink/ci: remove old gfx1200 lists - radv/ci: fix list of expected failures for VEGA10/NAVI10 - radv: fix a memleak with GS copy shader NIR - radv: emit PGM_HI_PS in the gfx preamble on GFX12 - radv: remove dead ES emit code on GFX12 - radv: invalidate compute/rt descriptors at pipeline bind time - radv: stop passing compute shader to radv_dispatch() - radv: rework graphics shaders/vbos prefetch sligthly - radv: handle compute/rt prefetch like graphics - radv: add radv_{before,after}_dispatch() functions - radv: replace DGC before/after dispatch helpers with the new ones - radv: fix fbfetch output with compresed FMASK on <= GFX9 - vulkan: fix missing presentId2/presentWait2 enable features - docs: add missing VK_KHR_present_id/2 to features.txt - ci: uprev VKCTS main to 9dd9a72b28218f1ca12777d9b73c2a85c5c60231 - ac/gpu_info,radv: use the maximum virtual address from the kernel - radv: invalidate compute/rt descriptors at dispatch time - zink/ci: skip spec\@arb_fragment_program\@fog-modes on RADV - radv/ci: fix GPU hang detection regex with recent kernels - zink/ci: reduce timeout of zink-radv-navi31-valve - zink/ci: make zink-radv-navi31-valve a pre-merge job - radv: precompute the mask for enabled color writes - radv: precompute the mask for color write attachments - radv: precompute color blend equations - radv: track more CB related context registers on < GFX12 - radv: regroup CB related states emission together - radv: tidy up radv_device_init_perf_counters() - radv: introduce radv_cmd_stream - radv: switch to radv_cmd_stream everywhere - radv: move buffered registers for GFX12 to radv_cmd_stream - radv: move context_roll_without_scissor_emitted to radv_cmd_stream - radv: move tracked registers to radv_cmd_stream - radv/ci: uprev kernel to 6.15.9 - radv: cleanup some redundant cmd_buffer->cs occurrences - radv: remove cs parameter for all opt context emit helpers - radv: remove cs parameter for gfx12 push SH reg helpers - radv: implement RB+ depth-only rendering for better perf - radv: fix destroying CS with RADV_PERFTEST=dmashaders - ac,radv,radeonsi: fix programming PA_SU_PRIM_FILTER_CNTL on GFX12 - radv/amdgpu: fix creation with different but unused RADV_PERFTEST flags - ac/descriptors: add a function to create a descriptor for HiZ surfaces - radv: allocate image metadata to implement a workaround for HiZ on GFX12 - radv: add a function to create an image view for HiZ surfaces - radv/meta: add a pass to clear HiZ surfaces - radv: initialize HiZ metadata during image layout transitions - radv/meta: update HiZ metadata after depth/stencil image clears - radv: validate dynamic states earlier - radv: implement an alternative workaround for HiZ on GFX12 - radv: fix reserving space for emitting push constants with DGC IES - radv: remove redundant push constant size alignment for DGC - radv: pass the IES struct when computing the DGC sequence size - radv: pre-compute more information when updating DGC IES - radv: optimize the preprocess buffer size for DGC IES compute - radv: use radv_write_sampler_descriptor() for combined image/sampler - radv: do not hardcode the combined image/sampler offset in the db path - radv: only write 32 bytes for combined image/sampler on GFX11+ - radv: reduce the combined image/sampler desc size on GFX11+ - radv: remove useless inline push constant emission with DGC IES - radv: stop using the pipeline layout for inlined push constants with DGC - radv: split uploading push constants with DGC in two parts - radv: stop using the pipeline layout for uploading push constants with DGC - radv: tidy up radv_flush_descriptors() - radv: slightly optimize indirect descriptor sets upload size - radv: invalidating push constants for compute<->rt during dispatches - radv: do not emit inlined SGPRs twice for merged shaders - radv: use radv_shader_need_indirect_descriptor_sets() more - radv: determine if push constants need to be uploaded earlier - radv: rework emitting push constants for less CPU overhead - radv: add a function that uploads push constants - radv: remove unused forwarded declarations of pipeline layout - radv: determine the push constant size from the shader itself - radv: add a function to get push constant layout info for DGC - radv: gather push constant size from shaders for DGC - radv: stop using the pipeline layout completely for DGC - radv: fix color attachment remapping with fast-GPL/ESO - radv: merge two similar loops in lookup_ps_epilog() - Revert "radv/ci: disable hang detection in navi31-vkcts" - zink/ci: skip one piglit subset that randomly hangs on RADV - zink/ci: update list of flakes for NAVI31/VANGOGH/CEZANNE - amd/drm-shim: add navi33 - radv: emit relocation for task shaders at the same place as other stages - radv: rework the helper to emit buffered regs on GFX12 - radv: emit compute pipeline with buffered SH regs on GFX12 - radv: emit descriptor pointers with buffered SH regs on GFX12 - radv: emit inlined push constants with buffered SH regs on GFX12 - radv/ci: update expected list of failures/flakes on GFX1201 - radv/ci: use 3 parallel jobs for radv-gfx1201-vkcts - radv/ci: reduce the timeout for radv-gfx1201-vkcts - radv/ci: make radv-gfx1201-vkcts a pre-merge job - radv/ci: document a very recent ACO regression on GFX12 - zink/ci: make zink-radv-gfx1201-valve a pre-merge job - zink/ci: update list of flakes for GFX1201 - radv: get the depth clamp mode earlier when emitting viewports - radv: emit depth clamp enable as part of the viewport state - radv: add a new dirty bit for the viewport state - radv: precompute the depth clamp mode - radv: precompute the depth clip enable - radv: dirty some states from graphics pipeline earlier - radv: do not emit few RADV_CMD_DIRTY_xxx based on dynamic states - radv: only re-emit needed states when PS inner coverage changes - radv: add a new dirty bit for the binning state - radv: optimize re-emitting the occlusion query state on GFX12 - radv: validate dynamic states for the occlusion query state earlier - radv: validate dynamic states for the db shader control state earlier - radv: add a new dirty bit for the ngg culling state - radv: add a new dirty bit for the FSR state - radv: add a new dirty bit for the rast samples state - radv: rename RADV_CMD_DIRTY_TESS_STATE to RADV_CMD_DIRTY_TCS_TES_STATE - radv: add a new dirty bit for the depth bias state - radv: dirty the depth stencil state when rendering begins - radv: dirty the cb render state when rendering begins - radv: dirty more states when rendering begins - radv: add a new dirty bit for the VS prolog state - radv: add a new dirty bit for the blend constants state - radv: add a new dirty bit for the sample locations state - radv: add a new dirty bit for the scissor state - radv: make radv_cmd_state::dirty a 64-bit field - radv: add missing L2 invalidate cache flush for non-coherent images - radv: add a new dirty bit for the tess domain origin state - radv: add a new dirty bit for the patch control points state - radv: add a new dirty bit for the VGT prim state - radv: remove radv_cmd_buffer_flush_dynamic_state() - radv: remove dead code when setting dynamic primitive topology - radv: dirty the rast sample states for VRS att/OOO rast - radv: dirty RADV_CMD_DIRTY_xx states when binding sample shading state - radv: dirty the rast samples state when VRS is forced to 1x1 - radv: rename rast_prim to vgt_outprim_type everywhere - radv: stop abusing dirty_dynamic when binding a NULL fragment shader - radv: clear RADV_CMD_DIRTY_xxx bits outside of the caller in most cases - radv: fix hashing graphics pipeline when no stages are compiled - radv: run nir_lower_memcpy after spirv->nir - radv: run nir_opt_memcpy before nir_opt_copy_prop_vars - radv/nir/lower_cmat: handle untyped pointers for load/store - radv: advertise VK_KHR_shader_untyped_pointers - radv: clear RADV_CMD_DIRTY_xxx bits outside of the caller in more cases - radv: handle fbfetch output after binding graphics shaders - radv: clear descriptors state dirty bit outside of the caller - radv: add a new state for forced VRS rates - radv: check if SQTT is enabled before calling radv_describe_draw() - radv: check flush_bits before calling radv_emit_cache_flush() in the draw path - radv: add radv_cmd_set_line_width() - radv: add radv_cmd_set_tessellation_domain_origin() - radv: add radv_cmd_set_patch_control_points() - radv: add radv_cmd_set_depth_clamp_range() - radv: add radv_cmd_set_depth_clip_negative_one_to_one() - radv: add radv_cmd_set_primitive_restart_enable() - radv: add radv_cmd_set_depth_bias() - radv: add radv_cmd_set_line_stipple() - radv: add radv_cmd_set_cull_mode() - radv: add radv_cmd_set_front_face() - radv: add radv_cmd_set_depth_bias_enable() - radv: add radv_cmd_set_rasterizer_discard_enable() - radv: add radv_cmd_set_polygon_mode() - radv: add radv_cmd_set_line_stipple_enable() - radv: add radv_cmd_set_depth_clip_enable() - radv: add radv_cmd_set_conservative_rasterization_mode() - radv: add radv_cmd_set_provoking_vertex_mode() - radv: add radv_cmd_set_depth_clamp_enable() - radv: add radv_cmd_set_line_rasterization_mode() - radv: add radv_cmd_set_alpha_to_coverage_enable() - radv: add radv_cmd_set_alpha_to_one_enable() - radv: add radv_cmd_set_sample_mask() - radv: add radv_cmd_set_rasterization_samples() - radv: add radv_cmd_set_sample_locations_enable() - radv: add radv_cmd_set_depth_bounds() - radv: add radv_cmd_set_stencil_compare_mask() - radv: add radv_cmd_set_stencil_write_mask() - radv: add radv_cmd_set_stencil_reference() - radv: add radv_cmd_set_logic_op() - radv: add radv_cmd_set_color_write_enable() - radv: add radv_cmd_set_color_write_mask() - radv: add radv_cmd_set_logic_op_enable() - radv: add radv_cmd_set_fragment_shading_rate() - radv: add radv_cmd_set_attachment_feedback_loop_enable() - radv: add radv_cmd_set_primitive_topology() - radv: add radv_cmd_set_blend_constants() - radv: add radv_cmd_set_discard_rectangle_mode() - radv: add radv_cmd_set_discard_rectangle_enable() - radv: add radv_cmd_set_depth_test_enable() - radv: add radv_cmd_set_depth_write_enable() - radv: add radv_cmd_set_depth_compare_op() - radv: add radv_cmd_set_depth_bounds_test_enable() - radv: add radv_cmd_set_stencil_test_enable() - radv: add radv_cmd_set_stencil_op() - radv: add radv_cmd_set_discard_rectangle() - radv: make use of RADV_DYNAMIC_{VIEWPORT,SCISSOR}_WITH_COUNT - radv: add radv_cmd_set_viewport_with_count() - radv: add radv_cmd_set_scissor_with_count() - radv: add radv_cmd_set_scissor() - radv: add radv_cmd_set_viewport() - radv: make radv_ps_epilog_state::color_blend_enable a 8-bit field - radv: pre-compute color blend enable - radv: add radv_cmd_set_color_blend_enable() - radv: add radv_cmd_set_rendering_attachment_locations() - radv: add radv_cmd_set_rendering_input_attachment_indices() - radv: add radv_cmd_set_sample_locations() - radv: add radv_cmd_set_color_blend_equation() - radv: only update vertex stride if pStrides is non-NULL when binding VBO - radv: use the dynamic state to store vertex binding strides - radv: bind the vertex binding strides like a normal dynamic state - radv: move radv_vertex_input_state to radv_pipeline_graphics.h - radv: move VBO misaligned/unaligned info to radv_vertex_input_state - radv: remove unused parameter to radv_pipeline_init_dynamic_state() - radv: use the dynamic state to store vertex input state - radv: replace an assertion with a check when emitting VS prolog - radv: bind the vertex input state like a normal dynamic state - radv: fix setting VBO misaligned mask in graphics pipelines - radv: allow to select a different HiZ workaround on GFX12 - radv: add RADV_GFX12_HIZ_WA to select the HiZ wa behavior on GFX12 - radv: rename NGG culling user SGPRs - radv: split RADV_CMD_DIRTY_NGGC_STATE in two states - radv: clear dynamic states earlier - radv: use radv_get_vgt_outprim_type() to disable NGGC for points/lines - radv: use radv_get_vgt_outprim_type() for the NGG SGPRs state - radv: add an early return to radv_flush_vertex_descriptors() - radv: emit BREAK_BATCH when the PS changes also for ESO - radv: cleanup configuring AUTO_RESET_CNTL - radv: dirty the raster state when setting the primitive topology - radv: pre-compute tessellation num patches/lds size earlier - radv: do not trigger PATCH_CONTROL_POINTS_STATE on GFX12 - radv: rename DIRTY_PATCH_CONTROL_POINTS_STATE to DIRTY_LS_HS_CONFIG - radv: remove unnecessary ternary expressions in radv_emit_depth_stencil_state() - radv: translate stencil op earlier - radv: fix compiler warnings when uploading cmdbuf data might fail - radv: remove unused radv_pipeline::user_data_0 - radv: remove set but unused has_nggc in radv_cmd_state - radv: remove set but unused radv_graphics_pipeline fields - radv: remove unnecessary radv_graphics_pipeline::is_ngg - radv: disable VK_EXT_image_compression_control on GFX12 - radv/rt: only use one user SGPR for the traversal shader addr - radv/rt: fix a potential issue with RADV_PERFTEST=dmashaders - radv/ci: remove RADV_DEBUG=novideo for radv-gfx1201-vkcts - radv: mark RADV_DEBUG=nodynamicbounds as deprecated - radv: mark RADV_DEBUG=invariantgeom as deprecated - radv: mark RADV_DEBUG=splitfma as deprecated - radv: mark RADV_DEBUG=nongg_gs as deprecated - radv: move drirc options to a separate struct - radv: move features related drirc to radv_drirc::features - radv: move performance related drirc to radv_drirc::performance - radv: move debug related drirc to radv_drirc::debug - radv: move misc related drirc to radv_drirc::misc - radv: fix vk_error in radv_update_preambles() - radv/amdgpu: add a function to query permitted context priorities - radv: only expose permitted global queue priorities - radv: rework the optimal packet order for "normal" draws - radv: rework the optimal packet order for task/mesh draws - radv: rework the optimal packet order for dispatches - radv: rename radv_flush_occlusion_query_state() - radv: simplify sample shading state tracking - radv: determine which shader is the last VGT shader using next stage - radv: trigger VS related states in radv_bind_pre_rast_shader() - radv/meta: use radv_CmdDispatchBase() directly for ASTC decode - radv: add small helper to dispatch RT - radv: remove unnecessary NULL check when creating PS epilogs - radv: add a function to bind a PS epilog - radv: add a new dirty bit for compiling/binding a PS epilog - radv: add a new dirty bit for emitting a PS epilog - radv: rename RADV_CMD_DIRTY_FS_STATE to RADV_CMD_DIRTY_PS_STATE - radv: exclude dynamic vertex input stride for the late scissor workaround - radv/amdgpu: return OOM device when BO mapping fails - radv/amdgpu: add more helpers for managing virtual BOs - radv: add RADV_DEBUG=bo_history - Revert "radv: handle fbfetch output after binding graphics shaders" - radv: emit more push shader registers on GFX12 - radv: report an message when RADV_GFX12_HIZ_WA value is invalid - radv: replace RADV_GFX12_HIZ_WA by a drirc option - radv: switch to the full HiZ workaround by default on GFX12 - radv: disable radv_disable_hiz_his_gfx12 for Mafia Definition Edition - radv: set radv_gfx12_hiz_wa=partial for some games to mitigate performance loss - zink/ci: mark one test as crash/flake for turnip a618 - radv: get NIR options after initializing the physical device cache key - radv: fix capture/replay with sampler border color - spirv: add missing non-uniform access for SSBO atomics - radv/meta: fix saving push constants for depth/stensil resolves on compute - radv/meta: rework depth/stencil resolves using compute - radv/meta: rework depth/stencil resolves using graphics - radv/meta: remove useless VK_ACCESS_2_SHADER_WRITE_BIT for subpass resolves - radv/meta: simplify barriers for resolves - radv/meta: simplify calling depth/stencil resolve helpers - radv/meta: remove useless assertion when choosing resolve method - radv: pre-compute the number of rasterization samples - radv: pre-compute the line rasterization mode - radv: pre-compute vgt_outprim_type - radv: remove redundant RADV_DYNAMIC_PRIMITIVE_TOPOLOGY - radv: remove redundant RADV_DYNAMIC_LINE_RASTERIZATION_MODE - radv: remove redundant RADV_DYNAMIC_POLYGON_MODE - radv: remove redundant RADV_DYNAMIC_RASTERIZATION_SAMPLES - radv: set DRLR mapping info from inheritance info when present - radv: add a helper whether shader fp16 is enabled - radv/ci: document recent unexpected failures on TAHITI - Revert "radv/ci: document recent unexpected failures on TAHITI" - radv: only expose AMD_device_coherent_memory if actually supported - radv: reserve more CS space when executing DGC calls - radv/ci: update expected list of failures for VEGA10/NAVI10 - radv: lower ycbcr tex instructions earlier - radv: lower embedded/immutable samplers earlier - radv: fix expected disk cache size for meta shaders - nir: adjust nir_tex_instr_need_sampler() for AMD FMASK instructions - radv: remove useless radeon_cmdbuf forwarded declaration - ac/sqtt: use void pointers for start/stop CS - ac/cmdbuf: introduce ac_cmdbuf - radeonsi: replace radeon_cmdbuf_chunk by ac_cmdbuf - radv: replace radeon_cmdbuf by ac_cmdbuf completely - radv,radeonsi: use new ac_cmdbuf macros - radv: do not initialize HiZ on transfer queue on RDNA4 - radv: use force_indirect_desc_sets when creating RT prologs - radv: rename indirect_descriptor_sets to indirect_descriptors - radv: rename shader arg descriptor_sets to descriptors - radv: make radv_descriptor_get_va() a static function - radv: rename radv_mark_descriptor_sets_dirty() - ac/surface: fix host image copies with 96-bits formats - ac/surface: fix host image copies with stencil-only - radv: allow VK_FORMAT_S8_UINT with host image copy - vulkan/runtime: fix memleak when creating ETC pipelines - radv/rt: fix memory leak in lower_rt_instructions_monolithic() - radv: fix shaders memleak when importing pipeline binaries with GPL - radv/meta: pass image formats to radv_meta_resolve_{hardware,fragment}_image() - radv/meta: re-use radv_meta_resolve_{fragment,hardware}_image() for subpass resolves - radv/meta: pass iview formats for subpass resolves - radv/meta: remove radv_cmd_buffer_resolve_rendering_{hw,cs,fs} - radv: enable the global BO list by default - radv: only return identicalMemoryLayout for linear images - radv: always return optimalDeviceAccess=TRUE for block-compressed formats - radv: declare a new user SGPR for dynamic descriptors - radv: upload and emit dynamic descriptors separately from push constants - radv: allow to inline all push constants even with dynamic descriptors - radv: use COPY_DATA_DST_MEM when writing timestamps - amd,radv: add ac_emit_cond_exec() - amd,radv: add ac_emit_write_data_imm() - amd,radv,radeonsi: add ac_emit_cp_wait_mem() - amd,radv,radeonsi: add ac_emit_cp_acquire_mem_pws() - amd,radv,radeonsi: add ac_emit_cp_release_mem_pws() - radv: use ac_emit_cp_{acquire,release}_mem_pws() when syncing GE rings - amd,radv,radeonsi: add ac_emit_cp_copy_data() - amd,radv,radeonsi: add ac_emit_cp_pfp_sync_me() - ci: uprev VKCTS main to db48c34bebaf3359453e44ab151a2ff9f9c58eb2 - radv/ci: bump timeout for radv-gfx1201-vkcts to 5 minutes more - radv: dirty dynamic descriptors when required - radv: ignore dual-source blending when blending isn't enabled for MRT0 - radv: add a workaround for illegal depth/stencil descriptors with No Man's Sky - aco: fix reserving VGPRs for 64-bit attributes in VS prologs - radv,aco: wait for all VMEM loads when the prolog loads large 64-bit attributes - radv: add vk_wsi_disable_unordered_submits and enable for GTK Serdar Kocdemir (2): - gfxstream: fix warnings about unused parameters - gfxstream: Enable VK_MVK_macos_surface for host dispatch Sergi Blanch Torne (19): - ci: fix gc2000 fails duplication - ci,crnm: migrate colorama to rich - Revert "ci: Temporarily hardcode S3 artifact path" - Revert "ci: Fix for GitLab 18.2.2 upgrade" - ci: disable Collabora's farm due to maintenance - ci: fix requirements file - Revert "ci: disable Collabora's farm due to maintenance" - ci,marge_queue: encapsulate monitor loop - ci,marge_queue: enhance script interruption - ci,marge_queue: objects to represent the queue - ci,marge_queue: refactor the get queue method - ci,marge_queue: protect form transient errors - ci,marge_queue: encapsulate GitLab module queries - ci,marge_queue: queue element formatting - docs,marge_queue: document the tool usage - ci,marge_queue: handle GitLab auth exception - ci,marge_queue: use rich module - ci,marge_queue: introduce testing - ci: Add missing aiohttp Python dependecy Sergi Blanch-Torne (3): - ci: disable Collabora's farm due to maintenance - Revert "ci: disable Collabora's farm due to maintenance" - ci: disable Collabora's farm due to maintenance Sergii Ushakov (1): - android: moving HMI symbol to separate file Sergio Lopez (1): - hk: fix instance reference in vk_free Seán de Búrca (14): - rusticl: move debug logging to the end of the build step - rusticl: disentangle \`ProgramBuild` state from kernel compilation - rusticl: clarify naming of program-related structs and fields - rusticl: release borrow on device build before linking - rusticl: consolidate linking code - rusticl: add abstraction for \`util_queue` - rusticl: introduce intermediate header object - rusticl: restructure program build to prepare for parallelization - rusticl: execute program builds as jobs on a worker thread - rusticl: adjust naming and assert usage for clarity - rusticl/kernel: delay calculation of CSO info until kernel creation - nak: remove boxing of instructions - rusticl/kernel: add Kernel::mut_ref_from_raw() - rusticl/kernel: remove mutexes from kernel structure Sid Pranjale (1): - docs: mark VK_KHR_depth_clamp_zero_one as done for NVK Sil Vilerino (16): - mediafoundation: Fix recon pic two pass VPBlit target - mediafoundation: Do GPU-GPU encoder sync for two-pass input vpblit - d3d12: Fix two pass flag setting and rate control dirty flag check - d3d12: Fix double video encode resource barrier for DPB/recon pic resources - d3d12: Implement d3d12_context_queue_priority_manager - mediafoundation: Implement d3d12_context_queue_priority_manager and related ICodecAPI - mediafoundation: Check driver caps for intra-refresh CodecAPI advertisement - d3d12: Check slice support for PIPE_VIDEO_CAP_ENC_INTRA_REFRESH support - d3d12: Fix leak d3d12_context::priority_manager_lock - mediafoundation: Fix leak mft_context_queue_priority_manager::m_lock - ci: Bump DirectX-Headers and Agility SDK dependencies to 1.618.1 - pipe: Add video encode spatial adaptive quantization interface - d3d12: Implement video encode spatial adaptive quantization interface - d3d12: Remove Agility v717 guards for features now available in v618 - mediafoundation: Remove Agility v717 guards for features now available in v618 - mediafoundation: Implement video encode spatial adaptive quantization interface Silvio Vilerino (8): - d3d12: Fix typo in cast when reading pipe_h265_enc_picture_desc::gpu_stats_psnr - mediafoundation: Use lower size estimations for compressed output bitstream sizes - d3d12: Use lower size estimations for compressed output bitstream sizes - d3d12: Allow frontends to set_video_encoder_max_async_queue_depth() to manage encoder memory overhead - d3d12: Fix video encoder async depth fence wait off by one bug - mediafoundation: Use d3d12 extension set_video_encoder_max_async_queue_depth to save memory in low latency (no async/in flight frames) - d3d12: Video encode - Check driver caps to determine which output stats are supported - mediafoundation: mftransform async slices parsing, avoid heap allocation inside loop Simon McVittie (2): - vulkan: Consistently form driver library names as prefix + name + suffix - vulkan: Compute path to write into JSON manifests once, use it everywhere Simon Perretta (251): - wsi/display: make HDR_OUTPUT_METADATA, Colorspace properties optional - nir/nir_lower_calls_to_builtins: trivially handle IA64 mangled functions - pvr: start moving over to using the vulkan runtime vertex input state - pco: handle replicated components when translating nir alu srcs - pvr: default varyings interpolation to smooth when not set - pco: amend index register mapping - pco: enable all expected types for vertex i/o - pvr: amend incorrect format assertions - pvr: support getting device info from public name - pco: pygen: support passing custom refs to enc_ops - pco, pygen: support more comparison ops and types - pco: support shift ops - pco, pygen: support integer add/mul/mad ops - pco, pygen: support gradient/derivative ops - pco: commonize and improve iteration helpers - pco: support re-indexing loops and ifs - pco: amend cf printing indentation - pco: pygen: amend op mod print strings - pco: fix idx reg print colors and sq brackets - pco: control-flow epilogue/interlogue/prologue boilerplate - pco: switch to glsl/list, add control flow boilerplate - pco: skip over empty blocks when iterating instructions - pco, pygen: differentiate between int and float ref mods - pco: add virtual register support - pco: primitive bool support - pco: pygen: propagate selected source for ops with multiple source selections - pco: pygen: support applying modifiers to OpRefs - pco: pygen: add control-flow and branch ops - pvr, pco: initial ssbo and atomics support - pco, pygen: support test predicate setting - pco: initial control-flow support - pco, pygen: expose enhanced logical ops with optional mask - pco: add support for various selection, complex, trig ops - pco: add support for more bitwise and bitfield ops - pvr, pco: add base compute support - pco: experimental regalloc changes - pvr: pack image/texture array size unconditionally - pvr: preliminary support for combined image samplers - pco: add uadd64_32 op - pco: add basic pass to shrink vecs with unused components - pco: initial texture/sampler compiler support - pvr: initial texture/sampler driver support - pco: add support for using index(ed) registers - pco, pvr: push constants support - pco: basic arrayed image/sampler descriptor support - pvr: storage image descriptor support - pco: add boilerplate code for legalizing pseudo-ops - pco: add helpers for phase iteration, print more igrp offset info - pvr, pco: add support for buffer size intrinsic - pco: rework nir processing and passes - pvr, pco: usc program (pre-)generation boilerplate - pco: add support for loops and ifs using predicated execution - pco: update virtual register support for bools and nir reg translation - pco: support integer abs/neg - pvr: temporarily tweak support required for query programs - pco, pygen: add mutex op - pco: add intrinsic for loading instance num in slot - pvr, pco: improve indexed reg support, add shared memory support - pvr, pco: temporarily add supporting code for VK_KHR_zero_initialize_workgroup_memory - pco: add initial support for shared atomics - pco: experimentally propagate olchk mod for fwd prop opt - pco: temporarily prevent shared mem (coeffs) and vregs from being copy proped - pco: basic support for undefs - pvr, pco: initial support for blend constants - pco: suppress uses_sample_shading changes from nir_lower_blend - pvr: enable logicOp feature - pvr, pco: point sampler support - pco: initial image support - pvr, pco: per frag/vertex input/output rework - pco: skip lowering fs outputs that aren't present - pco: add support for sscaled8* formats - pvr: add descriptor copy support - pco: lower {insert,extract}_[ui]{8,16} to bitfield ops - pvr, pco: temporarily add legacy tq shader gen code - pco: initial image write support - pvr: initial texel buffer support - pvr, pco: basic depth feedback/discard/terminate support - pvr, pco: add input attachment sampler and initial support - pvr: use mrt_resource output size for fs outputs and input attachments - pvr: skip setting up unused fragment shader outputs - pvr, pco: temporarily add legacy loadop shader gen code - pvr: check for unused attachments - pco, pvr: account for early frag testing - pvr: sampler and sampled image descriptor support - pco, pvr: sample mask out support - pco: support combined depth/discard isp feedback - pvr, pco: initial texture gather support with gather sampler - pco: fully switch over to common smp emission code - pco: basic image array support - pco: branching fence support, simple ditr insertion logic - pvr, pco: simple end-of-tile/render nir shader gen - pvr, pco: switch to new nop shader - pvr: drop legacy rogue compiler - pco: support dce for vregs - pco: further commonize iteration instruction emission - pco: support indirect function temp refs - pvr: initial sample rate shading support - pco: add pass to split shader in/out struct/array vars across more slots - pco: enable shrink vec opt - pco: support shader i/o arrays of structs - pco: temporarily treat already overridden refs as comps during regalloc - pvr: remove vertex position output assertion - pco: force image/texture array coordinate f2i32 conversions to be rtne - pco: add pass to expand out vecs only used by comps - pvr, pco: add support for gl_FrontFacing - pvr: dynamically handle shademodel for flat shaded varyings - pvr, pco: z-replicate support - pvr, pco: image size query support - pvr, pco: improved image write (with format) support, handle 111110 - pco: support render target/layer id intrinsic - pco: add render target awareness to input attachments - pco: temporarily make vecs interfere with their components during regalloc - pco: restrict regalloc debug printing - pco: add helpers for finding non-empty blocks, apply - pco: skip comp-only opt on collated vecs - pvr, pco: clip/cull distance support - pco: temporarily prevent vectorization of vertex outputs - pvr, pco: add support for robust buffer access - pvr: texture swizzle depth/stencil fix - pco: experimentally pre-propagate vectors during regalloc - pco: remap buffer samplers to be 2d - pco: basic image/texture cube support - pco: add remaining texture buffer support - pvr, pco: dynamic buffer and immutable sampler support - pco: handle vector ra via parallel copy - pvr: temporarily dword align \*all* descriptors - pco: temporarily aggressively prevent isp feedback reordering by opt passes - pvr, pco: fragment shader metadata boilerplate code - pvr, pco: additional multisample support - pvr, pco: tile buffer support - pco: experimentally transfer olchk to ops with refs requiring it - pvr, pco: add dummy stores for tilebuffer-only loadops - pvr: dynamic depth bias support - pco: remove modifiers from instructions with variable src/dests - pvr, pco: alpha to coverage support - pco: full shared atomics support - pco: improve image write using pck.prog - pvr: fix multi-type varying allocations - pco: fix split-type vertex attrib allocations/nir vars - pco: lower vertex attrib vars first - pco: add lower_io_array_vars_to_elements_no_indirects to preprocessing - pco: legalize between movs1/mbyp without emitting additional ops - pco: temporarily switch to basic lowering for [iu]mulextended - pco: add ops needed to support fquantize2f16 - pco: support accessing shareds/coeffs >= 256 - pco: lower nir phi undefs to zero - pco: handle offset calculation for empty blocks - pco: support break/continue in loop body/outside if/else - pvr: handle num workgroups in indirect compute - pco: uncoalesce vecs that can't be propagated - pvr, pco: handle stencil input attachments - pvr, pco: full support for tile buffer eot handling - pco: temporarily don't propagate pixout accesses in opt - nir, asahi: commonize interleave_agx - pco: image atomics support - pco: scalarize push constant accesses - pco: add write memory check before processing nir - pco: add early nir opt pass - pvr: select SPM EOT state words from render index - pco: rematerialize load consts to reduce register pressure - pco: amend early frag test/depthf logic for isp feedback - pco: support skipping overlap check emission, enable for eot shader - pvr: fix valgrind warnings for 64-bit unaligned access - pco: ensure srcs/dests interfere for instructions with repeat > 1 - pvr: spilling enablement - allow empty uploads - pco: spilling enablement - track barrier usage - pvr, pco: experimental temp spilling - pco: temporary spilling workarounds - pvr, pco: temporary initial scratch memory support - pvr, pco: implement VK_EXT_image_2d_view_of_3d - pvr, pco: add VK_EXT_image_2d_view_of_3d sampled image support - pvr: add support for VK_EXT_provoking_vertex - pvr, pco: implement VK_EXT_depth_clamp_zero_one - pvr, pco: implement alphaToOne feature - pvr, pco: implement VK_EXT_color_write_enable - pvr, pco: basic write without format support - pco: support 1010102 snorm, [us]scaled formats - pco: replace {un,}packing alu ops with intrinsics - pvr: add a2b10g10r10 formats - pvr: enable VK_EXT_extended_dynamic_state - pco: handle remaining loadop depth formats - pvr: width-based tq depth format selection - pco: lower nir_b2b* ops - pco: use nir_cf_{extract,reinsert} instead of inlining compute instance check - pco: fix missing csbgen dependency - pvr: fix missing types in x86 builds - pco/opt: disable back-propagation of indexed registers - pco/ra: properly handle non-dced instrs with unused defs - vulkan: setup max_subgroup_size for drivers without varying/max/min size support - nir: print loop unroll info if present - pco: store additional metadata for precompiled shaders - pvr, pco: enable pre-generated header string functions to work with clc - pvr/csbgen: use stdint macro for unsigned 64-bit constants - pco/usclib: switch to common defs - pco: move uses_usclib flag into shader data - pvr, pco: switch to clc state update shader - pvr, pco: switch to clc nop shader - pco/usclib: add some preprocessor helper macros - pvr, pco: switch to clc vertex passthrough shaders - pvr, pco: switch to clc query shaders - pvr, pco: switch to usc generated clear attachment shaders - pvr, pco: switch to usc generated zero-init workgroup memory shaders - pvr: switch to usc generated spm load shaders - pco/usclib: disable predicate control-flow in generated shaders - pvr, pco: switch to clc load/store sr and idfwdf shaders - pco: switch to using csbgen and clc helpers for tex/smp state {un,}packing - pvr: merge legacy uscgen code into pvr_usc - pvr/wsi: don't advertise supports_modifiers - docs/pvr: drop GX6250 from the active development hardware list - vulkan/runtime: only set shader subgroup info if non-zero - pco: add usclib build dependency on generated files - mesa/st, nir: commonize unlower_io_to_vars pass - pvr, pco: implement prerequisites for sampleRateShading - pco: use interpolated input intrinsics for shader io - pco: use nir_unlower_io_to_vars - pvr, pco: track and implement workaround for brn74056 - pvr: add debug for missing sysvals - pvr: enable sampleRateShading feature - pvr, pco: allow fs sample rate to be dynamically set - pco: discard invalid instances depending on the sample & valid masks - pvr: enable independentBlend feature - pvr: enable VK_FORMAT_D32_SFLOAT_S8_UINT - pvr, pco: add multiview compiler support, advertise extension - pco: treat all load_consts as 32-bit - pvr, pco: support imageCubeArray feature - pco: fully support Vulkan 1.2 image atomics - pvr, pco: add minimal support required for Vulkan 1.2 subgroups - pco: set lower_device_index_to_zero - pvr: add support for VK_KHR_shader_draw_parameters, drawIndirectFirstInstance - pvr, pco: add remaining support for eds2 & 3 - nir/lower_alpha: extend to support dynamic a2c - pvr, pco: add primitive support for VK_KHR_robustness2.nullDescriptor - pvr, pco: add primitive support for terminate,demote_to_helper}_invocation - nir/unlower_io_to_vars: keep io bases intact when keeping intrinsics - pco: apply rounding mode to relevant conversion ops - pco: tidy and commonize conversion ops - pco: improve early and late algebraic pass ordering - pvr: amend tile buffer size calculation for eot - pvr: amend num temps calculation when wg_size is not provided - pco: ensure a variable exists for the multiview index - docs/pvr: update hardware list - pvr: advertise VK_KHR_sampler_mirror_clamp_to_edge - pvr: advertise VK_KHR_shader_non_semantic_info - pvr: advertise VK_KHR_shader_relaxed_extended_instruction - pvr: advertise VK_EXT_shader_replicated_composites - pvr: advertise VK_KHR_device_group_creation - pvr: support VK_KHR_map_memory2 - pvr: support VK_EXT_map_memory_placed - pvr: support VK_EXT_map_memory_placed.memoryUnmapReserve - pco: add support for global memory - pco/ra: abort if spilling fails SoroushIMG (5): - pvr: fix transfer fast clear color for srgb formats - pvr: remove unnecessary asserts - pvr: fix color values and crash for soft bg load ops - pvr: add more helper format function for tq pbe formats - pvr: set nn coords in sampler state for tq shaders when needed Surafel Assefa (1): - wsi: Implements scaling controls for DRI3 presentation. Sushma Venkatesh Reddy (6): - intel/compiler: apply sqrt workaround for Horizon Forbidden West shader - intel/compiler: generalize workaround script name for broader applicability - intel/compiler: Initial bits for SRND instruction - brw: Add assembler support for SRND - intel/compiler: Validation for SRND instructions - intel/executor: Add examples for srnd Sviatoslav Peleshko (3): - anv: Always disable Color Blending for unused Render Targets - mesa,driconf: Add WA to initialize vertex program outputs to vec4(0,0,0,1) - driconf: Add vertex_program_default_out option for Penumbra: Overture Tapani Pälli (17): - isl/blorp: handle failing 96bpp linear blit case - compiler/types: handle BFLOAT16 when decoding blob - iris: remove stage_from_pipe and pipe_from_stage helpers - intel/genxml: update CACHE_MODE_0 register for gfx200 - intel/dev: provide a helper to detect bmg g31 device - iris/anv: toggle on CACHE_MODE_0::MsaaFastClearEnabled on BMG G31 - anv: change some image qualifiers as coherent for Last Of Us - egl: allocate device info lazily only when queried - anv: remove assert, group can have 0 shaders in it - iris: setup bits for ARB_texture_filter_minmax with gfx9+ - blorp: add missing pipecontrol after 3DSTATE_WM_HZ_OP for Xe2+ - intel/blorp: add restriction for gfx12 - iris: add a check if blorp can support blitter copy - anv: add cs stall for any pipe control on compute - anv/blorp: add missing cs stall on compute pipe control - anv: bring back some lost game drirc workarounds for subgroups - anv: fix issues found with indirect data stride Taras Pisetskyi (1): - drirc/anv: force_vk_vendor=-1 for Wuthering Waves TellowKrinkle (2): - hk: Enable caching on memory marked with HOST_CACHED_BIT - hk: Add non-cached memory type Thibault Payet (1): - venus: Use SYS_thr_self on FreeBSD instead of SYS_gettid Thomas H.P. Andersen (4): - anti-lag: pass a proper dataSize - zink: do not overwrite existing error for miptail on uncommit - nvk: implement VK_AMD_buffer_marker - nvk: allow host image copy on non host visible heaps Tim Van Patten (2): - intel/ds: Skip expensive timestamp query until necessary - intel: Convert getenv() to os_get_option() Timothy Arceri (33): - util: add workaround for Interstellar Rift - glsl: move mark_array_elements_referenced() with ubo code - glsl: add mark_array_elements_referenced() fast path - glsl: rename setup_uniform_remap_tables() - util: remove recursion from bitset helpers - st/glsl: encapsulate more in st_nir_state_variable_create() - st/glsl: fix packed uniform handling in st_nir_lower_fog() - st/glsl: fix nir_lower_position_invariant() - nir: move nir_lower_drawpixels() to the state tracker - st/glsl: set driver locations in nir_lower_drawpixels() - nir: move nir_lower_alpha_test() to the st - st/glsl: set driver location in nir_lower_alpha_test() - nir: move nir_lower_point_size_mov() to st - st/glsl: set driver location in nir_lower_point_size_mov() - st/glsl: set driver loc after lowering clipplane - st/glsl_to_nir: dont add duplicate state tokens - util: add range remap util - glsl: make use of u_range_remap for uniform remapping - glsl: remove now unused NumUniformRemapTable - nir: fix uniform cloning helper again - util: add shortcut for range remap inserts - util: rewrite remap util to avoid looping list - Revert "ci/freedreno: Skip overly-slow trace" - Reapply "ci/freedreno: Skip overly-slow trace" - util/range_remap: dont overwrite entry if ptr is NULL - glsl/util: update util_range_remap to use range_remap struct - util/range_remap: split list node from range entry - util/range_remap: use child memory context for list - util/range_remap: add util_range_switch_to_sorted_array() helper - util/range_remap: switch to using sorted array - Revert "Reapply "ci/freedreno: Skip overly-slow trace"" - mesa: skip redundant uniform update optimisation if unsafe - glsl: assign block indices in the order they appear Timur Kristóf (41): - radv/amdgpu: Fix crash with RADV_DEBUG=noibs - radv/amdgpu: Use correct NOP packets when unchaining a CS - radv/amdgpu: Don't use IB2 on GFX6 (for now) - radv: Don't set SWITCH_ON_EOI without tessellation - radv: Don't use EVENT_WRITE_EOS on GFX7 - radv: Clean up use of RELEASE_MEM on GFX7 MEC - radv: Don't use V_370_PFP or V_028A90_PS_DONE on compute queues - radeonsi: Flush L2 for render condition when CP can't use L2 - radeonsi: Fix some comments to also include GFX11.5 - radv: Add comment to document CP DMA prefetch - radv: Flush L2 before CP DMA copy/fill when CP DMA doesn't use L2 - docs: Add more details about the contribution process - spirv: Always mark FS layer and viewport index inpus as flat - ac/nir/ngg: Remove dead code for 64-bit mesh shader variables - ac/nir/ngg: Fix scalarized mesh primitive indices - radv/amdgpu: Rename use_ib to chain_ib - radv: Rename RADV_DEBUG=noibs to noibchaining - radv/amdgpu: Don't assert chaining match when copying secondary IB - radv/amdgpu: Add a helper function to emit NOP packets - radv/amdgpu: Emit a single 4 dword NOP in chainable CS buffers - radv/amdgpu: Small cleanup of counting submitted IBs - ac/gpu_info: Add can_chain_ib2 field to ac_gpu_info - radv/amdgpu: Support IB2 without chaining, enable on GFX6 - radv/amdgpu: Allow IB2 when primary CS isn't chained - radv: Pass correct queue family to radv_cs_emit_write_event_eop - radv: Pass correct queue family in radv_emit_cache_flush - radv: Call transfer copy functions from API functions, not helpers - radv: Clarify image and image/buffer copy helper functions - radv: Add amd_ip_type to radv_cmd_stream - radv: Remove qf argument from radv_cs_emit_write_event_eop - radv: Remove qf argument from radv_cp_wait_mem - radv: Remove qf argument from radv_cs_emit_cache_flush - radv: Remove qf argument from radv_cs_write_data (and _head) - radv: Remove unneeded forward declaration of qf from dgc header - radv: Remove qf from radv_spm/sqtt/perfcounter where applicable - radeonsi: Don't use compute queue with regalloc hang bug - radv: Disable compute queues when the regalloc bug is present - radv: Mitigate GPU hang on Hawaii in Dota 2 and RotTR - radv: Document SWITCH_ON_EOP and WD_SWITCH_ON_EOP - ac/nir/ngg_mesh: Lower num_subgroups to constant - ac/nir/ngg: Fix scratch space for NGG GS streamout Tomeu Vizoso (29): - teflon: Reformat with clang-format - pipe-loader: Implement loading of /dev/accel devices - teflon/tests: Increase tolerance - teflon: Query drivers on what operations they support - etnaviv/ml: Implement ml_operation_supported() callback - rocket: Initial commit of a driver for Rockchip's NPU - pipe-loader: Load the rocket accel driver - teflon: Link to the rocket driver - teflon: Add support for Reshape operations - etnaviv/ml: Add support for no-op Reshape operations - teflon: Add support for non-fused Relu operations - etnaviv/ml: Add support for non-fused ReLU - teflon: Add support for Absolute - etnaviv/ml: Add support for Absolute - teflon: Add support for Logistic - etnaviv/ml: Add support for Logistic - teflon: Add support for Subtract - etnaviv/ml: Add support for Subtract - teflon: Add support for Transpose - etnaviv/ml: Support Transpose operation - etnaviv/ml: Remove some skips that pass now - teflon/tests: Remove dependency on xtensor - teflon/tests: Replace YOLOX model with that from TI - teflon: Add support for the MaxPool operation - teflon: Add support for the StridedSlice operation - teflon: Add support for the ResizeNearestNeighbor operation - ethos: Initial commit of a driver for the Arm Ethos-U65 NPU. - pipe-loader: Load the ethos accel driver - teflon: Link to the ethos driver Torge Matthies (2): - wsi/display: Factor drmModeObjectProperties retrieval out of find_properties. - wsi/display: Fix vkGetRandROutputDisplayEXT when connector is not leased yet. Trigger Huang (2): - virtio/vdrm: add ENABLE_DRM_AMDGPU for c_args - radeonsi: Fix u_log_ctx for aux_context recreation Utku Iseri (1): - panvk: override can_present_on_device Val Packett (1): - radv: detect platform:virtio-mmio devices for virtgpu native context Valentine Burley (101): - ci/lava: Use UART for non-Chromebooks - freedreno/ci: Increase concurrency for a618 jobs - turnip/ci: Increase coverage of a618-vk, reduce parallelism - freedreno/ci: Re-enable a618-gl job - zink/ci: Run full zink-tu-a618 job pre-merge - freedreno,zink+tu/ci: Document Piglit bug - ci: Disable Valve keywords farm - ci: Always save the artifacts for performance traces - ci/angle: Update gn arg to avoid warning message - lavapipe/ci: Add Android Hardware Buffer test set - freedreno/ci: Update a6xx kernel to msm-next - freedreno/ci: Remove a630 jobs - freedreno/ci: Streamline using common a6xx-skips - zink/ci: Only enable VVL for deqp on RADV - zink/ci: Fix enabling VVL for RADV jobs - zink/ci: Enable more VVL on ANV - radeonsi/ci: Convert Fluster job to deqp-runner suite - radeonsi/ci: Remove Fluster flakes, document failures - ci/lava: Only keep structured_logger in lava-trigger container - ci/lava: Use init-stage1 from Mesa build instead of inlining it - vulkan/wsi/wayland: Enable 4444 formats - zink/ci: Add pre-merge EGL coverage on ANV - zink/ci: Drop duplicate full ANV deqp-runner suites - ci/lava: Add x86_64 ASan job templates - ci: Build more drivers in debian-x86_64-asan - radv/ci: Use same deqp-runner suite for all RADV jobs - radv/ci: Add an ASan RADV job on Cezanne - intel/ci: Fix acer-chromebox-cxi4-puff concurrency - zink/ci: Add an ASan job on CML - radeonsi/ci: Increase Fluster job concurrency - ci: Drop obsolete EGL skips - zink/ci: Use Weston's Xwayland instead of Xvfb - softpipe/ci: Use Weston's Xwayland instead of Xvfb - virgl/ci: Use Weston's Xwayland instead of Xvfb - ci: Remove xvfb from test-base container - freedreno/ci: Move a660-gl-cl job to nightly - zink/ci: Skip flaky tests on CML due to HW deficiency - zink/ci: Document flakes on ANV - zink/ci: Add a prefix for X11 dEQP-EGL on ANV - zink/ci: Document more flakes on ANV - ci: Separate build and test container tags - zink/ci: Run full zink-lavapipe job pre-merge - zink/ci: Add EGL coverage on lavapipe - zink/ci: Document recent flakes on TGL - ci/fluster: Uprev Fluster - ci/lava: Make Fluster vectors an optional overlay - ci: Temporarily hardcode S3 artifact path - anv/ci: Lower concurrency for nightly jobs - anv/ci: Update expectations from nightly jobs - zink/ci: Switch to quick_gl profile for nightly ANV jobs - zink/ci: Update expectations from nightly jobs - anv/ci: Run full anv-adl-angle job pre-merge - anv/ci: Add a job replaying traces with ANGLE - iris/ci: Add a new iris deqp job on Alder Lake - zink/ci: Add EGL coverage on Turnip - zink/ci: Document recent flakes on a618 with Turnip - radeonsi/ci: Fix radeonsi-vangogh-glcts job definition - freedreno/ci: Add missing caching proxy for traces - tu: Advertise VK_EXT_shader_atomic_float - ci/crosvm: Retry all curl errors when downloading kernel - zink/ci: Disable zink-anv-cml-asan - tu: Enable robustBufferAccessUpdateAfterBind - zink/ci: Enable VVL for Turnip on a618 - zink/ci: Document recent a618 EGL flakes - zink/ci: Add a new Minecraft restricted trace - ci/crosvm: Add log sections for crosvm - zink/ci: Disable ASan leak detection and re-enable zink-anv-cml-asan - llvmpipe: Initialize src array in generate_fs_twiddle - r300/compiler: Silence array-bounds warning - imgui: Mark imgui dependencies as system includes - imgui: Silence build warnings for imgui - util: Update BLAKE3 from 1.5.1 to 1.8.2 - util: Disable Werror for BLAKE3 - meson: Relax -Wmaybe-uninitialized errors - lavapipe/ci: Disable stack-use-after-return detection for ASan - ci/gfxreconstruct: Bump version for compatibility with Debian 13 - ci/skqp: Add missing include to fix compilation errors on Debian 13 - ci/vkd3d: Disable Werror for vkd3d-proton - ci/mold: Bump version for compatibility with Debian 13 - ci/lava: Update \`fire` for compatibility with Debian 13 - ci/va: Bump va-tools version for compatibility with Debian 13 - ci: Bump ci-kdl version for compatibility with Debian 13 - ci: Update to Debian 13 (trixie) - ci/android: Use aapt from Debian packages again - ci: Uprev ci-templates to pull in new helpers - zink/ci: Document flakes on Cezanne - zink/ci: Re-enable ASan leak detection and drop VVL filter on CML - ci/lava: Use lava-job-submitter from gfx-ci repo - ci: Remove lava-job-submitter, LAVA containers, and tests - ci/android: Upload arm64 Mesa driver builds - ci: Rename ANDROID_GPU_MODE to CUTTLEFISH_GPU_MODE - ci/android: Make Vulkan driver replacement conditional - ci: Disable broken MR check in sanity job - ci/lava: Make fastboot commands customizable - freedreno/ci: Update kernel to pull in updated dtb - freedreno/ci: Update expectations for a306 and a530 - freedreno/ci: Move a306 and a530 jobs to LAVA - freedreno/ci: Remove baremetal job templates - docs: Update LAVA caching setup - tu: Fix indexing with variable descriptor count - tu: Fix maxVariableDescriptorCount with inline uniform blocks Vasily Khoruzhick (1): - lima: ppir: index SSA nodes the same way as we index registers Vignesh Raman (7): - ci/lava: default CI_JOB_TIMEOUT to 3600 if unset - ci/lava: add main() function to fix entry point - ci/lava: make rootfs shell prompt configurable - ci/lava: Move lava_job_submitter tests to lava folder - ci/lava: bump ALPINE_X86_64_LAVA_TRIGGER_TAG - ci/init-stage1: avoid duplicate mounts - ci/container: add comment to bump image tag Vinson Lee (2): - panfrost: Remove duplicate variable ret - gfxstream: Fix build error Vitaliy Triang3l Kuzmin (6): - .gitignore: Add KDevelop \*.kdev4 - radv,ac: GFX10 depth/stencil HTILE mipmap bug info variable - radv,ac: Split has_tc_compat_zrange_bug into Z and ZS, document it - radeonsi: Disable TC-compatible HTILE when bug workarounds conflict - radeonsi: Use radeon_info bug flags in TILE_STENCIL_DISABLE setup - ac: Enable HTILE TC Z clear value bug workaround on GFX1013 Vlad Schiller (6): - pvr: Enable VK_FORMAT_FEATURE_2_TRANSFER_SRC_BIT flag - pvr: Enable VK_FORMAT_FEATURE_2_TRANSFER_DST_BIT flag - pvr: implement dynamically set vertex buffer strides - pvr: Enable KHR_swapchain_mutable_format - pvr: Implement VK_KHR_imageless_framebuffer - pvr: Implement EXT_separate_stencil_usage Wenfeng Gao (2): - mediafoundation: support CODECAPI_AVEncVideoSatdMapBlockSize and MFSampleExtension_VideoEncodeSatdMap for SATD map. - mediafoundation: look into using texture pool for metadata retrieval, e.g SATD, Bitsused map, etc. X512 (1): - NVK: report \`VK_KHR_unified_image_layouts` extenstion support Xaver Hugl (2): - vulkan/wsi: require extended target volume support for scRGB - vulkan/wsi: remove support for VK_COLOR_SPACE_EXTENDED_SRGB_NONLINEAR_EXT Yinjie Yao (3): - radeonsi/vcn: Enable preencode on VCN5.0 - ac,radeonsi/vcn: Use correct swizzle_mode for vcn4 - ac/parse_ib: Update vcn ib parser to include missing commands Yiwei Zhang (152): - doc: fix section and android instruction linking for install page - venus/virtgpu: drop mappable if blob size is smaller than requested - venus: drop force_unmappable hack - venus: refactor ahb import interface to take whole alloc info - venus/virtgpu: use size zero to request mapping the entire blob mem - venus: requests whole blob mem size for non-dedicated import - venus/ci: udpate expectations from venus-lavapipe-full runs - vulkan/android: add vk_android_get_ahb_image_properties - vulkan/android: add vk_android_get_ahb_buffer_properties - venus: adopt vk_android_get_ahb_buffer_properties - venus/wsi: move wsi image format info validation to vn_wsi - venus: adopt vk_android_get_ahb_image_properties - venus: clean up post vk_android_get_ahb_image_properties adoption - turnip: adopt vk_android_get_ahb_image_properties - turnip: amend AHB buffer support - vulkan/android: make vk_ahb_probe_format private to android runtime - v3dv: adopt vk_android_get_ahb_image_properties - v3dv: amend AHB buffer support - lvp: hook up AHB image and buffer properties queries - vulkan/android: improve AHB image format check logging - lavapipe: allow AHB export allocation - lavapipe: implement GetMemoryAndroidHardwareBufferANDROID - lavapipe: do not close import fd on error and amend an error code - lavapipe: properly handle AHB release - lavapipe: populate AHB memory mapping - lavapipe: do not short-circuit AHB export alloc (non-import) - lavapipe: amend missing object finish on mem alloc failure - lavapipe: adopt common vk_device_memory - lavapipe: do not early return for mem alloc size being zero - lavapipe: use common vk_device_memory::ahardware_buffer - lavapipe: drop redundant memory type index tracking - lavapipe: use common host ptr info - lavapipe: use common export and import info tracked - lavapipe: use common tracked size and override if needed - u_gralloc/mapper4: properly expose ChromaSiting types based on api level - lavapipe: ensure to use zero memoryOffset for wsi image alias binding - lavapipe: improve image memory binding - lavapipe: fix a leak on a lvp_image_create exit path - lavapipe: fix maint4 vkGetDeviceBufferMemoryRequirements - lavapipe: fix maint4 vkGetDeviceImageMemoryRequirements - venus: add code owners - vulkan/android: improve memoryTypeBits reporting in AHB props query - venus: adopt vk_common_GetAndroidHardwareBufferPropertiesANDROID - venus: rework AHB memory import - venus: drop cached ahb buffer memory types - venus: drop is_wsi tracking and some asserts - venus: set wsi alias binding memoryOffset to zero - nvk: clean up existing nvk_android frontend - nak: do not hide drm header on Android - nvk: clean up direct u_gralloc dep - Revert "android: moving HMI symbol to separate file" - venus/android: clean up leftovers from common AHB helpers adoption - docs/android: add docs for preparing offline compilers - docs/android: fix meson setup for Android cross-compilation - docs/android: update cross file and add nvk instructions - docs/android: drop pkg-config workaround from cross-file - util/perf: amend missing atrace_init - venus: drop vn_trace_init - vulkan/wsi/headless: allow explicit modifiers - vulkan/wsi/headless: drop redundant chain struct members - venus: fix a race condition in ring shmem reuse - vulkan/wsi/headless: acquire the most likely idle image - vulkan/wsi/headless: drop the wsi_create_null_image_mem override - vulkan/wsi/headless: clean up headless wsi device and headers - vulkan/util: add missing vulkan header - vulkan/util: no need to hide ANB property itself behind Android - vulkan/util: update common properties code gen to use platform guard - venus: stop consuming wsi_memory_signal_submit_info - venus: layer vkQueueSubmit2 over vkQueueSubmit w/o sync2 - meson/android: drop redundant libdisplay-info dep - venus: use VK_USE_PLATFORM_ANDROID_KHR when applicable - venus: hide swapchainMaintenance1 behind wsi guard - venus: expose KHR_present_id(2)/wait(2) support - hasvk: advertise present_id/wait behind ANV_USE_WSI_PLATFORM - anv: advertise present_id/wait behind ANV_USE_WSI_PLATFORM - nvk: advertise present_id/wait and the 2 version - panvk: no need to set DRI_CONF_VK_KHR_PRESENT_WAIT - turnip: advertise present_id/wait behind TU_USE_WSI_PLATFORM - radv: advertise present_id/wait behind RADV_USE_WSI_PLATFORM - hk: no need to set DRI_CONF_VK_KHR_PRESENT_WAIT - vulkan/wsi: drop obsolete wsi_common_vk_instance_supports_present_wait - driconf: drop obsolete DRI_CONF_VK_KHR_PRESENT_WAIT - venus: misc sync2 emulation fixes - panvk: stub out Android ANB and AHB image handling - panvk: resolve ANB (pre spec v8) - panvk: implement deferred image creation - panvk: ensure wsi memory is bound at offset 0 - panvk: add panvk_android_get_wsi_memory for AHB spec v8+ - panvk: add shared image support and advertise VK_ANDROID_native_buffer - panvk: implement AHB image deferred init and memory alloc - panvk: support VK_ANDROID_external_memory_android_hardware_buffer - vulkan/android: amend a missing case for IMPLEMENTATION_DEFINED AHB - anv: drop obsolete anv_create_ahw_memory - anv: avoid setting image format twice for AHB image - anv: adopt vk_android_get_ahb_image_properties - anv: drop anv_ahb_format_for_vk_format - anv: adopt common GetAndroidHardwareBufferPropertiesANDROID - vulkan/android: support AHARDWAREBUFFER_FORMAT_YCbCr_P010 format mapping - vulkan/android: refactor to retrieve AHB format properties once - vulkan/android: support AHB query for VK_ANDROID_external_format_resolve - panvk: drop an obsolete assert of explicit mod plane count - docs/android: default to use -Dandroid-libbacktrace=disabled - meson/android: amend the condition for libbacktrace - nvk: refactor nvk_CreateImage error path - vulkan/android: add an early return when there's no wait semaphores - vulkan/android: switch to vkQueueSubmit2 - vulkan/runtime: silence a -Wsometimes-uninitialized warning - vulkan/android: skip queue submit with copy_sync_payloads - vulkan/android: improve stage masks for semaphore ops - mailmap: add Yiwei Zhang - v3dv: use stack image for v3dv_GetDeviceImageSubresourceLayout - vulkan: handle wsi private data properly - anv: fix broken utrace - radv: bind aliased wsi image at memory offset zero - nvk: bind aliased wsi image at memory offset zero - tu: drop redundant Android headers - tu: simplify AHB image view format resolving for external format - vulkan/util: drop unused vk_select_android_external_format - tu: bind aliased wsi image at memory offset zero - tu: properly implement VkBindMemoryStatus from maint6 - panvk: fix broken clock sync after using CLOCK_MONOTONIC_RAW - intel/ds: VulkanApiEvent doesn't rely on interning data - intel/ds: simplify clock sync emit - intel/ds: minor code clean up - intel/ds: update GPU clock to be sequence-scoped when applicable - panvk: fix blackhole bo error path to use MODE_IMMEDIATE for unmap - panvk: fix image/buffer destroy to use MODE_IMMEDIATE for unmap - vulkan/util: drop workaround for ANB struct - panvk: use os_get_option instead of getenv - pan/genxml: improve pandecode_dump_file_open logging - pan/genxml: fall back to stderr when unable to create CS dump file - pan/genxml: use process name to distinguish CS dumps - panvk: add PANVK_DEBUG(category) to simplify debug control - panvk: adopt PANVK_DEBUG(category) - ci/panfrost: udpate panfrost-g610-fails to reflect latest stats - panvk: fix to clear FPK with incompatible blend modes - calendar: fix 25.3 branch names - panvk: use mesa_logi for startup info logs - panvk: log device and driver info for startup - panvk: allow panvk_pool_alloc_mem to use full slab_size - panvk: improve big_bo_pool bo utilization - panvk: drop panvk_pool_upload helper - panvk: improve error propagation in panvk_pool_upload_aligned - panvk: fix to advance vs driver_set properly - panvk: fix to advance vs res_table properly - panvk: fix sample shading of internal blend shader for MSAA - llvmpipe: zero is also a valid fd - llvmpipe: fix udmabuf mmap error check - llvmpipe: add a missing alloc error handling in fd import - llvmpipe: misc fixes for sparse binding - glcpp/meson: fix libglcpp generated header dependency - panvk: fix mem alloc size for VkBuffer backed by imported blob AHB Yonggang Luo (82): - radv: Move the amdgpu.h defines for Win32 to ac_linux_drm.h - addrlib: __debugbreak only present on Windows and from intrin.h - util: Refactoring util_dl_get_path_from_proc out of clc/clc_helpers.cpp - util: Add namespace over float16_t in half_float.h - util: Upgrade xxhash.h to v0.8.3 - renderdoc: Upgrade to v1.5 - util: Remove usage of WIN32 macro for DETECT_OS_WINDOWS - broadcom: gl_shader_stage_to_broadcom => mesa_shader_stage_to_broadcom - gallium: Remove unused TRACE_FLAG_USER_BUFFER - gallium/mesa: Change type of tgsi_shader_info::processor st_init_limits::sh to mesa_shader_stage - microsoft/clc: {} for struct initialize to avoid warning - microsoft/clc: Improve clc_compiler_test.cpp to use defined expect value - microsoft/compiler: Fixes dxcapi.h compiling warning with mingw64-clang - util: Remove dbghelp.h that already comes with winsdk and mingw for fix warning with mingw - virgl: Fixes warning: cast to smaller integer type 'unsigned long' from 'void \*' [-Wvoid-pointer-to-int-cast] - virgl: Fixes differs in parameter lists - ci/windows: Enable virgl for MSVC - aco: Fixes warning note: ambiguity is between a regular call to this operator and a call with the argument order reversed - lavapipe: Revise HAVE_LIBDRM to guard on drm only variables - util: Update DETECT_ARCH_X86_64 to exclude _M_ARM64EC - util: Add DETECT_ARCH_ARM64EC for defined(_M_ARM64EC) equivalent - util: Now DETECT_ARCH_X86_64 can be safely used in rounding.h - d3d10umd: Fixes building with mingw/gcc and windows sdk/ddk 10.0.26100.0 - va: Remove unused variable pscreen - va: Use { 0 } initialize struct - amdcommon: Use { 0 } initialize struct for .c files - radv: Fixes warning implicit conversion from enum type - radv: Fixes warning C5287: operands are different enum types 'VkShaderStageFlagBits' and ''; use an explicit cast - radv: Fixes warning C5287: operands are different enum types 'rgp_sqtt_marker_event_type' and 'rgp_sqtt_marker_general_api_type'; - mesa: Remove unused assyntax.h and update related files - ci: remove non-existent files in ci watch list - meson: Remove redundant TODO: - util: Add DETECT_ARCH_SPARC64 for sparc - mesa: Remove usage of USE_*ASM in mesa/main/debug.c - util: Remove usage of USE_**_ASM macros - vc4: Remove the usage of USE_ARM_ASM - mesa: refactor the glapi/tls includes into a single, reused header - mesa: Remove duplicated deceleration of _mesa_glapi_tls_Dispatch _mesa_glapi_tls_Context - meson: Remove unused with_asm_arch and USE_*_ASM macros - microsoft/clc: Fixes gcc 14 compile warning about sign-compare - microsoft/clc: Fixes gcc 14 compile warning about narrowing conversion - d3d12: Fixes warning: enumeration value 'PIPE_FORMAT_NONE' not handled in switch - d3d12: Fixes warning: comparison of integer expressions of different signedness - d3d12: Fixes warnings: format '%x' expects argument of type 'unsigned int', but argument 2 has type 'HRESULT' - d3d12: Fixes warning: format '%d' expects argument of type 'int', but argument 3 has type 'LONG' - meson: Use build_always_stale instead of build_always - util/format: u_format_gen.h are using UTIL_ARCH_LITTLE_ENDIAN, include util/u_endian.h for it - util: Always generate u_format_gen.h as docs need it - Revert "glsl: Work around MSVC arm64 optimizer bug" - Revert "nir: Temporarily disable optimizations for MSVC ARM64" - docs: Update requirement for MSVC - util: Remove the __declspec(dllexport) on win32 for PUBLIC export macro - util: Implement p_atomic_read for C++ properly. - d3d10umd: Fixes gcc warning: enumeration value 'D3D11_SB_OPERAND_TYPE_FUNCTION_BODY' not handled in switch [-Wswitch] - dzn: -DVK_USE_PLATFORM_WIN32_KHR is already comes from idep_vulkan_wsi_defines that depends by idep_vulkan_wsi - tgsi: Fixes ntt_should_vectorize_io parameters - tgsi/nir: Handling TGSI_OPCODE_RET in tgsi_to_nir - clang-format: Update the .clang-format files to conformance clang-format json-schema - clang-format: Move ForEachMacros into src/.clang-format for freedreno - meson: mingw do not need _USE_MATH_DEFINES, only MSVC need it - meson: Remove unused predefined macros for windows msvc/gcc - meson: Remove redundant '/wd4996' option for MSVC - meson: For windows, the with_ld_version_script won't take effect - aco: Fixes warning: function get_branch_target/to_clrx_device_name defined but not used - glsl: Fixes warning: deprecated directive: ‘%pure-parser’, ‘%error-verbose’ - meson: Remove non-unused inc_d3d9 - util: Fixes gcc warning: declaration of 'strndup' shadows a built-in function [-Wshadow] - meson: Getting symbols-check.py works for mingw - etnaviv: The relative path to build dir is not always valid, fix it - lavapipe: fixes warning C5286: implicit conversion from enum 'type1' to 'type2'; use an explicit cast to silence this warning - ci/window: Fixes LLVM error Lexer.cpp(1578): error C2065: 'C11AllowedIDCharRanges': undeclared identifier - ci/windows: Strip misleading release/15.x - ci/windows: Building gallium-d3d10umd with MSVC - ci/windows: Improve ci scripts - ci/windows: Rename to mesa_deps_packages.ps1 - ci/windows: Now building the deps with MSVC 2019 - ci/windows: Use winget to install packages and install Microsoft.WindowsWDK.10.0.26100 - ci/windows: Bump llvm and SPIRV-LLVM-Translator version tag - ci/windows: Bump image tag for enable d3d10umd building - ci/windows: Update documents to use winget - meson: Update comment to be clear - meson/util: Define _GNU_SOURCE for mingw Yurii Kolesnykov (2): - Guard double include of libdrm.h by defining LIBDRM_H - Guard call to free_zombie_glx_drawable with condition from its definition Zach Battleman (1): - brw: Initial bits of BFN support Zan Dobersek (7): - tu: disable LRZ writes also for alpha-to-coverage, FS sample coverage output - tu: prevent tu_bo unmapping during destruction while being dumped - tu/drm: avoid has_set_iova-specific util_vma_heap freeing in tu_bo_init - tu/drm: msm backend shouldn't use util_vma_heap in the !has_set_iova codepaths - tu/drm: msm's has_set_iova codepath should avoid freeing zombified tu_sparse_vma - tu: limit query pool types logged into RMV - fd: allow limiting RD dumps to specific frames and submits Zhao, Jiali (2): - amd/vpelib: Extend TMZ value to 8 bit - amd/vpelib: Create Function to Check for Blending Feature Zhou Qiankang (2): - anv: Use os_get_page_size for mmap offset alignment to work with page size other than 4K - meson: use pointer size for 64-bit detection instead of architecture names abdelhadi (2): - aco, radv: remove line duplicate - aco: fix debug info offset bbhtt (1): - meson: Clearly print error when distutils or packaging is missing fossdd (1): - bin/symbols-check: add __(de)register_frame_info_bases to platform symbols jglrxavpok (1): - radv: Avoid calls to strlen when parsing umr output to speed up hang progressing leonperianu (2): - pvr: Advertise KHR_separate_depth_stencil_layouts - pvr: add support for VK_KHR_depth_stencil_resolve llyyr (2): - radv: don't set HOST_IMAGE_TRANSFER_BIT if host_image_copy not enabled - vulkan: Update enum_to_str conversion to handle AMDX enum names nihui (2): - aco: gfx940 has no mad f32 instruction - aco: set program->dev.fused_mad_mix=true for GFX940 no92 (1): - gallivm: support LLVM 21 norablackcat (2): - rusticl: fix unit tests - rusticl: add Test targets sarbes (4): - lima: move RSW packing/unpacking to genxml - lima: clean up unused PP struct - lima: implement logicops - lima: wire up anisotropic filtering sergiuferentz (1): - gfxstream: VirtGpuDevice can be null for Goldfish. serguei (1): - Revert "ci: disable Collabora's farm due to maintenance" sjfricke (1): - nir: Fix gnu-empty-initializer warning stefan11111 (1): - glx: Fix segfault when Nvidia PRIME render offload is enabled, but not used swscm, z1 (1): - amd/vpelib: Ensures type-safe comparison for callback assignment Mesa 26.0.5 Release Notes / 2026-04-15 ====================================== Mesa 26.0.5 is a bug fix release which fixes bugs found since the 26.0.4 release. Mesa 26.0.5 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.0.5 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: SHA256: d229c9937d9a25ca0a8958c59f425174563d300ec42acbea2dbe84a055023368 mesa-26.0.5.tar.xz SHA512: 8aa03a46269b2443be15cbd516d523af78fdd35e5273f5346b9142d2d21e245f6d5fd47e7e90176b0444ea967540eaa7c62d217f599ca5a216ef83244ee97d5c mesa-26.0.5.tar.xz New features ------------ - None Bug fixes --------- - Is maxFragmentCombinedOutputResources=16 in Honeykrisp reflects an actual HW limit? - Mesa LLVMpipe Memory Leak Changes ------- Ahmed Hesham (1): - rusticl: fix flag validation when creating an image Daniel Schürmann (1): - aco/lower_branches: Don't remove branches which jump over loops David Rosca (1): - radeonsi: Set multi plane format also for imported textures Eric Engestrom (4): - docs: add sha sum for 26.0.4 - .pick_status.json: Update to 7e163fb79377c0fdf6d4e99ca4775fa7e1a4299e - .pick_status.json: Mark 9ff879441f91a8296891e2e13264a7a015a11a7d as denominated - .pick_status.json: Mark 4b3bd6b0b54d998a31356bf049911004683ea64f as denominated Eric Guo (1): - panfrost: disable round_to_nearest_even for NEAREST samplers Faith Ekstrand (6): - pan/bi: Support more swizzle aliases in the bifrost pack code - pan/bi: Delete a few instruction encodings - pan/bi/ra: Allow offsets on tied sources - pan/bi: Use bi_half() for texture MS indices - pan/bi: Add BI_SWIZZLE_NONE - pan/bi: Support all the swizzles in the packer Georg Lehmann (2): - nir/opt_load_skip_helpers: don't skip helpers for store_scratch data - aco/optimizer: do not try to create 3 byte constant operands Ian Romanick (2): - brw/const: Don't allow type changes when accumulators are involved - brw: brw_reg::nr for an accumulator is not part of the offset Icenowy Zheng (2): - pvr: fix pvr_clear_vdm_state_get_size_in_dw() inverted feature condition - pvr: set has_usc_alu_roundingmode_rne for all B-series Rogue cores Janne Grunau (1): - hk: Increase maxFragmentCombinedOutputResources to HK_MAX_DESCRIPTORS Job Noorman (4): - nir/opt_varyings: fix alu def cloning - nir/gather_info: clear interpolation qualifiers before gathering - ir3: fix handle_partial_const with vectorized src - nir/opt_uniform_subgroup: fix ballot_bit_count components Karol Herbst (4): - radeonsi: set valid_buffer_range for CL buffers - radeonsi: properly report unified memory on APUs - rusticl/kernel: implement CL_KERNEL_GLOBAL_WORK_SIZE for custom devices - rusticl/device: Fix reporting of global memory on mixed memory devices Konstantin Seurer (1): - radv/bvh: Prefer selecting quads as the first pair of a HW node Lionel Landwerlin (3): - anv: don't relocate memory from blob - brw: don't support frontfacing ternary optimization on != 32bit - elk: don't support frontfacing ternary optimization on != 32bit Marc Alcala Prieto (1): - pan/cs: Fix cs_run_fragment() calls with swapped arguments Mary Guillemard (2): - nvk: Adjust maxFragmentCombinedOutputResources to match max descriptors limit - hk: Add HK_MAX_RTS to maxFragmentCombinedOutputResources Mixie (1): - xlib: clear currentDpy when releasing the current context Natalie Vock (1): - radv/rt: Don't enable midpoint sorting Olivia Lee (1): - panfrost: don't try to emit varying shader stats on v12+ Pavel Ondračka (2): - st/bitmap: release the temporary bitmap sampler view - gallium/u_blitter: remove unused CONST declaration when using IMM Rhys Perry (3): - util: fix UBSan error with _mesa_bfloat16_bits_to_float - ir3/array_to_ssa: skip remove_trivial_phi for non-array phis - ir3/ra: fix copy-paste error Samuel Pitoiset (3): - spirv: fix OpUntypedVariableKHR with optional data type parameter - radv/meta: fix computing extent for image->image with both compressed formats - vulkan: mark RP attachments as invalid when no rendering create info Timothy Arceri (1): - radeonsi: add Gun Godz workaround Valentine Burley (2): - zink/ci: Move zink-tu-a618 to sc7180-trogdor-kingoftown - ci/freedreno: Move remaining lazor a618 jobs, retire device type Vinson Lee (1): - d3d12: Fix MinGW cross-build error in resource_state_if_promoted Wujian Sun (1): - mesa: Fix inconsistent multisampled CopyTexImage checks Xianzhong Li (1): - panfrost: Fix GEM handle refcount leak in panfrost_bo_import Yuxuan Shui (1): - wsi/display: initialize Xlib display connector property IDs in all cases Mesa 26.0.6 Release Notes / 2026-04-29 ====================================== Mesa 26.0.6 is a bug fix release which fixes bugs found since the 26.0.5 release. Mesa 26.0.6 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.0.6 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: SHA256: 1d3c3b8a8363b8cc354175bb4a684ad8b035211cc1d6fa17aeb9b9623c513f89 mesa-26.0.6.tar.xz SHA512: 7f298234cb7b353ac40dc4d116866115160d1a52de6da8d4f1343af7ed47ca11e1a07b70aad8d24221906e0b520a4c140c12ffa7e7e706e3f25f301842a37e68 mesa-26.0.6.tar.xz New features ------------ - None Bug fixes --------- - glcpp: incorrect macro expansion in token pasting - nir: possible exactness bug in reassociate Changes ------- Alyssa Rosenzweig (1): - nir/opt_reassociate: fix exactness bug Arjob Mukherjee (1): - pvr: increase value of maxPerStageDescriptorStorageBuffers Benjamin Cheng (1): - radv/wsi: Re-use transfer queue if it exists Dave Airlie (2): - nvk: don't set sector promotion on texture headers - nouveau: drop sector promotion. David Rosca (3): - d3d12: Use HEVC RefPicSet order from frontend - radv/video: Fix initializing rc structs with default rate control - frontends/va: Fix finding LTRs from POCs in HEVC decode Derek Lesho (1): - zink: Guard bo map/unmap on map_count. Duncan Brawley (1): - pco: Fix pco_last_igrp returning the first element instead of the last Eric Engestrom (8): - docs: add sha sum for 26.0.5 - .pick_status.json: Update to d4d7055aee547f452689f8165e21ca100869e6fe - .pick_status.json: Mark c2708afbc7e91da54183c72d5e8c89a452356073 as denominated - .pick_status.json: Mark a5ec9b7892bd45e028cbf302eeced65864d6da22 as denominated - .pick_status.json: Mark 24849eef9f55a072eca342a48ff46471e3b89422 as denominated - .pick_status.json: Mark f78541b765e42125b846bb936ead4439ff5fe915 as denominated - .pick_status.json: Mark 78e2bbc70f55f1cf6ef922e2052b1ce6b879b952 as denominated - .pick_status.json: Mark 3256fab5a363f9bf4e749cd8f7473405334a2816 as denominated Erik Faye-Lund (7): - panvk: drop out-of-date TODO - pan/lib: fix up afbc and linear layout - pan/lib: emit high bits of buffer-size - nouveau: do not report unsupported feature - radeonsi: remove old, unsupported cap - panvk: do not enable extension without required feature - panvk: do not enable extension without required feature Faith Ekstrand (2): - pan/ci: Mark couple of WSI crashes as flakes - panvk/csf: Emit INDEX_BUFFER[_SIZE] even for non-indexed draws GKraats (1): - crocus: Fix shader precompilation on Gen6 and higher Georg Lehmann (1): - intel/nir_opt_peephole_ffma: fix fp_math_ctlr for modifiers Icenowy Zheng (5): - pvr: finalize query_indices array after ending last sub_cmd - pvr: fix the code copying query_indices to sub_query_indices - pvr: propagate get_vis_results flag from secondary cmdbuf gfx jobs - pvr: follow other drivers' practice for copying build ID - pvr: skip emitting query program when copy result / reset with 0 queries Janne Grunau (1): - nir/gather_info: clear interpolation qualifiers only in fragment stage Jesse Natalie (1): - wgl: Use an hwnd xor hdc for framebuffers Job Noorman (1): - ir3/shared_ra: fix live-out reload after src reload Jose Maria Casanova Crespo (1): - broadcom/compiler: really enable branch in delay slots validation Karol Herbst (2): - nak: the MS location comes last in TLD, same spot as depth compare in TEX - mesa/st: do not advertise CL subgroup features on the GL side Lionel Landwerlin (4): - anv: avoid C23 - anv: fix compute push constant allocations on pre Gfx12.5 platforms - anv: fix debug printfs on hang - anv: fixup compute queue detection Liu, Mengyang (1): - aco: fix broken VGPRs reservation for 64-bit attributes in VS prologs Matt Turner (2): - intel/elk: Remove dead TXL_LZ/TXF_LZ opcodes - radv: fix UB in radv_format_pack_clear_color for snorm formats Natalie Vock (2): - nir/deref: Elide loads/stores from deref cast of undef - radv: Run nir_opt_deref after first optimization loop Nick Hamilton (1): - pco: fix clamping the array index when shaderImageGatherExtended is enabled Patrick Lerda (4): - r600: fix alpha-to-coverage and alpha-to-one used together - r600: fix atomic buffer offset - r600: update vertex emit_varying_pos - r600: fix atomic_counter_post_dec Pavel Ondračka (2): - r300: fix MSAA resolve COLORPITCH tiling after pipe_surface de-pointerization - r300: dirty VS state when switching variants Ryan Zhang (1): - panvk: add VK_IMAGE_LAYOUT_DEPTH_READ_ONLY_OPTIMAL to host copy layouts Sagar Ghuge (1): - anv: Fix Wa_14021821874, Wa_14018813551, Wa_14026600921 Samuel Pitoiset (2): - radv: fix GPU hangs with PS epilogs and secondaries properly - radv: re-introduce DGC+multiview support and enable it for vkd3d-proton only Simon Perretta (2): - pco: reserve additional outputs for trilinear sampled coeffs - pco: amend tg4 lowering Tapani Pälli (4): - drirc/anv: add flag to disable VK_EXT_subgroup_size_control - drirc: set anv_disable_subgroup_size_control for bg3 - drirc: use anv_disable_drm_ccs_modifiers for any GTK version - anv: do not use resource barrier with split barriers Timothy Arceri (1): - glcpp: fix paste within macro function expansion Valentine Burley (3): - zink/ci: Remove Cezanne job - tu/drm/virtio: Fix tu_wait_fence timeout handling - freedreno/drm/virtio: Fix wait_fence ret ordering Vinson Lee (1): - zink: remove unused variable in zink_instance.py Mesa 26.0.7 Release Notes / 2026-05-14 ====================================== Mesa 26.0.7 is a bug fix release which fixes bugs found since the 26.0.6 release. Mesa 26.0.7 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.0.7 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: SHA256: 0c56bbcf1947e1a6a90ac09b129b0ca0cb52cc31145b94595e57c8804cf02496 mesa-26.0.7.tar.xz SHA512: a60aaed37907bcf9edbd68e2a95e5cc95893215c64e91f8fceffa0d3f67fc63e8eb20877fbe9a59dd9df0b7f8bff4e73392085e11414eed8ae4939c4c8691f93 mesa-26.0.7.tar.xz New features ------------ - None Bug fixes --------- - None Changes ------- Adrián Larumbe (2): - pan/kmod: Fix minor version number check for USER_MMIO_OFFSET ioctl - pan/kmod: fix double syncop count sum when populating vm_bind syncs Ahmed Hesham (1): - pan/bi: Restore b3210 as a valid swizzle Caio Oliveira (1): - brw: Fix max_dispatch_width collection for CS with variable size Calder Young (3): - anv: Fix address bit masking for indirect SBTs - anv: Fix support for indirect SBTs on Xe3+ - anv: Fix some usage flags not propagated to ISL for explicit layouts Christoph Pillmayer (1): - pan/kmod: Fix uninitialized timestamp info Connor Abbott (2): - tu: Fix LRZ+FDM offset+secondaries - tu: Disable LRZ when resuming if the GPU doesn't support tracking Danylo Piliaiev (1): - tu: Fix CP_CCHE_INVALIDATE not being applied at the right point Dave Airlie (3): - gallivm: handle llvm 22 coroutine end change - gallivm: handle llvm 22 scatter/gather intrinsic changes. - lavapipe: treat NULL pColorAttachmentLocations as no handles David Rosca (2): - frontends/va: Fix setting output color properties from color standard - frontends/va: Add missing NULL check for additional output surface Emma Anholt (3): - ir3: Fix shared IMAD24 lowering. - tu: Add capture/replay for sparse buffers and descriptor buffer. - screenshot-layer: Fix leftover VK queues in the map at DeviceDestroy. Eric Engestrom (2): - docs: add sha sum for 26.0.6 - .pick_status.json: Update to aee10432272f77fd5979de084f4f64f7374c3278 Eric R. Smith (2): - panfrost: make sure INDEX_OFFSET is cleared - panfrost: add helper function for checking for active queries Erik Faye-Lund (4): - mesa/main: remove stale prototypes - mesa/main: remove incorrect debug-output - Revert "mesa: check for ARB_ES3_compatibility in format checks" - mesa/main: remove unused array Georg Lehmann (2): - radv: fix amount of sample shading with required sample shaded inputs - ac/nir/lower_tex_coords: fix optimizing cube txd to tex Icenowy Zheng (7): - pvr: wait for graphics jobs in CopyQueryPoolResults - pvr: increase maxPerStageResources for new maxPerStageDescriptorStorageBuffers - pvr: do not setup deferred RTA clear for active render targets - pvr: properly handle deferred RTA clears for 2D array view of 3D image - pvr: add deferred RTA clear command to list after checking it's not NULL - pvr: record deferred RTA clears for secondary cmdbuf subcmds - pvr: setup viewindex if the shader wants it even when multiview disabled Job Noorman (4): - ir3/cf: fix rewriting uses with different dst types - ir3/shared_ra: use ir3_cursor instead of instr in reload helpers - ir3/shared_ra: insert reloads before tied dst pcopies - ir3: don't cache driver param instructions Jon Turney (1): - ddebug: Fix use of alloca() without #include "c99_alloca.h" Jose Maria Casanova Crespo (2): - broadcom/compiler: move nir_lower_undef_to_zero out of optimization loop - v3dv: include mem_offset in vkCmdFillBuffer destination Karol Herbst (5): - nir/lower_cl_images: call nir_progress on every function - gallivm/nir/soa: use uint for booleans - llvmpipe: never pass a NULL function name to LLVMAddFunction - ci: install libstdc++-static on fedora - rusticl: link the C++ runtime statically Lionel Landwerlin (3): - anv: fix null pointer access - anv: fix arc artifacts on Farming simulator 2022 - anv: fixup null address check Lorenzo Rossi (1): - panvk/jm: Fix tls_size overwrite in indirect draws Louis Montagne (1): - zink: relax build-id length assertion for Mach-O Marek Olšák (1): - radeonsi: fix a typo in si_shader_update_spi_shader_formats Mel Henning (2): - nvk: Add a wfi for blackwell in CmdDispatchIndirect - nvk: Disable compression on Turing Mike Blumenkrantz (13): - llvmpipe: fix min_samples + A2C - lavapipe: fix indirect memory copies - lavapipe: fix pushconst data updating - util/format: support 256-bit formats in util_format_get_tilesize() - lavapipe: use the right type for DGC mesh draws - lavapipe: rework immutable samplers - lavapipe: allow fbfetch with shader objects - llvmpipe: always set view_index for linear rasterizer - lavapipe: update cbuf count when remapping attachments - lavapipe: unset attachment remap state if pColorAttachmentLocations==NULL - lavapipe: fix setting colormasks when attachments get remapped - zink: fix mixing of mesh descriptor bindings with gfx bindings - meson: fix renderdoc integration define Nick Hamilton (1): - pvr: Revert don't csb emit multi-layer clear attachments without rta support Paulo Zanoni (2): - intel/isl: fix assert when surf->size_B is > UINT_MAX - intel/isl: warn about excessive num_elements only once Raviraj Uppal (1): - driconf: disable allow_rgb16_configs for SPECviewperf Rohit Athavale (1): - mediafoundation: Test compile steps v/s step , and set build flag Samuel Pitoiset (5): - radv: fix determining needed dynamic states when rasterization is disabled - radv: allow DGC+multiview by default - radv: do not fallback to compute for image->buffer copies with emulated formats - spirv: preserve the explicit stride for untyped pointers with matrices - radv: fix another case of VRS with mipmaps on GFX10.3 Vinson Lee (2): - st/mesa: fix implicit conversion warning in st_atom_framebuffer - vulkan/screenshot-layer: initialize info to NULL llyyr (1): - vulkan/wsi/wayland: use mtx helpers in wait_for_present2 Mesa 26.0.8 Release Notes / 2026-05-27 ====================================== Mesa 26.0.8 is a bug fix release which fixes bugs found since the 26.0.7 release. Mesa 26.0.8 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.0.8 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: SHA256: caf1c0061a68e88dfa74967a7e780c0e85d65b6c4e334cd69095a5dc54ad78bc mesa-26.0.8.tar.xz SHA512: 3a43648a86c1bc48161a1669733b6c6a9294bf27ffb529f2ce078c2daa3b90b8c53b9cc06312197fb3a0830303d0b326d7535fb15849e6c26fad58009f3a6112 mesa-26.0.8.tar.xz New features ------------ - None Bug fixes --------- - None Changes ------- Caio Oliveira (1): - nir: Add print for other cmat_description slots Calder Young (1): - spirv: Fix debugPrintfEXT not working with multiple arguments Danylo Piliaiev (2): - tu/a8xx: Fix reading border_color from sampler memory - tu: Always lazy_init_vsc for tiler rendering Dave Airlie (2): - nak: fix image size for multisample arrays - nak: add more sizes to assert in bindless_image_sparse_load David Rosca (1): - radeonsi/uvd_enc: Skip extra padding bytes in output bitstream Eric Engestrom (3): - docs: add sha sum for 26.0.7 - .pick_status.json: Update to af8c3eb3d6537eeda79258dd9fcc4178933d2ad8 - etnaviv: initialize value before calling etna_gpu_get_param(), in case it fails Erik Faye-Lund (4): - pan/ci: add missing gitlab rules - pan/ci: remove outdated gitlab rule - pan/ci: add missing gitlab rule - pan/ci: fix gitlab rules after move Faith Ekstrand (1): - panvk/csf: fix VERTEX_SPD dirty tracking when topology changes Frank Binns (1): - pvr/ci: drop two tests from bxs-4-64-{fails,flakes} Georg Lehmann (4): - aco/tests: use explicit lod in sparse texture test - radv: use radv_get_sampled_image_desc_size instead of open coding it - radv: add radv_force_64_byte_sampled_image dri conf option - radv: enable radv_force_64_byte_sampled_image for Forza Horizon 6 Iago Toral Quiroga (1): - pan/bi: TEX_GRADIENT may need helper invocations Icenowy Zheng (6): - pvr: fix handling of invalid attachment info in pvr_init_fs_outputs_mrt - pvr: copy sub_cmd flags except owned when executing subcmds out of pass - pvr: stop to derive rt datasets based on geometry_terminate - pvr: add a structure containing data kept for suspended renderpasses - pvr: preserve and pass more data for suspending render passes - llvmpipe: stub other functions inside compute shaders for ORCJIT Iván Briano (2): - anv: fix return of cmd_buffer_set_indirect_stride() function - anv, iris: fix MOCS Index setting of EXECUTE_INDIRECT_* commands Job Noorman (2): - freedreno/computerator: fix UAV view size - ir3: mark __alias_n as UNUSED in foreach_src_in_alias_group_n Jon Turney (4): - glx/windows: Avoid shadowing 'type' parameter of driwindowsCreateDrawable() - glx/windows: Fix compilation of driwindows_glx after driscreen changed from pointer to member - glx/windows: Fix compliation after code motion to put event base in 'dri' context - glx/windows: Drop static from driwindowsCreateScreen() Jose Maria Casanova Crespo (2): - v3dv: avoid duplicate bo_handles between cpu_job and CSD lists - v3dv: avoid 16F TLB usage for B10G11R11_UFLOAT copies Karol Herbst (2): - clc: do not use std::filesystem - Revert "rusticl: link the C++ runtime statically" Lionel Landwerlin (5): - anv: add SIMD32 requirement heuristic for Dragon Dogma 2 - anv: bump max compute workgroup count - anv: fix missing bindless flag hashing - anv: fix render target remapping tracking at the beginning of render passes - spirv: fixup infinite recursion with shader replacement Lone_Wolf (1): - ac/llvm: fix build with LLVM 23 (MCSubtargetInfo) Lorenzo Rossi (1): - pan/valhall: fuse_cmp skip when fusing the same instruction Mary Guillemard (4): - nir/lower_bit_size: Preserve float controls when lowering alu ops - nvk: Handle foreign queue dependencies - nvk: Handle host accesses barrier - nvk: Multiply by local_size for CS invocations in DGC codepath Matthieu Oechslin (1): - r600: Fix crash on R600/R700 with custom border color Michael Cheng (1): - brw: Fix ordered dependency exec_all handling on Xe2+ Mike Blumenkrantz (2): - zink: fix unbinding vertex buffers from null VS state - zink: create views for samplers lazily Nemallapudi, Jaikrishna (1): - intel/dev: fix timebase_scale ticks-to-ns precision loss across 2^32 Patrick Lerda (1): - i915: fix emit_hw_vertex() unbounded memory access Rhys Perry (2): - aco/ra: test the register file in get_reg_specified() when necessary - aco/ra: don't rename phi operands in get_reg_phi() Samuel Pitoiset (2): - nir: fix shuffling local IDs for quad derivatives with larger workgroup sizes - radv: enable radv_wait_for_vm_map_updates for Forza Horizon 6 UMUTech (1): - wsi: correct the erroneous assertion Valentine Burley (1): - tu/kgsl: Fix memory type support detection for unsupported flags Yiwei Zhang (1): - venus: fix a renderer side queue timeline bound race hwandy (1): - Revert "intel/decoder: make libvulkan_intel to depend on stub decoder when buildtyle=release." johniyoods (1): - egl/dri2: require valid render fd before advertising EGL_WL_bind_wayland_display yserrr (1): - llvmpipe: fix UB and incorrect value in compute caps shift Mesa 26.1.5 Release Notes / 2026-07-15 ====================================== Mesa 26.1.5 is a bug fix release which fixes bugs found since the 26.1.4 release. Mesa 26.1.5 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.1.5 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: SHA256: 79e421c7ce18cd9e790b8375920325779f10798630bf30e0b22f1a21c8617122 mesa-26.1.5.tar.xz SHA512: 8b1714e0fe63d3077c0d761245680724537a94db65e17dedb7124c7c861f070a13493a060695e81b2ca9b6889007b52c7b5788dace8fe5d8e2e1ed67efdbbf07 mesa-26.1.5.tar.xz New features ------------ - None Bug fixes --------- - A702: assertions at A6XX_TEX_MEMOBJ_4_BASE_LO - Celestia hit fallback on r300g from git? - Copy paste bug in \`gallium/drivers/radeonsi/si_texture.c` - Issues with shadows on DOOM: The Dark Ages - Revelations DLC - OpenGL app deadlocks in loader_dri3_swap_buffers_msc → xcb_wait_for_special_event on radeonsi / Mesa 26.1 / X11 (Telegram Desktop) - Qt application TrenchBroom hangs in glXSwapBuffers in recent version of Mesa - [ANV][PTL] - Elden Ring (1245620) - Blue artifacts and flickering lights with raytracing enabled - [ANV][PTL] - Persona 3 Reload (2161700) - Blue color on some objects + reflections do not reflect - [RADV] REGRESSION: Lighting Bleed-through in Need for Speed games on mesa 1:26.1.1-2 ArchLinux - [RADV] REGRESSION: Weird graphics glitches and Lighting Bleed-through in Saints Row 2 - [RADV] Video color artifacts in mpv - [radv] Regression causes black patches on the ground in DOOM The Dark Ages Revelations DLC - frontends/va: Build fails with libva 2.15.0 - r600: SFN assersion failed in Transport fever 2 - v3dv: Compute shader crashes for unknown reason Changes ------- Benjamin Cheng (1): - ac/surface: Remove GFX10 limitation for FORCE_SWIZZLE_MODE Boris Brezillon (1): - pan/kmod: Add a pan_kmod_timestamp_cycles_to_ns() helper Caio Oliveira (1): - anv: Initialize shader debug archive key size Christian Gmeiner (2): - etnaviv: Fix sampler view leak on unsupported texture target - panvk: Restore push descriptor dirty bit after meta operations Daivik Bhatia (1): - nir/opt_copy_prop_vars: kill stale entries when source deref is written Danylo Piliaiev (1): - tu: Enable tu_dont_care_as_load for all Kex Engine games David Rosca (2): - r600/uvd: Set correct h264 chroma format - [26.1 only] va: Fix build with free codecs only and libva < 2.16 Eric Engestrom (5): - docs: add sha sum for 26.1.4 - .pick_status.json: Update to 0ab29c1d21d56f1780bd057176c998fda738402a - zink/ci: drop leftover anv-cml deqp suite - loader: move variable to correct scope - radeonsi: fix truncated cache key Faith Ekstrand (2): - vulkan/meta: Use z_off/scale for 2D array images as well - vulkan/meta: Allow resolving a 2D MSAA image to a 3D image Filip Gawin (1): - nv30: fix truncated values in line_stipple_pattern Frank Binns (1): - pvr: define PVR_USE_WSI_PLATFORM for xcb and xlib Georg Lehmann (3): - spirv: add option to treat FMax/FMin/FClamp like NMax - radv: add radv_force_nan_preserve_min_max option - radv: enable radv_force_nan_preserve_min_max for DOOM: The Dark Ages Gert Wollny (1): - r600/sfn: Drop assertions when emitting IF asm instruction Hans-Kristian Arntzen (1): - radv: Consider VkImageView usage rather than VkImage usage in feedback. Iván Briano (3): - brw/rt: fix max_t selection on intersection report - brw/rt: split HitAttribute area in pending/committed - brw/rt, anv: reduce maxRayHitAttributeSize JaeHoon Lee (8): - v3d: fix slot and input indexing in v3d_set_global_binding - v3dv: report maxDrawIndirectCount of 1 without multiDrawIndirect - v3d: clamp transform feedback offset to buffer size - v3dv: report the correct dynamic storage buffer UAB limit - v3dv: fix blake3 key truncated to 20 bytes in pipeline cache - v3d: fix blake3 key truncated to 20 bytes in shader cache - broadcom/compiler: really enable GFXH-1625 TMUWT validation - vc4: fix incorrect resource unref in vc4_flush_resource Jeremy Huddleston (1): - glx/apple: silence OpenGL deprecation warnings Jose Maria Casanova Crespo (1): - v3dv: close the primary node fd on physical device destruction Karmjit Mahil (1): - freedreno/decode,ir3: Mark decoded dwords as const Karol Herbst (7): - gallium: remove PIPE_BARRIER_GLOBAL_BUFFER - rusticl/device: fix long vector_width queries on devices without int64 support - nak/sm20: fix immediate encoding for F2I and F2F - nir: add nir_shader_fully_linked helper - rusticl/kernel: handle nir shader compilation failures gracefully - rusticl: return Result instead of Option from convert_spirv_to_nir - rusticl: abort compilation if the nir shader is not fully linked Konstantin Seurer (2): - vulkan: Fix ROOT_FLAGS_OFFSET_ID offset - radv: Store root_flags for BVH8 Lionel Landwerlin (4): - anv/brw: limit push constant promotion in vertex shaders - anv: fixup RT building barrier - anv: fill min_array_element with indirect descriptors - anv: fixup max push data delivered to shaders Marek Olšák (5): - radv: disable AMD_device_coherent_memory on gfx12 due to out of order behavior - ac/nir: fix incorrect upper bound for view_index - radv: fix a rare crash with NULL PS and force_vrs_per_vertex - radv: use radeon_opt_set_context_reg for PA_CL_VRS_CNTL to fix random behavior - ac: fix a GPU hang with LLVM due to incorrect VGPRS decoding of LLVM output Mario Kleiner (3): - wsi/display: Actually fix vblank-less systems for VK_EXT_present_timing. - hasvk: Expose VK_KHR_calibrated_timestamps. - wsi/display: Don't update connector last_frame/nsec in vkGetSwapchainCounterEXT. Mary Guillemard (1): - nvk: Do not enable remap in nvk_copy_indirect Matt Turner (5): - nir/tests: allow relative error in compare_inexact - nir: fix f2u/f2i constant folding to poison NaN and out-of-range inputs - nir: use i2f32 for patterns with signed-extraction opcodes - nir: fix pack_uvec4_to_uint to mask input components to 8 bits - nir/tests: fall back to integer comparison when float interpretation is NaN Mel Henning (2): - nvk: Fix DGC localsize computation - nvk: Handle large indirect stride pre-Turing Michel Dänzer (1): - dri3: Increment draw->send_sbc after waiting for last presentation Mike Blumenkrantz (6): - zink: unset unordered access on ordered transfer ops - gallium/cso: make unbind_context an explicit call - st/context: unbind gs shader before deleting hw select gs shaders - zink: always un-suspend queries on end - zink: fix the fix for ZINK_RENDERDOC=all - util/blitter: fix blitting multiple array layers Pavel Ondračka (1): - r300: fix swtcl per-vertex point size Pierre-Eric Pelloux-Prayer (1): - radeonsi: fix typo in si_copy_from_staging_texture Pohsiang (John) Hsu (1): - d3d12: fix build break in staging/26.1 due to DirectX header issue Rhys Perry (1): - radv: set RADV_CMD_DIRTY_GFX12_HIZ_WA_STATE around attachment clears Rob Clark (2): - freedreno/a6xx: Don't forget UBO driver params - freedreno/decode: Fix shader stats in summary mode Ryan Mckeever (1): - pan/bi: check if preds are dominated by header in bi_find_loop_blocks Sagar Ghuge (2): - brw: Track if CS uses fences - intel: Fix async compute thread limit Sahitya Kandru (1): - freedreno: Modify reg_size_vec4 for a608 and a612 to 32 Samuel Pitoiset (12): - vulkan,anv,radv: do not crash when querying descriptor size for unsupported type - zink: fix a memleak in zink_init_format_props() - glsl: fix a memleak in link_assign_subroutine_types() - loader: fix a memleak - kopper: fix a memleak - zink: fix a memleak with fences - zink: fix a memleak with the emulated GS NIR shader - zink: fix a memleak with sampler state - pipe-loader: fix a global-buffer-overflow ASAN error when getting driconf - radv/meta: fix restoring descriptor heaps - Revert "spirv: allow mapping readonly buffers with struct members" - radv: force late-Z with fragment shaders that use fbfetch Simon Perretta (3): - pco: allow non-pure integer formats for image xchg atomics - pco: lower sysvals early for fragment shaders - pco: allow fence ops to be legalized if they come last in a block Stijn Tintel (1): - rocket: fix mmap leak in buffer map/unmap Timur Kristóf (3): - radv: Wait for idle after every submission on GFX6-7 - radeonsi: Wait for shaders and flush L2 after every submission on GFX6-7 - Revert "radv: Mitigate GPU hang on Hawaii in Dota 2 and RotTR" Valentine Burley (4): - tu: Fix uninitialized gmem_offset when a GMEM layout is impossible - tu: Fix capture/replay with sampler custom border color - tu: Disable VK_EXT_extended_dynamic_state2 patch control points on A702 - zink: Gate tess/geom barrier stages on feature support Vincent Cloutier (1): - etnaviv: use buffer resource accessor for indirect draws Vinson Lee (4): - util/tests: replace sprintf with snprintf in cache tests - util/tests: fix unused variable warnings in cache List test - vulkan/screenshot-layer: replace itoa/sprintf with snprintf - vulkan/screenshot-layer: fix globalLock mutex leak Wujian Sun (1): - mesa: Allow GL_SRGB_ALPHA_EXT as color-renderable when EXT_sRGB is supported Yannis Juglaret (1): - nouveau: fix data race in nouveau_fence_ref Yiwei Zhang (3): - venus/virtgpu: amend a missing sim mutex init - venus: fix imported sync fence payload reset upon export - venus: avoid renderer semaphore wait upon temp payload export inspector-ambitious (1): - loader: fix loader_open_render_node_platform_devices result allocation Mesa 26.1.6 Release Notes / 2026-07-29 ====================================== Mesa 26.1.6 is a bug fix release which fixes bugs found since the 26.1.5 release. Mesa 26.1.6 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.1.6 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: TBD. New features ------------ - None Bug fixes --------- - Ambient occlusion is broken in The Chronicles of Riddick - Assault on Dark Athena - Blockland crashes when shader quality turned up - Doom Eternal - Page Fault - somewhat reproducable. (RX9070) - Horizon Forbidden West misrendred lighting effects on BMG - Regression. VTK polydata. gl_PrimitiveID requires explicit geometry shader with versions 25.3.6 and later. - [ANV][PTL] - Horizon: Forbidden West regression in shader quirk - [VAOn12] Thread safety issues - [amdgpu] Little Inferno rendering issues starting with Mesa 21.1 - [radeonsi] eglinfo 'libEGL warning: failed to get driver name for fd -1' - anv: bottom-of-pipe timestamp latched before vkCmdDispatch completes (Arc B390 / Panther Lake, Mesa 26.1.4), breaking wgpu compute-pass timings - anv: cooperative matrix loads from shared memory return the wrong tile on Arc B390 (Xe3/PTL), matmul results are wrong - radeonsi/VCN: AV1/HEVC encode reports a coded size larger than the coded buffer, client segfaults reading the mapped buffer - radv: acceleration structure update (refit) of AABB geometry loses intersection candidates (NAVI32, Mesa 26.1.4) - segfault v3d raspberry Pi5 h265 hw decoding / regression mesa 26.1.3/26.1.4 - vulkan/runtime: GetPipelineBinaryDataKHR incorrectly assumes \`pPipelineBinaryDataSize` must be 0 initialized Changes ------- Adam Stylinski (1): - nv30: fix an issue when this push buffer is NULL Benjamin Cheng (1): - ac/vcn_enc: Disable var slice for preencode with VCN4 Caio Oliveira (2): - nir: Account for cmat memory accesses in copy_prop_vars - anv: Include build and device identity in shader binary UUID Connor Abbott (2): - tu: Fix resetting command streams with writeable BOs - tu: Fix condition for skipping emitting aprons Danylo Piliaiev (6): - tu: Custom resolve should always use AVOID_CCU layout - tu: Fix subsampled metadata and blit emission for separate stencil - tu: Fix gfx_write_access checking for TRANSFORM_FEEDBACK_COUNTER_READ_BIT - tu: Fix blit_cache_cleaned never being set to true - nir: Include scalarized component offsets in UBO ranges - tu: Dirty LRZ after changing attachment locations disable LRZ writes David Rosca (6): - vulkan/video: Fix coding AV1 decoder/encoder_buffer_delay - vulkan/video: Don't code AV1 decoder model info when not present - vulkan/video: Fix coding AV1 operating points - d3d12/video: Don't reset batches in fence_wait - vulkan/video: Fix coding H265 ref pic list modification lists - vulkan/video: Fix coding H265 SPS pcm block sizes, inter ref pic set and lt refs Dmitry Baryshkov (1): - tu: limit VALVE_fragment_density_map_layered to Vulkan 1.1 devices Emma Anholt (1): - freedreno: Don't force image component A=1 substitution on R/RG textures. Eric Engestrom (9): - docs: add sha sum for 26.1.5 - .pick_status.json: Update to 112793b456b492bf2a6e925210c16229f3643a8a - .pick_status.json: Mark cb52837735756ddcac12b3c6b9b32ed6cfb6347e as denominated - .pick_status.json: Mark 7999060992e9cee91d1962faf65dc4e5c6fe4f69 as denominated - .pick_status.json: Mark 2515024a5919ed14fe05471e3f1f89c54a454610 as denominated - .pick_status.json: Mark 4fd93a0039a07ec2027f2a6d1d252c73ed033e0a as denominated - .pick_status.json: Mark adca1f255fd5cb9702db1e0a2bc25d13a521ae5f as denominated - pick-ui: turn commit.date into a (cached) property - pick-ui: show MR number for additional context Erico Nunes (1): - ci: lima farm maintenance Filip Gawin (2): - nv30: fix 1 << 31 issues - nv30: fix another left shift cannot be represented in type 'int' Frank Binns (2): - pvr: rearrange some functions in pvr_arch_border.c - pvr: setup all format fields for custom border color entries Frank Bouwer (1): - pvr: Fix for depth stencil 2d array writes. Georg Lehmann (3): - nir/opt_dead_write_vars: handle atomics as reads - nir: fix divergence for deref_cast - aco/live_var_analysis: make sure shared vgprs are within the encodable vgprs Ian Romanick (1): - brw: Handle empty top block in brw_nir_move_interpolation_to_top Iván Briano (1): - anv: fix 2d-array to 3d blits Jaakko Jokinen (1): - nir: Add cases to nir_get_io_offset_src_number() JaeHoon Lee (7): - v3dv: make room in the descriptor map for the no-sampler entries - v3dv: use the binning VS variant for the binning VPM config - v3dv: record the multiview geometry shader with its Vulkan stage bit - v3dv: record the no-op fragment shader with its Vulkan stage bit - nvk: free copy_memory_indirect_temps on command buffer destroy - v3dv: avoid restoring stale descriptor state after a meta op - nvk: report fills from memory correctly Jan Meisel (2): - nir/range_analysis: handle msad_4x8 in unsigned upper bound - radeonsi/vcn: fail feedback for truncated encodes Jiyu Yang (1): - panfrost: cleanup precomp_cache on screen destroy Jose Maria Casanova Crespo (1): - broadcom/compiler: don't leak the compile on assembly allocation failure Juan A. Suarez Romero (1): - v3d: save fragment constants on sand8/sand30 blit Karol Herbst (3): - rusticl/memory: return 0 for CL_IMAGE_SLICE_PITCH also for 2d images - rusticl/kernel: add libclc source hash to kernel shader keys - vtn/opencl: fix libclc needing fp16 lowering to fp32 Konstantin Seurer (1): - radv/bvh: Fix updating acceleration structures containing AABBs Lionel Landwerlin (6): - anv: flush accumulated barriers for top of TOP_OF_PIPE - anv: fix Wa_18040903259 - vulkan/runtime: fixup vk_shader leak on RT group recompile - anv: add workaround for atomics on R11G11B10 images - brw: fix wa_18019110168 lowering - anv: fix leak in RT binding point Loïc Molinari (1): - u_blitter: Fix out-dated draw_rectangle handler doc Marek Olšák (5): - radv: fix determining the raster prim for guardband - radv: fix determining the raster prim for line mode - radv: fix determining the dynamic raster prim for FS barycentrics - radv: fix determining the static raster prim for FS barycentrics and front_face - radv: don't expose memory types from AMD_device_coherent_memory without the ext Natalie Vock (3): - nir/opt_loop: Don't peel header blocks that jump - radv/nir: Clean up descriptor index lowering - radv: Expose mutable acceleration structure descriptors Olivia Lee (1): - panvk/csf: flush primitives generated query writes from CSF Pavel Ondračka (1): - nv50_ir_ra: align B96 spill slots to vec4 Peng_Lx (1): - turnip/kgsl: close the dma-buf fd of our own allocations Pohsiang (John) Hsu (1): - d3d12: fix msvc build warning C4819 Rhys Perry (5): - vulkan/bvh: fix to_emulated_float(-0.0) - vulkan/bvh: don't update min/max_bounds with inactive nodes - vulkan/bvh: enable SignedZeroInfNanPreserve in leaf.h - vulkan/bvh: limit valid 32-bit node keys to 0xfffffe00 - ac/nir/ngg: don't shrink device-scope memory barriers Ryan Mckeever (1): - docs: advertise VK_KHR_multiview support for Bifrost Sushma Venkatesh Reddy (1): - intel/dev: Clamp PTL+ CS workgroup threads to 32 Tacodiva (1): - vulkan/runtime: Fix bad assumption in GetPipelineBinaryDataKHR Timothy Arceri (6): - zink: fix swap interval changes being dropped - util: add Blockland workaround for crash - nir/opt_dead_write_vars: handle memcpy_deref as reads - util/mesa: add workaround to zero invalidated buffers - util: add workaround for Riddick using round() in glsl 1.20 - llvmpipe: emit FS input vertex attributes in driver location order Toshinari Morikawa (1): - egl: avoid calling loader_get_driver_for_fd with fd = -1 Valentine Burley (5): - ir3: Fix ballot_components for subgroups smaller than 32 - ir3: Derive max_variable_workgroup_size from device limits - ir3: Compute subgroup_size from threadsize_base - tu: Use computed subgroup size - tu: Report correct maxFragmentInputComponents limits Vinson Lee (1): - util/tests: silence unused iterator warning in sparse_bitset_test Yiwei Zhang (3): - venus: properly check sync2 enablement - venus: host image copy to scrub present_src layout if needed - venus: ensure cached vn_ring_submit batches are bounded utzcoz (1): - virtio: magma-gpu-rs: accept a null device in virtgpu_kumquat_finish Mesa 26.2.0 Release Notes / 2026-08-05 ====================================== Mesa 26.2.0 is a new development release. People who are concerned with stability and reliability should stick with a previous release or wait for Mesa 26.2.1. Mesa 26.2.0 implements the OpenGL 4.6 API, but the version reported by glGetString(GL_VERSION) or glGetIntegerv(GL_MAJOR_VERSION) / glGetIntegerv(GL_MINOR_VERSION) depends on the particular driver being used. Some drivers don't support all the features required in OpenGL 4.6. OpenGL 4.6 is **only** available if requested at context creation. Compatibility contexts may report a lower version depending on each driver. Mesa 26.2.0 implements the OpenCL 3.1 API, but the version reported by the CL_DEVICE_VERSION, CL_DEVICE_NUMERIC_VERSION and CL_DEVICE_OPENCL_C_ALL_VERSIONS clGetDeviceInfo queries depends on the particular driver being used. Mesa 26.2.0 implements the Vulkan 1.4 API, but the version reported by the apiVersion property of the VkPhysicalDeviceProperties struct depends on the particular driver being used. SHA checksums ------------- :: TBD. New features ------------ - cl_khr_subgroup_rotate on radeonsi - cl_khr_subgroup_rotate on iris - VK_EXT_shader_uniform_buffer_unsized_array on panvk - VK_KHR_shader_constant_data on RADV - VK_EXT_dynamic_rendering_unused_attachments on panvk - protectedMemory support on RADV/GFX10+ and VEGA10 - VK_KHR_performance_query on RADV/GFX11 - VK_EXT_conservative_rasterization on panvk - shaderImageGatherExtended on pvr - static C++ stdlib required on rusticl to workaround applications using their own C++ stdlib - VK_EXT_pipeline_protected_access on RADV - VK_EXT_extended_dynamic_state3 on panvk - GL_ARB_texture_query_lod on panfrost/v9+ - VK_KHR_maintenance11 on RADV - OpenCL 3.1 support for rusticl on asahi, iris, radeonsi, llvmpipe and zink - VK_KHR_workgroup_memory_explicit_layout on pvr - VK_KHR_maintenance5 on pvr - VK_KHR_calibrated_timestamps on hasvk - VK_KHR_present_id on pvr - VK_KHR_present_wait on pvr - VK_KHR_present_id2 on hasvk - VK_KHR_present_wait2 on hasvk - VK_KHR_shader_fma on RADV - VK_EXT_shader_split_barrier on RADV/GFX12 - VK_KHR_shader_fma on nvk - Support for G1-Ultra, G1-Premium and G1-Pro GPUs on Panfrost and PanVK - VK_EXT_shader_atomic_float on nvk - VK_{KHR,EXT}_index_type_uint8 on pvr - VK_EXT_debug_marker in vulkan runtime - VK_EXT_mesh_shader on NVK - VK_KHR_device_fault on RADV - VK_EXT_present_timing now also on wsi/x11 - VK_EXT_present_timing on hasvk - VK_GOOGLE_display_timing for KHR_display (and opt-in for x11, wayland) on anv, hasvk, hk, nvk, panvk, pvr, radv, tu, v3dv - VK_KHR_shader_abort on RADV - VK_KHR_shader_fma on panvk - VK_NV_shader_atomic_float16_vector on NVK - VK_KHR_compute_shader_derivatives on panvk - VK_EXT_shader_subgroup_ballot on pvr - VK_EXT_shader_subgroup_vote on pvr - VK_KHR_shader_subgroup_rotate on pvr - VK_KHR_shader_subgroup_uniform_control_flow on pvr - VK_EXT_subgroup_size_control on pvr - VK_EXT_rasterization_order_attachment_access on panvk - VK_ARM_rasterization_order_attachment_access on panvk - GL_OES_texture_float on etnaviv/HALF_FLOAT - GL_ARB_texture_float on etnaviv/HALF_FLOAT - VK_NVX_binary_import on NVK - VK_EXT_image_sliced_view_of_3d on panvk - VK_EXT_shader_tile_image on panvk - VK_KHR_internally_synchronized_queues on panvk - VK_KHR_workgroup_memory_explicit_layout on panvk - VK_EXT_device_memory_report on pvr - VK_EXT_descriptor_heap enabled by default on anv, RADV - VK_EXT_device_address_binding_report on panvk - VK_EXT_multisampled_render_to_single_sampled on RADV/Android - VK_KHR_shader_fma on anv - VK_KHR_extended_flags on RADV - VK_EXT_display_surface_counter on pvr - VK_EXT_display_control on pvr - VK_EXT_direct_mode_display on pvr - VK_KHR_surface_maintenance1 on pvr - VK_EXT_surface_maintenance1 on pvr - VK_KHR_swapchain_maintenance1 on pvr - VK_EXT_swapchain_maintenance1 on pvr - VK_EXT_swapchain_colorspace on pvr - VK_EXT_acquire_drm_display on pvr - VK_KHR_unified_image_layouts on pvr - VK_KHR_shader_quad_control on v3dv - VK_KHR_shader_subgroup_rotate on v3dv - VK_KHR_shader_maximal_reconvergence on v3dv - VK_EXT_shader_image_atomic_int64 on panvk - VK_EXT_host_image_copy on RADV/GFX10.3+ Bug fixes --------- - "bad target in _mesa_select_tex_object()" when using glXBindTexImageEXT - 7900XT, hardware acceleration crashes chromium-based apps unless LIBVA_DRIVER_NAME=null - A702: assertions at A6XX_TEX_MEMOBJ_4_BASE_LO - ANV: dEQP ASTC tests crash w/ FPE_INTDIV on Xe3 - AV1 videos dropping frames with AMD card - After updating GStreamer, all videos in Showtime are green/purple - Ambient occlusion is broken in The Chronicles of Riddick - Assault on Dark Athena - Blockland crashes when shader quality turned up - Broken rendering in Isonzo (Unity) since vulkan-radeon 26.1.0 - Celestia hit fallback on r300g from git? - Confidential issue #15670 - Copy paste bug in \`gallium/drivers/radeonsi/si_texture.c` - Copy paste bug in \`vulkan/runtime/vk_debug_utils.c` - Crash linking cached GLSL shader with optimized array with explicit uniform location - D3D12: last vertex's non-zero-offset attribute fetched as 0 with a tightly-packed interleaved vertex buffer (AMD D3D12 only, regression since 24.1) - DaVinci Resolve 20.2 opencl-mesa crash - Dishonored 2 flickering menus on BMG and DG2 - Doom Eternal - Page Fault - somewhat reproducable. (RX9070) - Gunfire Reborn crashes with ring gfx_0.0.0 timeout since Mesa 26.0.0 - Horizon Forbidden West misrendred lighting effects on BMG - Intel binaries using GLX crash on ARM Macs using llvmpipe - Is maxFragmentCombinedOutputResources=16 in Honeykrisp reflects an actual HW limit? - Issues with shadows on DOOM: The Dark Ages - Revelations DLC - Mesa fails to build due to rust bindings error - Minecraft with Complementary Shaders and Voxy LOD rendering issue on Radeon 680M on latest mesa main branch - NIR: loop unrolling should not create phis for constants defined in the loop - NVK The Surge 2 misrendering near text - NVK: Rendering corruption in Shadow of the Tomb Raider - OpenGL app deadlocks in loader_dri3_swap_buffers_msc → xcb_wait_for_special_event on radeonsi / Mesa 26.1 / X11 (Telegram Desktop) - Qt application TrenchBroom hangs in glXSwapBuffers in recent version of Mesa - RADV NIR optimization bug - RADV/RX 9070 XT: WUCHANG: Fallen Feathers gpu ring reset - RADV: RDNA2 Raytracing regression after "aco/lower_branches: Add try_rotate_latch_block() optimization" - RADV_DEBUG=bo_history crashes any vkd3d/dxvk game at startup with heap corruption (double free) - RADV_PERFTEST=transfer_queue causes assertion failure in prime scenario - Regression. VTK polydata. gl_PrimitiveID requires explicit geometry shader with versions 25.3.6 and later. - Rusticl causes a crash after compiling an OpenCL program - SIGABRT: Invalid free in zink_destroy_resource_surface_cache() - Spike in mmap count for vk_cmd_queue regressed dEQP-VK.api.object_management.max_concurrent#command_buffer_* - System freeze when seeking in h264 files with gst-play-1.0 using the VA plugin - The End is Nigh (Wine): No lighting in The Hollows - Uplink text rendering still bugged out - [AC/NIR] XPlane 12 failure to start - [ANV] Intel arc b580 | Halo Infinite misrenders and GPU hangs - [ANV][ARC] Space engineers 2 Artifacts - [ANV][LNL] - Split Fiction (2001120) - Vertex explosion on white cloth during tutorial. - [ANV][PTL] - Elden Ring (1245620) - Blue artifacts and flickering lights with raytracing enabled - [ANV][PTL] - Horizon: Forbidden West regression in shader quirk - [ANV][PTL] - Persona 3 Reload (2161700) - Blue color on some objects + reflections do not reflect - [Arc B580][Baldurs Gate 3] Hang on opening inventory - [BMG] regression: Artifacts in GTK applications in 26.0.x - [GR Breakpoint] Inverted frustum culling for grass meshes on Linux - [PTL] dEQP-VK.synchronization2.op.single_queue.event.write_fill_buffer_read_ubo_tess_control.buffer_16384_maintenance9 fails - [RADV/ACO]: (Bisected) Regression causes flashing reflections in CyperPunk 2077 - [RADV] REGRESSION: Lighting Bleed-through in Need for Speed games on mesa 1:26.1.1-2 ArchLinux - [RADV] REGRESSION: Weird graphics glitches and Lighting Bleed-through in Saints Row 2 - [RADV] Regression: Infinite loop in NIR compiler during nir_opt_dead_write_vars loading Gemma 4 via llama.cpp on gfx1100 - [RADV] Regression: Sackboy - A big Adventure, crash if RT is enabled - [RADV] Video color artifacts in mpv - [RADV][RDNA3][regression] GPU context lost in Sushi Ben (Steam 2419240) when using in-game camera - [Security] gallium VA-API AV1 tile_info(): unbounded loop writes past fixed tile_col_start_sb/width_in_sbs arrays on heap - [VAAPI] VRAM leak on Polaris, 26.1 regression - [VAAPI][Feature Request] Support of VAProcPipelineCaps.blend_flags in radeonsi (vf_overlay_vaapi) - [VAOn12] Thread safety issues - [VC4/V3D] GL_EXT_shadow_samplers exposed but GL_TEXTURE_COMPARE_MODE/FUNC returns GL_INVALID_ENUM in GLES2 - [Windows][arm64] Building windows ARM64 with MSVC fails on src\\util\\u_math.c - [Xe][ARC770][Baldurs Gate 3] Hang during shader compilation - [amdgpu] Little Inferno rendering issues starting with Mesa 21.1 - [anv] Intel ARC B390 | Hitman 2 | DX11 | Blinking corruptions seen - [anv] Intel ARC B390 | Starfield | DX12 | Crash after starting the game - [anv] [bmg] Halo Infinite will not start rendering on an Arc B580 - [anv] dpas/Vulkan coopmat uses ~double registers due to scalar lowering - [anv] negate of INT_MAX calculated as INT_MIN instead of INT_MIN + 1 - [anv] wrong value for local variable after nested switch merge - [meson][etnaviv] Build target etnaviv_isa_rs has no sources - [r300] [big-endian] bad colors under memory pressure - [radeonsi/VCN] RX 9070 XT (Navi 48, gfx1201): VCN unified ring timeout during VAAPI HEVC encode from Steam Game Recording - [radeonsi] SIGSEGV in gallivm during software fallback for GL_FEEDBACK with lighting - [radeonsi] eglinfo 'libEGL warning: failed to get driver name for fd -1' - [radv] Regression causes GPU page faults in Crimson Desert - [radv] Regression causes black patches on the ground in DOOM The Dark Ages Revelations DLC - android14 gpu virgl: Shader memory leak - anv: Assert in alias vkCreateImage - anv: Missing null check in vkCmdEndTransformFeedback - anv: NULL deref in anv_AllocateMemory when importing DMA-BUF without VkMemoryDedicatedAllocateInfo on Xe2 (regression) - anv: World of Warcraft dx11 CMAA 2 option buggy on a b580 - anv: anv_load_fp64_shader consumes 12MB of heap at device creation - anv: atan shader compiler bug leads to negative infinity - anv: bottom-of-pipe timestamp latched before vkCmdDispatch completes (Arc B390 / Panther Lake, Mesa 26.1.4), breaking wgpu compute-pass timings - anv: cooperative matrix loads from shared memory return the wrong tile on Arc B390 (Xe3/PTL), matmul results are wrong - asahi: broken image? rendering in firefox with active hw-acceleration since mesa-26.0.5 - brw/decoder: fragment shader decoding issue - brw: Blender performance regression for Barbershop benchmark scene - build: intel_eu_stall_viewer fails on 32-bit platform - ci: arm64 lava kernel missing zstd support - d3d12: stretched + black-bar artifact after fast offscreen FBO resize on Adreno (regression in MR 41322) - ethosu: Build error on 32-bit due to %lu on uint64_t - gallium/vl: scale_vaapi limited range RGB to limited range YUV broken with red tint - glcpp: incorrect macro expansion in token pasting - intel/blorp: VK_ANDROID_external_format regression on buildtype=debug - intel/blorp: shader_pipeline initialization results in incorrect blorp_key values - intel: Investigate HIZ Plane Optimization disable bit for gfx12.5 - ir3: THREAD64→THREAD128 heuristic change in 25.2.0 causes vertical line artifacts in No Man's Sky on Adreno 740 (a7xx) - ir3: ir3_cf brokenness - iris: Blender wireframe rendering broken - iris: unhandled -EAGAIN in iris_batch_flush (26.1 regression) - kk: Implement timestamps - kk: stencil test unexpectedly not working - lavapipe: dynamic state for advanced blend is broken - llvm23 breaks build of clc_helpers.cpp - llvm23 commit d50631f breaks build of ac_llvm_helper.cpp - macOS build stuck in infinite loop since 26.1 - mediafoundation: error C2039: 'step': is not a member of '_inputQPSettings' - mesa 26.1.3 does not compile with ARM64 MSVC - mesa: clean up st_context caps flags - mesa: glthread crash with MESA_VERBOSE=api - nir: possible exactness bug in reassociate - nir_opt_copy_prop_vars validation failure with Rusticl OpenCL kernel shaders - nvk: various float_controls cts test fails on Turing only - nvk: zcull causes SAVE_RESTORE_ADDR_OOB crashes in Horizon Forbidden West - panvk: Failures in new (1.4.5.0) texturequerylod*trilinear CTS tests - r300 bisected : Broken rendering with applications using MSAA visuals - r300: dEQP-GLES2.functional.shaders.struct.uniform.(not\_)equal_fragment regression - r600, sfn: Lowering to assembly failed on R700 while running ShooterGame demo - r600: SFN assersion failed in Transport fever 2 - radeonsi+ACO jobs should use ACO_DEBUG=validatera - radeonsi/VCN: AV1/HEVC encode reports a coded size larger than the coded buffer, client segfaults reading the mapped buffer - radeonsi: sqtt missing data for viewperf - radv: Use more efficient cache uuid for hardware identifiers - radv: acceleration structure update (refit) of AABB geometry loses intersection candidates (NAVI32, Mesa 26.1.4) - radv: descriptor heaps do not support non-uniform indexing on buffer pointers - radv: incorrect memory accounting for imported bufffers - radv: use preload sgprs for 16/8bit push constant loads - regression;bisected;radeonsi/video: corrupted VA-API H.264 video encoding (bisected to 4487162a) - rusticl + v3d assertion failure - rusticl/radeonsi OpenCL behavior regression after 5a298f3560629943e9140c74547183da1635352e - rusticl: incorrect float16 constant folding - segfault v3d raspberry Pi5 h265 hw decoding / regression mesa 26.1.3/26.1.4 - some dEQP-VK.image.host_image_copy failures on xe2/Xe3 with 32bit - std430 layout not supported on older GLSL profiles with GL_ARB_shader_storage_buffer_object extension - the new mesa 26.1.0-1 is causing mutliple graphical issues on intel i5-2400 and higher (integrated graphics) - tu: assertions in ir3_ra.c: Assertion \`physreg != (physreg_t)~0' failed - turnip: Blender: viewport contents invisible with UBWC - v3dv: Compute shader crashes for unknown reason - venus: rare flake in dEQP-VK.wsi.android.swapchain.render.basic - venus: typo in vn_queue_submit_2_to_1 - vulkan/runtime: GetPipelineBinaryDataKHR incorrectly assumes \`pPipelineBinaryDataSize` must be 0 initialized - vulkan/wsi/win32: bgcolor of Vulkan app window becomes lighter under venus - vulkan: 5.10 and 5.15 LTS kernels require implicit sync support - wsi: Suspicious assertion in wsi_create_buffer_blit_context - zink/ci: switch all traces jobs to surfaceless+gbm Changes ------- Adam Jackson (7): - zink: consolidate resource_create error paths - zink: extract format list setup from create_image - zink: extract pNext chain construction from create_image - zink: extract memory binding from create_image - zink: replace image negotiation with candidate-based approach - zink: stop find_good_mod from mutating ici in place - nvk: use unsigned comparison in UBO bounds checks Adam Stylinski (1): - nv30: fix an issue when this push buffer is NULL Aditya Swarup (2): - anv/pps: Use counter block to stay consistent with Perfetto - intel/test: Add support for Perfetto counter groups Adrián Larumbe (10): - pan/kmod: Fix minor version number check for USER_MMIO_OFFSET ioctl - pan/kmod: fix double syncop count sum when populating vm_bind syncs - drm-uapi: Sync the panthor header - pan/kmod: Use kernel-reported page sizes for new VM when available - pan/kmod: Pass signal and wait syncs separately - pan/kmod: Handle sync object signals in Panthor's vm_bind - pan/kmod: Introduce sparse binding - pan/kmod: Introduce vm_op buffering and sparse mapping emulation - panvk: Use pankmod instead of panthor drm interfaces in bind queues - panvk: Talk directly to pankmod when binding sparse resources Agate, Jesse (2): - amd/vpelib: separating frontend programming - amd/vpelib: Fix blending hang issue Ahmed Hesham (10): - pan/bi: Restore b3210 as a valid swizzle - pan/nir: Fix 8 and 16 bool reduction lowering - pan/bi: Fix function temp lowering with 64-bit pointers - pan/bi: Fix MKVEC.v2i8 src2 swizzle lowering - pan: report async CSF group faults via context reset status - clc: fix fp16 fallback mask for remquo - nir: fix vectorising phis with mixed chased sources - rusticl: enable panfrost by default - ci: Add OpenCL-CTS to GL test infrastructure - pan/ci: Add OpenCL-CTS quick job for Mali-G610 Aitor Camacho (37): - kk: Add poly dependency to KK - kk: Reuse as much poly utilities as possible for unrolling - kk: Increase maxFragmentCombinedOutputResources to KK_MAX_DESCRIPTORS - kk: Add ds state to fragment key since it's part of the pipeline we compile - kk: Fix subgroup failures on M1/2 due to bcsel - kk: Fix global_store writemask - kk: Rewrite force position output pass to use lowered io - kk: Add residency set to queues - kk: Add grid struct for dispatches for convenience - kk: Correctly report failures when compiling precompiled shaders - kk: Rework draw dispatch - kk: Rework shader compilation to handle more than 2 stages - kk: Implement tessellation - kk: Use index element size instead of Metal enum to avoid asserts - kk: Use subgroups for tessellation prefix count since they are now fixed - kk: Move poly data out of root buffer - kk: Clean up per draw upload for tessellation stage - kk: Move to Metal4 command encoding - kk: Add GPU hang detection - kk: Disable workarounds 1-6 in macOS 27 - kk: Reduce root buffer pointer by replacing it with the GPU address - kk: Expose texture max dimensions based on GPU family - poly/lower_tcs: Preserve TCS barriers against lowered outputs - kk: refold combined image/sampler packing after vars_to_ssa - kk: defer cmd buffer submission and lighten compute barriers - kk: Record command buffers live and replay only on resubmit - kk: Fix flrp signed-zero preservation with float_controls2 - nir: compute acos in 32-bit then downgrade to 16-bit - kk: Fix metal import assert - kk: Handle per alu math controls in MSL - kk: Use isnan(x) for x != x and !isnan(x) for x == x - kk: Compile all shaders with fast math - kk: Expose shaderSignedZeroInfNanPreserveFloat16/32 - kk: Implement VK_KHR_shader_float_controls2 - kk: Metal's precise functions are only fp32 - kk: expose Vulkan 1.4 - kk: Fix icd json api version Alejandro Piñeiro (9): - pan/midgard: reorder nir_shader_compiler_options alphabetically - panfrost: add explicit casts when assigning ~0 and 0 to enum-typed fields - cs_builder: fix trailing comma in cs_builder_init - panfrost: track active endpoint scoreboard slot - panfrost: define scoreboard slots - panfrost: add Perfetto render stage tracing support (v10+) - panfrost: add u_trace indirect capture for compute dispatches (v10+) - panfrost: add Batch, Barrier and CacheFlush Perfetto render stage (v10+) - docs/perfetto: panfrost now supports render stages Aleksi Sapon (2): - llvmpipe: remove unused SSE rasterization code - llvmpipe: fix overflow in rasterizer Alessandro Astone (4): - gallivm: Fix armhf build against LLVM 22 - radv: Support explicit DRM format modifier from android gralloc - anv: Support VK_ANDROID_native_buffer older than version 11 - hasvk: Fix android build with android-strict=false Alexander Slobodeniuk (1): - radeonsi: fix conformance window emission in the SPS Ali, Nawwar (1): - amd/vpelib: update shaper config size Allen Ballway (2): - vulkan/android: Set COLOR_ATTACHMENT_BIT for external format resolve - vulkan/android: Map AHARDWAREBUFFER_FORMAT_Y8 to VK_FORMAT_R8_UNORM Alyssa Rosenzweig (288): - nir/opt_generate_bfi: avoid trivial instructions - jay: strengthen assert - jay: drop dead code - jay: generalize last kill code - jay: reduce zeroing - jay: reduce calloc to malloc when memsetting after - jay: fix SEL implied pipe - jay: fix the source pinning code - jay/register_allocate: use standard builder name - jay/opt_dead_code: handle predication - jay/lower_pre_ra: skip predication - jay/assign_flags: handle predicated CMP - jay/register_allocate: tie predicated-defaults - jay: allow predication of pure-flag instrs - jay: fold logic ops - jay/test-optimizer: fuse before/after cases - jay: test logic op fusing - jay: fix simd32 deswizzle - jay: improve spiller debug - jay: fix spiller coupling code - jay: don't print internal without the flag - jay/register_allocate: don't depend on indexing - jay: call DCE an extra time - jay: refuse to propagate ADDRESS copies - jay: relax mov type check - jay: fix SEL types - jay/print: deal with bare r0 copies - jay/to_binary: handle packing accumulators - jay: validate non-SSA accumulators - jay/register_allocate: start using accumulators - jay/lower_post_ra: remove SWAP macro - jay/lower_post_ra: drop old 2<-->8 lowering - jay/ra: don't reserve registers when not spilling - jay/ra: use accumulator for memory copies - jay/ra: use accumulator for memory swaps - jay/ra: use accumulator for stride=4 swaps - jay/ra: drop memory copy reordering - jay/ra: only use stride=4 temps - gallium: Drop users of post-processing filters - gallium: Drop post-processing filters - nir/opt_algebraic: add redundant u2u32/unpack_64_2x32_split_x patterns - brw/nir_lower_cs_intrinsics: do some math at 16-bit - nir/opt_reassociate: fix exactness bug - jay/assign_flags: refactor for next commit - jay/assign_flags: don't burn a null flag - jay/assign_flags: don't burn a flag for ballots - jay: shrink stack allocation - jay: jayize swsb print - jay: consolidate file prefixes - jay: fix 16-bit predicated compares - jay: drop jay_exec_mask - jay: inline jay_control() - jay/opt_propagate: fold uflag copies - jay/opt_propagate: disable f64 opts for now - jay: introduce a physical control flow graph - jay: drop UGPR->UMEM spilling path - jay: convert to LCSSA - jay: do not copyprop ballots globally - jay: propagate inverse-ballots only locally - jay: check for inverse-ballots in jay_uses_flag - jay: smarten predication pass - jay: adjust flag replication - jay: predicate NoMask instructions in uniform IF's - jay: drop a bunch of stale TODO and XXX - jay/to_binary: rename grf -> phys_reg - jay/to_binary: fix packing of simd-split accumulators - jay: model MAC - jay: do moves on the float pipe where possible - jay: move simd32 deswizzling to float pipe - jay: assign accumulators post-RA - jay/lower_scoreboard: elide more dependencies - jay/lower_scoreboard: refactor wait pipe code - jay/lower_scoreboard: fix tracking for A\@* and \*\@7 - jay/lower_scoreboard: refactor SYNC.nop insertion - jay/lower_scoreboard: use .src annotations - jay/lower_scoreboard: be the sole emitter of SYNC - jay/lower_scoreboard: use SYNC.allrd/allwr - jay: swap predication/acc pass order - jay: add JAY_DEBUG=noacc option - jay: fix bfn with 0xffff constant - jay: elide atomic dests - jay: optimize pack_32_2x16_split(#0, x) - jay: make indirect push data blow up more obviously - jay: fix comment - jay: have proper UNDEF - jay: clarify development model - jay/lower_scoreboard: fix trivial scheduling - jay/lower_scoreboard: refactor - jay/lower_scoreboard: run RegDist globally - jay/lower_scoreboard: factor regdist logic out - jay/lower_scoreboard: control flow is int pipe - jay/lower_scoreboard: compact inst_exec_pipe - jay/lower_scoreboard: rename gpr_range -> key - jay/lower_scoreboard: use CFG for RegDist scoreboarding - jay/lower_scoreboard: use sbid syncs to elide regdist deps - jay/register_allocate: set num_regs[MEM] properly - jay/register_allocate: tweak roundrobin heuristic - jay/opt_propagate: fix NOT propagation - jay/opt_propagate: propagate undefs - pan/mdg: make clang warning quiet - jay: relax fragment payload layout - jay: insert simd32 deswizzle in a dedicated pass - jay/liveness: remove pointless bitset init - jay/liveness: speed up physical CFG merging - jay/liveness: drop redundant source filtering - jay/lower_scoreboard: handle accumulator hazard - jay/lower_scoreboard: add asserts on key bounds - jay/opt_propagate: avoid branching on poison - jay: annotate pure sends - jay: factor jay_op_(starts,ends)_block queries - jay: schedule for pressure - jay: fix omask on single sample - jay: hack for sample position - jay: use new fs payload variable more - brw/eu_validate: relax EOT requirements on Xe2 - CODEOWNERS: add Jay - brw,jay: add use_src_xy prog data field - jay/liveness: use jay_foreach_preload - jay: legalize shuffle(ugpr) for now - jay/spill: fix reload array size issues - jay/register_allocate: inline silly helper - jay/register_allocate: remove out of date comment - jay/lower_spill: use 1 less temporary - jay: fix FS reading too many sysvals - jay: add zip_ugpr16 instruction - jay: pack jay_stride - jay: stop asking for stride=4 ugpr's - jay/validate_ra: use jay_def_stride - jay: drop unneeded #include - jay: rewrite partition handling - jay: merge partition blocks - jay: remove send split hack - jay: allow simd32 gl_SamplePosition - jay: avoid overflow affinities with large UGPR vecs - jay: allow SIMD1 imageStore() - jay: hide MAD->MAC behind !JAY_DEBUG=strict - jay: gate early EOT code behind =strict - jay: renumber reg files predictably - jay: simplify uniformity checks - jay: allow npot operands in RA - brw: nir_lower_constant_convert_alu_types only once - intel/gen: remove dead #include - intel/gen: drop noisy build spam - jay/validate_ra: validate against partition - jay/register_allocate: do not treat reserved regs as free - jay/register_allocate: remove remnant of old partition code - jay/register_allocate: split out jay_stride.c - jay/register_allocate: drop #include - jay/partition: validate we don't generate g127<2> - jay/partition: pick better partitions - jay: limit stencil export to simd16 - jay/to_binary: relax packed float restriction - jay: generalize jay_extract_range_post_ra - jay: rework lane ID calculations - jay: replace GPR_FROM_UGPRs with a simple CVT - jay: replace BYTE/WORD_PACK with a simple MOV - jay/register_allocate: don't hang if a block is missing - jay/partition: reduce 16-bit partitioning more - jay: drop dead if - jay: clang-format - jay/lower_scoreboard: allow multiple jumps - jay: lower JAY_OPCODE_LOOP_ONCE earlier - jay/to_binary: big clean up post-gen - jay: avoid bogus copyprop with cmods - jay: add unit test for bogus copyprop case - jay/validate: add validation for bogus uflag cases - jay: fix 8/16-bit inline_data loads - jay/assign_flags: require ballots to be in the balloted src - jay: fix mismatched files with predication - jay: workaround the while bug - jay: follow source order for mad/bfe - jay: fix last-use accounting with ARF sources - jay: allow null in jay_collect_vectors - jay: uniformize bti indirects - jay: optimize out more early eot related copies - jay/register_allocate: make phi webs conservative - bin: add drm-shim script - jay/lower_pre_ra: allow immediate on bfe - anv: enable jay ray query - nir/opt_sink: sink more Intel block instructions - nir/lower_terminate_to_demote: tweak terminate_if lowering - jay/spill: spill at definitions - jay/spill: do initial find-and-replace for ugpr spilling - jay/spill: unstub rematerialization - jay/spill: refactor - jay/spill: drop sketchy heuristic - jay/spill: simplify limit() - jay: fix bogus unit tests - jay/validate: check for mixed ugpr/gpr problems - jay/register_allocate: simplify split copy logic - jay/register_allocate: don't search a 2nd UGPR temp - jay: allow more 3-src imms - jay: introduce accumulators into the partition - jay: pool constants per block under pressure - jay: remove #include - jay: cache message headers locally - jay: forbid 8-bit immediate prop - jay: autopep8 - jay: manually format jay_type_for_glsl_base_type - jay: clang-format - jay: track skip_helpers - jay: rewrite demote/terminate/helper/halt handling - nir/opt_dead_cf: delete redundant returns/halts - intel, nir: Add {load,store}_global_intel intrinsics - jay: distinguish physical & logical loop headers - jay/validate: validate backedges - jay/lower_scoreboard: simplify trivial swsb - jay/lower_scoreboard: fix barriers in trivial SWSB - jay/test: drop SSA repair tests - jay/to_binary: use dedicated addr reg for shuffle - jay/lower_pre_ra: fix oob read - jay/register_allocate: fix file prefixes - jay/opt_dead_code: drop stale todo - jay/opt_dead_code: handle phi properly - jay/spill: add an assert() - jay/spill: repair as we go - jay/spill: implement ugpr spilling - jay/spill: don't try to remat mov_imm64 - jay/spill: do lazy reloading instead - jay: remove unused SSA repair pass - jay: drop #include - jay/assign_flags: fix ballot handling - jay: fix barycentrics - jay: improve the stride partition heuristic - jay/lower_spill: rename to make easier to follow regs - jay/lower_pre_ra: fix f64 negate - jay: clang-format - Revert "rusticl: fix leak in \`util_queue`" - jay: fix EOT with indirect message descriptors - intel: make more multisampling/coarse state static - jay/to_binary: avoid overflowing address reg - jay: add jay_bare_regs helper - jay: use jay_bare_regs - jay: drop jay_extract_range_post_ra - jay: rename post-sched lowering - jay: zero a0 before divergent shuffles - jay: consolidate spilling call into a single file - jay: fix JAY_DEBUG=spill - jay/lower_spill: set cursor explicitly - jay: fix ugpr reloads in divergent control flow - jay: fix printing internal shaders - people: add Lucas Fryzek - util: add __bitset_zero, __bitset_copy helpers - treewide: use __bitset_{zero,copy} - jay: remove #define - jay: unify inline_data and push_data loads - jay: remove duplicated bf16 check - jay: remove pointless (and wrong) bf16 check - jay: drop #include - jay: treat init_helpers as a start instruction - jay: print the shader after spilling before RA - jay/print: make predication syntax more explicit - jay/schedule: move schedule into ctx - jay/schedule: move function into context - jay/dag: inline dag code - jay/dag: defer parent_count calculation - jay/dag: split into a mutable and immutable part - jay/dag: add jay_dag_print helper for debug - jay/dag: add jay_dag_iterator_reset helper - jay/schedule: add per-block data structure - jay/schedule: edit comment about pressure - jay/schedule: split block analysis from scheduling - jay/schedule: flip the signs on pressure calcs - jay/schedule: account for demand per-file - jay/schedule: fix accounting for dead flags - jay/schedule: add missing whitespace - jay/schedule: early-out invalid schedules as we go - jay: add mlen to SEND - jay: add real cycle model - nir: lower boolean shuffle without subgroup size - nir/opt_algebraic: optimize ~x != x - gen/print: include pipes for jay_print - intel/gen: allow dead loops - iris: use u_default_get_sample_position - crocus: use u_default_get_sample_position - intel: remove intel_get_sample_positions getter - anv: use vk_standard_sample_locations - hasvk: use vk_standard_sample_locations - jay/spill: fix variable shadowing - jay/lower_post_ra: fix simd16 flag zeroing - jay/opt_propagate: fix inverse_ballot(bfn) - jay/lower_helpers: fix unconditional discard - jay: lower boolean shuffles - jay: make uniformity explicit - jay: support as_uniform - jay: drop #include - jay: rewrite flags - intel: fuse off Jay in Mesa 26.2 Andrzej Datczuk (2): - radv: enable advertising of VK_KHR_pipeline_library under llvm - radv/rra,rmv: fix device id written into trace files Anna Maniscalco (1): - ir3: Skip preftech and and warmups for non bindless earlier Arjob Mukherjee (2): - pvr: increase value of maxPerStageDescriptorStorageBuffers - pvr: increase maxPerStageDescriptorStorageBuffers to 16 Arkady Shlykov (1): - Add offset getter for image intrinsics Arzaq Naufail Khan (1): - spirv: fix resource leak in spirv shader replacement Ashley Smith (2): - panfrost: Fix tiler_desc assignment - panfrost: Avoid race between bo import and unref Assadian, Navid (1): - amd/vpelib: FP16 non linear handling Autumn Ashton (5): - nak: Expose max_warps_per_sm - nvk: Allow nvk_cmd_upload_qmd to take a custom root descriptor - nvk: Add nvk_cmd_dispatch_with_root - nouveau/cubin: Add cubin and fatbin parsers - nvk: Implement VK_NVX_binary_import Benjamin Cheng (19): - ac/vcn: Rename VCN5 swizzle mode to GFX12 - radv/video_enc: Use correct swizzle mode for VCN5 with GFX11 - radv/wsi: Re-use transfer queue if it exists - ac/parse_ib: Add parsing for variable slice mode - radeonsi/video: Cleanup dpb buffer - radeonsi/mm: Disable variable slices when bad input is found - util/ycbcr: Fix adjust_to_range - util/ycbcr: Add a narrow range RGB coeff helper - gallium/vl: Fix RGB narrow range conversions - draw: Add lower_opcodes NIR pass - mesa/st: run the lower_opcodes pass for draw shaders - radv/video: Set accurate minQp/QIndex - radv/video: Report MULTIPLE_SLICE_SEGMENTS_PER_TILE_BIT - ac/video: Add {min,max}_qp to video enc caps - radv/video: Use {min,max}_qp caps from ac - mesa/st: use col0 attrib from provoking vertex for feedback - ac/surface: Remove GFX10 limitation for FORCE_SWIZZLE_MODE - ac/vcn_enc: Disable var slice for preencode with VCN4 - vl/proc: Guard compositor creation with HAVE_GFX_COMPUTE Benjamin Gaignard (1): - pan/format: Advertise support for AFBC(32x8,sparse) Benjamin Otte (1): - Revert "lavapipe: Don't advertise support for multiplane drm formats" Benoît du Garreau (1): - docs: Add many missing features Bo Hu (3): - codegen: update scripts/cereal/decoder.py - vk-snapshot: do not generate code to save vkQueueFlushCommandsGOOGLE - gfxstream: vk-snapshot: update handling of bufferview in vkUpdateDescriptorSets Boris Brezillon (9): - pan/kmod: Don't pass drmVersionPtr objects around - kraid: Fix cross-build - pan/kmod: Add a pan_kmod_timestamp_cycles_to_ns() helper - pan/props: Make pan_query_core_count() safe with wide shader_present bitmaps - pan/props: Split core_count/core_id_range into two helpers - pan/props: Add pan_query_perf_counter_per_block() - pan/perf: Start relying on canonical mali_perf definitions - pan/perf: Replace pan_perf_init() by pan_perf_{create,destroy}() - pan/perf: Transition to auto-generated derived counters Boyuan Zhang (1): - radeonsi/mm: use non-tmz buf when session tmz size is 0 Brandon Jones (1): - nir/opt_algebraic: fix fabs optimization Brendan King (1): - pvr: move some asserts in pvr_srv_alloc_display_pmr Caio Oliveira (115): - brw: Don't set saturate for SYNC instruction - brw: Use brw prefix to LSC helpers tied to brw - brw: Remove various unused fields - brw: Fix max_dispatch_width collection for CS with variable size - brw: Stop tracking inline parameter usage in prog_key/prog_data - intel/dev: Expose list of known platform names - brw: Fix some indentation in brw_generator.cpp - brw: Move brw_prog_data_init to a different file - brw: Remove references to SIMD4x2 - intel/executor: Map the DPAS check to has_systolic - brw/tests: Remove redundant parser test - brw/tests: Stop using regions/type for null in assembler tests - brw/tests: Stop using regions/type for non-null SEND sources in tests - anv: Remove saturating cmat configurations when INTEL_LOWER_DPAS=1 - anv: When using INTEL_LOWER_DPAS disable BFloat16 cmat configurations - intel: Move cmat configurations to anv_physical_device - intel/perf: Use intel_perf_context as ralloc parent of sample buffers - nir/instr_set: Fix multi-slot intrinsic index equality - intel/perf: Add helpers to get names of enums - intel/perf: Show type, data type and units in intel_perf_query_layout - nir/instr_set: Consider normalization when calculating hash - brw/scoreboard: Add disabled tests for RegDist baking on Xe2+ - brw: Save original regs_written() value in register coalesce - brw: Call size_read() once in regs_read() - brw: Avoid unnecessary calls to size_read() in flags_read() - brw: Pass VGRF numbers to liveness helpers - brw: Don't directly use regs_read/regs_written/size_read as bound for non-trivial loops - brw: Bound register coalesce rewrites by live range - nir: Add print for other cmat_description slots - compiler: Support more than 255 cols/rows in cmat descriptions - intel/compiler: Move bison command to shared meson.build - intel/executor: Add an overflow check for alloc function - intel/executor: Add performance counter support - brw: Use a single brw_compile entrypoint - brw: Move key and prog_data to base compile params - anv: Simplify code that calls brw/jay - iris: Simplify code that calls brw/jay - spirv: Stop warning about ignored invalid ArrayStride decorations - anv, brw: Use previous shader VUE map for FS input layout when available - anv: Fill VERTEX_ELEMENT_STATE before further emissions - anv: Use empty_vs_input for default VERTEX_ELEMENT_STATE - jay: Use TGL_PIPE_NONE when RegDist is zero - intel/gen: Add gen encoding module - intel/gen: Add various to_string/from_string functions - intel/gen: Add validation - intel/gen: Add function to finish structured control flow - intel/gen: Add gen_print() - intel/gen: Add gen_parse() - Revert "intel/dev: Remove unused intel_get_device_info_for_build() function" - intel/gen: Add integrated \`gentool` CLI for asm/disasm - brw: Add temporary workarounds for compatibility with old parser - brw/tests: Port assembler tests to gen module - intel/compiler: Stop replicating narrow immediate values in brw_asm "compat" mode - intel/compiler: Stop forcing null source <0;1,0> region in brw_asm "compat" mode - intel/compiler: Stop forcing null destination HSTRIDE=1 on pre-xe SEND in brw_asm "compat" mode - brw, jay: Use lsc_* symbols from gen - brw, jay: Use gen_swsb and related enums - brw, jay: Use gen_sfid instead of brw_sfid - brw, jay: Use various enums from gen instead of brw - jay: Use gen_condition enum and helper - jay: Use gen module - intel/decoder: Convert to use gen module - intel/executor: Update to use gen module - brw: Make a copy of brw_generator for using gen - brw: Port the copy of generator to use gen_encoding - brw: Use the new gen based generator - brw: Add brw_to_binary() as the single codegen entrypoint - brw: Make brw_generator an implementation detail of brw_to_binary.cpp - anv: Use gen_print instead of brw_disasm - intel/tools: Use gen_print instead of brw_disasm - iris: Use gen_print instead of brw_disasm - intel/compiler: Add and use gen_update_reloc_imm() - brw: Remove old encoder, generator and related tools - intel/gen: Change validation test code to use the parser - jay: Unify macro for NIR passes - jay: Add INTEL_DEBUG=mda support - util: Add runtime parser for boolean lookup tables - intel/gen: Support symbolic print/parse of BFN function - jay: Add helpers for managing unordered instructions - jay: Handle dpas_intel intrinsic - util: Fix float8 denorm rounding to min-normal - intel: build compiler before blorp - intel/gen: Generate opcodes and their metadata - intel/gen: Drop unused format parameter from gen_inst_has_dst - intel/gen: Don't encode zero exec_size - intel/gen: Add a gentool 'check-roundtrip' subcommand - intel/gen: Replace gen_parse_print_test.cpp with text tests - intel/gen: Import test cases from the brw assembler tests - brw: Remove the brw assembler tests - nir: Handle nir_var_mem_push_const in divergence analysis - intel: Change dpas_intel source order to follow DPAS - jay: Handle convert_cmat_intel intrinsic - jay: Add SIMD restriction for Math with HF - brw: Drop dead spill_writes in brw_opt_fill_and_spill - brw: Add unit tests for brw_opt_predicate_logic - brw: Refactor DPAS lowering for HF case - brw: Fix INTEL_LOWER_DPAS=1 for Xe2+ - util: Add util_is_half_subnormal() - brw, elk: Fix invalid case using float-negation in combine constants - intel/gen: Use lookup tables for Gfx12+ short type encoding/decoding - anv: Initialize shader debug archive key size - brw: Fix comma placement when printing memory logical sources - brw: Use shared LSC opcode names in IR printing - brw: Remove dead mixed-size MOV special case from def copy propagation - brw: Fix initializing matrix with uniform but non-constant values - intel/gen: Fix case where IF is a loop header - brw: Report scratch memory size in shader stats - brw: Track logical scratch offsets for spill cleanup - brw: Reuse scratch slots between non-interfering spilled VGRFs - nir: Account for cmat memory accesses in copy_prop_vars - anv: Include build and device identity in shader binary UUID - brw: Match fill/spill optimization scratch accesses by logical offset - intel/gen: Fix Gfx11 3-src accumulator file encoding - nir, spirv: Match debug printf argument alignment with u_printf - nir, spirv: Pad 3-component debug printf arguments to 4 components Caius-Moldovan-img (5): - pco: Replace nir_shader_lower_instructions with nir_shader_*_pass - pvr: set SMP component count for TQ frag load shaders - pco: Fix metadata invalidation - nir: Fix trailing comment generation for variable naming - pco: Remove hardcoded metadata location Calder Young (40): - anv: Fix address bit masking for indirect SBTs - anv: Fix support for indirect SBTs on Xe3+ - anv: Store batch buffers in a null-initialized VMA heap - anv: Add padding to the shader heap to manage EU prefetch - isl: Add usage flag to force SurfaceArray to false - isl: Add additional alignment/padding requirements to prevent overfetch - isl: Optimize the sampler cache to overlap as few 64B cachelines as possible - isl: Add function to calculate the amount of overfetch for an unpadded surface - isl: Add and use isl_tiling_get_intratile_range_el/sa - blorp: Work around sampler overfetch for buffer copies - anv: Make sure robust UBO access does not fault - brw: Avoid rounding every convergent block load up to a full register - brw: Avoid vectorizing loads in NIR if it could extend into a different page - anv: Disable scratch page by default on Xe KMD - intel_hang_replay: Don't force scratch page on Xe KMD unless explicitly requested - isl: Make sure isl_device::requires_padding is always initialized - anv: Fix some usage flags not propagated to ISL for explicit layouts - brw: Allow instruction reordering around memory writes - brw: Add support for ACCESS_CAN_REORDER memory ordering - spirv: Fix debugPrintfEXT not working with multiple arguments - brw: Add workaround pass for shaders using derivatives in control flow - anv: Add workaround for vertex explosions in Split Fiction - jay: Use gen_arf enums instead of jay_arf - jay: Use gen_names.h to print CMODs and ARFs - jay: Do not propagate ARF src unless its src0 - jay: Add support for saturating f2i16 and f2i8 NIR opcodes - jay: Disable avoid_ternary_with_two_constants when using jay - brw: Move ray payload bitfield generation to NIR - brw: Move topology id helper intrinsics to NIR - jay: Disable SIMD32 if ray queries are used - jay: Implement ray tracing topology id intrinsics - jay: Implement ray tracing trace intrinsics - nir: Do not mask helper lanes of writes if ACCESS_INCLUDE_HELPERS is set - anv: Track more error codes from certain IOCTLs - intel: Add common utils for page fault reporting - anv: Add function to get the list of page faults - anv: Print page faults whenever a queue gets banned - anv: Enable support for VK_EXT_device_fault/VK_KHR_device_fault - intel: Add function to get the number of SBIDs from device info - jay: Implement the global SBID scoreboarding pass Caleb Callaway (6): - docs: fix Intel tracepoints.py path - perfetto: v56.1 update - pps: generate forward-looking counters - perfetto: suppress array-bounds warning - perfetto: suppress stringop-overflow warning - anv: hide Intel vendor ID for Cyberpunk Casey Bowman (2): - anv: Add option to disable HiZ via drirc - anv: Set anv_disable_hiz for Sons of the Forest Charmaine Lee (1): - nir_to_tgsi: fix shared memory index Christian Gmeiner (149): - mesa/st: Extend st_context_invalidate_state with meta-op flags - mesa/st: Convert st_cb_drawtex to use st_context_invalidate_state - mesa/st: Convert st_cb_bitmap to use st_context_invalidate_state - mesa/st: Convert st_cb_clear to use st_context_invalidate_state - mesa/st: Convert st_cb_drawpixels to use st_context_invalidate_state - mesa/st: Convert st_cb_readpixels to use st_context_invalidate_state - mesa/st: Convert st_cb_texture to use st_context_invalidate_state - panvk: Advertise VK_EXT_shader_uniform_buffer_unsized_array - panvk: Advertise VK_EXT_dynamic_rendering_unused_attachments - panvk: Implement vkCmdFillBuffer with panlib kernels - panvk: Wire up VK_EXT_conservative_rasterization on v11+ - egl: Switch to mesa_log(..) - etnaviv: Bypass BGRA-internal optimization for shared resources - etnaviv: Convert PE-internal BGRA to RGBA when flushing shared resources - etnaviv: Select texture format dynamically for shared RB_SWAP resources - etnaviv: Add per-RT frag_rb_swap shader key and NIR lowering - etnaviv: Use shader R/B swap for LINEAR_PE shared resources - panvk: Apply sample mask in single-sample mode - panvk: Advertise VK_EXT_extended_dynamic_state3 - compiler/rust: Move VecPair from NAK to shared compiler crate - st/mesa: Zero MaxTextureImageUnits for unsupported stages - etnaviv: blt: Add sRGB support to blt_imginfo - etnaviv: Map R8G8B8A8_SRGB to BLT_FORMAT_A8R8G8B8 - etnaviv: blt: Add BLT format conversion support - lavapipe: Skip advanced blend lowering when blending is disabled - lavapipe: Lower advanced blend at draw time when its state is dynamic - lavapipe: Enable extendedDynamicState3ColorBlendAdvanced - mesa: Allow GL_TEXTURE_IMMUTABLE_LEVELS query on GLES3 - mesa/main: Add trace dispatch plumbing for MESA_VERBOSE=api - mesa/main: Auto-generate MESA_VERBOSE=api trace dispatch - etnaviv: Update headers from rnndb - etnaviv: Use the full 64-bit clear value for 64bpp render targets - etnaviv: Use integer texture formats for R32/RG32 integer textures - etnaviv: blt: Don't sRGB-roundtrip same-encoding copies - etnaviv: Drop unused num_loops shader stat - vulkan/wsi: Constify wsi_instance_supports_google_display_timing(..) - panvk: Advertise VK_GOOGLE_display_timing - panvk: Advertise VK_KHR_shader_fma - panvk: Shuffle local ids for quad derivatives - panvk: Advertise VK_KHR_compute_shader_derivatives - compiler/rust: move ACORN PRNG to shared location - panvk: Move maxFramebuffer limits to defines - panvk: Derive viewport limits from the framebuffer dimension - vulkan/runtime: Track rasterization_order_access in pipeline state - vulkan/runtime: Add rasterization_order_access to dynamic graphics state - panvk: Disable FPK and force late ZS for rasterization order access - panvk: Advertise VK_EXT_rasterization_order_attachment_access - etnaviv: Flush texture caches after clears - etnaviv: Add a helper for the 128-bit second-plane offset - etnaviv: Wrap pipe_framebuffer_state in etna_framebuffer_state - etnaviv: Lay out 128-bit color FBO as paired G32R32F render targets - etnaviv: Emit paired 128-bit sampler descriptors - etnaviv: NIR pass to lower 128-bit color RT and texture access - etnaviv: blt: Use block-layout offset for 128-bit second-plane blit - etnaviv: Save the framebuffer without 128-bit companion slots - etnaviv: rs: Support 128-bit color clears - etnaviv: Limit nir_lower_fragcolor(..) to advertised render targets - etnaviv: Advertise 128-bit color formats as renderable and samplable - etnaviv: Disable TS per render target on mixed TS modes - etnaviv: Update headers from rnndb - etnaviv: Set per-RT sRGB bit on non-zero render target slots - etnaviv: Support split sampler for 128-bit formats on the state path - etnaviv: Gate 128-bit render targets on HALF_FLOAT - pan/texture: Reuse the layer range as a Z-slice range on 3D views - panvk: Slice 3D storage image views on Valhall - panvk: Slice 3D storage image views on Bifrost - panvk: Advertise VK_EXT_image_sliced_view_of_3d - nir: Add load_tile_image intrinsic - spirv: Implement SPV_EXT_shader_tile_image - panvk: Lower tile image reads - panvk: Advertise VK_EXT_shader_tile_image - mesa/main: Keep RealPublished in sync when glthread toggles with api trace - panvk: Iterate the common queue list instead of a per-family array - panvk: Advertise VK_KHR_internally_synchronized_queues - pan/va: Only widen constants on 32-bit instructions - panvk: Advertise VK_KHR_workgroup_memory_explicit_layout - etnaviv: blt: Zero-initialize conv_swizzle - mr-label-maker: Add rule for rust files - nir: Fix lower_fround_even to round half to even - panvk: Add panvk_address_binding_report() helper - panvk: Report address binding for device memory - panvk: Report address binding for buffers - panvk: Report address binding for images - panvk: Report address binding for internal allocations - panvk: Report address binding for descriptor sets - panvk: Advertise VK_EXT_device_address_binding_report - u_transfer_helper: Add U_TRANSFER_HELPER_Z32F_S8_IN_Z24S8 - u_transfer_helper: Convert Z32_FLOAT in the Z32F_S8_IN_Z24S8 path - etnaviv: Route resources through u_transfer_helper - gallium: Add native_fp32_depth cap to gate ARB_depth_buffer_float - etnaviv: Emulate Z32_FLOAT (DEPTH_COMPONENT32F) as D24S8 - etnaviv: Lower depth32f shadow compare in the shader - etnaviv: Emulate Z32_FLOAT_S8X24_UINT (DEPTH32F_STENCIL8) as D24S8 - etnaviv: blt: Address emulated depth32f as D24S8 - etnaviv: Size depth32f transfer staging by the internal format - etnaviv: Support stencil blit of depth32f_stencil8 - etnaviv: rs: Resolve depth surfaces - etnaviv: blt: Resolve depth_component24 - etnaviv: Address emulated depth32f as D24S8 in resource copies - etnaviv: Decide transfer tileability on the physical format - etnaviv: Support MSAA resolve of emulated depth32f - etnaviv/isa: Add bit_insert instruction - nir: Add bitfield_insert_etna opcode - etnaviv: Support native bitfield_insert - etnaviv: hwdb: Add UNIFIED_SAMPLERS feature - etnaviv: Detect unified sampler support - etnaviv: Extract sampler descriptor emit helpers - etnaviv: Rework descriptor emit into a single loop - etnaviv: Implement unified sampler allocation - etnaviv: Allow depth only or stencil only MSAA resolves - etnaviv: Allow MSAA resolve of stencil only buffers - etnaviv: Fix sampler view leak on unsupported texture target - etnaviv: Extract texture descriptor fill helper - etnaviv: Keep texture descriptor template CPU-side - etnaviv: Compose texture descriptors at emit time - etnaviv: Rename etna_sampler_view_update_descriptor() - etnaviv: Add perf debug for texture descriptor recompose - util: Add u_shader_variant_cache - util/u_shader_variant_cache: Add hash + equal callback pair - etnaviv: Adopt u_shader_variant_cache - draw/llvm: Adopt u_shader_variant_cache - llvmpipe: Move FS variant LLVM types to temporary JIT struct - util/u_shader_variant_cache: Add cap + refcount + per-list eviction - llvmpipe: Migrate FS variants to u_shader_variant_cache - llvmpipe: Move setup variant LLVM function pointer to a local - llvmpipe: Migrate setup variants to u_shader_variant_cache - llvmpipe: Move CS variant LLVM types to temporary JIT struct - llvmpipe: Migrate CS/task/mesh variants to u_shader_variant_cache - llvmpipe: Remove dead USE_GLOBAL_LLVM_CONTEXT path - llvmpipe: Move shader caches and LLVMContext to screen - draw: Decouple GS/TCS/TES shader CSOs from draw_context - draw: Decouple VS shader CSO from draw_context - draw: Move current_variant slot from CSO to draw_context - draw/llvm: Bound per-stage variant caches with cap + pin slots - draw: Move per-draw state off the GS/TCS/TES shader CSOs - draw: Decouple mesh shader CSO from draw_context - llvmpipe: Enable PIPE_CAP_SHAREABLE_SHADERS - panvk: Restore push descriptor dirty bit after meta operations - panfrost: Add pan_nir_lower_image_64bit NIR pass - panvk: Alias R64 storage images to R32G32_UINT for color clear - panvk: Bounds-check SSBO accesses for robust storage buffer access - pan/format: Add R64_UINT/R64_SINT format entries for v9+ - nir/lower_robust_access: drop out-of-bounds buffer stores - panvk: Wire up VK_EXT_shader_image_atomic_int64 on v9+ - etnaviv: Only emit index buffer state when it changed - etnaviv: Skip state emission when nothing is dirty - mesa/st: Invalidate FS sampler views after PBO texture transfers - nir/opt_vectorize: Remove the phis that were combined - etnaviv: Drop stale stream output relocs in etna_update_hwxfb(..) Christian Meissl (1): - nir/lower_tex: skip external texture YUV lowering for query instructions Christoph Neuhauser (1): - anv: Add compute only divergent atomics fusion optimization for Blender Blender uses atomic operations as part of its virtual shadow mapping implementation. Virtual shadow mapping page tagging in compute shaders benefits from divergent atomics fusion, while fragment shaders doing the atomic raster step in general have worse performance with this optimization turned on. Thus, an option is added to only apply divergent atomics fusion to compute shaders in ANV, and this option is enabled for Blender. Christoph Pillmayer (19): - pan/bi: Fix source swizzle in bi_repair_ssa - pan/bi: Fix format in bi_repair_ssa - pan/kmod: Fix uninitialized timestamp info - pan: Set nir_shader::source_blake3 for internal shaders - pan/clc: Set source_blake3 for each precompiled variant - pan: Add BIFROST_MESA_DUMP_DIR option to dump shader binaries - pan/kmod: Add l2 features to pan_kmod_dev_props - pan/props: Add BUS_WIDTH query to pan_props - pan/perf: Generate derived counter definition from Arm's XMLs - pan/kmod: Add perf counter api - pan/kmod: Add a perf implementation to the panfrost backend - pan/pps: Delegate more tasks to PanfrostPerf - pan/perf: Make counter sum across block instances optional - pan/pps: Output counters per block - pan/perf: Add timing related getters/setters and use them - pan/perf: Use new kmod api - pan: Fix BIFROST_MESA_DUMP_DIR - pan/kraid: Fix 32bit hw_runner builds - pan/perf: Fix 32bit build of panquick Collabora's Gfx CI Team (21): - Uprev VVL to 8474616c3095756c52c1b810b21bd1366b3fc909 - Uprev VVL to 4acd00c7a0665c9b1d01604e5fe1454837f87134 - Uprev VVL to 6a6182c0edb35cba7bab0abc61eaff82d11022fb - Uprev VVL to d55be6264a17cd28f436805973b12f12a5d22f2f - Uprev ANGLE to 7772c5602d59140204494967ba8ebdf801180054 - Uprev Piglit to 6fd29fe44f8857b876a67bee962919635f22ecc8 - Uprev VVL to 36187ee9f2074609d3ec56fa1a315c191366b688 - Uprev ANGLE to a793c75398c746f3f8a08fd2e74dfc4dff07a0c9 - Uprev VVL to 315d28985ebd1ff9a2e4380e34ed2d8ebe487531 - Uprev ANGLE to 196d1b79eadbd8fdbf0590092266bb2d87264988 - Uprev VVL to d2b091858d12802cc1c3722c81a9f3d865d833d4 - Uprev VVL to 2ab77a01659e3e46d6ca8a25425b19b3425adb11 - Uprev ANGLE to 8e09325ebad45c7e11630a79754361e965e5fab0 - Uprev VVL to e17d63f8fcd967b2ff91efcb8607d2c9ab962e23 - Uprev ANGLE to 836636df1b06034b39a4a2a68b811ecf6f9b674f - Uprev VVL to 5e72d40395d07930c935d0624cd7db5f1a144c5e - Uprev ANGLE to a4eea1fbedace7a03bae52cd1bb9d6ebdfffb0f7 - Uprev VVL to 6906fd94f2f422beb682d43d8b1af872aaff97b0 - Uprev Piglit to 9a0eab5e1f7f009b4f72c25d23faf937b38354a6 - Uprev VVL to e875181c4f0fab1bd0c4926c21a21de0bd967bb3 - Uprev ANGLE to 7a960f346c57daae4b1615ead302b4c58af12258 Connor Abbott (35): - ir3: Don't reset immediate count to 0 after lowering - ir3: Use correct immediate size for constlen calculation - tu: Optimize sync2 event handling in the non-asymmetric case - tu: Don't zero-initialize query pool - turnip, ir3: Use shader for vertex input count - tu: Support VK_KHR_maintenance9 - tu: Fix LRZ+FDM offset+secondaries - tu: Disable LRZ when resuming if the GPU doesn't support tracking - tu: Zero out unused parts of descriptors - compiler/shader_info: Introduce occupancy_bounded_workgroup_fairness - spirv: Remove redundant OpExtInst handling - spirv: Use correct opcode in non-semantic OpExtInst handling - spirv: Implement concurrent workgroup hint from vkd3d-proton - freedreno: Document SP_CS_CNTL_0::COMPUTERRMODEEN - ir3, tu, freedreno: Plumb through round-robin mode - tu: Add TU_DEBUG=computeroundrobin - freedreno: Add round_robin_errata to device info - ir3: Implement round-robin workaround - nir/lower_amul: fix infinity recursion with different phi source order - freedreno/qrisc: Refactor section disassembly - freedreno/qrisc: Emulate multiple firmwares more accurately - freedreno/qrisc: ISA changes for gen8 - freedreno/qrisc: Support DDE on gen8 - freedreno/qrisc: Don't write CP_LPAC_SQE_CNTL on a7xx+ - freedreno/qrisc: Add support for gen8 firmwares - freedreno/qrisc: Add missing to #sqe-base - freedreno/qrisc: Add extra preempt entrypoint - freedreno: Add some gen8 control registers - freedreno/qrisc: Allow limited relocation expressions - freedreno/qrisc: Allow labels on two-src ALU instructions - freedreno/qrisc: Bump label/instruction count - freedreno/qrisc: Add support for "absolute" label references - freedreno/qrisc: Test new label features - tu: Fix resetting command streams with writeable BOs - tu: Fix condition for skipping emitting aprons Daivik Bhatia (6): - broadcom/compiler: Add explicit NOP instruction at page boundaries - pan/nir: fix GNU compilation error with clang - broadcom/compiler: add support for null descriptors - v3dv: Implement and enable nullDescriptor support - nir/opt_copy_prop_vars: kill stale entries when source deref is written - rocket: simplify input/output tensor creation Daniel Lang (3): - radv/meta: fix samples datatype in radv_meta_nir - nir: change type_size return type to unsigned in nir_lower_{amul,io} - docs: Add GL_ARB_map_buffer_range and GL_ARB_vertex_array_object to etnaviv Daniel Schürmann (39): - nir: add nir_loop::do_while to indicate do-while loops - vtn: set nir_loop::do_while during spirv_to_nir() - glsl_to_nir: set nir_loop::do_while - nir/opt_loop: Don't peel initial break from do-while loops - nir/opt_loop: stop recursion at loop header phi in can_constant_fold() - nir/opt_loop: always try to peel initial break from loops with unrolling hint - nir/opt_algebraic: use imul24_relaxed for lowered dot4x8_add - nir/opt_algebraic: add some imul24_relaxed pattern - nir/opt_constant_folding: create const_value_for_alu() helper - nir/opt_constant_folding: constant-fold op(bcsel(), #c) -> bcsel(.., #c1, #c2) - nir/opt_algebraic: optimize downcast followed by upcast to extract - nir/opt_algebraic: extend some extract_u8 pattern to extract_i8 - nir/builder: constant-fold nir_mov_alu() if requested - nir/lower_bit_size: use nir_builder::constant_fold_alu - nir/lower_bit_size: use nir_def_replace() instead of nir_def_rewrite_uses() - nir/lower_bit_size: skip conversion for more opcodes - anti-lag: rework wait time calculation - aco/assembler: Fix s_inst_prefetch insertion after loop latch rotation - aco/assembler: pass std::vector to insert_code - nir: remove fixed-sized nir_op_bitz / nir_op_bitnz - nir: disallow converting to sized booleans from nir_type_convert() - nouveau: don't handle 8- and 16-bit comparisons - panfrost: remove fixed-sized binop reductions - panfrost: replace fixed-sized with unsized comparison opcodes - panfrost: replace fixed-sized bcsel with unsized bcsel_pan - nir,panfrost: remove 8-bit and 16-bit booleans - aco/ra: Fix get_reg_impl() for operand registers - aco: encode unused VOP3 operands as inline constant 0 on RDNA - aco/isel: move add64_32() to aco_isel_helpers.cpp - aco/isel: use add64_32() for nir_iadd(nir_u2u64(), ..) - nir/range_analysis: handle read_first_invocation and friends in nir_def_num_lsb_zero() - nir/range_analysis: handle phis in nir_def_num_lsb_zero() - amd/lower_global_access: refactor using state struct - amd/lower_global_access: only consider constant offsets that are aligned - amd/lower_global_access: Only consider 32-bit offsets which are aligned - amd/lower_global_access: lower SMEM offsets according to hw capabilities - aco: remove alignment handling for global SMEM loads - aco: emit global SMEM loads directly - aco: Remove SMEM offset optimization for non-buffer loads Daniel Stone (9): - pan/afbc: Code motion for split modifier queries - pan/mod: Protect against no usage flags for 64k - pan/mod: Reorder linear modifier checks - pan/afbc: Properly validate format/parameter combinations - ci/panfrost: Switch T860 jobs to another RK3399 device type - ci/panfrost: Add two T860 OpenCL fails - symbols-check: Ignore more pthread symbols - doc/ci: Add custom-kernel testing workflow - draw: Avoid warnings for maybe-unused variable Danylo Piliaiev (54): - tu: Fix draw call offset for LRZ warnings in secondaries - tu/perfetto: Move away from single timeline for all apps - tu: Fix CP_CCHE_INVALIDATE not being applied at the right point - freedreno: Fix CP_CCHE_INVALIDATE not being applied at the right point - tu/u_trace: Use correct u_trace destination in tu_clone_trace_range - tu/u_trace: Prevent cloning stale RB_DONE_TS results - tu/u_trace: Correct the order of tracepoints clonning for binning - tu/u_trace: Fix explicit toggle_name not being used - tu/perfetto: Add a performance warning track to perfetto - tu/perfetto: Add performance warning tracepoints - tu: Fix tu_bo_make_zombie without queues - tu: Fix double free of timestamp_copy_data->trace - tu: Don't leak pre_chain.rp_trace, and correct u_trace_move - tu: Fix BV/BR race in tu_clone_trace_range when waiting on barrier - tu/a8xx: Fix reading border_color from sampler memory - tu: Don't disable UBWC for D24S8+USAGE_SAMPLED+customBorderColorWithoutFormat - tu: Disable concurrent binning by default due to perf regressions - tu: Don't enable FDM when there is FDM attachment is UNUSED - tu: Always lazy_init_vsc for tiler rendering - u_trace: Lazy init ut->linear_alloc - tu: Fix TU_CMD_DIRTY_DRAW_STATE value collision - tu: Start/End occlusion query should force depth state recalculation - tu: Change of disable_fs state should force depth state recalculation - tu/a7xx: Don't force enable IJ_LINEAR_PIXEL for FragFace/FragCoord - freedreno/a7xx: Don't force enable IJ_LINEAR_PIXEL for FragFace/FragCoord - tu: Disable FS in some cases even when FS explicitly writes D/S - ir3: Add resbase_ir3 intrinsic - tu: Add allow_oob_indirect_ubo_loads to device cache uuid - tu: Specify max texel buffer and storage buffer limits via GPU props - tu/a8xx: Set real storage/texel buffer size limits - tu: Add option to raise the maximum texel buffer size - tu: Enable texel buffer / SSBO emulation for known problematic games - tu: Match SW depth clear value packing with HW - tu: Match SW color clear value packing with HW - tu: Don't process A2R10G10B10 clear values via new pack function - tu/lrz: Pick correct depth attachments in msrtss case - tu: Force GMEM mode when renderpass has MSRTSS attachments - tu: Refactor separate D32S8 and ignore D/S aspect mask for RP attachments - tu: Remove depth/stencil-specific blit src/dst helpers - tu: Remove event_blit_dst_view - tu: Enable tu_dont_care_as_load for all Kex Engine games - tu: Use application_name_match instead of exe match for workarounds - tu/a6xx: Work around D32S8 EARLY_Z_LATE_Z hang - tu: Enable tu_allow_oob_indirect_ubo_loads for Clausewitz engine - tu: Custom resolve should always use AVOID_CCU layout - tu: Fix subsampled metadata and blit emission for separate stencil - tu: Fix gfx_write_access checking for TRANSFORM_FEEDBACK_COUNTER_READ_BIT - tu: Fix blit_cache_cleaned never being set to true - tu: Fix LRZ handling for VK_EXT_custom_resolve - tu: Dirty LRZ after changing attachment locations disable LRZ writes - nir: Include scalarized component offsets in UBO ranges - tu: Fix tu_event not being reset on creation - tu: Merge disable_write_for_rp from secondary to primary - tu: Fix memory leak of FDM patch-points Dave Airlie (15): - nouveau: drop sector promotion. - gallivm: handle llvm 22 coroutine end change - gallivm: handle llvm 22 scatter/gather intrinsic changes. - lavapipe: treat NULL pColorAttachmentLocations as no handles - nak: fix image size for multisample arrays - nak: add more sizes to assert in bindless_image_sparse_load - nvk: enable subgroupQuadOperationsInAllStages - ci: vmware farm is offline, stop using it - st: fix get tex subimage fallback for 1D ARRAY - st: drop ununsed arguments to copy_to_staging_dest. - u_blitter: only set texcoord.w to sample for multisample sources - nak: block pipe_format from nak bindings. - ir3: use the correct builder for adding preamble to main. - nir: add impl pointer to block to avoid recursive linked list - nir: remove a lot of nir_cf_node_get_function calls. David Airlie (3): - nir/coopmat: refactor the split vars to clean it up - nir/coopmat: move the row/col into a box and add some helpers. - nir/coopmat: rename the box split variables. David Rosca (70): - d3d12: Use HEVC RefPicSet order from frontend - ac/parse_ib: Fix printing enc recon VAs on VCN5 - radv: Fix uint32 overflow in slice offset calculation - radv/video: Fix initializing rc structs with default rate control - radeonsi: Always use 2D tiling for video dpb - frontends/va: Fix finding LTRs from POCs in HEVC decode - frontends/va: Fix out of bounds write in AV1 decode tile info - frontends/va: Fix setting output color properties from color standard - frontends/va: Fix dereference before NULL check in postproc - frontends/va: Add missing NULL check for additional output surface - vl: Use NV12 as deint format instead of preferred format - vl: Don't check npot textures support when creating buffers - pipe/video: Remove unused PIPE_VIDEO_CAP_PREFERRED_FORMAT - pipe/video: Remove unused PIPE_VIDEO_CAP_NPOT_TEXTURES - pipe/video: Remove unused PIPE_VIDEO_CAP_MAX_LEVEL - pipe/video: Remove unused PIPE_VIDEO_CAP_STACKED_FRAMES - radeonsi/uvd_enc: Skip extra padding bytes in output bitstream - radeonsi: Move si_vpe.* to mm subfolder - ac/info: Add video codec caps - radeonsi/video: Use new video codec caps - radv/video: Use new video codec caps - ac/info: Remove old video codec caps - ac/info: Print number of VPE instances - radeonsi: Add RADEON_FLUSH_FORCE and use it to force flush - ac/vcn_dec: Add ac_vcn_dec_init_regs to get register offsets - ac/vcn: Add ac_vcn_sq_header/tail and use it for decode - ac/cmdbuf: Add ac_emit_video_write_memory - radv: Use ac_emit_video_write_memory - ac/vcn_dec: Move register defines to ac_vcn_dec.c - ac/cmdbuf: Add ac_emit_video_write_timestamp - radv: Add support for timestamps on video queue - radeonsi/mm: Add support for 2-ref H264 encode - radeonsi/mm: Remove comment about kernel AV1 instance scheduling bug - ac/parse_ib: Add VCN decode queue parsing - ac/parse_ib: Add VCN timestamp command - radeonsi/mm: Add si_vid_create_buffer and use it - radeonsi/mm: Set PIPE_RESOURCE_FLAG_UNMAPPABLE for buffers - va: Set contiguous_planes for DMA-BUF imported surfaces - ac/vcn_dec: Add 10 to 8 bit dithering support - radeonsi/mm: Select DPB format independently from decode surface format - vl: Skip transfer function and primaries conversion when not needed - va: Always reset compositor chroma location - va: Use RGB format with matching bit depth for YUV->YUV matrices - radeonsi/mm: Return error when decoding H264 P/B frame with no refs - radeonsi/mm: Only setup ref surfaces with tier3 - radeonsi/mm: Set correct usage in si_dec_fill_surface - radeonsi/mm: Fix setting VPE rotation when horizontal flip is enabled - va: Implement vaPutImage for derived images - pipe/video: Add out_pipe_fence to pipe_picture_desc - vl: Support blending with gfx compositor - vl: Add pipe_video_codec proc using vl_compositor - vl: Add vl_proc pipe_video_codec using vl_compositor as fallback - va: Use vl_proc for processing context - va: Add vlVaDestroySurface - va: Add vlVaPostProc and use it instead of compositor and vid engine blit - va: Use vlVaPostProc in vlVaPutSurface and for subpictures - va: Stop using vl_compositor - pipe: Remove pipe_video_codec::expect_chunked_decode - r600/uvd: Set correct h264 chroma format - pipe: Remove pipe_video_codec::chroma_format - vulkan/video: Fix coding AV1 decoder/encoder_buffer_delay - vulkan/video: Don't code AV1 decoder model info when not present - vulkan/video: Fix coding AV1 operating points - d3d12/video: Don't reset batches in fence_wait - vulkan/video: Fix coding H265 ref pic list modification lists - vulkan/video: Fix coding H265 SPS pcm block sizes, inter ref pic set and lt refs - va: Fix leak when vlVaUploadImage fails - va: Ensure templat is valid for temporary surfaces - radeonsi: Stop forcing GTT with no gfx/compute - ac/video: Fix number of AV1 single refs Derek Lesho (2): - zink: Guard bo map/unmap on map_count. - zink: Fix zink_bo_unmap synchronization for client pointer support. Dhruv Mark Collins (7): - tu/autotune: Fail gracefully when CP counters are unavailable - fd/pps: Allocate performance counters from high-to-low - tu/autotune: Allocate performance counters from low-to-high - tu/query_pool: Avoid CP counter conflict with autotune - freedreno: Update A6XX_PC_MODE_CNTL definition and values - tu/util: Fix tile division algorithm - tu: Propagate allocation failures for tu_cs_* functions Dmitry Baryshkov (3): - rusticl: enable freedreno by default - tu: limit KHR_internally_synchronized_queues to Vulkan 1.1+ - tu: limit VALVE_fragment_density_map_layered to Vulkan 1.1 devices Dmitry Osipenko (3): - intel/virtio: Preserve errno properly when handling ioctl - drm-uapi: Update virtio-gpu with new hinting field - intel/virtio: Support DRM_VIRTGPU_BLOB_FLAG_HINT_DEFER_MAPPING Dorinda Bassey (1): - util/rust: Add atomic memory synchronization support Duncan Brawley (7): - pco: Fix pco_last_igrp returning the first element instead of the last - pco: Refactor internal shader pass skipping - pco: Add propagating coherent/volatile access qualifiers - pco: Add DMA ld/st caching support for ssbo/ubo operations - pco: Add DMA sampling caching support - pco: Fix smp instruction encoding map order - pco: Add DMA ld/st caching support for all ld/st instructions Dylan Baker (3): - intel/brw: Add assert for error case - meson: ensure that libdrm auto-features match requirements - intel/gen: decode type of src1 in basic 2 source after setting IMM Emma Anholt (87): - spirv: Demote the SPIRV 1.6 OpTypeSampledImage on Buffer failure to a warning. - ci: Bump apitrace version to 14.0. - ci: Don't set wine vars in deqp-runner.sh/vkd3d-runner.sh. - ci: Build a working wine installation in build-wine.sh. - ci/test-vk: Install win64 apitrace 14.0 along with setting up wine. - ci/test-vk: Install DXVK 2.7.1 to our wine installation. - ci/lava: Fix the name of the fluster overlay. - ci/lava: Add a note about an otherwise-mysterious error you can encounter. - ci/gfxreconstruct: Disable OpenXR support. - ci: Build gpu-trace-perf and include a script to use it. - ci: Bump the image tags for the previous build script changes. - bin/update-traces_checksum.py: Pull out per-job work to a helper function. - ci/update_traces_checksum: Parse gpu-trace-perf's format for hash changes. - ci/update_traces_checksum: Default to updating for the current HEAD. - ci/update_traces_checksum: Make it work on restricted traces jobs, too. - ci/llvmpipe: Use anholt's new GPU trace snapshot comparison tool. - ci/lavapipe: Use anholt's new GPU trace snapshot comparison tool. - ci/turnip: Drop two 660 vk jobs and tune down the vk coverage fraction. - ci/turnip: add an a660 VK restricted traces job. - ci: Delete references to various broken traces. - ci/amd: Switch radv-raven-traces-restricted over to gpu-trace-replay.sh - ci/intel: Switch over to the new tool for restricted traces. - ci/piglit-traces: Remove ANGLE trace support. - ci/llvmpipe: Disable some traces too close to the timeout. - tu: Set HALF_PRECISION on blits to R11G11B10. - ir3: Fix shared IMAD24 lowering. - tu: Add capture/replay for sparse buffers and descriptor buffer. - screenshot-layer: Fix leftover VK queues in the map at DeviceDestroy. - screenshot-layer: Fix a bunch of unused variable warnings. - screenshot-layer: Fix rename() to final png before the file is flushed. - screenshot-layer: Clean up the lifetime management of the copyDone fence. - screenshot-layer: Fix race on writing the .pngs vs device destroy. - screenshot-layer: Wait on the fence before fallible operations. - tu/ci: Drop some old xfails that don't trigger any more. - tu: Report missing layout support for host_image_copy with unifiedLayouts. - zink/ci/tu: Fix up skips/xfails for GLCTS testcases that got divided up. - tu: Disable storage image support for depth/stencil. - ci/panfrost: Drop a set of flakes whose fix had landed. - panfrost/ci: Skip dEQP-VK.wsi.wayland.swapchain.render.10swapchains on g52. - lvp/ci: Drop an old skip long since fixed in the CTS. - tu/ci: Drop a750 VKCTS to 50% coverage. - ci: Update VK CTS to 1.4.5.3 with fixes. - ir3: Add an env var to prefer single wavesize. - ir3: Fix shader bisect crashing out when too many shaders get bisected. - ir3: Give some feedback as we shader bisect. - ir3/shader_bisect: Allow a 'r' response to retry a run mid-bisect. - ir3: Drop the "SIMD0" debug print that was apparently added for frameretrace. - ir3: Deduplicate shader disassembly generation. - ir3: If we're dumping IR3_SHADER_BISECT=[hash] disasm, include the NIR. - tu: Disable 128-wide subgroups on No Man's Sky. - screenshot-layer: Log when we can't open the output directory. - weston: Run at a more reasonable 1920x1080 resolution, not 1024x640. - ci: Include Windows renderdoc in with wine and update gpu-trace-perf. - ci: Build the Vulkan screenshot layer as part of VK test builds. - freedreno/ci: Add restricted traces testing of D3D11 traces on a660. - freedreno/ci: Add more explanation of a trace failure that's not our fault. - tu: Always set the kernel's name for BOs. - util/drirc_gen: Add a little documentation of what this does. - radv/drirc_gen: Clean up the dependency handling. - util/drirc_gen: Reduce manual importing of functions. - util/drirc_gen: Move the common VK WSI options to a core helper function. - util/drirc_gen: Make the header usable from C++. - tu: Move to using drirc_gen. - drm-shim: Include the hex of the driver ioctl for unimplemented ioctls. - drm-shim/freedreno: Provide a dummy set of UBWC config params. - drm-shim/freedreno: report a 48-bit address space. - drm-shim/freedreno: Report VM_BIND support. - vulkan: Enable GOOGLE_display_timing on KHR_display across multiple drivers. - drm-shim/freedreno: Fix VM_BIND support. - docs: Link in particular to the difficulty: * issue tags in Help Wanted. - zink: Use the new common code for nearest consistency in blits. - zink: Also enable the nearest consistency workaround on turnip. - zink: Also enable the nearest consistency workaround on anv. - .mailmap: Switch to anholt's current work address. - intel/device_info_override_test: Make sure we actually find our device. - drm-shim: Share common code for PCI and platform device setup. - drm-shim: Generalize overriding of links. - drm-shim: Give the device/subsystem links real link values. - drm-shim: Lock access to shim_device.fd_map. - drm-shim: Remove drm_shim_driver_prefers_first_render_node. - drm-shim: Remove unnecessary runtime setup of drm_device_path_prefix. - drm-shim: Remove unnecessary runtime setup of various device strings. - drm-shim: Fix racy initialization. - freedreno/ci: Clear the xfail for texture-immutable-levels. - etnaviv/ci: Fix flakes lists that are breaking gc2000 CI. - intel/ci: Fix xfails for nightlies. - freedreno: Don't force image component A=1 substitution on R/RG textures. Emre Cecanpunar (1): - jay: allocate shader under memctx Eric Engestrom (131): - VERSION: bump to 26.2 - docs: reset new_features.txt - docs: update calendar for 26.1.0-rc1 - docs: update calendar for 26.0.5 - docs: add release notes for 26.0.5 - docs: add sha sum for 26.0.5 - docs: add stub of vk_struct_type_cast.h for vk_util.h - ci/bare-metal: drop duplicate timestamps now that gitlab-runner has per-line timestamps - docs: update calendar for 26.1.0-rc2 - docs: update calendar for 26.1.0-rc3 - docs: update calendar for 26.0.6 - docs: add release notes for 26.0.6 - docs: add sha sum for 26.0.6 - docs: update calendar for 26.1.0 - docs: add release notes for 26.1.0 - docs: add sha sum for 26.1.0 - docs: add calendar for the 26.1 cycle, and 26.2 branchpoint and release candidates - docs: fix unescaped \`*` - docs/submittingpatches: fix section nesting - docs/ci: explain what Marge saying "Manual Step encountered" means - zink+nvk/ci: update expected fails - docs: update calendar for 26.0.7 - docs: add release notes for 26.0.7 - docs: add sha sum for 26.0.7 - docs/ci: ignore docs.redhat.com & registry.khronos.org links - etnaviv: initialize value before calling etna_gpu_get_param(), in case it fails - meson/libmesa: ensure shader_replacement.h is generated before using it - meson/amd: only build libaco when requested - meson/asahi: only build libagx2_disasm when requested - meson/freedreno: only build libfreedreno_common when requested - meson/intel: only build libblorp_elk when requested - ci/build: restore riscv64 build as it works again - Revert "ci/build: restore riscv64 build as it works again" - docs: update calendar for 26.1.1 - docs: add release notes for 26.1.1 - docs: add sha sum for 26.1.1 - docs: update calendar for 26.0.8 - docs: add release notes for 26.0.8 - docs: add sha sum for 26.0.8 - util/meson: simplify list of per-driver drirc files - drirc: move 00-$drv-defaults.conf to each driver's folder - Revert "drirc: move 00-$drv-defaults.conf to each driver's folder" - docs: update calendar for 26.1.2 - docs: add release notes for 26.1.2 - docs: add sha sum for 26.1.2 - rusticl: skip bindgen for pipe_shader_state_from_tgsi - meson: exclude known buggy versions of bindgen - ci: bump rust version from 1.90 to 1.96 - ci: bump bindgen version from 0.71.1 to 0.72.1 - ci: bump fedora from 42 to 44 - meson: drop non-existent platforms=xcb check - Revert "egl: fix _EGL_NATIVE_PLATFORM fallback for unrecognized native displays" - docs: update calendar for 26.1.3 - docs: add release notes for 26.1.3 - docs: add sha sum for 26.1.3 - gen_release_notes_test: don't evaluate backslash - gen_release_notes: add support for "work_items" links - docs: fix release notes for 26.1.0 - docs: fix release notes for 26.1.1 - docs: fix release notes for 26.1.2 - docs: fix release notes for 26.1.3 - ci: fix perfetto download in \`make-git-archive` nightly job - ci: fix perfetto download in build-perfetto.sh - ci: fix the fix for perfetto download in \`make-git-archive` nightly job - etnaviv/ci: document two fixed tests - nvk/ci: document fixed tests, new failures, and recent flakes - zink+nvk/ci: fix duplicate fails - docs: drop x.org -> x.org/wiki/ redirect and expected url - docs: s/issues/work_items/ - docs/ci: mark yet another domain as blocking linkcheck - docs/ci: disable auto-retry on nightly linkcheck - docs/ci: use full/explicit option names in linkcheck job - docs/ci: only print the linkcheck issues, not the thousands of non-issues - mr-label-maker: add ~drirc label on all drirc files - util: add support for multiple colon-separated DRIRC_CONFIGDIR entries - drirc: move 00-$drv-defaults.conf to each driver's folder - meson: merge two consecutive \`if with_egl` - meson: add native platform to the summary - meson: ensure native platform is one of the undetectable ones - meson: drop misleading \`-D egl-native-platform` values - zink/ci: drop leftover anv-cml deqp suite - docs: update calendar for 26.1.4 - docs: add release notes for 26.1.4 - docs: add sha sum for 26.1.4 - docs: fix x.org url - etnaviv/ci: update nightly job expectations - zink+nvk/ci: update nightly job expectations - lvp/ci: update nightly job expectations - llvmpipe/ci: update nightly job expectations - nvk/ci: update nightly job expectations - drm-shim: name the driver name \`driver_name` consistently - drm-shim: set \`driver_name` in \`drm_shim_*_device_setup()` - docs: gitignore the contents of the \`_generated` folder - ci: disable auto-retry on rustfmt job - docs/helpwanted: url-encode \`[]` to avoid a pointless redirection - rusticl: document api\@clgetmemobjectinfo as fixed for all drivers - lavapipe/ci: document fixed dEQP-VK.mesh_shader.ext.misc.emit_in_control_flow_bad_emit_last - freedreno/ci: document fixed KHR-GL46.copy_image.smoke_test - freedreno/ci: document a recent flake - zink+nvk/ci: document a couple of recent flakes - ci/piglit: fix nightly expectations after piglit uprev - img/ci: add \`farm:imagination` tag to all jobs - loader: move variable to correct scope - broadcom/ci: mark fixed tests as such - ci/video: move two single-thread tests to global list - ci/video: install the current version of gstreamer - ci/video: uprev fluster - ci/video: download fluster test suites by codec name - ci/video: update comment with the new blocker for AV1 support - radv/ci: enable VP9 testing in fluster - anv/ci: enable VP9 testing in fluster - nvk/ci: document fixed dEQP-VK test - nvk/ci: document fixed vkd3d tests - nvk/ci: document two vkd3d regressions - freedreno/ci: document fixed tests - radeonsi: fix truncated cache key - mailmap: update my email address - zink+nvk/ci: document two fixed tests - VERSION: bump for 26.2.0-rc1 - .pick_status.json: Update to d49a00bdf15fd48b31af93aaf5feed3eebcbda12 - VERSION: bump for 26.2.0-rc2 - .pick_status.json: Update to 8b00adbe72f2705985146b057f6fde9256d0dcb0 - .pick_status.json: Mark 47efd739121d51e2f9049cec715a70c13767a67c as denominated - .pick_status.json: Mark 7999060992e9cee91d1962faf65dc4e5c6fe4f69 as denominated - .pick_status.json: Mark 2515024a5919ed14fe05471e3f1f89c54a454610 as denominated - .pick_status.json: Mark 4fd93a0039a07ec2027f2a6d1d252c73ed033e0a as denominated - pick-ui: turn commit.date into a (cached) property - pick-ui: show MR number for additional context - VERSION: bump for 26.2.0-rc3 - .pick_status.json: Update to 85c082ddbed727940535911e6bf87f7d274525bf - [26.2 only] docs/new_features: mention that VK_EXT_host_image_copy was exposed on RADV/GFX10.3+ Eric Guo (3): - compiler: Add missing MESA_SHADER_KERNEL case for SPIR-V dump - pan/compiler: Clamp fp16 ldexp exponent range - pan/bi: Lower 64-bit hadd on v9/v10 Eric R. Smith (4): - glsl, spirv: Improve accuracy of asin() and acos() - panfrost: add some sanity checks - panfrost: make sure INDEX_OFFSET is cleared - panfrost: add helper function for checking for active queries Erico Nunes (4): - ci: lima farm maintenance - Revert "ci: lima farm maintenance" - CODEOWNERS: add lima maintainers - ci: lima farm maintenance Erik Faye-Lund (88): - panvk: drop out-of-date TODO - panfrost: use perf-trilinear when doing anisotropic sampling - panvk: use perf-trilinear when doing anisotropic sampling - pan/lib: fix up afbc and linear layout - pan/lib: emit high bits of buffer-size - pan/lib: validate data_size_B in drivers - panvk: do not artificially limit image dimensions - panvk: increase maxResourceSize on v11 and later - panvk: increase maxBufferSize on v11 and later - nouveau: do not report unsupported feature - radeonsi: remove old, unsupported cap - d3d12: remove benign but unsupported cap - iris,crocus: remove benign but unsupported cap - llvmpipe: drop support for tgsi_tex_txf_lz cap - ntt: stop emitting TXF_LZ - gallium/u_blitter: stop emitting TEX_LZ - gallium: remove defunct pipe-cap - ttn: do not handle T{EX,XF}_LZ - gallium: completely remove T{EX,XF}_LZ opcode - panvk: do not enable extension without required feature - panvk: do not enable extension without required feature - haiku: remove unfinished post-processing support - gallium: delete leftovers of post-processing infrastructure - pan/ci: add a flake from nightly - util/format: make Y8_UNORM an alias of Y8_400_UNORM - util/format: make subsampling explicit - util/format: mark subsampled RGB formats as actually subsampled - util/format: verify subsampling in name - pan/va: do not allow force_delta_enable on v9 - pan/bi: correct computation of lod.x - panfrost: enable ARB_texture_query_lod on v9+ - mesa/main: remove stale prototypes - mesa/main: remove incorrect debug-output - mesa/main: do not gate performance warning - mesa/main: remove low-value debug-output - mesa/main: remove unused verbose-flags - mesa/main: remove VERBOSE_API - mesa/main: remove mesa_print_display_list function - mesa/main: remove low-value verbose-switch - Revert "mesa: check for ARB_ES3_compatibility in format checks" - mesa/main: remove unused array - pan/ci: update flakes based on nightly ci - pan/ci: remove benign typoed flake - meson: update libdrm wrap - pan/ci: add missing gitlab rules - pan/ci: remove outdated gitlab rule - pan/ci: add missing gitlab rule - pan/ci: fix gitlab rules after move - pan/genxml: correct size of field - pan/genxml: add missing modifier - pan/genxml: correct size of field - pan/genxml: correct size of field - pan/genxml: correct size of field - pan/genxml: add missing enum value - pan/genxml: sort CS structs by enum-value - pan/genxml: use consistent name for scissor - pan/genxml: use an enum for progress increment - pan/genxml: consistently use bool for error reject - pan/genxml: consistently use hex for masks - pan/genxml: consistently use uint for signal slot - pan/genxml: consistently use uint for chunk indexes - pan/genxml: remove needless defaults - pan/genxml: keep enum ordering from v10 - pan/genxml: correct casing of names/types - pan/genxml: consistently use hex for uint immediates - pan/genxml: consistently set default - pan/genxml: clean up whitespace - pan/genxml: make field consistent - pan/genxml: remove some pointless comments - pan/genxml: use consistent attribute order - pan/ci: add a couple of flakes - pan/ci: use slow-skips to only skip slow tests for merge-requests - pan/ci: move cts-bug-fails to skips - pan/ci: stuff some breadcrumbs in the fails-list - pan/ci: reenable passing tests - pan/ci: add a few new g925 flakes - pan/ci: add back missing skip-list heading - pan/ci: drop needless skips - pan/ci: move skip to flakes - pan/ci: skip slow test - pan/ci: move common flake to common flake-file - pan/ci: recognize flaking test - pan/ci: mark missing xfails - pan/ci: move longprim flake into common flake-file - pan/ci: add new flake - pan/ci: just mark all random-max draw-tests as flakes - ci/vulkan: remove long outdated skips - panvk: simplify non_polygon calculation Etaash Mathamsetty (4): - vulkan/wsi/wayland: Fix error handling for tearing control. - vulkan/wsi/wayland: Move drm syncobj to swapchain. - vulkan/wsi/wayland: Move color management surface to swapchain. - vulkan/wsi/wayland: Do a roundtrip after retiring the old swapchain. Faith Ekstrand (399): - panvk/csf: Emit INDEX_BUFFER[_SIZE] even for non-indexed draws - pan/bi: Improve swizzle propagation - zink: Assert if we try to use a dedicated allocation with offset > 0 - panfrost: Add and use a new pan_nir_res_handle() helper - pan,nir: Add cube face intrinsics - nir/builder: Allow backend1/2 in nir_build_tex() - nir: Add a new nir_texop_gradient_pan - panvk: Implement bitfield_select - pan/nir: Add a pass for lowering texture ops in NIR on Valhall+ - pan/nir: Use the NIR lowering on Valhall+ - nir: Add a new nir_op_f2u32_rtne - pan/bi: Implement nir_op_f2[iu]32_rtne - pan,nir: Add Bifrost texturing intrinsics - pan/nir: Add bifrost support to pan_nir_lower_tex() - pan/nir: Lower texturing ops in NIR on Bifrost - pan/nir: Load texel buffer conversion descriptors in NIR - pan/bi: Allow setting the table on lea_attr_pan - pan/nir: Use HW NIR intrinsics for texel buffer addresses - pan/bi: Delete the old texel buffer intrinsics - pan/nir: Lower texel buffers in nir_lower_tex() - pan/nir: Lower texture queries in nir_lower_tex() on Valhall+ - panfrost: Also remap image handles for image_size/samples - pan/nir: Lower image queries in NIR on Valhall+ - panvk: Let the compiler handle texture queries on v9+ - pan/nir/tex: Support full index+offset - panvk: Add MAX_VS_ATTRIBS to image indices in panvk_nir_lower_descriptors - panfrost: Take texture/sampler_index into account in lower_res_indices - panfrost: Prefix valhall bits of lower_res_indices - panfrost: Handle pre-Valhall images and texel buffers in lower_res_indices - pan/bi: Drop lower_index_to_offset from preprocess - util/half: Use explicit RTNE rounding for the C++ float16_t - util/half: Stop whacking CPU flags to test float_to_half_slow() - util/half: Rename the tests - util/half: Re-organize the tests a bit - util/half: Add float_to_half rounding tests - util/half: Add double_to_half tests - util/half: Add a simpler double_to_float16() - util/half: Add double_to_float16_ru/rd helpers - nak: Move Srcs/DstsAsSlice implementations - nak: Implement Srcs/DstsAsType directly for Op - nak: Implement Srcs/DstsAsSlice directly on ops - nak,compiler: Move AttrList into NAK - nak: Don't use the proc macro to implement auto-boxing of ops - nak,compiler: Move FromVariants to common code - pan/bi: Use LOD_MODE_EXPLICIT for the 2nd half of textureGrad() on Bifrost - docs: Move and rename "Development Notes" - docs: Add docs with Vulkan/SPIR-V extensions basics - docs: Add docs for drafting new MESA extensions - panvk/csf: fix VERTEX_SPD dirty tracking when topology changes - panvk/csf: Inline the SPD addr helpers - nouveau/push: Rename push_method to push_mthd - nouveau: Don't build NAK tests on Android - compiler/rust: Add a float16 wrapper - etnaviv: Remove f32_to_f16_fallback() in favor of float16::F16 - meson: Bump the minimum rust version to 1.85.0 - compiler/rust: Add LowerBoundedU32[Array] types - nak: Use LowerBoundedU32 for SSAValue - nak: Allow SSA value 0 again - nak: Simplify SSARef construction with try_push() - meson: Suffix compiler/rust bindings with _compiler_rs_extern - compiler/rust/bindings: Add util_dyarray - compiler/rust: Add a nir_shader::get_entrypoint() helper - compiler/rust: Add a nir_shader::to_string() - compiler/rust/nir: Add structured block iterators - compiler/rust/nir: Add helpers for getting ALU input/output types - compiler/rust/bitset: Add a BitIndex helper struct - compiler/rust/bitset: Don't reserve space in remove() - compiler/rust/bitset: Add find_next_[un]set() helpers - compiler/rust/bitset: Generalize BitSetIterator - compiler/rust/bitset: Implement Into/FromBitIndex for more types - compiler/rust/bitset: Add a new ConstBitSet type - compiler/rust: Add an EnumAsU8 trait - nak: Use EnumAsU8 for RegFile - panfrost: Initial rust build system support - panfrost: Add the basis for the new Kraid compiler - kraid: Add a GPU model abstraction - kraid: Add DataType and NumericType enums - kraid: Add a swizzle struct - kraid: Add SSAValue and SSARef structs - kraid: Add Src/Dst data types - kraid: Add an Opcode trait and Op enum - kraid: Add Instr, BasicBlock, and Shader structs - kraid: Add a builder - kraid: Start parsing NIR shaders - kraid: Parse the NIR CFG - kraid: Handle load_const instructions - kraid: Handle nir_op_mov/vec/[un]pack - kraid: Add some float alu ops - kraid: Implement nir_op_iadd - kraid: Handle a few NIR intrinsics - Kraid: re-indent shaders for prettier printing - kraid: Add a validator to check IR invariants - kraid: Add a super simple register allocator - kraid: Plumb through Model::encode_shader() - kraid: Rework swizzles - kraid: Print ASM swizzles when we have them - kraid: Copy the bitview module from nouveau - kraid: Add a FlowCtrl struct - kraid: Replace OpEnd with OpNop.end - kraid: Move proc/lib.rs to proc/macros.rs - subprojects: Pull in the Rust xml crate - kraid: Add ISA XML for v9-15 - kraid: Add the start of encoder code-gen - kraid/isa: Add a simple XML parser - kraid/isa: Generate enums with [Try]Encode/Decode - kraid/isa: Add an encoder for expressiosn - kraid/isa: Add an encoder for instructions - kraid/isa: Add support for field modifiers - kraid: Add the start of a v9 encoder - kraid/isa: Specially handle small_constant_t - kraid: Add a SmallConstant struct and a Model::small_constants() hook - kraid: Add a lower_small_constants() pass - kraid: Add a very dumb message slot assignment pass - kraid: Implement shifts and logic ops - kraid: Implement integer comparisons - krai/isa: Expose a new InstructionInfo struct per-instruction - kraid: Break v9 instruction encoding out into traits - kraid: Use instruction info to implement op_is_message() - kraid/isa: Add a special case in to_snake/camel_case() for data types - kraid/isa: Emit TryFrom for all data-type-like enums - kraid: Clean up the data type mess in the encoder - kraid: Implement OpCSel and nir_op_[ui]min/max - kraid: Claim we use 64 registers - kraid: Support signless IAdd - kraid: Implement nir_op_u2u/i2i - kraid: Add a SrcRef::Zero - kraid: Add a 16-bit ALU lowering pass - kraid: Implement nir_op_extract_* - kraid: Be more lax about immediates - kraid: Map H01 and B0123 to None in the encoder - kraid: Implement nir_op_f2f* - kraid: Make Instruction::get_info() more ergonamic - kraid: Add a Model::op_src_supports_imm32() query - kraid/isa: Handle field restrictions - kraid: Box ops inside Op - compiler/rust/smallvec: Implement Clone, Default, and new() - compiler/rust/smallvec: Add a push_mut() method - compiler/rust/smallvec: Implement Deref[Mut] - compiler/rust/smallvec: Implement Extend for SmallVec - compiler/rust/smallvec: Implement From> - compiler/rust/smallvec: Implement FromIterator and From<[T; N]> - compiler/rust/smallvec: Implement IntoIterator - compiler/rust/smallvec: Implement From> for Vec - nak: Simplify BasicBlock::map_instrs() - nak/builder: Use some of the SmallVec improvements - nak: Simplify our SmallVec usage - compiler/rust/smallvec: Hide the enum - compiler/rust/smallvec: Optimize extend() - nir: Allow atomic intrinsics to have multiple components - spirv,nir: Add support for AtomicFloat16VectorNV - nak/nir: Lower f16vec4 atomics to 2xf16v2 - nak: Rename AtomType::F16x2 to F16v2 - nak/from_nir: Handle f16v2 atomics - nvk: Advertise VK_NV_shader_atomic_float16_vector - kraid: Make SrcRef::Imm32 explicitly non-zero - kraid: Make SrcRef PartialEq - kraid: Add map_instrs() methods to Shader and BasicBlock - kraid/builder: Store the model in builders - kraid: Split DataType into two enums - kraid/v9: Fix encoding of high register numbers - kraid/v9: Rework the shift_lop encode macro - kraid/v9: Add the rest of the shift/lop ops - kraid: Add None logic and shift ops - kraid/v9: Allow immediates in logic ops - compiler/rust/bitset: Implement Eq and PartialEq for ConstBitSet - compiler/rust/enum_as_u8: Add an EnumAsU8::MAX_DISCRIMINANT - compiler/rust/enum_as_u8: Add an ConstU8EnumSet struct - compiler/rust/as_slice: Document AsSlice - compiler/rust/as_slice: Add a new AsArray trait - kraid: Add a VirtualOpcode trait - kraid: Add a Model::op_is_supported() query - kraid: Add a lanes to Dst - kraid/isa: Make Enum::meta a weak reference - kraid/isa: Rework enum literals - kraid/isa: Make Swizzle EnumAsU8 - kraid/isa: Treat exact= as a field restriction - kraid/isa: Expose allowed swizzles through InstructionInfo - kraid/isa: Expose allowed lanes through InstructionInfo - kraid: Add a Model::op_src_supports_swizzle() helper - kraid: Add a Model::op_dst_supports_lanes() helper - kraid: Add the hardware MkVec ops - kraid/v9: Fix OpShiftLop::src_supports_imm32() - kraid/v9: Fold swizzles and modifiers on imm1w sources - kraid: Add a virtual OpCopy and the relevant lowering pass - kraid/nir: Emit OpCopy instead of OpMov - kraid: Allow 8-bit SSA values - kraid: RA per-byte - kraid/ops: Claim even more variants - kraid/nir: Emit 8-bit ops - kraid: Widen ALU ops before RA - kraid: Expose the guts of Swizzle - kraid/builder: Add copy_iN_to() helpers - kraid: Add an OpSwz and a lower_mkvec_swz() pas - kraid/nir: Use OpSwz for nir_op_u2uN and nir_op_i2iN - kraid/nir: Fix 2x16 extract_[iu]8 - kraid/nir: Use OpSwz op_extract_* - kraid/nir: Implement nir_op_unpack_32_* - kraid: Add a new legalize_src_swizzles() pass - kraid: Add word() helpers to Src/Dst types - kraid/nir: Implement nir_op_unpack_64_* - kraid: Better document swizzles - kraid: Only dump shaders if KRAID_DEBUG=print is set - kraid: Fix RA for dead destinations - kraid: Add a Model::op_src_is_staging_reg() helper - kraid: Add a Model::op_dst_is_staging_reg() helper - kraid: Allocate whole registers for staging destinations - kraid: Re-materialize constants - panfrost: Set the rustfmt edition to 2024 - kraid/swizzle: Add a Swizzle::is_none() helper - kraid/swizzle: Add an is_none() special case in fold_u32() - kraid/swizzle: Take a src_bytes param in Swizzle::bytes_read() - kraid/validate: Fix 64-bit destination validation - kraid/hw_tests: Allow the test to specify swizzles and lanes - kraid/swizzle: Return Option from AsmSwizzleWiden::to_swizzle() - kraid: OpShiftLop is unsigned - kraid: Add an SSAValue::bytes() helper - kraid: Use a tuple struct for SSAValue - kraid: Add OpRegIn and OpRegOut - panvk/jm: De-duplicate most of cmd_draw[_indirect] - panvk/jm: Re-group setting desc tables and SSBOs - panvk/jm: Take a desc_info in meta_get_copy_desc_job - panvk/jm: Take a desc_info in prepare_desc/dyn_ssbo() - panvk/csf: Take a desc_info in fill_dyn_bufs() and prepare_res_table() - panvk: Move desc_info to panvk_shader - panvk: Call panvk_lower_nir() before lowering multiview - panvk: Add a central panvk_cmd_draw() helper - panvk/csf: Make various panvk_draw_info pointers const - panvk: Plumb index buffers through panvk_draw_info - panvk/jm: Plumb IA state through draw_info - panvk/csf: Plumb IA state through draw_info - panvk: Improve base instance tracking for indirect draws - panvk/csf: Add some sanity assertions in prepare_push_uniforms - panvk/csf: Break FS descriptor setup into a new helper - panvk/csf: Break VS descriptor setup into a new helper - panvk/csf: Prepare descriptors first - panvk: Patch VS attribute descriptors as a separate step - panvk: Improve panvk_shader_foreach_variant() - panvk: Add a helper for uploading to cmd mem - panvk/csf: Add a helper for dispatching compute shaders with 3D state - compiler/rust: Re-add From> to FromVariants - compiler/rust: Only allow FromVariants on enums - kraid: Make PAN_USE_KRAID per-stage - kraid: Use unsafe with no_mangle - kraid/v9: Simplify DstLanes logic for staging registers - kraid/ir: Rework some RegRange helpers - kraid: Automatically swizzle in From for Src - kraid: Add lowering for COPY.i64 - kraid: Add OpFMul and plumb it through - kraid/nir: Implement nir_op_inot - kraid/data_types: Add message types - kraid/data_types: Add unit tests - kraid: Add OpLea/LdTex and plumb them through - kraid: Add OpLd/StCvt and plumb them through - compiler/rust: Implement Eq/Hash/PartialEq for LowerBoundedU32Array - compiler/rust/bitset: Add an iteration test - compiler/rust/bitset: Further generalize find_next_set() - compiler/rust/bitset: Add a next_set() method - compiler/rust/bitset: Further generalize find_aligned_unset_range() - compiler/rust/bitset: Add a find_aligned_set_range() method - compiler/rust/bitset: Don't write past the end in insert_range() - compiler/rust/bitset: Generalize ConstBitSet::insert_range() - compiler/rust/bitset: Add some range methods to BitSet - compiler/rust/bitset: Add an iter_bit_indices() method - compiler/rust: Add more methods/traits to U8EnumSet - kraid/nir: Implement load_local_invocation_id - kraid: Add OpMux and plumb it through - kraid/isa: Handle 16-bit replicated destinations - kraid: Add OpFrcp/Frsq and plumb them through - kraid/ir: Add a Opcode::set_variant() method - kraid: Widen more ops - kraid/hw_tests: Use a single basic block - kraid: Store blocks in a CFG - compiler/rust/bitset: Enable From/IntoBitSet for u32 - kraid: Copy the SimpleLiveness and LiveSet from NAK - kraid: Add a parallel copy builder - kraid: Use Swizzle::is_none() more - kraid: Add new Phi label type and OpPhiSrc/Dst - kraid/nir: Handle nir_phi_instr - kraid: Implement EnumAsU8 for DstLanes - kraid: Rework supported DstLanes queries - kraid: Allow RegRef::word() on subregs - Revert "compiler/rust/bitset: Add an iter_bit_indices() method" - compiler/rust/bitset: Fix a unit test - compiler/rust/bitset: Add a count_set_in_range() method - compiler/rust/bitset: Implement Eq and PartialEq - compiler/rust/bitset: Add a retain() method - kraid: Better RA - kraid/nir: Use correct zero sizes for unused ALU components - kraid/swizzle: Enable Swizzle::swizzle() on word swizzles - kraid/ir,v9: Fix swizzles for the accum source of OpMkVecV2I8I16 - kraid/ra: Re-swizzle 64-bit sources that read 32-bit values - kraid/ra: More accurately compute source constraints - kraid/nir: Allow i8v3 ops - kraid: Use a tuple struct for SSARef - kraid: Implement FromIterator for SSARef - kraid: Implement load_ubo - kraid/lower_copy: Use Src::imm_u8() for shifts - kraid: Use a Builder in ParallelCopy - kraid/parallel_copy: Emit small constants directly - kraid/ra: Delete a left-over debug check - kraid/nir: implement nir_op_[ui](add|sub)_sat - kraid/data_type: Add more auto types - kraid: Add a DataType::SR special case - kraid: Add OpLeaBuf and plumb it through - kraid: Add OpTex* - kraid/nir: Plumb through texture ops - kraid: Use flat_map() instead of map().flatten() - kraid/data_type: Handle SR in as_data_type() - kraid/v9: Actually encode OpTexGradient - kraid/v9: Fix src_supports_imm32() for Op[IF]Add - kraid: Add a vec src legalization pass - kraid/model: Add an op_src_supports_mod() query - kraid: Add a word-based copy propagation pass - kraid/widen: Don't widen messages - kraid/nir: Enable load_global_constant - kraid/v9: Use the right data type for OpShiftLop::src_supports_imm32() - kraid: Add OpAtom* and plumb them through - kraid/v9: Don't allow src0 swizzles in OpShiftLop::src_supports_imm32() - kraid: Run copy-prop after legalizing_src_swizzles() - kraid/copy-prop: Don't propagate SSA values with mismatched sizes - kraid/swizzle: Expose the guts of swizzle composition - kraid/copy-prop: Add byte-based copy propagation - kraid: Pass the immediate to Model::op_src_supports_imm32() - kraid/v9: Support immediate buffer/texture handles - kraid/nir: Always use a destination for AtomOp::Xchg - kraid/ra: Handle OpPhiSrc with a swizzle - kraid: Call pan_shader_update_info() - kraid/nir: Respect FLOAT_CONTROLS_ROUNDING_MODE_RTZ - pan/nir: Lower read_invocation to 32 bits - compiler/rust/cfg: Assert that nodes are in a dominance-respecting order - compiler/rust/cfg: Unexpose CFG::from_blocks_edges() - compiler/rust/cfg: Make sorting optional in CFGBuilder::as_cfg() - kraid: Stop re-sorting blocks with CFGBuilder - kraid: Return an Option from Model::preload_reg() - kraid/nir: Use FAURef::user_i32() - kraid: Add special FAUs - kraid: Add OpBarrier and plumb it through - kraid/nir: Respect access flags on loads/store ops - kraid: Plumb TLS size through to pan_shader_info - kraid/nir: Implement load_scratch/shared_base_ptr - pan/nir: Lower scratch and shared to global for Kraid - kraid: Add OpWMask and plumb it through - kraid/nir: Implement load_subgroup_invocation - kraid: Add OpClper and plumb it through - kraid: Don't report Src::is_zero() with a BNot modifier - kraid/copy-prop: Trivialize zero copies - kraid/ir: Rename the raw src/dst type helpers - kraid: Add a DataType::total_bytes() helper - kraid/validate: Fix source swizzle validation - kraid/ra: Fix W1 widens - kraid: Fix lower_small_constants() for 64-bit sources - kraid: Take a DataType in Opcode::is_valid_variant() - kraid/ra: Also handle OpPhi swizzles in the pre-existing live-out case - kraid/copy-prop: Try to re-type opcodes for more widening - kraid/copy-prop: Fold widen ops into 64-bit sources - kraid/copy-prop: Treat F16ToF32 as a widening copy - compiler/rust/bitset: Improve test_find_aligned_unset_range() - compiler/rust/bitset: Enhance find_aligned_[un]set_range() - nvk/image: Style nits - nvk/image: Rewrite nvk_image_can_compress() to use early returns - nvk/image: Take an nvk_physical_device in can_compress() - nvk: Add an NVK_DEBUG=no_compression flag - vulkan/meta: Use z_off/scale for 2D array images as well - vulkan/meta: Allow resolving a 2D MSAA image to a 3D image - kraid/ra: Fix find_unpinned_bytes() for unaligned ranges - kraid/ra: Relax alignment requirements for staging registers - kraid/nir: Rework mov/vec handling - kraid/nir: Implement nir_op_insert_* - kraid/nir: Implement as_uniform - kraid: Add a Src::fneg_zero() helper - kraid: Use FMA instead of FMUL - kraid/ir: Don't compare labels in FAU/RegRef.eq() - kraid/nir: Add a special_fau() helper - kraid: Add a new FAUModel - kraid: Merge legalize_immmmediates and legalize_vec_srcs - compiler/rust: Add U8EnumSet::len() and ConstBitSet::len() - kraid/legalize: Add a move_src_to_tmp() helepr - kraid: Legalize FAU sources - kraid: Add OpIDpAdd and plumb it through - nvk: Replace nvk_addr_range with VkDeviceAddressRange - pan: Take a stage parameter to get_nir_shader_compiler_options() - pan: Move PAN_USE_KRAID into pan_compiler.c/h - kraid: Expose our own NIR compiler options - pan: Use Kraid's NIR options when it's enabled - kraid/isa,model: Add a op_srs_is_64bit() query - kraid: Align registers based on the new ISA query - kraid/copy-prop: Handle 64-bit OpShiftLop - kraid/nir: Implement 64-bit op_bitfield_select - kraid/nir: Enable more 64-bit ops - nir: Add combined shift-logic ops for panfrost - kraid: Use the new NIR shift+logic ops - kraid: Optimize shift+logic ops - kraid: Document a couple passes - kraid/nir: Implement nir_op_[iu]mul_2x32_64 - nvk: Advertise minStorageBufferOffsetAlignment=4 for VKD3D - docs: Add a note about Vulkan implicit sync in the 25.3.0 release notes - compiler/rust/cfg: Remap node edges in remove_unreachable() - compiler/rust/nir: Implement Send+Sync for nir_shader_compiler_options - kraid: Use nir_shader_compiler_options directly Feelthepain77 (1): - freedreno: add Adreno 613 (Snapdragon 4 Gen 2) to device list Filip Gawin (5): - r300: avoid UB through implicit conversions on 32bit - r300: use uint32_t instead of long in vertprog - nv30: fix truncated values in line_stipple_pattern - nv30: fix 1 << 31 issues - nv30: fix another left shift cannot be represented in type 'int' Francisco Jerez (4): - nir/divergence: Allow local_invocation_id.z to be treated as uniform. - intel/brw: Sort scheduling modes by performance after initial RA failure. - intel/brw/swsb: Omit redundant read-after-read synchronization for back-to-back DPAS. - intel/brw: Add NIR pass to vectorize dot products into DPAS matrix multiplications. Frank Binns (21): - pvr/ci: drop two tests from bxs-4-64-{fails,flakes} - pvr: re-enable {EXT,KHR}_index_type_uint8 - pvr/ci: add AXE-1-16M nightly Vulkan CTS testing - pvr/ci: skip timing out VK reconvergence test for AXE-1-16M - pvr/ci: add some timing out tests on AXE-1-16M to skips list - pvr: drop unused struct member from pvr_render_pass_attachment - pvr: drop unused pvr_descriptor struct - pvr: enable KHR_external_semaphore{,_fd} unconditionally - pvr: define PVR_USE_WSI_PLATFORM for xcb and xlib - pvr: move PVR_USE_WSI_PLATFORM_DISPLAY into a header - pvr: advertise VK_EXT_display_surface_counter - pvr: advertise VK_EXT_display_control - pvr: advertise VK_EXT_direct_mode_display - pvr: advertise VK_{KHR,EXT}_surface_maintenance1 - pvr: advertise VK_{KHR,EXT}_swapchain_maintenance1 - pvr: advertise VK_EXT_swapchain_colorspace - pvr: advertise support for VK_EXT_acquire_drm_display - pvr: advertise VK_KHR_unified_image_layouts - pvr: rearrange some functions in pvr_arch_border.c - pvr: setup all format fields for custom border color entries - zink: gate some EXT_descriptor_indexing related code Frank Bouwer (3): - pvr: Fix for depth stencil 2d array writes. - Revert "pvr: Fix for depth stencil 2d array writes." - pvr: Fix for depth stencil 2d array writes. Fyodor Kyslov (1): - mesa3d: gfxstream: Add P210 format support GKraats (2): - hasvk: unbreak assert format != ISL_FORMAT_UNSUPPORTED - crocus: Fix shader precompilation on Gen6 and higher Ganesh Belgur Ramachandra (8): - amd: import gfx11.7 addrlib - amd: add initial common code for gfx11.7 - radeonsi: add gfx11.7 - radv: add gfx11.7 - amd: use gfx_level instead of family_id to choose addrlib - amd/llvm: fix target feature setting (DumpCode -> dumpcode) - amd/llvm: fix LLVM asserts for signed integer constants - amd/llvm: truncate const intergers to bitwidth Georg Lehmann (118): - nir: remove nir_link_xfb_varyings - radv: allow input attachment to use pixel coord optimization - radv: move per-primitive fixup closer to radv_nir_lower_io - radv: move fs view_index handling after lowering io - radv: remove unused vs/tes num_outputs from shader info - radv: never call nir_assign_io_var_locations - radv: remove draw_id from mesh shader a bit later - radv: export multi view index as layer after lowering io - radv: remove radv_graphics_shaders_link - nir: disable fp class analysis for 64bit transcendentals - intel/nir_opt_peephole_ffma: fix fp_math_ctlr for modifiers - nir/instr_set: allow cse with fp_math_ctrl mismatches for intrinsics - nir/opt_varyings: back propagate signed zero information to outputs - nir/opt_varyings: do no_signed_zero linking even for non removable stores - nir/opt_algebraic: add more fmulz pattern - ac/nir/lower_tex_coord: fix moving wqm coordinates - nir: fix fp_math_ctrl in fisnan - nir/opt_peephole_select: do not count fmul towards the limit when only used by fadd - nir/loop_analyze: do not count fmul towards the limit when only used by fadd - nir,amd: reassociate fadd to create more fma/mad - radv/ci: update restricted trace checksums - radv: fix amount of sample shading with required sample shaded inputs - ac/nir/lower_tex_coords: fix optimizing cube txd to tex - aco: add tests for cube txd to tex opt - nir/opt_uniform_subgroup: preserve divergence during optimization - tgsi: delete unused lowering pass - aco/tests: use explicit lod in sparse texture test - spirv: always preserve infinities for FMin, FMax and FClamp - radv: use radv_get_sampled_image_desc_size instead of open coding it - radv: add radv_force_64_byte_sampled_image dri conf option - radv: enable radv_force_64_byte_sampled_image for Forza Horizon 6 - aco/optimizer: only create v_fma_legacy_f32 when denorms are disabled - nir: seperate ffmaz from has_fmulz - ac/llvm: don't assert on 32bit ffma before gfx9 - ac/llvm: never create ffmaz for broken llvm - radv: support VK_KHR_shader_fma - aco/gfx8: fix 16bit nir_op_ffma - nir/deref: consider atomics that store derefs as complex use - aco/gfx6: fix fp64 floor lowering - radv: don't lower dfloor in NIR - aco/gfx6: fix fceil lowering - aco/gfx6: use shorter lowering for ftrunc - aco: add rtne pseudo opcodes for fp64 add and fract - aco/gfx6: always use rtne for floor/ceil lowering - aco/gfx6: fix fround_even(-0.0) - aco/gfx6: always use rtne adds for fround_even lowering - radv: enable fp64 float controls on gfx6-7 - aco/isel: never manually flush denorms after 32bit fma - nir: preserve infinities and signed zero during atan2 - amd/gpu_info: precompute instruction prefetch distance - amd/common: don't pass radeon_info to ac_align_shader_binary_for_prefetch - amd/common: add helper for INST_PREF_SIZE - radv: remove gfx6 code from ngg emission - radv/gfx11+: program INST_PREF_SIZE for compute - radv/gfx11+: program INST_PREF_SIZE for pixel shaders - aco: add exec_size to prolog/epilog callback - radv/gfx12: program SPI_SHADER_PGM_RSRC4_GS for seperately compiled gs - radv/gfx11+: program INST_PREF_SIZE for NGG and HS - radeonsi: use ac_get_instr_prefetch_size - radeonsi: use exec_size from the aco prolog/epilog callback - aco/ra: fix inline constants with v_dot2c_f32_f16 - aco/sched_vopd: fix v_dual_dot2acc_f32_f16 created from VOP2 with inline constant - radv: fix setting inline push constants when only the last one is used - radv: inline 8 and 16bit push constant loads - aco/tests: test v_pk_fmac_f16 and v_dotc_f32_f16 with inline constants - aco/tests: test creating v_dual_dot2acc_f32_f16 from v_dot2c_f32_f16 with inline constant - nir/skip_helpers: fix stores with ACCESS_INCLUDE_HELPERS - nir/skip_helpers: handle vendored store_scratch - nir/skip_helpers: keep descriptors uniform even for stores that skip helpers - nir/skip_helpers: don't require helpers for non uniform descriptors - aco/assembler: chain branches in emit order - aco/assembler: do not abort when exec is written after position exports - ac/nir/mem_vectorize: never create vec5 stores - aco/assembler: don't reorder branch insertion block index twice - radv/gfx11+: do not use s[0:1] for unused scratch VA in compute shaders - radv: remove some dead compute scratch code - zink/ci: skip unvanquished-ultra trace on van gogh too - panfrost/lower_bool_to_bitsize: do not assume loop phi source order - nir/phi_builder: do not sort predecessors for phi sources - nir/to_lcssa: do not sort predecessors for phi sources - nir: generalize loop simplification - nir: add pass to optimize shared variables to subgroup operations - nir/opt_algebraic: fix vkd3d-proton pack_half_rtz pattern - aco/isel: emit v_mul_i32_i24 for imul with negative constant - aco/isel: emit v_mul_hi_i32_i24 for imul_high if possible - aco: remove isel setup code for no longer implemented intrinsics - nir,amd: split SGPR input intrinsic to specify workgroup divergence - ac/lower_intrinsics_to_args: use workgroup divergent ttmp intrinsic for subgroup id - nir/divergence: always consider load_ttmp_register_amd uniform - amd: use load_scalar_arg_wg_div_amd for workgroup divergent sgprs - nir/divergence: alyways consider load_scalar_arg_amd uniform - spirv: add option to treat FMax/FMin/FClamp like NMax - radv: add radv_force_nan_preserve_min_max option - radv: enable radv_force_nan_preserve_min_max for DOOM: The Dark Ages - nir: remove explict num_components from nir_def_rewrite_uses_with_alu_src - nir/opt_vectorize: prefer to swizzle vector phis at the destination, not the source - ac/nir: vectorize phis - radv: call nir_opt_phi_precision - nir: clean up weird qsort_r usage - nir/opt_shrink_vectors: restore load_const deduplication - nir/opt_sink: don't sink comparisons that use ballot(true) - radv: run nir_opt_reassociate_for_fma for VS/GS too - vulkan/nir_lower_heaps: assume no heap addressing can overflow - nir: add num_lsb_zero analysis for 64bit pack and u2u - ac/nir_lower_global_access: assume both addition operands are aligned if one is - nir: add num_lsb intrinsic index for amd arg loads - radv: add dword alignment information to descriptor set/heap pointers - ac/nir/lower_ngg: use workgroup divergence analysis for culling - ac/nir/lower_ngg: allow reuse of workgroup divergent variables even when subgroup ops are used - nir/opt_dead_write_vars: handle atomics as reads - nir: fix divergence for deref_cast - aco/live_var_analysis: make sure shared vgprs are within the encodable vgprs - nir/unsigned_upper_bound: fix float to int conversions - nir: mark some AMD specific shuffles as subgroup ops - aco/optimizer: fix skip_smem_offset_align - nir: support phi sources in nir_rematerialize_deref_in_use_blocks - nir/to_lcssa: fix progress for derefs - nir/to_lcssa: move constants before the loop instead of creating a phi Gert Wollny (67): - r600/sfn: Add lowering of tess inner and outer default intrinsics - r600: replace TGSI TCS passthrough with NIR version - r600: replace TGSI query shader with nir - r600/sfn: run nir_opt_idiv_const - r600/sfn: Avoid creating group-tagged registers for ALU dests - r600/sfn: signal progress when splitting address loads - r600/sfn: run additional optimization only after successful address split - r600/sfn: Extract some helpers from schedule_alu - r600/sfn: don't use return parameters in extracted method - r600/sfn: Extract schedule alu groups first - r600/sfn: extract fill_alu_group - r600/sfn: pass reference to group when possible - r600/sfn: Extract group fill failure handling - r600/sfn: extract t-slot allocation when filling ALU groups - r600/sfn: extract idx load state handling in scheduler - r600/sfn: make ALU scheduling return values more meaningful - r600/sfn: collaps no_schedule and scheduled - r600/sfn: simplify ALU scheduling failure handling - r600/sfn: split kcache evaluation into try and commit - r600/sfn: make try_kcache_reservation const - r600/sfn: move tracking of kcache reservation failure to scheduler - r600/sfn: Move tracking of kcache reservation to AluScheduleContext - r600/sfn: extract kcache check out of schedule_alu_to_group_vec - r600/sfn: refactor BlockScheduler::schedule_block - r600/sfn: Move exports emission to helper - r600/sfn: extract check and report for unscheduled instructions - r600/sfn: deduplicate some code in DCE - r600/sfn: deduplicate optimizer logging code - r600/sfn: deduplicate fixpoint loop for optimizers - r600/sfn: refactor CopyPropFwdVisitor::propagate_to - r600/sfn: refactor CopyPropFwdVisitor::visit(AluInsr*) - r600/sfn: extract logging from CopyPropFwdVisitor::visit(AluInstr*) - r600/sfn: refactor CopyPropBackVisitor::visit(AluInstr*) - r600/sfn: Fix typo with AssemberVisitor - r600/sfn: Make some member variables references - r600/sfn: Refactor AssemblerVisitor emit_alu_op - r600/sfn: de-duplicate emit_wait_ack - r600/sfn: use c++ pattern for zero-init of structs - r600/sfn: extract some byte code emission from assembler - r600/sfn: Drop index register handler in assembler - r600/sfn: simplify fill bytecode - r600/sfn: extract emitting the bytecode of Rat Instr too - r600/sfn: extract and decouple ALU post-emit state update - r600/sfn: return LDS opcode properties as tuple - r600/sfn: use opcode switch in ALU post-emit update - r600/sfn: use local opcode consistently in emit_alu_op - r600/sfn: Move lds_queue_read decrement out of prepare_alu_src to caller - r600/sfn: Move copy_src to sfn_fill_bytecode.cpp, rename to fill_alu_src - r600/sfn: Move prepare_alu_src to sfn_fill_bytecode.cpp, rename to fill_alu_src_operands - r600/sfn: Move prepare_alu_dst/copy_dst to sfn_fill_bytecode.cpp, rename to fill_alu_dst - r600/sfn: drop unused literals tracking in assembler - r600/sfn: minor reordering of operations in assembler - r600/sfn: Move last_addr handling out of fill_alu_dst - r600/sfn: Simplify m_last_addr tracking in emit_alu_op - r600/sfn: Validate ALU dst writes in emit_alu_op - r600/sfn: Extract ALU bytecode emission helper - r600/sfn: Handle dst write checks before mova setup split - r600/sfn: Extract ALU dst state update into AssemblerVisitor - r600/sfn: Move LDS ALU emission to fill_bytecode - r600/sfn: Make emit_alu_op return success status - r600/sfn: Move opcode_map to fill_bytecode, pass EAluOp to emit_bytecode_alu - r600/sfn: Move ds_opcode_map ownership to fill_bytecode - r600/sfn: Drop unused AssemblerVisitor members - r600/sfn: Add pin_to_chan method to Register and use it - r600/sfn: rename pin_dest_to_chan to pin_registers - r600/sfn: Pin alu sources as well when registers are pinned - r600/sfn: Drop assertions when emitting IF asm instruction Gleb Mazovetskiy (1): - os_misc.c: add missing include for mach_host_self() Gleb Popov (1): - Rename the CACHE_LINE_SIZE define to MESA_CACHE_LINE_SIZE Grant Nichol (1): - ethosu: Fix -Werror=format build error on 32-bit Gu, Wangfeng (3): - radv/sqtt: add instruction timing SE mask controls - radv/sqtt: emit pending barrier end before API markers - ac/spm: clamp cache miss counts in derived counters Gurchetan Singh (14): - gfxstream: fix string array marshalling - gfxstream: emit global state wrapped decoding for vkCmdEvent - subprojects: update libc-rs to 0.2.185 - subprojects: update to rustix 1.1.4 + downstream patches - freedreno: fix ignored qualifier - tu: fix -Wmissing-prototypes errors - tu: fix implicit fallthrough - tu: kgsl: fix -Wgnu-alignof-expression warning with Clang - freedreno: explicitly declare required depend_files, part 1 - freedreno: explicitly declare required depend_files, part 2 - util: rust: sync error handling fixes from downstream - util: rust: minor fixups - virtio: add magma-gpu-rs subdirectory - docs: fix references to moved crates Han, Mike (3): - amd/vpelib: complete 16bpc RGBA format mapping for 10/12bpc msb/lsb support - amd/vpelib: add format support check - amd/vpelib: Add missing argb variant support Hans-Kristian Arntzen (21): - wsi/common: Report correct time domain in VkPresentTimingInfo. - loader: Separate out X11 specific screen queries from dri_helper.h. - loader: Clear screen resources struct on init. - wsi/x11: Setup screen resources on x11_connection creation. - wsi/x11: Add helper to find appropriate screen resources for a window. - wsi/x11: Set up screen resources on swapchain creation. - wsi/x11: Add helper to compute xrandr rate estimate. - wsi/x11: Update xrandr refresh estimate on geometry change. - wsi/x11: Update refresh rate estimate based on MSC feedback. - wsi/x11: Implement main body of present timing. - wsi/x11: Add Xwl support for present timing. - wsi/x11: Only accept VRR refresh rates when we're flipping. - wsi/common: Prefer host query resets when available. - wsi/common: Pass along requested timing feedback as well. - wsi/x11: Avoid non-causal present timings when not flipping. - wsi/common: Refactor out the search for a present_timing struct. - wsi/common: Ensure that google display timing results propagate. - wsi/common: Always ensure that we can get a GPU done timestamp. - wsi/x11: Be more adaptive in how much the sleep is pulled back. - radv: Consider VkImageView usage rather than VkImage usage in feedback. - radv: Only consider default feedback loops for appropriate layouts. Hsieh, Mike (2): - amd/vpelib: add optional __stdcall calling convention via build option - amd/vpelib: add indirect shaper config support Hyunjun Ko (13): - anv/video: fix up H.264/H.265 encode session parameters to match advertised caps - anv/video: fix to set the upper bound of the bitstream of h265. - anv/video: Add to check size mismatch during motion field estimation. - anv/video: define ANV_VIDEO_AV1_MAX_DPB_SLOTS - anv/video: Change size of the cached array of recently decoded AV1 frames. - intel/genxml: update VDENC commands for gen125 - anv/video: Add h264 vdenc tables from media-driver - anv/video: Make H264 encoder work on Gen125 - anv/video: fix to set valid coded size for the source pictures. - anv/video: Add h265 vdenc tables from media-driver - anv/video: Make H265 encoder work on Gen125 - anv/video: Enable video encoding on gen125 - anv/video: Support H265 10-bit encoding Iago Toral Quiroga (2): - pan/bi: TEX_GRADIENT may need helper invocations - CODEOWNERS: update broadcom maintainers Ian Romanick (21): - brw: Lower all phis to scalar - brw: Don't lower phis involved in DPAS instructions to scalar - brw: Calcuate divergence before brw_from_nir - nir/opt_constant_folding: Don't fight with nir_lower_bit_size - nir: Use nir_instr_remove_v in nir_def_replace - nir/opt_if: use nir_def_replace() instead of nir_def_rewrite_uses() - nir/opt_if: Merge if-statements with inverted conditions - nir/algebraic: Convert bcsel of addition to addition of b2i or b2f - nir/opt_shrink_stores: Don't shrink ivec2 stores to int64 images - brw: Use nir_opt_shrink_stores - brw: Use nir_opt_shrink_vectors - brw: Add functions to calculate flags usage without a brw_inst - brw: Replace logical operations with predication - brw: Use nir_opt_uub - brw: Use nir_opt_fp_math_ctrl - elk: Use nir_opt_uub - elk: Use nir_opt_fp_math_ctrl - nir/divergence: Handle SYSTEM_VALUE_INSTANCE_INDEX - brw/predicate: Add missing test with farther_flags - brw: Handle empty top block in brw_nir_move_interpolation_to_top - brw/validate: Gfx11 can't have accumulator src0 in 3-src instructions Icenowy Zheng (36): - pvr: follow other drivers' practice for copying build ID - pvr: skip emitting query program when copy result / reset with 0 queries - isaspec: decode: manually print the sign when printing NaN float values - pvr: wait for graphics jobs in CopyQueryPoolResults - pvr: increase maxPerStageResources for new maxPerStageDescriptorStorageBuffers - pvr: do not setup deferred RTA clear for active render targets - pvr: properly handle deferred RTA clears for 2D array view of 3D image - pvr: add deferred RTA clear command to list after checking it's not NULL - pvr: record deferred RTA clears for secondary cmdbuf subcmds - pvr: ignore DS attachment's D or S when it's unused in dynamic rendering - dri: try to enable GL_ARB_compatiblity when supported GL core version is 3.1 - pvr: setup viewindex if the shader wants it even when multiview disabled - pvr: prohibit clang-format from touching the dri options list - pvr: add dri options used by common WSI code - pvr: fix handling of invalid attachment info in pvr_init_fs_outputs_mrt - pvr: copy sub_cmd flags except owned when executing subcmds out of pass - pvr: stop to derive rt datasets based on geometry_terminate - pvr: add a structure containing data kept for suspended renderpasses - pvr: preserve and pass more data for suspending render passes - pvr: remove dEQP-VK.pipeline.monolithic.misc.no_rendering from fail list - pvr: return FORMAT_NOT_SUPPORTED for unknown image types - pvr: prevent direct access to VkImageSubresourceLayers::layerCount - pvr: implement CmdBindIndexBuffer2 - pvr: implement GetRenderingAreaGranularity - pvr: implement GetImageSubresourceLayout2 - pvr: implement GetDeviceImageSubresourceLayout - pvr: advertise VK_KHR_maintenance5 - zink: move maint5 to gl21_baseline capabilities set - docs/zink: add maint5 to the list of required extensions - llvmpipe: stub other functions inside compute shaders for ORCJIT - pvr: bump conformance version to 1.4.3.3 - Revert "pipe-loader: fallback to zink instead of kmsro for render nodes" - pipe-loader: use zink for powervr device nodes - zink: check Z/S aspect before creating Z/S image view - pvr: apply the culling everything viewport shift for only triangles - vulkan: update spec to 1.4.354 Iván Briano (15): - anv: silence warning - intel/brw: add load_coverage_mask_intel intrinsic - intel/brw: add load_msaa_rate_intel intrinsic - intel/brw: add load_frag_shading_rate_intel - anv/brw: add conservative raster on/off to FS_CONFIG - anv/brw: handle FullyCoveredEXT - anv: add and use a drirc option to enable FullyCovered for vkd3d - anv: fix return of cmd_buffer_set_indirect_stride() function - anv, iris: fix MOCS Index setting of EXECUTE_INDIRECT_* commands - intel/dev: ARL-H supports EXECUTE_INDIRECT_* - anv: don't try to clear d/s attachments not backed by an image - brw/rt: fix max_t selection on intersection report - brw/rt: split HitAttribute area in pending/committed - brw/rt, anv: reduce maxRayHitAttributeSize - anv: fix 2d-array to 3d blits Jaakko Jokinen (1): - nir: Add cases to nir_get_io_offset_src_number() JaeHoon Lee (25): - v3d: release the texture reference if shadow resource creation fails - v3d: drop the tiled temporary when bailing on unsupported blits - v3d: free the cache buffer when loading a corrupt disk cache entry - v3d: create the compute job after the zero-sized dispatch check - v3dv: only report 16-bit float formats as blendable at 32/64 bpp - vc4: fix last_layer selection in the blit sampler view - v3d: fix slot and input indexing in v3d_set_global_binding - v3dv: report maxDrawIndirectCount of 1 without multiDrawIndirect - v3dv: honor wait dependencies for job-less submissions - v3d: clamp transform feedback offset to buffer size - v3dv: report the correct dynamic storage buffer UAB limit - v3dv: fix blake3 key truncated to 20 bytes in pipeline cache - v3d: fix blake3 key truncated to 20 bytes in shader cache - vc4: fix incorrect resource unref in vc4_flush_resource - vc4: free vertex and constant buffers on context destroy - broadcom/compiler: really enable GFXH-1625 TMUWT validation - broadcom/compiler: validate magic waddr writes - broadcom/qpu: remove empty qpu_validate.c - v3dv: make room in the descriptor map for the no-sampler entries - v3dv: use the binning VS variant for the binning VPM config - v3dv: record the multiview geometry shader with its Vulkan stage bit - v3dv: record the no-op fragment shader with its Vulkan stage bit - nvk: free copy_memory_indirect_temps on command buffer destroy - v3dv: avoid restoring stale descriptor state after a meta op - nvk: report fills from memory correctly Jaishankar Rajendran (2): - vulkan/runtime: enable parametrization of ASTC software decode - anv: tune parameters of the ASTC software decoding Jakob Sinclair (13): - panvk: Enable scissor_mode for draws - panvk: Remove unnecessary functions - vulkan/meta: Don't issue a full drawcall for clears - pan: Support lowering D24X8 to D24 - gallium: fix type size in z24_unorm_packed_pack_z_32unorm - pan/va: Decode support for ARSHIFT_OR on Valhall - pan: Add G52 skip for xlib wsi failure - pan/compiler: fix spilling for 64-bit values - pan: Add missing v14 primitive flag - panvk/draw: Separate build from prepare functions - panvk/csf: Use RUN_FULLSCREEN for cmd_draw_rects - panvk/csf: Use RUN_FULLSCREEN for cmd_draw_volume - pan/crc: Fix CRC check for sparse AFBC images Jan Meisel (3): - nir/range_analysis: handle msad_4x8 in unsigned upper bound - radeonsi/vcn: fail feedback for truncated encodes - radv: fix RADV_PERFTEST=nircache enablement Janne Grunau (7): - nir/gather_info: clear interpolation qualifiers only in fragment stage - asahi: nir: lower flrp64 - asahi: ci: Drop no longer failing VK.wsi.xcb.present_timing test - panfrost: ci: Drop no longer failing VK.wsi.xcb.present_timing test - asahi: ci: Add failing b10g11r11 and e5b9g9r9 copy tests - hk: xfb: Avoid assertions in nir_slot_num_components - poly: Fix comment after moving passthrough_gs Jason Macnak (6): - gfxstream: Override VkDeviceDeviceMemoryReportCreateInfoEXT vk.xml - virtgpu_kumquat_ffi: replace mutex.get_mut() with mutex.lock() - gfxstream: support testing d32 s8 - gfxstream: kumquat: validate device dmabuf support before use - gfxstream: route vkGet*ProcAddr to VkDecoderGlobalState - gfxstream: Avoid transfering VkAllocationCallbacks between guest and host Jeremy Gebben (5): - kk: Implement VK_KHR_dynamic_rendering_local_read - kk: Refactor encoder state - kk: Set availability for extra multiview queries in vkCmdEndQuery() - kk: Implement VK_QUERY_TYPE_TIMESTAMP - kk: Fix Vulkan to Metal stage translation for timestamps Jeremy Huddleston (38): - bin/install_megadrivers: Bail out if libname suffix is never reached - gallium/targets: Use libname_suffix for installed driver names - glx/apple: Convert K&R-style declarations to ANSI prototypes - glx/apple: Switch logging to os_log on macOS 10.12+ - glx/apple: Replace apple_glx_diagnostic with apple_glx_log_* - glx: free visinfo on BadMatch in glXCreateWindow's AppleGL path - glx: Fix stale end-comment on __glXInitialize direct-rendering block - glx: drop redundant __glXErrorString forward declaration - glx: fix DRI3-not-available diagnostic skip on macOS - glx: NULL-check frontend_screen in glXCreateContextAttribsARB - glx: simplify FBConfig wire decode - glx: free glx_drawable on CreateDRIDrawable failure - glx: bail bind_extensions on screens without frontend_screen - glx: drop dead AppleGL glXGetProcAddressARB fallback - glx/apple: Add create_context_attribs entry to the applegl_screen_vtable - zink: fix GLX_USE_APPLE typo (should be GLX_USE_APPLEGL) - glx: drop dead GLX_USE_APPLE check inside glXSwapBuffers - glx: Guard declaration of glx_accel and kopper to match use - glx/apple: free gc in applegl_destroy_context - glx: fix per-display drawHash / zombieGLXDrawable / dri2Hash leak on GLX_USE_APPLE builds - glx/apple: silence OpenGL deprecation warnings - glx/apple: return CGLError from apple_visual_create_pfobj instead of aborting - glx: route copy_context through a vtable slot - glx: route swap_buffers through a vtable slot - glx: extract drawable lifecycle into a vtable - glx/apple: allow selection between AppleGL and Gallium at runtime for GLX_USE_APPLE=1 builds - glx/apple: skip AppleGL election on macOS 26 and newer - glx: remove GLX_USE_APPLE and collapse the guards it gated - glx: Fold __glXGetDrawableAttribute and __glXQueryDrawable together - glx: Fold CreatePbuffer/DestroyPbuffer into CreateDrawable/DestroyDrawable - zink: Add missing link against libxcb-present - zink: Address libvulkan.1.dylib dlopen failure on macOS - glx/apple: honor the client-requested GLX context version and profile - llvmpipe: link all LLVM targets on Apple to fix build failure when using static LLVM libraries - dri/st: Fall back to Z32_FLOAT depth configs when Z32_UNORM is unsupported - glx/apple: only skip AppleGL election on macOS 26.0 through 26.5 - llvmpipe: don't create a screen when the process is not allowed to JIT - glx/apple: silence OpenGL deprecation warnings in libglx Jesse Natalie (23): - d3d12: Handle THREAD_SAFE maps and use them for async query results - microsoft/compiler: Back-propagate interpolator modes from FS - wgl: Use an hwnd xor hdc for framebuffers - d3d12: add screen pending-free list plumbing - d3d12: clear stale per-context BO state at context destroy - d3d12: transfer batch local_bos refs to screen at submit - d3d12: transfer batch->bos refs to screen at submit - d3d12: reclaim in-flight BO memory on allocation failure - d3d12: implement pb_fence vtbl for cache/slab reuse - d3d12: drop peer-batch peeking in resource_is_busy / wait_idle - d3d12: proactively trim completed pending-free entries - nir_lower_non_uniform_access: Add ASSERTED for assert-only var - va: Wrap assert-only code in NDEBUG - microsoft/compiler: Don't assume phi ordering - util: Fix u_math on MSVC arm64 - mesa/st: PBO memory barriers imply image barrier if PBO download goes through compute - d3d12: Use enhanced barriers for memory_barrier when we can - d3d12: Fix transition_array_size for 3D textures - d3d12: Fix WARP version detection for broken int64 - ci/windows: Update WARP to 1.0.20 - wgl: Move sub-8bpc pixel formats to extended format list to match other Windows drivers - d3d12: Disable vao fast path for AMD - mesa: Fix shared state bookkeeping for dynamic share list changes (wglShareLists) Jhanani Thiagarajan (1): - intel/mda: Change the default output directory Jianfeng Liu (1): - freedreno/drm: Fix uninitialized read of BO metadata on import Jianxun Zhang (1): - intel/decoder: Print more information in shader's headline Jiyu Yang (3): - nir/loop_analyze: Use pass_flags for memoization in is_only_uniform_src - panfrost: cleanup precomp_cache on screen destroy - egl/dri2: exclude >8bpc configs from GLES1 renderable/conformant bits Job Noorman (60): - ir3/ra: fix killed src detection while spilling - ir3/shared_ra: fix live-out reload after src reload - nir/get_io_offset_src_number: support \@load/store_global_ir3 - ir3/isa: use same src for ldg.a OFF field on a6xx/a7xx - ir3: always use byte offset for \@load/store_global_ir3 - nir/opt_offsets: add support for \@load/store_global_ir3 - ir3: move feature check down in ir3_nir_max_imm_offset - ir3: enable opt_offsets for load/store_global_offset - ir3: mark __alias_n as UNUSED in foreach_src_in_alias_group_n - ir3/cf: fix rewriting uses with different dst types - ir3/shared_ra: use ir3_cursor instead of instr in reload helpers - ir3/shared_ra: insert reloads before tied dst pcopies - ir3/cp: support propagating const vecs - ir3: allow const src0 for ldg.a/stg.a/ray_intersection - ir3: don't cache driver param instructions - ir3: allow (ss) on all cat7 instructions - freedreno/computerator: fix UAV view size - ir3/spill: extract child intervals for live-in reloads - ir3/ra: add ir3_ra_src_is_killed helper - ir3/ra: fix killed src detection for spillall min limit - ir3: use a1.x addressing for ldg.k with dst 256 - ir3: don't use bitfields in ir3_shader_output - ir3: don't store shader_options in the cache - freedreno/drm-shim: allow chip selection by chip_id - tu: use chip_id instead of gpu_id for the cache UUID - tu: add option to override the build ID - vulkan: add vk_shader_module_hash helper - vulkan: use consistent module hashing for pipeline stages - nir/lower_undef_to_zero: add filter argument - ir3: lower undef booleans to zero - nir/get_io_index_src_number: support \@load_ssbo_address - nir/lower_ssbo: take offset_shift into account - nir/lower_ssbo: add option to only lower large SSBOs - nir/lower_ssbo: add option to insert bounds checks - tu: Add option to raise the maximum SSBO size - ir3: fix possible signed overflow in ir3_link_add - ir3/opt_prefetch_descriptors: rematerialize defs at preamble start - nir/lower_vars_to_scratch_global: make callback deterministic - ir3/lower_vars_to_scratch_global: use stable sort for variables - nir: add nir_shader_deref_pass - nir: add nir_src_as_{alu,tex,phi}_src helpers - nir: add nir_convert_address_format pass - nir/lower_explicit_io: add support for 64bit_global_32bit_offset vars - nir/lower_explicit_io: support shifting non-const array index - nir: add load/store_global_offset intrinsics - nir/lower_explicit_io: add support for load/store_global_offset - nir/set_io_offset: add support for adjusting BASE - nir/lower_explicit_io: support offset_shift for 64bit_global_32bit_offset - rusticl/kernel: add support for 64bit_global_32bit_offset - ir3: add support load/store_global_offset - ir3: enable opt_offsets for load/store_global_offset - tu,ir3: use 64bit_global_32bit_offset for global memory - ir3/lower_tess: use load/store_global_offset - ir3/lower_shader_clock: use load_global_offset - ir3: don't manually lower load/store_global - tu/lower_ray_query: use load_global_offset - tu,ir3/analyze_ubo_ranges: use load_global_offset - tu/lower_ssbo_address_size: use load/store_global_offset - nir,ir3: remove load/store_global_ir3 - ir3/opt_preamble: lower load_global_offset to preamble Joe Wang (2): - ac/spm: bump AC_SPM_MAX_COUNTERS_PER_GROUP to 16 - radv,ac/spm: add user-defined raw counter collection Jon Turney (7): - ddebug: Fix use of alloca() without #include "c99_alloca.h" - glx/windows: Avoid shadowing 'type' parameter of driwindowsCreateDrawable() - glx/windows: Add stdbool.h include to 'direct GLX via WGL' implementation - glx/windows: Fix compilation of driwindows_glx after driscreen changed from pointer to member - glx/windows: Fix compliation after code motion to put event base in 'dri' context - glx/windows: Add GLX_USE_WINDOWSGL in new places it's needed to build libGL - glx/windows: Drop static from driwindowsCreateScreen() Jordan Justen (48): - brw: Don't set header_size at init since it will be re-set in later code - brw/compact: Precompact using 2src fields on 3src instructions - intel/gen: Add gen 9 through Xe2 instruction formats in JSON - intel/gen: Add gen_inst_info.py script to generate C++ headers - intel/gen: Create gen_info_util.h - intel/gen: Make use of generated instruction info - intel/gen/compact: Add compact tables from brw/brw_eu_compact.c - intel/gen: Add gen_raw_compact_inst type - intel/gen: Add gen_compact_accessor for compact/uncompact - intel/gen: Implement compact support - intel/gen: Split out type decode functions for use with uncompact - intel/gen: Implement uncompact support - intel/gen: Account for compact nop pad instruction in gen_scan_raw_layout() - intel/gen: Support declaring ISA fields with disconnected bits - intel/gen: Support accessing fields & sub-fields with disconnected bits - intel/gen: Merge THREE_SRC0_VSTRIDE HI/LO into a gen_split_range - intel/gen/xe: Merge THREE_SRC1_VSTRIDE HI/LO into a gen_split_range - intel/gen/xe: Merge BFN_FUNC_CONTROL HI/LO into a gen_split_range - intel/gen: Merge uncompat control bits into a gen_split_range - intel/gen: Merge uncompat datatype bits into a gen_split_range - intel/gen: Merge uncompat subreg bits into a gen_split_range - intel/gen: Merge uncompat src0 bits into a gen_split_range - intel/gen: Merge uncompat src1 bits into a gen_split_range - intel/gen: Merge uncompat 3src control bits into a gen_split_range - intel/gen: Merge uncompat 3src source bits into a gen_split_range - intel/gen/xe: Merge uncompat 3src subreg bits into a gen_split_range - intel/gen: Merge SRC_A16_SWIZZLE HI/LO ranges - intel/gen/xe: Merge Xe2 DATATYPE_INDEX HI2/LO3 into a gen_split_range - intel/gen/xe: Merge Xe2 compact 3src subreg HI2/LO3 into a gen_split_range - intel/gen: Assert that the gen opcode is supported by this platform - intel/gen: Start Xe3P support - intel/gen: Add gen_byte_stride() - intel/gen: Add Xe3P validation for src1 byte stride matching dst - intel/gen/validation: Start enabling Xe3P tests, but skip for now - intel/gen/validation: Update tests for new Xe3P src1 restriction - intel/gen/validation: Drop WA 22016140776 on Xe3P - intel/gen/validation: Enable running validation tests for Xe3P - intel/gen: Remove mac/mach/macl on Xe3P - intel/gen: Add mullh instruction for Xe3P - intel/gen/xe: Rename decode/encode_type_3src to decode/encode_type_short - intel/gen: Disable compact on Xe3P for now - intel/gen: Add Xe3P two source encoding changes - intel/gen/tests/basic: Add 2src round-trip tests covering nvl src1 changes - intel/gen: Add Xe3P three source src1 encoding changes - intel/gen/tests/basic: Add 3src round-trip tests covering nvl src1 changes - intel/gen/compact: Split datatype into 1src / 2src versions - intel/gen: Add Xe3P compact support - intel/executor: Enable Xe3P Jose Maria Casanova Crespo (59): - broadcom/compiler: Add V3D 7.1 v8dot dot product QPU instructions - broadcom/compiler: hardware-accelerated 4x8-bit dot products on V3D 7.1+ - broadcom/compiler: Add v8dot and setnnmode scheduler dependencies. - broadcom/compiler: Eliminate redundant setnnmode instructions - v3dv: Expose hardware-accelerated integer dot products on V3D 7.1+ - broadcom/compiler: move nir_lower_undef_to_zero out of optimization loop - v3dv: bump maxComputeSharedMemorySize to 32 KB - v3d/v3dv: Use new V3D_MAX_CSD_WG_SIZE = 256 - v3dv: lower oversized compute workgroups to 256 invocations - v3dv: include mem_offset in vkCmdFillBuffer destination - v3dv: Enable KHR_shader_subgroup_extended_types - v3dv: expose maxFragmentOutputAttachments as max_rts - v3dv: avoid duplicate bo_handles between cpu_job and CSD lists - v3dv: assert timestamp pool BO is disjoint from dst buffer BO - broadcom/ci: skip SSBO tests close to the 60s threshold on rpi4 - v3dv: avoid 16F TLB usage for B10G11R11_UFLOAT copies - v3dv: advertise VK_EXT_scalar_block_layout on V3D 7.1+ - v3dv: Enable meta_copy_buffer with TFU for V3D 7.1 - v3dv: move destroy_update_buffer_cb to a generic helper - v3dv: use TFU copy with stride-0 for vkCmdFillBuffer - v3dv: extract TFU helpers for format-plane and slice-stride args - v3dv: rename copy_buffer_to_image_tfu to copy_buffer_image_tfu - v3dv: implement TFU image-to-buffer copy on V3D 7.1 - v3dv: relax buffer padding in TFU buffer<->image copy - v3dv: share zero-fill TFU staging BO at device level - v3dv: expose the full simulator memory to applications - broadcom/qpu: support output pack on itof/utof - v3d: move nir_lower_frexp after nir_lower_bit_size - v3dv: lower flrp16 for consistency with flrp32 - v3d: widen sub-32-bit subgroup arithmetic and vote ops - v3d: improve liveness analysis for packed partial writes - broadcom/qpu: expose V3D 7.1 packed-f16 instructions - v3d: emit packed-f16 ALU ops natively on V3D 7.1 - v3dv: enable lowered shaderFloat16/Int16/Int8 + VK_KHR_shader_float16_int8 - broadcom/compiler: fix payload-register liveness condition - v3d: Enables GL_ARB_clip_control for v71+ - v3d: use NO_GUARDBAND clipper for near-zero viewport Z scale - v3dv: set non-zero array stride in null texture descriptor state - broadcom: add and use max_render_targets to devinfo - broadcom: raise framebuffer size to 7680 on V3D 7.1 - v3dv: gate Dawn-required limits and features behind V3D_WEBGPU_OVERRIDE - ci: igalia farm maintenance - v3dv: allow TFU readahead padding above maxMemoryAllocationSize - v3dv: route blending of UNORM16/SNORM16 RTs through software lowering - v3dv: rename format_plane unorm/snorm flags to sw_unorm/sw_snorm - v3dv: fix crash on device creation failure before meta initialization - v3dv: close the primary node fd on physical device destruction - broadcom/compiler: reduce the compile-strategy fallback ladder - broadcom/compiler: split v3d_nir_to_vir_finish out of v3d_nir_to_vir - broadcom/compiler: support probing a compile's pre-spill register pressure - broadcom/compiler: consolidate the compile-strategy logging - broadcom/compiler: move the 2-thread strategies to compile_2t_strategies - broadcom/compiler: pick the 2-thread compile strategy by register pressure - broadcom/compiler: don't leak the compile on assembly allocation failure - broadcom/compiler: drop V3D_DEBUG=opt_compile_time - broadcom/ci: unskip CTS tests that are no longer slow - broadcom/compiler: abort 2-thread spill loops over the best result so far - vc4: save the fragment constant buffer around the YUV blit - vc4: unbind the textures around the blitter clears José Roberto de Souza (40): - anv: Change fill_inline_params() first parameter from struct GENX(COMPUTE_WALKER_BODY) to uint32_t * - anv: Move VMA heaps init and finish of vma heaps to anv_va.c - anv: Move init and finish of state pools to its own functions - anv: Move code to load color border to memory to a function - intel/brw: Explicitly upcast UB to UW for SHR with vector immediates - intel/tools: Fix parse of '[HWCTX].replay_*' in aubinator_error_decode_xe - intel: Sync xe_drm.h - intel: Add support for madvise purgeable VMAs in Xe KMD - iris: Improve and standardize the behavior of madvice in i915 - intel/brw: Fix nir_intrinsic_load_inline_data_intel register offset calculation - intel/dev: Remove unused intel_get_device_info_for_build() function - intel/dev: Add URB max entries values - intel/dev: Use URB mesh/task min/max values in intel_device_info - intel/dev: Add a Xe2+ table of URB min and max entries - anv: Add assert to make sure we don't push more than max_push_regs to push constants - anv: Replace most parameters of fill_inline_param() by a struct - anv: Add function to get each anv_state_pool - anv: Use anv_device_get_general_state_pool() - anv: Use anv_device_get_aux_tt_pool() - anv: Use anv_device_get_dynamic_state_pool() - anv: Use anv_device_get_binding_table_pool() - anv: Use anv_device_get_scratch_surface_state_pool() - anv: Use anv_device_get_internal_surface_state_pool() - anv: Use anv_device_get_bindless_surface_state_pool() - anv: Use anv_device_get_indirect_push_descriptor_pool() - anv: Use anv_device_get_push_descriptor_buffer_pool() - anv: Replace va.bindless_surface_state_pool access with a function - anv: Replace va.indirect_descriptor_pool access with a function - anv: Replace va.dynamic_visible_pool access with a function - anv: Replace va.internal_surface_state_pool access with a function - anv: Replace va.dynamic_state_pool access with a function - anv: Replace va.indirect_push_descriptor_pool access with a function - anv: Replace va.push_descriptor_buffer_pool access with a function - anv: Replace va.aux_tt_pool access with a function - anv: Replace va.binding_table_pool access with a function - anv: Replace va.scratch_surface_state_pool access with a function - anv: Fix memcpy overflows around sampler state - anv: Support sampler state of different sizes - anv: Replace anv_descriptor_set_binding_layout::descriptor_data_sampler_size by a local variable - anv: Fix copy of sampler state for bindless Juan A. Suarez Romero (36): - v3d: mark mapped BO as initialized for valgrind - v3d/ci: add OpenCL regressions - ci: igalia farm maintenance - Revert "ci: igalia farm maintenance" - broadcom/ci: update kernel for nightly runs - vc4/ci: update expected results - loader: check if the kernel driver is amdgpu - broadcom/simulator: V3D is always 4.2 or above - v3dv: allow device with only render node - v3d/drm-shim: add GPU selection - v3dv: disable threadeded submissions under drm-shim - broadcom/ci: update kernel for nightly jobs - Revert "ci: igalia farm maintenance" - mesa: allow GL_TEXTURE_COMPARE_{MODE,FUN} with EXT_shadow_samplers - v3dv: increase max push constants size - Revert "people: update Marek's email" - broadcom/ci: upgrade kernel in DuTs - v3d/ci: update expected results and document failures - v3dv: fix assertion on push constants - st/mesa: release sampler view - rusticl: fix leak in \`util_queue` - v3d: initialize value in query info - v3d: free vertex and constant buffers on context destroy - v3d: add more blitter ops for saving resources - vc4: mark mapped BO as initialized for valgrind - vc4: initialize value in query info - vc4: move util_copy_constant_buffer to the function beginning - vc4: add blitter operations - doc/features.txt: enable VK_KHR_shader_float16_int8 for v3dv - v3dv: fix buffer creation usage flags validation - doc/features.txt: fix VK_KHR_shader_float16_int8 for v3dv - v3dv: enable VK_KHR_shader_quad_control / clustered subgroups - v3dv: enable VK_KHR_shader_subgroup_rotate / rotate subgroups - v3dv: enable VK_KHR_shader_maximal_reconvergence - v3d: add support for array of textures blit with TFU - v3d: save fragment constants on sand8/sand30 blit Julia Zhang (15): - radv: add new option RADV_DEBUG=notmz - radv: allocate encrypted rings BOs - radv: create encrypted BOs for protected cmd_buffers - radv: enable surface protected capability - radv: set TMZ bit in sdma_copy packet - radv: save protected queue and non-protected queue seperately - radv: advertise VK_EXT_pipeline_protected_access - vulkan/wsi: return image compression properties for surface formats - vulkan/wsi: pass compression control when filtering DRM modifiers - vulkan/wsi: copy swapchain compression fixed-rate flags - radv: reject DCC modifiers when image compression is disabled - radv: advertise EXT_image_compression_control_swapchain - radv/ci: skip compression_control cases - radv: implement bo_wait_for_idle - radeonsi: avoid unmatched Perfetto events Julien Schueller (4): - glx: avoid crash on glXBindTexImageEXT when no texture target set - egl: fix _EGL_NATIVE_PLATFORM fallback for unrecognized native displays - drisw_glx: handle XGetGeometry failure in get_drawable_geometry - st/drawpixels: tile images larger than max texture size instead of clamping Kajal Kajal (1): - freedreno/blitter: copy full depth of src box in resource_copy_region Karmjit Mahil (34): - freedreno/decode,ir3: Mark decoded dwords as const - freedreno/decode: Fix error() in script.c - freedreno: Don't set UCHE_CLIENT_PF - gbm: Remove unused ARRAY_SIZE macro - gbm: Replace VER_MIN with common MIN2 - docs: Fix struct redefinition errors - util: Add heap_memory_percent driconf option - vulkan: Add heap budget helper function - hk: Add heap_memory_percent driconf support - asahi: Add heap_memory_percent driconf support - v3dv: Add heap_memory_percent driconf support - v3d: Add heap_memory_percent driconf support - tu: Add heap_memory_percent driconf support - freedreno: Add heap_memory_percent driconf support - panvk: Add heap_memory_percent driconf support - panfrost: Add heap_memory_percent driconf support - nvk: Add heap_memory_percent driconf support - crocus: Add heap_memory_percent driconf support - pvr: Add heap_memory_percent driconf support - vc4: Use os_get_gpu_heap_size() - freedreno/computerator: Remove VLA giving a build warning - util/u_trace: Fix copy_func indentation in generated code - util/u_trace: Fix indentation of generated _trace function - util/u_trace: Evaluate copy_func expression in python - util/u_trace: Refactor TracepointArgStruct - util/u_trace: Commonize emitted print functions - util/u_trace: Move copyright into a Mako def - util/u_trace: Commonize trace function header with a Mako def - util/u_trace: Avoid sprintf when we already have a string - util/u_trace: Add ArgBlob for storing structs in traces - util/u_trace: Add u_trace_backend_type - tu: Emit bin info in Perfetto render_pass - drm-shim/freedreno: Fix fprintf format specifier - android_stub: Replace __ANDROID_API_V__ with 35 Karol Herbst (195): - nak: the MS location comes last in TLD, same spot as depth compare in TEX - mesa/st: do not advertise CL subgroup features on the GL side - radeonsi: advertise support for subgroup rotate - iris: advertise support for subgroup rotate - nak/lower_cf: remove single src phis - nak: call nir_opt_fp_math_ctrl - nak: call nir_opt_algebraic_distribute_src_mods - ci: install libstdc++-static on fedora - rusticl: link the C++ runtime statically - softfloat: make sign bit an unsigned int - nir: add fmul_rtz - nir: handle fmul_rtz in a couple of places - nak: handle nir_op_fmul_rtz - nak: use fmul_rtz for NAK_INTERP_MODE_PERSPECTIVE - nir: add fmul_rtz optimizations - nir/lower_cl_images: call nir_progress on every function - gallivm/nir/soa: use uint for booleans - llvmpipe: never pass a NULL function name to LLVMAddFunction - nvk: Use nvk_cmd_fill_memory in CmdResetQueryPool when possible - include: update CL headers - rusticl/program: handle CL_INVALID_CONTEXT for clCompileProgram and clLinkProgram - rusticl/kernel: update error code handling for clSetKernelExecInfo - rusticl/kernel: return CL_INVALID_WORK_GROUP_SIZE in clEnqueueNDRangeKernel for an explicit 0 workgroup - rusticl: start implementing CL 3.1 support - rusticl: implement CL 3.1 platform features - rusticl: implement CL 3.1 device features - docs/features: add OpenCL 3.1 section - bin/gen_release_notes: add paragraph on OpenCL support - rusticl: update names of types now core in 3.1 - ci: update OpenCL 3.1 piglit fails - clc: do not use std::filesystem - Revert "rusticl: link the C++ runtime statically" - glsl/softfp: rename ffma to fmad - lima: rename ppir_op_ffma to ppir_op_fmad - nir: rename ffma to ffma_old - nir: rename nir_fmad to nir_fmad_old - nir: add new float multiply-add opcodes - nir: validate new float_mul_add options - nir/tests: handle new multadd opcodes - nir/opt_algebraic: add fmad and ffma_weak lowering rules - nir: handle new multadd opcodes in lowerings and opts - nir: handle new multadd opcodes in helpers - nir: duplicate old ffma opts where necessary for new multadd ones - nir/tests: use ffma_weak - ntt: use ffma_weak - llvmpipe: port over to ffma_weak - softpipe: keep weak_ffmas around - intel/elk: port over to nir_op_ffma - intel/jay: support nir_op_ffma - intel/brw: port over to nir_op_ffma - i915: support nir_op_fmad - nv50/ir: port over to new multadd opcodes - nv30: advertize new float multadd options - nak: port over to nir_op_ffma - zink: port over to nir_op_ffma_weak - ir3: port to nir_op_fmad - freedreno/ir2: use nir_op_fmad - tu: use nir_op_ffma_weak in lowering - ac: handle new float multadd opcodes - ac: use nir_op_ffma_weak - ac/llvm: support new multadd opcodes - aco: support new multadd opcodes - radv: use nir_op_ffma_weak - radeonsi: advertize new float multadd options - r600,sfn: support new multadd opcodes - r300: port over to nir_op_fmad - agx: port over to nir_op_ffma - kk: support nir_op_ffma - bitfrost: support nir_op_ffma - microsoft/compiler: support nir_op_ffma - d3d12: use nir_op_ffma_weak - etnaviv: port over to nir_op_fmad - pco: port over to nir_op_ffma - pvr: use ffma_weak for lowering - lima: support nir_op_fmad - svga: use weak_ffma - virgl: advertise new muladd options - nir: add fmad_or_ffma helpers and use it in lower_double_ops - nir: update ffma helpers to use new opcodes - nir: make lowering use new ffma opcodes - mesa: use ffma_weak - vulkan/meta: use nir_op_ffma_weak - tgsi_to_nir: translate MAD as ffma_weak - glsl: translate fma as fma_weak - vtn/glsl: translate fma as ffma_weak - vtn: handle OpFmaKHR - vtn: use ffma_weak - vtn/opencl: map mad to ffma_weak and fma to ffma - vtn_bindgen2: keep ffma_weak - nir: remove ffma_old - ci: update traces due to ffma rework - ci/windows: add dEQP-VK.glsl.builtin.precision_double.mix.compute.vec3 fail - zink: keep ffma_weak and use GLSLstd450Fma for it - zink: support nir_op_ffma - nir: add nir_intrinsic_cmat_load_shared_nv to nir_get_io_offset_src_number - nak/sm70: add helper for memory load store addresses - nak: wire up UGPR Ld/St/Atom encoding - nir: add uniform address to nvidia IO intrinsics - nak: add UGPR/GPR lowering for load/store/atom instructions - nak: optimize iadds with an uniform operand in iadds of address calculations - zink: proper advertise keep_weak_ffma for fp16 - rusticl/kernel: handle nir shader compilation failures gracefully - rusticl: more intel compat stuff - rusticl/spirv: add SPIRVToNirOptions type - rusticl/spirv: silence GenericPointer cap warning - rusticl/spirv: properly set float execution mode at spirv_to_nir time - gallium: add fp16_no_denorms cap - nvk: enable VK_KHR_shader_fma - nir/opt_algebraic: add missing fmadz lowering for lower_fmulz_with_abs_min - nir/opt_dead_write_vars: cache is_entrypoint of the function - nir/lower_alu: fix lower_fminmax_signed_zero for denorms - asahi: fix dst range in buffer copy region - asahi: fix compute blitter for float16 image copies - asahi: move batch flushing into agx_launch_internal - asahi: fix fdiv lowering - meson: enable more rust 2024 lints - rusticl/util: add Traits to help with usage of CString - rusticl/kernel: store kernel names as CString - rusticl/program: store log as a CString - rusticl/program: wrap compiler option parsing - rusticl/util: fix rustc-1.95 compilation error - vtn/opencl: convert libclc workaround handling to a switch statement - vtn/opencl: fix edge case behavior for cospi - vtn/opencl: fix edge case behavior for sinpi - vtn/opencl: fix edge case behavior for tanpi - rusticl/program: print compiler output as Rust string - rusticl/util: add CStrExt trait - rusticl/util: add CStrExt::from_ptr_or_empty - rusticl/program: add CompileOptions::get_clang_args - rusticl/program: construct __OPENCL_VERSION__ inside CompileOptions::get_clang_args - rusticl/program: turn iter map into loop inside CompileOptions::new - rusticl/program: handle -create-library inside CompileOptions::new - rusticl/program: move -cl-std handling inside CompileOptions::get_clang_args - rusticl/program: set __OPENCL_C_VERSION__ ourselves - rusticl/program: store build options as CString - rusticl/program: implement CL_PROGRAM_BUILD_OPTIONS without a copy - rusticl/program: implement CL_PROGRAM_BUILD_LOG without a copy - spirv: set num_components for OpAtomicFlagTestAndSet - rusticl/kernel: override libclc shader config helpers - nak/instr_sched_prepass: Take predicate spilling into account when scheduling instrucitons - nak: normalize lop3 constant sources - nak: convert base to iadd for non-uniform ldcx lowering - nak: run nir_opt_constant_folding after nak_nir_lower_load_store - nir/opt_phi_precision: bail on load_const conversions between float and ints - rusticl: move the worker queue into the Platform - Reapply "rusticl: fix leak in \`util_queue`" - rusticl/kernel: updated dim_threads in Kernel::suggest_local_size - rusticl/kernel: extract impl of suggest_local_size - rusticl/kernel: adjust grid at the end of suggest_local_size_impl - rusticl/kernel: add suggest_local_size tests - rusticl/kernel: add suggest_local_size_impl_gcd - rusticl/kernel: remove code to fill non full subgroups - rusticl/kernel: rework block size selection - gallium: remove PIPE_BARRIER_GLOBAL_BUFFER - rusticl/device: fix long vector_width queries on devices without int64 support - nak: implement shfr - nir: rework float compare late algebraic opts - nir: enable more opts for unordered and neo float compares - nak: implement and enable has_fneo_fcmpu - nak: implement ford and funord - nir/algebraic: pattern-match manual iadd64 - brw: advertise fp64 fma on hw with fp64 support - anv: enable VK_KHR_shader_fma - anv: fix wrong rebase conflict resolution from VK_KHR_shader_fma MR - nak/sm20: fix immediate encoding for F2I and F2F - nak/hw_tests: add F2I test for NaN behavior - nir: use function foreach helpers inside nir_cleanup_functions - nir: add nir_shader_fully_linked helper - rusticl: return Result instead of Option from convert_spirv_to_nir - rusticl: abort compilation if the nir shader is not fully linked - gallium: add pipe_caps::hw_clear_buffer_sizes - rusticl/util: implement Debug for CLVec - rusticl/kernel: convert Queue parameter to Device in launch - rusticl/kernel: add interface to launch kernel with arguments without binding them - rusticl/kernel: add offset to buffer bindings - rusticl/meta: add builtin kernel support - rusticl/meta: add builtin kernels for buffer fills - rusticl/mem: make Image::fill return a closure - rusticl/mem: use meta for clEnqueueFillBuffer and clEnqueueSVMMemFill - rusticl/mem: implement 1Dbuffer fills on top of a plain buffer fill - asahi: update agx_get_cl_cts_version for submission 471 - mesa_clc: support 32 bit targets - meson/rusticl: fix typo in depfile for builtin shaders - rusticl/meta: mark builtin kernels SPIR-V as 64 bit - rusticl/mesa: compile 32 bit version - rusticl/program: add Program::from_spirv_with_devs - rusticl/meta: support 32 bit devices - rusticl/meta: split out ulong kernels - rusticl/memory: return 0 for CL_IMAGE_SLICE_PITCH also for 2d images - rusticl/kernel: add libclc source hash to kernel shader keys - vtn/opencl: fix libclc needing fp16 lowering to fp32 - nouveau: Fix return of dangling pointer in nouveau_fence_new - clc: make libclc optional for configs not needing it - clc: use our downstream fork of libclc - rusticl: warn if we load not our own libclc fork Ken Cunningham (1): - llvmpipe: fix arch of LLVM JIT when cross compiling on Apple Ken Xue (1): - radv: remove checking on the gralloc handle->numFds Kenneth Graunke (113): - jay: Add missing ROR case - jay: Don't forget UACCUM! - iris: Implement force_dual_color_blend_by_location via NIR - iris: Call elk_nir_lower_fs_outputs for Gen8 RT reads, not brw - nir: Set FRAG_RESULT_DUAL_SRC_BLEND in outputs_written when lowering - brw: Switch FS outputs to semantic IO and FRAG_RESULT_DUAL_SRC_BLEND - brw: Set prog_data::dual_src_blend from NIR outputs written bitfield - brw: Drop dead code from dispatch limit check for dual source blending - brw: Limit SIMD width based on NIR rather than first backend compile - nir: Allow bias for nir_texop_sparse_residency_intel - nir: Lower SSBO helper writes too - intel/nir: Only add an explicit LOD 0 when lod/bias don't already exist - anv: Delete anv_instance::mesh_conv_prim_attrs_to_vert_attrs - anv: Use device->info.has_mesh_shading in key->mesh_input check - jay: Include depth and stencil on all MRT stores - jay: Add a TODO for coarse pixel shading - jay: Gripe more clearly about dual source blending - brw: Lower sample_pos for non-per-sample shaders in NIR - jay: Move render target store payload/descriptor construction to backend - jay: Implement fragment shader stencil writes - jay: Implement sample mask writes - jay: Add comments summarizing the PS thread payload layout - jay: Set Dispatch GRF Start Register in jay_setup_payload() - jay: Add a GPR_FROM_UGPRS opcode - jay: Implement sample position - jay, nir: Make a dispatch_mask_intel intrinsic - jay: Implement coverage mask - jay: Implement load_fs_config_intel - jay: Prohibit JAY_STRIDE_8 for EXPAND_QUAD - jay: Call constant folding before collecting FS outputs - jay: add a hack until we munge barycentrics dynamically - jay: Don't skip sampler payload copies for 2 or fewer sources - anv: Drop TES dispatch mode asserts - brw: Fix URB read length for tessellation evaluation shaders - brw: Ensure entire input load fits in push data - brw: Refactor urb_read_length setting for TES - brw: Fix mistake in brw_nir_lower_deferred_urb_writes - brw: Fold constants after nir_lower_io for VS/GS/TES outputs - jay: Add URB load support - jay: Fix scratch surface address save/restore - jay: Remember sp_delta_B when rematerializing stack pointer lane 0 - jay: Generalize EXTRACT_LAYER to take an arbitrary mask - jay: Handle facing that differs across subspans - jay: Implement viewport index FS input - jay: Fix null render target writes - jay: Drop render target stores with unconditional discards - jay: Implement dual color blending (but require SIMD16) - jay: Don't swap FS interpolation .yz deltas - jay: Implement fragment shader barycentrics - jay: Pass proper simd_width to brw_nir_apply_key for fragment shaders - jay: Add tessellation evaluation shader support - jay: Add an INTEL_JAY=all option - jay: Unroll loops before lowering deferred URB writes - jay: Assert FS input deltas exist - jay: Fix hard coded number of FS inputs - jay: Ignore RT store condition if there are no outputs - jay: Improve unconditional discard removal - jay: Fix rewrite_without_flags for SEL with other flag sources - anv: Fix shader stats when using jay for non-compute stages - jay: Still predicate Null RT store if everything is discarded - jay: Implement load_subgroup_size - intel/nir: Improve address reuse in brw_nir_lower_immediate_offsets - intel/nir: Turn load_global_constant into load_global_intel too - jay: Call intel_nir_lower_shading_rate_output earlier - jay: implement load_frag_shading_rate - jay: Store the FS config def - jay: Store a test of the dynamic "is coarse?" FS config bit. - jay: Set prog_data->uses_fs_config when coarse pixel shading is dynamic - jay: Set the coarse pixel render target descriptor bit - jay: Don't run the entire optimization loop before prog data - jay: Use nir_lower_frag_coord_to_pixel_coord - jay: Lower to pixel_coord_intel and frag_coord_w_rcp after prog data - jay: fix frag coord .z lowering with coarse pixel shading - jay: Allow BFN on U16 types - jay: Make a builder local in setup_fragment_payload - jay: Implement coarse pixel coordinate calculations - jay: Fix stack smashing with more than 16 FS inputs - jay: Rewrite FS output gathering - jay: Run brw_nir_lower_alpha_to_coverage earlier - jay: Optimize out noop samplemask writes - jay: Add missing HF conversion stride restrictions - jay: Tighten mixed stride restrictions - jay: Make a jay_clobbers_address_reg() helper - jay: Add a new VECTOR_EXTRACT opcode for indirect moves - jay: Implement indirect push constant loads for 32-bit - jay: Implement indirect push constant loads for 8/16-bit sizes - jay: Move brw_nir_apply_key call to be shared among stages - nir: Early out in nir_opt_shrink_vectors if all components are read - nir: Don't shrink intrinsics and undefs to vec5s - jay: Speed up shuffles and vector extracts with uniform offsets - nir: Add an option for whether TCS invocation_id should be uniform - intel/compiler: Set nir_divergence_tcs_invocation_id_uniform - jay: Emit a noop URB write for EOT if there isn't one to reuse - jay: Increase JAY_NUM_LAST_USE_BITS to 64 - jay: Implement tessellation control shaders - jay: Use LOOP_ONCE if a loop ends in HALT too, not just BREAK - brw: Fix GS EOTs to not have an empty channel mask on LSC platforms - brw: Update comment that's so old it makes no sense - brw: Inline brw_do_emit_fb_writes - brw: Assert that repclears aren't used on Gfx12+ - brw: Drop brw_compile_fs_params::allow_spilling - jay: Use INTEL_SIMD_DEBUG=cs for compute shaders, not fs - intel: Drop INTEL_DEBUG=no{8,16,32} flags - intel: Fix multipolygon flags in INTEL_SIMD_DEBUG default handling - intel: Refactor SIMD selection's debug flag handling - intel: Replace INTEL_DEBUG=do32 with INTEL_SIMD_DEBUG - brw: Switch to INTEL_SIMD_FORCE for multipolygon modes - brw: Respect subgroup size requirements even with INTEL_SIMD_DEBUG - brw: Allow spilling and other poor decisions when using INTEL_SIMD_DEBUG - brw: Fix INTEL_SIMD_DEBUG=fs32 to work at all - brw: Rework FS SIMD selection to follow requirements over debug flags - jay: Add u16 and f16 support to CSEL - intel: Temporarily disable madvise on iris on xe.ko Koch, Pawel (1): - Update docs regarding anv shader dumps Reviewed-by: Lionel Landwerlin Konstantin (3): - vulkan/cmd_queue: Handle struct copies that are not pointers - lavapipe: Re-emit push constants if the size changed - lavapipe: Fix push_constant_size for shader objects Konstantin Seurer (64): - vulkan/radix_sort: Add support for 96-bit keys - vulkan: Rename radix_sort to radix_sort_u64 - vulkan: Rename key_id_pair to key32_id_pair - vulkan: Implement 64-bit morton codes - radv/rt: Use 64-bit keys for gfx11- - util/u_trace: Add an option to emit additional code - util/u_trace: Rework resource management - util/u_trace: Print tracepoints with indentation - vulkan: Fixes for a spec update - vulkan,spirv: Update spec to 1.4.352 - radv: Move debug options to radv_instance.h - radv: Move a whole bunch of debug/profiling related into a subdir - radv/tools: Rename radv_debug to radv_debug_hang - radv: Move radv_find_memory_index to radv_debug.c - radv: Add and use helpers for managing internal allocations - llvmpipe: Use i1 for sparce residency and expand as needed - llvmpipe: Implement sparse residency feedback for buffers - llvmpipe: Fix sparse binding large areas - lavapipe: Bump maxRayDispatchInvocationCount to the min requirement - nir: Duplicate the name in nir_def_set_name - llvmpipe: Remove lp_llvm_descriptor_base - lavapipe: Reduce descriptor sizes even further - lavapipe: Implement VK_KHR_shader_untyped_pointers - lavapipe: Add lvp_nir_lower_push_constants - llvmpipe: Fix memory leak when allocating sample functions - lavapipe: Perform shader object compatibility check early - lavapipe: Ignore src_plane for samplers - tools: Update imgui to the docking branch and add backends - meson: Add some include directories - vulkan/bvh: Add defines for acceleration structure types - radv: Use a separate BLAS pointer copy pass for BVH4 - radv: Rename copy_blas_addrs to copy_addrs - radv: Rename serialization fields in radv_accel_struct_header - util: Add RTI file format definitions - radv/tools: Add RTI file dumping - rti: Initial commit - vulkan: Handle arbitrary build flag counts in vk_build_stage - vulkan: Make vk_build_stage non-static - vulkan: Filter for updates in vk_build_stage - vulkan: Move build_flags to vk_build_config - vulkan: Add vk_accel_struct_cmd_begin_debug_marker - vulkan: Use vk_build_stage for encode/update passes - vulkan: Shuffle around bvh build code - meson: Add mesa_python_path - vulkan: Move capture_key_pressed to vk_device - util/u_trace: Add an option for accumulating tracepoint ranges - util/u_trace: Do not generate empty structs - radv: Add u_trace support - util/u_trace: Include payload in the range accumulation key - radv: Include build_flags in the range key - util/u_trace: Release memory for reused timestamps - radv: Ignore entrypoints inside meta OPs for utrace - radv,anv: Enable BVH updates - tool: Rename RTI to gamma - radv: Fix generating ray history code - gamma: Set the window title to gamma - u_trace: Initialize fuzzy_* callbacks correctly - vulkan: Fix ROOT_FLAGS_OFFSET_ID offset - radv: Store root_flags for BVH8 - radv: Use 64bit keys on GFX12 - gamma: Do not render minimized viewports - gamma/radv: Display the ray launch ID - radv: Delay lowering printf - radv/bvh: Fix updating acceleration structures containing AABBs Kovac, Krunoslav (3): - amd/vpelib: fix custom color space handling - amd/vpelib: Enable VPE cap for 3DLUT - amd/vpelib: Fixes for external lut compound Lakshman Chandu Kondreddy (3): - zink: Query external memory handle type compatibility - freedreno: Add support for A704 - zink: Set can_do_invalid_linear_modifier workaround for QCOM blob driver Lars-Ivar Hesselberg Simonsen (9): - pan/genxml: Print shader hex in trace for Valhall - pan/va/disasm: Print 64 bit src/dest regs as reg pairs - pan/va/disasm: Align FAU printing - pan/va/disasm: Align indentation - panvk: Fix debug flag overlap - panvk/v10+: Align allocations >= 64k to 64k - panvk: Ensure 64k alignment for sparse images - pan/format: Prefer 16X16_BLOCK_U over INTERLEAVED_64K - panvk/v10+: Fix size gt -> gte for 64k alignment Leandro Dorileo (1): - intel/executor: inform oa not available if that's the case Leder, Brendan Steve (Brendan) (1): - radeonsi/vpe: Update DCC API and programming Lei Huang (1): - amd/virtio: enable Android amdgpu-virtio build option Leon Perianu (1): - pvr: enable VK_EXT_device_memory_report Lin, Ricky (1): - amd/vpelib: DPM detect first frame action LingMan (6): - rusticl: Drop custom \`addr` implementation - mr-label-maker: Add \`Rust` label for \`src/compiler/rust` - mr-label-maker: Drop rule applying \`Rust` to all .rs files - mr-label-maker: Apply \`Rust` label to \`clippy.toml` and \`build-rust.sh` - mr-label-maker: Apply \`Rust` label to \`rustfmt.toml` - mr-label-maker: Apply \`Rust` label whenever crate dependencies are changed Lionel Landwerlin (160): - intel/dev: fixup intel_needs_workaround() macro - anv: avoid C23 - anv: fix compute push constant allocations on pre Gfx12.5 platforms - anv: fix invalid value for push block index - anv: fix debug printfs on hang - anv: fixup compute queue detection - anv: rework debug flag - anv: switch from INTEL_DEBUG to ANV_DEBUG for shader-print - anv: remove unused defines - anv: fix relocations into internal shaders - anv: simplify inline uniform descriptor loads - nir: expose nir_opt_dce_impl - anv: run a single impl loop for apply_pipeline_layout - anv/apply_layout: move some helpers around - ci/zink/intel: disable TGL demo-v2 trace - brw: track push constants shader stats - anv: promote push constant pointers to push buffers - anv: add a pass to realign global loads on DX CBV resources - intel/ci: update expectation for RPL - vulkan: add tracking for VK_EXT_primitive_restart_index - anv: implement VK_EXT_primitive_restart_index - anv: expose VK_KHR_shader_constant_data - anv: fix null pointer access - anv: stop using queue priority KHR aliases - anv: remove a bunch of KHR alias uses - anv/docs: update environment variable docs - anv: reorder debug options - anv: add a shader-dump debug option - imgui: update copy and port all tools using it - intel/tools: add eu stall viewer - brw/lower_texel_address: add heap support - anv: split sampler state packing from API object creation - brw: add heap support to brw_lower_storage_image - intel: add resource intrinsic support for heaps - anv: add lowering of descriptor heap intrinsics - anv: implement EXT_descriptor_heap entry points - anv: add descriptor heap binding support - docs: document ANV_DEBUG=desc-dirty - anv: enable EXT_descriptor_heap - anv: fix arc artifacts on Farming simulator 2022 - anv: print out the content of the printf buffer at vkDestroyDevice - anv/brw/nir: fix wa_18019110168 - anv: expose non binding-table/push-pointer flushing - anv: expose RT state flushing - anv: enable compute state flushing with indirect state - anv: move a bunch of structures to anv_types.h - anv/intel: add device generated commands shaders - anv/apply_layout: use the resource index to compute descriptor buffer addresses - anv: add apply_layout support for device bindable shaders/pipelines - anv: add a helper to flush the descriptors for indirect compute execution - anv: program relative push set offset for descriptor buffers device bindable shaders - vulkan: add pipeline helper to retrieve scratch-size/ray-queries - anv: add support for indirect execution set - anv: add indirect command layout support - anv: add unspecified internal kernel send count support - anv: allow simple shader spilling for complex ones - anv: enable generation shader calls - anv: handle descriptor binding with DGC - anv: implement generated preprocess & execute - anv: add barrier flags handling for preprocess buffers - anv: handle preprocess buffer creation on <= Gfx12.0 - anv: track generated commands work with perfetto - anv: expose VK_EXT_device_generated_commands by default on Gfx12.5+ - anv: add a device generated command debug option - anv: add Gfx9 support VK_EXT_device_generated_commands - anv: expose VK_KHR_maintenance11 - anv: group all performance drirc together - anv/iris: stop using 3DSTATE_PUSH_CONSTANT_ALLOC_PS on Gfx12.5 - anv: rename push constant allocation helper - anv: add an option to disable push constant space reallocation - vulkan/runtime: fix invalid address flags value for CmdCopyBufferToImage2 - anv: fixup null address check - anv: implement VK_KHR_device_address_commands - anv: remove old entrypoints - anv: implement missing device image property compression filtering - vulkan/wsi: write VkImageCompressionControlEXT from swapchain to image creation - anv: enable VK_EXT_swapchain_compression_control when possible - blorp: stop requesting the fp64 shader for ELK - blorp: only request fp64 shader on when required - anv: sweep the NIR fp64 shader before keeping it on the device - anv: only load fp64 software shader when needed - anv: add an option to disable allocation over subscription - brw: simplify VF component packing code - anv: add SIMD32 requirement heuristic for Dragon Dogma 2 - brw/jay: move some coarse lowering to NIR - brw/jay: move sample_mask_in handling to NIR - docs/features: updates for Anv - anv: temporarily reenable scratch page by default - anv: bump max compute workgroup count - anv: further optimize dirty state after secondary emission - anv: only reprogram line-stipple if enabled - util: add a script to auto-generate a drirc infrascture per driver - util/drirc_gen: enable validation for a specific driver - hasvk: rename a couple of drirc options - hasvk: add a driver section for drirc - drirc: remove non Anv option in the Anv section - anv: use the new generation script for drirc - anv: fix missing bindless flag hashing - anv: fix render target remapping tracking at the beginning of render passes - brw: avoid requiring a valid render target for empty fragment shaders - spirv: fixup infinite recursion with shader replacement - anv: use shader source hash rather than cmd_buffer fields - intel: switch shader hash to 64bit value - mi_builder: mi_umax2 tests - anv: rename drirc script - anv: move fake_sparse drirc to feature category - anv: move compression control drirc to feature section - anv: fake VK_EXT_image_compression_control on Xe2+ - iris: only call brw_nir_fs_needs_null_rt() with no render targets - brw: fix null render target decision - anv: fix assert/crash in import of compressed local memory on xe2+ - anv: align storage texel buffer support on image support - anv: add missing condition to update 3DSTATE_RASTER - anv: fix 3DSTATE_SF line width programming with Bresenham lines - anv: don't forget dataport flush for ANV_DEBUG=dgc-dump - elk: assert always/never on some of the FS config flags - brw: remove always true condition - brw: remove interpolator coarse bit setting - brw/jay: track usage of fs_config by backend - brw: only check for shader_info::fs.uses_sample_shading - anv/brw/jay: de-dynamify per-sample interpolation - anv: hash binding tables for EXT_descriptor_heap too - brw: add shader key to enable robust SLM accesses - anv: add hitman2 workaround for SLM load vectorization - anv: fix descriptor heap indexing of YCbCr embedded samplers - vulkan/runtime: fixup group building with shaders from libraries - spirv: add parsing of vkd3d-proton shader hashes - drirc: add a callback mechanism do deal with shader hash & options - anv: enable VK_EXT_descriptor_heap by default - vulkan: condition cmd_queue initialization to driver need - anv: fix push constant address emission for gfx commands - anv: add missing handling of push pointers in gfx dgc - anv: use vkd3d-proton provided shader hashes if available - anv: add infrastructure to deal with missing barriers in applications - anv/brw: fixup 64bit array image accesses - anv: fix 64bit image atomic emulation with EXT_descriptor_heap - anv: more dgc push constant fix - anv: fix push pointer optimization with DGC - anv: add memory heap budget tracking across VkInstance - anv/brw: limit push constant promotion in vertex shaders - anv: introduce an option to disable disk cache - anv: fixup RT building barrier - anv: fix compression control reporting on xe2+ - anv/ci: turn on astc emulation testing - anv: fill min_array_element with indirect descriptors - anv: fixup max push data delivered to shaders - anv: add workaround for atomics on R11G11B10 images - anv: remove previous Horizon Forbidden West workaround - anv: flush accumulated barriers for top of TOP_OF_PIPE - anv: fix Wa_18040903259 - vulkan/runtime: fixup vk_shader leak on RT group recompile - anv: fix push buffer descriptor address relocation - brw: add missing INTEL_FS_CONFIG_PER_PRIMITIVE_REMAPPING handling - brw: fix wa_18019110168 lowering - anv: fix leak in RT binding point - anv: fix barrier for Wa_1508744258 / Wa_14024015672 - iris: fix barrier for Wa_1508744258 / Wa_14024015672 - jay: copy resource_intel surface handle value - anv: fixup the logic dealing with STATE_BYTE_STRIDE - anv: only consider active view-capable queues for image views Lishin (3): - mesa/st: relax shader_has_one_variant checks for GLES2 - broadcom/qpu: add V3D 7.1 disasm tests - v3d/v3dv: use common compute limits Liu, Mengyang (2): - aco: fix broken VGPRs reservation for 64-bit attributes in VS prologs - amd: disable reset_filter_cam for mec Lone_Wolf (2): - ac/llvm: fix build with LLVM 23 (MCSubtargetInfo) - clc: fix build with LLVM23 (TargetRegistry::lookupTarget) Lorenzo Rossi (104): - nir: Extract float_is_half tests in common code - nir/opt_algebraic: optimize fadd/fmul with 16-bit source and constant - pan/compiler: Allow 16-bit alpha for atest_pan - pan/compiler: Fix WaRaR hazard in pressure scheduler - pan/compiler: Lower unaligned scratch memory accesses - pan/compiler: Handle ssbo_atomics in lower_vs_atomics - nir/lower_point_size: Handle 16-bit point sizes - nir/opt_sink: Add pan-specific load_input - panfrost: Constant-fold io locations after lowering - pan/compiler: Sort preprocess - panvk/jm: Fix tls_size overwrite in indirect draws - pan/compiler: Rework scratch memory strategy - panvk,panfrost: Pass inputs and info to postprocess - panvk: Remove pan_optimize_nir call - pan/compiler: Rename bifrost_optimize_nir - pan/compiler: Sort postprocess - pan/compiler: Collect nopersp varyings in lower_noperspective_fs - pan/compiler: Add better documentation for second lower_int64 - pan/compiler/lower_fs_inputs: Do not trust slot->alu_type - panfrost: Split default key creation in helper function - panfrost: Plumb VS varying_layout in FS - pan/bi: Vectorize f2f16 on v10 and earlier - pan/bi: Switch old license texts to SPDX - pan/mid/fuse_io_cvt: Disable fusion on highp - pan/bi: Add nir_fuse_io pass - pan/bi: Add a printing helper for pan_varying_layout - pan/valhall: fuse_cmp skip when fusing the same instruction - pan/bifrost: Make CSE independent of liveliness labels - nir/opt_algebraic: Optimize mediump fadd/fmul done in highp - pan/bifrost: Fix 16-bit demote_if - kraid: Fix out-of-tree build issue - kraid/tests: Edit meson to help rust-analyzer provide IDE suggetsions - panfrost: Separate the compiler from libpanfrost - panfrost: Reorder meson definitions - compiler/rust/lower_bounded: Add FromIterator impl - kraid/swizzle: Add fold_u64 - kraid/ir: Add SrcMod::fold_u64 - kraid/ir: Add FauRef UserPage creation utilities - kraid: Add alloc_vec utility - kraid: Fix FauRef Display bug - kraid/ops: Add a small crate documentation for conventions - kraid: Add OpIMul - kraid/model: Ensure dyn Model is Send + Sync - kraid: Add hw_runner - kraid: Add basic hw_tests - kraid: Add Foldable and initial tests for OpShiftLop - kraid/hw_runner: Unmap buffers on drop - kraid: Sort opcodes and keep them sorted - kraid: Replace alloc_vec with alloc_ref - compiler/rust/float16: Implement total_cmp as present in f32 - kraid/encode_v9: Implement accumulator ops for FCmp - kraid/ops: Fix Display for OpFCmp - kraid: Add tests for OpFCmp - kraid: Add Foldable impl for OpCSel - kraid/encode_v9: Implement accumulator ops in ICmp - kraid: Add tests for OpICmp - kraid: Add tests for OpIAdd - kraid: Add tests for OpIMul - kraid: Add OpISub with tests - kraid: Add OpClz with tests - kraid: Add OpIToF32 - kraid/nir: Add IMul - kraid: Add ufind_msb - kraid: Add ShaderInfo - kraid: Add register preloading - kraid: Add load_push_constant - kraid/hw_tests: Use preloaded registers - pan,nir: Add Panfrost image intrinsics - pan/bi: Lower image load/store/lea in NIR and fix OOB access - pan/bi: Remove unused backend lowering - panfrost/model: Add var,cvt,sfu rates for Valhall architectures - panfrost/compiler: Properly compute Valhall ALU bound - drm-shim.py: Add more panfrost models - panfrost/drm-shim: Fix assertion on v12+ - kraid: Legalize immediates - kraid: Add OpIAbs and plumb it through - kraid/nir: Fix nir_op_extract*16 - kraid: Add BitRev and wire it up - kraid: Add OpPopCount and wire it up - kraid: Add support for clamp and round to OpFAdd - pan/bi: Add kraid-specific algebraic rules - kraid: Add F32ToI32 and wire it up - kraid/algebraic: Lower b2i conversions - kraid/algebraic: Lower nir_op_pack_uvecX_to_uint - kraid: Wire up fneg - kraid: Add OpFma and wire it up - kraid: Add OpFlush - kraid: Add FRound and wire it up - kraid: Add Frexp and wire it up - kraid: Add OpFMin/OpFMax and wire them up - kraid: Add OpLdExp and wire it up - kraid: Add a pass macro for validation and debug - kraid: dead-code elimination pass - pan/bi: Fix f2f16(a\@16) in shader-db run on v13 - pan/bi: Move pan_nir_fuse_io in bi_optimize_late - pan/nir_fuse_io_cvt: Add texture cvt fusion - pan/compiler: Don't widen unaligned push-constants too much - panfrost: Unify FAU constants and relocation handling - panfrost: Promote constants to FAU - panfrost/midgard: Fix fau max not initialized - panfrost: Fix wrong layout reuse in user clip planes - pan/nir: Fix header static inline function - pan/compiler/stats: Fix ALU not being used in instruction bounds - pan/bi: Propagate swizzle in bi_optimizer_result_type Louis Montagne (2): - zink: relax build-id length assertion for Mach-O - meson: allow DRI on darwin to enable Zink + EGL builds Loïc Molinari (17): - pan/crc: Restrict CRC buffer creation to 1st RT mipmap level - pan/crc: Introduce pan_fb_info_is_fully_covered() - pan/crc: Check AFBC renderblock size on v5 and v6 too - pan/crc: Check CRC buffer validity and coverage on v5 and v6 too - pan/crc: Simplify CRC buffer selection logic - pan/crc: Use RT selection loop in single RT case - pan/crc: Check CRC requirements in dedicated function - pan/crc: Disallow CRC on sparse AFBC images - pan/crc: Simplify CRC buffer initialization - pan/crc: Cache temporary CRC info - pan/crc: allow setting a NULL pointer to the CRC validity state - pan/crc: Store CRC state in a struct - pan/crc: Enable CRC for multiple RTs on v6 - pan/crc: Disable CRC on v4 - pan/crc: Enable Empty Tile Elimination - pan/crc: Optimize clear color hashing - panfrost/ci: Mark "spec\@!opengl 1.4\@copy-pixels" as flake Lucas Francisco Fryzek (1): - util/u_trace: Don't use empty initializer list Lucas Fryzek (1): - Modify x11_xcb_display_supports_xshm to get xshm opcode Lucas Stach (4): - etnaviv: clean up index buffer handling code a bit - etnaviv: move index buffer handling in draw_vbo after derived state handling - etnaviv: reserve state emission space early in draw_vbo - etnaviv: move sampler source update before draw space reservation Luigi Santivetti (3): - pvr: de-dup strncmp in pvrsrvkm winsys - pvr: add missing multi-arch support for pipeline exec and stats - pvr: re-use texture state words for each load op Lukas Zapolskas (2): - pan/pps: Move PanfrostDevice to a separate file - pps: Add the Primitive, Instruction, Pixel and Fragment unit types Maaz Mombasawala (3): - Revert "ci: vmware farm is offline, stop using it" - Revert "ci-farms/vmware: Disable vmware tests for now" - svga: Update CI expectations. Marc Alcala Prieto (41): - pan/genxml: Add performance-trilinear enum values - pan/genxml: Add missing enum values on v9-v13 - pan/genxml: Add v14 definition - pan/genxml: Implement RUN_FRAGMENT2 - pan/decode: Remove progress-related decoding logic - pan/genxml: Build libpanfrost_decode for v14 - pan/clc: Build for v14 - pan/fb: Implement pan_emit_fb_desc for v14+ - pan/desc: Implement pan_emit_fbd for v14+ - pan/texture: Add v14+ YUV pipe format mappings - pan/format: Add v14+ YUV pipe format mappings - pan/afbc: Add v14+ AFBC YUV compression mappings - pan/afrc: Add v14+ AFRC YUV compression mappings - pan/lib: Build for v14 - panvk: Implement RUN_FRAGMENT2 - panvk: Handle provoking vertex and simultaneous reuse on v14 - panvk: Build for v14 - pan: Add v14 support - pan/va: Fix packing test for LdVarBufImmF16 on v11 - pan/bi,va: Use dedicated LD_VAR_BUF_FLAT* opcodes on v14+ - panfrost: Implement RUN_FRAGMENT2 on the Gallium driver - panfrost: Build the Gallium driver for v14 - panfrost: Advertize Mali-G1-Pro support - docs/panfrost: Advertize Mali-G1-Pro support - pan/decode: Support INTERLEAVED_64K Z/S target dumps - pan: Layer offset is not longer available starting on v14 - panvk/csf: Allow 256 layers per tiler descriptor on v14+ - panfrost: Advertise Mali-G1-Premium and Mali-G1-Ultra support - panfrost: Remove duplicated flushes before RUN_FRAGMENT[2] - panvk/csf: Emit fragment layer state just before RUN_FRAGMENT2 - panvk/csf: Implement incremental rendering on v14+ - pan/csf: Fix incremental rendering on v14+ - pan/bi: Load vertex view index from preload on v14+ - panvk: Fix multiview support on v14+ - pan: Add helper for max multiview view count and rise it to 16 on v14+ - pan/compiler: Rename multiview to per_view_outputs - pan/va: Fix serialization of atomic operations using BI_ATOM_OPC_AUMIN - pan/va: Unit test BI_ATOM_OPC_AUMIN - pan/ci: Remove GLES shader image load/store atomic flake - panvk: Fix DRM format modifiers for multi-planar YUV formats - panvk/csf: Avoid poisoning read-only fragment SRs Marek Olšák (167): - nir: add back color0/1 system values and VARYING_SLOT_PARAM_GEN_AMD - ac/nir: add ac_nir_get_io_driver_location as replacement for IO bases - ac,radeonsi: don't use nir_intrinsic_base for FS outputs - radeonsi: don't recompute IO bases for FS outputs - radeonsi: stop setting si_shader_info::output_semantic for FS - radeonsi: stop using si_shader_info::output_semantic for passthrough TCS - radeonsi: stop using output_semantic[] for LS outputs passed via VGPRs - radeonsi: remove si_shader_info::output_semantic[] - radeonsi: remove si_shader_info::num_outputs - ac,radeonsi: stop using nir_intrinsic_base for TCS inputs passed via VGPRs - ac/llvm: correctly load 16-bit TCS inputs from VGPRs and simplify - ac/llvm: reorder/remove variables in visit_load_input - radeonsi: update shader info in si_nir_lower_color_flatshade_twoside - ac,radv,radeonsi: don't use nir_intrinsic_base for FS inputs - radv: remove radv_recompute_fs_input_bases - radeonsi: compute si_shader_info::color_attr_index without input_semantic[] - radeonsi: compute si_shader_info::inputs_read without input_semantic[] - radeonsi: remove si_shader_info::input_semantic[] - radeonsi: don't call nir_recompute_io_bases for FS - radeonsi: set num_vs_inputs from nir->num_inputs and use it more - radeonsi: just get si_shader_info::num_inputs from NIR - amd: remove unnecessary and transitive #includes - ac/nir: add ac_nir_assign_fs_input_locations to set PS input locations in stone - nir/opt_licm: add a private state structure for the pass - nir/opt_licm: use nir_metadata_control_flow - nir/opt_licm: hoist instructions across multiple levels of nested loops - radeonsi/ci: remove the fixed XFB test from fails/flakes - radeonsi/ci/build: also fetch video decode/encode sample for VK CTS - nir/opt_dce: factor out dead instruction removal into a helper - nir/opt_dce: add shader_info::assert_inputs_not_dead - ac/nir: factor out ac_nir_lower_tex_coords from ac_nir_lower_image_tex - ac/nir/lower_tex_coords: move input loads instead of cloning them - aco/tests: update ACO tests for ac_nir_lower_tex_coords refactoring - glsl,gallium: add pipe_caps::glsl_bindless_handles_are_32bit - radeonsi: set glsl_bindless_handles_are_32bit - nir: add frag_coord_xy - nir/lower_wpos_ytransform: handle frag_coord_xy - nir/opt_frag_coord_to_pixel_coord: handle frag_coord_xy - nir: add direct lowered frag_coord building to replace lowering passes - nir: use nir_build_frag_coord everywhere - amd: add a tool that prints tiling layouts for all shim devices - winsys/amdgpu: revert invalid changes from CS functions - winsys/amdgpu: fix memory leaks when amdgpu_cs_create fails - amd/tools: rewrite ac_print_tiling_layouts to print all layouts, including XORs - nir/opt_licm: add filter callback - nir/tests: add nir_opt_licm tests - nir/licm: allow speculative hoisting across terminate if the filter is set - nir: add an option to ignore INTERP_MODE_NONE in nir_shader_gather_info - radeonsi: fix a typo in si_shader_update_spi_shader_formats - ac,radeonsi: add a helper to print PS input VGPR layout - ac,radeonsi: add helpers to print SPI_SHADER_COL/Z_FORMAT - aco,radeonsi: use enums for color barycentrics instead of input VGPR indices - radeonsi: remove dead get_frag_coord_from_pixel_coord optimization - radeonsi: use shader_info::fs::uses_sample_shading for ac_nir_lower_ps_early - ac: add ac_shader_args::line_stipple_tex_ena - radeonsi: move SI_SPI_PS_INPUT_ADDR_FOR_PROLOG into a helper function - aco,radeonsi: don't forward LINE_STIPPLE_TEX_ENA VGPR from the PS prolog - aco,radeonsi: declare prolog CENTROID VGPRs only if used - radeonsi: declare prolog ANCILLARY & SAMPLE_COVERAGE VGPRs only if used - radeonsi: declare prolog LINE_STIPPLE_TEX_ENA VGPR only if needed - radeonsi: declare prolog LINEAR_SAMPLE/CENTER VGPRs only if used - radeonsi: simplify get_interp_info_from_input_load - radeonsi: stop using TGSI definitions for interpolation - radeonsi: handle any size of shader args in the LLVM PS prolog - nir/opt_move_to_top: add an option to exclude moving at_offset/at_sample loads - nir: generalize nir_vertex_divergence_analysis -> nir_custom_divergence_analysis - nir/opt_varyings: use workgroup divergence to identify convergent mesh outputs - util/set: add helper _mesa_set_equal - nir/opt_varyings: rewrite elimination of duplicated outputs - nir/tests: don't leave "namespace {" unclosed in nir_opt_varyings_tests.h - nir/tests: test new output deduplication cases - nir/opt_varyings: always report progress when calling nir_remove_varying - nir/tests: use ASSERT_EQ instead of ASSERT_TRUE in nir_opt_varyings tests - nir: add missing SYSTEM_VALUE_FRAG_COORD_W_RCP - nir: handle load_frag_coord_w_rcp in multiple passes, same as non-rcp - nir: change nir_frag_coord_form options to a bitmask - nir: add nir_frag_coord_use_pixel_coord for OpenGL - radv: switch to nir_frag_coord_xy_z_w_separate with w_rcp - radeonsi: switch to nir_frag_coord_xy_z_w_separate with w_rcp - radeonsi: enable nir_frag_coord_use_pixel_coord - radeonsi: don't treat sample_pos as using frag_coord - ac,radeonsi: remove all frag_coord_xy code - nir/opt_frag_coord_to_pixel_coord: factor out helper nir_all_uses_of_float_are_integer - radv: add a pass that selects either frag_coord_xy or pixel_coord, but not both - radv: remove dead load_sample_pos code - radv: move SPI_PS_INPUT_ENA emission into radv_emit_ps_state - radv: select frag_coord_xy and pixel_coord conditionally based on dynamic state - nir/opt_algebraic: add more ffract/ffloor/ftrunc/f2u/f2i patterns - radeonsi/tests: add an ordered append bandwidth test - radv: ignore color attachment samples for ps_iter_samples - ac/nir/lower_ps_early: remove obsolete comment - ac/nir/lower_ps_early: assume frag_coord_is_center is always true - ac/nir: add a new pass ac_nir_lower_sample_mask_in - radv: switch to ac_nir_lower_sample_mask_in - radv: enable SAMPLE_COVERAGE PS VGPR dynamically - radv: fix an inefficiency where the ANCILLARY PS VGPR was enabled but unused - radeonsi: use ac_nir_lower_sample_mask_in - ac/nir/lower_ps_early: remove now-unused lowering of sample_mask_in - radv: make RAST_SAMPLES_STATE dirty in CmdBeginRendering only on gfx12+ - radv: emit_rast_samples_state uses uses_vrs_attachment only on gfx11+ - radv: don't leave SPI_PS_INPUT_ENA uninitialized with NULL PS to fix a hang - ac/surface: print the modifier in ac_surface_print_info - ac: add basic HTILE dword printing - radv: bump the sparse alignment requirement to 64K - radv: fix VK_MEMORY_PROPERTY_DEVICE_COHERENT_BIT_AMD with sparse buffers - radv,radeonsi: disallow VRS flat shading if SubgroupInvocationID is used - radv: rename vrs_coarse_shading -> vrs_flat_shading - nir/opt_idiv_const: a / uint_max -> b2i(a == uint_max) - radv: stop using set_sh_reg_idx(3) to reduce CP overhead - radv: fix setting COMPUTE_DISPATCH_INTERLEAVE on the gfx queue - radeonsi: remove unnecessary and indirect #includes - radeonsi/ci: allow glcts to be in the cts directory - radeonsi/ci: change DEQP_TARGET to default for Wayland - ac,radeonsi: remove uses_kernel_cu_mask and associated code - nir/opt_varyings: rewrite indirect IO tracking and dead IO elimination - nir/opt_varyings: split tidy_up_indirect_varyings - nir/opt_varyings: shrink pathological varying arrays to 1 element - nir: fix ibfe handling in ssa_def_bits_used - nir: extend ssa_def_bits_used to allow getting bits for any src component - nir: add a comp parameter into nir_def_bits_used - nir: change nir_all_uses_of_float_are_integer to return type masks and bits used - nir: change nir_def_bits_used to accept nir_scalar - nir/opt_idiv_const: a / b (where b > uint_max / 2) -> b2i(a >= b) - radv/lower_opt_fs_frag_pos: optimize f2u32(frag_coord_xy) & 0x1 to subgroup ops - radv: use quad_pos for pixel_coord conditionally based on dynamic state - radv: disable AMD_device_coherent_memory on gfx12 due to out of order behavior - ac/nir: fix incorrect upper bound for view_index - radv: lower view_index to a user SGPR instead of layer_id - radv: use PKT3_SET_SH_REG_PAIRS for setting multiple view_index SGPRs on gfx12 - nir/tests: restructure opt_varyings_tests_bicm_sysval - nir/opt_varyings: move (c ? interp_input0 : interp_input1) into the prev shader - nir: add shader_info::sample_mask_in_declared because it has side effects - radv: fix a rare crash with NULL PS and force_vrs_per_vertex - radv: use radeon_opt_set_context_reg for PA_CL_VRS_CNTL to fix random behavior - radv: use VRS flat shading even if other VRS state is enabled - radv: remove the VRS rate output if VRS flat shading overrides it - radv: cosmetic VRS changes - radv: don't use PS_ITER_SAMPLE to force VRS 1x1, use SC/DB VRS override instead - radv: remove the VRS rate output if VRS is force-disabled by FS - radv: don't execute pre-rast shader info code for FS - radv: remove no-op code from radv_consider_force_vrs for POPS - radv: disallow force_vrs_per_vertex with FragCoord when using GPL & ESO - radv: move force_vrs_per_vertex to emit_fsr_state to make it robust (rewrite) - radv: remove redundant PA_CL_VRS_CNTL setting from the initial state - radv: reduce duplication in gfx103_emit_vrs_override_state - radv: disable the VRS image on gfx11.x if the VRS rate is overridden - radv: fold gfx103_pipeline_vrs_flat_shading into its only use - radv: always set EN_VRS_RATE=1 because GE_VRS_RATE can also disable it - radv: inline radv_is_vrs_enabled - radv: use shader_info::fs::sample_mask_in_declared - radv: ignore VRS state for sample_mask_in lowering and optimizations - radv: take sample_mask_in_declared into account when lowering to pixel_coord - radv: set key.ps.force_vrs_enabled and key.vrs_may_be_enabled more accurately - bin/drm-shim: forward the error code from the command to the user - radv: fix low pixel throughput with NULL DS on GFX11.x - radv: don't set DB_Z_INFO.NUM_SAMPLES = 3 on gfx12 - radv/nir_trim_fs_color_exports: use a state structure to pass parameters - radv/nir_trim_fs_color_exports: remove mrt0.w if alpha_to_one makes it dead - ac: fix a GPU hang with LLVM due to incorrect VGPRS decoding of LLVM output - ac/llvm: rename ac_parse_shader_binary_config -> ac_parse_llvm_binary_config - radeonsi: use wave64_vgpr_encode_granularity - radv: don't expose memory types from AMD_device_coherent_memory without the ext - radv: fix determining the raster prim for guardband - radv: fix determining the raster prim for line mode - radv: fix determining the dynamic raster prim for FS barycentrics - radv: fix determining the static raster prim for FS barycentrics and front_face - nir/opt_varyings: fix incorrect counting of emit_vertex within a block Mario Kleiner (22): - wsi/display: Expose VK_FORMAT_B8G8R8A8_UNORM before VK_FORMAT_B8G8R8A8_SRGB - wsi/display: Improve connector->last_nsec timestamping. - wsi/display: Add workaround for all-zero valued pageflip events. - wsi/display: Deal with vblank-less systems for VK_EXT_present_timing. - wsi/common: Small compliance fixes for VK_EXT_present_timing. - wsi/common: Allow VK_EXT_present_timing present without presentStageQueries. - wsi/common: Allow to return queue_done_time in host time domain. - wsi/wayland: Unconditionally assign present_timing.time_domain. - wsi/common: Add VK_GOOGLE_display_timing support for KHR_display. - wsi: Don't try to create a timestamp query pool without driver support. - vulkan/wsi: Optionally expose VK_GOOGLE_display_timing on wsi wayland+x11. - wsi/wayland: Always use clock monotonic domain for GOOGLE_display_timing. - vulkan/wsi: Add hk, nvk as VK_GOOGLE_display_timing supported drivers. - docs/features: Add missing VK_EXT_present_timing enabled for X11. - wsi/display: Actually fix vblank-less systems for VK_EXT_present_timing. - pvr: Expose VK_KHR_present_id and VK_KHR_present_wait. - hasvk: Expose VK_KHR_present_id2 and VK_KHR_present_wait2. - hasvk: Expose VK_KHR_calibrated_timestamps. - hasvk: Expose VK_EXT_present_timing and VK_GOOGLE_display_timing. - wsi/display: Don't update connector last_frame/nsec in vkGetSwapchainCounterEXT. - wsi/x11: Skip next_present_ust_lower_bound assignment in certain FRR mode. - wsi/x11: Refine VRR vs. FRR detection a bit. Martin Roukala (né Peres) (18): - zink/ci: mark blender-demo-cube_diorama as flaky on gfx1201 - turnip/ci: document recent flakes - ci: disable the valve-kws farm - Revert "ci: disable the valve-kws farm" - freedreno/ci: reduce the parallelism of the a750-vk job - freedreno/ci: document more failures for the a750-gl-cl job - radv/ci: reduce parallelism for radv-gfx1201-vkcts - radv/ci: document more flakes - radeonsi/ci: document new flakes - amd/ci: tighten the timeouts of the Valve jobs - zink/ci: document a recent regression on navi10 - zink/ci: document more flakes - zink/ci: tighten the timeouts of the valve jobs - nvk/ci: tighten the timeouts of the valve jobs - radv/ci: bump the timeout of the valve vkd3d-asan jobs - ci: allow controlling which hw test jobs to create at pipeline creation - panfrost/ci: turn bifrost / valhall rules into per-kernel driver - radv/ci: document another WSI flake in radv-renoir-vkcts-full Mary Guillemard (45): - nvk: Use SET_REFERENCE in nvk_CmdResetQueryPool - nvk: use MME shadow RAM in nvk_meta begin/end - nvk: Move nv_push closer to their uses in nvk_cmd_begin_end_query - nvk: Clear counters at the begin of a query - nvk: Remove delta handling from query pool - nvk: Conditionally enable counters when needed - nvk: Move report offset to reports_start for nvk_CmdCopyQueryPoolResults - nvk: Handle zero queries in CmdCopyQueryPoolResults and CmdResetQueryPool - nvk: Store available and timestamps packed together - nir/lower_bit_size: Preserve float controls when lowering alu ops - nvk: Handle foreign queue dependencies - nvk: Handle host accesses barrier - nvk: Multiply by local_size for CS invocations in DGC codepath - nak: Allow YY swizzle for SM20 and SM32 asserts - nir/nir_format_convert: Add missing u2f32 in nir_format_unpack_r9g9b9e5 - nir,nak: Add match_any_nv - nak: Add a lowering pass for shared memory atomics in mesh stages - nvk: Prepare nvk_shader for GS header upload for mesh shaders - nvk: Prepare cbuf for mesh shader support - nvk: Add support for mesh and task shader binding - nvk: Implement mesh draw commands - nak: Implement mesh and task shader stages - nvk: Do not set lower_cs_local_index_to_id - nvk: Only lower shared memory for compute shaders - nvk: Lower mesh and task shaders - nvk: Advertises VK_EXT_mesh_shader - docs/nvk: Add some notes about mesh shading and ISBE layout - nvk: Do not report task and mesh stages as supported on pre-Turing - nvk/nvkmd: Do not merge bind operations across VA mappings - nvk: Implement support for non graphics timestamp - nouveau/mme: Add some simple MME shadow RAM dumper - nouveau/mme: Add a test for MME Shadow RAM behavior - nvk: Increase maxStorageBufferRange and maxBufferSize - nvk: Default to output primitives as lines for tesselation parameters - nvk/ci: Update expectations and document failures - nvk: Use I2M in CmdUpdateBuffer when possible - panvk: Split cmd_prepare_push_uniforms logic - drm-shim/nouveau: Report proper values in DRM_NOUVEAU_GET_ZCULL_INFO - drm-shim/nouveau: Stop using nouveau gallium names for classes - drm-shim/nouveau: Add Ada A to Blackwell B support - nvk: add a build option to override the build ID - nvk: Only increment CS counters when query is active in CmdDispatchBase - nvk: Do not enable remap in nvk_copy_indirect - nvk: Do not take base into account when lowering emulated attributes - nvk: Reenable compression support on Turing with nouveau 1.4.3 Matt Turner (15): - intel/elk: Remove some dead code - intel/elk: Remove dead TXL_LZ/TXF_LZ opcodes - radv: fix UB in radv_format_pack_clear_color for snorm formats - radv/perfcounter: guard select1 access in radv_emit_select - radv/perfcounter: add GFX11 performance counter selectors - radv: expose VK_KHR_performance_query on GFX11 - util, llvmpipe: flush subnormals to zero on ARM/AArch64 - nir: fix dedup_entry memcmp on structs with padding - gallivm: fix lp_build_round on altivec/VSX - gallivm: fix small_unorm -> unorm8 fetch path on big-endian - nir/tests: allow relative error in compare_inexact - nir: fix f2u/f2i constant folding to poison NaN and out-of-range inputs - nir: use i2f32 for patterns with signed-extraction opcodes - nir: fix pack_uvec4_to_uint to mask input components to 8 bits - nir/tests: fall back to integer comparison when float interpretation is NaN Matthieu Oechslin (7): - r600: Fix crash on R600/R700 with custom border color - r600: Improve and document R600_TRACE - r600: Stop emitting relocs with virtual address enabled - r600: Workaround GPU hang with compute shaders when VA is enbaled - r600: Calculate address at emit time for SSBOs - r600: Fix MSAA 2D view from array with VA enabled - r600: Document RADEON_VA and remove SB options references Mauro Rossi (5): - radv: Fix gnu-empty-initializer errors in 480a94fb - radv: Fix gnu-empty-initializer errors in 8c10eab1 - radv: Fix gnu-empty-initializer errors in ca9191a8 - intel/common: remove fallthrough annotation in unreachable code - pan/perf: fix building error due to 'Mali G1.xml' file name with space Maíra Canal (2): - etnaviv/ml: derive stride-2 destriding offsets from padding - v3dv: Drop legacy comments about single-sync support Mel Henning (36): - nak: Use shader_info->var_copies_lowered - nak: Use NIR_LOOP_PASS - nvk: Split out nvk_cmd_fill_memory - nvk: Allocate a zcull save region in fewer cases - nvk: Zero zcull data in layout transition - nvk: Don't LOAD_ZCULL w/ VK_RENDERING_RESUMING_BIT - nvk: Re-enable zcull save/restore - nvk: Add a wfi for blackwell in CmdDispatchIndirect - nvk: Disable compression on Turing - compiler/rust: Fix inline wrapper include dir - nak/nvdisasm_tests: Fix expected value of F16v2 - nak: Fix encoding of f16x2 min/max on sm90+ - vk/meta: Move get_uint_format_for_blk_size to common - nvk: Make nvk_cmd_buffer_queue_flags non-static - nvk: Split out aligned_for_linear_attachment - nvk: Use meta for image copies where possible - nil: Pass ImageDim to Tiling::choose() - nil: Pick tiling params closer to proprietary - nvk: Interp frag_coord at centroid for min_sample_shading - nvk: Fix DGC localsize computation - nvk: Serialize shaders with asm - nvk/meta: Rename begin/end with a _gfx suffix - nvk/meta: Implement save/restore for compute - nvk/meta: Add save_generic helpers - nvk: Use compute meta for some vkCmdCopyImage2 - nvk: Use meta for vkCmdCopyImageToBuffer2 - nvk: Use meta for vkCmdCopyBufferToImage2 - nvk: Use meta for vkCmdCopyBuffer2 - nvk: Add _ce suffix to nvk_cmd_fill_memory - nvk: Use meta for vkCmdFillBuffer - nvk: Handle large indirect stride pre-Turing - nvk: Prepare indirect draws for 64-bit stride - nvk: Convert draw/dispatch to device_address_commands - nvk: Don't re-align ssbo size/address - nvk: Move ssbo_4b_align to drirc - nvk: Use ?: in nvk_physical_device_compiler_flags Michael Cheng (11): - intel/ds: Add end_event_dyn() and CREATE_DUAL_EVENT_CALLBACK_DYN macro - intel/ds: Label compute events with dispatch dimensions in Perfetto - intel/ds: Label selected draw events with vertex count - brw: Fix ordered dependency exec_all handling on Xe2+ - intel/brw: allow baking more SBID dependencies into instructions on Xe2+ - intel/brw: Don't bake a long-pipe RegDist with an SBID dependency - intel/brw: Factor out combinable_ordered_pipe() helper - nir/opt_gcm: add option to keep texture ops in large loops - intel/brw: keep texture ops in large loops - intel: Fix DEBUG_FS_SIMD mask - intel: Fix operator precedence in intel_simd_overridden Michal Krol (7): - gallium: add pipe_sampler_view::min_lod_clamp - lavapipe: implement VK_EXT_image_view_min_lod with fractional minLod - gallivm/llvmpipe: fix VK_EXT_image_view_min_lod via texture handle path - lavapipe: lower array-deref-of-vec for mesh shader outputs - lavapipe: fix format properties for R10X6G10X6B10X6A10X6_UNORM_4PACK16 - gallivm: honour exec mask in EmitMeshTasksEXT - gallivm: don't deref a NULL buffer descriptor with an empty exec mask Michel Dänzer (14): - winsys/amdgpu: Use render node only as fallback - mr-label-maker: Label src/gallium/winsys/amdgpu as radeonsi - mr-label-maker: Label src/gallium/winsys/radeon as r300, r600 & radeonsi - egl/gbm: Do not destroy BO of current front buffer - egl/gbm: Use local variable for better readability - egl/gbm: Eliminate max_age local variable - egl/gbm: Ignore buffers with no BO for destroying excess BOs - egl/gbm: Ignore current front buffer in get_back_bo - egl/gbm: Use local variable for better readability in get_back_bo - egl/gbm: Eliminate local variable "age" in get_back_bo - egl/gbm: Use continue instead of nested block - egl/gbm: Eliminate local variable "max_age" in get_back_bo - dri3: Increment draw->send_sbc after waiting for last presentation - dri3: Simplify target_msc calculation in loader_dri3_swap_buffers_msc Mike Blumenkrantz (104): - lavapipe: KHR_device_address_commands - radv: add RADV_QUEUE_DISABLE env var for selectively disabling queues - llvmpipe: fix min_samples + A2C - lavapipe: fix indirect memory copies - lavapipe: fix pushconst data updating - lavapipe: null out local var to avoid uninit warning - util/format: support 256-bit formats in util_format_get_tilesize() - lavapipe: use the right type for DGC mesh draws - lavapipe: rework immutable samplers - lavapipe: allow fbfetch with shader objects - vk/cmd_queue: always ceil() param lens - vulkan: update spec to 1.4.350 - lavapipe: maintenance11 - llvmpipe: always set view_index for linear rasterizer - llvmpipe: unify setting raster_state for thread data - lavapipe: update cbuf count when remapping attachments - lavapipe: unset attachment remap state if pColorAttachmentLocations==NULL - lavapipe: fix setting colormasks when attachments get remapped - ci: stop skipping HIC tests on lavapipe - zink: use maintenance5 to more effectively set storage texel usage for bufferviews - zink: delete zink_resource_object::storage_buffer - zink: remove remaining maint5 checks - aux/trace: silence -Waddress warnings in macros - zink: delete unused descriptor variable - zink: fix mixing of mesh descriptor bindings with gfx bindings - meson: fix renderdoc integration define - vulkan: move vk_shader_stages_from_bind_point() to vk_util - zink: disable implicit sync handling for qcom proprietary - zink: rework custom sample locations - lavapipe: enable some forgotten ds3 states - zink: fix unbinding vertex buffers from null VS state - zink: add another anv/adl flake - zink: create views for samplers lazily - lavapipe: correctly disable depth/stencil in secondaries - vk/cmd_queue: simplify gross struct duplication - zink: use custom sample locations to (mostly) handle multisample=disabled - zink: link up vs COLx vars -> fs BFCx - zink: be more conservative about query pool sizing - lavapipe: stop using pipeline layouts in some places - lavapipe: Implement VK_EXT_descriptor_heap - zink: handle uint wrapping with batch submit count - zink/bo: reduce wasted memory due to the size tolerance in pb_cache - zink/bo: add an enum to disambiguate bo types - zink/bo: stop using pb_buffer vtable for destroy - zink/bo: use only a single layer of slabs - zink/bo: use pb_buffer_lean to save a little mem - zink/bo: check for usage before completion when reclaiming bos - zink: use maint11 for sso shader object compile - zink/clear: fix full_clear condition in texture clear - zink/clear: handle texture clears on current fb texture - llvmpipe: create a zeroed payload for use without task shaders - zink: always return DMA_BUF type handles from resource_get_handle - zink: tag tc info update in a few more places - util/tc: iterate the rp info more accurately during batch execution - aux/tc: enforce strict resolve semantics - vulkan/wsi: pass VkSurfaceCapabilities2KHR to get_capabilities - vulkan/wsi: add VK_IMAGE_CREATE_MULTISAMPLED_RENDER_TO_SINGLE_SAMPLED_BIT_EXT where supported - lavapipe: EXT_multisampled_render_to_swapchain - zink: fix import2d sampler view creation - zink: when triggering zink_blit_barriers() for src==dst, apply separate barriers - zink: stop forcing barriers if previous access was write - zink: properly invalidate fb attachments on dontcare stores - zink: proactively apply transfer sync when tracking renderpasses - zink: don't invalidate cbufs without inlined resolve - zink: add some ci flakes - tu: handle partially set resolve attachment info without crashing - util/tc: store resolve geometry to rp info - zink: use tc info to handle partial resolves - util/tc: unset TC_RESOLVE_STRICT - zink: set NO_TASK_SHADER for pipeline layouts with shader objects - zink: a618 ci updates - zink: always use src stages when flushing glMemoryBarrier calls - zink: always flush specified memory access for glMemoryBarrier calls - zink: reset usage following SHADER_WRITE access - zink: drop imageless_framebuffer requirement - zink: add a vb param to vertex buffer binding - zink: move vb binding out of c++ - zink: hook up VK_KHR_device_address_commands - zink: use DAC for vertex binding - zink: fix a missing case of zink_batch_submit_count_diff() - zink: stop unsetting resource usage on batch reset - zink: free nir if cs program create fails - zink: enable signed vbs - st/pbo_compute: account for drivers failing to create cs shaders - zink: split more read/write barriers - zink: stop adding usage with last-ref tracking - zink: unset unordered access on ordered transfer ops - zink: use bigger hammer to force sync between unordered->main cmdbufs - zink: add api for disabling reordered read/write - zink: unset ordered_access_is_copied when disabling unordered access - zink: noop per-resource synchronization for unordered->ordered access - gallium/cso: make unbind_context an explicit call - lavapipe: handle depth blit aspect masking - st/context: unbind gs shader before deleting hw select gs shaders - zink: always un-suspend queries on end - lavapipe: advertise dynamicRenderingLocalReadDepthStencilAttachments - zink: fix the fix for ZINK_RENDERDOC=all - zink: don't increment unique_id for reused bos - zink: translate depth write ALWAYS to GE/LE if possible - util/blitter: fix blitting multiple array layers - zink: start ZINK_DEBUG=perfinfo - zink: revert cached mem handling for staging uploads - zink: add anv ci flake - zink/ci: switch zink/anv jobs to surfaceless+i915 Mohamed Ahmed (8): - nil/modifiers: Clarify drm_format_mods_for_format rejecting modifiers for unsupported color formats - nvk: Calculate and stash the plane offset and alignment at create time - nvk: Extend tiled_shadow to be multiplanar - nvk: Defer tiled shadow plane memory allocation to draw time - nvk: Enable multiplanar YCbCr linear modifiers - nvk: Use the pre-calculated offsets for sparse binds - nvk: Remove nvk_image_plane_size_align_B() - nil: enable PLC for compressed data Nanley Chery (23): - intel/blorp: Halve max bpp for some redescribed blits - anv: Add transfer_src usage for ANDROID_external_format_resolve - anv: Improve the fast clear layout perf-warn - anv: Improve the CCS_E-incompatible perf-warn - anv: Avoid aux-disabling paths for block-compression - anv: Allow CCS on more storage images for gfx12.5 - anv: Move storage check out of CCS-compat helper - anv: Flush previous aux-mode changes - intel/isl: Define a CMF for ASTC formats - intel/isl: Fix the initial state HiZ state for Xe2+ - anv: Dedent a closing curly brace - anv: Skip some CCS performance warnings on gfx9-11 - anv: Allow partial depth fast clears on gfx12+ - anv: Set TRANSFER_DST_BIT for HiZ operations - intel: Add and use ISL_AUX_USAGE_ZCS - hasvk: Delete enum anv_depth_reg_mode - iris: Rework HiZ plane optimization disabling - iris: Disable HiZ planes for some read-only tests - anv: Track the depth buffer aux usage - anv: Enable overriding HiZ in depth stencil state - anv: Rework HiZ plane optimization disabling - anv: Bypass HiZ planes for read-only depth tests - anv: Drop the anv_disable_hiz drirc option Natalie Vock (9): - radv/rt: Don't overwrite bvh_base at the start of the traversal loop - radv: Dump printf buffer after detecting a GPU hang - radv/rt: Cache stack sizes of ahit/isec shaders from imported NIR - radv: Work around midpoint sorting issues instead of disabling - radv: Fix destroying address binding reports - mailmap: Update my email - nir/opt_loop: Don't peel header blocks that jump - radv/nir: Clean up descriptor index lowering - radv: Expose mutable acceleration structure descriptors Nataraj Deshpande (1): - intel/perf: map ray tracing counters to RAYTRACING group Neha Bhende (1): - svga: fix shared memory index for svga driver Nemallapudi, Jaikrishna (1): - intel/dev: fix timebase_scale ticks-to-ns precision loss across 2^32 Nick Hamilton (5): - pco: fix clamping the array index when shaderImageGatherExtended is enabled - pvr: Enable shaderImageGatherExtended - pvr: Revert don't csb emit multi-layer clear attachments without rta support - pvr: Fix load-op shader when loading from a 2d image view of a 3d image - glthread: fix check for unroll draws using user VBOs when the ctx supports GLES Okenczyc, Andrzej (1): - amd/vpelib: Report unsupported status if streams target rect equals 0 Olivia Lee (23): - pan/bi: fix memory access alignment - pan/genxml: add definitions for adjacency draw modes - panvk: add support for adjacency primitive topologies - panvk: fix executable properties handling for IDVS varying shaders - pan/csf: rename immediate CS add builder functions - pan/v13: add CS builder functions for reg/reg add and sub instructions - pan/v13: add CS builder functions for shift instructions - pan/v13: implement constant integer multiplication CS helper - pan/v13: implement CS udiv - panvk: remove redundant invalid primitive topology cases - panvk/csf: allow SYNC_WAIT-style synchronization in launch_gfx_cs - panvk: add create_shader helper for compiling full meta shaders - panvk: return both gpu and cpu pointers from cmd_prepare_*_push_uniforms - poly: allow VS outputs with <32 bits per component - poly: refactor GS lowering to store output and selected variables together - poly: preserve src_type in GS rast shader outputs - poly: preserve output component counts in GS - poly: move hk passthrough GS code to libpoly - poly: allow specifying output types in passthrough GS - nir: add nir_slot_num_components helper - poly: allow specifying component count in passthrough GS key - poly: clarify assertion failure message in lower_store_to_var - panvk/csf: flush primitives generated query writes from CSF Omar Rashwan (2): - intel: Fix bit width of int literal in eu stall viewer - intel: define type for std::max in eu stall viewer Patrick Lerda (19): - r600: refactor r600_shader_buffer_info_sel - r600: refactor eg_setup_buffer_constants - r600: rename sh_txs_cube_array_comp to sh_resinfo_via_uniform - r600: cypress resinfo buffer size workaround - r600: fix alpha-to-coverage and alpha-to-one used together - r600: add sample_lz and sample_c_lz opcodes compatibility - r600: update r600 nir for sample_lz and sample_c_lz - r600: enable EXT_texture_shadow_lod - docs/features: add GL_EXT_texture_shadow_lod - r600: remove r600_get_hw_atomic_count - r600: fix atomic buffer offset - r600: implement tes and tcs instanced gl_PrimitiveID support - r600: make r600_copy_region_with_blit global - r600: implement msaa 2d view from array - r600: update vertex emit_varying_pos - r600: fix atomic_counter_post_dec - r600: update memory barrier operations - i915: fix emit_hw_vertex() unbounded memory access - r600: update muladd support configuration Paulo Zanoni (30): - intel/isl: fix assert when surf->size_B is > UINT_MAX - intel/isl: warn about excessive num_elements only once - anv: don't silently convert view ranges from u64 to u32 then u64 - docs/envvars: remove ANV_SPARSE and ANV_SPARSE_USE_TRTT - docs/envvars: document ANV_SYS_MEM_LIMIT - docs/envvars: update the ANV_DEBUG documentation - intel/nir: fix sparse shadow comparison for BRW - anv/sparse: bring back our (limited) support for depth/stencil - brw: evict memory for workgroup scope in Xe2 and newer - intel/mi_builder: add mi_ixor() - intel/mi_builder: add mi_umax2() - intel/blorp: prepare for usage of mi_builder.h - libcl/vk: add aligned(4) to VkCopyMemoryIndirectCommandKHR - libcl/vk: add VkCopyMemoryToImageIndirectCommandKHR and its members - anv: implement VK_KHR_copy_memory_indirect - intel/tools: fix stall_csv_filename maybe-unitialized error - intel/brw: move cache_mode assignment to after send->sfid choice - brw: split cache mode selection into atomic, load and store modes - brw: have a single if-ladder to pick cache_modes - brw: control cache_mode through bypass_{l1,l3} variables - intel/blorp: don't silently ignore compilation failures - intel/blorp: fix blorp base key initialization - intel/blorp: move struct blorp_blit_prog_key to blorp_blit.c - intel/blorp: don't include "util/format_rgb9e5.h" - anv: give anv_ensure_fp64_shader() a chance to be called - brw: don't preprocess software doubles if opts->softfp64 is not set - anv: don't put clear colors for aliased images in private bindings - intel/blorp: rearrange struct blorp_blit_prog_key - intel/blorp: pack every blorp key struct - intel/blorp: memset(0) blorp keys during initialization Pavel Ondračka (111): - r300,i915/ci: update expectations - r300/ci: update expectations - i915/ci: update expectations - r300: fix MSAA resolve COLORPITCH tiling after pipe_surface de-pointerization - r300: dirty VS state when switching variants - dri3: add big-endian 8888 fourccs to dri3_cpp_for_fourcc - dri: add big-endian 8888 entries to dri2_format_table - dri: add big-endian 8888 entries to driImageFormatToSizedInternalGLFormat - nir: fix partial loop unroll OOB check for loops not starting at 0 - nir/tests: add helpers for counting used/unused instructions - nir/tests: add partial unroll OOB tests - r300: drop unused input arrays from ntr - r300: drop multiple ubo support from ntr - r300: drop GS/tess and load_draw_id support from ntr - r300: drop framebuffer fetch handling from ntr - r300: drop GLSL 4.x texture ops from ntr - r300: drop unsupported sampler dimensions from ntr - r300: drop the i915g vertex_id/instance_id U2F branch from ntr - r300: drop GLSL 4.x interpolation intrinsics from ntr - r300: drop opcode paths lowered before emission from ntr - r300: drop TEX2/TXB2/TXL2 dead path from ntr - r300/ci: run EGL deqp tests - r300/ci: update expectations - i915/ci: update expectations - r300: remove unused LIT opcode - r300: pack immediates more aggressively to avoid running out of constant slots - r300: reuse positive and negative immediate values - r300: remove extra newline for compiler errors - r300: remove the redundant control flow checks - r300: stop dumping TGSI - r300: always convert to NIR and move ntr later - r300: collect input/output info directly from NIR - r300: move compiler init earlier and to a helper - r300: fix use-after-free of remap_table in rc_remove_unused_constants - r300: build the dummy fragment shader directly as NIR - r300: emit one full vec4 immediate per NIR load_const - r300: move r300_transform_*_trig_input out of nir_to_rc - r300: extract TGSI->RC translation helpers into nir_to_rc.h - r300: emit RC instructions directly from nir_to_rc - r300: drop the TGSI opcode middle step from ntr - r300: allocate FS outputs from NIR locations - r300: use NIR varying locations directly in ntr - r300: stop declaring samplers with ureg in ntr - r300: lower sysvals to varyings - r300: use backend texture targets directly in ntr - r300: drop ureg shader properties in ntr - r300: drop ureg_DECL_temporary in ntr - r300: drop ureg_DECL_vs_input in ntr - r300: drop dead tg4_offsets and query_levels paths in ntr - r300: drop ureg_DECL_address - r300: get rid of user_src and ureg_dst - r300: stop using ureg_dst_undef in ntr - r300: get rid of ureg_program in ntr - r300: get rid of ureg_src_undef in ntr - r300: get rid of ntr_emit_load_output - r300: get rid of the precise modifier - r300: use RC registers directly in nir_to_rc - r300: drop more dead ntr code - nir/algebraic: prevent ffract optimization on lowered ffloor - r300: fix R300_VAP_TCL_BYPASS state leak - r300: fix vs->first leak in swtcl delete path - r300: add NIR pass to append the wpos output - r300: add NIR pass to add required color outputs - r300: convert swtcl vertex shader setup to NIR - r300: remove dead first-time build path from r300_pick_vertex_shader - r300: clean up some dead draw/TGSI leftovers - r300: remove draw support on big endian - meson: require r300 LLVM draw only on x86 - gallivm: add NIR pass to lower load_ubo_vec4 to load_ubo - gallivm: add no_integers intrinsic fixup pass - gallivm: add NIR pass to lower float if conditions - gallivm: add algebraic NIR pass for no_integers - gallivm: pass base NIR ALU types to cast_type - r300: enable VS instance ID in draw - r300: prepare for the the draw NIR path - draw: use gallivm NIR for no_integers vertex shaders - i915: use gallivm NIR for vertex shaders - r300: use signed index offset for index translation - r300: always use 32-bit indices on big endian - r300: use R32_FLOAT as 32-bit dummy vertex format - r300: fix occlusion query results on big endian - r300: fix BE 8888 render-to-texture endian state - r300: fix BE RGB565/RGB5 render-to-texture formats - r300: fix BE constant blend color for colorbuffer formats - r300: fix BE depth/stencil raw transfer endian state - r300: clean up endian swap selection - r300: add NIR LICM pass to enable removal of residual loops - i915: add NIR LICM pass to enable removal of residual loops - r300: remove dead state constants - r300: add private NIR state constant plumbing - r300: handle texture destination constraints in nir_to_rc translation - r300: lower backend texture coordinate handling in NIR - r300: lower fragment position in NIR - r300: lower VS seq/sne in NIR - r300: lower FS alpha-to-one in NIR instead of backend - r300: handle FS depth output channel mapping at nir_to_rc time - r300: lower r300 FS derivative stubs in r300_optimize_nir - r300: move r500 FS derivative fixup to nir_to_rc translation - r300: extend the wined3d A0 rounding pattern recognition - r300: some post-int/bool lowering optimizations - r300: don't split ALU instructions on R5xx - r300: run rgb alpha conversion after RC optimize - r300: penalize presubtract NOP in pair scheduling - r300: run late CSE after lowering vectors to registers - r300: fix swtcl per-vertex point size - r300: remove redundant FACE input workaround - r300: keep NIR output count consistent - r300: unify WPOS output handling between swtcl and hwtcl - r300: share vertex shader variant handling in hwtcl and swtcl paths - r300: emulate gl_FrontFacing on R3xx/R4xx - nv50_ir_ra: align B96 spill slots to vec4 Peng_Lx (1): - turnip/kgsl: close the dma-buf fd of our own allocations Peyton Lee (14): - amd/vpelib: add alpha fill support check - amd/vpelib: Support vpe 2.0 - amd/gmlib: add tm_generate_formatted_3DLut - radeonsi/vpe: add VPE 2.0 support - frontends/va: add ABGR format mappings - amd: validate and expose VPE 2.0.0 - radeonsi: gate format and rotate/flip support by VPE version - amd/vpelib: support vpe 2.2 - ac/gpu_info: add VPE_2_2 support - radeonsi/vpe: adjust message - amd/vpelib: tighten external LUT compound color pipeline updates - amd/vpelib: refine coding style - amd/vpelib: Fix Color Corruption AV1 Issue - amd/vpelib: Replace hardcoded bg format table size with sizeof Philipp Zabel (2): - etnaviv/isa: Fix Meson warning about etnaviv_isa_rs dummy library - meson: add rusticl to with_driver_using_cl Physics Enthusiast (1): - venus: allow to use vtest as a fallback for virtgpu Pierre-Eric Pelloux-Prayer (40): - radeonsi: clamp cp prefetch size - ac/info: add gfx12.1 identification - radeonsi/tests: update expectations - amd/virtio: use AMDGPU_VA_MGR_RESERVE_HALF_VA_FOR_PRT - amd/virtio: fix amdgpu_sw_info_address_prt_wa_control_bit handling - radeonsi: handle NULL return value from amdgpu_cs - gallium/vl: only release created sampler views - radeonsi: delay aux context initialization to first use - radeonsi: add has_gfx_compute property to si_screen - radeonsi: don't use staging texture when we can't blit - radeonsi/vce: deal with has_gfx_compute being false - radeonsi: create a mm subfolder for multimedia code - radeonsi: add si_init_screen_nir_options - radeonsi: add gfx subfolder - radeonsi: move shader cache code to new file - radeonsi: extract si_init_gfx_caps from si_init_screen_caps - radeonsi: add si_resource_copy_buffer - radeonsi/gfx: add si_gfx_screen.c - radeonsi/gfx: move code from si_get to si_gfx_screen - radeonsi: add si_gfx_context.c and move code from si_pipe.c - radeonsi: add si_context.c - radeonsi: move all multimedia files to mm - radeonsi: move more code to gfx subfolder - radeonsi: move function prototypes from si_pipe.h to si_gfx.h - radeonsi/gfx: remove unnecessary u_stub usage - radeonsi/gfx: move static inline helpers to si_gfx.h - radeonsi: add tests subfolder and move AMD_TEST code inside - gallium/u_blitter: remove unused skip_viewport_restore - radeonsi: fix sqtt setup - radeonsi/sqtt: hash only the relevant part of the shader key - radv, radeonsi: do sqtt buffer_size calc using uint64 - ac/sqtt: add ac_sqtt_update_bo_size - radeonsi: fix sdma copy for gfx10 - radeonsi: consolidate aux context creation into si_get_aux_context - radeonsi: use aux context locks in si_destroy_screen - ac/parse_ib: initialize data variables to 0 - radeonsi: delay si_disk_create_cache call - radv/rra: bump rt_driver_interface_version - radv: restore RRA capture support - radeonsi: fix typo in si_copy_from_staging_texture Pohsiang (John) Hsu (15): - mediafoundation: periodic clang-format - mediafoundation: code clean up - meson: Make with_gfx_compute depend on video encode support (mediafoundation) - d3d12: add av1 handling to d3d12_video_encoder_get_encode_headers and d3d12_video_encoder_update_current_encoder_config_state_av1 - mediafoundation: extract code to ProcessDX12EncodeContext - mediafoundation: fix a few minor variant bool handling - mediafoundation: detach xThreadProc frame processing from apiLock to unblock concurrentt ProcessOutput calls - d3d12: fix infinite gop handling in d3d12_video_enc_av1.cpp - mediafoundation: define AVC_LOG2_MAX_FRAME_NUM_MINUS4, HEVC_LOG2_MAX_PIC_ORDER_CNT_LSB_MINUS4 instead of using number. - mediafoundation: preserve low latency ping pong behavior between ProcessInput and ProcessOutput - mediafoundation: change default value for HEVC_LOG2_MAX_PIC_ORDER_CNT_LSB_MINUS4 to 12 - mediafoundation: initial av1 dx12 hmft prototype - d3d12: add support to output temporal delimiter for AV1 via raw_header - mediafoundation: ask for temporal delimiter for AV1 via raw header - d3d12: fix msvc build warning C4819 Qiang Yu (2): - ac,radeonsi,radv: fix print IB assertion fail for reserved fields - ac,radeonsi,radv: use V_581A_* engine sel for non-pws acquire_mem packet QwertyChouskie (3): - docs/features: Remove mentions of r300 and nv30 - docs/features: Fix typo - docs/features: Mark VK_EXT_descriptor_heap done for anv Radu Costas (6): - pco: Set register classes for vec refs - pco: Move preproc_vecs out of loop - pco: Add debug variables for RA - pco: Move RA context handling to state-based - pco: Try allocating with optimal temp registers - pvr, ci: Update axe and bxs failure list Rahul Mahantappa Bhadrashette (1): - radeonsi: emit MESHLET registers for mesh shaders in gfx10_emit_shader_ngg Raviraj Uppal (2): - driconf: disable allow_rgb16_configs for SPECviewperf - radv: app workaround implemented using internal layers for GFXBench 5.0 Rhys Perry (122): - ac/nir_lower_global_access: perform range analysis if useful - ac: add gfx11.7 enums - aco/gfx11.7: add opcode numbers - aco: adjust some gfx_level checks for gfx11.7 - aco/gfx11.7: don't create v_dot2c_f32_f16 - aco/gfx11.7: don't use v_pack_b32_f16 in do_pack_2x16 - aco/gfx11.7: allow any src VGPR for VOPD with two v_dual_mov_B32 - ac/gpu_info/gfx11.7: enable has_point_sample_accel - aco/gfx11.7: claim support - radv/gfx11.7: take GFX12 paths in radv_nir_lower_cooperative_matrix - radv/gfx11.7: enable float8 - radv/gfx11.7: enable shaderMixedFloatDotProductFloat8AccFloat32 - radv/gfx11.7: don't advertise shaderImageFloat32AtomicMinMax - aco: refactor spiller to use spills_needed variable - aco: prefer spilling smaller temporaries if it finishes spilling - aco: use RegisterDemand::operator[] more - radv: move ac_nir_lower_indirect_derefs to end of radv_shader_spirv_to_nir - radv: lower indirect derefs after linking - radv: don't use radv_optimize_nir after lowering indirect derefs for RT - ac: move lds_size_per_workgroup to ac_compiler_info - ac: move has_cs_regalloc_hang_bug to ac_compiler_info - radv: assert there is no padding in cache keys - radv: initialize nir_shader_compiler_options directly in compiler info - radv: move load_grid_size_from_user_sgpr to radv_physical_device - radv: move use_llvm to radv_compiler_info::key - radv: move fields to radv_compiler_info::key - radv: add fields to radv_compiler_info from radv_physical_device_cache_key - radv: remove radv_compiler_info::cache_key - radv: hash radv_compiler_info::key into the cache key - radv: remove most fields from radv_physical_device_cache_key - radv: remove radv_physical_device_cache_key - radv: remove radv_device_cache_key - ac/llvm: fix isub image atomic - radv: inline shader_compile() - radv: remove radv_aco_convert_opts - radv: replace radv_nir_compiler_options with a LLVM one - radv: split radv_compiler_info's family into debug::family and key::family - aco/gfx11.7: fix v_pk_min_f16/v_pk_max_f16 opcode numbers - nir/search: fix nir_algebraic_automaton after constant folding op(bcsel) - nir: rename nir_src_parent_instr to nir_src_use_instr - radv,ac: make rembrandt and vangogh cache compatible - radv: don't pass GPU name to disk_cache_create - radv: remove family from cache key - aco/validate: fix some RA validator error messages - aco/ra: fix v3b VALU at byte>0 - aco/ra: test the register file in get_reg_specified() when necessary - aco: add helpers to get instruction subdword capabilities - aco: rework subdword definition RA validation a bit - aco/ra: fix compact_relocate_vars path for get_reg_for_operand - aco/ra: fix fill() with certain subdword cases - aco: fix regclasses for spill/reload subdword temporaries - aco/ra: don't rename phi operands in get_reg_phi() - aco/ra: use phi_dummy instead of is_phi() - aco/ra: remove precolored checks in get_reg_impl() - nir/algebraic: optimize ishl(iadd(ishl, ishl)) - nir/algebraic: optimize ishl(iadd(iadd(iadd(a, #b), c), d), #e) - nir: make cmat_muladd_amd a subgroup intrinsic - nir: add load_deref_transpose_amd - nir: add load_global_transpose_amd - nir,ac/nir,aco: add load_global_tr_amd - radv: track cooperative matrix robustness - radv: use load_deref_transpose_amd for transposed cooperative matrix loads - radv: fix usage of radv_nir_cmat_length - aco: add cost estimation of s_barrier - aco: don't emit workgroup-scope p_barrier for single-wave workgroups - aco/waitcnt: always use uint32_t for event masks - aco: fix printing of primitive exports - aco: optimize redundant s_wait_alu vm_vsrc(0) during waitcnt insertion - aco: only assume load/store with semantic_atomic is atomic - aco: don't emit waitcnts before subgroup-scope execution barriers - aco: add split barrier instructions - aco: use split barrier instructions - aco: schedule split barriers - ac/gpu_info: add has_smem_partial_oob_access_bug - radv: workaround has_smem_partial_oob_access_bug - nir/opt_undef: fix prefer_nan - aco: don't increase barrier exec scope to subgroup - radv/bvh: use atomic load/store in update_gfx12.comp - ac/lower_global_access: set cursor earlier - ac/lower_global_access: rewrite try_extract_additions - ac/lower_global_access: parse u2u64 even if \*out_offset!=NULL - ac/lower_global_access: extract constants after ishl/imul - ac/lower_global_access: combine multiple 32-bit offsets - radv: use radv_shader_stage_key::keep_{statistic,executable}_info more - radv: move nir_debug_info from debug to key - radv: don't create nir_string if dump_shader=true - radv: simplify radv_declare_shader_args parameters - radv: add radv_shader_stage_key::keep_shader_arg_info - radv: cache shader IR, asm and spir-v - radv: parse stats from binary in radv_parse_binary_debug_info - radv: make raytracing radv_shader_stage_key array per-shader - radv: make raytracing radv_shader_stage_key initialization per-shader - radv: merge radv_shader_stage_key for combined ahit/isec shaders - radv: rework creation of traveral radv_shader_stage_key - radv: inline some helpers used in radv_pipeline_get_shader_key - radv: use nir_opt_uub - radv: repeat loop in radv_optimize_nir_algebraic_early more - radv: do nir_opt_algebraic last in radv_optimize_nir_algebraic_early - radv: fix barriers in decompress shaders - drm-shim: implement most readlink() without initializing the shim - drm-shim: skip init_shim() if drm_shim_fd_lookup() would be NULL - nir: add NIR_MEMORY_CONTROL_ARRIVE and NIR_MEMORY_CONTROL_WAIT - nir: add nir_lower_disordered_control_barriers - aco: implement NIR_MEMORY_CONTROL_ARRIVE and NIR_MEMORY_CONTROL_WAIT - vtn: implement SplitBarrierEXT - radv: implement VK_EXT_shader_split_barrier - radv: do radv_parse_binary_debug_info in radv_shader_dump_asm - ac/nir: skip SMEM fixup for more 32-bit load_global addresses - radv: set RADV_CMD_DIRTY_GFX12_HIZ_WA_STATE around attachment clears - vtn,nir: print shader filename and spec constants - radv: include debug information in bvh shaders - radv: shorten internal shader names - vulkan/bvh: fix to_emulated_float(-0.0) - vulkan/bvh: don't update min/max_bounds with inactive nodes - vulkan/bvh: enable SignedZeroInfNanPreserve in leaf.h - vulkan/bvh: limit valid 32-bit node keys to 0xfffffe00 - ac/nir/ngg: don't shrink device-scope memory barriers - ac/nir/ngg: track uniformity with multiple set_vertex_and_primitive_count - nir/load_store_vectorize: rework barriers to use entries - nir/load_store_vectorize: recreate entry key when adding from predecessor - vtn: don't fail at uniformity decorations on variables and OpBufferPointer - radv: drop support for cooperativeMatrixRobustBufferAccess Rob Clark (65): - tu: Remove use of fd_perfcntr_type - freedreno: Remove use of fd_perfcntr_type/result_type - freedreno/perfcntr: Remove type and result_type - freedreno/registers: Add json to describe perfctr groups - freedreno/registers: Generate perfcntr tables - freedreno/perfcntrs: Switch to generated perfcntr tables - freedreno/registers: Small reg32 vs reg64 fixes - freedreno/registers: Sync back xml changes from kernel - freedreno/registers: Add pipe to perfcntr group - freedreno/registers: Add gen8 perfcntr support - freedreno/registers: Correct register name - freedreno/registers: Add gen8 perfcntrs - pps: Re-emit time clock_sync more regularly - freedreno/ds: Use gpu timestamps - freedreno/ds: Split a6xx/a7xx counters out - tu: Fix preemption latency selector values - freedreno/a6xx: Expose subgroup ops - rusticl: Support/ignore -qcom-accelerate-16-bit - freedreno/common: Fix X2-90, add X2-85 - freedreno/registers: Skip deprecated warns for kernel - freedreno/registers: Add a6xx CMP counter group - freedreno/registers: Gen8 perfcntr fixes - drm-uapi: Sync msm_drm.h - freedreno/common: Add ioctl ptr helpers - freedreno/fdperf: Move where we setup counter groups - freedreno/fdperf: Prepare for partial-counter usage - freedreno/fdperf: Add PERFCNTR_CONFIG support - freedreno/ds: PERFCNTR_CONFIG support - freedreno/ds: Add a8xx derived counters - freedreno/perfcntrs: Add helpers to resolve group and countable - freedreno/perfcntrs: Add helper to assign counters - tu: Use counter allocation helper - freedreno/a6xx: Use counter allocation helper - freedreno/perfcntrs: Refactor derived counter setup - freedreno/perfcntrs: Use helper for derived counters - freedreno: Skip BV perfcntrs - tu: Disable preemption for counters on gen8 - tu/gen8: Program slice selector regs - freedreno/a6xx: Program gen8+ slice SEL regs - freedreno/perfcntrs: Expose gen8 counters - freedreno/a6xx: Push RB_A2D_PIXEL_CNTL magic into blitter - freedreno/registers: Improve A2D docs - freedreno/registers: Add RB_RESOLVE_CNTL_0.YUV_PLANE_ID - freedreno/a6xx: Un-open-code RB_A2D_PIXEL_CNTL - tu: Un-open-code RB_A2D_PIXEL_CNTL - perfetto: Add API to flush track events - util/thread: Flush traces at thread exit - util/queue: Flush perfetto before blocking - rusticl: Flush perfetto track events - perfetto: Increase SMB size - perfetto: Use BufferExhaustedPolicy::kStall - freedreno/perfetto: Add non-draw stage - freedreno/perfetto: serialize clk snapshots - freedreno/perfetto: Use sequence-scoped clk - freedreno/crashdec: Update gpu revision parsing - freedreno/crashdec: Add additional HFI queue - freedreno/a6xx: Don't clamp 32b clear values - freedreno/a6xx: Fallback for conditional blits - freedreno/a6xx: Handle R9G9B9E5 blits as R32_UINT - freedreno/a6xx: Set HALF_PRECISION for R11G11B10_FLOAT - freedreno/a6xx: Don't forget UBO driver params - freedreno/decode: Fix shader stats in summary mode - freedreno/ci: Update a660-vk-traces-restricted checksums - nir/convert_address_format: Split convert_def into two passes - nir/convert_address_format: Handle non-deref sources Rob Herring (Arm) (34): - ethosu: Make quantization shift signed - ethosu: Add a common initializer for struct ethosu_operation - ethosu: Store ethosu_tensor struct ptr in feature map - ethosu: Move stride calculation to lowering - ethosu: Fix concatenation OFM scaling - ethosu: Support axis 1 concatention - ethosu: Add fully-connected operation - ethosu: Rename ethosu_lower_add to ethosu_lower_eltwise - ethosu: Support element wise op with constant IFM buffer - teflon: Add multiply operation - ethosu: Add multiply operation support - teflon: Add TANH operation support - ethosu: Add logistic and TANH operations - teflon: Add hard swish operation - ethosu: Add hard swish operation - teflon: Add LeakyRelu operation - ethosu: Add LeakyRelu operation - teflon: Add quantize operation - ethosu: Add quantize operation - ethosu: Add reshape operation - teflon: Add minimum and maximum operations - ethosu: Add minimum and maximum operators - ethosu: Add performance counter debug output - teflon: Ensure all TfLiteRegistration fields are 0 - meson: Skip NIR tests with headers-only NIR - teflon/tests: Use reference kernels - ethosu: Compute elementwise broadcasts from OFM shape - ethosu: Fix depthwise conv layout for IFM depth 1 - ethosu: Flatten fully connected inputs - ethosu: Fix command dependency tracking - ethosu: Improve equal-cost block selection - ethosu: Use full scale for NOP pooling - ethosu: Preserve spatial dimensions for FC lowering - ethosu: Preserve fused pad extents Robert Mazur (6): - ci: update firmware tag to ff46ce35 - ci: update kernel tag to v6.19-mesa-712d - imagination/ci: Use standard CI-tron gfx-ci/linux kernel - pvr: switch core count mesa_logw() to pvr_finishme() - pvr: introduce PVR_IGNORE_FINISHME_WARNINGS envvar - pvr/ci: enable PVR_IGNORE_FINISHME_WARNINGS Rohit Athavale (1): - mediafoundation: Test compile steps v/s step , and set build flag Roland Scheidegger (1): - gallivm: fix subtle filtering issue with different min/mag filter for cube maps Roman Stratiienko (3): - v3dv/android: Add deferred ANB allocation support - v3dv: move noop_job creation to device scope - v3dv: Emulate multi-queue support via vk_queue for Android Romaric Jodin (1): - anv: Declare 00-mesa-defaults.conf as an input to anv_dricrc_gen.py Rouf, Farhan (4): - amd/vpelib: Introduced reset to frontend - amd/vpelib: Changed cmd_info input for background segment - amd/vpelib: Chroma coefficient select for sampled formats corrected - amd/vpelib: Refactoring Reset Pipes Function Rudraksha Gupta (1): - freedreno: add Adreno 225 Ryan Houdek (1): - turnip: Add an override to uncached memory type Ryan Mckeever (2): - pan/bi: check if preds are dominated by header in bi_find_loop_blocks - docs: advertise VK_KHR_multiview support for Bifrost Ryan Neph (1): - anv/xe: prevent WaitIdle optimization for fences with exported sync_fd Ryan Zhang (4): - panvk: add VK_IMAGE_LAYOUT_DEPTH_READ_ONLY_OPTIMAL to host copy layouts - gfxstream/platform: add missing inc_include to platform_virtgpu build - panvk: set cfg cull status according to primitive topology - panvk: Drop empty SYNC_ONLY bind queue ops Saeed, Ghamr (1): - amd/vpelib: variable was accumulating size and not reset properly Sagar Ghuge (49): - anv/rt: Copy 16bytes at once instead of copying 8bytes - anv: Fix Wa_14021821874, Wa_14018813551, Wa_14026600921 - brw: Pass write back register for ray query messages - intel/genxml: Update xml for dynamic stack ID control fields - anv: Enable dynamic stack ID control on Xe3+ - intel/genxml: Disable compute walker mid-thread preemption - util: Increase array size to 20 - intel/genxml: Added dispatch timeout counter extended field - anv: Update values for DispatchTimeoutCounter - anv: Set execution mask based on SIMD size - brw/rt: Commit hit even if we are skipping closest hit shader - brw/rt: Update committed hit leaf type properly - brw/rt: Use BLAS(Object) level to get the ray address - jay: Implement halt - anv/rt: Skip invalid node in child block count - anv: Pass vk_acceleration_structure_build_state as param - anv/rt: Extract common code in separate header - anv/rt: Use constant BVH offset instead of pushing - anv: Track parent-child map for BVH update - anv: Track leaf block offset map - intel: Add debug option to dump out parent-child map - anv: Implement update BVH - intel: Add debug hook to dump out BVH after update - ci-farms/vmware: Disable vmware tests for now - anv: Allocate lookup maps for update based on mode and flag - intel: Add drirc option to write lookup maps unconditionally - Revert "anv: Fix Wa_14021821874, Wa_14018813551, Wa_14026600921" - anv: Workaround game bug for Witcher3 - jay: Extend CS payload to handle BTD stack IDs - jay: Handle nir_intrinsic_load_btd_stack_id_intel intrinsic - jay: Factor out RT message header build part - jay: Implement btd_retire instrinsic - jay: Implement BTD Spawn intrinsic - anv: Compile init RT shader with Jay - anv: Bump subgroup size for histogram and prefix shader - vulkan: use center/extent form for instance node AABB transform - jay: Control cache_mode through bypass_{l1,l3} variables - jay: Setup bindless thread payload - jay: Handle nir_intrinsic_load_btd_global_arg_addr_intel - jay: Handle nir_intrinsic_load_btd_local_arg_addr_intel - jay: Set simd width for bindless shader - jay: Handle EOT for RT shader - jay: Init header with zero for all components - jay: Stuff stackIDs for trace_ray message - brw: Track if CS uses fences - intel: Fix async compute thread limit - anv: No need to flush RT cache if we update buffer via CS - jay: Track max stack size for bindless shaders - jay: Add loop_once_halt opcode Sahitya Kandru (1): - freedreno: Modify reg_size_vec4 for a608 and a612 to 32 Sam James (4): - glx: append extra_ld_args_libgl, not clobber - src: add -Wl,--no-fatal-rwx-sections for two libraries - gallium/dri: fix redundant Meson condition - util: remove bogus const attribute Samuel Pitoiset (263): - vulkan: add an option to lower SHADER_RECORD_INDEX to non-uniform - radv: lower SHADER_RECORD_INDEX to non-uniform - radv/ci: document some HIC failures since addrlib uprev for GFX11.7 - radv: add enable_mrt_output_nan_fixup to the physical cache key - ac/surface: add stencil-only support for host mem->surf copies - radv: add depth+stencil formats support with host image copy - radv: allow depth+stencil formats with host image copy - amd: allow addrlib to enable SIMD if possible - radv: advertise VK_EXT_host_image_copy by default on GFX10.3+ - radv/ci: document more HIC regressions on NAVI10 - vulkan: refactor vk_pipeline_robustness_state_fill() slightly - vulkan: pre-compute the default robustness state in the device - vulkan,treewide: stop passing vk_device to vk_pipeline_robustness_state_fill() - spirv,treewide: rework specialization constant - radv: fix GPU hangs with PS epilogs and secondaries properly - radv/rt: pass more parameters to radv_rt_nir_to_asm() - radv: add a radv_compiler_info object - radv: use radv_compiler_info everywhere during compilation - spirv: add support for SPV_KHR_constant_data - radv: advertise VK_KHR_shader_constant_data - radv: move queue related cmd buffer state to a new struct - radv: move uses_perf_counters to radv_cmd_buffer_queue_state - radv: move shader_upload_seq to radv_cmd_buffer_queue_state - radv: remove redundant initialization when beginning a cmdbuf - radv: zero-initialize radv_cmd_state only when a cmdbuf is reset - radv: pass radv_compiler_info to radv_pipeline_get_shader_key() - radv: store the number of PS params heuristic to radv_compiler_info - radv: re-introduce DGC+multiview support and enable it for vkd3d-proton only - radv: fix a potential NULL pointer dereference when emitting VBOs - radv: remove an useless check when emitting the index buffer - radv: only emit the "normal" index buffer when needed with DGC - radv: stop dirtying some states after DGC execute - radv: cleanup invalidating vertex draw state - radv: replace use_ngg_streamout by gfx_level checks - ac,radv,radeonsi: replace mesh_fast_launch_2 by gfx_level checks - vulkan: add missing VkMemoryRangeBarriersInfoKHR support - radv: add missing VkMemoryRangeBarriersInfoKHR from DAC - radv: simplify resetting pipeline state for ESO - radv: rename RADV_CMD_DIRTY_PIPELINE to RADV_CMD_DIRTY_GRAPHICS_PIPELINE - radv: stop tracking the last emitted graphics pipeline - radv: add RADV_CMD_DIRTY_COMPUTE_PIPELINE - radv: add RADV_CMD_DIRTY_RAY_TRACING_PIPELINE - radv: remove useless tracking about non-coherent RBs with secondaries - radv: slightly rework initializing the default graphics state - ci: bump libdrm to 2.4.133 - meson: bump required libdrm to 2.4.133 for AMDGPU - radv: move suspend_streamout to radv_streamout_state - radv: move streamout bindings to radv_streamout_state - radv: remove unnecessary radv_cmd_state::mesh_shading - radv: move index buffer state to radv_index_buffer_state - radv: cleanup suspending/resuming cond rendering with DGC - radv: move conditional rendering state to radv_cond_render_state - radv: move vertex buffer state to radv_cmd_state - radv: re-organize radv_cmd_state slightly - radv/ci: bump timeouts for radv-{navi21,gfx1201}-vkcts-full - ac/gpu_info: store more addr space info - ac/gpu_info: add has_smem_with_null_prt_bug - ac/gpu_info: query the PRT workaround control bit from libdrm - ac/nir: add a pass to fixup SMEM loads with NULL PRT pages - radv: run the pass to fixup SMEM loads with NULL PRT pages - radv: use the "LOW" address space for UBOs - radv/amdgpu: emulate sparse residency for the SMEM loads with NULL PRT workaround - radv: set RADEON_FLAG_EMULATE_SPARSE_RESIDENCY for sparse SSBO/UBO buffers - radv/ci: update list of skipped tests - ci: uprev vkd3d - radv: fix printing image format with RADV_DEBUG=img - radv/meta: fix expanding HTILE on compute with multisampling - docs: describe the contributions workflow for RADV - radv: bump VkConformanceVersion to 1.4.5.3 - radv: fix determining needed dynamic states when rasterization is disabled - radv: make optimalTilingLayoutUUID driver and chip specific - ac/surface: allow to select hybrid/block memcpy path for host copies - radv: take advantage of VK_HOST_IMAGE_COPY_MEMCPY_BIT - vulkan: replace VK_SHADER_CREATE_INDEPENDENT_SETS_BIT_MESA with the maint11 flag - vulkan: stop forcing independent sets for shader object - ci: uprev vkd3d - radv/tests: add tests for global pipeline keys compatibility - radv: allow DGC+multiview by default - radv: fix an assertion with RADV_DEBUG=fullsync on GFX11+ - radv: do not fallback to compute for image->buffer copies with emulated formats - spirv: preserve the explicit stride for untyped pointers with matrices - radv: add support for VK_SHADER_CREATE_INDEPENDENT_SETS_BIT_KHR - radv: adjust minImageTransferGranularity for transfer queue - radv: advertise VK_KHR_maintenance11 - radv: fix another case of VRS with mipmaps on GFX10.3 - radv: remove a TODO about layeredShadingRateAttachments - radv: invalidate command buffer state after executing secondaries - radv/meta: adjust an assertion for HTILE expand on SDMA with compute fallback - radv: clear the follower gang semaphore when a cmdbuf is reset - radv: destroy the gang CS when a cmdbuf is reset - radv: fix copying acceleration structure with DAC - nir: fix shuffling local IDs for quad derivatives with larger workgroup sizes - radv: enable radv_wait_for_vm_map_updates for Forza Horizon 6 - radv: advertise VK_EXT_device_fault by default - radv: remove an outdated comment in radv_GetDeviceFaultInfoEXT() - radv: move radv_GetDeviceFaultInfoEXT() to radv_device.c - util: pass a struct to driParseConfigFiles() - util: do not generate drirc options that shouldn't be parsed - util: fix declaring drirc options as string - radv: rename few drirc options for consistency - radv: use the new generation script for drirc - radv/ci: cleanup list of expected failures - util: add very basic way to validate drirc files - radv: validate drirc option names at compile time - radv: close the local fd immediately after the winsys is created - radv: rename radv_zero_vram to vk_zero_vram - radv: use radv_device::ws directly for quering sync payloads - radv: pre-compute a mask of supported global queue priorities - radv: add a separate function to query allocated/usage for each heap - radv: remove declared but unused create_null_physical_device() - radv/amdgpu: simplify syncobj verifications during submissions - radv: determine supported syncobj types directly in the physical device - radv/amdgpu: fix releasing the mutex for virtio and RADV_PERFTEST=localbos - nir: add new intrinsics for SPV_KHR_abort - spirv: implement SPV_KHR_abort - nir: add nir_lower_abort - radv: close the local fd slightly later when enumerating physical devices - radv: remove useless checks when creating a physical_device - radv: rename master_fd to wsi_master_fd - util: share the DOCTYPE for all driconf files - util: add a separate file for Zink drirc - util: add a separate file for RadeonSI drirc - util: add a separate file for turnip drirc - util: add a separate file for ANV drirc - util: add a separate file for NVK drirc - util: add a separate file for r300 drirc - util: add a separate file for iris drirc - util: add a separate file for asahi drirc - util: add a separate file for asahi vulkan drirc - util: add a separate file for panvk drirc - util: add a separate file for panfrost drirc - util: add a separate file for crocus drirc - util: add a separate file for dozen drirc - util: add a separate file for virgl drirc - util: add a separate file for r600 drirc - util: add a separate file for msm drirc - util: add a separate file for d3d12 drirc - util: add a separate file for vmgfx drirc - util: add a separate file for v3d drirc - util: add a separate file for hasvk drirc - util: remove useless comments in 00-mesa-defaults.conf - util,turnip: move drirc entries with vk_dont_care_as_load to Turnip - util,asahi: move drirc entries with no_fp16 to asahi - util: remove declared but unused drirc options - ci: adjust time-trace.sh to not exceed the limit of 255 chars - radv/ci: fix list of expected failures - radv/amdgpu: rework tracking allocated memory for budget - radv/amdgpu: stop deduplicating winsys - ac/nir,radv: lower task payload to zeroes when the mesh shader has no task - radv: enable radv_force_64_byte_sampled_image for Crimson Desert - radv: cleanup conditional header includes - radv: implement VK_KHR_device_fault - radv: advertise VK_KHR_device_fault - radv: fix DGC with conditional rendering and task+mesh shaders - radv/amdgpu: allow RADV_PERFTEST=localbos with virtio - radv: cleanup occurrences of radeon_info::has_vm_always_valid - radv: return VK_ERROR_INITIALIZATION_FAILED if VM_ALWAYS_VALID isn't supported - ci: uprev vkd3d - util: remove declared but unused DRIC_CONF_VK_REQUIRE_ASTC - util/drirc_gen: add a function to declare commmon VK options - radv: declare common VK drirc options using the helper - anv: declare common VK drirc options using the helper - turnip: declare common VK drirc options using the helper - radv/rt: fix a memory leak with hash tables - radv/rt: fix a memory leak with the RT prolog NIR - radv/rt: fix a memory leak with ahit/isec group - radv: fix a memory leak with perfcounters - radv/amdgpu: destroy the BO for the NULL PRT workaround earlier - vulkan: Update spec to 1.4.353 - ci: uprev vkd3d - ci/vkd3d: add support for running with ASAN - radv/ci: run vkd3d jobs with ASAN by default - radv: add the mesh scratch ring BO to the preambles BO list - radv: handle errors correctly when creating gang waits - aco: emit nir_jump_halt - radv: implement VK_KHR_shader_abort - radv: advertise VK_KHR_shader_abort - radv/amdgpu: defer allocating the NULL PRT BO - radv/ci: skip all WSI tests on GFX1201 - util/drirc_gen: change the driconf DTD to not require one app/engine entry - util/drirc_gen: fix generating 64-bit driconf options - util/drirc_gen: prevent generating empty structs - util/drirc_gen: add heap_memory_percent to common VK options - util/drirc_gen: allow to override the defaults VK WSI common options - dzn: use drirc_gen - pvr: use drirc_gen - panvk: use drirc_gen - hk: use drirc_gen - v3dv: use drirc_gen - venus: use drirc_gen - radv,anv: remove useless includes for drirc stuff - util/drirc: remove the driver option in drirc_validate - ci: add a new option called profile in ci_run_n_monitor.py - radv/ci: skip all WSI tests also on NAVI21/NAVI31 - util: remove useless entries for Intel hasvk - hasvk: use drirc_gen - ac/video: drop an useless drm_minor check - ac/descriptors: fix setting CB_COLOR_ATTRIB3.RESOURCE_LEVEL - ac/gpu_info: only initialize has_desc_resource_level on GFX10+ - nir/print: add a missing UNREACHABLE for unknown jump instructions - nir: add a new nir_jump_abort - nir,aco: use nir_jump_abort instead of nir_jump_halt for abort - radv/ci: add more flakes for RAPHAEL - radv/ci: update the list of expected failures for NAVI10 - radv: prevent closing the render node fd twice for AMD_FORCE_VPIPE=1 - radv: allow to query GPU info without creating a winsys - radv/amdgpu: add a function to query heap info - radv: query heap info without using the winsys - radv: duplicate the fd used for syncobj with KHR_display - radv: create one winsys for each logical device - radv: fix REPLAYED shader arena blocks not being marked as holes on free - radv/amdgpu: fix padding by one VM page - vulkan: fix lowering untyped accel struct with descriptor heap - spirv: mark UBO/SSBO array accesses as always in-bounds - radv: fix a synchronization issue with taskmesh and pending cache flushes - radv: fix a synchronization bug with DGC preprocess and taskmesh - radv: clear gang cache flushes when the command buffer is reset - vulkan: fix incorrect sType for VkDebugUtilsObjectTagInfoEXT - radv: workaround game bugs with Sniper Elite 5 - spirv: allow mapping readonly buffers with struct members - radv: fix clearing the streamout state on GFX12 - radv: remove redundant memory initialization on the CPU - vulkan: fix initializing address flags - vulkan: add vk_buffer_usage_flags() - radv: cleanup pCreateInfo uses for VkBuffer - radv: cleanup pCreateInfo uses for VkImage - radv: allow ptr to be NULL in radv_cmd_buffer_upload_alloc() - radv/amdgpu: fix computing allocated VRAM for imported BOs from fd - radv: fix a memleak with embedded samplers and descriptor heap - radv: store copying embedded samplers for heap to the shader layout - radv: enable VK_EXT_descriptor_heap by default - radv: disable VRS with MSAA 8x on GFX11-11.7 to prevent GPU hangs - radv: disable VRS for flat shading with MSAA 8X to prevent GPU hangs on GFX11 - radv: disable VRS with MSAA 8x also on GFX10.3 - radv: remove the deprecated warning for RADV_FORCE_FAMILY - docs,radv: auto-generate driconf documentation from drirc_gen - util: remove declared but unused driconf vulkan-related options - vulkan,anv,radv: do not crash when querying descriptor size for unsupported type - zink: fix a memleak in zink_init_format_props() - glsl: fix a memleak in link_assign_subroutine_types() - vulkan: search VkImageViewUsageCreateInfo in pNext - vulkan: implement VK_KHR_extended_flags - vulkan/wsi: implement VK_KHR_extended_flags - zink/ci: update lists for RADV - loader: fix a memleak - kopper: fix a memleak - zink: fix a memleak with fences - zink: fix a memleak with the emulated GS NIR shader - zink: fix a memleak with sampler state - radv: use the image view usage for MSRSTT transient iviews - radv: implement VK_KHR_extended_flags - radv: advertise VK_KHR_extended_flags - ci: apply patches to fix memleaks for GL/GLES CTS - zink/ci: add a new job for NAVI31 with ASAN enabled - pipe-loader: fix a global-buffer-overflow ASAN error when getting driconf - radv/meta: fix restoring descriptor heaps - Revert "spirv: allow mapping readonly buffers with struct members" - radv: always consider some outputs as invariant - radv: remove radv_invariant_geom driconf option - zink/ci: update trace checksums - radv: force late-Z with fragment shaders that use fbfetch - spirv: fix handling OpAbortKHR - radv: fix invalid assertions in DGC when queues aren't enabled Serdar Kocdemir (14): - gfxstream: add gitignore for generated code - gfxstream: Add VK_EXT_pipeline_protected_access - gfxstream: allow VK_KHR_maintenance extensions - gfxstream: some cleanup on device extension allow list - Set driver ID for gfxstream - gfxstream: allow VK_GOOGLE_display_timing - gfxstream: remove android conditioning for sampler extensions - gfxstream: use VK_DRIVER_ID_MESA_GFXSTREAM as driver id - gfxstream: update codegen for host side vulkan header update to v1.4.350 - gfxstream: Fix codegen causing missing vulkan structures - gfxstream: disallow maintenance6 extension due to serialization bugs - gfxstream: correctly ignore timeline semaphore info - gfxstream: allow VK_EXT_border_color_swizzle - gfxstream: check in auto-generated guest code Sergi Blanch Torne (15): - ci: disable Collabora's farm due to maintenance - Revert "ci: disable Collabora's farm due to maintenance" - ci: disable Collabora's farm due to maintenance - Revert "ci: disable Collabora's farm due to maintenance" - xfiles: update before uprev - ci: disable Collabora's farm due to maintenance - Revert "ci: disable Collabora's farm due to maintenance" - ci: review initial ANGLE flakes - ci,crnm: information from pipeline url - ci,crnm: search MR pipelines in forks - ci,crnm: handle exception when auth fail - xfiles: update before uprev Piglit - xfiles: update expectations based on 2026-7-8 nightly - ci: disable Collabora's farm due to maintenance - Revert "ci: disable Collabora's farm due to maintenance" Sergi Blanch-Torne (1): - ci,crnm: bugfix project default Sergio Sanchez Valencia (1): - d3d12/wgl: reclaim deferred BOs before ResizeBuffers Shih, Jude (4): - amd/vpelib: Alpha blending enhancement - amd/vpelib: Refactor DPP function table layout - amd/vpelib: Fix Compiler Warnings - amd/vpelib: Realign DPP callback initialization with the updated interface layout Sid Pranjale (12): - nvk: Implement VK_EXT_shader_atomic_float - vulkan: implement VK_EXT_debug_marker - v3dv: drop legacy CPU queue fallback paths - nak/nir: lower f16vec2 shared atomics - v3dv: replace single-field options struct with bool - v3dv: directly use v3d_has_feature instead of caps struct - broadcom/common: add multisync helpers - gallium/v3d: use common multisync code - v3dv: use common multisync code - v3dv: implement CPU-side fence merging for queue signaling - v3dv: simplify queue submission - v3dv: remove unused no-op job allocation setup Silvio Vilerino (18): - d3d12: Create PIPE_BIND_SHARED resources with D3D12_RESOURCE_FLAG_ALLOW_SIMULTANEOUS_ACCESS - mediafoundation: Create readable dpb buffers with PIPE_BIND_RENDER_TARGET and PIPE_BIND_SHARED for DX11 sharing - Revert "d3d12: Video sliced encode: Use same ID3D12Fence/different per slice values as optimization" - d3d12: Flush stale video encode wait registrations when reusing ID3D12Fence objects - d3d12: Support video encode AUTO slice/tile only capable hardware - mediafoundation: check for AUTO slice/tile only capable hardware - d3d12: Use sequential video enc subregion signaling - d3d12: d3d12_create_fence_raw to lazily register fence event on waits - d3d12/video: fix comparison-with-wider-type warnings - d3d12: avoid signed integer overflow in copy staging box setup - util: u_trace.c: Fix error C4189: buffer_count: local variable is initialized but not referenced - d3d12: Fix NULL dereference check in d3d12_video_buffer_destroy - d3d12: Use res_device to import resource from different device via handle - pipe: Expose new fence_wait_multiple operation - d3d12: Implement fence_wait_multiple with SetEventOnMultipleFenceCompletion - mediafoundation: Use eventless fence_wait_multiple instead of WaitForMultipleObjects - d3d12: Remove event cleanup since d3d12_fence now uses lazy SEOC - d3d12: Check fence values before SEOC in d3d12_fence_wait_multiple Simon Perretta (29): - pco: reserve additional outputs for trilinear sampled coeffs - pco: amend tg4 lowering - pco: track how many tg4/raw sample comps are needed - pvr: consider barriers when calculating compute instances - pvr, pco: add support for spilling shared memory to global memory - pvr, pco: store device runtime info in compiler context - pco: conditionally spill shared memory to global memory - pco, pvr: finish and enable VK_KHR_workgroup_memory_explicit_layout - pco: drop global path for null descriptor checking - pco: add mappings for setl, savl ops - pvr, pco: add "real" basic subgroup support - pco: handle mov offset special regs - pco: add support for read_invocation via shared memory - pco: add subgroup ballot support via shared memory - pvr: advertise VK_EXT_shader_subgroup_ballot and ballot feature - pco: add br.skip_next op - pco: commonize execution mask counter ref helper function - pco: add support for subgroup vote_{all,any} ops - pvr: advertise VK_EXT_shader_subgroup_vote and vote feature - pco: add support for reduce/scan ops with cluster awareness - pvr: advertise subgroup arithmetic and clustered features - pco: add support for subgroup shuffle ops - pvr: advertise subgroup shuffle and shuffle relative features - pvr, pco: add support for VK_KHR_shader_subgroup_rotate - pvr: advertise VK_KHR_shader_subgroup_uniform_control_flow - pvr, pco: advertise support for VK_EXT_subgroup_size_control - pco: allow non-pure integer formats for image xchg atomics - pco: lower sysvals early for fragment shaders - pco: allow fence ops to be legalized if they come last in a block Skyth (1): - spirv2dxil: Replace UAV_FENCE_THREAD_GROUP usage with UAV_FENCE_GLOBAL. Sonny Jiang (2): - radeonsi: always set is_format_supported in screen create - radeonsi/vcn: Add vcn_5_0_2 support Stijn Tintel (1): - rocket: fix mmap leak in buffer map/unmap Stéphane Cerveau (1): - vulkan/video: Reject interlaced picture layout for H.264 baseline profile Suresh Guttula (1): - ac: Add vcn_5_3_0 support Sushma Venkatesh Reddy (2): - intel/perf: Add WCL OA support - intel/dev: Clamp PTL+ CS workgroup threads to 32 Tacodiva (1): - vulkan/runtime: Fix bad assumption in GetPipelineBinaryDataKHR Tanner Van De Walle (5): - draw: add lower-bound assert on shader_stage - gallium/u_blitter: add lower-bound assert on target - util/format: add lower-bound assert on format - dzn: silence PREfast C33010 warnings - nir/nir_builder: inline dst_bit_size calculation in assert Tapani Pälli (13): - intel/compiler: implement macl part of Wa_18035690555 - drirc: use anv_disable_drm_ccs_modifiers for any GTK version - drirc/anv: add flag to disable VK_EXT_subgroup_size_control - drirc: set anv_disable_subgroup_size_control for bg3 - anv: do not use resource barrier with split barriers - intel/dev: update mesa_defs.json from workaround database - iris: use INTEL_NEEDS_WA_14025112257 define for workaround - anv: use INTEL_NEEDS_WA_14025112257 define for workaround - anv: allocate tile sized temporary copy instead of whole size - anv: skip writing xfb buffer if we get null information - iris: align down the max_shader_buffer_size - anv: fix a null pointer access with isl_mod_info - anv: optimization for Wa_14025112257 case Thomas H.P. Andersen (5): - nvk: set queryResultStatusSupport - nvk: use the new generation script for drirc - nouveau/cubin: use libelf 64 bit instead of gelf - nvk: hide NVX_binary_import behind NVK_EXPERIMENTAL=dlss env var - nvk: add env var to allow backwards compat in dlss Thong Thai (25): - util: move u_stub to src/util, add u_stub_gfx_compute.h - util: allow for overriding u_stub tail - meson: update default build option for libva subproject - meson: check if video encoding support is to be built - frontends/va: decode only stubs - radeonsi: move si_get video functions to si_video - amd: make ac_ib_parser an amd tool build option - gallium/auxiliary/vl: Fix typo in cs_create_shader pseudo-code comment - pipe: Add PIPE_VIDEO_VPP_BLEND_MODE_PREMULTIPLIED_ALPHA - vl/video_buffer: Set alpha swizzle if format has alpha - gallium/video: Add enabled flag to vpp interface - gallium/vl: Implement compositor shader-based alpha blending - frontends/va: Enable shader-based alpha blending - amd: Build nir files only when with_gfx_compute - radeonsi: Remove ACO dependency for non-GFX/compute builds - nir: Only build NIR headers when with_gfx_compute is false - gallium/auxiliary: Simplify auxiliary for non-gfx/compute builds - meson: Make with_gfx_compute depend on video encode support - meson: Don't require libelf for radeonsi when with_gfx_compute is false - radeonsi: Allow call to stub'd si_init_gfx_context to continue - radeonsi: Store SQTT cb_id - radeonsi: Handle SQTT timestamps - radeonsi: Store SQTT device_id - radeonsi: Implement SQTT CB_START and CB_END - radeonsi: Setup SQTT sampling clocks Timothy Arceri (17): - glcpp: update out of date comment - glcpp: fix paste within macro function expansion - amd/radeonsi: dont clamp packed user varyings - mesa: fix typo in validation string - ac/nir/lower_tex_coord: update cursor when moving wqm coordinates - ac/nir/lower_tex_coord: basic lower tex coord test - mesa: flush bitmap cache when scissor box changes - nir: use the correct induction var when guessing loop iterations - glsl: allow uniform block layout qualifiers when SSBO enabled - util/u_range_remap: allow insert to truncate range - glsl: treat temp globals wrappers as roots when resolving function calls - zink: fix swap interval changes being dropped - nir/opt_dead_write_vars: handle memcpy_deref as reads - util: add Blockland workaround for crash - util/mesa: add workaround to zero invalidated buffers - util: add workaround for Riddick using round() in glsl 1.20 - llvmpipe: emit FS input vertex attributes in driver location order Timur Kristóf (10): - nir/divergence: Consider ACCESS_SMEM_AMD divergence across subgroups - nir/divergence: Consider uniformity of read_invocation accross subgroups - nir/divergence: Consider ttmp_register_amd and load_scalar_arg_amd as workgroup divergent - ac/nir: When loading an arg, assert that it's used - ac/nir: Fix SMEM workaround with emulated RT - radv: Wait for idle after every submission on GFX6-7 - radeonsi: Wait for shaders and flush L2 after every submission on GFX6-7 - Revert "radv: Mitigate GPU hang on Hawaii in Dota 2 and RotTR" - ac/nir/ngg: Remember if a mesh shader has non-API waves. - ac/nir/ngg: Use workgroup divergence analysis for mesh output counts. Tomeu Vizoso (5): - teflon/tests: avoid loading build-tree tensorflow-lite stub at runtime - teflon/tests: make tflite stubs fail loudly with diagnostics - teflon: remove synthetic model generation and flatbuffers dependency - ci: Remove flatbuffers from builds - teflon/tests: Remove leftover files from synthetic tests Toshinari Morikawa (2): - virgl: fix memory leak on shader translation - egl: avoid calling loader_get_driver_for_fd with fd = -1 Trigger Huang (12): - radv: supports protected memory allocation - radv: allow creation of protected queues - radv: support secure submission - radv: add protected type bits for memory requirements - radv: enable protected memory - radv: emulate MSRTSS via implicit MSAA resolve - radv/meta: thread separate src/dst sample counts through gfx copy - radv/meta: derive gfx copy dst sample count from the destination - radv/meta: add MSRTSS attachment replicate helper - radv: replicate MSRTSS attachments on LOAD_OP_LOAD - radv: handle VkSubpassResolvePerformanceQueryEXT - radv: enable VK_EXT_multisampled_render_to_single_sampled UMU618 (1): - venus: fix typo in vn_queue_submit_2_to_1 UMUTech (1): - wsi: correct the erroneous assertion Utku Iseri (1): - v3dv: close display_fd on incompatible_driver path Val Packett (3): - util: rust: align API with real eventfd capabilities - util: rust: Support detecting socket file descriptors - util: rust: Add a way to create a Tube from an existing OwnedFd Valentine Burley (102): - mr-label-maker: Label Collabora farm with driver tags - ci/zink/intel: Disable flaky TGL canvas_moire-v2 trace - zink/ci: Document recent flakes - anv/ci: Revert ADL VKCTS job to stable 6.17 kernel - zink/ci: Move Turnip flakes to correct list - tu/drm/virtio: Fix tu_wait_fence timeout handling - freedreno/drm/virtio: Fix wait_fence ret ordering - zink/ci: Remove Cezanne job - radv/ci: Add more ASAN VKCTS jobs on Cezanne - vulkan/android: Add deferred image helper - panvk: Use vk_android deferred image helper - vulkan/android: Add vk_android_import_anb_memory helper - vulkan: Query memory requirements in vk_android_import_anb_memory - tu: Implement deferred image creation for ANB and AHB - ci/crosvm: Sanitize CROSVM_RET in crosvm-runner.sh - tu: Fix D16 depth clear rounding mismatch in sysmem mode - panfrost/ci: Update kernel to pick up ZSTD support for ZRAM - venus/ci: Skip more robustness tests on ANV - tu: Move Android extensions into main list - tu: Add shared image support on Android - panfrost/ci: Move t860 jobs to nightly - panfrost/ci: Document recent g610 flake - ci/android: Remove SurfaceFlinger wait in get_surfaceflinger_pid - ci/android: Fix intermittent adb root failures - ci/android: Update Cuttlefish build - ci/deqp: Add Android WSI support - lavapipe/ci: Enable WSI testing on Android - turnip/ci: Enable WSI testing on Android - venus/ci: Enable WSI testing on Android - ci/android: Remove CtsDeqpTestCases from Android CTS - ci/deqp: Backport host_image_copy fix - ci/lava: Reduce LAVA job timeout to 20 minutes for Marge - mr-label-maker: Add rule for new trace replay config files - ci: Add missing rule for new trace replay config files - tu/autotune: Clear active_batches before history objects are freed - ci/deqp: Backport validation error fix - ci/deqp: Backport landed patch - ci/deqp: Rewrite headless Android WSI patch - venus/ci: Skip more even more robustness and synchronization2 tests on ANV - ci: Disable debian-riscv64 - panvk/ci: Mark dEQP-VK.subgroups.* as flaky on G925 - ci: Bump ci-deb-repo revision to update aapt - ci/android: Update Android CTS to android-cts-16.0_r5 - ci/android: Add arm64 support for Android CTS - turnip/ci: Add nightly Android CTS job - tu: Disable -Wmisleading-indentation when compiling with GCC - panvk: Fix ignored qualifier warnings - meson: Add Soong compatibility compiler flags to Vulkan drivers - tu/kgsl: Fix memory type support detection for unsupported flags - tu: Merge tu_image_init and tu_image_update_layout - venus/ci: Widen the ANV skips - tu: Advertise VK_KHR_internally_synchronized_queues - turnip/ci: Update ANGLE trace checksum - panfrost/ci: Switch traces over to gpu-trace-perf - tu: Fix vk_queue leak on submitqueue creation failure - vulkan/queue: Add common queue emulation support - tu: Emulate second graphics queue for skiavk on Android - tu/ci: Add coverage for emulated second graphics queue - ci/android: Update Cuttlefish build - anv/ci: Disable anv-adl-vk job - venus/ci: Move pre-merge ANV coverage from Comet Lake to Alder Lake - venus/ci: Retire Intel Comet Lake runner - venus/ci: Revert ADL jobs to stable 6.17 kernel - intel/gen: Explicitly declare gen_opcodes_private.h dependency - perfetto: Centralize perfetto header include in u_perfetto.h - bin: Expose drm-shim in meson devenv - doc/ci: Add drm-shim CI reproduction guide - zink/ci: Increase zink-lavapipe parallelism - virgl/ci: Retire disabled virgl-iris jobs - virgl/ci: Retire disabled android-virgl-llvmpipe - pipe-loader: Enable null winsys on Android - zink/ci: Remove zink-anv-cml-asan job - intel/ci: Remove nightly CML jobs, retire runner - intel/ci: Increase iris-apl-egl parallelism - zink/ci: Drop old VVL filters - panfrost/ci: Fix typo in .panfrost-vk-manual-panthor-rules template - tu: Fix uninitialized gmem_offset when a GMEM layout is impossible - ci: Update kernel to Linux 7.1.2 - panfrost/ci: Use Linux 7.1 kernel for more jobs - tu: Fix capture/replay with sampler custom border color - ci/lava: Uprev lava-job-submitter - drm-shim/freedreno: Add support for Adreno 610 - drm-shim/freedreno: Shim perf counter config ioctl - bin/drm-shim: Add more freedreno GPUs - tu: Disable VK_EXT_extended_dynamic_state2 patch control points on A702 - zink: Gate tess/geom barrier stages on feature support - ci: Bump ci-deb-repo revision to update vulkan-loader - ci/deqp: Backport landed Android WSI patch - ci/deqp: Update VK CTS to 1.4.6.1 - ci/deqp: Backport -frounding-math default for GCC builds - tu: Implement VK_EXT_primitive_restart_index - ir3: Fix ballot_components for subgroups smaller than 32 - ir3: Derive max_variable_workgroup_size from device limits - ir3: Compute subgroup_size from threadsize_base - tu: Use computed subgroup size - panvk/ci: Update expectations for g610-vk-asan following VK CTS uprev - panvk: Disable VK_KHR_internally_synchronized_queues on Vulkan 1.0 - tu: Report correct maxFragmentInputComponents limits - zink: Use ShaderLayer capability for gl_Layer when available - freedreno: Increase reg_size_vec4 for A702 - freedreno: Fix VPC_RAST_STREAM_CNTL register layouts - ci/piglit: Switch all trace jobs to surfaceless+gbm Vincent Cloutier (2): - etnaviv: use buffer resource accessor for indirect draws - etnaviv: support native bitfield extract/reverse/count ops Vinson Lee (14): - st/mesa: fix implicit conversion warning in st_atom_framebuffer - vulkan/screenshot-layer: initialize info to NULL - gfxstream: codegen: drop const from let-param scalar cast - mesa/main: cast GLhandleARB to unsigned int in api trace - ethosu/mlw_codec: silence warnings in the vendored Regor encoder - ethosu/mlw_codec: silence -Wunused-const-variable in vendored encoder - ethosu: use FALLTHROUGH macro in ethosu_emit_operation_accesses - radeonsi: remove duplicate '.bpp' initializer in si_sdma_copy_image - gfxstream: link goldfish_address_space against perfetto - util/tests: replace sprintf with snprintf in cache tests - util/tests: fix unused variable warnings in cache List test - vulkan/screenshot-layer: replace itoa/sprintf with snprintf - vulkan/screenshot-layer: fix globalLock mutex leak - util/tests: silence unused iterator warning in sparse_bitset_test Virgile Bello (3): - microsoft/compiler: sink load_invocation_id in TCS split even for single-use - microsoft/compiler, d3d12: flip tess winding at caller, not in nir_to_dxil - microsoft/compiler, d3d12: preserve TCS outputs and pad TES inputs for cross-stage signature matching Vishnu Vardan (27): - mesa/st: remove redundant has_stencil_export from st_context - mesa/st: remove redundant astc_void_extents_need_denorm_flush from st_context - mesa/st: remove redundant has_shareable_shaders from st_context - mesa/st: remove redundant needs_texcoord_semantic from st_context - mesa/st: remove emulate_gl_clamp from st_context - mesa/st: remove has_time_elapsed from st_context - mesa/st: remove has_multi_draw_indirect from st_context - mesa/st: remove has_indirect_partial_stride from st_context - mesa/st: remove has_occlusion_query from st_context - mesa/st: remove has_single_pipe_stat from st_context - mesa/st: remove has_pipeline_stat from st_context - mesa/st: remove has_indep_blend_enable from st_context - mesa/st: remove has_indep_blend_func from st_context - mesa/st: remove can_dither from st_context - mesa/st: remove lower_flatshade from st_context - mesa/st: remove lower_alpha_test from st_context - mesa/st: remove lower_two_sided_color from st_context - mesa/st: remove lower_ucp from st_context - mesa/st: remove prefer_real_buffer_in_constbuf0 from st_context - mesa/st: remove has_conditional_render from st_context - mesa/st: remove lower_rect_tex from st_context - mesa/st: remove allow_st_finalize_nir_twice from st_context - mesa/st: remove can_bind_const_buffer_as_vertex from st_context - mesa/st: remove validate_all_dirty_states from st_context - mesa/st: remove can_null_texture from st_context - mesa/st: remove redundant has_hw_atomics from st_context - anv/rt: reorder encode_internal_node to only process valid children Vlad Zahorodnii (1): - wsi/wayland: Add support for wl_fixes.ack_global_remove Wig Cheng (2): - rocket: pad weight packing input channels to FEATURE_ATOMIC_SIZE - rocket: compute element-wise ADD requant instead of LUT Wujian Sun (2): - mesa: Fix clipping order in _mesa_clip_blit() - mesa: Allow GL_SRGB_ALPHA_EXT as color-renderable when EXT_sRGB is supported Xinju Li (1): - nir: resolve functions: only resolve functions that are reachable from main Yannis Juglaret (1): - nouveau: fix data race in nouveau_fence_ref Yiwei Zhang (146): - venus: adopt vk_android_init_deferred_image - venus: adopt vk_android_get_ahb_layout - venus: refactor vn_android_get_wsi_memory to return VkDeviceMemory - venus: adopt common vk_image::anb_memory - venus: adopt common ANB helpers - panvk: adopt common ANB helpers - lvp/android: use common ANB implementations - util/android_stub: drop legacy atrace - panvk: drop panvk_android_create_deferred_image - util/os_misc: use ndk api __system_property_get - egl/android: use ndk api __system_property_get - android_stub: drop cutils/properties dependency - CODEOWNERS: update owners for Android components - util/os_misc: use stable NDK __android_log_write helper - intel: use stable NDK __android_log_print helper - broadcom: remove unused Android log utils - android_stub: purge unused log utils - ci: uprev virglrenderer - android_stub: fix update-android-headers.sh for libbacktrace - android_stub: fix libhardware source include path - android_stub: avoid vending in unused headers - android_stub: sync Android 16 headers - pan/nir/tex: use unsigned type for texture op lod_or_fetch - venus: fix a renderer side queue timeline bound race - panvk: fix to report device memory with heapIndex - tu: fix to report device memory with heapIndex - venus: update create_from_device_memory to take a cmd payload - venus: let resource_create_blob wait for mem alloc - venus: fix unbound malloc leak in vn_ring_get_submits - anv: fix lock scope in anv_ensure_fp64_shader - anv: amend missing shader dump finish upon device destruction - venus: amend roundtrip between fence submit and wait idle - panvk: fix plane indexing for afbc image subres layout - pan: handle downscaling of plane view for multiplanar yuv textures - panvk: reject interleaved_64k for multiplanar yuv - panvk: avoid separate reconstruction filter for YUV texturing - panvk: use vk_component_mapping_to_pipe_swizzle - panvk: apply YUV swizzle to the view swizzle for YUV texturing - panvk: override default chroma siting for YUV texturing - panvk: enforce strict import for yuv images - panvk: add get_pan_image_props helper - panvk: use binding layout textures_per_desc to write image view descs - panvk: add pan_texture_get_payload_alignment to help with tex emit - panvk: add and use panvk_image_get_tex_count helper - panvk: properly set up image and view planes for YUV texturing - panvk: lower YUV texturing to do SW CSC - pan: add 8bit multi-planar 420 and 422 format for Vulkan - panvk: enable 8bit multiplanar YUV formats on v9+ to v13 - panvk: add P010 native YUV support - vulkan/android: force linear for mutable format - venus/wsi: skip VkPresentRegionsKHR when pRegions is NULL - venus: avoid touching sfb dst slot upon resume - venus: always check device lost on sfb warn order - venus: rename vn_semaphore_feedback_cmd to vn_sync_feedback_cmd - venus: move sfb helpers into vn_feedback - venus: wrap sfb cmd preparation with vn_sync_feedback_command - venus: extract sync feedback host write and query - venus: move sfb suspend resume handling over to vn_feedback - venus: add vn_sync_feedback_enabled - venus: always check device lost on ffb warn order - venus: migrate ffb to use vn_sync_feedback - venus: recycle fence sfb in post submission - venus: drop vk_xwayland_wait_ready override - v3dv: drop vk_xwayland_wait_ready - hasvk: drop vk_xwayland_wait_ready - vulkan/wsi/util: purge vk_xwayland_wait_ready - venus/virtgpu: amend a missing sim mutex init - venus/virtgpu: drop obsolete SIMULATE_BO_SIZE_FIX - venus/virtgpu: implicit fencing is gone - venus/virtgpu: simplify to drop virtgpu_sync - venus: refactor vn_renderer_submit to only take a single batch - venus/virtgpu: merge SIMULATE_SUBMIT into SIMULATE_SYNCOBJ - venus/virtgpu: drop signaled_fd - venus/virtgpu: drop cpu sync timeout - venus/virtgpu: drop wait available - venus/virtgpu: use STACK_ARRAY for syncobj handles - venus/virtgpu: split out sim_syncobj - venus/virtgpu: refactor sim_submit - venus/virtgpu: use uAPI to signal syncobjs - venus/virtgpu: adopt u_sync_provider - venus/virtgpu: use virtgpu syncobj on supported kernels - venus/virtgpu: avoid pretending timeline sync support - venus/virtgpu: sim_syncobj to implement u_sync_provider - venus/virtgpu: simplify sim_syncobj - venus/virtgpu: refactor virtgpu_submit - venus/virtgpu: flatten all the virtgpu_ioctl_syncobj wrappers - venus: properly clean up driver internal sim syncobj allocs - venus: fix imported sync fence payload reset upon export - venus: avoid renderer semaphore wait upon temp payload export - venus: refactor external fence and semaphore advertisment - venus: drop obsolete zink performance workaround - venus: track can_feedback in struct vn_queue_submission - venus: purge sync feedback for sparse binding - venus: split fence/semaphore/event commands to vn_sync.(c|h) - venus: refactor imported semaphore check and wait - venus/wsi: refactor args of vn_wsi_fence_wait and vn_wsi_flush - venus: queue submit to take a single batch - venus: sparse binding to use vn_queue_submit - venus: simplify vn_queue_submission to handle single batch - venus: simplify feedback cmd setup - venus: simplify vn_queue_submission_alloc_storage - venus: extract pNext chain fixup out from feedback cmds init - venus: prepare to handle sparse binding batch interception - venus: properly drop imported semaphores from submission - venus: relax SYNC_FD semaphore import requirement for WSI - venus/renderer: improve renderer backend init logs - venus: recycle idle sfb cmds only after async wait - venus: properly check sync2 enablement - venus: host image copy to scrub present_src layout if needed - venus: clean up sync2 treatment leftovers - venus: simplify sync feedback tracking - venus: update pnext fix tracking - venus: explicitly track if need to fix batch - venus: flatten sync feedback cmd counting - venus: deprecate fence feedback - venus: split ring submission to vn_queue_submission_do_submit - venus: skip empty batch submission - venus: drop vn_sync_payload_external from submission tracking - venus: extract vn_timeout_to_poll_timeout - venus/virtgpu: only signal non-zero initial value - venus/virtgpu/vtest: drop initial value from sync reset - venus: prepare for VN_SYNC_TYPE_SYNC - venus: migrate fence over to VN_SYNC_TYPE_SYNC - venus: relax SYNC_FD fence export requirement - venus: deprecate queue idle wait workaround - venus: rename to be explicit about sync fd semaphore - venus: vn_semaphore_(is|wait)_sync_fd to support VN_SYNC_TYPE_SYNC - venus: rename existing queue submission wait semaphore tracking - venus: count and prepare storage to scrub SYNC_FD signal semaphore - venus: extract syncs and scrub SYNC_FD signal semaphores - venus: ensure renderer sync fence is submitted between queue batches - venus: migrate SYNC_FD semaphore over to VN_SYNC_TYPE_SYNC - venus: relax SYNC_FD semaphore export requirement - venus/virtgpu: hide has_timeline_sync behind a new perf option - venus/vtest: advertise timeline syncobj support - venus: drop vn_renderer_sync_flags - venus: add VN_SYNC_TYPE_TIMELINE_SYNC - venus: track queue internal array index within vn_device::queues - venus: count and init syncs from timeline semaphores - venus: implement host signal for TIMELINE_SYNC - venus: implement counter query for TIMELINE_SYNC - venus: add vn_wait_semaphores_legacy for legacy wait - venus: implement semaphore wait for TIMELINE_SYNC - venus: migrate timeline semaphore to VN_SYNC_TYPE_TIMELINE_SYNC - venus: document timeline semaphore implementation - venus: ensure cached vn_ring_submit batches are bounded Yogesh Mohan Marimuthu (5): - ac,radeonsi,radv: add has_desc_resource_level var instead of gfx_level check - radv: Program RESOURCE_LEVEL bit in descriptor for dgc - amd: add initial code for gfx1156 - ac: set has_smem_with_null_prt_bug to false for gfx1156 - ac: set has_desc_resource_level to true for gfx1156 You, Min-Hsuan (1): - amd/vpelib: fix FROD alignment handling after interface change Zan Dobersek (5): - fd: add a8xx perfcntr countables - tu: only support userspace-managed perfcounters on a7xx and earlier - tu/a8xx: remove enforced TU_DEBUG_FLUSHALL - tu/kgsl: initialize dump bo state in kgsl_bo_init sooner - fd: lrz_block in fdl6_lrz_layout_init() should not be static Zeyang Lyu (1): - radv: Add base array layer to htile offset Zhao, Jiali (2): - amd/vpelib: revert predication fix - amd/vpelib: fix HDR external monitor video black content ZhengMing (1): - vulkan/wsi/win32: Prefer the more popular surface format on Windows Zoltán Böszörményi (1): - radv: Advertise msrtss in features.txt adrian baker (4): - jay: add bfloat16 support - jay: fix jay bf16 comment formatting - nir: remove incorrect algebraic properties from intel mixed bf ops - jay: add geometry shader support. gyeyoung (2): - panvk: fix flags2-only bit leak in legacy format features - panvk: report DRM format modifiers through List2EXT gyeyoung baek (1): - rocket: drop wrong assert(input_op_1) in ADD fuse path hmtheboy154 (15): - pvr: add support for driconf for the Vulkan driver - driconf: Add an option to override Vulkan's deviceName - anv: driconf: Add an option to override Vulkan's deviceName - hasvk: driconf: Add an option to override Vulkan's deviceName - nvk: driconf: Add an option to override Vulkan's deviceName - radv: driconf: Add an option to override Vulkan's deviceName - venus: driconf: Add an option to override Vulkan's deviceName - v3dv: driconf: Add an option to override Vulkan's deviceName - lvp: add support for driconf - lvp: driconf: Add an option to override Vulkan's deviceName - tu: driconf: Add an option to override Vulkan's deviceName - panvk: driconf: Add an option to override Vulkan's deviceName - pvr: driconf: Add an option to override Vulkan's deviceName - dzn: driconf: Add an option to override Vulkan's deviceName - hk: driconf: Add an option to override Vulkan's deviceName hwandy (1): - Revert "intel/decoder: make libvulkan_intel to depend on stub decoder when buildtyle=release." inspector-ambitious (1): - loader: fix loader_open_render_node_platform_devices result allocation jglrxavpok (1): - RADV: Add object names inside address binding report and vm_fault jiajia Qian (6): - rusticl: extract tokenize() and fix UTF-8 handling in compile options - rusticl: add LinkOptions struct with validation - rusticl: validate build/compile options before passing to backend - rusticl: validate input_programs binary type in clLinkProgram - ci/panfrost: add piglit OpenCL testing for G610 - rusticl/device: use OpenCL spec minimum for mem_base_addr_align jinmiliu (2): - mesa/st: Set protected content context flag based on pipe context attributes - radeonsi: enable protected context support for Android johniyoods (1): - egl/dri2: require valid render fd before advertising EGL_WL_bind_wayland_display jyotiranjan (1): - radv/sqtt: forward zero-submit-count vkQueueSubmit2 for SQTT capture ljohnson (1): - venus/wsi: deep copy pRectangles when cloning presentation info llyyr (2): - radeonsi: don't init screen state functions twice - vulkan/wsi/wayland: use mtx helpers in wait_for_present2 nyanmisaka (1): - intel/dev: update PTL device names sergiuferentz (1): - gfxstream: Prevent LINUX_GUEST_BUILD from being added to android platforms squidbus (89): - kk: Use device limits for buffers and compute shared memory. - kk: Enable VK_AMD_shader_image_load_store_lod - kk: Update dynamic depth stencil state regardless of set attachments. - kk: Add type inference for additional built-in intrinsics. - kk: Fix VK_CULL_MODE_FRONT_AND_BACK with points and lines. - asahi,nir: Move asahi dynamic clipz pass to common. - kk: Add support for VK_EXT_depth_clip_control. - kk: Fix issues with maximal reconvergence - kk: Enable VK_EXT_extended_dynamic_state3 - kk: Enable VK_EXT_buffer_device_address - kk: Enable VK_(EXT/KHR)_global_priority and VK_EXT_global_priority_query - kk: Workaround for GPU capture under Rosetta 2. - nir: Only attempt subgroups lower_boolean_reduce for single component. - kk: Expand workaround 3 to cover general use of ballot/vote ops - kk: Fix emitting negative infinity - kk: Support subgroup rotate ops - kk: Enable remaining subgroup operations - kk: Fix geometry unroll for list primitives. - kk: Support VK_(KHR/EXT)_index_type_uint8 - kk: Enable VK_EXT_multi_draw - kk: Split per-draw data to separate binding - kk: Fix handling of sample mask and sample rate shading - kk: Support VK_EXT_post_depth_coverage - kk: Support robustBufferAccess2 - kk: Support nullDescriptor - kk: Enable VK_(EXT/KHR)_robustness2 and VK_EXT_pipeline_robustness - kk: Enable VK_(EXT/KHR)_line_rasterization - kk: Support shaderCullDistance - kk: Query device for supported sample counts - kk: Complete VK_EXT_memory_budget - kk: Create image layout from vk_image - kk: Implement index buffer robustness for BindIndexBuffer2 - kk: Handle accurate OpSMod and default point size requirements - kk: Disable A8_UNORM format - kk: Support device without queue - kk: Fix image copies for depth/stencil<->color and differing subresources - kk: Support new query pool and dynamic rendering flags - kk: Enable maintenance extensions through VK_KHR_maintenance10 - kk: Separate linear and GPU optimized image layout properties - kk: Support VK_EXT_host_image_copy - kk: Support VK_KHR_shader_fma - kk: Support attachment feedback loop extensions - kk: Support VK_KHR_unified_image_layouts - kk: Fix some missed NIR debug asserts - kk: Allocate temporary command memory from pool - kk: Fix pre-compiled compute grid size - kk: Enable code formatting enforcement - poly: Refactor poly_unroll_restart for general purpose unrolling - poly: Fix range used for index unroll bounds checks - kk: Fix compute system value and algebric lowering in pre-compiles - kk: De-duplicate geometry unroll logic - kk: Support VK_KHR_shader_untyped_pointers - kk: Refactor multi-draws and predicates into kk_draw_data - kk: Enable VK_EXT_nested_command_buffer - kk: Support VK_EXT_conditional_rendering - kk: Fix precomp data buffer alignment - kk: Sanitize image copy through buffer extents - kk: Do not use image-to-image copies for 1D compressed textures - kk: Handle index robustness for fully bound buffers manually - kk: Accurately declare supported samples in image format properties - kk: Support VK_EXT_vertex_attribute_robustness - kk: Support VK_EXT_blend_operation_advanced - kk: Support VK_EXT_custom_resolve - kk: Support VK_EXT_primitive_restart_index - kk: Support VK_EXT_primitive_topology_list_restart - kk: Fence read-write images after write - kk: Support VK_IMAGE_CREATE_BLOCK_TEXEL_VIEW_COMPATIBLE_BIT - kk: Support VK_EXT_external_memory_host - kk: Advertise additional tessellation dynamic state - kk: Perform sink-and-move of instructions - kk: Do not force render image view to all subresources - kk: Fix divide by 0 in non-indexed draw unroll - kk: Support VK_EXT_sample_locations - kk: Remove unused deprecated Metal APIs - kk: Ensure some vertex lowerings happen on hardware stage - kk: Enable shaderTessellationAndGeometryPointSize - kk: Migrate to Metal 4 pipelines - kk: Work around crash with multiple concurrent MTL4Compiler - wsi/metal: Support HDR10 color spaces - kk,wsi/metal: Support VK_EXT_hdr_metadata - kk,wsi/metal: Support VK_(KHR/EXT)_swapchain_maintenance1 - kk: Respect precomp-compiler options when setting up kk_clc - kk: Implement draw-related commands using device addresses - kk: Work around Metal index robustness gaps - kk: Pre-declare texture SSA variables - kk: Use safe math for nir_fp_no_reassoc - kk: Enable shaderRoundingModeRTEFloat16/32 - kk: Only support 1 sample for storage images - kk: Remove deprecated MTL4CommandQueueErrorDeviceRemoved utzcoz (4): - gfxstream: Validate guest mapped-memory ranges in flush/invalidate - ci/amd: enable ACO validation on radeonsi jobs - radeonsi: convert gather_instruction to nir_function_instructions_pass - virtio: magma-gpu-rs: accept a null device in virtgpu_kumquat_finish xueyuli2 (1): - amd/virtio: fix bo use-after-free race condition in amdvgpu_bo_free yserrr (4): - llvmpipe: fix UB and incorrect value in compute caps shift - v3d: fix stencil blit layer selection - v3d: lower more 64-bit integer operations - v3d: remove duplicate util_blitter_save_so_targets() call