shadPS4

mirror of https://github.com/shadps4-emu/shadPS4.git synced 2025-06-26 12:26:18 +00:00

Author	SHA1	Message	Date
squidbus	3f949d2b6c	amdgpu: Handle 32-bit Unorm formats. (#2974 ) Some checks are pending Build and Release / reuse (push) Waiting to run Details Build and Release / clang-format (push) Waiting to run Details Build and Release / get-info (push) Waiting to run Details Build and Release / windows-sdl (push) Blocked by required conditions Details Build and Release / windows-qt (push) Blocked by required conditions Details Build and Release / macos-sdl (push) Blocked by required conditions Details Build and Release / macos-qt (push) Blocked by required conditions Details Build and Release / linux-sdl (push) Blocked by required conditions Details Build and Release / linux-qt (push) Blocked by required conditions Details Build and Release / linux-sdl-gcc (push) Blocked by required conditions Details Build and Release / linux-qt-gcc (push) Blocked by required conditions Details Build and Release / pre-release (push) Blocked by required conditions Details	2025-05-22 03:16:20 -07:00
squidbus	f4eb0b9b9e	shader_recompiler: Fix buffer type reading from step rate attribute. (#2973 )	2025-05-22 03:03:24 -07:00
Lander Gallastegi	1b952bf173	ReadConst debug msg (#2964 ) Some checks are pending Build and Release / reuse (push) Waiting to run Details Build and Release / clang-format (push) Waiting to run Details Build and Release / get-info (push) Waiting to run Details Build and Release / windows-sdl (push) Blocked by required conditions Details Build and Release / windows-qt (push) Blocked by required conditions Details Build and Release / macos-sdl (push) Blocked by required conditions Details Build and Release / macos-qt (push) Blocked by required conditions Details Build and Release / linux-sdl (push) Blocked by required conditions Details Build and Release / linux-qt (push) Blocked by required conditions Details Build and Release / linux-sdl-gcc (push) Blocked by required conditions Details Build and Release / linux-qt-gcc (push) Blocked by required conditions Details Build and Release / pre-release (push) Blocked by required conditions Details	2025-05-21 19:08:27 +03:00
Marcin Mikołajczyk	6abda17532	Avoid post-increment of SGPR in S_*_LOAD_DWORD (#2928 )	2025-05-13 14:31:14 -07:00
Marcin Mikołajczyk	484fbcc320	Handle -1 as V_CMP_NE_U64 argument (#2919 )	2025-05-13 13:19:56 -07:00
squidbus	8909d9bb89	shader_recompiler: Always mark buffers as storage buffers. (#2914 )	2025-05-12 10:46:40 -07:00
Missake212	8d7cbf9943	Adding opcode IMAGE_SAMPLE_B_O (#2894 ) * Adding opcode IMAGE_SAMPLE_B_O: * fix clang (my first time !)	2025-05-09 09:01:34 -07:00
squidbus	b130fe6ed5	vulkan: Handle incompatible depth format using null binding. (#2892 ) Co-authored-by: kalaposfos13 <153381648+kalaposfos13@users.noreply.github.com>	2025-05-09 08:43:20 -07:00
MajorP93	c7fb3ebd93	shader_recompiler: Widen num_conversion bitfield (#2886 ) Some checks are pending Build and Release / reuse (push) Waiting to run Details Build and Release / clang-format (push) Waiting to run Details Build and Release / get-info (push) Waiting to run Details Build and Release / windows-sdl (push) Blocked by required conditions Details Build and Release / windows-qt (push) Blocked by required conditions Details Build and Release / macos-sdl (push) Blocked by required conditions Details Build and Release / macos-qt (push) Blocked by required conditions Details Build and Release / linux-sdl (push) Blocked by required conditions Details Build and Release / linux-qt (push) Blocked by required conditions Details Build and Release / linux-sdl-gcc (push) Blocked by required conditions Details Build and Release / linux-qt-gcc (push) Blocked by required conditions Details Build and Release / pre-release (push) Blocked by required conditions Details We do this in order to be able to actually fit in all possible values from AmdGpu::NumberConversion. Fixes gcc compiler warnings: warning: ‘Shader::PsColorBuffer::num_conversion’ is too small to hold all values of ‘enum class AmdGpu::NumberConversion’	2025-05-06 17:11:32 -07:00
Mahmoud Adel	b0e4e87ff3	Implement SnormNz conversion (#2841 ) * + * + * Unpack Snorm 2x16 * + * SintToSnormNz * all is broken ig.... * review changes * my stupid ass messed all while trying to resolve the conflicts.. * + * + * fix rebase * clang-format fix (1) * clang-format fix (2) --------- Co-authored-by: squidbus <175574877+squidbus@users.noreply.github.com>	2025-05-01 02:12:15 -07:00
squidbus	5fd5b62539	shader_recompiler: Few fixes for buffer number conversions. (#2869 ) * liverpool: Pass correct color buffer number type for conversion mapping. * shader_recompiler: Apply number conversion to vertex inputs.	2025-04-30 20:46:16 -07:00
squidbus	10b24d04bc	fix: Add new image atomic instructions to relevant lists.	2025-04-30 17:55:50 -07:00
squidbus	ede60e8f7f	fix: Do not declare atomic float capability when not supported.	2025-04-30 11:43:51 -07:00
Marcin Mikołajczyk	c08f92aca1	Implement IMAGE_ATOMIC_FMIN and IMAGE_ATOMIC_FMAX for 32bit floats (#2820 ) * Implement IMAGE_ATOMIC_FMIN and IMAGE_ATOMIC_FMAX for 32bit floats * Handle missing VK_EXT_shader_atomic_float2	2025-04-30 11:42:08 -07:00
squidbus	a3bbf2274f	fix: Mistake in store bounds check index.	2025-04-30 11:39:42 -07:00
squidbus	81fa9b7fff	shader_recompiler: Add lowering pass for when 64-bit float is unsupported. (#2858 ) * shader_recompiler: Add lowering pass for when 64-bit float is unsupported. * shader_recompiler: Fix PackDouble2x32/UnpackDouble2x32 type. * shader_recompiler: Remove extra bit cast implementations.	2025-04-28 00:04:16 -07:00
squidbus	385c5a4507	fix: Add missing OpSelectionMerge in bounds check.	2025-04-27 21:53:36 -07:00
squidbus	b505829e16	lower_buffer_format_to_raw: Fix handling of format remapping. (#2857 )	2025-04-27 16:52:52 -07:00
baggins183	e816bc4b99	Use GetSrc in VALU insts instead of assuming vector reg (was vcc_lo) (#2845 ) * Use GetSrc in v_add_i32 instead of assuming vector reg (was vcc_lo) * some other cases	2025-04-25 19:44:03 -07:00
squidbus	aeee7706ee	renderer_vulkan: Restore Vulkan version to 1.3 (#2827 ) Co-authored-by: georgemoralis <giorgosmrls@gmail.com>	2025-04-23 13:28:31 +03:00
Fire Cube	0c86c54d48	Implement SET_PC_B64 instruction (#2823 ) * basic impl * minor improvements * clang * more clang * improvements requested by squidbus	2025-04-21 14:25:15 -07:00
Dmugetsu	ddc05e8a5f	Implementing DS_SUB_U32, DS_INC_U32, DS_DEC_U32. (#2797 ) * Implementing DS_SUB_U32, DS_INC_U32, DS_DEC_U32, DS_WRITE_SRC2_B32, DS_WRITE_SRC2_B64. * Added ir instructions for new opcodes. Removing Write implementations. Maping operation S_BFE_I32 as it was added in translate but wasnt pointing to anything. * Suggestions	2025-04-16 17:56:27 -07:00
squidbus	62a4182aca	shader_recompiler: Fill in IMAGE_GATHER4_* variants in table. (#2798 )	2025-04-16 17:35:14 -07:00
Fire Cube	aa8dab5371	shader_recompiler: Implement S_SUBB_U32 instruction (#2796 ) * add S_SUBB_U32 instruction * add missing case * move case to match enum	2025-04-16 14:24:18 -07:00
Missake212	243ee04b1c	Implement DS_READ2ST64_B64 (#2795 )	2025-04-16 09:54:05 -07:00
squidbus	52ab1ed04b	shader_recompiler: Implement S_FLBIT_I32_B32 and V_MUL_HI_I32. (#2793 )	2025-04-16 18:08:09 +03:00
squidbus	4bea00135d	resource_tracking_pass: Add heuristic to detect incorrectly tracked buffer sharp. (#2786 )	2025-04-14 20:58:49 -07:00
squidbus	bec1b9056f	shader_recompiler: Misc shader fixes. (#2781 ) * shader_recompiler: Fix frexp exponent type. * shader_recompiler: Implement V_CMP_CLASS_F32 negative class mask. * shader_recompiler: Define operands for DS_ORDERED_COUNT.	2025-04-13 23:46:30 -07:00
squidbus	afd0251dd2	shader_recompiler: Use VK_AMD_shader_trinary_minmax when available. (#2739 ) * shader_recompiler: Use VK_AMD_shader_trinary_minmax when available. * shader_recompiler: Simplify signed/unsigned trinary instruction variants.	2025-04-02 23:36:54 +03:00
Stephen Miller	b0a12c02e1	Add S_SETPRIO to EmitFlowControl (#2727 ) By squidbus' suggestion, I've added a warning log for this case.	2025-03-31 12:55:14 +03:00
TheTurtle	1f9ac53c28	shader_recompiler: Improve divergence handling and readlane elimintation (#2667 ) * control_flow_graph: Improve divergence handling * recompiler: Simplify optimization passes Removes a redudant constant propagation and cleans up the passes a little * ir_passes: Add new readlane elimination pass The algorithm has grown complex enough where it deserves its own pass. The old implementation could only handle a single phi level properly, however this one should be able to eliminate vast majority of lane cases remaining. It first performs a traversal of the phi tree to ensure that all phi sources can be rewritten into an expected value and then performs elimintation by recursively duplicating the phi nodes at each step, in order to preserve control flow. * clang format * control_flow_graph: Remove debug code	2025-03-23 00:35:42 +02:00
TheTurtle	2a05af22e1	emit_spirv: Fix comparison type (#2658 )	2025-03-19 23:20:00 +02:00
baggins183	c59d5eef45	Specialize vertex attributes on dst_sel (#2580 ) * Specialize vertex attributes on dst_sel * compare vs attrib specs by default, ignore NumberFmt when vertex input dynamic state is supported * specialize data_format when attribute uses step rates * use num_components in data fmt instead of fmt itself	2025-03-02 19:17:11 -08:00
Missake212	a583a9abe0	fixes to get in game (#2583 )	2025-03-02 22:13:23 +02:00
baggins183	7a4244ac8b	Misc Cleanups (#2579 ) -dont do trivial phi removal during SRT pass, that's now done in ssa_rewrite -remove unused variable when walking tess attributes -fix some tess comments	2025-03-02 21:52:32 +02:00
Missake212	76483f9c7b	opcode implementation test (#2567 )	2025-03-02 21:31:49 +02:00
Paris Oplopoios	fd3bfdae80	Implement some RDNA flags (#2510 )	2025-02-24 19:52:57 +02:00
TheTurtle	76b4da6212	video_core: Various small improvements and bug fixes (#2525 ) * ir_passes: Add barrier at end of block too * vk_platform: Always assign names to resources * texture_cache: Better overlap handling * liverpool: Avoid resuming ce_task when its finished * spirv_quad_rect: Skip default attributes Fixes some crashes * memory: Improve buffer size clamping * liverpool: Relax binary header validity check * liverpool: Stub SetPredication with a warning * Better than outright crash * emit_spirv: Implement round to zero mode * liverpool: queue::pop takes the front element * image_info: Remove obsolete assert The old code assumed the mip only had 1 layer thus a right overlap could not return mip 0. But with the new path we handle images that are both mip-mapped and multi-layer, thus this can happen * tile_manager: Fix size calculation * spirv_quad_rect: Skip default attributes --------- Co-authored-by: poly <47796739+polybiusproxy@users.noreply.github.com> Co-authored-by: squidbus <175574877+squidbus@users.noreply.github.com>	2025-02-24 14:31:12 +02:00
squidbus	9424047214	shader_recompiler: Proper support for inst-typed buffer format operations. (#2469 )	2025-02-21 03:01:18 -08:00
¥IGA	8447412c77	Bump to Clang 19 (#2434 )	2025-02-18 15:55:13 +02:00
squidbus	fd3d3c4158	shader_recompiler: Implement AMD buffer bounds checking behavior. (#2448 ) * shader_recompiler: Implement AMD buffer bounds checking behavior. * shader_recompiler: Use SRT flatbuf for bounds check size. * shader_recompiler: Fix buffer atomic bounds check. * buffer_cache: Prevent false image-to-buffer sync. Lowering vertex fetch to formatted buffer surfaced an issue where a CPU modified range may be overwritten with stale GPU modified image data. * Address review comments.	2025-02-17 16:13:39 +02:00
TheTurtle	82cacec8eb	shader_recompiler: Remove special case buffers and add support for aliasing (#2428 ) * shader_recompiler: Move shared mem lowering into emitter * IR can be quite verbose during first stages of translation, before ssa and constant prop passes have run that drastically simplify it. This lowering can also be done during emission so why not do it then to save some compilation time * runtime_info: Pack PsColorBuffer into 8 bytes * Drops the size of the total structure by half from 396 to 204 bytes. Also should make comparison of the array a bit faster, since its a hot path done every draw * emit_spirv_context: Add infrastructure for buffer aliases * Splits out the buffer creation function so it can be reused when defining multiple type aliases * shader_recompiler: Merge srt_flatbuf into buffers list * Its no longer a special case, yay * shader_recompiler: Complete buffer aliasing support * Add a bunch more types into buffers, such as F32 for float reads/writes and 8/16 bit integer types for formatted buffers * shader_recompiler: Remove existing shared memory emulation * The current impl relies on backend side implementaton and hooking into every shared memory access. It also doesnt handle atomics. Will be replaced by an IR pass that solves these issues * shader_recompiler: Reintroduce shared memory on ssbo emulation * Now it is performed with an IR pass, and combined with the previous commit cleanup, is fully transparent from the backend, other than requiring workgroup_index be provided as an attribute (computing this on every shared memory access is gonna be too verbose * clang format * buffer_cache: Reduce buffer sizes * vk_rasterizer: Cleanup resource binding code * Reduce noise in the functions, also remove some arguments which are class members * Fix gcc	2025-02-15 14:06:56 +02:00
squidbus	6e12642151	shader_recompiler: Lower non-compute shared memory into spare VGPRs. (#2403 )	2025-02-12 20:10:13 -08:00
DanielSvoboda	98eb8cb741	Fix S_LSHR_B32 (#2405 ) the shift value should be extracted from the 5 least significant bits of the second operand (S1.u[4:0]), to ensure that the shift is limited to values from 0 to 31, suitable for 32-bit operations Instruction S_LSHR_B32 Description D.u = S0.u >> S1.u[4:0]. SCC = 1 if result is non-zero.	2025-02-12 06:31:19 -08:00
squidbus	40eef6a066	shader_recompiler: Exclude defaulted fragment inputs from quad/rect passthrough. (#2383 )	2025-02-10 21:33:30 -08:00
psucien	04fe3a79b9	fix: lower UBO max size to account buffer cache offset (#2388 ) * fix: lower UBO max size to account buffer cache offset * review comments * remove UBO size from spec and always set it to max on shader side	2025-02-09 22:03:20 +01:00
squidbus	cfe249debe	shader_recompiler: Replace texel buffers with in-shader buffer format interpretation (#2363 ) * shader_recompiler: Replace texel buffers with in-shader buffer format interpretation * shader_recompiler: Move 10/11-bit float conversion to functions and address some comments. * vulkan: Remove VK_KHR_maintenance5 as it is no longer needed for buffer views. * shader_recompiler: Add helpers for composites and bitfields in pack/unpack. * shader_recompiler: Use initializer_list for bitfield insert helper.	2025-02-06 20:40:49 -08:00
squidbus	b879dd59c6	shader_recompiler: Add workaround for drivers with unexpected unorm rounding behavior. (#2310 )	2025-02-04 01:01:59 -08:00
makigumo	fffd373652	Fix shader type names (#2336 ) Names didn't match definition in type.h	2025-02-03 23:24:56 -08:00
makigumo	f8f732e78c	fix ASSERT_MSG arguments (#2337 )	2025-02-04 08:51:07 +02:00

1 2 3 4 5 ...

322 commits