Commit Graph

8670 Commits

Author SHA1 Message Date
goldislead
66779fb873 [Vulkan] Fix adaptive tessellation issues
The domain shader reads the patch index from the hull shader output instead of gl_PrimitiveID, which bypassed the endian swap, offset, wrap and clamp already applied upstream.

The tessellator winds clockwise now. Clip space Y is not flipped on Vulkan, so counterclockwise winding inverted the facing and guest backface culling removed whole surfaces.

Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-18 08:32:23 +02:00
goldislead
aed81ca93a [Memory] Rework guest access resolution and unwatched invalidation
This is a combination of three different Edge commits.

Guest page access is resolved with system page granularity so that anything deciding on protection now takes permissive access of every guest page a system page covers, which matters when the host is larger than the guest.

Access violations: Write faults on pages with no watch armed are only reported handled if guest mapping allows the write.

Invalidation of unwatched ranges when made writable or freed: Decommit, Release and Protect-to-writable (including write-combine) raise invalidation callbacks even with no watch armed.

Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-18 08:32:23 +02:00
Adrian
59c08cd462 [Kernel] Replaced X_DISPATCHER_FLAGS with X_OBJECT_TYPES 2026-08-18 08:00:29 +02:00
The-Little-Wolf
cdd60494d7 [XEX] - Add support for dash 1746
- Add support for dash 1746
2026-08-18 07:26:06 +02:00
Gliniak
5cf409d6e5 [Misc] C++20'ify Part 1
- Use ranges
2026-08-17 18:23:52 +02:00
AurisDSP
8ffe24e372 [CPU] Publish JIT entry results under the entry table lock
GetOrCreate reads entry->status while holding the global critical region,
but ResolveFunction published compile results by writing entry->function,
end_address and status directly, with no lock held. Nothing establishes a
happens-before edge between the two, so a thread that observes
STATUS_READY can still read a stale entry->function -- which it does
unlocked, right after GetOrCreate returns.

Benign on x86's TSO in practice; reachable on AArch64.

Add MarkReady/MarkFailed, which publish under that same lock, and route
Processor through them.

ThreadSanitizer against the real EntryTable: 2 data races before, 0 after.
The race is undefined behaviour by definition; no torn pointer was
actually observed in those runs.
2026-08-17 14:23:12 +02:00
goldislead
0d395ce9ab [GPU] Move debug cvars to a new TOML block, misc cleanup
native_2x_msaa is now debug_msaa_4x_as_2x, hopefully clearing up any user ambiguity.
2026-08-17 13:55:01 +02:00
Adrian
907d92bf8c [XAM] Fixed potential OOB memory write in copy utils 2026-08-13 21:54:31 +02:00
Michael Oliver
92ada8ebc0 [GPU] Fix scalar ALU swizzles with three-source vector ops 2026-08-13 20:24:57 +02:00
WawWeFix
ec5c875122 [CPU/X64] Optimize VMX dot products, vrsqrtefp, partial stores, and permutes 2026-08-13 07:25:39 +02:00
Gliniak
e31142bd79 [Linux/XSocket] Added error ID mapping for error 11. 2026-08-12 16:44:33 +02:00
bomabomabomaboma
fc48d37cdc [GPU] Remove num_format/decode cvars and dead shader code
Both of these have been around long enough to be probably prove safe and correct.

As a reminder, color resolves only take the full shader path if the destination number format matches EDRAM encoding && full 8_8_8_8_GAMMA resolves always decode PWL to linear before MSAA sample averaging.

Keeping decode_pwl_gamma bit for testing, but it's probably superfluous.
2026-08-12 09:24:36 +02:00
Adrian
8c98ef0280 [XAM] Cleanup XamUserCreateTitlesPlayedEnumerator 2026-08-11 19:27:27 +02:00
The-Little-Wolf
367d22d269 [XBOXKNRL/MISC] - Stub MicDeviceRequest
- Stub MicDeviceRequest
- Games:
  - Rock Band 2 now exits game instead of being stuck on boot when mic set to false
  - Guitar Hero World Tour now properly responds when you tell it mic is connected
2026-08-11 11:53:40 +02:00
NicknineTheEagle
12f5084b93 [XAM] Allow deleting titles with achievements earned from profiles 2026-08-11 07:21:55 +02:00
bomabomabomaboma
052cb95f24 [GPU] Give integer scales the channel width that feeds each lane 2026-08-10 22:42:38 +02:00
The-Little-Wolf
7d8db5a2cd [XBOXKRNL] - Replace lpvoid_t with pointer_t
- Replace lpvoid_t with pointer_t
- Use TypedGuestPointer for X_DEVICE_OBJECT struct
2026-08-10 08:54:54 +02:00
goldislead
2b3f0cb456 [GPU] Select alpha for A8 resolves
Treat k_8 + LOW_BLUE as alpha selection. 5451080D uses this to resolve its opacity plane, and 4D530808 does the same for a fog lighting pass. Their blue channels are black, so treating LOW_BLUE as an R/B exchange drops data.

Informed by XGCopySurface decompilation, which combines source and inverse dest swizzle, plus notcing how L8 and A8 share the same texture format.
2026-08-10 07:43:43 +02:00
Gliniak
b98037bed2 [CI] Disable running pipeline for draft PRs
There is no need to run them on main repo, when PR is not ready.
They still should be functional on personal fork.
2026-08-07 17:12:04 +02:00
bomabomabomaboma
6a45452087 [GPU] Walk the guest swizzle for integer scales; update comments
Fixed integer fetches are updated to look widths up through the guest swizzle, so every output channel is scaled by the width it came from.

Previously, host swizzle was being walked, and it was causing some 6 bit channels to be scaled as 5 bits and vice versa.

Co-authored-by: philtimmes <5494151+philtimmes@users.noreply.github.com>
2026-08-06 06:41:21 +02:00
goldislead
25597a5465 [GPU] Allow depth clamping instead of clipping
For now, this adds a depth clamp override to both backends that's kept disabled by default.

494707EE needs this for its setup draws that feed its lighting passes. It could be that guest clipping / host near and flare Z planes aren't cleanly interchangeable at the edge of the clip volume.
2026-08-05 21:52:24 +02:00
Adrian
6ff2a56f6a [XAM] Fixed returning incorrect max property size 2026-08-04 19:37:53 +02:00
bomabomabomaboma
0f2980de44 [GPU] Apply gradient exponent bias during fetch that's per-axis
The fetch constant carries independent signed exponent biases in [-16, 15] for the horizontal and vertical LOD gradients. Scale the H and V gradients by exp2(lod + bias_h) and exp2(lod + bias_v) respectively in the computed LOD sample path. getCompTexLOD keeps returning the raw queried LOD, treating the adjustment like the fetch-constant LOD bias, which is also not folded into it. Zero, the common case, is a no-op.

On the cube implicit-LOD path, where no explicit gradients exist to scale, the greater of the two biases is added to the LOD bias instead - exact when both are equal, erring towards a blurrier mip otherwise.

Co-Authored-By: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
090cecd1b8 [Vulkan] Implement vertex kill (oPts.z) in the translator
A non-zero value in bits 0:30 of the kill flag kills the vertex, tested on the integer bits rather than as a float comparison so denormal flushing can't affect the result, matching the DXBC translator. With PA_CL_CLIP_CNTL::vtx_kill_or, the position W is set to NaN, killing the whole primitive if any of its vertices requests the kill, with the "and" operator, a dedicated cull distance after the user clip plane cull distances is set to -1, culling the primitive only when all of its vertices request it.

Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
3254ac20f8 [Vulkan] Sample auto-LOD cube fetches with implicit LOD + bias
Explicit gradients of the direction reconstructed from the guest S/T/face coordinates pick the wrong mip on Vulkan, so sample cube fetches that use computed LOD without register gradients with implicit LOD and the accumulated LOD bias (fetch constant + register LOD + instruction bias) as the Bias image operand instead. This also makes tfetchCube consistent with getCompTexLOD, which already queries the implicit LOD. Register-gradient cube fetches keep explicit gradients in cube space, and other dimensions keep explicit gradients, matching the DXBC path. Gradient setup is skipped entirely on the implicit-LOD path.

Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
658cdbb5de [Vulkan] Disable anisotropy for ineligible texture formats
The check is moved after the anisotropy resolution (and the anisotropic override) so it sees, and can reset, the final filter state.

Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
1fdbe569e4 [Vulkan] Barrier consecutive uploads to the same image
Avoids a TRANSFER_WRITE -> TRANSFER_WRITE hazard identified by the Vulkan validation layer when a texture is reuploaded (after guest memory invalidation) without having been used for drawing in between, leaving it in the transfer destination usage with no layout transition to order the copies.

Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
cde5d85ec9 [GPU] Host RT polygon offset for Z-fighting decals 2026-08-04 10:34:42 +02:00
bomabomabomaboma
2a6e9f4e2a [Vulkan] Add depth_float24_convert_in_pixel_shader support
Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
947075f880 [GPU] Implement wide 1D texture support
Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
d0dd989238 [GPU] Honor force_bc_w_to_max: force sampler border color alpha to 1.0
Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
084f14ec41 [Vulkan] Use VK_EXT_custom_border_color for YCbCr texture border colors
Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
ac42250688 [GPU] Fix 20e4 bit span in Depth20e4To32 comments 2026-08-04 10:34:42 +02:00
bomabomabomaboma
e519d59e40 [Vulkan] Fix stacked-texture inter-layer lerp base in SampleTexture
Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
bomabomabomaboma
f3e42609a2 [GPU] Fix mantissa placement in CPU Float7e3To32
Co-authored-by: Herman S. <429230+has207@users.noreply.github.com>
2026-08-04 10:34:42 +02:00
goldislead
fbdb1f2817 [GPU] Handle two component tfetch1D coordinates
545407D4's UI shader uses tfetch1D with a nonreplicated source swizzle and a 2D constant. The way we have it set up is causing the translators to sample with Y set to 0.

Handle this as 2D in the translators. Normal scalar tfetch1D untouched.
2026-08-04 07:22:48 +02:00
Gliniak
fe40c0d68f [Kernel] Fixed regression related to X_OBJECT_HEADER struct 2026-08-03 21:07:14 +02:00
Gliniak
7ef873b0d5 [Audio] Added mutex for each audio client
This should prevent crashing when there is TOCTOU case in audio worker loop
2026-08-03 09:05:27 +02:00
Adrian
889d93e134 [Kernel] Fixed X_OBJECT_HEADER 2026-08-03 09:04:10 +02:00
The-Little-Wolf
a0e1ba8313 [XAM/INFO] - Stub XamQueryLiveHive & XamGetLiveHiveValue functions
- Stub XamQueryLiveHive & XamGetLiveHiveValue functions
- Stub XNetLogonGetMachineID
2026-08-03 09:03:15 +02:00
The-Little-Wolf
6d530a1ab9 [XBOXKRNL/IO] - Add missing structs for IoCreateDevice
- Add missing structs for IoCreateDevice
2026-08-03 08:42:12 +02:00
The-Little-Wolf
131503f68d [XEX] - add more file signatures
- Add older xex file signatures
- Add xex1 devkit key
- Add xex1 support
2026-08-02 22:00:14 +02:00
goldislead
8486e97a06 [GPU] Sample locked mip unnormalized fetches in the mip's grid
Fixes visibility popping in 555308B6 and 5553080B.

Unnormalized texture coordinates address texels of the mip level being sampled, not always the base level. The titles in question lock the fetch constant to one mip and address that mip's grid. The denominator was still the base level size, so each reduction after the first read garbage.

So the locked size is now used for 2D fetches (point filter, clamp-to-edge) with MipMinLevel == MipMaxLevel. Everything else keeps the base level denominator until a shared effective LOD model exists.
2026-08-02 20:53:25 +02:00
The-Little-Wolf
364dc1d770 [XBOXKRNL/IO] - add structs for NtDeviceIoControlFile
- Add structs for NtDeviceIoControlFile
2026-08-02 19:12:54 +02:00
goldislead
da47dfaacf [GPU/DXBC] Fix signed round bias breaking memexport
Write the bias into round_bias_temp as intended in both packed 8/16/32bpp path and 16_16_16_16.

The copysign rounding bias was targeted the eM register, and destroying the scaled value, then added round_bias_temp, which nothing had written.
2026-08-02 18:05:24 +02:00
Gliniak
7010c86fb1 Revert "[Kernel] Initial XMP Notifications"
This reverts commit 0c6dc9b628.
2026-07-30 20:02:22 +02:00
Adrian
e4f28f6984 [Kernel] Fixed object type 2026-07-30 16:55:27 +02:00
The-Little-Wolf
2585d3cd69 [XAM/NET] - Stub XampXAuth functions
- Stub XampXAuthStartup, XampXAuthShutdown, XampXAuthGetTitleBuffer
2026-07-30 09:52:59 +02:00
Adrian
1807da06db [Kernel] Added usage of X_DISPATCHER_FLAGS 2026-07-30 08:54:23 +02:00
bomabomabomaboma
2ddc5ef737 [GPU] 8_8_8_8_GAMMA resolve update; fix unsigned-biased fetch scaling
8_8_8_8_GAMMA resolve update:
This hopefully settles how safe resolves of gamma RTs are handled, because it's almost certainly not settled by destination.

It's now believed that gamma sources seen by resolve are always being decoded, and that's a decode that's keyed only on the source format, where every destination is fed linear values, and any encode is now just removed rather than conditional.

These sources are re-alised as plain before resolving, which is the same exact idea that preserves float 2_10_10_10 bits through a resolve too.

Integer texture fetch scaling update:
Unsigned-biased components were using the regular unsigned scale even though the conversion had remapped the sample [0, 1] to [-1, 1].

Scale bits data is now increased from 5 to 6 bits and use the new bit to mark unsigned-biased components. DXBC/SPIR-V use half of the unsigned scale and add a -0.5 offset.

This fixes post-processing in 415608B2 and will likely improve other titles w/ EDRAM transports through affected textures, but perhaps not so dramatically.
2026-07-30 07:34:28 +02:00