ryujinx/Ryujinx

mirror of https://github.com/Ryujinx/Ryujinx.git synced 2024-11-08 06:18:38 +00:00

Author	SHA1	Message	Date
gdkchan	851f56b08a	Support Array/3D depth-stencil render target, and single layer clears (#3400 ) * Support Array/3D depth-stencil render target, and single layer clears * Alignment	2022-06-14 13:30:39 -03:00
gdkchan	9a9349f0f4	Fix instanced indexed inline draw index count (#3389 )	2022-06-10 23:44:49 -03:00
gdkchan	46cc7b55f0	Fix instanced indexed inline draws (#3383 )	2022-06-05 21:24:28 -03:00
gdkchan	a3e7bb8eb4	Copy dependency for multisample and non-multisample textures (#3382 ) * Use copy dependency for textures that differs in multisample but are otherwise compatible * Remove allowMs flag as it's no longer required for correctness, it's just an optimization now * Dispose intermmediate pool	2022-06-05 14:06:47 -03:00
Billy Laws	d03124a992	Fix 3D semaphore counter type 0 handling (#3380 ) Counter type 0 actually releases the semaphore payload rather than a constant zero as was previously thought. This is required by Skyrim.	2022-06-02 19:51:36 -03:00
Emmanuel Hansen	deb99d2cae	Avalonia UI - Part 1 (#3270 ) * avalonia part 1 * remove vulkan ui backend * move ui common files to ui common project * get name for oading screen from device * rebase. * review 1 * review 1.1 * review * cleanup * addressed review * use cancellation token * review * review * rebased * cancel library loading when closing window * remove star image, use fonticon instead * delete render control frame buffer when game ends. change position of fav star * addressed @Thog review * ensure the right ui is downloaded in updates * fix crash when showing not supported dialog during controller request * add prefix to artifact names * Auto-format Avalonia project * Fix input * Fix build, simplify app disposal * remove nv stutter thread * addressed review * add missing change * maintain window size if new size is zero length * add game, handheld, docked to local * reverse scale main window * Update de_DE.json * Update de_DE.json * Update de_DE.json * Update italian json * Update it_IT.json * let render timer poll with no wait * remove unused code * more unused code * enabled tiered compilation and trimming * check if window event is not closed before signaling * fix atmospher case * locale fix * locale fix * remove explicit tiered compilation declarations * Remove ) it_IT.json * Remove ) de_DE.json * Update it_IT.json * Update pt_BR locale with latest strings * Remove ')' * add more strings to locale * update locale * remove extra slash * remove extra slash * set firmware version to 0 if key's not found * fix * revert timer changes * lock on object instead * Update it_IT.json * remove unused method * add load screen text to locale * drop swap event * Update de_DE.json * Update de_DE.json * do null check when stopping emulator * Update de_DE.json * Create tr_TR.json * Add tr_TR * Add tr_TR + Turkish * Update it_IT.json * Update Ryujinx.Ava/Input/AvaloniaMappingHelper.cs Co-authored-by: Ac_K <Acoustik666@gmail.com> * Apply suggestions from code review Co-authored-by: Ac_K <Acoustik666@gmail.com> * Apply suggestions from code review Co-authored-by: Ac_K <Acoustik666@gmail.com> * addressed review * Update Ryujinx.Ava/Ui/Backend/OpenGl/OpenGlRenderTarget.cs Co-authored-by: gdkchan <gab.dark.100@gmail.com> * use avalonia's inbuilt renderer on linux * removed whitespace * workaround for queue render crash with vsync off * drop custom backend * format files * fix not closing issue * remove warnings * rebase * update avalonia library * Reposition the Text and Button on About Page * Assign build version * Remove appveyor text Co-authored-by: gdk <gab.dark.100@gmail.com> Co-authored-by: Niwu34 <67392333+Niwu34@users.noreply.github.com> Co-authored-by: Antonio Brugnolo <36473846+AntoSkate@users.noreply.github.com> Co-authored-by: aegiff <99728970+aegiff@users.noreply.github.com> Co-authored-by: Ac_K <Acoustik666@gmail.com> Co-authored-by: MostlyWhat <78652091+MostlyWhat@users.noreply.github.com>	2022-05-15 13:30:15 +02:00
riperiperi	9ba73ffbe5	Prefetch capabilities before spawning translation threads. (#3338 ) * Prefetch capabilities before spawning translation threads. The Backend Multithreading only expects one thread to submit commands at a time. When compiling shaders, the translator may request the host GPU capabilities from the backend. It's possible for a bunch of translators to do this at the same time. There's a caching mechanism in place so that the capabilities are only fetched once. By triggering this before spawning the thread, the async translation threads no longer try to queue onto the backend queue all at the same time. The Capabilities do need to be checked from the GPU thread, due to OpenGL needing a context to check them, so it's not possible to call the underlying backend directly. * Initialize the capabilities when setting the GPU thread + missing call in headless * Remove private variables	2022-05-14 11:58:33 -03:00
riperiperi	43b4b34376	Implement Viewport Transform Disable (#3328 ) * Initial implementation (no specialization) * Use specialization * Fix render scale, increase code gen version * Revert accidental change * Address Feedback	2022-05-12 10:47:13 -03:00
gdkchan	9eb5b7a10d	Restrict cases where vertex buffer size from index buffer type is used (#3304 )	2022-05-01 11:12:34 -03:00
riperiperi	d64594ec74	Fix various issues with texture sync (#3302 ) * Fix various issues with texture sync A variable called _actionRegistered is used to keep track of whether a tracking action has been registered for a given texture group handle. This variable is set when the action is registered, and should be unset when it is consumed. This is used to skip registering the tracking action if it's already registered, saving some time for render targets that are modified very often. There were two issues with this. The worst issue was that the tracking action handler exits early if the handle's modified flag is false... which means that it never reset _actionRegistered, as that was done within the Sync() method called later. The second issue was that this variable was set true after the sync action was registered, so it was technically possible for the action to run immediately, set the flag to false, then set it to true. Both situations would lead to the action never being registered again, as the texture group handle would be sure the action is already registered. This breaks the texture for the remaining runtime, or until it is disposed. It was also possible for a texture to register sync once, then on future frames the last modified sync number did not update. This may have caused some more minor issues. Seems to fix the Xenoblade flashing bug. Obviously this needs a lot of testing, since it was random chance. I typically had the most luck getting it to happen by switching time of day on the event theatre screen for a while, then entering the equipment screen by pressing X on an event. May also fix weird things like random chance air swimming in BOTW, maybe a few texture streaming bugs. * Exchange rather than CompareExchange	2022-04-29 18:34:11 -03:00
gdkchan	43ebd7a9bb	New shader cache implementation (#3194 ) * New shader cache implementation * Remove some debug code * Take transform feedback varying count into account * Create shader cache directory if it does not exist + fragment output map related fixes * Remove debug code * Only check texture descriptors if the constant buffer is bound * Also check CPU VA on GetSpanMapped * Remove more unused code and move cache related code * XML docs + remove more unused methods * Better codegen for TransformFeedbackDescriptor.AsSpan * Support migration from old cache format, remove more unused code Shader cache rebuild now also rewrites the shared toc and data files * Fix migration error with BRX shaders * Add a limit to the async translation queue Avoid async translation threads not being able to keep up and the queue growing very large * Re-create specialization state on recompile This might be required if a new version of the shader translator requires more or less state, or if there is a bug related to the GPU state access * Make shader cache more error resilient * Add some missing XML docs and move GpuAccessor docs to the interface/use inheritdoc * Address early PR feedback * Fix rebase * Remove IRenderer.CompileShader and IShader interface, replace with new ShaderSource struct passed to CreateProgram directly * Handle some missing exceptions * Make shader cache purge delete both old and new shader caches * Register textures on new specialization state * Translate and compile shaders in forward order (eliminates diffs due to different binding numbers) * Limit in-flight shader compilation to the maximum number of compilation threads * Replace ParallelDiskCacheLoader state changed event with a callback function * Better handling for invalid constant buffer 1 data length * Do not create the old cache directory structure if the old cache does not exist * Constant buffer use should be per-stage. This change will invalidate existing new caches (file format version was incremented) * Replace rectangle texture with just coordinate normalization * Skip incompatible shaders that are missing texture information, instead of crashing This is required if we, for example, support new texture instruction to the shader translator, and then they allow access to textures that were not accessed before. In this scenario, the old cache entry is no longer usable * Fix coordinates normalization on cubemap textures * Check if title ID is null before combining shader cache path * More robust constant buffer address validation on spec state * More robust constant buffer address validation on spec state (2) * Regenerate shader cache with one stream, rather than one per shader. * Only create shader cache directory during initialization * Logging improvements * Proper shader program disposal * PR feedback, and add a comment on serialized structs * XML docs for RegisterTexture Co-authored-by: riperiperi <rhy3756547@hotmail.com>	2022-04-10 10:49:44 -03:00
gdkchan	e44a43c7e1	Implement VMAD shader instruction and improve InvocationInfo and ISBERD handling (#3251 ) * Implement VMAD shader instruction and improve InvocationInfo and ISBERD handling * Shader cache version bump * Fix typo	2022-04-08 12:42:39 +02:00
gdkchan	3139a85a2b	Allow copy texture views to have mismatching multisample state (#3152 )	2022-04-08 11:26:48 +02:00
merry	a4e8bea866	Lop3Expression: Optimize expressions (#3184 ) * lut3 * bugfixes * TruthTable * false/true -> 0/-1 * add or to expressions * fix inversions * increment cache version	2022-04-08 11:17:38 +02:00
gdkchan	952f6f8a65	Calculate vertex buffer size from index buffer type (#3253 ) * Calculate vertex buffer size from index buffer type * We also need to update the size if first vertex changes	2022-04-08 11:02:06 +02:00
gdkchan	d4b960d348	Implement primitive restart draw arrays properly on OpenGL (#3256 )	2022-04-04 18:43:24 -03:00
gdkchan	b2a225558d	Do not force scissor on clear if scissor is disabled (#3258 )	2022-04-04 18:30:43 -03:00
gdkchan	1402d8391d	Support NVDEC H264 interlaced video decoding and VIC deinterlacing (#3225 ) * Support NVDEC H264 interlaced video decoding and VIC deinterlacing * Remove unused code	2022-03-23 17:09:32 -03:00
gdkchan	79408b68c3	De-tile GOB when DMA copying from block linear to pitch kind memory regions (#3207 ) * De-tile GOB when DMA copying from block linear to pitch kind memory regions * XML docs + nits * Remove using * No flush for regular buffer copies * Add back ulong casts, fix regression due to oversight	2022-03-20 13:55:07 -03:00
gdkchan	e5ad1dfa48	Implement S8D24 texture format and tweak depth range detection (#2458 )	2022-03-15 03:42:08 +01:00
gdkchan	79becc4b78	Dynamically increase buffer size when resizing (#2861 ) * Grow buffers by 1.5x of its size when resizing * Further restrict the cases where the dynamic expansion is done	2022-03-15 03:33:53 +01:00
gdkchan	0bcbe32367	Only initialize shader outputs that are actually used on the next stage (#3054 ) * Only initialize shader outputs that are actually used on the next stage * Shader cache version bump	2022-03-06 20:42:13 +01:00
gdkchan	0a24aa6af2	Allow textures to have their data partially mapped (#2629 ) * Allow textures to have their data partially mapped * Explicitly check for invalid memory ranges on the MultiRangeList * Update GetWritableRegion to also support unmapped ranges	2022-02-22 13:34:16 -03:00
riperiperi	c9c65af59e	Perform unscaled 2d engine copy on CPU if source texture isn't in cache. (#3112 ) * Initial implementation of fast 2d copy TODO: Partial copy for mismatching region/size. * WIP * Cleanup * Update Ryujinx.Graphics.Gpu/Engine/Twod/TwodClass.cs Co-authored-by: gdkchan <gab.dark.100@gmail.com> Co-authored-by: gdkchan <gab.dark.100@gmail.com>	2022-02-22 11:21:29 -03:00
Berkan Diler	644b497df1	Collapse AsSpan().Slice(..) calls into AsSpan(..) (#3145 ) * Collapse AsSpan().Slice(..) calls into AsSpan(..) Less code and a bit faster * Collapse an Array.Clear(array, 0, array.Length) call to Array.Clear(array)	2022-02-22 10:32:10 -03:00
gdkchan	72e543e946	Prefer texture over textureSize for sampler type (#3132 ) * Prefer texture over textureSize for sampler type * Shader cache version bump	2022-02-18 02:44:46 +01:00
gdkchan	3bd357045f	Do not allow render targets not explicitly written by the fragment shader to be modified (#3063 ) * Do not allow render targets not explicitly written by the fragment shader to be modified * Shader cache version bump * Remove blank lines * Avoid redundant color mask updates * HostShaderCacheEntry can be null * Avoid more redundant glColorMask calls * nit: Mask -> Masks * Fix currentComponentMask * More efficient way to update _currentComponentMasks	2022-02-16 23:15:39 +01:00
gdkchan	7bfb5f79b8	When copying linear textures, DMA should ignore region X/Y (#3121 )	2022-02-16 11:13:45 +01:00
Berkan Diler	8f35345729	Use Enum and Delegate.CreateDelegate generic overloads (#3111 ) * Use Enum generic overloads * Remove EnumExtensions.cs * Use Delegate.CreateDelegate generic overloads	2022-02-13 10:50:07 -03:00
gdkchan	f861f0bca2	Fix missing geometry shader passthrough inputs (#3106 ) * Fix missing geometry shader passthrough inputs * Shader cache version bump	2022-02-11 19:52:20 +01:00
Mary	6dffe0fad4	misc: Make PID unsigned long instead of long (#3043 )	2022-02-09 17:18:07 -03:00
gdkchan	b944941733	Fix bug that could cause depth buffer to be missing after clear (#3067 )	2022-01-31 00:11:43 -03:00
riperiperi	c52158b733	Add timestamp to 16-byte/4-word semaphore releases. (#3049 ) * Add timestamp to 16-byte semaphore releases. BOTW was reading a ulong 8 bytes after a semaphore return. Turns out this is the timestamp it was trying to do performance calculation with, so I've made it write when necessary. This mode was also added to the DMA semaphore I added recently, as it is required by a few games. (i think quake?) The timestamp code has been moved to GPU context. Check other games with an unusually low framerate cap or dynamic resolution to see if they have improved. * Cast dma semaphore payload to ulong to fill the space * Write timestamp first Might be just worrying too much, but we don't want the applcation reading timestamp if it sees the payload before timestamp is written.	2022-01-27 22:50:32 +01:00
riperiperi	fd6d3ec88f	Fix res scale parameters not being updated in vertex shader (#3046 ) This fixes an issue where the render scale array would not be updated when technically the scales on the flat array were the same, but the start index for the vertex scales was different.	2022-01-27 14:17:13 -03:00
gdkchan	42c75dbb8f	Add support for BC1/2/3 decompression (for 3D textures) (#2987 ) * Add support for BC1/2/3 decompression (for 3D textures) * Optimize and clean up * Unsafe not needed here * Fix alpha value interpolation when a0 <= a1	2022-01-22 19:23:00 +01:00
gdkchan	7e967d796c	Stop using glTransformFeedbackVaryings and use explicit layout on the shader (#3012 ) * Stop using glTransformFeedbackVarying and use explicit layout on the shader * This is no longer needed * Shader cache version bump * Fix gl_PerVertex output for tessellation control shaders	2022-01-21 12:35:21 -03:00
gdkchan	0e59573f2b	Add capability for BGRA formats (#3011 )	2022-01-20 08:37:21 -03:00
gdkchan	fb853f13e9	Scale scissor used for clears (#3002 )	2022-01-16 20:23:00 -03:00
gdkchan	6e0799580f	Fix render target clear when sizes mismatch (#2994 )	2022-01-11 20:15:17 +01:00
riperiperi	ef24c8983d	Fix adjacent 3d texture slices being detected as Incompatible Overlaps (#2993 ) This fixes some regressions caused by #2971 which caused rendered 3D texture data to be lost for most slices. Fixes issues with Xenoblade 2's colour grading, probably a ton of other games. This also removes the check from TextureCache, making it the tiniest bit smaller (any win is a win here).	2022-01-11 09:37:40 +01:00
gdkchan	7f6b3d234a	Implement IMUL, PCNT and CONT shader instructions, fix FFMA32I and HFMA32I (#2972 ) * Implement IMUL shader instruction * Implement PCNT/CONT instruction and fix FFMA32I * Add HFMA232I to the table * Shader cache version bump * No Rc on Ffma32i	2022-01-10 12:08:00 -03:00
gdkchan	952c6e4d45	Fix sampled multisample image size (#2984 )	2022-01-10 08:45:25 +01:00
riperiperi	cda659955c	Texture Sync, incompatible overlap handling, data flush improvements. (#2971 ) * Initial test for texture sync * WIP new texture flushing setup * Improve rules for incompatible overlaps Fixes a lot of issues with Unreal Engine games. Still a few minor issues (some caused by dma fast path?) Needs docs and cleanup. * Cleanup, improvements Improve rules for fast DMA * Small tweak to group together flushes of overlapping handles. * Fixes, flush overlapping texture data for ASTC and BC4/5 compressed textures. Fixes the new Life is Strange game. * Flush overlaps before init data, fix 3d texture size/overlap stuff * Fix 3D Textures, faster single layer flush Note: nosy people can no longer merge this with Vulkan. (unless they are nosy enough to implement the new backend methods) * Remove unused method * Minor cleanup * More cleanup * Use the More Fun and Hopefully No Driver Bugs method for getting compressed tex too This one's for metro * Address feedback, ASTC+ETC to FormatClass * Change offset to use Span slice rather than IntPtr Add * Fix this too	2022-01-09 13:28:48 -03:00
riperiperi	79adba4402	Add support for render scale to vertex stage. (#2763 ) * Add support for render scale to vertex stage. Occasionally games read off textureSize on the vertex stage to inform the fragment shader what size a texture is without querying in there. Scales were not present in the vertex shader to correct the sizes, so games were providing the raw upscaled texture size to the fragment shader, which was incorrect. One downside is that the fragment and vertex support buffer description must be identical, so the full size scales array must be defined when used. I don't think this will have an impact though. Another is that the fragment texture count must be updated when vertex shader textures are used. I'd like to correct this so that the update is folded into the update for the scales. Also cleans up a bunch of things, like it making no sense to call CommitRenderScale for each stage. Fixes render scale causing a weird offset bloom in Super Mario Party and Clubhouse Games. Clubhouse Games still has a pixelated look in a number of its games due to something else it does in the shader. * Split out support buffer update, lazy updates. * Commit support buffer before compute dispatch * Remove unnecessary qualifier. * Address Feedback	2022-01-08 14:48:48 -03:00
gdkchan	15131d4350	Force crop when presentation cached texture size mismatches (#2957 )	2021-12-31 12:00:42 -03:00
gdkchan	c05c8e09d4	Add support for the R4G4 texture format (#2956 )	2021-12-30 17:10:54 +01:00
gdkchan	ef39b2ebdd	Flip scissor box when the YNegate bit is set (#2941 ) * Flip scissor box when the YNegate bit is set * Flip scissor based on screen scissor state, account for negative scissor Y * No need for abs when we already know the value is negative	2021-12-28 08:37:23 -03:00
gdkchan	a87f7f2029	Fix DMA copy fast path line size when xCount < stride (#2942 )	2021-12-26 13:05:26 -03:00
gdkchan	451673ada5	Fix I2M texture copies when line length is not a multiple of 4 (#2938 ) * Fix I2M texture copies when line length is not a multiple of 4 * Do not copy padding bytes for 1D copies * Nit	2021-12-26 12:39:07 -03:00
gdkchan	e7c2dc8ec3	Fix for texture pool not being updated when it should + buffer texture related fixes (#2911 )	2021-12-19 11:50:44 -03:00

1 2 3 4 5 ...

380 commits