dynarmic

Author	SHA1	Message	Date
Lioncash	349d4b577a	General: Remove unnecessary includes Removes unnecessary header dependencies that have accumulated over time as changes have been made. Lessens the amount of files that need to be rebuilt when the headers change.	2020-04-22 21:04:22 +01:00
Lioncash	43fd2b400a	frontend/ir_emitter: Add half-precision opcode for FPVectorEquals	2020-04-22 21:04:22 +01:00
Lioncash	dd315e89eb	A64/translate/*: Apply const where applicable Just some tidying up for consistency	2020-04-22 21:04:22 +01:00
Lioncash	4f47861669	A64/translate/impl: Mark DecodeBitMasks and AdvSIMDExpandImm as static These don't rely on instance state to perform their behavior. They're just helper functions.	2020-04-22 21:04:22 +01:00
Lioncash	dddba94c17	disassembler_arm: Apply const where applicable	2020-04-22 21:04:22 +01:00
Lioncash	9365487797	frontend/A32/ir_emitter: Remove unnecessary includes std::initializer_list isn't used anywhere in here, and we can just forward declare the CoprocReg enum to avoid needing to include the header.	2020-04-22 21:04:22 +01:00
Lioncash	bfa8035414	A32/A64: Make public header inclusions consistent For all public header inclusions, we use the <> form of including them as opposed to "", which we typically use for internal headers.	2020-04-22 21:04:22 +01:00
Lioncash	fb7d33830c	A32: Make includes consistent Normalizes includes to be relative to the project root, like the rest of the includes in the project.	2020-04-22 21:04:22 +01:00
Lioncash	b57ed8917a	frontend/A32/types: Remove redundant std::string initializer std::string initializes to empty by default. While we're at it, brace a lone unbraced if statement.	2020-04-22 21:04:22 +01:00
Lioncash	182ceb2807	General: Make parameter names from declarations and implementations consistent Most of the time when this occurs, it's a bug. Thankfully this isn't the case. However, we can resolve these cases to make the codebase more consistent.	2020-04-22 21:04:22 +01:00
Lioncash	b301fcd520	A32/translate/translate: Add missing doxygen parameter string	2020-04-22 21:04:22 +01:00
Lioncash	6b9bf7868a	General: Correct typos is code comments	2020-04-22 21:04:22 +01:00
MerryMage	7d20f3b861	A32/translate_thumb: Split off implementation into thumb16 and thumb32	2020-04-22 21:04:22 +01:00
Lioncash	b79ce71b0f	ir/basic_block: std::move Terminal within SetTerminal and ReplaceTerminal A terminal isn't a trivial type (and boost::variant is allowed to heap allocate), so we can std::move it here to avoid a redundant copy.	2020-04-22 21:04:22 +01:00
MerryMage	e639aa1583	A32/translate: Rename translate_arm directory to impl Mirror what the A64 frontend does.	2020-04-22 21:04:22 +01:00
Lioncash	63eff4e7cc	ir/terminal: std::move constructor parameters where applicable Allows the compiler to choose the most suitable code in this scenario, given a Terminal isn't a trivial type.	2020-04-22 21:04:22 +01:00
MerryMage	5f8eb7c51c	A32/location_descriptor: Add CPSR.IT to A32::LocationDescriptor	2020-04-22 21:04:22 +01:00
MerryMage	13f65f55eb	PSR: Use Common::ModifyBit{,s}	2020-04-22 21:04:22 +01:00
MerryMage	74633301c1	A32: Add ITState	2020-04-22 21:04:22 +01:00
MerryMage	6e2cd35e4f	a32_jitstate: Optimize runtime location descriptor calculation Calculation is now one unaligned 64-bit load.	2020-04-22 21:04:22 +01:00
MerryMage	f178562ee7	a32_jitstate: Remove exception trap enables from FPSCR_MODE_MASK We don't currently use this for anything (we do not currently trap floating point exceptions). This frees these bits up for other purposes.	2020-04-22 21:04:21 +01:00
Merry	fd6222f0a1	Merge pull request #500 from lioncash/cbz A32: Implement Thumb-1's CBZ/CBNZ instructions	2020-04-22 21:04:21 +01:00
Merry	bab4e29075	Merge pull request #498 from lioncash/ahp A32/location_descriptor: Add AHP bit to the FPSCR mask	2020-04-22 21:04:21 +01:00
Lioncash	87083af733	general: Remove trailing spaces General code-related cleanup. Gets rid of trailing spaces in the codebase.	2020-04-22 21:04:21 +01:00
Lioncash	03e6899fd7	A32: Implement Thumb-1's CBZ/CBNZ instructions Introduced in ARMv6T2, this allows for short forward branches.	2020-04-22 21:02:47 +01:00
Lioncash	d02a4e6fc9	A32/location_descriptor: Add AHP bit to the FPSCR mask Ensures the alternate half-precision state is preserved within the location descriptors, which will be necessary when implementing the half-precision extensions for VFP and NEON.	2020-04-22 21:02:47 +01:00
Merry	f4990a5f6b	Merge pull request #499 from lioncash/movw A32: Implement ARM-mode MOVW	2020-04-22 21:02:47 +01:00
Lioncash	bd755ae494	frontend/ir/ir_emitter: Add A32 equivalent to A64's SetCheckBit This will be used in a subsequent change to implement ARMv6T2's CBZ/CBNZ Thumb-1 instructions.	2020-04-22 21:02:47 +01:00
Lioncash	106c8c2473	A32: Implement ARM-mode MOVW Introduced to the ISA in ARMv6T2	2020-04-22 21:02:47 +01:00
Lioncash	9935f3aa28	A32: Implement Thumb-1 variant of SEVL While we're at it, also add the Thumb-2 encoding to the encoding table to make sure it isn't forgotten about in the future.	2020-04-22 21:02:47 +01:00
Lioncash	9a097e307f	A32: Implement the ARM-mode variant of SEVL	2020-04-22 21:02:47 +01:00
Lioncash	e89ca42048	A32: Implement Thumb-1 variant of YIELD	2020-04-22 21:02:47 +01:00
Lioncash	ebab7ede55	A32: Implement Thumb-1 variant of WFI	2020-04-22 21:02:47 +01:00
Lioncash	b4110af22a	A32: Implement Thumb-1 variant of WFE	2020-04-22 21:02:47 +01:00
Lioncash	57675fe592	A32: Implement Thumb-1 variant of SEV	2020-04-22 21:02:47 +01:00
Lioncash	07699b47ba	A32/translate_thumb: Add helper function for raising exceptions Similar to the variant within the ARM-mode translator visitor. This will be used in subsequent changes to implement the hint instructions introduced in ARMv7.	2020-04-22 21:02:47 +01:00
Lioncash	f74762ae4e	frontend/decoder/decoder_detail: Replace std::is_same, with std::is_same_v Same thing, same readability, less characters.	2020-04-22 21:02:47 +01:00
Lioncash	64879396f6	A32: Implement Thumb-1 variant of NOP	2020-04-22 21:02:47 +01:00
Merry	6a67da1225	Merge pull request #493 from lioncash/ir frontend/ir/ir_emitter: Remove unnecessary logical shift overloads	2020-04-22 21:02:47 +01:00
Merry	81b908b077	Merge pull request #495 from lioncash/bkpt A32: Implement Thumb-16's variant of BKPT	2020-04-22 21:02:47 +01:00
Merry	30d28029a8	Merge pull request #492 from lioncash/vfp A32: Rename vfp2-related files to vfp	2020-04-22 21:02:47 +01:00
Lioncash	b17a5d3365	A32: Implement Thumb-16's variant of BKPT	2020-04-22 21:02:47 +01:00
Lioncash	b902f72001	A32/disassembler_arm: Remove <unimplemented> from hint instruction output Given we now support hooking these hint instructions, we can consider them implemented.	2020-04-22 21:02:47 +01:00
Lioncash	0fa0bca22a	A32: Handle different variants of PLD	2020-04-22 21:02:47 +01:00
Lioncash	c6f99235e1	frontend/ir/ir_emitter: Remove unnecessary logical shift overloads These aren't necessary anymore, now that the U32U64 overload already exists.	2020-04-22 21:02:46 +01:00
Merry	9ba503e394	Merge pull request #491 from lioncash/hint A32: Allow hooking of hint instructions in ARM mode.	2020-04-22 21:02:46 +01:00
Lioncash	97277c598b	A32: Rename vfp2-related files to vfp Now that we fuzz against Unicorn, we aren't just restricted to VFPv2. VFPv3 and VFPv4 facilities can now be implemented. This renames constructs mentioning VFPv2 to just refer to VFP.	2020-04-22 21:02:46 +01:00
Lioncash	8c3122ff46	A64/translate/impl/impl: Mark locals const where applicable in DecodeBitMasks() Follows the convention of making immutable state explicit.	2020-04-22 21:02:46 +01:00
Merry	a132b56d57	Merge pull request #490 from lioncash/crc32 A32: Implement ARM-mode CRC32 instructions	2020-04-22 21:02:46 +01:00
Lioncash	966e04d03d	A32: Allow hooking of hint instructions in ARM mode. Mirrors the hooking functionality from the AArch64 frontend to make the behavior of both consistent.	2020-04-22 21:02:46 +01:00
Lioncash	134b586c5c	frontend/ir/ir_emitter: Amend arguments to conversion opcodes Accidentally caused within 967d1fcc8d6f60749a162a96b997439450fed687. That one's on me. My bad.	2020-04-22 21:02:46 +01:00
Lioncash	e37689315d	A32: Implement ARM-mode CRC32 instructions Implements the ARM-mode variants of the CRC32 instructions introduced within ARMv8. This is also one of the instruction cases where there is UNPREDICTABLE behavior that is constrained (we must do one of the options indicated by the reference manual). In both documented cases of constrained unpredictable behavior, we treat the instructions as unpredictable in order to allow library users to hook the unpredictable exception to provide the intended behavior they desire.	2020-04-22 21:02:46 +01:00
Lioncash	95d9baea67	{A32, A64}/types: Use std::array deduction guides where applicable We also make the arrays static here, as MSVC tends to load the whole array every time the function is called, instead of storing the data within rodata. This also line breaks the elements a little earlier for readability.	2020-04-22 21:02:46 +01:00
Lioncash	bac945f2d8	A32: Resolve parameter discrepancies discovered via use of the Imm template	2020-04-22 21:02:46 +01:00
Lioncash	e4c65721fe	frontend/ir/type: Generify std::array declaration With deduction guides, we can eliminate the need to explicitly size the array. Also newlines the elements based off their relation, making it slightly nicer to read.	2020-04-22 21:02:46 +01:00
Lioncash	4ba2318b2e	A32: Replace immediate type aliases with the Imm template Replaces type aliases of raw integral types with the more type-safe Imm template, like how the AArch64 frontend has been using it. This makes the two frontends more consistent with one another.	2020-04-22 21:02:46 +01:00
Lioncash	f96036b3f1	A32/barrier: Correct PC assignment within ISB The SetRegister() IR function doesn't allow specifying the PC as a register. This is a discrepancy that slipped through (my bad). Instead, we can use BranchWritePC(), like how the other similar PC modifying locations do it.	2020-04-22 21:02:46 +01:00
Lioncash	196e7b5e35	frontend/A32/ir_emitter: Mark locals as const where applicable Makes const usage consistent within the source file.	2020-04-22 21:02:46 +01:00
Lioncash	511613c736	frontend/A32/types: Use helper function in operator+ overload Allows deduplicating an assert and a cast.	2020-04-22 21:02:46 +01:00
Lioncash	8103652a91	frontend: Move imm.h to the top-level directory of the frontends Preparation to utilize the immediate type within the A32 backend as well, which will allow eliminating numerous type aliases like Imm4, Imm5, etc.	2020-04-22 21:02:46 +01:00
Lioncash	796bb8a7f7	frontend/A64/types: Make RegNumber() and VecNumber() constexpr Given they simply perform casting, they can be safely made constexpr.	2020-04-22 21:02:46 +01:00
Lioncash	64e51a6d4d	A32/disassembler_arm: Mark utility functions as static where applicable These don't depend on class state and can be marked static to make that explicit.	2020-04-22 21:02:46 +01:00
Lioncash	0c43228ad5	frontend/A64/types: Use helper functions in operator+ overloads Allows us to get rid of another explicit cast.	2020-04-22 21:02:46 +01:00
Lioncash	a1cace21a9	frontend/ir/ir_emitter: Apply const to locals where applicable Makes const usage consistent with all other functions in the source file.	2020-04-22 21:02:46 +01:00
Lioncash	0a35836998	frontend/ir/ir_emitter: Use switch constructs in floating point opcodes where applicable This'll reduce the amount of noise necessary in changes implementing half-precision instructions, as the type can just be prepended to the switch cases, instead of rewriting the whole if/else branch.	2020-04-22 21:02:46 +01:00
Lioncash	8316d231e9	A32: Implement barrier instructions introduced in ARMv7 Provides basic implementations of the barrier instruction introduced within ARMv7. Currently these simply mirror the behavior of the AArch64 equivalents.	2020-04-22 21:02:46 +01:00
Lioncash	7fc3bd689d	A32: Implement ARM-mode MLS	2020-04-22 21:02:46 +01:00
Lioncash	8b338b7def	A32: Implement ARM-mode MOVT	2020-04-22 21:02:46 +01:00
Lioncash	877fa0f8c3	A32: Implement ARM-mode SBFX	2020-04-22 21:02:46 +01:00
Lioncash	47218ee65d	A32: Implement ARM-mode UBFX	2020-04-22 21:02:46 +01:00
Lioncash	2970b34e3c	A32: Implement ARM-mode BFI	2020-04-22 21:02:46 +01:00
Lioncash	fab3a59e05	A32: Implement ARM-mode BFC	2020-04-22 21:02:46 +01:00
Lioncash	7305d13221	A32: Implement ARM-mode RBIT	2020-04-22 21:02:46 +01:00
Lioncash	b2f7a0e7ba	A32: Implement ARM-mode SDIV/UDIV Now that we have Unicorn in place, we can freely implement instructions introduced in newer versions of the ARM architecture.	2020-04-22 21:02:46 +01:00
Lioncash	c0ae23bbb7	A32/translate_thumb: Clean up formatting Performs a similar tidying up of the Thumb translator, like what was done with the regular ARM translator to make it consistent with the rest of the codebase. The A32 backend (both Thumb and ARM), will likely see more changes to it in the near future, so this just acts as a "dusting off".	2020-04-22 21:02:46 +01:00
Merry	837c23a8ec	Merge pull request #483 from lioncash/invert frontend/ir/cond: Remove unused invert() function	2020-04-22 21:02:46 +01:00
Merry	09ee64ea98	Merge pull request #482 from lioncash/fixedfp A64: Handle half-precision variants of FP->Fixed instructions	2020-04-22 21:02:45 +01:00
Lioncash	06ec6ab0da	frontend/ir/cond: Remove unused invert() function This is no longer used by anything in the codebase, so it can be removed.	2020-04-22 21:01:46 +01:00
Lioncash	64e3d233f4	A64: Handle half-precision variants of FP->Fixed-point instructions	2020-04-22 21:01:45 +01:00
Lioncash	4fc531f71b	ir/basic_block: Forward declare headers where applicable Now that the constructor and destructors have been placed within the cpp file, we can forward declare the memory pool data structures. Now, a change to the memory pool code won't ripple across the entirety of the IR emitter.	2020-04-22 21:01:45 +01:00
Lioncash	427b7afd66	frontend/ir/microinstruction: Add missing fixed-point opcodes to ReadsFromAndWritesToFPSRCumulativeExceptionBits()	2020-04-22 21:01:45 +01:00
Lioncash	9309d95b17	ir/block: Default ctor and dtor in the cpp file Prevents potentially inlining allocation code everywhere. While we're at it, also explicitly delete/default the copy/move constructor/assignment operators to be explicit about them.	2020-04-22 21:01:45 +01:00
Lioncash	604f39f00a	frontend/ir_emitter: Add half-precision->fixed-point opcodes	2020-04-22 21:01:45 +01:00
Lioncash	471eb77bc9	A64: Implement FRSQRTS' half-precision vector variant	2020-04-22 21:01:45 +01:00
Lioncash	f9b2862217	A64: Implement FRSQRTS' half-precision scalar variant With the necessary machinery in place, we can now handle the half-precision variant.	2020-04-22 21:01:45 +01:00
Lioncash	96356fac93	frontend/ir_emitter: Add half-precision opcode variant of FPVectorRSqrtStepFused	2020-04-22 21:01:45 +01:00
Merry	45864133f5	Merge pull request #478 from lioncash/stepfused A64: Handle half-precision variants of FRECPE and FRECPS	2020-04-22 21:01:44 +01:00
Lioncash	824c551ba2	frontend/ir_emitter: Add half-precision opcode variant of FPRSqrtStepFused	2020-04-22 21:01:44 +01:00
Lioncash	3739d92097	A64: Implement half-precision vector variant of FRECPE	2020-04-22 21:01:44 +01:00
Lioncash	7b212ec8ae	A64: Implement half-precision variant of FRSQRTE's vector variant	2020-04-22 21:01:44 +01:00
Lioncash	0945a491bd	A64: Implement half-precision scalar variant of FRECPE	2020-04-22 21:01:44 +01:00
Lioncash	77c84bcf9b	A64: Implement half-precision variant of FRSQRTE's scalar variant	2020-04-22 21:01:44 +01:00
Lioncash	86b7626a2f	A64: Implement half-precision vector variant of FRECPS	2020-04-22 21:01:44 +01:00
Lioncash	037acb17b9	frontend/ir_emitter: Add half-precision opcode variant for FPVectorRSqrtEstimate	2020-04-22 21:01:44 +01:00
Lioncash	de43f011a7	A64: Implement half-precision scalar variant of FRECPS	2020-04-22 21:01:44 +01:00
Lioncash	5dba99b4f4	frontend/ir_emitter: Add half-precision opcode variant for FPRSqrtEstimate	2020-04-22 21:01:44 +01:00
Lioncash	825a3ea16f	frontend/ir_emitter: Add half-precision opcode for FPVectorRecipEstimate	2020-04-22 21:01:44 +01:00
Lioncash	2184d24e8f	frontend/ir_emitter: Add half-precision opcode for FPRecipEstimate	2020-04-22 21:01:44 +01:00
Lioncash	d7f394fc1a	A64: Enable half-precision vector FRINT* variants	2020-04-22 21:01:44 +01:00
Lioncash	5d5c9f149f	frontend/ir_emitter: Add half-precision opcode for FPVectorRecipStepFused	2020-04-22 21:01:44 +01:00
Lioncash	24f583c498	A64: Enable half-precision variants of floating-point FRINT* variants With all the backing machinery in place, we can remove the fallback check for half-precision.	2020-04-22 21:01:44 +01:00
Lioncash	6da0411111	frontend/ir_emitter: Add half-precision opcode for FPRecipStepFused	2020-04-22 21:01:44 +01:00
Lioncash	fb829b9525	frontend/microinstruction: Add FPVectorRoundInt types to ReadsFromAndWritesToFPSRCumulativeExceptionBits() All variants were previously missing from this.	2020-04-22 21:01:44 +01:00
Lioncash	5b4673da4b	frontend/ir_emitter: Add half-precision variant of FPVectorRoundInt	2020-04-22 21:01:44 +01:00
Lioncash	ad0c698f89	frontend/ir_emitter: Add half-precision variant of FPRoundInt	2020-04-22 21:01:44 +01:00
Merry	cb9a1b18b6	Merge pull request #475 from lioncash/muladd A64: Enable half-precision variants of floating-point multiply-add instructions	2020-04-22 21:01:44 +01:00
Merry	d6db7ad46c	Merge pull request #474 from lioncash/bracing load_store_*: Make bracing consistent and variables const where applicable	2020-04-22 21:01:44 +01:00
Merry	1b6520f5dd	A64/location_descriptor: Ensure FZ16 is included in the FPCR mask	2020-04-22 21:01:44 +01:00
Merry	13f421c27d	Merge pull request #473 from lioncash/sqshlu A64: Implement SQSHLU	2020-04-22 21:01:44 +01:00
Lioncash	b5bf890584	load_store_*: Make bracing consistent and variables const where applicable Makes bracing consistent, and variables const where applicable to be consistent with the rest of the codebase. In most bracing cases, they'd need to be added to conditionals that would involve checking stack pointer alignment in the future anyways.	2020-04-22 21:01:44 +01:00
Lioncash	9a58c3f1c7	A64: Implement FMLA/FMLS' half-precision vector indexed variants	2020-04-22 21:01:44 +01:00
Merry	d7da53a74b	Merge pull request #472 from lioncash/exception general: Mark hash functions as noexcept	2020-04-22 21:01:44 +01:00
Lioncash	9dcc04e106	A64: Implement SQSHLU's scalar variant	2020-04-22 21:01:44 +01:00
Merry	b91c6c8bae	Merge pull request #471 from lioncash/sqrdmulh A64: Implement SQRDMULH's scalar vector variant	2020-04-22 21:01:44 +01:00
Lioncash	1fdd3ef8a0	A64: Implement FMLA/FMLS' half-precision scalar indexed variants	2020-04-22 21:01:44 +01:00
Lioncash	2d59d10ac8	A64: Implement SQSHLU's vector variant The vector shift by immediate category is now fully implemented.	2020-04-22 21:01:44 +01:00
Merry	b5e25959d9	Merge pull request #470 from lioncash/assert general: Replace unreachable-imitating assertions with UNREACHABLE()	2020-04-22 21:01:44 +01:00
Lioncash	d6606deda2	A64: Implement half-precision vector variants of FMLA/FMLS	2020-04-22 21:01:44 +01:00
Lioncash	a4cadf1cd9	frontend/ir_emitter: Add opcodes for signed saturated left shifts with unsigned saturation	2020-04-22 21:01:44 +01:00
Lioncash	ec6b3ae084	ir/frontend: Add half-precision opcode for FPVectorMulAdd	2020-04-22 21:01:44 +01:00
Lioncash	5f74d25bf7	A64: Enable half-precision floating point variants of FP data-processing three register instructions This handles half-precision floating point for: - FMADD - FMSUB - FNMADD - FNMSUB	2020-04-22 21:01:44 +01:00
Lioncash	bd82513199	frontend/ir_emitter: Add half-precision opcode for FPMulAdd	2020-04-22 21:01:44 +01:00
Lioncash	7bb5440507	general: Mark hash functions as noexcept Generally hash functions shouldn't throw exceptions. It's also a requirement for the standard library-provided hash functions to not throw exceptions. An exception to this rule is made for user-defined specializations, however we can just be consistent with the standard library on this to allow it to play nicer with it. While we're at it, we can also make the std::less specializations noexcpet as well, since they also can't throw.	2020-04-22 21:01:43 +01:00
Lioncash	3b46b4a37d	A64: Implement SQRDMULH's scalar vector variant Implements the scalar variant in terms of the vector variant for the time being.	2020-04-22 21:01:43 +01:00
Lioncash	fe95575b95	general: Replace unreachable-imitating assertions with UNREACHABLE() We can just use the self-documenting assertion for indicating unreachable paths, instead of manually passing false and providing a message.	2020-04-22 21:01:43 +01:00
Lioncash	64de80839e	A64/impl: Reorganize peculiar void use in V_scalar To a reader this might look particularly strange, given the function itself has a void return value, but this is actually valid, given the function in the return statement also has a void return value. This instead alters it to be a little easier to parse and potentially be a little less confusing at a glance.	2020-04-22 21:01:43 +01:00
Merry	9a4e3b24e4	Merge pull request #467 from lioncash/reserved A64: Handle reserved instruction cases more specifically where applicable	2020-04-22 21:01:43 +01:00
Lioncash	8309d49588	A64: Handle reserved instruction cases more specifically where applicable These are cases that are defined as reserved within the ARMv8 reference manual, so we can handle them as such instead of as unallocated encodings. While this doesn't actually change emulated behavior, it does at least allow the JIT to generate the more appropriate exception.	2020-04-22 21:00:47 +01:00
Lioncash	6c2c68bce6	A64: Implement FCMLA's indexed element variant With this, all of the instructions introduced with ARMv8.3-CompNum have an implementation.	2020-04-22 21:00:47 +01:00
Lioncash	7bc7042104	simd_scalar_shift_by_immediate: Change UnallocatedEncoding() path in SaturatingShiftLeft to ReservedValue() Strictly speaking, immh being zero is defined as reserved in the ARMv8 reference manual. This was just an error on my part when introducing the SQSHL immediate scalar variant.	2020-04-22 21:00:47 +01:00
Lioncash	b1b4487e4d	A64: Implement UQSHL (immediate)'s scalar variant Like SQSHL's immediate scalar variant, we can also implement UQSHL's immediate scalar variant in terms of the vector variant for the time being.	2020-04-22 21:00:47 +01:00
Lioncash	e1b4ff1068	simd_scalar_shift_by_immediate: Migrate SQSHL implementation to file-scope function This will allow it to be reused for the implementation of UQSHL.	2020-04-22 21:00:47 +01:00
Lioncash	3649dc6d9a	A64: Implement scalar variant of SQSHL (immediate) This can be handled in terms of the vector variant for the time being.	2020-04-22 21:00:47 +01:00
Merry	01bb1cdd88	Merge pull request #458 from lioncash/float-op A64: Handle half-precision floating point in FABS, FNEG, and scalar FMOV	2020-04-22 20:58:12 +01:00
Lioncash	28a8b4d210	A64: Handle half-precision floating point in scalar FMOV This is simply performing a scalar value transfer between registers without conversions, so this is trivial to handle as-is.	2020-04-22 20:58:12 +01:00
Lioncash	d7ac5a664f	A64: Handle half-precision floating point in FCVTL Like FCVTN, now that we have half-precision floating point conversion functions available, we can go ahead and use those to eliminate the interpreter fallback.	2020-04-22 20:58:12 +01:00
Lioncash	fe84ecb780	A64: Handle half-precision floating point in scalar FABS Now that we have the half-precision variant of the opcode added, we can simply handle the instruction instead of treating it as undefined.	2020-04-22 20:58:12 +01:00
Lioncash	fac9224d5e	A64: Handle half-precision floating point in FCVTN Now that we have IR instructions for performing conversions with half-precision floating point, we can also handle half-precision values within FCVTN.	2020-04-22 20:58:12 +01:00
Lioncash	8309ec7a9f	frontend/ir_emitter: Add half-precision variant of FPAbs	2020-04-22 20:58:12 +01:00
Lioncash	16de99d3e3	A64: Enable FCVT floating-point conversions for half-precision With this, we no longer have to fall back to the interpreter in any of the FCVT floating-point conversion instructions.	2020-04-22 20:58:12 +01:00
Lioncash	10abc77fad	A64: Handle half-precision floating point in scalar FNEG With the half-precision variant of the FPNeg opcode added, we can utilize it here to emulate the half-precision variant of FNEG.	2020-04-22 20:58:12 +01:00
Lioncash	e4c259d69f	frontend/ir_emitter: Add half->{single, double} and {double, single}->half conversion opcodes	2020-04-22 20:58:12 +01:00
Lioncash	c97efcb978	frontend/ir_emitter: Add half-precision variant of FPNeg	2020-04-22 20:58:12 +01:00
Merry	f01afc5ae6	Merge pull request #456 from lioncash/mov A64: Enable FMOV (general) for half-precision floating point	2020-04-22 20:58:12 +01:00
Merry	c1ce94872d	Merge pull request #455 from lioncash/sqrdmulh-scalar A64: Implement SQRDMULH and SQDMULL's scalar indexed variants	2020-04-22 20:58:11 +01:00
Lioncash	25a7256ee1	A64: Enable FMOV (general) for half-precision floating point This just transfers values between vector registers and general-purpose registers with no conversions performed, so this is trivial to add support for half-precision to.	2020-04-22 20:58:11 +01:00
Lioncash	97dd3d0596	A64: Implement SQRDMULH's scalar indexed element variant	2020-04-22 20:58:11 +01:00
Lioncash	49b51e34f1	simd_vector_x_indexed_element: Deduplicate index and Vm operand construction	2020-04-22 20:58:11 +01:00
Lioncash	692aba91b6	A64: Implement SQDMULL{2}'s scalar indexed element variant	2020-04-22 20:58:11 +01:00
Lioncash	c043b831d5	A64: Implement SQDMULL{2}'s by-element variant	2020-04-22 20:58:11 +01:00
Lioncash	72af5a3dff	simd_scalar_x_indexed_element: Factor out index and Vm argument construction This will be useful in the implementations of SQRDMULH and SQDMULL{2} as well.	2020-04-22 20:58:11 +01:00
Lioncash	224ff0afaa	A64: Implement SQRDMULH's by-index vector variant	2020-04-22 20:58:11 +01:00
Lioncash	3a3542414b	A64: Implement FRECPX's half-precision floating point variant	2020-04-22 20:58:11 +01:00
Lioncash	bd892ec4ef	frontend/ir/ir_emitter: Amend FPRecipExponent to handle half-precision floating point	2020-04-22 20:58:11 +01:00
Lioncash	974fbf0677	frontend/ir/value: Add U16U32U64 type to represent floating point types	2020-04-22 20:58:11 +01:00
Lioncash	126c29a9e9	A64: Implement SQSHRN, SQSHRUN, and UQSHRN's scalar variants These can just be implemented in terms of the vector variants for the time being.	2020-04-22 20:58:11 +01:00
Lioncash	dd7433f9d3	A64: Amend prototypes of some SIMD scalar shift by immediate opcodes These take a vector for a destination.	2020-04-22 20:58:11 +01:00
Merry	bbd5330ad2	Merge pull request #447 from lioncash/flag A64: Implement CFINV, RMIF, AXFlag and XAFlag	2020-04-22 20:58:11 +01:00
Merry	fb039e232c	Merge pull request #442 from lioncash/fcvtxn A64: Implement scalar and vector variants of FCVTXN	2020-04-22 20:58:11 +01:00
Merry	4f937c1ee1	Merge pull request #446 from lioncash/sqshl A64: Implement scalar variants of SQSHL (register) and UQSHL (register)	2020-04-22 20:58:11 +01:00
Lioncash	aa22db534b	A64: Implement AXFlag and XAFlag	2020-04-22 20:58:11 +01:00
Merry	d74cccbc84	Merge pull request #445 from lioncash/sqrt A64: Implement single and double-precision vector variant of FSQRT	2020-04-22 20:58:11 +01:00
Lioncash	20ffe568d0	A64: Implement RMIF	2020-04-22 20:58:11 +01:00
Merry	6d7e7c3269	Merge pull request #443 from lioncash/flag A64: Rearrange flag format/manipulation instructions	2020-04-22 20:58:11 +01:00
Lioncash	51b526e453	A64: Implement CFINV	2020-04-22 20:58:11 +01:00
Lioncash	597a8be5d5	ir: Add A64-specific opcodes for getting and setting raw NZCV values This will be necessary to implement the flag manipulation and flag format instructions.	2020-04-22 20:58:11 +01:00
Lioncash	d3515279df	A64: Implement the vector version of FCVTXN	2020-04-22 20:58:10 +01:00
Lioncash	17aea0b997	A64: Implement UQSHL (register)'s scalar variant This can be implemented in terms of the vector variant.	2020-04-22 20:58:10 +01:00
Lioncash	c99d4b762e	A64: Implement single and double-precision vector variant of FSQRT	2020-04-22 20:58:10 +01:00
Lioncash	54e0b487f3	A64: Rearrange flag format/manipulation instructions Gives these instructions better categorical labeling.	2020-04-22 20:58:10 +01:00
Lioncash	302f56b36a	A64: Fall back to interpreting for FCADD and FCMLA half-precision variants Rather than straight-up treating them as undefined, we can fall back to an interpreter in this case.	2020-04-22 20:58:10 +01:00
Lioncash	4339a8fff6	A64: Implement the scalar version of FCVTXN	2020-04-22 20:58:10 +01:00
Lioncash	35ddf68ad5	A64: Implement SQSHL (register)'s scalar variant We can implement this in terms of the vector variant.	2020-04-22 20:58:10 +01:00
Lioncash	5cf1478620	frontend/ir: Add opcodes for vector square roots	2020-04-22 20:58:10 +01:00
Lioncash	36027ebef5	frontend/ir/microinstruction: Add missing cases for FPRecipExponent{32,64} for ReadsFromAndWritesToFPSRCumulativeExceptionBits() This was intended to be added within #437, but was missed	2020-04-22 20:58:10 +01:00
Lioncash	7c81a58ed3	frontend/ir/ir_emitter: Alter parameters of FPDoubleToSingle() and FPSingleToDouble() to pass along desired rounding mode This will be necessary to special-case the non-IEEE Von Neumann rounding to odd rounding mode.	2020-04-22 20:58:10 +01:00
Merry	40b081438a	Merge pull request #439 from lioncash/fcmla A64: Implement FCADD and FCMLA	2020-04-22 20:58:10 +01:00
Merry	d91192681a	Merge pull request #438 from lioncash/fmulx A64: Implement scalar double/single precision FMULX (by element)	2020-04-22 20:58:10 +01:00
Lioncash	ed29ef8cca	A64: Implement FCMLA	2020-04-22 20:58:10 +01:00
Merry	9f11720a69	Merge pull request #437 from lioncash/frecpx A64: Implement FRECPX (single, double precision)	2020-04-22 20:58:10 +01:00
Lioncash	bdcea0b0dc	A64: Implement scalar double/single precision FMULX (by element)	2020-04-22 20:58:10 +01:00
Lioncash	5ce17574f9	A64: Implement FCADD	2020-04-22 20:58:10 +01:00
Merry	34d917f34e	Merge pull request #436 from lioncash/no-alloc A64: Implement LDNP/STNP	2020-04-22 20:58:10 +01:00
Lioncash	e44730ba6d	A64: Implement FRECPX (single, double precision)	2020-04-22 20:58:10 +01:00
Lioncash	bfaeb08d3c	A64: Implement LDNP/STNP LDNP and STNP indicate that a memory access is non-temporal/streaming (i.e. unlikely to be repeated), allowing data caching to not be performed. However, given this is only a hint, we can treat these two instructions as regular LDP and STP instructions for the time being.	2020-04-22 20:58:10 +01:00
Lioncash	9cf3c25811	frontend/ir/ir_emitter: Add opcodes for floating point reciprocal exponents	2020-04-22 20:58:10 +01:00
Lioncash	05a6ab691d	translate_arm/coprocessor: Minor tidying up	2020-04-22 20:58:10 +01:00
Lioncash	1e32a09c03	translate_arm/vfp2: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	e209b31073	translate_arm/synchronization: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	9514e3602e	translate_arm/status_register_access: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	c6aa1a708a	translate_arm/saturated: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	a72813599a	translate_arm/reversal: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	7be56e6b67	translate_arm/parallel: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	3c00a616d6	translate_arm/packing: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	c711188f46	translate_arm/multiply: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	c8dad40d81	translate_arm/misc: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	a7bf5ff77d	translate_arm/load_store: Invert conditionals where applicable	2020-04-22 20:58:10 +01:00
Lioncash	f4b19a7393	translate_arm/extension: Invert conditionals where applicable	2020-04-22 20:58:09 +01:00
Lioncash	c2de6ecfd0	translate_arm/exception_generating: Invert conditionals where applicable	2020-04-22 20:58:09 +01:00
Lioncash	d8a8d3b073	translate_arm/data_processing: Invert conditionals where applicable	2020-04-22 20:58:09 +01:00
Lioncash	df5c51ff47	translate_arm/branch: Invert conditionals where applicable Allows unindenting code a bit.	2020-04-22 20:58:09 +01:00
Lioncash	ee973f13c7	frontend/A32/ir_emitter: Mark PC() and AlignPC() as const-qualified member functions These don't modify instance state, so they can be const-qualified member functions.	2020-04-22 20:57:38 +01:00
Lioncash	3a2dd09122	frontend/A64/ir_emitter: Mark PC() and AlignPC() as const qualified member functions These don't actually alter any instance state.	2020-04-22 20:57:38 +01:00
MerryMage	e3898e628e	A64: Implement FMULX (by element), single and double precision variants	2020-04-22 20:57:37 +01:00
MerryMage	c106d8cedf	A64: Implement FMULX, vector single-precision and double-precision variant	2020-04-22 20:57:37 +01:00
MerryMage	fa8925c4df	IR: Implement FPVectorMulX	2020-04-22 20:57:37 +01:00
Michał Janiszewski	bbd8abaa25	Provide justification for always-true condition (#412 )	2020-04-22 20:57:37 +01:00
V.Kalyuzhny	764a93bf5a	Switch boost::optional to std::optional	2020-04-22 20:57:37 +01:00
Lioncash	f1a66c37ba	a64: Add ARMv8.4+ instructions encodings to the encoding table Keeps the table up to date with the ARM specification.	2020-04-22 20:57:37 +01:00
Lioncash	0583d401e3	ir/value: Add IsSignedImmediate() and IsUnsignedImmediate() functions to Value's interface This allows testing against arbitrary values while also simultaneously eliminating the need to check IsImmediate() all the time in expressions.	2020-04-22 20:57:37 +01:00
Lioncash	e3258e8525	ir/value: Add a GetImmediateAsS64() function Provides a signed analogue to GetImmediateAsU64() for consistency with both integral classes when it comes to signed/unsigned..	2020-04-22 20:57:37 +01:00
Lioncash	4a3c064b15	ir/value: Add an IsZero() member function to Value's interface By far, one of the most common things to check for is whether or not a value is zero, as it typically allows folding away unnecesary operations (other close contenders that can help with eliding operations are 1 and -1). So instead of requiring a check for an immediate and then actually retrieving the integral value and checking it, we can wrap it within a function to make it more convenient.	2020-04-22 20:57:37 +01:00
Merry	c649f11c0a	Merge pull request #401 from lioncash/folding constant_propagation_pass: Fold &, \|, ^, and ~ operations where applicable	2020-04-22 20:56:01 +01:00
MerryMage	2524d536b0	A32/ir_emitter: Bugfix: ExceptionRaised was producing incorrect PC Use actual PC and not pipelined PC.	2020-04-22 20:56:01 +01:00
Lioncash	d69fceec55	value: Move ImmediateToU64() to be a part of Value's interface This'll make it slightly nicer to do basic constant folding for 32-bit and 64-bit variants of the same IR opcode type. By that, I mean it's possible to inspect immediate values without a bunch of conditional checks beforehand to verify that it's possible to call GetU32() or GetU64, etc.	2020-04-22 20:55:50 +01:00
Lioncash	f40fcda1f6	ir/value: Add member function to check whether or not all bits of a contained value are set This is useful when we wish to know if a contained value is something like 0xFFFFFFFF, as this helps perform constant folding. For example the operation: x & 0xFFFFFFFF can be folded to just x in the 32-bit case.	2020-04-22 20:55:50 +01:00
MerryMage	f0920c0ded	Fix VShift terminology An arithmetic shift is by definition a signed shift, and a logical shift is by definition an unsigned shift. - Rename VectorLogicalVShiftS* -> VectorArithmeticVShift* - Rename VectorLogicalVShiftU* -> VectorLogicalVShift*	2020-04-22 20:55:50 +01:00
VelocityRa	c30b8dbe99	decoders: Cast to correctly-sized type before shifting Fixes decoding for 64-bit instructions Does not help/apply to any currently supported ARM versions (since all are 32-bit length or below), it's for future-proofing should such an arch be supported.	2020-04-22 20:55:50 +01:00
MerryMage	09bf273bc8	A64: Implement SCVTF, UCVTF (vector, fixed-point), scalar variant	2020-04-22 20:55:06 +01:00
MerryMage	f9129db6fd	A64: Implement FCVTZS, FCVTZU, UCVTF, SCVTF (vector, fixed-point), vector variant	2020-04-22 20:55:06 +01:00
Lioncash	48df9b9a7d	A64: Implement UQSHL's vector immediate and register variants	2020-04-22 20:55:06 +01:00
Lioncash	d426dfe942	ir: Add opcodes for unsigned saturating left shifts	2020-04-22 20:55:06 +01:00
Lioncash	ab60720418	A64/translate/impl: Make signatures consistent for unimplemented by-element SIMD variants Makes them all consistent, so it isn't necessary to change the prototypes over when implementing them.	2020-04-22 20:55:06 +01:00
Lioncash	6b5ea6ee66	A64: Implement BRK Currently, we can just implement this as part of the exception interface, similar to how it's done for the A32 interface with BKPT.	2020-04-22 20:55:06 +01:00
Lioncash	b915364c16	A64/imm: Add full range of comparison operators to Imm template Makes the comparison interface consistent by providing all of the relevant members. This also modifies the comparison operators to take the Imm instance by value, as it's really only a u32 under the covers, and it's cheaper to shuffle around a u32 than a 64-bit pointer address.	2020-04-22 20:55:06 +01:00
MerryMage	02150bc0b7	IR: Add fbits argument to FPVectorFrom{Signed,Unsigned}Fixed	2020-04-22 20:55:06 +01:00
MerryMage	027b0ef725	A64: Implement SCVTF, UCVTF (scalar, fixed-point)	2020-04-22 20:55:06 +01:00
MerryMage	8051f60db0	opcodes.inc: Align columns to a tabstop of 4	2020-04-22 20:55:06 +01:00
MerryMage	90193b0e3d	IR: Add fbits argument to FixedToFP-related opcodes	2020-04-22 20:55:06 +01:00
Lioncash	616a153c16	A64: Implement SQSHL's vector immediate variant	2020-04-22 20:55:06 +01:00
Lioncash	e8b0f25dff	A64: Implement SQSHL's vector register variant	2020-04-22 20:55:06 +01:00
Lioncash	b14eaaec46	ir: Add opcodes for left signed saturated shifts	2020-04-22 20:55:06 +01:00
Lioncash	da55ed7b31	branch: Make variables const where applicable	2020-04-22 20:55:06 +01:00
Lioncash	867b666285	move_wide: Make variables const where applicable	2020-04-22 20:55:06 +01:00
Lioncash	78024a9dc4	load_store_register_unprivileged: Make variables const where applicable	2020-04-22 20:55:06 +01:00
Lioncash	e45e5da610	load_store_register_immediate: Place conditional bodies on their own line Makes the conditionals visually consistent with the rest of the codebase.	2020-04-22 20:55:06 +01:00
Lioncash	b586cf3f56	load_store_load_literal: Make variables const where applicable	2020-04-22 20:55:06 +01:00
Lioncash	c3a3b9687e	data_processing_logical: Move datasize declarations after early-exit conditionals While we're at it, make variables const where applicable.	2020-04-22 20:55:06 +01:00
Lioncash	ed797e6540	data_processing_conditional_select: Make variables const where applicable Makes CSEL's function consistent with all of the others.	2020-04-22 20:55:06 +01:00
Lioncash	c82fa5ec5a	data_processing_addsub: Move datasize declarations after early-exit conditionals While we're at it, also make relevant variables const where applicable	2020-04-22 20:55:06 +01:00
Lioncash	f4a66d2477	data_processing_bitfield: Move datasize variables after early-exit conditionals Moves the declaration of datasize to the scope that it's used within. This also takes the opportunity to apply const where applicable, and make early-exits all vertically consistent with one another.	2020-04-22 20:55:06 +01:00
Lioncash	2e0fcd6161	A64: Implement CLS's vector variant Leverages CLZ like the integral variant does.	2020-04-22 20:55:06 +01:00
MerryMage	12243692f5	A64: Implement SQRDMULH (vector), vector variant	2020-04-22 20:55:06 +01:00
MerryMage	a9ffcf08b1	A64: Implement SQDMULL (vector), vector variant	2020-04-22 20:55:06 +01:00
MerryMage	3e447614c6	IR: Add VectorSignedSaturatedDoublingMultiplyLong	2020-04-22 20:55:06 +01:00
MerryMage	06b31448aa	emit_x64_vector: Changes to VectorSignedSaturatedDoublingMultiply * Return both the upper and lower parts of the multiply if required * SSE2 does not support the pmuldq instruction, do sign correction to an unsigned result instead * Improve port utilisation where possible (punpck instructions were a bottleneck)	2020-04-22 20:55:06 +01:00
MerryMage	08c0e017a5	IR: Implement Vector{Signed,Unsigned}Multiply{16,32}	2020-04-22 20:55:06 +01:00
Lioncash	112cff9ab9	A64: Implement CLZ's vector variant	2020-04-22 20:55:06 +01:00
Lioncash	e739624296	ir: Add opcodes for vector CLZ operations We can optimize these cases further for with the use of a fair bit of shuffling via pshufb and the use of masks, but given the uncommon use of this instruction, I wouldn't consider it to be beneficial in terms of amount of code to be worth it over a simple manageable naive solution like this. If we ever do hit a case where vectorized CLZ happens to be a bottleneck, then we can revisit this. At least with AVX-512CD, this can be done with a single instruction for the 32-bit word case.	2020-04-22 20:55:05 +01:00
MerryMage	d4c37a68a8	A64/translate: VectorZeroUpper for V(64) stores Ensures correctness.	2020-04-22 20:55:05 +01:00
MerryMage	b8daa4feac	simd_two_register_misc: FNEG (vector) with Q == 0 had dirty upper	2020-04-22 20:55:05 +01:00
Lioncash	14e026a7f0	A64: Implement USQADD's scalar and vector variants	2020-04-22 20:55:05 +01:00
Lioncash	d4a76aaa04	ir: Add opcodes form unsigned saturated accumulations of signed values	2020-04-22 20:55:05 +01:00
Lioncash	18ad7f237d	A64: Implement SUQADD's scalar and vector variants	2020-04-22 20:55:05 +01:00
Lioncash	6f911a26da	ir: Add opcodes for signed saturated accumulations of unsigned values	2020-04-22 20:55:05 +01:00
Lioncash	9a3d38d2ee	A64: Implement SMLAL{2}, SMLSL{2}, UMLAL{2}, and UMLSL{2}'s vector by-element variants We can simply modify the general function made for SMULL{2} and UMULL{2}'s by-element variants to also handle the other multiply-based by-element variants.	2020-04-22 20:55:05 +01:00
Lioncash	6ccfbc9b39	A64: Implement UMULL{2}'s vector by-element variant	2020-04-22 20:55:05 +01:00
Lioncash	58e21f175c	A64: Implement SMULL{2}'s vector by-element variant	2020-04-22 20:55:05 +01:00
Lioncash	134bb02e19	ir/value: Replace includes with forward declarations enum classes are still considered complete types when forward declared (as the compiler knows the exact size of the type from the declaration alone). The only difference in this case being that the members of the enum class aren't visible. Given we don't use the members within this header in any way, we can simply forward declare them here and remove the inclusions.	2020-04-22 20:55:05 +01:00
Lioncash	2c8e07e7d0	ir/cond: Migrate to C++17 nested namespace specifiers	2020-04-22 20:55:05 +01:00
Lioncash	0a3976059f	A64: Implement URSQRTE	2020-04-22 20:55:05 +01:00
Lioncash	b6e74fd17d	ir: Add opcodes for performing unsigned reciprocal square root estimates	2020-04-22 20:55:05 +01:00
Lioncash	bd3582e811	A64: Implement URECPE	2020-04-22 20:55:05 +01:00
Lioncash	af83360f89	ir: Add opcodes for unsigned reciprocal estimate	2020-04-22 20:55:05 +01:00
Lioncash	740ffa52ae	A64: Implement SQNEG's scalar and vector variant	2020-04-22 20:53:46 +01:00
Lioncash	fca7eddb9e	A64: Add opcodes for signed saturating negations	2020-04-22 20:53:46 +01:00
Lioncash	f5fb496e7e	A64: Implement SQDMULH's by-element scalar variant	2020-04-22 20:53:46 +01:00
Lioncash	40f0576995	A64: Implement SQDMULH's by-element vector variant	2020-04-22 20:53:46 +01:00
MerryMage	9b65100660	A64: Implement FastDispatchHint	2020-04-22 20:53:46 +01:00
MerryMage	f96c43d422	A32: Implement FastDispatchHint	2020-04-22 20:53:46 +01:00
MerryMage	aa8d826c13	ir/terminal: Add FastDispatchHint	2020-04-22 20:53:46 +01:00
Lioncash	1a69a61cb4	A64: Implement SQDMULH's scalar variant	2020-04-22 20:53:46 +01:00
Lioncash	7ebfd0f31c	ir: Add opcodes for scalar signed saturated doubling multiplies	2020-04-22 20:53:46 +01:00
Lioncash	9c03311fed	A64: Implement SQDMULH's vector variant	2020-04-22 20:53:46 +01:00
Lioncash	a0231e5546	ir: Add opcodes for signed saturated doubling multiplies	2020-04-22 20:53:46 +01:00
Lioncash	db24e1f09b	A64: Implement SQABS' scalar variant	2020-04-22 20:53:46 +01:00
Lioncash	bda5d14c7f	A64: Implement SQABS' vector variant.	2020-04-22 20:53:46 +01:00
Lioncash	0507e47420	ir: Add opcodes for signed saturated absolute values	2020-04-22 20:53:46 +01:00
MerryMage	3415828fb4	IR: Simplify FP{Single,Double}ToFixed{U,S}{32,64}	2020-04-22 20:53:46 +01:00
Lioncash	e30f9816ec	A32/decoder: Add missing <algorithm> includes These includes should be present, as we use std::find_if() within these headers.	2020-04-22 20:53:46 +01:00
Lioncash	053175f69b	ir_emitter: Rename fpscr_controlled parameters to fpcr_controlled Part of addressing #333	2020-04-22 20:53:46 +01:00
MerryMage	f0184c4b8d	a32/exception_generating: BPKT: Define unpredictable behaviour Define unpredictable behaviour to be BKPT executes conditionally	2020-04-22 20:53:46 +01:00
MerryMage	a12854857b	A32: Add define_unpredictable_behaviour option	2020-04-22 20:53:46 +01:00
MerryMage	b0abaa8312	A32/location_descriptor: Change formatting to use hex	2020-04-22 20:53:46 +01:00
MerryMage	ccbf6c7f63	microinstruction: A32ExceptionRaised causes CPU exception	2020-04-22 20:53:46 +01:00
MerryMage	6595e49a31	A32/types: CondToString: Add nv	2020-04-22 20:53:46 +01:00
MerryMage	f73104633b	a32_emit_x64: Fix incorrect BMI2 implementation for SetCpsr * The MSB for each byte in cpsr_ge were not being appropriately set. * We also expand test coverage to test this case. * We fix the disassembly of the MSR (imm) and MSR (reg) instructions as well.	2020-04-22 20:53:46 +01:00
MerryMage	3b13f1eb12	A64/translate: Standardize arguments of helper functions Don't pass in IREmitter when TranslatorVisitor is already available.	2020-04-22 20:53:45 +01:00
MerryMage	a4e556d59c	A64/translate: Standardize TranslatorVisitor abbreviation Prefer v to tv.	2020-04-22 20:53:45 +01:00
Lioncash	3d465e2c36	A64: Implement SQXTN, SQXTUN, and UQXTN's scalar variants We can implement these in terms of the vector variants	2020-04-22 20:53:45 +01:00
Lioncash	4ff39c6ea8	A64: Implement SDOT and UDOT's (by element) variants Gets all of the dot product instructions out of the way.	2020-04-22 20:53:45 +01:00
MerryMage	0c18b85c27	A64: Implement TBL and TBX	2020-04-22 20:53:45 +01:00
MerryMage	89d08c7d61	IR: Add VectorTable and VectorTableLookup IR instructions	2020-04-22 20:53:45 +01:00
MerryMage	0288974512	opcodes: Cleanup opcodes table * Remove T:: prefix from types. * Add another column for a 4th argument.	2020-04-22 20:53:45 +01:00
Lioncash	d9fc6cf31f	A64: Implement SDOT and UDOT's vector variant	2020-04-22 20:53:45 +01:00
Lioncash	cb5e5c5d49	A64: Implement SADALP and UADALP While we're at it we can join the code for SADDLP and UADDLP with these instructions, since the only difference is we do an accumulate at the end of the operation.	2020-04-22 20:53:45 +01:00
Lioncash	29f8b30634	A64: Implement SRSHL and URSHL Implements both scalar and vector variants.	2020-04-22 20:53:45 +01:00
Lioncash	0efa2ce3b0	ir: Add opcodes for performing rounding left shifts	2020-04-22 20:53:45 +01:00
Lioncash	f3f60cd179	A64: Implement ISB Given we want to ensure that all instructions are fetched again, we can treat an ISB instruction as a code cache flush.	2020-04-22 20:53:45 +01:00
Lioncash	be53e356a2	A64: Implement FCVTN{2}	2020-04-22 20:53:45 +01:00
Lioncash	4c3d7c5a8d	A64: Implement FCVTL{2}	2020-04-22 20:53:45 +01:00
Lioncash	7eb6be7a6a	A64: Implement FMAXNM and FMINNM vector variants. Currently we can implement these in terms of the scalar IR variants.	2020-04-22 20:53:45 +01:00
Lioncash	8b65ea68c0	A64: Implement FMAXP, FMAXNMP, FMINP, and FMINNMP's vector variants We can just implement these in terms of scalars for the time being.	2020-04-22 20:53:45 +01:00
MerryMage	8a3b6364c2	load_store_exclusive: Define s == t state to be Constraint_NONE Downstream (yuzu) mentioned that the instruction: STXR W9, W9, [X0] was executed in the program "Crash N-Sane Trilogy".	2020-04-22 20:53:45 +01:00
MerryMage	cd40e4dae0	A64/translate: Allow for unpredictable behaviour to be defined	2020-04-22 20:53:45 +01:00
MerryMage	d1d6f4feb5	system: Implement MRS CNTFRQ_EL0	2020-04-22 20:53:45 +01:00
Lioncash	7ef7def661	A64: Implement SQ{ADD, SUB}, and UQ{ADD, SUB}'s vector variants Currently we implement these in terms of the scalar variants. Falling back to the interpreter is slow enough to make it more effective than doing that.	2020-04-22 20:46:23 +01:00
Lioncash	a4b0e2ace6	A64: Implement UQADD/UQSUB's scalar variants	2020-04-22 20:46:23 +01:00
Lioncash	acbaf04fef	ir: Add opcodes for unsigned saturating add and subtract	2020-04-22 20:46:23 +01:00
Lioncash	2188765e28	ir/value: Use type alias CoprocessorInfo for std::array<u8, 8> Provides a more descriptive label for the interface, and avoids the need to hardcode the array size in multiple places.	2020-04-22 20:46:23 +01:00
MerryMage	71e137715d	status_register_access: Add support for bits 0 and 1 of mask to MSR	2020-04-22 20:46:23 +01:00
MerryMage	ac51c2547d	A32/translate/load_store: Correct detection of writeback	2020-04-22 20:46:23 +01:00
MerryMage	d345220251	A32/translate: Add TranslateSingleInstruction	2020-04-22 20:46:23 +01:00
MerryMage	5fc197c564	A32/ir_emitter: Bug fix: IREmitter::ExceptionRaised using incorrect opcode	2020-04-22 20:46:23 +01:00
MerryMage	ff3805e332	A32/decoders: Split instruction list into include file	2020-04-22 20:46:23 +01:00
MerryMage	3f4d118d73	microinstruction: Improve assert messages	2020-04-22 20:46:23 +01:00
MerryMage	f5e11d117a	A64: Implement FMULX, scalar single/double variant	2020-04-22 20:46:23 +01:00
MerryMage	17f73974f2	IR: Implement FPMulX IR instruction	2020-04-22 20:46:23 +01:00
MerryMage	9669e49817	A64: Implement FRINT{N,M,P,Z,A,X,I} (vector), single/double variant	2020-04-22 20:46:23 +01:00
MerryMage	f976c47008	IR: Initial implementation of FPVectorRoundInt	2020-04-22 20:46:23 +01:00
MerryMage	f2393488fe	A64: Implement SQADD and SQSUB, scalar variant	2020-04-22 20:46:23 +01:00
MerryMage	10e196480f	IR: Generalise SignedSaturated{Add,Sub} to support more bitwidths	2020-04-22 20:46:23 +01:00
Lioncash	d0fdd3c6e6	simd_three_same: Extract non-paired SMAX, SMIN, UMAX, UMIN code to a common function Deduplicates a bit of code and makes its layout consistent with the paired variants	2020-04-22 20:46:23 +01:00
Lioncash	2bea2d0512	A64: Implement SMAXP, SMINP, UMAXP, UMINP	2020-04-22 20:46:23 +01:00
Lioncash	463b9a3d02	ir: Add opcodes for vector paired maximum and minimums For the time being, we can just do a naive implementation which avoids falling back to the interpreter a bit. Horizontal operations aren't necessarily x86 SIMD's forte anyways.	2020-04-22 20:46:23 +01:00
Lioncash	43344c5400	A64: Implement SMAXV, SMINV, UMAXV, and UMINV	2020-04-22 20:46:23 +01:00
Lioncash	2501bfbfae	ir: Add opcodes for performing scalar integral min/max	2020-04-22 20:46:23 +01:00
Lioncash	7fdd8b0197	A64: Implement PMULL{2}	2020-04-22 20:46:23 +01:00
Lioncash	5ebf496d4e	translate: Deduplicate GetDataSize() functions Avoids defining the same function multiple times in different files.	2020-04-22 20:46:22 +01:00
Lioncash	f83cd2da9a	floating_point_{conditional}_compare: Deduplicate code Deduplicates the implementation code of instructions by extracting the code to a common function.	2020-04-22 20:46:22 +01:00
Lioncash	b48fb8ca6b	A64: Implement PMUL	2020-04-22 20:46:22 +01:00
Lioncash	affa312d1d	ir: Add opcode for performing polynomial multiplication	2020-04-22 20:46:22 +01:00
MerryMage	dd4ac86f8e	A64: Implement FCVT{N,M,A,P}{U,S} (vector), FCVTZU (vector, integer), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	28b38916a8	A64: Implement FCVTZS (vector, integer), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	507bcd8b8b	IR: Implement FPVectorTo{Signed,Unsigned}Fixed	2020-04-22 20:46:22 +01:00
Lioncash	c778c7b868	A64: Implement FMAX's vector single and double precision variants	2020-04-22 20:46:22 +01:00
Lioncash	009879d92b	A64: Implement FMIN's vector single and double precision variants	2020-04-22 20:46:22 +01:00
MerryMage	7b03da86c2	IR: Implement FPVector{Max,Min}	2020-04-22 20:46:22 +01:00
MerryMage	ddcff86f9c	microinstruction: Update ReadsFromAndWritesToFPSRCumulativeExceptionBits	2020-04-22 20:46:22 +01:00
MerryMage	10de36394e	A64: Implement FRECPS, vector/scalar single/double variants	2020-04-22 20:46:22 +01:00
MerryMage	901bd9b4e2	IR: Implement FPRecipStepFused, FPVectorRecipStepFused	2020-04-22 20:46:22 +01:00
MerryMage	f66f61d8ab	A64: Implement FRECPE, vector single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	939f5f5c7a	IR: Implement FPVectorRecipEstimate	2020-04-22 20:46:22 +01:00
MerryMage	27c73dd56a	A64: Implement FRECPE, scalar single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	c1dcfe29f7	IR: Implement FPRecipEstimate	2020-04-22 20:46:22 +01:00
MerryMage	642b6c31d2	A64: Implement MLA, MLS (by element), vector single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	0de37b11ad	A64: Implement FMLS (vector), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	04f325a05e	IR: Implement FPVectorNeg	2020-04-22 20:46:22 +01:00
MerryMage	934132e0c5	A64: Implement FMLA (vector), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	771a4fc20b	IR: Implement FPVectorMulAdd	2020-04-22 20:46:22 +01:00
MerryMage	1edd0125b2	mp: rename mp.h to mp/function_info.h	2020-04-22 20:46:22 +01:00
MerryMage	ecbf9dbae5	IR: Implement A64OrQC	2020-04-22 20:46:22 +01:00
MerryMage	f0fecf2615	A64: Implement UQSHRN, UQRSHRN (vector)	2020-04-22 20:46:22 +01:00
MerryMage	8f4c1a8558	emit_x64_vector: -0x80000000 isn't -0x80000000	2020-04-22 20:46:22 +01:00
MerryMage	b455b566e7	A64: Implement UQXTN (vector)	2020-04-22 20:46:22 +01:00
MerryMage	3874cb37e3	A64: Implement SQXTN (vector)	2020-04-22 20:46:22 +01:00
MerryMage	712c6c1d7e	A64: Implement SQSHRUN, SQRSHRUN (vector)	2020-04-22 20:46:22 +01:00
MerryMage	c5722ec963	simd_shift_by_immediate: Simplify ShiftRight	2020-04-22 20:46:22 +01:00
MerryMage	f020dbe4ed	A64: Implement SQXTUN	2020-04-22 20:46:22 +01:00
MerryMage	6918ef7360	microinstruction: Reorganize FPSCR related instruction queries	2020-04-22 20:46:22 +01:00
Lioncash	a639fa5534	microinstruction: Add missing FP scalar opcodes to ReadsFromFPSCR() and WritesToFPSCR() These were forgotten when the opcodes were added.	2020-04-22 20:46:22 +01:00
MerryMage	b2e4c16ef8	A64: Implement FRSQRTS (vector), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	45dc5f74f3	A64: Implement FRSQRTE (vector), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	b74d5520f9	A64: Implement FRSQRTS (scalar), single/double variant	2020-04-22 20:46:22 +01:00
MerryMage	506e544bfe	IR: Implement FPRSqrtStepFused	2020-04-22 20:46:22 +01:00
Lioncash	ace7d2ba50	A64: Implement FMAXP, FMINP, FMAXNMP and FMINNMP's scalar double/single-precision variant	2020-04-22 20:46:21 +01:00
Lioncash	49c7edf7c6	A64: Implement FMLA and FMLS (by element)'s double/single-precision scalar variant	2020-04-22 20:46:21 +01:00
Lioncash	c704acafe4	A64: Implement FMUL (by element)'s scalar double/single-precision variant	2020-04-22 20:46:21 +01:00
Lioncash	b7bd70fd19	A64: Implement FMAXV, FMINV, FMAXNMV, and FMINNMV	2020-04-22 20:46:21 +01:00
Lioncash	3447c82656	translate: Return by bool in helpers where applicable Gets rid of a bit of duplication regarding the early-out cases and makes all helpers functions consistent (previously some had a return type of bool, while others had a return type of void).	2020-04-22 20:46:21 +01:00
MerryMage	f837ce8e78	simd_scalar_two_register_misc: Implement FRSQRTE, scalar variant	2020-04-22 20:46:21 +01:00
MerryMage	bde58b04d4	IR: Implement FPRSqrtEstimate	2020-04-22 20:46:21 +01:00
MerryMage	16061c28f3	simd_vector_x_indexed_element: Implement FMUL (by element), vector variant	2020-04-22 20:46:21 +01:00
MerryMage	55eaa16615	a64_emit_x64: Ensure host has updated ticks in EmitA64GetCNTPCT Discovered by @Subv. Fixes incomplete fix begun in 5a91c94dca47c9702dee20fbd5ae1f4c07eef9df. That fix fails to take into account that LinkBlock doesn't update ticks until there are no remaining ticks to be executed. Test added to confirm fix.	2020-04-22 20:46:21 +01:00
Lioncash	e5d80e998e	A64: Implement SADDLV	2020-04-22 20:46:21 +01:00
Lioncash	a1bc8ddb53	A64: Implement UADDLV	2020-04-22 20:46:21 +01:00
Subv	4606a081c9	A64: The A64SetTPIDR IR instruction writes to a system register and should not be eliminated by the dead code elimination pass. Previously this instruction was alway eliminated, resulting in incorrect values for TPIDR_EL0.	2020-04-22 20:46:21 +01:00
MerryMage	b53127600b	fp: A64::FPCR -> FP::FPCR	2020-04-22 20:46:21 +01:00
MerryMage	699c5f36d5	system: Simplify static_cast	2020-04-22 20:46:21 +01:00
MerryMage	3f602129f4	system: Ensure value of CNTPCT_EL0 is accurate Since we currently only update the host's tick count at the end of a block, we force an end-of-block before executing a MRS %, CNTPCT_ELO instruction.	2020-04-22 20:46:21 +01:00
Lioncash	af3e23b224	simd_scalar_shift_by_immediate: Implement FCVT{ZS, ZU} (vector, fixed-point)'s scalar double/single-precision variant	2020-04-22 20:46:21 +01:00
Lioncash	91abf87169	simd_scalar_two_register_misc: Implement FCVT{AS, AU, MS, MU, NS, NU, PS, PU, ZS, ZU} (vector)'s scalar double/single-precision variants We can simply implement this in terms of the fixed-point IR opcodes.	2020-04-22 20:46:21 +01:00
MerryMage	e18fca17dc	A64: Implement FABD in terms of existing IR instructions Fixes NaN issue. Closes #306.	2020-04-22 20:46:21 +01:00
MerryMage	a40127a054	A64: Implement FRINTX, FRINTI (scalar)	2020-04-22 20:46:20 +01:00
MerryMage	962fa3b65e	A64: Implement FRINTP, FRINTM, FRINTZ (scalar)	2020-04-22 20:46:20 +01:00
MerryMage	5200bf41cf	A64: Implement FRINTN (scalar)	2020-04-22 20:46:20 +01:00
MerryMage	8718dc1692	A64: Implement FRINTA (scalar)	2020-04-22 20:46:20 +01:00
MerryMage	b228694012	IR: Implement FPRoundInt	2020-04-22 20:46:20 +01:00
Lioncash	f7f83b76b7	simd_scalar_two_register_misc: Implement scalar double/single-precision variants of FCM{EQ, GE, GT, LE, LT} (zero)	2020-04-22 20:46:20 +01:00
Lioncash	9db6d1e98b	translate_arm: Remove unnecessary rotr() function We already have RotateRight() in our common code, so we can remove this function and replace it with it. We can also implement ArmExpandImm_C() in terms of ArmExpandImm().	2020-04-22 20:46:20 +01:00
MerryMage	89e43867c1	A64: Implement FADDP (scalar)	2020-04-22 20:46:19 +01:00
MerryMage	33fa65de23	A64: Implement FADDP (vector)	2020-04-22 20:46:19 +01:00
MerryMage	9dba273a8c	A64: Implement SADDLP	2020-04-22 20:46:19 +01:00
MerryMage	70ff2d73b5	A64: Implement UADDLP	2020-04-22 20:46:19 +01:00
MerryMage	5563bbbd79	A64: Implement EXT	2020-04-22 20:46:19 +01:00
MerryMage	3d9677d094	A64: Implement FCVTMU (scalar)	2020-04-22 20:46:19 +01:00
MerryMage	79c9018d60	A64: Implement FCVTMS (scalar)	2020-04-22 20:46:19 +01:00
MerryMage	49c4499a87	A64: Implement FCVTPU (scalar)	2020-04-22 20:46:19 +01:00
MerryMage	af661ef5a6	A64: Implement FCVTPS (scalar)	2020-04-22 20:46:19 +01:00
MerryMage	27319822bb	A64: Implement FCVTAU (scalar)	2020-04-22 20:46:19 +01:00

... 6 7 8 9 10 ...

1449 commits