Re: [PATCH 28/76] target/arm: Implement FPCR.FIZ handling

qemu-devel

[Date Prev][Date Next][Thread Prev][Thread Next][Date Index][Thread Index]

Re: [PATCH 28/76] target/arm: Implement FPCR.FIZ handling

From:	Richard Henderson
Subject:	Re: [PATCH 28/76] target/arm: Implement FPCR.FIZ handling
Date:	Sat, 25 Jan 2025 09:25:56 -0800
User-agent:	Mozilla Thunderbird

On 1/24/25 08:27, Peter Maydell wrote:

Part of FEAT_AFP is the new control bit FPCR.FIZ.  This bit affects
flushing of single and double precision denormal inputs to zero for
AArch64 floating point instructions.  (For half-precision, the
existing FPCR.FZ16 control remains the only one.)

FPCR.FIZ differs from FPCR.FZ in that if we flush an input denormal
only because of FPCR.FIZ then we should *not* set the cumulative
exception bit FPSR.IDC.

FEAT_AFP also defines that in AArch64 the existing FPCR.FZ only
applies when FPCR.AH is 0.

We can implement this by setting the "flush inputs to zero" state
appropriately when FPCR is written, and by not reflecting the
float_flag_input_denormal status flag into FPSR reads when it is the
result only of FPSR.FIZ.

Signed-off-by: Peter Maydell <peter.maydell@linaro.org>
---
  target/arm/vfp_helper.c | 58 ++++++++++++++++++++++++++++++++++-------
  1 file changed, 48 insertions(+), 10 deletions(-)

diff --git a/target/arm/vfp_helper.c b/target/arm/vfp_helper.c
index 8c79ab4fc8a..5a0b389f7a3 100644
--- a/target/arm/vfp_helper.c
+++ b/target/arm/vfp_helper.c
@@ -61,19 +61,29 @@ static inline uint32_t vfp_exceptbits_from_host(int 
host_bits)

static uint32_t vfp_get_fpsr_from_host(CPUARMState *env)

  {
-    uint32_t i = 0;
+    uint32_t a32_flags = 0, a64_flags = 0;

- i |= get_float_exception_flags(&env->vfp.fp_status_a32);

-    i |= get_float_exception_flags(&env->vfp.fp_status_a64);
-    i |= get_float_exception_flags(&env->vfp.standard_fp_status);
+    a32_flags |= get_float_exception_flags(&env->vfp.fp_status_a32);
+    a32_flags |= get_float_exception_flags(&env->vfp.standard_fp_status);
      /* FZ16 does not generate an input denormal exception.  */
-    i |= (get_float_exception_flags(&env->vfp.fp_status_f16_a32)
+    a32_flags |= (get_float_exception_flags(&env->vfp.fp_status_f16_a32)
            & ~float_flag_input_denormal_flushed);
-    i |= (get_float_exception_flags(&env->vfp.fp_status_f16_a64)
+    a32_flags |= (get_float_exception_flags(&env->vfp.standard_fp_status_f16)
            & ~float_flag_input_denormal_flushed);
-    i |= (get_float_exception_flags(&env->vfp.standard_fp_status_f16)
+
+    a64_flags |= get_float_exception_flags(&env->vfp.fp_status_a64);
+    a64_flags |= (get_float_exception_flags(&env->vfp.fp_status_f16_a64)
            & ~float_flag_input_denormal_flushed);
-    return vfp_exceptbits_from_host(i);
+    /*
+     * Flushing an input denormal only because FPCR.FIZ == 1 does
+     * not set FPSR.IDC. So squash it unless (FPCR.AH == 0 && FPCR.FZ == 1).
+     * We only do this for the a64 flags because FIZ has no effect
+     * on AArch32 even if it is set.
+     */
+    if ((env->vfp.fpcr & (FPCR_FZ | FPCR_AH)) != FPCR_FZ) {
+        a64_flags &= ~float_flag_input_denormal_flushed;
+    }

It might be worth pointing to FPUnpackBase pseudocode to say if both FZ and FIZ set, FZtakes precedence for setting IDC.


Anyway,
Reviewed-by: Richard Henderson <richard.henderson@linaro.org>


r~

[Prev in Thread]

Current Thread

[Next in Thread]

Re: [PATCH 24/76] fpu: allow flushing of output denormals to be after rounding, (continued)
- [PATCH 16/76] target/arm: Use FPST_FPCR_F16_A32 in A32 decoder, Peter Maydell, 2025/01/24
  - Re: [PATCH 16/76] target/arm: Use FPST_FPCR_F16_A32 in A32 decoder, Richard Henderson, 2025/01/25
- [PATCH 14/76] target/arm: Use fp_status_f16_a32 in AArch32-only helpers, Peter Maydell, 2025/01/24
  - Re: [PATCH 14/76] target/arm: Use fp_status_f16_a32 in AArch32-only helpers, Richard Henderson, 2025/01/25
- [PATCH 19/76] fpu: Rename float_flag_input_denormal to float_flag_input_denormal_flushed, Peter Maydell, 2025/01/24
  - Re: [PATCH 19/76] fpu: Rename float_flag_input_denormal to float_flag_input_denormal_flushed, Richard Henderson, 2025/01/25
- [PATCH 18/76] target/arm: Remove now-unused vfp.fp_status_f16 and FPST_FPCR_F16, Peter Maydell, 2025/01/24
  - Re: [PATCH 18/76] target/arm: Remove now-unused vfp.fp_status_f16 and FPST_FPCR_F16, Richard Henderson, 2025/01/25
- [PATCH 28/76] target/arm: Implement FPCR.FIZ handling, Peter Maydell, 2025/01/24
  - Re: [PATCH 28/76] target/arm: Implement FPCR.FIZ handling, Richard Henderson <=
- [PATCH 20/76] fpu: Rename float_flag_output_denormal to float_flag_output_denormal_flushed, Peter Maydell, 2025/01/24
  - Re: [PATCH 20/76] fpu: Rename float_flag_output_denormal to float_flag_output_denormal_flushed, Richard Henderson, 2025/01/25
- [PATCH 23/76] fpu: Implement float_flag_input_denormal_used, Peter Maydell, 2025/01/24
  - Re: [PATCH 23/76] fpu: Implement float_flag_input_denormal_used, Richard Henderson, 2025/01/25
- [PATCH 30/76] target/arm: Adjust exception flag handling for AH = 1, Peter Maydell, 2025/01/24
  - Re: [PATCH 30/76] target/arm: Adjust exception flag handling for AH = 1, Richard Henderson, 2025/01/25
- [PATCH 35/76] target/arm: Use FPST_FPCR_AH for BFMLAL*, BFMLSL* insns, Peter Maydell, 2025/01/24
  - Re: [PATCH 35/76] target/arm: Use FPST_FPCR_AH for BFMLAL*, BFMLSL* insns, Richard Henderson, 2025/01/25
- [PATCH 06/76] target/arm: Define new fp_status_a32 and fp_status_a64, Peter Maydell, 2025/01/24
  - Re: [PATCH 06/76] target/arm: Define new fp_status_a32 and fp_status_a64, Richard Henderson, 2025/01/25

Prev by Date: Re: [PATCH 27/76] target/arm: Define FPCR AH, FIZ, NEP bits
Next by Date: Re: [PATCH 29/76] target/arm: Adjust FP behaviour for FPCR.AH = 1
Previous by thread: [PATCH 28/76] target/arm: Implement FPCR.FIZ handling
Next by thread: [PATCH 20/76] fpu: Rename float_flag_output_denormal to float_flag_output_denormal_flushed
Index(es):
- Date
- Thread