diff options
author | Joe Ramsay <Joe.Ramsay@arm.com> | 2023-06-28 12:19:36 +0100 |
---|---|---|
committer | Szabolcs Nagy <szabolcs.nagy@arm.com> | 2023-06-30 09:04:10 +0100 |
commit | aed39a3aa3ea68b14dce3395fb14b1416541e6c6 (patch) | |
tree | 38f866205e31b1bef745122636dbaa61922cb1cc /sysdeps/ieee754 | |
parent | 84e93afc734a3c30e35ed2d21466a44259ac577e (diff) | |
download | glibc-aed39a3aa3ea68b14dce3395fb14b1416541e6c6.tar.gz glibc-aed39a3aa3ea68b14dce3395fb14b1416541e6c6.tar.xz glibc-aed39a3aa3ea68b14dce3395fb14b1416541e6c6.zip |
aarch64: Add vector implementations of cos routines
Replace the loop-over-scalar placeholder routines with optimised implementations from Arm Optimized Routines (AOR). Also add some headers containing utilities for aarch64 libmvec routines, and update libm-test-ulps. Data tables for new routines are used via a pointer with a barrier on it, in order to prevent overly aggressive constant inlining in GCC. This allows a single adrp, combined with offset loads, to be used for every constant in the table. Special-case handlers are marked NOINLINE in order to confine the save/restore overhead of switching from vector to normal calling standard. This way we only incur the extra memory access in the exceptional cases. NOINLINE definitions have been moved to math_private.h in order to reduce duplication. AOR exposes a config option, WANT_SIMD_EXCEPT, to enable selective masking (and later fixing up) of invalid lanes, in order to trigger fp exceptions correctly (AdvSIMD only). This is tested and maintained in AOR, however it is configured off at source level here for performance reasons. We keep the WANT_SIMD_EXCEPT blocks in routine sources to greatly simplify the upstreaming process from AOR to glibc. Reviewed-by: Szabolcs Nagy <szabolcs.nagy@arm.com>
Diffstat (limited to 'sysdeps/ieee754')
-rw-r--r-- | sysdeps/ieee754/dbl-64/math_config.h | 3 | ||||
-rw-r--r-- | sysdeps/ieee754/flt-32/math_config.h | 2 |
2 files changed, 0 insertions, 5 deletions
diff --git a/sysdeps/ieee754/dbl-64/math_config.h b/sysdeps/ieee754/dbl-64/math_config.h index 499ca7591b..19af33fd86 100644 --- a/sysdeps/ieee754/dbl-64/math_config.h +++ b/sysdeps/ieee754/dbl-64/math_config.h @@ -148,9 +148,6 @@ make_double (uint64_t x, int64_t ep, uint64_t s) return asdouble (s + x + (ep << MANTISSA_WIDTH)); } - -#define NOINLINE __attribute__ ((noinline)) - /* Error handling tail calls for special cases, with a sign argument. The sign of the return value is set if the argument is non-zero. */ diff --git a/sysdeps/ieee754/flt-32/math_config.h b/sysdeps/ieee754/flt-32/math_config.h index a7fb332767..d1b06a1a90 100644 --- a/sysdeps/ieee754/flt-32/math_config.h +++ b/sysdeps/ieee754/flt-32/math_config.h @@ -151,8 +151,6 @@ make_float (uint32_t x, int ep, uint32_t s) return asfloat (s + x + (ep << MANTISSA_WIDTH)); } -#define NOINLINE __attribute__ ((noinline)) - attribute_hidden float __math_oflowf (uint32_t); attribute_hidden float __math_uflowf (uint32_t); attribute_hidden float __math_may_uflowf (uint32_t); |