freebsd-skq

Author	SHA1	Message	Date
Bruce Evans	85c309021f	Oops, when merging from the float version to the double versions, don't forget to translate "float" to "double". ucbtest didn't detect the bug, but exhaustive testing of the float case relative to the double case eventually did. The bug only affects args x with \|x\| ~> 2*19(pi/2) on non-i386 (i386 is broken in a different way for large args).	2008-01-20 04:09:44 +00:00
Bruce Evans	a9b721d6b2	Remove the float version of the kernel of arg reduction for pi/2, since it should never have existed and it has not been used for many years (floats are reduced faster using doubles). All relevant changes (just the workaround for broken assignment) have been merged to the double version.	2008-01-19 22:50:50 +00:00
Bruce Evans	5b62c3808e	Do an ordinary assignment in STRICT_ASSIGN() except for floats until there is a problem with non-floats (when i386 defaults to extra precision). This essentially restores yesterday's behaviour for doubles on i386 (since generic rint() isn't used and everywhere else assumed working assignment), but for arches that use the generic rint() it finishes restoring some of 1995's behaviour (don't waste time doing unnecessary store/load).	2008-01-19 22:05:14 +00:00
Bruce Evans	684217d889	Use STRICT_ASSIGN() for exp2f() and exp2() instead of a volatile variable hack for exp2f() only. The volatile variable had a surprisingly large cost for exp2f() -- 19 cycles or 15% on i386 in the worst case observed. This is only partly explained by there being several references to the variable, only one of which benefited from it being volatile. Arches that have working assignment are likely to benefit even more from not having any volatile variable. exp2() now has a chance of working with extra precision on i386. exp2() has even more references to the variable, so it would have been pessimized more by simply declaring the variable as volatile. Even the temporary volatile variable for STRICT_ASSIGN costs 5-10% on i386, (A64) so I will change STRICT_ASSIGN() to do an ordinary assignment until i386 defaults to extra precision.	2008-01-19 21:37:14 +00:00
Bruce Evans	fa7fdac725	Use STRICT_ASSIGN() for _kernel_rem_pio2f() and _kernel_rem_pio2f() instead of a volatile cast hack for the float version only. The cast hack broke with gcc-4, but this was harmless since the float version hasn't been used for a few years. Merge from the float version so that the double version has a chance of working on i386 with extra precision. See k_rem_pio2f.c rev.1.8 for the original hack. Convert to _FBSDID().	2008-01-19 20:02:55 +00:00
Bruce Evans	0814af48f7	Use STRICT_ASSIGN() for log1pf() and log1p() instead of a volatile cast hack for log1pf() only. The cast hack broke with gcc-4, resulting in ~1 million errors of more than 1 ulp, with a maximum error of ~1.5 ulps. Now the maximum error for log1pf() on i386 is 0.5034 ulps again (this depends on extra precision), and log1p() has a chance of working with extra precision. See s_log1pf.c 1.8 for the original hack. (It claims only 62343 large errors). Convert to _FBSDID(). Another thing broken with gcc-4 is the static const hack used for rcsids.	2008-01-19 18:13:21 +00:00
Bruce Evans	6a876b92fb	Use STRICT_ASSIGN() instead of assorted direct volatile hacks to work around assignments not working for gcc on i386. Now volatile hacks for rint() and rintf() don't needlessly pessimize so many arches and the remaining pessimizations (for arm and powerpc) can be avoided centrally. This cleans up after s_rint.c 1.3 and 1.13 and s_rintf.c 1.3 and 1.9: - s_rint.c 1.13 broke 1.3 by only using a volatile cast hack in 1 place when it was needed in 2 places, and the volatile cast hack stopped working with gcc-4. These bugs only affected correctness tests on i386 since i386 normally uses asm rint() and doesn't support the extra precision mode that would break assignments of doubles. - s_rintf.c 1.9 improved(?) on 1.3 by using a volatile variable hack instead of an extra-precision variable hack, but it declared 2 variables as volatile when only 1 variable needed to be volatile. This only affected speed tests on i386 since i386 uses asm rintf().	2008-01-19 16:37:57 +00:00
David Schultz	86c2e0c047	Use volatile hacks to make sure these functions generate an underflow exception when they're supposed to. Previously, gcc -O2 was optimizing away the statement that generated it.	2008-01-18 22:19:04 +00:00
David Schultz	3d2cc91218	Hook up exp2l() and related docs to the build.	2008-01-18 21:43:10 +00:00
David Schultz	5526551600	Introduce a new log(3) manpage and move the relevant functions there. Document exp2l() in exp(3), and remove the quaint discussion of topics such as what these functions were called on the HP-71B's variant of BASIC.	2008-01-18 21:43:00 +00:00
David Schultz	968b39e3b9	Implement exp2l(). There is one version for machines with 80-bit long doubles (i386, amd64, ia64) and one for machines with 128-bit long doubles (sparc64). Other platforms use the double version. I've only done runtime testing on i386. Thanks to bde@ for helpful discussions and bugfixes.	2008-01-18 21:42:46 +00:00
Bruce Evans	1880ccbd79	Add a macro STRICT_ASSIGN() to help avoid the compiler bug that assignments and casts don't clip extra precision, if any. The implementation is to assign to a temporary volatile variable and read the result back to assign to the original lvalue. lib/msun currently 2 different hard-coded hacks to avoid the problem in just a few places and needs it in a few more places. One variant uses volatile for the original lvalue. This works but is slower than necessary. Another temporarily casts the lvalue to volatile. This broke with gcc-4.2.1 or earlier (gcc now stores to the lvalue but doesn't load from it).	2008-01-17 17:02:11 +00:00
David Schultz	a2d171e440	Optimize this a bit better. Submitted by: bde (although these aren't all of his changes)	2008-01-15 23:31:24 +00:00
David Schultz	d3f9671a7d	Implement rintl(), nearbyintl(), lrintl(), and llrintl(). Thanks to bde@ for feedback and testing of rintl().	2008-01-14 02:12:07 +00:00
David Schultz	73b2958b94	- Correct the range check in the double version to catch negative values that would overflow. - Style fixes and improved handling of NaNs suggested by bde.	2008-01-11 04:18:25 +00:00
David Schultz	45310fdb5d	Grumble. DO declare logbl(), DON'T declare logl() just yet. bde is going to commit logl() Real Soon Now. I'm just trying to slow him down with merge conflicts. Noticed by: bde	2007-12-20 03:16:55 +00:00
David Schultz	58c9a67ed7	Remove the declaration of logl(). The relevant bits haven't been committed yet, but the declaration leaked in when I added nan() and friends. Reported by: pav	2007-12-20 00:06:33 +00:00
David Schultz	7cd4a83267	Since nan() is supposed to work the same as strtod("nan(...)", NULL), my original implementation made both use the same code. Unfortunately, this meant libm depended on a vendor header at compile time and previously- unexposed vendor bits in libc at runtime. Hence, I just wrote my own version of the relevant vendor routine. As it turns out, mine has a factor of 8 fewer of lines of code, and is a bit more readable anyway. The strtod() and *scanf() routines still use vendor code. Reviewed by: bde	2007-12-18 23:46:32 +00:00
David Schultz	0ba1fd2f72	Remove z_abs(). The z_*() functions were in libf77, and for some reason someone thought it would be a good idea to copy z_abs() to libm in 1994. However, it's never been declared or documented anywhere, and I'm reasonably confident that nobody uses it. Discussed with: bde, deischen, kan	2007-12-18 01:15:20 +00:00
Bruce Evans	ccef8c4fcb	Oops, the previous commit was not needed -- the file was committed but not checked out due to my checkout error.	2007-12-17 18:21:23 +00:00
Bruce Evans	a18b106ffc	Translate from the i386 so that this compiles and runs. I hope that this and the i386 version of it will not be needed, but this is currently about 16 cycles or 36% faster than the C version, and the i386 version is about 8 cycles or 19% faster than the C version, due to poor optimization of the C version.	2007-12-17 18:12:06 +00:00
Bruce Evans	9ed67737f2	Don't try to build s_nanl.c before it is committed.	2007-12-17 13:20:38 +00:00
David Schultz	6821aba9e5	Add logbl(3) to libm.	2007-12-17 03:53:38 +00:00
David Schultz	3be0479b4c	Document the fact that we have nan(3) now, and make some minor clarifications in other places.	2007-12-17 01:04:43 +00:00
David Schultz	4b6b574455	Implement and document nan(), nanf(), and nanl(). This commit adds two new directories in msun: ld80 and ld128. These are for long double functions specific to the 80-bit long double format used on x86-derived architectures, and the 128-bit format used on sparc64, respectively.	2007-12-16 21:19:28 +00:00
David Schultz	3c27af2a44	1. Add csqrt{,f}(3). 2. Put carg{,f}(3) under the FBSD_1.1 namespace where it belongs (requested by kan@)	2007-12-15 08:39:03 +00:00
David Schultz	aaf70b2314	Implement and document csqrt(3) and csqrtf(3).	2007-12-15 08:38:44 +00:00
David Schultz	ce448a2e74	Update the standards section, and make a minor clarification about the return value of sqrt.	2007-12-14 07:53:09 +00:00
David Schultz	80f974729f	Typo in previous commit	2007-12-14 03:08:10 +00:00
David Schultz	39ebc398b6	Symbol.map additions for carg and cargf. (They're in C99, so I didn't add a new version for them.)	2007-12-14 03:06:50 +00:00
David Schultz	9768c3fea8	s/C90/C99/	2007-12-12 23:50:00 +00:00
David Schultz	367d55260f	Add a "STANDARDS" section.	2007-12-12 23:49:40 +00:00
David Schultz	205bd64894	Implement carg(3) and cargf(3). Rotting in an old src tree since: March 2005	2007-12-12 23:43:51 +00:00
Bruce Evans	b5e547df33	Oops, back out previous commit since it was backwards to a wrong branch.	2007-06-14 05:57:13 +00:00
Bruce Evans	d382c5ebb4	MFC: 1.11: fix the threshold for (not) using the simple Taylor approximation.	2007-06-14 05:51:00 +00:00
Bruce Evans	a8a2e00ebf	Fix an aliasing bug which was finally detected by gcc-4.2. fdlibm has hundreds of similar aliasing bugs, but all except this one seem to have been fixed by Cygnus and/or NetBSD before the modified version of fdlibm was imported into FreeBSD in 1994. PR: standards/113147 Submitted by: Steve Kargl <sgk@troutmask.apl.washington.edu>	2007-06-11 07:48:52 +00:00
Bruce Evans	20a990117d	Merge the relevant part of rev.1.14 of s_cbrt.c (a micro-optimization involving moving the check for x == 0). The savings in cycles are smaller for cbrtf() than for cbrt(), and positive in all measured cases with gcc-3.4.4, but still very machine/compiler-dependent.	2007-05-29 07:13:07 +00:00
Daniel Eischen	419ecd5dee	Bump library versions in preparation for 7.0. Ok'd by: kan	2007-05-21 02:49:08 +00:00
Daniel Eischen	00fb440c1a	Enable symbol versioning by default. Use WITHOUT_SYMVER to disable it. Warning, after symbol versioning is enabled, going back is not easy (use WITHOUT_SYMVER at your own risk). Change the default thread library to libthr. There most likely still needs to be a version bump for at least the thread libraries. If necessary, this will happen later.	2007-05-13 14:12:40 +00:00
Bruce Evans	9698b3b564	Don't assume that int is signed 32-bits in one place. Keep assuming that ints have >= 31 value bits elsewhere. s/int/int32_t/ seems to have been done too globally for all other files in msun/src before msun/ was imported into FreeBSD. Minor fixes in comments. e_lgamma_r.c: Describe special cases in more detail: - exception for lgamma(0) and lgamma(neg.integer) - lgamma(-Inf) = Inf. This is wrong but is required by C99 Annex F. I hope to change this.	2007-05-02 16:54:22 +00:00
Bruce Evans	e95cc9b700	Fix tgamma() on some special args: (1) tgamma(-Inf) returned +Inf and failed to raise any exception, but should always have raised an exception, and should behave like tgamma(negative integer). (2) tgamma(negative integer) returned +Inf and raised divide-by-zero, but should return NaN and raise "invalid" on any IEEEish system. (3) About half of the 252 negative intgers between -253 and -2**52 were misclassified as non-integers by using floor(x + 0.5) to round to nearest, so tgamma(x) was wrong (+-0 instead of +Inf and now NaN) on these args. The floor() expression is hard to use since rounding of (x + 0.5) may give x or x + 1, depending on \|x\| and the current rounding mode. The fixed version uses ceil(x) to classify x before operating on x and ends up being more efficient since ceil(x) is needed anyway. (4) On at least the problematic args in (3), tgamma() raised a spurious inexact. (5) tgamma(large positive) raised divide-by-zero but should raise overflow. (6) tgamma(+Inf) raised divide-by-zero but should not raise any exception. (7) Raise inexact for tiny \|x\| in a way that has some chance of not being optimized away. The fix for (5) and (6), and probably for (2), also prevents -O optimizing away the exception. PR: 112180 (2) Standards: Annex F in C99 (IEC 60559 binding) requires (1), (2) and (6).	2007-05-02 15:24:49 +00:00
Bruce Evans	dd936b27fc	Document (in a comment) the current (slightly broken) handling of special values in more detail, and change the style of this comment to be closer to fdlibm and C99: - tgamma(-Inf) was undocumented and is wrong (+Inf, should be NaN) - tgamma(negative integer) is as intended (+Inf) but not best for IEEE-754 (NaN) - tgamma(-0) was documented as being wrong (+Inf) but was correct (-Inf) - documentation of setting of exceptions (overflow, etc.) was more complete here than in most of libm, but was further from matching the actual setting than in most of libm, due to various bugs here (primarily, always evaluating +Inf one/zero and getting unwanted divide-by-zero exceptions from this). Now the actual behaviour with gcc -O0 is documented. Optimization still breaks setting of exceptions all over libm, so nothing can depend on this working. - tgamma(NaN)'s exception was documented as being wrong (invalid) but was correct (no exception with IEEEish NaNs). Finish (?) rev.1.5. gamma was not renamed to tgamma in one place. Finish (?) rev.1.6. errno.h was not completely removed.	2007-05-02 13:49:28 +00:00
Daniel Eischen	5f864214bb	Use C comments since we now preprocess these files with CPP.	2007-04-29 14:05:22 +00:00
Warner Losh	ee7093a640	Remove California Regent's clause 3, per letter	2007-01-09 01:02:06 +00:00
David Schultz	9abb1ff616	Implement modfl().	2007-01-07 07:54:21 +00:00
David Schultz	8185b32b5a	Fix a problem relating to fesetenv() clobbering i387 register stack. Details: As a side-effect of restoring a saved FP environment, fesetenv() overwrites the tag word, which indicates which i387 registers are in use. Normally this isn't a problem because the calling convention requires the register stack to be empty on function entry and exit. However, fesetenv() is inlined, so we need to tell gcc explicitly that the i387 registers get clobbered. PR: 85101	2007-01-06 21:46:23 +00:00
David Schultz	93e0663877	Fix a cut-and-paste-o.	2007-01-06 21:23:20 +00:00
David Schultz	829d55ac9c	Correctly handle NaN.	2007-01-06 21:22:57 +00:00
David Schultz	9fa229fc8d	Correctly handle inf/nan. This routine is currently unused because we seem to have assembly versions for all architectures, but it can't hurt to fix it.	2007-01-06 21:22:38 +00:00
David Schultz	6642c3fa74	Remove modf from libm's symbol map. It's actually in libc for hysterical raisins.	2007-01-06 21:18:17 +00:00

1 2 3 4 5 ...

433 Commits