Faithfully Rounded Floating-point Computations

Lange, Marko; Rump, Siegfried M.

Faithfully Rounded Floating-point Computations

Publikationstyp

Journal Article

Date Issued

2020-09

Sprache

English

Author(s)

Lange, Marko

Rump, Siegfried M.

Institut

Zuverlässiges Rechnen E-19

TORE-URI

http://hdl.handle.net/11420/7628

Journal

ACM transactions on mathematical software

Volume

46

Issue

3

Article Number

3290955

Citation

ACM Transactions on Mathematical Software 3 (46): 3290955 (2020-09)

Publisher DOI

10.1145/3290955

Scopus ID

2-s2.0-85092335707

We present a pair arithmetic for the four basic operations and square root. It can be regarded as a simplified, more-efficient double-double arithmetic. The central assumption on the underlying arithmetic is the first standard model for error analysis for operations on a discrete set of real numbers. Neither do we require a floating-point grid nor a rounding to nearest property. Based on that, we define a relative rounding error unit u and prove rigorous error bounds for the computed result of an arbitrary arithmetic expression depending on u, the size of the expression, and possibly a condition measure. In the second part of this note, we extend the error analysis by examining requirements to ensure faithfully rounded outputs and apply our results to IEEE 754 standard conform floating-point systems. For a class of mathematical expressions, using an IEEE 754 standard conform arithmetic with base β, the result is proved to be faithfully rounded for up to 1 / √βu-2 operations. Our findings cover a number of previously published algorithms to compute faithfully rounded results, among them Horner's scheme, products, sums, dot products, or Euclidean norm. Beyond that, several other problems can be analyzed, such as polynomial interpolation, orientation problems, Householder transformations, or the smallest singular value of Hilbert matrices of large size.

Subjects

Double-double

inaccurate cancellation

rigorous error bounds

Options

Faithfully Rounded Floating-point Computations