Algorithm 908: Online Exact Summation of Floating-Point Streams

Zhu Yong Kang<sup>*</sup>; Hayes Wayne B

doi:10.1145/1824801.1824815

摘要

We present a novel, online algorithm for exact summation of a stream of floating-point numbers. By "online" we mean that the algorithm needs to see only one input at a time, and can take an arbitrary length input stream of such inputs while requiring only constant memory. By "exact" we mean that the sum of the internal array of our algorithm is exactly equal to the sum of all the inputs, and the returned result is the correctly-rounded sum. The proof of correctness is valid for all inputs (including nonnormalized numbers but modulo intermediate overflow), and is independent of the number of summands or the condition number of the sum. The algorithm asymptotically needs only 5 FLOPs per summand, and due to instruction-level parallelism runs only about 2-3 times slower than the obvious, fast-but-dumb "ordinary recursive summation" loop when the number of summands is greater than 10,000. Thus, to our knowledge, it is the fastest, most accurate, and most memory efficient among known algorithms. Indeed, it is difficult to see how a faster algorithm or one requiring significantly fewer FLOPs could exist without hardware improvements. An application for a large number of summands is provided.

出版日期2010-9

全文

访问全文

收藏分享被引(7) 浏览

更新时间：2018-01-19 21:17

Algorithm 908: Online Exact Summation of Floating-Point Streams

摘要

全文

产品服务

站内浏览

服务支持

联系方式

科研之友