Design tradeoff analysis of floating-point adders in FPGAs

被引:7
作者
Malik, Ali [1 ]
Chen, Dongdong [1 ]
Choi, Younhee [1 ]
Lee, Moon Ho [2 ]
Ko, Seok-Bum [1 ]
机构
[1] Univ Saskatchewan, Dept Elect & Comp Engn, Saskatoon, SK S7N 5A9, Canada
[2] Chonbuk Natl Univ, Elect & Informat Engn Dept, Jeonju 561756, Jeonbuk, South Korea
来源
CANADIAN JOURNAL OF ELECTRICAL AND COMPUTER ENGINEERING-REVUE CANADIENNE DE GENIE ELECTRIQUE ET INFORMATIQUE | 2008年 / 33卷 / 3-4期
关键词
floating-point adder; FPGA;
D O I
10.1109/CJECE.2008.4721634
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
With gate counts of ten million, field-programmable gate arrays (FPGAs) are becoming suitable for floating-point computations. Addition is the most complex operation in a floating-point unit and can cause major delay while requiring a significant area. Over the years, the VLSI community has developed many floating-point adder algorithms aimed primarily at reducing the overall latency. An efficient design of the floating-point adder offers major area and performance improvements for FPGAs. Given recent advances in FPGA architecture and area density, latency has become the main focus in attempts to improve performance. This paper studies the implementation of standard; leading-one predictor (LOP); and far and close datapath (2-path) floating-point addition algorithms in FPGAs. Each algorithm has complex sub-operations which contribute significantly to the overall latency of the design. Each of the sub-operations is researched for different implementations and is then synthesized onto a Xilinx Vertex-II Pro FPGA device. Standard and LOP algorithms are also pipelined into five stages and compared with the Xilinx IP. According to the results, the standard algorithm is the best implementation with respect to area, but has a large overall latency of 27.059 ns while occupying 541 slices. The LOP algorithm reduces latency by 6.5% at the cost of a 38% increase in area compared to the standard algorithm. The 2-path implementation shows a 19% reduction in latency with an added expense of 88% in area compared to the standard algorithm. The five-stage standard pipeline implementation shows a 6.4% improvement in clock speed compared to the Xilinx IP with a 23% smaller area requirement. The five-stage pipelined LOP implementation shows a 22% improvement in clock speed compared to the Xilinx IP at a cost of 15% more area.
引用
收藏
页码:169 / 175
页数:7
相关论文
共 50 条
[31]   Evaluation of a Floating-Point Intensive Kernel on FPGA [J].
Jin, Zheming ;
Finkel, Hal ;
Yoshii, Kazutomo ;
Cappello, Franck .
EURO-PAR 2017: PARALLEL PROCESSING WORKSHOPS, 2018, 10659 :664-675
[32]   FPGA Implementation of a Custom Floating-Point Library [J].
Campos, Nelson ;
Edirisinghe, Eran ;
Fatima, Shaheen ;
Chesnokov, Slava ;
Lluis, Alexis .
INTELLIGENT SYSTEMS AND APPLICATIONS, VOL 2, 2023, 543 :527-542
[33]   Floating Point Hardware for Embedded Processors in FPGAs: Design Space Exploration for Performance and Area [J].
Rodolfo, Taciano A. ;
Calazans, Ney L. V. ;
Moraes, Fernando G. .
2009 INTERNATIONAL CONFERENCE ON RECONFIGURABLE COMPUTING AND FPGAS, 2009, :24-29
[34]   ASIC Design of Nanoscale Artificial Neural Networks for Inference/Training by Floating-Point Arithmetic [J].
Niknia, Farzad ;
Wang, Ziheng ;
Liu, Shanshan ;
Reviriego, Pedro ;
Louri, Ahmed ;
Lombardi, Fabrizio .
IEEE TRANSACTIONS ON NANOTECHNOLOGY, 2024, 23 :208-216
[35]   Improving Synthesis of Fixed-Point Adders on FPGAs using Primitive Instantiations [J].
Khurshid, Burhan ;
Naaz, Roohie .
2015 ANNUAL IEEE INDIA CONFERENCE (INDICON), 2015,
[36]   An Autonomous Vector/Scalar Floating Point Coprocessor for FPGAs [J].
Kathiara, Jainik ;
Leeser, Miriam .
2011 IEEE 19TH ANNUAL INTERNATIONAL SYMPOSIUM ON FIELD-PROGRAMMABLE CUSTOM COMPUTING MACHINES (FCCM), 2011, :33-36
[37]   A floating-point multiplier based on angle representation method [J].
Gan, Bo ;
Wang, Kang ;
Wang, Guangsen ;
Zheng, Huiji ;
Chen, Guoyong .
INTERNATIONAL JOURNAL OF CIRCUIT THEORY AND APPLICATIONS, 2024, 52 (12) :6479-6487
[38]   An IEEE floating-point adder for multi-operation [J].
Huang, P ;
Wen, ZP ;
Yu, LX .
2004: 7TH INTERNATIONAL CONFERENCE ON SOLID-STATE AND INTEGRATED CIRCUITS TECHNOLOGY, VOLS 1- 3, PROCEEDINGS, 2004, :2098-2101
[39]   Enhanced Floating-Point Adder with Full Denormal Support [J].
Sohn, Jongwook ;
Dean, David K. ;
Quintana, Eric ;
Wong, Wing Shek .
2022 IEEE 29TH SYMPOSIUM ON COMPUTER ARITHMETIC (ARITH 2022), 2022, :35-42
[40]   CuFP: An HLS Library for Customized Floating-Point Operators [J].
Hajizadeh, Fahimeh ;
Ould-Bachir, Tarek ;
David, Jean Pierre .
ELECTRONICS, 2024, 13 (14)