SpiceQA
Questions Tags Users Badges

fma

11 Questions
Newest Active Unanswered Frequent
Score
View
Card Compact
Fast fixed-size polynomial evaluation: MSVC vs GCC
user_43494070
• asked Aug 16, 2022
3
0
43
compiler-optimization fma gcc visual-c++ c++
Fastest way to multiply and sum/add two arrays (dot product) - unaligned surprisingly faster than FMA
Petoj_585530
• asked Mar 26, 2022
5
1
334
.net-6.0 avx2 fma intrinsics c#
Terminology: why "floating multiply-add" instead of "fused multiply-add"?
user_17782750
• asked Feb 2, 2022
4
1
89
fma language-lawyer floating-point c terminology
Why Fma code is performing worse than Avx?
user_171060330
• asked Oct 8, 2021
3
0
179
avx fma benchmarking c#
CUDA half float operations without explicit intrinsics
user_3011660
• asked Jan 7, 2021
2
1
292
half-precision-float nvcc fma intrinsics cuda
How to refine floating-point division on FMA-capable GPUs?
user_47550750
• asked Dec 24, 2020
2
1
178
fma floating-point math division gpu
More aggresive optimization for FMA operations
user_145763520
• asked Nov 4, 2020
3
3
142
fma clang gcc c++
How advantageous is using fused multiply-accumulate for double-precision?
user_31169360
• asked Jun 9, 2020
3
1
546
fma x86-64 assembly c++ performance
Using FMA instructions for an FFT algorithm
user_25929170
• asked Mar 26, 2020
5
1
303
fma fft signal-processing c++
fmad=false gives good performance
user_3569220
• asked Aug 17, 2012
5
1
2788
fma nvidia cuda
  • 1 (current)
  • 2
  • NextNext
Hot Questions
Terms of service Privacy policy
Powered by Answer