In ARM Cortex-A8 processor, I understand what NEON is, it is an SIMD co-processor.

Question

0

Asked: May 17, 20262026-05-17T23:02:30+00:00 2026-05-17T23:02:30+00:00

In ARM Cortex-A8 processor, I understand what NEON is, it is an SIMD co-processor.

0

In ARM Cortex-A8 processor, I understand what NEON is, it is an SIMD co-processor.

But is VFP(Vector Floating Point) unit, which is also a co-processor, works as a SIMD processor? If so which one is better to use?

I read few links such as –

Link1
Link2.

But not really very clear what they mean. They say that VFP was never intended to be used for SIMD but on Wiki I read the following – “The VFP architecture also supports execution of short vector instructions but these operate on each vector element sequentially and thus do not offer the performance of true SIMD (Single Instruction Multiple Data) parallelism.“

It so not so clear what to believe, can anyone elaborate more on this topic?

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-05-17T23:02:30+00:00

There are quite some difference between the two. Neon is a SIMD (Single Instruction Multiple Data) accelerator processor as part of the ARM core. It means that during the execution of one instruction the same operation will occur on up to 16 data sets in parallel. Since there is parallelism inside the Neon, you can get more MIPS or FLOPS out of Neon than you can a standard SISD processor running at the same clock rate.

The biggest benefit of Neon is if you want to execute operation with vectors, i.e. video encoding/decoding. Also it can perform single precision floating point(float) operations in parallel.

VFP is a classic floating point hardware accelerator. It is not a parallel architecture like Neon. Basically it performs one operation on one set of inputs and returns one output. It’s purpose is to speed up floating point calculations. It supports single and double precision floating point.

You have 3 possibilities to use Neon:

use intrinsics functions #include “arm_neon.h”
inline the assembly code
let the gcc to do the optimizations for you by providing -mfpu=neon as argument (gcc 4.5 is good on this)

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

In ARM Cortex-A8 processor, I understand what NEON is, it is an SIMD co-processor.

Leave an answerCancel reply

1 Answer

Leave an answer
Cancel reply