EP1447752A2

Method and system for multi-processor FFT/IFFT with minimum inter-processor data communication

Abstract

The present invention provides a scalable method for implementing FFT/IFFT computations in multiprocessor architectures that provides improved throughput by eliminating the need for inter-processor communication after the computation of the first "log2P" stages for an implementation using "P" processing elements, comprising computing each butterfly of the first "log2P" stages on either a single processor or each of the "P" processors simultaneously and distributing the computation of the butterflies in all the subsequent stages among the "P" processors such that each chain of cascaded butterflies consisting of those butterflies that have inputs and outputs connected together, are processed by the same processor. The invention also provides a system for obtaining scalable implementation of FFT/IFFT computations in multiprocessor architectures that provides improved throughput by eliminating the need for inter-processor communication after the computation of the first "log2P" stages for an implementation using "P" processing elements.

EP1447752A2, drawing sheet 1
Sheet 1 of 5

Term

Term ended

Projected expiry passed 16 February 2024, 2.6 years ago.

  1. Priority
  2. Filed
  3. Published
  4. Projected expiry
  5. Today

8 claims: 2 independent, 6 dependent

  1. 1
    A scalable method for implementing FFT/IFFT computations in multiprocessor architectures that provides improved throughput by eliminating the need for inter-processor communication after the computation of the first "log 2 P" stages for an implementation using "P" processing elements, comprising the steps of :- computing each butterfly of the first "log 2 P" stages on either a single processor or each of the "P" processors simultaneously, - distributing the computation of the butterflies in all the subsequent stages among the "P" proce ssors such that each chain of cascaded butterflies consisting of those butterflies that have inputs and outputs connected together, are processed by the same processor.
  2. 5
    A system for obtaining scalable implement ation of FFT/IFFT computations in multiprocessor architectures that provides improved throughput by eliminating the need for inter-processor communication after the computation of the first "log 2 P" stages for an implementation using "P" processing elements, comprising :- a means for computing each butterfly of the first "log 2 P" stages on either a single processor or each of the "P" processors simultaneously, - an addressing means for distributing the computation of the butterflies in all the subsequent stages among the "P" processors such that each chain of cascaded butterflies consisting of those butterflies that have inputs and outputs connected together, are processed by the same processor.