Add FFT related operators and APIs #35665

iclementine · 2021-09-10T17:22:03Z

PR types

New features

PR changes

OPs

Describe

Add fft related functionalities.

paddle.fft subpackage

differentiable 1d to nd complex to complex fast fourier transform. (fft, fft2, fftn, ifft, ifft2, ifftn)
differentiable 1d to nd real to complex fast fourier transform. (rfft, rfft2, rfftn, ihfft, ihfft2, ihfftn)
differentiable 1d to nd complex to real fast fourier transform. (hfft, hfft2, hfftn, irfft, irfft2, irfftn)
differentiable helper functions. (fftfreq, rfftfreq, fftshift, ifftshift)

paddle.signal subpackage

short-timw fast fourier transform( stft)
inverse short-time fast fourier transform( istft)

Operators

fft_c2c, fftr2c, fft_c2r, fft_c2c_grad, fftr2c_grad, fft_c2r_grad with backends on CPU(mkl_cdft and pocket_fft) and GPU(cufft for nvidia gpus and hipfft for amd gpus.) for differentiable fast fourier transforms.
frame and overlap_add for stft and istft functionalities.

Others

add kernel with complex data type for many operators.
add to the functionality for complex data type and transforms between complex and real data types.

Examples

It works like scipy.fft

import numpy as np
import paddle
x = np.random.randn(2, 4) + np.random.randn(2, 4)
xt = paddle.to_tensor(x)
print(np.allclose(np.fft.fftn(x), paddle.fft.fftn(xt).numpy()))

import scipy.fft
import paddle
x = np.random.randn(2, 4)
xt = paddle.to_tensor(x)
print(np.allclose(scipy.fft.rfftn(x), paddle.fft.rfftn(xt).numpy()))

import scipy.fft
import paddle
x = np.random.randn(2, 4) + np.random.randn(2, 4)
xt = paddle.to_tensor(x)
print(np.allclose(scipy.fft.hfftn(x), paddle.fft.hfftn(xt).numpy()))

And more than that.

2. add data type predicate; 3. fix paddle.roll.

add interface for fft

add fft c2c cufft kernel

add common code for implementing FFT, add pocketfft as a dependency

add fft c2c cufft kernel function

…ctors.

2. add complex<float>, complex<double> support for concat and flip.

Advance: done with pockfft, 1d, forward for fft, rfft, hfft cases.

2. shape_op: add support for complex data types.

fft c2c cufft kernel done with compiling and linking

add mkl placeholder

remove mkl

complete fft c2c on gpu

2. change the design, add input and output typename as template parameter for all FFTFunctors, update pocketfft-based implementation.

add mkl-based implementation, change design

…nto fft_c2c_cufft

…dtype with paddle dtype. (#40) * always convert numpy array to paddle.Tensor to avoid comparing numpy dtype with paddle dtype.

…ble (#41) 1. always convert numpy array to paddle.Tensor to avoid comparing numpy dtype with paddle dtype. 2. promote floating point tensor to complex tensor ior fft_c2c and fft_c2r; 3. fix unittest to catch UnImplementedError and RuntimeError; 4. fix compile error by avoid using thrust when cuda is not available. 5. fix sample code, use paddle.fft instead of paddle.tensor.fft

* Add doc strings. * Update overlap_add op unittest

* fix MKL-based FFT implementation, MKL CDFT's FORWARD DOMAIN is always REAL for R2C and C2R

* use std::ptrdiff_t as datatype of stride (instead of int64_t) to avoid argument mismatch on some platforms. * add complex support for fill_zeros_like * use dynload for cufft

* Add doc of frame op and overlap_add op. * Update unittest.

1. use dynload for cufft 2. fix unittest; 3. temporarily disable Rocm.

fix conflicts and merge upstream

* fix compile error: only link dyload_cuda when cuda is available

1. fix dynload for cufft on windows; 2. fix unittests.

add NOMINMAX to compile on windows

explicitly specify capture mode for lambdas

* fix fft sample

… project_fft

update scipy and numpy version for unittests of fft

1. replace numpy.fft with scipy.fft as numpy<1.20 not support ortho norm 2. remove cache of cufft plans; 3. enhance error checking. 4. default WITH_ONEMKL to OFF

… project_fft

chenwhql

LGTM for PADDLE_THROW(status)，我们的cuda关联第三方库都提供了官方的错误信息作为提示，希望后续能够补充，只报一个错误码还不太友好，可以参考 #33003

XiaoguangHu01

LGTM

Xreki

LGTM for op benchmark ci

iclementine · 2021-09-17T14:55:02Z

LGTM for PADDLE_THROW(status)，我们的cuda关联第三方库都提供了官方的错误信息作为提示，希望后续能够补充，只报一个错误码还不太友好，可以参考 #33003

目前的做法是报错误信息的，但是没有加到 proto 里面，我们后续试试把 cufft 和 mkl-cdft 的报错也都加到里面去。

* 1. add interface for fft; 2. add data type predicate; 3. fix paddle.roll. * add fft c2c cufft kernel * implement argument checking & op calling parts for fft_c2c and fftn_c2c * add operator and opmaker definitions * only register float and double for cpu. * add common code for implementing FFT, add pocketfft as a dependency * add fft c2c cufft kernel function * fix bugs in python interface * add support for c2r, r2c operators, op makers, kernels and kernel functors. * test and fix bugs * 1. fft_c2c function: add support for onesided=False; 2. add complex<float>, complex<double> support for concat and flip. * 1. fft: fix python api bugs; 2. shape_op: add support for complex data types. * fft c2c cufft kernel done with complie and link * fix shape_op, add mkl placeholder * remove mkl * complete fft c2c in gpu * 1. implement mkl-based fft, FFTC2CFunctor and common function exec_fft; 2. change the design, add input and output typename as template parameter for all FFTFunctors, update pocketfft-based implementation. * complete fft c2c on gpu in ND * complete fft c2c on gpu in ND * complete fft c2c backward in ND * fix MKL-based implementation * Add frame op and CPU/GPU kernels. * Add frame op forward unittest. * Add frame op forward unittest. * Remove axis parameter in FrameFunctor. * Add frame op grad CPU/GPU kernels and unittest. * Add frame op grad CPU/GPU kernels and unittest. * Update doc string. * Update after review and remove librosa requirement in unittest. * Update grad kernel. * add fft_c2r op * Remove data allocation in TransCompute function. * add fft r2c onesided with cpu(pocketfft/mkl) and gpu * last fft c2r functor * fix C2R and R2C for cufft, becase the direction is not an option in these cases. * add fft r2c onesided with cpu(pocketfft/mkl) and gpu * fix bugs in python APIs * fix fft_c2r grad kernal * fix bugs in python APIs * add cuda fft c2r grad kernal functor * clean code * fix fft_c2r python API * fill fft r2c result with conjugate symmetry (#19) fill fft r2c result with conjugate symmetry * add placeholder for unittests (#24) * simple parameterize test function by auto generate test case from parm list (#25) * miscellaneous fixes for python APIs (#26) * add placeholder for unittests * resize fft inputs before computation is n or s is provided. * add complex kernels for pad and pad_grad * simplify argument checking. * add type promotion * add int to float or complex promotion * fix output data type for static mode * fix fft's input dtype dispatch, import fft to paddle * fix typos in axes checking (#27) * fix typos in axes checking * fix argument checking (#28) * fix argument checking * Add C2R Python layer normal and abnormal use cases (#29) * documents and single case * test c2r case * New C2R Python layer normal and exception use cases * complete rfft,rfft2,rfftn,ihfft,ihfft2,ihfftn unittest and doc string (PaddlePaddle#30) * Documentation of the common interfaces of c2r and c2c (PaddlePaddle#31) * Documentation of the common interfaces of c2r and c2c * clean c++ code (PaddlePaddle#32) * clean code * Add numpy-based implementation of spectral ops (PaddlePaddle#33) * add numpy reference implementation of spectral ops * Add fft_c2r numpy based implementation for unittest. (PaddlePaddle#34) * add fft_c2r numpy implementation * Add deframe op and stft/istft api. (#23) * Add frame api * Add deframe op and kernels. * Add stft and istft apis. * Add deframe api. Update stft and istft apis. * Fix bug in frame_from_librosa function when input dims >= 3 * Rename deframe to overlap_add. * Update istft. * Update after code review. * Add overlap_add op and stft/istft api unittest (PaddlePaddle#35) * Add overlap_add op unittest. * Register complex kernels of squeeze/unsquuze op. * Add stft/istft api unittest. * Add unittest for fft helper functions (PaddlePaddle#36) * add unittests for fft helper functions. add complex kernel for roll op. * complete static graph unittest for all public api (PaddlePaddle#37) * Unittest of op with FFT C2C, C2R and r2c added (PaddlePaddle#38) * documents and single case * test c2r case * New C2R Python layer normal and exception use cases * Documentation of the common interfaces of c2r and c2c * Unittest of op with FFT C2C, C2R and r2c added Co-authored-by: lijiaqi <[email protected]> * add fft related options to CMakeLists.txt * fix typos and clean code (PaddlePaddle#39) * fix invisible character in mkl branch and fix error in error message * clean code: remove docstring from unittest for signal.py. * always convert numpy array to paddle.Tensor to avoid comparing numpy dtype with paddle dtype. (PaddlePaddle#40) * always convert numpy array to paddle.Tensor to avoid comparing numpy dtype with paddle dtype. * fix CI Errors: numpy dtype comparison, thrust when cuda is not available (PaddlePaddle#41) 1. always convert numpy array to paddle.Tensor to avoid comparing numpy dtype with paddle dtype. 2. promote floating point tensor to complex tensor ior fft_c2c and fft_c2r; 3. fix unittest to catch UnImplementedError and RuntimeError; 4. fix compile error by avoid using thrust when cuda is not available. 5. fix sample code, use paddle.fft instead of paddle.tensor.fft * remove inclusion of thrust, add __all__ list for fft (PaddlePaddle#42) * Add api doc and update unittest. (PaddlePaddle#43) * Add doc strings. * Update overlap_add op unittest * fix MKL-based FFT implementation (PaddlePaddle#44) * fix MKL-based FFT implementation, MKL CDFT's FORWARD DOMAIN is always REAL for R2C and C2R * remove code for debug (PaddlePaddle#45) * use dynload for cufft (PaddlePaddle#46) * use std::ptrdiff_t as datatype of stride (instead of int64_t) to avoid argument mismatch on some platforms. * add complex support for fill_zeros_like * use dynload for cufft * Update doc and unittest. (PaddlePaddle#47) * Add doc of frame op and overlap_add op. * Update unittest. * use dynload for cufft (PaddlePaddle#48) 1. use dynload for cufft 2. fix unittest; 3. temporarily disable Rocm. * fix conflicts and merge upstream (PaddlePaddle#49) fix conflicts and merge upstream * fix compile error: only link dyload_cuda when cuda is available (PaddlePaddle#50) * fix compile error: only link dyload_cuda when cuda is available * fix dynload for cufft on windows (PaddlePaddle#51) 1. fix dynload for cufft on windows; 2. fix unittests. * add NOMINMAX to compile on windows (PaddlePaddle#52) add NOMINMAX to compile on windows * explicitly specify capture mode for lambdas (PaddlePaddle#55) explicitly specify capture mode for lambdas * fix fft sample (PaddlePaddle#53) * fix fft sample * update scipy and numpy version for unittests of fft (PaddlePaddle#56) update scipy and numpy version for unittests of fft * Add static graph unittests of frame and overlap_add api. (PaddlePaddle#57) * Remove cache of cuFFT & Disable ONEMKL (PaddlePaddle#59) 1. replace numpy.fft with scipy.fft as numpy<1.20 not support ortho norm 2. remove cache of cufft plans; 3. enhance error checking. 4. default WITH_ONEMKL to OFF Co-authored-by: jeff41404 <[email protected]> Co-authored-by: root <[email protected]> Co-authored-by: KP <[email protected]> Co-authored-by: lijiaqi <[email protected]> Co-authored-by: Xiaoxu Chen <[email protected]> Co-authored-by: lijiaqi0612 <[email protected]>

chenfeiyu and others added 30 commits August 10, 2021 19:40

1. add interface for fft;

5dff7f9

2. add data type predicate; 3. fix paddle.roll.

add fft c2c cufft kernel

53c2448

implement argument checking & op calling parts for fft_c2c and fftn_c2c

3c22959

add operator and opmaker definitions

b07ee4c

only register float and double for cpu.

579eda0

Merge pull request #1 from iclementine/fft

f0d0413

add interface for fft

Merge pull request #2 from jeff41404/fft_c2c_cufft

8e51eea

add fft c2c cufft kernel

add common code for implementing FFT, add pocketfft as a dependency

6451ac8

Merge pull request #3 from iclementine/fft

4cf05fb

add common code for implementing FFT, add pocketfft as a dependency

add fft c2c cufft kernel function

0ba4495

fix bugs in python interface

b187832

Merge pull request #4 from jeff41404/fft_c2c_cufft

c3820f2

add fft c2c cufft kernel function

add support for c2r, r2c operators, op makers, kernels and kernel fun…

034cddb

…ctors.

test and fix bugs

e4bcaed

1. fft_c2c function: add support for onesided=False;

b00a0f1

2. add complex<float>, complex<double> support for concat and flip.

Merge pull request #5 from iclementine/fft

322b9e3

Advance: done with pockfft, 1d, forward for fft, rfft, hfft cases.

1. fft: fix python api bugs;

e691dca

2. shape_op: add support for complex data types.

fft c2c cufft kernel done with complie and link

9f0cc98

Merge pull request #6 from jeff41404/fft_c2c_cufft

96e8b09

fft c2c cufft kernel done with compiling and linking

Merge branch 'project_fft' of github.com:iclementine/Paddle into fft

6864616

fix shape_op, add mkl placeholder

7402387

Merge pull request #7 from iclementine/fft

5866c75

add mkl placeholder

remove mkl

5f3166d

Merge pull request #8 from iclementine/fft

bce9488

remove mkl

complete fft c2c in gpu

8349129

Merge pull request #9 from jeff41404/fft_c2c_cufft

dea0e79

complete fft c2c on gpu

1. implement mkl-based fft, FFTC2CFunctor and common function exec_fft;

d224cc4

2. change the design, add input and output typename as template parameter for all FFTFunctors, update pocketfft-based implementation.

Merge pull request #10 from iclementine/fft

7242079

add mkl-based implementation, change design

complete fft c2c on gpu in ND

c200985

Merge branch 'project_fft' of https://github.com/iclementine/Paddle i…

721e532

…nto fft_c2c_cufft

Feiyu Chan and others added 21 commits September 14, 2021 17:53

always convert numpy array to paddle.Tensor to avoid comparing numpy …

a714cfe

…dtype with paddle dtype. (#40) * always convert numpy array to paddle.Tensor to avoid comparing numpy dtype with paddle dtype.

remove inclusion of thrust, add __all__ list for fft (#42)

d2eebba

Add api doc and update unittest. (#43)

c0289d1

* Add doc strings. * Update overlap_add op unittest

fix MKL-based FFT implementation (#44)

b3d5f13

* fix MKL-based FFT implementation, MKL CDFT's FORWARD DOMAIN is always REAL for R2C and C2R

remove code for debug (#45)

2cb21c0

use dynload for cufft (#46)

5ccdf98

* use std::ptrdiff_t as datatype of stride (instead of int64_t) to avoid argument mismatch on some platforms. * add complex support for fill_zeros_like * use dynload for cufft

Update doc and unittest. (#47)

5e33e7f

* Add doc of frame op and overlap_add op. * Update unittest.

use dynload for cufft (#48)

cc9d3a0

1. use dynload for cufft 2. fix unittest; 3. temporarily disable Rocm.

fix conflicts and merge upstream (#49)

150da5c

fix conflicts and merge upstream

Merge branch 'develop' into project_fft

0279d4a

fix compile error: only link dyload_cuda when cuda is available (#50)

1e16889

* fix compile error: only link dyload_cuda when cuda is available

fix dynload for cufft on windows (#51)

e804bd5

1. fix dynload for cufft on windows; 2. fix unittests.

add NOMINMAX to compile on windows (#52)

ffcf187

add NOMINMAX to compile on windows

explicitly specify capture mode for lambdas (#55)

6c3322c

explicitly specify capture mode for lambdas

fix fft sample (#53)

d700f45

* fix fft sample

Merge branch 'develop' of https://github.com/PaddlePaddle/Paddle into…

b85788b

… project_fft

update scipy and numpy version for unittests of fft (#56)

2c615bb

update scipy and numpy version for unittests of fft

Add static graph unittests of frame and overlap_add api. (#57)

e968c20

Remove cache of cuFFT & Disable ONEMKL (#59)

76401f4

1. replace numpy.fft with scipy.fft as numpy<1.20 not support ortho norm 2. remove cache of cufft plans; 3. enhance error checking. 4. default WITH_ONEMKL to OFF

Merge branch 'develop' of https://github.com/PaddlePaddle/Paddle into…

f8c2a2e

… project_fft

chenwhql approved these changes Sep 17, 2021

View reviewed changes

XiaoguangHu01 approved these changes Sep 17, 2021

View reviewed changes

Xreki approved these changes Sep 17, 2021

View reviewed changes

raindrops2sea approved these changes Sep 18, 2021

View reviewed changes

XiaoguangHu01 merged commit 11518a4 into PaddlePaddle:develop Sep 18, 2021

cxxly mentioned this pull request Oct 13, 2021

add rocm support for fft api #36415

Merged

iclementine mentioned this pull request Oct 14, 2021

dynamic load mkl as a fft backend when it is avaialble and requested #36414

Merged

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Add FFT related operators and APIs #35665

Add FFT related operators and APIs #35665

iclementine commented Sep 10, 2021 •

edited

Loading

chenwhql left a comment

XiaoguangHu01 left a comment

Xreki left a comment

iclementine commented Sep 17, 2021

Add FFT related operators and APIs #35665

Add FFT related operators and APIs #35665

Conversation

iclementine commented Sep 10, 2021 • edited Loading

PR types

PR changes

Describe

paddle.fft subpackage

paddle.signal subpackage

Operators

Others

Examples

chenwhql left a comment

Choose a reason for hiding this comment

XiaoguangHu01 left a comment

Choose a reason for hiding this comment

Xreki left a comment

Choose a reason for hiding this comment

iclementine commented Sep 17, 2021

iclementine commented Sep 10, 2021 •

edited

Loading