opencv

Commit Graph

Author	SHA1	Message	Date
CNClareChen	d142a796d8	Merge pull request #23929 from CNClareChen:4.x * Optimize some function with lasx. Optimize some function with lasx. #23929 This patch optimizes some lasx functions and reduces the runtime of opencv_test_core from 662,238ms to 633603ms on the 3A5000 platform. ### Pull Request Readiness Checklist See details at https://github.com/opencv/opencv/wiki/How_to_contribute#making-a-good-pull-request - [x] I agree to contribute to the project under Apache 2 License. - [x] To the best of my knowledge, the proposed patch is not based on a code under GPL or another license that is incompatible with OpenCV - [x] The PR is proposed to the proper branch - [x] There is a reference to the original bug report and related work - [x] There is accuracy test, performance test and test data in opencv_extra repository, if applicable Patch to opencv_extra has the same branch name. - [x] The feature is well documented and sample code can be built with the project CMake	1 year ago
Vadim Pisarevsky	ba4d6c859d	added detection & dispatching of some modern NEON instructions (NEON_FP16, NEON_BF16) (#24420 ) * added more or less cross-platform (based on POSIX signal() semantics) method to detect various NEON extensions, such as FP16 SIMD arithmetics, BF16 SIMD arithmetics, SIMD dotprod etc. It could be propagated to other instruction sets if necessary. * hopefully fixed compile errors * continue to fix CI * another attempt to fix build on Linux aarch64 * * reverted to the original method to detect special arm neon instructions without signal() * renamed FP16_SIMD & BF16_SIMD to NEON_FP16 and NEON_BF16, respectively * removed extra whitespaces	1 year ago
ashadrina	3889dcf3f8	Merge pull request #24286 from ashadrina:intel_icx_compiler_support Add Intel® oneAPI DPC++/C++ Compiler (icx) #24286 Intel® C++ Compiler Classic (icc) is deprecated and will be removed in a oneAPI release in the second half of 2023 ([deprecation notice](https://community.intel.com/t5/Intel-oneAPI-IoT-Toolkit/DEPRECATION-NOTICE-Intel-C-Compiler-Classic/m-p/1412267#:~:text=Intel%C2%AE%20C%2B%2B%20Compiler%20Classic%20(icc)%20is%20deprecated%20and%20will,the%20second%20half%20of%202023.)). This commit is intended to add support for the next-generation compiler, Intel® oneAPI DPC++/C++ Compiler (icx) (the documentation for the compiler is available on the [link](https://www.intel.com/content/www/us/en/docs/dpcpp-cpp-compiler/developer-guide-reference/2023-2/overview.html)). ### Pull Request Readiness Checklist See details at https://github.com/opencv/opencv/wiki/How_to_contribute#making-a-good-pull-request - [x] I agree to contribute to the project under Apache 2 License. - [x] To the best of my knowledge, the proposed patch is not based on a code under GPL or another license that is incompatible with OpenCV - [x] The PR is proposed to the proper branch - [ ] There is a reference to the original bug report and related work - [ ] There is accuracy test, performance test and test data in opencv_extra repository, if applicable Patch to opencv_extra has the same branch name. - [ ] The feature is well documented and sample code can be built with the project CMake	1 year ago
Alexander Alekhin	941d89e06d	cmake: fix RISC-V toolchains - RVV options are moved to configuration scripts instead of toolchains	2 years ago
Alexander Alekhin	d480e2e51b	cmake(opt): force separate targets for dispatched code - PCH may not pass compilation flags properly	2 years ago
wxsheng	4154bd0667	Add Loongson Advanced SIMD Extension support: -DCPU_BASELINE=LASX * Add Loongson Advanced SIMD Extension support: -DCPU_BASELINE=LASX * Add resize.lasx.cpp for Loongson SIMD acceleration * Add imgwarp.lasx.cpp for Loongson SIMD acceleration * Add LASX acceleration support for dnn/conv * Add CV_PAUSE(v) for Loongarch * Set LASX by default on Loongarch64 * LoongArch: tune test threshold for Core/HAL.mat_decomp/15 Co-authored-by: shengwenxue <shengwenxue@loongson.cn>	2 years ago
Tomoaki Teshima	b3269b08a1	neon: add dotprod dispatch implementation * read vector at runtime * add enum	2 years ago
Alexander Alekhin	5dfe65d53a	cmake: fix popcnt detection with Intel Compiler	3 years ago
Zhangyin	ff4c3873f2	Added cmake toolchain for RISC-V with clang. - Added cross compile cmake file for target riscv64-clang - Extended cmake for RISC-V and added instruction checks - Created intrin_rvv.hpp with C++ version universal intrinsics	4 years ago
Alexander Alekhin	6ea29a7696	cmake: prefer using CMAKE_SYSTEM_PROCESSOR / CMAKE_SIZEOF_VOID_P Drop: - discouraged CMAKE_CL_64 - MSVC64 - MINGW64	5 years ago
Alexander Alekhin	21c38bbdaf	cmake(cpu optmizations): fix cleanup of OPENCV_DEPENDANT_TARGETS_* vars	5 years ago
Fei Wu	90af2835a2	Fix issue 15730.	5 years ago
Alexander Alekhin	bdc097495a	fix avx512 detection - renamed Cascade Lake AVX512_CEL => AVX512_CLX (align with Intel SDE tool) - fixed CLX instruction sets (no IFMA/VBMI) - added flag to bypass CPU baseline check: OPENCV_SKIP_CPU_BASELINE_CHECK	5 years ago
mipsopen-fwu	b1ea91d8bd	Merge pull request #15422 from mipsopen-fwu:msa-dev * Added MSA implementations for mips platforms. Intrinsics for MSA and build scripts for MIPS platforms are added. Signed-off-by: Fei Wu <fwu@wavecomp.com> * Removed some unused code in mips.toolchain.cmake. Signed-off-by: Fei Wu <fwu@wavecomp.com> * Added comments for mips toolchain configuration and disabled compiling warnings for libpng. Signed-off-by: Fei Wu <fwu@wavecomp.com> * Fixed the build error of unsupported opcode 'pause' when mips isa_rev is less than 2. Signed-off-by: Fei Wu <fwu@wavecomp.com> * 1. Removed FP16 related item in MSA option defines in OpenCVCompilerOptimizations.cmake. 2. Use CV_CPU_COMPILE_MSA instead of __mips_msa for MSA feature check in cv_cpu_dispatch.h. 3. Removed hasSIMD128() in intrin_msa.hpp. 4. Define CPU_MSA as 150. Signed-off-by: Fei Wu <fwu@wavecomp.com> * 1. Removed unnecessary CV_SIMD128_64F guarding in intrin_msa.hpp. 2. Removed unnecessary CV_MSA related code block in dotProd_8u(). Signed-off-by: Fei Wu <fwu@wavecomp.com> * 1. Defined CPU_MSA_FLAGS_ON as "-mmsa". 2. Removed CV_SIMD128_64F guardings in intrin_msa.hpp. Signed-off-by: Fei Wu <fwu@wavecomp.com> * Removed unused msa_mlal_u16() and msa_mlal_s16 from msa_macros.h. Signed-off-by: Fei Wu <fwu@wavecomp.com>	5 years ago
luz.paz	57ccf14952	FIx misc. source and comment typos Found via `codespell -q 3 -S ./3rdparty,./modules -L amin,ang,atleast,dof,endwhile,hist,uint` backporting of commit: `32aba5e64b`	5 years ago
luz.paz	32aba5e64b	FIx misc. source and comment typos Found via `codespell -q 3 -S ./3rdparty,./modules -L amin,ang,atleast,dof,endwhile,hist,uint`	5 years ago
Tomoaki Teshima	db6a6ccaba	re-enable CPU_BASELINE=FP16 on Armv7 platform	6 years ago
Vitaly Tuzov	d2aadabc5e	Merge pull request #14743 from terfendail:wui512_fixvswarn Fix for MSVS2019 build warnings (#14743) * AVX512 arch support for MSVS * Fix for MSVS2019 build warnings: updated integral() AVX512 implementation * Fix for MSVS2019 build warnings: reworked v_rotate_right AVX512 implementation * fix indentation	6 years ago
Alexander Alekhin	d8b42792a6	cmake: update ENABLE_FAST_MATH option	6 years ago
Alexander Alekhin	6b6222bfbb	cmake: support CPU_DISPATCH=ALL, fix misused CPU_DISPATCH - CPU_DISPATCH_FINAL should be used for filtering	6 years ago
Sayed Adel	5a77f4cee3	Merge pull request #14007 from seiko2plus:core_avx512_infa * core: improve AVX512 infrastructure by adding more CPU features groups * cmake: use groups for AVX512 optimization flags * core: remove gap in CPU flags enumeration * cmake: restore default CPU_DISPATCH	6 years ago
Alexander Alekhin	fab0eb0d75	cmake: fix compiler flags (CPU_BASELINE_REQUIRED=xxx + CPU_BASELINE=DETECT)	6 years ago
Sayed Adel	474a0dac49	core: several improves and fixes on ppc64le infrastructure - add infrastructure support for Power9/VSX3 - fix missing VSX flags on GCC4.9 and CLANG4(#13210, #13222) - fix disable VSX optimzation on GCC by using flag ENABLE_VSX - flag ENABLE_VSX is deprecated now, use CPU_BASELINE, CPU_DISPATCH instead - add VSX3 to arithmetic dispatchable flags	6 years ago
Alexander Alekhin	c54676d625	cmake: fix supporting of legacy flags	6 years ago
Alexander Alekhin	0f07edded6	cmake: don't change baseline compiler flags in 'detection' mode	6 years ago
Alexander Alekhin	d6a8e08acc	cmake: fix variable expand in CMake conditions	6 years ago
Alexander Alekhin	3f302cabb8	core(test): intrinsic tests for all dispatched CPU optimizations - tests for both SIMD128 / SIMD256 - different dispatched + baseline(SIMD128) intrinsics	6 years ago
Maksim Shabunin	597db69151	ts: test case list is printed after cmd line parsing, refactored	6 years ago
Dmitry Kurtaev	0c4d5ffecd	Do not copy cv_cpu_helper.h to parent if OpenCV is a submodule	6 years ago
Alexander Alekhin	56222f35bb	cmake: fix CPU_BASELINE_FINAL filling - remove duplicates - restore "always on" missing entries - fix FP16 detection on MSVC	7 years ago
Alexander Alekhin	ff6ce6cd01	cmake: change CPU_BASELINE=DETECT for MacOSX	7 years ago
Alexander Alekhin	97882d03cc	core: fix FP16 conversion with CV_DISABLE_OPTIMIZATION option Reproducer: cmake -DCPU_BASELINE=AVX2 -DCV_DISABLE_OPTIMIZATION=ON ...	7 years ago
Alexander Alekhin	5b867b6f1f	cmake: fix CPU_BASELINE=NATIVE on MSVS	7 years ago
Alexander Alekhin	08941b7890	cmake: avoid amending of CMAKE_COMPILER_IS_[GNUCXX\|CLANGCXX\|CCACHE] vars - Recommended compiler checks: - GCC: CV_GCC - Clang: CV_CLANG - fixed problem with CMAKE_CXX_COMPILER_ID=Clang/AppleClang mess on MacOSX Details: cmake --help-policy CMP0025 - do not declare Clang as GCC compiler	7 years ago
Alexander Alekhin	6c051a55e5	cmake: don't add include <module>/src directory to avoid conflicts during opencv_world builds	7 years ago
Alexander Alekhin	0b4428e92f	cmake: AVX512 with clang	7 years ago
Alexander Alekhin	14032c6653	cmake: reset __content variable if file doesn't exist Resolves CMake error after relaunch with updated source code: Cannot find source file: modules/dnn/layers/layers_common.avx512_skx.cpp	7 years ago
Alexander Alekhin	5a791e6e06	cmake: update reporting of excluded dispatching files (#10711 ) * cmake: add ocv_get_smart_file_name() macro * cmake: avoid adding files for unavailable dispatch modes	7 years ago
Alexander Alekhin	73891d619a	Merge pull request #10700 from alalek:cpu_dispatch_axv512 * cmake: enable CPU dispatching for AVX512 (SKX) * cmake: update handling of unsupported flags/modes	7 years ago
Alexander Alekhin	7d67d60fb1	cmake(opt): AVX512_SKX	7 years ago
Alexander Alekhin	898ca38257	cmake: AVX512 -> AVX_512F	7 years ago
Arjan van de Ven	fc8e848a54	Add basic plumbing for AVX512 support The opencv infrastructure mostly has the basics for supporting avx512 math functions, but it wasn't hooked up (likely due to lack of users) In order to compile the DNN functions for AVX512, a few things need to be hooked up and this patch does that Signed-off-by: Arjan van de Ven <arjan@linux.intel.com>	7 years ago
Alexander Alekhin	89d855c0b7	cmake: update optimization filter	7 years ago
Maksim Shabunin	93813fec6e	VisualStudio: Added solution folders for dispatched optimization targets	7 years ago
Gregory Morse	d30a0c6f03	Merge pull request #9856 from GregoryMorse:patch-1 * Update OpenCVCompilerOptimizations.cmake Neon not supported on MSVC ARM breaking build fix * Update OpenCVCompilerOptimizations.cmake Whitespace * Update intrin.hpp Many problems in MSVC ARM builds (at least on VS2017) being fixed in this PR now. C:\Users\Gregory\DOCUME~1\MYLIBR~1\OPENCV~3\opencv\sources\modules\core\include\opencv2/core/hal/intrin.hpp(444): error C3861: '_tzcnt_u32': identifier not found * Update hal_replacement.hpp Passing variadic expansion in a macro to another macro does not work properly in MSVC and a famous known workaround is hereby applied. Discussion of it: https://stackoverflow.com/questions/5134523/msvc-doesnt-expand-va-args-correctly Only needed the fix for ARM builds: TEGRA_ macros are used for cv_hal_ functions in the carotene library. C:\Users\Gregory\Documents\My Libraries\opencv330\opencv\sources\modules\core\src\arithm.cpp(2378): warning C4003: not enough actual parameters for macro 'TEGRA_ADD' C:\Users\Gregory\Documents\My Libraries\opencv330\opencv\sources\modules\core\src\arithm.cpp(2378): error C2143: syntax error: missing ')' before ',' C:\Users\Gregory\Documents\My Libraries\opencv330\opencv\sources\modules\core\src\arithm.cpp(2378): error C2059: syntax error: ')' * Update hal_replacement.hpp All hal_replacement's using carotene\hal\tegra_hal.hpp TEGRA_ functions as macros preprocessed by variadic macros should be changed, identical as was done in core. C:\Users\Gregory\Documents\My Libraries\opencv330\opencv\sources\modules\imgproc\src\color.cpp(9604): warning C4003: not enough actual parameters for macro 'TEGRA_CVTBGRTOBGR' C:\Users\Gregory\Documents\My Libraries\opencv330\opencv\sources\modules\imgproc\src\color.cpp(9604): error C2059: syntax error: '==' * Update OpenCVCompilerOptimizations.cmake * Update hal_replacement.hpp * Update hal_replacement.hpp	7 years ago
Sayed Adel	d077778074	Added support for VSX	7 years ago
Alexander Alekhin	4f558e8b89	cmake: added "SSE4_2" into default CPU dispatch	8 years ago
Alexander Alekhin	f8a75c4361	dispatch: added CV_TRY_${OPT} macro, fix dnn build - 1: OPT is available directly or via dispatcher - 0: optimization is not compiled at all	8 years ago
Alexander Alekhin	6b7a1d4dde	build: disable AVX512 Currently it is not supported. All builds are broken with enabled AVX512 option.	8 years ago
Alexander Alekhin	3e3e2dd512	android: make optional "cpufeatures", build fixes for NDK r15	8 years ago

1 2

59 Commits (a478757483c644666ad36219db0aa81550abefc8)