opencv

Commit Graph

Author	SHA1	Message	Date
Andrew Ryrie	ea7d4be3f8	Merge pull request #20658 from smbz:lstm_optimisation * dnn: LSTM optimisation This uses the AVX-optimised fastGEMM1T for matrix multiplications where available, instead of the standard cv::gemm. fastGEMM1T is already used by the fully-connected layer. This commit involves two minor modifications: - Use unaligned access. I don't believe this involves any performance hit in on modern CPUs (Nehalem and Bulldozer onwards) in the case where the address is actually aligned. - Allow for weight matrices where the number of columns is not a multiple of 8. I have not enabled AVX-512 as I don't have an AVX-512 CPU to test on. * Fix warning about initialisation order * Remove C++11 syntax * Fix build when AVX(2) is not available In this case the CV_TRY_X macros are defined to 0, rather than being undefined. * Minor changes as requested: - Don't check hardware support for AVX(2) when dispatch is disabled for these - Add braces * Fix out-of-bounds access in fully connected layer The old tail handling in fastGEMM1T implicitly rounded vecsize up to the next multiple of 8, and the fully connected layer implements padding up to the next multiple of 8 to cope with this. The new tail handling does not round the vecsize upwards like this but it does require that the vecsize is at least 8. To adapt to the new tail handling, the fully connected layer now rounds vecsize itself at the same time as adding the padding(which makes more sense anyway). This also means that the fully connected layer always passes a vecsize of at least 8 to fastGEMM1T, which fixes the out-of-bounds access problems. * Improve tail mask handling - Use static array for generating tail masks (as requested) - Apply tail mask to the weights as well as the input vectors to prevent spurious propagation of NaNs/Infs * Revert whitespace change * Improve readability of conditions for using AVX * dnn(lstm): minor coding style changes, replaced left aligned load	3 years ago
Alexander Alekhin	1aacb9bb15	dnn(perf): update convolution tests	3 years ago
Alexander Alekhin	28aab134db	dnn(test): update tests for OpenVINO 2021.2	4 years ago
Sergei Slashchinin	61144f935e	Merge pull request #18783 from sl-sergei:fix_conv1d Add support for Conv1D on OpenCV backend * Add support for Conv1D on OpenCV backend * disable tests on other targets/backends * Fix formatting * Restore comment * Remove unnecessary flag and fix test logic * Fix perf test * fix braces * Fix indentation, assert check and remove unnecessary condition * Remove unnecessary changes * Add test cases for variable weights and bias * dnn(conv): fallback on OpenCV+CPU instead of failures * coding style	4 years ago
Alexander Alekhin	6da05f7086	dnn(test): update tests for OpenVINO 2021.1	4 years ago
Alexander Alekhin	81e027eef7	dnn: fix OpenCL implementation of Slice layer	4 years ago
Alexander Alekhin	1c371d07b5	dnn(test): adjust tests for OpenVINO 2020.4	4 years ago
Alexander Alekhin	99c4b76a6d	dnn(test): add YOLOv4-tiny tests	4 years ago
Dmitry Kurtaev	d9bada9867	dnn: EfficientDet	5 years ago
Alexander Alekhin	6b89154afd	dnn(test): add YOLOv4 tests	5 years ago
Dmitry Kurtaev	d8e10f3a8d	Enable MaxPooling with indices in Inference Engine	5 years ago
Lubov Batanina	7523c777c5	Merge pull request #15537 from l-bat:ngraph * Support nGraph * Fix resize	5 years ago
Dmitry Kurtaev	6193e403e7	Enable some tests for 2019R2	5 years ago
Dmitry Kurtaev	a0c3bb70a9	Modify SSD from TensorFlow graph generation script to enable MyriadX	5 years ago
Alexander Alekhin	416c693b3f	dnn(test): OpenVINO 2019R2	5 years ago
Lubov Batanina	8bcd7e122a	Merge pull request #14842 from l-bat:ocv_conv3d * Support Conv3D on OCV backend * Add header * Add perf tests * Support pool3d * Enable Resnet34_kinetics on OCV backend * Add test * Fix conv * Optimize Conv2D	5 years ago
Alexander Alekhin	13a782c039	test: fix usage of findDataFile() misused 'optional' mode	6 years ago
Dmitry Kurtaev	9c0af1f675	Enable more deconvolution layer configurations with IE backend	6 years ago
Dmitry Kurtaev	44d21e5a79	Enable Slice layer on Inference Engine backend	6 years ago
Alexander Alekhin	cafa010389	dnn(test): skip tests	6 years ago
Alexander Alekhin	fcb07c64f3	cmake: fix build of dnn tests with shared common code - don't share .cpp files (PCH support is broken)	6 years ago
Lubov Batanina	7d3d6bc4e2	Merge pull request #13932 from l-bat:MyriadX_master_dldt * Fix precision in tests for MyriadX * Fix ONNX tests * Add output range in ONNX tests * Skip tests on Myriad OpenVINO 2018R5 * Add detect MyriadX * Add detect MyriadX on OpenVINO R5 * Skip tests on Myriad next version of OpenVINO * dnn(ie): VPU type from environment variable * dnn(test): validate VPU type * dnn(test): update DLIE test skip conditions	6 years ago
Dmitry Kurtaev	ed710eaa1c	Make Inference Engine R3 as a minimal supported version	6 years ago
Liubov Batanina	183c0fcab1	Changed condition for resize and lrn layers	6 years ago
Dmitry Kurtaev	f0ddf302b2	Move Inference Engine to new API	6 years ago
Maksim Shabunin	fe459c82e5	Merge pull request #13332 from mshabunin:dnn-backends DNN backends registry (#13332) * Added dnn backends registry * dnn: process DLIE/FPGA target	6 years ago
Dmitry Kurtaev	0d117312c9	DNN_TARGET_FPGA using Intel's Inference Engine	6 years ago
Alexander Alekhin	96c71dd3d2	dnn: reduce set of ignored warnings	6 years ago
tompollok	0b77600718	change area() emptiness checks to empty()	6 years ago
Alexander Alekhin	c557193b8c	dnn(test): use dnnBackendsAndTargets() param generator	6 years ago
Alexander Alekhin	3e6b3a6856	dnn(perf): fix and merge Convolution tests - OpenCL tests didn't run any OpenCL kernels - use real configuration from existed models (the first 100 cases) - batch size = 1	6 years ago
Dmitry Kurtaev	8e034053af	Faster-RCNN from TensorFlow on CPU with Intel's Inference Engine backend	6 years ago
Dmitry Kurtaev	2c291bc2fb	Enable FastNeuralStyle and OpenFace networks with IE backend	7 years ago
Dmitry Kurtaev	40765c5f8d	Enable SSD models from TensorFlow with OpenCL plugin of Intel's Inference Engine	7 years ago
David	7175f257b5	Added ResizeBilinear op for tf (#11050 ) * Added ResizeBilinear op for tf Combined ResizeNearestNeighbor and ResizeBilinear layers into Resize (with an interpolation param). Minor changes to tf_importer and resize layer to save some code lines Minor changes in init.cpp Minor changes in tf_importer.cpp * Replaced implementation of a custom ResizeBilinear layer to all layers * Use Mat::ptr. Replace interpolation flags	7 years ago
Dmitry Kurtaev	f3a6ae5f00	Wrap Inference Engine init to try-catch	7 years ago
Alexander Alekhin	6816495bee	dnn(test): reuse test/test_common.hpp, eliminate dead code warning	7 years ago
Dmitry Kurtaev	b781ac7346	Make Intel's Inference Engine backend is default if no preferable backend is specified.	7 years ago
Dmitry Kurtaev	f96f934426	Update Intel's Inference Engine deep learning backend (#11587 ) * Update Intel's Inference Engine deep learning backend * Remove cpu_extension dependency * Update Darknet accuracy tests	7 years ago
Li Peng	1b517a45ae	add fp16 accuracy and perf test Signed-off-by: Li Peng <peng.li@intel.com>	7 years ago
Dmitry Kurtaev	bd77d100e1	Enable some tests for clDNN plugin from Intel's Inference Engine	7 years ago
Dmitry Kurtaev	97fec07d96	Support YOLOv3 model from Darknet	7 years ago
Dmitry Kurtaev	709cf5d038	OpenCL GPU target for Inference Engine deep learning backend Enable FP16 GPU target for DL Inference Engine backend.	7 years ago
Dmitry Kurtaev	7972f47ed4	Load networks from intermediate representation of Intel's Deep learning deployment toolkit.	7 years ago
Dmitry Kurtaev	7fe97376c2	MobileNet-SSD from TensorFlow 1.3 and Inception-V2-SSD using Inference Engine backend	7 years ago
Dmitry Kurtaev	ed94136548	OpenCV face detection network using Inference Engine backend	7 years ago
Dmitry Kurtaev	10e1de74d2	Intel Inference Engine deep learning backend (#10608 ) * Intel Inference Engine deep learning backend. * OpenFace network using Inference Engine backend	7 years ago
Alexander Alekhin	4a297a2443	ts: refactor OpenCV tests - removed tr1 usage (dropped in C++17) - moved includes of vector/map/iostream/limits into ts.hpp - require opencv_test + anonymous namespace (added compile check) - fixed norm() usage (must be from cvtest::norm for checks) and other conflict functions - added missing license headers	7 years ago
Alexander Alekhin	9b131b5f7e	dnn(test): avoid calling of cv::setNumThreads() in tests directly It is not necessary by default. Also it breaks test system command-line parameters: --perf_threads / --test_threads	7 years ago
Dmitry Kurtaev	6aabd6cc7a	Remove cv::dnn::Importer	7 years ago

1 2

62 Commits (e1ce2146f52bf0e0b9822b321aacbd6217b3bc26)