* Optimize CompressQuantizeWeights transformation
- remove CoordinateTransform usage from FakeQuantize reference implementation
- move ZeroPointOptimizer functionality inside CompressQuantizeWeights
- compute scale and zero point in the same loop
Ticket: CVS-119273
* review comments
* clang format
* fix comments
* Merge opt_kernel into reference
* Remove get_default_order
* Use ov:: in jit generators
* Remove unused template function
* Add reshape parameter for consistency with Op spec
* Add brief description and such
* Remove unused param from reshape ref
* Use C++ casting
* Remove ngraph namespace from operations without namespace
* Try to fix build
* Additional fixes
* More fixes
* More fixs
* Fix reverse op
* Fixed tests
* Throw an exception if somebody tries to reallocate tensor
* Revert "Throw an exception if somebody tries to reallocate tensor"
This reverts commit 8e06d6d576.
* Remove python test
* Revert "Remove python test"
This reverts commit 37b12148d3.
* Changed evaluate model behavior
* Migrate ReduceL1, ReduceL2 to new API
- add some new utils which are deprecated
* Hide helper functions from public API
* Migrate reductions ops to new API
* Migrate get_constant_from_source to dev API
* Rename ref max to reduce_max
* Rename ref min to reduce_min
* Rename ref mean to reduce_mean
* Rename ref sum to reduce_sum
* Rename ref product to reduce_prod
- minor optimization in ReduceProd operator
* Restore custom isfinite for ov float types
* Fix type name in reduce_max.hpp
* Add missing include in shape_util.hpp
* Make count same type as data type in reduce mean
* Correct reduce sum doxy comment
* Use FindPython3.cmake
* Fixed compilation on macOS 14 with new core development tools
* Try to use Python3_SOABI instead of PYTHON_MODULE_EXTENSION
* Use Development.Module
* Keep specifying only Python3_EXECUTABLE
* Print PYTHON_MODULE_EXTENSION
* Added check for minimal cmake version for python API
* Returned Python3_INCLUDE_DIR for cross-compilation case
* Try to allow cmake older than 3.18
* Use build python interpreter to check cython dependency
* revert changes in .ci/openvino-onnx/Dockerfile
* removed unused code
* Fixed issue with variables scope
* Experiment: remove include dirs
* Corrected docs
* Use pybind11 function to set extension
* Revert "Experiment: remove include dirs"
This reverts commit 6f7f90211c.
* Refactor ConvolutionBackpropDataLayerTest, ConvolutionLayerTest, DeformableConvolutionLayerTest (#19810)
* Refactor ConvolutionBackpropDataLayerTest
* Refactor ConvolutionLayerTest
* Refactor DeformableConvolutionLayerTest
* Apply comments
* Apply comments
* Fix
* Updated minimum cmake version for Windows
* Simplified check
* Removed useless message status
* Use puiblic option
---------
Co-authored-by: Oleg Pipikin <oleg.pipikin@intel.com>
* Migrate ops evaluate
* Remove using ngraph and std from ops
* Use OPENVINO_ASSERT instead of NGRAPH_CHECK
* Move `shape_util.hpp` to `dev_api/openvino/core/`
* Remove visit_attributes, same as base impl
* Fix build issues
* Fix build issues
* Symbolic shape inference and graph optimizations
- Prepares a place in CommonOptimizations pipeline for symbolic optimizations
- Introduces symbolic propagation and symbolic optimizations for ChainedMaximum, NopBroadcast and shape sub-graph optimization
- Introduces utility runtime info for TableOfEquivalence passing and disabling of value invalidation during shape inference
* Executes NgramFusion in a symbolic environment. Relaxes Ngram fusion pattern utilizing symbolic knowledge
* Remove debug model visualization
* rt_info copying to new Add operation
* Fix visualization and place validation in nicer place in symbolic transformation
* Fix Slice operation not to propagate labels if input and output dimension is fully dynamic
* Covering Vladislav comments
* Replace value invalidation followed by validation to revalidation since it does the same thing
* Adding back invalidation of cached values to Symbolic Propagation pass
* Fix StridedSlice label propagation. Code style
* Update src/common/transformations/tests/symbolic_transformations/nop_broadcast.cpp
* Avoid Constant casting / printing when OV_VISUALIZE_TREE_CONST_MAX_ELEMENTS==0
Cast only requested amount of elements in Constant::cast_vector<>
* Refactor
* Revert style back
* Fix signed/unsigned comparison
* test
* Style
* Style
* lpt transformations: transition to api 2.0, ngraph -> openvino
* use ov namespace for lpt transformations
* fix low_precision usings
* includes refactoring
* delete RecurrentGraphRewrite and RecurrentMatcher as unused classes
* use ov header for itt; delete the disabled test
* delete the unused function
* suppress doxygen warning
* fix link in the documentation
* Migrate ReduceL1, ReduceL2 to new API
- add some new utils which are deprecated
* Add missing include
* Remove debug message
* Hide helper functions from public API
* Reduce number of rank checks
* Preserve data shape if signal_size input is not provided
* Add bounds propagation on fft input
* Improved preserving bounds on fft input
* Remove size_t rank cast and have_axes variable
* Check refactor
* Use ge helper for rank comparison
* Make bounds constexpr
* Pass raw pointer instead of unique_ptr ref
* Use normalize_axes helper
* Ensure to call set label if it's not zero
* Restored opset1::Reshape label peropagation for -1 special value
* Lets opset1::Reshape keep same shape infer. Makes FindBatch transformation keep labels in output shapes of Result node
* uses Parameter from correct namespace
* Use API 2.0 in operators evaluate
- Drop ngraph namespace in ops
- Refactor reference implementation for modified ops
* Apply code style
* Fix build issue in reference impl
* Fix code style
* Fix compile warnings
* Add inputs check and set output shape in evaluates
* Reuse common shape validation for fft base
* Align helper names
* Common test class for fft ops
* Move all (I)DFT test cases to the common test class
* More test cases for param axes
* Init labels validation
* More label tests
* Labels validation for non const signal size
* Init label tests for IRDFT
* More label test for irdft
* Labels tests for RDFT
* Remove duplicated tests
* Rename common validation file
* Rename shape infer tests file
* Use node shape infer check
* Headers order alignment
* Add const to the test params vector
* Use this make_op
* Use OV_EXPECT_THROW in common fft tests
* Use OV_EXPECT_THROW iin rdft an irdft tests
* Pass input shapes and use SHAPE_INFER_CHECK
* Shorter error messages
* Update to use ov namespace in typeprop tests
* Move files to new directories
* Use quotes for openvino includes
* Provide proxy calls for transition
of dependant components.
* Correct includes style
* Redo proxies
* Fix deprecated
* Move aliases to proxy files
* Apply code style
* WIP Postpone fp16 in CompressFloatConstantsImpl
* Apply suggestions from code review
Co-authored-by: Ilya Lavrenov <ilya.lavrenov@intel.com>
* WIP: Compression to FP16 in Serialize
* Prepared for efficient fp32 to fp16 conversion
* Update src/core/reference/src/runtime/reference/convert.cpp
* Called real slow reference implementations in the place where the optimized versions are supposed to be implemented
* Code style
* Fixed 0 values in the fast f64 to f16 compression
* Optimized convert_from_f32_to_f16_with_clamp
* Added optimized f32->f16 instance of change_constant_precision
* compression transformation Python test
* use tmp dir, minor corrections
* Update src/bindings/python/tests/test_transformations/test_compression.py
* Update src/bindings/python/tests/test_transformations/test_compression.py
* style fix
* define rt_info for postponed_fp16_compression
* remove redundant class
* fix temp dir for Win in test_compression.py
* update definitions in convert.hpp
* Update implementation in convert.cpp
* Update serialize.cpp
* Update compress_float_constants.cpp
* added macros for ARM/non_x86 in convert.cpp
* fix macros in convert.cpp
* change fixme placement in serialize.cpp
* style_fix
* Update src/core/reference/src/runtime/reference/convert.cpp
* style_fix
* Optimized count_out_of_f16_range
* Code style
* Revert unused
* Update src/core/src/pass/serialize.cpp
Co-authored-by: Ilya Lavrenov <ilya.lavrenov@intel.com>
* Update src/core/reference/src/runtime/reference/convert.cpp
Co-authored-by: Ilya Lavrenov <ilya.lavrenov@intel.com>
* use optimized convert_from_f32_to_f16_with_clamp for non postponed
* minor corrections
* Update src/common/transformations/src/transformations/common_optimizations/compress_float_constants.cpp
* Update compress_float_constants.cpp
* Switched mo and ovc to save_model instead of serialize to leverage performance improvements in fp32->fp16
* Applied minor code imporvements to address review feedback
* Minor changes in code
* Update tools/ovc/openvino/tools/ovc/main.py
* Apply suggestions from code review
* Fixed failed test in case when both usual xml compression and fp16 compression are applied simultaneously (disabled for now)
* Added description for CompressFloatConstantImpl postponed parameter
* Description of postponed parameter for CompressFloatConstants
* Reverted switching to save_model in mo as the compression can be applied not only via CLI and old code should be kept for Python path (not applicable for ovc)
* Removed remining committed test artefacts and reverted remaining changes in mo
---------
Co-authored-by: Ilya Lavrenov <ilya.lavrenov@intel.com>
Co-authored-by: dmitrygo <dmitry.gorokhov@intel.com>
Co-authored-by: Vladimir Paramuzov <vladimir.paramuzov@intel.com>
Co-authored-by: Pavel Esir <pavel.esir@intel.com>
Co-authored-by: Pavel Esir <pavel.esir@gmail.com>
* Remove ngraph headers from some core source files
* Suppress some warnings
* Suppress more warnings
* Try to fix some compilation issues
* Suppress more warnings
* Supress clone_model
* Suppress warnings for Windows
* Suppress more warnings for Windows
* Suppress more warnings for Windows
* Additional suppress
* More Windows warnings
* Additional warning
* Suppress more warnings
* Suppress warning from python API
* ResolveNamesCollisions transformation refactoring; enable it in MOC
* fix the description
* call ResolveNamesCollisions transformation in the frontends; resolve review comments
* Resolve review comments
* fix EliminateUnsqueezeGather and AlignMixedTypes transformations