Skip to content

Bump openvino from 2026.3.1 to 2026.4.0 in /openvino_notebooks/gpu-device - #661

Open
dependabot[bot] wants to merge 1 commit into
mainfrom
dependabot/uv/openvino_notebooks/gpu-device/openvino-2026.4.0
Open

dependabot[bot] wants to merge 1 commit into
mainfrom
dependabot/uv/openvino_notebooks/gpu-device/openvino-2026.4.0

Conversation

@dependabot

@dependabot dependabot Bot commented on behalf of github Oct 1, 2026

Copy link
Copy Markdown
Contributor

Bumps openvino from 2026.3.1 to 2026.4.0.

Release notes

Sourced from openvino's releases.

2026.4.0

Summary of major features and improvements  

  • More GenAI coverage and framework integrations to minimize code changes

    • New models supported:
      • On CPU: Gemma-3n
      • On CPU, GPU: Kokoro-82M, Qwen3-VL-4B with EAGLE-3, Qwen3-ASR, Muse Glimmer 30B, Qwen3.8 27B, Gemma 4 12B, Hy-MT2-1.8B, DeepSeek OCR-2, and Granite 4.0 H Micro
      • On NPUs: FLUX.2-Klein 4B and Kokoro-82M
    • Additional CPU and GPU-enabled models available as early releases: Qwen-Image, Z-Image-Turbo, Granite 4.0 H Tiny, Fun-ASR-Nano, LFM2.5-8B-A1B, MiniCPM5-2B, RF-DETR, BGE Reranker-V2-M3, and BGE M3
  • Broader LLM model support and more model compression techniques

    • OpenVINO™ GenAI adds Multi-Token Prediction (MTP) speculative decoding for Gemma 4, Qwen3.5, and Qwen3.6 on CPUs and GPUs, enabling higher throughput and lower latency without sacrificing accuracy.
    • Preview: OpenVINO™ GenAI introduces DFlash acceleration for Qwen models on GPUs, and visual-token support to reduce latency and speed up GenAI pipelines on Intel® Core™ Ultra Series 3 processors.
    • With Tree Drafting (Top-K) for EAGLE-3 now supported in OpenVINO™ GenAI, developers can unlock higher throughput on VLM pipelines compared to Chain Drafting (Top-1).
    • Xe3 integrated graphics optimizations in OpenVINO™ GenAI improve AI inference performance for Gemma 4 models processing long-context inputs on Intel® Core™ Ultra Series 3 processors.
    • OpenVINO extends Instrumentation and Tracing Technology (ITT) profiling support to the NPU, enabling developers to use Intel® VTune™ Profiler to analyze CPU, GPU, and NPU execution through one consistent toolchain.
  • More portability and performance to run AI at the edge, in the cloud, or locally.

    • OpenVINO™ GenAI introduces support for ASRPipeline in Node.js, enabling JavaScript developers to run automatic speech recognition (such as Whisper and Qwen3-ASR) with streaming and performance metrics using a pipeline API similar to C++ and Python.
    • Preview: Bounded dynamic-shape support on NPUs now extends to vision workloads such as image super-resolution. This capability has been validated with the ESPCN model.
    • Preview: OpenVINO™ Model Server now includes preview support for idle model management, which unloads models when they are not in use to reduce memory usage.
    • OpenVINO™ Model Server adds support for new agentic models such as Muse Glimmer 30B and Qwen3.8 27B.

Support Change and Deprecation Notices

  • Discontinued in 2026.0:

    • The deprecated openvino.runtime namespace has been removed. Please use the openvino namespace directly.
    • The deprecated openvino.Type.undefined has been removed. Please use openvino.Type.dynamic instead.
    • Support for Debian 10 has been discontinued due to the end of its standard support.
    • The PostponedConstant constructor signature has been updated for improved usability:
      • Old (removed): Callable[[Tensor], None]
      • New: Callable[[], Tensor]
    • The deprecated OpenVINO™ GenAI predefined generation configs were removed.
    • The deprecated OpenVINO GenAI support for whisper stateless decoder model has been removed. Please use a stateful model.
    • The deprecated OpenVINO GenAI StreamerBase put method, bool return type for callbacks, and ChunkStreamer class has been removed.
    • NNCF create_compressed_model() method is now deprecated and removed in 2026. Please use nncf.prune() method for unstructured pruning and nncf.quantize() for INT8 quantization.
    • NNCF optimization methods for TensorFlow models and TensorFlow backend in NNCF are deprecated and removed in 2026. It is recommended to use PyTorch analogous models for training-aware optimization methods and OpenVINO™ IR, PyTorch, and ONNX models for post-training optimization methods from NNCF.
    • The following experimental NNCF methods are deprecated and removed: NAS, Structural Pruning, AutoML, Knowledge Distillation, Mixed-Precision Quantization, Movement Sparsity.
    • CPU plugin now requires support for the AVX2 instruction set as a minimum system requirement. The SSE instruction set will no longer be supported.
    • Dropped support for the TensorFlow Serving (TFS) API in OpenVINO Model Server 2026.3. KServe API is recommended for classic model deployments.
    • Support for Stateful Models was removed in OpenVINO Model Server 2026.3. These capabilities were originally introduced for Kaldi audio models which are no longer applicable. Current audio model support relies on the OpenAI API and pipelines implemented with the OpenVINO™ GenAI library.
    • Support for Python 3.10 will be discontinued in OpenVINO 2026.5, due to its end-of-life (EOL) status.
  • Deprecated and to be removed in the future:

    • auto shape and auto batch size (reshaping a model in runtime) will be removed in the future. OpenVINO’s dynamic shape models are recommended instead.
    • Starting with 2026.0 release major internal refactoring of the graph iteration mechanism has been implemented for improved performance and maintainability. The legacy path can be enabled by setting the ONNX_ITERATOR=0 environment variable. This legacy path is deprecated and will be removed in future releases.
    • Modifying runtime information (RTMap) through read-only ConstOutput objects in the Python API is deprecated and will be removed in a 2027.0 release. Users should use non-const outputs when updating runtime metadata; attempts to modify RTMap through ConstOutput now emit a deprecation warning to ease migration.
    • OpenVINO Model Server:
      • The dedicated OpenVINO operator for Kubernetes and OpenShift is now deprecated in favor of the recommended KServe operator. The OpenVINO operator will remain functional in upcoming OpenVINO™ Model Server releases but will no longer be actively developed. Since KServe provides broader capabilities, no loss of functionality is expected. On the contrary, more functionalities will be accessible and migration between other serving solutions and OpenVINO Model Server will be much easier.
      • Directed Acyclic Graph Scheduler will be deprecated in favor of pipelines managed by MediaPipe scheduler and will be removed in 2026.3. That approach gives more flexibility, includes wider range of calculators and has support for using processing accelerators.

... (truncated)

Commits
  • 99c8149 [NPUW]Fixed issues with Pyramid + Block-KV. (#37783)
  • f608735 docs: sync master documentation changes into releases/2026/4 (#38071)
  • 6c38f03 Add section about npu dynamism (#36938) (#38069)
  • d9ef3b5 Fix cmake install for static libraries (#38018)
  • 227c337 [NPU] Update docs for NPU 6010 (#37903)
  • 21ed7c2 [GPU] Reject in-place crop optimization when it feeds a contiguous-input SDPA...
  • 2557b55 [AUTO]Fix low power device matching, logging, and telemetry details (#37909)
  • be7cf2e [NPU] Don't throw before releasing mutex and object (#37898)
  • b0f34e1 [GPU] reverting out-of-bounds fixes to fsv16 convolution kernels (perf drop a...
  • 3eb27a0 [NPU] Fix potential NPU DMA OOB write via forged OpenVINO state-tensor names ...
  • Additional commits viewable in compare view

Dependabot compatibility score

Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting @dependabot rebase.


Dependabot commands and options

You can trigger Dependabot actions by commenting on this PR:

  • @dependabot rebase will rebase this PR
  • @dependabot recreate will recreate this PR, overwriting any edits that have been made to it
  • @dependabot show <dependency name> ignore conditions will show all of the ignore conditions of the specified dependency
  • @dependabot ignore this major version will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself)
  • @dependabot ignore this minor version will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself)
  • @dependabot ignore this dependency will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)

Bumps [openvino](https://github.com/openvinotoolkit/openvino) from 2026.3.1 to 2026.4.0.
- [Release notes](https://github.com/openvinotoolkit/openvino/releases)
- [Changelog](https://github.com/openvinotoolkit/openvino/blob/master/docs/RELEASE.MD)
- [Commits](openvinotoolkit/openvino@2026.3.1...2026.4.0)

---
updated-dependencies:
- dependency-name: openvino
  dependency-version: 2026.4.0
  dependency-type: direct:production
  update-type: version-update:semver-minor
...

Signed-off-by: dependabot[bot] <support@github.com>
@dependabot dependabot Bot added dependencies Pull requests that update a dependency file python:uv Pull requests that update python:uv code labels Oct 1, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

dependencies Pull requests that update a dependency file python:uv Pull requests that update python:uv code

Projects

None yet

Development

Successfully merging this pull request may close these issues.

0 participants