openvino/docs/articles_en/learn-openvino/interactive-tutorials-python/notebooks-section-2-model-d...

573 lines
22 KiB
ReStructuredText
Raw Blame History

This file contains ambiguous Unicode characters

This file contains Unicode characters that might be confused with other characters. If you think that this is intentional, you can safely ignore this warning. Use the Escape button to reveal them.

.. {#notebooks_section_2_model_demos}
Model Demos
===========
.. toctree::
:maxdepth: 1
:hidden:
Demos that demonstrate inference on a particular model.
.. showcase::
:title: 284-openvoice
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/284-openvoice/284-openvoice.png
Voice tone cloning with OpenVoice and OpenVINO.
.. showcase::
:title: 283-photo-maker
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/283-photo-maker/283-photo-maker.gif
Text-to-image generation using PhotoMaker and OpenVINO.
.. showcase::
:title: 282-siglip-zero-shot-image-classification
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/228-clip-zero-shot-image-classification/228-clip-zero-shot-convert.png
Zero-shot Image Classification with SigLIP.
.. showcase::
:title: 281-kosmos2-multimodal-large-language-model
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/281-kosmos2-multimodal-large-language-model/281-kosmos2-multimodal-large-language-model.png
Kosmos-2: Multimodal Large Language Model and OpenVINO.
.. showcase::
:title: 280-depth-anything
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/280-depth-anything/280-depth-anything.gif
Depth estimation with DepthAnything and OpenVINO.
.. showcase::
:title: 279-mobilevlm-language-assistant
:img: _static/images/notebook_eye.png
Mobile language assistant with MobileVLM and OpenVINO.
.. showcase::
:title: 278-stable-diffusion-ip-adapter
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/278-stable-diffusion-ip-adapter/278-stable-diffusion-ip-adapter.png
Image Generation with Stable Diffusion and IP-Adapter.
.. showcase::
:title: 277-amused-lightweight-text-to-image
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/277-amused-lightweight-text-to-image/277-amused-lightweight-text-to-image.png
Lightweight image generation with aMUSEd and OpenVINO.
.. showcase::
:title: 276-stable-diffusion-torchdynamo-backend
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/276-stable-diffusion-torchdynamo-backend/276-stable-diffusion-torchdynamo-backend.png
Image Generation with Stable Diffusion using OpenVINO TorchDynamo backend.
.. showcase::
:title: 275-llm-question-answering
:img: _static/images/notebook_eye.png
LLM Instruction-following pipeline with OpenVINO.
.. showcase::
:title: 274-efficient-sam
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/274-efficient-sam/274-efficient-sam.png
Object segmentations with EfficientSAM and OpenVINO.
.. showcase::
:title: 273-stable-zephyr-3b-chatbot
:img: _static/images/notebook_eye.png
LLM-powered chatbot using Stable-Zephyr-3b and OpenVINO.
.. showcase::
:title: 272-paint-by-example
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/272-paint-by-example/272-paint-by-example.png
Paint by Example using Stable Diffusion and OpenVINO.
.. showcase::
:title: 271-sdxl-turbo
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/271-sdxl-turbo/271-sdxl-turbo.png
Single step image generation using SDXL-turbo and OpenVINO.
.. showcase::
:title: 270-sound-generation-audioldm2
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/270-sound-generation-audioldm2/270-sound-generation-audioldm2.png
.. showcase::
:title: 269-film-slowmo
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/269-film-slowmo/269-film-slowmo.gif
Frame interpolation using FILM and OpenVINO.
.. showcase::
:title: 268-table-question-answering
:img: _static/images/notebook_eye.png
Table Question Answering using TAPAS and OpenVINO.
.. showcase::
:title: 267-distil-whisper-asr
:img: _static/images/notebook_eye.png
Automatic speech recognition using Distil-Whisper and OpenVINO.
.. showcase::
:title: 266-speculative-sampling
:img: _static/images/notebook_eye.png
Text Generation via Speculative Sampling, KV Caching, and OpenVINO.
.. showcase::
:title: 265-wuerstchen-image-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/265-wuerstchen-image-generation/265-wuerstchen-image-generation.png
Image generation with Würstchen and OpenVINO.
.. showcase::
:title: 264-qrcode-monster
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/264-qrcode-monster/264-qrcode-monster.png
Generate creative QR codes with ControlNet QR Code Monster and OpenVINO.
.. showcase::
:title: 263-latent-consistency-models-image-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/263-latent-consistency-models-image-generation/263-latent-consistency-models-image-generation.png
Image generation with Latent Consistency Model and OpenVINO.
.. showcase::
:title: 263-lcm-lora-controlnet
:img: https://user-images.githubusercontent.com/29454499/284292122-f146e16d-7233-49f7-a401-edcb714b5288.png
Text-to-Image Generation with LCM LoRA and ControlNet Conditioning.
.. showcase::
:title: 262-softvc-voice-conversion
:img: _static/images/notebook_eye.png
SoftVC VITS Singing Voice Conversion and OpenVINO.
.. showcase::
:title: 261-fast-segment-anything
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/261-fast-segment-anything/261-fast-segment-anything.gif
Object segmentation with FastSAM and OpenVINO.
.. showcase::
:title: 259-decidiffusion-image-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/259-decidiffusion-image-generation/259-decidiffusion-image-generation.png
Image generation with DeciDiffusion and OpenVINO.
.. showcase::
:title: 258-blip-diffusion-subject-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/258-blip-diffusion-subject-generation/258-blip-diffusion-subject-generation.png
Subject-driven image generation and editing using BLIP Diffusion and OpenVINO.
.. showcase::
:title: 257-llava-multimodal-chatbot
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/257-llava-multimodal-chatbot/257-llava-multimodal-chatbot.png
Visual-language assistant with LLaVA and OpenVINO.
.. showcase::
:title: 257-videollava-multimodal-chatbot.ipynb
:img: _static/images/notebook_eye.png
Visual-language assistant with Video-LLaVA and OpenVINO.
.. showcase::
:title: 256-bark-text-to-audio
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/256-bark-text-to-audio/256-bark-text-to-audio.png
Text-to-speech generation using Bark and OpenVINO.
.. showcase::
:title: 254-rag-chatbot
:img: _static/images/notebook_eye.png
Create an LLM-powered RAG system using OpenVINO.
.. showcase::
:title: 254-llm-chatbot
:img: _static/images/notebook_eye.png
Create an LLM-powered Chatbot using OpenVINO.
.. showcase::
:title: 253-zeroscope-text2video
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/253-zeroscope-text2video/253-zeroscope-text2video.gif
Text-to video synthesis with ZeroScope and OpenVINO™.
.. showcase::
:title: 252-fastcomposer-image-generation
:img: _static/images/notebook_eye.png
Image generation with FastComposer and OpenVINO™.
.. showcase::
:title: 251-tiny-sd-image-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/251-tiny-sd-image-generation/251-tiny-sd-image-generation.png
Image Generation with Tiny-SD and OpenVINO™.
.. showcase::
:title: 250-music-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/250-music-generation/250-music-generation.png
Controllable Music Generation with MusicGen and OpenVINO™.
.. showcase::
:title: 249-oneformer-segmentation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/249-oneformer-segmentation/249-oneformer-segmentation.png
Universal segmentation with OneFormer and OpenVINO™.
.. showcase::
:title: 248-segmind-vegart
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/248-stable-diffusion-xl/248-stable-diffusion-xl.png
High-resolution image generation with Segmind-VegaRT and OpenVINO.
.. showcase::
:title: 248-ssd-b1
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/248-stable-diffusion-xl/248-stable-diffusion-xl.png
Image generation with Stable Diffusion XL and OpenVINO™.
.. showcase::
:title: 248-stable-diffusion-xl
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/248-stable-diffusion-xl/248-stable-diffusion-xl.png
Image generation with Stable Diffusion XL and OpenVINO™.
.. showcase::
:title: 247-code-language-id
:img: _static/images/notebook_eye.png
Identify the programming language used in an arbitrary code snippet.
.. showcase::
:title: 246-depth-estimation-videpth
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/246-depth-estimation-videpth/246-depth-estimation-videpth.png
Monocular Visual-Inertial Depth Estimation with OpenVINO™.
.. showcase::
:title: 245-typo-detector
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/245-typo-detector/245-typo-detector.png
English Typo Detection in sentences with OpenVINO™.
.. showcase::
:title: 244-named-entity-recognition
:img: _static/images/notebook_eye.png
Named entity recognition with OpenVINO™.
.. showcase::
:title: 243-tflite-selfie-segmentation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/243-tflite-selfie-segmentation/243-tflite-selfie-segmentation.gif
Selfie Segmentation using TFLite and OpenVINO™.
.. showcase::
:title: 242-freevc-voice-conversion
:img: _static/images/notebook_eye.png
High-Quality Text-Free One-Shot Voice Conversion with FreeVC and OpenVINO™
.. showcase::
:title: 241-riffusion-text-to-music
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/241-riffusion-text-to-music/241-riffusion-text-to-music.png
Text-to-Music generation using Riffusion and OpenVINO™.
.. showcase::
:title: 240-dolly-2-instruction-following
:img: _static/images/notebook_eye.png
Instruction following using Databricks Dolly 2.0 and OpenVINO™.
.. showcase::
:title: 239-image-bind-convert
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/239-image-bind/239-image-bind-convert.png
Binding multimodal data, using ImageBind and OpenVINO™.
.. showcase::
:title: 238-deep-floyd-if-optimize
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/238-deepfloyd-if/238-deep-floyd-if-optimize.png
Text-to-image generation with DeepFloyd IF and OpenVINO™.
.. showcase::
:title: 237-segment-anything
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/237-segment-anything/237-segment-anything.png
Prompt based object segmentation mask generation, using Segment Anything and OpenVINO™.
.. showcase::
:title: 236-stable-diffusion-v2-text-to-image
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/236-stable-diffusion-v2/236-stable-diffusion-v2-optimum-demo.png
Text-to-image generation with Stable Diffusion v2 and OpenVINO™.
.. showcase::
:title: 236-stable-diffusion-v2-text-to-image-demo
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/236-stable-diffusion-v2/236-stable-diffusion-v2-optimum-demo.png
Stable Diffusion Text-to-Image Demo.
.. showcase::
:title: 236-stable-diffusion-v2-optimum-demo
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/236-stable-diffusion-v2/236-stable-diffusion-v2-optimum-demo.png
Stable Diffusion v2.1 using Optimum-Intel OpenVINO.
.. showcase::
:title: 236-stable-diffusion-v2-optimum-demo-comparison
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/236-stable-diffusion-v2/236-stable-diffusion-v2-optimum-demo.png
Stable Diffusion v2.1 using Optimum-Intel OpenVINO and multiple Intel Hardware
.. showcase::
:title: 236-stable-diffusion-v2-infinite-zoom
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/236-stable-diffusion-v2/236-stable-diffusion-v2-infinite-zoom.gif
Text-to-image generation and Infinite Zoom with Stable Diffusion v2 and OpenVINO™.
.. showcase::
:title: 235-controlnet-stable-diffusion
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/235-controlnet-stable-diffusion/235-controlnet-stable-diffusion.png
A text-to-image generation with ControlNet Conditioning and OpenVINO™.
.. showcase::
:title: 234-encodec-audio-compression
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/234-encodec-audio-compression/234-encodec-audio-compression.png
Audio compression with EnCodec and OpenVINO™.
.. showcase::
:title: 233-blip-convert
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/233-blip-visual-language-processing/233-blip-convert.png
Visual Question Answering and Image Captioning using BLIP and OpenVINO.
.. showcase::
:title: 233-blip-optimize
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/233-blip-visual-language-processing/233-blip-convert.png
Post-Training Quantization and Weights Compression of OpenAI BLIP model with NNCF.
.. showcase::
:title: 232-clip-language-saliency-map
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/232-clip-language-saliency-map/232-clip-language-saliency-map.png
Language-visual saliency with CLIP and OpenVINO™.
.. showcase::
:title: 231-instruct-pix2pix-image-editing
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/231-instruct-pix2pix-image-editing/231-instruct-pix2pix-image-editing.png
Image editing with InstructPix2Pix.
.. showcase::
:title: 230-yolov8-optimization
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/230-yolov8-optimization/230-yolov8-object-detection.png
Optimize YOLOv8, using NNCF PTQ API.
.. showcase::
:title: 229-distilbert-sequence-classification
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/229-distilbert-sequence-classification/229-distilbert-sequence-classification.png
Sequence classification with OpenVINO.
.. showcase::
:title: 228-clip-zero-shot-quantize
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/228-clip-zero-shot-image-classification/228-clip-zero-shot-quantize.png
Post-Training Quantization of OpenAI CLIP model with NNCF.
.. showcase::
:title: 228-clip-zero-shot-convert
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/228-clip-zero-shot-image-classification/228-clip-zero-shot-convert.png
Zero-shot Image Classification with OpenAI CLIP and OpenVINO™.
.. showcase::
:title: 227-whisper-subtitles-generation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/227-whisper-subtitles-generation/227-whisper-convert.png
Generate subtitles for video with OpenAI Whisper and OpenVINO.
.. showcase::
:title: 226-yolov7-optimization
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/226-yolov7-optimization/226-yolov7-optimization.png
Optimize YOLOv7, using NNCF PTQ API.
.. showcase::
:title: 225-stable-diffusion-text-to-image
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/225-stable-diffusion-text-to-image/225-stable-diffusion-text-to-image.png
Text-to-image generation with Stable Diffusion method.
.. showcase::
:title: 224-3D-segmentation-point-clouds
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/224-3D-segmentation-point-clouds/224-3D-segmentation-point-clouds.png
Process point cloud data and run 3D Part Segmentation with OpenVINO.
.. showcase::
:title: 223-text-prediction
:img: _static/images/notebook_eye.png
Use pre-trained models to perform text prediction on an input sequence.
.. showcase::
:title: 222-vision-image-colorization
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/222-vision-image-colorization/222-vision-image-colorization.png
Use pre-trained models to colorize black & white images using OpenVINO.
.. showcase::
:title: 221-machine-translation
:img: _static/images/notebook_eye.png
Real-time translation from English to German.
.. showcase::
:title: 220-cross-lingual-books-alignment
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/220-cross-lingual-books-alignment/220-cross-lingual-books-alignment.png
Cross-lingual Books Alignment With Transformers and OpenVINO™
.. showcase::
:title: 219-knowledge-graphs-conve
:img: _static/images/notebook_eye.png
Optimize the knowledge graph embeddings model (ConvE) with OpenVINO.
.. showcase::
:title: 218-vehicle-detection-and-recognition
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/218-vehicle-detection-and-recognition/218-vehicle-detection-and-recognition.png
Use pre-trained models to detect and recognize vehicles and their attributes with OpenVINO.
.. showcase::
:title: 216-attention-center
:img: _static/images/notebook_eye.png
The attention center model with OpenVINO™
.. showcase::
:title: 215-image-inpainting
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/215-image-inpainting/215-image-inpainting.gif
Fill missing pixels with image in-painting.
.. showcase::
:title: 214-grammar-correction
:img: _static/images/notebook_eye.png
Grammatical error correction with OpenVINO.
.. showcase::
:title: 213-question-answering
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/213-question-answering/213-question-answering.png
Answer your questions basing on a context.
.. showcase::
:title: 212-pyannote-speaker-diarization
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/212-pyannote-speaker-diarization/212-pyannote-speaker-diarization.png
Run inference on speaker diarization pipeline.
.. showcase::
:title: 211-speech-to-text
:img: _static/images/notebook_eye.png
Run inference on speech-to-text recognition model.
.. showcase::
:title: 210-slowfast-video-recognition
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/210-slowfast-video-recognition/210-slowfast-video-recognition.gif
Video Recognition using SlowFast and OpenVINO™
.. showcase::
:title: 209-handwritten-ocrn
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/209-handwritten-ocr/209-handwritten-ocr.png
OCR for handwritten simplified Chinese and Japanese.
.. showcase::
:title: 208-optical-character-recognition
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/208-optical-character-recognition/208-optical-character-recognition.png
Annotate text on images using text recognition resnet.
.. showcase::
:title: 207-vision-paddlegan-superresolution
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/207-vision-paddlegan-superresolution/207-vision-paddlegan-superresolution.png
Upscale small images with superresolution using a PaddleGAN model.
.. showcase::
:title: 206-vision-paddlegan-anime
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/206-vision-paddlegan-anime/206-vision-paddlegan-anime.png
Turn an image into anime using a GAN.
.. showcase::
:title: 205-vision-background-removal
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/205-vision-background-removal/205-vision-background-removal.png
Background Removal Demo.
.. showcase::
:title: 204-segmenter-semantic-segmentation
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/204-segmenter-semantic-segmentation/204-segmenter-semantic-segmentation.png
Semantic segmentation with OpenVINO™ using Segmenter.
.. showcase::
:title: 203-meter-reader
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/203-meter-reader/203-meter-reader.png
PaddlePaddle pre-trained models to read industrial meters value.
.. showcase::
:title: 202-vision-superresolution-video
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/202-vision-superresolution/202-vision-superresolution-video.gif
Turn 360p into 1080p video using a super resolution model.
.. showcase::
:title: 202-vision-superresolution-image
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/202-vision-superresolution/202-vision-superresolution-image.png
Upscale raw images with a super resolution model.
.. showcase::
:title: 201-vision-monodepth
:img: https://raw.githubusercontent.com/openvinotoolkit/openvino_notebooks/main/notebooks/201-vision-monodepth/201-vision-monodepth.gif
Monocular depth estimation with images and video.