{"id":773,"date":"2025-07-31T05:50:37","date_gmt":"2025-07-31T05:50:37","guid":{"rendered":"https:\/\/blog.aetherix.com\/?p=773"},"modified":"2026-03-27T13:18:30","modified_gmt":"2026-03-27T13:18:30","slug":"jetson-generative-ai-live-llava","status":"publish","type":"post","link":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/","title":{"rendered":"Jetson Generative AI \u2013 Live LLaVA"},"content":{"rendered":"<div class=\"fusion-fullwidth fullwidth-box fusion-builder-row-1 fusion-flex-container has-pattern-background has-mask-background nonhundred-percent-fullwidth non-hundred-percent-height-scrolling\" style=\"--awb-border-radius-top-left:0px;--awb-border-radius-top-right:0px;--awb-border-radius-bottom-right:0px;--awb-border-radius-bottom-left:0px;--awb-flex-wrap:wrap;\" ><div class=\"fusion-builder-row fusion-row fusion-flex-align-items-flex-start fusion-flex-content-wrap\" style=\"max-width:1331.2px;margin-left: calc(-4% \/ 2 );margin-right: calc(-4% \/ 2 );\"><div class=\"fusion-layout-column fusion_builder_column fusion-builder-column-0 fusion_builder_column_1_1 1_1 fusion-flex-column\" style=\"--awb-bg-size:cover;--awb-width-large:100%;--awb-margin-top-large:0px;--awb-spacing-right-large:1.92%;--awb-margin-bottom-large:20px;--awb-spacing-left-large:1.92%;--awb-width-medium:100%;--awb-order-medium:0;--awb-spacing-right-medium:1.92%;--awb-spacing-left-medium:1.92%;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:1.92%;--awb-spacing-left-small:1.92%;\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-text fusion-text-1\"><div>\n<div>Vision-Language Models reach new heights when applied to <strong>live video streams<\/strong>\u2014<strong>Live LLaVA<\/strong> demonstrates real-time multimodal AI that can see, understand, and describe what&#8217;s happening in your camera feed <strong>continuously<\/strong>\u00a0on your Jetson device.<\/div>\n<div><\/div>\n<div>In this article you&#8217;ll learn how to run Live LLaVA with optimized vision-language models like LLaVA and VILA, featuring hardware-accelerated video processing and real-time inference capabilities.<\/div>\n<\/div>\n<p>&nbsp;<\/p>\n<\/div><div class=\"fusion-title title fusion-title-1 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-margin-bottom:-10px;\"><h3 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\">Requirements<\/h3><\/div>\n<div class=\"table-1\">\n<p>&nbsp;<\/p>\n<table width=\"100%\">\n<thead>\n<tr>\n<th align=\"left\">\n<div>\n<div>Hardware \/ Software<\/div>\n<\/div>\n<\/th>\n<th align=\"left\">\n<div>\n<div>Notes<\/div>\n<\/div>\n<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Jetson AGX Orin (64GB)<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Recommended for best performance<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Jetson AGX Orin (32GB)<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Good performance for most use cases<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Jetson Orin NX (16GB)<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Solid performance<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Jetson Orin Nano (8GB)<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Minimum requirement &#8211; use smaller models<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>JetPack 6 (L4T r36.x)<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Required for latest optimizations<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>USB camera or CSI camera<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>For live video input<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>NVMe SSD highly recommended<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>For storage speed and space<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>22GB for nano_llm container<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Container image storage<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>&gt;10GB for models<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Vision-language model storage<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<div class=\"fusion-text fusion-text-2\" style=\"--awb-margin-top:30px;\"><div>\n<div><em><strong>Note:<\/strong>\u00a0Follow the NanoVLM tutorial first to familiarize yourself with vision\/language models, and see Agent Studio for an interactive pipeline editor.<\/em><\/div>\n<\/div>\n<\/div><div class=\"fusion-title title fusion-title-2 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><div>\n<h4>Supported Models<\/h4>\n<\/div><\/h1><\/div><div class=\"fusion-text fusion-text-3\"><p>The following vision-language models are optimized for Live LLaVA:<\/p>\n<div>\n<div>\n<div><strong>LLaVA Models:<\/strong><\/div>\n<blockquote>\n<div>`liuhaotian\/llava-v1.5-7b`,<\/div>\n<div>`liuhaotian\/llava-v1.5-13b`<\/div>\n<div>`liuhaotian\/llava-v1.6-vicuna-7b`<\/div>\n<div>`liuhaotian\/llava-v1.6-vicuna-13b`<\/div>\n<\/blockquote>\n<div><strong>VILA Models:<\/strong><\/div>\n<blockquote>\n<div>`Efficient-Large-Model\/VILA-2.7b`<\/div>\n<div>`Efficient-Large-Model\/VILA-7b`<\/div>\n<div>`Efficient-Large-Model\/VILA-13b`<\/div>\n<div>`Efficient-Large-Model\/VILA1.5-3b`<\/div>\n<div>`Efficient-Large-Model\/Llama-3-VILA1.5-8B`<\/div>\n<div>`Efficient-Large-Model\/VILA1.5-13b`<\/div>\n<\/blockquote>\n<div><strong>Jetson Orin Nano Compatible Models:<\/strong><\/div>\n<blockquote>\n<div>VILA-2.7b<\/div>\n<div>VILA1.5-3b<\/div>\n<div>VILA-7b<\/div>\n<div>Llava-7b<\/div>\n<div>Obsidian-3B<\/div>\n<\/blockquote>\n<\/div>\n<\/div>\n<\/div><div class=\"fusion-title title fusion-title-3 fusion-sep-none fusion-title-text fusion-title-size-three\"><h3 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><div>\n<h3>Step-by-Step Setup<\/h3>\n<\/div><\/h3><\/div><div class=\"fusion-title title fusion-title-4 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><div>\n<h4>1. Verify Camera Connection<\/h4>\n<\/div><\/h1><\/div><div class=\"fusion-text fusion-text-4 fusion-text-no-margin\" style=\"--awb-margin-bottom:10px;\"><p>Check that your camera is properly connected and detected:<\/p>\n<\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-1 > .CodeMirror, .fusion-syntax-highlighter-1 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-1 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_1\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_1\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_1\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\"># List available video devices\nls \/dev\/video*\n\n# Test camera with GStreamer (optional)\ngst-launch-1.0 v4l2src device=\/dev\/video0 ! autovideosink<\/textarea><\/div><div class=\"fusion-title title fusion-title-5 fusion-sep-none fusion-title-text fusion-title-size-four\"><h4 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>2. Clone and setup jetson-containers<\/h4><\/h4><\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-2 > .CodeMirror, .fusion-syntax-highlighter-2 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-2 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_2\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_2\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_2\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\">git clone https:\/\/github.com\/dusty-nv\/jetson-containers\nbash jetson-containers\/install.sh<\/textarea><\/div><div class=\"fusion-title title fusion-title-6 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><div>\n<p>&nbsp;<\/p>\n<h4>3. Launch Live LLaVA<\/h4>\n<\/div><\/h1><\/div><div class=\"fusion-text fusion-text-5 fusion-text-no-margin\" style=\"--awb-margin-bottom:10px;\"><p>Start the VideoQuery agent with your camera:<\/p>\n<\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-3 > .CodeMirror, .fusion-syntax-highlighter-3 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-3 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_3\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_3\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_3\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\">jetson-containers run $(autotag nano_llm) \\\n  python3 -m nano_llm.agents.video_query --api=mlc \\\n    --model Efficient-Large-Model\/VILA1.5-3b \\\n    --max-context-len 256 \\\n    --max-new-tokens 32 \\\n    --video-input \/dev\/video0 \\\n    --video-output webrtc:\/\/@:8554\/output<\/textarea><\/div><div class=\"fusion-title title fusion-title-7 fusion-sep-none fusion-title-text fusion-title-size-four\"><h4 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>4. Access the Web Interface<\/h4><\/h4><\/div><div class=\"fusion-text fusion-text-6\"><div>\n<div>Navigate your browser to:<\/div>\n<div>\n<div>\n<blockquote>\n<div>https:\/\/&lt;jetson-ip&gt;:8050<\/div>\n<\/blockquote>\n<div>\n<div>\n<div><strong>\u26a0\ufe0f Chrome Recommended:<\/strong>\u00a0For best WebRTC performance, use Chrome with `chrome:\/\/flags#enable-webrtc-hide-local-ips-with-mdns` disabled.<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div><div class=\"fusion-title title fusion-title-8 fusion-sep-none fusion-title-text fusion-title-size-four\"><h4 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>5. Configure Prompts<\/h4><\/h4><\/div><div class=\"fusion-text fusion-text-7 fusion-text-no-margin\" style=\"--awb-margin-bottom:10px;\"><p>In the web interface, you can:<\/p>\n<div>&#8211; <strong>Set custom prompts<\/strong>\u00a0for continuous analysis<\/div>\n<div>&#8211; <strong>Adjust inference frequency<\/strong>\u00a0for real-time performance<\/div>\n<div>&#8211; <strong>Monitor live video feed<\/strong>\u00a0with AI descriptions<\/div>\n<\/div><div class=\"fusion-image-element awb-imageframe-style awb-imageframe-style-below awb-imageframe-style-1\" style=\"--awb-margin-top:10px;--awb-margin-bottom:20px;--awb-caption-title-font-family:var(--h2_typography-font-family);--awb-caption-title-font-weight:var(--h2_typography-font-weight);--awb-caption-title-font-style:var(--h2_typography-font-style);--awb-caption-title-size:var(--h2_typography-font-size);--awb-caption-title-transform:var(--h2_typography-text-transform);--awb-caption-title-line-height:var(--h2_typography-line-height);--awb-caption-title-letter-spacing:var(--h2_typography-letter-spacing);\"><span class=\" fusion-imageframe imageframe-none imageframe-1 hover-type-none\"><img decoding=\"async\" width=\"1024\" height=\"632\" src=\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_face_detection-1024x632.png\" alt class=\"img-responsive wp-image-781\" srcset=\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_face_detection-200x123.png 200w, https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_face_detection-400x247.png 400w, https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_face_detection-600x370.png 600w, https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_face_detection-800x494.png 800w, https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_face_detection-1200x741.png 1200w\" sizes=\"(max-width: 640px) 100vw, 1024px\" \/><\/span><div class=\"awb-imageframe-caption-container\" style=\"text-align:center;\"><div class=\"awb-imageframe-caption\"><p class=\"awb-imageframe-caption-text\">Live LLaVA Face Detection<\/p><\/div><\/div><\/div><div class=\"fusion-title title fusion-title-9 fusion-sep-none fusion-title-text fusion-title-size-one\" style=\"--awb-margin-top:-30px;\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>Real-time Object Detection<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-8\"><div>\n<div>Live LLaVA can continuously analyze your video feed, detecting and describing objects, people, and activities in real-time:<\/div>\n<\/div>\n<\/div><div class=\"fusion-image-element awb-imageframe-style awb-imageframe-style-below awb-imageframe-style-2\" style=\"text-align:center;--awb-aspect-ratio:16 \/ 9;--awb-caption-title-font-family:var(--h2_typography-font-family);--awb-caption-title-font-weight:var(--h2_typography-font-weight);--awb-caption-title-font-style:var(--h2_typography-font-style);--awb-caption-title-size:var(--h2_typography-font-size);--awb-caption-title-transform:var(--h2_typography-text-transform);--awb-caption-title-line-height:var(--h2_typography-line-height);--awb-caption-title-letter-spacing:var(--h2_typography-letter-spacing);\"><span class=\" fusion-imageframe imageframe-none imageframe-2 hover-type-none has-aspect-ratio\"><img decoding=\"async\" width=\"500\" height=\"282\" src=\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/live_llava_object_detection.gif\" class=\"img-responsive wp-image-782 img-with-aspect-ratio\" data-parent-fit=\"cover\" data-parent-container=\".fusion-image-element\" alt \/><\/span><div class=\"awb-imageframe-caption-container\" style=\"text-align:center;\"><div class=\"awb-imageframe-caption\"><p class=\"awb-imageframe-caption-text\">Live LLaVA Object Detection<\/p><\/div><\/div><\/div><div class=\"fusion-title title fusion-title-10 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><div>\n<h4>Custom Prompting<\/h4>\n<\/div><\/h1><\/div><div class=\"fusion-text fusion-text-9 fusion-text-no-margin\" style=\"--awb-margin-bottom:10px;\"><p>You can customize the analysis with specific prompts:<\/p>\n<\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-4 > .CodeMirror, .fusion-syntax-highlighter-4 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-4 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_4\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_4\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_4\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\"># Example prompts\n\"Describe what you see in detail\"\n\"What objects are on the desk?\"\n\"Count the number of people in the scene\"\n\"What is the person doing?\"\n\"Describe the lighting and environment\"<\/textarea><\/div><div class=\"fusion-title title fusion-title-11 fusion-sep-none fusion-title-text fusion-title-size-three\"><h3 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>Pre-recorded Video Analysis<\/h4><\/h3><\/div><div class=\"fusion-text fusion-text-10\"><div>\n<div>Process existing video files instead of live camera feeds:<\/div>\n<\/div>\n<\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-5 > .CodeMirror, .fusion-syntax-highlighter-5 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-5 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_5\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_5\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_5\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\">jetson-containers run \\\n  -v \/path\/to\/your\/videos:\/mount \\\n  $(autotag nano_llm) \\\n    python3 -m nano_llm.agents.video_query --api=mlc \\\n      --model Efficient-Large-Model\/VILA1.5-3b \\\n      --max-context-len 256 \\\n      --max-new-tokens 32 \\\n      --video-input \/mount\/my_video.mp4 \\\n      --video-output \/mount\/output.mp4 \\\n      --prompt \"What does the weather look like?\"<\/textarea><\/div><div class=\"fusion-title title fusion-title-12 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>Supported Formats<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-11\"><div>\n<div><strong>Input Formats:<\/strong><\/div>\n<blockquote>\n<div>MP4, MKV, AVI, FLV (with H.264\/H.265 encoding)<\/div>\n<div>Live network streams (RTP, RTSP, WebRTC)<\/div>\n<div>USB\/CSI cameras<\/div>\n<\/blockquote>\n<div><strong>Output Formats:<\/strong><\/div>\n<blockquote>\n<div>Video files (MP4, AVI, etc.)<\/div>\n<div>Network streams (WebRTC, RTSP)<\/div>\n<div>Display output<\/div>\n<\/blockquote>\n<\/div>\n<\/div><div class=\"fusion-title title fusion-title-13 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>NanoDB Integration<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-12 fusion-text-no-margin\" style=\"--awb-margin-bottom:10px;\"><p>Enable reverse-image search and database tagging by integrating with NanoDB:<\/p>\n<\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-6 > .CodeMirror, .fusion-syntax-highlighter-6 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-6 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_6\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_6\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_6\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\">jetson-containers run $(autotag nano_llm) \\\n  python3 -m nano_llm.agents.video_query --api=mlc \\\n    --model Efficient-Large-Model\/VILA1.5-3b \\\n    --max-context-len 256 \\\n    --max-new-tokens 32 \\\n    --video-input \/dev\/video0 \\\n    --video-output webrtc:\/\/@:8554\/output \\\n    --nanodb \/data\/nanodb\/coco\/2017<\/textarea><\/div><div class=\"fusion-text fusion-text-13\" style=\"--awb-margin-top:10px;\"><p>This enables:<\/p>\n<div>&#8211; <strong>Reverse-image search<\/strong>\u00a0against your database<\/div>\n<div>&#8211; <strong>One-shot recognition<\/strong>\u00a0tasks via web UI<\/div>\n<div>&#8211; <strong>Automatic tagging<\/strong>\u00a0of incoming images<\/div>\n<\/div><div class=\"fusion-title title fusion-title-14 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><div>\n<h4>Video VILA &#8211; Multi-frame Analysis<\/h4>\n<\/div><\/h1><\/div><div class=\"fusion-text fusion-text-14 fusion-text-no-margin\" style=\"--awb-margin-bottom:10px;\"><div>\n<div>VILA-1.5 models can analyze multiple frames simultaneously for temporal understanding:<\/div>\n<\/div>\n<\/div><style type=\"text\/css\" scopped=\"scopped\">.fusion-syntax-highlighter-7 > .CodeMirror, .fusion-syntax-highlighter-7 > .CodeMirror .CodeMirror-gutters {background-color:#2d3748;}<\/style><div class=\"fusion-syntax-highlighter-container fusion-syntax-highlighter-7 fusion-syntax-highlighter-theme-dark\" style=\"opacity:0;margin-top:0px;margin-right:0px;margin-bottom:0px;margin-left:0px;font-size:14px;border-width:1px;border-style:solid;border-color:rgba(242,243,245,0);\"><div class=\"syntax-highlighter-copy-code\"><span class=\"syntax-highlighter-copy-code-title\" data-id=\"fusion_syntax_highlighter_7\" style=\"font-size:14px;\">Copy to Clipboard<\/span><\/div><label for=\"fusion_syntax_highlighter_7\" class=\"screen-reader-text\">Syntax Highlighter<\/label><textarea class=\"fusion-syntax-highlighter-textarea\" id=\"fusion_syntax_highlighter_7\" data-readOnly=\"nocursor\" data-lineNumbers=\"\" data-lineWrapping=\"\" data-theme=\"oceanic-next\" data-mode=\"text\/x-sh\">jetson-containers run $(autotag nano_llm) \\\n  python3 -m nano_llm.vision.video \\\n    --model Efficient-Large-Model\/VILA1.5-3b \\\n    --max-images 8 \\\n    --max-new-tokens 48 \\\n    --video-input \/data\/my_video.mp4 \\\n    --video-output \/data\/my_output.mp4 \\\n    --prompt 'What changes occurred in the video?'<\/textarea><\/div><div class=\"fusion-title title fusion-title-15 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-margin-bottom:-20px;\"><h3 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h3>Troubleshooting<\/h3><\/h3><\/div><div class=\"fusion-title title fusion-title-16 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>How to fix freezing issues while loading the model?<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-15\"><p>The documentation uses the old <code data-start=\"154\" data-end=\"160\">awq4<\/code>; instead, use the <code data-start=\"179\" data-end=\"203\"><span style=\"color: #38c92e;\">--quantization q4f16_1<\/span><\/code> parameter.<br data-start=\"214\" data-end=\"217\" data-is-only-node=\"\" \/>The 13B model eventually freezes on the Jetson AGX Orin 32GB due to running out of tokens; if speed is needed, we recommend using <b>VILA-7B<\/b> or <b>VILA-2.7B <\/b>instead.<\/p>\n<\/div><div class=\"fusion-title title fusion-title-17 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>How to fix the issue of the camera not being detected<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-16\"><p>To make a USB camera accessible inside the container, add the parameter <code data-start=\"210\" data-end=\"232\"><span style=\"color: #38c92e;\">--device \/dev\/video0<\/span><\/code> when running the container. This maps the host&#8217;s camera device into the container, allowing applications inside to access the video stream as if it were running natively on the host system.<\/p>\n<\/div><div class=\"fusion-title title fusion-title-18 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4 data-start=\"105\" data-end=\"183\">How to Avoid Color Distortion Issues with Logitech C505e Using MJPEG Codec ?<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-17\"><p>To prevent color distortion problems on the Logitech C505e camera, we recommend using the <code data-start=\"275\" data-end=\"302\"><span style=\"color: #38c92e;\">--video-input-codec mjpeg<\/span><\/code> parameter. This forces the camera to use the MJPEG codec, which is better supported and helps maintain accurate color reproduction.<\/p>\n<\/div><div class=\"fusion-title title fusion-title-19 fusion-sep-none fusion-title-text fusion-title-size-one\"><h1 class=\"fusion-title-heading title-heading-left\" style=\"margin:0;\"><h4>Resolution Limitation<\/h4><\/h1><\/div><div class=\"fusion-text fusion-text-18 fusion-text-no-margin\" style=\"--awb-margin-bottom:-10px;\"><p>For stable FPS performance, use the parameters <code data-start=\"139\" data-end=\"165\"><span style=\"color: #38c92e;\">--video-input-width 1280<\/span><\/code><span style=\"color: #38c92e;\"> and <\/span><code data-start=\"170\" data-end=\"196\"><span style=\"color: #38c92e;\">--video-input-height 720<\/span><\/code><span style=\"color: #38c92e;\">.<\/span> These settings limit the video resolution to 1280&#215;720, helping maintain smoother and more consistent frame rates.<\/p>\n<\/div>\n<div class=\"table-1\">\n<p>&nbsp;<\/p>\n<table width=\"100%\">\n<thead>\n<tr>\n<th align=\"left\">\n<div>\n<div>Issue<\/div>\n<\/div>\n<\/th>\n<th align=\"left\">\n<div>\n<div>Fix<\/div>\n<\/div>\n<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Camera not detected<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Check USB connection, verify with `ls \/dev\/video*`<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>WebRTC not working<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Use Chrome, disable WebRTC local IP hiding flag<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Out of memory errors<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Use smaller model (VILA1.5-3b), reduce context length<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Low frame rate<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Reduce max-new-tokens, use smaller model, check camera resolution<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<tr>\n<td align=\"left\">\n<div>\n<div><strong>Video codec errors<\/strong><\/div>\n<\/div>\n<\/td>\n<td align=\"left\">\n<div>\n<div>Verify input format is H.264\/H.265, check jetson_utils installation<\/div>\n<\/div>\n<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<div class=\"fusion-text fusion-text-19\" style=\"--awb-margin-top:20px;\"><p><em><strong>For more information about Live LLaVA and advanced configurations, visit the <a style=\"color: #38c92e;\" href=\"https:\/\/github.com\/dusty-nv\/NanoLLM\"><span style=\"color: #38c92e;\"><span style=\"color: #38c92e;\">NanoLLM <span style=\"color: #38c92e;\">G<\/span><\/span><\/span><span style=\"color: #38c92e;\"><span style=\"color: #38c92e;\">itHub<\/span> repository.<\/span><\/a><\/strong><\/em><\/p>\n<\/div><\/div><\/div><\/div><\/div>\n","protected":false},"excerpt":{"rendered":"","protected":false},"author":3,"featured_media":1581,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[17],"tags":[53,55,52,56,54],"class_list":["post-773","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-generative-ai","tag-jetson-edge-visionlanguage-agent","tag-jetson-nanollm-live-llava-setup","tag-live-llava-on-jetson","tag-multimodal-stream-inference-jetson","tag-realtime-vlm-camera-pipeline"],"yoast_head":"<!-- This site is optimized with the Yoast SEO Premium plugin v25.3.1 (Yoast SEO v25.3.1) - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>Jetson Generative AI \u2013 Live LLaVA Run Live Llava locally on Jetson - OpenZeka EN Blog<\/title>\n<meta name=\"description\" content=\"Run Live LLaVA on Jetson with WebUI\u2014a powerful local real-time vision-language AI that understands your camera input and responds instantly.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"Jetson Generative AI \u2013 Live LLaVA\" \/>\n<meta property=\"og:description\" content=\"Run Live LLaVA on Jetson with WebUI\u2014a powerful local real-time vision-language AI that understands your camera input and responds instantly.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\" \/>\n<meta property=\"og:site_name\" content=\"OpenZeka EN Blog\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/www.facebook.com\/profile.php?id=61576911356211\" \/>\n<meta property=\"article:published_time\" content=\"2025-07-31T05:50:37+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-03-27T13:18:30+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg\" \/>\n\t<meta property=\"og:image:width\" content=\"1920\" \/>\n\t<meta property=\"og:image:height\" content=\"1500\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/jpeg\" \/>\n<meta name=\"author\" content=\"Enhar\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:creator\" content=\"@Aetherixnl\" \/>\n<meta name=\"twitter:site\" content=\"@Aetherixnl\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Enhar\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"4 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\/\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#article\",\"isPartOf\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\"},\"author\":{\"name\":\"Enhar\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/#\/schema\/person\/62c964376839cf2c4b2eb682bf14d3cb\"},\"headline\":\"Jetson Generative AI \u2013 Live LLaVA\",\"datePublished\":\"2025-07-31T05:50:37+00:00\",\"dateModified\":\"2026-03-27T13:18:30+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\"},\"wordCount\":4122,\"publisher\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/#organization\"},\"image\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage\"},\"thumbnailUrl\":\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg\",\"keywords\":[\"Jetson edge vision\u2011language agent\",\"Jetson NanoLLM Live LLaVA setup\",\"Live\u202fLLaVA on Jetson\",\"Multimodal stream inference Jetson\",\"Real\u2011time VLM camera pipeline\"],\"articleSection\":[\"Generative AI\"],\"inLanguage\":\"en-US\"},{\"@type\":\"WebPage\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\",\"url\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\",\"name\":\"Jetson Generative AI \u2013 Live LLaVA Run Live Llava locally on Jetson - OpenZeka EN Blog\",\"isPartOf\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage\"},\"image\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage\"},\"thumbnailUrl\":\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg\",\"datePublished\":\"2025-07-31T05:50:37+00:00\",\"dateModified\":\"2026-03-27T13:18:30+00:00\",\"description\":\"Run Live LLaVA on Jetson with WebUI\u2014a powerful local real-time vision-language AI that understands your camera input and responds instantly.\",\"breadcrumb\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage\",\"url\":\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg\",\"contentUrl\":\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg\",\"width\":1920,\"height\":1500},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\/\/blog.openzeka.com\/en\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Jetson Generative AI \u2013 Live LLaVA\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/#website\",\"url\":\"https:\/\/blog.openzeka.com\/en\/\",\"name\":\"Aetherix B.V.\",\"description\":\"NVIDIA Jetson Developer Kits &amp;Edge Devices\",\"publisher\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\/\/blog.openzeka.com\/en\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/#organization\",\"name\":\"Aetherix B.V.\",\"url\":\"https:\/\/blog.openzeka.com\/en\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/#\/schema\/logo\/image\/\",\"url\":\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/06\/aetherix-site-icon.webp\",\"contentUrl\":\"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/06\/aetherix-site-icon.webp\",\"width\":421,\"height\":398,\"caption\":\"Aetherix B.V.\"},\"image\":{\"@id\":\"https:\/\/blog.openzeka.com\/en\/#\/schema\/logo\/image\/\"},\"sameAs\":[\"https:\/\/www.facebook.com\/profile.php?id=61576911356211\",\"https:\/\/x.com\/Aetherixnl\",\"https:\/\/www.instagram.com\/aetherixnl\/\",\"https:\/\/www.tiktok.com\/@aetherixnl\"],\"description\":\"Aetherix provides a full range of NVIDIA Jetson-based edge AI solutions\u2014including Developer Kits, AI Kits, industrial-grade Carrier Boards, and fully integrated Boxed AI Systems.\",\"email\":\"info@aetherix.com\",\"legalName\":\"Aetherix B.V.\",\"vatID\":\"NL867727688B01\"},{\"@type\":\"Person\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/#\/schema\/person\/62c964376839cf2c4b2eb682bf14d3cb\",\"name\":\"Enhar\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/blog.openzeka.com\/en\/#\/schema\/person\/image\/\",\"url\":\"https:\/\/secure.gravatar.com\/avatar\/189d567adce3bb0c8d438b4586bf861ec04980f2e451003975e3cf871781d0f4?s=96&d=mm&r=g\",\"contentUrl\":\"https:\/\/secure.gravatar.com\/avatar\/189d567adce3bb0c8d438b4586bf861ec04980f2e451003975e3cf871781d0f4?s=96&d=mm&r=g\",\"caption\":\"Enhar\"}}]}<\/script>\n<!-- \/ Yoast SEO Premium plugin. -->","yoast_head_json":{"title":"Jetson Generative AI \u2013 Live LLaVA Run Live Llava locally on Jetson - OpenZeka EN Blog","description":"Run Live LLaVA on Jetson with WebUI\u2014a powerful local real-time vision-language AI that understands your camera input and responds instantly.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/","og_locale":"en_US","og_type":"article","og_title":"Jetson Generative AI \u2013 Live LLaVA","og_description":"Run Live LLaVA on Jetson with WebUI\u2014a powerful local real-time vision-language AI that understands your camera input and responds instantly.","og_url":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/","og_site_name":"OpenZeka EN Blog","article_publisher":"https:\/\/www.facebook.com\/profile.php?id=61576911356211","article_published_time":"2025-07-31T05:50:37+00:00","article_modified_time":"2026-03-27T13:18:30+00:00","og_image":[{"width":1920,"height":1500,"url":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg","type":"image\/jpeg"}],"author":"Enhar","twitter_card":"summary_large_image","twitter_creator":"@Aetherixnl","twitter_site":"@Aetherixnl","twitter_misc":{"Written by":"Enhar","Est. reading time":"4 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#article","isPartOf":{"@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/"},"author":{"name":"Enhar","@id":"https:\/\/blog.openzeka.com\/en\/#\/schema\/person\/62c964376839cf2c4b2eb682bf14d3cb"},"headline":"Jetson Generative AI \u2013 Live LLaVA","datePublished":"2025-07-31T05:50:37+00:00","dateModified":"2026-03-27T13:18:30+00:00","mainEntityOfPage":{"@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/"},"wordCount":4122,"publisher":{"@id":"https:\/\/blog.openzeka.com\/en\/#organization"},"image":{"@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage"},"thumbnailUrl":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg","keywords":["Jetson edge vision\u2011language agent","Jetson NanoLLM Live LLaVA setup","Live\u202fLLaVA on Jetson","Multimodal stream inference Jetson","Real\u2011time VLM camera pipeline"],"articleSection":["Generative AI"],"inLanguage":"en-US"},{"@type":"WebPage","@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/","url":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/","name":"Jetson Generative AI \u2013 Live LLaVA Run Live Llava locally on Jetson - OpenZeka EN Blog","isPartOf":{"@id":"https:\/\/blog.openzeka.com\/en\/#website"},"primaryImageOfPage":{"@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage"},"image":{"@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage"},"thumbnailUrl":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg","datePublished":"2025-07-31T05:50:37+00:00","dateModified":"2026-03-27T13:18:30+00:00","description":"Run Live LLaVA on Jetson with WebUI\u2014a powerful local real-time vision-language AI that understands your camera input and responds instantly.","breadcrumb":{"@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#primaryimage","url":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg","contentUrl":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/07\/12.jpg","width":1920,"height":1500},{"@type":"BreadcrumbList","@id":"https:\/\/blog.openzeka.com\/en\/jetson-generative-ai-live-llava\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/blog.openzeka.com\/en\/"},{"@type":"ListItem","position":2,"name":"Jetson Generative AI \u2013 Live LLaVA"}]},{"@type":"WebSite","@id":"https:\/\/blog.openzeka.com\/en\/#website","url":"https:\/\/blog.openzeka.com\/en\/","name":"Aetherix B.V.","description":"NVIDIA Jetson Developer Kits &amp;Edge Devices","publisher":{"@id":"https:\/\/blog.openzeka.com\/en\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/blog.openzeka.com\/en\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/blog.openzeka.com\/en\/#organization","name":"Aetherix B.V.","url":"https:\/\/blog.openzeka.com\/en\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/blog.openzeka.com\/en\/#\/schema\/logo\/image\/","url":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/06\/aetherix-site-icon.webp","contentUrl":"https:\/\/blog.openzeka.com\/en\/wp-content\/uploads\/2025\/06\/aetherix-site-icon.webp","width":421,"height":398,"caption":"Aetherix B.V."},"image":{"@id":"https:\/\/blog.openzeka.com\/en\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/www.facebook.com\/profile.php?id=61576911356211","https:\/\/x.com\/Aetherixnl","https:\/\/www.instagram.com\/aetherixnl\/","https:\/\/www.tiktok.com\/@aetherixnl"],"description":"Aetherix provides a full range of NVIDIA Jetson-based edge AI solutions\u2014including Developer Kits, AI Kits, industrial-grade Carrier Boards, and fully integrated Boxed AI Systems.","email":"info@aetherix.com","legalName":"Aetherix B.V.","vatID":"NL867727688B01"},{"@type":"Person","@id":"https:\/\/blog.openzeka.com\/en\/#\/schema\/person\/62c964376839cf2c4b2eb682bf14d3cb","name":"Enhar","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/blog.openzeka.com\/en\/#\/schema\/person\/image\/","url":"https:\/\/secure.gravatar.com\/avatar\/189d567adce3bb0c8d438b4586bf861ec04980f2e451003975e3cf871781d0f4?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/189d567adce3bb0c8d438b4586bf861ec04980f2e451003975e3cf871781d0f4?s=96&d=mm&r=g","caption":"Enhar"}}]}},"_links":{"self":[{"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/posts\/773","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/users\/3"}],"replies":[{"embeddable":true,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/comments?post=773"}],"version-history":[{"count":30,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/posts\/773\/revisions"}],"predecessor-version":[{"id":1580,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/posts\/773\/revisions\/1580"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/media\/1581"}],"wp:attachment":[{"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/media?parent=773"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/categories?post=773"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/blog.openzeka.com\/en\/wp-json\/wp\/v2\/tags?post=773"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}