OpenVino Model Server Demo // The Easiest Way to Run LLMs Locally​ | Intel Software

Intel Devs Guide 1 months ago

Description

Ezequiel Lanza, Intel AI software evangelist, demonstrates how developers can locally run a modern LLM using familiar APIs. In this episode, Ezequiel explains how developers can reroute an existing OpenAI‑based application to local, optimized inference while examining the design decisions behind the setup.​

Try it yourself: https://docs.openvino.ai/2025/model-server/ovms_docs_rest_api_chat.html ​

Deploy OVMS on Docker: https://docs.openvino.ai/2025/model-server/ovms_docs_deploying_server.html

About Intel Software:
Intel® Developer Zone is committed to empowering and assisting software developers in creating applications for Intel hardware and software products. The Intel Software YouTube channel is an excellent resource for those seeking to enhance their knowledge. Our channel provides the latest news, helpful tips, and engaging product demos from Intel and our numerous industry partners. Our videos cover various topics; you can explore them further by following the links.

Connect with Intel Software:
INTEL SOFTWARE WEBSITE: https://intel.ly/2KeP1hD
INTEL SOFTWARE on FACEBOOK: http://bit.ly/2z8MPFF
INTEL SOFTWARE on TWITTER: http://bit.ly/2zahGSn
INTEL SOFTWARE GITHUB: http://bit.ly/2zaih6z
INTEL DEVELOPER ZONE LINKEDIN: http://bit.ly/2z979qs
INTEL DEVELOPER ZONE INSTAGRAM: http://bit.ly/2z9Xsby
INTEL GAME DEV TWITCH: http://bit.ly/2BkNshu

#intelsoftware

OpenVino Model Server Demo // The Easiest Way to Run LLMs Locally​ | Intel Software