Articles with the tag “LLM”

Reference Architecture: Deploying a vision-language model with vLLM on OVHcloud MKS for high performance inference and full observability
EngineeringEléa Petton10/04/2026

Reference Architecture: Custom metric autoscaling for LLM inference with vLLM on OVHcloud AI Deploy and observability using MKS
EngineeringEléa Petton10/02/2026

Reference Architecture: build a sovereign n8n RAG workflow for AI agent using OVHcloud Public Cloud solutions
EngineeringEléa Petton27/01/2026

GPU for LLM Inferencing Guide
EngineeringDavid Tonda24/07/2025

Deep Dive into DeepSeek-R1
EngineeringFabien Ric06/03/2025

Mistral Small 24B served with vLLM and AI Deploy - a single command to deploy an LLM
EngineeringEléa Petton24/02/2025