Articles with the tag “LLM”

Reference Architecture: Deploying a vision-language model with vLLM on OVHcloud MKS for high performance inference and full observability
OVHcloud EngineeringEléa Petton10/04/2026

Reference Architecture: Custom metric autoscaling for LLM inference with vLLM on OVHcloud AI Deploy and observability using MKS
OVHcloud EngineeringEléa Petton10/02/2026

Reference Architecture: build a sovereign n8n RAG workflow for AI agent using OVHcloud Public Cloud solutions
OVHcloud EngineeringEléa Petton27/01/2026

Deep Dive into DeepSeek-R1 - Part 1
OVHcloud EngineeringFabien Ric06/03/2025

Mistral Small 24B served with vLLM and AI Deploy - a single command to deploy an LLM (Part 1)
OVHcloud EngineeringEléa Petton24/02/2025