Articles with the tag “GPU”

Reference Architecture: Deploying a vision-language model with vLLM on OVHcloud MKS for high performance inference and full observability
EngineeringEléa Petton10/04/2026

GPU for LLM Inferencing Guide
EngineeringDavid Tonda24/07/2025

Pushing beyond the limits of embedded real-time AI for edge devices
Startup ProgramKatya Guez03/04/2025

Enhancing Customer Service with Interactive Avatars
Startup ProgramLeonard Pommereau20/03/2025

Five ways to develop sovereign, sustainable AI solutions
Startup ProgramCezary Skarzynski27/01/2025

How to serve LLMs with vLLM and OVHcloud AI Deploy
EngineeringMathieu Busquet29/05/2024

Fine-Tuning LLaMA 2 Models using a single GPU, QLoRA and AI Notebooks
EngineeringMathieu Busquet21/07/2023

Using GPU on Managed Kubernetes Service with NVIDIA GPU operator
EngineeringMaxime Hurtrel19/01/2022

Managing GPU pools efficiently in AI pipelines
GeneralBastien Verdebout22/12/2020