GitHub / InseeFrLab / auto-tuning-vllm
Auto-tuning for vllm. Getting the best performance out of your LLM deployment (vllm+guidellm+optuna)
JSON API: https://ecosystem.code.gouv.fr/api/v1/hosts/GitHub/repositories/InseeFrLab%2Fauto-tuning-vllm
étoiles: 6
forks: 0
issues ouvertes: 15
licence: apache-2.0
langage: Python
taille: 2,99 Mo
dépendances analysées:
0
date de création: il y a 6 mois
date de mise à jour: il y a 18 jours
enregistré: il y a 23 jours
dernière synchronisation: il y a 6 jours
Sujets: benchmarking, gpu, guidellm, hyperparameter-optimization, llm-inference, llm-serving, llmops, optuna, speculative-decoding, vllm
No dependencies found