
Featured
How vLLM Semantic Router Trains Embedding Models and Publishes Them to Hugging Face
End-to-end walkthrough of fine-tuning cache LoRA and domain-adapted embeddings, evaluating them, uploading to llm-semantic-router on Hugging Face, and using them for routing and semantic cache.








