Views
No views yet
Deploy a (fine-tuned) LLM for fast, batched, OpenAI-compatible serving.
| Base model | Any (fine-tuned) LLM you serve |
| Task | fast LLM serving / deployment |
| Training objective | High-throughput batched inference (PagedAttention) — no training. |
| Track | LM · Language & multimodal |
| Built on | vllm-project/vllm |
| Notebook | |
| Compute / storage / time | GPU required — see the Compute · storage · time table in the notebook |
HfApi().upload_folder(...)) — the checkpoint + metrics.json + figures replace this placeholder.metrics.json · [ ] add figures · [ ] swap in the real results card1@misc{ropedia_academy,
2 title = {Ropedia Academy: an interactive course on embodied & spatial AI},
3 author = {Ropedia Academy},
4 year = {2026},
5 howpublished = {\url{https://chaoyue0307.github.io/ropedia-academy/}}
6}