Methodology · Wild GitHub
serving-llms-vllm
Serves LLMs with high throughput using vLLM's PagedAttention and continuous batching.
Composite
4.3
C 4.3 · A 0.0
How we got there
1 source verified
- Best source
github:SKILL.md - Authority tier Tier 3 — Wild GitHub
- Source link https://github.com/majiayu000/claude-skill-registry/blob/b2fb5f555f4285b8f35a58f3179a04dbdce9f170/skills/ai-ml/vllm/SKILL.md ↗
- First published 2026-08-31
Use this skill
/plugin install serving-llms-vllm More in Methodology
claude-api
Reference for the Claude API / Anthropic SDK — model ids, pricing, params, streaming, tool use, MCP, agents, caching, token counting, model migration.
prompt-engineering
Universal prompt engineering techniques for any LLM.
github-swyxio-ai-notes
notes for software engineers getting up to speed on new AI developments.
hatch-pet
Create, repair, validate, preview, and package Codex-compatible animated pet spritesheets from character art, screenshots, generated images, or visual references.
Auto-indexed. Editorial review pending — score is based on the rubric only.