home/categories/llm-ai/atrawog-bazzite-ai-plugins-bazzite-ai-jupyter-skills-inference-skill-md
llm-aidata-ai

inference

Fast inference with Unsloth and vLLM backend. Covers model loading, fast_generate(), thinking model output parsing, and memory management for efficient inference.

atrawog
maintainer
atrawog
์—…๋ฐ์ดํŠธ๋จ 1/12/2026
์Šคํƒ€
0
ํฌํฌ
0
quick start

Installation and usage

Fast inference with Unsloth and vLLM backend. Covers model loading, fast_generate(), thinking model output parsing, and memory management for efficient inference.

์„ค์น˜
$ install --globalskills.sh
์‚ฌ์šฉ๋ฒ•

์„ค์น˜ ํ›„ ํ„ฐ๋ฏธ๋„์—์„œ ๋‹ค์Œ ๋ช…๋ น์„ ์‹คํ–‰ํ•˜์—ฌ ์ด ์Šคํ‚ฌ์„ ์‚ฌ์šฉํ•  ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค:

skills use inference