mirror of
https://github.com/wassname/vllm.git
synced 2026-10-02 12:50:41 +08:00
1 parent
b9bcdc7158
commit
c0ce15dfb2
1 file changed
+1
-1
@@ -55,7 +55,7 @@ Start the serving the LLaMA-13B model on an A100 GPU:
|
||||
|
||||
$ sky launch serving.yaml
|
||||
|
||||
Check the output of the command. There will be a sharable gradio link (like the last line of the following). Open it in your browser to use the LLaMA model to do the text completion.
|
||||
Check the output of the command. There will be a shareable gradio link (like the last line of the following). Open it in your browser to use the LLaMA model to do the text completion.
|
||||
|
||||
.. code-block:: console
|
||||
|
||||
|
||||
Reference in new issue
Block a user