Update quickstart.rst (#2369)

827cbcd3 · Simon · GitHub · cb7a1c1c · 827cbcd3
Unverified Commit 827cbcd3 authored Jan 12, 2024 by Simon Committed by GitHub Jan 12, 2024
Hide whitespace changes
Inline Side-by-side

Showing with 1 addition and 1 deletion

docs/source/getting_started/quickstart.rst docs/source/getting_started/quickstart.rst +1 -1

No files found.
--- a/docs/source/getting_started/quickstart.rst
+++ b/docs/source/getting_started/quickstart.rst
@@ -95,7 +95,7 @@ OpenAI-Compatible Server
 ------------------------

 vLLM can be deployed as a server that mimics the OpenAI API protocol. This allows vLLM to be used as a drop-in replacement for applications using OpenAI API.
-By default, it starts the server at ``http://localhost:8000``. You can specify the address with ``--host`` and ``--port`` arguments. The server currently hosts one model at a time (OPT-125M in the above command) and implements `list models <https://platform.openai.com/docs/api-reference/models/list>`_, `create chat completion <https://platform.openai.com/docs/api-reference/chat/completions/create>`_, and `create completion <https://platform.openai.com/docs/api-reference/completions/create>`_ endpoints. We are actively adding support for more endpoints.
+By default, it starts the server at ``http://localhost:8000``. You can specify the address with ``--host`` and ``--port`` arguments. The server currently hosts one model at a time (OPT-125M in the command below) and implements `list models <https://platform.openai.com/docs/api-reference/models/list>`_, `create chat completion <https://platform.openai.com/docs/api-reference/chat/completions/create>`_, and `create completion <https://platform.openai.com/docs/api-reference/completions/create>`_ endpoints. We are actively adding support for more endpoints.

 Start the server: