> For the complete documentation index, see [llms.txt](https://docs.cloudeka.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.cloudeka.ai/reference/deployment-llama-3.1-70b-with-vllm-on-kubernetes/accessing-the-llama-service.md).

# Accessing the LLaMA Service

Once the service is running and has an EXTERNAL-IP, you can access the service using the EXTERNAL-IP and port 80.

## List Models

To list available models, use the following curl command.

```sh
curl -X GET http://<EXTERNAL-IP>/v1/models
```

## Create a Completion

To create a completion, use the following `curl` command.

```sh
curl -X POST http://<EXTERNAL-IP>/v1/completions \
     -H "Content-Type: application/json" \
     -d '{
           "model": "meta-llama/Llama-3.1-70B-Instruct",
           "prompt": "Once upon a time",
           "max_tokens": 50
         }'
```
