Azure / Azure/azure-sdk-for-python
Error "Invalid URL (POST /v1/chat/completions)" when using gpt-4o-mini with azure-ai-inference library
- Dominant language
- Python
- Stars
- 5.6k
- Forks
- 3.4k
- Avg merge
- 1d 21h
- Merged PRs (30d)
- 193
Description
- **Package Name**: azure-ai-inference
- **Package Version**: 1.0.0b9
- **Python Version**: 3.11
**Describe the bug**
When attempting to use the gpt-4o-mini model (Deployment type: Standard) with ChatCompletionsClient, an error occurs:
```bash
Error: (None) Invalid URL (POST /v1/chat/completions)
Code: None
Message: Invalid URL (POST /v1/chat/completions)
```
**To Reproduce**
Steps to reproduce the behavior:
1. Set the environment variables AZURE_INFERENCE_ENDPOINT and AZURE_INFERENCE_KEY correctly.
2. Run the following code:
```python
import os
from azure.ai.inference import ChatCompletionsClient
from azure.core.credentials import AzureKeyCredential
from azure.ai.inference.models import SystemMessage, UserMessage
try:
client = ChatCompletionsClient(
endpoint=os.getenv("AZURE_INFERENCE_ENDPOINT"),
credential=AzureKeyCredential(os.getenv("AZURE_INFERENCE_KEY")),
)
response = client.complete(
messages=[
SystemMessage(content="You are a helpful assistant."),
UserMessage(content="What is the capital of France?"),
],
model="gpt-4o-mini",
)
print("azure-ai-inference response:", response.choices[0].message.content)
except Exception as e:
print(f"Error: {e}")
```
**Expected behavior**
I expected the API call to succeed and return a chat completion response from the gpt-4o-mini model, similar to how it works with other models like gpt-4o.
**Screenshots**
N/A (Error is text-based and printed to console)
**Additional context**
- I have verified that the endpoint and key are correct.
- The same code structure works with other models.
- The same model works when using the AzureOpenAI library
Contributor guide
Assessment
This issue has not been assessed yet.