@@ -51,13 +51,13 @@ copy .env.example .env # Windows
5151Open ` .env ` and fill in your real values:
5252
5353```
54- AZURE_MODEL_ROUTER_ENDPOINT=https://your-resource.services.ai.azure.com/models
54+ AZURE_MODEL_ROUTER_ENDPOINT=https://your-resource.services.ai.azure.com/openai/v1
5555AZURE_MODEL_ROUTER_KEY=your-model-router-key
5656AZURE_MODEL_ROUTER_DEPLOYMENT=model-router
57- AZURE_OPENAI_ENDPOINT=https://your-resource.openai. azure.com
57+ AZURE_OPENAI_ENDPOINT=https://your-resource.services.ai. azure.com/openai/v1
5858AZURE_OPENAI_KEY=your-azure-openai-key
5959AZURE_BASELINE_DEPLOYMENT=your-baseline-deployment
60- AZURE_JUDGE_ENDPOINT=https://your-resource.openai. azure.com
60+ AZURE_JUDGE_ENDPOINT=https://your-resource.services.ai. azure.com/openai/v1
6161AZURE_JUDGE_KEY=your-azure-openai-key
6262AZURE_JUDGE_DEPLOYMENT=your-judge-deployment
6363AZURE_PRICING_REGION=eastus
@@ -71,6 +71,11 @@ AZURE_PRICING_REGION=eastus
7171- * Azure Portal → your Foundry resource → Keys and Endpoint*
7272- * Azure Portal → your Azure OpenAI resource → Keys and Endpoint*
7373
74+ Use the resource name exactly as shown by the portal. If the portal gives a
75+ target URI ending in ` /chat/completions ` or ` /responses ` , remove that final
76+ operation segment and keep ` /openai/v1 ` . The live presets select the operation
77+ with ` api_mode ` .
78+
7479## Step 3: Configure the evaluation (optional)
7580
7681The default config (` configs/default.yaml ` ) works out of the box. The settings most people change first:
@@ -80,6 +85,7 @@ The default config (`configs/default.yaml`) works out of the box. The settings m
8085| Baseline model | ` endpoints.baseline.deployment_name ` | ` gpt-5 ` | Which model the router is compared against |
8186| Number of prompts | ` evaluation.sample_size ` | ` null ` (all) | How many dataset prompts to use |
8287| Judge enabled | ` judge.enabled ` | ` true ` | Whether to run quality scoring (costs extra API calls) |
88+ | Endpoint API | ` api_mode ` | Per endpoint | ` chat_completions ` or ` responses ` |
8389| Concurrency | ` concurrency.max_parallel_requests ` | ` 5 ` | How many prompts run in parallel |
8490
8591To change the baseline model, edit ` configs/default.yaml ` :
0 commit comments