Files
librefang-registry/providers/microsoft.toml
T
Evan 9a7a2751e4 feat(providers): split microsoft api_key_env from github-copilot (#83)
Both `microsoft` (GitHub Models / Azure AI Inference at
models.inference.ai.azure.com) and `github-copilot` (the IDE subscription
product) previously declared `api_key_env = "GITHUB_TOKEN"`, so a single
PAT silently activated both providers — users with only the IDE
subscription saw GitHub Models entries appear in their model picker
without intent, and vice versa.

Rename `microsoft` to `GITHUB_MODELS_TOKEN`. `github-copilot` keeps
`GITHUB_TOKEN` as the established convention for the IDE side.

Existing users who set `GITHUB_TOKEN` to use the GitHub Models endpoint
will see the `microsoft` provider become unavailable until they set the
new env var. The librefang daemon will be updated separately to
recognize `GITHUB_TOKEN` as a deprecated fallback for `microsoft` during
a migration window, mirroring the legacy-fallback infrastructure added
in librefang/librefang#3279.

Refs librefang/librefang#3278
2026-04-27 16:34:37 +09:00

35 lines
994 B
TOML

# microsoft — auto-generated from OpenRouter API
#
# This entry is GitHub Models / Azure AI Inference (https://models.inference.ai.azure.com),
# distinct from the `github-copilot` IDE-subscription product even though both
# physically accept a GitHub Personal Access Token. Each now has its own env so
# users with only one product configured don't see the other appear in the model
# picker. See librefang/librefang#3278.
[provider]
id = "microsoft"
display_name = "Microsoft"
api_key_env = "GITHUB_MODELS_TOKEN"
base_url = "https://models.inference.ai.azure.com"
key_required = true
[[models]]
id = "phi-4"
display_name = "Microsoft: Phi 4"
tier = "fast"
context_window = 16384
max_output_tokens = 16384
input_cost_per_m = 0.065
output_cost_per_m = 0.14
supports_streaming = true
[[models]]
id = "wizardlm-2-8x22b"
display_name = "WizardLM-2 8x22B"
tier = "smart"
context_window = 65535
max_output_tokens = 8000
input_cost_per_m = 0.62
output_cost_per_m = 0.62
supports_streaming = true