Files
librefang-registry/providers/huggingface.toml
T
Evan 85db9c8d78 feat(providers): add supports_thinking to thinking-capable models (#53)
* feat(providers): add supports_thinking field to thinking-capable models

Mark models that support extended thinking / reasoning with
supports_thinking = true so the dashboard can conditionally show
thinking toggles.

Providers updated: anthropic (8), codex-cli (7), gemini (6),
openai (3), qwen (2), deepseek (1). Schema updated accordingly.

* feat(providers): add supports_thinking to remaining thinking-capable models

Cover 19 additional providers: alibaba-coding-plan, allenai, arcee-ai,
fireworks, gemini-cli, groq, liquid, nvidia-nim, ollama, openai (codex),
openrouter, perplexity, qwen-code, replicate, sambanova, tngtech,
venice, vertex-ai, xai.

Total: 62 models across 24 providers now have supports_thinking = true.

* feat(providers): add supports_thinking to chutes, huggingface, together

Missed in prior commits: DeepSeek-R1 on chutes/huggingface/together,
Qwen3-235B on chutes. Total now 66 models across 27 providers.

* feat(providers): add supports_thinking to bedrock, claude-code, aider, moonshot, stepfun

- bedrock: all 5 Claude models
- claude-code: all 3 models (opus/sonnet/haiku wrappers)
- aider: aider/sonnet (Claude-backed)
- moonshot: kimi-k2.5 (reasoning mode)
- stepfun: step-1o-turbo-vision (reasoning model)

Total: 77 models across 32 providers.

* feat(providers): add supports_thinking to alibaba kimi-k2.5, openrouter claude-sonnet-4
2026-04-15 00:46:51 +09:00

50 lines
1.1 KiB
TOML

# Hugging Face — https://huggingface.co
# Models: 3 (static entries; additional models discovered dynamically at runtime)
[provider]
id = "huggingface"
display_name = "Hugging Face"
api_key_env = "HF_API_KEY"
base_url = "https://api-inference.huggingface.co/v1"
key_required = true
[[models]]
id = "hf/meta-llama/Llama-3.3-70B-Instruct"
display_name = "Llama 3.3 70B (HF)"
tier = "balanced"
context_window = 128000
max_output_tokens = 4096
input_cost_per_m = 0.30
output_cost_per_m = 0.30
supports_tools = false
supports_vision = false
supports_streaming = true
aliases = []
[[models]]
id = "hf/deepseek-ai/DeepSeek-R1"
display_name = "DeepSeek R1 (HF)"
tier = "smart"
context_window = 64000
max_output_tokens = 4096
input_cost_per_m = 0.30
output_cost_per_m = 0.30
supports_tools = false
supports_vision = false
supports_streaming = true
supports_thinking = true
aliases = []
[[models]]
id = "hf/Qwen/Qwen2.5-72B-Instruct"
display_name = "Qwen 2.5 72B (HF)"
tier = "balanced"
context_window = 32768
max_output_tokens = 4096
input_cost_per_m = 0.30
output_cost_per_m = 0.30
supports_tools = false
supports_vision = false
supports_streaming = true
aliases = []