Commit Graph
8 Commits
Author SHA1 Message Date
github-actions[bot] b0e0d2e8f4 chore: sync model pricing from OpenRouter API 2026-04-27 08:33:11 +00:00
Evan 7398983350 fix: use 127.0.0.1 instead of localhost for local provider base URLs (#73)
On dual-stack hosts (notably macOS), `localhost` resolves to both ::1
and 127.0.0.1 with IPv6 tried first. Local LLM servers (Ollama, vLLM,
LM Studio) installed via the standard scripts bind IPv4 only, so the
IPv6 connection attempt fails immediately and Happy Eyeballs fallback
to IPv4 isn't reliably triggered for connection-refused errors,
producing spurious "Configured local provider offline" warnings in
the daemon even when the server is up and reachable via curl.

Companion to librefang/librefang#3112 which fixes the hardcoded URL
constants in the main repo. After both land, existing installs pick
up the fix on their next registry sync.
2026-04-25 18:58:59 +09:00
github-actions[bot] 6d44be4277 chore: sync model pricing from OpenRouter API 2026-04-16 07:55:20 +00:00
Evan 4f8dd2404d feat(providers): expand ollama catalog with thinking-capable models (#58)
* feat(providers): expand ollama model catalog with thinking-capable models

Add commonly used local models with accurate capability flags:
- gemma4, gemma3: supports_thinking, supports_vision
- deepseek-r1: supports_thinking (fix missing flag)
- deepseek-v3: supports_tools
- qwen3, qwq: supports_thinking
- llama4: supports_vision
- llama3.3: supports_tools
- phi4: supports_tools

Previously only 6 models were listed and none had supports_thinking
(except deepseek-r1), causing the dashboard to hide thinking toggles
for models that actually support it.

* chore(providers): major cleanup — remove defunct providers and old models

Delete 21 defunct/obscure providers:
aion-labs, arcee-ai, deepcogito, eleutherai, essentialai, ibm-granite,
inception, inflection, kwaipilot, lemonade, liquid, morph, nex-agi,
nousresearch, prime-intellect, reka, relace, switchpoint, tngtech,
upstage, writer

Clean up 10 major providers — keep only latest generation models:
- anthropic: remove claude-3.5-sonnet (superseded by 4.x)
- openai: remove gpt-4o/4-turbo/3.5/o1/o3-mini (superseded by gpt-5/4.1/o3/o4-mini)
- gemini: remove 1.5-*/2.0-flash (superseded by 2.5/3.x)
- deepseek: remove coder/chat-v3-0324 (superseded by r1/v3)
- qwen: remove turbo/2.5-coder (superseded by qwen3)
- groq: remove old llama/mixtral/gemma (keep latest only)
- mistral: remove medium/nemo/pixtral-large (keep large/small/codestral)
- xai: remove grok-2 (superseded by grok-3/4)
- meta-llama: remove 3.x/guard (keep llama-4 + 3.3)
- ollama: rewrite with current models (gemma4, qwen3, qwq, llama4, etc)

Total: 90 → 48 models across major providers. All thinking-capable
models now have supports_thinking = true.

* chore: add pre-commit hook for automatic TOML formatting

- .githooks/pre-commit: runs taplo fmt on staged .toml files
- Makefile: add setup target + auto-configure hooks on first make
- .gitignore: add .make-setup-done and .sync_marker
2026-04-15 22:18:01 +09:00
github-actions[bot] 4979c355b2 chore: sync model pricing from OpenRouter API 2026-04-15 07:55:35 +00:00
Evan 85db9c8d78 feat(providers): add supports_thinking to thinking-capable models (#53)
* feat(providers): add supports_thinking field to thinking-capable models

Mark models that support extended thinking / reasoning with
supports_thinking = true so the dashboard can conditionally show
thinking toggles.

Providers updated: anthropic (8), codex-cli (7), gemini (6),
openai (3), qwen (2), deepseek (1). Schema updated accordingly.

* feat(providers): add supports_thinking to remaining thinking-capable models

Cover 19 additional providers: alibaba-coding-plan, allenai, arcee-ai,
fireworks, gemini-cli, groq, liquid, nvidia-nim, ollama, openai (codex),
openrouter, perplexity, qwen-code, replicate, sambanova, tngtech,
venice, vertex-ai, xai.

Total: 62 models across 24 providers now have supports_thinking = true.

* feat(providers): add supports_thinking to chutes, huggingface, together

Missed in prior commits: DeepSeek-R1 on chutes/huggingface/together,
Qwen3-235B on chutes. Total now 66 models across 27 providers.

* feat(providers): add supports_thinking to bedrock, claude-code, aider, moonshot, stepfun

- bedrock: all 5 Claude models
- claude-code: all 3 models (opus/sonnet/haiku wrappers)
- aider: aider/sonnet (Claude-backed)
- moonshot: kimi-k2.5 (reasoning mode)
- stepfun: step-1o-turbo-vision (reasoning model)

Total: 77 models across 32 providers.

* feat(providers): add supports_thinking to alibaba kimi-k2.5, openrouter claude-sonnet-4
2026-04-15 00:46:51 +09:00
Evan 553ecc6947 feat: add pricing sync script and update model prices from OpenRouter (#27)
* fix: pin npm package versions in MCP integration templates

Prevent supply chain attacks by pinning exact versions instead of
using unpinned `npx -y @package` which pulls latest on every run.

23 of 25 integrations pinned. sqlite-mcp and aws skipped (packages
not found on npm registry).

* fix: use stable azure/mcp version instead of beta

* feat: add pricing sync script and update model prices from OpenRouter API

- scripts/sync-pricing.py fetches real-time pricing from OpenRouter
- Updated 64 price fields across 13 provider files
- Run periodically or in CI to keep prices current
2026-03-25 23:43:13 +09:00
Evan 21c82e335c feat: initial model catalog with 196 models across 39 providers
Community-maintained TOML catalog for LibreFang. New models can be added
via PR without requiring a LibreFang binary release.

Includes validation script, bilingual docs, and GitHub templates.
2026-03-14 11:52:10 +09:00