Repository navigation
feat(sync): add Novita AI model catalog sync - #7050
Closed
Alex-yang00 wants to merge 27 commits into
Closed
Alex-yang00 wants to merge 27 commits into
Alex-yang00 wants to merge 27 commits into
Conversation
Alex-yang00
marked this pull request as ready for review
September 18, 2026 05:37
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
Action items
|
Contributor
|
No actionable findings. |
Contributor
Author
|
@rekram1-node Could you help take a look? |
Resolve conflicts with anomalyco#8029's corrected Novita output limits: - MiniMax M2.5 / M2.5 Highspeed and Kimi K2 0905 now match the lab limit (131_072 / 98_304), so the provider override is dropped. - Kimi K2 Thinking keeps a 98_304 output override (lab is 262_144). - Keep both NOVITA_API_KEY and MISTRAL_API_KEY in the sync workflow. Pin those verified limits in the Novita sync so the hourly sync does not restore the overstated /models values, and update the Vercel output-limit test now that the DeepSeek V3.1 Terminus lab entry supplies the 32_768 limit. Co-Authored-By: Claude <noreply@anthropic.com>
Contributor
|
No actionable findings. |
Contributor
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
https://api.novita.ai/openai/v1/modelsdeleteMissing: false), because visibility may be account- or tier-scopedNOVITA_API_KEYinto the hourly sync workflowData sources
Live reasoning verification
Requests were made directly to Novita
POST /openai/v1/chat/completionsusing the exact provider model IDs.thinking.type = enabled|disabledwas checked per allowlisted toggle model by observingreasoning_content.reasoning_effort = low|high|max; V4 Pro acceptshigh|max.moonshotai/kimi-k3acceptedreasoning_effort = low|high|max; each request returned reasoning tokens andreasoning_content.thinking_budget; a budget of 64 produced 64 reasoning tokens on the verified routes.qwen/qwen3-maxreturnsreasoning_contentwhen thinking is enabled, despite the canonical Alibaba route being non-reasoning.qwen/qwen3-235b-a22b-fp8returned no reasoning content or reasoning tokens with thinking both enabled and disabled, so Novita overrides the lab reasoning capability to false.qwen/qwen3-235b-a22b-thinking-2507honorsthinking_budget; a budget of 64 produced 64 reasoning tokens.deepseek/deepseek-r1-turboandbaidu/ernie-4.5-vl-424b-a47breturnedreasoning_contentwiththinking.type=enabledand none withdisabled.zai-org/glm-5.2acceptedreasoning_effort=none|high|max;noneomitted reasoning content,high|maxreturned it.deepseek/deepseek-r1-distill-llama-70bemitted reasoning in ordinary content under both thinking settings, without a separate reasoning side channel./modelslabelsopenai/gpt-oss-20bandopenai/gpt-oss-120bas image-input routes, but live image requests produced answers that the models cannot view the supplied image, so both are represented as text-only.moonshotai/kimi-k2.5accepted eventemperature=99; this does not demonstrate that sampling temperature is respected, so the Novita override was dropped in favor of the lab default.minimax/minimax-m2.1returnedreasoning_contentwiththinking.typeboth enabled and disabled, so it is represented as always-on reasoning withreasoning_options = [].Stale route verification
The following retained routes were removed only after direct chat-completion requests returned exact
404 MODEL_NOT_FOUND; absence from the account-scoped model list alone was not used as deletion evidence:xiaomimimo/mimo-v2-flashxiaomimimo/mimo-v2-proqwen/qwen3-4b-fp8qwen/qwen3-8b-fp8qwen/qwen3-30b-a3b-fp8qwen/qwen3-32b-fp8qwen/qwen3-next-80b-a3b-thinkingqwen/qwen3-vl-30b-a3b-thinkingdeepseek/deepseek-r1-distill-qwen-14bdeepseek/deepseek-r1-distill-qwen-32bbaidu/ernie-4.5-21B-a3b-thinkingbaidu/ernie-4.5-vl-28b-a3b-thinkingbaidu/ernie-4.5-vl-28b-a3binclusionai/ring-2.6-1tzai-org/glm-4.5deepseek/deepseek-prover-v2-671binclusionai/ling-2.6-1tinclusionai/ling-2.6-flashkwaipilot/kat-coder-proqwen/qwen2.5-vl-72b-instructsao10K/L3-8B-stheno-v3.2sao10K/l3-70b-euryale-v2.1sao10K/l3-8b-lunarissao10K/l31-70b-euryale-v2.2qwen/qwen3-vl-8b-instructSafety behavior
baidu/ernie-4.5-21B-a3breturned 429 anddeepseek/deepseek-r1-0528-qwen3-8breturned 503; neither was deleted based on transient responses.featuresarrays only add positive capability claims and cannot erase inherited capabilities.Required repository secret
The sync job requires the
NOVITA_API_KEYrepository secret.Validation
bun test packages/core/test/novita-ai.test.ts- 31 passedbun validate- passedbun models:sync novita-ai --dry-run --no-issues- 0 created, 0 updated, 0 removed, 101 unchangedpackages/core/test/sync.test.ts- new header regression test passes; two unrelated pre-existing failures remain (Hyper reasoning inheritance and DeepInfra live modalities)