Skip to content

feat(deepinfra/nvidia/NVIDIA-Nemotron-3.5-Lightning): add new models [bot] - #2138

Open
models-bot[bot] wants to merge 1 commit into
mainfrom
bot/add-deepinfra-nvidia-NVIDIA-Nemotron-3.5-Lightning-20260813-000526
Open

feat(deepinfra/nvidia/NVIDIA-Nemotron-3.5-Lightning): add new models [bot]#2138
models-bot[bot] wants to merge 1 commit into
mainfrom
bot/add-deepinfra-nvidia-NVIDIA-Nemotron-3.5-Lightning-20260813-000526

Conversation

@models-bot

@models-bot models-bot Bot commented Aug 13, 2026

Copy link
Copy Markdown
Contributor

Auto-generated by model-addition-agent for deepinfra/nvidia/NVIDIA-Nemotron-3.5-Lightning.


Note

Low Risk
Metadata-only addition with no runtime or auth changes; incomplete YAML vs other models could affect discovery or routing if validators require extra fields.

Overview
Adds a new DeepInfra provider YAML for nvidia/NVIDIA-Nemotron-3.5-Lightning (auto-generated by model-addition-agent).

The entry defines per-token costs ($0.05/M input, $0.20/M output), a 28,672 context and max output limit, and mode: unknown. It is thinner than sibling Nemotron files (no status, provisioning, sources, modalities, or supportedModes), which may matter if downstream tooling expects those fields.

Reviewed by Cursor Bugbot for commit c75e20e. Bugbot is set up for automated code reviews on this repo. Configure here.

@cursor cursor Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Cursor Bugbot has reviewed your changes using default effort and found 1 potential issue.

Fix All in Cursor

❌ Bugbot Autofix is OFF. To automatically fix reported issues with cloud agents, enable autofix in the Cursor dashboard.

Reviewed by Cursor Bugbot for commit c75e20e. Configure here.

limits:
context_window: 28672
max_tokens: 28672
mode: unknown

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Wrong mode for chat model

Medium Severity

mode is set to unknown for nvidia/NVIDIA-Nemotron-3.5-Lightning, but this is an OpenAI-compatible chat model with tool calling and reasoning. Consumers that filter on mode: chat will omit it, so it will not appear or route as a chat model.

Fix in Cursor Fix in Web

Reviewed by Cursor Bugbot for commit c75e20e. Configure here.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

0 participants