LLMWildling's picture
Add files using upload-large-folder tool
0ac2e10 verified
|
Raw
History Blame Contribute Delete
2.24 kB
metadata
library_name: transformers
license: other
license_name: nvidia-nemotron-open-model-license
license_link: >-
  https://www.nvidia.com/en-us/agreements/enterprise-software/nvidia-nemotron-open-model-license/
pipeline_tag: text-generation
base_model:
  - nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4
language:
  - en
tags:
  - nvidia
  - nemotron-3
  - coding
  - agentic
  - tool-use
  - reasoning
  - quantized
  - NVFP4

NVIDIA-Nemotron-3-Super-155B-A13B-Coder-NVFP4

A coding-focused expansion of NVIDIA Nemotron 3 Super 120B-A12B NVFP4.

This is a community model and is not an official NVIDIA release. It was inspired by NVIDIA's open-model, open-data, and open-tooling work around Nemotron.

Model Summary

Total Parameters Approximately 155B
Active Parameters Approximately 13B per token
Quantization NVFP4 mixed-precision checkpoint
Architecture Nemotron hybrid Mamba-2, LatentMoE, Attention, and MTP
Validated Context Length 131,072 tokens
Specialization Agentic coding, reasoning, tool use, and multi-turn software-engineering workflows
Base Model NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4

Does This Work?

The public Nemotron 130B LLMWildling Canary NVFP4 provides a smaller proof point. It demonstrates direct recall of newly added domain knowledge and carries that knowledge into a follow-up task without RAG or prompt-injected context.

Intended Use

This checkpoint is intended for production coding assistants, repository analysis, agentic software-engineering systems, and tool-using workflows.

License

This model is derived from NVIDIA Nemotron 3 Super. Use is governed by the NVIDIA Nemotron Open Model License. Review the upstream model card for its full terms, safety information, limitations, and base-model details.