Skip to content

Add Gemma4 and Nemotron3(bfcl) model to H200#31

Merged
khluu merged 1 commit into
vllm-project:mainfrom
tarukumar:gemma4_nemotron
Jun 30, 2026
Merged

Add Gemma4 and Nemotron3(bfcl) model to H200#31
khluu merged 1 commit into
vllm-project:mainfrom
tarukumar:gemma4_nemotron

Conversation

@tarukumar

@tarukumar tarukumar commented Jun 29, 2026

Copy link
Copy Markdown
Contributor

Add support for RedHatAI/gemma-4-31B-it-FP8-dynamic and BFCL nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8 on H200.

@tarukumar
tarukumar force-pushed the gemma4_nemotron branch 2 times, most recently from 855e7cb to 89137f4 Compare June 29, 2026 12:05
@tarukumar tarukumar changed the title Add Gemma4 and Nemotron-ultra model to H200 Add Gemma4 and Nemotron3(bfcl) model to H200 Jun 29, 2026

@dougbtv dougbtv left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

lgtm!

Signed-off-by: Tarun Kumar <takumar@redhat.com>
@sfeng33

sfeng33 commented Jun 29, 2026

Copy link
Copy Markdown

Thank you!!

@khluu
khluu merged commit 9133b80 into vllm-project:main Jun 30, 2026
2 checks passed
@tarukumar
tarukumar deleted the gemma4_nemotron branch June 30, 2026 19:52
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

4 participants