Skip to content

NaN/Inf Occurs When Running MiniMax H3 with FP16 Precision (Tesla V100 16GB) #15262

Description

@aaalll12322

Custom Node Testing

Your question

I am using a Tesla V100 16GB graphics card (SM70 architecture) on vanilla ComfyUI with all custom nodes disabled. When I set the calculation precision of MiniMax H3 to FP16 via the ModelComputeDtype node, NaN and Inf values appear in tensors during inference, leading to failed generation.
Switching the precision to FP32 enables normal operation, yet the inference speed of FP32 is only 1/4 of the expected FP16 speed, resulting in severe performance degradation.
I would like to ask if there is a fix available for this FP16 numerical overflow issue.

Logs

Other

No response

Metadata

Metadata

Assignees

No one assigned

    Labels

    User SupportA user needs help with something, probably not a bug.

    Type

    No type

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions