Skip to content

Questions regarding GGUF handling #91

Description

@danishblunt

Custom Node Testing

Your question

Hello there, first of all, great fork, love how well it works thus far, but do want to ask regarding GGUF's.

When using a qwen 2.5 vl scaled fp8 text model in combination with an qwen image edit 2511 Q4 guff, the memory footprint in my vRAM is significantly smaller than when using a qwen 2.5 vl Q2 model, which in theory should be much smaller, however I still ended up with a notably larger footprint when using the GGUF vairant, around 2.3GB vRAM less to be precise.

I understand this has a lot do with how ROCM generally handles GGUF's poorly and AMD natively seems to lack features to do quantization efficiently.

Are there any recommendations and or explanations from your side to possibly fix this behaviour or is it simply expected when using GGUF formats?

Logs

Other

No response

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions