Skip to content

GPTQ pseudo-quantization saved weights (pt format) How load Re-evaluation #50

Description

@CXiaorong

GPTQ pseudo-quantization saved weights (pt format) How load re-evaluation, I set the --load parameter after the execution, found an error:
gptq-main\quant.py", line 636, in forward
raise ValueError('Only supports a single token currently.')

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions