Skip to content

applying O3 quant config to llama #107

Description

@cjm715

how can I go about applying o3 quant config (static per-tensor a8w8) to llama? I don't see this in example notebooks

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions