Skip to content

Gw/dllm int operation - #33

Open
GeorgeWu1204 wants to merge 36 commits into
blou_dllmfrom
gw/dllm_int_operation
Open

Gw/dllm int operation#33
GeorgeWu1204 wants to merge 36 commits into
blou_dllmfrom
gw/dllm_int_operation

Conversation

@GeorgeWu1204

Copy link
Copy Markdown
Collaborator

The main changes are:

  • Rewrite the VectorSRAM, now based on binary instead of simple vector of fp. This is to make it support storing int32 while preserving the fp storage.
  • Add funct2 to the PREFETCH_V instr so that it support either "mx" or "normal" load pattern. "normal" pattern means without needs to load the separate scales as mx.
  • Change env build function to map integers to the HBM.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants