Skip to content

Q4_1 quantization compiling to vmfb megacommit - #2

Merged
stellaraccident merged 1 commit into
stellaraccident:mainfrom
Max191:turbine_llamacpp_i4_quantization
Feb 23, 2024
Merged

Q4_1 quantization compiling to vmfb megacommit#2
stellaraccident merged 1 commit into
stellaraccident:mainfrom
Max191:turbine_llamacpp_i4_quantization

Conversation

@Max191

@Max191 Max191 commented Feb 22, 2024

Copy link
Copy Markdown
Contributor

No description provided.

@Max191

Max191 commented Feb 22, 2024

Copy link
Copy Markdown
Contributor Author

I should split this into multiple PRs, but my commits accidentally got too jumbled up so I just squashed everything :P

I'll split it up tomorrow, but I'll leave this PR here in case anyone wants to see it or cherry pick it

@Max191

Max191 commented Feb 22, 2024

Copy link
Copy Markdown
Contributor Author

nod-ai/AMD-SHARK-ModelDev#473 is also needed for this PR

@stellaraccident stellaraccident left a comment

Copy link
Copy Markdown
Owner

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

It's fine. Prototype code we'll clean it in a future revision.

@stellaraccident
stellaraccident marked this pull request as ready for review February 23, 2024 00:14
@stellaraccident
stellaraccident merged commit e2189c7 into stellaraccident:main Feb 23, 2024
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants