Johannes Uusikuu
Johneeee
AI & ML interests
Small omlx models without vision & mtp for m1 m2 or low memory macs. Try the oQ63 quants they have 5 bit base but bump a lot of the model to 8 and 6 bit. My experimentation shows that giving a big bit budget bumps up 24 % of the tensors to 8 bit and around 15 percent to 6 bit. I do also experimentation with the last 4 layers in either 8 bit or bf16.
Recent Activity
updated a model about 2 hours ago
Johneeee/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-oQ5e-6target-bf16-last4_8bit-vision-mtp published a model about 2 hours ago
Johneeee/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-oQ5e-6target-bf16-last4_8bit-vision-mtp updated a model about 2 hours ago
Johneeee/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-uncher-oQ5e-6target-bf16-last4_8bit-vision-mtpOrganizations
None yet