6 Commits
Author SHA1 Message Date
City 06a936813f Consolidate loader logic 2024-12-11 17:32:57 +01:00
City 7e96bf676c Add force offload node 2024-08-02 15:27:46 +02:00
a-One-Fan 8cff976421 Intel support for Pixart-Sigma (#64)
* Fix IPEX attn_drop

* IPEX SDPA 4GB fix

Alchemist GPUs lack the 64 bit emulation needed for >4GB in this case. Attention slicing workaround copied from https://github.com/vladmandic/automatic/blob/master/modules/intel/ipex/attention.py

* Do 4GB fix only when necessary

~20% performance boost to not use the 4GB fix when not necessary

* Move attention.py inside utils

* Better 4GB estimation

* Remove unnecessary warning
2024-07-06 20:26:29 +02:00
City 451c7588c7 dtype cleanup 2023-12-22 15:54:16 +01:00
City 3782b16606 Re add BF16 for VAE 2023-12-15 00:33:25 +01:00
City 389c16f2f5 Add dtype selection to T5 2023-11-28 00:49:06 +01:00