This website requires JavaScript.
Explore
Help
Register
Sign In
EngineX-Ascend
/
enginex-ascend-910-llama.cpp
Watch
10
Star
0
Fork
0
You've already forked enginex-ascend-910-llama.cpp
Code
Issues
Pull Requests
Actions
4
Projects
Releases
Wiki
Activity
2,943
Commits
1
Branch
1
Tag
e932094d58f513d5996c3efc9f6fed8238894c57
Commit Graph
2 Commits
Author
SHA1
Message
Date
Johannes Gäßler
133d99c599
CUDA: deduplicate FlashAttention code (
#7352
)
2024-05-18 12:36:25 +02:00
Johannes Gäßler
0fc1e820a9
CUDA: faster large batch FA without tensor cores (
#7314
)
2024-05-17 18:54:52 +02:00