Initial K3 snapshot: 0.5B KDA/MLA/MoE train path
Standalone tree split from LLMRL/projects/kda. Includes Triton dt_bias backward fix, train_k3 --preset 0.5b, SFT, Docker runtime, and tests.
This commit is contained in:
@@ -0,0 +1,9 @@
|
||||
# Vendored FLA KDA kernels
|
||||
|
||||
Subset of [flash-linear-attention](https://github.com/fla-org/flash-linear-attention)
|
||||
used by `kda.ops` `backend="triton"`.
|
||||
|
||||
- License: MIT (see `LICENSE`)
|
||||
- Upstream version tag in `__init__.py`
|
||||
- Import path is `kda._fla.*`, not `fla.*`
|
||||
- Not included: context parallel, Ascend, TileLang, `flash_kda`, non-KDA ops
|
||||
Reference in New Issue
Block a user