KTransformers as the framework for heterogeneous LLM inference and fine-tune optimizations is out today with its v0.7 feature release...

KTransformers as the framework for heterogeneous LLM inference and fine-tune optimizations is out today with its v0.7 feature release.
With KTransformers 0.7 there is now full AVX-512 support for LoRA fine-tuning without depending upon Advanced Matrix Extensions (AMX) also being present. This AVX-512-only without AMX benefits AMD EPYC Zen 4 / Zen 5 / Zen 6 servers with excellent AVX-512 support while lacking AMX and also older Intel Xeon processors with AVX-512 prior to the introduction of AMX with Sapphire Rapids.
The merge request noted the testing on AMD hardware and the foxus on AVX-512 without AMX platforms. The KTransformers runtime will automatically select the proper CPU implementation and in turn allowing MoE expert training to happen on a wider range of large-memory servers.
KTransformers 0.7 also adds VLM fine-tuning support, native FP8 LoRA support, improved DeepSeek V4 deployment, and CPU activation reuse.
More details on KTransformers 0.7 for those using it for LLM inference optimizations and fine-tuning can find all the details via the release announcement on GitHub.
| # | Наименование новости | Тональность | Информативность | Дата публикации |
|---|---|---|---|---|
| 1 | Intel-Scaler-vLLM 0.21.0-b1 Delivers Latest Features For vLLM On Intel GPUs | 0 | 7 | 10-07-2026 |
| 2 | Goal-Aware Adaptive Regulation Across Modalities: Evolutionary Architecture Search for Transformer Language Models on a Consumer AMD GPU | 0 | 5 | 17-07-2026 |
| 3 | AMD Working On A New Backend For Improving ROCm Compute In QEMU/VMs | 0 | 8.49 | 17-08-2026 |
| 4 | LLVM 23.1-rc1 Released With AMD Zen 6 & AVX-512 BMM Support, Other Compiler Enhancements | 0 | 12.99 | 23-07-2026 |
| 5 | Zlib-rs 0.6.6 Released With Updated Zlib API Support | 5 | 7 | 09-07-2026 |
| 6 | AWS Tunes Up Graviton5 For Agentic AI, Boosts Bang For The Buck Bigtime | 0 | 8.52 | 11-06-2026 |
| 7 | AMD ZenDNN 6.0 Brings Many Improvements For Accelerating Inference On Ryzen/EPYC CPUs | 5 | 7 | 08-07-2026 |
| 8 | Rückkehr von AVX-512: Nova Lake unterstützt Befehlssatzerweiterung in P- und E-Kernen | 0 | 5 | 07-07-2026 |
| 9 | Ryzen AI Software 1.8 Released With New Model Support, More Optimizations | 0 | 9.53 | 23-07-2026 |