KTransformers
A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations
About
KTransformers is a research project focused on efficient inference and fine-tuning of large language models through CPU-GPU heterogeneous computing. The project now exposes two user-facing capabilities from the kt-kernel source tree: Inference and SFT.
Open Source Health
- Stars
- 19,505
- Forks
- 1,573
- License
- Apache-2.0
- Last commit
- 28 days ago
Resources & Links
Related Categories
Vendor
kvcache.ai
Open-source project on GitHub: KTransformers
Quick Links
Open Source
Related Products
Peft
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
Top category match
Sophia Script For Windows
zap: The most powerful PowerShell module for fine-tuning Windows 10 & Windows 11 on GitHub
Top category match
LongWriter
[ICLR 2025] LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs
Top category match
Regress Lm
Library for sequence-to-sequence numeric prediction, applicable to any tokenizable input, and allows pretraining and fine-tuning over multiple tasks.
Top category match
nanoGPT
The simplest, fastest repository for training/finetuning medium-sized GPTs.
Top category match