ktransformers

by kvcache-ai · Developer Tools

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

New0 ratings19,471 starsActive
Developer ToolsDeveloper ToolPythonC++
Open project ↗↓ Download v0.7.0GitHub⚑ Report
ktransformers preview

About this project

A Flexible Framework for Experiencing Cutting-edge LLM Inference/Fine-tune Optimizations 🎯 Overview 🚀 Inference 🎓 SFT 🔥 Citation 🚀 Roadmap(2026Q2) 🎯 Overview KTransformers is a research project focused on efficient inference and fine-tuning of large language models through CPU-GPU heterogeneous computing. The project now exposes two user-facing capabilities from the kt-kernel source tree: Inference and SFT. 🔥 Updates Aug 26, 2026: Added native support for GLM-5.3-flash, bringing 1M-token context and multimodal input to consumer GPUs. (Tutorial) Aug 25, 2026: Uploaded a new easy-to-use KTransformers × LlamaFactory MoE Fine-Tuning Cookbook, covering hardware checks,…

Technologies

ShellPythonC++CudaCMake

Project health

Actively maintained
Last update6 days ago
Contributors116
Latest releasev0.7.0
Open issues & PRs507
LicenseApache-2.0
On GitHubsince 2024

GitHub

19,471
stars
1,562
forks
116
contributors
507
open issues & PRs
Python
language
6 days ago
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review ktransformers.

Built by

kvcache-ai
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain kvcache-ai/ktransformers? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like