KTransformers

Technical Work

This section separates technical background from user task guides. Use the Inference and Fine-Tuning sections when you want to run a model. Use Technical Work when you want to understand the system ideas, implementation direction, or public talks behind KTransformers.

Topics

TopicPage
CPU-GPU heterogeneous MoE inferenceHeterogeneous Inference
Workstation LoRA SFT directionLocal Fine-Tuning
Public talks and slide decksTalks and Slides
GitHub docs reading mapGitHub Docs Reading Map

Technical pages may discuss ideas before they become polished product workflows. To run a model directly, prefer pages that include package versions, commands, hardware configurations, and validation notes.