@metaharness/weight-eft
esm
Fine-tune cheap open-source LLMs (GLM, Qwen, DeepSeek) on your AI coding agent's successful runs with LoRA (SFT + DPO) so your model cascade escalates to expensive frontier models (GPT, Claude) less often — cutting cost-per-resolved. Turns run history int
Version 0.1.1 License MIT
Keywords
llmlorafine-tuningpeftsftdporlhf-alternativemodel-distillationknowledge-distillationai-agentscoding-agentagenticllm-agentswe-benchllm-routing
INSTALL