learnweak-evocua-8b-lora-r32-gimp

This repository contains a domain-specialized LoRA adapter for meituan/EvoCUA-8B-20260105 specialized for the GIMP domain. It was trained using the LearnWeak framework.

LearnWeak is an annotation-free specialization framework for small computer-use agents (CUAs) that uses a stronger reference agent to identify the student's weaknesses in the target domain, synthesize targeted tasks, and construct supervision automatically.

How to Get Started with the Model

Serve with vLLM

You can serve this adapter along with its base model using vLLM:

vllm serve meituan/EvoCUA-8B-20260105 \
  --enable-lora \
  --max-lora-rank 32 \
  --lora-modules learnweak-gimp=SujiKim/learnweak-evocua-8b-lora-r32-gimp

Use the LoRA module name, such as learnweak-gimp, when calling the served model.

Citation

If you find this work helpful, please cite:

@article{kim2026learnweaknessesautomateddomain,
  title   = {Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents},
  author  = {Kim, Suji and Kim, Kangsa and Hwang, Sung Ju},
  journal = {arXiv preprint arXiv:2605.28775},
  year    = {2026}
}

Acknowledgments

This project builds on OSWorld, LlamaFactory, and EvoCUA.

Downloads last month
13
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for SujiKim/learnweak-evocua-8b-lora-r32-gimp

Adapter
(8)
this model

Paper for SujiKim/learnweak-evocua-8b-lora-r32-gimp