negotiator402 Formulator LoRA

LoRA adapter trained to improve plain-language strategic scenarios -> strict numeric 2x2 payoff matrix JSON.

Training data: Alogotron/GameTheory-Formulator, 565 examples. Max steps: 1200.

Downloads last month
6
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Alogotron/negotiator402-formulator-lora-v2-1200

Base model

Qwen/Qwen2.5-3B
Adapter
(1407)
this model