AdapterOps · recorded demo

Four adapters, one base model, the negative results kept

Four LoRA adapters over one Qwen2.5-1.5B base, served together with vLLM multi-LoRA, scored against a prompted base model and GPT-4o-mini on the same frozen golden sets.

Recorded outputs, not live inference. Every output below was saved during a run whose score is in the repository — the adapters on vLLM (A10, greedy, pinned revisions), GPT-4o-mini with the escalation arm's prompts — and the page refuses to build if these rows do not reproduce the published scores. The live Gradio app is in the repository as demo/app.py.