Target model:
MJPansa/MiniMax-M2.7-REAP-172B-A10B-AutoRound-W4A16/mnt/fast-ai/llm-models/minimax-m2.7-reap-autoround-w4a16/home/steve/src/vllm/mnt/fast-ai/llm-models/minimax-m2.7-int4-autoroundWhy this lane exists:
MiniMaxM2ForCausalLM architecture and AutoRound W4A16 quantization path, so it should exercise the existing XPU MiniMax/vLLM work before any new DeepSeek or STEP implementation work.Initial fit estimate:
Run order:
scripts/check-hf-metadata.shscripts/download-model.shscripts/quality-smoke.shscripts/bench-decode.shDo not submit LocalMaxxing results from this lane until quality gates pass against the existing MiniMax canaries and repeatability is clean.