integration-direct: route to MiniMax M3 max-effort instead of Anthropic OAuth

Integration runs were burning the personal Anthropic 5-hour window
(429 storm + window drain, 2026-09-02). New env-only context
minimax-m3-max/v1 pins ANTHROPIC_MODEL=MiniMax-M3[1m], maxes the
thinking budget, and sets the 1M auto-compact window; layered on
minimax/v1 for base URL + apiKeyHelper auth.

Claude-Session: https://claude.ai/code/session_019tJk7P8tZzJ24PvtgoGhLN
This commit is contained in:
Paul O'Reilly
2026-09-02 21:33:50 +12:00
parent e89208f44b
commit 4201b1ac3a
2 changed files with 24 additions and 1 deletions

View File

@@ -0,0 +1,19 @@
kind: context
name: minimax-m3-max
version: 1
description: "Pin MiniMax M3 (1M context) at max reasoning effort — layer AFTER minimax/v1 (auth + base URL come from there)"
requires: [claude-code]
provides: []
# Model pin + effort for the MiniMax Anthropic-compatible proxy.
# - ANTHROPIC_MODEL forces every main-loop request to M3 (1M context variant)
# instead of the proxy's default Claude-name mapping.
# - MAX_THINKING_TOKENS maxes the extended-thinking budget; the MiniMax
# /anthropic layer translates thinking budget to M3 reasoning effort.
# - CLAUDE_CODE_AUTO_COMPACT_WINDOW matches M3's real 1M window (the proxy's
# model metadata under-reports 200K, which triggers premature compaction).
env:
ANTHROPIC_MODEL: "MiniMax-M3[1m]"
ANTHROPIC_SMALL_FAST_MODEL: "MiniMax-M3"
MAX_THINKING_TOKENS: "32000"
CLAUDE_CODE_AUTO_COMPACT_WINDOW: "1000000"