integration-direct: route to MiniMax M3 max-effort instead of Anthropic OAuth
Integration runs were burning the personal Anthropic 5-hour window (429 storm + window drain, 2026-09-02). New env-only context minimax-m3-max/v1 pins ANTHROPIC_MODEL=MiniMax-M3[1m], maxes the thinking budget, and sets the 1M auto-compact window; layered on minimax/v1 for base URL + apiKeyHelper auth. Claude-Session: https://claude.ai/code/session_019tJk7P8tZzJ24PvtgoGhLN
This commit is contained in:
19
harnesses/contexts/minimax-m3-max/v1/harness.yaml
Normal file
19
harnesses/contexts/minimax-m3-max/v1/harness.yaml
Normal file
@@ -0,0 +1,19 @@
|
||||
kind: context
|
||||
name: minimax-m3-max
|
||||
version: 1
|
||||
description: "Pin MiniMax M3 (1M context) at max reasoning effort — layer AFTER minimax/v1 (auth + base URL come from there)"
|
||||
requires: [claude-code]
|
||||
provides: []
|
||||
|
||||
# Model pin + effort for the MiniMax Anthropic-compatible proxy.
|
||||
# - ANTHROPIC_MODEL forces every main-loop request to M3 (1M context variant)
|
||||
# instead of the proxy's default Claude-name mapping.
|
||||
# - MAX_THINKING_TOKENS maxes the extended-thinking budget; the MiniMax
|
||||
# /anthropic layer translates thinking budget to M3 reasoning effort.
|
||||
# - CLAUDE_CODE_AUTO_COMPACT_WINDOW matches M3's real 1M window (the proxy's
|
||||
# model metadata under-reports 200K, which triggers premature compaction).
|
||||
env:
|
||||
ANTHROPIC_MODEL: "MiniMax-M3[1m]"
|
||||
ANTHROPIC_SMALL_FAST_MODEL: "MiniMax-M3"
|
||||
MAX_THINKING_TOKENS: "32000"
|
||||
CLAUDE_CODE_AUTO_COMPACT_WINDOW: "1000000"
|
||||
Reference in New Issue
Block a user