- hidden_size: 256 (smaller for faster iteration) - num_layers: 2 - color_attention_layers: 2, color_attention_heads: 8 - max_iterations: 1000 - advantage/strategy_updates: 256 (reduced from 512) - Ready for experimental training run Co-Authored-By: Claude Haiku 4.5 <noreply@anthropic.com>