Files
foxhunt/config/gpu/default.toml
jgrusewski 5546a45bc3 refactor: remove gpu_n_episodes override — auto-scale from VRAM everywhere
gpu_n_episodes was manually overridden in GPU profiles, training configs,
test files, and hyperopt — all set to 0 or small fixed values that
bypassed the auto-scaling logic, causing a div-by-zero crash in
train_baseline_rl.

Now: single auto-scaling path via optimal_n_episodes() from VRAM/SM
count. No manual override field. Cap at 16384 (consistent with
AutoBatchSizer's 8192 cap pattern). Floor at 32 for small GPUs.

Removed gpu_n_episodes from:
- DQNHyperparameters, PpoHyperparameters structs
- All 4 GPU profiles (rtx3050, h100, a100, default)
- Training profiles (smoketest, localdev)
- ExperienceProfile struct + serde
- Hyperopt adapter
- All test overrides

Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
2026-04-02 14:21:45 +02:00

14 lines
317 B
TOML

# Default GPU profile -- conservative settings for unknown GPUs
[training]
batch_size = 0 # 0 = auto-compute from VRAM
num_atoms = 21
buffer_size = 0 # 0 = auto from VRAM
hidden_dim_base = 256
replay_buffer_vram_fraction = 0.50
[experience]
gpu_timesteps_per_episode = 200
[cuda]
cuda_stack_bytes = 32768 # 32KB