Skip to main content
Ctrl+K
pimm pimm

pimm 0.5.2

  • Get started
  • Models
  • Data
  • Training
  • Develop
  • Reference
  • GitHub
  • Hugging Face
  • Get started
  • Models
  • Data
  • Training
  • Develop
  • Reference
  • GitHub
  • Hugging Face

Section Navigation

  • Train and pretrain
  • Configs and overrides
  • Scale up
  • Checkpoints and resume
  • Evaluate
  • Monitor and debug
  • Troubleshooting
  • Training

Training#

Run recipes, change them, run them on more GPUs and clusters, and monitor your runs.

  • Train and pretrain
    • Pick a recipe
    • Check it on a few events
    • Keep a variant as a child config
    • Start the run
    • From a notebook
  • Configs and overrides
    • Experiment configs
    • Overrides on the command line
    • Precedence
    • Common experiment fields
    • Check before running
  • Scale up
    • More GPUs on one machine
    • Run on Slurm
    • Interactive allocations and chains
    • Launch recipes
    • exex
  • Checkpoints and resume
    • What’s saved
    • Warm start and resume
    • Resume a run
    • Changing the number of GPUs
    • model_best.pth
    • Errors
  • Evaluate
    • During training
    • On a finished run
  • Monitor and debug
    • Weights & Biases
    • TensorBoard
    • Diagnostic hooks
    • Structured traces
  • Troubleshooting
    • Installation
    • Data
    • Launcher and Slurm
    • Training
    • Checkpoints and evaluation
    • Reporting a problem

previous

Transforms

next

Train and pretrain

© Copyright 2026, DeepLearnPhysics.

Created using Sphinx 8.1.3.

Built with the PyData Sphinx Theme 0.19.0.