Struct GradientAccumulationCallback

Source

pub struct GradientAccumulationCallback { /* private fields */ }

Expand description

Gradient Accumulation callback.

Simulates larger batch sizes by accumulating gradients over multiple mini-batches before updating parameters. This is useful when GPU memory is limited but you want to train with effectively larger batches.

Effective batch size = mini_batch_size * accumulation_steps

Implementations§

Source §

impl GradientAccumulationCallback

Source

pub fn new(accumulation_steps: usize) -> TrainResult<Self>

Create a new Gradient Accumulation callback.

§Arguments

accumulation_steps - Number of mini-batches to accumulate (e.g., 4, 8, 16)

Source

pub fn accumulate( &mut self, gradients: &HashMap<String, Array<f64, Ix2>>, ) -> TrainResult<()>

Accumulate gradients.

Source

pub fn should_update(&self) -> bool

Check if we should perform an optimizer step.

Source

pub fn get_and_reset(&mut self) -> HashMap<String, Array<f64, Ix2>>

Get averaged accumulated gradients and reset.

Trait Implementations§

Source §

impl Callback for GradientAccumulationCallback

Source §

fn on_epoch_begin( &mut self, _epoch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the beginning of an epoch.

Source §

fn on_train_begin(&mut self, _state: &TrainingState) -> TrainResult<()>

Called at the beginning of training.

Source §

fn on_train_end(&mut self, _state: &TrainingState) -> TrainResult<()>

Called at the end of training.

Source §

fn on_epoch_end( &mut self, _epoch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the end of an epoch.

Source §

fn on_batch_begin( &mut self, _batch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the beginning of a batch.

Source §

fn on_batch_end( &mut self, _batch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the end of a batch.

Source §

fn on_validation_end(&mut self, _state: &TrainingState) -> TrainResult<()>

Called after validation.

Source §

fn should_stop(&self) -> bool

Check if training should stop early.

Auto Trait Implementations§

§

impl UnwindSafe for GradientAccumulationCallback

Blanket Implementations§

Source §

impl<T> Any for T
where T: 'static + ?Sized,

Source §

fn type_id(&self) -> TypeId

Gets the TypeId of self. Read more

Source §

impl<T> Borrow<T> for T
where T: ?Sized,

Source §

fn borrow(&self) -> &T

Immutably borrows from an owned value. Read more

Source §

impl<T> BorrowMut<T> for T
where T: ?Sized,

Source §

fn borrow_mut(&mut self) -> &mut T

Mutably borrows from an owned value. Read more

Source §

impl<T> From<T> for T

Source §

fn from(t: T) -> T

Returns the argument unchanged.

Source §

impl<T, U> Into for T
where U: From<T>,

Source §

fn into(self) -> U

Calls U::from(self).

That is, this conversion is whatever the implementation of From<T> for U chooses to do.

Source §

impl<T> IntoEither for T

Source §

fn into_either(self, into_left: bool) -> Either<Self, Self>

Converts self into a Left variant of Either<Self, Self> if into_left is true. Converts self into a Right variant of Either<Self, Self> otherwise. Read more

Source §

fn into_either_with<F>(self, into_left: F) -> Either<Self, Self>
where F: FnOnce(&Self) -> bool,

Converts self into a Left variant of Either<Self, Self> if into_left(&self) returns true. Converts self into a Right variant of Either<Self, Self> otherwise. Read more

Source §