Struct GradientAccumulationCallback

Source

pub struct GradientAccumulationCallback { /* private fields */ }

Expand description

Gradient Accumulation callback with advanced features.

Simulates larger batch sizes by accumulating gradients over multiple mini-batches before updating parameters. This is useful when GPU memory is limited but you want to train with effectively larger batches.

Effective batch size = mini_batch_size * accumulation_steps

§Features

Memory-efficient in-place accumulation
Multiple scaling strategies
Gradient overflow detection
Memory usage tracking
Automatic gradient zeroing

§Example

use tensorlogic_train::{GradientAccumulationCallback, GradientScalingStrategy};

let mut grad_accum = GradientAccumulationCallback::new(
    4, // accumulate over 4 mini-batches
    GradientScalingStrategy::Average,
).unwrap();

Implementations§

Source §

impl GradientAccumulationCallback

Source

pub fn new(accumulation_steps: usize) -> TrainResult<Self>

Create a new Gradient Accumulation callback with default average scaling.

§Arguments

accumulation_steps - Number of mini-batches to accumulate (e.g., 4, 8, 16)

Source

pub fn with_strategy( accumulation_steps: usize, scaling_strategy: GradientScalingStrategy, ) -> TrainResult<Self>

Create a new Gradient Accumulation callback with specified scaling strategy.

§Arguments

accumulation_steps - Number of mini-batches to accumulate
scaling_strategy - How to scale accumulated gradients

Source

pub fn with_grad_clipping(self, max_norm: f64) -> Self

Enable gradient clipping during accumulation.

§Arguments

max_norm - Maximum gradient norm before clipping

Source

pub fn accumulate( &mut self, gradients: &HashMap<String, Array<f64, Ix2>>, ) -> TrainResult<()>

Accumulate gradients with optional clipping and overflow detection.

Source

pub fn should_update(&self) -> bool

Check if we should perform an optimizer step.

Source

pub fn get_and_reset(&mut self) -> HashMap<String, Array<f64, Ix2>>

Get scaled accumulated gradients and reset state.

Source

pub fn get_stats(&self) -> GradientAccumulationStats

Get statistics about gradient accumulation.

Source

pub fn reset(&mut self)

Reset all state without returning gradients (useful for error recovery).

Trait Implementations§

Source §

impl Callback for GradientAccumulationCallback

Source §

fn on_epoch_begin( &mut self, _epoch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the beginning of an epoch.

Source §

fn on_train_begin(&mut self, _state: &TrainingState) -> TrainResult<()>

Called at the beginning of training.

Source §

fn on_train_end(&mut self, _state: &TrainingState) -> TrainResult<()>

Called at the end of training.

Source §

fn on_epoch_end( &mut self, _epoch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the end of an epoch.

Source §

fn on_batch_begin( &mut self, _batch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the beginning of a batch.

Source §

fn on_batch_end( &mut self, _batch: usize, _state: &TrainingState, ) -> TrainResult<()>

Called at the end of a batch.

Source §

fn on_validation_end(&mut self, _state: &TrainingState) -> TrainResult<()>

Called after validation.

Source §

fn should_stop(&self) -> bool

Check if training should stop early.

Auto Trait Implementations§

§

impl UnwindSafe for GradientAccumulationCallback

Blanket Implementations§

Source §

impl<T> Any for T
where T: 'static + ?Sized,

Source §

fn type_id(&self) -> TypeId

Gets the TypeId of self. Read more

Source §

impl<T> Borrow<T> for T
where T: ?Sized,

Source §

fn borrow(&self) -> &T

Immutably borrows from an owned value. Read more

Source §

impl<T> BorrowMut<T> for T
where T: ?Sized,

Source §

fn borrow_mut(&mut self) -> &mut T

Mutably borrows from an owned value. Read more

Source §

impl<T> From<T> for T

Source §

fn from(t: T) -> T

Returns the argument unchanged.

Source §

impl<T, U> Into for T
where U: From<T>,

Source §

fn into(self) -> U

Calls U::from(self).

That is, this conversion is whatever the implementation of From<T> for U chooses to do.

Source §

impl<T> IntoEither for T

Source §

fn into_either(self, into_left: bool) -> Either<Self, Self>

Converts self into a Left variant of Either<Self, Self> if into_left is true. Converts self into a Right variant of Either<Self, Self> otherwise. Read more

Source §

fn into_either_with<F>(self, into_left: F) -> Either<Self, Self>
where F: FnOnce(&Self) -> bool,

Converts self into a Left variant of Either<Self, Self> if into_left(&self) returns true. Converts self into a Right variant of Either<Self, Self> otherwise. Read more

Source §