Struct aws_sdk_comprehend::model::AugmentedManifestsListItem

source · [−]

#[non_exhaustive]pub struct AugmentedManifestsListItem {
    pub s3_uri: Option<String>,
    pub split: Option<Split>,
    pub attribute_names: Option<Vec<String>>,
    pub annotation_data_s3_uri: Option<String>,
    pub source_documents_s3_uri: Option<String>,
    pub document_type: Option<AugmentedManifestsDocumentTypeFormat>,
}

Expand description

An augmented manifest file that provides training data for your custom model. An augmented manifest file is a labeled dataset that is produced by Amazon SageMaker Ground Truth.

Fields (Non-exhaustive)

This struct is marked as non-exhaustive

Non-exhaustive structs could have additional fields added in future. Therefore, non-exhaustive structs cannot be constructed in external crates using the traditional Struct { .. } syntax; cannot be matched against without a wildcard ..; and struct update syntax will not work.

s3_uri: Option<String>

The Amazon S3 location of the augmented manifest file.

split: Option<Split>

The purpose of the data you've provided in the augmented manifest. You can either train or test this data. If you don't specify, the default is train.

TRAIN - all of the documents in the manifest will be used for training. If no test documents are provided, Amazon Comprehend will automatically reserve a portion of the training documents for testing.

TEST - all of the documents in the manifest will be used for testing.

attribute_names: Option<Vec<String>>

The JSON attribute that contains the annotations for your training documents. The number of attribute names that you specify depends on whether your augmented manifest file is the output of a single labeling job or a chained labeling job.

If your file is the output of a single labeling job, specify the LabelAttributeName key that was used when the job was created in Ground Truth.

If your file is the output of a chained labeling job, specify the LabelAttributeName key for one or more jobs in the chain. Each LabelAttributeName key provides the annotations from an individual job.

annotation_data_s3_uri: Option<String>

The S3 prefix to the annotation files that are referred in the augmented manifest file.

source_documents_s3_uri: Option<String>

The S3 prefix to the source files (PDFs) that are referred to in the augmented manifest file.

document_type: Option<AugmentedManifestsDocumentTypeFormat>

The type of augmented manifest. PlainTextDocument or SemiStructuredDocument. If you don't specify, the default is PlainTextDocument.

PLAIN_TEXT_DOCUMENT A document type that represents any unicode text that is encoded in UTF-8.
SEMI_STRUCTURED_DOCUMENT A document type with positional and structural context, like a PDF. For training with Amazon Comprehend, only PDFs are supported. For inference, Amazon Comprehend support PDFs, DOCX and TXT.

Struct aws_sdk_comprehend::model::AugmentedManifestsListItem

Fields (Non-exhaustive)

Implementations

impl AugmentedManifestsListItem

pub fn s3_uri(&self) -> Option<&str>

pub fn split(&self) -> Option<&Split>

pub fn attribute_names(&self) -> Option<&[String]>

pub fn annotation_data_s3_uri(&self) -> Option<&str>

pub fn source_documents_s3_uri(&self) -> Option<&str>

pub fn document_type(&self) -> Option<&AugmentedManifestsDocumentTypeFormat>

impl AugmentedManifestsListItem

pub fn builder() -> Builder

Trait Implementations

impl Clone for AugmentedManifestsListItem

fn clone(&self) -> AugmentedManifestsListItem

fn clone_from(&mut self, source: &Self)

impl Debug for AugmentedManifestsListItem

fn fmt(&self, f: &mut Formatter<'_>) -> Result

impl PartialEq<AugmentedManifestsListItem> for AugmentedManifestsListItem

fn eq(&self, other: &AugmentedManifestsListItem) -> bool

fn ne(&self, other: &AugmentedManifestsListItem) -> bool

impl StructuralPartialEq for AugmentedManifestsListItem

Auto Trait Implementations

impl RefUnwindSafe for AugmentedManifestsListItem

impl Send for AugmentedManifestsListItem

impl Sync for AugmentedManifestsListItem

impl Unpin for AugmentedManifestsListItem

impl UnwindSafe for AugmentedManifestsListItem

Blanket Implementations

impl<T> Any for T where T: 'static + ?Sized,

fn type_id(&self) -> TypeId

impl<T> Borrow<T> for T where T: ?Sized,

fn borrow(&self) -> &T

impl<T> BorrowMut<T> for T where T: ?Sized,

fn borrow_mut(&mut self) -> &mut T

impl<T> From<T> for T

fn from(t: T) -> T

impl<T> Instrument for T

fn instrument(self, span: Span) -> Instrumented<Self>

fn in_current_span(self) -> Instrumented<Self>

impl<T, U> Into<U> for T where U: From<T>,

fn into(self) -> U

impl<T> ToOwned for T where T: Clone,

type Owned = T

fn to_owned(&self) -> T

fn clone_into(&self, target: &mut T)

impl<T, U> TryFrom<U> for T where U: Into<T>,

type Error = Infallible

fn try_from(value: U) -> Result<T, <T as TryFrom<U>>::Error>

impl<T, U> TryInto<U> for T where U: TryFrom<T>,

type Error = <U as TryFrom<T>>::Error

fn try_into(self) -> Result<U, <U as TryFrom<T>>::Error>

impl<T> WithSubscriber for T

fn with_subscriber<S>(self, subscriber: S) -> WithDispatch<Self> where S: Into<Dispatch>,

fn with_current_subscriber(self) -> WithDispatch<Self>

impl<T> Any for T where
T: 'static + ?Sized,

impl<T> Borrow<T> for T where
T: ?Sized,

impl<T> BorrowMut<T> for T where
T: ?Sized,

impl<T, U> Into<U> for T where
U: From<T>,

impl<T> ToOwned for T where
T: Clone,

impl<T, U> TryFrom<U> for T where
U: Into<T>,

impl<T, U> TryInto<U> for T where
U: TryFrom<T>,

fn with_subscriber<S>(self, subscriber: S) -> WithDispatch<Self> where
S: Into<Dispatch>,