- 概要
- Document Processing Contracts
- リリース ノート
- Document Processing Contracts について
- Box クラス
- IPersistedActivity インターフェイス
- PrettyBoxConverter クラス
- IClassifierActivity インターフェイス
- IClassifierCapabilitiesProvider インターフェイス
- ClassifierDocumentType クラス
- ClassifierResult クラス
- ClassifierCodeActivity クラス
- ClassifierNativeActivity クラス
- ClassifierAsyncCodeActivity クラス
- ClassifierDocumentTypeCapability クラス
- ContentValidationData クラス
- EvaluatedBusinessRulesForFieldValue クラス
- EvaluatedBusinessRuleDetails クラス
- ExtractorAsyncCodeActivity クラス
- ExtractorCodeActivity クラス
- ExtractorDocumentType クラス
- ExtractorDocumentTypeCapabilities クラス
- ExtractorFieldCapability クラス
- ExtractorNativeActivity クラス
- ExtractorResult クラス
- FieldValue クラス
- FieldValueResult クラス
- ICapabilitiesProvider インターフェイス
- IExtractorActivity インターフェイス
- ExtractorPayload クラス
- DocumentActionPriority 列挙型
- DocumentActionData クラス
- DocumentActionStatus 列挙型
- DocumentActionType 列挙型
- DocumentClassificationActionData クラス
- DocumentValidationActionData クラス
- UserData クラス
- Document クラス
- DocumentSplittingResult クラス
- DomExtensions クラス
- Page クラス
- PageSection クラス
- Polygon クラス
- PolygonConverter クラス
- Metadata クラス
- WordGroup クラス
- Word クラス
- ProcessingSource 列挙型
- ResultsTableCell クラス
- ResultsTableValue クラス
- ResultsTableColumnInfo クラス
- ResultsTable クラス
- Rotation 列挙型
- ルール クラス
- RuleResult クラス
- RuleSet クラス
- RuleSetResult クラス
- SectionType 列挙型
- WordGroupType 列挙型
- IDocumentTextProjection インターフェイス
- ClassificationResult クラス
- ExtractionResult クラス
- ResultsDocument クラス
- ResultsDocumentBounds クラス
- ResultsDataPoint クラス
- ResultsValue クラス
- ResultsContentReference クラス
- ResultsValueTokens クラス
- ResultsDerivedField クラス
- ResultsDataSource 列挙型
- ResultConstants クラス
- SimpleFieldValue クラス
- TableFieldValue クラス
- DocumentGroup クラス
- DocumentTaxonomy クラス
- DocumentType クラス
- Field クラス
- FieldType 列挙型
- FieldValueDetails クラス
- LanguageInfo クラス
- MetadataEntry クラス
- TextType 列挙型
- TypeField クラス
- ITrackingActivity インターフェイス
- ITrainableActivity インターフェイス
- ITrainableClassifierActivity インターフェイス
- ITrainableExtractorActivity インターフェイス
- TrainableClassifierAsyncCodeActivity クラス
- TrainableClassifierCodeActivity クラス
- TrainableClassifierNativeActivity クラス
- TrainableExtractorAsyncCodeActivity クラス
- TrainableExtractorCodeActivity クラス
- TrainableExtractorNativeActivity クラス
- BasicDataPoint クラス
- BasicValue クラス
- ComponentCollectionFacade クラス
- DataPointFacade基本クラス
- ExtractionResultHandler クラス
- FieldGroupDataPoint クラス
- FieldGroupValue クラス
- FieldLookup 基本クラス
- FieldRedactionSettings クラス
- RedactionOptions クラス (プレビュー)
- RedactionType 列挙型
- ResultsValueFacadeBase クラス
- TableDataPoint クラス
- TableRow クラス
- TableValue クラス
- WildcardDataPoint クラス
- WildcardDataPointCollection クラス
- Document Understanding ML
- Document Understanding OCR ローカル サーバー
- Document Understanding
- IntelligentOCR
- リリース ノート
- IntelligentOCR アクティビティ パッケージについて
- プロジェクトの対応 OS
- タクソノミーを読み込み
- ドキュメントをデジタル化
- ドキュメント分類スコープ
- キーワード ベースの分類器
- Document Understanding プロジェクト分類器
- インテリジェント キーワード分類器
- ドキュメント分類アクションを作成
- ドキュメント検証成果物を作成
- ドキュメント検証成果物を取得
- ドキュメント分類アクション完了まで待機し再開
- 分類器トレーニング スコープ
- キーワード ベースの分類器トレーナー
- インテリジェント キーワード分類器トレーナー
- データ抽出スコープ
- Document Understanding プロジェクト抽出器
- Document Understanding プロジェクト抽出器トレーナー
- 正規表現ベースの抽出器
- フォーム抽出器
- インテリジェント フォーム抽出器
- ドキュメントを墨消し
- ドキュメント検証アクションを作成
- ドキュメント検証アクション完了まで待機し再開
- 抽出器トレーニング スコープ
- 抽出結果をエクスポート
- マシン ラーニング抽出器
- マシン ラーニング抽出器トレーナー
- マシン ラーニング分類器
- マシン ラーニング分類器トレーナー
- 生成 AI 分類器
- 生成 AI 抽出器
- 認証を構成する
- ML サービス
- OCR
- OCR Contracts
- リリース ノート
- OCR コントラクトについて
- プロジェクトの対応 OS
- IOCRActivity インターフェイス
- OCRAsyncCodeActivity クラス
- OCRCodeActivity クラス
- OCRNativeActivity クラス
- Character クラス
- OCRResult クラス
- Word クラス
- FontStyles 列挙型
- OCRRotation 列挙型
- OCRCapabilities クラス
- OCRScrapeBase クラス
- OCRScrapeFactory クラス
- ScrapeControlBase クラス
- ScrapeEngineUsages 列挙型
- ScrapeEngineBase
- ScrapeEngineFactory クラス
- ScrapeEngineProvider クラス
- OmniPage
- PDF
- [リストから削除済] ABBYY
- [リストから削除済] ABBYY Embedded
[ドキュメントを墨消し] アクティビティ: 抽出結果と指定した単語を適用して、墨消しされた PDF を生成します。
UiPath.IntelligentOCR.Activities.Redaction.RedactDocument
説明
[ドキュメントを墨消し] アクティビティは、元の入力 PDF (ドキュメント パスとして指定) と、[抽出結果] および [墨消しする単語] 入力フィールドに基づいて、墨消し PDF を生成します。
[ドキュメントを墨消し] アクティビティは、ドキュメント オブジェクト モデルを使用して PDF 内で特定されたすべての単語の場所にアクセスします。一方、[抽出結果] フィールドと [墨消しする単語] フィールドは以下のように、墨消しする必要があるデータの入力として使用されます。
- [墨消しする単語] 入力配列の各エントリは、ドキュメント内で墨消しする単語の連続検索を行うための文字列と見なされます。大文字と小文字は区別されません。
- 参照がある抽出結果の値は、この参照値に基づいて編集されます (値の参照としての顧客領域の選択を含む)。標準フィールドと表のセルの両方が編集されています。
- 参照のない [抽出結果] の値 ([参照が必要] が False に設定されているフィールドに参照なしで追加された値) は、[墨消しする単語] フィールドのエントリと同様に見なされます。つまり、入力ドキュメント内に出現するその特定のテキストはすべて墨消しされます。
このアクティビティはドキュメント オブジェクト モデルを使用して単語を検索します。あいまい一致は使用できません。
墨消しの表示方法を制御できます。既定では、墨消しされた各領域は塗りつぶしのボックスで覆われています。RedactionOptions 入力 (プレビュー) を使用すると、この外観をフィールドごとに制御できます。たとえば、塗りつぶしのボックスまたは取り消し線のスタイル、カスタムの塗りつぶしと枠線の色と不透明度、墨消しされた各領域に表示される墨消しコード ラベルなどです。詳しくは、 RedactionOptions クラスをご覧ください。
非常に機密性の高いドキュメントを扱う場合は、抽出結果の人間による検証を行い、参照ベースの値と選択を使用することを強くお勧めします。これにより、墨消しが必要なすべてのデータが包括的にレビューされ、OCR エラーや語順の問題が最終的な墨消しの結果に影響を与える可能性が最小限に抑えられます。
プロジェクトの対応 OS
Windows
構成
デザイナー パネル
入力
- ドキュメント パス: 墨消しするドキュメントへのパスです。
- ドキュメント オブジェクト モデル: [ドキュメントをデジタル化] アクティビティから取得される、文書化された入力のドキュメント オブジェクト モデルです。
- 抽出結果 (任意): データ抽出プロセスの抽出結果です。
ExtractionResult変数に格納されます。これは、[ データ抽出スコープ] アクティビティから取得できます。 - 墨消しする単語 (任意): [抽出結果] 入力フィールドから取得されるデータに加えて、墨消しされる文字列のリスト。
- 出力ファイル: 墨消しした PDF を保存する出力ファイル パスです。
プロパティ パネル
共通
- 表示名: アクティビティの表示名です。
入力
- ドキュメント パス: 墨消しするドキュメントへのパスです。
- ドキュメント オブジェクト モデル: [ドキュメントをデジタル化] アクティビティから取得される、文書化された入力のドキュメント オブジェクト モデルです。
- 抽出結果 (任意): データ抽出プロセスの抽出結果です。
ExtractionResult変数に格納されます。これは、[ データ抽出スコープ] アクティビティから取得できます。 - 墨消しする単語 (任意): [抽出結果] 入力フィールドから取得されるデータに加えて、墨消しされる文字列のリスト。
- 出力ファイル: 墨消しした PDF を保存する出力ファイル パスです。
その他
- プライベート: オンにすると、変数および引数の値が Verbose レベルでログに出力されなくなります。
出力
- 出力ファイル: 墨消しされた情報を含む出力ファイルです。
墨消しの設定
- Border color: The color of the border used for redaction.
- Border thickness: The thickness of the border used for redaction.
- 解像度: 墨消しされた PDF に埋め込まれた画像の品質を表す DPI の値です。
- Fill color: The fill color used for redaction.
- Redaction code (preview): The default text to draw over each redaction that does not define its own redaction code. To split the text across multiple lines, use
\nwhere you want a line break, for exampleREDACTED\nCONFIDENTIAL. To display a literal backslash, type it twice, as inC:\\temp\\notes. The text is resized automatically to fit inside the redaction box. - Redaction code font color (preview): The default hex color of the redaction code text, used for any redaction that does not define its own font color.
- Redaction code font size (preview): The default font size of the redaction code text, used for any redaction that does not define its own font size. Use
0to auto-fit the text to the redaction box. - Redaction options (preview): Per-field redaction settings that override the global redaction appearance. Each entry uses a field id to style that field's redactions, or an empty field id to style the Words To Redact matches. Fields without a matching entry use the global FillColor, BorderColor, and BorderThickness values. For details, see the RedactionOptions class.
The Redaction code, Redaction code font color, and Redaction code font size properties set the defaults for the whole document. They apply to every redaction that does not already define its own value, including fields, table cells, and Words To Redact matches. Each one is applied independently, so a field keeps the values you set for it, and uses the defaults only for the values you left empty.