Skip to content

ファイル

レジストリのファイルタブは、構造化データや非構造化データ、ドキュメント、バイナリーなど、あらゆるタイプのファイルを保存、共有、検索するための一元的な場所です。 通常、非構造化ナレッジワークフローの基盤となるデータを保管するために使用されますが、管理や再利用が必要なあらゆるファイルに活用することができます。

こちらもご覧ください。

DataRobot内のワークフローの一部として非構造化データが追加された場合、登録が完了すると、このページからファイルにアクセスできるようになります(関連するユースケースが削除された場合でも同様です)。

ここから、次のことができます。

  要素 説明
1 フィルター / 検索 ファイルをソース、オーナー、タグで絞り込むか、ファイル名で検索します。
2 + ファイルを追加 Register files using the selected method:
  • Upload file: Upload a single file stored locally.
  • Upload folder: Upload a folder stored locally.
  • Add from URL: Reference a file using a URL.
  • Browse data: Connect to and add files from an external data source. See Connectivity workflows.
3 アクションメニュー メニューを開くと、ファイルのダウンロードや削除ができます。

登録されているファイルごとに、以下の列がテーブルに表示されます。

説明
名前 登録されているファイルの名前。 ファイル名を変更するには、そのファイルをクリックします。
タイプ そのアイテムがファイルなのかフォルダーなのかを示します。
ソース そのファイルがどのようにアップロードされたかを示します(例:ローカルファイル)。
ファイル コンテナ内のファイルの数。
作成日時 そのファイルが最初に追加された日時。
最終変更 そのファイルが最後に更新された日時。
作成者 そのファイルを最初に追加したユーザー。
タグ そのファイルに関連付けられたユーザー定義のタグ。

1つ以上のファイルを選択すると、ページ上部に追加のファイル操作が表示されます。

現在選択されているアイテムの数がカウンターに表示され、その右側には以下のオプションがあります。

  • ユースケースへのリンク:ファイルを1つ以上のユースケースにリンクします。
  • 共有:ファイルをユーザー、グループ、組織と共有します。
  • タグ:フィルターに使用するタグをファイルに追加します。
  • 削除:レジストリからファイルを削除します。 これにより、関連付けられているユースケースからもファイルが削除されます。

ファイルの表示

ファイルの行をクリックすると、ファイルの詳細ペインが開き、プレビューやバージョン履歴など、特定のアイテムに関する追加情報が表示されます。 中央のパネルには、表形式のプレビューと、ファイルに対する操作が表示されます。

  要素 説明
1 ファイル名 クリックするとファイル名を編集できます。
2 アクション Opens the Actions dropdown, where you can:
  • Convert to folder and add files: Convert the file to a folder, then upload a file or folder from your local machine or add a file from a URL to the folder.
  • Link to Use Case: Link the files to one or more Use Cases.
  • Download: Download the file to your local machine.
  • Share: Share the files with a user, group, and/or organization.
  • Tag: Add tags to the files to use for filtering.
  • Delete: Delete the file from Registry and all Use Cases linked to it.

The new version appears under File versions on the right.
3 プレビューを展開 プレビューの全画面表示を開きます。
4 ダウンロード ファイルをローカルマシンにダウンロードします。
5 削除 レジストリおよびそれにリンクされているすべてのユースケースからファイルを削除します。
フォルダーの表示

フォルダーを表示する際、2つの表示オプションがあります。

  • リスト:新しいフォルダーを選択すると、その特定のフォルダーの情報のみが表示されます。 上部のパンくずリストを使用して、現在の階層を確認します。 この例では、メインフォルダーtest_fold_2が選択され、次にフォルダーnest_fold_1が選択されます。

  • :新しいフォルダーを選択すると、右側に新しいパネルが開きます。 このビューは、より多くのデータを一度に表示したり、フォルダー間を簡単に移動したりできるため、参照に適しています。

Viewing Jira and Confluence files

Files added from the Jira and Confluence connectors are stored as structured JSON. Selecting one opens an interactive preview that displays the content as a collapsible tree rather than raw text.

Expand and collapse the items to move through the structure—issue fields and comments for Jira, page content and metadata for Confluence. If a value contains additional JSON, the preview unwraps it so you can expand it in place instead of reading an escaped string.

As with any other file, click Expand preview to open the full-page view and Download to save the original JSON.

ファイルのメタデータ

右側のパネルにある情報ドロップダウンには、ファイルのメタデータが表示されます。

説明
タグ ファイルに追加されたユーザー定義のタグ。
サイズ ファイルのサイズ。
作成 ファイルがアップロードされた日時で、YYYY-MM-DD HH:MM:SSの形式で表示されます。
変更 ファイルが最後に更新された日時で、YYYY-MM-DD HH:MM:SSの形式で表示されます。
オーナー ファイルをアップロードしたユーザー。
カタログID プログラムでファイルを参照するために使用される一意の識別子。 各カタログアイテムには、固有のID、権限、およびバージョン履歴があります(つまり、カタログは1つですが、バージョンは複数存在する場合があります)。
バージョンID 特定のファイルバージョンを識別するための一意の識別子。

Schedule file refreshes

本機能の提供について

The ability to create a scheduled file refresh is only available for files that allow DataRobot to reread the original source to add the latest version. This includes files that were originally:

  • Added from an external data source via the Browse data modal.
  • Added from a URL.

Files uploaded from a local machine cannot be automatically refreshed because there is no source for DataRobot to re-read.

Scheduling allows you to automatically refresh a file on a specified schedule, so the latest version in Registry remains current with the source. Each time a scheduled refresh runs, DataRobot re-reads the file from its original source and adds it as a new version—all previous versions are listed under File versions, allowing you to compare what changed or use a previous version.

To schedule a file refresh:

  1. Expand Scheduling in the right panel and click + Schedule file refresh.
  2. To set a date for scheduled refreshes to begin, click the field below Start date and use the calendar picker to choose a date. The start date and cadence default to your current time.

  3. Under Cadence, select how often you want file refreshes to occur—daily, weekly, monthly, or annually.

  4. To finish creating the schedule, click Save.

Access and view the status of all scheduled refreshes by expanding Scheduling.

Open the Actions menu next to a refresh to:

  • Pause/Resume: Pause or resume the scheduled refresh.
  • Edit: Edit the schedule of the refresh.
  • Delete: Delete the scheduled refresh.

All scheduled file refreshes appear with one of the following status badges:

バッジ 説明
スケジュール済み The schedule is active and DataRobot will perform the refresh at the specified date/time/cadence.
一時停止 The schedule is not active and must be manually unpaused to resume the schedule.
進行中 The schedule was recently edited and is being saved.

トラブルシューティング

If DataRobot cannot reach the source to perform a refresh, for example, if credentials expired or access was revoked, the schedule is disabled rather than retrying because reauthorization requires users to sign in. When this happens, you cannot activate the same scheduled refresh. Instead, you must reconnect to the source and create the schedule again.

ファイルのバージョン管理

右側のパネルにあるバージョンドロップダウンには、バージョンが追加された日時や作成者など、ファイルのバージョン履歴が表示されます。

各ファイルエントリーは複数のバージョンをサポートしているため、個別のエントリーを作成しなくてもファイルの新しいバージョンをアップロードでき、ファイルの変更履歴を保持できます。 クリックすると、ファイルの特定のバージョンを表示できます。

次のステップ

Once files are registered, use them as source data for other workflows or manage them alongside your other registered assets.

  • Create a vector database: Use a dataset from the File Registry as the knowledge source when building a vector database for RAG.
  • Create custom models: Assemble a custom model in the workshop using a collection of registered files.
  • Data: Manage structured datasets in the companion Data Registry, which converts uploaded files into CSV format for use in experiments.