- Crawlers: Add and edit crawlers to ingest data, and monitor the status and run history of all crawlers in this knowledge graph.
- Projects: Manage the projects that have access to this knowledge graph.
- Access: Manage the users who can contribute to this knowledge graph.
- In the left navigation, click , then Context Engine.
- Click the knowledge graph you want to access.
- In the top right, click Set up.
Request or grant access to a knowledge graph
If you don’t have an account role with the Create knowledge graph permission, follow these steps to request access to a knowledge graph:- Click the name of the knowledge graph on the Context Engine dashboard.
- In the Viewer access only box at the top of the visualization, click Request contributor access. This sends an email to all contributors on the knowledge graph notifying them that you have requested access.
- Click the name of the knowledge graph on the Context Engine dashboard.
- Click the Access tab on the left.
- In the Pending access requests section above the list of users, click Approve or Reject to answer the request.
Crawlers
A crawler is a continuous process that connects, discovers, and maps data structures to populate the knowledge graph and keep them updated. In the Crawlers tab, you can view and manage the crawlers that populate the knowledge graph with data from your warehouse and pipelines. You can add any number of crawlers to a knowledge graph. There are two types of crawler, each with their own prerequisites:- Warehouse data crawlers harvest your warehouse and structured sources supported via connectors to populate the graphs. This helps keep an accurate view of the data landscape. To run a Warehouse data crawler, your Maia project must contain a schema that is populated with data.
- Pipeline execution crawlers harvest your pipeline executions for your chosen project and environment to build an operational understanding of how data flows through your organization and workflow. To run a Pipeline execution crawler, your Maia project must contain at least one pipeline.
For warehouse data, after the first crawl is completed, subsequent crawls of the data you selected in your warehouse (for example, the selected datasets in your Google BigQuery warehouse) only crawl new data added since the last crawl.
Add a crawler
- In the Crawlers tab of a knowledge graph, click Add crawler.
-
In the Add crawler step:
- In Name, enter a name.
-
In Type, select Warehouse data or Pipeline execution, then follow the steps in the corresponding tab below.
- Warehouse data
- Pipeline execution
- In Project, select a project that is connected to the warehouse you want to harvest data from.
- In Environment, select the environment you want to use to harvest data.
- Click Continue.
- In the Select data step, select the data to crawl:
- If you’re crawling a Snowflake warehouse, select the database and schemas to crawl.
- If you’re crawling a Databricks warehouse, select the catalog and schemas to crawl.
- If you’re crawling an Amazon Redshift warehouse, select the schemas to crawl.
- If you’re crawling a Google BigQuery warehouse, select the datasets to crawl.
- Click Continue.
-
In the Schedule crawl step, choose how often this crawl should run. The crawl will run immediately after you create this knowledge graph, then follow this schedule.
- Select Standard or Advanced schedule settings.
- In Timezone, select the timezone for the crawl schedule.
- If you selected Standard settings, use the repeat drop-downs to set how often the crawler runs.
- If you selected Advanced settings, enter a cron expression to set how often the crawler runs.
- Click Add crawler.
Manage crawlers
To edit, reschedule, or monitor a crawler, in the Crawlers tab of a knowledge graph, click the name of the crawler you want to manage. Each crawler is configured in four tabs on the left, where you can perform the following actions. Click the Set up tab to:- View the crawler’s details (the data it crawls, the user who created it, and the user who last edited it).
- Edit the crawler’s configuration: Click Edit and make your changes to its name, project, environment, and/or the data it crawls. Then, click Save.
- Delete the crawler: Click Delete. Then, in the confirmation dialog, click Yes, delete.
- View and change the crawler’s schedule: Click Edit, make your changes, and then click Save.
- Pause a scheduled crawler: Click Pause in the top right. This crawler will not run until you resume it.
- Resume a paused crawler: Click Resume in the top right. This crawler will run at its next scheduled time.
- View the status, start and end time, and duration of the crawler’s most recent crawl.
- View a detailed crawl log: Click View logs to see a detailed log of the entities that the most recent crawl linked and added to your knowledge graph. You can filter the log to only show warnings and errors.
- View a list of all crawls performed by this crawler, including their status, start and end time, and duration.
- View detailed crawl logs: Click View logs for a specific crawl to learn more about the entities that this crawl linked and added to your knowledge graph. You can filter the log to only show warnings and errors.
Crawl statuses
The table below lists the possible crawl statuses you might see in the Crawlers tab or when viewing the Crawl history for a specific crawler.Projects
The Projects tab shows whether the knowledge graph is restricted, or available to all projects. If the knowledge graph is restricted, this tab lists the projects that can use this knowledge graph. If the knowledge graph is restricted:- To add a project to the allow list, click Add project, select one or more projects, then click Add.
- To remove a project from the allow list, click the three dots in the project’s row, then click Yes, remove.
- To make the knowledge graph available to all projects, click Restricted in the top right, then click Yes, allow all projects.
- To restrict access to the knowledge graph, click Restrict access. You will then need to add projects to the allow list as described above.
Access
The Access tab lists the users who can contribute to this knowledge graph. By default, the user who added a knowledge graph and all users with an account role with the Create knowledge graph permission can contribute to knowledge graphs. Users added to a knowledge graph can perform all available actions on the knowledge graph and its crawlers. To add a contributor to a knowledge graph, click Add contributor, search for and select one or more users, then click Add. To remove a contributor from a knowledge graph, click the trash can icon in their row. This immediately removes the contributor.Delete a knowledge graph
To delete a knowledge graph:- Click the name of the knowledge graph on the Context Engine dashboard.
- Click Delete graph in the top right.
- In the confirmation dialog, click Yes, delete.
