google/skills

datalineage-bigquery-asset-impact-analysis

- Analyzes the downstream impact (blast radius) when a BigQuery table or view is broken, stale, or modified.

查看源码
仓库原始内容

按源仓库内容呈现,保留标题、案例、代码、表格、链接以及原文引用的演示图片。

BigQuery Asset Impact Analysis

This skill guides the agent in performing a downstream impact analysis (blast radius assessment) when a BigQuery table or view is reported as broken, stale, missing, or when a user is planning maintenance and wants to know the consequences of modifying or pausing updates to an asset.

It relies primarily on the Google Cloud Data Lineage (Knowledge Catalog) MCP Server to discover relationships between assets.

Prerequisites

This skill requires access to the Google Cloud Data Lineage API and an active client connection to the Data Lineage MCP Server. For detailed connection configurations and tool schemas, refer to MCP Usage.

Analysis Workflow

1. Resolve the Asset's Fully Qualified Name (FQN)

  • Ensure you have the correct FQN format for the BigQuery asset:
  • Format: bigquery:{project_id}.{dataset_id}.{table_or_view_id}
  • Example: bigquery:my-prod-project.analytics.orders

2. Determine Locations and Parent Path

Identify the locations to search and construct the Data Lineage API request:

  • Discover Asset Location: Run the command `bq show --format=json

{projectid}:{datasetid} and extract the location field (e.g., us-central1 or us`). If location discovery fails due to permissions or missing tools, prompt the user for the dataset's location.

  • Set Parent Path: Set the parent path using the project ID and the

MCP server's location. Consult the DataLineageServer tool definition to find the configured region or location (e.g., us). The format is: projects/{project_id}/locations/{mcp_server_location}.

  • Configure Search Scope: Include the discovered asset location in the

locations array of the payload (e.g., ["us-central1"] or ["us", "us-central1"]).

3. Retrieve the Downstream Lineage Graph

Call the DataLineageServer:search_lineage tool to fetch downstream relationships.

  • Direction: Set to DOWNSTREAM.
  • Search Parameters: Use max_depth = 10 and max_process_per_link = 5

as robust defaults.

4. Identify the Blast Radius

Traverse the returned lineage links to build the impact graph:

  • Affected Assets: The target of each link represents a downstream asset

that depends on your source asset.

  • Transform Processes: Inspect the processes field on each link. This

identifies the ETL pipelines, BigQuery Views, or Scheduled Queries that propagate the data.

  • Direct vs. Indirect Impact:
  • Direct Impact (Depth 1): Assets directly consuming the source asset.

If a link has dependency_type: EXACT_COPY, mark the target as "Directly Stale / Identical Copy".

  • Indirect Impact (Depth > 1): Assets further down the stream that

will experience cascading stale data or failures.

5. Summarize and Format the Output

Present your findings clearly to the user using the following structure:

  1. Executive Summary: State the total number of downstream assets affected

and the maximum depth of the impact.

  1. Critical Path: Highlight high-priority downstream assets (e.g., assets

containing "prod", "dashboard", "reporting", or "master" in their names).

  1. Blast Radius Table: A clean Markdown table listing the dependencies. You

MUST include all columns:

Downstream AssetTransform ProcessDepthImpact Type
bigquery:project.dataset.tableprojects/p/locations/l/processes/proc1Direct
bigquery:project.dataset.viewprojects/p/locations/l/processes/view2Indirect
  1. Analysis Metadata: Provide transparency on the parameters and boundaries

of your search so the user can choose to expand them:

  • Locations Searched: {list_of_locations_queried}
  • Parent Location: {parent_path}
  • Depth Limit: {max_depth}
  • Process per Link Limit: {max_process_per_link}
  • Tip for User: Let the user know they can request to rerun the analysis

with expanded locations or larger depth limits.

Crucial Constraints & Guardrails

  1. Interpret Empty Responses Correctly:
  • If the lineage response is empty, immediately assume that no

dependencies exist in the queried locations and report this to the user.

  1. Strictly Banned Bypasses:
  • Exclusively retrieve downstream relationships using the

DataLineageServer:search_lineage tool.

  1. Verify Asset Existence First:
  • If bq show indicates the source table does not exist, stop and report

this directly to the user. Do not attempt to guess alternative table names unless the user explicitly instructs you to do so.

  1. No Output Shortcutting or Hallucinated Artifacts:
  • Present the complete downstream blast radius table directly in your

final response. Avoid telling the user you have created a separate Markdown file or artifact containing the details unless you have explicitly executed file-writing tools to create it.

Reference Directory

  • MCP Usage: Using the Google Cloud Data Lineage

remote MCP server and tool preferences.

External Documentation

来自同一仓库

更多 Skills

全部 Skills
google
社区

google-analytics-admin-api-basics

- Manages Google Analytics account and property settings, enables the Analytics Admin API via the Cloud CLI, lists accounts and properties, and manages data streams, custom dimensions, conversion events, and integrations. Use when you need to programmatically configure Google Analytics accounts, provision properties, manage data retention, configure Measurement Protocol secrets, or manage Firebase and Google Ads links.

安装量
4
GitHub Stars
2万
最近更新
9月22日
google
社区

gke-workload-security

- Audits, configures, and hardens workload-level security controls for Google Kubernetes Engine (GKE) applications and namespaces. Covers running cluster security audits (auditcluster.sh), configuring Workload Identity Federation (impersonation, KSA/GSA binding, and pod setup), enforcing Network Policies (default-deny and Dataplane V2 logging), isolating high-risk pods inside GKE Sandbox (gVisor), enforcing Pod Security Standards (restricted labeling), and mounting Secret Manager secrets via CSI (SecretProviderClass). Use when auditing cluster security posture, isolating namespaces, applying pod security standards, setting up Workload Identity, or configuring network policies and secret volume mounts. Don't use for cluster-wide control plane security, RBAC hardening, Binary Authorization, Shielded Nodes, or enabling platform-level GKE add-ons (use gke-platform-security instead).

安装量
3
GitHub Stars
2万
最近更新
9月22日
google
社区

google-ads-api-account-diagnostics

- Diagnoses Google Ads account performance issues such as conversion loss (value or volume), low lead flow/volume, and lost impression share (opportunities) due to ad rank, bids, or budgets. Use when troubleshooting sudden performance drops, analyzing campaign impression share metrics, investigating low lead flow, or searching for bidding and budget constraints. Don't use for setting up new campaigns, uploading conversion events directly, or general Google Mobile Ads SDK integration issues (use gma-android-integrate instead).

安装量
4
GitHub Stars
2万
最近更新
9月22日
google
社区

google-ads-api-mcp-setup

Guides developers through downloading, configuring, and installing the official open-source Google Ads MCP Server. Use this skill when a user wants to connect their AI assistant (such as Gemini, Claude Code, or Cursor) to their Google Ads account to query campaigns or retrieve reporting metrics using natural language.

安装量
4
GitHub Stars
2万
最近更新
9月22日