BigQueryExampleGen
E1338344
UNEXPLORED
BigQueryExampleGen is a TFX ExampleGen component variant that reads training examples directly from Google BigQuery for use in TensorFlow Extended pipelines.
All labels observed (1)
| Label | Occurrences |
|---|---|
| BigQueryExampleGen canonical | 1 |
How this entity was disambiguated
This entity first appeared as the object of triple T18704859 — resolving that mention is where its identity was fixed. The disambiguator weighed these candidate entities and picked the highlighted one (or “None”, minting a new entity). This is how homonymy is resolved: the same surface form can point to different entities.
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: BigQueryExampleGen Context triple: [ExampleGen, hasSubcomponent, BigQueryExampleGen]
-
A.
Google BigQuery
Google BigQuery is a fully managed, serverless cloud data warehouse from Google Cloud designed for fast SQL-based analytics on large-scale datasets.
-
B.
Google Cloud Dataflow
Google Cloud Dataflow is a fully managed service for developing and executing batch and streaming data processing pipelines, based on Apache Beam, within the Google Cloud ecosystem.
-
C.
Apache Beam
Apache Beam is an open-source unified programming model for defining and executing batch and streaming data processing pipelines across multiple execution engines.
-
D.
Google Cloud Dataproc
Google Cloud Dataproc is a managed cloud service for running Apache Hadoop, Spark, and other big data workloads on scalable, automated clusters in Google Cloud.
-
E.
Bigtable
Bigtable is Google's distributed, scalable NoSQL database designed to handle massive amounts of structured data with high performance and reliability.
- F. None of above. chosen
- G. Unsure - the case is ambiguous/there is not enough information to decide.
NED2
Entity disambiguation (via description)
gpt-5-mini-2025-08-07
Target entity: BigQueryExampleGen Target entity description: BigQueryExampleGen is a TFX ExampleGen component variant that reads training examples directly from Google BigQuery for use in TensorFlow Extended pipelines.
-
A.
Google BigQuery
Google BigQuery is a fully managed, serverless cloud data warehouse from Google Cloud designed for fast SQL-based analytics on large-scale datasets.
-
B.
Google Cloud Dataflow
Google Cloud Dataflow is a fully managed service for developing and executing batch and streaming data processing pipelines, based on Apache Beam, within the Google Cloud ecosystem.
-
C.
Apache Beam
Apache Beam is an open-source unified programming model for defining and executing batch and streaming data processing pipelines across multiple execution engines.
-
D.
Google Cloud Dataproc
Google Cloud Dataproc is a managed cloud service for running Apache Hadoop, Spark, and other big data workloads on scalable, automated clusters in Google Cloud.
-
E.
Bigtable
Bigtable is Google's distributed, scalable NoSQL database designed to handle massive amounts of structured data with high performance and reliability.
- F. None of above. chosen
Referenced by (1)
Full triples — surface form annotated when it differs from this entity's canonical label.