Tez

E829915

Tez is a generalized data processing framework from the Apache Hadoop ecosystem designed to execute complex data-processing tasks efficiently, often used as an underlying engine for higher-level tools like Apache Pig and Hive.

All labels observed (1)

Label Occurrences
Tez canonical 2

How this entity was disambiguated

Statements (48)

Predicate Object
instanceOf Apache Hadoop ecosystem project
data processing framework
open-source software
advantageOverMapReduce enables better query optimization in Hive
reduces job startup overhead
supports complex DAGs instead of fixed map and reduce phases
comparedTo MapReduce
designedFor DAG-based data processing
generalized data processing
developer Apache Software Foundation
ecosystem Apache Hadoop
linked to: Hadoop
feature container reuse
counters and metrics
customizable data processing pipelines
directed acyclic graph execution
fault tolerance
input and output processors
pluggable shuffle handlers
session reuse
speculative execution
task-level optimization
timeline events
vertex parallelism
hasComponent Application Master
DAGAppMaster
linked to: ApplicationMaster

Edge
Task
Vertex
license Apache License 2.0
name Tez
optimizedFor efficient resource utilization
high throughput
low-latency execution
partOf Apache Big Data ecosystem
programmingLanguage Java
repository https://github.com/apache/tez
runsOn Hadoop YARN
linked to: YARN
supports ETL workloads
SQL-on-Hadoop workloads
batch processing
complex data-processing tasks
graph processing patterns
interactive processing
usedAs execution engine
usedBy Apache Cascading
Apache Hive
Apache Pig
website https://tez.apache.org/

How these facts were elicited

Referenced by (2)

Full triples — surface form annotated when it differs from this entity's canonical label.

Apache Tez name Tez
subject linked to: Tez