Tez

E829915

Tez is a generalized data processing framework from the Apache Hadoop ecosystem designed to execute complex data-processing tasks efficiently, often used as an underlying engine for higher-level tools like Apache Pig and Hive.

All labels observed (1)

Label Occurrences
Tez canonical 2

How this entity was disambiguated

Statements (48)

Predicate Object
instanceOf Apache Hadoop ecosystem project ⓘ
data processing framework ⓘ
open-source software ⓘ
advantageOverMapReduce enables better query optimization in Hive ⓘ
reduces job startup overhead ⓘ
supports complex DAGs instead of fixed map and reduce phases ⓘ
comparedTo MapReduce ⓘ
designedFor DAG-based data processing ⓘ
generalized data processing ⓘ
developer Apache Software Foundation ⓘ
ecosystem Apache Hadoop ⓘ
linked to: Hadoop
feature container reuse ⓘ
counters and metrics ⓘ
customizable data processing pipelines ⓘ
directed acyclic graph execution ⓘ
fault tolerance ⓘ
input and output processors ⓘ
pluggable shuffle handlers ⓘ
session reuse ⓘ
speculative execution ⓘ
task-level optimization ⓘ
timeline events ⓘ
vertex parallelism ⓘ
hasComponent Application Master ⓘ
DAGAppMaster ⓘ
linked to: ApplicationMaster

Edge ⓘ
Task ⓘ
Vertex ⓘ
license Apache License 2.0 ⓘ
name Tez ⓘ
optimizedFor efficient resource utilization ⓘ
high throughput ⓘ
low-latency execution ⓘ
partOf Apache Big Data ecosystem ⓘ
programmingLanguage Java ⓘ
repository https://github.com/apache/tez ⓘ
runsOn Hadoop YARN ⓘ
linked to: YARN
supports ETL workloads ⓘ
SQL-on-Hadoop workloads ⓘ
batch processing ⓘ
complex data-processing tasks ⓘ
graph processing patterns ⓘ
interactive processing ⓘ
usedAs execution engine ⓘ
usedBy Apache Cascading ⓘ
Apache Hive ⓘ
Apache Pig ⓘ
website https://tez.apache.org/ ⓘ

How these facts were elicited

Referenced by (2)

Full triples — surface form annotated when it differs from this entity's canonical label.

Apache Tez → name → Tez ⓘ
subject linked to: Tez