Triple

T23080442
Position Surface form Disambiguated ID Type / Status
Subject Bashkir Latin alphabet E575453 entity
Predicate basedOn P98 FINISHED
Object Uniform Turkic Latin alphabet
The Uniform Turkic Latin alphabet was a standardized Latin-based script used in the early Soviet Union to write multiple Turkic languages before being replaced by Cyrillic and other writing systems.
E1570242 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Uniform Turkic Latin alphabet | Statement: [Bashkir Latin alphabet, basedOn, Uniform Turkic Latin alphabet]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Uniform Turkic Latin alphabet
Context triple: [Bashkir Latin alphabet, basedOn, Uniform Turkic Latin alphabet]
  • A. Uyghur Latin alphabet
    The Uyghur Latin alphabet is a romanized writing system developed for the Uyghur language, used primarily in digital communication and linguistic transcription.
  • B. Turkmen alphabet
    The Turkmen alphabet is the standardized script used to write the Turkmen language, currently based on a modified Latin script adopted after the Soviet era.
  • C. Kazakh Latin alphabet
    The Kazakh Latin alphabet is a modern script based on the Latin writing system that has been adopted for writing the Kazakh language as part of Kazakhstan’s language reform and modernization efforts.
  • D. Turkmen Cyrillic alphabet
    The Turkmen Cyrillic alphabet is a Cyrillic-based writing system formerly used for the Turkmen language, particularly during the Soviet era before its replacement by a Latin-based script.
  • E. Turkmen Arabic alphabet
    The Turkmen Arabic alphabet is a historical writing system based on the Arabic script that was formerly used to write the Turkmen language before being replaced by Latin- and Cyrillic-based alphabets.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Uniform Turkic Latin alphabet
Triple: [Bashkir Latin alphabet, basedOn, Uniform Turkic Latin alphabet]
Generated description
The Uniform Turkic Latin alphabet was a standardized Latin-based script used in the early Soviet Union to write multiple Turkic languages before being replaced by Cyrillic and other writing systems.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Uniform Turkic Latin alphabet
Target entity description: The Uniform Turkic Latin alphabet was a standardized Latin-based script used in the early Soviet Union to write multiple Turkic languages before being replaced by Cyrillic and other writing systems.
  • A. Uyghur Latin alphabet
    The Uyghur Latin alphabet is a romanized writing system developed for the Uyghur language, used primarily in digital communication and linguistic transcription.
  • B. Turkmen alphabet
    The Turkmen alphabet is the standardized script used to write the Turkmen language, currently based on a modified Latin script adopted after the Soviet era.
  • C. Kazakh Latin alphabet
    The Kazakh Latin alphabet is a modern script based on the Latin writing system that has been adopted for writing the Kazakh language as part of Kazakhstan’s language reform and modernization efforts.
  • D. Turkmen Cyrillic alphabet
    The Turkmen Cyrillic alphabet is a Cyrillic-based writing system formerly used for the Turkmen language, particularly during the Soviet era before its replacement by a Latin-based script.
  • E. Turkmen Arabic alphabet
    The Turkmen Arabic alphabet is a historical writing system based on the Arabic script that was formerly used to write the Turkmen language before being replaced by Latin- and Cyrillic-based alphabets.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e245be28d48190ad1348d5a73db37d completed April 17, 2026, 2:37 p.m.
NER Named-entity recognition batch_69f18c66a80481909ebc2ba69f1e4bd9 completed April 29, 2026, 4:43 a.m.
NED1 Entity disambiguation (via context triple) batch_6a0c15b0a6ac819085f13e7bdc82840f completed May 19, 2026, 7:48 a.m.
NEDg Description generation batch_6a0c175e246081909d6451f245ccfe66 completed May 19, 2026, 7:55 a.m.
NED2 Entity disambiguation (via description) batch_6a0c1820a4d88190aed59f50b4e7ec47 completed May 19, 2026, 7:58 a.m.
Created at: April 17, 2026, 3:56 p.m.