Triple
T9031945
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Pochury Naga |
E216392
|
entity |
| Predicate | language |
P15
|
FINISHED |
| Object |
Pochury language
Pochury language is a Sino-Tibetan language spoken by the Pochury Naga people of Nagaland in northeastern India.
|
E772412
|
NE FINISHED |
How this triple was built (4 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Pochury language | Statement: [Pochury Naga, language, Pochury language]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: Pochury language Context triple: [Pochury Naga, language, Pochury language]
-
A.
Poqomam language
The Poqomam language is a Mayan language spoken primarily in Guatemala by the Poqomam people, recognized as part of the country’s indigenous linguistic heritage.
-
B.
Chumburung language
The Chumburung language is a Niger-Congo language spoken primarily by the Chumburung people in Ghana.
-
C.
Pokomo language
The Pokomo language is a Bantu language spoken primarily by the Pokomo people along Kenya’s Tana River.
-
D.
Potohari language
The Potohari language is an Indo-Aryan language spoken primarily in Pakistan’s Potohar Plateau and surrounding regions, closely related to Punjabi and used in various local dialects.
-
E.
Puyuma language
The Puyuma language is an endangered Austronesian language spoken by the Puyuma Indigenous people of southeastern Taiwan.
- F. None of above. chosen
- G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg
Description generation
gpt-5.1
Instruction
Generate a one-sentence description of the target entity. You are given a context triple in the form (subject, predicate, object), where the object is the target entity. # Instructions Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. Avoid repeating the information from the triple, unless really essential. # Response Format Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Pochury language Triple: [Pochury Naga, language, Pochury language]
Generated description
Pochury language is a Sino-Tibetan language spoken by the Pochury Naga people of Nagaland in northeastern India.
NED2
Entity disambiguation (via description)
gpt-5-mini-2025-08-07
Target entity: Pochury language Target entity description: Pochury language is a Sino-Tibetan language spoken by the Pochury Naga people of Nagaland in northeastern India.
-
A.
Poqomam language
The Poqomam language is a Mayan language spoken primarily in Guatemala by the Poqomam people, recognized as part of the country’s indigenous linguistic heritage.
-
B.
Chumburung language
The Chumburung language is a Niger-Congo language spoken primarily by the Chumburung people in Ghana.
-
C.
Pokomo language
The Pokomo language is a Bantu language spoken primarily by the Pokomo people along Kenya’s Tana River.
-
D.
Potohari language
The Potohari language is an Indo-Aryan language spoken primarily in Pakistan’s Potohar Plateau and surrounding regions, closely related to Punjabi and used in various local dialects.
-
E.
Puyuma language
The Puyuma language is an endangered Austronesian language spoken by the Puyuma Indigenous people of southeastern Taiwan.
- F. None of above. chosen
Provenance (5 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69ca83d10b608190b2b2f8e0a7faaf14 |
completed | March 30, 2026, 2:08 p.m. |
| NER | Named-entity recognition | batch_69cc6a9f2c7481909b4a272183f20585 |
completed | April 1, 2026, 12:45 a.m. |
| NED1 | Entity disambiguation (via context triple) | batch_69cfdbc9c6e08190aa71d84316afc6d5 |
completed | April 3, 2026, 3:24 p.m. |
| NEDg | Description generation | batch_69cfdcb95b508190b9d5562f4248e074 |
completed | April 3, 2026, 3:28 p.m. |
| NED2 | Entity disambiguation (via description) | batch_69cfdd3293888190a9cd5e0621c2fe10 |
completed | April 3, 2026, 3:30 p.m. |
Created at: March 30, 2026, 7:08 p.m.