Triple

T21653437
Position Surface form Disambiguated ID Type / Status
Subject Yaeyama Ryukyuan E534396 entity
Predicate hasDialect P4251 FINISHED
Object Hatoma dialect
The Hatoma dialect is a regional variety of the Yaeyama Ryukyuan language traditionally spoken on Hatoma Island in Okinawa, Japan.
E1495470 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: Hatoma dialect | Statement: [Yaeyama Ryukyuan, hasDialect, Hatoma dialect]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: Hatoma dialect
Context triple: [Yaeyama Ryukyuan, hasDialect, Hatoma dialect]
  • A. Fang-Ntumu dialect
    The Fang-Ntumu dialect is a regional variety of the Beti-Fang language continuum spoken primarily by Fang communities in Central Africa.
  • B. Maututu dialect
    The Maututu dialect is a regional variety of the Nakanai language spoken in parts of New Britain, Papua New Guinea.
  • C. Oporoma dialect
    The Oporoma dialect is a regional variety of the Izon (Ijaw) language spoken by communities in and around Oporoma in Nigeria’s Niger Delta.
  • D. Noatia dialect
    The Noatia dialect is a regional variety of the Kokborok language spoken primarily by the Noatia community in Tripura, India.
  • E. Harauti dialect
    The Harauti dialect is an Indo-Aryan variety spoken primarily in the Hadoti (Harauti) region of Rajasthan and neighboring areas of India.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: Hatoma dialect
Triple: [Yaeyama Ryukyuan, hasDialect, Hatoma dialect]
Generated description
The Hatoma dialect is a regional variety of the Yaeyama Ryukyuan language traditionally spoken on Hatoma Island in Okinawa, Japan.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: Hatoma dialect
Target entity description: The Hatoma dialect is a regional variety of the Yaeyama Ryukyuan language traditionally spoken on Hatoma Island in Okinawa, Japan.
  • A. Fang-Ntumu dialect
    The Fang-Ntumu dialect is a regional variety of the Beti-Fang language continuum spoken primarily by Fang communities in Central Africa.
  • B. Maututu dialect
    The Maututu dialect is a regional variety of the Nakanai language spoken in parts of New Britain, Papua New Guinea.
  • C. Oporoma dialect
    The Oporoma dialect is a regional variety of the Izon (Ijaw) language spoken by communities in and around Oporoma in Nigeria’s Niger Delta.
  • D. Noatia dialect
    The Noatia dialect is a regional variety of the Kokborok language spoken primarily by the Noatia community in Tripura, India.
  • E. Harauti dialect
    The Harauti dialect is an Indo-Aryan variety spoken primarily in the Hadoti (Harauti) region of Rajasthan and neighboring areas of India.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69e0c466aec88190ba39c7543dbc8ba2 completed April 16, 2026, 11:13 a.m.
NER Named-entity recognition batch_69ef59164fe081908abd2e33dcd67def completed April 27, 2026, 12:39 p.m.
NED1 Entity disambiguation (via context triple) batch_6a0a15960e348190b944ee34d80cd8a8 completed May 17, 2026, 7:23 p.m.
NEDg Description generation batch_6a0a16c278188190855335930666cede completed May 17, 2026, 7:28 p.m.
NED2 Entity disambiguation (via description) batch_6a0a17a2c81481908e229ea58e61843e completed May 17, 2026, 7:31 p.m.
Created at: April 16, 2026, 6:36 p.m.