Triple

T18591448
Position Surface form Disambiguated ID Type / Status
Subject ICL 1900 series E454376 entity
Predicate characterEncoding P7661 FINISHED
Object ICL 1900 6-bit character set
The ICL 1900 6-bit character set is a compact, machine-specific encoding scheme used on ICL 1900 series computers to represent a limited repertoire of characters within 6-bit words.
E1332101 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: ICL 1900 6-bit character set | Statement: [ICL 1900 series, characterEncoding, ICL 1900 6-bit character set]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: ICL 1900 6-bit character set
Context triple: [ICL 1900 series, characterEncoding, ICL 1900 6-bit character set]
  • A. ISO 646
    ISO 646 is an international standard for 7-bit character encodings that defines a set of basic Latin characters and allows national variants, serving as a foundation for many early computer character sets.
  • B. ISO/IEC 8859
    ISO/IEC 8859 is a family of 8-bit character encoding standards that define various single-byte coded character sets for different languages and scripts, widely used before the adoption of Unicode.
  • C. ISO/IEC 2022
    ISO/IEC 2022 is an international standard that defines mechanisms for encoding and switching between multiple character sets within a single byte-oriented data stream.
  • D. Unicode 1.0
    Unicode 1.0 was the first published version of the Unicode standard, establishing a unified character encoding system that laid the foundation for modern multilingual text representation in computing.
  • E. Unicode Standard code charts
    Unicode Standard code charts are official visual reference tables published by the Unicode Consortium that display every encoded character, its code point, and related annotations for each script and symbol block in the Unicode Standard.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: ICL 1900 6-bit character set
Triple: [ICL 1900 series, characterEncoding, ICL 1900 6-bit character set]
Generated description
The ICL 1900 6-bit character set is a compact, machine-specific encoding scheme used on ICL 1900 series computers to represent a limited repertoire of characters within 6-bit words.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: ICL 1900 6-bit character set
Target entity description: The ICL 1900 6-bit character set is a compact, machine-specific encoding scheme used on ICL 1900 series computers to represent a limited repertoire of characters within 6-bit words.
  • A. ISO 646
    ISO 646 is an international standard for 7-bit character encodings that defines a set of basic Latin characters and allows national variants, serving as a foundation for many early computer character sets.
  • B. ISO/IEC 8859
    ISO/IEC 8859 is a family of 8-bit character encoding standards that define various single-byte coded character sets for different languages and scripts, widely used before the adoption of Unicode.
  • C. ISO/IEC 2022
    ISO/IEC 2022 is an international standard that defines mechanisms for encoding and switching between multiple character sets within a single byte-oriented data stream.
  • D. Unicode 1.0
    Unicode 1.0 was the first published version of the Unicode standard, establishing a unified character encoding system that laid the foundation for modern multilingual text representation in computing.
  • E. Unicode Standard code charts
    Unicode Standard code charts are official visual reference tables published by the Unicode Consortium that display every encoded character, its code point, and related annotations for each script and symbol block in the Unicode Standard.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69d8d38ae7e081908a98df1251842402 completed April 10, 2026, 10:40 a.m.
NER Named-entity recognition batch_69e545b5bc688190bdfe3911ac6b2d76 completed April 19, 2026, 9:14 p.m.
NED1 Entity disambiguation (via context triple) batch_6a04f898b59481909aa72291b24a4cd0 completed May 13, 2026, 10:18 p.m.
NEDg Description generation batch_6a04fc705b3481908bea3e5f6ce33f1e completed May 13, 2026, 10:34 p.m.
NED2 Entity disambiguation (via description) batch_6a04fcc66ddc819094d427390e7c3774 completed May 13, 2026, 10:35 p.m.
Created at: April 10, 2026, 11:44 a.m.