Triple

T7723429
Position Surface form Disambiguated ID Type / Status
Subject Normes ortogràfiques E175068 entity
Predicate partOf P40 FINISHED
Object IEC normative corpus for Catalan
The IEC normative corpus for Catalan is the official set of linguistic standards issued by the Institut d’Estudis Catalans that codifies the language’s spelling, grammar, and usage.
E684436 NE FINISHED

How this triple was built (4 steps)

Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.

NER Named-entity recognition gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: IEC normative corpus for Catalan | Statement: [Normes ortogràfiques, partOf, IEC normative corpus for Catalan]
NED1 Entity disambiguation (via context triple) gpt-5-mini-2025-08-07
Target entity: IEC normative corpus for Catalan
Context triple: [Normes ortogràfiques, partOf, IEC normative corpus for Catalan]
  • A. Corpus del Español del Siglo XXI
    Corpus del Español del Siglo XXI is a large, modern reference corpus of contemporary Spanish designed to represent current usage across different regions, genres, and registers.
  • B. Corpus de Referencia del Español Actual
    Corpus de Referencia del Español Actual is a large, balanced reference corpus of contemporary Spanish used for linguistic research and lexicographic work.
  • C. CORDE corpus
    The CORDE corpus is a large historical Spanish language corpus compiled by the Royal Spanish Academy, used for studying the evolution and usage of Spanish over time.
  • D. Gramàtica de la llengua catalana
    Gramàtica de la llengua catalana is the authoritative modern reference grammar that codifies the standard norms of the Catalan language as established by the Institut d’Estudis Catalans.
  • E. Corpus
    Corpus is a common shortened name for Corpus Christi College, one of the historic constituent colleges of the University of Cambridge.
  • F. None of above. chosen
  • G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg Description generation gpt-5.1
Instruction
Generate a one-sentence description of the target entity. 
You are given a context triple in the form (subject, predicate, object), where the object is the target entity. 
# Instructions
Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. 
Avoid repeating the information from the triple, unless really essential.
# Response Format
Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: IEC normative corpus for Catalan
Triple: [Normes ortogràfiques, partOf, IEC normative corpus for Catalan]
Generated description
The IEC normative corpus for Catalan is the official set of linguistic standards issued by the Institut d’Estudis Catalans that codifies the language’s spelling, grammar, and usage.
NED2 Entity disambiguation (via description) gpt-5-mini-2025-08-07
Target entity: IEC normative corpus for Catalan
Target entity description: The IEC normative corpus for Catalan is the official set of linguistic standards issued by the Institut d’Estudis Catalans that codifies the language’s spelling, grammar, and usage.
  • A. Corpus del Español del Siglo XXI
    Corpus del Español del Siglo XXI is a large, modern reference corpus of contemporary Spanish designed to represent current usage across different regions, genres, and registers.
  • B. Corpus de Referencia del Español Actual
    Corpus de Referencia del Español Actual is a large, balanced reference corpus of contemporary Spanish used for linguistic research and lexicographic work.
  • C. CORDE corpus
    The CORDE corpus is a large historical Spanish language corpus compiled by the Royal Spanish Academy, used for studying the evolution and usage of Spanish over time.
  • D. Gramàtica de la llengua catalana
    Gramàtica de la llengua catalana is the authoritative modern reference grammar that codifies the standard norms of the Catalan language as established by the Institut d’Estudis Catalans.
  • E. Corpus
    Corpus is a common shortened name for Corpus Christi College, one of the historic constituent colleges of the University of Cambridge.
  • F. None of above. chosen

Provenance (5 batches)

The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.

Step Stage Batch ID Status When
creating Elicitation batch_69c6995d541c81909eaa646b1a8369a9 completed March 27, 2026, 2:51 p.m.
NER Named-entity recognition batch_69c702f39fa48190b7b8a09446b5cf78 completed March 27, 2026, 10:21 p.m.
NED1 Entity disambiguation (via context triple) batch_69c8b51faa348190b4fa0b5a307c83db completed March 29, 2026, 5:14 a.m.
NEDg Description generation batch_69c8b74ee6d081908454b2d4774a3a7b completed March 29, 2026, 5:23 a.m.
NED2 Entity disambiguation (via description) batch_69c8b7af4c58819097360e89e7ea6062 completed March 29, 2026, 5:25 a.m.
Created at: March 27, 2026, 4:05 p.m.