Triple
T7723429
| Position | Surface form | Disambiguated ID | Type / Status |
|---|---|---|---|
| Subject | Normes ortogràfiques |
E175068
|
entity |
| Predicate | partOf |
P40
|
FINISHED |
| Object |
IEC normative corpus for Catalan
The IEC normative corpus for Catalan is the official set of linguistic standards issued by the Institut d’Estudis Catalans that codifies the language’s spelling, grammar, and usage.
|
E684436
|
NE FINISHED |
How this triple was built (4 steps)
Every LLM step that produced this triple, in pipeline order — named-entity classification, the disambiguation choices (the exact options shown, with the pick highlighted), and the generated description. The batch + timestamp of each is in the Provenance table below.
NER
Named-entity recognition
gpt-5-mini
Instruction
Given a phrase, classify it is english named entity (e.g., persons, organizations, works of art) in Latin script, or not (e.g., literals, dates, URLs, verbose phrases). For disambiguation, the statement where the phrase occurs as object is also given. Please return a JSON object with `phrase` (string, the phrase being analyzed) and `is_ne` (boolean, indicating whether the phrase is a Named Entity).
Input
Phrase: IEC normative corpus for Catalan | Statement: [Normes ortogràfiques, partOf, IEC normative corpus for Catalan]
NED1
Entity disambiguation (via context triple)
gpt-5-mini-2025-08-07
Target entity: IEC normative corpus for Catalan Context triple: [Normes ortogràfiques, partOf, IEC normative corpus for Catalan]
-
A.
Corpus del Español del Siglo XXI
Corpus del Español del Siglo XXI is a large, modern reference corpus of contemporary Spanish designed to represent current usage across different regions, genres, and registers.
-
B.
Corpus de Referencia del Español Actual
Corpus de Referencia del Español Actual is a large, balanced reference corpus of contemporary Spanish used for linguistic research and lexicographic work.
-
C.
CORDE corpus
The CORDE corpus is a large historical Spanish language corpus compiled by the Royal Spanish Academy, used for studying the evolution and usage of Spanish over time.
-
D.
Gramàtica de la llengua catalana
Gramàtica de la llengua catalana is the authoritative modern reference grammar that codifies the standard norms of the Catalan language as established by the Institut d’Estudis Catalans.
-
E.
Corpus
Corpus is a common shortened name for Corpus Christi College, one of the historic constituent colleges of the University of Cambridge.
- F. None of above. chosen
- G. Unsure - the case is ambiguous/there is not enough information to decide.
NEDg
Description generation
gpt-5.1
Instruction
Generate a one-sentence description of the target entity. You are given a context triple in the form (subject, predicate, object), where the object is the target entity. # Instructions Use the triple to infer relevant information about the entity. Describe the entity based on what is most defining, well-known. Avoid repeating the information from the triple, unless really essential. # Response Format Return only the sentence: "Description: [one-sentence description of the target entity]"
Input
Entity: IEC normative corpus for Catalan Triple: [Normes ortogràfiques, partOf, IEC normative corpus for Catalan]
Generated description
The IEC normative corpus for Catalan is the official set of linguistic standards issued by the Institut d’Estudis Catalans that codifies the language’s spelling, grammar, and usage.
NED2
Entity disambiguation (via description)
gpt-5-mini-2025-08-07
Target entity: IEC normative corpus for Catalan Target entity description: The IEC normative corpus for Catalan is the official set of linguistic standards issued by the Institut d’Estudis Catalans that codifies the language’s spelling, grammar, and usage.
-
A.
Corpus del Español del Siglo XXI
Corpus del Español del Siglo XXI is a large, modern reference corpus of contemporary Spanish designed to represent current usage across different regions, genres, and registers.
-
B.
Corpus de Referencia del Español Actual
Corpus de Referencia del Español Actual is a large, balanced reference corpus of contemporary Spanish used for linguistic research and lexicographic work.
-
C.
CORDE corpus
The CORDE corpus is a large historical Spanish language corpus compiled by the Royal Spanish Academy, used for studying the evolution and usage of Spanish over time.
-
D.
Gramàtica de la llengua catalana
Gramàtica de la llengua catalana is the authoritative modern reference grammar that codifies the standard norms of the Catalan language as established by the Institut d’Estudis Catalans.
-
E.
Corpus
Corpus is a common shortened name for Corpus Christi College, one of the historic constituent colleges of the University of Cambridge.
- F. None of above. chosen
Provenance (5 batches)
The batch behind each pipeline step, in order, with when it ran. Timestamps are batch-level — stages were processed in waves, so the object chain (NER → NED1 → NEDg → NED2) reads in order, but predicate / elicitation batches can sit in a different wave.
| Step | Stage | Batch ID | Status | When |
|---|---|---|---|---|
| creating | Elicitation | batch_69c6995d541c81909eaa646b1a8369a9 |
completed | March 27, 2026, 2:51 p.m. |
| NER | Named-entity recognition | batch_69c702f39fa48190b7b8a09446b5cf78 |
completed | March 27, 2026, 10:21 p.m. |
| NED1 | Entity disambiguation (via context triple) | batch_69c8b51faa348190b4fa0b5a307c83db |
completed | March 29, 2026, 5:14 a.m. |
| NEDg | Description generation | batch_69c8b74ee6d081908454b2d4774a3a7b |
completed | March 29, 2026, 5:23 a.m. |
| NED2 | Entity disambiguation (via description) | batch_69c8b7af4c58819097360e89e7ea6062 |
completed | March 29, 2026, 5:25 a.m. |
Created at: March 27, 2026, 4:05 p.m.