
RUNA-2 is a directional revision of RUNA-1, a typed biosemiotic knowledge-graph embedding (BoxE, PyKEEN) of European ecological interactions. What changed (v1.4 → v2.0): RUNA-1 was direction-blind on its most important relations — it scored reversed trophic triples higher than correct ones (overall directionality 0.53, barely above chance; preysOn 0.26, eats 0.45, pollinates 0.24). The root cause, verified from the shipped model, was training with create_inverse_triples=True, which symmetrizes the per-relation boxes and destroys direction. Retraining with create_inverse_triples=False restores directionality to 0.98 (preysOn 0.96, eats 0.99, pollinates 1.00) with no hard negatives and no architecture change, and with no regression on already-correct symmetric relations. Provenance: the retrain, the leak-free MRR benchmark, the per-relation directionality figures and the architecture sweep were produced by the model-training session (RTX 5090) and are documented in RUNA_v2_recommendation.md in the bundle; they are reported as-measured by that session. The directionality battery in RUNA_v2_verification_and_generation.md was independently run against the live v2.0 endpoint. Ranking is unchanged. On an honest leak-free split (split by unordered entity-pair), v1.4 scores MRR 0.114 and v2.0 scores 0.118 — the same ranker. v1.4's originally-reported ~0.30 MRR was measured on a leaky random split that both inflated MRR and hid the direction flaw. An architecture sweep (RotatE, PairRE) confirms the ~0.12 MRR is a data ceiling, not an architecture limit. Use: v2.0 fixes direction, not ranking — score_interaction is now directionally trustworthy, but RUNA remains a candidate seeder, not an arbiter. A confident forward>reverse margin means the model is sure which way the arrow points, not that the edge is real; candidates must still be gated by external evidence. This deposit includes the retrained model, the full retrain/diagnosis writeup, an independent verification battery against the live endpoint, and a directed candidate-generation report.
v1.4.0 adds the navigatesBy relation (senses->navigation: magnetic/sun/star/olfaction), ingests literature-confirmed magnetoreception facts, removes a 15-edge ash-dieback contamination cluster (Hymenoscyphus fraxineus now resolves ash correctly), and anchors sparse endemics (Saimaa seal, aspen bracket, stag beetle). 53,099 entities / 33 relations / 421,952 edges. Validation: MRR 0.301 held; temperature rho 0.576, pH 0.316 - within seed variance of v1.3 (same-data control under seed 43: 0.569 / 0.298). New VERSION of concept 10.5281/zenodo.20630499. Novelty is the relation schema + frozen triples + derivation/verification code, hashed in MANIFEST.sha256.
multi-model verification, knowledge graph embedding, Umwelt, electroreception, Ellenberg indicator values, EIVE, BoxE, PyKEEN, GLOBI, GBIF, agent-curated knowledge graph, biosemiotics, magnetoreception, animal navigation, RUNA-1, bioindicators, SoilGrids, ecological networks, animal communication, xeno-canto, link prediction
multi-model verification, knowledge graph embedding, Umwelt, electroreception, Ellenberg indicator values, EIVE, BoxE, PyKEEN, GLOBI, GBIF, agent-curated knowledge graph, biosemiotics, magnetoreception, animal navigation, RUNA-1, bioindicators, SoilGrids, ecological networks, animal communication, xeno-canto, link prediction
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
