arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:2609.00002v1 [cs.AI] 12 Jun 2026

HyperWorld: Hypergraph-Structured State Serialization
Improves Learned Textual World Models

Yun-Jian Zhang1 Chen-Wei Liang2 Tian-Yi Zhang3 Jian Ding1 Yi-Lun Wu1
Ao-Bo Li1 Wei-Cong Su1 Saifullah1 Hong-Yu An4 Mu-Jiang-Shan Wang1,5,*
1 Shenzhen Kaihong Digital Industry Development Co., Ltd.
{zhangyunjian,wuyilun,Liaobo}@kaihong.com
{suweicong,saifullah}@kaihong.com
; dj19@tsinghua.org.cn
2 School of Mathematics and Statistics, Faculty of Science, University of New South Wales,
Sydney, NSW 2052, Australia; z5537371@ad.unsw.edu.au
3 College of Computer Science and Technology, Zhejiang University,
38 Zheda Road, Hangzhou 310027, China; tianyizhang0213@zju.edu.cn
4 State Key Laboratory of Internet of Things for Smart City (SKL-IoTSC),
University of Macau, Macau; yc48116@connect.um.edu.mo
5 Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences,
Shenzhen, China; mjs.wang@siat.ac.cn
* Corresponding author: mjs.wang@siat.ac.cn
Abstract

World models, which predict how an environment evolves under actions, are increasingly used to equip language-model agents with the ability to plan before acting. In text environments, a world model must learn symbolic dynamics from serialized descriptions of the state, yet how the structure of this serialization affects learning remains largely unexamined: prior work serializes states either as flat text or as pairwise relational triples, both of which fragment the joint, higher-order structure of environment states. We present HyperWorld, a systematic study of state serialization structure for learning textual world models with small language models. We compare four information-equivalent representations of the same ground-truth symbolic state—raw observations, independent sentences, pairwise triples, and entity-centred hyperedge units that jointly describe multiple entities and their relations—under an identical effect-prediction objective: given a state and an action, predict the symbolic effects or judge the action infeasible. Across model scales (0.5B–3B) and data budgets, hyperedge serialization improves effect prediction most clearly at 0.5B–1.5B and on out-of-distribution (OOD) test worlds, where it retains more of its in-distribution accuracy than raw text and, in most settings, outperforms pairwise triples. At 3B, larger models narrow the gap—triples even match or slightly exceed hyperedge EM on IID data—but hyperedge grouping still achieves the strongest OOD fact F1. Raw observations support strong feasibility detection but weak effect prediction; hyperedge units recover near-raw feasibility while improving effects, yielding the best overall trade-off at small-to-medium scale. A greedy planner using the hyperedge world model also attains the highest success rate among the representations we test. These results suggest that higher-order state structure is a cheap inductive bias for learned symbolic world models, especially when capacity is limited or test worlds differ from training.

1 Introduction

World models—internal predictive models of environment dynamics—are a central ingredient of deliberative agents: they allow an agent to imagine the consequences of candidate actions before executing them, enabling lookahead planning, risk assessment, and sample-efficient policy learning [10, 11, 14]. Recent large-scale systems extend this paradigm to pixels and video [5, 1], while a parallel line of work equips language-model agents with textual world models that predict environment feedback or symbolic state changes in interactive text environments [2, 6, 9].

A textual world model must consume a serialized description of the current state. Existing approaches serialize states either as raw textual observations or as a set of pairwise relational triples [2, 3, 8]—the same pairwise decomposition used by knowledge graphs. However, environment states are intrinsically higher-order: the conditions and effects of an action typically involve several entities jointly. Unlocking a chest, for instance, simultaneously concerns the player’s location, the key in the inventory, the chest’s lock state, and the key–lock correspondence. Pairwise decomposition scatters this joint structure across disconnected triples, and flat text scatters it across sentences, leaving the model to re-assemble the relevant context from fragments at every prediction step.This concern is conceptually related to classical graph-theoretic studies of reachability and navigation, where orientation, Hamiltonicity, and path embedding provide structural ways to reason about feasible transitions and robust routes in discrete state spaces [24, 17, 21]. Here we ask an analogous question for textual world models: whether the serialization of a symbolic state exposes the joint transition structure needed for learning action effects. Hypergraphs, whose hyperedges connect arbitrary numbers of vertices, are a natural formalism for such joint structure, and have recently proven useful for retrieval-augmented generation [16] and knowledge representation. Whether higher-order structure helps a model learn dynamics is, to our knowledge, unexplored.

This paper asks a deliberately focused question:

Holding information content fixed, does hypergraph-structured state serialization improve a small language model’s ability to learn textual world dynamics?

To answer it we build HyperWorld, a controlled experimental framework on procedurally generated TextWorld games [7]. From ground-truth symbolic states we derive four serializations with identical information content but different structure: (i) raw observations, (ii) independent natural-language sentences, (iii) pairwise triples, and (iv) entity-centred hyperedge units that jointly render an entity’s location, internal state, contents, and key bindings as a single unit. A small language model (Qwen2.5, 0.5B–3B) [23] is LoRA-fine-tuned [12] as a world model that maps a serialized state and an action to symbolic effects (ADD/REMOVE facts) or an INFEASIBLE verdict. Because targets are identical across serializations, any performance difference is attributable to input structure alone.

Our contributions are threefold:

  • •

    A controlled testbed for studying state-serialization structure in learned textual world models, with information-equivalent flat, pairwise, and hypergraph renderings of ground-truth symbolic states, spanning IID and out-of-distribution splits.

  • •

    Empirical evidence that hyperedge serialization improves effect prediction and OOD robustness at 0.5B–1.5B, recovers much of raw text’s feasibility signal that flat fact renderings lose at 0.5B, and remains competitive at 3B where triples are strong on IID metrics.

  • •

    A planning demonstration showing that, under the same greedy search procedure, the hyperedge world model attains a higher live-game success rate than sentence- or triple-based world models.

2 Related Work

World models.

Learning predictive models of environment dynamics has a long history in model-based reinforcement learning [10, 11]. Recent foundation-scale world models generate interactive visual environments [5, 1], while JEPA-style architectures learn latent dynamics without pixel reconstruction [14, 4]. Our work concerns a complementary, low-compute regime: small language models learning symbolic dynamics of text environments.

Textual world models for LLM agents.

Equipping LLM agents with world models improves decision making by allowing lookahead before execution [6, 9, 15]. Worldformer [2] predicts knowledge-graph transitions in text games; AriGraph [3] maintains a memory graph for agent reasoning; Graph World Models [8] formalize world models over graph-structured states. StateFactory [18] factorizes observations into object–attribute hierarchies for reward prediction. All of these rely on flat text or pairwise relations; none examines higher-order serialization structure, which is our focus.

Graph-theoretic robustness and structural recoverability.

A related discrete perspective comes from graph-theoretic studies of network robustness and diagnosability. Diagnosability asks whether faulty components can be identified from comparison information, while edge connectivity measures the tolerance of a networked structure to link failures [19, 20]. Although these notions differ from learned state serialization, they reflect a common theme: useful representations should preserve enough structure to support recovery, robustness, and reliable reasoning under perturbations.

Hypergraphs for language systems.

Hypergraph representations capture nn-ary associations that pairwise graphs fragment, with demonstrated benefits in retrieval-augmented generation [16] and multi-hop question answering. We transfer this insight from knowledge retrieval to dynamics learning, and test whether hyperedge-grouped states make forward simulation easier to learn.

Refer to caption
Figure 1: Overview of HyperWorld. The same ground-truth symbolic state is rendered by four information-equivalent serializers; a LoRA-fine-tuned small LM learns to predict symbolic action effects (or infeasibility) from each rendering. Learned dynamics are evaluated by single-step and rollout metrics and by a downstream world-model-guided planner on live games.

3 Method

3.1 Problem Setup

We model a text environment as a partially observable MDP whose underlying state is a set of ground facts st={f1,…,fm}s_{t}=\{f_{1},\dots,f_{m}\}, where each fact f=p​(e1,…,ek)f=p(e_{1},\dots,e_{k}) applies a predicate pp to entities eie_{i} (e.g., in​(key,chest)\mathrm{in}(\text{key},\text{chest}), locked​(chest)\mathrm{locked}(\text{chest})). Executing action ata_{t} yields the successor state st+1=(st∖Rt)∪Ats_{t+1}=(s_{t}\setminus R_{t})\cup A_{t}, where AtA_{t} and RtR_{t} are the added and removed fact sets. Actions outside the admissible set leave the state unchanged.

A textual world model is a function

fθ​(σ​(st),at)⟶{(At,Rt)if ​at​ is admissible in ​st,INFEASIBLEotherwise,f_{\theta}\bigl(\sigma(s_{t}),\,a_{t}\bigr)\;\longrightarrow\;\begin{cases}(A_{t},\,R_{t})&\text{if }a_{t}\text{ is admissible in }s_{t},\\ \texttt{INFEASIBLE}&\text{otherwise,}\end{cases} (1)

where σ​(⋅)\sigma(\cdot) is a serializer that renders the symbolic state as text. The output vocabulary (canonical fact strings) is fixed across serializers, so the choice of σ\sigma is the only manipulated variable.

3.2 Information-Equivalent State Serializations

Given the same fact set ss, we compare four serializers (Figure 2):

Sentences (σsent\sigma_{\mathrm{sent}}) Triples (σtri\sigma_{\mathrm{tri}})
player is at kitchen.
rusty key is in inventory.
chest is at kitchen.
chest is closed.
chest is locked.
gold coin is in chest.
rusty key matches chest.
pantry is west of kitchen.
(player, at, kitchen)
(rusty key, in, inventory)
(chest, at, kitchen)
(chest, is, closed)
(chest, is, locked)
(gold coin, in, chest)
(rusty key, match, chest)
(pantry, west_of, kitchen)
Hyper (σhyp\sigma_{\mathrm{hyp}}, ours) Target (identical for all)
[player | at: kitchen | holds: rusty key]
[chest | at: kitchen | state: closed, locked | contains: gold coin | key: rusty key]
[kitchen | contains: chest | exits: west -> pantry]
ACTION: unlock chest with rusty key
EFFECTS: ADD: none |
 REMOVE: locked(chest)
Figure 2: Information-equivalent serializations of the same symbolic state. The hyperedge form groups all facts relevant to an entity into a single nn-ary unit, so the joint precondition of an action (key held, chest locked, key–lock match, co-location) appears in one place rather than scattered across lines.

Raw (σraw\sigma_{\mathrm{raw}}): the textual observation produced by the environment (room description and inventory). This is what a purely text-based world model sees.

Sentences (σsent\sigma_{\mathrm{sent}}): each fact rendered as an independent natural-language sentence (“The key is in the chest.”). A flat control condition that carries exactly the facts, with no relational structure.

Triples (σtri\sigma_{\mathrm{tri}}): each fact rendered as a binary triple (head,relation,tail)(\text{head},\text{relation},\text{tail}); unary predicates become (entity,is,predicate)(\text{entity},\texttt{is},\text{predicate}). This is the pairwise-graph representation used by graph-based world models [2, 3, 8].

Hyper (σhyp\sigma_{\mathrm{hyp}}, ours): facts are grouped into entity-centred nn-ary units by a deterministic procedure. Each unit is a hyperedge over several entities, rendered on one line:

[chest | at: kitchen | state: closed, locked | contains: coin | key: rusty key]

The grouping covers the player (location and inventory), every object (location, unary states, contents, key bindings), and room connectivity (exits); a catch-all unit preserves any remaining facts, keeping the rendering lossless.

All fact-based serializers (σsent,σtri,σhyp\sigma_{\mathrm{sent}},\sigma_{\mathrm{tri}},\sigma_{\mathrm{hyp}}) are bijective renderings of the same fact set: they contain identical information and differ only in how facts are grouped on the page. This isolates structure as the experimental variable.

3.3 World-Model Learning

Transitions are collected from procedurally generated games by mixing goal-directed walkthrough trajectories with ϵ\epsilon-random branches, which yields diverse on-path and off-path dynamics. Infeasible actions are sampled from syntactically valid commands that are not admissible in the current state, providing negatives for feasibility learning.

The world model is a pretrained decoder-only LM fine-tuned with LoRA on the unified objective of Eq. (1), serialized as

STATE: σ​(st)\sigma(s_{t}) ACTION: ata_{t} EFFECTS: ADD: … | REMOVE: …

with cross-entropy on the target tokens only. Identical targets, hyperparameters, and data across serializers guarantee a controlled comparison.

3.4 World-Model-Guided Planning

To test whether better dynamics translate into better behaviour, we use the learned world model inside a greedy planner on live, unseen games. At each step, for every admissible command aa, the planner queries the world model for imagined effects (A^,R^)(\hat{A},\hat{R}) and scores

score​(a)=wg​Δgoal​(a)+wf​ 1​[feasible^]+wn​ 1​[novel imagined state]+wc​c​(a),\mathrm{score}(a)=w_{g}\,\Delta_{\mathrm{goal}}(a)+w_{f}\,\mathbb{1}[\hat{\text{feasible}}]+w_{n}\,\mathbb{1}[\text{novel imagined state}]+w_{c}\,c(a), (2)

where Δgoal\Delta_{\mathrm{goal}} is the imagined gain in satisfied goal facts, and c​(a)c(a) is the model’s mean token log-probability, a confidence signal that down-weights unreliable imaginations. The action with the highest score is executed. The planner is deliberately simple: it isolates the quality of the learned dynamics rather than the sophistication of search.

4 Experiments

4.1 Setup

Environments and data.

We generate 410 TextWorld games in four splits: train (300 games; 4 rooms, 8 objects, quest length 3), validation (30), test-IID (40 unseen games with training parameters), and test-OOD (40 games with 8 rooms, 16 objects, quest length 6). Rolling out one walkthrough trajectory and two ϵ\epsilon-random-branching trajectories per game (ϵ=0.35\epsilon{=}0.35, max 25 steps) yields 8,375 train / 797 validation / 1,058 test-IID / 2,001 test-OOD transitions, roughly one third of which are infeasible-action negatives.

Models and training.

Qwen2.5-Instruct at 0.5B, 1.5B, and 3B with LoRA (r=16r{=}16), 2 epochs, learning rate 2×10−42\times 10^{-4}, identical across conditions. All experiments run on a single RTX 5090.

Metrics.

(i) Feasibility accuracy/F1: detecting inadmissible actions; (ii) Effect exact match: predicted delta exactly equals the gold delta; (iii) Fact F1: precision/recall over predicted added/removed facts; (iv) Rollout state F1: F1 between the simulated and true fact sets after kk imagined steps; (v) Planning success rate on live games.

4.2 Main Results

Table 1 reports effect exact match (EM) and fact-level F1 on test-IID and test-OOD splits across three model scales (bold = best per column). hyper is strongest overall at 0.5B and on most OOD columns, but it is not uniformly dominant: at 3B IID, triples reach the highest EM (0.956 vs. 0.952 for hyper), and sentences and triples tie hyper on IID fact F1 (0.980 vs. 0.984). The clearest advantage appears at 1.5B on OOD data, where hyper reaches EM 0.914 and fact F1 0.939—7.6 and 6.6 points above triples and well above raw text (EM 0.603). Under distribution shift, degradation also differs: at 1.5B, raw falls from 0.715 to 0.603 EM (IID→\toOOD), whereas hyper falls only from 0.936 to 0.914; sentences actually exceeds triples on OOD EM at this scale (0.850 vs. 0.838), but both remain below hyper.

A complementary pattern emerges for feasibility detection (Table 2). Raw observations achieve the highest feasibility accuracy (0.965 IID / 0.944 OOD at 1.5B) because environment text mentions visible objects directly. Independent sentences and triples degrade feasibility sharply at 0.5B (∼\sim0.72 accuracy) by scattering joint preconditions across lines. hyper recovers much of this signal (0.944 / 0.913 at 1.5B) while predicting effects more accurately than fact-based flat or pairwise renderings in the same setting—a favorable trade-off, though not the best on feasibility alone.

At 3B, all fact-based serializers improve substantially. hyper achieves the best OOD fact F1 (0.975) and OOD EM (0.939), but IID EM favors triples. This suggests that higher-capacity models partially compensate for fragmented structure, narrowing—without eliminating—the benefit of hyperedge grouping.

Table 1: World-model prediction on test-IID and test-OOD (effect exact match / fact F1). Bold: best value within each split and model scale.
0.5B 1.5B 3B
Repr EM F1 EM F1 EM F1
test-IID
Raw 0.710 0.853 0.715 0.861 0.726 0.871
Sentences 0.865 0.866 0.848 0.877 0.946 0.980
Triples 0.889 0.892 0.856 0.878 0.956 0.980
Hyper 0.943 0.963 0.936 0.956 0.952 0.984
test-OOD
Raw 0.617 0.814 0.603 0.805 0.621 0.823
Sentences 0.784 0.811 0.850 0.873 0.913 0.952
Triples 0.790 0.818 0.838 0.873 0.927 0.961
Hyper 0.909 0.943 0.914 0.939 0.939 0.975
Table 2: Feasibility detection accuracy (1.5B). Raw is best; hyper is second and substantially above sentences/triples.
Repr test-IID test-OOD
Raw 0.965 0.944
Sentences 0.923 0.874
Triples 0.922 0.887
Hyper 0.944 0.913

Three random seeds at 1.5B confirm stability: OOD EM for hyper is 0.914±0.0040.914\pm 0.004 (mean±\pmstd), versus 0.826±0.0730.826\pm 0.073 for sentences and 0.805±0.0380.805\pm 0.038 for triples.

4.3 Sample Efficiency

Figure 3 varies the training fraction at 1.5B. With only 10% of training data, all representations perform similarly on OOD EM (∼\sim0.63–0.64), so structure alone is insufficient in the extreme low-data regime. With 25% data, hyper reaches OOD EM 0.824, clearly above triples (0.703) and sentences (0.709); on IID EM at the same fraction, however, sentences (0.885) still exceed hyper (0.836). Hyperedge grouping therefore helps most on harder OOD generalization once a modest amount of training data is available, but does not dominate every metric at every budget.

Refer to caption
Figure 3: OOD fact F1 vs. training data fraction (1.5B). hyper leads on OOD from 25% data upward; at 10% all methods are close.

4.4 Multi-Step Rollouts

We apply predicted deltas iteratively along recorded trajectories and measure state-set F1 at horizons 1–5. Figure 4 shows modest but consistent rollout gains for hyper at 1.5B: at horizon 5 on test-OOD, state F1 is 0.989 for hyper versus 0.980 for sentences and 0.983 for triples. Better single-step prediction therefore carries over, albeit with small margins, to multi-step imagination.

Refer to caption
Figure 4: Multi-step rollout state F1 vs. horizon (1.5B). Left: test-IID; right: test-OOD.

4.5 World-Model-Guided Planning

Table 3 evaluates the learned world models inside a greedy planner on 30 held-out test-IID games. With the default scoring rule, hyper attains 76.7% success, clearly above random (23.3%), sentences (53.3%), and triples (56.7%), and uses fewer steps on average (14.3 vs. 22.9–34.3). Better learned dynamics therefore transfer to downstream task completion under matched search.

Removing the confidence term wc​c​(a)w_{c}\,c(a) in Eq. (2) raises hyper success further to 93.3% (9.9 steps), indicating that the default confidence weight is miscalibrated rather than helpful. We report both settings for transparency and leave calibrated confidence–planning integration to future work.

Table 3: World-model-guided planning on 30 test-IID games (1.5B world models).
Planner WM repr Success Avg. steps
Random — 23.3% 34.3
WM-guided sentences 53.3% 22.9
WM-guided triples 56.7% 20.5
WM-guided hyper 76.7% 14.3
WM-guided (no conf.) hyper 93.3% 9.9

4.6 Case Study

Consider a test game requiring: take rusty key, unlock chest with rusty key, take gold coin. In the initial state the planner must choose among admissible commands including open chest (infeasible while locked) and go east (irrelevant).

The raw world model often predicts feasible effects for open chest because the observation mentions the chest, yielding misleading positive imaginations. The triples model represents locked(chest) and match(rusty key, chest) on separate lines; it may predict partial unlock effects without jointly satisfying all preconditions. The hyper model renders [chest | state: closed, locked | key: rusty key] as one unit and correctly predicts INFEASIBLE for open chest and ADD: none | REMOVE: locked(chest) for unlock chest with rusty key, guiding the planner along the shortest successful path in 10 steps rather than exhausting the 40-step budget.

5 Conclusion

We presented HyperWorld, a controlled study of how state-serialization structure affects learned textual world models. Hyperedge grouping is not uniformly optimal—at 3B IID, triples match or exceed it on EM—but it offers the strongest overall profile at 0.5B–1.5B, the best OOD generalization in most settings, near-raw feasibility at 1.5B, and the highest planning success among matched world models (76.7% vs. 53–57% for flat/pairwise alternatives). On TextWorld, this yields up to 31 percentage-point OOD EM gains over raw text at 1.5B. More broadly, the result connects state-serialization design with structured prediction in non-stationary dynamical systems, where recent spatio-temporal graph attention models use graph structure to capture evolving dependencies among interacting variables [22]. It also suggests a possible bridge to long-horizon vision–language–action manipulation, in which symmetry-aware decision-making and structured state abstraction are important for planning over extended action sequences [13]. Future work includes learned hyperedge grouping, calibrated confidence for planning, and extension to richer environments such as ALFWorld, ScienceWorld, and embodied VLA manipulation tasks.

References

  • [1] N. Agarwal, A. Ali, M. Bala, et al. (2025) Cosmos world foundation model platform for physical AI. arXiv preprint arXiv:2501.03575. Cited by: §1, §2.
  • [2] P. Ammanabrolu and M. Riedl (2021) Learning knowledge graph-based world models of textual environments. In Advances in Neural Information Processing Systems, Vol. 34. Cited by: §1, §1, §2, §3.2.
  • [3] P. Anokhin, N. Semenov, A. Sorokin, D. Evseev, A. Kravchenko, M. Burtsev, and E. Burnaev (2024) AriGraph: learning knowledge graph world models with episodic memory for LLM agents. arXiv preprint arXiv:2407.04363. Note: Published at IJCAI 2025 Cited by: §1, §2, §3.2.
  • [4] M. Assran, A. Bardes, D. Fan, Q. Garrido, R. Howes, M. Komeili, M. Muckley, A. Rizvi, C. Roberts, K. Sinha, A. Zholus, et al. (2025) V-JEPA 2: self-supervised video models enable understanding, prediction and planning. arXiv preprint arXiv:2506.09985. Cited by: §2.
  • [5] J. Bruce, M. Dennis, A. Edwards, et al. (2024) Genie: generative interactive environments. In International Conference on Machine Learning, Cited by: §1, §2.
  • [6] H. Chae, N. Kim, K. T. Ong, M. Gwak, G. Song, J. Kim, S. Kim, D. Lee, and J. Yeo (2025) Web agents with world models: learning and leveraging environment dynamics in web navigation. In International Conference on Learning Representations, Cited by: §1, §2.
  • [7] M. Côté, Á. Kádár, X. Yuan, et al. (2018) TextWorld: a learning environment for text-based games. In Workshop on Computer Games, pp. 41–75. Cited by: §1.
  • [8] T. Feng, Y. Wu, G. Lin, and J. You (2025) Graph world model. arXiv preprint arXiv:2507.10539. Cited by: §1, §2, §3.2.
  • [9] Y. Gu, B. Zheng, B. Gou, K. Zhang, C. Chang, S. Srivastava, Y. Xie, P. Qi, H. Sun, and Y. Su (2024) Is your LLM secretly a world model of the internet? model-based planning for web agents. arXiv preprint arXiv:2411.06559. Cited by: §1, §2.
  • [10] D. Ha and J. Schmidhuber (2018) World models. arXiv preprint arXiv:1803.10122. Cited by: §1, §2.
  • [11] D. Hafner, J. Pasukonis, J. Ba, and T. Lillicrap (2023) Mastering diverse domains through world models. arXiv preprint arXiv:2301.04104. Cited by: §1, §2.
  • [12] E. J. Hu, Y. Shen, P. Wallis, et al. (2022) LoRA: low-rank adaptation of large language models. In International Conference on Learning Representations, Cited by: §1.
  • [13] Y. Jian, D. Tian, X. Chen, Z. Wei, C. Liang, and M. Wang (2026) PI-vla: adaptive symmetry-aware decision-making for long-horizon vision–language–action manipulation. Symmetry 18 (3), pp. 394. Cited by: §5.
  • [14] Y. LeCun (2022) A path towards autonomous machine intelligence. Note: OpenReview External Links: Link Cited by: §1, §2.
  • [15] Y. Liu, J. Wang, H. Wang, B. Guo, and W. Li (2026) Imagine-then-plan: agent learning from adaptive lookahead with world models. arXiv preprint arXiv:2601.08955. Cited by: §2.
  • [16] H. Luo, H. E, G. Chen, Y. Zheng, X. Wu, Y. Guo, Q. Lin, Y. Feng, Z. Kuang, M. Song, Y. Zhu, and A. T. Luu (2025) HyperGraphRAG: retrieval-augmented generation via hypergraph-structured knowledge representation. In Advances in Neural Information Processing Systems, Vol. 39. Cited by: §1, §2.
  • [17] W. Mu-Jiang-shan, Y. Jun, L. Shang-wei, et al. (2010) Ordered and hamilton digraphs. Chinese Quarterly Journal of Mathematics 25 (3), pp. 317–326. Cited by: §1.
  • [18] Y. Shen, D. Chen, X. Hu, J. Mi, H. Zhao, K. Zhang, and P. Fung (2026) Reward prediction with factorized world states. arXiv preprint arXiv:2603.09400. Cited by: §2.
  • [19] M. Wang and S. Wang (2016) Diagnosability of cayley graph networks generated by transposition trees under the comparison diagnosis model. Annals of Applied Mathematics 32 (2), pp. 166–173. Cited by: §2.
  • [20] S. Wang and M. Wang (2018) The edge connectivity of expanded k-ary n-cubes. Discrete Dynamics in Nature and Society 2018 (1), pp. 7867342. Cited by: §2.
  • [21] S. Wang, J. Wangmu, Z. Qi, and Y. Ren (2011) Embedding paths into the 4-ary n-cube with faulty nodes. In 2011 International Conference on Consumer Electronics, Communications and Networks (CECNet), pp. 4949–4951. Cited by: §1.
  • [22] Z. Wei, H. An, Y. Yao, W. Su, G. Li, Saifullah, B. Sun, and M. Wang (2025) FSTGAT: financial spatio-temporal graph attention network for non-stationary financial systems and its application in stock price prediction. Symmetry 17 (8), pp. 1344. Cited by: §5.
  • [23] A. Yang, B. Yang, B. Zhang, et al. (2024) Qwen2.5 technical report. arXiv preprint arXiv:2412.15115. Cited by: §1.
  • [24] L. Zhao, M. Wang, X. Zhang, Y. Lin, and S. Wang (2017) An algorithm for the orientation of complete bipartite graphs. In 2017 International Conference on Applied Mathematics, Modelling and Statistics Application (AMMSA 2017), pp. 361–364. Cited by: §1.