This sounds like one agent at each junction point, that may or may not be running the same model and architecture as the previous one, and they need to constantly crawl each other or just a protocol to call each other. But then you have a million of then constructing a docuverse web all having an agent who all need to know each other and its all connections and each a database repository.
Well you only need links that exist, and those that are created.
In a paragraph it may exist in one node that is referenced and that node may be a collection of other nodes, or each sentence or even each word points to other nodes, how much does each model know and what does it keep in its database. A point to the next node is not necessarily durable or versioned, and that goes for each node is connects to.
A single paragraph could have thousands or more sources, references, and each source would need versioning and depending on version it may well point to different nodes in prior or future versions.
A proper version of a paragraph would then need to know this, or expect that all nodes with an agent will be durable and rational and interoperable but if you build on that premise it will not work in the real world.
All unless you keep the docuverse limited in scope to a few data stores and agents who comprehend that spatial universe, and it is a form of an intranet