A
D
V
E
R
T
I
S
E
M
E
N
T
ADVERTISEMENT
Universal strong Kadison–Kastler stability
expertly designed by an internal OpenAI model  ·  released 2026-09-23  ·  original PDF
Theorems: 6 Lemmas: 35 Proofs: 48
Formulas: 3,065 Words: 28,981 Play time: ~3 hours

>>> How to Play <<<
We prove that sufficiently close unital von Neumann algebras are conjugate by a unitary arbitrarily close to the identity. The tolerance depends only on the prescribed distance of that unitary from the identity, uniformly over all algebras, representations, and Hilbert spaces. This resolves the strong Kadison–Kastler conjecture.

>>> Level Map <<<
  1. Introduction
  2. Historical context
  3. Proof overview
  4. Perturbation tools and a completely positive criterion
  5. Metrics, amplification, and limits
  6. The amenable embedding theorem
  7. From commutators to intertwiners
  8. The completely positive criterion
  9. Tensor products and averaging
  10. Uniform amplification of commutators
  11. Columns and trace limits
  12. Free projection words
  13. Free unitaries with controlled lifts
  14. Randomized coefficients
  15. Recovering a scalar operator
  16. Extracting the matrix commutator
  17. Boundary averaging and measurable correction
  18. Finite factors in their original representations
  19. An irreducible hyperfinite subfactor
  20. Common tracial position
  21. From the boundary to a Jordan isomorphism
  22. Return to the original Hilbert space
  23. Preparing infinite factors and comparing modular graphs
  24. A fixed tensor factor and common position
  25. Flattening rescaled Tomita graphs
  26. Separated spectral cuts
  27. Aligning the modular action
  28. Distance to the infinitesimal normalizers
  29. Analytic approximation of the modular comparison
  30. Correction within the normalizer group
  31. Common crossed products and descent to factors
  32. A concrete fixed-point description
  33. A finite corner and its inflation
  34. Undoing preparation and the auxiliary tensor factor
  35. Centers and arbitrary Hilbert spaces
  36. Aligning the centers
  37. The common-center decomposition
  38. Separable tests in an arbitrary representation
  39. Uniformity and commutator estimates
  40. Uniformity over factors and representations
  41. Two commutator arguments
  42. Matrix gaps and commutants at the conclusion

Introduction

For a von Neumann algebra \(M\subseteq\mathcal B(H)\), write \(M_1\) for its closed operator-norm unit ball. The Kadison–Kastler distance between two such algebras is \[ d(M,N)=\max\left\{ \sup_{x\in M_1}\inf_{y\in N_1}\lVert x-y\rVert,\quad \sup_{y\in N_1}\inf_{x\in M_1}\lVert x-y\rVert\right\}. \tag{1}\] Kadison and Kastler introduced this metric in 1972 (Kadison and Kastler 1972). Their perturbation problem asks whether closeness of the unit balls forces spatial isomorphism, implemented by a unitary close to the identity. We prove the uniform form of this assertion.

Theorem 1 (Strong Kadison–Kastler stability). For every \(\varepsilon>0\) there is \(\delta=\delta(\varepsilon)>0\) such that the following holds. If \(H\) is any complex Hilbert space and \(M,N\subseteq\mathcal B(H)\) are unital von Neumann algebras with identity \(I_H\) and \(d(M,N)<\delta\), there is a unitary \(u\in\mathcal B(H)\) satisfying \[uMu^*=N,\qquad \lVert u-I_H\rVert<\varepsilon.\]

The tolerance is independent of the algebras, their type, their representations, and the cardinalities of their preduals and Hilbert spaces. Thus 1 resolves the strong Kadison–Kastler conjecture positively. Its hypothesis concerns the ordinary distance (1).

Historical context

Early positive results treated type I algebras: Phillips proved spatial isomorphism, and Christensen obtained a conjugating unitary close to the identity (Phillips 1974; Christensen 1975). The work of Christensen, Johnson, and Raeburn–Taylor established the amenable case (Christensen 1977a; Johnson 1977; Raeburn and Taylor 1977). Christensen subsequently developed the one-sided theory of near inclusions: an amenable von Neumann algebra sufficiently nearly contained in another can be moved into it by a small unitary (Christensen 1980). His earlier work also treated close von Neumann subalgebras of a common finite von Neumann algebra (Christensen 1977b). The common finite ambient algebra is an essential hypothesis of that result; merely assuming that both represented algebras are finite gives a different problem. In a different common-ambient setting, Ino proved near-identity unitary conjugacy for sufficiently close subalgebras of a countably decomposable ambient algebra when both admit conditional expectations of finite probabilistic index from that algebra (Ino 2016, Theorem 3.2).

The corresponding development for \(C^*\)-algebras clarifies the role of the conclusion. Christensen, Sinclair, Smith, White, and Winter proved that sufficiently close \(C^*\)-algebras acting on a separable Hilbert space are spatially isomorphic if one of them is norm separable and nuclear (Christensen, Sinclair, Smith, White, and Winter 2010, 2012). Johnson’s examples exclude a uniform bound on the displacement of an implementing unitary that tends to zero with the algebra distance; they do not exclude spatial isomorphism (Johnson 1982).

For nonamenable von Neumann algebras, Cameron, Christensen, Sinclair, Smith, White, and Wiggins constructed strongly stable factors of the form \((L^\infty(X)\rtimes\mathrm{SL}_n(\mathbb Z))\mathbin{\overline\otimes}R\), where \(n\geq3\), the action is free, ergodic, and probability preserving on a standard nonatomic probability space, and \(R\) is the hyperfinite type \(\mathrm{II}_1\) factor (Cameron et al. 2012, 2014). Their proof combines amenable near-inclusion methods, a common standard representation, and bounded cohomology. Their later work transfers property \(\Gamma\) and Cartan structure between sufficiently close type \(\mathrm{II}_1\) factors with separable predual, and also studies stability of solidity, strong solidity, and tensor-product structure (Cameron et al. 2017).

The passage between representations is tied to Kadison’s similarity problem: whether a bounded homomorphism from a \(C^*\)-algebra into \(\mathcal B(H)\) is conjugate by an invertible operator to a \(*\)-homomorphism. Christensen’s work on property \(\Gamma\) factors provides a foundational special case (Christensen 1986, 2001). Christensen, Sinclair, Smith, and White established perturbation results for the similarity property and operator-algebraic invariants (Christensen, Sinclair, Smith, and White 2010); Cameron and collaborators then proved that the similarity property is equivalent to continuity of commutants, with the required uniformity over representations (Cameron et al. 2013). These results explain why ordinary norm closeness, control at all matrix levels, and implementation on the original Hilbert space must be kept distinct throughout a proof of strong stability. The companion manuscript (OpenAI 2026) also proves a uniform commutator estimate, using invariant tensors and a free-corner Hankel test. The amplification argument here instead uses free-word randomization and is proved independently. Neither commutator estimate alone constructs a nearby multiplicative correspondence between two algebras; Appendix 10 compares their scope and this remaining perturbation step.

The unit-ball metric supplies nearby elements but does not specify a linear or multiplicative correspondence. We must construct such a correspondence and control its implementation on the given Hilbert space, with estimates independent of the algebras and representations.

Proof overview

We first work on separable Hilbert spaces. For finite factors, two arguments meet in 5. The first supplies matrix norm control: for a separably acting type \(\mathrm{II}_1\) factor \(M\subseteq\mathcal B(H)\) and \(h=h^*\), 8 bounds all matrix amplifications of \(x\mapsto[h,x]\) by an absolute multiple of its ordinary operator norm. It combines Haagerup’s cyclic similarity estimate (Haagerup 1983) and Popa’s independence theorem for tracial ultraproducts (Popa 2014) with a free-word randomization argument. The constant is independent of matrix size and representation.

A separate argument constructs an isomorphism close in ordinary norm to the inclusion. Using the common-position construction of (Cameron et al. 2014, Lemma 4.10), we move the factors to a tracial representation in which one vector represents both traces, and the left and right actions of a common hyperfinite subfactor \(R\) agree. The inclusion is irreducible, meaning \(R'\cap M=\mathbb CI\).

For a countable strongly dense subgroup \(G\) of the unitary group of \(M\), choose nearby unitaries \(v_g\in N\). Their multiplication errors are uniformly small. Section 4 constructs a probability space \(D\) with an amenable \(G\)-action preserving null sets. An adaptation of the averaging and polar-correction method for approximate representations (Kazhdan 1982; Burger et al. 2013) corrects the nearby unitaries to measurable fields \(W_g:D\to\mathcal U(N)\) satisfying \[W_g(x)W_h(g^{-1}x)=W_{gh}(x).\] Here \(G\) need not be amenable: the correction produces a cocycle over its amenable action, not a representation on the original space. On \(D^2\), a tracial averaging argument and double ergodicity force a suitable equivariant vector field to be constant. We give the unitary-coefficient double-ergodicity argument in full, using the bilateral random-walk viewpoint of Kaimanovich; Björklund gives the corresponding unitary formulation and another proof (Kaimanovich 2003; Björklund 2014). The resulting linear bijection preserves adjoints and squares and is close to the inclusion; proximity excludes reversal of multiplication. Transporting this isomorphism back and applying the commutator estimate gives a small unitary on the original Hilbert space.

For infinite factors, a common copy of \(\mathcal B(\ell^2)\) identifies the ordinary gap with the supremum of all matrix gaps. Tensoring with one fixed hyperfinite factor gives type \(\mathrm{III}_1\) factors with an irreducible hyperfinite finite subfactor admitting a conditional expectation. This preparation uses Marrakchi’s absorption criterion for a trivial bicentralizer and Haagerup’s construction of the expected hyperfinite subfactor (Marrakchi 2020; Haagerup 1987). Houdayer and Marrakchi’s September 2026 preprint proves triviality of the bicentralizer for every type \(\mathrm{III}_1\) factor with a faithful normal state (Houdayer and Marrakchi 2026, Theorem G); the preparation here uses only the earlier absorption criterion.

Small adjustments align the subfactor’s left and right actions and provide a common cyclic separating vector \(\Omega\) whose two vector states centralize \(R\). For either algebra, the closure of \(x\Omega\mapsto x^*\Omega\) is its Tomita operator. 33 shows that small matrix gaps force its two graphs to be close when the output is multiplied by any \(c>0\), uniformly in \(c\). This uniformity permits comparison of the modular actions. A uniform infinitesimal estimate for type \(\mathrm{III}\) factors controls norm-continuous paths of unitaries whose conjugations nearly preserve the algebra. Using Arveson’s nest distance formula (Arveson 1975), analytic smoothing brings the modular comparison into this setting. Limits of conjugation maps yield, at each time, a unitary whose conjugation preserves one algebra and which is close to the other algebra’s modular unitary. Averaging corrects their multiplication law. Section 7 thereby obtains a single small change of position making the first algebra invariant under that modular group.

We form crossed products using this common action. Their unit balls are close at every matrix size, and continuous decomposition identifies one crossed product as a type \(\mathrm{II}_\infty\) factor (Takesaki 1973, 2003). After aligning matrix units, its finite corner and the corresponding nearby corner are both finite factors. The finite-factor theorem conjugates these corners and hence the crossed products. Averaging gives a unital completely positive (UCP) map between the base algebras, meaning that it preserves the identity and positivity at every matrix size. A state slice gives a second UCP map between the original factors. The criterion in 5, derived from Ricard–Roydor’s multiplication perturbation theorem (Ricard and Roydor 2014, Proposition 3.2), turns both maps into small spatial isomorphisms. It does not require the averaging maps to be normal.

Finally, aligning centers and selecting measurable unitaries assembles the factor results on separable spaces. For an arbitrary Hilbert space, finite sets of operators and vectors lie in suitable separable restrictions. Extend their small conjugating unitaries by the identity. A subnet of the conjugation maps converges pointwise in the weak operator topology to a UCP map. The finite tests put its range in the required algebra, and the same criterion finishes the proof. All tolerances depend only on the requested final unitary displacement.

1 displays how the finite-factor result enters both the finite and infinite cases and how they meet in the final reductions.

The finite-factor theorem has two inputs and is reused in the crossed-product argument. The UCP criterion supplies pointwise normalizers in Section 7, both descents in Section 8, and the final limit in Section 9.

Section 2 proves the transfer tools. Sections 3–5 establish the finite case. Sections 6–8 treat infinite factors. Section 9 removes the factor and separability assumptions and gathers the final quantifiers. Appendix 10 distinguishes uniformity over representations from uniformity over factors and compares the commutator methods of this paper and the similarity companion.

Perturbation tools and a completely positive criterion

Several steps of the proof produce an averaged map instead of a homomorphism. The main result of this section, 5, turns such a map into a small unitary conjugacy. All constants in this section are independent of the algebras and Hilbert spaces.

Metrics, amplification, and limits

We use the unit-ball formula (1) also for unital \(C^*\)-algebras. For \(A\subseteq\mathcal B(H)\), write \(\mathcal U(A)\) for its unitary group, \(A'=\{T\in\mathcal B(H):Ta=aT\text{ for every }a\in A\}\) for its commutant, and \(Z(A)=A\cap A'\) for its center. The symbol \(A\mathbin{\overline\otimes}B\) denotes the spatial von Neumann tensor product.

For a linear map \(T:A\to\mathcal B(H)\), let \(T_n\) act entrywise on \(M_n(A)\), and put \(\lVert T\rVert_{\mathrm{cb}}=\sup_n\lVert T_n\rVert\). The same notation for rectangular matrices is obtained by adding zero rows and columns and compressing. For concretely represented von Neumann algebras define \[d_{\mathrm{cb}}(A,B)=\sup_{n\ge1}d(M_n(A),M_n(B)).\] Here every matrix unit ball uses its operator norm on \(H^n\). We write \(A\subset_\gamma B\) if there is \(\gamma'<\gamma\) such that every \(x\in A_1\) has \(\operatorname{dist}(x,B)\le\gamma'\); the approximant in \(B\) need not be contractive. If \(\lVert x\rVert\le1\) and \(\lVert x-y\rVert\le r\), then \[y_0=\frac{y}{\max\{1,\lVert y\rVert\}}\in B_1, \qquad \lVert x-y_0\rVert\le2r.\] We use this normalization whenever passing from near inclusions to unit-ball distances. Conversely, a strict upper bound for \(d(A,B)\) gives both near inclusions with that bound. For a unitary \(v\), \[ \begin{aligned} d(A,vBv^*)&\le d(A,B)+2\lVert v-I\rVert,\\ d_{\mathrm{cb}}(A,vBv^*)&\le d_{\mathrm{cb}}(A,B)+2\lVert v-I\rVert. \end{aligned} \tag{2}\] These inequalities follow at every matrix level from \(\lVert v_n x v_n^*-x\rVert\le2\lVert v-I\rVert\lVert x\rVert\). Also \(\lVert v_1\cdots v_k-I\rVert\le\sum_j\lVert v_j-I\rVert\).

Inner products are linear in their first variable. A normal map between von Neumann algebras is an ultraweakly continuous map. A map is unital completely positive, abbreviated UCP, if it preserves the identity and each of its matrix amplifications is positive. Such maps are completely contractive.

We repeatedly use the following elementary compactness facts. A bounded operator ball is compact in the weak operator topology (WOT). A net of UCP maps with fixed domain and codomain \(\mathcal B(H)\) has a subnet converging pointwise in WOT: use the product of the operator balls indexed by domain elements. The limit is UCP, since linearity, the identity, and every finite matrix positivity inequality pass to the limit. If the maps are within \(a\) in cb norm of a fixed map, so is their limit; apply WOT closedness of norm balls at each fixed matrix size. If their values lie in a fixed von Neumann algebra, the limit has the same property. Normality need not pass to this limit and will not be required.

For clarity, \(O(t)\) denotes a bound \(Ct\) for an absolute constant \(C\), valid below an absolute positive threshold for \(t\). Constants may change between displays. A function tending to zero will always be chosen uniformly over the explicitly stated class of algebras.

The amenable embedding theorem

The following form of Christensen’s theorem supplies the small changes of position used in the proof (Christensen 1980).

Theorem 2 (Amenable near inclusions). Let \(R,Q\subseteq\mathcal B(H)\) be von Neumann algebras with identity \(I\), with \(R\) amenable. If \(R\subset_\gamma Q\) and \(\gamma<1/100\), there is a unitary \(v\in(R\cup Q)''\) such that \[vRv^*\subseteq Q,\qquad \lVert v-I\rVert\le150\gamma.\] If also \(Q\subset_\gamma R\) and \(\gamma<1/101\), the same conclusion holds with \(vRv^*=Q\).

This is (Cameron et al. 2014, Theorem 3.3(i)), with its strict near-inclusion convention. The location of \(v\) in \((R\cup Q)''\) is useful: any algebra commuting with both \(R\) and \(Q\) also commutes with the change of position. We apply the theorem to abelian algebras, hyperfinite finite factors, and type \(\mathrm I\) algebras, all of which are amenable.

From commutators to intertwiners

For \(h\in\mathcal B(H)\) write \(\operatorname{ad}_h(x)=[h,x]=hx-xh\).

Lemma 3. For every unital von Neumann algebra \(A\subseteq\mathcal B(H)\) and \(h\in\mathcal B(H)\), \[ \operatorname{dist}(h,A')\le\lVert\operatorname{ad}_h|_A\rVert_{\mathrm{cb}}. \tag{3}\] Consequently \(d_{\mathrm{cb}}(A,B)\le\gamma\) implies \(d(A',B')\le4\gamma\).

Proof. Since \(A'\) is weak-star closed, its preannihilator in the trace class computes the quotient norm \(\operatorname{dist}(h,A')\). Let \(\varphi\) be a trace-class functional of norm at most one vanishing on \(A'\). A singular-value decomposition gives vectors \(\xi,\eta\in\ell^2\otimes H\), each of norm at most one, with \[\varphi(x)=\langle (I\otimes x)\xi,\eta\rangle\quad(x\in\mathcal B(H)).\] This uses only countably many vectors even when \(H\) is nonseparable. Let \(p\) project onto \(\overline{(I\otimes A')\xi}\). This subspace reduces \(I\otimes A'\), so every matrix entry of \(p\) belongs to \(A\). Moreover \(p\xi=\xi\) and \(p\eta=0\). Therefore \[\varphi(h)=\langle [I\otimes h,p]\xi,\eta\rangle.\] If \(q_n\) projects onto the first \(n\) coordinates of \(\ell^2\), the compression of this commutator is \([I_n\otimes h,q_npq_n]\), of norm at most \(\lVert\operatorname{ad}_h|_A\rVert_{\mathrm{cb}}\) because \(q_npq_n\in M_n(A)_1\). Testing on finitely supported vectors and then using density proves the same bound for the full commutator. Taking the supremum over \(\varphi\) proves (3).

For \(h\in A'_1\), approximate each contraction in \(M_n(B)\) by one in \(M_n(A)\), with arbitrary slack above \(\gamma\). It follows that \(\lVert\operatorname{ad}_h|_B\rVert_{\mathrm{cb}}\le2\gamma\). The first assertion gives distance at most \(2\gamma\) to \(B'\), and normalization of approximants gives distance at most \(4\gamma\) to \(B'_1\). Reverse the algebras to finish. ◻

Lemma 4 (Close representations). Let \(\rho,\theta:P\to\mathcal B(H)\) be normal unital \(*\)-representations of a von Neumann algebra. If \(\lVert\rho-\theta\rVert_{\mathrm{cb}}\le e<1/4\), there is a unitary \(u\in\mathcal B(H)\) with \[u\rho(x)u^*=\theta(x)\quad(x\in P),\qquad \lVert u-I\rVert\le4e.\]

Proof. If \(e=0\), take \(u=I\). Otherwise let \[\pi(x)=\begin{pmatrix}\rho(x)&0\\0&\theta(x)\end{pmatrix}, \qquad S=\begin{pmatrix}0&I\\I&0\end{pmatrix}.\] The von Neumann image \(\pi(P)\) has the quotient matrix norms of \(P\). The commutator of \(S\) with \(\pi\) has cb norm at most \(e\), hence the same is true on \(\pi(P)\) by quotient lifts with arbitrarily small norm slack. By 3 choose \(D\in\pi(P)'\) with \(\lVert D-S\rVert<2e\). Its lower off-diagonal block \(T\) satisfies \[T\rho(x)=\theta(x)T,\qquad \lVert T-I\rVert<2e<1.\] Thus \(T\) is invertible and \(T^*T\) commutes with \(\rho(P)\). Its polar unitary \(u=T|T|^{-1}\) intertwines the representations. The singular values of \(T\) lie between \(1-r\) and \(1+r\), where \(r=\lVert T-I\rVert\), so \[\lVert u-I\rVert\le\lVert u-T\rVert+r=\lVert I-|T|\rVert+r\le2r<4e.\] ◻

The completely positive criterion

The norm relevant to perturbing a product is the completely bounded norm on the Haagerup tensor product. We recall only the form needed here. For a bilinear map \(b:P\times P\to P\), it is the least \(K\) such that, at all rectangular sizes, \[ \left\|\left[\sum_{j=1}^r b(x_{ij},y_{jk})\right]_{i,k}\right\| \le K\lVert X\rVert\lVert Y\rVert, \quad X\in M_{p,r}(P),\quad Y\in M_{r,q}(P). \tag{4}\] Indeed the matrix Haagerup norm is the infimum of \(\lVert X\rVert\lVert Y\rVert\) over such row-column factorizations. We write \(\lVert b\rVert_{\mathrm{cb}}\) for this norm as well.

Write \(m_P(x,y)=xy\) for the original multiplication on \(P\). Correction of multiplication by cohomological methods goes back to Johnson and Raeburn–Taylor (Johnson 1977; Raeburn and Taylor 1977). We use the completely bounded version supplied by Ricard–Roydor’s multiplication perturbation theorem (Ricard and Roydor 2014, Proposition 3.2): on a von Neumann algebra \(P\), an associative product \(m\) with \(\lVert m-m_P\rVert_{\mathrm{cb}}\le1/11\) admits a cb linear isomorphism \(L\) with \[L(m(x,y))=L(x)L(y),\qquad \lVert L-\operatorname{id}\rVert_{\mathrm{cb}}\le10\lVert m-m_P\rVert_{\mathrm{cb}}.\] If \(m(x,y)^*=m(y^*,x^*)\), \(L\) can preserve the involution. There is no normality or separability assumption on the new product.

Proposition 5 (UCP criterion). Let \(P,Q\) be von Neumann algebras, let \(\rho:P\to\mathcal B(H)\) be a faithful normal unital representation, and let \(Q\subseteq\mathcal B(H)\) have the same identity. Suppose that \(F:P\to Q\) is UCP and \[\lVert F-\rho\rVert_{\mathrm{cb}}\le e<10^{-3},\qquad \sup_{y\in Q_1}\operatorname{dist}(y,\rho(P)_1)\le g,\qquad g+e<1.\] Then there is a normal unital \(*\)-isomorphism \(\theta:P\to Q\) and a unitary \(u\in\mathcal B(H)\) such that \[\lVert\theta-\rho\rVert_{\mathrm{cb}}\le32e,\qquad u\rho(x)u^*=\theta(x)\ (x\in P),\qquad \lVert u-I\rVert\le128e.\] The map \(F\) need not be normal.

Proof. Complete isometry of \(\rho\) gives \(\lVert F_n(X)\rVert\ge(1-e)\lVert X\rVert\) at every matrix size. Thus \(F\) is injective, its range \(V\) is norm closed, and \(\lVert F^{-1}:V\to P\rVert_{\mathrm{cb}}\le(1-e)^{-1}\). For \(y\in Q_1\), match \(y\) with \(\rho(x)\) for \(x\in P_1\), allowing arbitrary slack, to obtain \(\operatorname{dist}(y,V)\le g+e<1\). If \(V\) were proper, the quotient map \(Q\to Q/V\) would have norm one, contradicting this bound on \(Q_1\). Hence \(F\) is onto.

Transport multiplication from \(Q\) to \(P\): \[m(x,y)=F^{-1}(F(x)F(y)).\] It is associative and preserves the involution. For rectangular \(X,Y\), insert \(\rho(XY)=\rho(X)\rho(Y)\) to see that \[\lVert F(XY)-F(X)F(Y)\rVert\le3e\lVert X\rVert\lVert Y\rVert.\] Here the first difference is bounded by \(e\lVert XY\rVert\) and the product difference by \(2e\lVert X\rVert\lVert Y\rVert\), since \(F\) and \(\rho\) are complete contractions. Equation (4) therefore gives \[q:=\lVert m-m_P\rVert_{\mathrm{cb}}\le\frac{3e}{1-e}<\frac1{11}.\] The cited theorem gives an involution-preserving \(L\) with \(\lVert L-\operatorname{id}\rVert_{\mathrm{cb}}\le\lambda:=30e/(1-e)<1\). Then \(\theta=FL^{-1}\) is a bijective \(*\)-homomorphism, and \[\lVert\theta-\rho\rVert_{\mathrm{cb}} \le e+\lVert L^{-1}-\operatorname{id}\rVert_{\mathrm{cb}} \le e+\frac{\lambda}{1-\lambda} = e+\frac{30e}{1-31e}\le32e.\] Bijective multiplicativity makes \(\theta\) unital. A \(*\)-isomorphism between von Neumann algebras is an order isomorphism on their selfadjoint parts, so it preserves bounded increasing suprema and is normal. Apply 4, since \(32e<1/4\). When \(e=0\), the preceding range argument gives \(Q=\rho(P)\) directly. ◻

We will use this criterion both to descend from a tensor product and to replace a pointwise weak limit by a spatial isomorphism. Its quantitative conclusion depends only on the cb error of the map; the ordinary reverse gap supplies surjectivity.

Tensor products and averaging

Lemma 6. Let \(A,B\subseteq\mathcal B(H)\) be unital von Neumann algebras with \(d_{\mathrm{cb}}(A,B)\le\gamma\). If \(K\) is separable, then \[d_{\mathrm{cb}}(A\mathbin{\overline\otimes}\mathcal B(K),B\mathbin{\overline\otimes}\mathcal B(K))\le\gamma.\] If \(D\subseteq\mathcal B(K)\) is a von Neumann algebra whose commutant is the strong closure of increasing finite-dimensional unital algebras, then \[d_{\mathrm{cb}}(A\mathbin{\overline\otimes}D,B\mathbin{\overline\otimes}D)\le\gamma.\]

Proof. Compress a contraction of \(M_n(A\mathbin{\overline\otimes}\mathcal B(K))\) by increasing finite-rank projections on \(K\). In a corner of rank \(r\), choose an opposite contraction in \(M_{nr}(B)\) within \(\gamma\) plus arbitrarily small slack, and extend it by zero on the complement. The compressions of the original contraction converge strongly to it. A bounded WOT cluster of the opposite contractions is in \(M_n(B\mathbin{\overline\otimes}\mathcal B(K))_1\) and is within \(\gamma\), by WOT closedness of the norm bound, choosing the approximation slack to tend to zero along the compressions. Reverse the argument and take the supremum over \(n\).

Now start with a contraction in \(M_n(A\mathbin{\overline\otimes}D)\). The first assertion supplies a nearby contraction in \(M_n(B\mathbin{\overline\otimes}\mathcal B(K))\). Average it by conjugation over the unitary group of each finite-dimensional stage of \(D'\). The input commutes with these groups, so the bound is preserved. A WOT cluster of the averages commutes with \(D'\) and still lies in \(M_n(B\mathbin{\overline\otimes}\mathcal B(K))\). The von Neumann tensor commutation theorem (Blackadar 2006, Theorem III.4.5.8 and Corollary III.4.5.9, pp. 319–320) places it in \(M_n(B\mathbin{\overline\otimes}D)\). Reverse and amplify as before. ◻

Lemma 7 (Invariant operator means). Let \(G\) be a discrete amenable group and \(m\) a left invariant mean on \(\ell^\infty(G)\): thus \(m\) is a positive unital linear functional with \(m(f(h^{-1}\,\cdot))=m(f)\) for every \(h\in G\). For a bounded family \(T_g\in\mathcal B(K)\), define \[\langle (m_gT_g)\xi,\eta\rangle=m_g\langle T_g\xi,\eta\rangle.\] This averaging map from bounded operator families to \(\mathcal B(K)\) is UCP, commutes with multiplication on either side by a fixed operator, and preserves weak-operator closed linear subspaces. If \(\lVert T_g-T\rVert\le a\) for every \(g\), then \(\lVert m_gT_g-T\rVert\le a\), also at rectangular matrix levels. No measurability of the family is required.

For a unitary representation \(g\mapsto V_g\) on \(K\), the map \[E_V(T)=m_gV_gTV_g^*\qquad(T\in\mathcal B(K))\] is a UCP projection onto \(\{V_g:g\in G\}'\). It preserves every von Neumann algebra invariant under these conjugations. More generally, for an action by automorphisms \(\alpha_g\) on a von Neumann algebra \(P\), the formula \[\varphi(E(x))=m_g\varphi(\alpha_g(x))\qquad(\varphi\in P_*)\] defines a UCP map onto \(P^G\) fixing that algebra. These averaging maps need not be normal.

Proof. The coefficient formula defines a bounded sesquilinear form and hence an operator. Positivity at every matrix size and unitality follow by applying the positive scalar mean to matrix quadratic forms; the domain has the pointwise operations on bounded operator families. Fixed left and right factors pass through the coefficient formula. Testing the linear equations defining a weak-operator closed subspace proves range preservation. Testing vector coefficients at each rectangular size gives the norm bound. For the action average, left invariance makes the range invariant, while averaging fixes each invariant operator. An invariant algebra is preserved by the preceding range property. Predual duality gives the same argument for the general action on \(P\), including preservation of matrix norm bounds. ◻

For \(G=\mathbb R\) with the discrete topology, an invariant mean is obtained as a weak-star cluster point of uniform averages on finite boxes in finitely generated subgroups. Direct the choices by a finite set of translations and an error tolerance. Each subgroup is free abelian, and the relative boundary of its boxes tends to zero for every prescribed translation; the cluster is therefore positive, unital, and translation invariant.

The boundary means in 4 require a separate construction: their equivariance involves a nonsingular action and their module coefficients vary over the boundary space.

Uniform amplification of commutators

The finite-factor argument will first produce an isomorphism that is close to the given representation in operator norm. To implement it by a small unitary, we need the same control at every matrix level. The following estimate supplies that passage.

Theorem 8 (Uniform amplification). There is an absolute constant \(C_{\mathrm{amp}}\) such that, for every type \(\mathrm{II}_1\) factor \(M\subseteq\mathcal B(H)\) on a separable Hilbert space and every selfadjoint \(h\in\mathcal B(H)\), \[\lVert\operatorname{ad}_h|_M\rVert_{\mathrm{cb}}\le C_{\mathrm{amp}}\lVert\operatorname{ad}_h|_M\rVert.\] The constant is independent of \(M\), its representation, and \(H\).

We first control columns by the cyclic similarity theorem. A free-word estimate then permits us to apply this column bound to randomized matrix coefficients. The randomization also recovers a fixed positive multiple of the identity in a weak limit; this will detect every matrix commutator. The auxiliary matrix sizes depend on the matrix level being tested, but every estimate that survives to the conclusion is absolute.

Columns and trace limits

For \(a_1,\ldots,a_l\in\mathcal B(H)\), write \(\lVert(a_i)_i\rVert_{\mathrm{col}}=\lVert\sum_i a_i^*a_i\rVert^{1/2}\).

Lemma 9 (Column estimate). Let \(A\subseteq\mathcal B(H)\) be a unital \(C^*\)-algebra, let \(h\in\mathcal B(H)\), and put \(\delta=\lVert\operatorname{ad}_h|_A\rVert\). Then \[\lVert([h,a_i])_i\rVert_{\mathrm{col}} \le 8\delta\,\lVert(a_i)_i\rVert_{\mathrm{col}} \qquad(a_1,\ldots,a_l\in A).\]

Proof. There is nothing to prove if \(\delta=0\). Otherwise we use the triangular homomorphism associated with a derivation, as in (Haagerup 1983, 216): \[\Phi(a)= \begin{pmatrix}a&\delta^{-1}[h,a]\\0&a\end{pmatrix} \quad(a\in A).\] This is a unital homomorphism with \(\lVert\Phi\rVert\le2\). For \(\xi\in H\), the closed subspace \[K_\xi=\overline{\{\Phi(a)(0,\xi):a\in A\}} \subseteq H\oplus H\] is invariant for \(\Phi(A)\), and the restricted homomorphism is cyclic. Haagerup’s cyclic similarity theorem gives \(\lVert\Phi|_{K_\xi}\rVert_{\mathrm{cb}}\le\lVert\Phi|_{K_\xi}\rVert^3\le8\); see (Haagerup 1983, Theorem 1.1 and its proof, pp. 217, 225). Apply its column amplification to \((0,\xi)\in K_\xi\), and retain the upper component of each output. This gives \[\left(\sum_i\lVert[h,a_i]\xi\rVert^2\right)^{1/2} \le8\delta\,\lVert(a_i)_i\rVert_{\mathrm{col}}\lVert\xi\rVert.\] Taking the supremum over unit vectors proves the claim. ◻

We use normalized traces on finite factors and the standard comparison and subdivision of projections by trace (Blackadar 2006, III.1.7.9–10, p. 257; III.2.5.4(iii) and Theorem III.2.5.7, p. 279). For a finite von Neumann algebra \((P,\tau)\), put \(\lVert x\rVert_2=\tau(x^*x)^{1/2}\) and \(\lVert x\rVert_1=\tau(|x|)\). Fix a free ultrafilter \(\omega\) on \(\mathbb N\). The tracial ultrapower \(P^\omega\) is the quotient of the bounded sequences in \(P\) by those with \(\lim_{j\to\omega}\lVert x_j\rVert_2=0\); its trace is \(\tau_\omega((x_j)_\omega)=\lim_{j\to\omega}\tau(x_j)\). We identify \(P\) with its constant sequences.

Lemma 10 (Transfer from the trace to a normal representation). Let \(P\subseteq\mathcal B(K)\) be a finite von Neumann algebra with faithful normal trace \(\tau\).

  1. If \((x_j)\) is bounded and \(\lim_{j\to\omega}\lVert x_j\rVert_2=0\), then \(\lim_{j\to\omega}\lVert x_j\xi\rVert=0\) for every \(\xi\in K\).

  2. If \(x=(x_j)_\omega\in P^\omega\), then the weak-operator ultralimit of \(x_j\) in this representation is \(E_P(x)\), where \(E_P:P^\omega\to P\) is the trace-preserving conditional expectation.

Proof. The functionals \(a\mapsto\tau(ba)\), \(b\in P\), are norm dense in the predual \(P_*\). Indeed, their annihilator in \((P_*)^*=P\) is zero, by faithfulness of the trace, so Hahn–Banach gives the density. For a normal vector functional \(\varphi\), approximate \(\varphi\) in norm by one of these functionals. If \(\lVert x_j\rVert\le K_0\), then \[|\varphi(x_j^*x_j)| \le K_0^2\lVert\varphi-\tau(b\,\cdot)\rVert +\lVert b\rVert\lVert x_j\rVert_2^2.\] This proves (i). For (ii), the defining identity of the conditional expectation gives \[\lim_{j\to\omega}\tau(bx_j) =\tau_\omega(bx)=\tau(bE_P(x))\qquad(b\in P).\] The same predual-density argument, using boundedness, extends the identity to every normal functional. ◻

Free projection words

The following estimate controls sums of projection words whose labels have the same multiplicities. In the later randomization, these sums will give coefficients to which we apply the column estimate separately. When the squared estimates are added, each diagonal vector component is counted once for every multiset containing its label. We therefore need squared word-norm bounds with a uniformly bounded sum through each fixed label.

Lemma 11 (Free projection words). For every integer \(r\geq 1\) there are constants \(C_r>0\) and an integer \(d_r\geq r\) with the following property. Let \(d\geq d_r\), and let \((\mathcal A,\tau)\) be a von Neumann algebra with a faithful normal tracial state. Suppose that \(A_1,\ldots,A_r\subset\mathcal A\) are unital copies of \(M_d(\mathbb C)\) which are scalar free: the trace of a product of centered letters from successively different copies is zero. For each \(j\), let \((e_a^{(j)})_{a=1}^d\) be the diagonal minimal projections of \(A_j\).

For a multiset \(W\) of cardinality \(r\) drawn from \(\{1,\ldots,d\}\), put \[B_W=\sum_{\substack{(a_1,\ldots,a_r)\in\{1,\ldots,d\}^r\\ \text{with multiset }W}} e_{a_r}^{(r)}\cdots e_{a_1}^{(1)}.\] Then \(\lVert B_W\rVert\leq\beta_W\), where \[ \beta_W= \begin{cases} 7\sqrt{(r-1)!}\,d^{-(r-1)/2}, &\text{if all labels in $W$ are distinct},\\ C_r d^{-(r-1)/2},&\text{otherwise}. \end{cases} \tag{5}\] The constants and threshold depend only on \(r\), not on the ambient algebra, the free copies, or \(W\).

Every word in \(B_W\) ends with a projection \(e_a^{(1)}\) whose label occurs in \(W\). Thus \(B_W\) has right support in the sum of those projections from the first copy, at most \(r\) of them. This support bound will let us reduce each coefficient to a column, once the auxiliary matrix dimension is large enough.

Proof. We estimate the centered words first, then restore their scalar parts. The decomposition below uses the creation, annihilation, and diagonal actions in the free-product Hilbert space, as in the framework developed by Voiculescu (Voiculescu 1985) and the reduced-free-product inequalities of Ricard and Xu (Ricard and Xu 2006, sec. 1 and Fact 2.6). The corresponding bound for words of fixed length in free groups is Haagerup’s inequality (Haagerup 1979, Lemma 1.4). Here the prescribed label multiplicities require the additional tuple-matrix and square-summation estimates proved below.

Reduced words and their cuts.

We first describe the Hilbert space on which the word norm will be estimated. Replace \(\mathcal A\) by \(W^*(A_1,\ldots,A_r)\) and use its faithful left representation on \(L^2(\mathcal A,\tau)\). Set \[H_j^\circ=L^2(A_j,\tau)\ominus\mathbb C1.\] Freeness identifies this representation space with \[ \mathcal F=\mathbb C\Omega\ \oplus\! \bigoplus_{\substack{\ell\geq1,\ i_1,\ldots,i_\ell\in\{1,\ldots,r\}\\ i_t\ne i_{t+1}}} H_{i_1}^\circ\otimes\cdots\otimes H_{i_\ell}^\circ. \tag{6}\] Indeed, map a tensor to the corresponding product applied to \(1\). Its inner products are the tensor inner products: in the trace of the adjoint of one reduced word times another, pair the adjacent letters at the junction, split their product into its trace and centered part, and iterate. Freeness eliminates the centered terms until the copy sequences agree and all letters have been paired. Distinct summands are orthogonal. Splitting each letter into its scalar and centered parts shows that reduced words span the algebra generated by the copies, hence a dense subspace of \(L^2(\mathcal A,\tau)\).

Write \(h_a^{(j)}=e_a^{(j)}-d^{-1}1\). Define the label map \[D_j:\ell^2(\{1,\ldots,d\})\longrightarrow H_j^\circ, \qquad D_j\delta_a=h_a^{(j)}.\] The normalized matrix trace gives \[ D_j^*D_j=\frac1d I-\frac1{d^2}\mathbf1\mathbf1^*, \qquad \lVert D_j\rVert\leq d^{-1/2}, \tag{7}\] where \(\mathbf1\) is the vector all of whose entries are one.

Left multiplication by \(h_a^{(j)}\) creates the first letter \(h_a^{(j)}\) on a reduced tensor not beginning in copy \(j\). On a tensor \(b\otimes\xi\) beginning in that copy, it has two parts: annihilation, equal to \(\tau(h_a^{(j)}b)\xi\), and preservation of the first leg, with operator \[ K_a^{(j)} =P_j^\circ L_{e_a^{(j)}}P_j^\circ -\frac1d I_{H_j^\circ} \quad\text{on }H_j^\circ. \tag{8}\] Here \(P_j^\circ\) is the orthogonal projection onto \(H_j^\circ\) inside \(L^2(A_j,\tau)\), and \(L\) denotes left multiplication. These formulas also apply to the vacuum, on which only creation occurs.

Consider a fully centered word \(h_{a_r}^{(r)}\cdots h_{a_1}^{(1)}\). The rightmost letter acts first. Since the copy indices \(1,\ldots,r\) are distinct, once a creation or preservation occurs, every later operation must be a creation. Thus the word is a sum of cuts of two kinds: \(p\) annihilations followed by \(q=r-p\) creations, or \(p\) annihilations, one preservation, and \(q=r-p-1\) creations.

For completeness, the domain restrictions in these cuts can be made explicit. For a set \(S\) of copy indices, let \(\mathcal F_{\notin S}\) denote the vacuum plus all reduced tensors whose first copy index is outside \(S\). For a cut without preservation put \[S_p=\bigl(\{p\}\text{ if }p\geq1\bigr) \cup\bigl(\{p+1\}\text{ if }p<r\bigr), \qquad E_p=\mathcal F_{\notin S_p}.\] Its source and target spaces are canonically embedded copies of \[H_1^\circ\otimes\cdots\otimes H_p^\circ\otimes E_p, \qquad H_r^\circ\otimes\cdots\otimes H_{p+1}^\circ\otimes E_p,\] respectively; empty tensor products mean \(\mathbb C\). Excluding copy \(p\) makes the input reduced, and excluding copy \(p+1\) permits the first creation. For a preservation cut put \(j=p+1\) and \(E=\mathcal F_{\notin\{j\}}\). Its source and target spaces are \[H_1^\circ\otimes\cdots\otimes H_p^\circ\otimes H_j^\circ\otimes E, \qquad H_r^\circ\otimes\cdots\otimes H_{j+1}^\circ\otimes H_j^\circ\otimes E.\] All embeddings are isometries; a cut is zero on the orthogonal complement of its indicated source space. In particular these restrictions depend on the cut alone and never on the labels.

Distinct labels and tuple matrices.

We now suppose that \(W\) has distinct labels and sum the fully centered words with multiset \(W\). In each cut, the annihilation map is the tensor product of \(D_1^*,\ldots,D_p^*\), with the identity on the remaining factors. The creation map uses the \(D_j\) in decreasing copy order. Their norms are at most \(d^{-p/2}\) and \(d^{-q/2}\). Between them is a matrix on label tuples, scalar-valued without preservation and operator-valued on \(H_j^\circ\) with preservation. Tensor order reversals are unitary. Together with the isometric source and target embeddings, this gives a factorization of each cut whose norm is bounded by \[ d^{-(p+q)/2} \lVert\text{matrix on label tuples}\rVert. \tag{9}\] The tail factor carries the identity throughout.

Without preservation, the tuple matrix has entry one exactly when the input and output tuples together list \(W\). Group tuples by their underlying subsets. Each input subset of size \(p\) is paired only with its complementary output subset of size \(q\), giving mutually orthogonal all-ones blocks of size \(q!\) by \(p!\). The matrix norm is therefore \(\sqrt{p!q!}\). Summing these cuts gives at most \[ \sum_{p=0}^r\sqrt{p!(r-p)!}\,d^{-r/2} \leq(r+1)\sqrt{r!}\,d^{-r/2}. \tag{10}\]

For a preservation cut, \(p+q=r-1\). Summing and replicating over the orders of the input and output tuples introduces a factor \(\sqrt{p!q!}\). The remaining matrix is indexed by subsets \(B\subset W\) of size \(p\) on the input and subsets \(A\subset W\) of size \(q\) on the output. Its entry is \(K_a^{(j)}\) when \(A\cap B\) is empty and \(\{a\}=W\setminus(A\cup B)\); other entries vanish.

To estimate this matrix, first replace \(K_a^{(j)}\) by the projection \(L_{e_a^{(j)}}\) on \(L^2(A_j,\tau)\) and call the resulting matrix \(T\). For a fixed row \(A\), the admissible columns \(B\) give distinct missing labels \(a\), so their output ranges are orthogonal. For a fixed column \(B\), the admissible rows also give distinct \(a\), so the projection domains are orthogonal. Thus, for an input \((\xi_B)_B\), \[\lVert T\xi\rVert^2 =\sum_A\sum_{\substack{B:\ A\cap B=\varnothing}} \lVert L_{e_{a(A,B)}^{(j)}}\xi_B\rVert^2 \leq\sum_B\lVert\xi_B\rVert^2.\] Consequently \(\lVert T\rVert\leq1\). Compressing every input and output leg to \(H_j^\circ\) preserves this bound. By (8), the remaining scalar correction is \(d^{-1}\) times the zero-one matrix of admissible pairs, tensored with \(I_{H_j^\circ}\). This matrix has \(p+1\) entries in each row and \(q+1\) in each column; Cauchy–Schwarz in each row bounds its norm by \(\sqrt{(p+1)(q+1)}\). The preservation tuple matrix therefore has norm at most \[\sqrt{p!q!} \left(1+\frac{\sqrt{(p+1)(q+1)}}d\right) \leq\sqrt{p!q!}\left(1+\frac rd\right).\] Using (9) and summing the preservation cuts gives \[ \left(1+\frac rd\right)\sqrt{(r-1)!}\,d^{-(r-1)/2} \sum_{p=0}^{r-1}\binom{r-1}{p}^{-1/2}. \tag{11}\]

The last sum is at most five, uniformly in \(r\). To see this, put \(m=r-1\). For \(m=0,1,2,3\) the sums are respectively \(1\), \(2\), \(2+1/\sqrt2\), and \(2+2/\sqrt3\). For \(m\geq4\), unimodality of the binomial coefficients gives \[\sum_{p=0}^m\binom mp^{-1/2} \leq2+\frac2{\sqrt m} +(m-3)\sqrt{\frac2{m(m-1)}} \leq3+\sqrt2<5.\]

Scalar parts and repeated labels.

It remains to restore the scalar parts of the projections and to treat repeated labels. We record a bound for a single centered word of length \(l\geq1\) in distinct copy positions, with arbitrary labels. The same cuts now have scalar coefficients rather than tuple sums. Each creation or annihilation has norm at most \(d^{-1/2}\), and each preservation has norm at most one, since it is a compression of multiplication by \(e_a-d^{-1}1\). There are \(l\) preservation cuts and \(l+1\) cuts without preservation. Hence \[ \lVert h_{a_l}^{(i_l)}\cdots h_{a_1}^{(i_1)}\rVert \leq l\,d^{-(l-1)/2}+(l+1)d^{-l/2} \leq(2l+1)d^{-(l-1)/2}. \tag{12}\] This estimate does not require the labels to be distinct.

Expand a projection word of length \(r\) using \(e_a=d^{-1}1+h_a\). A term retaining \(l<r\) centered letters contains the scalar factor \(d^{-(r-l)}\). For \(1\leq l<r\), (12) bounds it by \[(2l+1)d^{-r+(l+1)/2}\leq(2l+1)d^{-r/2}.\] The term retaining no centered letters has norm \(d^{-r}\leq d^{-r/2}\). Thus the difference between \(B_W\) and its fully centered sum has norm at most \[ r! H_r d^{-r/2}, \qquad H_r=1+\sum_{l=1}^{r-1}\binom rl(2l+1), \tag{13}\] because at most \(r!\) ordered tuples have multiset \(W\).

For distinct labels, combining (10), (11), and (13) yields \[\lVert B_W\rVert \leq5\left(1+\frac rd\right) \sqrt{(r-1)!}\,d^{-(r-1)/2} +J_r d^{-r/2}, \qquad J_r=(r+1)\sqrt{r!}+r!H_r.\] For example, choose an integer \[d_r\geq\max\left\{5r,\frac{J_r^2}{(r-1)!}\right\}.\] Then the two excess terms, after division by \(\sqrt{(r-1)!}\,d^{-(r-1)/2}\), are each at most one, proving the constant seven in (5).

For repeated labels, apply (12) to every fully centered summand and use (13). Since \(d^{-r/2}\leq d^{-(r-1)/2}\), it suffices to take \[C_r=r!\bigl(2r+1+H_r\bigr).\] These choices depend only on \(r\) and complete the proof. ◻

Corollary 12 (Sum of squared word bounds through one label). For every integer \(N\geq1\) there is an integer \(d_0(N)\) such that, whenever \(d\geq d_0(N)\), the constants (5) satisfy \[ \max_{1\leq a\leq d} \sum_{\substack{W\text{ a multiset from }\{1,\ldots,d\}\\ |W|=r,\ a\in W}} \beta_W^2\leq64 \qquad(1\leq r\leq N). \tag{14}\] Membership \(a\in W\) here means that \(a\) occurs at least once; each multiset is counted once, regardless of that multiplicity. In particular, for free copies as in 11, the norm bounds and (14) hold simultaneously through length \(N\) after choosing this one threshold for \(d\).

Proof. Fix \(r\) and a label \(a\). There are \(\binom{d-1}{r-1}\) distinct-label multisets containing \(a\), and their contribution is \[49(r-1)!d^{-(r-1)}\binom{d-1}{r-1}\leq49.\] For \(r=1\) this is the entire sum. Suppose \(r\geq2\). A repeated-label multiset has support of cardinality \(t\leq r-1\). If it contains \(a\), its other \(t-1\) support labels can be chosen in \(\binom{d-1}{t-1}\) ways. Its positive multiplicities, with total \(r\), can be chosen in \(\binom{r-1}{t-1}\) ways. Consequently the number of such multisets is at most \[\sum_{t=1}^{r-1}\binom{d-1}{t-1}\binom{r-1}{t-1} \leq A_r d^{r-2}, \qquad A_r=\sum_{u=0}^{r-2}\frac1{u!}\binom{r-1}{u}.\] Their contribution to (14) is therefore at most \(C_r^2 A_r/d\). Choose \(d_0(N)\) to be an integer at least all the \(d_r\) for \(1\leq r\leq N\) and all the numbers \(C_r^2A_r/15\) for \(2\leq r\leq N\). The latter condition is empty when \(N=1\). The sum is then at most \(49+15=64\), uniformly in \(a\) and simultaneously for the finitely many lengths \(r\leq N\). ◻

Free unitaries with controlled lifts

For varying finite tracial algebras \((F_j,\tau_j)\), the same bounded-sequence construction defines the tracial ultraproduct \(\prod_\omega(F_j,\tau_j)\). A unitary \(s\) is Haar if \(\tau(s^m)=0\) for every nonzero \(m\in\mathbb Z\). Two unital subalgebras are scalar free if every alternating product of trace-zero elements from them has trace zero. Subalgebras containing \(B\) are free over \(B\) if the corresponding assertion holds with the trace replaced by the trace-preserving conditional expectation \(E_B\).

Lemma 13 (Free Haar unitaries with matrix-algebra lifts). Let \(M\subseteq\mathcal B(H)\) be a separably acting \(\mathrm{II}_1\) factor, let \(n\geq1\), and put \(P=M_n(M)\). For every free ultrafilter \(\omega\) there is a Haar unitary \(s\in P^\omega\) such that \(W^*(s)\) and the constant copy of \(P\) are scalar free. Moreover, \(s\) has unitary representatives \[s=(s_j)_\omega,\qquad s_j\in\mathcal U(M_n(R_j)),\] where each \(R_j\subseteq M\) is a unital full matrix algebra. Consequently, for every \(h=h^*\in\mathcal B(H)\), with \(\delta=\lVert\operatorname{ad}_h|_M\rVert\) and \(h_n=I_n\otimes h\), these representatives satisfy \[ \lVert[h_n,s_j]\rVert\leq2\delta\qquad(j\geq1). \tag{15}\]

Proof. We construct the free unitary in two stages, then control its representatives.

Ultraproduct facts and the independence input.

We first record the two facts about tracial ultraproducts used in the construction. A tracial ultraproduct of finite factors is a factor. Indeed, in any finite factor \(F\) the scalar \(\tau(x)1\) belongs to the \(2\)-norm closed convex hull of \(\{uxu^*:u\in\mathcal U(F)\}\). To see this, take the unique element of least \(2\)-norm in that closed convex hull. It is fixed by every unitary conjugation. It still belongs to \(F\), because the convex hull is uniformly bounded in operator norm and the same is true of its \(2\)-norm closure. It is therefore scalar, and its trace identifies it as \(\tau(x)1\). In particular, \[\lVert x-\tau(x)1\rVert_2 \leq\sup_{u\in\mathcal U(F)}\lVert[x,u]\rVert_2.\] If \((x_j)_\omega\) is central in an ultraproduct, the supremum on the right tends to zero along \(\omega\): otherwise choosing an almost maximizing unitary in each offending coordinate would give a unitary sequence not commuting with \((x_j)_\omega\). Thus every central element is scalar. The ultraproduct of matrix factors whose sizes tend to infinity is diffuse, since their projections have traces approaching any prescribed number in \([0,1]\). An ultrapower of a \(\mathrm{II}_1\) factor is likewise diffuse, since it contains the constant factor. These ultraproducts are therefore \(\mathrm{II}_1\) factors.

Second, every unitary in such an ultraproduct lifts to coordinate unitaries. If \((x_j)_\omega\) is unitary, then \(\lim_\omega\lVert x_j^*x_j-1\rVert_2=0\). Extend the partial isometry in the polar decomposition of \(x_j\) to a unitary \(u_j\). This extension exists because the two complementary support projections have the same trace in the finite factor. Functional calculus gives \[\lVert x_j-u_j\rVert_2 =\lVert|x_j|-1\rVert_2 \leq\lVert x_j^*x_j-1\rVert_2,\] so \((u_j)_\omega=(x_j)_\omega\).

We use the following form of Popa (2014, Theorem 0.1(b)). Let \(\mathcal F=\prod_\omega F_j\), where the \(F_j\) are finite factors with \(\dim F_j\longrightarrow\infty\), allowing infinite dimensions. For a separable amenable von Neumann subalgebra \(B_1\subseteq\mathcal F\), put \(Q=B_1'\cap\mathcal F\) and \(D=Q'\cap\mathcal F\). If \(\mathcal X\subseteq\ker E_D\) is a separable linear subspace, there is a diffuse abelian von Neumann algebra \(A\subseteq Q\) such that \[ E_D(x_0a_1x_1\cdots a_rx_r)=0 \tag{16}\] whenever \(r\geq1\), \(x_0\in\mathcal X\cup\{1\}\), \(x_1,\ldots,x_r\in\mathcal X\), and \(a_1,\ldots,a_r\in A\cap\ker\tau\). Here separability is for the \(2\)-norm, as specified in (Popa 2014, sec. 1.1). We use the displayed form with its final \(x_r\) and establish the additional endpoints needed below.

In our applications \(D\) is a full matrix algebra, \(A\subseteq D'\cap\mathcal F\), and \(\mathcal X\) is selfadjoint. For \(a\in A\), bimodularity implies \(E_D(a)\in Z(D)\), whence \(E_D(a)=\tau(a)1\). Taking adjoints in (16) also treats alternating words that start with an \(\mathcal X\)-letter and end with an \(A\)-letter. For a word with both endpoints in \(A\), write \[W=a_0x_1a_1\cdots x_ra_r, \qquad r\geq1,\] with all displayed letters centered in their respective senses. For \(b\in D\), traciality and \([A,D]=0\) give \[\tau(bW)=\tau\bigl(b\,a_ra_0x_1a_1\cdots x_r\bigr).\] Split \(a_ra_0\) into its trace and a trace-zero element of \(A\). The latter term vanishes by (16) with \(x_0=1\); the scalar term vanishes by the same formula when \(r\geq2\), and by \(E_D(x_1)=0\) when \(r=1\). Thus \(E_D(W)=0\). A single centered \(A\)-letter has already been covered. Hence all alternating endpoint patterns used below are valid.

A finite-matrix model.

We first construct a matrix-algebra model that is free from \(B=M_n(\mathbb C)\otimes1\). In \[\mathcal F_0=\prod_\omega\bigl(M_n(\mathbb C)\otimes M_j(\mathbb C)\bigr)\] apply Popa’s theorem with \(B_1=\mathbb C1\), so that \(Q=\mathcal F_0\) and \(D=\mathbb C1\), and with \(\mathcal X=B\cap\ker\tau\). The resulting diffuse abelian algebra is scalar free from \(B\) by the endpoint argument. It contains a Haar unitary \(\widetilde s_0\): successively bisect projections in this diffuse abelian algebra to obtain a selfadjoint element uniformly distributed on \([0,1]\), and exponentiate it by \(t\mapsto e^{2\pi it}\). Choose unitary lifts \(\widetilde s_{0,j}\in M_n(\mathbb C)\otimes M_j(\mathbb C)\).

For each \(j\), choose a unital matrix subalgebra \(R_j^0\cong M_j(\mathbb C)\) of \(M\). Such a subalgebra is obtained by subdividing \(1\) into \(j\) projections of trace \(1/j\) and choosing matrix units using equivalence of equal-trace projections in a finite factor. The coordinate embeddings \(M_n(\mathbb C)\otimes M_j(\mathbb C)\hookrightarrow M_n(M)\) preserve the normalized traces and fix \(B\). They induce a trace-preserving embedding of \(\mathcal F_0\) into \[\mathcal P=P^\omega=M_n(M^\omega).\] Thus the image \(s_0\) of \(\widetilde s_0\) is Haar and scalar free from \(B\), and it has representatives \(s_{0,j}\in\mathcal U(M_n(R_j^0))\).

Freeness from the constant algebra.

Next we move this model into free position relative to the whole constant algebra \(P\). The equality \(P^\omega=M_n(M^\omega)\) follows entry by entry from \(\lVert(x_{ab})\rVert_2^2=n^{-1}\sum_{a,b}\lVert x_{ab}\rVert_2^2\). Since \(M^\omega\) is a factor, \[B'\cap\mathcal P=I_n\otimes M^\omega, \qquad (B'\cap\mathcal P)'\cap\mathcal P=B.\] Put \(T=W^*(P,s_0)\subseteq\mathcal P\) and \(\mathcal X=T\cap\ker E_B\). This is a \(2\)-norm separable selfadjoint linear subspace: the separably acting algebra \(P\) is countably generated as a von Neumann algebra, and adjoining one unitary preserves this property. Equivalently, bounded polynomial approximations in a countable generating set give \(2\)-norm density. Apply Popa’s theorem with \(B_1=B\) and this \(\mathcal X\). Choose a Haar unitary \(w\) in the resulting abelian algebra \(A\subseteq I_n\otimes M^\omega\).

Let \(P_1=W^*(B,s_0)\) and \(P_2=wP_1w^*\). Both contain \(B\), since \(w\) commutes with \(B\). Trace testing against \(B\) shows that \[E_B(wyw^*)=E_B(y)\qquad(y\in\mathcal P).\] We claim that \(P_2\) and \(P\) are free over \(B\). In an alternating product of \(B\)-centered elements from these algebras, write each \(P_2\)-letter as \(wyw^*\) with \(y\in P_1\cap\ker E_B\). The resulting word alternates letters \(w,w^*\) from \(A\cap\ker\tau\) with letters in \(\mathcal X\). It has zero \(E_B\) by (16) and the endpoint extensions proved above. A product consisting of one letter has zero expectation by its assumed centering. This proves the claim.

Set \(s=ws_0w^*\) and \(C_s=W^*(s)\). It is Haar, and \(C_s\) is scalar free from \(B\), because conjugation by \(w\) fixes \(B\) and preserves the trace. To deduce scalar freeness from \(P\), first observe that \[ E_B(c_1b_1c_2\cdots b_{m-1}c_m)=0 \tag{17}\] for \(m\geq1\), \(c_i\in C_s\cap\ker\tau\), and \(b_i\in B\cap\ker\tau\). Indeed, test the trace against an arbitrary \(b\in B\) and split \(b\) into its scalar and centered parts; scalar freeness makes both resulting traces zero. By bimodularity, arbitrary \(B\)-letters may also be attached at the two endpoints in (17).

Single letters have trace zero by centering. Now take any mixed alternating scalar-centered word from \(C_s\) and \(P\), and split each \(P\)-letter \(x\) as \[x=(x-E_B(x))+E_B(x).\] The first summand is \(B\)-centered, and the second lies in \(B\cap\ker\tau\). In every expanded term, group maximal blocks of \(C_s\)-letters and the selected \(B\)-letters. Each such block lies in \(P_2\) and is \(B\)-centered by (17), including any \(B\)-letters at the endpoints. These blocks alternate with \(B\)-centered \(P\)-letters, so freeness over \(B\) makes their product have zero expectation. If there is no remaining \(B\)-centered \(P\)-letter, the term is one block covered by (17). Taking traces proves that \(C_s\) and \(P\) are scalar free.

Controlled representatives.

Finally write \(w=I_n\otimes z\) with \(z\in\mathcal U(M^\omega)\), and choose unitary lifts \(z_j\in\mathcal U(M)\). Then \[R_j=z_jR_j^0z_j^*,\qquad s_j=(I_n\otimes z_j)s_{0,j}(I_n\otimes z_j^*)\] have all the asserted properties. For the commutator estimate, average \(h\) over the compact unitary group of \(R_j\): \[h^{(j)}=\int_{\mathcal U(R_j)}u h u^*\,du.\] This operator commutes with \(R_j\), and \(\lVert h-h^{(j)}\rVert\leq\delta\). Therefore \(I_n\otimes h^{(j)}\) commutes with \(s_j\), and \[\lVert[h_n,s_j]\rVert =\lVert[I_n\otimes(h-h^{(j)}),s_j]\rVert \leq2\lVert h-h^{(j)}\rVert\leq2\delta.\] ◻

Lemma 14 (Unitary twists of a free Haar unitary). Let \((F,\tau)\) be a finite tracial von Neumann algebra, let \(P\subseteq F\) be a unital von Neumann subalgebra, and let \(s\in F\) be Haar and scalar free from \(P\). For every \(U\in\mathcal U(P)\), both \(sU\) and \(Us\) are Haar and scalar free from \(P\).

Proof. Put \(v=sU\). It suffices first to show \[\tau(v^k)=0\quad(k\neq0),\qquad \tau(v^{k_1}x_1\cdots v^{k_r}x_r)=0\] for \(r\geq1\), nonzero integers \(k_i\), and \(x_i\in P\cap\ker\tau\). These cyclic mixed words suffice: if an alternating word starts and ends in the same algebra, cyclically bring those two letters together, split their product into its scalar and centered parts, and induct on word length. Regard an expanded trace word as a cyclic sequence of letters \(s\) and \(s^*\) separated by elements of \(P\). The expansion of \((sU)^k\) uses \((sU)^k\) itself for \(k>0\) and \((U^*s^*)^{-k}\) for \(k<0\). At a transition from an \(s\)-letter to an \(s^*\)-letter the separator is \(Ux_iU^*\); at a transition from \(s^*\) to \(s\) it is \(x_i\). Thus every separator between opposite signs is centered, including the cyclic junction.

Split every other separator into its scalar and centered parts. In each resulting term, removal of a scalar separator merges only powers of \(s\) with the same sign. Consequently every surviving power has nonzero exponent, and every surviving separator is a centered element of \(P\). If any separator remains, freeness of \(s\) from \(P\) gives trace zero. If no separator remains, all signs were the same and the term is a scalar multiple of a nonzero power of \(s\), again with trace zero. The argument for \(\tau(v^k)\) is identical, with all signs the same from the outset. This proves the displayed identities.

Centered Laurent polynomials in \(v\) are \(2\)-norm dense, with bounded approximants, in the centered part of \(W^*(v)\): use the Fejér sums of the corresponding bounded functions on the circle. The displayed identities therefore imply scalar freeness of \(W^*(v)\) and \(P\). Here bounded \(2\)-norm approximation can be passed through each trace word by telescoping and the inequality \(\lvert\tau(a y b)\rvert\leq\lVert a\rVert\lVert y\rVert_2\lVert b\rVert\). Finally apply the proved assertion to \(s^*\) and \(U^*\) and take adjoints to obtain the assertion for \(Us\). ◻

Randomized coefficients

We now prove 8. Fix \(M\subseteq\mathcal B(H)\) and selfadjoint \(h\), and set \(\delta=\lVert\operatorname{ad}_h|_M\rVert\). We may assume \(\delta>0\). It suffices first to bound the commutator with an arbitrary unitary at an arbitrary matrix level. Fix \(n\ge2\), \(U\in\mathcal U(P)\), where \(P=M_n(M)\), and a unit vector \(X\in H^n\). Write \[h_n=I_n\otimes h,\qquad Y=UX,\qquad B=M_n(\mathbb C)\otimes1.\] The following order of choices will be used throughout this proof: \[R_*=100\pi,\qquad N\ge100nR_*^3,\qquad k\ge N,\qquad d=nk. \tag{\(\ast\)}\] Here \(N\) is an integer, and then \(k\) is chosen so large that 12 holds at every length \(1\le r\le N\). Choose a unital matrix algebra \(R\cong M_k(\mathbb C)\) in \(M\). Such an algebra is obtained by subdividing the identity into \(k\) equivalent projections of trace \(1/k\) and choosing matrix units. Let \(e_1,\ldots,e_d\) be the diagonal minimal projections of \(D=M_n(R)\).

Take the free unitary \(s=(s_j)_\omega\) of 13. Each \(s_j\) lies in \(M_n(R_{0,j})\) for a unital matrix algebra \(R_{0,j}\subset M\), and therefore \[ \lVert[h_n,s_j]\rVert\le2\delta. \tag{18}\] Put \(v_j=s_jU\) and \(v=sU\). By 14, \(v\) is a Haar unitary free from the constant algebra \(P\).

Fix an integer \(q>2N\). Independently and uniformly choose \(q\)-th roots of unity \(\theta_1,\ldots,\theta_d\), and set \[t=\sum_{a=1}^d\theta_ae_a,\qquad L_j=tv_j.\] For \(1\le r\le N\), consider the unitary expressions \[Q^+_{r,j}=(tv_j)^{r-1}t,\qquad Q^-_{r,j}=(v_j^*t^*)^r.\] In either expression, group its phase monomials by the multiset \(W\) of the \(r\) labels that occur. Thus \[Q^\pm_{r,j}=\sum_{|W|=r}\chi^\pm_W(\theta)K^\pm_{W,j}, \qquad \chi^\pm_W(\theta)=\prod_{a\in W}\theta_a^{\pm1}.\] Multiplicities are included in the product. Within either sign, these characters are distinct and orthogonal, since \(q>2N\). For the next estimates we suppress the superscript and \(r\). Orthogonality and unitarity give \[ \sum_WK_{W,j}^*K_{W,j}=I. \tag{19}\] Moreover, \[ K_{W,j}=K_{W,j}p_W,\qquad p_W=\sum_{a\in W\text{ distinct}}e_a, \tag{20}\] because the rightmost factor carrying a phase is a diagonal projection with a label in \(W\).

Their classes \(K_W=(K_{W,j})_\omega\) have norm at most \(\beta_W\) from 11. Indeed, moving all powers of \(v\) to the left rewrites each coefficient, up to a left unitary, as the corresponding sum of projection words from distinct algebras \(v^{-l}Dv^l\). These algebras are scalar free: in a product of centered letters from alternating conjugates, consecutive unequal exponents leave nonzero powers of the Haar unitary between centered \(P\)-letters; scalar freeness of \(v\) and \(P\) makes its trace zero. Consequently, 12 gives \[ \max_a\sum_{\substack{|W|=r\\a\in W}}\beta_W^2\le64 \qquad(1\le r\le N). \tag{21}\]

The coordinate coefficients need not themselves obey their limiting norm bounds. We enforce those bounds by spectral truncation. For \(b>0\), put \[f_b(u)=\min\{1,b/u\}\quad(u>0),\qquad f_b(0)=1,\] and define \[K'_{W,j}=K_{W,j}f_{\beta_W}(|K_{W,j}|),\qquad E_{W,j}=K_{W,j}-K'_{W,j}.\] Then \(\lVert K'_{W,j}\rVert\le\beta_W\), its right support is still contained in \(p_W\), and \[ \sum_W(K'_{W,j})^*K'_{W,j}\le I,\qquad \sum_WE_{W,j}^*E_{W,j}\le I. \tag{22}\] These follow term by term from [eq:coefficient-column]. Continuous functional calculus in the quotient gives \[\lim_{j\to\omega}\lVert E_{W,j}\rVert_2=0,\] since the positive function \((u-\beta_W)_+\) vanishes on the spectrum of \(|K_W|\).

The purpose of these truncated coefficients is to replace a matrix commutator by a column commutator with a small support. Average \(h\) over \(\mathcal U(R)\) with Haar probability measure, obtaining \[h^0=\int_{\mathcal U(R)}uhu^*\,du\in R',\qquad \lVert h-h^0\rVert\le\delta,\qquad \lVert\operatorname{ad}_{h^0}|_M\rVert\le\delta.\] Write \(h_n^0=I_n\otimes h^0\). The last inequality follows by averaging \([uhu^*,a]=u[h,u^*au]u^*\). For each \(W\), there is a rectangular partial isometry \(V\in M_{n,1}(R)\) with \(VV^*=p_W\): the support of \(p_W\) has at most \(r\le k\) minimal diagonal projections, so its rank does not exceed the dimension \(k\) of the source column. Its entries commute with \(h^0\), and hence \[[h_n^0,K'_{W,j}]Z =\bigl(h_n^0K'_{W,j}V-K'_{W,j}Vh^0\bigr)V^*Z.\] 9 bounds this by \(8\delta\beta_W\lVert p_WZ\rVert\). Orthogonality of the diagonal projections and [eq:beta-through-label] give the weighted support estimate \[\sum_W\beta_W^2\lVert p_WZ\rVert^2 =\sum_{a=1}^d\left(\sum_{\substack{|W|=r\\a\in W}}\beta_W^2\right) \lVert e_aZ\rVert^2 \le64\lVert Z\rVert^2.\] Squaring the column bounds, summing over \(W\), and taking a square root therefore yields \[\left(\sum_W\lVert[h_n^0,K'_{W,j}]Z\rVert^2\right)^{1/2} \le64\delta\lVert Z\rVert.\] Replacing \(h^0\) by \(h\) costs at most \(2\delta\lVert Z\rVert\), by the first column contraction in [eq:clipped-columns]. Thus \[ \left(\sum_W\lVert[h_n,K'_{W,j}]Z\rVert^2\right)^{1/2} \le66\delta\lVert Z\rVert. \tag{23}\]

For \(Z=X\), restoring the coefficients \(E_{W,j}\) costs a quantity tending to zero along \(\omega\), by 10. There are only finitely many \(W\). For the moving vector \(Z=v_jX=s_jY\), write \[[h_n,E_{W,j}]s_jY =h_nE_{W,j}s_jY-E_{W,j}s_jh_nY -E_{W,j}[h_n,s_j]Y.\] The first two terms tend to zero in column norm: \(E_{W,j}s_j\) is bounded and 2-small, and \(Y,h_nY\) are fixed. The last term has column norm at most \(2\delta\), by [eq:clipped-columns,eq:cheap-lift-commutator]. Phase orthogonality now converts these column bounds back to root-mean-square bounds for the unitary words: \[\begin{align*} \lim_{j\to\omega} \left(\mathbb E_\theta\lVert[h_n,Q^+_{r,j}]v_jX\rVert^2\right)^{1/2} &\le68\delta,\tag{24}\\ \lim_{j\to\omega} \left(\mathbb E_\theta\lVert[h_n,Q^-_{r,j}]X\rVert^2\right)^{1/2} &\le66\delta. \tag{25}\end{align*}\] Both bounds hold for every \(r\le N\).

Recovering a scalar operator

The word estimates control commutator errors independently of the matrix size. We now combine the words into rows and choose a contraction of \(M\) so that their averaged product converges weakly to a positive scalar multiple of the identity. This scalar will detect the matrix commutator.

Independently choose a uniform \(q\)-th root of unity \(z\) and a uniform standard basis column \(\alpha\in\mathbb C^n\). Write \(p=\alpha\alpha^*\in B\), and define \[f_j=N^{-1/2}\sum_{r=1}^N(zL_j)^r,\qquad F_j=\sqrt n\,\alpha^*f_j,\qquad G_j=\sqrt n\,\alpha^*f_j^*.\] Here \(F_j,G_j:H^n\to H\) are rows over \(M\). In the remainder of the proof \(\mathbb E\) averages over all three finite random choices \(\theta,z,\alpha\). Averaging first over \(\alpha\) and then using orthogonality of \(z,\ldots,z^N\) gives \[ \mathbb E\lVert F_jX\rVert^2=\mathbb E\lVert G_jX\rVert^2=1. \tag{26}\]

The same orthogonality removes cross terms between different powers in the following errors: \[\begin{align*} \lim_{j\to\omega} \left(\mathbb E\lVert hF_jX-F_jU^*h_nY\rVert^2\right)^{1/2} &\le70\delta,\tag{27}\\ \lim_{j\to\omega} \left(\mathbb E\lVert hG_jX-G_jh_nX\rVert^2\right)^{1/2} &\le70\delta. \tag{28}\end{align*}\] To check the first estimate, use \(L_j^r=Q^+_{r,j}s_jU\). The corresponding error before the scalar phase is \[[h_n,Q^+_{r,j}]s_jY+Q^+_{r,j}[h_n,s_j]Y,\] which is bounded by [eq:positive-drift] and [eq:cheap-lift-commutator]. For the second estimate, the adjoint power is \(Q^-_{r,j}\), so [eq:negative-drift] applies. Averaging the normalized sum of the \(N\) squared bounds produces no factor depending on \(N\).

We next choose contractions \(x_j\in M\) so that the average of \(G_j^*x_jF_j\) has a nonzero scalar weak limit. Put \[A_j=pf_j^2p,\qquad \widetilde x_j=A_j^*(A_jA_j^*+.01p)^{-1/2} =\alpha x_j\alpha^*.\] The inverse square root is taken in the unital corner \(pPp\). Functional calculus gives \(\lVert x_j\rVert\le1\). Let \[C_j=G_j^*x_jF_j=nf_j\widetilde x_jf_j.\] Since \(\lVert f_j\rVert\le\sqrt N\), these operators have norm at most \(nN\), uniformly in \(j\) and the random choices.

Use the same letters without the subscript for their ultrapower classes. The unitary \(u=ztv\) is Haar and free from \(P\), by 14. Since \(C\in W^*(u,p)\), freeness gives \[ E_P(C)=\lambda p+\mu(1-p). \tag{29}\] For completeness, expand a polynomial in \(u,u^*,p\) into alternating centered letters from \(W^*(u)\) and \(W^*(p)\). Every reduced term containing a centered \(u\)-letter has zero expectation onto \(P\), as is seen by testing its trace against elements of \(P\). The remaining terms lie in \(W^*(p)\). Polynomial approximation, including the continuous corner functional calculus defining \(\widetilde x\), proves [eq:scalar-expectation]. The numbers \(\lambda,\mu\) depend only on \(n,N\): the free joint law of \(u,p\) is the same for all choices, since \(\tau(p)=1/n\). As \(\mathbb E_\alpha p=I/n\), it follows that \[ \mathbb EE_P(C)=c_*I,\qquad c_*=\tau(C). \tag{30}\]

We give the lower bound on \(c_*\), since its independence from \(n\) is essential. Traciality and the support of the corner functional calculus show that \[c_*=n\tau\bigl(AA^*(AA^*+.01p)^{-1/2}\bigr) \ge n\lVert pf^2p\rVert_1-.1.\] The inequality follows from \(t^2(t^2+.01)^{-1/2}\ge t-.1\) for \(t\ge0\), and \(n\tau(p)=1\).

Let \(e\) be the spectral projection of \(u\) for the arc \(\{e^{it}:|t|\le R_*/N\}\), and put \(\eta=\tau(e)=R_*/(\pi N)\). Haar measure gives \(\tau(|f|^2)=1\), and \(\lVert f^2\rVert\le N\). For \(0<|t|\le\pi\), the geometric-sum formula gives \[\left|\sum_{r=1}^Ne^{irt}\right|\le\frac{\pi}{|t|}.\] Integrating its square outside the arc yields \[ \tau(|f|^2(1-e)) \le \frac1{2\pi}\int_{|t|>R_*/N} \frac{\pi^2}{Nt^2}\,dt \le\frac{\pi}{R_*}. \tag{31}\] Since \(f\) commutes with \(e\), tracial Cauchy–Schwarz and freeness of \(p\) from \(f,e\) give \[\begin{align*} n\lVert p(f^2-ef^2e)p\rVert_1 &=n\lVert pf(1-e)fp\rVert_1\\ &\le n\lVert pf(1-e)\rVert_2\lVert(1-e)fp\rVert_2\\ &=\tau(|f|^2(1-e)) \le\frac{\pi}{R_*}. \end{align*}\]

Put \(s_e=nepe\), a positive operator in the corner \(eP^\omega e\). The polar decomposition of \(\sqrt n\,ep\) identifies the nonzero parts of the following two operators, and hence gives \[ n\lVert pef^2ep\rVert_1 =\lVert s_e^{1/2}f^2s_e^{1/2}\rVert_1. \tag{32}\] Explicitly, if \(\sqrt n\,ep=s_e^{1/2}w\), the left-hand operator inside the norm is \(w^*s_e^{1/2}f^2s_e^{1/2}w\), and the middle operator is supported on \(ww^*=\operatorname{supp}(s_e)\) on both sides.

The elementary free moment calculation gives \[\lVert s_e^{1/2}\rVert_2=\sqrt\eta,\qquad \lVert s_e-e\rVert_2^2=(n-1)\eta^2.\] For the second identity, put \(a=np-I\). Then \(\tau(a)=0\), \(\tau(a^2)=n-1\), and freeness gives \(\tau(aeae)=\eta^2\tau(a^2)\); expansion of \(\tau((nepe-e)^2)\) gives the same quantity. Within the corner with identity \(e\), \(|\sqrt t-1|\le|t-1|\) for \(t\ge0\), so \(\lVert s_e^{1/2}-e\rVert_2\le\sqrt{n-1}\eta\). Consequently, \[\begin{align*} \lVert s_e^{1/2}f^2s_e^{1/2}-ef^2e\rVert_1 &\le\lVert s_e^{1/2}-e\rVert_2\lVert f^2\rVert \bigl(\lVert s_e^{1/2}\rVert_2+\lVert e\rVert_2\bigr)\\ &\le2N\sqrt{n-1}\eta^{3/2}. \end{align*}\] Since \(\lVert ef^2e\rVert_1=\tau(e|f|^2)\), we obtain \[ c_*\ge .9-\frac{2\pi}{R_*} -2N\sqrt{n-1}\left(\frac{R_*}{\pi N}\right)^{3/2} \ge\frac12. \tag{33}\] Indeed, by [eq:amplification-choices], the last term is at most \(.2/\pi^{3/2}\), and \(2\pi/R_*=.02\).

For each fixed random choice, 10 identifies the weak ultralimit of \(C_j\) with \(E_P(C)\). All random spaces are finite, and their sizes are fixed before taking the ultralimit. Thus [eq:scalar-average] gives \[ \underset{j\to\omega}{\operatorname{WOT}\!-\!\lim}\,\mathbb EC_j=c_*I. \tag{34}\]

Extracting the matrix commutator

The randomized rows now have two properties: their commutator drift is bounded independently of \(n\), and their product with a contraction \(x_j\in M\) has a positive scalar weak limit. Apply the defining scalar commutator bound to that contraction: \[\left|\mathbb E\langle [h,x_j]F_jX,G_jX\rangle\right| \le\delta\,\mathbb E\bigl(\lVert F_jX\rVert\lVert G_jX\rVert\bigr)\le\delta,\] using [eq:row-rms]. As \(h=h^*\), the expression inside the average equals \[\langle x_jF_jX,hG_jX\rangle-\langle x_jhF_jX,G_jX\rangle.\] Substitute the main terms from [eq:forward-row-drift,eq:backward-row-drift]. By Cauchy–Schwarz and [eq:row-rms], the sum of the two average errors has ultralimit at most \(140\delta\). The remaining expression is \[\langle C_jX,h_nX\rangle-\langle C_jU^*h_nY,X\rangle.\] [eq:weak-scalar-limit] therefore implies \[c_*\left|\langle h_nX,X\rangle-\langle h_nUX,UX\rangle\right| \le141\delta.\] Here the quadratic forms are real because \(h_n\) is selfadjoint. By [eq:positive-scalar], their difference has absolute value at most \(282\delta\). The choices of randomization were allowed to depend on \(U,X,n\), but this bound does not. Taking the supremum over unit vectors gives \[\lVert h_n-U^*h_nU\rVert\le282\delta,\qquad \lVert[h_n,U]\rVert\le282\delta.\] Every selfadjoint contraction \(a\) in a unital \(C^*\)-algebra is the average of the unitaries \(a\pm i(1-a^2)^{1/2}\). Writing a general contraction as its real part plus \(i\) times its imaginary part therefore bounds its commutator by \(564\delta\). The level \(n=1\) is already covered by the definition of \(\delta\). Taking the supremum over \(n\) proves 8, for example with \(C_{\mathrm{amp}}=600\).

Boundary averaging and measurable correction

The finite-factor argument begins with nearby unitaries whose products are only approximately respected. For their countable indexing group \(G\), we construct a probability space \((D,\nu)\) with a nonsingular \(G\)-action and an equivariant mean that corrects these multiplication errors. On \(D^2\) two actions have different roles. The separate action \((g,h)(x,y)=(gx,hy)\) permits averaging of cocycles. The diagonal action \(g(x,y)=(gx,gy)\) has a rigidity property: bounded measurable vector fields transforming under a unitary representation are constant. For cocycles with values in a finite algebra, a tracial convex-hull construction will connect these two uses of \(D^2\). All estimates are independent of \(G\).

For a nonsingular action of a countable group \(G\) on a probability space \((D,\nu)\), write \[\alpha_g F(x)=F(g^{-1}x),\qquad (\lambda_g F)(t,x)=F(g^{-1}t,g^{-1}x).\] Nonsingular means that the action preserves the class of null sets; measure preservation is not required. On \(G\times D\) we use counting measure in the first variable. A bounded operator field on a separable Hilbert space \(K\) means an essentially bounded, weakly measurable map into \(\mathcal B(K)\). Its norm is the essential supremum of its operator norms. Separability makes this a measurable function: the operator norm can be tested on countably many unit vectors. Products of such fields are again weakly measurable, since their values on fixed vectors are measurable \(K\)-valued maps and can be approximated by countably valued maps.

The scalar mean below is the conditional-expectation formulation of amenability for a measured group action (Zimmer 1978; Adams et al. 1994); see in particular (Adams et al. 1994, Theorem 3.4, p. 810). We give the construction and the operator-field extension required here.

Lemma 15 (Equivariant boundary means). For every countable group \(G\) there are a standard probability space \((D,\nu)\) with a nonsingular \(G\)-action and a positive unital contraction \[m:L^\infty(G\times D)\longrightarrow L^\infty(D)\] satisfying \[ m(\lambda_g F)=\alpha_g(mF),\qquad m(fF)=f\,m(F) \quad(f\in L^\infty(D)). \tag{35}\] There is a mean \(m^{(2)}\) with the same properties for the separate \(G\times G\)-action on \(D\times D\) and functions on \(G^2\times D^2\). Each mean extends to a contraction on bounded operator fields on any separable \(K\). The extension preserves adjoints and von Neumann algebra values, and is a bimodule map over bounded operator fields independent of the group variable: \[ m(BFC)=B\,m(F)\,C. \tag{36}\]

Proof. Choose a symmetric probability measure \(\mu\) on \(G\) with \(\mu(g)>0\) for every \(g\). Let \(X=G^{\mathbb N\cup\{0\}}\) with product probability \(\mathbb P=\mu^{\mathbb N\cup\{0\}}\). On \(X\) define \[g(\omega_0,\omega_1,\ldots)=(g\omega_0,\omega_1,\ldots),\qquad S(\omega_0,\omega_1,\ldots)=(\omega_0\omega_1,\omega_2,\ldots).\] These maps commute. The first-coordinate laws of their pushforwards are respectively \(g\mu\) and \(\mu*\mu\), while all later coordinates remain independent with law \(\mu\). Full support therefore makes both pushforwards equivalent to \(\mathbb P\).

Let \(\mathcal I\) be the completed sigma algebra of sets \(E\) satisfying \(S^{-1}E=E\) modulo null sets. It is countably generated modulo null sets. Indeed the measure algebra of \(X\) is separable for the metric \(\mathbb P(E\mathbin\triangle F)\), so its subspace \(\mathcal I\) is separable. Choose a countable dense family \(E_j\) in that subspace. For each \(E\in\mathcal I\), a sequence from this family with summable symmetric-difference errors converges almost everywhere in indicators to \(1_E\). Thus the \(E_j\) generate \(\mathcal I\) after completion. The \(G\)-action preserves \(\mathcal I\), because it commutes with \(S\). Include all translates \(gE_j\) and set \[b(\omega)_{j,g}=1_{gE_j}(\omega),\qquad D=\{0,1\}^{\mathbb N\times G},\qquad \nu=b_*\mathbb P.\] Equip \(D\) with its Borel probability measure and, when taking measurable functions, its completion. The coordinate-permutation action \((h x)_{j,g}=x_{j,h^{-1}g}\) satisfies \(b(h\omega)=h b(\omega)\). It is nonsingular, since the action on \(X\) is nonsingular. We have \(b\circ S=b\) almost everywhere. Pullback by \(b\) identifies \(L^\infty(D)\) with the \(S\)-fixed subalgebra of \(L^\infty(X)\): one inclusion follows from \(b\circ S=b\), and the other from the generation of \(\mathcal I\). This description also applies to scalar \(L^2\) functions. Countably many representative identities can always be made simultaneous by deleting a null set.

For \(F\in L^\infty(G\times D)\) put \[(JF)(\omega)=F(\omega_0,b(\omega)),\qquad T_nF=\frac1n\sum_{j=0}^{n-1}(JF)\circ S^j.\] The map \(J\) is well defined on equivalence classes: there are only countably many group coordinates, and each section has a \(\nu\)-null exceptional set. These maps are positive unital contractions. Take a simultaneous pointwise weak-star cluster subnet of the maps \(T_n\), using the product of the weak-star compact balls \(\prod_F\{a\in L^\infty(X):\lVert a\rVert\le\lVert F\rVert\}\). Denote its limit by \(T\); it is linear, positive, unital and contractive.

Pullback by a nonsingular measurable map is weak-star continuous on \(L^\infty\). To see this directly, for \(q\in L^1\) the measure \(E\mapsto\int_{S^{-1}E}q\,d\mathbb P\) is absolutely continuous with respect to \(\mathbb P\); its Radon–Nikodym density is the required preadjoint applied to \(q\). Its \(L^1\) norm is at most \(\lVert q\rVert_1\). Consequently the telescoping bound \[\lVert(T_nF)\circ S-T_nF\rVert\le 2\lVert F\rVert/n\] implies \((TF)\circ S=TF\). Write \(TF=(mF)\circ b\). Moreover \(J\lambda_g F=(JF)\circ g^{-1}\) and the group action commutes with \(S\). Weak-star continuity of its nonsingular pullback gives the first identity in (35). The identity \(b\circ S=b\) gives \(T_n(fF)=(f\circ b)T_nF\); multiplication by a fixed bounded function is weak-star continuous, proving the second identity.

For the product mean use \(X^2\), its two commuting shifts \(S_1,S_2\), and rectangular averages \[\frac1{nm}\sum_{i=0}^{n-1}\sum_{j=0}^{m-1} F\bigl((S^i\omega)_0,(S^j\omega')_0,b(\omega),b(\omega')\bigr), \qquad n,m\longrightarrow\infty.\] A simultaneous weak-star cluster is fixed by both shifts. These common fixed functions are exactly the functions of \((b(\omega),b(\omega'))\). For completeness, a fixed scalar \(L^2(X^2)\) function has almost every section in the closed subspace \(L^2(\mathcal I)\subset L^2(X)\) in each variable, by Fubini and the preceding identification. The two orthogonal projections onto \(L^2(\mathcal I)\otimes L^2(X)\) and \(L^2(X)\otimes L^2(\mathcal I)\) commute; their product has range \(L^2(\mathcal I)\otimes L^2(\mathcal I)\). This proves the assertion. The preceding equivariance and module arguments now apply in both coordinates.

Here are details of the operator-field extension; they also explain why no normality of \(m\) is needed. On a countable rational complex vector subspace dense in \(K\), define the coefficients of \(m(F)(x)\) by applying the scalar \(m\) to \(\langle F(t,x)\xi,\eta\rangle\). Outside a single null set these coefficients are sesquilinear and bounded by \(\lVert F\rVert\lVert\xi\rVert\lVert\eta\rVert\). They therefore define a bounded operator, and extension by density supplies all coefficients. Positivity of the scalar mean implies that this extension preserves adjoints; the norm bound gives contractivity. The coefficient identities also prove equivariance.

The scalar module identity implies the same coefficient formula when the test vectors depend measurably on \(x\). First take countably valued bounded vector fields: on each member of the countable joint partition, localize using the scalar module identity and use the fixed-vector formula. Then approximate arbitrary bounded measurable vector fields uniformly by countably valued ones, using separability of \(K\) and the norm bound. Testing \(BFC\) on fixed vectors with the variable vectors \(C(x)\xi\) and \(B(x)^*\eta\) proves (36). In particular, if \(F\) takes values in a von Neumann algebra \(A\subset\mathcal B(K)\), its mean commutes with each constant operator in \(A'\). Choose a countable strongly dense subset of the unit ball of \(A'\) to make this simultaneous almost everywhere; then \(m(F)(x)\in A''=A\). Every step applies also to \(m^{(2)}\). ◻

The next proof uses the averaging and polar-decomposition iteration that appears in stability of approximate unitary representations (Kazhdan 1982; Burger et al. 2013); compare (Burger et al. 2013, Theorem 3.2 and its proof, pp. 116–117). Here the mean belongs to an amenable action, and the corrected objects are measurable cocycles.

Lemma 16 (Correction of a measurable unitary cocycle). Let \(D\) and \(m\) be as in 15, and let \(A\subset\mathcal B(K)\) be a von Neumann algebra on a separable Hilbert space. Suppose \(V_g\in L^\infty(D,\mathcal U(A))\), \(g\in G\), and \[e=\sup_{g,h\in G}\lVert V_g\alpha_g(V_h)-V_{gh}\rVert_\infty \le\frac1{64}.\] Then there are measurable unitary fields \(W_g\) with \[ W_g\alpha_g(W_h)=W_{gh},\qquad \sup_g\lVert W_g-V_g\rVert_\infty\le 2e. \tag{37}\] In particular \(W_1=I\). The identities hold simultaneously almost everywhere for every \(g,h\), and may be used on a \(G\)-invariant conull subset of \(D\).

Proof. We give a quadratic improvement and then iterate it. For the current family define \[H_g(t,x)=V_t(x)V_{g^{-1}t}(g^{-1}x)^*,\qquad A_g=m_t H_g.\] The defect inequality implies \(\lVert H_g-V_g\rVert_\infty\le e\), hence \(\lVert A_g-V_g\rVert_\infty\le e\) and \(\lVert H_g-A_g\rVert_\infty\le2e\). All \(H_g\) are unitary and all \(A_g\) are contractions. There is the exact factorization \[H_{gh}=H_g\lambda_g(H_h),\qquad m(\lambda_g H_h)=\alpha_g(A_h).\] Using the left and right module identities to cancel the linear terms in the product gives \[\begin{align*} A_{gh}-A_g\alpha_g(A_h) &=m\bigl((H_g-A_g)(\lambda_g H_h-\alpha_g(A_h))\bigr),\\ I-A_g^*A_g&=m\bigl((H_g-A_g)^*(H_g-A_g)\bigr). \end{align*}\] Both right sides have norm at most \(4e^2\). The same calculation gives \(\lVert I-A_gA_g^*\rVert\le4e^2\). Since \(A_g\) is within \(e<1\) of a unitary, it is invertible in \(L^\infty(D,A)\). Its unitary polar part \(V'_g=A_g(A_g^*A_g)^{-1/2}\) satisfies \[\lVert V'_g-A_g\rVert_\infty =\lVert I-|A_g|\rVert_\infty\le4e^2.\] Functional calculus preserves measurability: here it can be obtained by uniform polynomial approximation on a fixed compact interval away from zero. The new defect is at most \(16e^2\), since changing the three terms in \(A_g\alpha_g(A_h)-A_{gh}\) to their polar parts costs at most \(12e^2\). Also \(\sup_g\lVert V'_g-V_g\rVert_\infty\le e+4e^2\).

Starting with \(e_0=e\), the successive defects satisfy \(e_{j+1}\le16e_j^2\le e_j/4\). Thus the families converge uniformly in \(g\) and in essential operator norm to unitary fields \(W_g\). Their defects tend to zero, and the total change is bounded by \[\sum_{j\ge0}(e_j+4e_j^2) \le\frac{17}{16}\sum_{j\ge0}e_j \le\frac{17}{12}e\le2e.\] If \(e=0\), the original family already works. Countability of \(G\) and nonsingularity permit deletion of all exceptional sets and their translates. Finally the cocycle identity with \(g=h=1\) makes the unitary \(W_1\) idempotent, so \(W_1=I\). ◻

Lemma 17 (Averaged intertwiner). Take either \((\Gamma,Z)=(G,D)\) or \((\Gamma,Z)=(G^2,D^2)\), with its mean from 15. Let \(U:\Gamma\to\mathcal U(K)\) be a unitary representation on a separable Hilbert space, and let \(\Sigma_s\in L^\infty(Z,\mathcal U(K))\) be an exact cocycle for the given action. If \[\sup_{s\in\Gamma}\lVert\Sigma_s-U_s\rVert_\infty\le\delta<1,\] there is \(L\in L^\infty(Z,\mathcal U(K))\) such that \[ \lVert L-I\rVert_\infty\le2\delta,\qquad L(z)U_s=\Sigma_s(z)L(s^{-1}z). \tag{38}\]

Proof. Set \(B=m_t(\Sigma_tU_t^*)\). Then \(\lVert B-I\rVert_\infty\le\delta\), so \(B\) is invertible. Equivariance, the module identities and the cocycle identity give \[\Sigma_s\alpha_s(B)U_s^* =m_t\bigl(\Sigma_s\alpha_s(\Sigma_{s^{-1}t}) U_{s^{-1}t}^*U_s^*\bigr) =m_t(\Sigma_tU_t^*)=B.\] Thus \(BU_s=\Sigma_s\alpha_s(B)\). Taking adjoints and products shows \(B^*B=U_s\alpha_s(B^*B)U_s^*\), so the same identity holds after replacing \(B\) by its unitary polar part \(L\). The inequalities \((1-\delta)I\le|B|\le(1+\delta)I\) give \(\lVert L-B\rVert_\infty\le\delta\), proving the stated bound. ◻

Applied on \(D^2\), the preceding lemma gives an intertwiner for the separate action. For the finite-factor application we also need an algebra-valued comparison of the two boundary variables under the diagonal action. The next construction uses the trace in place of an invariant mean for that action.

Lemma 18 (Diagonal comparison in a finite algebra). Let \(A\) be a finite von Neumann algebra with faithful normal tracial state \(\tau\) and separable trace Hilbert space \(L^2(A,\tau)\). Let \(W_g\in L^\infty(D,\mathcal U(A))\) be an exact cocycle. Suppose \[\sup_{g\in G}\mathop{\rm ess\,sup}_{x,y\in D} \lVert W_g(x)W_g(y)^*-I\rVert\le r<1.\] Then there is a measurable field \(a:D^2\to\mathcal U(A)\) such that \[ \lVert a-I\rVert_\infty\le2r,\qquad a(x,y)=W_g(x)a(g^{-1}x,g^{-1}y)W_g(y)^*. \tag{39}\] In particular, if \(\sup_g\lVert W_g-u_g\rVert_\infty\le\delta\) for any constant unitaries \(u_g\) on a common representation space, the conclusion holds with \(\lVert a-I\rVert_\infty\le4\delta\) when \(2\delta<1\).

Proof. For almost every \((x,y)\) let \(C_{x,y}\) be the closed convex hull in \(L^2(A,\tau)\) of the countable set \[\{W_t(x)W_t(y)^*:t\in G\}.\] It has a unique vector \(c(x,y)\) of least Hilbert norm. To prove measurability explicitly, enumerate its rational convex combinations as \(z_j(x,y)\) and put \(d(x,y)=\inf_j\lVert z_j(x,y)\rVert_2\). These are measurable functions: bounded weakly measurable operator fields are strongly measurable as vectors in the separable trace Hilbert space. For each \(n\), choose the first \(j\) with \(\lVert z_j(x,y)\rVert_2^2<d(x,y)^2+2^{-n}\). The choices are measurable. For two such choices, the parallelogram identity and the lower bound \(d\) on the norm of their midpoint give \[\lVert z_j-z_k\rVert_2^2 \le2(\lVert z_j\rVert_2^2-d^2)+2(\lVert z_k\rVert_2^2-d^2).\] They converge pointwise in \(L^2\) to the measurable field \(c\).

Every \(z_j\) is an operator contraction with \(\lVert z_j-I\rVert\le r\). These inequalities persist under bounded weak-operator limits, and weak-operator compactness in the trace representation identifies any such limit with the \(L^2\) limit by its value on the trace vector. Consequently \(c(x,y)\in A\), \(\lVert c(x,y)\rVert\le1\) and \(\lVert c(x,y)-I\rVert\le r\). On bounded subsets, trace-Hilbert measurability implies weak operator measurability in the trace representation: test first on the dense vectors \(p,q\in A\subset L^2(A,\tau)\), where the coefficients are \(\tau(q^*cp)\), and then use the uniform operator bound. Thus \(c\) is also a measurable bounded operator field.

The linear isometry \[z\longmapsto W_g(x)zW_g(y)^*\] from the trace Hilbert space to itself sends \(C_{g^{-1}x,g^{-1}y}\) onto \(C_{x,y}\). Indeed the cocycle identity sends its generator with index \(t\) to the generator with index \(gt\). Uniqueness of the least-norm vector gives the asserted equivariance for \(c\). Since \(\lVert c-I\rVert\le r<1\), its unitary polar part \(a\) is measurable, satisfies \(\lVert a-I\rVert\le2r\), and has the same transformation law: polar decomposition respects multiplication on either side by unitaries. Finally \(\lVert W_g(x)W_g(y)^*-I\rVert\le2\delta\) follows by comparing both \(W_g(x)\) and \(W_g(y)\) with \(u_g\). ◻

It remains to establish the rigidity of the diagonal action. The bilateral-shift argument below proves the unitary-coefficient form of the double-ergodicity theorem for Poisson boundaries (Kaimanovich 2003); see also (Björklund 2014, Theorem 2.1 and Remark 2.2, p. 7). Symmetry of the random-walk measure identifies the forward and reflected boundary laws used in the proof.

Lemma 19 (Double ergodicity with unitary coefficients). For the space \(D\) constructed in 15, let \(\sigma:G\to\mathcal U(K)\) be a unitary representation on a separable Hilbert space. If a bounded measurable map \(f_0:D^2\to K\) satisfies \[f_0(gx,gy)=\sigma(g)f_0(x,y)\quad(g\in G)\] almost everywhere, then \(f_0\) is almost everywhere a constant vector fixed by every \(\sigma(g)\).

Proof. Use bilateral independent coordinates \((\omega_j)_{j\in\mathbb Z}\) with law \(\mu\), and let \(\theta\) be the left shift, so \((\theta\omega)_j=\omega_{j+1}\). Define \[b_+(\omega)=b(\omega_0,\omega_1,\ldots),\qquad b_-(\omega)=b(\omega_{-1}^{-1},\omega_{-2}^{-1},\ldots),\qquad f(\omega)=f_0(b_+(\omega),b_-(\omega)).\] Symmetry of \(\mu\) makes \(b_+\) and \(b_-\) independent with common law \(\nu\). The identities \(b\circ S=b\) and equivariance give \[b_+(\theta\omega)=\omega_0^{-1}b_+(\omega),\qquad b_-(\theta\omega)=\omega_0^{-1}b_-(\omega).\] For the first identity apply \(b\circ S=b\) to the forward sequence; for the second, apply it to \((\omega_0^{-1},\omega_{-1}^{-1},\ldots)\) and then use equivariance. All these identities hold almost everywhere, including after every integer shift. Consequently \[ f(\theta^n\omega) =\sigma(\omega_{n-1}^{-1}\cdots\omega_0^{-1})f(\omega) \quad(n\ge1). \tag{40}\]

Let \(f_m\) be the Hilbert-valued conditional expectation of \(f\) on coordinates \([-m,m]\). Then \(\varepsilon_m=\lVert f-f_m\rVert_2\to0\) by density of finite-coordinate functions in the product \(L^2\) space. Fix \(n>2m+1\) and put \[\begin{align*} Q&=\sigma(\omega_m^{-1}\cdots\omega_0^{-1}),\qquad A=Qf_m,\\ P&=\sigma(\omega_{n-m-1}^{-1}\cdots\omega_{m+1}^{-1}),\\ R&=\sigma(\omega_{n-1}^{-1}\cdots\omega_{n-m}^{-1}),\qquad B=R^{-1}f_m\circ\theta^n. \end{align*}\] Thus \(A\), \(P\) and \(B\) depend respectively on the disjoint coordinate sets \([-m,m]\), \([m+1,n-m-1]\) and \([n-m,n+m]\); they are independent. Equation (40) and unitarity yield \[ \lVert PA-B\rVert_2\le2\varepsilon_m. \tag{41}\] Set \(a_0=\mathbb EA\) and \(b_0=\mathbb EB\). Independence and orthogonality of centered Hilbert-valued random variables give \[\mathbb E\lVert PA-B\rVert^2 =\mathbb E\lVert A-a_0\rVert^2+\mathbb E\lVert B-b_0\rVert^2 +\mathbb E\lVert Pa_0-b_0\rVert^2.\] For example, condition first on \(P\): the centered terms \(P(A-a_0)\) and \(B-b_0\) have zero mean, are independent, and their squared norms are unchanged by \(P\). It follows that \[\lVert A-a_0\rVert_2\le2\varepsilon_m,\qquad \lVert B-b_0\rVert_2\le2\varepsilon_m.\] Undoing the first transport shows \(\lVert f-Q^{-1}a_0\rVert_2\le3\varepsilon_m\). This approximation depends only on coordinates \([0,m]\). Undoing the other transport shows \(\lVert f\circ\theta^n-Rb_0\rVert_2\le3\varepsilon_m\). After shifting back by \(n\), the approximating function depends only on coordinates \([-m,-1]\). Hence \(f\) belongs both to the closed \(L^2\) subspace of functions of the nonnegative coordinates and to that of functions of the negative coordinates. The two sigma algebras are independent, so their common subspace consists of constants: if \(f\) is measurable with respect to the second one, its conditional expectation on the first is its constant expectation, and this conditional expectation also equals \(f\). Write \(f=v\) almost everywhere. The joint law of \((b_+,b_-)\) is \(\nu\times\nu\), so \(f_0=v\) almost everywhere. Its equivariance and nonsingularity imply \(\sigma(g)v=v\) for every \(g\). ◻

Finite factors in their original representations

We now apply the boundary construction to close finite factors. First we place them on one tracial Hilbert space and identify their two actions of a hyperfinite subfactor. This gives the norm control needed for the boundary argument. That argument produces a nearby isomorphism; the amplification estimate then implements it on the original Hilbert space.

Theorem 20 (Uniform stability of finite factors). There are a number \(t_{\mathrm{fin}}>0\) and a nondecreasing function \(\eta_{\mathrm{fin}}:(0,t_{\mathrm{fin}})\to(0,\infty)\), with \(\eta_{\mathrm{fin}}(t)\to0\) as \(t\downarrow0\), having the following property. Let \(H\) be a separable nonzero Hilbert space, let \(M\subseteq\mathcal B(H)\) be a unital finite factor, and let \(N\subseteq\mathcal B(H)\) be a von Neumann algebra with the same identity. If \(d(M,N)<t<t_{\mathrm{fin}}\), then \(N\) is a finite factor and there is a unitary \(u\in\mathcal B(H)\) such that \[uMu^*=N,\qquad \lVert u-I_H\rVert\le\eta_{\mathrm{fin}}(t).\] The threshold and function are independent of \(M,N\) and their representations.

An irreducible hyperfinite subfactor

We use the normalized trace \(\tau\) of a finite factor, with \(\lVert x\rVert_2=\tau(x^*x)^{1/2}\) and trace vector \(\Omega=1\in L^2(M,\tau)\). We use the trace-preserving expectations onto von Neumann subalgebras of a finite algebra. An inclusion \(R\subseteq M\) is irreducible when \(R'\cap M=\mathbb CI\).

Lemma 21. If \(M\subseteq\mathcal B(H)\) is a unital finite factor and \(N\subseteq\mathcal B(H)\) is a von Neumann algebra with \(d(M,N)<1/10\), then \(N\) is a finite factor.

Proof. Choose \(t\) with \(d(M,N)<t<1/10\). If \(v\in N\) were a proper isometry, choose \(x\in M_1\) with \(\lVert x-v\rVert<t\). Then \(x^*x\ge(1-t)^2I\), so the polar part \(w\) of \(x\) is an isometry in \(M\). Finiteness makes \(w\) unitary. Since \(1-t\le |x|\le1\), \(\lVert w-v\rVert\le\lVert w-x\rVert+\lVert x-v\rVert<2t<1\). This is impossible: a proper isometry within distance less than one of a unitary would be invertible. Thus \(N\) is finite.

Suppose that \(q\in Z(N)\) is a projection. Choose \(a\in M_1\) with \(\lVert a-q\rVert<t\). For \(v\in\mathcal U(M)\) choose \(b\in N_1\) with \(\lVert v-b\rVert<t\). Since \(q\) commutes with \(b\), \[\lVert vav^*-q\rVert\le\lVert a-q\rVert+\lVert[v,q]\rVert<3t.\] The trace-Hilbert closed convex hull of the conjugates \(vav^*\) contains \(\tau(a)I\). Indeed, its unique vector of least norm is conjugacy invariant, is bounded by one by weak compactness of the operator ball, and is therefore scalar; its trace is \(\tau(a)\). Consequently \(\lVert q-\tau(a)I\rVert\le3t<1/2\). A nontrivial projection has distance \(1/2\) from the scalars, so \(q\) is zero or one. The spectral projections of every central selfadjoint element are therefore trivial, proving \(Z(N)=\mathbb CI\). ◻

We recall Popa’s irreducible hyperfinite subfactor theorem (Popa 1981) and give the finite-factor construction needed here.

Lemma 22 (Increasing matrix algebras with scalar relative commutant). Every separable-predual \(\mathrm{II}_1\) factor \(M\) contains an irreducible unital hyperfinite \(\mathrm{II}_1\) subfactor \(R\). One can write \(R=(\bigcup_j Q_j)''\), where \(Q_j\) are increasing unital full matrix algebras and their sizes tend to infinity.

Proof. We first make a single selfadjoint element nearly scalar under a matrix average. Let \(P\) be a \(\mathrm{II}_1\) factor, \(y=y^*\in P\), and \(\varepsilon>0\). Choose a spectral step approximation \(y_0=\sum_{i=1}^l\lambda_i p_i\), where \(\sum_i p_i=I\). For a sufficiently large integer \(b\), split \(\lfloor b\tau(p_i)\rfloor\) orthogonal projections of trace \(1/b\) from each \(p_i\). The remaining projection has trace \[\frac{b-\sum_i\lfloor b\tau(p_i)\rfloor}{b}<\frac lb,\] and also splits into projections of trace \(1/b\). The resulting \(b\) equivalent projections extend to matrix units for a unital \(F\cong M_b(\mathbb C)\) in \(P\). Let \(y_1\in F\) be diagonal, equal to \(\lambda_i\) on the pieces taken from \(p_i\) and zero on the remainder. Then \[\lVert y-y_1\rVert_2\le\lVert y-y_0\rVert_2+\lVert y_0\rVert\sqrt{l/b}.\] Haar averaging over \(\mathcal U(F)\) sends \(y_1\) to \(\tau(y_1)I\). Contractivity of this average in \(L^2\) gives \[ \lVert E_{F'\cap P}(y)-\tau(y)I\rVert_2 \le2\lVert y-y_1\rVert_2<\varepsilon, \tag{42}\] by first choosing \(y_0\) and then \(b\ge2\).

Choose an \(L^2\)-dense sequence \((x_j)\) in the selfadjoint unit ball of \(M\), and enumerate the requirements \((j,k)\in\mathbb N^2\). Start with \(Q_0=\mathbb CI\). Given a full matrix algebra \(Q_n\), its relative commutant \(P_n=Q_n'\cap M\) is a \(\mathrm{II}_1\) factor: matrix units identify \(M\) with a matrix algebra over one of its corners and \(P_n\) with that corner. For the next requirement \((j,k)\) apply (42) in \(P_n\) to \(E_{P_n}(x_j)\), with error \(1/k\). If \(F\subseteq P_n\) is the resulting matrix algebra, set \(Q_{n+1}=Q_n\vee F\). These commuting matrix algebras generate a full matrix algebra of at least twice the size of \(Q_n\), and \[P_{n+1}=F'\cap P_n,\qquad E_{P_{n+1}}=E_{P_{n+1}}E_{P_n}.\] Hence \(\lVert E_{P_{n+1}}(x_j)-\tau(x_j)I\rVert_2<1/k\), and all earlier bounds persist under subsequent expectations. For \(R=(\bigcup_nQ_n)''\) the expectation onto \(R'\cap M=\bigcap_nP_n\) thus sends every \(x_j\) to \(\tau(x_j)I\). Density gives \(R'\cap M=\mathbb CI\). In particular \(R\) is a factor; it is finite and has unital matrix subalgebras of unbounded size, so it is of type \(\mathrm{II}_1\). Its increasing matrix presentation makes it hyperfinite. ◻

We will repeatedly use the following consequence of this construction. If \(E_j\) denotes Haar averaging over \(\mathcal U(Q_j)\), then \(E_j\) is the orthogonal projection in \(L^2(M,\tau)\) onto \(L^2(Q_j'\cap M)\). These decreasing subspaces intersect in \(\mathbb CI\), so \[ E_j(x)\longrightarrow\tau(x)I\quad\hbox{in }L^2(M,\tau) \qquad(x\in M). \tag{43}\] Here is a direct justification avoiding any boundedness assumption on a Hilbert-space vector. Decreasing orthogonal projections have a strong Hilbert-space limit. For bounded \(x\), the operators \(E_j(x)\) are uniformly bounded, and every weak-operator cluster belongs to all \(Q_j'\cap M\), hence is scalar; its trace identifies it as \(\tau(x)I\). Applying the cluster to \(\Omega\) identifies the Hilbert-space limit. Density of the bounded vectors then gives the assertion on the whole trace Hilbert space.

Common tracial position

A factor \(P\subseteq\mathcal B(K)\) is in standard tracial position with unit vector \(\Omega\) when \(x\mapsto x\Omega\) identifies \(K\) with \(L^2(P,\tau_P)\). Its conjugation \(J_P\) is determined by \(J_P(x\Omega)=x^*\Omega\).

The representation theorem we use is the following part of Cameron et al. (2014, Lemma 4.10). We state it separately because it changes the Hilbert space and does not itself give the unitary in 20.

Proposition 23 (Common-position theorem). Let \(M,N\) be separable-predual \(\mathrm{II}_1\) factors acting nondegenerately on one Hilbert space, with mutual strict \(g\)-near inclusions, where \(0<g<1.74\cdot10^{-13}\). Suppose that an amenable von Neumann subalgebra \(R\subseteq M\cap N\) satisfies \(R'\cap M\subseteq R\). There are faithful normal representations \(\pi:M\to\mathcal B(K)\) and \(\rho:N\to\mathcal B(K)\), a common unit vector \(\Omega\), and a number \(\beta<50948\sqrt g\) such that:

  1. both images are in standard tracial position with vector \(\Omega\);

  2. \(\pi|_R=\rho|_R\), and the images have mutual strict \(\beta\)-near inclusions;

  3. for original contractions \(x\in M\), \(y\in N\), \[\lVert\pi(x)-\rho(y)\rVert\le\beta+\lVert x-y\rVert;\]

  4. if \(e_R\) is the projection onto \(\overline{\pi(R)\Omega}\), then \((\pi(M)\cup\{e_R\})''=(\rho(N)\cup\{e_R\})''\).

The space \(K\) can be taken separable.

Below, after applying the proposition, we suppress \(\pi,\rho\) when discussing its images. Normalizing near-inclusion approximants gives \[ d(M,N)\le 2\beta. \tag{44}\]

Lemma 24 (The common right action). In the setting of 23, the right actions of \(R\) agree elementwise: \[J_M rJ_M=J_N rJ_N\qquad(r\in R).\] In particular \(T=J_NJ_M-I\) commutes with the left and right actions of \(R\).

Proof. The commutant of \((M\cup\{e_R\})''\) is \(J_MRJ_M\). Indeed, an element of \(M'\) is right multiplication by an element of \(M\). If it commutes with \(e_R\), applying it to \(\Omega\) shows that its right multiplier lies in \(L^2(R)\cap M=R\), where the equality follows from the trace expectation. Conversely right multiplication by \(R\) reduces \(L^2(R)\). The same argument applies to \(N\). Equality of the basic constructions therefore gives \(J_MRJ_M=J_NRJ_N\). For each \(r\in R\) the two displayed operators take \(\Omega\) to \(r^*\Omega\). The vector \(\Omega\) is separating for the common right algebra, since it is cyclic for \(M\), so the operators are equal. Applying these identities twice proves the last assertion. ◻

Lemma 25 (Testing a bimodular operator on bounded algebra elements). Let \(M\) be a \(\mathrm{II}_1\) factor and let \(R=(\bigcup_jQ_j)''\subseteq M\) be as in 22. If a bounded operator \(V\) on \(L^2(M,\tau)\) commutes with left and right multiplication by \(R\), then \[\lVert V\rVert\le384\sup_{a,b\in M_1} \lvert\langle Va\Omega,b\Omega\rangle\rvert.\]

Proof. Denote the supremum by \(K\). Start with bounded \(a,b\in M\) with \(\lVert a\rVert_2,\lVert b\rVert_2\le1\), without an operator-norm bound. By (43), finite convex averages of conjugations by unitaries of \(R\) can bring both \(aa^*\) and \(bb^*\) within one of their scalar traces in \(L^2\). One first takes a sufficiently high Haar average and then approximates its integral by a finite sum; rational weights and repetitions give equal weights. Choose such a family \((d_i)\). Independently choose a family \((e_l)\) so that the averages of \(e_l^*a^*ae_l\) and \(e_l^*b^*be_l\) have the same property. Enumerate the Cartesian product of the two families by \(k=1,\ldots,m\) and put \(a_k=d_i a e_l\), \(b_k=d_i b e_l\). Their covariance averages satisfy \[L_a=\frac1m\sum_k a_ka_k^*,\quad Q_a=\frac1m\sum_k a_k^*a_k,\qquad \lVert L_a\rVert_2,\lVert Q_a\rVert_2\le2,\] and the same bounds hold for \(L_b,Q_b\).

Use the same independent fair signs in the two sums \[A=m^{-1/2}\sum_k\epsilon_k a_k,\qquad B=m^{-1/2}\sum_k\epsilon_k b_k.\] Bimodularity and sign orthogonality give \[ \mathbb E\langle VA\Omega,B\Omega\rangle=\langle Va\Omega,b\Omega\rangle, \qquad \mathbb E\lVert A\rVert_2^2,\ \mathbb E\lVert B\rVert_2^2\le1. \tag{45}\] Put \(\lVert x\rVert_4^4=\tau((x^*x)^2)\). The three pairings in the fourth moment give exactly \[\mathbb E\lVert A\rVert_4^4 =\tau(Q_a^2)+\tau(L_a^2) +\frac1{m^2}\sum_{k,l}\tau((a_k^*a_l)^2) -\frac2{m^2}\sum_k\tau((a_k^*a_k)^2).\] Tracial Cauchy–Schwarz gives \(\lvert\tau((a_k^*a_l)^2)\rvert\le\tau(a_ka_k^*a_la_l^*)\). Dropping the last nonpositive term therefore yields \[\mathbb E\lVert A\rVert_4^4\le\lVert Q_a\rVert_2^2+2\lVert L_a\rVert_2^2\le12, \qquad \mathbb E\lVert B\rVert_4^4\le12.\] If \(A=w|A|\), set \(A_s=w|A|1_{[0,s]}(|A|)\) and define \(B_s\) similarly. These operators have norm at most \(s\), their \(L^2\) norms do not increase, and \[\mathbb E\lVert A-A_s\rVert_2^2\le12/s^2, \qquad \mathbb E\lVert B-B_s\rVert_2^2\le12/s^2.\] Substituting them in (45) and applying Cauchy–Schwarz in probability gives \[\lvert\langle Va\Omega,b\Omega\rangle\rvert \le s^2K+\frac{2\sqrt{12}}s\lVert V\rVert.\] Take \(s=4\sqrt{12}\), and then take the supremum over the dense bounded \(L^2\) unit vectors \(a,b\). We obtain \(\lVert V\rVert\le192K+\lVert V\rVert/2\), proving the assertion. ◻

Lemma 26 (Norm control of tracial conjugations). Let \(M,N\subseteq\mathcal B(K)\) be \(\mathrm{II}_1\) factors in common standard tracial position with vector \(\Omega\). Suppose that \(R\subseteq M\cap N\) is an irreducible hyperfinite subfactor of \(M\), presented by increasing matrix algebras, and that its right actions agree. If \(d(M,N)<b\), then \[\lVert J_M-J_N\rVert\le C_Jb,\qquad C_J=768.\]

Proof. Set \(T=J_NJ_M-I\). The common right-action identity implies that \(T\) commutes with both actions of \(R\). For \(x\in M_1\) choose \(y\in N_1\) with \(\lVert x-y\rVert<b\). Then \[\lVert Tx\Omega\rVert =\lVert J_Nx^*\Omega-x\Omega\rVert \le\lVert(x^*-y^*)\Omega\rVert+\lVert(x-y)\Omega\rVert<2b.\] Thus the supremum in 25 is at most \(2b\). That lemma gives \(\lVert T\rVert\le768b\), and \(\lVert T\rVert=\lVert J_M-J_N\rVert\). ◻

From the boundary to a Jordan isomorphism

We now work entirely in the common tracial representation. Our immediate objective, when \(d(M,N)<b\) is sufficiently small, is to construct a Hilbert-space unitary \(S_0\) fixing \(\Omega\) such that, for every \(u\in\mathcal U(M)\), \[S_0u\Omega=b_u\Omega,\qquad b_u\in\mathcal U(N),\qquad \lVert b_u-u\rVert=O(b),\] with the error uniform in \(u\). Unitaries linearly span \(M\). Linearity of \(S_0\) and the separating property of \(\Omega\) will make the assignment \(u\mapsto b_u\) extend to a unital linear map from \(M\) to \(N\). We will first prove that it preserves the Jordan product \(x\circ y=(xy+yx)/2\); a linear map with this property is called a Jordan homomorphism. After proving surjectivity, Herstein’s orientation theorem (Herstein 1956) will leave only the multiplicative and antimultiplicative possibilities, and closeness will exclude the latter. We record its algebraic proof before constructing \(S_0\). All closeness estimates here are in ordinary operator norm; complete boundedness enters when we return to the original representation.

Lemma 27 (The orientation of a Jordan map into a factor). A surjective linear Jordan homomorphism between unital complex algebras, whose target is a factor von Neumann algebra, is either a homomorphism or an antihomomorphism.

Proof. Write the map as \(\theta:\mathcal A\to N\), where \(N\) is the target factor. Preservation of the Jordan product gives preservation of \(xyx\), since \(xyx=2x\circ(x\circ y)-x^2\circ y\), where \(x\circ y=(xy+yx)/2\). Polarizing in \(x\) gives preservation of \(xyz+zyx\). Fix \(a,b\), and set \[A=\theta(a),\quad B=\theta(b),\quad D=\theta(ab),\quad E=\theta(ba)=AB+BA-D, \qquad h=D-AB,\quad k=D-BA.\] Compute \(\theta(ab\,c\,ba+ba\,c\,ab)\) in two ways. The polarized triple identity gives \(DCE+ECD\), where \(C=\theta(c)\); applying the unpolarized identity first to the inner and then to the outer products gives \(ABCBA+BACAB\). Expanding their equality yields \[hCk+kCh=0\qquad(C\in N),\] where surjectivity allows every \(C\) in the target. Substitute \(C=C_1kC_2\), and use the same identity with \(C_1\) and \(C_2\) to move \(h\) past the adjacent factors. The two resulting terms are equal, so \[hC_1kC_2k=0\qquad(C_1,C_2\in N).\] A factor is prime: \(xNy=\{0\}\) implies \(x=0\) or \(y=0\), since the central supports of the support projections of nonzero \(x,y\) are both one. If \(k\ne0\), primeness first gives \(hC_1k=0\) for every \(C_1\), and then \(h=0\). Thus, for each pair \((a,b)\), either \(\theta(ab)=\theta(a)\theta(b)\) or \(\theta(ab)=\theta(b)\theta(a)\).

For fixed \(a\), the sets of \(b\) satisfying these two respective identities are additive subgroups whose union is the whole additive group. A group cannot be the union of two proper subgroups: choose an element in each but not the other, and their sum belongs to neither. One identity therefore holds for all \(b\). The same two-subgroup argument, now for those \(a\) for which each identity holds for every \(b\), makes one orientation hold globally. ◻

Proposition 28. There are absolute constants \(b_0>0\) and \(C_{\mathrm{st}}>0\) such that the following holds. Let \(M,N\subseteq\mathcal B(K)\) be type \(\mathrm{II}_1\) factors on a separable Hilbert space in common standard tracial position with vector \(\Omega\). Suppose they contain a common hyperfinite \(\mathrm{II}_1\) subfactor \(R\), irreducible in \(M\) and generated by increasing matrix algebras, and that \(J_MrJ_M=J_NrJ_N\) for every \(r\in R\). If \(d(M,N)<b<b_0\), then there is a normal unital star-isomorphism \(\vartheta:M\to N\) satisfying \[\sup_{x\in M_1}\lVert\vartheta(x)-x\rVert\le C_{\mathrm{st}}b.\]

Proof. Choose a countable strongly dense subgroup \(G\subseteq\mathcal U(M)\). For each \(g\in G\) choose a unitary \(v_g\in N\) with \(\lVert v_g-g\rVert\le3b\): a contraction approximant to \(g\) is invertible for \(b<1\), and its polar part has this bound. Thus \[\sup_{g,h\in G}\lVert v_gv_h-v_{gh}\rVert\le9b.\] Apply 16 to these constant fields on the boundary space \((D,\nu)\) of 15. For sufficiently small absolute \(b\) it gives measurable unitary fields \(W_g:D\to\mathcal U(N)\) satisfying \[ W_g(x)W_h(g^{-1}x)=W_{gh}(x),\qquad \sup_g\lVert W_g-g\rVert_{L^\infty(D)}\le C_1b, \tag{46}\] where \(C_1\) is absolute.

To obtain this correspondence on unitary vectors, we will compare the cocycle with the original left and right actions, and then compare its two boundary values within \(N\). The first comparison uses the separate action on \(D^2\). Form its product cocycle and the constant representation \[\Sigma_{g,h}(x,y)=W_g(x)J_NW_h(y)J_N, \qquad U_{g,h}=gJ_MhJ_M.\] The left and right \(N\) actions commute, so \(\Sigma\) is an exact cocycle for the separate \(G^2\) action. By 26, \[\sup_{g,h}\lVert\Sigma_{g,h}-U_{g,h}\rVert_\infty \le(2C_1+2C_J)b.\] The product version of 17 gives a measurable unitary field \(L:D^2\to\mathcal U(K)\), with \(\lVert L-I\rVert_\infty\le C_2b\), such that \[ L(x,y)U_{g,h}=\Sigma_{g,h}(x,y)L(g^{-1}x,h^{-1}y). \tag{47}\] The constants \(C_1,C_2\), and all subsequent \(C_i\) in this proof, are absolute.

The field \(L\) takes values in \(\mathcal U(K)\); its intertwining identity alone does not supply trace vectors of unitaries in \(N\). For this we use the diagonal comparison. Equation (46) gives \[\sup_g\mathop{\rm ess\,sup}_{x,y\in D} \lVert W_g(x)W_g(y)^*-I\rVert\le2C_1b.\] For sufficiently small \(b\), 18 therefore gives a measurable \(a:D^2\to\mathcal U(N)\) with \(\lVert a-I\rVert_\infty\le C_3b\) and \[ a(x,y)=W_g(x)a(g^{-1}x,g^{-1}y)W_g(y)^*. \tag{48}\] We can now combine the two comparisons: the vector field \(f(x,y)=L(x,y)^*a(x,y)\Omega\) is equivariant for the diagonal representation \(g\mapsto U_{g,g}\). Indeed right multiplication by \(W_g(y)^*\) on trace vectors is \(J_NW_g(y)J_N\), so (48) and (47) give \(f(x,y)=U_{g,g}f(g^{-1}x,g^{-1}y)\). By 19, \(f\) is a constant invariant vector. The invariant vectors for \(g\mapsto gJ_MgJ_M\) are exactly \(\mathbb C\Omega\). To see this, strong density extends invariance from \(G\) to \(\mathcal U(M)\); averaging over \(\mathcal U(Q_j)\) acts on \(x\Omega\) as \(E_j(x)\Omega\) and converges to the projection onto \(\mathbb C\Omega\) by (43). Hence \[ L(x,y)^*a(x,y)\Omega=\lambda\Omega, \qquad |\lambda|=1, \qquad |\lambda-1|\le(C_2+C_3)b. \tag{49}\]

Discard one null set and all its countably many \(G^2\) translates, so that the preceding identities hold at a chosen pair \((x,y)\) and at all its translates. Define the Hilbert-space unitary \[S_0=\lambda a(x,y)^*L(x,y).\] It fixes \(\Omega\). Using (47) for \((g,1)\) and (49) at \((g^{-1}x,y)\) gives \[ S_0g\Omega =a(x,y)^*W_g(x)a(g^{-1}x,y)\Omega =b_g\Omega, \tag{50}\] where \(b_g\in\mathcal U(N)\) and \(\lVert b_g-g\rVert\le C_4b\) uniformly in \(g\).

Every \(u\in\mathcal U(M)\) is the strong limit of a sequence \(g_j\in G\). The vectors \(b_{g_j}\Omega=S_0g_j\Omega\) are Cauchy in trace \(L^2(N)\). For unitaries, traciality gives \(\lVert b_{g_j}^*-b_{g_l}^*\rVert_2=\lVert b_{g_j}-b_{g_l}\rVert_2\). Uniformly bounded \(L^2\) convergence in the standard representation implies strong convergence: test first on the dense right-multiplier vectors \(N'\Omega\) and then use boundedness. Thus \(b_{g_j}\) and their adjoints converge strongly, and their limits are a unitary \(b_u\in N\) and its adjoint. We obtain \[S_0u\Omega=b_u\Omega,\qquad \lVert b_u-u\rVert\le C_4b,\] where the norm inequality follows by weak closedness of the appropriate operator ball.

Every element of \(M\) is a linear combination of unitaries, so the rule \[ \vartheta(z)\Omega=S_0z\Omega\qquad(z\in M) \tag{51}\] defines a linear map into \(N\). It is well-defined because \(\Omega\) is separating for \(N\). Writing a contraction as its real and imaginary selfadjoint parts, each an average of two unitaries, gives \[ \vartheta(I)=I,\qquad \lVert\vartheta(z)-z\rVert\le 2C_4b\lVert z\rVert\quad(z\in M). \tag{52}\] In particular \(\vartheta\) is bounded, and it sends every unitary to a unitary. For \(z=z^*\) expand \(\vartheta(e^{itz})^*\vartheta(e^{itz})=I\) in operator norm. The coefficient of \(t\) says \(\vartheta(z)^*=\vartheta(z)\). Applying this also to \(z^2\) and comparing the coefficient of \(t^2\) gives \(\vartheta(z^2)=\vartheta(z)^2\). Complex linearity and polarization show that \(\vartheta\) preserves adjoints and the Jordan product.

When \(2C_4b<1\), (52) makes its range closed and makes \(\vartheta\) injective. Every contraction of \(N\) has distance at most \((1+2C_4)b\) from this range, by the original gap and (52). For \((1+2C_4)b<1\) the range must be all of \(N\): otherwise the quotient map by this proper closed linear subspace has norm one, contrary to that bound.

27 shows that this surjective Jordan map is either multiplicative or antimultiplicative. The latter possibility is excluded by (52). Indeed, for unitaries \(z,w\in M\) antimultiplicativity would give \[\lVert zw-wz\rVert\le\lVert zw-\vartheta(zw)\rVert +\lVert\vartheta(w)\vartheta(z)-wz\rVert \le6C_4b.\] A unital copy of \(M_2(\mathbb C)\) in \(M\) contains anticommuting selfadjoint unitaries, whose commutator has norm two. Taking \(6C_4b<2\) excludes the anti case. Thus \(\vartheta\) is a unital star-isomorphism, and it is normal because a bijective order isomorphism of von Neumann algebras preserves all bounded increasing suprema. This proves the proposition, with \(C_{\mathrm{st}}=2C_4\) and one absolute sufficiently small \(b_0\). ◻

Return to the original Hilbert space

Proof of 20. Choose \(t> d(M,N)\) below the absolute thresholds specified below. By 21, \(N\) is a finite factor. If either factor is a matrix algebra, 2 for the mutual near inclusions gives the required near-identity conjugacy, with an absolute linear bound. We may therefore suppose both factors are of type \(\mathrm{II}_1\).

Choose \(R\subseteq M\) from 22. By 2, there is a unitary \(w\) with \(\lVert w-I\rVert\le150t\) and \(wRw^*\subseteq N\). Replacing \(N\) by \(N_0=w^*Nw\) makes \(R\) common, and \[d(M,N_0)<301t.\] Take \(g=302t\) and apply 23, provided \(g<1.74\cdot10^{-13}\). Write \(M_s=\pi(M)\), \(N_s=\rho(N_0)\). Their common subfactor remains irreducible in \(M_s\) by faithfulness. By 24 its right actions agree, and \(d(M_s,N_s)\le2\beta\), with \(\beta<50948\sqrt g\). Choose a strict standard-gap bound \(b=3\cdot50948\sqrt g\). If \(b<b_0\), 28 supplies a normal star-isomorphism \(\vartheta_s:M_s\to N_s\) with \(\lVert\vartheta_s-\operatorname{id}\rVert\le C_{\mathrm{st}}b\) on \(M_s\).

Transport it to the original representation: \[\theta=\rho^{-1}\vartheta_s\pi:M\longrightarrow N_0.\] For \(x\in M_1\) choose \(y\in(N_0)_1\) with \(\lVert x-y\rVert<g\). The original-contraction comparison in 23 and faithfulness of \(\rho\) give \[\begin{align*} \lVert\theta(x)-x\rVert &\le\lVert\theta(x)-y\rVert+\lVert y-x\rVert\\ &\le\lVert\vartheta_s\pi(x)-\pi(x)\rVert +\lVert\pi(x)-\rho(y)\rVert+g\\ &\le C_{\mathrm{st}}b+\beta+2g. \end{align*}\] Let \(e(t)=C_{\mathrm{st}}b+50948\sqrt{302t}+604t\). This is an absolute function tending to zero.

On the original \(H\oplus H\), consider the faithful normal paired representation \(x\mapsto x\oplus\theta(x)\) of \(M\), and the selfadjoint switch \(s(\xi,\eta)=(\eta,\xi)\). Its scalar commutator norm on that factor is at most \(e(t)\). 8 therefore gives \[\lVert\operatorname{ad}_s|_{\{x\oplus\theta(x):x\in M\}}\rVert_{\mathrm{cb}} \le C_{\mathrm{amp}}e(t).\] At each matrix level the switch commutator has off-diagonal block \([x_{ij}-\theta(x_{ij})]\), so equivalently \(\lVert\theta-\operatorname{id}\rVert_{\mathrm{cb}}\le C_{\mathrm{amp}}e(t)\). For sufficiently small absolute \(t\), 4 gives a unitary \(v\in\mathcal B(H)\) with \(vMv^*=N_0\) and \(\lVert v-I\rVert\le C_{\mathrm{pair}}C_{\mathrm{amp}}e(t)\), for an absolute constant \(C_{\mathrm{pair}}\). Then \(u=wv\) conjugates \(M\) onto \(N\), and \[\lVert u-I\rVert\le150t+C_{\mathrm{pair}}C_{\mathrm{amp}}e(t).\] All thresholds used above are absolute. Besides \(t<1/10\), they are the bounds in [thm:amenable,lem:paired] and the two bounds in [prop:finite-common-position,prop:finite-standard-isomorphism]. Choose \(t_{\mathrm{fin}}\) so that they hold for every \(0<t<t_{\mathrm{fin}}\), and take \(\eta_{\mathrm{fin}}\) to be the maximum of the displayed bound and the matrix-factor bound. These bounds are nondecreasing functions of \(t\) and tend to zero. This establishes the claimed uniformity on the original Hilbert space. ◻

Preparing infinite factors and comparing modular graphs

We now pass from infinite factors to a common modular setting. The input is ordinary closeness on a separable Hilbert space. After one small adjustment, proper infiniteness makes that closeness completely bounded. A fixed hyperfinite tensor factor then supplies an irreducible finite subfactor in a state centralizer. Aligning its left and right actions gives a common cyclic separating vector. The main estimate of this section compares the two Tomita graphs uniformly under rescaling. In 7, this estimate will permit a small change of position after which the modular group of one algebra preserves both algebras.

A fixed tensor factor and common position

Lemma 29 (Amplification by common isometries). Let \(M,N\subseteq\mathcal B(H)\) contain the same unital copy of \(\mathcal B(\ell^2)\). Then \(d_{\mathrm{cb}}(M,N)=d(M,N)\).

Proof. For each \(n\geq1\) choose isometries \(s_1,\ldots,s_n\) in that common copy such that \(s_i^*s_j=\delta_{ij}1\) and \(\sum_i s_is_i^*=1\). The operator \(V:H^n\to H\) given by \(V(\xi_i)=\sum_i s_i\xi_i\) is unitary. For either \(P=M\) or \(P=N\), \[V M_n(P)V^*=P, \qquad V[x_{ij}]V^*=\sum_{i,j}s_ix_{ij}s_j^*.\] Indeed the inverse coefficient formula is \(x_{ij}=s_i^*xs_j\). The same unitary therefore identifies the two amplified unit balls with the two original unit balls. Taking the supremum over \(n\) proves the assertion. ◻

We fix, once and for all, \[\lambda=e^{-1},\qquad \mu=e^{-\sqrt2}.\] For \(0<\nu<1\), let \(\rho_\nu\) be the state of \(M_2(\mathbb C)\) with density \((1+\nu)^{-1}\operatorname{diag}(1,\nu)\). Form a countable infinite tensor product containing infinitely many sites of each of the two types \((M_2(\mathbb C),\rho_\lambda)\) and \((M_2(\mathbb C),\rho_\mu)\), and denote its von Neumann algebra and product state by \((D,\rho)\). We use the tensor product of the site standard representations, on a separable Hilbert space \(K\). The product vector \(\Omega_\rho\) is cyclic for both left and right local matrix actions, and hence is cyclic and separating for \(D\). Thus \(\rho\) is faithful and normal. Both \(D\) and its standard commutant are hyperfinite: their finite site matrix algebras increase strongly to them.

We recall the modular notation. In a standard representation with a cyclic separating vector \(\Omega\) implementing a faithful normal state \(\psi\), the closure of \(x\Omega\mapsto x^*\Omega\) has polar decomposition \(S=J\Delta^{1/2}\). The antilinear isometry \(J\) is the modular conjugation, and \(\sigma_t^\psi(x)=\Delta^{it}x\Delta^{-it}\) is the modular automorphism group. The state determines this action independently of the chosen standard representation; see (Takesaki 2003) for Tomita–Takesaki theory.

For a faithful normal state \(\psi\) on an algebra \(P\), its centralizer is \[P_\psi=\{a\in P:\psi(ax)=\psi(xa)\text{ for every }x\in P\}.\] It is also the fixed algebra of the modular automorphism group \(\sigma^\psi\). We use the following precise continuous-core facts. The modular crossed product \(P\rtimes_{\sigma^\psi}\mathbb R\) is semifinite, with a faithful normal semifinite trace scaled by its dual action (Blackadar 2006, III.4.8.1–4, p. 329). For a type \(\mathrm{III}\) factor \(P\), this crossed product is a factor exactly when \(P\) is of type \(\mathrm{III}_1\); in that case it is of type \(\mathrm{II}_\infty\) (Takesaki 1973, Lemma 8.2 and Corollary 9.7). For a semifinite factor, choose a faithful normal semifinite trace \(\tau\). The modular crossed product is independent, up to isomorphism, of the faithful normal semifinite weight used to form it (Blackadar 2006, III.4.8.1, p. 329). Since \(\sigma^\tau\) is trivial, this crossed product is \(P\mathbin{\overline\otimes}L(\mathbb R)\cong P\mathbin{\overline\otimes}L^\infty(\mathbb R)\) and is not a factor. Here \(L(\mathbb R)\) is the von Neumann algebra generated by translations on \(L^2(\mathbb R)\), identified with \(L^\infty(\mathbb R)\) by Fourier transform. Thus a factor has a factorial modular crossed product precisely when it is of type \(\mathrm{III}_1\).

Lemma 30 (The fixed tensor factor). One has \(D_\rho'\cap D=\mathbb C1\). If \(A\) is any factor with separable predual, then \(A\mathbin{\overline\otimes}D\) is a type \(\mathrm{III}_1\) factor and \[A\mathbin{\overline\otimes}D\cong(A\mathbin{\overline\otimes}D)\mathbin{\overline\otimes}R_\lambda,\] where \(R_\lambda\) is the Powers factor obtained using only \((M_2(\mathbb C),\rho_\lambda)\) sites.

Proof. Swapping two sites of the same type is implemented by a unitary in a finite site algebra. This unitary preserves the product state, so it belongs to \(D_\rho\). Suppose \(x\in D_\rho'\cap D\). Given a finite site element \(a\), choose bounded finite site approximants \(x_j\) with \(\lVert(x-x_j)\Omega_\rho\rVert\to0\). A product of the indicated swaps moves \(a\) to an element \(a_j\) supported away from \(x_j\). Since \(x\) commutes with the implementing unitary and the state is invariant, \[\rho(ax)=\rho(a_jx),\qquad \rho(a_jx_j)=\rho(a)\rho(x_j).\] State Cauchy–Schwarz bounds the error in the first product by \(\rho(aa^*)^{1/2}\lVert(x-x_j)\Omega_\rho\rVert\); the other error tends to zero as well. Hence \(\rho(ax)=\rho(a)\rho(x)\) for every finite site \(a\). Density gives \((x-\rho(x)1)\Omega_\rho=0\), and faithfulness gives \(x=\rho(x)1\). In particular \(D\) is a factor.

Choose a faithful normal state \(\psi\) on \(A\) and put \(P=A\mathbin{\overline\otimes}D\) with state \(\psi\otimes\rho\). In the regular representation of its modular crossed product on \(L^2(\mathbb R,H_P)\), the generators are \[(\pi(x)\xi)(t)=\sigma^{\psi\otimes\rho}_{-t}(x)\xi(t), \qquad (\lambda_s\xi)(t)=\xi(t-s).\] Here \(H_P\) can be the standard space of \(P\). Let \(Z\) be central in the crossed product. After Fourier transform in \(t\), commutation with every \(\lambda_s\) makes \(Z\) a decomposable operator. The crossed product also commutes with the constant representation of \(P'\). Consequently its decomposable field takes values in \(P\). Commutation with the constant operators \(1\otimes D_\rho\) puts this field in \(A\otimes1\): normal slices and \(D_\rho'\cap D=\mathbb C1\) prove this relative-commutant identity.

An off-diagonal matrix unit at a \(\nu\)-site is a modular eigenoperator with frequency \(\log\nu\) or \(-\log\nu\). Its regular representation, after Fourier transform, is that matrix unit followed by a frequency translation. If \(a(\zeta)\otimes1\) denotes the field of \(Z\), commutation therefore gives \[a(\zeta+\log\lambda)=a(\zeta),\qquad a(\zeta+\log\mu)=a(\zeta) \quad\text{almost everywhere}.\] The generated subgroup is countable and dense in \(\mathbb R\). Test the field on countably many vector coefficients. Translation continuity in local \(L^1\) extends its invariance to every real translation, and convolution with compactly supported approximate identities shows that each coefficient is constant. Thus \(Z=a_0\otimes1\) is a constant field. Commutation with \(\pi(A\otimes1)\), using strong continuity of the modular action and a countable strongly dense set, gives \(a_0\in A'\cap A=\mathbb C1\). The core is a factor, so the continuous-core criterion makes \(P\) a type \(\mathrm{III}_1\) factor.

Finally, adjoining another countable family of \(\lambda\)-sites and reindexing that family gives the asserted absorption. All these tensor products preserve the displayed product states. The usual identification of the two-ratio product with the hyperfinite \(\mathrm{III}_1\) factor is also recorded in (Haagerup 1987, Example 1.6, p. 103); the argument above establishes the additional assertion for every \(A\). ◻

For a faithful normal state \(\psi\) on \(P\), its bicentralizer \(\mathrm{B}(P,\psi)\) consists of those \(a\in P\) for which \(\lVert[a,x_n]\rVert_\psi\to0\) whenever \((x_n)\) is a bounded sequence in \(P\) and \(\lVert x_n\psi-\psi x_n\rVert\to0\). Here \(\lVert x\rVert_\psi=\psi(x^*x)^{1/2}\), and the last norm is the predual norm of the functional \(y\mapsto\psi(yx_n-x_ny)\). We use the following consequence of two established results. If a \(\sigma\)-finite type \(\mathrm{III}_1\) factor \(P\) absorbs \(R_\lambda\) for some \(0<\lambda<1\), then its bicentralizer is \(\mathbb C1\) (Marrakchi 2020, Theorem D(ii)). For a type \(\mathrm{III}_1\) factor with separable predual, trivial bicentralizer gives an irreducible hyperfinite \(\mathrm{II}_1\) subfactor \(R\subseteq P\) with a faithful normal conditional expectation \(E:P\to R\). The finite subfactor is constructed in (Haagerup 1987, Lemma 3.8 and the proof of Theorem 3.1, \((2)\Rightarrow(3)\), pp. 142–145): the construction first gives such a subfactor in a nonzero corner, which is isomorphic to \(P\). This statement imposes no injectivity assumption on \(P\). The state \(\tau_R\circ E\) centralizes \(R\), by bimodularity of \(E\).

Definition 31 (Prepared pair). A prepared pair consists of type \(\mathrm{III}_1\) factors \(M,N\subseteq\mathcal B(H)\) on a separable Hilbert space, a common cyclic separating unit vector \(\Omega\), and a hyperfinite \(\mathrm{II}_1\) factor \(R\subseteq M\cap N\) such that

  1. \(R'\cap M=R'\cap N=\mathbb C1\);

  2. the vector state \(\varphi(x)=\langle x\Omega,\Omega\rangle\) centralizes \(R\) when restricted to either algebra;

  3. \(J_MrJ_M=J_NrJ_N\) for every \(r\in R\), where \(J_M,J_N\) are the modular conjugations associated with \(\Omega\).

Proposition 32 (Preparing an infinite pair). There are absolute constants \(c,C>0\) with the following property. Suppose \(M_0,N_0\subseteq\mathcal B(H)\) are factors on a separable Hilbert space, \(M_0\) is infinite, and \(d(M_0,N_0)<t<c\). For the fixed \(D\subseteq\mathcal B(K)\) above, there are unitaries \(u_0\in\mathcal B(H)\) and \(v\in\mathcal B(H\otimes K)\) satisfying \[\lVert u_0-1\rVert+\lVert v-1\rVert\leq Ct\] such that \[M=M_0\mathbin{\overline\otimes}D,\qquad N=v\bigl((u_0N_0u_0^*)\mathbin{\overline\otimes}D\bigr)v^*\] form a prepared pair and \(d_{\mathrm{cb}}(M,N)\leq Ct\).

Proof. A countably decomposable infinite factor is properly infinite and contains a unital copy of \(\mathcal B(\ell^2)\): split its identity into countably many equivalent projections and choose corresponding matrix units (Blackadar 2006, III.1.3.5 and III.1.5.1, pp. 243, 247). 2 embeds this amenable subfactor of \(M_0\) into \(N_0\) by a unitary of norm distance at most \(150t\) from 1. Conjugating \(N_0\) in the inverse direction gives \(u_0\) for which the subfactor is common. This also proves that \(N_0\) is infinite. The adjusted ordinary gap is at most \(301t\). 29 makes this also a cb gap. 6 transfers it to the tensor pair because \(D'\) is hyperfinite. By 30, both tensor factors are of type \(\mathrm{III}_1\) and absorb \(R_\lambda\).

For the moment write \(N_1=(u_0N_0u_0^*)\mathbin{\overline\otimes}D\). The bicentralizer and expected-subfactor results just stated give an irreducible hyperfinite \(\mathrm{II}_1\) factor \(R\subseteq M\) and a faithful normal state \(\varphi\) on \(M\) centralizing \(R\). The given representation of \(M\) has a cyclic separating vector realizing \(\varphi\). Here is the representation point needed for this assertion. A normal representation on a separable space splits as a countable sum of cyclic normal representations. Each cyclic summand embeds in the standard representation by the standard-form vector implementing its vector functional (Haagerup 1975, Lemma 2.10(1), p. 278). A countable amplification of the standard representation is unitarily equivalent to it, since its type \(\mathrm{III}\) commutant contains a countable family of isometries with ranges summing to 1. Finally, every nonzero projection in that countably decomposable type \(\mathrm{III}\) commutant is equivalent to 1 (Blackadar 2006, Theorem III.1.7.9). Thus every nonzero normal representation here is equivalent to the standard one. Transport the standard vector for \(\varphi\) to obtain the desired \(\Omega\) on the actual space \(H\otimes K\). No conjugation of either algebra is made in this choice.

Put \(S=J_MRJ_M\subseteq M'\). Centralization gives \[ r\Omega=J_Mr^*J_M\Omega\qquad(r\in R). \tag{53}\] First apply 2 to \(R\subseteq M\) and its near inclusion in \(N_1\). It gives a unitary \(v_1\) with \(v_1Rv_1^*\subseteq N_1\) and \(\lVert v_1-1\rVert=O(t)\). Put \(N_2=v_1^*N_1v_1\), so \(R\subseteq M\cap N_2\). Cb gaps change by at most twice the norm distance of the adjusting unitary from 1. Thus 3 gives \(d(M',N_2')=O(t)\). Apply 2 to \(S\subseteq M'\) and its near inclusion in \(N_2'\). It gives a unitary \(v_2\) with \[v_2Sv_2^*\subseteq N_2',\qquad \lVert v_2-1\rVert=O(t),\qquad v_2\in(S\cup N_2')''\subseteq R'.\] Set \(N=v_2^*N_2v_2\). Since \(v_2\) commutes with \(R\), this algebra still contains \(R\), and now \(S\subseteq N'\). The algebra \(M\), its state \(\varphi\), and the vector \(\Omega\) have remained fixed. The unitary in the statement is \(v=v_2^*v_1^*\), with \(\lVert v-1\rVert=O(t)\), and the resulting cb gap is \(O(t)\).

We next check irreducibility and the vector, rather than assume they survive an arbitrary small perturbation. Choose increasing matrix subfactors \(R_j\) strongly generating \(R\). Haar averaging over \(\mathcal U(R_j)\) converges weakly on \(M\) to \(x\mapsto\varphi(x)1\). Indeed every bounded cluster point commutes with all \(R_j\), hence is scalar, and the averages preserve \(\varphi\). If a nontrivial projection \(p\in R'\cap N\) existed, a close approximant \(x\in M\) would therefore give \[\operatorname{dist}(p,\mathbb C1)\leq\lVert p-x\rVert=O(t).\] The left side is \(1/2\). Taking \(t\) sufficiently small proves \(R'\cap N=\mathbb C1\). The same argument in the commutants, with \(S\) and the vector state on \(M'\), proves \(S'\cap N'=\mathbb C1\).

Let \(p\) be the projection onto \(\overline{N'\Omega}\). It belongs to \(N\). By (53), this subspace reduces \(R\): for \(r\in R\) and \(x'\in N'\), the vector \(rx'\Omega=x'r\Omega\) lies in \(\overline{N'\Omega}\), and the same holds for \(r^*\). Thus \(p\in R'\cap N\), so \(p=1\). Similarly the projection onto \(\overline{N\Omega}\) belongs to \(S'\cap N'\) and is 1. Hence \(\Omega\) is cyclic and separating for \(N\). The common commuting right action in (53) shows that its state on \(N\) is invariant under conjugation by \(\mathcal U(R)\), and hence centralizes \(R\). The two operators \(J_MrJ_M,J_NrJ_N\in N'\) now agree on \(\Omega\); since \(\Omega\) is separating for \(N'\), they agree everywhere. All conditions in 31 are satisfied. ◻

Flattening rescaled Tomita graphs

Fix a prepared pair on a space denoted again by \(H\). Choose increasing unital matrix subfactors \(R_j\) strongly generating its common subfactor \(R\). For \(P=M,N\), let \(S_P\) be the closure of \(x\Omega\mapsto x^*\Omega\) and write \(S_P=J_P\Delta_P^{1/2}\). For \(c>0\), let \(E_c^P\) be the orthogonal projection in \(H\oplus\overline H\) onto the closure of \[ g_c(x)=(x\Omega,c\,\overline{x^*\Omega}),\qquad x\in P. \tag{54}\] The conjugate Hilbert space makes this a complex linear graph. Our next result gives uniform control even as \(c\) tends to zero or infinity. Bounding \(x\) in the ordinary operator norm directly would lose this uniformity. We instead flatten a graph vector and then replicate it into a row, where cb closeness supplies the needed estimate.

Proposition 33 (Uniform graph comparison). There is an absolute constant \(C\) such that every prepared pair with \(d_{\mathrm{cb}}(M,N)\leq\gamma\) satisfies \[\sup_{c>0}\lVert E_c^M-E_c^N\rVert\leq C\gamma.\]

Proof. For \(s\in R\) put \(R_s=J_Ms^*J_M=J_Ns^*J_N\). On either graph, the operation \(g_c(x)\mapsto g_c(rxs)\) for \(r,s\in\mathcal U(R)\) is the restriction of the common unitary \[(rR_s)\oplus\overline{s^*R_{r^*}} \quad\text{on }H\oplus\overline H.\] Both graph subspaces reduce these unitaries. Thus \(T=(1-E_c^N)E_c^M\) commutes with them.

A bounded element retaining the graph signal.

First take \(c=\sqrt{k}\), \(k\in\mathbb N\). If \(T\ne0\), choose \(x\in M\) with \[\lVert g_c(x)\rVert=1, \qquad\lVert Tg_c(x)\rVert\geq\tfrac12\lVert T\rVert.\] Such vectors are dense in the graph. Write \[\alpha=\varphi(x^*x),\qquad\beta=\varphi(xx^*), \qquad\alpha+k\beta=1.\] Choose independent Haar unitaries \(u_i,v_i\in\mathcal U(R_j)\) and independent signs \(\epsilon_i\in\{-1,1\}\), and put \[x_i=u_ixv_i,\qquad y=m^{-1/2}\sum_{i=1}^m\epsilon_ix_i.\] For fixed unitaries, the vectors \(Tg_c(x_i)\) all have the same norm. The sign second moment and the elementary Hilbert fourth-moment bound give \[\mathbb E_\epsilon\lVert Tg_c(y)\rVert^2=\lVert Tg_c(x)\rVert^2, \qquad \mathbb E_\epsilon\lVert Tg_c(y)\rVert^4\leq3\lVert Tg_c(x)\rVert^4.\] Consequently, with sign probability at least \(1/12\), \[ \lVert Tg_c(y)\rVert\geq\frac{\lVert T\rVert}{2\sqrt2}. \tag{55}\] For completeness, this probability bound follows by applying Cauchy–Schwarz to the part of a nonnegative random variable above half its mean: \(\mathbb P(Z\geq\mathbb EZ/2)\geq(\mathbb EZ)^2/(4\mathbb EZ^2)\).

We need a simultaneous bound on the fourth moments of \(y\). The Haar averages \[\mathcal E_j(a)=\int_{\mathcal U(R_j)}uau^*\,du\qquad(a\in M)\] converge weakly to \(\varphi(a)1\) as in the preceding proof. For distinct \(i,l\), centralizer cycling and Haar invariance give the exact identities \[\begin{align*} \mathbb E\varphi(x_i^*x_i x_l^*x_l) &=\varphi\bigl((x^*x)\mathcal E_j(x^*x)\bigr),\\ \mathbb E\varphi(x_i^*x_l x_l^*x_i) &=\varphi\bigl(x^*\mathcal E_j(xx^*)x\bigr). \end{align*}\] For the first identity, cycle \(v_i^*\) to the right and average over the Haar unitary \(v_iv_l^*\). For the second, cycle \(v_i^*\) and \(v_i\) away and average over \(u_i^*u_l\). Each formula contains just one weakly convergent average. The three sign-pairing terms in \(\varphi((y^*y)^2)\) have expected limits \[\begin{align*} \mathbb E\varphi(x_i^*x_i x_l^*x_l)&\longrightarrow\alpha^2,\\ \mathbb E\varphi(x_i^*x_l x_l^*x_i)&\longrightarrow\alpha\beta,\\ \lvert\mathbb E\varphi(x_i^*x_l x_i^*x_l)\rvert &\leq\mathbb E\varphi(x_i^*x_l x_l^*x_i) \longrightarrow\alpha\beta. \end{align*}\] For the third, state Cauchy–Schwarz gives \(\lvert\varphi(a^2)\rvert\leq\varphi(aa^*)^{1/2}\varphi(a^*a)^{1/2}\), and the two expected factors agree by interchanging \(i,l\). The all-equal indices contribute \(m^{-1}\varphi((x^*x)^2)\). The same calculation for \(yy^*\) interchanges \(\alpha,\beta\). For the fixed \(x,k\), first choose \(m\) so that the weighted diagonal contribution satisfies \[\frac{\varphi((x^*x)^2)+k\varphi((xx^*)^2)}{m}\leq1.\] The limiting weighted off-diagonal contribution is at most \(\alpha^2+k\beta^2+2(k+1)\alpha\beta\leq6\), because \(\alpha+k\beta=1\) and \(k\geq1\). Now choose \(j\) large enough that the weighted off-diagonal contribution is at most 7. Since \(\mathbb E\lVert g_c(y)\rVert^2=1\), these choices give \[ \mathbb E\left[\lVert g_c(y)\rVert^2+ \varphi((y^*y)^2)+k\varphi((yy^*)^2)\right]\leq10. \tag{56}\] Markov’s inequality bounds the probability that the bracket in (56) exceeds \(240\) by \(1/24\). Together with (55), this gives a realization with both properties. Fix such a \(y\), and clip its singular values at a fixed constant \(K_1\): \[z=y f_{K_1}(|y|),\qquad f_b(t)=\min(1,b/t),\quad f_b(0)=1.\] Functional calculus on both sides of the polar decomposition gives \[\lVert g_c(y-z)\rVert^2 \leq K_1^{-2} \left[\varphi((y^*y)^2)+k\varphi((yy^*)^2)\right].\] Choose \(K_1\) once so that \(\sqrt{240}/K_1\leq1/(4\sqrt2)\). Writing \(a_0=1/(4\sqrt2)\), the signal loss satisfies \[\begin{aligned} \lVert Tg_c(z)\rVert &\geq\lVert Tg_c(y)\rVert-\lVert T\rVert\lVert g_c(y-z)\rVert\\ &\geq\left(\frac1{2\sqrt2}-\frac{\sqrt{240}}{K_1}\right)\lVert T\rVert \geq a_0\lVert T\rVert. \end{aligned}\] Thus the choice of \(K_1\) is independent of the size of \(\lVert T\rVert\), and \[ \lVert z\rVert\leq K_1,\quad \lVert Tg_c(z)\rVert\geq a_0\lVert T\rVert,\quad \varphi(z^*z)\leq240,\quad\varphi(zz^*)\leq240/k. \tag{57}\]

A bounded row retaining the normalized signal.

We have produced a bounded element retaining the signal, but a single bounded-element approximation would still cost \(\sqrt{k}\gamma\) on the second graph leg. To remove that factor, replicate \(z\) into the row \[Y=[r_1z\ \cdots\ r_kz],\qquad A=YY^*,\] where the \(r_i\) are independent Haar unitaries in a sufficiently large \(R_j\). For rows of graph vectors use the normalized Hilbert norm \[\lVert(\eta_i)_{i=1}^k\rVert_{\mathrm{av}}^2 =k^{-1}\sum_i\lVert\eta_i\rVert^2.\] Equivariance makes the normalized signal of this row at least \(a_0\lVert T\rVert\) for every choice of the \(r_i\). Put \(\alpha_z=\varphi(z^*z)\) and \(\beta_z=\varphi(zz^*)\). Haar averaging and centralizer cycling give \[\begin{align*} \mathbb E\varphi(A^2)&\longrightarrow k\varphi((zz^*)^2)+k(k-1)\beta_z^2,\\ \mathbb E\left[k^{-1}\sum_i\varphi(Y_i^*AY_i)\right] &\longrightarrow \varphi((z^*z)^2)+(k-1)\alpha_z\beta_z. \end{align*}\] For example, the second off-diagonal term averages \(zz^*\) under \(r_i^*r_l\) between \(z^*\) and \(z\). By (57), the limits are bounded by an absolute constant: use \(\varphi((zz^*)^2)\leq K_1^2\beta_z\) and \(\varphi((z^*z)^2)\leq K_1^2\alpha_z\) for the diagonals. Choose the stage and then a realization with \[ \varphi(A^2)+k^{-1}\sum_i\varphi(Y_i^*AY_i)\leq C_1, \tag{58}\] where \(C_1\) is absolute.

Clip the whole row on the left: \[X=f_{K_2}(A^{1/2})Y,\] so \(\lVert X\rVert\leq K_2\). If \(Q=1-f_{K_2}(A^{1/2})\), then \(Q^2\leq A/K_2^2\) and \(Y-X=QY\). As \(c^2=k\), its normalized squared graph error is \[\begin{align*} k^{-1}\sum_i\lVert g_c(Y_i-X_i)\rVert^2 &=k^{-1}\sum_i\varphi(Y_i^*Q^2Y_i)+\varphi(QAQ)\\ &\leq C_1/K_2^2. \end{align*}\] Choose the absolute \(K_2\) so that \(\sqrt{C_1}/K_2\leq a_0/2\). The same proportional error estimate for the row gives \[ \begin{aligned} \lVert(Tg_c(X_i))_i\rVert_{\mathrm{av}} &\geq\lVert(Tg_c(Y_i))_i\rVert_{\mathrm{av}} -\lVert T\rVert\lVert(g_c(Y_i-X_i))_i\rVert_{\mathrm{av}}\\ &\geq\left(a_0-\frac{\sqrt{C_1}}{K_2}\right)\lVert T\rVert \geq(a_0/2)\lVert T\rVert. \end{aligned} \tag{59}\]

Comparison at one matrix level and all rescalings.

Cb closeness gives a row \(Z\in M_{1,k}(N)\) with \(\lVert X-Z\rVert\leq K_2\gamma\), with arbitrarily small slack if needed. For any row \([e_i]\) of norm at most \(\eta\), \[k^{-1}\sum_i\lVert g_c(e_i)\rVert^2 =k^{-1}\sum_i\varphi(e_i^*e_i)+ \varphi\left(\sum_i e_ie_i^*\right) \leq2\eta^2.\] Here the graph-vector formula itself makes sense for all operators on \(H\), whether or not they lie in \(M\) or \(N\). Because \(g_c(X_i)\in\operatorname{ran}E_c^M\) and \(g_c(Z_i)\in\operatorname{ran}E_c^N\), \[\lVert(Tg_c(X_i))_i\rVert_{\mathrm{av}} =\lVert((1-E_c^N)g_c(X_i))_i\rVert_{\mathrm{av}} \leq\sqrt2K_2\gamma.\] Together with (59), this bounds \(\lVert(1-E_c^N)E_c^M\rVert\) by an absolute multiple of \(\gamma\). If \(T=0\) the same assertion is immediate. Interchanging \(M,N\) gives the reverse bound, hence the claimed projection bound for \(c=\sqrt{k}\).

The antiunitary flip \((\xi,\overline\eta)\mapsto(\eta,\overline\xi)\) takes \(g_c(x)\) to \(c g_{1/c}(x^*)\), so the bound also holds for \(c=1/\sqrt{k}\). For \(c\geq1\) choose \(k=\lfloor c^2\rfloor\); then \(c/\sqrt{k}\in[1,\sqrt2]\). Scaling the second leg by this ratio has condition number at most \(\sqrt2\) and increases either one-sided subspace gap by at most that factor. Reciprocal scaling handles \(0<c<1\). Finally, for two orthogonal projections \(E,F\), \[\lVert E-F\rVert=\max\{\lVert(1-F)E\rVert,\lVert(1-E)F\rVert\}.\] This proves the uniform assertion. ◻

Separated spectral cuts

Corollary 34 (Logarithmic modular cuts). For a prepared pair with \(d_{\mathrm{cb}}(M,N)\leq\gamma\), put \(L_P=\log\Delta_P\), \(P=M,N\). There is an absolute constant \(C\) such that, for every \(a\in\mathbb R\), \[\lVert 1_{(-\infty,a-1/2]}(L_N) 1_{[a+1/2,\infty)}(L_M)\rVert\leq C\gamma.\] The same estimate holds with \(M,N\) interchanged.

Proof. The upper-left block of the projection onto a closed operator graph is \((1+A^*A)^{-1}\). Applied to (54), this gives \((1+c^2\Delta_P)^{-1}\). With \(g(s)=(1+e^s)^{-1}\), 33 consequently yields \[\lVert g(L_M-a)-g(L_N-a)\rVert\leq C\gamma\qquad(a\in\mathbb R).\] Let \(p=1_{(-\infty,a-1/2]}(L_N)\) and \(q=1_{[a+1/2,\infty)}(L_M)\), and regard \(pq\) as an operator \(qH\to pH\). Let \(A\) and \(B\) be the restrictions of \(g(L_N-a)\) and \(g(L_M-a)\) to those respective subspaces. Then \[A\geq g(-1/2)1,\qquad B\leq g(1/2)1, \qquad\lVert A(pq)-(pq)B\rVert\leq C\gamma.\] For \(D=A(pq)-(pq)B\), integration of the derivative of \(e^{-tA}(pq)e^{tB}\) gives \[pq=\int_0^\infty e^{-tA}D e^{tB}\,dt, \qquad \lVert pq\rVert\leq\frac{\lVert D\rVert}{g(-1/2)-g(1/2)}.\] The denominator is a fixed positive number. Reversing the two algebras proves the other estimate. ◻

Aligning the modular action

The preceding section compares the spectral subspaces of two modular operators. We now turn that comparison into an algebra invariant under one modular group. Distances are measured in operator norm, and all constants are absolute. The selected unitaries need not depend continuously on time; the proof specifies the topology whenever continuity is used.

Proposition 35 (A common modular action). There are absolute constants \(c_{\rm mod},C_{\rm mod}>0\) with the following property. Let \((M,N,\Omega,R)\) be a prepared pair as in 31, acting on the same separable Hilbert space \(H\), and suppose that \(d_{\mathrm{cb}}(M,N)<\gamma<c_{\rm mod}\). There is a unitary \(q\in\mathcal B(H)\) such that \[\lVert q-I\rVert\le C_{\rm mod}\gamma, \qquad d_{\mathrm{cb}}(qMq^*,N)\le C_{\rm mod}\gamma, \qquad \Delta_N^{it}(qMq^*)\Delta_N^{-it}=qMq^* \quad(t\in\mathbb R).\] The algebra \(N\) and its modular group are unchanged.

The proof has three steps. A uniform infinitesimal estimate controls norm-continuous paths of approximate normalizers. Analytic smoothing puts the modular comparison into that setting, and a completely positive limit recovers exact normalizers at each time. Finally an averaging iteration corrects their multiplication law and produces the near-identity change of position.

We use two groups associated with a concrete factor \(P\subset\mathcal B(H)\): \[\mathcal G_P=\mathcal U(P)\mathcal U(P'),\qquad \mathcal N_P=\{v\in\mathcal U(\mathcal B(H)):vPv^*=P\}.\] The two factors in a product defining \(\mathcal G_P\) commute, so \(\mathcal G_P\) is a subgroup of \(\mathcal N_P\). We first prove a uniform estimate near this subgroup. This estimate uses only that \(P\) is a separably acting type \(\mathrm{III}\) factor.

Distance to the infinitesimal normalizers

Lemma 36. For a separably acting properly infinite von Neumann algebra \(P\subset\mathcal B(H)\) and \(h=h^*\in\mathcal B(H)\), \[\operatorname{dist}(h,P')\le4\lVert\operatorname{ad}_h|_P\rVert.\]

Proof. Put \(d=\lVert\operatorname{ad}_h|_P\rVert\). Choose a unital copy \(S\) of \(\mathcal B(\ell^2)\) in \(P\), with matrix units \((e_{ij})\) whose diagonal sum is \(I\). The finite-dimensional unital algebras \[S_n=\operatorname{span}\{e_{ij}:1\le i,j\le n\} +\mathbb C\Bigl(I-\sum_{i=1}^ne_{ii}\Bigr)\] increase and generate \(S\). Average \(h\) over \(\mathcal U(S_n)\) and take a weak-operator cluster point \(h_0\). Each average is self-adjoint, within \(d\) of \(h\), and has scalar commutator bound at most \(d\) on \(P\). These properties pass to the cluster point. Haar invariance at every earlier stage gives \(h_0\in S'\).

For any \(k\), choose isometries \(s_1,\ldots,s_k\in S\) with \(s_i^*s_j=\delta_{ij}I\) and \(\sum_i s_is_i^*=I\). The unitary \(V:H^k\to H\), \(V(\xi_i)=\sum_i s_i\xi_i\), identifies \(M_k(P)\) with \(P\) and \(h_0^{(k)}\) with \(h_0\). Consequently \[\lVert\operatorname{ad}_{h_0}|_P\rVert_{\mathrm{cb}}\le d.\] By 3, \(\operatorname{dist}(h_0,P')\le d\), so in fact \(\operatorname{dist}(h,P')\le2d\). The stated larger constant will be convenient. ◻

Lemma 37 (Uniform infinitesimal estimate). There is an absolute \(K<\infty\) such that every separably acting type \(\mathrm{III}\) factor \(P\subset\mathcal B(H)\) and every self-adjoint \(h\in\mathcal B(H)\) satisfy \[ \operatorname{dist}(h,P+P') \le K\sup_{x\in P_1}\operatorname{dist}([h,x],P). \tag{60}\] The sum \(P+P'\) here is a linear subspace; no closedness assertion about that sum is needed.

Proof. Suppose there is no common \(K\). By scaling a counterexample and subtracting a self-adjoint sum from \(P+P'\), we can choose factors \(P_j\subset\mathcal B(H_j)\) and self-adjoint \(h_j\) with \[\operatorname{dist}(h_j,P_j+P_j')=1, \qquad \lVert h_j\rVert\le2, \qquad \varepsilon_j:=\sup_{x\in(P_j)_1}\operatorname{dist}([h_j,x],P_j)\longrightarrow0.\] Taking real parts justifies using a self-adjoint approximating sum. Its subtraction does not change either distance in the assertion, because its commutator with \(P_j\) belongs to \(P_j\).

Fix a free ultrafilter \(\omega\) on \(\mathbb N\) and form the norm ultraproducts \[A=\prod_\omega P_j\subset \prod_\omega\mathcal B(H_j)=E.\] Here a norm ultraproduct is the algebra of bounded sequences modulo the sequences whose operator norms tend to zero along \(\omega\). The algebra \(A\) is simple. Indeed, a nonzero positive element \(a\) has uniformly bounded positive representatives \(a_j\) and a number \(t>0\) such that \(e_j=1_{[t,\infty)}(a_j)\) is nonzero on a set in \(\omega\). In a separably acting type \(\mathrm{III}\) factor each nonzero projection is equivalent to \(I\). Choose isometries \(v_j\) with \(v_jv_j^*=e_j\) on that set. Then \(v_j^*a_jv_j\ge tI\) there. The class \(v^*av\) is invertible in \(A\) and belongs to the ideal generated by \(a\). Thus every nonzero closed ideal of \(A\) contains \(I\).

Commutation with \([h_j]\in E\) defines a bounded derivation \(D:A\to A\). To see that its range is in \(A\), choose, for every bounded sequence \(x_j\in P_j\), elements \(y_j\in P_j\) with \[\lVert[h_j,x_j]-y_j\rVert \le (\varepsilon_j+j^{-1})\sup_i\lVert x_i\rVert.\] The \(y_j\) are uniformly bounded, and their class is the commutator in \(E\). Its definition is therefore independent of the choices; linearity and the derivation identity follow in \(E\).

Sakai’s theorem for simple unital \(C^*\)-algebras gives \(a\in A\) with \(D(x)=[a,x]\) (Sakai 1971, sec. 1, p. 259; Section 2, p. 260). No separability of the norm ultraproduct is required. Since the \(h_j\) are self-adjoint, replacing \(a\) by \((a+a^*)/2\) leaves this identity unchanged. Choose uniformly bounded self-adjoint representatives \(a_j\in P_j\). The equality of the two derivations implies \[\lim_\omega\lVert\operatorname{ad}_{h_j-a_j}|_{P_j}\rVert=0.\] Otherwise contractions witnessing a fixed positive commutator norm could be selected at each coordinate of an ultrafilter-large set, contradicting equality on their class in \(A\). 36 now gives \[1=\operatorname{dist}(h_j,P_j+P_j') \le\operatorname{dist}(h_j-a_j,P_j') \le4\lVert\operatorname{ad}_{h_j-a_j}|_{P_j}\rVert\longrightarrow_\omega0,\] a contradiction. ◻

Lemma 38 (A path of approximate normalizers). There are absolute constants \(e_0,C_{\rm path}>0\) such that the following holds. Let \(P\subset\mathcal B(H)\) be a separably acting type \(\mathrm{III}\) factor, let \(I\subset\mathbb R\) be an interval containing \(0\), and let \(v:I\to\mathcal U(\mathcal B(H))\) be norm continuous with \(v(0)=I\). If \(0\le e<e_0\) and \[\sup_{t\in I}\sup_{x\in P_1}\operatorname{dist}(v(t)xv(t)^*,P)\le e,\] then \[\sup_{t\in I}\operatorname{dist}(v(t),\mathcal G_P)\le C_{\rm path}e.\] The constants are independent of the interval’s length.

Proof. Write \(b(t)=\operatorname{dist}(v(t),\mathcal G_P)\). This is a continuous function, since distance to any nonempty subset is \(1\)-Lipschitz. Suppose \(b=b(t)\) is smaller than a fixed absolute radius, and choose \(p\in\mathcal G_P\) with \(\lVert p^*v(t)-I\rVert<b+\eta\). For sufficiently small \(b+\eta\), principal functional calculus writes \(p^*v(t)=e^{ih}\) with \(h=h^*\) and \(\lVert h\rVert\le2(b+\eta)\). As \(p\) normalizes \(P\), this conjugation has the same one-sided approximation bound \(e\). The twice-integrated conjugation formula gives \[\lVert e^{ih}xe^{-ih}-x-i[h,x]\rVert\le2\lVert h\rVert^2\lVert x\rVert.\] It follows that \[\sup_{x\in P_1}\operatorname{dist}([h,x],P)\le e+2\lVert h\rVert^2.\] Apply 37 and approximate \(h\) by a self-adjoint sum \(a+b'\) with \(a\in P\), \(b'\in P'\). Since \(a\) and \(b'\) commute, \(e^{i(a+b')}=e^{ia}e^{ib'}\) belongs to \(\mathcal G_P\). The inequality \(\lVert e^{ih}-e^{ik}\rVert\le\lVert h-k\rVert\) for self-adjoint \(h,k\) therefore shows, on letting both approximation slacks vanish, that \[ b(t)\le K e+8K b(t)^2 \quad\text{whenever }b(t)\text{ is in that radius}. \tag{61}\] Choose \(r>0\) within the radius and with \(8Kr\le1/4\). If \(Ke<r/2\), the continuous function \(b\), starting at zero, cannot first reach \(r\): (61) would give \(r<3r/4\). Thus \(b(t)<r\) throughout the interval between \(0\) and any chosen \(t\in I\). Absorbing the quadratic term in (61) gives \(b(t)\le2Ke\) uniformly. This proves the assertion with suitably reduced \(e_0\). ◻

Analytic approximation of the modular comparison

The next argument converts separated spectral-cut estimates into actual normalizers. Its analytic approximation takes place in operator norm away from the real axis. Its passage to the real axis uses the strong-* topology at each fixed time; it does not require norm continuity of modular unitaries.

We first record explicitly the analytic fact needed below. On \(\mathcal K=L^2(\mathbb R,H)\) put \(D=-i\,d/dt\), with Fourier-transform convention \[\widehat f(\xi)=(2\pi)^{-1/2}\int_\mathbb Re^{-it\xi}f(t)\,dt, \qquad P_a=1_{(a,\infty)}(D).\] A bounded operator-valued function always means a field measurable on each vector, modulo a common null set, with essentially bounded operator norm. Here and below \(H\) is separable. Integrals of such fields are defined by matrix coefficients, equivalently by integration on each vector when the kernel is integrable. No norm measurability as a \(\mathcal B(H)\)-valued function is required.

Lemma 39 (Analytic extension of a frequency-preserving multiplier). Suppose multiplication by a bounded field \(B(t)\) on \(L^2(\mathbb R,H)\) preserves \(P_0\mathcal K\). Then it is the boundary field of a bounded, operator-norm holomorphic function on the upper half-plane, given by \[ B(t+ir)=\int_\mathbb R\frac{r}{\pi((t-s)^2+r^2)}B(s)\,ds, \qquad r>0. \tag{62}\] Its norm is at most \(\mathop{\rm ess\,sup}_s\lVert B(s)\rVert\). For two such fields, and any constant \(x\in\mathcal B(H)\), the extension of the boundary product \(B(s)xS(s)\) is \(B(z)xS(z)\).

Proof. Write \(w=(t-i)/(t+i)\) and give the unit circle normalized measure \(d\theta/(2\pi)\). The transformation \[(Jf)(t)=\frac{1}{\sqrt\pi(t+i)} f\left(\frac{t-i}{t+i}\right)\] is unitary from \(L^2(\mathbb T,H)\) onto \(\mathcal K\): the change of variables has \(d\theta/(2\pi)=dt/(\pi(1+t^2))\). For \(n\geq0\), \(J(w^n)\) is a linear combination of \((t+i)^{-k}\), \(1\leq k\leq n+1\), whose Fourier transforms are supported in \([0,\infty)\). For \(n<0\) its expression is a linear combination of \((t-i)^{-k}\), whose Fourier transforms are supported in \((-\infty,0]\). These support statements follow directly from \[(t+i)^{-k}=\frac{(-i)^k}{(k-1)!} \int_0^\infty u^{k-1}e^{-u}e^{itu}\,du\] and its complex conjugate, interpreted under the \(L^2\) Fourier transform. Since the images of all circle modes form an orthonormal basis, \(J\) identifies the nonnegative-mode subspace with \(P_0\mathcal K\).

Conjugating the multiplier by \(J\) gives the circle multiplier \(\widetilde B(w)=B(t)\). Apply its invariant-subspace property to every constant vector. All negative Fourier coefficients of \(\widetilde B\) vanish. If \(b_n\) denotes its \(n\)th coefficient, defined by matrix coefficients, then \[F(w)=\sum_{n\geq0} b_nw^n,\qquad |w|<1,\] converges in operator norm locally uniformly and is operator-norm holomorphic. Its value at \(w=\rho e^{i\theta}\) is the circle Poisson integral of \(\widetilde B\), since all its negative coefficients vanish. The Poisson kernel is positive and has integral one, so \(\lVert F(w)\rVert\leq\lVert\widetilde B\rVert_\infty\). Changing variables in the circle Poisson integral gives exactly (62); explicitly, harmonic measure at \(t+ir\) pushes forward to \(r\,ds/(\pi((t-s)^2+r^2))\).

For completeness, multiplication of these extensions can be checked without presupposing pointwise boundary limits. The radial Poisson integrals converge in \(L^2(\mathbb T,H)\) on each fixed vector, for both \(\widetilde B\) and \(\widetilde B^*\), and likewise for \(\widetilde S\) and \(\widetilde S^*\). Choose a sequence of radii tending to one along which these convergences hold pointwise for a countable dense family of vectors. Uniform boundedness extends the pointwise convergence to every vector. Along this sequence the products converge strongly-* almost everywhere to \(\widetilde B(w)x\widetilde S(w)\). Each Fourier coefficient of the bounded analytic product \(F(w)xG(w)\) therefore converges, by dominated convergence on matrix coefficients, to the corresponding coefficient of that boundary product. In particular its negative coefficients vanish, and its nonnegative coefficients are precisely the Taylor coefficients of \(F xG\). This proves the asserted product identity. The same argument applies to products with a scalar bounded analytic function. ◻

Lemma 40 (From modular cuts to near normalizers). Let \(M,N\subseteq\mathcal B(H)\) be a prepared pair on a separable Hilbert space, with \(M\) of type \(\mathrm{III}\) and \(d_{\mathrm{cb}}(M,N)\leq\gamma\). Write \[U_P(t)=e^{itL_P},\qquad L_P=\log\Delta_P\quad(P=M,N), \qquad W(t)=U_N(t)U_M(t)^*.\] Assume the separated-cut estimate of 34, with its universal constant, and the local path estimate of 38. There are universal constants \(\gamma_0>0\) and \(C\) such that, if \(0<\gamma<\gamma_0\), then for every \(t\in\mathbb R\) there is a unitary \(n_t\) satisfying \[n_tMn_t^*=M,\qquad \lVert n_t-W(t)\rVert\leq C\gamma.\] No regularity of the choices \(t\mapsto n_t\) is asserted or needed.

Proof. The letters \(C\) below denote constants depending only on the universal constants in the stated inputs, and can increase between occurrences. All required smallness conditions are met by decreasing one universal \(\gamma_0\) at the end.

The initial approximate normalizers.

Both \(U_M(t)\) and \(U_N(t)\) normalize their respective algebras. For every matrix size \(k\) and \(x\in M_k(M)_1\), first move \(U_M(t)^*xU_M(t)\) into \(M_k(N)_1\), conjugate by \(U_N(t)\), and move back into \(M_k(M)_1\). The two approximation errors are at most \(\gamma\) each, with arbitrarily small slack. Thus \[ \operatorname{dist}\bigl(W(t)^{(k)}xW(t)^{(k)*},M_k(M)_1\bigr)\leq2\gamma. \tag{63}\] Here \(T^{(k)}=I_{\mathbb C^k}\otimes T\). Applying the same two approximations before conjugating by \(U_M(t)\) gives the corresponding assertion for \(W(t)^*xW(t)\). Consequently \[ d_{\mathrm{cb}}\bigl(W(t)MW(t)^*,M\bigr)\leq2\gamma \qquad(t\in\mathbb R). \tag{64}\]

The frequency nest and its signs.

Use the notation \(\mathcal K,D,P_a\) introduced above. Let \(C_s\) denote multiplication by \(e^{ist}\), and let \(\mathcal U_P\) be multiplication by \(U_P(t)\). Spectral calculus gives \[\mathcal U_P^*D\mathcal U_P=D+L_P,\qquad C_s^*DC_s=D+s.\] The sum on the right is the selfadjoint operator obtained from the commuting spectral measures on the two tensor factors; these identities therefore include their domains. They can also be checked in the spectral representation of \(L_P\), where \(\mathcal U_P\) is a scalar frequency translation on each fiber. Put \(\mathcal W=\mathcal U_N\mathcal U_M^*\). Directly moving the spectral projections past these unitaries gives \[ (1-P_a)C_1\mathcal W P_a =C_1\mathcal U_N 1_{(-\infty,a-1]}(D+L_N) 1_{(a,\infty)}(D+L_M)\mathcal U_M^*. \tag{65}\] After Fourier transformation the product in the middle has fiber \[1_{(-\infty,a-1-\xi]}(L_N) 1_{(a-\xi,\infty)}(L_M).\] Its two thresholds have separation one, with midpoint \(a-\tfrac12-\xi\). The cut estimate of 34 bounds its norm by \(C\gamma\), uniformly in \(a\) and \(\xi\). The same argument with \(M,N\) interchanged applies to \(C_1\mathcal W^*\). In particular \[ \sup_a\lVert(1-P_a)C_1\mathcal W P_a\rVert\leq C\gamma, \qquad \sup_a\lVert(1-P_a)C_1\mathcal W^*P_a\rVert\leq C\gamma. \tag{66}\]

Adjoin \(0,I\) to the complete nest \(\{P_a:a\in\mathbb R\}\). Arveson’s distance formula for its nest algebra (Arveson 1975, Theorem 1.1, p. 212) says that the distance of an operator \(T\) to the operators preserving every \(P_a\mathcal K\) equals \(\sup_a\lVert(1-P_a)TP_a\rVert\). Choose preserving operators \(T_B,T_S\) with \[\lVert T_B-C_1\mathcal W\rVert\leq\delta, \qquad\lVert T_S-C_1\mathcal W^*\rVert\leq\delta, \qquad \delta=C\gamma,\] enlarging \(C\) to allow slack in taking the infimum.

Character averaging produces multipliers.

The averaging below parallels Arveson’s construction of an expectation onto multiplication operators that preserves the frequency nest algebra on \(L^2(\mathbb T)\) (Arveson 1975, Proposition 5.1, pp. 228–229). Use the invariant mean on the discrete group \(\mathbb R\) from 7. For \(T\in\mathcal B(\mathcal K)\) put \[\mathcal E(T)=m_s C_sTC_s^*.\] This UCP map has range in the commutant of the scalar characters. Since \(C_s^*P_aC_s=P_{a-s}\), conjugation by \(C_s\) preserves the entire nest algebra. The coefficient average preserves it as well: each identity \((1-P_a)C_sTC_s^*P_a=0\) remains zero after taking the mean. Both target multipliers \(C_1\mathcal W\) and \(C_1\mathcal W^*\) commute with every \(C_s\), so averaging retains both approximation bounds.

The scalar characters generate the scalar multiplication algebra \(L^\infty(\mathbb R)\). Indeed, an \(L^1\) function annihilating all characters has zero Fourier transform and hence is zero. Their commutant on \(L^2(\mathbb R,H)\) is therefore the algebra of decomposable operators. One may see the latter assertion by taking matrix coefficients relative to a countable orthonormal basis of \(H\): each coefficient commutes with scalar multiplication and is itself a scalar multiplier; the operator norm bound on a countable dense set of finite vectors gives a common almost-everywhere bounded operator-valued field. Thus \(\mathcal E(T_B)\) and \(\mathcal E(T_S)\) are multipliers by fields \(B(s),S(s)\) satisfying \[ \mathop{\rm ess\,sup}_s\lVert B(s)-e^{is}W(s)\rVert\leq\delta, \qquad \mathop{\rm ess\,sup}_s\lVert S(s)-e^{is}W(s)^*\rVert\leq\delta. \tag{67}\] They preserve \(P_0\mathcal K\), so 39 applies. Write their extensions as \(B(z),S(z)\).

Analytic products retain the approximate algebra relations.

The boundary products satisfy, in both orders, \[\lVert B(s)S(s)-e^{2is}I\rVert,\ \lVert S(s)B(s)-e^{2is}I\rVert\leq2\delta+\delta^2 \quad\hbox{almost everywhere}.\] By the product identity and the positive Poisson kernel, \[ \lVert B(z)S(z)-e^{2iz}I\rVert,\ \lVert S(z)B(z)-e^{2iz}I\rVert\leq2\delta+\delta^2. \tag{68}\] Similarly, for every \(k\) and \(x\in M_k(M)_1\), \[ \operatorname{dist}\bigl(B(z)^{(k)}xS(z)^{(k)},M_k(M)\bigr)\leq\beta, \quad \operatorname{dist}\bigl(S(z)^{(k)}xB(z)^{(k)},M_k(M)\bigr)\leq\beta, \qquad \beta=2\gamma+2\delta+\delta^2. \tag{69}\] Here is a justification that avoids choosing measurable nearest points in \(M_k(M)\). On the boundary, the inequalities follow from (63) and (67). Test against any normal functional annihilating \(M_k(M)\). The product identity identifies its value in the upper half-plane with the Poisson average of its boundary values, so the same bound holds there. Since \(M_k(M)\) is weak-* closed, its norm distance is the supremum over these annihilating normal functionals of norm at most one. This last distance formula follows from Hahn–Banach applied to the preannihilator in the trace class.

For \(z=t+ir\) with \(0<r\leq\gamma\) define \[A_r(t)=e^{-iz}B(z),\qquad C_r(t)=e^{-iz}S(z).\] Because \(|e^{-iz}|=e^r\), equations (68)–(69) give, uniformly in \(r,t,k\), \[\begin{align*} \lVert A_r(t)\rVert,\ \lVert C_r(t)\rVert&\leq e^\gamma(1+\delta) \leq1+C\gamma,\tag{70}\\ \lVert A_r(t)C_r(t)-I\rVert,\ \lVert C_r(t)A_r(t)-I\rVert &\leq C\gamma,\tag{71}\\ \operatorname{dist}\bigl(A_r(t)^{(k)}xC_r(t)^{(k)},M_k(M)\bigr),\quad \operatorname{dist}\bigl(C_r(t)^{(k)}xA_r(t)^{(k)},M_k(M)\bigr) &\leq C\gamma. \tag{72}\end{align*}\] The last line is for \(x\in M_k(M)_1\).

Polar smoothing gives a norm-continuous path.

For small enough \(\gamma\), both products in (71) are invertible. They provide a right and a left inverse for \(A_r(t)\), which agree, and \[\lVert A_r(t)^{-1}\rVert\leq\frac{1+C\gamma}{1-C\gamma} \leq1+C\gamma.\] Together with (70), this places the spectrum of \(|A_r(t)|\) in \([1-C\gamma,1+C\gamma]\). Its polar unitary \[V_r(t)=A_r(t)\bigl(A_r(t)^*A_r(t)\bigr)^{-1/2}\] therefore satisfies \[ \lVert V_r(t)-A_r(t)\rVert\leq C\gamma, \qquad \lVert V_r(t)^*-C_r(t)\rVert\leq C\gamma. \tag{73}\] For the second bound, first use \(\lVert C_r(t)-A_r(t)^{-1}\rVert \leq\lVert A_r(t)^{-1}\rVert\lVert A_r(t)C_r(t)-I\rVert\), and then compare \(A_r(t)^{-1}\) with \(V_r(t)^*\). It follows from (72) that \(V_r(t)\) approximately normalizes \(M\) in both directions, with error \(C\gamma\) at every matrix level.

For fixed \(r>0\), the fields \(A_r(t)\) are norm continuous by holomorphy, and their inverses are uniformly bounded. Continuous functional calculus makes \(V_r(t)\) norm continuous. Put \[E_r(t)=V_r(t)V_r(0)^*.\] This is a norm-continuous unitary path with \(E_r(0)=I\) and \[ \sup_{x\in M_1}\operatorname{dist}(E_r(t)xE_r(t)^*,M)\leq C\gamma. \tag{74}\] To check composition here, approximate \(V_r(0)^*xV_r(0)\) by an \(m\in M\) with \(\lVert m\rVert\leq1+C\gamma\), and apply the corresponding estimate for \(V_r(t)\) to \(m\); the sum of the two errors is still \(C\gamma\). Apply 38 to the portion of this path joining \(0\) to \(t\), in either time direction. For each \(r,t\) it gives \(a_r(t)\in\mathcal U(M)\) and \(b_r'(t)\in\mathcal U(M')\) such that \[ \lVert E_r(t)-a_r(t)b_r'(t)\rVert\leq C\gamma, \tag{75}\] where a further slack of \(\gamma\) allows for an unattained infimum. No continuous choice of these factors is needed.

The boundary comparison holds at every fixed time.

Set \(F_0(s)=e^{is}W(s)\), a bounded strongly-* continuous field, and \[D_r(t)=e^{-i(t+ir)}\int_\mathbb R \frac{rF_0(s)}{\pi((t-s)^2+r^2)}\,ds.\] The Poisson formula and (67) give \(\lVert A_r(t)-D_r(t)\rVert\leq e^r\delta\). Thus \[ \lVert V_r(t)-D_r(t)\rVert\leq C\gamma, \qquad \lVert D_r(t)\rVert\leq e^r. \tag{76}\] At every fixed \(t\), the Poisson approximate identity gives \(D_r(t)\to W(t)\) strongly-* as \(r\downarrow0\). For clarity, on a fixed vector split the integral into \(|s-t|<h\) and its complement. Strong continuity controls the first piece, while the Poisson mass of the complement tends to zero. Apply the same argument to the adjoint field; the scalar prefactor converges to \(e^{-it}\). This proves both strong limits everywhere, not just almost everywhere.

Consequently \[Q_r(t)=D_r(t)D_r(0)^*\longrightarrow W(t) \quad\hbox{strongly-* at every fixed }t,\] since \(W(0)=I\) and all factors are uniformly bounded. By (76), \[ \lVert E_r(t)-Q_r(t)\rVert\leq C\gamma. \tag{77}\] These estimates compare \(E_r(t)\) with a convergent family; they do not assert convergence of \(E_r(t)\) itself.

A UCP limit recovers an actual normalizer.

Fix \(t\). For a sequence \(r\downarrow0\) define the UCP maps \[\Phi_r:M\longrightarrow M,\qquad \Phi_r(x)=a_r(t)x a_r(t)^*.\] Since \(b_r'(t)\) commutes with \(M\), this is the restriction of \(\operatorname{Ad}(a_r(t)b_r'(t))\). Equations (75) and (77) therefore imply, at every matrix level, \[ \lVert\Phi_r^{(k)}(x)-Q_r(t)^{(k)}xQ_r(t)^{(k)*}\rVert \leq C\gamma\lVert x\rVert,\qquad x\in M_k(M). \tag{78}\] Indeed, for any bounded operators \(T,R\) the cb norm of the difference of the maps \(x\mapsto TxT^*\) and \(x\mapsto RxR^*\) is at most \((\lVert T\rVert+\lVert R\rVert)\lVert T-R\rVert\).

Compactness of the product of the weak-operator compact balls \(\{y\in M:\lVert y\rVert\leq\lVert x\rVert\}\), indexed by \(x\in M\), supplies a subnet for which \(\Phi_r(x)\) converges weakly for every \(x\). Denote its limit by \(\Phi(x)\). Linearity, unitality and positivity at every matrix level pass to this limit, so \(\Phi:M\to M\) is UCP. No normality of \(\Phi\) is needed. Bounded strong-* continuity of multiplication gives \(Q_r(t)xQ_r(t)^*\to W(t)xW(t)^*\) strongly-* for each fixed \(x\), and likewise at every fixed matrix level. Weak-operator lower semicontinuity of norm in (78) now yields \[\lVert\Phi-\rho_t\rVert_{\mathrm{cb}}\leq C\gamma, \qquad \rho_t(x)=W(t)xW(t)^*.\] The representation \(\rho_t\) is faithful and normal, and its range has two-sided cb gap at most \(2\gamma\) from \(M\) by (64). In particular the reverse ordinary gap required by 5 is small. That proposition supplies a unitary \(q_t\) with \[\lVert q_t-I\rVert\leq C\gamma,\qquad q_t\rho_t(M)q_t^*=M.\] Then \(n_t=q_tW(t)\) is the required unitary normalizer and \(\lVert n_t-W(t)\rVert=\lVert q_t-I\rVert\leq C\gamma\). All estimates and smallness thresholds used only universal constants, which proves the uniform assertion. ◻

Correction within the normalizer group

The preceding construction produces pointwise normalizers close to a modular group. Their selections need not be measurable or continuous. We correct their multiplication law using an invariant mean on the discrete group \(\mathbb R\). The final comparison with the original modular group will then supply strong continuity.

The quadratic correction follows the averaging method for approximate unitary representations (Kazhdan 1982; Burger et al. 2013). Here the extra requirement is that every corrected unitary still normalizes the prescribed algebra. The local factorization proved next ensures that the averaging iteration preserves this requirement.

Lemma 41 (Small exact normalizers). There are absolute \(r_0,C_0>0\) such that, if \(P\) is a separably acting type \(\mathrm{III}\) factor, \(v\in\mathcal N_P\), and \(\lVert v-I\rVert<r_0\), then \[v=ab',\qquad a\in\mathcal U(P),\quad b'\in\mathcal U(P'), \qquad \lVert a-I\rVert+\lVert b'-I\rVert\le C_0\lVert v-I\rVert.\]

Proof. Write \(v=e^{ih}\) by the principal logarithm, with \(h=h^*\) and \(\lVert h\rVert\le2\lVert v-I\rVert\). For \(r_0\) small, the norm-convergent logarithm series on \(\mathcal B(\mathcal B(H))\) gives \[\log(\operatorname{Ad}_v)=i\operatorname{ad}_h.\] Indeed \(\operatorname{Ad}_v=\exp(i\operatorname{ad}_h)\) and \(\lVert\operatorname{ad}_h\rVert\) is within the radius where logarithm and exponential are inverse analytic functions. Each term of the logarithm series preserves \(P\), since \(vPv^*=P\). Hence \(i\operatorname{ad}_h\) restricts to a bounded \(*\)-derivation of \(P\).

The factor \(P\) is a simple unital \(C^*\)-algebra, by the same projection argument as in 37. Sakai’s theorem therefore gives a self-adjoint \(a_0\in P\) such that \(h-a_0\in P'\). A scalar may be subtracted from \(a_0\) and added to this commutant element. Choose the scalar at the midpoint of the spectrum of \(a_0\). We claim that the resulting \(a_0\) has \(\lVert a_0\rVert\le\lVert h\rVert\).

For any self-adjoint \(a\in P\), its spectral diameter is at most \(\lVert\operatorname{ad}_a|_P\rVert\). In fact, choose nonzero spectral projections near the two spectral endpoints and a norm-one partial isometry between nonzero subprojections of them. Compressing \([a,w]\) by \(w^*\) shows that its norm is at least the endpoint separation minus the two spectral tolerances. Let those tolerances tend to zero. Apply this to \(a_0\), using \(\lVert\operatorname{ad}_{a_0}|_P\rVert=\lVert\operatorname{ad}_h|_P\rVert\le2\lVert h\rVert\). After the midpoint shift, \(\lVert a_0\rVert\le\lVert h\rVert\) and \(b_0'=h-a_0\) has norm at most \(2\lVert h\rVert\). They commute, so \(v=e^{ia_0}e^{ib_0'}\); these are the required factors, with total displacement at most \(3\lVert h\rVert\). ◻

In the next proof all operator means use 7 for the discrete group \(\mathbb R\). Thus arbitrarily selected factors may be averaged separately in \(P\) and \(P'\), without a measurability assumption.

Lemma 42 (Correction of normalizer representations). There are absolute \(e_1,C_1>0\) such that the following holds. Let \(P\subset\mathcal B(H)\) be a separably acting type \(\mathrm{III}\) factor. For any family \(v_t\in\mathcal N_P\), \(t\in\mathbb R\), with \[e:=\sup_{t,s\in\mathbb R}\lVert v_tv_s-v_{t+s}\rVert<e_1,\] there is an exact unitary representation \(z:\mathbb R\to\mathcal N_P\) of the discrete group with \[\sup_t\lVert z_t-v_t\rVert\le C_1e.\]

Proof. Let \(m_s\) denote the invariant operator mean just described. Put \[F_t(s)=v_{t+s}v_s^*,\qquad A_t=m_sF_t(s).\] Every \(F_t(s)\) is a unitary normalizer and \(\lVert F_t(s)-v_t\rVert\le e\). Hence \(\lVert A_t-v_t\rVert\le e\). Translation invariance and the exact factorization \[F_{t+r}(s)=F_t(r+s)F_r(s)\] give \[\begin{align*} A_{t+r}-A_tA_r &=m_s\bigl[(F_t(r+s)-A_t)(F_r(s)-A_r)\bigr],\\ \lVert A_{t+r}-A_tA_r\rVert&\le4e^2. \end{align*}\] The centered terms have norm at most \(2e\), and their means vanish; expanding verifies the first equality with the written operator order. Likewise \[I-A_t^*A_t =m_s(F_t(s)-A_t)^*(F_t(s)-A_t),\] so its norm is at most \(4e^2\). Taking polar parts in \(\mathcal B(H)\) alone would not ensure membership in \(\mathcal N_P\). We instead use the local product structure of exact normalizers.

By 41, choose, for every \(s,t\), \[F_t(s)v_t^*=a_{t,s}b_{t,s}', \quad a_{t,s}\in\mathcal U(P),\quad b_{t,s}'\in\mathcal U(P'), \quad \lVert a_{t,s}-I\rVert+\lVert b_{t,s}'-I\rVert\le C_0e.\] All these choices can be arbitrary. Their averages \(a_t=m_sa_{t,s}\in P\) and \(b_t'=m_sb_{t,s}'\in P'\) satisfy \[\lVert m_s(a_{t,s}b_{t,s}')-a_tb_t'\rVert\le4C_0^2e^2.\] This follows by expanding the centered product; its two factors have norm at most \(2C_0e\). Unitarity of the individual factors also gives \[\lVert I-a_t^*a_t\rVert\le4C_0^2e^2, \qquad \lVert I-(b_t')^*b_t'\rVert\le4C_0^2e^2.\] For small \(e\) the averages are invertible, since they are within \(C_0e\) of \(I\). Their polar unitaries \(\widehat a_t\in P\) and \(\widehat b_t'\in P'\) differ from them by at most \(Ce^2\), with an absolute \(C\): use \(\lVert|a_t|-I\rVert\le \lVert a_t^*a_t-I\rVert/(1+\sqrt{1-4C_0^2e^2})\), and similarly on the other side. Therefore \[w_t=\widehat a_t\widehat b_t'v_t\in\mathcal N_P, \qquad \lVert w_t-A_t\rVert\le C_2e^2.\] It follows, uniformly in \(t,r\), that \[\lVert w_t-v_t\rVert\le e+C_2e^2, \qquad \lVert w_tw_r-w_{t+r}\rVert\le(4+3C_2)e^2.\] For the second inequality use \(\lVert A_t\rVert\le1\) and \(\lVert w_t\rVert=1\) to compare products, then the preceding bound on the defect of \(A\).

Iterate this operation, always with the same absolute constants. Choose \(e_1\) so small that every new defect is at most half the preceding one and every displacement is at most twice the preceding defect. The total displacement is then at most \(4e\). The families converge uniformly in \(t\) in operator norm to unitaries \(z_t\), satisfying \(z_tz_r=z_{t+r}\). The group \(\mathcal N_P\) is norm closed: if unitaries \(u_j\to u\) normalize \(P\), norm closure of \(P\) gives \(uPu^*\subset P\) and, by applying the same argument to \(u_j^*\), the reverse inclusion. Thus \(z_t\in\mathcal N_P\) for all \(t\). The exact multiplication law implies \(z_0=I\) and \(z_{-t}=z_t^*\), as required. ◻

Proof of 35. Put \(U_t=\Delta_N^{it}\) and \(W(t)=\Delta_N^{it}\Delta_M^{-it}\). 40 supplies \(n_t\in\mathcal N_M\) with \(\sup_t\lVert n_t-W(t)\rVert\le C\gamma\). Set \(v_t=n_t\Delta_M^{it}\), so \(v_t\in\mathcal N_M\) and \(\sup_t\lVert v_t-U_t\rVert\le C\gamma\). Since \(U\) is an exact representation, \(\sup_{t,s}\lVert v_tv_s-v_{t+s}\rVert\le3C\gamma\). For sufficiently small universal \(\gamma\), apply 42 to obtain an exact representation \(z_t\in\mathcal N_M\) with \[\sup_t\lVert z_t-U_t\rVert\le C'\gamma.\]

Take the invariant mean \(T=m_tz_tU_t^*\). It is within \(C'\gamma\) of \(I\) and hence invertible for small \(\gamma\). For every \(s\), translation invariance and the two exact representation laws imply \[z_sTU_s^*=m_tz_{s+t}U_{s+t}^*=T.\] Its polar unitary \(w=T|T|^{-1}\) therefore also satisfies \(z_sw=wU_s\). Moreover \(\lVert w-I\rVert\le2\lVert T-I\rVert\le2C'\gamma\). In particular \(z_s=wU_sw^*\); this proves strong continuity of \(z\) without any regularity assumption on the selected normalizers.

Set \(q=w^*\). Since \(z_s\) normalizes \(M\), the intertwining identity gives \(U_s(qMq^*)U_s^*=qMq^*\) for every \(s\). At each matrix level conjugation by \(q^{(k)}\) moves contractions by at most \(2\lVert q-I\rVert\). Consequently \[d_{\mathrm{cb}}(qMq^*,N)\le\gamma+2\lVert q-I\rVert\le C_{\rm mod}\gamma.\] Enlarging \(C_{\rm mod}\) and reducing \(c_{\rm mod}\) to meet the finitely many universal smallness conditions proves all assertions. ◻

Common crossed products and descent to factors

The preceding section puts the adjusted tensor pair under one strongly continuous normalizing unitary group. We now pass to its crossed products, compare finite corners there, and descend the resulting conjugacy twice: first to the adjusted tensor pair, and then through the auxiliary tensor factor to the original Hilbert space. Only the crossed product of the algebra whose modular group we use is initially known to be a factor.

A concrete fixed-point description

Let \(P\subseteq\mathcal B(H)\) be a unital von Neumann algebra on a separable Hilbert space, and let \((U_s)_{s\in\mathbb R}\) be a strongly continuous unitary group with \(U_sPU_s^*=P\). On \(\mathcal K=L^2(\mathbb R,H)\) define \[(\lambda_s\xi)(t)=\xi(t-s),\qquad (\rho_s\xi)(t)=\xi(t+s),\qquad (\pi(x)\xi)(t)=U_t^*xU_t\xi(t).\] We use the same symbol \(\pi\) for this faithful normal representation of all of \(\mathcal B(H)\). It is a complete isometry: at every matrix level its norm is the essential supremum of unitary conjugates of the input. Write \[C(P,U)=\bigl(\pi(P)\cup\{\lambda_s:s\in\mathbb R\}\bigr)''.\] This is the usual concrete crossed product for the action \(\operatorname{Ad}(U_s)|_P\). For a second unital algebra normalized by the same group, both its crossed product and the map \(\pi\) are on this same space.

Lemma 43 (Common-action fixed points). With this notation, \[ C(P,U)= \bigl(P\mathbin{\overline\otimes}\mathcal B(L^2(\mathbb R))\bigr)^\alpha, \qquad \alpha_s=\operatorname{Ad}(U_s\otimes\rho_s). \tag{79}\] For \((Q_a\xi)(t)=e^{iat}\xi(t)\), character conjugation \(\beta_a=\operatorname{Ad}(Q_a)\) preserves \(C(P,U)\), and \[ C(P,U)^\beta=\pi(P). \tag{80}\]

Proof. The generators \(\pi(x)\) and \(\lambda_r\) lie in \(P\mathbin{\overline\otimes}\mathcal B(L^2(\mathbb R))\) and are fixed by every \(\alpha_s\). For the reverse inclusion, take a fixed operator \(X\) in that tensor product. If \(f,g\in C_c(\mathbb R)\), put \[\lambda(f)=\int_\mathbb Rf(s)\lambda_s\,ds, \qquad Y=\lambda(f)X\lambda(g)^*.\] These integrals are understood on vectors. The operator \(Y\) remains fixed because translations commute with \(U_s\otimes\rho_s\). Define bounded operators from \(\mathcal K\) to \(H\) by \[R_{f,t}\xi=\int_\mathbb Rf(t-s)\xi(s)\,ds.\] They have norm at most \(\lVert f\rVert_2\) and depend continuously on \(t\) in operator norm. On compactly supported vector functions, Fubini’s theorem gives the integral kernel \[ k(t,r)=R_{f,t}XR_{g,r}^*,\qquad (Y\xi)(t)=\int_\mathbb Rk(t,r)\xi(r)\,dr. \tag{81}\] The kernel is bounded and jointly norm continuous. Each value belongs to \(P\): the evaluation operators intertwine the constant action of \(P'\), and \(X\) commutes with that action. Fixedness of \(Y\) gives \[k(t+s,r+s)=U_s^*k(t,r)U_s.\] For fixed \(s\) the identity first holds almost everywhere by testing the two kernels on compactly supported vectors. Norm continuity then makes it hold at every \((t,r)\). Consequently \[ k(t,r)=U_t^*k(0,r-t)U_t. \tag{82}\]

To reconstruct a bounded operator from this kernel inside the generated algebra, use the masks \[\chi_j(v)=(1-|v|/j)_+\qquad(j\ge1).\] The function \(\chi_j\) is the Fourier transform of the probability measure with density \[\frac{j}{2\pi} \left(\frac{\sin(ja/2)}{ja/2}\right)^2\,da,\] with the continuous value at \(a=0\). Thus multiplication of an operator kernel by \(\chi_j(t-r)\) is a contraction: it is the weak integral of \(\operatorname{Ad}(Q_a)\) against this measure. Let \(Y_j\) be the masked operator. Using (82) and \(\rho_v\xi(t)=\xi(t+v)\), \[ Y_j=\int_\mathbb R \chi_j(v)\pi(k(0,v))\rho_v\,dv\ \in C(P,U). \tag{83}\] Indeed the integrand has compact support and is strongly continuous, since \(v\mapsto k(0,v)\) is norm continuous. Its vector integral is a strong limit of finite sums of elements of \(C(P,U)\). Formula (81) verifies the equality on compactly supported vector functions. On these tests the kernels are bounded on a set of finite measure and \(\chi_j(t-r)\to1\), so dominated convergence gives \(Y_j\to Y\) weakly. The uniform norm bound extends this convergence to all vector tests. Hence \(Y\in C(P,U)\).

Choose nonnegative \(f_n\in C_c(\mathbb R)\) of integral one with supports shrinking to zero. Strong continuity of translations gives \(\lambda(f_n)\to I\) and \(\lambda(f_n)^*\to I\) strongly. Applying the preceding conclusion with \(f=g=f_n\) gives \(\lambda(f_n)X\lambda(f_n)^*\to X\) strongly. This proves (79).

For the second assertion, \[\beta_a(\pi(x))=\pi(x),\qquad \beta_a(\lambda_s)=e^{ias}\lambda_s.\] Thus \(\beta\) preserves \(C(P,U)\). An operator fixed by all \(\beta_a\) commutes with the scalar multiplication algebra on \(L^2(\mathbb R)\). If it also belongs to \(P\mathbin{\overline\otimes}\mathcal B(L^2(\mathbb R))\), it is multiplication by a bounded measurable \(P\)-valued field \(F(t)\). Here and below field measurability is weak operator measurability; on bounded sets, separability of \(H\) permits countably many coefficient tests. The remaining \(\alpha\)-fixedness says, for every \(s\), \[F(t+s)=U_s^*F(t)U_s\quad\hbox{for almost every }t.\] It follows that \(G(t)=U_tF(t)U_t^*\) is translation invariant in \(L^\infty(\mathbb R;\mathcal B(H))\). Such a field is almost everywhere constant. To see this without choosing uncountably many exceptional sets, test against a countable dense family of vector pairs. Each scalar coefficient is translation invariant as an \(L^\infty\) class; convolving with a smooth compactly supported approximate identity gives continuous translation-invariant functions, hence constants. Their local \(L^1\) convergence proves the assertion for the coefficients and then for the field. Since \(U_t\) normalizes \(P\), the constant value \(x\) belongs to \(P\). Therefore \(F(t)=U_t^*xU_t\) almost everywhere, which proves (80). ◻

For each unitary action \(V_s\) below, use the discrete invariant mean from 7 and write \[E_V(T)=m_sV_sTV_s^*.\] This UCP map fixes invariant operators and preserves every von Neumann algebra invariant under the action.

Lemma 44 (Completely bounded gap of common crossed products). Let \(P,Q\subseteq\mathcal B(H)\) be normalized by the same group \(U\). If \(d_{\mathrm{cb}}(P,Q)<\gamma\), then \[d_{\mathrm{cb}}(C(P,U),C(Q,U))\le\gamma.\] The conclusion holds with any strict upper bound in place of \(\gamma\).

Proof. By 6, tensoring the pair with \(\mathcal B(L^2(\mathbb R))\) preserves the upper bound. Average on the common ambient operator algebra using \(V_s=U_s\otimes\rho_s\) in 7. For a contraction \(X\) in any matrix level of \(C(P,U)\), choose a contraction \(Y\) in the same matrix level of \(Q\mathbin{\overline\otimes}\mathcal B(L^2(\mathbb R))\) within the chosen gap bound. The matrix amplification of \(E_V\) fixes \(X\), takes \(Y\) into \(C(Q,U)\) by 43, and preserves both the norm bound and the approximation error. Reverse \(P,Q\). This proves the assertion at every matrix level with the same bound. ◻

A finite corner and its inflation

We shall use 21 to check the other corner: its proof first excludes proper isometries by approximation in a finite algebra and then excludes nontrivial central projections by tracial conjugacy averaging. Thus no factor hypothesis on the other crossed product is needed.

Proposition 45 (Conjugacy through a common crossed product). There are an absolute \(\gamma_0>0\) and a nondecreasing function \(\zeta:(0,\gamma_0)\to(0,\infty)\), with \(\zeta(\gamma)\to0\) as \(\gamma\downarrow0\), with the following property. Suppose \(P,Q\subseteq\mathcal B(H)\) are unital von Neumann algebras on a separable Hilbert space, \(d_{\mathrm{cb}}(P,Q)<\gamma<\gamma_0\), and a strongly continuous unitary group \(U\) normalizes both. If \(C(Q,U)\) is a type \(\mathrm{II}_\infty\) factor, then a unitary \(a\in\mathcal B(H)\) satisfies \[aPa^*=Q,\qquad \lVert a-I\rVert\le\zeta(\gamma).\] No factoriality or modular interpretation of \(C(P,U)\) is assumed.

Proof. Write \(\mathcal A=C(P,U)\) and \(\mathcal B=C(Q,U)\). Use a slightly smaller gap bound if necessary so that 44 gives \(d_{\mathrm{cb}}(\mathcal A,\mathcal B)<\gamma\). The semifinite factor \(\mathcal B\) contains matrix units \((e_{ij})_{i,j\ge1}\) with \(\sum_i e_{ii}=I\) strongly and \(e=e_{11}\) finite and nonzero. Here is the usual projection construction. A faithful normal semifinite trace, together with separability of the representation, gives an increasing sequence of finite-trace projections with supremum \(I\). Split their successive differences by trace continuity, and group the resulting pieces into countably many projections of one fixed positive finite trace. The infinite total trace ensures that the grouping exhausts \(I\). Equal-trace comparison gives the required matrix units. Their weak closed span \(S\) is a unital copy of \(\mathcal B(\ell^2)\), hence amenable.

Apply 2 to \(S\subset_\gamma\mathcal A\). For \(\gamma<1/100\) it gives a unitary \(w\) with \[wSw^*\subseteq\mathcal A,\qquad \lVert w-I\rVert\le150\gamma.\] Put \(\mathcal A_1=w^*\mathcal A w\), so \(S\subseteq\mathcal A_1\cap\mathcal B\) and \[ d_{\mathrm{cb}}(\mathcal A_1,\mathcal B)<301\gamma=:\kappa. \tag{84}\] Compression by the common projection \(e\) gives \[A_0=e\mathcal A_1e|_{e\mathcal K},\qquad B_0=e\mathcal Be|_{e\mathcal K},\qquad d(A_0,B_0)<\kappa.\] The algebra \(B_0\) is a type \(\mathrm{II}_1\) factor. When \(\kappa<1/10\), 21 first makes \(A_0\) finite and then makes it a factor. Thus 20 applies to this pair on the separable Hilbert space \(e\mathcal K\). Choosing its orientation appropriately gives a unitary \(v_0\) on \(e\mathcal K\) with \[v_0A_0v_0^*=B_0, \qquad \lVert v_0-e\rVert\le\eta_{\rm fin}(\kappa).\]

Inflate this unitary using the common matrix units: \[ V=\sum_{i=1}^{\infty}e_{i1}v_0e_{1i}. \tag{85}\] The sum converges strongly and is a unitary, because it is the direct sum of copies of \(v_0\) on the orthogonal spaces \(e_{ii}\mathcal K\). Moreover \[\lVert V-I\rVert=\lVert v_0-e\rVert,\qquad Ve_{ij}=e_{ij}V.\] For \(x\in\mathcal A_1\), its transported matrix entries \(e_{1i}xe_{j1}\) belong to \(A_0\). The entries of \(VxV^*\) therefore belong to \(B_0\). Its finite matrix compressions lie in \(\mathcal B\) and converge strongly to \(VxV^*\); hence \(V\mathcal A_1V^*\subseteq \mathcal B\). The reverse argument gives equality. Thus \[ Z=Vw^*,\qquad Z\mathcal A Z^*=\mathcal B, \qquad \lVert Z-I\rVert\le150\gamma+\eta_{\rm fin}(301\gamma). \tag{86}\]

It remains to descend to \(H\). Average the character action \(\beta\) on the entire \(\mathcal B(\mathcal K)\), obtaining a unital completely positive map \(E_\beta\) from 7, with \(V_a=Q_a\). By 43, it maps \(\mathcal B\) into \(\pi(Q)\) and fixes \(\pi(x)\) for every \(x\in\mathcal B(H)\). Therefore \[ F(x)=\pi^{-1}\bigl(E_\beta(Z\pi(x)Z^*)\bigr),\qquad x\in P, \tag{87}\] defines a unital completely positive map \(P\to Q\). Since the same map \(\pi\) is completely isometric on all of \(\mathcal B(H)\), \[\lVert F-\operatorname{id}_P\rVert_{\mathrm{cb}} \le2\lVert Z-I\rVert \le300\gamma+2\eta_{\rm fin}(301\gamma).\] Here \(\operatorname{id}_P\) denotes the original inclusion into \(\mathcal B(H)\). The ordinary gap \(d(P,Q)\) is less than \(\gamma\). Apply 5. Its absolute thresholds hold for all sufficiently small \(\gamma\), because \(\eta_{\rm fin}(t)\to0\). Its displacement bound supplies \(\zeta(\gamma)\). A constant multiple of \(\gamma+\eta_{\rm fin}(301\gamma)\), enlarged to a nondecreasing function if needed, is sufficient. This proves the proposition. ◻

Undoing preparation and the auxiliary tensor factor

We finish the separable factor case. The continuous-core input used here is the one in 6: the crossed product of a type \(\mathrm{III}_1\) factor by the modular action of a faithful normal state is a type \(\mathrm{II}_\infty\) factor. We apply it only to the algebra whose modular group remains unchanged during alignment (Takesaki 1973, Corollary 9.7, p. 297).

Theorem 46 (Uniform stability of separably acting factors). There are an absolute \(t_{\rm fac}>0\) and a nondecreasing function \(\eta_{\rm fac}:(0,t_{\rm fac})\to(0,\infty)\) with \(\eta_{\rm fac}(t)\to0\) as \(t\downarrow0\) such that the following holds. For every separable complex Hilbert space \(H\) and every pair of unital factors \(M_0,N_0\subseteq\mathcal B(H)\) with the same identity and \(d(M_0,N_0)<t<t_{\rm fac}\), there is a unitary \(u\in\mathcal B(H)\) with \[uM_0u^*=N_0,\qquad \lVert u-I_H\rVert\le\eta_{\rm fac}(t).\] The function and the threshold are independent of the factors, their types, and their representations.

Proof. If either factor is finite, 20 applies, with the roles reversed if necessary. Its finite-neighbor conclusion also shows that for a sufficiently small gap both factors are finite in that case. We may therefore suppose that both are infinite.

Use 32, retaining its changes of representation explicitly. On \(H\otimes K\) it gives the fixed hyperfinite tensor factor \(D\subseteq\mathcal B(K)\) and unitaries \(u_0\in\mathcal B(H)\) and \(v\in\mathcal B(H\otimes K)\) such that \[M=M_0\mathbin{\overline\otimes}D,\qquad N=v\bigl((u_0N_0u_0^*)\mathbin{\overline\otimes}D\bigr)v^*,\] are a prepared type \(\mathrm{III}_1\) pair, and, for absolute constants, \[\lVert u_0-I\rVert+\lVert v-I\rVert\le C_1t, \qquad d_{\mathrm{cb}}(M,N)<C_2t.\] An enlargement of \(C_2\) absorbs strict upper-bound slack. Let \(U_s=\Delta_N^{is}\) be the modular group belonging to the common vector in that preparation. By 35, a unitary \(q\) satisfies \[\lVert q-I\rVert\le C_3t,\qquad P=qMq^*\text{ is normalized by every }U_s, \qquad d_{\mathrm{cb}}(P,N)<C_4t.\] The algebra \(N\) and its modular group have not changed. Consequently \(C(N,U)\) is its type \(\mathrm{II}_\infty\) modular core. No modular claim about the action on \(P\) is necessary. Apply 45 with \(\gamma=C_4t\) to obtain \[aPa^*=N,\qquad \lVert a-I\rVert\le\zeta(C_4t).\]

All these operations have taken place on the actual tensor product of the original Hilbert space. Their inverse adjustments give \[ b=(u_0^*\otimes I_K)v^*aq, \qquad b(M_0\mathbin{\overline\otimes}D)b^*=N_0\mathbin{\overline\otimes}D, \tag{88}\] and \[ \lVert b-I\rVert\le(C_1+C_3)t+\zeta(C_4t)=:r(t). \tag{89}\] The right side tends to zero universally.

Choose a unit vector \(\xi\in K\), and let \(\omega_\xi\) be its normal vector state on all of \(\mathcal B(K)\). Slicing the conjugation in (88) defines \[F_0(x)=(\operatorname{id}\mathbin{\overline\otimes}\omega_\xi) \bigl(b(x\otimes I_K)b^*\bigr),\qquad x\in M_0.\] This is a unital completely positive map into \(N_0\). The slice is defined on the entire ambient tensor product, so comparison with \(x\otimes I_K\) gives at every matrix level \[\lVert F_0-\operatorname{id}_{M_0}\rVert_{\mathrm{cb}}\le2\lVert b-I\rVert\le2r(t).\] The reverse ordinary gap is less than \(t\). A second application of 5 therefore gives the desired conjugating unitary on the original \(H\), with displacement at most an absolute constant times \(r(t)\), for all sufficiently small \(t\).

Choose \(t_{\rm fac}\) below the finite, preparation, modular-alignment, core-descent, and two UCP thresholds just used. Every threshold is absolute, and the only non-linear modulus used here is \(\eta_{\rm fin}\), which is uniform and tends to zero. The maximum of \(\eta_{\rm fin}(t)\) and a sufficiently large constant multiple of \(r(t)\), enlarged if needed to be positive and nondecreasing, defines \(\eta_{\rm fac}\). This establishes the stated uniform quantifiers. ◻

Centers and arbitrary Hilbert spaces

We now pass from 46 to arbitrary von Neumann algebras. There are two steps. On a separable Hilbert space, aligning the centers permits a measurable assembly of the factor implementers. On an arbitrary Hilbert space, separable restrictions instead supply a completely positive map into the neighboring algebra. The UCP criterion then gives a unitary on the original Hilbert space. The uniformity of 46 is essential in both steps.

Aligning the centers

Lemma 47. For unital von Neumann algebras \(M,N\subseteq\mathcal B(H)\), \[d\bigl(Z(M),Z(N)\bigr)\leq 3d(M,N).\] Consequently, whenever \(d(M,N)<g<1/303\), there is a unitary \(w\in\mathcal B(H)\) such that \[wZ(M)w^*=Z(N),\qquad \lVert w-I\rVert\leq450g.\] For \(N_0=w^*Nw\) one then has \(Z(N_0)=Z(M)\) and \(d(M,N_0)<901g\).

Proof. Fix \(z\in Z(M)_1\) and choose \(y\in N_1\) with \(\lVert z-y\rVert<g\). For \(v\in\mathcal U(N)\), choose \(a\in M_1\) with \(\lVert v-a\rVert<g\). Since \(z\) commutes with \(a\), \[\lVert[z,v]\rVert<2g, \qquad \lVert vyv^*-z\rVert \leq\lVert y-z\rVert+\lVert v z v^*-z\rVert<3g.\] Dixmier’s averaging theorem for von Neumann algebras (Blackadar 2006, Definition III.2.5.16 and Theorem III.2.5.18) says that the norm-closed convex hull of \(\{vyv^*:v\in\mathcal U(N)\}\) meets \(Z(N)\). An element of this intersection is a contraction and is within \(3g\) of \(z\). Interchanging \(M,N\) and then letting \(g\downarrow d(M,N)\) proves the gap estimate.

The centers are amenable. Apply the mutual conclusion of 2, with near-inclusion bound \(3g\), to obtain \(w\). Finally, conjugation by \(w\) changes an ordinary gap by at most \(2\lVert w-I\rVert\), giving the asserted bound for \(N_0\). ◻

The common-center decomposition

We use the following precise form of central decomposition. If \(M\) acts on a separable Hilbert space and \(Z=Z(M)\), there is a standard measure space \((X,\mu)\) and a measurable field of separable Hilbert spaces such that \[H=\int_X^\oplus H_t\,d\mu(t),\qquad Z=L^\infty(X,\mu),\qquad M=\int_X^\oplus M_t\,d\mu(t),\] where \(M_t\) is a factor almost everywhere. The measure may be chosen finite and completed. A bounded measurable field with value in \(M_t\) almost everywhere represents an element of \(M\), and its norm is the essential supremum of its fiber norms. These are the central case of (Blackadar 2006, III.1.6.2–4). For two algebras with the same center, we use the same decomposition of \(H\) and the same diagonal algebra. Countably generated \(C^*\)-subalgebras can be represented simultaneously on the fibers; see also (Blackadar 2006, III.5.1.10–14).

Lemma 48. Let \(M,N\subseteq\mathcal B(H)\) act on a separable Hilbert space and have the same center. In their common central decomposition, \[d(M_t,N_t)\leq d(M,N) \quad\text{for almost every }t.\]

Proof. Choose separable unital \(C^*\)-subalgebras \(A\subseteq M\) and \(B\subseteq N\) whose strong closures are \(M,N\). Such subalgebras exist because the strong topology on each bounded operator ball is metrizable and separable when \(H\) is separable. Choose countable norm-dense subsets \(\{a_j\}\subseteq A_1\) and \(\{b_j\}\subseteq B_1\). Off one null set, their fiber representations satisfy \[M_t=A(t)'',\qquad N_t=B(t)''.\]

We first check the unit-ball density needed below. The map \(A\to A(t)\) is a quotient \(*\)-homomorphism, possibly with a nonzero kernel. A contraction in \(A(t)\) has lifts of norm at most \(1+s\) for every \(s>0\). Rescaling these lifts to contractions shows that the image of \(A_1\) is norm dense in \(A(t)_1\). It follows that \(\{a_j(t)\}\) is norm dense in \(A(t)_1\) and, by Kaplansky density, strongly dense in \((M_t)_1\). The same argument applies to \(B\).

Put \(d_0=d(M,N)\). For every \(j,n\geq1\), choose \(y_{j,n}\in N_1\) such that \[\lVert a_j-y_{j,n}\rVert<d_0+1/n.\] The essential-supremum norm formula and the countability of these choices give a common conull set on which \[y_{j,n}(t)\in(N_t)_1, \qquad \lVert a_j(t)-y_{j,n}(t)\rVert\leq d_0+1/n \quad(j,n\geq1).\] Fix such a \(t\), an \(x\in(M_t)_1\), and an \(n\). Choose a net of \(a_j(t)\) converging strongly to \(x\). Along a subnet, the corresponding \(y_{j,n}(t)\) converge weakly to a contraction \(y\in N_t\). Operator-norm balls are weak-operator closed, so \(\lVert x-y\rVert\leq d_0+1/n\). Letting \(n\) increase gives \(\operatorname{dist}(x,(N_t)_1)\leq d_0\). Make the same countable choices in the opposite direction and discard their null exceptions. This proves both near inclusions on one conull set. ◻

Lemma 49. In the setting of 48, suppose that for some \(a>0\) and almost every \(t\) there is a unitary \(v_t\in\mathcal B(H_t)\) satisfying \[v_t M_t v_t^*=N_t,\qquad\lVert v_t-I\rVert<a.\] Then there is a unitary \(v\in\mathcal B(H)\) such that \(vMv^*=N\) and \(\lVert v-I\rVert\leq a\).

Proof. Work on the underlying standard Borel model of \(X\), retaining the countable contraction fields \(a_j(t),b_j(t)\) from the preceding proof. Choose Borel versions of the dimension function, the trivializations on its countably many level sets, and the countably many contraction fields. Discard a Borel null set containing all exceptions to these identifications, the asserted density and algebraic properties, and the implementer hypothesis. On each remaining dimension stratum, the fibers are identified with one fixed separable Hilbert space \(K\). Zero-dimensional fibers contribute nothing to the direct integral and may be omitted. It suffices to select implementers on each of these standard Borel strata.

The unitary group \(\mathcal U(\mathcal B(K))\), with the strong-\(*\) topology, is Polish. Fix a dense sequence \((\xi_l)\) in the unit sphere of \(K\). Consider the relation on pairs \((t,u)\) defined by \(\lVert u-I\rVert\leq a\) and the following two families of conditions: \[\begin{align*} &\text{for every }j,m,k\text{ there is }q\text{ such that } \max_{l\leq m} \lVert\bigl(u a_j(t)u^*-b_q(t)\bigr)\xi_l\rVert<1/k,\\ &\text{for every }j,m,k\text{ there is }q\text{ such that } \max_{l\leq m} \lVert\bigl(u^* b_j(t)u-a_q(t)\bigr)\xi_l\rVert<1/k. \end{align*}\] All these predicates are Borel. Indeed, multiplication of bounded operator fields and strong-\(*\) unitary fields is Borel, and the norm condition on \(u-I\) is a countable collection of vector-norm tests. The strong density of the contraction fields shows that the first family is equivalent to \(uM_tu^*\subseteq N_t\), and the second to the reverse inclusion. Thus the relation describes exactly the implementers with the imposed bound. Its sections are nonempty by hypothesis.

The Jankov–von Neumann uniformization theorem (Kechris 1995, Theorem 18.1) gives a selection measurable for the sigma algebra generated by analytic sets, for an analytic relation between standard Borel spaces; our relation is Borel and hence analytic. Analytic sets are universally measurable (Kechris 1995, Theorem 21.10), so the selector is measurable for the completed measure \(\mu\). Its direct integral is a unitary \(v\) with \(\lVert v-I\rVert\leq a\). For \(x\in M\), the field \(v_t x_t v_t^*\) is a bounded measurable field in \(N_t\), so the field-reconstruction part of central decomposition puts \(vxv^*\) in \(N\). The reverse inclusion follows in the same way. ◻

Theorem 50. For every \(a>0\) there is \(\delta_{\mathrm{sep}}(a)>0\) such that whenever \(M,N\) are unital von Neumann algebras on a separable Hilbert space and \(d(M,N)<\delta_{\mathrm{sep}}(a)\), there is a unitary \(u\) with \[uMu^*=N,\qquad \lVert u-I\rVert<a.\] The tolerance is independent of the algebras and the representation.

Proof. The zero Hilbert space is immediate. By 46, choose \(\tau>0\) such that any pair of factors on a separable Hilbert space at distance less than \(\tau\) has an implementer within \(a/2\) of the identity. Choose \(\delta_{\mathrm{sep}}(a)>0\) so small that 47 applies with \(g=\delta_{\mathrm{sep}}(a)\) and \[450\delta_{\mathrm{sep}}(a)<a/2, \qquad 901\delta_{\mathrm{sep}}(a)<\tau.\] For a pair satisfying the hypothesis, align the centers by the unitary \(w\) in that lemma, and set \(N_0=w^*Nw\). The common-center pair \(M,N_0\) has distance less than \(\tau\). By 48, almost every fiber pair has distance less than \(\tau\) as well. Apply the factor theorem on each fiber and then 49, with bound \(a/2\). The resulting unitary \(v\) satisfies \[vMv^*=N_0,\qquad\lVert v-I\rVert\leq a/2.\] Now \(u=wv\) implements the original conjugacy and \(\lVert u-I\rVert\leq\lVert w-I\rVert+\lVert v-I\rVert<a\). Every tolerance used above is independent of the fibers and of \(H\). ◻

Separable tests in an arbitrary representation

The preceding theorem provides one uniform unitary bound for all separable restrictions. These restrictions will not be chosen compatibly. Instead, finite collections of operator coefficients will produce a UCP limit map, to which 5 applies.

Lemma 51. Let \(M,N\subseteq\mathcal B(H)\) be unital von Neumann algebras, and let \(g>d(M,N)\). Given finite sets \(E\subseteq M\), \(F\subseteq N\), and \(V\subseteq H\) with \(V\) containing a nonzero vector, there are separable unital \(C^*\)-subalgebras \(A\subseteq M\), \(B\subseteq N\) and a nonzero separable closed subspace \(K\subseteq H\) such that:

  1. \(E\subseteq A\), \(F\subseteq B\), and \(V\subseteq K\);

  2. \(K\) reduces both \(A\) and \(B\);

  3. \(d(A,B)<g\) and \[d\bigl((A|_K)'',(B|_K)''\bigr)<g.\]

Here the restriction homomorphisms are not required to be faithful.

Proof. Choose \(r\) with \(d(M,N)<r<g\), and start with \(A_0=C^*(E,I)\), \(B_0=C^*(F,I)\). Suppose that separable unital \(A_j\subseteq M\), \(B_j\subseteq N\) have been constructed. For each element of a countable norm-dense subset of \((A_j)_1\), choose an \(N_1\)-approximant within \(r\), and adjoin all these approximants to \(B_j\). In the same way, adjoin to \(A_j\) approximants in \(M_1\) for a countable dense subset of \((B_j)_1\). Let \(A_{j+1},B_{j+1}\) be the unital \(C^*\)-algebras generated by these additions. Put \[A=\overline{\bigcup_j A_j}^{\,\lVert\cdot\rVert}, \qquad B=\overline{\bigcup_j B_j}^{\,\lVert\cdot\rVert}.\] They are separable and contain the required tests. The stage unit balls are norm dense in the unit balls of these closures: approximate in norm and rescale if necessary. The choices of approximants therefore give \(d(A,B)\leq r<g\).

Let \(K=\overline{\operatorname{span}(C^*(A,B)V)}\). Because the algebra is separable and \(V\) finite, \(K\) is separable. It contains \(V\) and is reducing for both algebras. To transfer the gap through restriction, take \(x\in(A|_K)_1\). For every \(s>0\) there is a quotient lift \(a\in A\) with \(a|_K=x\) and \(\lVert a\rVert\leq1+s\). The contraction \(a/(1+s)\) has a \(B_1\)-approximant within \(r\) plus arbitrary slack. Restricting this approximant gives \[\operatorname{dist}\bigl(x,(B|_K)_1\bigr) \leq \frac{s}{1+s}+r.\] Let \(s\downarrow0\) and reverse the roles. Thus \(d(A|_K,B|_K)\leq r\).

Finally, for a contraction in \((A|_K)''\), Kaplansky density gives a strongly convergent net of contractions from \(A|_K\). Approximate these in \((B|_K)_1\) with errors at most \(r\) plus arbitrary slack, and take a bounded weak-operator cluster subnet. Its limit lies in \(((B|_K)'')_1\) and has the same error bound. Interchanging the algebras proves the final assertion. ◻

Proof of 1. Let \(\varepsilon>0\). The conclusion of 5 is uniform in the algebras and the Hilbert space, and its implementing-unitary bound tends to zero with the cb error. Choose \(a>0\) and \(g_0>0\) so small that this proposition gives an implementer within \(\varepsilon\) of the identity whenever the reverse ordinary gap is at most \(g_0\) and the cb error of the UCP map is at most \(2a\). All its smallness requirements are included in this choice. Choose \[0<\delta<\min\{g_0,\delta_{\mathrm{sep}}(a)\}.\] We prove the theorem for any \(M,N\subseteq\mathcal B(H)\) with \(d(M,N)<\delta\).

If \(H=\{0\}\), take its identity unitary. Otherwise fix a unit vector \(\xi_0\in H\) and a number \(g\) satisfying \[d(M,N)<g<\delta_{\mathrm{sep}}(a).\] Consider the directed set of quadruples \(i=(E_i,F_i,V_i,k_i)\), where \(E_i\subseteq M\), \(F_i\subseteq N\), and \(V_i\subseteq H\) are finite, \(\xi_0\in V_i\), and \(k_i\in\mathbb N\) is positive. The order is inclusion in the three finite sets and the usual order in \(k_i\).

For each \(i\), apply 51 to obtain \(A_i,B_i,K_i\). Set \[P_i=(A_i|_{K_i})'',\qquad Q_i=(B_i|_{K_i})''.\] These are unital von Neumann algebras on the nonzero separable space \(K_i\), and \(d(P_i,Q_i)<g\). By 50, there is a unitary \(u_i\in\mathcal B(K_i)\) with \[u_iP_iu_i^*=Q_i,\qquad\lVert u_i-I\rVert<a.\] Extend it by the identity on \(K_i^\perp\) to a unitary \(U_i\in\mathcal B(H)\). For all matrix sizes, conjugation satisfies \[ \lVert\operatorname{Ad}(U_i)|_M-\operatorname{id}_M\rVert_{\mathrm{cb}}\leq2\lVert U_i-I\rVert<2a. \tag{90}\]

We next record what the local conjugacy says about the finite tests on the original space. If \(x\in E_i\), then \(K_i\) reduces \(x\), and the restriction of \(U_ixU_i^*\) belongs to \(Q_i\) with norm at most \(\lVert x\rVert\). Kaplansky density approximates this restriction on the finitely many coefficients from \(V_i\) by an element of \(B_i|_{K_i}\) of norm at most \(\lVert x\rVert\) (with arbitrarily small slack if needed). Lift that element through the restriction quotient. Choosing the slack small enough gives \(b_{i,x}\in B_i\subseteq N\) such that \[ \lVert b_{i,x}\rVert\leq\lVert x\rVert+1, \qquad \lvert\langle (U_ixU_i^*-b_{i,x})\xi,\eta\rangle\rvert<1/k_i \quad(\xi,\eta\in V_i). \tag{91}\] The displayed coefficients are coefficients on \(H\), since all the vectors belong to \(K_i\). At this stage the whole conjugated algebra need not lie in \(N\); (91) is the property that will survive as the tests increase.

For each \(x\in M\), the operators \(U_ixU_i^*\) lie in the weak-operator compact ball of radius \(\lVert x\rVert\) in \(\mathcal B(H)\). Compactness of the product of these balls gives a subnet for which \[F(x)=\operatorname*{WOT\!\!-\!lim}_i U_ixU_i^* \quad\text{exists for every }x\in M.\] Linearity and unitality pass to this limit. At each matrix size, positive operators form a weak-operator closed cone, so \(F\) is completely positive. Also, the norm ball of radius \(2a\lVert X\rVert\) is weak-operator closed for every fixed matrix \(X\) over \(M\). Taking limits in (90) therefore gives \[ \lVert F-\operatorname{id}_M\rVert_{\mathrm{cb}}\leq2a. \tag{92}\]

It remains to verify the range of \(F\). Fix \(x\in M\). Cofinality of the subnet implies that eventually \(x\in E_i\), so choose \(b_{i,x}\) as in (91) on this tail. Every fixed pair \(\xi,\eta\in H\) eventually belongs to \(V_i\), while \(k_i\to\infty\). Hence (91) and the definition of \(F\) imply \(b_{i,x}\to F(x)\) in the weak-operator topology. As \(N\) is weak-operator closed, \(F(x)\in N\). This holds for every \(x\in M\), and proves that \(F:M\to N\) is UCP. Normality of this cluster map is neither asserted nor needed.

The original inclusion \(M\hookrightarrow\mathcal B(H)\) is faithful, normal, and unital. Its range has reverse ordinary gap less than \(\delta<g_0\) from \(N\), and (92) supplies the cb bound. Apply 5. It yields a unitary \(u\in\mathcal B(H)\) with \(uMu^*=N\) and \(\lVert u-I\rVert<\varepsilon\).

The choices of \(a,g_0\), and \(\delta\) depend only on \(\varepsilon\). The intervening choices of finite tests, separable subspaces, and quotient lifts change none of their bounds. Thus the same tolerance works for all algebras, all representations, and all Hilbert spaces, as claimed. ◻

Uniformity and commutator estimates

The proof distinguishes ordinary norm control, control at every matrix size, and implementation on the original Hilbert space. We compare the uniformity required for these steps with the commutator estimate in the similarity companion, and then record how the small-unitary conclusion controls both matrix gaps and commutants.

Uniformity over factors and representations

Fix an abstract type \(\mathrm{II}_1\) factor \(M\) with separable predual. A stability statement uniform over its representations would assign, for each \(\varepsilon>0\), a number \(\delta_M(\varepsilon)>0\) that works for every faithful normal unital representation \(\rho:M\to\mathcal B(H)\) and every unital von Neumann algebra \(N\subseteq\mathcal B(H)\): if \(d(\rho(M),N)<\delta_M(\varepsilon)\), then a unitary within \(\varepsilon\) of the identity conjugates \(\rho(M)\) onto \(N\). Even with this representation uniformity, the tolerance may depend on \(M\).

The quantifier order for universal stability is stronger: the tolerance must be chosen before the factor as well as before its representation. Positivity of each \(\delta_M(\varepsilon)\) does not imply \[\inf_M\delta_M(\varepsilon)>0.\] Consequently, a representation-transfer argument with absolute constants preserves an already universal tolerance but does not by itself remove factor dependence from its input.

In 20, uniformity is obtained before the concluding reductions. The common-position theorem has absolute constants, the boundary argument in 28 gives an isomorphism with an absolute ordinary-norm bound, and 8 supplies an absolute amplification constant. Their combination gives the single modulus \(\eta_{\rm fin}\) for all separably acting finite factors. Sections 8 and 9 propagate this uniform control through finite corners, measurable assembly, and separable restrictions.

Two commutator arguments

The amplification theorem needed here concerns selfadjoint implementers of commutators on separably acting type \(\mathrm{II}_1\) factors. The similarity companion proves a broader estimate: for every unital von Neumann algebra \(P\subseteq\mathcal B(K)\), every \(Y\in\mathcal B(K)\), every integer \(n\ge1\), and every \(X\in M_n(P)\), \[\lVert[Y^{(n)},X]\rVert \le C\lVert\operatorname{ad}_Y|_P\rVert\lVert X\rVert,\] where \(Y^{(n)}\) is the diagonal amplification on \(K^n\) and \(C\) is absolute, with no separability assumption (OpenAI 2026, Theorem 1.2).

Both arguments use cyclic representation estimates and tracial free products to obtain bounds independent of matrix size. In the present paper, the column estimate of 9 controls randomized matrix coefficients. The free-word estimates bound their commutators, while the averaged products converge weakly to a positive scalar multiple of the identity in (34). Together these facts yield 8.

The companion’s central intermediate result is a corner estimate. It works in a tracial free product \(\mathcal L=D*A_0\), where \(D\) is a hyperfinite \(\mathrm{II}_1\) factor generated by increasing full matrix algebras and \(A_0\) is a finite von Neumann algebra with a faithful normal tracial state. Equip \(\mathcal L\) with the free-product trace. Writing \(L_x\) for left multiplication by \(x\) on \(L^2(\mathcal L)\), suppose that \(G\in\mathcal B(L^2(\mathcal L))\) commutes with \(L_D\) and that, for some \(e\ge0\), \[\lVert[G,L_x]\rVert\le e\lVert x\rVert\qquad(x\in p\mathcal Lp),\] where \(p\in D\) is a projection of trace \(1/m\) for an integer \(m\ge2\). For a trace-zero selfadjoint unitary \(s\in A_0\), the companion bounds the compression of \([G,L_s]\) to \(L^2(A_0)\) by \(C_0e\), with \(C_0\) independent of \(m\) (OpenAI 2026, Theorem 3.1). Its proof analyzes invariant tensors of reduced words, approximates operators by right multipliers on small right supports, and obtains the trace-independent bound from a Hankel compression (OpenAI 2026, secs. 3–5). Using Popa’s independence theorem, the companion transfers this estimate to finite factors; passage through a countable amplification treats their arbitrary separable normal representations (OpenAI 2026, sec. 6 and Proposition 7.1).

The finite stability proof here uses 8 directly. For the near-identity isomorphism \(\theta:M\to N\) constructed there, the switch operator on \(H\oplus H\) has commutator with \(x\mapsto x\oplus\theta(x)\) whose off-diagonal entries are \(x-\theta(x)\) and its negative. Applying the amplification theorem to this paired representation converts the ordinary bound on \(\theta-\operatorname{id}\) into a cb bound. 4 then supplies the unitary on the original \(H\). This identifies the precise role of commutator control in the stability argument.

Matrix gaps and commutants at the conclusion

Corollary 52. For every \(\varepsilon>0\) there is \(\delta>0\), independent of the algebras, their representations, and their Hilbert spaces, such that unital von Neumann algebras \(M,N\subseteq\mathcal B(H)\) with the same identity and \(d(M,N)<\delta\) satisfy \[d_{\mathrm{cb}}(M,N)<\varepsilon,\qquad d(M',N')<\varepsilon.\]

Proof. Choose \(\delta\) from 1 with unitary displacement \(\varepsilon/2\). The resulting unitary satisfies \(uMu^*=N\) and \(\lVert u-I\rVert<\varepsilon/2\). Conjugation by its diagonal amplification identifies the unit balls of \(M_n(M)\) and \(M_n(N)\) and moves each contraction by at most \(2\lVert u-I\rVert\), independently of \(n\). The same estimate applies to the commutants because \(uM'u^*=N'\). ◻

Since \(d(M,N)\le d_{\mathrm{cb}}(M,N)\), this corollary shows that the ordinary and completely bounded distances induce the same uniform structure on the unital von Neumann algebras in each \(\mathcal B(H)\), with tolerances uniform over \(H\). It also gives uniform continuity of the commutant operation for the ordinary distance. These controls follow at the conclusion from one small implementing unitary; the separate estimates in the proof are what make construction of that unitary possible.

Adams, Scot, George A. Elliott, and Thierry Giordano. 1994. “Amenable Actions of Groups.” Transactions of the American Mathematical Society 344 (2): 803–22. https://doi.org/10.1090/S0002-9947-1994-1250814-5.
Arveson, William. 1975. “Interpolation Problems in Nest Algebras.” Journal of Functional Analysis 20: 208–33. https://doi.org/10.1016/0022-1236(75)90041-5.
Björklund, Michael. 2014. Five Remarks about Random Walks on Groups. https://arxiv.org/abs/1406.0763v2.
Blackadar, Bruce. 2006. Operator Algebras: Theory of \(C^*\)-Algebras and von Neumann Algebras. Vol. 122. Encyclopaedia of Mathematical Sciences. Springer. https://doi.org/10.1007/3-540-28517-2.
Burger, Marc, Narutaka Ozawa, and Andreas Thom. 2013. “On Ulam Stability.” Israel Journal of Mathematics 193 (1): 109–29. https://doi.org/10.1007/s11856-012-0050-z.
Cameron, Jan, Erik Christensen, Allan M. Sinclair, Roger R. Smith, Stuart White, and Alan D. Wiggins. 2012. “Type \(\mathrm{II}_1\) Factors Satisfying the Spatial Isomorphism Conjecture.” Proceedings of the National Academy of Sciences of the United States of America 109 (50): 20338–43. https://doi.org/10.1073/pnas.1217792109.
Cameron, Jan, Erik Christensen, Allan M. Sinclair, Roger R. Smith, Stuart White, and Alan D. Wiggins. 2013. “A Remark on the Similarity and Perturbation Problems.” Comptes Rendus Mathématiques de l’Académie Des Sciences, La Société Royale Du Canada 35 (2): 70–76. https://arxiv.org/abs/1206.5405.
Cameron, Jan, Erik Christensen, Allan M. Sinclair, Roger R. Smith, Stuart White, and Alan D. Wiggins. 2014. “Kadison–Kastler Stable Factors.” Duke Mathematical Journal 163 (14): 2639–86. https://doi.org/10.1215/00127094-2819736.
Cameron, Jan, Erik Christensen, Allan M. Sinclair, Roger R. Smith, Stuart White, and Alan D. Wiggins. 2017. “Structural Properties of Close \(\mathrm{II}_1\) Factors.” Münster Journal of Mathematics 10: 19–37. https://doi.org/10.17879/33249460964.
Christensen, Erik. 1975. “Perturbations of Type I von Neumann Algebras.” Journal of the London Mathematical Society, 2nd series, vol. 9 (3): 395–405. https://doi.org/10.1112/jlms/s2-9.3.395.
Christensen, Erik. 1977a. “Perturbation of Operator Algebras.” Inventiones Mathematicae 43: 1–13. https://doi.org/10.1007/BF01390201.
Christensen, Erik. 1977b. “Perturbations of Operator Algebras II.” Indiana University Mathematics Journal 26 (5): 891–904. https://doi.org/10.1512/iumj.1977.26.26072.
Christensen, Erik. 1980. “Near Inclusions of \(C^*\)-Algebras.” Acta Mathematica 144: 249–65. https://doi.org/10.1007/BF02392125.
Christensen, Erik. 1986. “Similarities of \(\mathrm{II}_1\) Factors with Property \(\Gamma\).” Journal of Operator Theory 15 (2): 281–88. https://jot.theta.ro/jot/archive/1986-015-002/1986-015-002-005.html.
Christensen, Erik. 2001. “Finite von Neumann Algebra Factors with Property \(\Gamma\).” Journal of Functional Analysis 186 (2): 366–80. https://doi.org/10.1006/jfan.2001.3812.
Christensen, Erik, Allan M. Sinclair, Roger R. Smith, and Stuart A. White. 2010. “Perturbations of \(C^*\)-Algebraic Invariants.” Geometric and Functional Analysis 20 (2): 368–97. https://doi.org/10.1007/s00039-010-0070-y.
Christensen, Erik, Allan M. Sinclair, Roger R. Smith, Stuart A. White, and Wilhelm Winter. 2010. “The Spatial Isomorphism Problem for Close Separable Nuclear \(C^*\)-Algebras.” Proceedings of the National Academy of Sciences of the United States of America 107 (2): 587–91. https://doi.org/10.1073/pnas.0913281107.
Christensen, Erik, Allan M. Sinclair, Roger R. Smith, Stuart A. White, and Wilhelm Winter. 2012. “Perturbations of Nuclear \(C^*\)-Algebras.” Acta Mathematica 208 (1): 93–150. https://doi.org/10.1007/s11511-012-0075-5.
Haagerup, Uffe. 1975. “The Standard Form of von Neumann Algebras.” Mathematica Scandinavica 37: 271–83. https://doi.org/10.7146/math.scand.a-11606.
Haagerup, Uffe. 1979. “An Example of a Non Nuclear \(C^*\)-Algebra, Which Has the Metric Approximation Property.” Inventiones Mathematicae 50 (3): 279–93. https://doi.org/10.1007/BF01410082.
Haagerup, Uffe. 1983. “Solution of the Similarity Problem for Cyclic Representations of \(C^*\)-Algebras.” Annals of Mathematics, 2nd series, vol. 118 (2): 215–40. https://doi.org/10.2307/2007028.
Haagerup, Uffe. 1987. “Connes’ Bicentralizer Problem and Uniqueness of the Injective Factor of Type \(\mathrm{III}_1\).” Acta Mathematica 158: 95–148. https://doi.org/10.1007/BF02392257.
Herstein, I. N. 1956. “Jordan Homomorphisms.” Transactions of the American Mathematical Society 81 (2): 331–41. https://doi.org/10.1090/S0002-9947-1956-0076751-6.
Houdayer, Cyril, and Amine Marrakchi. 2026. The Classification of Flows on \(\mathrm{II}_1\) Factors and Connes’ Bicentralizer Problem. https://arxiv.org/abs/2609.11462v1.
Ino, Shoji. 2016. “Perturbations of von Neumann Subalgebras with Finite Index.” Canadian Mathematical Bulletin 59 (2): 320–25. https://doi.org/10.4153/CMB-2015-081-8.
Johnson, Barry E. 1977. “Perturbations of Banach Algebras.” Proceedings of the London Mathematical Society, 3rd series, vol. 34 (3): 439–58. https://doi.org/10.1112/plms/s3-34.3.439.
Johnson, Barry E. 1982. “A Counterexample in the Perturbation Theory of \(C^*\)-Algebras.” Canadian Mathematical Bulletin 25 (3): 311–16. https://doi.org/10.4153/CMB-1982-043-4.
Kadison, Richard V., and Daniel Kastler. 1972. “Perturbations of von Neumann Algebras. I. Stability of Type.” American Journal of Mathematics 94 (1): 38–54. https://doi.org/10.2307/2373592.
Kaimanovich, Vadim A. 2003. “Double Ergodicity of the Poisson Boundary and Applications to Bounded Cohomology.” Geometric and Functional Analysis 13: 852–61. https://doi.org/10.1007/s00039-003-0433-8.
Kazhdan, David. 1982. “On \(\varepsilon\)-Representations.” Israel Journal of Mathematics 43 (4): 315–23. https://doi.org/10.1007/BF02761236.
Kechris, Alexander S. 1995. Classical Descriptive Set Theory. Vol. 156. Graduate Texts in Mathematics. Springer. https://doi.org/10.1007/978-1-4612-4190-4.
Marrakchi, Amine. 2020. “Full Factors, Bicentralizer Flow and Approximately Inner Automorphisms.” Inventiones Mathematicae 222: 375–98. https://doi.org/10.1007/s00222-020-00971-w.
OpenAI. 2026. Kadison’s similarity theorem through uniform derivation estimates. OpenAI Math Release preprint OAI:Kadisons-similarity-theorem-through-uniform-derivation-estimates-September-23-2026.
Phillips, John. 1974. “Perturbations of Type I von Neumann Algebras.” Pacific Journal of Mathematics 52 (2): 505–11. https://msp.org/pjm/1974/52-2/pjm-v52-n2-p19-s.pdf.
Popa, Sorin. 1981. “On a Problem of R. V. Kadison on Maximal Abelian \(*\)-Subalgebras in Factors.” Inventiones Mathematicae 65: 269–81. https://doi.org/10.1007/BF01389015.
Popa, Sorin. 2014. “Independence Properties in Subalgebras of Ultraproduct \(\mathrm{II}_1\) Factors.” Journal of Functional Analysis 266 (9): 5818–46. https://doi.org/10.1016/j.jfa.2014.02.004.
Raeburn, Iain, and Joseph L. Taylor. 1977. “Hochschild Cohomology and Perturbations of Banach Algebras.” Journal of Functional Analysis 25 (3): 258–66. https://doi.org/10.1016/0022-1236(77)90072-6.
Ricard, Éric, and Jean Roydor. 2014. “A Noncommutative Amir–Cambern Theorem for von Neumann Algebras and Nuclear \(C^*\)-Algebras.” Journal of Functional Analysis 267 (4): 1121–36. https://doi.org/10.1016/j.jfa.2014.05.018.
Ricard, Éric, and Quanhua Xu. 2006. “Khintchine Type Inequalities for Reduced Free Products and Applications.” Journal für Die Reine Und Angewandte Mathematik 599: 27–59. https://doi.org/10.1515/CRELLE.2006.077.
Sakai, Shôichirô. 1971. “Derivations of Simple \(C^*\)-Algebras. II.” Bulletin de La Société Mathématique de France 99: 259–63. https://doi.org/10.24033/bsmf.1719.
Takesaki, Masamichi. 1973. “Duality for Crossed Products and the Structure of von Neumann Algebras of Type III.” Acta Mathematica 131: 249–310. https://doi.org/10.1007/BF02392041.
Takesaki, Masamichi. 2003. Theory of Operator Algebras II. Vol. 125. Encyclopaedia of Mathematical Sciences. Springer. https://doi.org/10.1007/978-3-662-10451-4.
Voiculescu, Dan. 1985. “Symmetries of Some Reduced Free Product \(C^*\)-Algebras.” In Operator Algebras and Their Connections with Topology and Ergodic Theory, edited by Huzihiro Araki, Calvin C. Moore, Şerban-Valentin Strătilă, and Dan-Virgil Voiculescu, vol. 1132. Lecture Notes in Mathematics. Springer. https://doi.org/10.1007/BFb0074909.
Zimmer, Robert J. 1978. “Amenable Ergodic Group Actions and an Application to Poisson Boundaries of Random Walks.” Journal of Functional Analysis 27 (3): 350–72. https://doi.org/10.1016/0022-1236(78)90013-7.
LEVEL 1 COMPLETE!
You read 28,981 words and 3,065 formulas. Your math teacher would be proud.
Converted from the LaTeX source. Something look off? The original PDF is the real thing.

Cool Links: openai/math   Lean   Mathlib   arXiv   the real Coolmath Games