Let $P=\{1,\dots,n\}$ be a set of parties with local Hilbert spaces $\mathcal H^{(a)}$ of finite dimensions $d_a$. Following the notation of the local-invariant correlation-sector decomposition introduced by Aschauer \emph{et al.}\cite{aschauer}, choose for each party $a$ traceless Hermitian generators
\begin{equation*}
\sigma^{(a)}_1,\dots,\sigma^{(a)}_{d_a^2-1}
\end{equation*}
together with the identity $\sigma_0^{(a)}=\id$, normalized by
for $i,j\in\{0,1,\dots,d_a^2-1\}$, so that the orthogonality relation already includes the identity generator. For each nonempty subset $S\subseteq P$ we denote by
They are local-unitary invariants and admit a clean convexity-based entanglement criterion, but they retain only the total quadratic size of each tensor block.
The point of the present note is that the old geometric picture has a natural multipartite continuation. Instead of compressing each block to one scalar, we keep the full family of source-indexed correlation-response maps produced by a chosen source party and stack the responses landing in the different orthogonal sectors of the complement. In this sense the main contribution is best viewed as a new object architecture rather than a new functional applied to a familiar full correlation tensor. The general construction and cut-separable bound below work for arbitrary finite local dimensions; the later biseparable benchmark and all numerical examples then specialize to qubits, where the constants and geometry are especially transparent.
A substantial literature already studies separability criteria in the Bloch-representation and correlation-tensor language, including bipartite correlation-matrix criteria \cite{devicente,chenwu}, multipartite unfoldings and matricizations of the full correlation tensor \cite{hassanjoag,devicentehuber,li2014,jingzhang2023}, nonlinear geometric tensor criteria \cite{laskowski2011}, scalar multi-sector norm criteria \cite{klocklhuber2015}, and more recent extended or partition-adapted mixed-order block constructions \cite{shen2016,sarbicki2020,zhao2020,huang2024extended,liyaoyangfei2025}. This note is intended to sit explicitly within that landscape rather than outside it.
Historically, it is useful to distinguish two nearby lines of work. The local-invariant sector decomposition of Aschauer \emph{et al.}\cite{aschauer} already gave an early multipartite entanglement criterion in terms of the coefficients of a local operator expansion; that expansion is carried out, from the first equations of that paper onward, in the full product basis of local generators together with the identity on each party, so the full tensor $\mathcal C(\rho)$ used in Section~\ref{sec:tensor-viewpoint} below is already present there, and the sector-restricted tensor $C_S(\rho)$ used for $L_S$ is obtained from it by exactly the restriction to nonidentity indices already carried out in that paper. Later work such as Hassan and Joag \cite{hassanjoag} made the Bloch-representation terminology explicit and developed criteria from full-tensor unfoldings. The present construction belongs to that broader correlation-tensor lineage, but its organization is closest in spirit to, and its starting tensor is literally the one already used in, \cite{aschauer}.
It is worth stating carefully what is and is not being claimed here. The recent correlation-tensor literature already contains powerful criteria based on full tensors, matricizations, Bloch tensors, and cut-aware mixed-order block trace norms. In particular, recent generalized-Bloch constructions already study multipartite cut-aware mixed-order block trace-norm criteria \cite{liyaoyangfei2025}. Thus the intended novelty claim is narrow: not the first multipartite block trace-norm criterion of this general kind, but the specific direct-sum organization obtained by fixing one source party $a$ and stacking
Simultaneous bigraduation of the same cut matrix by \emph{both} source sector $V$ and target sector $T$, rather than one collapsed block & Definition~\ref{def:bigraduated}\\
Sub-block and sector-profile witnesses that this bigraduation makes available automatically, with no separate proof & Corollary~\ref{cor:sub-block}, profiles $\Phi_{a\to T}$, $\Phi^{(2)}_{a\to T}$\\
Sector profiles are demonstrably \emph{not} invariant under unitaries acting collectively on the coarse complement, even though the full cut norm is & Proposition following Corollary~\ref{cor:projection-bound}, Bell-product/CNOT example \\
Every shadow map, at every graduation level, is one matricization of a single augmented Bloch tensor, so the $\le1$ bound is inherited rather than reproved at each level & Theorem~\ref{thm:one-fact}, Section~\ref{sec:tensor-viewpoint}\\
An explicit, fully proved three-qubit biseparable threshold in closed form & Eq.~\eqref{eq:bisep-threshold}\\
A qutrit PPT-entangled benchmark and a systematic scan of $38$ four-qubit graph states and $n=3,4,5$ state families & qutrit benchmark in Section~\ref{sec:tensor-viewpoint}, Section~\ref{sec:qubit-numerics}\\
\bottomrule
\end{tabular}
\caption{What this note recovers from the existing correlation-tensor literature (top) versus what it contributes on top of that baseline (bottom). The headline bound $\norm{\mathcal M_S(\rho)}_*\le1$ itself belongs to the top half; the substantive claims of the note are the structural and numerical items in the bottom half.}
A guiding thread throughout the note is that this architecture is not tied to a single source party. Sections~\ref{sec:response-maps}--\ref{sec:multiparty-sources} show that fixing a source \emph{cluster}$S\subseteq P$ instead of a single party costs nothing in the proof: the rank-one mechanism behind the cut-separable bound is agnostic to how many parties sit on the source side. Section~\ref{sec:tensor-viewpoint} then makes precise in what sense this is not a coincidence: every construction in this note---the single-party map $\mathcal M_a$, its cluster generalization $\mathcal M_S$, and every sub-block compression of either---is a matricization or sub-block restriction of one and the same full correlation tensor, and the bound $\le1$ is a single rank-one fact about that tensor, inherited unchanged through each linear operation. The single-party case is simply the version of this fact that is cheapest to state first.
As a concrete demonstration that this stacked structure detects entanglement invisible to the standard PPT test, Section~\ref{sec:tensor-viewpoint} exhibits a two-qutrit state built from the Tiles unextendible product basis of \cite{bennettUPB}: it is numerically PPT to machine precision, so the Peres--Horodecki criterion is silent on it \cite{peres,horodeckiPPT}, yet the shadow-map witness detects its entanglement outright. We flag this example here because it is, in our view, the single clearest piece of evidence in the note that the construction has practical bite beyond reorganizing known bounds.
\subsection*{Notation guide}
The construction accumulates several closely related maps and scalars as it is refined step by step; Table~\ref{tab:notation} collects the main ones for reference, in the order they are introduced.
\begin{table}[h]
\centering
\small
\begin{tabular}{@{}lp{0.72\textwidth}@{}}
\toprule
Symbol & Meaning \\
\midrule
$C_S(\rho)$, $L_S(\rho)$& sector correlation tensor and its squared Frobenius norm (Aschauer \emph{et al.}\cite{aschauer}) \\
$\widetilde{\mathcal M}_a(\rho)$& intrinsic (basis-free) one-vs-rest response map, source party $a$\\
$\mathcal M_a(\rho)$& combined, normalized shadow map for a single source party $a$ (Definition~\ref{def:combined-shadow}) \\
$\widehat M_{a\to T}(\rho)$& sector shadow map, normalized independently of the other sectors \\
$\Phi_{a\to T}(\rho)$, $\Phi^{(2)}_{a\to T}(\rho)$& sector nuclear and Frobenius shadow profiles, $\norm{\widehat M_{a\to T}}_*$ and $\norm{\widehat M_{a\to T}}_\fro^2$\\
$\Phi_a^{(\le k)}(\rho)$, $\Phi_a^{(\ge k)}(\rho)$& nuclear norm after projecting onto grouped target sectors of order $\le k$ or $\ge k$\\
$\mathcal M_S(\rho)$& bigraduated shadow map for a source \emph{cluster}$S$ (Definition~\ref{def:bigraduated}); reduces to $\mathcal M_a(\rho)$ when $S=\{a\}$\\
$M_{V\to T}(\rho)$& bigraduated block from source sector $V\subseteq S$ to target sector $T\subseteq S^c$\\
$\Phi_{\mathrm{sym}}(\rho)$, $\Phi_{\max}(\rho)$& source-aggregated functionals, average and maximum of $\norm{\mathcal M_a(\rho)}_*$ over $a\in P$\\
$\mathcal C(\rho)$& full order-$n$ Bloch tensor of which every map above is a matricization or slice (Definition~\ref{def:full-tensor}) \\
$A_\lambda$& reduced shadow map on the multiplicity space of isotype $\lambda$, once $\rho$ carries a compatible symmetry (Section~\ref{sec:symmetry-blocks}) \\
\bottomrule
\end{tabular}
\caption{Main notation, in order of introduction. $\Phi_{\mathrm{sym}}$, $\Phi_{\max}$, and the biseparable threshold are qubit-specific; everything above them in the table is defined for arbitrary finite local dimensions.}
Fix a party $a\in P$, and write $\bar a=P\setminus\{a\}$. For each party $b$, let $\V^{(b)}$ be the real Hilbert space of Hermitian operators on $\mathcal H^{(b)}$, equipped with the Hilbert-Schmidt inner product, and let
After choosing orthonormal bases in $\V_0^{(a)}$ and in each $\V_T^{(\bar a)}$, these become the coordinate maps used below. We will use the shorter term \emph{shadow maps} for this family. We emphasize that this usage is unrelated to the classical-shadows measurement protocols of \cite{huangkuengpreskill2020}: the shadow maps of this note are linear response operators built directly from the correlation tensor of $\rho$, not estimators reconstructed from randomized single-copy measurements.
Under the orthogonal decomposition in Eq.~\eqref{eq:complement-bloch-decomposition}, this is just the matrix representation of the intrinsically defined map $\widetilde{\mathcal M}_a(\rho)$ in sector-adapted orthonormal coordinates. We refer to this direct-sum response operator as the \emph{combined shadow map}. The image of the unit ball in $\R^{d_a^2-1}$ under $\mathcal M_a(\rho)$ is the corresponding response ellipsoid in correlation space, which we also call the combined shadow ellipsoid of the party $a$.
\end{definition}
For qubits this reduces to the earlier normalization, since $(d_a-1)(d_{\bar a}-1)=2^{n-1}-1$. In particular, for three qubits this direct sum is
\begin{equation*}
\mathcal W_A=\R^3\oplus\R^3\oplus\R^9,
\end{equation*}
corresponding to the $B$, $C$, and $BC$ response sectors.
\section{Cut-separable states}
\label{sec:cut-separable}
\subsection*{The cut-separable bound and its refinements}
The key fact is easiest to see first for states that are product across $a\mid\bar a$: then the full response operator is rank one, and the sector maps are simply its orthogonal components. The general cut-separable case follows by convexity.
\begin{theorem}
\label{thm:cut-bound}
Let $\rho$ be separable across the cut $a\mid\bar a$. Then
\begin{equation}
\norm{\mathcal M_a(\rho)}_*\le 1.
\label{eq:cut-bound}
\end{equation}
Consequently,
\begin{equation*}
\norm{\mathcal M_a(\rho)}_*>1
\qquad\Longrightarrow\qquad
\rho\text{ is entangled across }a\mid\bar a.
\end{equation*}
\end{theorem}
\begin{proof}
It is enough to begin with a product state across the cut,
\begin{equation*}
\rho=\rho_a\otimes\sigma_{\bar a}.
\end{equation*}
Let $r^{(a)}\in\V_0^{(a)}$ be the Bloch vector of $\rho_a$, defined by
\begin{equation*}
\langle X,r^{(a)}\rangle=\tr(\rho_a X),
\qquad X\in\V_0^{(a)},
\end{equation*}
and let $v_{\bar a}\in\V_0^{(\bar a)}$ be the traceless Bloch vector of $\sigma_{\bar a}$, defined analogously. Then
is rank one, and in sector-adapted coordinates this becomes the direct sum of the component maps. Equivalently, for each nonempty $T\subseteq\bar a$ let
using the correlation-sum identity from the Aschauer \emph{et al.} framework for the $(n-1)$-party state $\sigma_{\bar a}$. Thus Eq.~\eqref{eq:cut-bound} holds for every product state across the cut.
Now let $\rho$ be mixed and separable across the cut,
Let $\Pi$ be any orthogonal projection on $\V_0^{(\bar a)}$, and let $\Pi\mathcal M_a(\rho)$ denote the corresponding projected map in any orthonormal coordinates adapted to the decomposition of $\V_0^{(\bar a)}$. If $\rho$ is separable across the cut $a\mid\bar a$, then
\begin{equation}
\norm{\Pi\mathcal M_a(\rho)}_*\le 1.
\label{eq:projection-bound}
\end{equation}
Consequently, every orthogonally selected target subspace of $\V_0^{(\bar a)}$ yields a valid cut witness.
\end{corollary}
\begin{proof}
Orthogonal projection is contractive for the operator norm and hence for singular values. Therefore
The claim follows from Theorem~\ref{thm:cut-bound}.
\end{proof}
\begin{remark}
Equation~\eqref{eq:projection-bound} produces a whole hierarchy of weaker but natural cut witnesses. Besides the individual sectors $\Pi=P_T$, one can project onto grouped sector subspaces. For instance, for $1\le k\le |\bar a|$ let
Now $C_T(\sigma_{\bar a})$ depends only on the reduced state $\sigma_T$, so
\begin{equation*}
\norm{v_T}^2 = L_T(\sigma_T)
\le
\sum_{\emptyset\neq U\subseteq T}L_U(\sigma_T)
= d_T\tr(\sigma_T^2)-1
\le d_T-1.
\end{equation*}
Together with $\norm{r^{(a)}}^2\le d_a-1$, this proves the product-state case. The mixed separable case again follows by linearity and convexity, and the Frobenius statement follows from $\norm{X}_{\fro}\le\norm{X}_*$.
\end{proof}
\begin{definition}
For each nonempty subset $T\subseteq\bar a$, define the sector nuclear shadow profile by
which depends not only on the sizes of the individual sector maps but also on how their right-singular directions align in the common source space. Thus the full nuclear-norm signal is not, in general, a linear combination of the individual sector nuclear norms. In the cut-product case all sector maps share one common right factor, so the full map is again rank one and the proof of Theorem~\ref{thm:cut-bound} reduces to one Euclidean bound on the stacked target vector.
\end{remark}
\subsection*{Sector profiles are not cut invariants}
The sector profiles $\Phi_{a\to T}(\rho)$ and $\Phi^{(2)}_{a\to T}(\rho)$ are therefore not invariants of the coarse cut $a\mid\bar a$ under general collective complement unitaries; they are invariants only under unitaries that preserve the chosen internal factorization of $\bar a$ into parties.
\end{proposition}
\begin{proof}
Choose orthonormal bases of traceless Hermitian operators on $\mathcal H^{(a)}$ and $\mathcal H^{(\bar a)}$, normalized by
\begin{equation*}
\tr(\sigma_i\sigma_j)=d_a\,\delta_{ij},
\qquad
\tr(\tau_i\tau_j)=d_{\bar a}\,\delta_{ij}.
\end{equation*}
The intrinsic response map transforms by the adjoint actions on source and target Bloch spaces:
In the chosen orthonormal bases these adjoint actions are represented by real orthogonal matrices $O_a(U_a)$ and $O_{\bar a}(U_{\bar a})$, giving Eq.~\eqref{eq:two-sided-collective-covariance}. Left and right multiplication by orthogonal matrices preserve both nuclear and Frobenius norms, so the norm equalities follow.
The sector maps arise only after choosing the product operator basis on $\mathcal H^{(\bar a)}$ determined by the internal decomposition of $\bar a$ into parties and then splitting that basis into orthogonal summands. A general collective unitary on $\bar a$ need not preserve those summands, so it reshuffles the sector profile even though the full cut norm is unchanged.
\end{proof}
\begin{remark}
The failure of sector invariance is already visible in the simplest three-qubit Bell-product example. Let
So a collective unitary on $BC$ does not push the signal purely into the highest-order sector. Instead it redistributes a strongly localized $A\to B$ witness into a mixed profile spread across $A\to B$, $A\to C$, and $A\to BC$, while leaving the full $A\mid BC$ witness unchanged.
This persists under white noise. For
\begin{equation*}
\rho_p=p\rho+(1-p)\frac{\id}{8},
\qquad
\rho'_p=p\rho'+(1-p)\frac{\id}{8},
\end{equation*}
the full-map threshold is the same in both forms,
\begin{equation*}
\norm{\mathcal M_A(\rho_p)}_*>1
\iff
\norm{\mathcal M_A(\rho'_p)}_*>1
\iff
p>\frac{1}{\sqrt 6}.
\end{equation*}
But the sector thresholds differ sharply: before the collective rotation the sector $A\to B$ already detects for $p>1/3$, whereas after the rotation the sectors $A\to B$ and $A\to C$ never strictly violate the cut-separable bound and the sector $A\to BC$ only detects for $p>\sqrt{3/8}$. In particular, at $p=0.60$ one has $\norm{\mathcal M_A(\rho'_p)}_*>1$ while all three individual sectors still satisfy $\Phi_{A\to T}(\rho'_p)\le1$.
The construction of Section~\ref{sec:response-maps} singles out one party $a$ as source and treats the entire complement $\bar a$ as target. Nothing in the argument in fact requires $|S|=1$ on the source side; the same object exists for any source cluster $\emptyset\neq S\subsetneq P$, with complement $S^c:=P\setminus S$. Making this explicit exposes a layer of internal structure on the source side that Definition~\ref{def:combined-shadow} discards by construction, and it costs nothing beyond re-reading the definitions and the proof of Theorem~\ref{thm:cut-bound} with $a$ replaced by $S$.
\subsection*{Source sector decomposition}
For nonempty $S\subseteq P$, set $\V^{(S)}:=\bigotimes_{a\in S}\V^{(a)}$ and apply exactly the decomposition of Eq.~\eqref{eq:complement-bloch-decomposition}, now with $S$ in the role previously played by $\bar a$:
This is not a new construction, only the source-side instance of the same orthogonal sector decomposition already used for the complement. The traceless source space is accordingly
graded by which parties within $S$ are active. For $S=\{a\}$ the only nonempty $V\subseteq S$ is $V=S$ itself, so this decomposition is trivial for a singleton source; the bigraduation below is genuinely new structure only once $|S|\ge2$.
is given by the same formula as Eq.~\eqref{eq:intrinsic-response}, with $a$ replaced by $S$. Both $\V_0^{(S)}$ and $\V_0^{(S^c)}$ carry an orthogonal sector decomposition, Eq.~\eqref{eq:source-bloch-decomposition} on the source side and Eq.~\eqref{eq:complement-bloch-decomposition} on the target side, so the coordinate representation of $\widetilde{\mathcal M}_S(\rho)$ is naturally \emph{bigraded} by source sector $V\subseteq S$ and target sector $T\subseteq S^c$ simultaneously. Writing $\iota_V:\V_V^{(S)}\hookrightarrow\V_0^{(S)}$ for the canonical inclusion of a source sector and $P_T:\V_0^{(S^c)}\to\V_T^{(S^c)}$ for the orthogonal projection onto a target sector, define
i.e.\ the matrix representation of $\widetilde{\mathcal M}_S(\rho)$, normalized exactly as in Eq.~\eqref{eq:combined-map}, in sector-adapted orthonormal coordinates on both sides.
\end{definition}
For $S=\{a\}$, Definition~\ref{def:bigraduated} reduces exactly to Definition~\ref{def:combined-shadow}: the source-side decomposition then has only the single summand $V=S$, so the bigraduation collapses to the ordinary target-only graduation of $\mathcal M_a$. Definition~\ref{def:combined-shadow} is thus the singleton case of this construction, not a separate object introduced in parallel to it.
The total number of scalar entries in $\mathcal M_S(\rho)$ is $(d_S-1)(d_{S^c}-1)$, exactly as for the collapsed cut matrix, so evaluating the norm $\norm{\mathcal M_S(\rho)}_*$ in Theorem~\ref{thm:cluster-cut} costs a single singular value decomposition of that size and is no more expensive than the realignment-type bounds it recovers. The bigraduation of Definition~\ref{def:bigraduated} does not change this cost; it only reorganizes the same entries into $(2^{|S|}-1)(2^{|S^c|}-1)$ combinatorial blocks $M_{V\to T}$. Consequently, using the full map $\mathcal M_S(\rho)$ as a single witness remains cheap, but exploiting the sub-block witnesses of Corollary~\ref{cor:sub-block} exhaustively --- inspecting every source sector against every target sector separately, rather than only the handful used in the examples below --- requires examining up to $O(2^{|S|+|S^c|})=O(2^n)$ individual blocks in the worst case. Section~\ref{sec:symmetry-blocks} shows that when $\rho$ carries a compatible symmetry, this exponential proliferation of combinatorial blocks collapses instead onto a typically much smaller number of representation-theoretic multiplicity spaces, giving a tractable route to the same sub-block information without enumerating every sector by hand.
Let $\rho$ be separable across the cut $S\mid S^c$. Then
\begin{equation}
\norm{\mathcal M_S(\rho)}_*\le1.
\label{eq:cluster-cut-bound}
\end{equation}
Consequently, $\norm{\mathcal M_S(\rho)}_*>1$ implies that $\rho$ is entangled across $S\mid S^c$.
\end{theorem}
\begin{proof}
The proof of Theorem~\ref{thm:cut-bound} goes through with $a\to S$ and $\bar a\to S^c$ without modification. For a product state $\rho=\rho_S\otimes\sigma_{S^c}$, let $r^{(S)}\in\V_0^{(S)}$ be the traceless Bloch vector of the $|S|$-party reduced state $\rho_S$ -- the same object as $r^{(a)}$ in the proof of Theorem~\ref{thm:cut-bound}, now for the composite system $S$ treated as a single $d_S$-dimensional party -- and let $v_{S^c}\in\V_0^{(S^c)}$ be defined analogously from $\sigma_{S^c}$. Tensor factorization gives $\widetilde{\mathcal M}_S(\rho)=v_{S^c}(r^{(S)})^T$, which is rank one, and
by the same correlation-sum identity used in the proof of Theorem~\ref{thm:cut-bound}, now applied to the $|S|$-party state $\rho_S$ and the $|S^c|$-party state $\sigma_{S^c}$ rather than to single-party marginals. Hence $\norm{\mathcal M_S(\rho)}_*\le1$ for every product state across the cut, and convexity of the nuclear norm extends the bound to mixtures exactly as in the proof of Theorem~\ref{thm:cut-bound}.
\end{proof}
\begin{corollary}[Sub-block witnesses]
\label{cor:sub-block}
Let $\mathcal V\subseteq\{V:\emptyset\neq V\subseteq S\}$ and $\mathcal T\subseteq\{T:\emptyset\neq T\subseteq S^c\}$ be any nonempty families, and let $P_{\mathcal V}$, $P_{\mathcal T}$ be the orthogonal projections onto $\bigoplus_{V\in\mathcal V}\V_V^{(S)}$ and $\bigoplus_{T\in\mathcal T}\V_T^{(S^c)}$ respectively. If $\rho$ is separable across $S\mid S^c$, then
Orthogonal projections are contractions for the operator norm, and the nuclear norm satisfies $\norm{AXB}_*\le\norm{A}_{\mathrm{op}}\norm{X}_*\norm{B}_{\mathrm{op}}$ for linear maps $A,B$ of compatible size. Taking $A=P_{\mathcal T}$ and $B=P_{\mathcal V}$ gives Eq.~\eqref{eq:sub-block-bound}; the claim then follows from Theorem~\ref{thm:cluster-cut}. This is the two-sided extension of Corollary~\ref{cor:projection-bound}, which only ever compressed the target side.
\end{proof}
In particular, taking $\mathcal V=\{S\}$ isolates the single block $M_{S\to T}$, which carries only the correlation attributable to the full cluster $S$ acting jointly rather than to any proper sub-cluster of $S$ -- a witness targeted specifically at ``genuine $S$'' structure landing in the sector $T$.
\subsection*{Example: the Smolin state under two different cuts}
where here and in what follows for qubits $X,Y,Z$ denote the Pauli matrices, i.e.\ $\sigma_1=X$, $\sigma_2=Y$, $\sigma_3=Z$ in the generator convention fixed formally in Section~\ref{sec:tensor-viewpoint} below. We evaluate it under two cuts side by side.
\emph{The $2\mid2$ cut.} Take the cluster $S=\{A,B\}$, $S^c=\{C,D\}$. Every block $M_{V\to T}$ with $V\neq S$ or $T\neq S^c$ vanishes identically, because the Smolin state carries no one- or three-body correlations. Only $M_{S\to S^c}$ survives, as the $9\times9$ matrix diagonal on the aligned Pauli directions $(x,x)\to(x,x)$, $(y,y)\to(y,y)$, $(z,z)\to(z,z)$ with unit coefficients and zero elsewhere. After the normalization $1/\sqrt{(d_S-1)(d_{S^c}-1)}=1/3$ this gives three singular values of $1/3$ each, so
\begin{equation*}
\norm{\mathcal M_{AB}(\rho_{\mathrm{Smo}})}_*=1
\quad\text{exactly --- saturating, not violating, the bound of Theorem~\ref{thm:cluster-cut}.}
\end{equation*}
\emph{The $1\mid3$ cut.} Take instead a single source party $a$, so $\bar a$ is the remaining three-qubit cluster. Again only the full four-body sector survives, giving three orthogonal response directions with unnormalized singular value $1$ each. The normalization is now $1/\sqrt{(d_a-1)(d_{\bar a}-1)}=1/\sqrt{1\cdot7}=1/\sqrt7$, so
strictly violating the bound of Theorem~\ref{thm:cut-bound}.
So the same state sits exactly at the boundary for every $2\mid2$ cut while clearly violating the $1\mid3$ bound --- the two cut types are correctly told apart within one framework. The refinement from singleton to cluster sources costs nothing in the proof yet makes this distinction available at all, since the singleton construction of Definition~\ref{def:combined-shadow} cannot even pose the $2\mid2$ question. Section~\ref{sec:qubit-numerics} below returns to the $1\mid3$ value in the source-aggregated language of $\Phi_{\mathrm{sym}}$ and $\Phi_{\max}$, and adds the white-noise robustness threshold.
%% \subsection*{Example: the Smolin state under a $2\mid2$ cut}
%% (examined again from the single-party viewpoint in Section~\ref{sec:qubit-numerics} below), and take the cluster $S=\{A,B\}$, $S^c=\{C,D\}$. Every block $M_{V\to T}$ with $V\neq S$ or $T\neq S^c$ vanishes identically, because the Smolin state carries no one- or three-body correlations. Only $M_{S\to S^c}$ survives, as the $9\times9$ matrix diagonal on the aligned Pauli directions $(x,x)\to(x,x)$, $(y,y)\to(y,y)$, $(z,z)\to(z,z)$ with unit coefficients and zero elsewhere. After the normalization $1/\sqrt{(d_S-1)(d_{S^c}-1)}=1/3$ this gives three singular values of $1/3$ each, so
%% \quad\text{exactly --- saturating, not violating, the bound of Theorem~\ref{thm:cluster-cut}.}
%% \end{equation*}
%% This is consistent with the Smolin state being separable across every $2\mid2$ cut while violating the $1\mid3$ bound at $3/\sqrt7\approx1.134$, computed below in Section~\ref{sec:qubit-numerics}. The refinement from singleton to cluster sources costs nothing in the proof yet correctly distinguishes the two cut types, where the singleton construction of Definition~\ref{def:combined-shadow} cannot even pose the $2\mid2$ question.
\section{The tensor viewpoint: shadow maps as unfoldings of one full Bloch tensor}
\label{sec:tensor-viewpoint}
The guiding thread announced in Section~\ref{sec:response-maps}--\ref{sec:multiparty-sources} can now be made precise. Every non-scalar object introduced so far --- $M_a$, its bigraduated extension $\mathcal M_S$, and the sub-block witnesses of Corollary~\ref{cor:sub-block} --- turns out to be a matricization or sub-block restriction of a single order-$n$ tensor built from the full correlation data of $\rho$; the source-aggregated functionals used in the qubit applications below are then averages or maxima of the corresponding one-vs-rest norms. The $\le1$ bound is therefore not a family of independently proved facts, but one rank-one statement about that tensor, observed through different linear lenses.
\begin{definition}[Full Bloch tensor]
\label{def:full-tensor}
For each party $a$, we augment the local index range by including $i_a =0$, which selects the identity $\sigma_0^{(a)}=\id$ already introduced in Section~\ref{sec:old-criterion}. This allows us to define the full (order-$n$) Bloch tensor
Since $\bigotimes_{a\in P}\R^{d_a^2}=\bigotimes_{a\in P}\bigl(\R\oplus\R^{d_a^2-1}\bigr)$ expands by distributivity into $2^n$ orthogonal summands indexed by which legs are trivial, every sector tensor of Eq.~\eqref{eq:corr-tensor-def} is simply a slice of $\mathcal C(\rho)$:
\begin{equation}
C_V(\rho)=\mathcal C(\rho)\big|_{\,i_a\neq0\text{ for }a\in V,\ i_a=0\text{ for }a\notin V},
The sector decomposition used throughout this note, on both the target side (Section~\ref{sec:response-maps}) and the source side (Section~\ref{sec:multiparty-sources}), is therefore not an additional structure imposed on the correlation data: it \emph{is} the tensor-product structure of $\mathcal C(\rho)$ in the identity-plus-generators basis.
\begin{remark}[This is already the tomography tensor of \cite{aschauer}]
\label{rem:aschauer-tensor}
Definition~\ref{def:full-tensor} introduces no object beyond what \cite{aschauer} starts from. Writing out the operator expansion of $\rho$ in the full product basis $\{\bigotimes_{a\in P}\sigma^{(a)}_{i_a}:0\le i_a\le d_a^2-1\}$ used there for state tomography gives exactly
with $\mathcal C(\rho)=(c_{i_1,\dots,i_n})$, over the same unrestricted index range. The sector-restricted tensor $C_S(\rho)$ of Eq.~\eqref{eq:corr-tensor-def}, on which the correlation strengths $L_S$ and every construction built on them in this note ultimately depend, is the further restriction of that same tomography tensor to $i_a>0$ for every $a\in S$. In this sense, Sections~\ref{sec:response-maps}--\ref{sec:multiparty-sources} never leave the object \cite{aschauer} already had in hand; what changes is only what is extracted from it, replacing the scalar sector norm $L_S$ with a matrix unfolding and its singular values. The point of the present section is that this change of extraction is itself best understood at the level of the tensor $\mathcal C(\rho)$, rather than sector by sector.
\end{remark}
We now switch perspectives: the shadow maps of the preceding sections will no longer be treated as separately constructed response operators, but as canonical unfoldings, sector restrictions, and identity-leg slices of the single full tensor $\mathcal C(\rho)$.
\begin{proposition}[Product states are exactly the states whose full Bloch tensor has CP-rank one]
\label{prop:cp-rank-one}
$\rho=\bigotimes_{a\in P}\rho_a$ if and only if $\mathcal C(\rho)=\bigotimes_{a\in P}w^{(a)}$ for vectors $w^{(a)}\in\R^{d_a^2}$ with $w^{(a)}_0=1$.
\end{proposition}
\begin{proof}
($\Rightarrow$) Immediate from multiplicativity of the trace over the tensor factors, with $w^{(a)}_{i_a}:=\tr(\rho_a\sigma^{(a)}_{i_a})$.
($\Leftarrow$) By the orthogonality relation~\eqref{eq:generator-orthogonality}, extended over $i,j\in\{0,\dots,d_a^2-1\}$, the map $\rho\mapsto\mathcal C(\rho)$ is a linear bijection with inverse
Substituting a rank-one $\mathcal C(\rho)=\bigotimes_a w^{(a)}$ into this inversion formula factorizes term by term into $\bigotimes_{a\in P}\rho_a$ with $\rho_a:=d_a^{-1}\sum_{i_a}w^{(a)}_{i_a}\sigma^{(a)}_{i_a}$. The condition $w_0^{(a)}=1$ gives $\tr(\rho_a)=1$ for each $a$; positivity of each $\rho_a$ then follows because $\rho=\bigotimes_a\rho_a$ is positive semidefinite and every factor is Hermitian with trace one and nonzero.
\end{proof}
\begin{proposition}[Shadow maps are unfoldings]
\label{prop:unfolding}
Fix a bipartition $P=S\sqcup S^c$. Grouping the $S$-legs of $\mathcal C(\rho)$ into a single row index and the $S^c$-legs into a single column index is the standard mode-$(S,S^c)$ matricization of $\mathcal C(\rho)$ in the sense of the multilinear singular value decomposition \cite{delathauwer}. Deleting the trivial ($i=0$) row and column --- equivalently, discarding the $V=\emptyset$ and $T=\emptyset$ sectors, which carry no information beyond normalization --- and rescaling by $[(d_S-1)(d_{S^c}-1)]^{-1/2}$ reproduces $M_S(\rho)$ exactly. The bigraduated shadow map $\mathcal M_S(\rho)$ of Definition~\ref{def:bigraduated} is the same unfolding with the row index additionally kept graded by $V\subseteq S$ instead of collapsed.
Fix a bipartition $P=S\sqcup S^c$. For every state separable across $S\mid S^c$, the normalized nuclear-norm bound $\norm{\cdot}_*\le1$ holds not only for the full shadow map $\mathcal M_S(\rho)$, but also for every witness obtained from the same mode-$(S,S^c)$ unfolding of $\mathcal C(\rho)$ by keeping or collapsing the source and target sector gradings, by taking two-sided sector restrictions as in Corollary~\ref{cor:sub-block}, or by taking target-side identity-leg slices corresponding to partial traces as in Corollary~\ref{cor:trace-is-slice} below, with the normalization appropriate to the remaining source and target systems. Thus Theorem~\ref{thm:cut-bound}, its cluster generalization (Theorem~\ref{thm:cluster-cut}), and the sub-block witnesses of Corollary~\ref{cor:sub-block} are not independently proved facts, but one algebraic statement observed through different linear lenses.
\end{theorem}
\begin{proof}
It is enough to consider a product state across the chosen cut, $\rho=\rho_S\otimes\sigma_{S^c}$. Grouping the legs in $S$ and $S^c$, trace multiplicativity gives
so the mode-$(S,S^c)$ unfolding is rank one. After deleting the identity row and column, this is the rank-one matrix $v_{S^c}(r^{(S)})^T$ appearing in the proof of Theorem~\ref{thm:cluster-cut}, and the same correlation-sum identity gives
\begin{equation*}
\norm{r^{(S)}}^2\le d_S-1,
\qquad
\norm{v_{S^c}}^2\le d_{S^c}-1.
\end{equation*}
Hence the normalized full unfolding has nuclear norm at most one for each product term.
Keeping the sector gradings is only a change of coordinates, while collapsing them gives the same matrix representation with grouped row or column indices. Two-sided sector restrictions have the form $Axy^TB=(Ax)(B^Ty)^T$ on each rank-one product term and cannot increase the nuclear norm when $A$ and $B$ are orthogonal projections. Target-side identity-leg slices give the corresponding reduced product tensor and obey the same estimate with the dimensions of the surviving source and target systems. Finally, linearity of $\rho\mapsto\mathcal C(\rho)$ and convexity of the nuclear norm extend the bound from product states to arbitrary mixtures separable across $S\mid S^c$.
\end{proof}
\begin{corollary}[Partial trace is a slice, not a sum]
\label{cor:trace-is-slice}
For $E\subseteq P$ and $\rho_{P\setminus E}:=\tr_E(\rho)$,
\begin{equation}
\mathcal C(\rho_{P\setminus E})=\mathcal C(\rho)\big|_{\,i_a=0\text{ for all }a\in E},
\label{eq:trace-is-slice}
\end{equation}
i.e.\ the marginal's full tensor is the slice of $\mathcal C(\rho)$ at the trivial index on every traced-out leg, not a contraction or summation over $E$. Consequently, if $P=S\sqcup R\sqcup E$ with $S$ a source cluster as in Definition~\ref{def:bigraduated}, $\rho_{SR}:=\tr_E(\rho)$, and $\Pi_R$ denotes the orthogonal projection that annihilates every target sector $T$ with $T\cap E\neq\emptyset$, then exactly
Eq.~\eqref{eq:trace-is-slice} is immediate from $\sigma_0^{(a)}=\id$: setting $i_a=0$ for $a\in E$ in the defining sum of $\mathcal C(\rho)$ inserts the identity on every traced-out leg, which is exactly $\tr_E(\rho)$ evaluated against the remaining generators. For the second claim, apply Proposition~\ref{prop:unfolding} to the slice~\eqref{eq:trace-is-slice}: because $S\cap E=\emptyset$, the source legs are untouched by the slicing, so for every $T\subseteq R$ the unnormalized block $M_{S\to T}$ computed from $\mathcal C(\rho)$ agrees exactly with the one computed from $\mathcal C(\rho_{SR})$. The two combined maps therefore differ only in their normalization constants, $[(d_S-1)(d_{S^c}-1)]^{-1/2}$ for $\mathcal M_S(\rho)$ against $[(d_S-1)(d_R-1)]^{-1/2}$ for $\mathcal M_S(\rho_{SR})$, since the complement of $S$ is $R\cup E$ in the first case and $R$ alone in the second, with $d_{S^c}=d_Rd_E$. Their ratio is exactly the stated factor.
\end{proof}
Eq.~\eqref{eq:trace-rescale} shows that discarding a residual cluster $E$ by tracing it out is a strictly weaker operation than the sub-block compression of Corollary~\ref{cor:sub-block}: the latter only ever shrinks the nuclear norm, while Eq.~\eqref{eq:trace-rescale} rescales it upward by the factor $\sqrt{(d_{S^c}-1)/(d_R-1)}\ge1$, so that a violation of the bound on $\rho_{SR}$ can certify entanglement across $S\mid R$ that survives the complete loss of $E$, a strictly stronger and operationally different statement than merely detecting entanglement somewhere across $S\mid RE$.
\begin{remark}[Why unfold at all]
The full tensor $\mathcal C(\rho)$ carries strictly more information than any single unfolding: two states can share every matricization $M_S$ over all bipartitions and still differ in genuine multi-way structure, exactly as a generic tensor is not determined by its unfoldings alone. The reason this note works with unfoldings rather than $\mathcal C(\rho)$ directly is computational, not conceptual. The CP-rank-one statement for products, Proposition~\ref{prop:cp-rank-one}, is exact and dimension-independent, but the associated \emph{tensor} nuclear norm --- the natural generalization of $\norm{\cdot}_*$ that would witness separability directly on $\mathcal C(\rho)$, as the infimum of $\sum_r|\lambda_r|$ over CP decompositions --- has no polynomial-time algorithm once three or more legs are grouped independently. Every matricization used in this note, by contrast, is an ordinary matrix with a computable singular value decomposition. The constructions of the preceding sections are thus best understood as the maximal set of efficiently computable shadows of one underlying rank-one fact, chosen along the cut structure that is operationally relevant to entanglement questions.
\end{remark}
\subsection*{A qutrit PPT-entangled benchmark}
The dimension-independent normalization is not only a formal convenience. As a two-qutrit test case, consider the Tiles unextendible product basis of Bennett
\emph{et al.}~\cite{bennettUPB}, consisting of the five orthonormal product vectors
This is the standard rank-four PPT-entangled state supported on the completely entangled complement of the UPB. Using Gell-Mann generators scaled by $\sqrt{3/2}$, so that $\tr(\sigma_i\sigma_j)=3\delta_{ij}$ as in Eq.~\eqref{eq:generator-orthogonality}, the bipartite shadow map is the $8\times8$ correlation matrix divided by
Thus the state is PPT up to numerical precision, so the Peres-Horodecki PPT test is silent \cite{peres,horodeckiPPT}, while the shadow-map witness detects its entanglement. This should be read as complementarity rather than domination: the realignment criterion of Chen and Wu \cite{chenwu} also detects this benchmark, with trace norm $1.087412465\ldots$. The script \texttt{scripts/tiles\_upb.py} reproduces the generator normalization, the PPT spectrum, the shadow value, and the realignment comparison.
\subsection*{Qubit Pauli-tensor form}
For qubits we use the convention already anticipated in the Smolin example of Section~\ref{sec:multiparty-sources}: $\sigma_0=\id$ and $\sigma_1=X$, $\sigma_2=Y$, $\sigma_3=Z$. In this notation, the full Bloch tensor $\mathcal C(\rho)$ becomes the familiar Pauli-correlation tensor
and each sector is specified simply by the support pattern of the nonidentity Pauli indices. Thus the one-vs-rest map for a source party $a$ is obtained by unfolding this Pauli tensor with the $a$-leg as source, discarding the all-identity sector on the complement, and multiplying by
\begin{equation*}
\frac{1}{\sqrt{2^{n-1}-1}}.
\end{equation*}
More generally, for a qubit source cluster $S$ the normalized cluster map is the Pauli-tensor unfolding across $S\mid S^c$, with the all-identity source and target sectors removed, scaled by
\begin{equation*}
\frac{1}{\sqrt{(2^{|S|}-1)(2^{|S^c|}-1)}}.
\end{equation*}
This form makes two features of the examples below transparent. First, adding white noise as
\begin{equation*}
\rho(p)=p\rho+(1-p)\frac{\id}{2^n}
\end{equation*}
leaves the all-identity coefficient fixed and multiplies every nonidentity Pauli coefficient by $p$, so every shadow map and every shadow norm scales linearly with $p$. Second, stabilizer and graph states have Pauli tensors supported on their stabilizer groups, with nonzero coefficients equal to $\pm1$. Their shadow maps are therefore normalized signed support-pattern unfoldings, which explains why the numerical graph-state values below are rigid singular-value facts rather than generic floating-point coincidences.
\section{Qubit specialization and source-aggregated benchmarks}
For the remaining benchmarks and numerics we stay in the qubit setting. The single-party response spaces are $\R^3$, the normalization in Eq.~\eqref{eq:combined-map} reduces to $1/\sqrt{2^{n-1}-1}$, and one can derive explicit constants that do not seem to be available so cleanly in higher dimensions.
Hence violating either bound certifies entanglement.
\end{corollary}
\begin{proof}
A fully separable state is separable across every one-vs-rest cut. The claim follows by applying Theorem~\ref{thm:cut-bound} to each party.
\end{proof}
In equal local dimensions these functionals are permutation invariant. More generally, they are source-aggregated cut-sensitive scalars. The average probes all one-vs-rest cuts simultaneously, whereas the maximum asks whether at least one cut exhibits a large combined shadow ellipsoid.
For three qubits the symmetric average can be optimized explicitly over the biseparable set. This is the first place where we use genuinely qubit-specific formulas rather than only the general response-map architecture. Suppose first that
\begin{equation*}
\rho=\rho_A\otimes\sigma_{BC}
\end{equation*}
is product across $A\mid BC$. If the total state is pure, then this is equivalent to being pure and separable across that cut. Let $a\in\R^3$ be the Bloch vector of $\rho_A$, and let
\begin{equation*}
b,c\in\R^3,
\qquad
T\in\R^{3\times 3}
\end{equation*}
be the one- and two-body correlation data of the pure two-qubit state $\sigma_{BC}$. Then $\norm{a}=1$, and Theorem~\ref{thm:cut-bound} gives
the relevant response vectors for each source party are orthogonal, with equal squared norm $2/3$ after the normalization in Eq.~\eqref{eq:combined-map}. Hence the singular values of each $\mathcal M_a$ are all equal to $\sqrt{2/3}$, so
so the criterion correctly detects entanglement across every $1\mid3$ cut, while --- as already noted --- this is not a genuine-multipartite conclusion, since the Smolin state is separable across every $2\mid2$ split. For the white-noise family $p\rho_{\mathrm{Smo}}+(1-p)\id/16$, the fully separable threshold would only be crossed for
\begin{equation*}
p>\frac{\sqrt 7}{3}\approx 0.882,
\end{equation*}
so this example is much more fragile under white noise than the graph-state families discussed below.
Using the \texttt{qtensor} package, we also evaluated the symmetric shadow functionals for all connected labeled graph states on four qubits. Numerically, all $38$ such graph states give the same value,
The same numerical scan shows that this shadow detection is not explained by pairwise entanglement in the reduced states. For those representative families, every two-qubit marginal remains PPT at the threshold $p=\sqrt7/6$, and for the line and ring graph states some of the two-qubit marginals are even maximally mixed. So the shadow functional is responding to multipartite correlation structure that is not visible in pair reductions. This is still not a four-qubit genuine-multipartite-entanglement proof, because the corresponding biseparable threshold is not yet known, but it makes the criterion promising as a genuinely multipartite diagnostic.
The ring graph state also gives a compact illustration of what the bigraduated $2\mid2$ map sees beyond the source-aggregated scalars. Let $\rho_{\square}$ be the four-qubit graph state on the cycle with edges $(1,2),(2,3),(3,4),(4,1)$. Evaluating the normalized cluster map of Eq.~\eqref{eq:bigraduated-map} across the two inequivalent $2\mid2$ cuts gives
\begin{center}
\begin{tabular}{lccc}
\toprule
cut $S\mid S^c$&$\norm{\mathcal M_S(\rho_{\square})}_*$&$\norm{P_{S^c}\mathcal M_S(\rho_{\square})P_S}_*$& marginal on $S$\\
Here $P_S$ and $P_{S^c}$ denote the projections onto the full source and full target sectors, so the middle column isolates the genuine two-body-to-two-body block within the same normalized witness. Thus both cuts violate the separable bound, but by different mechanisms: for the adjacent cut the full-sector block already violates, while for the diagonal cut that block only saturates the bound and the excess comes from lower source or target sectors. The script \texttt{scripts/grraph\_state\_cuts.py} reproduces these values directly from the Pauli correlation tensor.
As a first systematic extension, we also scanned the families $\GHZ_n$, $W_n$, the line graph state, and the ring graph state for $n=3,4,5$, together with D\"ur states for $n=4,5$. Numerically,
throughout that range, with common values $\sqrt6$, $6/\sqrt7$, and $2.19089\ldots$ for $n=3,4,5$, respectively. The $W_n$ family is consistently slightly lower but still well above the fully separable bound, while the D\"ur family already lies below $1$ for $n=4,5$. So the symmetric shadow functional strongly favors graph-like and GHZ-like global correlation structure, but it is not simply a monotone of party number.
At $n=4$, for instance, $W_4$ still crosses the fully separable white-noise threshold at about $p\approx0.469$, whereas the D\"ur value is already below the fully separable benchmark even before white noise is added.
To probe the unresolved four-qubit biseparable benchmark, we performed a small random search over pure biseparable states across all inequivalent cuts, using $120$ samples per cut. The largest sampled value was
\begin{equation*}
\Phi_{\mathrm{sym}}\approx 2.235
\end{equation*}
for a state separable across a $2\mid2$ partition, while the best sampled $1\mid3$ values were only around $1.94$. This is not a proof of the true biseparable threshold, but it suggests two useful heuristics: first, the most dangerous competitors to the graph-state value $6/\sqrt7\approx2.268$ come from $2\mid2$ cuts rather than $1\mid3$ cuts; second, the connected four-qubit graph-state value sits slightly above the best random biseparable samples we found.
\section{Symmetry-adapted block decomposition of the shadow map}
\label{sec:symmetry-blocks}
The constructions of Sections~\ref{sec:response-maps}--\ref{sec:multiparty-sources} treat the sector grading $V\subseteq S$, $T\subseteq S^c$ as the only available organizing structure on either side of a cut. When $\rho$ carries an additional symmetry --- exactly, or after an appropriate twirl --- this grading can be refined further, along representation-theoretic rather than combinatorial lines. The refinement is worth having for four distinct reasons, which we state before developing the formal statements, since they motivate different parts of what follows and are not equally strong.
\begin{itemize}
\item\emph{Explanatory power.} Several numerical facts already reported in this note --- the triple degeneracy of the singular values of $\mathcal M_a$ for $\GHZ_3$ (Section~\ref{sec:qubit-numerics}), the triple degeneracy at the Smolin state under both the $1\mid3$ and $2\mid2$ cuts (Section~\ref{sec:multiparty-sources}), and the rigid common value across all $38$ four-qubit graph states --- have so far been recorded as numerical observations. The block-diagonality and stabilizer-support results below show that such degeneracies are not coincidental: they are forced exactly, by two complementary mechanisms depending on whether $\rho$ is symmetry-invariant in a representation-theoretic sense or Pauli-diagonal in a stabilizer sense. This converts a family of separately verified numerical facts into structural statements.
\item\emph{Diagnostic resolution.} The sub-block witnesses of Corollary~\ref{cor:sub-block} resolve a violation only down to the level of \emph{which parties} are involved ($V,T$). A representation-theoretic decomposition, where applicable, resolves it further, down to \emph{which symmetry channel} within a fixed $(V,T)$ sector is responsible --- for instance, whether a two-party source correlation block is carrying its signal in a totally symmetric or in an antisymmetric combination of its constituents. This is invisible to the party-indexed grading alone.
\item\emph{Computational cost.} For a source or target cluster respecting a symmetry group $G$, an isotypic decomposition replaces one singular value decomposition on the full sector space by several independent, much smaller singular value decompositions on the multiplicity spaces $M_\lambda$, whose dimensions grow far more slowly than the ambient sector dimension as the cluster size increases. This is the natural computationally tractable foothold for the symmetric sub-family of the higher-order tensor construction discussed in the closing remark of Section~\ref{sec:tensor-viewpoint}.
\item\emph{A one-sided extension via twirling.} If $\rho$ itself lacks the relevant symmetry but a twirl $T_G(\rho)$ is cheap to evaluate, entanglement detected on $T_G(\rho)$ certifies entanglement of $\rho$ (twirling by local unitaries preserves separability), so a block-diagonal criterion can serve as an inexpensive pre-test, with the standard one-sided caveat that a negative result on $T_G(\rho)$ is uninformative about $\rho$.
\end{itemize}
What this refinement does \emph{not} deliver is a numerically sharper detection threshold: Remark~\ref{rem:no-universal-sharpening} below shows that no universal, state-independent improvement over Theorems~\ref{thm:cut-bound} and \ref{thm:cluster-cut} exists at the level of an individual representation-theoretic block, except in the single-isotype case where the improvement is one of concentration rather than of threshold value.
\subsection*{Setup}
Let $G$ be a compact group acting on the system by local unitaries, $g\mapsto\bigotimes_{a\in P}U_g^{(a)}$, and suppose $\rho$ is $G$-invariant: $(\bigotimes_a U_g^{(a)})\,\rho\,(\bigotimes_a U_g^{(a)})^\dagger=\rho$ for all $g\in G$. Fix a cut $S\mid S^c$ preserved by $G$ as a set partition, so that $G$ acts on $\V_0^{(S)}$ and on $\V_0^{(S^c)}$ separately, via the adjoint representations $\mathrm{Ad}^{(S)}_g$, $\mathrm{Ad}^{(S^c)}_g$. Decompose both traceless spaces into isotypic components,
where $\lambda$ ranges over the irreducible representations of $G$ appearing on either side, $V_\lambda$ denotes a fixed model of the irreducible representation of dimension $d_\lambda$, and $M_\lambda^{(S)}$, $M_\lambda^{(S^c)}$ are the corresponding multiplicity spaces. This is exactly the same type of decomposition already used for the source and target sector gradings in Sections~\ref{sec:response-maps}--\ref{sec:multiparty-sources}, now taken with respect to a representation-theoretic rather than a combinatorial grading; the two coincide only in special cases.
The general strategy --- twirl a state over a symmetry group and use Schur's lemma to collapse the resulting computation onto the much smaller multiplicity spaces --- is the same one used classically to simplify the computation of entanglement measures for symmetric states \cite{vollbrechtwerner2001}; the present section applies it to the shadow map itself rather than to a scalar entanglement measure.
\begin{proposition}[Exact block-diagonality]
\label{prop:block-diagonal}
Under the hypotheses above, $\widetilde{\mathcal M}_S(\rho)$ is block diagonal with respect to the isotypic decomposition~\eqref{eq:isotypic-decomp}: writing $\Pi_\lambda^{(S^c)}$, $\Pi_\mu^{(S)}$ for the isotypic projections,
for a unique linear map $A_\lambda:M_\lambda^{(S)}\to M_\lambda^{(S^c)}$, the reduced shadow map at $\lambda$.
\end{proposition}
\begin{proof}
$G$-invariance of $\rho$ gives $\mathrm{Ad}^{(S^c)}_g\,\widetilde{\mathcal M}_S(\rho)=\widetilde{\mathcal M}_S(\rho)\,\mathrm{Ad}^{(S)}_g$ for all $g$, i.e.\ $\widetilde{\mathcal M}_S(\rho)$ is a $G$-equivariant map between the two representations~\eqref{eq:isotypic-decomp}. Both claims are then Schur's lemma applied to the isotypic decomposition: equivariant maps vanish between inequivalent irreducible summands, and act as a fixed scalar multiple of the identity on the irreducible factor $V_\lambda$ within a matching pair, leaving exactly the freedom recorded in $A_\lambda$ on the multiplicity spaces.
\end{proof}
Proposition~\ref{prop:block-diagonal} organizes $\widetilde{\mathcal M}_S(\rho)$ by symmetry channel alone, treating $\V_0^{(S)}$ and $\V_0^{(S^c)}$ as undifferentiated representations of $G$. But these spaces already carry the combinatorial sector grading of Sections~\ref{sec:response-maps}--\ref{sec:multiparty-sources}, $\V_0^{(S)}=\bigoplus_{\emptyset\neq V\subseteq S}\V_V^{(S)}$. The next result shows that this grading and the isotypic one are never in tension: for any symmetry acting locally by conjugation, one refines the other, and both can be imposed simultaneously without contradiction.
\begin{proposition}[Compatibility of sector and symmetry decompositions]
\label{prop:sector-symmetry-compatible}
Let $G$ act on the system by local unitaries, $g\mapsto\bigotimes_{a\in P}U_g^{(a)}$, and let $S\subseteq P$. Writing $R(g):=\bigotimes_{a\in S}\mathrm{Ad}_{U_g^{(a)}}$ for the induced action on $\V^{(S)}$, we have
\begin{equation}
R(g)\,\V_V^{(S)} = \V_V^{(S)}
\qquad\text{for every } g\in G \text{ and every nonempty } V\subseteq S.
\label{eq:sector-invariance}
\end{equation}
Consequently the sector projectors $P_V$ of Eq.~\eqref{eq:source-bloch-decomposition} commute with $R(g)$ for every $g\in G$, hence with every isotypic projector $P_\lambda=d_\lambda\int_G\chi_\lambda(g)^*R(g)\,dg$ of Eq.~\eqref{eq:isotypic-decomp}, and $\V^{(S)}$ decomposes simultaneously as
Since $\mathrm{Ad}_{U_g^{(a)}}$ fixes the identity, $\mathrm{Ad}_{U_g^{(a)}}\id^{(a)}=\id^{(a)}$, and maps the traceless Hermitian subspace $\V_0^{(a)}$ to itself (conjugation by a unitary preserves both Hermiticity and tracelessness), each local factor space splits $G$-invariantly as $\mathbb C\,\id^{(a)}\oplus\V_0^{(a)}$. The sector space $\V_V^{(S)}$ is, by definition, exactly the span of product basis elements with the $a$-th factor in $\V_0^{(a)}$ for $a\in V$ and equal to $\id^{(a)}$ for $a\notin V$. Since $R(g)$ acts factor-wise and each local factor preserves its own identity-versus-traceless split, $R(g)$ cannot move an element of $\V_V^{(S)}$ out of $\V_V^{(S)}$: it maps trivial legs to trivial legs and active legs to (possibly rotated, but still traceless) active legs, without ever changing which legs are trivial. This gives $R(g)\V_V^{(S)}\subseteq\V_V^{(S)}$; the same argument applied to $g^{-1}$ gives the reverse inclusion, hence Eq.~\eqref{eq:sector-invariance}. Commutation of $P_V$ with $R(g)$, and therefore with the integral defining $P_\lambda$, follows immediately.
\end{proof}
\begin{corollary}[Joint refinement of the bigraduated shadow map]
\label{cor:joint-refinement}
Let $S\mid S^c$ be a cut preserved by $G$ as a set partition, with $G$ acting locally by conjugation on both $S$ and $S^c$ as above. Then, whenever $\rho$ is $G$-invariant, every bigraduated block $M_{V\to T}(\rho)$ of Definition~\ref{def:bigraduated} further decomposes as
for reduced maps $A_{V,T,\lambda}$ on the corresponding multiplicity spaces. In particular, the combinatorial sector grading of Section~\ref{sec:multiparty-sources} and the representation-theoretic isotypic grading of the present section are not competing organizations of $\mathcal M_S(\rho)$, but two orthogonal refinements of the same operator, and can be applied jointly: one may first restrict to a sub-block witness $M_{V\to T}$ as in Corollary~\ref{cor:sub-block} and then further resolve it by symmetry channel, or apply the two refinements in the opposite order, with the same result.
\end{corollary}
\begin{proof}
By Proposition~\ref{prop:sector-symmetry-compatible} applied on the source side to $\V_V^{(S)}$ and, symmetrically, on the target side to $\V_T^{(S^c)}$, the restrictions $R_S(g)\big|_{\V_V^{(S)}}$ and $R_{S^c}(g)\big|_{\V_T^{(S^c)}}$ are themselves well-defined representations of $G$. Since $\rho$ is $G$-invariant, $\widetilde{\mathcal M}_S(\rho)$ intertwines $R_S(g)$ and $R_{S^c}(g)$ on the full spaces (Proposition~\ref{prop:block-diagonal}), and since the sector inclusion $\iota_V$ and sector projection $P_T$ used in Definition~\ref{def:bigraduated} are themselves $G$-equivariant by Proposition~\ref{prop:sector-symmetry-compatible}, the composite $M_{V\to T}(\rho)=P_T\,\widetilde{\mathcal M}_S(\rho)\,\iota_V$ intertwines the restricted representations on $\V_V^{(S)}$ and $\V_T^{(S^c)}$. Schur's lemma applied to this restricted intertwiner gives Eq.~\eqref{eq:joint-block}.
\end{proof}
\subsection*{Exact formula and its consequence}
\begin{proposition}[Reduced matrix element formula]
\label{prop:reduced-formula}
Let $\rho=T_G(\rho_S\otimes\sigma_{S^c})$ be the $G$-twirl of a product state across $S\mid S^c$, so $\rho$ is separable (a mixture, over $g\in G$, of product states) and $G$-invariant. Let $r^{(S)}\in\V_0^{(S)}$, $v_{S^c}\in\V_0^{(S^c)}$ be the traceless Bloch vectors of $\rho_S$, $\sigma_{S^c}$, with isotypic components $r_\lambda$, $v_\lambda$. Then
where $\tilde r_\lambda$, $\tilde v_\lambda$ are $r_\lambda$, $v_\lambda$ reshaped as $(\dim M_\lambda^{(S)})\times d_\lambda$ and $(\dim M_\lambda^{(S^c)})\times d_\lambda$ matrices in a basis of $V_\lambda$ shared by both sides. Consequently,
For fixed $g$, the product term $(\mathrm{Ad}^{(S)}_g r^{(S)})(\mathrm{Ad}^{(S^c)}_g v_{S^c})^T$ (unnormalized) is rank one, exactly as in the proof of Theorem~\ref{thm:cluster-cut}. Averaging over $g$ and expanding both factors in the isotypic bases gives, by the Schur orthogonality relation $\int_G D^\lambda(g)_{cb}D^\lambda(g)_{da}\,dg=\tfrac1{d_\lambda}\delta_{cd}\delta_{ab}$ for the (real, orthogonal) irreducible matrix elements $D^\lambda$, exactly Eq.~\eqref{eq:reduced-element} on each isotypic block, with all cross-$\lambda$ contributions vanishing by the same orthogonality relation applied to inequivalent irreducibles. Equation~\eqref{eq:per-block-bound} then follows from the standard nuclear-norm bound $\norm{AB}_*\le\norm{A}_\fro\norm{B}_\fro$ applied to Eq.~\eqref{eq:reduced-element}, using $\norm{\tilde r_\lambda}_\fro=\norm{r_\lambda}$, $\norm{\tilde v_\lambda}_\fro=\norm{v_\lambda}$.
\end{proof}
\begin{corollary}[Consistency with the cut-separable bound]
\label{cor:consistency}
Under the hypotheses of Proposition~\ref{prop:reduced-formula},
\begin{equation}
\sum_\lambda\dim(V_\lambda)\,\norm{A_\lambda}_*
\;\le\;
\norm{r^{(S)}}\,\norm{v_{S^c}}
\;\le\;
\sqrt{(d_S-1)(d_{S^c}-1)},
\label{eq:cs-recovery}
\end{equation}
the first inequality by Cauchy--Schwarz over $\lambda$ applied to Eq.~\eqref{eq:per-block-bound}, and the second by the correlation-sum identity used in Theorem~\ref{thm:cluster-cut}. By linearity and convexity of the nuclear norm, Eq.~\eqref{eq:cs-recovery} extends to arbitrary $G$-invariant separable states (finite or continuous mixtures of twirled product terms), recovering the bound of Theorem~\ref{thm:cluster-cut} through the block decomposition rather than around it.
\end{corollary}
\begin{remark}[No universal per-block sharpening, except by concentration]
\label{rem:no-universal-sharpening}
The first inequality in Eq.~\eqref{eq:cs-recovery} is generically strict: equality in Cauchy--Schwarz requires $\norm{r_\lambda}\propto\norm{v_\lambda}$ across all $\lambda$, which independently chosen $\rho_S,\sigma_{S^c}$ have no reason to satisfy. A direct numerical check on a twirled random product state gives $\norm{\mathcal M_S(\rho)}_*=0.429$ against the bound $\norm{r^{(S)}}\norm{v_{S^c}}=0.750$ from Eq.~\eqref{eq:cs-recovery} --- a strict, and generic, gap. Since Eq.~\eqref{eq:per-block-bound} bounds $\norm{r_\lambda}$ only by the global $\norm{r^{(S)}}^2\le d_S-1$, with no constraint on how the purity budget distributes across $\lambda$ for a generic $\rho_S$, no universal constant improving on $\sqrt{(d_S-1)(d_{S^c}-1)}$ holds for an individual block $\lambda$ in general.
The exception is the case where $\V_0^{(S)}$ itself carries only a single isotypic component under $G$: then $r_\lambda=r^{(S)}$ trivially, the full purity budget sits in the one available block, and Eq.~\eqref{eq:per-block-bound} becomes
numerically identical to Theorem~\ref{thm:cut-bound}, but now a statement about a matrix of size $\dim M_\lambda^{(S^c)}\times1$ rather than the full target space. This single-isotype case is a genuine sharpening of concentration, not of threshold, and applies whenever $\rho$ is actually $G$-invariant for a group under which the source side is irreducible. It does \emph{not}, however, apply to the two states used to motivate this section: $\GHZ_3$ and the Smolin state carry no continuous collective symmetry (a direct check shows both fail to be invariant already under a one-parameter collective rotation), so the representation-theoretic mechanism above is not the explanation for their observed degeneracies. The correct explanation for those two states, and more generally for any Pauli-diagonal state, is combinatorial rather than representation-theoretic, and is developed next.
\end{remark}
\subsection*{Axial $U(1)$ symmetry: an exact worked example}
\label{sec:u1-example}
The isotypic mechanism of Proposition~\ref{prop:block-diagonal} is easiest to see concretely for an abelian symmetry group, where it reduces to an ordinary charge-conservation selection rule.
\begin{remark}[Abelian symmetries as selection rules]
\label{rem:abelian-selection-rule}
Let $G=U(1)$ act by collective $z$-rotation, $g=\theta\mapsto\bigotimes_{a\in P}R_z^{(a)}(\theta)$ with $R_z(\theta)=e^{-i\theta Z/2}$. Every irreducible representation of $U(1)$ is one-dimensional, so the isotypic decomposition of Proposition~\ref{prop:block-diagonal} coincides exactly with the eigenspace decomposition of the conserved total charge $S_z^{\mathrm{tot}}$. Writing $E_\pm=(X\pm iY)/\sqrt2$ for the weight-$(\pm1)$ combinations and $E_0=Z$ for the weight-$0$ direction, so that $\mathrm{Ad}_{R_z(\theta)}E_\pm=e^{\pm i\theta}E_\pm$ and $\mathrm{Ad}_{R_z(\theta)}E_0=E_0$, $G$-invariance of $\rho$ forces
where $q_S,q_T$ are the total weights of $\sigma_{i_S}$, $\sigma_{i_T}$ in the $\{E_+,E_0,E_-\}$ basis. Hence a source block of total charge $q_S$ couples \emph{only} to a target block of charge $q_T=-q_S$, not $q_T=q_S$: this is the familiar rule that a correlation function survives only between charge-conjugate directions (e.g.\ $\langle S^+S^-\rangle$ need not vanish, $\langle S^+S^+\rangle$ must). For a non-abelian group such as $SU(2)$ this simple picture does \emph{not} extend by using a Cartan generator (e.g.\ $J_z$) in place of the Casimir: diagonalizing $J_z$ only refines each isotypic block further, by magnetic quantum number, and neither reproduces the isotypic block structure of Proposition~\ref{prop:block-diagonal} nor explains why the reduced block $A_\lambda$ of Proposition~\ref{prop:reduced-formula} is independent of that quantum number --- that collapse is a genuine consequence of full non-abelian equivariance, not of commuting with a single conserved charge. The abelian case is special precisely because weight and isotype coincide there.
\end{remark}
To see this mechanism at work on a state that is not simultaneously covered by the stabilizer mechanism of Lemma~\ref{lem:stabilizer-degeneracy}, consider four qubits $A,B,C,D$ and
with $S=\{A,B\}$, $S^c=\{C,D\}$. Both computational-basis terms have Hamming weight $2$, so $\rho_\varphi$ commutes with the collective $U(1)_z$ generated by $R_z(\theta)^{\otimes4}$ for every $\varphi$, but for generic $\varphi$ it is invariant under no larger continuous collective symmetry. A direct computation gives
\begin{align*}
\tr(\rho_\varphi\,X^{\otimes4})&=\sin(2\varphi),
&
\tr(\rho_\varphi\,X_AX_BY_CY_D)&=-\sin(2\varphi),
\\
\tr(\rho_\varphi\,X_AY_BY_CX_D)&=\sin(2\varphi),
&
\tr(\rho_\varphi\,Z^{\otimes4})&=1,
\end{align*}
with all remaining nonzero-weight correlators vanishing identically; the first three are $\pm1$ only at the isolated points $\varphi=\pi/4\ (\mathrm{mod}\ \pi/2)$, so $\rho_\varphi$ is generically not Pauli-flat and Lemma~\ref{lem:stabilizer-degeneracy} does not apply to it.
For each two-qubit cluster $S$ and $S^c$, the local weight decomposition $\{-1,0,+1\}\otimes\{-1,0,+1\}$ gives total-charge sectors $q=-2,-1,0,1,2$ of dimension $1,2,3,2,1$ respectively (the coefficients of $(x^{-1}+1+x)^2$), on both the source and the target side. By Remark~\ref{rem:abelian-selection-rule} the normalized bigraduated map $\mathcal M_{AB}(\rho_\varphi)$ of Eq.~\eqref{eq:bigraduated-map} block-diagonalizes according to $q_S=-q_T$, and an exact computation of every block gives
The $q=0$ value $\tfrac13$ is $\varphi$-independent, because $ZZZZ$ is diagonal in the computational basis and both terms of $\lvert\psi_\varphi\rangle$ give it the same eigenvalue $+1$; the entire $\varphi$-dependence, and with it the entanglement signal, sits in the $q=\pm2$ sectors. Equation~\eqref{eq:u1-example-formula} exceeds the cut-separable bound of Theorem~\ref{thm:cluster-cut} exactly when $\sin(2\varphi)>1/2$, i.e.\ for $\varphi\in(\pi/12,\,5\pi/12)$; it vanishes at $\varphi=0$, where $\rho_\varphi$ is itself a product state across $S\mid S^c$, and reaches $\norm{\mathcal M_{AB}(\rho_\varphi)}_*=5/3$ at the maximally superposed point $\varphi=\pi/4$.
This example illustrates Corollary~\ref{cor:joint-refinement} concretely: the full $9\times9$ block factors, without any state-specific input, into one $3\times3$ block of rank $1$, two mutually zero $2\times2$ blocks, and two $1\times1$ blocks, purely by charge conservation --- before any computation is needed to know which entries can possibly be nonzero. Since $\rho_\varphi$ is generically not a stabilizer state, this degeneracy pattern cannot be attributed to Lemma~\ref{lem:stabilizer-degeneracy}, giving a concrete instance of the separation of mechanisms described in general terms by Remark~\ref{rem:two-mechanisms}.
\subsection*{Exact degeneracy from stabilizer structure}
The degeneracies recorded for $\GHZ_3$, the Smolin state, and all $38$ four-qubit graph states share no continuous symmetry of the kind used above. Their common origin is instead a discrete, combinatorial fact about \emph{Pauli-diagonal} states, requiring only elementary group theory over $\mathbb F_2$, and it is this mechanism --- not Proposition~\ref{prop:block-diagonal} --- that is responsible for every degeneracy reported in Sections~\ref{sec:multiparty-sources} and~\ref{sec:qubit-numerics}. That reduced density matrices of stabilizer states are maximally mixed on their support, with the entanglement across any cut given by the rank of the corresponding restriction of the stabilizer group, is a classical fact \cite{fattal2004stabilizer}; Lemma~\ref{lem:code-support} below is the flat-support statement underlying that result, and Lemma~\ref{lem:stabilizer-degeneracy} translates it directly into the singular-value structure of the shadow map itself, rather than into an entanglement entropy.
\paragraph{Setup.} Identify each single-party Pauli index with $\mathbb F_2^2$ via $I\mapsto(0,0)$, $X\mapsto(1,0)$, $Y\mapsto(1,1)$, $Z\mapsto(0,1)$, so that an $n$-party Pauli string $\sigma_{\vec i}$ corresponds to $\vec i\in\mathbb F_2^{2n}$, and string multiplication (up to phase) becomes addition. Let $H\le\mathbb F_2^{2n}$ be an isotropic subgroup (i.e.\ its elements pairwise commute as operators) not containing $-I$, and let
the maximally mixed state on the joint $+1$-eigenspace of $H$ (a stabilizer code state; $\rho_H$ is pure iff $|H|=2^n$).
\begin{lemma}[Support of a stabilizer-code state]
\label{lem:code-support}
$\operatorname{Tr}[\rho_H\,\sigma_{\vec i}]=\mathbb1[\vec i\in H]$ for every $\vec i\in\mathbb F_2^{2n}$.
\end{lemma}
\begin{proof}
For $\vec i\in H$: $\Pi_H\sigma_{\vec i}=\Pi_H$ since $h\Pi_H=\Pi_H$ for every $h\in H$, so $\operatorname{Tr}[\Pi_H\sigma_{\vec i}]=\operatorname{Tr}[\Pi_H]=\operatorname{rank}\Pi_H$, giving $\operatorname{Tr}[\rho_H\sigma_{\vec i}]=1$. For $\vec i\notin H$: either $\sigma_{\vec i}$ anticommutes with some $h\in H$, whence $\operatorname{Tr}[\Pi_H\sigma_{\vec i}]=\operatorname{Tr}[h\Pi_H\sigma_{\vec i}]=-\operatorname{Tr}[\Pi_H\sigma_{\vec i}h]=-\operatorname{Tr}[\Pi_H\sigma_{\vec i}]$ (using $h\Pi_H=\Pi_H$ and cyclicity), forcing it to vanish; or $\sigma_{\vec i}$ commutes with all of $H$ without belonging to it (a logical operator), in which case it acts as a nonzero-weight, traceless operator on the logical subspace on which $\rho_H$ restricts to a multiple of the identity, again giving zero.
\end{proof}
\begin{lemma}[Forced degeneracy of $\widetilde{\mathcal M}_S(\rho_H)$]
\label{lem:stabilizer-degeneracy}
Let $S\mid S^c$ be a cut and let $\varphi:H\to\mathbb F_2^{2|S|}$, $\psi:H\to\mathbb F_2^{2|S^c|}$ be the two restriction homomorphisms, so that $H\hookrightarrow\mathbb F_2^{2|S|}\times\mathbb F_2^{2|S^c|}$ via $h\mapsto(\varphi(h),\psi(h))$. If $\psi$ is injective, then every nonzero row of $\widetilde{\mathcal M}_S(\rho_H)$, expressed directly in the (non-orthonormal) raw Pauli-string basis $\{\sigma_{\vec i}\}$ on both sides --- the basis and normalization in which the correlation tensor $C_S(\rho)$ of Eq.~\eqref{eq:corr-tensor-def} and the combined shadow map $\mathcal M_S(\rho)$ of Definition~\ref{def:bigraduated} are actually computed --- has exactly $|\ker\varphi|$ nonzero entries, all of magnitude $1$, with pairwise disjoint column supports across distinct rows; consequently
\[
\widetilde{\mathcal M}_S(\rho_H) \text{ has exactly } |\operatorname{im}\varphi|-1 \text{ equal nonzero singular values, each } =\sqrt{|\ker\varphi|}.
\]
After the normalization of Eq.~\eqref{eq:bigraduated-map}, the singular values of $\mathcal M_S(\rho_H)$ are therefore
\[
\sqrt{\frac{|\ker\varphi|}{(d_S-1)(d_{S^c}-1)}},
\]
with multiplicity $|\operatorname{im}\varphi|-1$.
\end{lemma}
\begin{proof}
By Lemma~\ref{lem:code-support}, the entry of $\widetilde{\mathcal M}_S(\rho_H)$ at row $\vec j\in\mathbb F_2^{2|S|}\setminus\{0\}$, column $\vec k$, is $1$ if $(\vec j,\vec k)\in H$ and $0$ otherwise. For fixed $\vec j\in\operatorname{im}\varphi$, the set $\{\vec k:(\vec j,\vec k)\in H\}$ is a coset of $\ker\varphi$ under the group structure of $H$ (standard fiber property of a homomorphism), hence has size $|\ker\varphi|$, giving the row weight and (since all entries are $\pm1$ in magnitude by Lemma~\ref{lem:code-support}) equal row norm $\sqrt{|\ker\varphi|}$ for every nonzero row. If rows for $\vec j\neq\vec j'$ shared a nonzero column $\vec k$, then $(\vec j,\vec k),(\vec j',\vec k)\in H$ would give $(\vec j-\vec j',0)\in H$ with $\vec j\neq\vec j'$, i.e.\ a nontrivial element of $\ker\psi$ --- excluded by injectivity of $\psi$. Rows are thus pairwise orthogonal with equal norm, hence (after normalizing) already the right singular vectors, and the singular values are all equal to the common row norm. The final rescaling is exactly the normalization already applied in Eq.~\eqref{eq:bigraduated-map}.
\end{proof}
\begin{corollary}
\label{cor:stabilizer-examples}
The Lemma applies uniformly to pure stabilizer states ($|H|=2^n$, including all graph states and $\GHZ_n$) and to uniform mixtures over a stabilizer code space with $|H|<2^n$ (including the Smolin state, $H=\{IIII,XXXX,YYYY,ZZZZ\}$). Injectivity of $\psi$ holds automatically whenever $H$ contains no element supported entirely on $S$ --- checkable by inspection of the generators, without any Lie-group input.
For $\GHZ_3$ with $S=\{1\}$: $H=\{III,XXX,ZZI,ZIZ,IZZ,YYX,YXY,XYY\}$ (up to signs), $|\ker\varphi|=2$, $|\operatorname{im}\varphi|=4$, $(d_S-1)(d_{S^c}-1)=1\times3=3$, giving $3$ singular values equal to $\sqrt{2/3}$ --- exactly the value reported in Section~\ref{sec:qubit-numerics}.
For the Smolin state, every $1\mid3$ and $2\mid2$ cut has $\varphi$ bijective ($|\ker\varphi|=1$), so $|\operatorname{im}\varphi|-1=3$ in both cases --- bijectivity of $\varphi$ holds for \emph{any} nonempty proper subset $S$ of the four legs given this particular $H$. This gives $3$ equal singular values throughout, with value
\[
\sqrt{\frac{1}{(d_S-1)(d_{S^c}-1)}}
=
\begin{cases}
1/\sqrt7 &\text{for the } 1\mid3 \text{ cut } ((d_S-1)(d_{S^c}-1)=1\times7),\\[2pt]
1/3 &\text{for the } 2\mid2 \text{ cut } ((d_S-1)(d_{S^c}-1)=3\times3),
\end{cases}
\]
matching Section~\ref{sec:multiparty-sources} and Section~\ref{sec:qubit-numerics} exactly.
\end{corollary}
\begin{remark}
The $y$-parity grading occasionally useful for real-in-the-computational-basis states is the special case $H=\{I^{\otimes n}\}$ acting trivially --- more precisely, it is not itself an instance of this Lemma but a compatible, coarser $\mathbb Z_2$-grading that commutes with any $H$-decomposition and can be applied on top of it without modification.
\end{remark}
\begin{remark}[Two complementary mechanisms]
\label{rem:two-mechanisms}
Proposition~\ref{prop:block-diagonal} and Lemma~\ref{lem:stabilizer-degeneracy} are complementary, not competing, and apply under disjoint hypotheses. The representation-theoretic mechanism applies whenever $\rho$ is genuinely $G$-invariant under some compact group $G$ acting by local unitaries and preserving the cut, regardless of whether $\rho$ is Pauli-diagonal; it says nothing about states, such as a generic finite-group-symmetric state built from a permutation representation, that are not Pauli-diagonal. The stabilizer mechanism applies whenever $\rho$ is (a uniform mixture over) a stabilizer code state, regardless of whether it possesses any continuous symmetry at all --- as is the case for $\GHZ_n$, the Smolin state, and every graph state used elsewhere in this note, none of which is invariant under a nontrivial continuous collective symmetry. In the (comparatively narrow) overlap where a state is both $G$-invariant for some continuous $G$ and Pauli-diagonal, both mechanisms apply and constrain the same block structure from different directions; outside that overlap, exactly one of the two is available, and it is this Lemma, not Proposition~\ref{prop:block-diagonal}, that accounts for every numerically observed degeneracy reported so far in this note.
The combined shadow map should be viewed as a structured refinement of the older correlation strengths $L_S$ introduced by Aschauer \emph{et al.}\cite{aschauer}. Those quantities keep one Frobenius norm per tensor block; the present construction keeps the common one-vs-rest channel structure across all orthogonal sectors on the complement. In that sense it preserves the geometric spirit of that local-invariant sector decomposition while extracting more information from the same correlation data.
There is a further degree of freedom this note has only partly explored: what to do with parties that belong to neither the chosen source cluster $S$ nor its target complement. Writing $P=S\sqcup R\sqcup E$ for a residual cluster $E$, there are three qualitatively different constructions. One can fold $E$ into the target side, $R':=R\cup E$, which is exactly the construction used throughout this note. One can discard $E$ by tracing it out and apply $\mathcal M_S$ to $\rho_{SR}=\tr_E(\rho)$, which Corollary~\ref{cor:trace-is-slice} shows is a rescaled restriction of the same object and certifies a strictly stronger, loss-of-$E$-robust form of $S\mid R$ entanglement. Or one can keep $E$ as a third graded tensor leg alongside $S$ and $R$, which leads out of matrix territory into the genuine higher-order tensor discussed in the remark closing Section~\ref{sec:tensor-viewpoint}, with the associated computability cost. The present note develops the first option in full and the second as an immediate corollary of the tensor viewpoint; a systematic treatment of families of source and target clusters with a nontrivial residual, including the genuinely tensorial third option, is left for future work.
There is also a natural exact continuation of this program. In the qubit case, the full Pauli correlation coefficients determine the density matrix linearly, so the separability problem can be formulated as a truncated moment problem on a product of Bloch spheres, in line with recent moment-based tensor criteria \cite{huang2024moments}. In higher local dimensions the same philosophy should run through generalized Bloch coordinates and products of local state spaces, with dimension-dependent response spaces and cut bounds. From that viewpoint the present shadow criteria are inexpensive front-end tests: they keep enough geometry to matter, but remain explicit and analytically tractable. The full moment hierarchy belongs to a larger project, so in the present note we retain it only as an outlook.