Variational quantum algorithms (VQAs) have established themselves as a central computational paradigm in the Noisy Intermediate-Scale Quantum (NISQ) era. By coupling parameterized quantum circuits (PQCs) with classical optimization, they operate effectively under strict hardware limitations. However, as quantum architectures transition toward early fault-tolerant (EFT) and ultimate fault-tolerant (FT) regimes, the foundational principles and long-term viability of VQAs require systematic reassessment. This review offers an insightful analysis of VQAs and their progression toward the fault-tolerant regime. We deconstruct the core algorithmic framework by examining ansatz design and classical optimization strategies, including cost function formulation, gradient computation, and optimizer selection. Concurrently, we evaluate critical training bottlenecks, notably barren plateaus (BPs), alongside established mitigation strategies. The discussion then explores the EFT phase, detailing how the integration of quantum error mitigation and partial error correction can sustain algorithmic performance. Addressing the FT phase, we analyze the inherent challenges confronting current hybrid VQA models. Furthermore, we synthesize recent VQA applications across diverse domains, including many-body physics, quantum chemistry, machine learning, and mathematical optimization. Ultimately, this review outlines a theoretical roadmap for adapting quantum algorithms to future hardware generations, elucidating how variational principles can be systematically refined to maintain their relevance and efficiency within an error-corrected computational environment.
While quantum computers promise to solve classically intractable problems, identifying the point at which fault-tolerant quantum computation outperforms the best classical algorithms for practical applications remains an outstanding challenge. Here we establish a concrete quantum-classical crossover for quantum many-body dynamics under realistic hardware conditions. We introduce a scalable fault-tolerant framework that combines coherent observable estimation with a space-time-efficient implementation of non-Clifford rotations, suppressing the residual logical errors that limit existing partially fault-tolerant approaches. A benchmark against state-of-the-art tensor-network and variational Monte Carlo algorithms reveals a concrete crossover for mixed-field Ising dynamics at modest system sizes. For a physical error rate of p=10^-3, fault-tolerant simulation requires approximately 2 hours and 3.7 × 10^5 physical qubits for a 100-site 1D system, whereas tensor network approaches would require about 100 years. For 2D models, where rapid entanglement growth limits the classical evolution time, we project quantum runtimes within minutes. A physical error rate of p=10^-4 leads to at least an order of magnitude reduction in qubit count (3.1 × 10^4 physical qubits) and runtime (minutes for 1D and seconds for 2D). The reduction in quantum runtime arises from our improved rotation-state injection and co-design of quantum error correction and observable-estimation protocols, which jointly suppress logical-error accumulation and reduce sampling overhead. Our results establish a scalable route towards practical quantum advantage and identify quantitative engineering targets for future fault-tolerant architectures.
We study the simultaneous estimation of partial-transpose moments p_j(ρ_AB)=Tr[(ρ_AB^T_B)^j], j=2,…,K, of an unknown bipartite n-qubit state from independent copies under an explicit active-memory constraint. We give a sequential qubit-reuse realization of the partial-transpose permutation that uses at most 2n+1 active qubits, independent of K, and estimates all moments p_2,…,p_K to uniform additive error ε with total copy complexity O(Klog K/ε^2). We also prove two converse bounds. First, any uniformly accurate simultaneous estimator requires Ω(K/ε^2) copies in the worst case. Second, the same scaling holds on an explicit isospectral two-qubit negative-partial-transpose (NPT) family whose ordinary moments are constant while the partial-transpose moments vary. These results characterize the copy complexity of the partial-transpose moment hierarchy up to a logarithmic factor and extend simultaneous nonlinear-functional estimation from ordinary state powers to partial-transpose spectral data under active quantum memory independent of the target moment order.
Fault tolerance is an indispensable prerequisite for constructing large-scale universal quantum computers. Drawing philosophies from classical computer architecture, this paper presents a hardware-agnostic three-layer high-level architectural framework for generic fault-tolerant quantum computation. Guided by the real execution workflows of fault-tolerant quantum algorithms, the proposed model is decoupled from specific physical qubit hardware platforms and quantum error correction codes, serving as a universal abstract standard rather than a platform-specific implementation scheme. Special attention is devoted to the intermediate Fault-Tolerance Layer, which serves as the architectural bridge between application-level logical programs and hardware-level execution. We systematically characterize its five internal components, the interfaces and data exchanged among them, and the execution, correction, and adaptation paths that together enable logical synthesis, fault-tolerant resources management, decoding, and runtime fault-tolerant control. An end-to-end example is further provided to illustrate the full-stack operating pipeline of fault-tolerant quantum algorithms under this framework. Given the increasing emphasis on modular, heterogeneous, and cross-layer fault-tolerant quantum systems, our architecture provides a unified foundational model for organizing such designs.
A fundamental task in quantum information science is to measure nonlinear functionals of quantum states, such as Tr(ρ^{k}O). Intuitively, one expects that computing a kth order quantity generally requires O(k) copies of the state ρ, and we rigorously establish this lower bound under sample access to ρ. Surprisingly, this limitation can be overcome when one has purified access via a unitary that prepares a purification of ρ, a scenario naturally arising in quantum simulation and computation. In this setting, we find a different lower bound of Θ(sqrt[k]) and present a quantum algorithm that achieves this bound, demonstrating a quadratic advantage over sample-based methods. The key technical innovation lies in a designed quantum algorithm and optimal polynomial approximation theory-specifically, Chebyshev polynomial approximations tailored to the boundary behavior of power functions. Our results unveil a fundamental distinction between sample and purified access to quantum states, with broad implications for estimating quantum entropies and quantum Fisher information, realizing quantum virtual distillation and cooling and evaluating other multiple nonlinear quantum observables with classical shadows.
Constrained combinatorial optimization with strict linear constraints underpins applications in drug discovery, power grids, logistics, and finance, yet remains computationally demanding for classical algorithms, especially at large scales. The quantum approximate optimization algorithm (QAOA) offers a promising quantum framework, but conventional penalty-based formulations distort optimization landscapes and demand deep circuits, undermining scalability on near-term hardware. In this work, we introduce Hamming weight operators, a new class of constraint-aware operators that confine quantum evolution strictly within the feasible subspace. Building on this idea, we develop an adaptive Hamming weight operator QAOA, which dynamically selects the most effective operators to construct shallow, problem-tailored circuits. We validate our approach on benchmark tasks from both finance and high-energy physics, specifically portfolio optimization and two-jet clustering with energy balance. Across these problems, our method inherently satisfies all constraints by construction, converges faster, and achieves higher approximation ratios than penalty-based QAOA, while requiring roughly half as many gates. By embedding constraint-aware operators into an adaptive variational framework, our approach establishes a scalable and hardware-efficient pathway for solving practical constrained optimization problems on near-term quantum devices.
Multipartite Bell tests provide a correlation-only route to benchmarking quantum processors, but their application at large scales is hindered by the rapid decay of many-body correlators under noise and exponentially many terms in conventional Bell expressions. Here we address these scalability obstacles by introducing a finite-setting generalized Mermin family of state-tailored Bell inequalities with analytic certification bounds, in which the measurement-setting number $m$ provides an additional certification dimension complementary to the system size $n$. We show that, for the powers-of-two setting choices considered here, increasing $m$ leaves the ideal normalized multipartite quantum value unchanged while lowering the relevant classical bounds, thereby strengthening the Bell-violation ratios and yielding an improved noise-robustness scaling compared to the standard Mermin inequality. We test this construction experimentally on a programmable superconducting processor by preparing Greenberger-Horne-Zeilinger (GHZ) states of up to 80 qubits. Using randomized sampling for direct Bell-operator estimation, we observe Bell ratios that grow exponentially with system size, certify a nonlocality depth of 14, and show that increasing $m$ strengthens both the Bell ratio and depth certification. All results are obtained solely from measured correlators and analytical bounds, without readout correction, tomography, or model-based mitigation. Generalized Mermin inequalities therefore provide a sharper Bell benchmark for noisy large-scale GHZ states.
Simulating large quantum circuits on hardware with limited qubit counts is often attempted through methods like circuit knitting, which typically incur sample costs that grow exponentially with the number of connections cut. In this work, we introduce a framework based on Cluster-level Light-cone analysis that leverages the natural locality of quantum workloads. We propose two complementary algorithms: the Causal Decoupling Algorithm, which exploits geometric disconnections in the light cone for sampling efficiency, and the Algebraic Decomposition Algorithm, which utilizes algebraic expansion to minimize hardware requirements. These methods allow simulation costs to depend on circuit depth and connectivity rather than system size. Together, our results generalize Lieb-Robinson-inspired locality to modular architectures and establish a quantitative framework for probing local physics on near-term quantum devices by decoupling the simulation cost from the global system size.
Logical T state preparation is a major overhead source in fault tolerant architectures built from stabilizer operations. Existing protocols, however, are reported under different code families, noise models, postselection rules, and cost conventions, making direct comparison difficult. We compare three representative preparation routes: magic state distillation, magic state cultivation, and code switching, using currently available results. Rather than reducing heterogeneous data to a single cost metric, we retain source native cost units and record output error, single attempt cost, expected cost per accepted output, footprint, latency, and reporting completeness for each configuration. Within the current dataset, distillation reaches the lowest output error regime; code switching achieves the lowest reported single attempt cost and the smallest explicit footprint among the compatible rows; and recent RP2 cultivation results add low cost cultivation points with output errors between 1e-6 and 1e-9. As a simple algorithm level case study, we also examine the reported preparation routes under an error budget motivated by Shor factoring algorithm, in order to relate single state preparation costs to full workload requirements. The resulting comparison clarifies the trade offs currently supported across the literature, while remaining bounded by the conventions and coverage of the underlying papers.
The universal scaling of critical behavior in phase transitions is a cornerstone of physics. Dynamical quantum phase transitions (DQPTs) are their nonequilibrium analogues: abrupt nonanalyticities that emerge as a quantum system evolves in time. Yet the hardness and cost of detecting this phenomenon remain largely unexplored. We prove that estimating DQPT to a certain precision is intractable even for quantum computers, whereas deciding a subsystem variant of DQPT is as hard as simulating generic quantum circuits, implying a provable exponential quantum advantage. Furthermore, to search for critical times of local DQPTs, we show a quadratically faster quantum algorithm that estimates observables of Hamiltonian dynamics at multiple time points with Heisenberg-limited precision and sublinear scaling in the number of time points. Moreover, through encoding classical evolution into quantum dynamics, our framework enables broader quantum speedups for detecting anomalous phenomena in classical systems.
In distributed fault-tolerant quantum computing, entanglement and magic are essential resources for quantum communication and universal fault-tolerant computation, respectively. Although they are usually treated as distinct resource currencies, whether they admit a unified resource-theoretic description remains an open question. Here, we introduce the distributed resource theory of entanglement and magic (DREAM). In this framework, the free states are convex mixtures of product local stabilizer states, and the free operations are local stabilizer circuits assisted by classical communication (LSCC). We show that DREAM contains nontrivial resource states that are neither entanglement nor magic, so it is strictly richer than treating the two resources independently. Surprisingly, such resources can enable the teleportation of magic states between distant parties without consuming entanglement, revealing a counterintuitive form of resource teleportation mediated entirely by separable states. We further generalize this result to quantum networks and investigate general quantum-state teleportation under LSCC. We show that the shared resource under DREAM is closely related to the teleportation capability, quantified by the magic of the teleported state. Our work establish a systematic framework for studying distributed quantum resources and uncover intrinsic relations among distinct resources within a unified resource theory.
Modern quantum physics now enables control of quantum systems at the level of individual trajectories, opening a new frontier that links quantum information theory, quantum many-body physics, and quantum thermodynamics, and uncovers novel non-equilibrium phenomena such as deep thermalization and measurement-induced entanglement. However, a central challenge remains: their characterization relies on measuring nonlinear properties of individual quantum states, a task tantamount to fine-grained cloning of a quantum ensemble. Here, the fundamental laws governing the cloning of quantum ensembles are investigated. First, a general no-cloning theorem for arbitrary ensembles is established from an information-theoretic perspective, even assuming multiple copies of the ensemble's purification. It is then shown that this barrier can be unexpectedly circumvented for physical ensembles generated by finite-time evolutions. Nevertheless, these tasks are proven to remain computationally intractable, even when the full circuit description of state preparation is known. This stands in sharp contrast to the conventional no-cloning theorem, which relies on the state being unknown. Together, these results establish new fundamental principles of quantum mechanics, reveal intrinsic trade-offs among sample complexity, computational complexity, and quantum measurements, and highlight the necessity of problem-specific strategies for probing measurement-induced quantum phenomena.
Adiabatic evolution is a central paradigm in quantum physics. Digital simulations of adiabatic processes are generally regarded as resource-intensive, not only because of the long evolution time required, but also because algorithmic errors typically accumulate throughout the evolution, thereby demanding exceptionally deep circuits to preserve accuracy. This work demonstrates that digital adiabatic evolution is intrinsically accurate and robust to simulation errors. We analyze two Hamiltonian simulation methods—Trotterization and generalized quantum signal processing—and prove that the simulation error does not increase with time. We further show that accurate time-dependent adiabatic evolution can be achieved using only time-independent Hamiltonian-simulation algorithms. Numerical simulations of the adiabatic algorithms for molecular systems and linear equations confirm the theory, revealing that digital adiabatic evolution is substantially more efficient than previously assumed. Remarkably, our estimation for the first-order Trotterization error can be 106 times tighter than previous analyses for the transverse field Ising model even with less than 6 qubits. The findings establish fundamental robustness of digital adiabatic evolution and provide a basis for accurate, efficient implementations on fault-tolerant-and potentially near-term–quantum platforms. In this study, the authors show that digital adiabatic evolution is highly robust to simulation errors. Rather than accumulating, these errors self-cancel and even decrease over time, enabling highly efficient algorithms on near-term and future quantum computers.
Abstract Spectroscopy underpins modern scientific discovery across diverse disciplines. While experimental spectroscopy probes material properties through scattering or radiation measurements, computational spectroscopy combines theoretical models with experimental data to predict spectral properties, essential for advancements in physics, chemistry, and materials science. However, quantum systems present unique challenges for computational spectroscopy due to their inherent complexity, and current quantum algorithms remain largely limited to static and closed quantum systems. Here, we present and demonstrate a generalised quantum computational spectroscopy that lifts these limitations by reconstructing the quantum autocorrelation function via an ancilla-assisted Hadamard test quantum circuit. Our method is applicable to a broad range of quantum systems, including closed, open, and time-dependent driven quantum systems. We experimentally validate this approach, which leverages arbitrary controlled quantum dynamics and efficient classical noise-mitigation strategy, on a programmable silicon-photonic quantum processing chip, capable of high-fidelity time-evolution simulations. The versatility of our method is demonstrated through spectroscopic computations for diverse quantum systems, revealing novel phenomena such as parity-time symmetry breaking and topological holonomy that are inaccessible to conventional spectroscopy or quantum eigenstate algorithms. This work establishes a noise-robust methodology for quantum spectral analysis.
Learning the generator of an open many-body system is more challenging than Hamiltonian learning: local responses, which can directly reveal coherent interaction terms in closed-system dynamics, may also contain dissipative contributions in open-system dynamics. In this paper, we address this challenge by developing an efficient Lindbladian learning framework for a known local candidate generator dictionary with bounded dissipative support and either bounded dual-interaction-graph degree or bounded unweighted local strength. The framework resolves the coherent-dissipative ambiguity by treating local Pauli responses as a linear system over both types of generator terms. Inverting this response system separates their contributions and makes the individual Lindbladian coefficients accessible from local response data in a fixed short-time window. Within this framework, we develop two efficient learning algorithms: Chebyshev–Lobatto response interpolation, which uses logarithmically many short evolution times and has a post-mean cost linear in M, with the stated dependence on ε, and Single-time projected response contraction, which uses a single fixed evolution time and globally inverts a truncated response function. Both procedures estimate M candidate coefficients to entrywise accuracy ε using 𝒪(M/ε^2) sample and classical post-processing complexity. Our theoretical results establish local response inversion as a scalable paradigm for learning, calibrating, and diagnosing complex quantum systems from experimentally accessible short-time data.
Quantum imaginary-time evolution (QITE) is a fundamental framework for preparing ground and thermal states, yet its computational cost scales significantly with the evolution duration τ. Reducing this duration is critical for practical quantum advantage. Here, we establish a unified theoretical framework for the Mpemba effect in QITE – a counterintuitive phenomenon where a state initially farther from the ground state relaxes to it faster than one initially closer. We derive a remarkably simple necessary and sufficient condition for the occurrence of this effect, showing it is uniquely determined by the population ratios of excited states to the ground state. For practical state preparation, we introduce a rigorous sufficient condition for the finite-time Mpemba effect, ensuring the crossing occurs before reaching a prescribed proximity threshold. Furthermore, we unveil unique dynamical features, including a multiple-crossing phenomenon in multi-level systems and simultaneous intersections for collinear initial states. Our results provide criteria for identifying favorable initial states in QITE and offer deep insights into the speed limit of quantum state preparation.
Quantum simulation is widely regarded as one of the most promising applications of quantum computing. A critical challenge in this domain is understanding and quantifying the accumulation of algorithmic errors over time, which is essential for designing more efficient simulation algorithms and for assessing the resources required to achieve quantum advantage. Conventional error analyses typically rely on the triangle inequality to bound the total simulation error, but such approaches tend to overestimate errors by ignoring error interference-a phenomenon in which errors from different simulation segments partially cancel. Here, we introduce a new framework for directly estimating long-time algorithmic errors in segmented quantum simulations. Our approach captures the full structure of error interference, enabling significantly tighter and more accurate error bounds. We identify both necessary and sufficient conditions for strict error interference and propose the notion of approximate error interference to account for realistic, imperfect cancellation. We demonstrate the broad applicability of our framework across a range of models and settings, including Heisenberg and Fermi-Hubbard systems, lattice Hamiltonians with power-law interactions, higher-order Trotter decompositions, and adiabatic evolution. By providing a unified and practical methodology for analyzing error interference, our Letter advances the theoretical understanding of quantum simulation and informs the design and benchmarking of algorithms for near-term and future quantum hardware.
Nonlinear spectroscopy is a cornerstone of quantum science, providing unique access to multi-point correlations, quantum coherence, and couplings that are invisible to linear methods. However, classical simulation of these phenomena is fundamentally limited by the exponential growth of the Hilbert space, and practical quantum algorithms for the nonlinear regime have remained largely unexplored. Here, we present a unified quantum algorithmic framework for computing n-th order nonlinear spectroscopies. By reformulating multi-time responses as a weighted sum of expectation values at finite pump amplitudes via a generalized parameter shift rule, our approach bypasses the costly evaluation of high-order commutators and time-dependent operator expansions. This reformulation enables efficient execution via real-time evolution on current quantum hardware, ensuring inherent noise resilience. We validate the framework on IBM's superconducting quantum processors, successfully obtain higher-order response functions of a 12-qubit XXZ spin-chain. Furthermore, the versatility of our method is demonstrated by resolving quasi-particle excitation spectra in spin-liquids and identifying interaction-induced cross-peaks in atomic systems. Our results establish a practical and scalable pathway for probing complex quantum dynamics on near-term quantum devices, extending the reach of quantum simulation into the nonlinear domain.
Neural quantum states offer expressive representations of quantum many-body wave functions, yet their practical accuracy can be limited by stochastic optimization rather than representational capacity. Here we identify a finite-sample instability, termed subspace trapping, in which physically important configurations become strongly underestimated, remain absent from successive sampling batches and receive insufficient gradient feedback. This self-reinforcing loss of sampled support can confine optimization to an effective subspace and produce apparently stationary states above the true ground state energy. To address this problem, we introduce annealed gradient descent (AGD), a sampling-aware update with annealing factor that temporarily increases the relative contribution of sampled low-probability configurations while limiting the dominance of high-probability ones. We establish the connection between finite-sample support loss and effective subspace optimization, and then evaluate the method across molecular systems, one and two-dimensional J_1-J_2 models. Annealed gradient descent suppresses metastable trapping, preserves physically relevant configurations and enables compact neural quantum states to attain chemical accuracy and competitive state-of-the-art performance. These results establish AGD as a lightweight complement to expressive neural architectures, improved sampling strategies for scalable quantum many-body optimization.
Observing the physical world is a foundational pursuit in science. In the quantum realm, however, observation necessitates a fundamental quantum-to-classical conversion: destructive measurements irreversibly project quantum states into classical data, inevitably incurring a loss of information. What physical principles govern this information loss, and how can we construct optimal measurements to maximize the readout? Here, we address these questions by establishing an intrinsic relationship between readout capability–quantified by the ratio of accessible classical Fisher information to the total quantum Fisher information (QFI), and measurement complexity–defined as the quantum circuit depth required prior to projection. Remarkably, we uncover a sudden emergence of observability: a sharp hidden-to-visible transition driven entirely by measurement complexity. We rigorously prove that below critical depth thresholds–Θ((log n)^1/δ) for δ-dimensional architectures and Θ(loglog n) for all-to-all connectivity–readout capability decays exponentially with system size n, rendering the quantum information fundamentally inaccessible. Surprisingly, immediately above this threshold, the system enters a visible regime: we demonstrate that randomized measurements universally recover a constant fraction of the QFI using approximate unitary 3-designs, for which we explicitly develop optimal-depth circuit constructions tailored to finite-dimensional architectures. By unveiling the fundamental scaling laws and transitions that govern quantum observation, our results delineate definitive resource boundaries for quantum learning, state certification, and quantum metrology.