arXiv · 2412.17497
Advantages of density in tensor network geometries for gradient based training
Abstract
Tensor networks are a very powerful data structure tool originating from quantum system simulations. In recent years, they have seen increased use in machine learning, mostly in trainings with gradient-based techniques, due to their flexibility and performance exploiting hardware acceleration. As ans\"atze, tensor networks can be used with flexible geometries, and it is known that for highly regular ones their dimensionality has a large impact in performance and representation power. For heterogeneous structures, however, these effects are not completely characterized. In this article, we train tensor networks with different geometries to encode a random quantum state, and see that densely connected structures achieve better infidelities than more sparse structures, with higher success rates and less time. Additionally, we give some general insight on how to improve memory requirements on these sparse structures and its impact on the trainings. Finally, as we use HPC resources for the calculations, we discuss the requirements for this approach and showcase performance improvements with GPU acceleration on a last-generation supercomputer.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sergi Masot-Llima, Artur Garcia-Saez. 2024-12-23. Advantages of density in tensor network geometries for gradient based training. https://arxiv.org/abs/2412.17497
Cite the original work for its findings. Save a collection to share your selection of sources.