Abstract:Quantum-centric supercomputing presents a compelling framework for large-scale hybrid quantum-classical tasks. Although quantum machine learning (QML) offers theoretical benefits in various applications, challenges such as large-size data encoding in the input stage and the reliance on quantum resources in the inference stage limit its practicality for tasks like fine-tuning large language models (LLMs). Quantum parameter generation, a novel approach of QML, addresses these limitations by using quantum neural networks (QNNs) to generate classical model weights (parameters) exclusively during training, thereby decoupling inference from quantum hardware. In this work, we introduce Quantum Parameter Adaptation (QPA) in the framework of quantum parameter generation, which integrates QNNs with a classical multi-layer perceptron mapping model to generate parameters for fine-tuning methods. Using Gemma-2 and GPT-2 as case studies, QPA demonstrates significant parameter reduction for parameter-efficient fine-tuning methods, such as Low-Rank Adaptation (LoRA), while maintaining comparable or improved performance in text generation tasks. Specifically, QPA reduces the number of parameters to $52.06\%$ of the original LoRA for GPT-2 with a slight performance gain of $0.75\%$, and to $16.84\%$ for Gemma-2, with a marginal performance improvement of $0.07\%$. These results highlight QPA's ability to achieve efficient parameter reduction without sacrificing performance in the quantum parameter generation framework. This work showcases the potential of quantum-enhanced parameter reduction, offering a scalable quantum-classical solution for fine-tuning LLMs while preserving the feasibility of inference on classical hardware.

GPT on a Quantum Computer

Unleashing the Potential of LLMs for Quantum Computing: A Study in Quantum Architecture Design

Quantum linear algebra is all you need for Transformer architectures

Quixer: A Quantum Transformer Model

Adapting Pre-trained Language Models for Quantum Natural Language Processing

Transformer Models for Quantum Gate Set Tomography

Quantum Machine Learning: An Interplay Between Quantum Computing and Machine Learning

Quantum Large Language Models via Tensor Network Disentanglers

A Quantum Circuit-Based Compression Perspective for Parameter-Efficient Learning

Application of Large Language Models to Quantum State Simulation

C4Q: A Chatbot for Quantum

Implementation Guidelines and Innovations in Quantum LSTM Networks

Quantum Machine Learning: Bridging Quantum Computing and Artificial Intelligence

Recent Advances for Quantum Neural Networks in Generative Learning

Grammar-aware sentence classification on quantum computers

Leveraging Pre-Trained Neural Networks to Enhance Machine Learning with Variational Quantum Circuits

Summary of ChatGPT/GPT-4 Research and Perspective Towards the Future of Large Language Models

A Quantum Machine Learning Algorithm Based on Generative Models

Power of Quantum Generative Learning

Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models

Quantum-Train: Rethinking Hybrid Quantum-Classical Machine Learning in the Model Compression Perspective