arXiv AI By Mohammad Arif Rasyidi, Syahirul Faiz

Benchmarking API Drift in LLM-Generated Quantum Code Across Successive SDK Versions

Read the original on arXiv AI →

arXiv:2607. 04072v1 Announce Type: cross Abstract: Large language models can generate plausible quantum code, but it is unclear whether they can reliably target the specific software development kit (SDK) version requested by the user.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Sep 7

Impact of Data Loss in Postprocessing on Training and Inference of Quantum Neural Networks

The paper investigates how postprocessing routines in quantum neural network software can cause significant data loss when run on large quantum hardware. In a case study of Qiskit’s “SamplerQNN”, a filter that assumes measurement bit‑strings are in virtual qubit space removed 85–99.6% of valid shots on IBM backends, leading to unnormalised probability vectors and distorted predictions. This loss caused inference accuracy to drop from 0.94 to 0.39 and compressed training loss signals by 22–27×, severely reducing optimizer sensitivity. The authors implemented a layout‑based marginalisation fix that was merged into the library to make “SamplerQNN” forward‑compatible with current and future hardware.

By Soraya V. Panambalom, Edoardo Altamura, Nick Chancellor, Jonte R. Hance
Hugging Face Trending Papers
Sep 4

Impact of Data Loss in Postprocessing on Training and Inference of Quantum Neural Networks

The paper investigates how postprocessing routines in quantum neural network software can inadvertently discard a large portion of valid measurement data when run on real quantum hardware. In a case study of Qiskit’s “SamplerQNN”, a filter that assumes virtual qubit space caused 85–99.6% of measurement shots to be lost on IBM backends, leading to unnormalised probability vectors, degraded inference accuracy (from 0.94 to 0.39), and a 22–27× compression of the training loss signal. The authors provide a layout‑based marginalisation fix that has been merged into the library to ensure forward‑compatibility with current and future hardware.

arXiv AI
Jul 24

PennySynth: RAG-Driven Data Synthesis for Automated Quantum Code Generation

arXiv:2605. 25572v2 Announce Type: replace-cross Abstract: The growing complexity of quantum programming frameworks has exposed a critical limitation in existing large language model (LLM)-based code assistants: general-purpose models hallucinate PennyLane-specific gate names, misplace device configurations, and produce structurally invalid circuits when faced with specialized quantum coding challenges.

By Minghao Shao, Nouhaila Innan, Hariharan Janardhanan, Muhammad Kashif, Alberto Marchisio, Muhammad Shafique
Hugging Face Trending Papers
5d ago

QC-Stark: A Multi-Task Benchmark Revealing Capability Dissociations in LLMs Evaluated on Quantum Computing Tasks

QC-Stark is a benchmark that evaluates large language models on 11 quantum computing tasks, including circuit construction, debugging, compilation, error correction, and simulation. It comprises 2,750 evaluations across 10 models, 5 difficulty levels, and 5 seeds, revealing that overall model rankings can hide significant per-task differences, with Spearman correlation being insignificant for 4 of the 11 tasks. The benchmark’s measurement quality is validated by a 2‑parameter Item Response Theory model, prompt sensitivity analysis confirms ranking robustness, and all tasks are auto‑verifiable via execution, eliminating the need for manual evaluation. The code and data are publicly available on Hugging Face.