Benchmarking

New Qiskit HumanEval Release: Qiskit 1.4 Compatibility and Benchmark Improvements featured image

New Qiskit HumanEval Release: Qiskit 1.4 Compatibility and Benchmark Improvements

Released a new version of Qiskit HumanEval compatible with Qiskit 1.4, featuring significant improvements to the benchmark including more robust and rigorous code execution tests …

avatar
Juan Cruz-Benito
•
Presenting Qiskit HumanEval at IEEE Quantum Week 2024 featured image

Presenting Qiskit HumanEval at IEEE Quantum Week 2024

Presenting the Qiskit HumanEval benchmark for LLMs at IEEE Quantum Week 2024 in the SYS-BNCH Benchmarking session. Available afterwards at the IBM Quantum booth to discuss AI and …

avatar
Juan Cruz-Benito
•
Qiskit HumanEval: Evaluation Benchmark for Quantum Code Generation Published featured image

Qiskit HumanEval: Evaluation Benchmark for Quantum Code Generation Published

Published research paper introducing Qiskit HumanEval dataset for evaluating Large Language Models capability to generate quantum computing code. The dataset comprises more than …

avatar
Juan Cruz-Benito
•