Requirement-Grounded LLM Agents for Functional Testing of React Applications: A Controlled Study Protocol, Benchmark Specification, and Artifact Audit

This work presents a controlled study protocol, benchmark specification, and artifact audit for investigating requirement-grounded large language model (LLM) agents for functional testing of React applications. The study design focuses on requirement-to-test traceability, LLM-assisted functional test generation, browser-based testing, Playwright-oriented test execution, mutation testing, test-oracle considerations, defect detection, and reproducibility. The manuscript defines the proposed methodology, benchmark structure, evaluation design, and artifact requirements for a controlled empirical study. This upload represents the protocol and benchmark specification stage and does not report completed experimental results.

Authors

Institutions

Publication Details

Journal
Zenodo (CERN European Organization for Nuclear Research)
Published
2026-09-29
DOI
https://doi.org/10.5281/zenodo.23033997
Primary Topic
Software Testing and Debugging Techniques
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Requirement-Grounded LLM Agents for Functional Testing of React Applications: A Controlled Study Protocol, Benchmark Specification, and Artifact Audit

Abdurrahman Khan
Zenodo (CERN European Organization for Nuclear Research)
Software Testing and Debugging Techniques
article

Requirement-Grounded LLM Agents for Functional Testing of React Applications: A Controlled Study Protocol, Benchmark Specification, and Artifact Audit

Abdurrahman Khan
article en

Abstract

This work presents a controlled study protocol, benchmark specification, and artifact audit for investigating requirement-grounded large language model (LLM) agents for functional testing of React applications. The study design focuses on requirement-to-test traceability, LLM-assisted functional test generation, browser-based testing, Playwright-oriented test execution, mutation testing, test-oracle considerations, defect detection, and reproducibility. The manuscript defines the proposed methodology, benchmark structure, evaluation design, and artifact requirements for a controlled empirical study. This upload represents the protocol and benchmark specification stage and does not report completed experimental results.

Zenodo (CERN European Organization for Nuclear Research)
Barkatullah University (IN)
Peace, Justice and strong institutions
Openalex Percentile: Top 6%
Software Testing and Debugging Techniques
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.