•  1
    BackgroundStatistical review is essential for research quality and integrity, yet traditional manual review is inefficient. Large language models (LLMs) offer potential support but are unreliable when used without guidance for precise calculations and raise concerns about accountability. This study evaluated whether a structured, rule-based prompt can reliably constrain an LLM to perform statistical review of comparative categorical data, and characterized both its feasibility and its inherent r…Read more