pairwiseLLM

Pairwise Comparison Tools for Large Language Model-Based Writing Evaluation

v1.1.0 · Dec 22, 2025 · MIT + file LICENSE

Description

Provides a unified framework for generating, submitting, and analyzing pairwise comparisons of writing quality using large language models (LLMs). The package supports live and/or batch evaluation workflows across multiple providers ('OpenAI', 'Anthropic', 'Google Gemini', 'Together AI', and locally-hosted 'Ollama' models), includes bias-tested prompt templates and a flexible template registry, and offers tools for constructing forward and reversed comparison sets to analyze consistency and positional bias. Results can be modeled using Bradley–Terry (1952) <doi:10.2307/2334029> or Elo rating methods to derive writing quality scores. For information on the method of pairwise comparisons, see Thurstone (1927) <doi:10.1037/h0070288> and Heldsinger & Humphry (2010) <doi:10.1007/BF03216919>. For information on Elo ratings, see Clark et al. (2018) <doi:10.1371/journal.pone.0190393>.

Downloads

550

Last 30 days

7330th

1.6K

Last 90 days

1.8K

Last year

Trend: +3.4% (30d vs prior 30d)

CRAN Check Status

14 OK

Show all 14 flavors

Flavor	Status	Time
r-devel-linux-x86_64-debian-clang	OK	132.7s
r-devel-linux-x86_64-debian-gcc	OK	89.3s
r-devel-linux-x86_64-fedora-clang	OK	200.7s
r-devel-linux-x86_64-fedora-gcc	OK	229.9s
r-devel-macos-arm64	OK	35s
r-devel-windows-x86_64	OK	163s
r-oldrel-macos-arm64	OK	43s
r-oldrel-macos-x86_64	OK	231s
r-oldrel-windows-x86_64	OK	197s
r-patched-linux-x86_64	OK	127.5s
r-release-linux-x86_64	OK	115.2s
r-release-macos-arm64	OK	42s
r-release-macos-x86_64	OK	277s
r-release-windows-x86_64	OK	154s

Check History

OK 14 OK · 0 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE Mar 30, 2026

ERROR 13 OK · 0 NOTE · 0 WARNING · 1 ERROR · 0 FAILURE Mar 27, 2026

ERROR r-devel-linux-x86_64-fedora-clang

whether package can be installed

Installation failed.
See ‘/data/gannet/ripley/R/packages/tests-clang/pairwiseLLM.Rcheck/00install.out’ for details.

OK 14 OK · 0 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE Mar 10, 2026

Dependency Network

Version History

new 1.1.0 Mar 10, 2026

Maintainer

Sterett H. Mercer

Dependencies

Depends

R (>= 4.1)

Imports

curl dplyr httr2 jsonlite rlang stats tibble tidyselect tools utils

Suggests

BradleyTerry2 EloChoice knitr mockery purrr readr rmarkdown sirt stringr testthat (>= 3.0.0) tidyr withr

Compilation

No compilation needed

First Published

Dec 22, 2025

RSS Feed

CRAN Checks

View on CRAN →