Resource

AI/Agent

  • your-voice: A skill that learns your writing style from your own samples and revises AI drafts across three passes so they read like you, not generic AI.

  • AInotator: Annotate computer-mediated discourse with LLMs to capture communicative acts, politeness strategies, and meta-acts. Supports multiple model backends and provides reproducible, theory-grounded annotations for large-scale CMC research.

  • RLAM: Make jargon-laden scientific abstracts accessible to those without a college degree.

  • NovEval: Assess scientific novelty in alignment with human evaluation.

Corpus

  • Blog-1K: A redistributable English authorship identification benchmark with roughly balanced samples per author and fixed data splits (train/val/test), allowing for fair comparison among deep learning–based models.

  • RAABT: The Reproducible Authorship Attribution Benchmark Tasks include five tasks for attributing authorship of contemporary non-fiction American English prose. Fixed training/testing splits prevent accuracy inflation from homogeneous corpora.

  • Cross-Register Authorship Attribution Corpus: Contains writings from eight authors who wrote in both vernacular and classical Chinese. With 4.2 million Chinese characters, it supports authorship identification research.

Python Package

Other