M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text Detection
arXiv · · Significant research
Summary
MBZUAI researchers introduce M4GT-Bench, a new benchmark for evaluating machine-generated text (MGT) detection across multiple languages and domains. The benchmark includes tasks for binary MGT detection, identifying the specific model that generated the text, and detecting mixed human-machine text. Experiments with baseline models and human evaluation show that MGT detection performance is highly dependent on access to training data from the same domain and generators.
Keywords
machine-generated text · MGT detection · benchmark · LLM · MBZUAI
Get the weekly digest
Top AI stories from the GCC region, every week.