Add Plagiat AI Detector RAID submission - #179
Conversation
|
Hello, could you please approve the workflow run for this RAID leaderboard submission when convenient? Thank you in advance. |
|
It looks like this eval run failed. Please check the workflow logs to see what went wrong, then push a new commit to your PR to rerun the eval. |
|
Apologies @cristeamarinescu-cpu seems like there's an issue on our end with the evaluation bot (GitHub updated some default permission setting). Will follow up on this tonight. |
|
It looks like this eval run failed. Please check the workflow logs to see what went wrong, then push a new commit to your PR to rerun the eval. |
|
Eval run succeeded! Link to run: link Here are the results of the submission(s): Plagiat AI DetectorRelease date: 2026-07-26 I've committed detailed results of this detector's performance on the test set to this PR. Warning No aggregate score across all settings is reported here as some domains/generator models/decoding strategies/repetition penalties/adversarial attacks were not included in the submission. This submission will not appear in the main leaderboard; it will only be visible within the splits in which all samples were evaluated. Warning No aggregate score across all non-adversarial settings is reported here as some domains/generator models/decoding strategies/repetition penalties were not included in the submission. If all looks well, a maintainer will come by soon to merge this PR and your entry/entries will appear on the leaderboard. If you need to make any changes, feel free to push new commits to this PR. Thanks for submitting to RAID! |
|
@cristeamarinescu-cpu bot is fixed now, seems like there were some predictions missing from your submission. Can you double check that you have outputs for every text in the dataset? |
Replaces the incomplete 56,000-row (test_none only) predictions with complete coverage of the official test.csv, including every adversarial variant, as requested in the PR review. 672,000 records, one score per id, no nulls, scores in [0, 1]. predictions.json sha256 81381eb424a599129bfcdf34ccf9630371f8609ae47fd487123750607f596cb3 Model unchanged: GeorgeDrayson/modernbert-ai-detection-raid-mage @ e047d9c5.
|
Hello @liamdugan — you were right. Sorry for the long silence. Our first submission only covered the 56,000 rows of We have now re-scored the complete
Could you re-approve the evaluation workflow when you have a moment? Happy to fix anything else you spot. Thanks for the benchmark and for the patience. |
|
Eval run succeeded! Link to run: link Here are the results of the submission(s): Plagiat AI DetectorRelease date: 2026-07-26 I've committed detailed results of this detector's performance on the test set to this PR. On the RAID dataset as a whole (aggregated across all generation models, domains, decoding strategies, repetition penalties, and adversarial attacks), it achieved an AUROC of 97.64 and a TPR of 94.19% at FPR=5% and 88.32% at FPR=1%. If all looks well, a maintainer will come by soon to merge this PR and your entry/entries will appear on the leaderboard. If you need to make any changes, feel free to push new commits to this PR. Thanks for submitting to RAID! |
|
Thank you, @liamdugan — The bot found our new relese an the evaluation ran cleanly this time. The results are exactly what we needed. We appreciate the benchmark. Best regards! |
Adds predictions and metadata for Plagiat AI Detector.