Birdwatch Archive

Birdwatch Note Rating

2023-07-21 19:17:24 UTC - HELPFUL

Rated by Participant: 1E68FDD5E51BE1805B4E02850CD8E34EEE5EF7F2CECDF5263C07CEB37C9A5381
Participant Details

Original Note:

GPT-4 was never really able to do the task they were testing - checking whether a number is prime. It just reversed from usually saying numbers were prime to usually saying they were non-prime. When tested on just prime numbers this made it look like it was getting worse. https://www.aisnakeoil.com/p/is-gpt-4-getting-worse-over-time

All Note Details

Original Tweet

All Information

  • noteId - 1682415491724222466
  • participantId -
  • raterParticipantId - 1E68FDD5E51BE1805B4E02850CD8E34EEE5EF7F2CECDF5263C07CEB37C9A5381
  • createdAtMillis - 1689967044295
  • version - 2
  • agree - 0
  • disagree - 0
  • helpful - 0
  • notHelpful - 0
  • helpfulnessLevel - HELPFUL
  • helpfulOther - 0
  • helpfulInformative - 0
  • helpfulClear - 1
  • helpfulEmpathetic - 0
  • helpfulGoodSources - 1
  • helpfulUniqueContext - 0
  • helpfulAddressesClaim - 1
  • helpfulImportantContext - 1
  • helpfulUnbiasedLanguage - 1
  • notHelpfulOther - 0
  • notHelpfulIncorrect - 0
  • notHelpfulSourcesMissingOrUnreliable - 0
  • notHelpfulOpinionSpeculationOrBias - 0
  • notHelpfulMissingKeyPoints - 0
  • notHelpfulOutdated - 0
  • notHelpfulHardToUnderstand - 0
  • notHelpfulArgumentativeOrBiased - 0
  • notHelpfulOffTopic - 0
  • notHelpfulSpamHarassmentOrAbuse - 0
  • notHelpfulIrrelevantSources - 0
  • notHelpfulOpinionSpeculation - 0
  • notHelpfulNoteNotNeeded - 0
  • ratingsId - 16824154917242224661E68FDD5E51BE1805B4E02850CD8E34EEE5EF7F2CECDF5263C07CEB37C9A5381