Categories: FAANG

ProText: A Benchmark Dataset for Measuring (Mis)gendering in Long-Form Texts

We introduce ProText, a dataset for measuring gendering and misgendering in stylistically diverse long-form English texts. ProText spans three dimensions: Theme nouns (names, occupations, titles, kinship terms), Theme category (stereotypically male, stereotypically female, gender-neutral/non-gendered), and Pronoun category (masculine, feminine, gender-neutral, none). The dataset is designed to probe (mis)gendering in text transformations such as summarization and rewrites using state-of-the-art Large Language Models, extending beyond traditional pronoun resolution benchmarks and beyond the…
AI Generated Robotic Content

Recent Posts

If AI tools had existed in the past

Not just a meme... submitted by /u/takayatodoroki [link] [comments]

5 hours ago

Asus ROG Swift RGB Stripe OLED Review: Clarity King

The Asus PG27UCWM brings a new sub-pixel layout to the world of OLED gaming monitors,…

6 hours ago

Chinese humanoid robots smash human records in 100m sprint and high jump at Beijing robot games

Chinese humanoid robots broke records set by humans, including beating Usain Bolt's 100-meter sprint world…

6 hours ago

Saily Ultra eSIM Premum Plan Review: Packed With Perks

For uninterrupted service as you country-hop, the Saily Ultra eSIM works well and comes with…

1 day ago

AI agents can build consensus on a scale humans can’t

Everyone is familiar with the situation: A larger group of people plans to visit a…

1 day ago

A Tale of Two Flink Autoscalers

Samuel Yeboah, Francesco Di Chiara and Mingliang LiuToday, Netflix runs two Flink autoscalers. That is…

2 days ago