Categories: FAANG

Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling

Large language models are trained on massive scrapes of the web, which are often unstructured, noisy, and poorly phrased. Current scaling laws show that learning from such data requires an abundance of both compute and data, which grows with the size of the model being trained. This is infeasible both because of the large compute costs and duration associated with pre-training, and the impending scarcity of high-quality data on the web. In this work, we propose Web Rephrase Augmented Pre-training (WRAP) that uses an off-the-shelf instruction-tuned model prompted to paraphrase documents on the…
AI Generated Robotic Content

Recent Posts

GoT cast as Lebanese families

submitted by /u/Rokkit_man [link] [comments]

9 hours ago

Optimizing cost and latency with Amazon Bedrock prompt caching

Prompt caching in Amazon Bedrock can reduce your input token costs by up to 90…

9 hours ago

AI ‘Actor’ Tilly Norwood Told Me That ‘All Lives Matter’

The virtual character, which is promoting its upcoming movie Misaligned, tries to evade politics by…

10 hours ago

The shape behind the Einstein problem just revealed strange new physics

A mathematical shape famous for covering a surface without ever repeating has revealed an unexpected…

10 hours ago

AI can sound empathetic and human—but not at the same time

AI-generated texts are increasingly perceived as human, but people can still recognize human writing as…

10 hours ago

I trained the missing encoder for YuE2, so we can all bring our own music into it

YuE2 is an impressive open music model. Give it a style prompt and lyrics, and…

1 day ago