AI models nearly erase female characters when they write kids stories about animals

Last year, Melanie Walsh, a University of Washington assistant professor in the Information School, wrote an article examining how 300 popular children’s books gendered their animal characters. Of the 13 most common animals, most were male—unless they happened to be cats, ducks or birds, which trended slightly more female. But a frog, a wolf? More …

Measuring Performance of Transformer Inference

This chapter is divided into eight parts; they are: • Metrics for LLM Inference • Measuring a Single Request • Warmup and Synchronization • Measuring GPU Work with CUDA Events • Measuring Memory Usage • Measuring Concurrent Requests • Multiple GPUs and Multiple Machines • Cost per Token The most common inference metrics are: • …

ML 21265 1

Introducing Web Search on Amazon Bedrock for foundation model grounding

When a foundation model needs to answer a question about last week’s earnings call, yesterday’s regulatory change, or this morning’s weather forecast, it needs knowledge it was never trained on. Grounding the model in current web knowledge closes that gap – whether it’s powering chatbots, coding assistants, CLI tools, or enterprise applications, grounding helps answer …

DtBank Apigee 1max 1000x1000 1

How Deutsche Bank unlocked agility with an API-ready ecosystem

When people think about digital transformation in banking, they often focus on the visible results: mobile apps and new digital services. But there’s an invisible infrastructure making all these services possible: APIs. At Deutsche Bank, we recognized that APIs aren’t just technical plumbing; they’re the nervous system of modern banking.  A few years ago, our …

People prefer stories written by AI—especially when told they’re written by a human

People gave the highest ratings to AI-generated stories they were told had been written by humans and were unable to tell the difference between human- and AI-created stories, according to new research published in Judgement and Decision Making. The study, led by researchers at Villanova University, U.S., adds to the growing body of evidence that …

Decoding Strategies and Output Control

This chapter is divided into nine parts; they are: • Reading Logits from a Model • Greedy Decoding • Temperature Sampling • Top-$k$ Sampling • Nucleus Sampling • Repetition Penalties • Beam Search • Stop Conditions • Structured Output Constraints The model returns a vector of logits for every position in the input sequence.