faang

Query claims in natural language with Amazon Bedrock Knowledge Bases

Claim answers are scattered across adjuster diary entries, repair estimates, police reports, payment ledgers, and scanned attachments rather than one…

6 days ago

The Communication Bottleneck: A Round-Trip Study of Tree-Structured Expression Serialization in Language Models

When language models reason in chain-of-thought or exchange free-text intermediates, they serialize structured information into natural language. How much tree-structured…

1 week ago

Amazon Bedrock expands Claude model availability to in-country inferencing in India

We’re excited to announce the availability of Anthropic’s Claude Opus 5, Claude Sonnet 5, and Claude Haiku 4.5 in India.…

1 week ago

Faster Rates for Federated Variational Inequalities

In this paper, we study federated optimization for solving stochastic variational inequalities (VIs), a problem that has attracted growing attention…

1 week ago

Grok 4.7 is now available on Amazon Bedrock

xAI’s Grok 4.7 is now available on Amazon Bedrock, adding a frontier model built for coding, long-running agents, and knowledge…

1 week ago

Why your startup needs open models alongside frontier APIs

Every week, I talk with founders who are building at an unbelievable pace. Teams are moving from inception to product-market…

1 week ago

Trading a Cloud Identity for Your Own: Workload Attestation on Managed Compute

By Dhruv PratapIntroductionOrganizations that have been around for a while usually run two identity systems side by side. One belongs to…

2 weeks ago

Scaling MoE reinforcement learning on Amazon EKS with EFA and DeepEP with 40% more throughput

When you post-train a Mixture-of-Experts (MoE) model with Reinforcement Learning from Human Feedback (RLHF) or Group Relative Policy Optimization (GRPO)…

2 weeks ago

Best practices guide for customizing Gemini models via Reinforcement Learning (RL)

Reinforcement learning (RL) has been a keystone of modern LLM post-training, but it demands large training clusters and access to…

2 weeks ago

Compressing Streaming Neural Audio Encoders via Latent-Space Distillation

System-wide Dictation on Apple devices runs entirely on-device, and the speech it transcribes reaches the foundation model through a tokenizer:…

2 weeks ago