image1 XYX8VCTmax 1000x1000 1

Building cost-effective, high-throughput gen AI workflows in Google Dataflow

Real-time streaming pipelines are the operational backbone of modern enterprises, continuously processing everything from customer support interactions to transaction logs. Traditionally, streaming DAGs are static; once deployed, their processing logic and execution paths are fixed. However, by integrating generative AI agents, we can move beyond static logic to adaptive execution. This allows streaming workflows to …

ML 21710 1

NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart

NVIDIA Nemotron 3.5 Lightning is designed for the fast, specialized model execution required by high-volume agentic workloads. With NVIDIA Nemotron 3.5 Lightning on Amazon SageMaker JumpStart, you can access an open model designed for high-volume agentic workloads. With this launch, you can deploy Nemotron 3.5 Lightning from Amazon SageMaker JumpStart without configuring the serving infrastructure …

Engineers make edge AI more efficient by redesigning both algorithm and hardware

Researchers in the Riccio College of Engineering at the University of Massachusetts Amherst have demonstrated that redesigning both hardware and algorithms can make AI applications on edge devices more efficient. As proof of concept, their system achieved 95.24% accuracy in language identification while reducing computing resources by 90%—the highest reported accuracy from a system of …