Categories: FAANG

PROOF-Gen: From Optimized Data to Better Distillation

Supervised fine-tuning on teacher-generated trajectories is the standard first stage for distilling tool-calling capabilities into deployable models. Post-training pipelines that drive shipped tool-calling agents re-run this stage on a daily or weekly cadence, paying the frontier-teacher cost each cycle, yet the mechanism is generate-and-filter (keep the teacher’s passing trajectories, discard the rest) and each cycle leaves behind the same hard scenarios because failures supply no signal. On τ 2-bench, 57% of teacher trials fail, two-thirds of them near-misses (most tool calls correct, undone…
AI Generated Robotic Content

Recent Posts

From 3D layout to compositing: new LTX VFX tools

During VFX Week last week, we released seven open-weight capabilities for LTX-2.5, covering high-res editing,…

7 hours ago

How (and Why) to Build an AI Agent from Scratch in Python

In this article, you will learn what an AI agent is and how to build…

7 hours ago

Negotiating Ontological Boundaries in User-Authored Personal Sensing Systems

Designed artifacts are ontological, shaping, and at times limiting, what becomes possible or imaginable. One…

7 hours ago

Introducing GLM 5.3 on Amazon Bedrock

Coding and agentic workloads are asking more of AI models than ever: refactor a repository…

7 hours ago

Elon Musk’s Exes Are the Most Damning Part of a Massive New Documentary

Interviews with Justine Wilson and Ashley St. Clair in the nearly four-hour-long Musk paint the…

8 hours ago

Browsing this sub in the past week

No hate. Just for fun. submitted by /u/the_bollo [link] [comments]

1 day ago