AI Agents for Data Analysis with Shreya Shankar - Listen - The

The TWIML AI Podcast (formerly This Week in...

AI Agents for Data Analysis with Shreya Shankar

Listen now

Description

Today, we're joined by Shreya Shankar, a PhD student at UC Berkeley to discuss DocETL, a declarative system for building and optimizing LLM-powered data processing pipelines for large-scale and complex document analysis tasks. We explore how DocETL's optimizer architecture works, the intricacies of building agentic systems for data processing, the current landscape of benchmarks for data processing tasks, how these differ from reasoning-based benchmarks, and the need for robust evaluation methods for human-in-the-loop LLM workflows. Additionally, Shreya shares real-world applications of DocETL, the importance of effective validation prompts, and building robust and fault-tolerant agentic systems. Lastly, we cover the need for benchmarks tailored to LLM-powered data processing tasks and the future directions for DocETL. The complete show notes for this episode can be found at https://twimlai.com/go/703.

More Episodes

See all »

Automated Reasoning to Prevent LLM Hallucination with Byron Cook

Today, we're joined by Byron Cook, VP and distinguished scientist in the Automated Reasoning Group at AWS to dig into the underlying technology behind the newly announced Automated Reasoning Checks feature of Amazon Bedrock Guardrails. Automated Reasoning Checks uses mathematical proofs to help...

Published 12/09/24

AI at the Edge: Qualcomm AI Research at NeurIPS 2024 with Arash Behboodi

Today, we're joined by Arash Behboodi, director of engineering at Qualcomm AI Research to discuss the papers and workshops Qualcomm will be presenting at this year’s NeurIPS conference. We dig into the challenges and opportunities presented by differentiable simulation in wireless systems, the...

Published 12/03/24

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Published 12/03/24