Ship GenAI with Confidence: Evaluation Pipelines w...
# announcements
e
Ship GenAI with Confidence: Evaluation Pipelines with Kedro and Langfuse In the next edition of the Kedro Coffee Chat 🔶, we’ll explore how to bring systematic evaluation into production GenAI workflows using Kedro and Langfuse. We’ll introduce
LangfuseEvaluationDataset
— a recently released dataset that makes evaluation a first-class part of your Kedro project. We’ll demo how to define test cases in configuration, run experiments across different prompt versions and models, and compare results side-by-side in the Langfuse UI. 📅 When: 10 Apr, 1 PM (GMT+1) youtube YouTube:

https://www.youtube.com/watch?v=7uZ-eiw43i0

linkedin LinkedIn: https://www.linkedin.com/events/7447960424517582848?viewAsMember=true 📣 These sessions are public and open to everyone. Join us live or catch the recording afterwards.
💯 2
l
v keen to see what the Kedro team came up with, we've been building our own system as well to handle LangFuse + Kedro
would be nice to exchange some learnings
e
Hey, sounds nice - very keen to see what you’ve come up with too! 🙂
We actually shipped quite a few integrations with Langfuse and Opik already for tracing and prompt management: https://docs.kedro.org/projects/kedro-datasets/en/kedro-datasets-9.3.0/api/kedro_datasets_experimental/
Let us know what you think if you give it a try 🙂
We’re starting Kedro Coffee Chat 🔶 soon!