Now, you can automatically find best prompts for any agentic workflow you're putting together, means manual prompt engineering isn't needed at all.
🟢 What is the simple idea?
1️⃣ Take a starting prompt and an eval dataset
2️⃣ Then, an optimizer iteratively improves the prompt
3️⃣ In the end, you get an optimal prompt automatically
And all of this in just a few lines of code.
➡️ Why Opik specifically?
Opik is a 100% open-source platform for evaluating LLMs.
It helps optimize LLM systems so that they work better, faster, and cheaper: from RAG chatbots to code assistants. Opik includes tracing, evaluations, and dashboards.
Best Part: Everything can be run completely locally, because you can use any local LLMs as optimizers and evaluators.
GitHub repository
••••••••••••••••••••••••••••••••••••••••••••••
🤖 Data & ML | @DataXplore