New top story on Hacker News: Show HN: Gentrace – connect to your LLM app code and run/eval it from a UI

Show HN: Gentrace – connect to your LLM app code and run/eval it from a UI
10 by dsaffy | 0 comments on Hacker News.
Hey HN - Doug from Gentrace here. We originally launched via Show HN in August of 2023 as evaluation and observability for generative AI: https://ift.tt/QJDLltX Since then, everyone from the model providers to LLM ops companies built a prompt playground. We had one too, until we realized this was totally the wrong approach: - It's not connected to your application code - They don't support all models - You have to rebuild evals for just this one prompt (can't use your end-to-end evals) In other words, it was a ton of work and time to use these to actually make your app better. So, we built a new experience and are relaunching around this idea: Gentrace is a collaborative LLM app testing and experimentation platform that brings together engineers, PMs, subject matter experts, and more to run and test your actual end-to-end app. To do this, use our SDK to: - connect your app to Gentrace as a live runner over websocket (local) / via webhook (staging, prod) - wrap your parameters (eg prompt, model, top-k) so they become tunable knobs in the front end - edit the parameters and then run / evaluate the actual app code with datasets and evals in Gentrace We think it's great for tuning retrieval systems, upgrading models, and iterating on prompts. It's free to trial. Would love to hear your feedback / what you think!

New top story on Hacker News: Show HN: Gentrace – connect to your LLM app code and run/eval it from a UI New top story on Hacker News: Show HN: Gentrace – connect to your LLM app code and run/eval it from a UI Reviewed by nasir khan on December 12, 2024 Rating: 5
Powered by Blogger.