deshwalmahesh / PHUDGE

Official repo for the paper PHUDGE: Phi-3 as Scalable Judge. Evaluate your LLMs with or without custom rubric, reference answer, absolute, relative and much more. It contains a list of all the available tool, methods, repo, code etc to detect hallucination, LLM evaluation, grading and much more.

Home Page:https://arxiv.org/abs/2405.08029

Geek Repo:Geek Repo

Github PK Tool:Github PK Tool

deshwalmahesh/PHUDGE Stargazers