Austin Xu is a Research Scientist working on LLM reasoning, automatic evaluation, and contextual understanding. He contributed to the development of SFRJudge, a family of automatic evaluators used the verify model-generated outputs. He received his PhD from the Georgia Institute of Technology.
AI is rapidly transforming industries, helping businesses enhance customer experiences, improve efficiency, and make smarter decisions. But an essential question arises: How can we ensure that AI is creating accurate and grounded answers?…
As the development and deployment of large language models (LLMs) accelerates, evaluating model outputs has become increasingly important. The established method of evaluating responses typically involves recruiting and training human evaluators, having them…