LangChain releases open-source auto-evaluator for QA chains

LangChain open-sourced an auto-evaluator tool that auto-generates QA test sets and grades answers for LLM question-answer chains, now available as a free hosted app and API. It combines Anthropic's model-written evaluation sets with OpenAI's model-graded evaluation.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- ChatGPT user reports 5-hour limit consumed by single prompt
- Google's Gemini 3.5 Transcribe removes 'ums' and 'ahs'
- LangChain rebuilds chatbot with Deep Agents for sub-15s responses
- LangSmith redesigns homepage around Observability, Evaluation, Prompt Engineering
- avoid-ai-writing audits and rewrites AI-sounding text