This blog post introduces self-speculative decoding, a technique that uses early layers of a large language model for drafting tokens and later layers for verification, achieving faster text…
BAAI introduces FlagEval Debate, a multilingual debate platform for evaluating LLMs through direct debate competitions in English, Chinese, Arabic, and Korean, with dual evaluation metrics and…
Hugging Face and AtlaAI announce Judge Arena, a platform for benchmarking LLMs as evaluators. Users can compare models side-by-side by having them judge responses and voting on the best evaluation.
LG AI Research held its annual LG AI Insight 2024 conference, recapping the past year's accomplishments and discussing future AI trends, including AI agents, LLMs, multimodal AI, bio/medical AI, and…
LG AI Research's AI Ethics Seminar discusses key AI toolkits (Microsoft HAX and Google PAIR) for designing trustworthy AI experiences and introduces LG's own AI Ethical Impact Assessment system.