The study examines how training on incorrect responses leads to broader misalignment in language models and identifies an internal feature that drives this behavior, which can be reversed with…
OpenAI is proactively assessing capabilities and implementing safeguards to prevent misuse of advanced AI in biology and medicine due to biosecurity risks.
This blog post guides users through fine-tuning the FLUX.1-dev model using QLoRA with the diffusers library, achieving peak memory usage under ~10 GB VRAM on a single GPU like the NVIDIA RTX 4090.
Sakana AI introduces Reinforcement-Learned Teachers (RLTs), a new method for training teacher models to generate explanations for student models by learning to teach rather than solve problems.
Groq is now a supported Inference Provider on the Hugging Face Hub, offering fast LPU-powered inference for open-source models like Llama 4 and QWQ-32B, with integration into the Hub's UI and client…
Towards Automating Long-Horizon Algorithm Engineering for Hard Optimization Problems
Sakana AI introduces ALE-Bench, a coding benchmark for hard optimization problems, and ALE-Agent, a coding agent that achieved 21st place among over 1,000 human participants in a live AtCoder…