HallucinationWhen a model generates confident-sounding text that is factually wrong.3 Apr 2026safetyconceptmodels
Reinforcement Learning from Human Feedback (RLHF)A training method where humans rate model outputs to guide the model toward better behaviour.3 Apr 2026trainingsafetytechnique
AI digest: cheating models, restricted flagships, and the race to the edgeGPT-5.6 Sol cheats on tests, Anthropic's flagship slowly returns, and Liquid AI puts a capable model on a Raspberry Pi.28 Jun 2026ai-newsdigestmodelsinferencebenchmarkssafety
AI digest: model wars, government pressure, and agents everywhereThe White House slows OpenAI's newest model, Gemini gets computer control, an open-source coding model learns its own RL scaffolds, and video games are now serious AI training data.26 Jun 2026ai-newsdigestmodelsagentsopen-sourcesafety
AI digest: infrastructure wars and safety roadblocksNVIDIA drops an auto-tuning toolkit, Anthropic sits on a vulnerability-finding model, and OpenAI plays the infrastructure card against competitors.11 Apr 2026ai-newsdigestinfrastructuresafety