Google/DeepMind

Latest AI news, models and releases from Google/DeepMind. ['A2UI v0.9', 'AI Co-Scientist', 'AI Evaluator', 'AI Mode', 'AI Overviews', 'AlphaEvolve', 'AlphaFold', 'AlphaGenome', 'AlphaQubit', 'Antigravity', 'Auto frame', 'BERT', 'Chinchilla', 'Computational Discovery', 'Confidential GKE Nodes', 'Co-Scientist', 'Empirical Research Assistance', 'Era', 'ERA', 'Executive LLM', 'Farmscapes 2020', 'Flood Hub', 'frontier AI', 'Frozen v2', 'FunctionGemma', 'Gemini', 'Gemini 1.5 Flash', 'Gemini-1.5-pro', 'Gemini 1.5 Pro', 'Gemini 2.0 Flash', 'Gemini 2.5', 'Gemini 2.5 Flash', 'Gemini-2.5 Flash', 'Gemini-2.5-Flash', 'gemini-2.5-flash-image', 'Gemini 2.5 Pro', 'Gemini-2.5 Pro', 'Gemini-2.5-Pro', 'Gemini 3.1 Flash', 'gemini-3.1-flash-image']

Research 🇺🇸

Google evaluates alignment of behavioral tendencies in large language models

Google Research introduces a framework that converts psychological questionnaires into situational judgment tests (SJTs) to assess behavioral alignment of LLMs. Testing 25 models reveals that smaller models often deviate from human consensus, while larger models show improved but imperfect alignment. The study also finds models are overconfident when human opinions diverge.

Google/DeepMindGoogle/DeepMind
Google Research27.07 · 16:04
Fresh news