Page 2: Research news on Large language models

Large language models are high-capacity neural sequence models trained on massive text and multimodal corpora to perform language understanding, generation, and reasoning. Current work examines their internal representations, cognitive and social behavior analogies to humans, and limitations in mathematical, causal, and strategic reasoning. Research also addresses alignment with human values and brain activity, safety and security vulnerabilities, privacy and de-anonymization risks, cross-lingual and sociocultural biases, scaling and efficiency laws, and frameworks for tool use, multi-agent interaction, and domain-specific deployment.

Computer Sciences

How AI models decide not to answer a question

Large language models (LLMs), the artificial intelligence (AI) systems supporting the functioning of ChatGPT, Gemini and similar conversational platforms, are often expected to answer queries, generate texts for specific ...

Machine learning & AI

Who should own the knowledge that underpins AI technology?

Russian mathematician Yurii Nesterov was recently awarded the Gauss Prize for outstanding mathematical contributions and his "groundbreaking work" on optimization. The award cited his work on gradient descent methods—and ...

page 2 from 34