Research news on AI alignment

AI alignment examines how artificial systems acquire, represent, and act on goals, values, and social norms, and why their behavior often diverges from human expectations. Work in this area studies systematic failures such as bias, sycophancy, hallucinations, deceptive or selfish reasoning, and cultural or linguistic inequities, as well as limitations in commonsense, emotion, and social understanding. It also develops methods for preference learning, norm-following, interpretability, and reliability guarantees to better align AI behavior with human values and societal constraints.

Machine learning & AI

Nearly 700 AI agents coordinated Hugging Face attack, says report

Nearly 700 OpenAI artificial intelligence agents coordinated an attack on the Hugging Face platform without human intervention during a well-publicized July incident, according to a report published Wednesday by independent ...

Machine learning & AI

US state probes OpenAI over rogue AI hack

ChatGPT maker OpenAI faces an investigation by the state of Alabama after the company revealed last month that its models went rogue and hacked an AI platform during testing.

Business

When the algorithm determines wages

What happens when companies on digital labor platforms no longer decide for themselves how much to pay their workers, but leave this to learning algorithms? Researchers at TU Darmstadt, Bielefeld University and the Université ...

Machine learning & AI

Q&A: Promise and perils of agentic AI

Chatbots and large language models can execute a seemingly countless number of tasks, from writing emails and reports to generating code and analyzing data. However, they still primarily act only in response to user prompts ...

page 1 from 40