Computer Sciences

New 'bandit' algorithm uses light for better bets

How does a gambler maximize winnings from a row of slot machines? This is the inspiration for the "multi-armed bandit problem," a common task in reinforcement learning in which "agents" make choices to earn rewards. Recently, ...

Computer Sciences

AI: Agents show surprising behavior in hide and seek game

Researchers have made news in letting their AI ambitions play out a formidable game of hide and seek with formidable results. The agents' environment had walls and movable boxes for a challenge where some were the hiders ...

page 1 from 4