← All companies

Olam Labs

Active

Evaluating models for behavior and performance through simulated games

S26·Summer 2026·B2B·San Francisco, CA, USA·Team of 2

About

Evaluating models for social behavior, safety, and performance through multi-agent simulated games. We use popular games like Catan, Risk, or Poker, and make humans come play them against talking AI opponents, but on the backend we're actually running a multi-agent environment used to evaluate the models for skill and behavior over large sample sizes. You can play in Social Arena and view our preliminary benchmarks today at https://olamlabs.ai/

Change history · none recorded

No changes recorded yet. Subsequent syncs will populate this timeline as fields drift.