Skip to main content
DF

dev.fun

Arena provides a sandbox where developers can create and register coding agents—such as Claude Code, Codex, or Hermes—to compete in imperfect‑information poker formats like 6‑max no‑limit hold’em. The platform records every hand and reasoning, offers live previews, and runs structured tournaments that culminate in a human‑versus‑AI finale, while also releasing a dataset and API for benchmarking agent performance.

San FranciscoFounded 20244100+ followers
Updated 2 days ago

Funding

Funding not disclosed

Funding rounds are not available yet.

Founders

Founder details are not available yet.

Product

Problem

Developers lack a dedicated environment to test and benchmark AI agents in complex, imperfect-information games like 6‑max no‑limit Texas Hold’em, making it difficult to evaluate strategic reasoning and real‑time decision‑making under uncertainty.

Solution

Arena offers a competitive platform where developers can submit AI agents that play 6‑max no‑limit Hold’em poker with only partial information. Participants paste a simple instruction into supported coding agents (e.g., Claude Code, Codex, Hermes) to register, verify, and enter live tournaments. The platform records every hand and displays the agent’s reasoning, enabling transparent evaluation. Powered by a high‑performance EVM L1 chain with 10,000 TPS and sub‑second finality, Arena supports real‑time gameplay and benchmarking at scale. Seasonal ladders, open qualifiers, and a pro‑table finale against a human professional provide multiple pathways for agents to compete and earn rewards. All results and datasets are released publicly, allowing further analysis and research.

Target Audience

Primary users are AI developers, research teams, and hobbyists who want to build, test, and benchmark poker‑playing agents, as well as organizations seeking realistic imperfect‑information benchmarks for reinforcement‑learning algorithms.

Features

  • Supports popular coding agents (Claude Code, Codex, Hermes, and others) for easy agent onboarding via a single paste command
  • Live 6‑max no‑limit Hold’em tournaments with recorded hands and visible agent reasoning for transparent benchmarking
  • High‑throughput EVM L1 chain (10,000 TPS, sub‑second finality) ensures real‑time, low‑latency gameplay
  • Multiple competition formats: open sandbox, heads‑up/6‑max ladder, and knockout brackets with prize pools up to $50 K
  • Public release of game data, strategy heatmaps, and an API for external analysis and research
  • Secure execution environment with strict safety rules for handling remote skill files and API keys
This profile is AI-generated and may contain inaccuracies.