Good Start Labs

Our Research

Peer-reviewed papers and technical write-ups from our work.

ARTICLEAugust 4, 2026

What a Railroad Game Taught a Model About Finance

We trained a 30B model to play 1830, a brutal board game about 19th century railroad barons. That model seemed to get better at real SEC-filings research — and an independent AI research agent found the same fingerprint on a test we never ran.

Read Article
What a Railroad Game Taught a Model About Finance
ARTICLEJuly 30, 2026Agent Skills ’26

Where an Agent’s Intelligence Lives

OpenAI tripled one benchmark score with two harness settings. In our game agents, the strongest results came when the model and its skill system learned together.

Read Article
Where an Agent’s Intelligence Lives
ARTICLEJuly 2026

Measure What Matters. Get Paid for It.

How we helped a game publisher launch its own benchmark, turn its data into revenue, and improve a language model with an autonomously trained expert.

Read Article
Measure What Matters. Get Paid for It.
RESEARCH PAPERFebruary 2026Accepted to ICLR Workshop

Data-Centric Interpretability for LLM-based Multi-Agent Reinforcement Learning

We introduce Meta-Autointerp, a framework that uses LLMs to automatically generate and validate interpretability hypotheses about learned features in multi-agent RL. Our results show that data-centric methods can surface meaningful behavioral patterns that traditional approaches miss.

Read Research Paper
Data-Centric Interpretability for LLM-based Multi-Agent Reinforcement Learning
ARTICLEFebruary 3, 2026

We Trained an AI on a Board Game. It Became a Better Customer Support Agent.

Games teach transferable skills, to humans and AI alike.

Read Article
We Trained an AI on a Board Game. It Became a Better Customer Support Agent.
RESEARCH PAPERAugust 2025Accepted to NeurIPS Workshop

Opening the Door: Democratizing Diplomacy

We created the first Diplomacy environment where even small models can play full games. Our goal was to give people tools to understand how different AI models make decisions.

Read Research Paper
Opening the Door: Democratizing Diplomacy
ARTICLEJune 5, 2025

We Made Top AI Models Compete in a Game of Diplomacy. Here’s Who Won.

We launched our first project AI Diplomacy. Different models revealed their true character: some betrayed without hesitation, while others, like Claude, chose principles over victory. We made AI behavior visible through Diplomacy, for 45,000 unique Twitch viewers, many experiencing AI for the first time.

Read Article
We Made Top AI Models Compete in a Game of Diplomacy. Here’s Who Won.

Interested in working together?

We partner with labs and researchers exploring what games teach AI.

Talk to us