sm

RESEARCH NOTES

Behind the papers.

Research note2024

Beyond Lines and Circles

Investigating the geometric reasoning gap in large language models.

Research note2023

Finding Biases in Code Generation

When generated code follows superficial cues instead of the underlying task.

Research note2022

Measuring CLEVRness

Black-box testing of visual reasoning models.

Subscribe via RSS ↗