Creative
Using large multimodal models to assess creativity and generate game levels in Physics Playground—toward scalable, valid measures of creative production.

Project Description
Overview
We train and prompt LMMs to rate novelty, usefulness, and surprise of human-authored levels, compare to expert rubrics, and study bias and reliability.
Methods
- Multimodal prompting pipelines (image + text)
- Rater-agreement calibration vs. human experts
- Auto-generation with constraint checks for playability
Project Details
- Lead: Dr. Seyedahmad Rahimi
- Sponsor: UF / Internal seed
- Funding: —
- Duration: 2024
