Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games
Shengyuan Ding
ChrisDing1105
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 hour ago
JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution upvoted a paper 3 days ago
GameXpert-Bench: How Far Are Coding Agents from Expert Game Development?Organizations
None yet