Running Reproduction: iWorld-Bench: A Benchmark for Interactive World Models with a Unified Action Generation Framework ๐ฏ Explore and sync benchmark logs with an AI coding agent
Running Reproduction: PRISM: Perception Reasoning Interleaved for Sequential Decision Making. ๐ฏ Explore and sync experiment logs with an AI agent
Running Reproduction: Telescope: Improving Zero Shot Detection of LLM Generated Content By Measuring Token Repetition Probability ๐ฏ Collaborate with an AI agent to manage and update your logbook