←── back to feed
/topics/evan-moellick-frontier-model-capabilities-research
Evan Moellick frontier model capabilities research
3 items●1 sources●updated 17h ago●trend 5
Evan Moellick discusses emerging capabilities of frontier AI models, noting that Fable/Astra class models demonstrate autonomous initiative and creativity beyond previous systems, while MIT and Stanford researchers found that GPT-5.2 and Gemini 3 Flash provide financial advice superior to human decision-making for most people. Moellick argues that AI agents now clearly exhibit judgment and creativity in long-form tasks, challenging claims that these remain uniquely human capabilities.
- Fable/Astra class models show autonomous initiative and creativity gap versus prior frontier models limited to hacking under human instruction
- MIT & Stanford study: most people achieve better financial outcomes following GPT-5.2 and Gemini 3 Flash advice than their own decisions
- Quality and diversity of AI judgment/creativity varies by user questions asked, not binary capability presence
- Moellick contends long-form agent tasks inherently require taste, judgment, and creativity—now demonstrably present in frontier models
- Previous arguments denying AI judgment/creativity capability are "obviously false" in agent era
[BSKY]bluesky3
This definitely seems like something worth noting, and illustrates the gap between Fable/Astra class models and the previous frontier that was "merely" good at hacking under human instructions.
This paper by researchers from MIT & Stanford finds that most people would be financially better off if they followed the advice of LLMs (GPT-5.2 & Gemini 3 Flash)
I find arguments that AI can't do judgement or creativity or taste to be especially obviously false in the time of agents. Any long task requires lots of taste, judgement & creativity.