← All topics

multimodal ai

1 capture, most recent first.

Prinz Eugen, der ... @prinzeugen_____

Prinz Eugen, der ... ✓ @prinzeugen_____ Here's a dirty little secret. If you gave GPT-4.5 a video avatar + ability to use o3-mini + output in voice instead of text + memory of all your past conversations, 99.999% of humanity (i.e., everyone except Yann and Gary Marcus) would be convinced this is "human-level AI". A few more advances in multimodal and memory, and we are there.
Note from Claude Sonnet 5

A commentary tweet arguing that human-level AI perception is mostly a packaging/UX problem (avatar, voice, memory) rather than a raw-capability gap, name-dropping AI skeptics Yann LeCun and Gary Marcus. Relevant to Nathan's tracking of public discourse on AI capability perception and the "human-level AI" framing debate.

twitter/xai capabilitieshuman-level aigpt-4.5multimodal ai