That is remarkably text-like without actually being text. Makes sense that reading would produce more distinct activity.
I would have expected more details on the faces, considering how much processing power is assigned to the task.
What's below the elephant is ground, what's above it is sky. Right? It's like the fMRI went "Hey, brain, what color is sky?" "Blue." "And what color is ground?" "Uh, green?"
http://sites.google.com/site/gallantlabucb/publications/nishimoto-et-al-2011