Multimedia content is provided continuously for multiple users each potentially moving to a different place while enjoying it together. Mesh data and texture co-ordinate data associating vertices of a mesh surface with locations in an image space of an image to be rendered onto the mesh surface are then processed to determine mapping between the co-ordinate space of the mesh data and the image space of an image to be rendered onto the mesh surface. Each eye-movement pattern is recognized by comparing the elementary features with a predetermined eye-movement pattern template.