This paper presents the analysis of a video excerpt of a professional meeting held in Paris. It studies how different kinds of linguistic resources (mainly lexical and syntactic) are mobilized with other embodied resources (gestures, gazes, manipulation of objects), within particular sequential environments in social interaction. Thus, the analysis demonstrates the interplay between multimodal resources and sequential organisation, and the necessity to take them into consideration in order to understand the detailed order of talk-in-interaction.