%0 Journal Article %9 ACL : Articles dans des revues avec comité de lecture répertoriées par l'AERES %A Paradis, Emmanuel %T Reduced multidimensional scaling %D 2022 %L fdi:010082101 %G ENG %J Computational Statistics %@ 0943-4062 %K Dimension reduction ; Distance data ; HIV ; Multidimensional scaling %M ISI:000658117600001 %P 91-105 %R 10.1007/s00180-021-01116-0 %U https://www.documentation.ird.fr/hor/fdi:010082101 %> https://www.documentation.ird.fr/intranet/publi/2021-07/010082101.pdf %V 37 %W Horizon (IRD) %X Dimension reduction is a common problem when analysing large data sets. The present paper proposes a method called reduced multidimensional scaling based on performing an initial standard multidimensional scaling on a reduced data set. This method faces the problem of finding a representative reduced sample. An algorithm is presented to perform this selection based on alternating sampling in outlier areas and observations in high density areas. A space is then constructed with the selected reduced sample by standard multidimentional scaling using pairwise distances. The observations not included in the reduced sample are then projected on the constructed space using Gower's formula in order to obtain a final representation of the whole data set. The only requirement is the ability to compute distances among observations. A simulation study showed that the proposed algorithm results performs well to detect outliers. Evaluation of running times suggests that the proposed method could run in a few hours with data sets that would take more than one year to analyse with standard multidimensional scaling. An application is presented with a dataset of 9547 DNA sequences of human immunodeficiency viruses. %$ 020