<?xml version="1.0" encoding="ISO-8859-1"?><article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id>0123-3033</journal-id>
<journal-title><![CDATA[Ingeniería y competitividad]]></journal-title>
<abbrev-journal-title><![CDATA[Ing. compet.]]></abbrev-journal-title>
<issn>0123-3033</issn>
<publisher>
<publisher-name><![CDATA[Facultad de Ingeniería, Universidad del Valle]]></publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id>S0123-30332013000200014</article-id>
<title-group>
<article-title xml:lang="en"><![CDATA[Evaluation of disparity maps]]></article-title>
<article-title xml:lang="es"><![CDATA[Evaluación de mapas de disparidad]]></article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname><![CDATA[Cabezas]]></surname>
<given-names><![CDATA[Ivan]]></given-names>
</name>
<xref ref-type="aff" rid="A01"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname><![CDATA[Trujillo]]></surname>
<given-names><![CDATA[Maria]]></given-names>
</name>
<xref ref-type="aff" rid="A02"/>
</contrib>
</contrib-group>
<aff id="A01">
<institution><![CDATA[,Universidad del Valle Escuela de Ingeniería de Sistemas y Computación ]]></institution>
<addr-line><![CDATA[Cali ]]></addr-line>
<country>Colombia</country>
</aff>
<aff id="A02">
<institution><![CDATA[,Universidad del Valle Escuela de Ingeniería de Sistemas y Computación ]]></institution>
<addr-line><![CDATA[Cali ]]></addr-line>
<country>Colombia</country>
</aff>
<pub-date pub-type="pub">
<day>00</day>
<month>12</month>
<year>2013</year>
</pub-date>
<pub-date pub-type="epub">
<day>00</day>
<month>12</month>
<year>2013</year>
</pub-date>
<volume>15</volume>
<numero>2</numero>
<fpage>151</fpage>
<lpage>161</lpage>
<copyright-statement/>
<copyright-year/>
<self-uri xlink:href="http://www.scielo.org.co/scielo.php?script=sci_arttext&amp;pid=S0123-30332013000200014&amp;lng=en&amp;nrm=iso"></self-uri><self-uri xlink:href="http://www.scielo.org.co/scielo.php?script=sci_abstract&amp;pid=S0123-30332013000200014&amp;lng=en&amp;nrm=iso"></self-uri><self-uri xlink:href="http://www.scielo.org.co/scielo.php?script=sci_pdf&amp;pid=S0123-30332013000200014&amp;lng=en&amp;nrm=iso"></self-uri><abstract abstract-type="short" xml:lang="en"><p><![CDATA[A disparity map is the output of a stereo correspondence algorithm. It is estimated in an intermediate step of a 3D information recovery process, from two or more images. A performance assessment of stereo correspondence algorithms may be addressed by a quantitative comparison of estimated disparity maps against ground-truth data. This assessment requires of the use of a methodology, which involves several evaluation elements and methods. Some elements and methods have been discussed with more attention than others in the literature. In the one hand, the quantity of used images and their relation to the application domain are topics rising large debate. On the other hand, there exist few publications on evaluation measures and error criteria. In practice, contradictory evaluation results may be obtained if different error measures are used, even on a same test-bed. In this paper, an evaluation methodology for stereo correspondence algorithms is presented. In contrast to conventional methodologies, it allows an interactive selection of multiple evaluation elements and methods. Moreover, it is based on a formal definition of error criteria based on set partitions. Experimental evaluation results showed that the proposed methodology allows a better understanding and analysis of algorithms performance than the Middlebury methodology. Final remarks highlights the relevance of discussing on the different elements and methods involved in an evaluation process]]></p></abstract>
<abstract abstract-type="short" xml:lang="es"><p><![CDATA[Un mapa de disparidad es la salida de un algoritmo de estimación de puntos correspondientes, el cual es estimado en una etapa intermedia del proceso de reconstrucción de la profundidad a partir de dos o máágenes. La comparación del desempeño de un grupo de algoritmos de estimación de correspondencia puede hacerse mediante una evaluación cuantitativa de mapas de disparidad contra mapas de referencia. Está evaluación requiere de una metodología, la cual involucra diversos elementos y métodos. Algunos de estos elementos y métodos han recibido más atención que otros en la literatura. La cantidad de imágenes utilizadas, y la relación entre el contenido de las mismas y los diferentes dominios de aplicación han sido temas de amplia discusión en la literatura. Por otra parte, existen pocas publicaciones que aborden los temas relacionados con las medidas y los criterios de evaluación. En la práctica, el uso de diferentes medidas podría conllevar a la obtención de resultados contradictorios, empleando inclusive un mismo conjunto de pruebas. Adicionalmente, las particularidades de diferentes dominios de aplicación pueden implicar requerimientos variables durante el proceso de evaluación. En este artículo se presenta una metodología de evaluación para algoritmos de estimación de correspondencia en imágenes estéreo. La metodología se considera como aumentada en la medida que, a diferencia de las metodologías convencionales, permite una selección interactiva de diferentes elementos y métodos de evaluación, con diferentes propiedades. En la presente metodología, se formaliza el concepto de criterios de error, mediante la teoría de conjuntos. La experimentación realizada mostró que el uso de la metodología propuesta provee resultados innovadores, realzando la relevancia de una discusión en los diferentes elementos y métodos involucrados en el proceso de evaluación]]></p></abstract>
<kwd-group>
<kwd lng="en"><![CDATA[Evaluation criteria]]></kwd>
<kwd lng="en"><![CDATA[evaluation methodologies]]></kwd>
<kwd lng="en"><![CDATA[evaluation models]]></kwd>
<kwd lng="en"><![CDATA[stereo correspondence]]></kwd>
<kwd lng="es"><![CDATA[Correspondencia estéreo]]></kwd>
<kwd lng="es"><![CDATA[criterios de evaluación]]></kwd>
<kwd lng="es"><![CDATA[metodologías de evaluación]]></kwd>
<kwd lng="es"><![CDATA[modelos de evaluación]]></kwd>
</kwd-group>
</article-meta>
</front><body><![CDATA[   <font size="2" face="Verdana, Geneva, sans-serif">      <p align="center"><font size="4"><b>Evaluation of disparity maps</b></font></p>      <p align="center"><font size="3"><b>Evaluaci&oacute;n de mapas de disparidad</b></font></p>      <p><i>Ivan Cabezas</i><br /> Escuela de Ingenier&iacute;a de Sistemas y Computaci&oacute;n. Universidad del Valle. Cali, Colombia<br /> E-mail: <a href="mailto:ivan.cabezas@correounivalle.edu.co">ivan.cabezas@correounivalle.edu.co</a></p>      <p><i>Maria Trujillo</i><br /> Escuela de Ingenier&iacute;a de Sistemas y Computaci&oacute;n. Universidad del Valle. Cali, Colombia<br /> E-mail: <a href="mailto:maria.trujillo@correounivalle.edu.co">maria.trujillo@correounivalle.edu.co</a></p>      <p><b>Eje tem&aacute;tico:</b> Systems engineering / Ingenier&iacute;a de sistemas<br /> Recibido: 26 de Abril de 2013<br /> Aceptado: 06 de Septiembre de 2013</p>  <hr />      <p><font size="3"><b>Abstract</b></font></p>      <p>A disparity map is the output of a stereo correspondence algorithm. It is estimated in an intermediate step of a 3D information recovery process, from two or more images. A performance assessment of stereo correspondence algorithms may be addressed by a quantitative comparison of estimated disparity maps against ground-truth data. This assessment requires of the use of a methodology, which involves several evaluation elements and methods. Some elements and methods have been discussed with more attention than others in the literature. In the one hand, the quantity of used images and their relation to the application domain are topics rising large debate. On the other hand, there exist few publications on evaluation measures and error criteria. In practice, contradictory evaluation results may be obtained if different error measures are used, even on a same test-bed. In this paper, an evaluation methodology for stereo correspondence algorithms is presented. In contrast to conventional methodologies, it allows an interactive selection of multiple evaluation elements and methods. Moreover, it is based on a formal definition of error criteria based on set partitions. Experimental evaluation results showed that the proposed methodology allows a better understanding and analysis of algorithms performance than the Middlebury methodology. Final remarks highlights the relevance of discussing on the different elements and methods involved in an evaluation process.</p>      <p><b>Keywords:</b> Evaluation criteria, evaluation methodologies, evaluation models, stereo correspondence.</p>  <hr />      <p><font size="3"><b>Resumen</b></font></p>      ]]></body>
<body><![CDATA[<p>Un mapa de disparidad es la salida de un algoritmo de estimaci&oacute;n de puntos correspondientes, el cual es estimado en una etapa intermedia del proceso de reconstrucci&oacute;n de la profundidad a partir de dos o m&aacute;s im&aacute;genes. La comparaci&oacute;n del desempe&ntilde;o de un grupo de algoritmos de estimaci&oacute;n de correspondencia puede hacerse mediante una evaluaci&oacute;n cuantitativa de mapas de disparidad contra mapas de referencia. Est&aacute; evaluaci&oacute;n requiere de una metodolog&iacute;a, la cual involucra diversos elementos y m&eacute;todos. Algunos de estos elementos y m&eacute;todos han recibido m&aacute;s atenci&oacute;n que otros en la literatura. La cantidad de im&aacute;genes utilizadas, y la relaci&oacute;n entre el contenido de las mismas y los diferentes dominios de aplicaci&oacute;n han sido temas de amplia discusi&oacute;n en la literatura. Por otra parte, existen pocas publicaciones que aborden los temas relacionados con las medidas y los criterios de evaluaci&oacute;n. En la pr&aacute;ctica, el uso de diferentes medidas podr&iacute;a conllevar a la obtenci&oacute;n de resultados contradictorios, empleando inclusive un mismo conjunto de pruebas. Adicionalmente, las particularidades de diferentes dominios de aplicaci&oacute;n pueden implicar requerimientos variables durante el proceso de evaluaci&oacute;n. En este art&iacute;culo se presenta una metodolog&iacute;a de evaluaci&oacute;n para algoritmos de estimaci&oacute;n de correspondencia en im&aacute;genes est&eacute;reo. La metodolog&iacute;a se considera como aumentada en la medida que, a diferencia de las metodolog&iacute;as convencionales, permite una selecci&oacute;n interactiva de diferentes elementos y m&eacute;todos de evaluaci&oacute;n, con diferentes propiedades. En la presente metodolog&iacute;a, se formaliza el concepto de criterios de error, mediante la teor&iacute;a de conjuntos. La experimentaci&oacute;n realizada mostr&oacute; que el uso de la metodolog&iacute;a propuesta provee resultados innovadores, realzando la relevancia de una discusi&oacute;n en los diferentes elementos y m&eacute;todos involucrados en el proceso de evaluaci&oacute;n.</p>      <p><b>Palabras clave:</b> Correspondencia est&eacute;reo, criterios de evaluaci&oacute;n, metodolog&iacute;as de evaluaci&oacute;n, modelos de evaluaci&oacute;n.</p>  <hr />      <p><font size="3"><b>1. Introduction</b></font></p>      <p>A quantitative comparison of estimated disparity maps allows evaluating, in a fair basis, the performance of Stereo Correspondence Algorithms (SCA) as well as algorithmic components and procedures, (Scharstein &amp; Szeliski, 2002; Neilson &amp; Yang, 2008; Cabezas and Trujillo, 2013), among others. An evaluation process can be addressed using either a qualitative or a quantitative approach. Although a qualitative evaluation approach, which is based on human viewing experiences, may properly take into account factors that are complex to quantify in an automatic process (Trucco &amp; Ruggeri, 2013), it is time and resources consuming. Moreover, obtained results may be not repeatable (Wang et al., 2004). A quantitative evaluation of SCA can be automatically addressed using Disparity Ground-Truth Data (DGTD). However, the generation of DGTD may impose constraints on content of captured imagery test-bed (Geiger et al., 2012). In practice, quantitative evaluation methodologies, which are based on comparing estimated disparity maps against DGTD, are widely adopted. In general, an evaluation methodology for assessing estimated disparity maps is composed by a set of elements and methods, which interact following a sequence of steps, as it is illustrated in <a href="#fig1">Figure 1</a>. Two fundamental evaluation element and method are error criteria and error measures, respectively. Error criteria define image regions of interest, on which errors are calculated, allowing a detailed evaluation according to the application domain. Error measures quantify differences among estimated data and ground-truth data. </P>      <p align="center"><a name="fig1"><img src="img/revistas/inco/v15n2/v15n2a14-fig1.jpg" /></a></p>       <p>Most of published papers introducing SCA rely on the use of the Middlebury's methodology (Scharstein &amp; Szeliski, 2002; 2013). This methodology can be analyzed based on the evaluation elements and methods depicted in <a href="#fig1">Figure 1</a>. In particular, three aspects require attention for the sake of the discussion presented in this paper: the error measure, the error criteria, and the evaluation model. Regarding the error measure, the Middlebury's methodology uses the percentage of the Bad Matched Pixels (BMP) as the error measure. The BMP is a binary function, which calculates an estimation error using a threshold and ignores the inverse relation between depth and disparity along with error magnitude (Cabezas et al., 2011). Consequently, it does not properly distinguish between a small and a large disparity error; neither considers if an evaluated point is close or far from the stereo camera system. These considerations are illustrated in <a href="#fig2">Figure 2</a>, assuming a canonical stereo rig. <a href="#fig2">Figure 2(a)</a> and <a href="#fig2">Figure 2(c)</a> illustrate how estimation errors of a same magnitude - represented by points p'<sub>r</sub> and q'<sub>r</sub> respectively - may cause different 3D reconstruction errors by triangulation -represented by points P' and Q&acute;, respectively. <a href="#fig2">Figure 2(b)</a> and <a href="#fig2">Figure 2(d)</a> illustrate how a larger estimation error magnitude increases the 3D reconstruction error by triangulation. These 3D reconstruction errors have to be taking into account during an evaluation process. The impact of the limitations of the BMP measure is illustrated in <a href="#fig3">Figure 3</a>. The left view of the Tsukuba stereo image pair and associated ground-truth disparity map are shown in <a href="#fig3">Figure 3(a)</a>, and <a href="#fig3">Figure 3(b)</a>, respectively (Scharstein &amp; Szeliski, 2013). <a href="#fig3">Figure 3(c)</a> and <a href="#fig3">Figure 3(d)</a> show erroneously estimated disparity maps. <a href="#fig3">Figure 3(e)</a> and <a href="#fig3">Figure 3(f)</a> show corrupted disparity maps by adding salt and pepper noise. It can be observed that the calculated disparity maps in <a href="#fig3">Figure 3(c)</a> and <a href="#fig3">Figure 3(d)</a> contain errors, in the background and the foreground, whilst the maps in <a href="#fig3">Figure 3(e)</a> and <a href="#fig3">Figure 3(f)</a> contain a similar quantity of errors, but with a small and a large magnitude, respectively. In addition, <a href="#fig3">Figure 3</a> includes the values of the BMP percentage and the Peak Signal-to-Noise Ratio (PSNR) for the estimated disparity maps. It is clear that the BMP is incapable of distinguishing errors of different  magnitudes. Regarding error criteria, three of them are simultaneously used in the Middlebury's methodology. The <I>disc</I> criterion considers errors in points near depth discontinuities. The <I>nonocc </I>criterion considers errors in non-occluded points. The <I>all</I> criterion includes the points in the whole image (i.e. for those which a disparity groundtruth value is available). Nevertheless, an image point may be included in more than one error criterion. This is illustrated in <a href="#fig4">Figure 4</a> using the Teddy stereo image. The ground-truth disparity map of the Teddy image is shown in <a href="#fig4">Figure 4(a)</a>. It can be observed that all points included in the <I>disc</I> criterion - <a href="#fig4">Figure 4(b)</a> - are also included in the <I>nonocc</I> criterion - <a href="#fig4">Figure 4(c)</a>. In addition, all points included in the <I>nonocc</I> criterion are also included in the <I>all</I> criterion - Figure 4(d). The relation among the points composing each one of the above criterion is illustrated in <a href="#fig4">Figure 4(e)</a> using a Venn diagram. Thus, a disparity estimation error is counted more than once, biasing final scores, in the Middlebury methodology. The multiple counting of errors is quantified in <a href="#tab1">Table 1</a>, using the evaluation of the Venus stereo image, a set of selected algorithms, and the threshold equal to 1 pixel. It can be observed that errors associated to the <I>nonocc</I> criterion are in fact errors associated to the <I>disc</I> criterion. Moreover, the total of score using for ranking is almost twice the number of errors in the disparity map. Regarding the evaluation model, the model of Middlebury's methodology can be seen as a linear function which relates ranks to weights. It is based on sorting BMP scores from each error criterion, ranking sorted positions, and averaging all rankings in order to obtain a final ranking. However, although there are different evaluation methodologies for evaluating SCA, a unique methodology that properly handles all evaluation requirements may not exist. This may be due to, in existing methodologies, evaluation elements and methods are fixed beforehand, assuming that different Research and Development (R&amp;D) processes will have similar evaluation requirements. Consequently, allowed evaluation scenarios are fixed. Nevertheless, requirements may indeed change according to some particularities such as the application domain, or in general, to evaluation goals. In practice, problems arise when specific components of a methodology do not fit to some evaluation requirements, or they do not provide proper feedback. Thus, there is a lack of an adaptive evaluation methodology allowing the selection of diverse evaluation elements and methods. Moreover evaluation results may be biased due to possible shortcomings of considered evaluation elements and methods. In this paper, an evaluation methodology for SCA based on DGTD is presented. The presented methodology introduces a set of error criteria in order to allow a proper analysis of disparity estimation errors. In addition, it includes two evaluation models capable of handling multiple evaluation measures: the A* Groups model (Cabezas et al., 2012a), which is based on the Pareto Dominance relation (Van Veldhuizen et al., 2003), and an extension to the Middlebury's evaluation model (Scharstein &amp; Szeliski, 2002). In this way, the proposed methodology suited to be used in different phases of R&amp;D processes. Experimental results are shown in order to exemplify the versatility of the presented methodology. </p>       <p align="center"><a name="fig2"><img src="img/revistas/inco/v15n2/v15n2a14-fig2.jpg" /></a></p>      <p align="center"><a name="fig3"><img src="img/revistas/inco/v15n2/v15n2a14-fig3.jpg" /></a></p>      <p align="center"><a name="fig4"><img src="img/revistas/inco/v15n2/v15n2a14-fig4.jpg" /></a></p>      <p align="center"><a name="tab1"><img src="img/revistas/inco/v15n2/v15n2a14-tab1.jpg" /></a></p>      ]]></body>
<body><![CDATA[<p><font size="3"><B>2. Related work</B></font></p>       <p>The state-of-the-art is briefly reviewed, in thissection, using as reference the methodology presented in <a href="#fig1">Figure 1</a>. An extensive review and a thorough discussion on some mentioned aspects can be found in (Cabezas &amp; Trujillo, 2013). </P>     <p><B>Test-bed images: </B>Ground-truth data can be generated by either using a ray tracing algorithm -synthetic data-, or, using an active vision technique such as structured light (Scharstein &amp; Szeliski, 2003) or laser scanning (Geiger et al., 2012), among others -real imagery data. Synthetic data, generated considering noise models, have been used in different approaches (Van der Mark &amp; Gavrila, 2006; Neilson &amp; Yang, 2008). However, introduced noise in a systematic way may not necessarily correspond to real image capturing conditions. The generation of DGTD for real imagery is a challenging task, which is not always possible. The Middlebury's methodology use a test-bed of four stereo images (the Tsukuba, the Venus, the Teddy and the Cones stereo images). It is available in an online evaluation platform (Scharstein &amp; Szeliski, 2013). Such test-bed is widely known and used by the stereo vision community. Regarding the quantity of images selected as test-bed, when a small number of images are used, obtained evaluation results may lack of statistical significance (Cabezas &amp; Trujillo, 2013). In general terms, is not convenient to consider the obtained results using a particular test-bed, as of general character (i.e. to be repeatable under a different imagery test-bed) (Cabezas &amp; Trujillo, 2011). </P>     <p><B>Evaluation criteria: </B>The concept of error criteria was introduced in (Scharstein &amp; Szeliski, 2002) as binary image segmentation. Most of methodologies consider evaluation criteria related to disparity estimation errors, which are related to challenging content for SCA (Scharstein &amp; Szeliski, 2002). However, considering aspects such as the consumed time and required resources (e.g. memory and the use of specialized hardware, among others) may enhance the evaluation process. </P>     <p><B>Evaluation measures:</B> Several evaluation measures are available in the literature. The BMP is based on counting disparity estimation errors exceeding a threshold &delta;, the most commonly used value is 1 pixel (Scharstein &amp; Szeliski, 2002). The Sigma-Z-Error (SZE) is based on the inverse relation between depth and disparity, and aims to measure the impact on depth estimation of disparity estimation errors (Cabezas et al., 2011). The Mean Absolute Error (MAE) is based on absolute differences. The Mean Square Error (MSE) is based on quadratic differences, and the Mean Relative Error (MRE) is based on the ratio between absolute differences and the ground-truth disparity (Van der Mark &amp; Gavrila, 2006). The MAE, the MSE, and the SZE are metric functions. However, they are unbounded. In (Cabezas et al., 2012b) is highlighted how obtained evaluation results may vary according to the selection of error measures. </P>       <p><B>Algorithms and/or algorithmic components: </B>In the one hand,the estimation of stereo corresponding point can be tackled as an optimization problem under a constrained scenario. Consequently, there are multiple approaches for estimating disparity maps. In the Middlebury's online benchmark (Scharstein &amp; Szeliski, 2013) all the reported algorithms are compared regardless the nature of their optimisation technique. Moreover, in some cases, compared algorithms correspond to unpublished works. This makes difficult an analysis of obtained results. On the other hand, for evaluation purposes, it is expected that SCA be executed under the same conditions (i.e. with the same information, or with fixed parameters for the entire test-bed). However, in practice, this may be beyond control of the evaluation methodology. </P>      <p><B>Evaluation models: </B>The linear approach used by the Middlebury's evaluation model has been used in other methodologies (Cabezas &amp; Trujillo, 2013). In contrast, the A* Groups evaluation model proposed in (Cabezas et al., 2012a) addresses the comparison of SCA as a multiobjective optimization problem. This non-linear model is based on the Pareto dominance relation (Van Veldhuizen et al., 2003), and extends the model introduced in (Cabezas &amp; Trujillo, 2011). It iteratively computes groups of SCA -A* sets- with comparable performance (i.e. not better, neither worst), according to scores of evaluation measures. Computed groups define a partition of the original set of SCA under evaluation. Among them, the A<sup>*</sup><sub>1</sub> group is of special interest since it is composed by the SCA of superior performance, under a specific evaluation scenario. </P>      <p><B>Interpret results: </B>In a raking based model, a higher ranking is associated to a superior performance. However, some issues may arise in such model. For instance, two algorithms may have the same error scores but they will not obtain the same ranking. Moreover, it is not clear when two rankings are close or distant enough to affirm that the performance of associated SCA may be considered as similar or different, respectively. In addition, in this model, the number of top-performer algorithms is a free parameter. In contrast, in the A* Groups model, the interpretation of results is defined, without ambiguity, based on the cardinality of the A* set and the group label assigned to it. In this way, researchers and practitioners may obtain an unambiguous feedback. </P>      <p><font size="3"><B>3. An evaluation methodology</B></font></p>      <p>The methodology follows the steps illustrated in <a href="#fig1">Figure 1</a>. It offers the possibility of choosing different evaluation elements and methods, according to evaluation requirements. Moreover, it includes an extension to the Middlebury's evaluation model by considering multiple error measures, and introduces a formalization of error criteria. </P>      ]]></body>
<body><![CDATA[<p><font size="3">3.1 Multiple error measures with different properties</font></p>      <p>The MAE, the MSE, the MRE and the SZE measures consider the disparity estimation error magnitude. In addition, the MRE and the SZE measures consider the inverse relation between depth and disparity. The use of multiple measures, with different properties makes an evaluation process less sensitive to the selection of error measure, since each one may capture a different aspect of the estimated maps. In this way, the possible weaknesses of a specific measure may be compensated by the strength of another one. Moreover, different measures can be used in a complementary way focusing on measuring specific aspects. </P>       <p><font size="3">3.2 An error criteria definition</font></p>      <p>An error criterion is conceived as a membership function defining a set partition of points (pixels) belonging to a disparity map. It identifies challenging image points associated to a specific and unique meaning. For the sake of completeness, required concepts are defined as follows. Let I be a set of image points composing the reference view from a stereo image pair: </P>  <img src="img/revistas/inco/v15n2/v15n2a14-for1.jpg" />      <p>where T is the amount of points composing a reference image for which ground-truth disparity value is known. Let M be a set defined as: </P>  <img src="img/revistas/inco/v15n2/v15n2a14-for2.jpg" />      <p>In addition, a set C of error criteria is considered. Let C be a set of onto-functions defined as: </P>  <img src="img/revistas/inco/v15n2/v15n2a14-for3.jpg" />      <p>where: </P>   <img src="img/revistas/inco/v15n2/v15n2a14-for4.jpg" />      <p>Thus, an error criterion <I>ci </I>is an onto-function with domain in I, and codomain in M: </P>  <img src="img/revistas/inco/v15n2/v15n2a14-for5.jpg" />      <p>where:</p>   <img src="img/revistas/inco/v15n2/v15n2a14-for6-7.jpg" />      <P>where Pis a <I>positive </I>subset of I, according to the onto-function <I>ci</I>, defined as:</P>  <img src="img/revistas/inco/v15n2/v15n2a14-for8.jpg" />      ]]></body>
<body><![CDATA[<P> Analogously, let N be a subset of I, according to the onto-function C, defined as: </P>  <img src="img/revistas/inco/v15n2/v15n2a14-for9.jpg" />      <P>Thus, an onto-function <I>ci </I>defines a partition over I, by fulfilling the following properties:</P>      <p><img src="img/revistas/inco/v15n2/v15n2a14-for10-13.jpg" /></p>     <p> Let P<sub>Ck</sub> be the <I>positive</I> set of the onto-function ck</p>      <p><img src="img/revistas/inco/v15n2/v15n2a14-for14-19.jpg" /></p>     <p>Based on the above definition, three meaningful error criteria are identified as follows. </p>      <p><B><I>boundary</I></B>: this criterion considers points near to both depth discontinuities and occluded regions. In this case the property of being near can be computed by a function using thresholds for determining a neighbourhood and what a depth discontinuity is. This criterion offers backward compatibility with the <I>disc</I> criterion used in the Middlebury's methodology. <B><I>interior</I></B>: this criterion considers image points which are visible in both stereo images, and far enough from depth discontinuities and occluded regions. <B><I>occluded</I></B>: this criterion considers occluded points, which can be detected by forward projecting the reference view, or applying the bi-directional constraint on DGTD. </P>       <p>In practice, each criterion can be represented and stored as a binary image mask. <a href="#fig5">Figure 5</a> shows the image masks associated to the three proposed error criteria, using the Teddy stereo image. Figure 5(a) illustrates the <I>boundary</I> criterion. <a href="#fig5">Figure 5</a> (b) illustrates the <I>interior </I>criterion, and Figure 5 (c) illustrates the <I>occluded</I> criterion. Figure 5(d) represents a generalisation of the introduced formalisation, applied to error criteria presented above, as disjoint sets which union compose the set of points to be evaluated in a disparity map. </P>      <p align="center"><a name="fig5"><img src="img/revistas/inco/v15n2/v15n2a14-fig5.jpg" /></a></p>      <p>Following the presented definition for evaluationcriteria, the <I>boundary</I>, the <I>interior</I> and the <I>occluded</I> criteria are composed by different image points. The amount of disparity estimation errors with a deviation of more than 1 pixel from the real disparity value, associated to introduced criteria are shown in <a href="#tab1">Table 1</a> using the Venus stereo image and a set of selected algorithms. Errors based on a criterion do not imply errors based on another criterion. </P>      ]]></body>
<body><![CDATA[<p><font size="3">3.3 Extending the middlebury's evaluation model</font></p>      <p>The Middlebury's evaluation model is based on averaged rankings. It considers exclusively the BMP measure. The introduced extension incorporates multiple measures. It is described as follows. The conventional Middlebury's evaluation model is applied, separately, for each selected error measure. Then, the final ranking of each SCA is computed by summing all the intermediate rankings, and sorting the sum. In this way, a discrete value is obtained as result. </P>       <p>A &tau; threshold (with the number of considered error measures as suggested value), is used to determine if two algorithms show a similar performance, in the following way: if the absolute distance between the sum of rankings is less than &tau;, the performance of the two algorithms can be considered as similar. This criterion is applied to two algorithms under comparison. It defines a reflexive and symmetric relation. The extension provides robustness by considering multiple error measures. </P>      <p><font size="3">3.4 An interactive evaluation process</font></p>      <p>The presented methodology considers several choices during the evaluation process, which are outlined as follows. Test-bed images can be selected, for instance, among the Tsukuba, the Venus, the Teddy, and the Cones stereo images (Scharstein &amp; Szeliski, 2013). The evaluation criteria can be selected either among the introduced criteria (the <I>boundary</I>, the <I>interior</I>, and the <I>occluded</I> error criteria), or among the Middlebury criteria (the <I>nonocc</I>, the <I>disc</I> and the <I>all </I>criteria). Evaluation measures can be selected among available evaluation measures, such as: the BMP, the SZE, the MAE, the MSE, and the MRE, or a combination of them. Evaluated algorithms can be chosen from, for instance, the repository available at (Scharstein &amp; Szeliski, 2013). The evaluation involves a set of algorithms, which can be selected algorithms according to user requirements in order to focus a comparison in SCA based on similar optimisation techniques. The evaluation model can be selected from the A* Groups model (Cabezas et al., 2012a), the Middlebury's model (Scharstein &amp; Szeliski, 2002), or from its introduced extended version. </P>       <p><font size="3"><b>4. Experimental validation</b></font></p>      <p>The presented methodology offers multiple evaluation possibilities. Only two different evaluation scenarios are presented in order to illustrate versatility, due to some constraints in space. Both scenarios consider the full set of testbed images available at (Scharstein &amp; Szeliski, 2013). The first scenario is devised for comparing performance of SCA in regions visible in both images. It uses the <I>boundary </I>evaluation criterion, the full set of error measures presented in section 3.4, the entire set of SCA, and the extended Middlebury's evaluation model. The top-fifteen ranked algorithms under this evaluation scenario are shown in <a href="#tab2">Table 2</a>. In this case, the suggested value for the threshold &tau; is five, which is applied to the sum of rankings. Multiple observations can be extracted from obtained results. For instance, it can be observed that several algorithms belonging to variations of the semi-global approach (Hirschmuller, 2005) are present. Regarding the inclusion of multiple error measures, each measure can be seen as an expert, contributing in a multi-expert evaluation approach, reducing the impact of selecting a single measure. The second scenario is devised to find algorithms showing the best performance (in terms of 3D reconstruction accuracy, by considering the impact of disparity estimation errors) in areas near to depth discontinuities. It uses the <I>occluded </I>evaluation criterion, the MRE and the SZE measures, the entire set of SCA, and the A* Groups evaluation model, which, by definition, properly handles multiple evaluation measures. The members of the A*1 set are: WarpMat, SurfaceStereo, </P>    <p>Unsupervised, AdaptingBP, Segm+visib, ObjectStereo, and AdaptingOverSegBP, CurveletSupWgt, and InteriorPtLP (Scharstein &amp; Szeliski, 2013). Several of them use an approach based on segmentation, assigning disparity values to each segment. Eight groups of SCA are selected by the used evaluation model. The cardinality of each A*group is: |A*<sub>1</sub>| = 9, |A*<sub>2</sub>|=24, |A*<sub>3</sub>|=23, |A*<sub>4</sub>|=14, |A*<sub>5</sub>|=9, |A*<sub>6</sub>|=11, |A*<sub>7</sub>|=8, |A*<sub>8</sub>|=7, and |A*<sub>9</sub>|=6 respectively.</p>      <p align="center"><a name="tab2"><img src="img/revistas/inco/v15n2/v15n2a14-tab2.jpg" /></a></p>      <p><font size="3"><B>5. Conclusions</B></font></p>      ]]></body>
<body><![CDATA[<p>An interactive evaluation methodology for comparing SCA, considering the use of multiple error measures with different evaluation properties was presented. Taking this into account, an extension to the Middlebury's evaluation model was introduced. Moreover, the presented methodology introduces a defintion of error criteria avoiding ambiguity during the gathering of score errors, along with an innovative criterion allowing the evaluation of disparity assignations in occluded regions. Error measures and error criteria are selected by a user, in order to obtain more reliable and useful evaluation results. In this way, users are allowed to define different evaluation scenarios, according to evaluation requirements. Thus the presented methodology provides state-of-the-art evaluation capabilities. Nevertheless, although the extension to the model introduces reliability to evaluation results, it may no alleviate the inherent issues of a ranking based model. Thus, the use of the A* Groups model, which is based on the Pareto dominance relation, arises as a proper alternative. </P>  <hr />      <p><font size="3"><b>6. References</b></font></p>      <!-- ref --><p>Cabezas, I., &amp; Trujillo M., (2013). Methodologies for evaluating disparity estimation algorithms. In, J. Garc&iacute;a-Rodriguez &amp; M. Cazorla (Eds.) <i>Robotic Vision - Technologies for machine Learning and Vision Applications, </i>(cap. 10). Hershey PA: Information Science Reference - IGI Global.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000065&pid=S0123-3033201300020001400001&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Cabezas, I., &amp; Trujillo, M. (2011). A non-linear quantitative evaluation approach for disparity estimation. <i>International Conference on Computer Vision, Theory and Applications </i> (pp. 704-709). Algarve, Portugal.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000066&pid=S0123-3033201300020001400002&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Cabezas, I., Padilla, V., &amp; Trujillo, M. (2011). A measure for accuracy disparity maps evaluation. <i>Iberoamerican Congress on Pattern Recognition </i>. Lecture Notes in Computer Science (vol. 7042) (pp. 223-231). Buenos Aires, Argentina: Springer.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000067&pid=S0123-3033201300020001400003&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Cabezas, I., Padilla, V., Trujillo, M., Florian, M. (2012). On the impact of the error measure selection in evaluating disparity maps. <i>World Automation Congress - WAC </i> (pp. 1-6). Puerto Vallarta, Mexico.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000068&pid=S0123-3033201300020001400004&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Cabezas, I., Trujillo, M., Florian, M. (2012). An evaluation methodology for stereo correspondence algorithms. <i>International Conference on Computer Vision, Theory and Applications </i> (pp. 154-163). Rome, Italy.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000069&pid=S0123-3033201300020001400005&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Geiger, A., Lenz, P., &amp; Urtasun, R., (2012). Are we ready for autonomous driving? The KITTI vision benchmark suite. <i>IEEE Conference on Computer Vision and Pattern Recognition </i> (pp. 3354-3361).    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000070&pid=S0123-3033201300020001400006&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Hirschmuller, H., (2005). Accurate and efficient stereo processing by semi-global matching and mutual information. <i>IEEE Conference on Computer Vision and Pattern Recognition </i> (pp. 807-814).    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000071&pid=S0123-3033201300020001400007&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Neilson, D., &amp; Yang, Y. (2008). <i>Evaluation of constructible match cost measures for stereo correspondence using cluster ranking </i>. Computer Vision and Pattern Recognition. IEEE Computer Society (pp. 1-8).    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000072&pid=S0123-3033201300020001400008&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Scharstein, D., &amp; Szeliski, R. (2002). A taxonomy and evaluation of dense two-frame stereo correspondence algorithms. <i>International Journal of Computer Vision </i>, 47, 7-42.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000073&pid=S0123-3033201300020001400009&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Scharstein, D., &amp; Szeliski, R. (2003/06/18-20). High&shy;accuracy stereo depth maps using structured light. In IEEE. <i>Computer Vision and Pattern Recognition, 2003. Proceedings. 2003 IEEE Computer Society Conference on </i>. (pp. 195-202).    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000074&pid=S0123-3033201300020001400010&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Scharstein, D., &amp; Szeliski, R. (2013). <i>Middlebury stereo evaluation - version 2 </i>. Recovered 2013/12/15 <a href="http://vision. middlebury.edu/stereo/eval/" target="_blank">http://vision. middlebury.edu/stereo/eval/</a>     &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000075&pid=S0123-3033201300020001400011&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Trucco, E., &amp; Ruggeri, A. (2013). Towards a multi-site International public dataset for the validation of retinal image analysis software <i>. International Conference of the IEEE Engineering in Medicine and Biology Society </i> (pp. 7152-7155). Osaka, Japan: IEEE.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000076&pid=S0123-3033201300020001400012&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Van der Mark, W., &amp; Gavrila, D. (2006). Real-time dense stereo for intelligent vehicles. IEEE <i>Transactions on Intelligent Transportation Systems </i>, 7 (1), 38-50.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000077&pid=S0123-3033201300020001400013&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --><!-- ref --><p> Van Veldhuizen, D., Zydallis, J., &amp; Lamont, G. (2003). Considerations in engineering parallel multiobjective evolutionary algorithms. <i>IEEE Transactions on Evolutionary Computation </i>, 7(2), 144-173.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000078&pid=S0123-3033201300020001400014&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --></p>        <!-- ref --><p>Wang, Z., Bovik, A., Sheikh, H., &amp; Simoncelli, E. (2004). Image quality assessment: from error visibility to structural similarity. <i>IEEE Transactions on Image Processing</i>, 13 (4), 600-612.    &nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;[&#160;<a href="javascript:void(0);" onclick="javascript: window.open('/scielo.php?script=sci_nlinks&ref=000080&pid=S0123-3033201300020001400015&lng=','','width=640,height=500,resizable=yes,scrollbars=1,menubar=yes,');">Links</a>&#160;]<!-- end-ref --></p>       <p><img src="img/revistas/inco/v15n2/cc.jpg">    ]]></body>
<body><![CDATA[<br> Revista Ingenier&iacute;a y Competitividad por Universidad del Valle se encuentra bajo una <a href="https://creativecommons.org/licenses/by/3.0/deed.es_ES" target="_blank">licencia Creative Commons Reconocimiento</a> - Debe reconocer adecuadamente la autor&iacute;a, proporcionar un enlace a la licencia e indicar si se han realizado cambios. Puede hacerlo de cualquier manera razonable, pero no de una manera que sugiera que tiene el apoyo del licenciador o lo recibe por el uso que hace.</p>  </font>      ]]></body><back>
<ref-list>
<ref id="B1">
<nlm-citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Cabezas]]></surname>
<given-names><![CDATA[I.]]></given-names>
</name>
<name>
<surname><![CDATA[Trujillo]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Methodologies for evaluating disparity estimation algorithms]]></article-title>
<person-group person-group-type="editor">
<name>
<surname><![CDATA[García-Rodriguez]]></surname>
<given-names><![CDATA[J.]]></given-names>
</name>
<name>
<surname><![CDATA[Cazorla]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
</person-group>
<source><![CDATA[Robotic Vision - Technologies for machine Learning and Vision Applications]]></source>
<year>2013</year>
<publisher-loc><![CDATA[Hershey PA ]]></publisher-loc>
<publisher-name><![CDATA[Information Science Reference - IGI Global]]></publisher-name>
</nlm-citation>
</ref>
<ref id="B2">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Cabezas]]></surname>
<given-names><![CDATA[I.]]></given-names>
</name>
<name>
<surname><![CDATA[Trujillo]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[A non-linear quantitative evaluation approach for disparity estimation]]></article-title>
<source><![CDATA[International Conference on Computer Vision, Theory and Applications]]></source>
<year>2011</year>
<page-range>704-709</page-range><publisher-loc><![CDATA[Algarve ]]></publisher-loc>
</nlm-citation>
</ref>
<ref id="B3">
<nlm-citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Cabezas]]></surname>
<given-names><![CDATA[I.]]></given-names>
</name>
<name>
<surname><![CDATA[Padilla]]></surname>
<given-names><![CDATA[V.]]></given-names>
</name>
<name>
<surname><![CDATA[Trujillo]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[A measure for accuracy disparity maps evaluation]]></article-title>
<source><![CDATA[Iberoamerican Congress on Pattern Recognition]]></source>
<year>2011</year>
<volume>7042</volume>
<page-range>223-231</page-range><publisher-loc><![CDATA[Buenos Aires ]]></publisher-loc>
<publisher-name><![CDATA[Springer]]></publisher-name>
</nlm-citation>
</ref>
<ref id="B4">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Cabezas]]></surname>
<given-names><![CDATA[I.]]></given-names>
</name>
<name>
<surname><![CDATA[Padilla]]></surname>
<given-names><![CDATA[V.]]></given-names>
</name>
<name>
<surname><![CDATA[Trujillo]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
<name>
<surname><![CDATA[Florian]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[On the impact of the error measure selection in evaluating disparity maps]]></article-title>
<source><![CDATA[World Automation Congress - WAC]]></source>
<year>2012</year>
<page-range>1-6</page-range><publisher-loc><![CDATA[Puerto Vallarta ]]></publisher-loc>
</nlm-citation>
</ref>
<ref id="B5">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Cabezas]]></surname>
<given-names><![CDATA[I.]]></given-names>
</name>
<name>
<surname><![CDATA[Trujillo]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
<name>
<surname><![CDATA[Florian]]></surname>
<given-names><![CDATA[M.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[An evaluation methodology for stereo correspondence algorithms]]></article-title>
<source><![CDATA[International Conference on Computer Vision, Theory and Applications]]></source>
<year>2012</year>
<page-range>154-163</page-range><publisher-loc><![CDATA[Rome ]]></publisher-loc>
</nlm-citation>
</ref>
<ref id="B6">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Geiger]]></surname>
<given-names><![CDATA[A.]]></given-names>
</name>
<name>
<surname><![CDATA[Lenz]]></surname>
<given-names><![CDATA[P.]]></given-names>
</name>
<name>
<surname><![CDATA[Urtasun]]></surname>
<given-names><![CDATA[R.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Are we ready for autonomous driving? The KITTI vision benchmark suite]]></article-title>
<source><![CDATA[IEEE Conference on Computer Vision and Pattern Recognition]]></source>
<year>2012</year>
<page-range>3354-3361</page-range></nlm-citation>
</ref>
<ref id="B7">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Hirschmuller]]></surname>
<given-names><![CDATA[H.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Accurate and efficient stereo processing by semi-global matching and mutual information]]></article-title>
<source><![CDATA[IEEE Conference on Computer Vision and Pattern Recognition]]></source>
<year>2005</year>
<page-range>807-814</page-range></nlm-citation>
</ref>
<ref id="B8">
<nlm-citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Neilson]]></surname>
<given-names><![CDATA[D.]]></given-names>
</name>
<name>
<surname><![CDATA[Yang]]></surname>
<given-names><![CDATA[Y.]]></given-names>
</name>
</person-group>
<source><![CDATA[Evaluation of constructible match cost measures for stereo correspondence using cluster ranking]]></source>
<year>2008</year>
<page-range>1-8</page-range><publisher-name><![CDATA[IEEE Computer Society]]></publisher-name>
</nlm-citation>
</ref>
<ref id="B9">
<nlm-citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Scharstein]]></surname>
<given-names><![CDATA[D.]]></given-names>
</name>
<name>
<surname><![CDATA[Szeliski]]></surname>
<given-names><![CDATA[R.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[A taxonomy and evaluation of dense two-frame stereo correspondence algorithms]]></article-title>
<source><![CDATA[International Journal of Computer Vision]]></source>
<year>2002</year>
<volume>47</volume>
<page-range>7-42</page-range></nlm-citation>
</ref>
<ref id="B10">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Scharstein]]></surname>
<given-names><![CDATA[D.]]></given-names>
</name>
<name>
<surname><![CDATA[Szeliski]]></surname>
<given-names><![CDATA[R.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[High­accuracy stereo depth maps using structured light]]></article-title>
<collab>IEEE</collab>
<source><![CDATA[Computer Vision and Pattern Recognition, 2003. Proceedings. 2003 IEEE Computer Society Conference on]]></source>
<year>2003</year>
<month>/0</month>
<day>6/</day>
<page-range>195-202</page-range></nlm-citation>
</ref>
<ref id="B11">
<nlm-citation citation-type="">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Scharstein]]></surname>
<given-names><![CDATA[D.]]></given-names>
</name>
<name>
<surname><![CDATA[Szeliski]]></surname>
<given-names><![CDATA[R.]]></given-names>
</name>
</person-group>
<source><![CDATA[Middlebury stereo evaluation - version 2]]></source>
<year>2013</year>
</nlm-citation>
</ref>
<ref id="B12">
<nlm-citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Trucco]]></surname>
<given-names><![CDATA[E.]]></given-names>
</name>
<name>
<surname><![CDATA[Ruggeri]]></surname>
<given-names><![CDATA[A.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Towards a multi-site International public dataset for the validation of retinal image analysis software]]></article-title>
<source><![CDATA[International Conference of the IEEE Engineering in Medicine and Biology Society]]></source>
<year>2013</year>
<page-range>7152-7155</page-range><publisher-loc><![CDATA[Osaka ]]></publisher-loc>
<publisher-name><![CDATA[IEEE]]></publisher-name>
</nlm-citation>
</ref>
<ref id="B13">
<nlm-citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Van der Mark]]></surname>
<given-names><![CDATA[W.]]></given-names>
</name>
<name>
<surname><![CDATA[Gavrila]]></surname>
<given-names><![CDATA[D.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Real-time dense stereo for intelligent vehicles]]></article-title>
<source><![CDATA[IEEE Transactions on Intelligent Transportation Systems]]></source>
<year>2006</year>
<volume>7</volume>
<numero>1</numero>
<issue>1</issue>
<page-range>38-50</page-range></nlm-citation>
</ref>
<ref id="B14">
<nlm-citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Van Veldhuizen]]></surname>
<given-names><![CDATA[D.]]></given-names>
</name>
<name>
<surname><![CDATA[Zydallis]]></surname>
<given-names><![CDATA[J.]]></given-names>
</name>
<name>
<surname><![CDATA[Lamont]]></surname>
<given-names><![CDATA[G.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Considerations in engineering parallel multiobjective evolutionary algorithms]]></article-title>
<source><![CDATA[IEEE Transactions on Evolutionary Computation]]></source>
<year>2003</year>
<volume>7</volume>
<numero>2</numero>
<issue>2</issue>
<page-range>144-173</page-range></nlm-citation>
</ref>
<ref id="B15">
<nlm-citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname><![CDATA[Wang]]></surname>
<given-names><![CDATA[Z.]]></given-names>
</name>
<name>
<surname><![CDATA[Bovik]]></surname>
<given-names><![CDATA[A.]]></given-names>
</name>
<name>
<surname><![CDATA[Sheikh]]></surname>
<given-names><![CDATA[H.]]></given-names>
</name>
<name>
<surname><![CDATA[Simoncelli]]></surname>
<given-names><![CDATA[E.]]></given-names>
</name>
</person-group>
<article-title xml:lang="en"><![CDATA[Image quality assessment: from error visibility to structural similarity]]></article-title>
<source><![CDATA[IEEE Transactions on Image Processing]]></source>
<year>2004</year>
<volume>13</volume>
<numero>4</numero>
<issue>4</issue>
<page-range>600­-612</page-range></nlm-citation>
</ref>
</ref-list>
</back>
</article>
