
Una tabla de área sumada es una estructura de datos y un algoritmo para generar de forma rápida y eficiente la suma de valores en un subconjunto rectangular de una cuadrícula. En el ámbito del procesamiento de imágenes , también se la conoce como imagen integral . Fue introducida en los gráficos por computadora en 1984 por Frank Crow para su uso con mipmaps . En visión por computadora, fue popularizada por Lewis [ 1 ] y posteriormente recibió el nombre de "imagen integral", utilizándose prominentemente dentro del marco de detección de objetos de Viola-Jones en 2001. Históricamente, este principio es muy conocido en el estudio de funciones de distribución de probabilidad multidimensionales, concretamente en el cálculo de probabilidades 2D (o ND) (área bajo la distribución de probabilidad) a partir de las respectivas funciones de distribución acumulativa . [ 2 ]
El algoritmo
Como su nombre indica, el valor en cualquier punto ( x , y ) de la tabla de área sumada es la suma de todos los píxeles situados por encima y a la izquierda de ( x , y ), inclusive: [ 3 ] [ 4 ] dóndees el valor del píxel en ( x , y ).
La tabla de áreas sumadas se puede calcular de manera eficiente en una sola pasada sobre la imagen, ya que el valor en la tabla de áreas sumadas en ( x , y ) es simplemente: [ 5 ] (Tenga en cuenta que la matriz resultante se calcula desde la esquina superior izquierda).

Once the summed-area table has been computed, evaluating the sum of intensities over any rectangular area requires exactly four array references regardless of the area size. That is, the notation in the figure at right, having A = (x0, y0), B = (x1, y0), C = (x0, y1) and D = (x1, y1), the sum of i(x,y) over the rectangle spanned by A, B, C, and D is:
Extensions
This method is naturally extended to continuous domains.[2]
The method can be also extended to high-dimensional images.[6] If the corners of the rectangle are with in , then the sum of image values contained in the rectangle are computed with the formula where is the integral image at and the image dimension. The notation correspond in the example to , , , and . In neuroimaging, for example, the images have dimension or , when using voxels or voxels with a time-stamp.
This method has been extended to high-order integral image as in the work of Phan et al.[7] who provided two, three, or four integral images for quickly and efficiently calculating the standard deviation (variance), skewness, and kurtosis of local block in the image. This is detailed below:
To compute variance or standard deviation of a block, we need two integral images: The variance is given by: Let and denote the summations of block of and , respectively. and are computed quickly by integral image. Now, we manipulate the variance equation as: Where and .
Similar to the estimation of the mean () and variance (), which requires the integral images of the first and second power of the image respectively (i.e. ); manipulations similar to the ones mentioned above can be made to the third and fourth powers of the images (i.e. .) for obtaining the skewness and kurtosis.[7] But one important implementation detail that must be kept in mind for the above methods, as mentioned by F Shafait et al.[8] is that of integer overflow occurring for the higher order integral images in case 32-bit integers are used.
Consideraciones para la implementación
Es posible que el tipo de datos para las sumas deba ser diferente y de mayor tamaño que el tipo de datos utilizado para los valores originales, para poder acomodar la suma máxima esperada sin desbordamiento . Para datos de punto flotante, el error se puede reducir mediante la suma compensada .
Véase también
Referencias
- ↑ Lewis, JP (1995). Fast template matching . Proc. Vision Interface . pp. 120– 123.
- 1 2 Finkelstein, Amir; neeratsharma (2010). "Integrales dobles mediante la suma de valores de la función de distribución acumulativa" . Proyecto de demostración de Wolfram .
- ↑ Crow, Franklin (1984). "Tablas de área sumada para mapeo de texturas" . SIGGRAPH '84: Actas de la 11.ª conferencia anual sobre gráficos por computadora y técnicas interactivas . págs. 207–212 . doi : 10.1145/800031.808600 .
- ↑ Viola, Paul; Jones, Michael (2002). "Detección robusta de objetos en tiempo real" (PDF) . Revista internacional de visión por computadora .
- ↑ BADGERATI (2010-09-03). "Visión por computadora: la imagen integral" . computersciencesource.wordpress.com . Recuperado el 2017-02-13 .
- ↑ Tapia, Ernesto (enero de 2011). "Una nota sobre el cálculo de imágenes integrales de alta dimensión". Pattern Recognition Letters . 32 (2): 197– 201. Bibcode : 2011PaReL..32..197T . doi : 10.1016/j.patrec.2010.10.007 .
- 1 2 Phan, Thien; Sohoni, Sohum; Larson, Eric C.; Chandler, Damon M. (22 de abril de 2012). "Aceleración de la evaluación de la calidad de imagen basada en el análisis del rendimiento". Simposio IEEE del Suroeste de 2012 sobre Análisis e Interpretación de Imágenes (PDF) . págs. 81–84 . CiteSeerX 10.1.1.666.4791 . doi : 10.1109/SSIAI.2012.6202458 . hdl : 11244/25701 . ISBN 978-1-4673-1830-3. S2CID 12472935 .
- ↑ Shafait, Faisal; Keysers, Daniel; M. Breuel, Thomas (enero de 2008). Yanikoglu, Berrin A.; Berkner, Kathrin (eds.). "Implementación eficiente de técnicas de umbralización adaptativa local utilizando imágenes integrales" (PDF) . Electronic Imaging . Document Recognition and Retrieval XV. 6815 : 681510–681510–6. Bibcode : 2008SPIE.6815E..10S . CiteSeerX 10.1.1.109.2748 . doi : 10.1117/12.767755 . S2CID 9284084 .
Enlaces externos
- Implementación de tablas sumadas en la detección de objetos
Vídeos de conferencias
- Introducción a la teoría que sustenta el algoritmo de imagen integral.
- Una demostración de una versión continua del algoritmo de imagen integral, del Proyecto de Demostraciones de Wolfram.
- Geometría digital
- Estructuras de datos de gráficos por computadora