Professional Documents
Culture Documents
1–12
Abstract
Direct volume rendering (DVR) is a technique that emphasizes structures of interest (SOIs) within an image volume visually,
while simultaneously depicting adjacent regional information, e.g., the spatial location of a structure concerning its neighbors.
In DVR, transfer function (TF) plays a key role by enabling accurate identification of SOIs interactively as well as ensuring ap-
propriate visibility of them. TF generation typically involves non-intuitive trial-and-error optimization of rendering parameters
(opacity and color), which is time-consuming and inefficient. Attempts at mitigating this manual process have led to approaches
that make use of a knowledge database consisting of pre-designed TFs by domain experts. In these approaches, a user navi-
gates the knowledge database to find the most suitable pre-designed TF for their input volume to visualize the SOIs. Although
these approaches potentially reduce the workload to generate the TFs, they, however, require manual TF navigation of the
knowledge database, as well as the likely fine tuning of the selected TF to suit the input. In this work, we propose a TF design
approach where we introduce a new content-based retrieval (CBR) to automatically navigate the knowledge database. Instead
of pre-designed TFs, our knowledge database contains image volumes with SOI labels. Given an input image volume, our CBR
approach retrieves relevant image volumes (with SOI labels) from the knowledge database; the retrieved labels are then used
to generate and optimize TFs of the input. This approach does not need any manual TF navigation and fine tuning. For our
CBR approach, we introduce a novel volumetric image feature which includes both a local primitive intensity profile along
the SOIs and regional spatial semantics available from the co-planar images to the profile. For the regional spatial semantics,
we adopt a convolutional neural network (CNN) to obtain high-level image feature representations. For the intensity profile,
we extend the dynamic time warping (DTW) technique to address subtle alignment differences between similar profiles (SOIs).
Finally, we propose a two-stage CBR scheme to enable the use of these two different feature representations in a complemen-
tary manner, thereby improving SOI retrieval performance. We demonstrate the capabilities of our approach with comparison
to a conventional CBR approach in visualization, where an intensity profile matching algorithm is used, and also with poten-
tial use-cases in medical image volume visualization where DVR plays an indispensable role for different clinical usages. (see
https://www.acm.org/publications/class-2012)
CCS Concepts
• Computing methodologies → Volumetric models; Image-based rendering;
database of SOIs. We note that our CBR-TF approach can be ap- contrast, and peeling. They used a 2D graph-cut algorithm to iden-
plied to any knowledge database of image volumes that includes tify SOIs. Yuan et al. [YZNC05] used 3D graph-cut algorithm for
SOI labels; with our automated process to build the knowledge the SOI identification. Users, however, often has difficulty in pre-
database, we can extend the current database when new labeled cisely inferring SOIs from their interaction, especially, when the
volumes are available. SOIs are fuzzy, semi-transparent, and multi-layered in the initial
visualization. As such, these image-centric approaches may require
manual and repetitive user interactions for desired visualizations.
2. Related Works
2.1. 2.1 Multi-dimensional TF Designs 2.3. Automated Parameter Optimization for TF Designs
TF design has been advanced from traditional intensity based 1D Some investigators [HHKP96,CM10,JKE∗ 13,JKK∗ 16] focused on
TFs toward multi-dimensional TFs to facilitate the identification using parameter optimization algorithms to lessen the need for ex-
of SOIs. A pioneer work by Kindlmann et al. [KKH01] proposed tensive user TF interaction. They assume if SOIs in a volume were
2D TFs using intensity with its first or second order derivatives. known, the initial TF parameters could be automatically adjusted.
Projecting these additional derived features as secondary dimen- These SOI based TF optimizations ensure that the visibility of the
sions on the TFs widget improved SOIs exploration; the new di- SOIs is maintained by reducing the opacity parameters of other
mension typically acted as additional ‘indicators’ to guide users structures / voxels that are occluding the SOIs. They generally re-
during TF design to help identify gradient information along the lied on ‘greedy implementation’ that searches for locally optimal
intensity (representing SOI) and visually emphasize the boundaries solutions, due to the unavailability of exhaustive search of huge TF
of SOIs. Similarly, Correa et al. [CM08, CM09] identified SOIs ac- parameters space. Such greedy solutions are sensitive to the initial
cording to the local size [CM08] or spatial relationship with its parameters, which are often not properly defined by the users. A
neighbors [CM09]. Some works used more than two dimensions suboptimal initialization may result in the undesirable visualization
(features), for example, the use of 20 local texture features by Ca- of a local minimum.
ban et al. [CR08], to enable greater differentiation among SOIs.
As the number of TF dimensions increases, there is a greater need
for manual optimizations in a complex multi-dimensional space 2.4. Knowledge Databases and CBR in DVR Visualization
[PLB∗ 01, LKG∗ 16]. Marks et al. [MAB∗ 97] firstly generated a knowledge database with
a collection of pre-designed TFs by domain experts for TF design.
This database was comprised of pairs of a DVR and its correspond-
2.2. Image-centric TF Designs
ing TF. It was then used manually by the users to search through
Image-centric approaches make TFs design intuitive by allowing the TF cases that were most suitable to the input volume, with aid
users to identify and optimize SOIs directly on an initial DVR of paired DVRs. In another work, Guo et al. [GLY14] structured
visualization through manual gestures. Ropinski et al. [RPSH08] their knowledge database into a 2D search representation of TF
enabled intuitive selection of SOIs by users drawing one or more cases via multi-dimensional reduction (MDS) for effective man-
strokes directly onto the DVR visualization near the silhouette of ual knowledge database navigation. In this search space, similar
the SOIs. Then based on the stokes(s), an intensity histogram anal- TF cases were grouped into clusters via the density field of their
ysis was done to identify the intended SOIs in the intensity 1D TF MDS. Although these works demonstrated the potential of how to
widget. Guo et al. [GMY11] manipulated the appearance of the in- use knowledge databases in TF design, they only partially reduced
tended SOIs through high-level ‘painting’ metaphors, e.g., eraser, the workload to generate TFs as they required manual exhaustive
navigations of the TF cases from the knowledge database, as well co-planar to the ray according to the primary axes; (ii) a profile
as the potential need for additional adjustment of the selected TFs of intensity values along the ray; (iii) the corresponding profile of
to be best mapped to the input volume. the SOI labels; (iv) location coordinates of entry- and exit-point of
the ray. We uniformly extracted multiple rays along a dimension
Apart from TF design, Kohlmann et al. [KBKG09] investigated
(axis), with the consistent number of rays from all the three axes.
the use of a knowledge database to provide complementary visual-
Empty backgrounds of an image volume were excluded during the
ization in 2D cross-sectional views and 3D DVR. They introduced
ray extraction. An image volume consisted of 192 rays in total and
CBR concept to automate search through the knowledge database,
64 rays from each of the three axes. The extraction interval, i.e.,
where the intensity profiles were used to represent the ‘content’,
(the number of rays) was experimentally determined to reduce the
such as a SOI along the profile. This is the only work that demon-
likelihood of extracting highly similar rays as well as lowering the
strated the use of the CBR based knowledge database for DVR vi-
computational complexity of our CBR process.
sualization.
3.1. Overview For a TIQ from a user-selected ray q: a pair of images qi and an
intensity profile q p , we retrieved the most similar counterpart (c)
An overview of our CBR-TF approach is shown in Figure 2 exem- from the knowledge database by calculating the similarity distances
plified with an input image volume of the human upper-abdomen. between the triplet pairs (q and c) in a two-stage manner. We first
We constructed, as an offline process, a knowledge database using matched qi with ci . Among the top N retrieval results based on the
a set of labeled image volumes (Section 3.2). In the quarter-view image matching, as the second stage, we re-ranked them according
visualization of the input image volume with 2D cross-sectional to the similarity of intensity profiles (q p and c p ).
views and a 3D DVR, a user drew a line (a ray) to visualize certain
SOIs in any of the 2D views. With the user-selected ray, we gen- CNN based Image Matching. We used a well-established CNN,
erated a TIQ comprising a pair of co-planar images to the ray and AlexNet, which was pre-trained using 1.2 million general images
its intensity profile. The TIQ was then used to query the knowledge from the ImageNet database [KSH17]. The architecture of AlexNet
database (Section 3.3). We used the SOI labels provided from the is shown in Figure 4. It consists of total 7 trainable layers with 60
retrieved results to generate an intensity based 1D TF for the SOIs million parameters; the first five are convolutional layers, and the
(Section 3.4), which can be further optimized with the input image remaining two are fully connected layers. The output of the last
volume to make sure that particular SOIs prioritize the visibility in fully connected layer is a vector of 4096 dimensions. We used this
the final 3D DVR (Section 3.5). as a feature vector for the image matching.
from the image pair from the knowledge database, and F b is bth other shapes depending on specific visualization and data needs
CNN image feature. [CKLG98]. We finally assigned each tent apex to unique colors
using Color Brewer [HB03] and all tent end points to a single
DTW based Intensity Profile Matching. We measured the in-
color (black) to improve visual color depth perception among SOIs
tensity profile similarity distance between q p and c p by calculating
[JKK∗ 18]. However, there is no limitation to apply different color
the cost of an optimal warping path using DTW. A Lq p × Lc p dis-
schemes, e.g., the tent end points and apex can have the same color.
tance matrix was created which presets all pairwise distances be-
tween each element in q p and c p , where Lq p is the length of q p and
Lc p is to c p . A warping path P was then defined to pass through the 3.5. Visibility based TF Parameter Optimization
contiguous low-cost parts of distance matrix to create a sequence of
Visibility based TF parameter optimization [CM10, JKE∗ 13] was
points =(p1 ,p2 ,. . . ,pK ) as shown in Figure 5. The optimal warping
an iterative process: the initial TF and target visibilities of SOIs
path P∗ was calculated as the path with the minimum cost:
v were entered as the inputs. Each iteration calculated an intermedi-
uK
u ate new TF and the resulting visibility of SOIs. The visibility calcu-
Intensity Pro f ile Similarity Distance(q p , c p ) = min(t ∑ pk /K) lation is explained later in this section. The iteration continues until
k=1 it converges to a minimum of an energy function E of visibility
(2) tolerance:
where K is the length of a sequence, and there are three constraints: M
(i) boundary condition: P starts and ends in diagonally opposite minθ E(θ) = ∑ (VT m −VRm )2 (5)
corners of distance matrix; (ii) continuity: each step in P proceeds m=1
from, and moves to, an adjacent cell; and (iii) monotonicity: P does where M is the total number of SOIs, and θ is a set of opacity
not take a step backwards in spatial locations. parameters of tent apexes forming the initial TF. The visibility tol-
erance is used to control how much the SOIs are visible in the op-
timized visualizations and it can be calculated via the sum of the
squared differences between target visibility VT and resultant visi-
bility VR for individual SOIs. The value of V is the visibility propor-
tion among all SOIs (where V = 1.0 means the corresponding SOI
exclusively priorities visibility); the assignment of VT was either
automatically uniformly set to all SOIs or user-defined to enhance
the visibility of particular SOIs.
We used a downhill simplex method [NM65] to solve the op-
timization problem. This method is effective in a variety of prac-
tical non-linear optimization problems with multiple local min-
ima [BIJ96, LRWW98]. However, there is no limitation to apply
Figure 5: Warping path between an intensity profile pair example
other gradient-based approaches, e.g., gradient-descent [Rud16].
(q p and c p ) in DTW.
The visibility F(p) for a voxel is the front-to-back opacity com-
position of all the voxels of a volume starting from a view-point h
to p according to [JKE∗ 13]:
3.4. SOI based TF Generation Rh
We generated a conventional intensity based 1D TF for SOIs of F(p) = e− p A(g)dg
(6)
a TIQ using SOI labels retrieved from the knowledge database.
We formulated a separate TF component Tm for each SOI Sm , m = A(p) = A(p − ∆p) + (1.0 − A(p − ∆p))O(p) (7)
1, . . . , M where M was the number of SOI labels. Each component where A(p) is the composited opacity of the voxel coordinate p,
was a tent-shape in TF parameter space (intensity-opacity), and was O(p) is its opacity, which is defined by a TF, and ∆p is the size
specified as follows: of the sampling step. The visibilities of all voxels, weighted as the
T j = {(Il,m , 0), (I m , σm ), (Ih,m , 0)} (3) product of their opacity, are then added to determine the visibility
of SOI (intensity range) given by:
where Il,m is the lowest intensity value in Sm , Ih,m is the highest
intensity value in Sm , I m is the average intensity in Sm , and σm is the VSOI = ∑ O(p)F(p) (8)
p∈SOI
initial opacity of a tent peak set to 0.3 (where σm =1.0 indicates full
opacity). The initial TF was given by the union of the tent shapes.
∪M
m=1 Tm (4) 3.6. Evaluation Procedure
We used the tent shapes since they revealed the iso-surface of their 3D-IRCADb-01 dataset [SHA∗ 10] was used for the evaluation. It
corresponding SOIs, allowing a TF to visualize multiple SOIs in consists of 20 CT image volumes of the upper-abdomen (10 males
a semi-transparent manner. This is a desirable property in the vi- and 10 females). Since the dataset was mainly designed for bench-
sualization of complex SOIs. Nevertheless, our TF generation is marking liver segmentation algorithms, there were limited manual
not limited to this particular shape and can be re-parameterized to labels available. We selected the six SOIs that commonly appear at
Table 1: Recall metrics for retrieval among six different SOIs for the ED baseline [KBKG09] and our two-stage method with the two
individual compartments.
SOIs Artery Bone Kidney Liver Lung Spleen All
ED (baseline) 0.453 0.491 0.344 0.649 0.552 0.363 0.518
Our Two-stage 0.612 0.692 0.648 0.778 0.829 0.644 0.719
DTW (individual compartment) 0.530 0.645 0.542 0.760 0.808 0.569 0.672
AlexNet (individual compartment) 0.549 0.681 0.615 0.745 0.795 0.454 0.679
Table 2: Recall metrics for retrieval among six different SOIs for our two-stage method with three different types of pre-trained CNNs
[KSH17, SLJ∗ 15, HZRS16] in a comparison to the ED baseline [KBKG09].
SOIs Artery Bone Kidney Liver Lung Spleen All
With Pre-trained AlexNet 0.612 0.692 0.648 0.778 0.829 0.644 0.719
With Pre-trained GoogleNet 0.591 0.681 0.616 0.770 0.836 0.608 0.707
With Pre-trained ResNet 0.543 0.624 0.547 0.682 0.668 0.487 0.620
ED (baseline) 0.453 0.491 0.344 0.649 0.552 0.363 0.518
Table 3: Recall metrics for retrieval among six different SOIs for the pre-trained AlexNet [KSH17] and its fine-tuning counterpart [SRG∗ 16].
SOIs Artery Bone Kidney Liver Lung Spleen All
With Pre-trained AlexNet 0.612 0.692 0.648 0.778 0.829 0.644 0.719
With Fine-tuned AlexNet 0.609 0.700 0.617 0.765 0.816 0.688 0.714
the abdomen, which consist of the artery (1058 occurrences among 0.0 to 1.0 where a value of 1.0 indicates that all instances of a SOI
all the images slices), bone (1824), kidney (506), liver (1645), lung type are identified by a retrieval approach. In our evaluation, recall
(1101), and spleen (306). was considered the most representative in evaluating retrieval per-
formances; if false-positive occurs, the user can manually eliminate
We believe that our chosen SOIs have sufficient complexity to
it in the subsequent TF design, but such manual optimization can be
measure the retrieval performances since the multiple SOIs, e.g.,
much more challenging in missing SOIs (i.e., false-negative). We,
kidney, liver, and spleen, are similar in image features such as in-
therefore, discussed the recall results and included the precision re-
tensity and shape, and accurate retrieval among the SOIs is chal-
sults in the Appendix (see Table A1). We conducted 10-fold cross
lenging. In total, our knowledge database was made of 3840 ray el-
validation for the evaluation, where for each fold, 18 CT image
ements derived from the 20 image volumes, with each having 192
volumes were acted as the knowledge database, and the other 2 CT
rays.
image volumes were used as the queries. We rotated this process 9
We compared our CBR-TF approach with the approach pub- times to cover all 20 CT image volumes.
lished in the visualization community by Kohlmann et al.
For all these experiments, we used a consumer PC with an Intel
[KBKG09]. This baseline approach relied on intensity profile
i7 CPU and an Nvidia GTX 980 Ti GPU, running Windows 7 (x64),
matching using ED for measuring the similarity between the in-
and implemented our CBR-TF approach using volume rendering
put and the cases from the knowledge database. For this compari-
engine (Voreen) [MSRMH09].
son, only the second stage of our CBR-TF approach, which is the
intensity profile matching using DTW, was applied in the knowl-
edge database retrieval. We evaluated the first stage of our CBR- 4. Results
TF approach, which is image matching using pre-trained AlexNet
4.1. CBR Analysis
[KSH17], by comparing to two established pre-trained CNNs of
GoogleNet [SLJ∗ 15] and ResNet [HZRS16] by measuring the re- We show the recall metrics for retrieval among six different SOIs
trieval accuracy. The pre-trained AlexNet was also compared with between the ED baseline [KBKG09] and our two-stage method in
the fine-tuned counterpart using a fine-tuning method [SRG∗ 16]. Table 1. We also present results from using individual compart-
We note that there are no comparable works that used image re- ments of our method: DTW and AlexNet to analyze their individ-
trieval in the TF designs as well as DVR visualization. ual performances and their contributions to the combined results.
Compared to the ED baseline method [KBKG09], our two-stage
The most common retrieval evaluation metrics of recall and pre-
method outperformed in every SOI retrieval by 0.201 on average
cision were used. Recall refers to the number of times the SOI was
and kidney retrieval by 0.304 as the largest improvement.
correctly identified out of the number of times the SOI occurred.
Precision refers to the number of times the SOI was correctly iden- Our two-stage method outperformed both the individual com-
tified out of the number of times the SOI was identified by a re- partments in all the SOIs. Within the two compartments, the
trieval approach. The values of both recall and precision range from AlexNet method was better than the DTW method for artery, bone,
and kidney, and the other three SOIs were better retrieved by the SOI retrieval with comparison to the ED baseline method using
DTW method. The DTW method improved upon the ED baseline only the intensity profile [KBKG09]. The results show that the ED
counterpart across all the SOIs. baseline method incorrectly labeled the arteries and kidney along
the TIQ. In contrast, all the SOIs along the TIQ were correctly re-
Our image retrieval compartment method is not limited to the
trieved when our two-stage method was applied.
particular CNN backbone of AlexNet [KSH17] and can be ap-
plied to other CNNs. In Table 2, we showed the retrieval accu-
racies of pre-trained CNNs including AlexNet [KSH17] and two
other CNNs with deeper layers, GoogleNet [SLJ∗ 15] and ResNet
[HZRS16]. Our results show that the pre-trained AlexNet, our de-
fault, had the highest average recall across all the SOIs except
for lung where the pre-trained GoogleNet resulted in marginal im-
provement with 0.007. Regardless of the type of CNNs, they out-
performed the ED baseline counterpart [KBKG09] across every
SOI. The comparison result in Table 3 shows that the fine-tuned
AlexNet was not always able to achieve higher recall accuracy
compared to the pre-trained AlexNet; only spleen produced some
meaningful improvement from the fine-tuned AlexNet by 0.044.
This means that the fine-tuning method used [SRG∗ 16] might not
be optimal for our application.
where our DTW compartment method correctly identified all three can see that these images were similar to the image query pair, with
SOIs while only one SOI was achieved in the ED baseline method. greater similarity (i.e., consistency in the same SOIs) in the coro-
nal views (the second column). With the inclusion of the intensity
profile matching results to the top-4 ranked image retrieval results,
we see that the best intensity profile matching is with the 4th best
image retrieval, among the only two results that were able to match
the spleen (with the other result being the 3rd best image retrieval).
Thus, in the combined result, the 4th image retrieval result will be-
come the 1st ranked final retrieval result and has the overall best
matching.
a retrieval method must be able to accommodate these variations. A fore, can avoid additional manual TF fine tuning that may be re-
representative example is shown in the Appendix (see Figure A1). quired for the cases where the TFs from knowledge databases are
not optimal for the particular SOIs of the input image volume. A
For image retrieval, we used a pre-trained AlexNet [KSH17] due
TF parameter optimization application example is shown in Figure
to its highest average recall of 0.719 when compared to the two pre-
11, where the visibility of the bones was prioritized over the lungs
trained deeper CNNs, GoogleNet [SLJ∗ 15] and ResNet [HZRS16]
by setting VT of the bones to 0.7 and VT of the lungs to the rest
(see Table 2). This could be because the deeper networks learned
(0.3).
more data-specific features that are relevant to general images, i.e.,
less generalizable to medical images. In our CBR-TF approach, TFs were comprised of individual
tent-shape components to represent each SOI. In the cases when
Our two-stage method, i.e., rank-based sequential combination
the false-positive from our SOI retrieval occurs, we note that the
of the two individual compartments, was experimentally designed
user needs to manually eliminate the TF component of the false-
based on the observation that the two individual compartments
positive SOI. Such manual TF optimization can be more challeng-
could be complimentary, but this advantage was not properly en-
ing in missing structures (i.e., false-negative). However, our CBR-
coded when two individual compartments were used at the same
TF approach is still effective because the user does not need to start
time (i.e., a single-stage retrieval). In the first stage of image re-
the TF design from scratch, and it is easier to add new SOIs to the
trieval, we used the top 40 retrieval results (i.e., N=40), with the
TF that already contains other SOIs.
second stage of intensity profile retrieval only executed within the
N retrievals. This was empirically derived with variations in N hav- In terms of the generation of knowledge databases, existing ap-
ing a minor impact on the overall retrieval performances. Full de- proaches [KBKG09, GLY14] collect pre-designed TFs by domain
tails of this experiment were included in the Appendix (see Table experts, whereas our CBR-TF approach relies on SOI labels of raw
A2). image volumes. In terms of the perspective of experts in knowledge
database generation, generating TFs may take much effort com-
All the results presented in this work used a well-established re- pared to SOI labels.
call metric to evaluate the retrieval performances. As a complemen-
tary metric, we also included precision results in the Appendix (see
Table A1); we note that, as with recall, our two-stage method per- 5.3. Future Works
formed better in precision compared to the ED baseline [KBKG09]. The image retrieval results (Table 2) demonstrated the adaptability
of our CBR-TF approach to different CNNs. As such, our CBR-TF
5.2. TF Design approach is not constrained to a particular CNN backbone. We note
that recent advances of CNNs, e.g., for segmentation or classifica-
Our results showed that our image-centric interaction for TF de- tion tasks, demonstrated the capabilities to better describe image
sign is intuitive and effective. Our CBR-TF approach required a features to represent SOIs [LKB∗ 17, SWS17, CWZ∗ 22, GLO∗ 16].
user to simply draw a line in 2D cross-sectional views (see Fig- These backbones could be leveraged for our CBR-TF approach, but
ure 9) or a point in 3D DVRs (see Figure 10) to select SOIs. The optimal adoption should be investigated to improve SOI retreival
usability of such image-centric interaction to directly select SOIs perforamance (see Table 3). We consider this challenging but inter-
in a visualization space, not in an indirect TF widget space, was esting investigation as our future work.
investigated by the early work [RPSH08], where positive informal
feedback from clinicians who conducted interaction using a medi- Efficient retrieval computation is an important requirement for
cal CT image volume was observed. the practical application of our CBR-TF approach. In this work,
we used an exhaustive search where every case was matched with
The use-case with the quarter-view visualization (see Figure 9) a TIQ. This serial matching was sufficient for 3896 cases in our
can be effective in medical image visualization where the user, i.e., knowledge database. We believe that less than 20 seconds to gen-
clinicians, generally has prerequisite knowledge of their data; they erate TFs (and retrievals) from scratch (see Table 4) is an ac-
can readily define a line (a ray) with SOIs on the 2D cross-sectional ceptable user experience. However, for larger knowledge database
views. The fact that they are a novice user for TF design makes sizes, we will investigate adopting GPU based parallelism libraries
our CBR-TF approach more valuable. They do not need to rely on [HCM13, MC11] to speed up retrieval computational efficiency
the TF widgets that are complicated to and unfamiliar with them. by enabling individual matching in parallel. Another approach is
Instead, the SOIs can be simply defined in the 2D cross-sectional to construct a knowledge database hierarchically, e.g., using a set
views where they are usually working on in the clinical routine. cover problem [AMS06], for rapid but approximated retrieval.
With another use-case (see Figure 10) where the single point se- Our CBR-TF approach does not require any dataset-specific pa-
lection was directly done in an initial 3D DVR for the SOI identi- rameter settings for the knowledge database construction. Further-
fication, we suggested that our CBR-TF approach enables a user to more, we automated the entire construction procedure including
obtain a more detailed visualization of specific SOIs intuitively and ray extraction, representation, and storage. Different types of imag-
automatically, thereby easily inspecting information that might be ing modalities can be easily integrated for different application do-
hidden in the initial DVR. mains. We limited the scope of this work to medical imaging data,
but we suggest that our CBR-TF approach can be applied to non-
In our CBR-TF approach, we adopted an automated visibility
medical imaging datasets.
based TF parameter optimization to ensure that some of the iden-
tified SOIs prioritize visibility in resulting DVRs. The user, there- We demonstrated our CBR-TF approach with intensity based 1D
TF design because they are commonly used in a wide range of vi- [GLO∗ 16] G UO Y., L IU Y., O ERLEMANS A., L AO S., W U S., L EW
sualization applications and have fewer variables that influence vi- M. S.: Deep learning for visual understanding: A review. Neurocomput-
sualization experiments. Our CBR-TF approach, however, is not ing 187 (2016), 27–48. 10
restricted to 1D TFs and could be adapted to multi-dimensional [GLY14] G UO H., L I W., Y UAN X.: Transfer function map. In 2014
IEEE Pacific visualization symposium (2014), IEEE, pp. 262–266. 2, 3,
TFs [KKH01, CM09, CM08] for greater differentiation of SOIs. 10
Huang et al. [HM03] proposed an approach to generate 2D TFs
[GMY11] G UO H., M AO N., Y UAN X.: Wysiwyg (what you see is what
when the identification of SOIs in the 2D cross-sectional views are you get) volume visualization. IEEE Transactions on Visualization and
available. We regard this as an important new avenue of research. Computer Graphics 17, 12 (2011), 2106–2114. 2, 3
[HB03] H ARROWER M., B REWER C. A.: Colorbrewer. org: an online
tool for selecting colour schemes for maps. The Cartographic Journal
6. Conclusion 40, 1 (2003), 27–37. 5
We presented a new TF design approach by proposing an auto- [HCM13] H EIDARI H., C HALECHALE A., M OHAMMADABADI A.: Par-
allel implementation of color based image retrieval using cuda on the
mated two-stage CBR for a knowledge database, and the use of gpu. International Journal of Information Technology and Computer
its retrieved results in generating TFs. Our experimental results Science (IJITCS) 6, 1 (2013), 33–40. 10
demonstrate that our CBR-TF approach was capable of identifying [HHKP96] H E T., H ONG L., K AUFMAN A., P FISTER H.: Generation of
and visualizing SOIs from the combination of the complementary transfer functions with stochastic search techniques. In Proceedings of
image semantics with the intensity profile matching. Seventh Annual IEEE Visualization’96 (1996), IEEE, pp. 227–234. 2, 3
[HM03] H UANG R., M A K.-L.: Rgvis: Region growing based tech-
niques for volume visualization. In 11th Pacific Conference onComputer
References Graphics and Applications, 2003. Proceedings. (2003), IEEE, pp. 355–
363. 11
[AL13] A SADI N., L IN J.: Effectiveness/efficiency tradeoffs for candi-
[HZRS16] H E K., Z HANG X., R EN S., S UN J.: Deep residual learn-
date generation in multi-stage retrieval architectures. In Proceedings of
ing for image recognition. In Proceedings of the IEEE conference on
the 36th international ACM SIGIR conference on Research and develop-
computer vision and pattern recognition (2016), pp. 770–778. 6, 7, 10
ment in information retrieval (2013), pp. 997–1000. 9
[JKE∗ 13] J UNG Y., K IM J., E BERL S., F ULHAM M., F ENG D. D.:
[AMS06] A LON N., M OSHKOVITZ D., S AFRA S.: Algorithmic con- Visibility-driven pet-ct visualisation with region of interest (roi) segmen-
struction of sets for k-restrictions. ACM Transactions on Algorithms tation. The Visual Computer 29, 6 (2013), 805–815. 2, 3, 5
(TALG) 2, 2 (2006), 153–177. 10
[JKK∗ 16] J UNG Y., K IM J., K UMAR A., F ENG D. D., F ULHAM M.: Ef-
[AUI19] A HMED K. T., U MMESAFI S., I QBAL A.: Content based image ficient visibility-driven medical image visualisation via adaptive binned
retrieval using image features information fusion. Information Fusion 51 visibility histogram. Computerized Medical Imaging and Graphics 51
(2019), 76–99. 9 (2016), 40–49. 2, 3
[BIJ96] BARTON R. R., I VEY J R J. S.: Nelder-mead simplex modifi- [JKK∗ 18] J UNG Y., K IM J., K UMAR A., F ENG D. D., F ULHAM
cations for simulation optimization. Management Science 42, 7 (1996), M.: Feature of interest-based direct volume rendering using contex-
954–973. 5 tual saliency-driven ray profile analysis. In Computer Graphics Forum
[CKLG98] C ASTRO S., K ÖNIG A., L ÖFFELMANN H., G RÖLLER E.: (2018), vol. 37, Wiley Online Library, pp. 5–19. 5
Transfer function specification for the visualization of medical data. Vi- [KBKG09] KOHLMANN P., B RUCKNER S., K ANITSAR A., G ROLLER
enne University of Technology (1998). 5 M. E.: Contextual picking of volumetric structures. In 2009 IEEE Pacific
Visualization Symposium (2009), IEEE, pp. 185–192. 2, 4, 6, 7, 9, 10
[CM08] C ORREA C., M A K.-L.: Size-based transfer functions: A new
volume exploration technique. IEEE transactions on visualization and [KKC∗ 13] K UMAR A., K IM J., C AI W., F ULHAM M., F ENG D.:
computer graphics 14, 6 (2008), 1380–1387. 2, 3, 11 Content-based medical image retrieval: a survey of applications to mul-
tidimensional and multimodality data. Journal of digital imaging 26, 6
[CM09] C ORREA C., M A K.-L.: The occlusion spectrum for volume (2013), 1025–1039. 2
classification and visualization. IEEE Transactions on Visualization and
Computer Graphics 15, 6 (2009), 1465–1472. 2, 3, 11 [KKH01] K NISS J., K INDLMANN G., H ANSEN C.: Interactive volume
rendering using multi-dimensional transfer functions and direct manip-
[CM10] C ORREA C. D., M A K.-L.: Visibility histograms and visibility- ulation widgets. In Proceedings Visualization, 2001. VIS’01. (2001),
driven transfer functions. IEEE Transactions on Visualization and Com- IEEE, pp. 255–562. 2, 3, 11
puter Graphics 17, 2 (2010), 192–204. 2, 3, 5
[KR05] K EOGH E., R ATANAMAHATANA C. A.: Exact indexing of dy-
[CR08] C ABAN J. J., R HEINGANS P.: Texture-based transfer functions namic time warping. Knowledge and information systems 7, 3 (2005),
for direct volume rendering. IEEE Transactions on Visualization and 358–386. 2
Computer Graphics 14, 6 (2008), 1364–1371. 2, 3 [KSH17] K RIZHEVSKY A., S UTSKEVER I., H INTON G. E.: Imagenet
[CWZ∗ 22] C HEN X., WANG X., Z HANG K., F UNG K.-M., T HAI T. C., classification with deep convolutional neural networks. Communications
M OORE K., M ANNEL R. S., L IU H., Z HENG B., Q IU Y.: Recent ad- of the ACM 60, 6 (2017), 84–90. 2, 4, 6, 7, 10
vances and clinical applications of deep learning in medical image anal- [LBH15] L E C UN Y., B ENGIO Y., H INTON G.: Deep learning. nature
ysis. Medical Image Analysis (2022), 102444. 10 521, 7553 (2015), 436–444. 2
[DJLW08] DATTA R., J OSHI D., L I J., WANG J. Z.: Image retrieval: [LKB∗ 17] L ITJENS G., KOOI T., B EJNORDI B. E., S ETIO A. A. A.,
Ideas, influences, and trends of the new age. ACM Computing Surveys C IOMPI F., G HAFOORIAN M., VAN D ER L AAK J. A., VAN G INNEKEN
(Csur) 40, 2 (2008), 1–60. 2 B., S ÁNCHEZ C. I.: A survey on deep learning in medical image analy-
[FBKC∗ 12] F EDOROV A., B EICHEL R., K ALPATHY-C RAMER J., sis. Medical image analysis 42 (2017), 60–88. 10
F INET J., F ILLION -ROBIN J.-C., P UJOL S., BAUER C., J ENNINGS [LKG∗ 16] L JUNG P., K RÜGER J., G ROLLER E., H ADWIGER M.,
D., F ENNESSY F., S ONKA M., ET AL .: 3d slicer as an image comput- H ANSEN C. D., Y NNERMAN A.: State of the art in transfer functions for
ing platform for the quantitative imaging network. Magnetic resonance direct volume rendering. In Computer Graphics Forum (2016), vol. 35,
imaging 30, 9 (2012), 1323–1341. 2 Wiley Online Library, pp. 669–691. 1, 3