SearcharxivSearch

arXiv subjects

Antonis Sakellarios

Publications and source records attributed to Antonis Sakellarios.

3 recordsLinked to original sources

OliveGemma: A 3 Billion Visual Language Model for Recognising the Mediterranean & European Diet

Image based dietary assessment offers a scalable alternative to self reported food diaries, yet fine-grained food recognition remains challenging due to high intra-class variability and visually similar dishes. This study presents OliveGemma, a vision language model for recognising and reasoning about Mediterranean and European cuisine. Built on the open-weight PaliGemma-2-3B architecture, OliveGemma is fine-tuned with LoRA on a unified corpus of 17,340 images from three European research project datasets (MedGR, ODIN, and VIPPSTAR), reconciled into a vocabulary of 216 composed dish categories and paired with 102,642 instruction style question-answer items covering dish recognition, likely and visible ingredients, class boundary discrimination, visual evidence and overall visual food understanding. Under a 3-fold cross-validation scheme, OliveGemma achieves a top-1 accuracy of 92.96% +/- 0.91%, exceeding the strongest CNN baseline (DenseNet-121) by 7.31% and outperforming zero-shot frontier models with exact instructions and bounded classes including Gemini Flash 3 and 3.5, GPT-5.4 Mini, and Claude Haiku 4.6 by 8%, 46%, and 64% respectively. Furthermore, OliveGemma demonstrates competitive performance on Top-3 and Top-5 accuracy, being second best across CNNs and frontier models, surpassed only by DenseNet-121. In addition, OliveGemma achieves 90.79% +/- 1.3% Exact-Set on the likely ingredients of the food categories. These results demonstrate that PEFT adaptation of a small VLM can surpass substantially larger proprietary models on specialised food recognition. The model is publicly available at https://huggingface.co/JamesZar/OliveGemma-3B and the experiments and results can be found at https://github.com/tsiokris/OliveGemma.

cs.CV

3-D Registration on Carotid Artery imaging data: MRI for different timesteps

A common problem which is faced by the researchers when dealing with arterial carotid imaging data is the registration of the geometrical structures between different imaging modalities or different timesteps. The use of the "Patient Position" DICOM field is not adequate to achieve accurate results due to the fact that the carotid artery is a relatively small structure and even imperceptible changes in patient position and/or direction make it difficult. While there is a wide range of simple/advanced registration techniques in the literature, there is a considerable number of studies which address the geometrical structure of the carotid artery without using any registration technique. On the other hand the existence of various registration techniques prohibits an objective comparison of the results using different registration techniques. In this paper we present a method for estimating the statistical significance that the choice of the registration technique has on the carotid geometry. One-Way Analysis of Variance(ANOVA) showed that the p-values were <0.0001 for the distances of the lumen from the centerline for both right and left carotids of the patient case that was studied.

physics.med-ph

3D Reconstruction of Coronary Arteries and Atherosclerotic Plaques based on Computed Tomography Angiography Images

The purpose of this study is to present a new semi-automated methodology for three-dimensional (3D) reconstruction of coronary arteries and their plaque morphology using Computed Tomography Angiography (CTA) images. The methodology is summarized in seven stages: pre-processing of the acquired CTA images, extraction of the vessel tree centerline, estimation of a weight function for lumen, outer wall and calcified plaque, lumen segmentation, outer wall segmentation, plaque detection, and finally 3D surfaces construction. The methodology was evaluated using both expert manual annotations and estimations of a recently presented Intravascular Ultrasound (IVUS) reconstruction method. As far as the manual annotation validation process is concerned, the mean value of the comparison metrics for the 3D segmentation were 0.749 and 1.746 for the Dice coefficient and Hausdorff distance, respectively. On the other hand, the correlation coefficients for the degree of stenosis 1, the degree of stenosis 2, the plaque burden, the minimal lumen area and the minimal lumen diameter, when comparing the derived from the proposed methodology 3D models with the IVUS reconstructed models, were 0.79, 0.77, 0.75, 0.85, 0.81, respectively. The proposed methodology is an innovative approach for reconstruction of coronary arteries, since it provides 3D models of the lumen, the outer wall and the CP plaques, using the minimal user interaction. Its first implementation demonstrated that it provides an accurate reconstruction of coronary arteries and thus, it may have a wide clinical applicability

eess.IV