SearcharxivSearch

arXiv subjects

Gianluca Setti

Publications and source records attributed to Gianluca Setti.

At least 19 recordsLinked to original sources

Text-conditioned Segmentation for Tomato Phenotyping via Procedural Synthetic Data

Vision-based automation is an excellent candidate for reducing manual labor in greenhouse crop production and phenotyping. However, progress is constrained by the lack of annotated training data. Recent advances in vision-based foundational models have shown promising results in zero-shot generalization to novel domains, but their performance drops in complex agricultural environments. In this work, we present a sim-to-real framework for tomato plant segmentation that combines synthetic data generation with fine-tuning of a foundation model. We model a commercial cherry tomato greenhouse and use it to generate a large-scale synthetic dataset under diverse viewpoints, lighting conditions, and plant morphology. Subsequently, we fine-tune the Segment Anything Model 3 (SAM 3) on the synthetic dataset, specializing its text-conditioned segmentation behavior for greenhouse crop organs while retaining the general visual prior that makes zero-shot transfer possible. By evaluating our framework on multiple real-world greenhouse datasets, we demonstrate that combining synthetic data with SAM 3 fine-tuning significantly improves segmentation performance and model confidence. To support community benchmarking, we publicly release the procedural model, the generated synthetic dataset, and our fine-tuned SAM 3 weights.

cs.CV

Vanishing Contributions: A Unified Framework for Smooth and Iterative Model Compression

The increasing scale of Deep Neural Networks (DNNs) introduces the need for compression techniques such as pruning, quantization, and low-rank decomposition. While these methods are very effective at reducing memory, computation, and energy consumption, they may introduce severe accuracy degradation, which is often mitigated by using iterative, gradual compression. However, different compression techniques require distinct iterative approaches, and some result in unstable, discontinuous model fine-tuning. We introduce Vanishing Contributions (VCON), a unified framework for the smooth, iterative transition of DNNs into a compressed form. Rather than replacing the original network directly with its compressed version, VCON executes both in parallel during fine-tuning. The contribution of the original (uncompressed) model is progressively reduced, while that of the compressed model is gradually increased. This affine combination allows the network to slowly adapt, improving stability and mitigating accuracy degradation. We evaluate VCON on computer vision and natural language processing benchmarks, using multiple compression strategies. In most settings, our framework improves accuracy over post-shot and iterative baselines. Typical gains exceed 1%, while some configuration exhibits improvements above 15%. VCON is thus compatible with existing compression techniques and consistently improves performance across diverse tasks.

cs.LG

LiDAR for Rehabilitation: A Comprehensive Survey of Applications, AI Techniques, and Future Directions

Rehabilitation aims to help patients with limited mobility regain their physical abilities through targeted movements, exercises, stimulation, and other therapeutic methods. Recent advances in technology have introduced sensor-based systems into rehabilitation and clinical practices, enabling real-time monitoring and providing accurate feedback on movement accuracy. Among these sensors, LiDAR has demonstrated strong potential, offering key advantages over conventional techniques such as camera-based systems, which raise privacy concerns, and wearable sensors, which can be uncomfortable and prone to errors. In this work, we review the applications of LiDAR in rehabilitation, post-injury care, and hospital environments, focusing on studies published between 2019 and 2025. Studies across several areas have been explored: 3D body scanning and gait analysis with standalone LiDAR, LiDAR mounted on robotic systems for rehabilitation, real-time monitoring and environment scanning for safe navigation, and activity and position recognition. We also analyze processing techniques, particularly learning-based approaches, and support the discussion with statistical analysis, highlighting trends, gaps, and future research opportunities. To the best of our knowledge, this is the first comprehensive survey dedicated to LiDAR for rehabilitation applications, providing an overview of current methods, AI-based processing techniques, and open challenges.

cs.RO

LiDAR for Crowd Management: Applications, Benefits, and Future Directions

Light Detection and Ranging (LiDAR) technology offers significant advantages for effective crowd management. This article presents LiDAR technology and highlights its primary advantages over other monitoring technologies, including enhanced privacy, performance in various weather conditions, and precise 3D mapping. We present a general taxonomy of four key tasks in crowd management: crowd detection, counting, tracking, and behavior classification, with illustrative examples of LiDAR applications for each task. We identify challenges and open research directions, including the scarcity of dedicated datasets, sensor fusion requirements, artificial intelligence integration, and processing needs for LiDAR point clouds. This article offers actionable insights for developing crowd management solutions tailored to public safety applications.

cs.CV

A Fully Analog Implementation of Model Predictive Control with Application to Buck Converters

This paper proposes a novel approach to design analog electronic circuits that implement Model Predictive Control (MPC) policies for dynamical systems described by affine models. Effective approaches to define a reduced-complexity Explicit MPC form are combined and applied to realize an analog circuit comprising a limited set of low-latency, commercially available components. The practical feasibility and effectiveness of the proposed approach are demonstrated through its application in the design of a novel MPC-based controller for DC-DC Buck converters. We formally analyze the stability of the resulting system and conduct extensive numerical simulations to demonstrate the control system's performance in rejecting line and load disturbances.

eess.SY

Multi-Layer Confidence Scoring for Detection of Out-of-Distribution Samples, Adversarial Attacks, and In-Distribution Misclassifications

The recent explosive growth in Deep Neural Networks applications raises concerns about the black-box usage of such models, with limited trasparency and trustworthiness in high-stakes domains, which have been crystallized as regulatory requirements such as the European Union Artificial Intelligence Act. While models with embedded confidence metrics have been proposed, such approaches cannot be applied to already existing models without retraining, limiting their broad application. On the other hand, post-hoc methods, which evaluate pre-trained models, focus on solving problems related to improving the confidence in the model's predictions, and detecting Out-Of-Distribution or Adversarial Attacks samples as independent applications. To tackle the limited applicability of already existing methods, we introduce Multi-Layer Analysis for Confidence Scoring (MACS), a unified post-hoc framework that analyzes intermediate activations to produce classification-maps. From the classification-maps, we derive a score applicable for confidence estimation, detecting distributional shifts and adversarial attacks, unifying the three problems in a common framework, and achieving performances that surpass the state-of-the-art approaches in our experiments with the VGG16 and ViTb16 models with a fraction of their computational overhead.

cs.LG

Field Free Spin-Orbit Torque Controlled Synapse and Stochastic Neuron Devices for Spintronic Boltzmann Neural Networks

Spintronics offers a promising approach to energy efficient neuromorphic computing by integrating the functionalities of synapses and neurons within a single platform. A major challenge, however, is achieving field-free spin orbit torque SOT control over both synaptic and neuronal devices using an industry-adopted spintronic materials stack. In this study, we present field-free SOT spintronic synapses utilizing a CoFeB ferromagnetic thin film system, where asymmetrical device design and specifically added lateral notches in the CoFeB thin film facilitate effective domain wall DW nucleation, movement, and pinning and depinning. This method yields multiple analog, nonvolatile resistance states with enhanced linearity and symmetry, resulting in programmable and stable synaptic weights. We provide a systematic measurement approach to improve the linearity and symmetry of the synapses. Additionally, we demonstrate nanoscale magnetic tunnel junctions MTJs that function as SOT-driven stochastic neurons, exhibiting current-tunable, Boltzmann-like probabilistic switching behavior, which provides an intrinsic in-hardware Gibbs sampling capability. By integrating these synapses and neurons into a Boltzmann machine complemented by a classifier layer, we achieve recognition accuracies greater than 98 percent on the MNIST dataset and 86 percent on Fashion MNIST. This work establishes a framework for field free synaptic and neuronal devices, setting the stage for practical, materials-compatible, and all spintronic neuromorphic computing hardware.

physics.app-ph

RDD: Pareto Analysis of the Rate-Distortion-Distinguishability Trade-off

Extensive monitoring systems generate data that is usually compressed for network transmission. This compressed data might then be processed in the cloud for tasks such as anomaly detection. However, compression can potentially impair the detector's ability to distinguish between regular and irregular patterns due to information loss. Here we extend the information-theoretic framework introduced in [1] to simultaneously address the trade-off between the three features on which the effectiveness of the system depends: the effectiveness of compression, the amount of distortion it introduces, and the distinguishability between compressed normal signals and compressed anomalous signals. We leverage a Gaussian assumption to draw curves showing how moving on a Pareto surface helps administer such a trade-off better than simply relying on optimal rate-distortion compression and hoping that compressed signals can be distinguished from each other.

eess.SP

Robust Load Disturbance Rejection in PWM DC-DC Buck Converters

This paper presents a novel approach to robust load disturbance rejection in DC-DC Buck converters. We propose a novel control scheme based on the design of two nested feedback loops. First, we design the controller in the outer loop using H infinity optimal control theory, and we show, by means of mu-analysis, that such a controller provides robust stability in the presence of uncertainty affecting the physical parameters of the circuit. Then, we introduce an inner feedback loop to improve the system's response to output load disturbances. As far as the inner loop is considered, we propose a novel load estimation-compensation (LEC) scheme, and we discuss under what conditions the insertion of such an inner loop preserves the robust stability of the entire control system. The LEC scheme is compared with the other two linear structures based on well-established disturbance rejection methods. The advantages of LEC in terms of both complexity of implementation and obtained performances are discussed and demonstrated by means of numerical simulation. Finally, we present experimental results obtained through the implementation of the proposed control scheme on a prototype board to demonstrate that the proposed approach significantly enhances disturbance rejection performances with respect to the approach commonly used in DC-DC buck converters.

eess.SY

Unsupervised Transcript-assisted Video Summarization and Highlight Detection

Video consumption is a key part of daily life, but watching entire videos can be tedious. To address this, researchers have explored video summarization and highlight detection to identify key video segments. While some works combine video frames and transcripts, and others tackle video summarization and highlight detection using Reinforcement Learning (RL), no existing work, to the best of our knowledge, integrates both modalities within an RL framework. In this paper, we propose a multimodal pipeline that leverages video frames and their corresponding transcripts to generate a more condensed version of the video and detect highlights using a modality fusion mechanism. The pipeline is trained within an RL framework, which rewards the model for generating diverse and representative summaries while ensuring the inclusion of video segments with meaningful transcript content. The unsupervised nature of the training allows for learning from large-scale unannotated datasets, overcoming the challenge posed by the limited size of existing annotated datasets. Our experiments show that using the transcript in video summarization and highlight detection achieves superior results compared to relying solely on the visual content of the video.

cs.CV

Deep learning-driven pulmonary artery and vein segmentation reveals demography-associated vasculature anatomical differences

Pulmonary artery-vein segmentation is crucial for disease diagnosis and surgical planning and is traditionally achieved by Computed Tomography Pulmonary Angiography (CTPA). However, concerns regarding adverse health effects from contrast agents used in CTPA have constrained its clinical utility. In contrast, identifying arteries and veins using non-contrast CT, a conventional and low-cost clinical examination routine, has long been considered impossible. Here we propose a High-abundant Pulmonary Artery-vein Segmentation (HiPaS) framework achieving accurate artery-vein segmentation on both non-contrast CT and CTPA across various spatial resolutions. HiPaS first performs spatial normalization on raw CT volumes via a super-resolution module, and then iteratively achieves segmentation results at different branch levels by utilizing the lower-level vessel segmentation as a prior for higher-level vessel segmentation. We trained and validated HiPaS on our established multi-centric dataset comprising 1,073 CT volumes with meticulous manual annotations. Both quantitative experiments and clinical evaluation demonstrated the superior performance of HiPaS, achieving an average dice score of 91.8% and a sensitivity of 98.0%. Further experiments showed the non-inferiority of HiPaS segmentation on non-contrast CT compared to segmentation on CTPA. Employing HiPaS, we have conducted an anatomical study of pulmonary vasculature on 11,784 participants in China (six sites), discovering a new association of pulmonary vessel anatomy with sex, age, and disease states: vessel abundance suggests a significantly higher association with females than males with slightly decreasing with age, and is also influenced by certain diseases, under the controlling of lung volumes.

cs.CV

Magnetic Field Gated and Current Controlled Spintronic Mem-transistor Neuron -based Spiking Neural Networks

Spintronic devices, such as the domain walls and skyrmions, have shown significant potential for applications in energy-efficient data storage and beyond CMOS computing architectures. In recent years, spiking neural networks have shown more bio-plausibility. Based on the magnetic multilayer spintronic devices, we demonstrate the magnetic field-gated Leaky integrate and fire neuron characteristics for the spiking neural network applications. The LIF characteristics are controlled by the current pulses, which drive the domain wall, and an external magnetic field is used as the bias to tune the firing properties of the neuron. Thus, the device works like a gate-controlled LIF neuron, acting like a spintronic Mem-Transistor device. We develop a LIF neuron model based on the measured characteristics to show the device integration in the system-level SNNs. We extend the study and propose a scaled version of the demonstrated device with a multilayer spintronic domain wall magnetic tunnel junction as a LIF neuron. using the combination of SOT and the variation of the demagnetization energy across the thin film, the modified leaky integrate and fire LIF neuron characteristics are realized in the proposed devices. The neuron device characteristics are modeled as the modified LIF neuron model. Finally, we integrate the measured and simulated neuron models in the 3-layer spiking neural network and convolutional spiking neural network CSNN framework to test these spiking neuron models for classification of the MNIST and FMNIST datasets. In both architectures, the network achieves classification accuracy above 96%. Considering the good system-level performance, mem-transistor properties, and promise for scalability. The presented devices show an excellent properties for neuromorphic computing applications.

physics.app-ph

Magnetic Skyrmion: From Fundamental Physics to Pioneering Applications

Skyrmionic devices exhibit energy-efficient and high-integration data storage and computing capabilities due to their small size, topological protection, and low drive current requirements. So, to realize these devices, an extensive study, from fundamental physics to practical applications, becomes essential. In this article, we present an exhaustive review of the advancements in understanding the fundamental physics behind magnetic skyrmions and the novel data storage and computing technologies based on them. We begin with an in-depth discussion of fundamental concepts such as topological protection, stability, statics and dynamics essential for understanding skyrmions, henceforth the foundation of skyrmion technologies. For the realization of CMOS-compatible skyrmion functional devices, the writing and reading of the skyrmions are crucial. We discuss the developments in different writing schemes such as STT, SOT, and VCMA. The reading of skyrmions is predominantly achieved via two mechanisms: the Magnetoresistive Tunnel Junction (MTJ) TMR effect and topological resistivity (THE). So, a thorough investigation into the Skyrmion Hall Effect, topological properties, and emergent fields is also provided, concluding the discussion on skyrmion reading developments. Based on the writing and reading schemes, we discuss the applications of the skyrmions in conventional logic, unconventional logic, memory applications, and neuromorphic computing in particular. Subsequently, we present an overview of the potential of skyrmion-hosting Majorana Zero Modes (MZMs) in the emerging Topological Quantum Computation and helicity-dependent skyrmion qubits.

cond-mat.mes-hall

Multilayer Ferromagnetic Spintronic Devices for Neuromorphic Computing Applications

Spintronics has gone through substantial progress due to its applications in energy-efficient memory, logic and unconventional computing paradigms. Multilayer ferromagnetic thin films are extensively studied for understanding the domain wall and skyrmion dynamics. However, most of these studies are confined to the materials and domain wall/skyrmion physics. In this paper, we present the experimental and micromagnetic realization of a multilayer ferromagnetic spintronic device for neuromorphic computing applications. The device exhibits multilevel resistance states and the number of resistance states increases with lowering temperature. This is supported by the multilevel magnetization behavior observed in the micromagnetic simulations. Furthermore, the evolution of resistance states with spin-orbit torque is also explored in experiments and simulations. Using the multi-level resistance states of the device, we propose its applications as a synaptic device in hardware neural networks and study the linearity performance of the synaptic devices. The neural network based on these devices is trained and tested on the MNIST dataset using a supervised learning algorithm. The devices at the chip level achieve 90\% accuracy. Thus, proving its applications in neuromorphic computing. Furthermore, we lastly discuss the possible application of the device in cryogenic memory electronics for quantum computers.

physics.app-ph

Anomalous and Topological Hall Resistivity in Ta/CoFeB/MgO Magnetic Systems for Neuromorphic Computing Applications

Topologically protected spin textures, such as magnetic skyrmions, have the potential for dense data storage as well as energy-efficient computing due to their small size and a low driving current. The evaluation of the writing and reading of the skyrmion's magnetic and electrical characteristics is a key step toward the implementation of these devices. In this paper, we present the magnetic heterostructure Hall bar device and study the anomalous Hall and topological Hall signals in the device. Using the combination of different measurements like magnetometry at different temperatures, Hall effect measurement from 2K to 300K, and magnetic force microscopy imaging, we investigate the magnetic and electrical characteristics of the magnetic structure. We measure the skyrmion topological resistivity at different temperatures as a function of the magnetic field. The topological resistivity is maximum around the zero magnetic field and it decreases to zero at the saturating field. This is further supported by MFM imaging. Interestingly the resistivity decreases linearly with the field, matching the behavior observed in the corresponding micromagnetic simulations. We combine the experimental results with micromagnetic simulations, thus propose a skyrmion-based synaptic device and show spin-orbit torque-controlled potentiation/depression in the device. The device performance as the synapse for neuromorphic computing is further evaluated in a convolutional neural network CNN. The neural network is trained and tested on the MNIST data set we show devices acting as synapses achieving a recognition accuracy close to 90%, on par with the ideal software-based weights which offer an accuracy of 92%.

cond-mat.mtrl-sci

Anomaly Detection based on Compressed Data: an Information Theoretic Characterization

We analyze the effect of lossy compression in the processing of sensor signals that must be used to detect anomalous events in the system under observation. The intuitive relationship between the quality loss at higher compression and the possibility of telling anomalous behaviours from normal ones is formalized in terms of information-theoretic quantities. Some analytic derivations are made within the Gaussian framework and possibly in the asymptotic regime for what concerns the stretch of signals considered. Analytical conclusions are matched with the performance of practical detectors in a toy case allowing the assessment of different compression/detector configurations.

cs.IT

From Chaos to Pseudo-Randomness: A Case Study on the 2D Coupled Map Lattice

Applying chaos theory for secure digital communications is promising and it is well acknowledged that in such applications the underlying chaotic systems should be carefully chosen. However, the requirements imposed on the chaotic systems are usually heuristic, without theoretic guarantee for the resultant communication scheme. Among all the primitives for secure communications, it is well-accepted that (pseudo) random numbers are most essential. Taking the well-studied two-dimensional coupled map lattice (2D CML) as an example, this paper performs a theoretical study towards pseudo-random number generation with the 2D CML. In so doing, an analytical expression of the Lyapunov exponent (LE) spectrum of the 2D CML is first derived. Using the LEs, one can configure system parameters to ensure the 2D CML only exhibits complex dynamic behavior, and then collect pseudo-random numbers from the system orbits. Moreover, based on the observation that least significant bit distributes more evenly in the (pseudo) random distribution, an extraction algorithm E is developed with the property that, when applied to the orbits of the 2D CML, it can squeeze uniform bits. In implementation, if fixed-point arithmetic is used in binary format with a precision of $z$ bits after the radix point, E can ensure that the deviation of the squeezed bits is bounded by $2^{-z}$ . Further simulation results demonstrate that the new method not only guide the 2D CML model to exhibit complex dynamic behavior, but also generate uniformly distributed independent bits. In particular, the squeezed pseudo random bits can pass both NIST 800-22 and TestU01 test suites in various settings. This study thereby provides a theoretical basis for effectively applying the 2D CML to secure communications.

cs.CR

On the security of a class of diffusion mechanisms for image encryption

The need for fast and strong image cryptosystems motivates researchers to develop new techniques to apply traditional cryptographic primitives in order to exploit the intrinsic features of digital images. One of the most popular and mature technique is the use of complex ynamic phenomena, including chaotic orbits and quantum walks, to generate the required key stream. In this paper, under the assumption of plaintext attacks we investigate the security of a classic diffusion mechanism (and of its variants) used as the core cryptographic rimitive in some image cryptosystems based on the aforementioned complex dynamic phenomena. We have theoretically found that regardless of the key schedule process, the data complexity for recovering each element of the equivalent secret key from these diffusion mechanisms is only O(1). The proposed analysis is validated by means of numerical examples. Some additional cryptographic applications of our work are also discussed.

cs.CR