Catalog footprint

What is connected

42works
28topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

42 published item(s)

preprint2026arXiv

MATEX: Multi-scale Attention and Text-guided Explainability of Medical Vision-Language Models

We introduce MATEX (Multi-scale Attention and Text-guided Explainability), a novel framework that advances interpretability in medical vision-language models by incorporating anatomically informed spatial reasoning. MATEX synergistically combines multi-layer attention rollout, text-guided spatial priors, and layer consistency analysis to produce precise, stable, and clinically meaningful gradient attribution maps. By addressing key limitations of prior methods, such as spatial imprecision, lack of anatomical grounding, and limited attention granularity, MATEX enables more faithful and interpretable model explanations. Evaluated on the MS-CXR dataset, MATEX outperforms the state-of-the-art M2IB approach in both spatial precision and alignment with expert-annotated findings. These results highlight MATEX's potential to enhance trust and transparency in radiological AI applications.

preprint2026arXiv

Predicting When to Trust Vision-Language Models for Spatial Reasoning

Vision-Language Models (VLMs) demonstrate impressive capabilities across multimodal tasks, yet exhibit systematic spatial reasoning failures, achieving only 49% (CLIP) to 54% (BLIP-2) accuracy on basic directional relationships. For safe deployment in robotics and autonomous systems, we need to predict when to trust VLM spatial predictions rather than accepting all outputs. We propose a vision-based confidence estimation framework that validates VLM predictions through independent geometric verification using object detection. Unlike text-based approaches relying on self-assessment, our method fuses four signals via gradient boosting: geometric alignment between VLM claims and coordinates, spatial ambiguity from overlap, detection quality, and VLM internal uncertainty. We achieve 0.674 AUROC on BLIP-2 (34.0% improvement over text-based baselines) and 0.583 AUROC on CLIP (16.1% improvement), generalizing across generative and classification architectures. Our framework enables selective prediction: at 60% target accuracy, we achieve 61.9% coverage versus 27.6% baseline (2.2x improvement) on BLIP-2. Feature analysis reveals vision-based signals contribute 87.4% of model importance versus 12.7% from VLM confidence, validating that external geometric verification outperforms self-assessment. We demonstrate reliable scene graph construction where confidence-based pruning improves precision from 52.1% to 78.3% while retaining 68.2% of edges.

preprint2024arXiv

CrisisViT: A Robust Vision Transformer for Crisis Image Classification

In times of emergency, crisis response agencies need to quickly and accurately assess the situation on the ground in order to deploy relevant services and resources. However, authorities often have to make decisions based on limited information, as data on affected regions can be scarce until local response services can provide first-hand reports. Fortunately, the widespread availability of smartphones with high-quality cameras has made citizen journalism through social media a valuable source of information for crisis responders. However, analyzing the large volume of images posted by citizens requires more time and effort than is typically available. To address this issue, this paper proposes the use of state-of-the-art deep neural models for automatic image classification/tagging, specifically by adapting transformer-based architectures for crisis image classification (CrisisViT). We leverage the new Incidents1M crisis image dataset to develop a range of new transformer-based image classification models. Through experimentation over the standard Crisis image benchmark dataset, we demonstrate that the CrisisViT models significantly outperform previous approaches in emergency type, image relevance, humanitarian category, and damage severity classification. Additionally, we show that the new Incidents1M dataset can further augment the CrisisViT models resulting in an additional 1.25% absolute accuracy gain.

preprint2022arXiv

An exact quantum hidden subgroup algorithm and applications to solvable groups

We present a polynomial time exact quantum algorithm for the hidden subgroup problem in $Z_{m^k}^n$. The algorithm uses the quantum Fourier transform modulo m and does not require factorization of m. For smooth m, i.e., when the prime factors of m are of size poly(log m), the quantum Fourier transform can be exactly computed using the method discovered independently by Cleve and Coppersmith, while for general m, the algorithm of Mosca and Zalka is available. Even for m=3 and k=1 our result appears to be new. We also present applications to compute the structure of abelian and solvable groups whose order has the same (but possibly unknown) prime factors as m. The applications for solvable groups also rely on an exact version of a technique proposed by Watrous for computing the uniform superposition of elements of subgroups.

preprint2022arXiv

An exact quantum order finding algorithm and its applications

We present an efficient exact quantum algorithm for order finding problem when a multiple $m$ of the order $r$ is known. The algorithm consists of two main ingredients. The first ingredient is the exact quantum Fourier transform proposed by Mosca and Zalka in [MZ03]. The second ingredient is an amplitude amplification version of Brassard and Hoyer in [BH97] combined with some ideas from the exact discrete logarithm procedure by Mosca and Zalka in [MZ03]. As applications, we show how the algorithm derandomizes the quantum algorithm for primality testing proposed by Donis-Vela and Garcia-Escartin in [DVGE18], and serves as a subroutine of an efficient exact quantum algorithm for finding primitive elements in arbitrary finite fields. .

preprint2022arXiv

Blockchain-based Collaborated Federated Learning for Improved Security, Privacy and Reliability

Federated Learning (FL) provides privacy preservation by allowing the model training at edge devices without the need of sending the data from edge to a centralized server. FL has distributed the implementation of ML. Another variant of FL which is well suited for the Internet of Things (IoT) is known as Collaborated Federated Learning (CFL), which does not require an edge device to have a direct link to the model aggregator. Instead, the devices can connect to the central model aggregator via other devices using them as relays. Although, FL and CFL protect the privacy of edge devices but raises security challenges for a centralized server that performs model aggregation. The centralized server is prone to malfunction, backdoor attacks, model corruption, adversarial attacks and external attacks. Moreover, edge device to centralized server data exchange is not required in FL and CFL, but model parameters are sent from the model aggregator (global model) to edge devices (local model), which is still prone to cyber-attacks. These security and privacy concerns can be potentially addressed by Blockchain technology. The blockchain is a decentralized and consensus-based chain where devices can share consensus ledgers with increased reliability and security, thus significantly reducing the cyberattacks on an exchange of information. In this work, we will investigate the efficacy of blockchain-based decentralized exchange of model parameters and relevant information among edge devices and from a centralized server to edge devices. Moreover, we will be conducting the feasibility analysis for blockchain-based CFL models for different application scenarios like the internet of vehicles, and the internet of things. The proposed study aims to improve the security, reliability and privacy preservation by the use of blockchain-powered CFL.

preprint2022arXiv

Effect of Measurement Errors on the Multivariate CUSUM CoDa Control Chart for the Manufacturing Process

Control charts, one of the main tools in Statistical Process Control (SPC), have been widely adopted in manufacturing sectors as an effective strategy for malfunction detection throughout the previous decades. Measurement errors (M.E's) are involved in the quality characteristic of interest. The authors explored the impact of a linear covariate error model on the multivariate cumulative sum (CUSUM) control charts for a specific kind of data known as compositional data(CoDa). The average run length ARL is used to assess the performance of the proposed chart. The results indicate that M.E's significantly affects the multivariate CUSUM-CoDa control charts. The authors have used the Markov chain method to study the impact of different involved parameters using four different cases for the variance-covariance matrix (i.e. uncorrelated with equal variances, negatively correlated with equal variances, uncorrelated with unequal variances, positively correlated with unequal variances). The authors concluded that the ARL of the multivariate CUSUM-CoDa chart increase with an increase in the value of error variance-covariance matrix, while the ARL decreases with an increase in the subgroup size m or the constant powering b. For the implementation of the proposal, two illustrated examples have been reported for multivariate CUSUM-CoDa control charts in the presence of M.E's. One deals with the manufacturing process of uncoated aspirin tablets, and the other is based on monitoring machines in the muesli manufacturing process.

preprint2022arXiv

Incidents1M: a large-scale dataset of images with natural disasters, damage, and incidents

Natural disasters, such as floods, tornadoes, or wildfires, are increasingly pervasive as the Earth undergoes global warming. It is difficult to predict when and where an incident will occur, so timely emergency response is critical to saving the lives of those endangered by destructive events. Fortunately, technology can play a role in these situations. Social media posts can be used as a low-latency data source to understand the progression and aftermath of a disaster, yet parsing this data is tedious without automated methods. Prior work has mostly focused on text-based filtering, yet image and video-based filtering remains largely unexplored. In this work, we present the Incidents1M Dataset, a large-scale multi-label dataset which contains 977,088 images, with 43 incident and 49 place categories. We provide details of the dataset construction, statistics and potential biases; introduce and train a model for incident detection; and perform image-filtering experiments on millions of images on Flickr and Twitter. We also present some applications on incident analysis to encourage and enable future work in computer vision for humanitarian aid. Code, data, and models are available at http://incidentsdataset.csail.mit.edu.

preprint2022arXiv

Is Blockchain for Internet of Medical Things a Panacea for COVID-19 Pandemic?

The outbreak of the COVID-19 pandemic has deeply influenced the lifestyle of the general public and the healthcare system of the society. As a promising approach to address the emerging challenges caused by the epidemic of infectious diseases like COVID-19, Internet of Medical Things (IoMT) deployed in hospitals, clinics, and healthcare centers can save the diagnosis time and improve the efficiency of medical resources though privacy and security concerns of IoMT stall the wide adoption. In order to tackle the privacy, security, and interoperability issues of IoMT, we propose a framework of blockchain-enabled IoMT by introducing blockchain to incumbent IoMT systems. In this paper, we review the benefits of this architecture and illustrate the opportunities brought by blockchain-enabled IoMT. We also provide use cases of blockchain-enabled IoMT on fighting against the COVID-19 pandemic, including the prevention of infectious diseases, location sharing and contact tracing, and the supply chain of injectable medicines. We also outline future work in this area.

preprint2022arXiv

Unravelling Token Ecosystem of EOSIO Blockchain

Being the largest Initial Coin Offering project, EOSIO has attracted great interest in cryptocurrency markets. Despite its popularity and prosperity (e.g., 26,311,585,008 token transactions occurred from June 8, 2018 to Aug. 5, 2020), there is almost no work investigating the EOSIO token ecosystem. To fill this gap, we are the first to conduct a systematic investigation on the EOSIO token ecosystem by conducting a comprehensive graph analysis on the entire on-chain EOSIO data (nearly 135 million blocks). We construct token creator graphs, token-contract creator graphs, token holder graphs, and token transfer graphs to characterize token creators, holders, and transfer activities. Through graph analysis, we have obtained many insightful findings and observed some abnormal trading patterns. Moreover, we propose a fake-token detection algorithm to identify tokens generated by fake users or fake transactions and analyze their corresponding manipulation behaviors. Evaluation results also demonstrate the effectiveness of our algorithm.

preprint2022arXiv

Wireless Powering Internet of Things with UAVs: Challenges and Opportunities

Unmanned aerial vehicles (UAVs) have the potential to overcome the deployment constraint of Internet of Things (IoT) in remote or rural area. Wirelessly powered communications (WPC) can address the battery limitation of IoT devices through transferring wireless power to IoT devices. The integration of UAVs and WPC, namely UAV-enabled Wireless Powering IoT (Ue-WPIoT) can greatly extend the IoT applications from cities to remote or rural areas. In this article, we present a state-of-the-art overview of Ue-WPIoT by first illustrating the working flow of Ue-WPIoT and discussing the challenges. We then introduce the enabling technologies in realizing Ue-WPIoT. Simulation results validate the effectiveness of the enabling technologies in Ue-WPIoT. We finally outline the future directions and open issues.

preprint2020arXiv

A First Look at Privacy Analysis of COVID-19 Contact Tracing Mobile Applications

Today's smartphones are equipped with a large number of powerful value-added sensors and features such as a low power Bluetooth sensor, powerful embedded sensors such as the digital compass, accelerometer, GPS sensors, Wi-Fi capabilities, microphone, humidity sensors, health tracking sensors, and a camera, etc. These value-added sensors have revolutionized the lives of the human being in many ways such, as tracking the health of the patients and movement of doctors, tracking employees movement in large manufacturing units, and monitoring the environment, etc. These embedded sensors could also be used for large-scale personal, group, and community sensing applications especially tracing the spread of certain diseases. Governments and regulators are turning to use these features to trace the people thought to have symptoms of certain diseases or virus e.g. COVID-19. The outbreak of COVID-19 in December 2019, has seen a surge of the mobile applications for tracing, tracking and isolating the persons showing COVID-19 symptoms to limit the spread of disease to the larger community. The use of embedded sensors could disclose private information of the users thus potentially bring threat to the privacy and security of users. In this paper, we analyzed a large set of smartphone applications that have been designed to contain the spread of the COVID-19 virus and bring the people back to normal life. Specifically, we have analyzed what type of permission these smartphone apps require, whether these permissions are necessary for the track and trace, how data from the user devices is transported to the analytic center, and analyzing the security measures these apps have deployed to ensure the privacy and security of users.

preprint2020arXiv

Analysis of Social Media Data using Multimodal Deep Learning for Disaster Response

Multimedia content in social media platforms provides significant information during disaster events. The types of information shared include reports of injured or deceased people, infrastructure damage, and missing or found people, among others. Although many studies have shown the usefulness of both text and image content for disaster response purposes, the research has been mostly focused on analyzing only the text modality in the past. In this paper, we propose to use both text and image modalities of social media data to learn a joint representation using state-of-the-art deep learning techniques. Specifically, we utilize convolutional neural networks to define a multimodal deep learning architecture with a modality-agnostic shared representation. Extensive experiments on real-world disaster datasets show that the proposed multimodal architecture yields better performance than models trained using a single modality (e.g., either text or image).

preprint2020arXiv

Artificial Intelligence and Machine Learning in 5G Network Security: Opportunities, advantages, and future research trends

Recent technological and architectural advancements in 5G networks have proven their worth as the deployment has started over the world. Key performance elevating factor from access to core network are softwareization, cloudification and virtualization of key enabling network functions. Along with the rapid evolution comes the risks, threats and vulnerabilities in the system for those who plan to exploit it. Therefore, ensuring fool proof end-to-end (E2E) security becomes a vital concern. Artificial intelligence (AI) and machine learning (ML) can play vital role in design, modelling and automation of efficient security protocols against diverse and wide range of threats. AI and ML has already proven their effectiveness in different fields for classification, identification and automation with higher accuracy. As 5G networks' primary selling point has been higher data rates and speed, it will be difficult to tackle wide range of threats from different points using typical/traditional protective measures. Therefore, AI and ML can play central role in protecting highly data-driven softwareized and virtualized network components. This article presents AI and ML driven applications for 5G network security, their implications and possible research directions. Also, an overview of key data collection points in 5G architecture for threat classification and anomaly detection are discussed.

preprint2020arXiv

Blockchain-enabled Internet of Medical Things to Combat COVID-19

We are experiencing an unprecedented healthcare crisis caused by newly-discovered corona-virus disease (COVID-19). The outbreaks of COVID-19 reveal the frailties of existing healthcare systems. Therefore, the digital transformation of healthcare systems becomes an inevitable trend. During this process, the Internet of Medical Things (IoMT) plays a crucial role while intrinsic vulnerabilities of security and privacy deter the wide adoption of IoMT. In this article, we present a blockchain-enabled IoMT to address the security and privacy concerns of IoMT systems. We also discuss the solutions brought by blockchain-enabled IoMT to COVID-19 from five different perspectives. Moreover, we outline the open challenges and future directions of blockchain-enabled IoMT.

preprint2020arXiv

Blockchain-enabled Resource Management and Sharing for 6G Communications

The sixth generation (6G) network must provide performance superior to previous generations in order to meet the requirements of emerging services and applications, such as multi-gigabit transmission rate, even higher reliability, sub 1 millisecond latency and ubiquitous connection for Internet of Everything. However, with the scarcity of spectrum resources, efficient resource management and sharing is crucial to achieve all these ambitious requirements. One possible technology to enable all of this is blockchain, which has recently gained significance and will be of paramount importance to 6G networks and beyond due to its inherent properties. In particular, the integration of blockchain in 6G will enable the network to monitor and manage resource utilization and sharing efficiently. Hence, in this article, we discuss the potentials of blockchain for resource management and sharing in 6G using multiple application scenarios namely, Internet of things, device-to-device communications, network slicing, and inter-domain blockchain ecosystems.

preprint2020arXiv

Composition, Size, and Surface Functionalization dependent Optical Properties of Lead Bromide Perovskite Nanocrystals

The photoluminescence (PL), color purity, and stability of lead halide perovskite nanocrystals depend critically on the surface passivation. We present a study on the temperature dependent PL and PL decay dynamics of lead bromide perovskite nanocrystals characterized by different types of A cations, surface ligands, and nanocrystal sizes. Throughout, we observe a single emission peak from cryogenic to ambient temperature. The PL decay dynamics are dominated by the surface passivation, and a post-synthesis ligand exchange with a quaternary ammonium bromide (QAB) results in a more stable passivation over a larger temperature range. The PL intensity is highest from 50K-250K, which indicates that the ligand binding competes with the thermal energy at ambient temperature. Despite the favorable PL dynamics of nanocrystals passivated with QAB ligands (monoexponential PL decay over a large temperature range, increased PL intensity and stability), the surface passivation still needs improvement toward increased emission intensity in nanocrystal films.

preprint2020arXiv

Detecting natural disasters, damage, and incidents in the wild

Responding to natural disasters, such as earthquakes, floods, and wildfires, is a laborious task performed by on-the-ground emergency responders and analysts. Social media has emerged as a low-latency data source to quickly understand disaster situations. While most studies on social media are limited to text, images offer more information for understanding disaster and incident scenes. However, no large-scale image datasets for incident detection exists. In this work, we present the Incidents Dataset, which contains 446,684 images annotated by humans that cover 43 incidents across a variety of scenes. We employ a baseline classification model that mitigates false-positive errors and we perform image filtering experiments on millions of social media images from Flickr and Twitter. Through these experiments, we show how the Incidents Dataset can be used to detect images with incidents in the wild. Code, data, and models are available online at http://incidentsdataset.csail.mit.edu.

preprint2020arXiv

Nano- and microscale apertures in metal films fabricated by colloidal lithography with perovskite nanocrystals

We demonstrate patterning of metal surfaces based on lift-off of perovskite nanocrystals that enables the fabrication of nanometer-size features without the use of resist-based nanolithography. The perovskite nanocrystals act as templates for defining the shape of the apertures in metal layers, and we exploit the variety of sizes and shapes that can be controlled in the colloidal synthesis to demonstrate the fabrication of nanoholes, nanogaps and guides with size smaller than the wavelength of light in the visible spectrum. The process can be readily integrated with standard lithography and etching techniques for the creation of more complex structures.

preprint2020arXiv

Rapid Damage Assessment Using Social Media Images by Combining Human and Machine Intelligence

Rapid damage assessment is one of the core tasks that response organizations perform at the onset of a disaster to understand the scale of damage to infrastructures such as roads, bridges, and buildings. This work analyzes the usefulness of social media imagery content to perform rapid damage assessment during a real-world disaster. An automatic image processing system, which was activated in collaboration with a volunteer response organization, processed ~280K images to understand the extent of damage caused by the disaster. The system achieved an accuracy of 76% computed based on the feedback received from the domain experts who analyzed ~29K system-processed images during the disaster. An extensive error analysis reveals several insights and challenges faced by the system, which are vital for the research community to advance this line of research.

preprint2020arXiv

Thermal vulnerability detection in integrated electronic and photonic circuits using IR thermography

Failure prediction of any electrical/optical component is crucial for estimating its operating life. Using high temperature operating life (HTOL) tests, it is possible to model the failure mechanisms for integrated circuits. Conventional HTOL standards are not suitable for operating life prediction of photonic components owing to their functional dependence on thermo-optic effect. This work presents an IR-assisted thermal vulnerability detection technique suitable for photonic as well as electronic components. By accurately mapping the thermal profile of an integrated circuit under a stress condition, it is possible to precisely locate the heat center for predicting the long-term operational failures within the device under test. For the first time, the reliability testing is extended to a fully functional microwave photonic system using conventional IR thermography. By applying image fusion using affine transformation on multimodal acquisition, it was demonstrated that by comparing the IR profile and GDSII layout, it is possible to accurately locate the heat centers along with spatial information on the type of component. Multiple IR profiles of optical as well as electrical components/circuits were acquired and mapped onto the layout files. In order to ascertain the degree of effectiveness of the proposed technique, IR profiles of CMOS RF and digital circuits were also analyzed. The presented technique offers a reliable automated identification of heat spots within a circuit/system.

preprint2020arXiv

Unmanned Aerial Vehicle for Internet of Everything: Opportunities and Challenges

The recent advances in information and communication technology (ICT) have further extended Internet of Things (IoT) from the sole "things" aspect to the omnipotent role of "intelligent connection of things". Meanwhile, the concept of internet of everything (IoE) is presented as such an omnipotent extension of IoT. However, the IoE realization meets critical challenges including the restricted network coverage and the limited resource of existing network technologies. Recently, Unmanned Aerial Vehicles (UAVs) have attracted significant attentions attributed to their high mobility, low cost, and flexible deployment. Thus, UAVs may potentially overcome the challenges of IoE. This article presents a comprehensive survey on opportunities and challenges of UAV-enabled IoE. We first present three critical expectations of IoE: 1) scalability requiring a scalable network architecture with ubiquitous coverage, 2) intelligence requiring a global computing plane enabling intelligent things, 3) diversity requiring provisions of diverse applications. Thereafter, we review the enabling technologies to achieve these expectations and discuss four intrinsic constraints of IoE (i.e., coverage constraint, battery constraint, computing constraint, and security issues). We then present an overview of UAVs. We next discuss the opportunities brought by UAV to IoE. Additionally, we introduce a UAV-enabled IoE (Ue-IoE) solution by exploiting UAVs's mobility, in which we show that Ue-IoE can greatly enhance the scalability, intelligence and diversity of IoE. Finally, we outline the future directions in Ue-IoE.

preprint2016arXiv

A Robust Framework for Classifying Evolving Document Streams in an Expert-Machine-Crowd Setting

An emerging challenge in the online classification of social media data streams is to keep the categories used for classification up-to-date. In this paper, we propose an innovative framework based on an Expert-Machine-Crowd (EMC) triad to help categorize items by continuously identifying novel concepts in heterogeneous data streams often riddled with outliers. We unify constrained clustering and outlier detection by formulating a novel optimization problem: COD-Means. We design an algorithm to solve the COD-Means problem and show that COD-Means will not only help detect novel categories but also seamlessly discover human annotation errors and improve the overall quality of the categorization process. Experiments on diverse real data sets demonstrate that our approach is both effective and efficient.

preprint2016arXiv

Airy plasmons in graphene based waveguides

In this paper, we propose that both the quasi-transverse-magnetic (TM) and quasi-transverseelectric (TE) Airy plasmons can be supported in graphene-based waveguides. The solution of Airy plasmons is calculated analytically and the existence of Airy plasmons is studied under the paraxial approximation. Due to the tunability of the chemical potential of graphene, the self-accelerating behavior of quasi-TM Airy plasmons can be steered effectively, especially in multilayer graphene based waveguides. Besides the metals, graphene provides an additional platform to investigate the propagation of Airy plasmons and to design various plasmonic devices.

preprint2016arXiv

Colloidal Synthesis of Strongly Fluorescent CsPbBr3 Nanowires with Width Tunable down to the Quantum Confinement Regime

We report the colloidal synthesis of strongly fluorescent CsPbBr3 perovskite nanowires (NWs) with rectangular section and with tuneable width, from 20 nm (exhibiting no quantum confinement, hence emitting in the green) down to around 3 nm (in the strong quan-tum-confinement regime, emitting in the blue), by introducing in the synthesis a short acid (octanoic acid or hexanoic acid) together with alkyl amines (octylamine and oleylamine). Temperatures below 70 °C promoted the formation of monodisperse, few unit cell thick NWs that were free from byproducts. The photoluminescence quantum yield of the NW samples went from 12% for non-confined NWs emitting at 524 nm to a maximum of 77% for the 5 nm diameter NWs emitting at 497 nm, down to 30% for the thinnest NWs (diameter ~ 3nm), in the latter sample most likely due to aggregation occurring in solution.

preprint2016arXiv

Cross-Language Domain Adaptation for Classifying Crisis-Related Short Messages

Rapid crisis response requires real-time analysis of messages. After a disaster happens, volunteers attempt to classify tweets to determine needs, e.g., supplies, infrastructure damage, etc. Given labeled data, supervised machine learning can help classify these messages. Scarcity of labeled data causes poor performance in machine training. Can we reuse old tweets to train classifiers? How can we choose labeled tweets for training? Specifically, we study the usefulness of labeled data of past events. Do labeled tweets in different language help? We observe the performance of our classifiers trained using different combinations of training sets obtained from past disasters. We perform extensive experimentation on real crisis datasets and show that the past labels are useful when both source and target events are of the same type (e.g. both earthquakes). For similar languages (e.g., Italian and Spanish), cross-language domain adaptation was useful, however, when for different languages (e.g., Italian and English), the performance decreased.

preprint2016arXiv

Enabling Digital Health by Automatic Classification of Short Messages

In response to the growing HIV/AIDS and other health-related issues, UNICEF through their U-Report platform receives thousands of messages (SMS) every day to provide prevention strategies, health case advice, and counsel- ing support to vulnerable population. Due to a rapid increase in U-Report usage (up to 300% in last 3 years), plus approximately 1,000 new registrations each day, the volume of messages has thus continued to increase, which made it impossible for the team at UNICEF to process them in a timely manner. In this paper, we present a platform designed to perform automatic classification of short messages (SMS) in real-time to help UNICEF categorize and prioritize health-related messages as they arrive. We employ a hybrid approach, which combines human and machine intelligence that seeks to resolve the information overload issue by introducing processing of large-scale data at high-speed while maintaining a high classification accuracy. The system has recently been tested in conjunction with UNICEF in Zambia to classify short messages received via the U-Report platform on various health related issues. The system is designed to enable UNICEF make sense of a large volume of short messages in a timely manner. In terms of evaluation, we report design choices, challenges, and performance of the system observed during the deployment to validate its effectiveness.

preprint2016arXiv

Hybrid Airy Plasmons with Dynamically Steerable Trajectories

With the intriguing properties of diffraction-free, self-accelerating, and self-healing, Airy plasmons are promising to be used in the trapping, transporting, and sorting of micro-objects, imaging, and chip scale signal processing. However, the high dissipative loss and the lack of dynamical steerability restrict the implementation of Airy plasmons in these applications. Here we reveal the hybrid Airy plasmons for the first time by taking a hybrid graphene-based plasmonic waveguide in the terahertz (THz) domain as an example. Due to the coupling between an optical mode and a plasmonic mode, the hybrid Airy plasmons can have large propagation lengths and effective transverse deflections, where the transverse waveguide confinements are governed by the hybrid modes with moderate quality factors. Meanwhile, the propagation trajectories of hybrid Airy plasmons are dynamically steerable by changing the chemical potential of graphene. These hybrid Airy plasmons may promote the further discovery of non-diffracting beams with the emerging developments of optical tweezers and tractor beams.

preprint2016arXiv

Rapid Classification of Crisis-Related Data on Social Networks using Convolutional Neural Networks

The role of social media, in particular microblogging platforms such as Twitter, as a conduit for actionable and tactical information during disasters is increasingly acknowledged. However, time-critical analysis of big crisis data on social media streams brings challenges to machine learning techniques, especially the ones that use supervised learning. The Scarcity of labeled data, particularly in the early hours of a crisis, delays the machine learning process. The current state-of-the-art classification methods require a significant amount of labeled data specific to a particular event for training plus a lot of feature engineering to achieve best results. In this work, we introduce neural network based classification methods for binary and multi-class tweet classification task. We show that neural network based models do not require any feature engineering and perform better than state-of-the-art methods. In the early hours of a disaster when no labeled data is available, our proposed method makes the best use of the out-of-event data and achieves good results.

preprint2016arXiv

Summarizing Situational and Topical Information During Crises

The use of microblogging platforms such as Twitter during crises has become widespread. More importantly, information disseminated by affected people contains useful information like reports of missing and found people, requests for urgent needs etc. For rapid crisis response, humanitarian organizations look for situational awareness information to understand and assess the severity of the crisis. In this paper, we present a novel framework (i) to generate abstractive summaries useful for situational awareness, and (ii) to capture sub-topics and present a short informative summary for each of these topics. A summary is generated using a two stage framework that first extracts a set of important tweets from the whole set of information through an Integer-linear programming (ILP) based optimization technique and then follows a word graph and concept event based abstractive summarization technique to produce the final summary. High accuracies obtained for all the tasks show the effectiveness of the proposed framework.

preprint2016arXiv

Twitter as a Lifeline: Human-annotated Twitter Corpora for NLP of Crisis-related Messages

Microblogging platforms such as Twitter provide active communication channels during mass convergence and emergency events such as earthquakes, typhoons. During the sudden onset of a crisis situation, affected people post useful information on Twitter that can be used for situational awareness and other humanitarian disaster response efforts, if processed timely and effectively. Processing social media information pose multiple challenges such as parsing noisy, brief and informal messages, learning information categories from the incoming stream of messages and classifying them into different classes among others. One of the basic necessities of many of these tasks is the availability of data, in particular human-annotated data. In this paper, we present human-annotated Twitter corpora collected during 19 different crises that took place between 2013 and 2015. To demonstrate the utility of the annotations, we train machine learning classifiers. Moreover, we publish first largest word2vec word embeddings trained on 52 million crisis-related tweets. To deal with tweets language issues, we present human-annotated normalized lexical resources for different lexical variations.

preprint2015arXiv

Electron Spin Resonance in a Two-Dimensional Fermi Liquid with Spin-Orbit Coupling

Electron spin resonance (ESR) is usually interpreted as a single-particle phenomenon protected from the effect of many-body correlations. We show that this is not the case in a two-dimensional Fermi liquid (FL) with spin-orbit coupling (SOC). Depending on whether the magnetic field is below or above some critical value, ESR in such a system probes up to three collective chiral-spin modes, augmented by the presence of the field, or the Larmor mode, augmented both by SOC and FL renormalizations. We argue that ESR can be used as a probe not only for SOC but also for many-body physics.

preprint2015arXiv

Processing Social Media Messages in Mass Emergency: A Survey

Social media platforms provide active communication channels during mass convergence and emergency events such as disasters caused by natural hazards. As a result, first responders, decision makers, and the public can use this information to gain insight into the situation as it unfolds. In particular, many social media messages communicated during emergencies convey timely, actionable information. Processing social media messages to obtain such information, however, involves solving multiple challenges including: handling information overload, filtering credible information, and prioritizing different classes of messages. These challenges can be mapped to classical information processing operations such as filtering, classifying, ranking, aggregating, extracting, and summarizing. We survey the state of the art regarding computational methods to process social media messages, focusing on their application in emergency response scenarios. We examine the particularities of this setting, and then methodically examine a series of key sub-problems ranging from the detection of events to the creation of actionable and useful summaries.

preprint2014arXiv

An Effective End-User Development Approach Through Domain-Specific Mashups for Research Impact Evaluation

Over the last decade, there has been growing interest in the assessment of the performance of researchers, research groups, universities and even countries. The assessment of productivity is an instrument to select and promote personnel, assign research grants and measure the results of research projects. One particular assessment approach is bibliometrics i.e., the quantitative analysis of scientific publications through citation and content analysis. However, there is little consensus today on how research evaluation should be performed, and it is commonly acknowledged that the quantitative metrics available today are largely unsatisfactory. A number of different scientific data sources available on the Web (e.g., DBLP, Google Scholar) that are used for such analysis purposes. Taking data from these diverse sources, performing the analysis and visualizing results in different ways is not a trivial and straight forward task. Moreover, people involved in such evaluation processes are not always IT experts and hence not capable to crawl data sources, merge them and compute the needed evaluation procedures. The recent emergence of mashup tools has refueled research on end-user development, i.e., on enabling end-users without programming skills to produce their own applications. We believe that the heart of the problem is that it is impractical to design tools that are generic enough to cover a wide range of application domains, powerful enough to enable the specification of non-trivial logic, and simple enough to be actually accessible to non-programmers. This thesis presents a novel approach for an effective end-user development, specifically for non-programmers. That is, we introduce a domain-specific approach to mashups that "speaks the language of users"., i.e., that is aware of the terminology, concepts, rules, and conventions (the domain) the user is comfortable with.

preprint2014arXiv

Engineering Crowdsourced Stream Processing Systems

A crowdsourced stream processing system (CSP) is a system that incorporates crowdsourced tasks in the processing of a data stream. This can be seen as enabling crowdsourcing work to be applied on a sample of large-scale data at high speed, or equivalently, enabling stream processing to employ human intelligence. It also leads to a substantial expansion of the capabilities of data processing systems. Engineering a CSP system requires the combination of human and machine computation elements. From a general systems theory perspective, this means taking into account inherited as well as emerging properties from both these elements. In this paper, we position CSP systems within a broader taxonomy, outline a series of design principles and evaluation metrics, present an extensible framework for their design, and describe several design patterns. We showcase the capabilities of CSP systems by performing a case study that applies our proposed framework to the design and analysis of a real system (AIDR) that classifies social media messages during time-critical crisis events. Results show that compared to a pure stream processing system, AIDR can achieve a higher data classification accuracy, while compared to a pure crowdsourcing solution, the system makes better use of human workers by requiring much less manual work effort.

preprint2014arXiv

Understanding Types of Users on Twitter

People use microblogging platforms like Twitter to involve with other users for a wide range of interests and practices. Twitter profiles run by different types of users such as humans, bots, spammers, businesses and professionals. This research work identifies six broad classes of Twitter users, and employs a supervised machine learning approach which uses a comprehensive set of features to classify users into the identified classes. For this purpose, we exploit users' profile and tweeting behavior information. We evaluate our approach by performing 10-fold cross validation using manually annotated 716 different Twitter profiles. High classification accuracy (measured using AUC, and precision, recall) reveals the significance of the proposed approach.

preprint2014arXiv

Weak total resolving sets in graphs

A set $W$ of vertices of $G$ is said to be a weak total resolving set for $G$ if $W$ is a resolving set for $G$ as well as for each $w\in W$, there is at least one element in $W-\{w\}$ that resolves $w$ and $v$ for every $v\in V(G)- W$. Weak total metric dimension of $G$ is the smallest order of a weak total resolving set for $G$. This paper includes the investigation of weak total metric dimension of trees. Also, weak total resolving number of a graph as well as randomly weak total $k$-dimensional graphs are defined and studied in this paper. Moreover, some characterizations and realizations regarding weak total resolving number and weak total metric dimension are given.

preprint2012arXiv

Electron transport through a diatomic molecule

Electron transport through a diatomic molecular tunnel junction shows wave like interference phenomenon. By using Keldysh non-equilibrium Green's function (NEGF) theory, we have explicitly presented current and differential conductance calculation for a diatomic molecular and two isolated atoms (two atoms having zero hybridization between their energy orbital) tunnel junctions. In case of a diatomic molecular tunnel junction, Green's function propagators entering into current and differential conductance formula interfere constructively for a molecular anti-bonding state and destructively for bonding state. Consequently, conductance through a molecular bonding state is suppressed, and to conserve current, conductance through anti-bonding state is enhanced. Therefore, current steps and differential conductance peaks amplitude show asymmetric correspondence between molecular bonding and anti-bonding states. Interestingly, for a diatomic molecule, comprising of two atoms of same energy level, these propagators interfere completely destructively for molecular bonding state and constructively for molecular anti-bonding state. Hence under such condition, a single step or a single peak is shown up in current versus voltage or differential conductance versus voltage studies.

preprint2012arXiv

Energy balancing through cluster head selection using K-Theorem in homogeneous wireless sensor networks

The objective of this paper is to increase life time of homogeneous wireless sensor networks (WSNs) through minimizing long range communication and energy balancing. Sensor nodes are resource constrained particularly with limited energy that is difficult or impossible to replenish. LEACH (Low Energy Adaptive Clustering Hierarchy) is most well-known cluster based architecture for WSN that aims to evenly dissipate energy among all sensor nodes. In cluster based architecture, the role of cluster head is very crucial for the successful operation of WSN because once the cluster head becomes non functional, the whole cluster becomes dysfunctional. We have proposed a modified cluster based WSN architecture by introducing a coordinator node (CN) that is rich in terms of resources. This CN take up the responsibility of transmitting data to the base station over longer distances from cluster heads. We have proposed a cluster head selection algorithm based on K-theorem and other parameters i.e. residual energy, distance to coordinator node, reliability and degree of mobility. The K-theorem is used to select candidate cluster heads based on bunch of sensor nodes in a cluster. We believe that the proposed architecture and algorithm achieves higher energy efficiency through minimizing communication and energy balancing. The proposed architecture is more scalable and proposed algorithm is robust against even/uneven node deployment and node mobility.

preprint2012arXiv

LNOS - Live Network Operating System

Operating Systems exists since existence of computers, and have been evolving continuously from time to time. In this paper we have reviewed a relatively new or unexplored topic of Live OS. From networking perspective, Live OS is used for establishing Clusters, Firewalls and as Network security assessment tool etc. Our proposed concept is that a Live OS can be established or configured for an organizations specific network requirements with respect to their servers. An important server failure due to hardware or software could take time for remedy of the problem, so for that situation a preconfigured server in the form of Live OS on CD/DVD/USB can be used as an immediate solution. In a network of ten nodes, we stopped the server machine and with necessary adjustments, Live OS replaced the server in less than five minutes. Live OS in a network environment is a quick replacement of the services that are failed due to server failure (hardware or software). It is a cost effective solution for low budget networks. The life of Live OS starts when we boot it from CD/DVD/USB and remains in action for that session. As soon as the machine is rebooted, any work done for that session is gone, (in case we do not store any information on permanent storage media). Live CD/DVD/USB is normally used on systems where we do not have Operating Systems installed. A Live OS can also be used on systems where we already have an installed OS. On the basis of functionality a Live OS can be used for many purposes and has some typical advantages that are not available on other operating systems. Vendors are releasing different distributions of Live OS and is becoming their sole identity in a particular domain like Networks, Security, Education or Entertainment etc. There can be many aspects of Live OS, but Linux based Live OS and their use in the field of networks is the main focus of this paper.

preprint2012arXiv

Personality wireless sensor networks (PWSNs)

In recent years, WSNs are garnering lot of interest from research community because of their unique characteristics and potential for enormous range of applications. Envision for new class of applications are being emerged such as human augmentation, enhancing social interaction etc. Misunderstanding or misinterpretation of behaviors from individuals leads to social conflicts. There are various theories that classify people into different personality types. Most of the existing theories rely on questionnaires, which is highly unreliable. Anyone can lead such theories in practice to incorrect classification intentionally or unintentionally. The objective of this research is to investigate existing solutions and propose a basic infrastructure for an automated context-aware psychological classification based on different parameters. The idea is to use wearable sensors to sense and measure various human body parameters (i.e. body temperature, blood pressure, perspiration, brain impulses etc) that coerce human psychological condition. The data collected from these parameters is transformed in to information, to determine personality type, mood and psychological condition of interacting parties. This information is shared among counterparts to better understand each other in order to avoid potential conflicting situations. We believe that it will help peoples understand each other, improve their quality of life and minimize possible conflicting situations.

preprint2011arXiv

Asymmetric propagation of electronic wave function through molecular bonding and anti-bonding states

Electron transport through molecular bridge shows novel quantum features. Propogation of electronic wave function through molecular bridge is completely different than individual atomic bridge employed between two contacts. In case of molecular bridge electronic wave propagators interfere and effect conduction through molecular bonding and anti-bonding states.In the present work i showed through simple calculation that interference of electronic wave propagators cause asymmetric propagation of electronic wave through bonding and anti-bonding state. While for hydrogenic molecule these propagators interfere completely destructively for bonding state and constructively for anti-bonding state, giving rise to only one peak in spectral function for anti- bonding state.