Source author record

Francisco Louzada

Francisco Louzada appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

19works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

19 published item(s)

preprint2020arXiv

Multiple repairable systems under dependent competing risks with nonparametric Frailty

The aim of this article is to analyze data from multiple repairable systems under the presence of dependent competing risks. In order to model this dependence structure, we adopted the well-known shared frailty model. This model provides a suitable theoretical basis for generating dependence between the components failure times in the dependent competing risks model. It is known that the dependence effect in this scenario influences the estimates of the model parameters. Hence, under the assumption that the cause-specific intensities follow a PLP, we propose a frailty-induced dependence approach to incorporate the dependence among the cause-specific recurrent processes. Moreover, the misspecification of the frailty distribution may lead to errors when estimating the parameters of interest. Because of this, we considered a Bayesian nonparametric approach to model the frailty density in order to offer more flexibility and to provide consistent estimates for the PLP model, as well as insights about heterogeneity among the systems. Both simulation studies and real case studies are provided to illustrate the proposed approaches and demonstrate their validity.

preprint2020arXiv

Power laws distributions in objective priors

The use of objective prior in Bayesian applications has become a common practice to analyze data without subjective information. Formal rules usually obtain these priors distributions, and the data provide the dominant information in the posterior distribution. However, these priors are typically improper and may lead to improper posterior. Here, we show, for a general family of distributions, that the obtained objective priors for the parameters either follow a power-law distribution or has an asymptotic power-law behavior. As a result, we observed that the exponents of the model are between 0.5 and 1. Understand these behaviors allow us to easily verify if such priors lead to proper or improper posteriors directly from the exponent of the power-law. The general family considered in our study includes essential models such as Exponential, Gamma, Weibull, Nakagami-m, Haf-Normal, Rayleigh, Erlang, and Maxwell Boltzmann distributions, to list a few. In summary, we show that comprehending the mechanisms describing the shapes of the priors provides essential information that can be used in situations where additional complexity is presented.

preprint2020arXiv

Power laws in the Roman Empire: a survival analysis

The Roman Empire shaped Western civilization, and many Roman principles are embodied in modern institutions. Although its political institutions proved both resilient and adaptable, allowing it to incorporate diverse populations, the Empire suffered from many internal conflicts. Indeed, most emperors died violently, from assassination, suicide, or in battle. These internal conflicts produced patterns in the length of time that can be identified by statistical analysis. In this paper, we study the underlying patterns associated with the reign of the Roman emperors by using statistical tools of survival data analysis. We consider all the 175 Roman emperors and propose a new power-law model with change points to predict the time-to-violent-death of the Roman emperors. This model encompasses data in the presence of censoring and long-term survivors, providing more accurate predictions than previous models. Our results show that power-law distributions can also occur in survival data, as verified in other data types from natural and artificial systems, reinforcing the ubiquity of power law distributions. The generality of our approach paves the way to further related investigations not only in other ancient civilizations but also in applications in engineering and medicine.

preprint2020arXiv

Random Machines Regression Approach: an ensemble support vector regression model with free kernel choice

Machine learning techniques always aim to reduce the generalized prediction error. In order to reduce it, ensemble methods present a good approach combining several models that results in a greater forecasting capacity. The Random Machines already have been demonstrated as strong technique, i.e: high predictive power, to classification tasks, in this article we propose an procedure to use the bagged-weighted support vector model to regression problems. Simulation studies were realized over artificial datasets, and over real data benchmarks. The results exhibited a good performance of Regression Random Machines through lower generalization error without needing to choose the best kernel function during tuning process.

preprint2016arXiv

Classification methods applied to credit scoring: A systematic review and overall comparison

The need for controlling and effectively managing credit risk has led financial institutions to excel in improving techniques designed for this purpose, resulting in the development of various quantitative models by financial institutions and consulting companies. Hence, the growing number of academic studies about credit scoring shows a variety of classification methods applied to discriminate good and bad borrowers. This paper, therefore, aims to present a systematic literature review relating theory and application of binary classification techniques for credit scoring financial analysis. The general results show the use and importance of the main techniques for credit rating, as well as some of the scientific paradigm changes throughout the years.

preprint2016arXiv

Cooperative Parallel Particle Filters for online model selection and applications to Urban Mobility

We design a sequential Monte Carlo scheme for the dual purpose of Bayesian inference and model selection. We consider the application context of urban mobility, where several modalities of transport and different measurement devices can be employed. Therefore, we address the joint problem of online tracking and detection of the current modality. For this purpose, we use interacting parallel particle filters, each one addressing a different model. They cooperate for providing a global estimator of the variable of interest and, at the same time, an approximation of the posterior density of each model given the data. The interaction occurs by a parsimonious distribution of the computational effort, with online adaptation for the number of particles of each filter according to the posterior probability of the corresponding model. The resulting scheme is simple and flexible. We have tested the novel technique in different numerical experiments with artificial and real data, which confirm the robustness of the proposed scheme.

preprint2016arXiv

MWStat: A Modulated Web-Based Statistical System

In this paper we present the development of a modulated web based statistical system, hereafter MWStat, which shifts the statistical paradigm of analyzing data into a real time structure. The MWStat system is useful for both online storage data and questionnaires analysis, as well as to provide real time disposal of results from analysis related to several statistical methodologies in a customizable fashion. Overall, it can be seem as a useful technical solution that can be applied to a large range of statistical applications, which needs of a scheme of devolution of real time results, accessible to anyone with internet access. We display here the step-by-step instructions for implementing the system. The structure is accessible, built with an easily interpretable language and it can be strategically applied to online statistical applications. We rely on the relationship of several free languages, namely, PHP, R, MySQL database and an Apache HTTP server, and on the use of software tools such as phpMyAdmin. We expose three didactical examples of the MWStat system on institutional evaluation, statistical quality control and multivariate analysis. The methodology is also illustrated in a real example on institutional evaluation.

preprint2016arXiv

Objective Bayesian Analysis for the Lomax Distribution

In this paper we propose to make Bayesian inferences for the parameters of the Lomax distribution using non-informative priors, namely the Jeffreys prior and the reference prior. We assess Bayesian estimation through a Monte Carlo study with 500 simulated data sets. To evaluate the possible impact of prior specification on estimation, two criteria were considered: the bias and square root of the mean square error. The developed procedures are illustrated on a real data set.

preprint2015arXiv

Maximum Likelihood Estimation for the Weight Lindley Distribution Parameters under Different Types of Censoring

In this paper the maximum likelihood equations for the parameters of the Weight Lindley distribution are studied considering different types of censoring, such as, type I, type II and random censoring mechanism. A numerical simulation study is perform to evaluate the maximum likelihood estimates. The proposed methodology is illustrated in a real data set.

preprint2015arXiv

Modeling Compositional Regression with uncorrelated and correlated errors: a Bayesian approach

Compositional data consist of known compositions vectors whose components are positive and defined in the interval (0,1) representing proportions or fractions of a "whole". The sum of these components must be equal to one. Compositional data is present in different knowledge areas, as in geology, economy, medicine among many others. In this paper, we introduce a Bayesian analysis for compositional regression applying additive log-ratio (ALR) transformation and assuming uncorrelated and correlated errors. The Bayesian inference procedure based on Markov Chain Monte Carlo Methods (MCMC). The methodology is illustrated on an artificial and a real data set of volleyball.

preprint2015arXiv

The Optimised Theta Method

Accurate and robust forecasting methods for univariate time series are very important when the objective is to produce estimates for a large number of time series. In this context, the Theta method called researchers attention due its performance in the largest up-to-date forecasting competition, the M3-Competition. Theta method proposes the decomposition of the deseasonalised data into two "theta lines". The first theta line removes completely the curvatures of the data, thus being a good estimator of the long-term trend component. The second theta line doubles the curvatures of the series, as to better approximate the short-term behaviour. In this paper, we propose a generalisation of the Theta method by optimising the selection of the second theta line, based on various validation schemes where the out-of-sample accuracy of the candidate variants is measured. The recomposition process of the original time series builds on the asymmetry of the decomposed theta lines. An empirical investigation through the M3-Competition data set shows improvements on the forecasting accuracy of the proposed optimised Theta method.

preprint2015arXiv

The zero-inflated cure rate regression model: Applications to fraud detection in bank loan portfolios

In this paper, we introduce a methodology based on the zero-inflated cure rate model to detect fraudsters in bank loan applications. In fact, our approach enables us to accommodate three different types of loan applicants, i.e., fraudsters, those who are susceptible to default and finally, those who are not susceptible to default. An advantage of our approach is to accommodate zero-inflated times, which is not possible in the standard cure rate model. To illustrate the proposed method, a real dataset of loan survival times is fitted by the zero-inflated Weibull cure rate model. The parameter estimation is reached by maximum likelihood estimation procedure and Monte Carlo simulations are carried out to check its finite sample performance.

preprint2015arXiv

The zero-inflated promotion cure rate regression model applied to fraud propensity in bank loan applications

In this paper we extend the promotion cure rate model proposed by Chen et al (1999), by incorporating excess of zeros in the modelling. Despite allowing to relate the covariates to the fraction of cure, the current approach, which is based on a biological interpretation of the causes that trigger the event of interest, does not enable to relate the covariates to the fraction of zeros. The presence of zeros in survival data, unusual in medical studies, can frequently occur in banking loan portfolios, as presented in Louzada et al (2015), where they deal with propensity to fraud in lending loans in a major Brazilian bank. To illustrate the new cure rate survival method, the same real dataset analyzed in Louzada et al (2015) is fitted here, and the results are compared.

preprint2014arXiv

A Modified Reference Prior for the Generalized Gamma Distribution

In this paper we propose an objective Bayesian estimation approach for the parameters of the generalized gamma distribution. Various reference priors are obtained, but showing that they lead to improper posterior distributions. We overcome this problem by proposing a modification in a reference priori distribution, allowing for a proper posterior distribution for the parameters of the generalized gamma distribution. We perform a simulation study in order to study the efficiency of the proposed methodology, which is also fully illustrated on a real data set.

preprint2014arXiv

A modified version of the inference function for margins and interval estimation for the bivariate Clayton copula SUR Tobit model: An simulation approach

This paper extends the analysis of bivariate seemingly unrelated regression (SUR) Tobit model by modeling its nonlinear dependence structure through the Clayton copula. The ability in capturing/modeling the lower tail dependence of the SUR Tobit model where some data are censored (generally, at zero point) is an additionally useful feature of the Clayton copula. We propose a modified version of the inference function for margins (IFM) method (Joe and Xu, 1996), which we refer to as MIFM method, to obtain the estimates of the marginal parameters and a better (satisfactory) estimate of the copula association parameter. More specifically, we employ the data augmentation technique in the second stage of the IFM method to generate the censored observations (i.e. to obtain continuous marginal distributions, which ensures the uniqueness of the copula) and then estimate the dependence parameter. Resampling procedures (bootstrap methods) are also proposed for obtaining confidence intervals for the model parameters. A simulation study is performed in order to verify the behavior of the MIFM estimates (we focus on the copula parameter estimation) and the coverage probability of different confidence intervals in datasets with different percentages of censoring and degrees of dependence. The satisfactory results from the simulation (under certain conditions) and empirical study indicate the good performance of our proposed model and methods where they are applied to model the U.S. ready-to-eat breakfast cereals and fluid milk consumption data.

preprint2014arXiv

An Evidence of Link between Default and Loss of Bank Loans from the Modeling of Competing Risks

In this paper, we propose a method that provides a useful technique to compare relationship between risks involved that takes customer become defaulter and debt collection process that might make this defaulter recovered. Through estimation of competitive risks that lead to realization of the event of interest, we showed that there is a significant relation between the intensity of default and losses from defaulted loans in collection processes. To reach this goal, we investigate a competing risks model applied to whole credit risk cycle into a bank loans portfolio. We estimated competing causes related to occurrence of default, thereafter, comparing it with estimated competing causes that lead loans to write-off condition. In context of modeling competing risks, we used a specification of Poisson distribution for numbers from competing causes and Weibull distribution for failures times. The likelihood maximum estimation is used to parameters estimation and the model is applied to a real data of personal loans

preprint2014arXiv

Analyzing Volleyball Data on a Compositional Regression Model Approach: An Application to the Brazilian Men's Volleyball Super League 2011/2012 Data

Volleyball has become a competitive sport with high physical and technical performance. Matches results are based on the players and teams'skills as technical and tactical strategies to succeed in a championship. At this point, some studies are carried out on the performance analysis of different match elements, contributing to the development of this sport. In this paper, we proposed a new approach to analyze volleyball data. The study is based on the compositional data methodology modeling in regression model. The parameters are obtained through the maximum likelihood. We performed a simulation study to evaluate the estimation procedure in compositional regression model and we illustrated the proposed methodology considering real data set of volleyball.

preprint2014arXiv

BayesDccGarch - An Implementation of Multivariate GARCH DCC Models

Multivariate GARCH models are important tools to describe the dynamics of multivariate times series of financial returns. Nevertheless, these models have been much less used in practice due to the lack of reliable software. This paper describes the {\tt R} package {\bf BayesDccGarch} which was developed to implement recently proposed inference procedures to estimate and compare multivariate GARCH models allowing for asymmetric and heavy tailed distributions.

preprint2014arXiv

Recovery Risk: Application of the Latent Competing Risks Model to Non performing Loans

This article proposes a method for measuring the latent risks involved in the recovery process of non performing loans in financial institutions and business firms that deal with collection and recovery processes. To that end, we apply the competing risks model referred to in the literature as the promotion time model. The result achieved is the probability of credit recovery for a portfolio segmented into groups based on the information available. Within the context of competing risks, application of the technique yielded an estimation of the number of latent events that concur to the credit recovery event. With these results in hand, we were able to compare groups of defaulters in terms of risk or susceptibility to the recovery event during the collection process, and thereby determine where collection actions are most efficient. We specify the Poisson distribution for the number of latent causes leading to recovery, and the Weibull distribution for the time up to recovery. To estimate the model parameters, we use the maximum likelihood method. Finally, the model was applied to a sample of defaulted loans from a financial institution.