Source author record

Rémi Louf

Rémi Louf appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

10works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

10 published item(s)

preprint2020arXiv

HuggingFace's Transformers: State-of-the-art Natural Language Processing

Recent progress in natural language processing has been driven by advances in both model architecture and model pretraining. Transformer architectures have facilitated building higher-capacity models and pretraining has made it possible to effectively utilize this capacity for a wide variety of tasks. \textit{Transformers} is an open-source library with the goal of opening up these advances to the wider machine learning community. The library consists of carefully engineered state-of-the art Transformer architectures under a unified API. Backing this library is a curated collection of pretrained models made by and available for the community. \textit{Transformers} is designed to be extensible by researchers, simple for practitioners, and fast and robust in industrial deployments. The library is available at \url{https://github.com/huggingface/transformers}.

preprint2016arXiv

Patterns of residential segregation

The spatial distribution of income shapes the structure and organisation of cities and its understanding has broad societal implications. Despite an abundant literature, many issues remain unclear. In particular, all definitions of segregation are implicitely tied to a single indicator, usually rely on an ambiguous definition of income classes, without any consensus on how to define neighbourhoods and to deal with the polycentric organization of large cities. In this paper, we address all these questions within a unique conceptual framework. We avoid the challenge of providing a direct definition of segregation and instead start from a definition of what segregation is not. This naturally leads to the measure of representation that is able to identify locations where categories are over- or underrepresented. From there, we provide a new measure of exposure that discriminates between situations where categories co-locate or repel one another. We then use this feature to provide an unambiguous, parameter-free method to find meaningful breaks in the income distribution, thus defining classes. Applied to the 2014 American Community Survey, we find 3 emerging classes -- low, middle and higher income -- out of the original 16 income categories. The higher-income households are proportionally more present in larger cities, while lower-income households are not, invalidating the idea of an increased social polarisation. Finally, using the density -- and not the distance to a center which is meaningless in polycentric cities -- we find that the richer class is overrepresented in high density zones, especially for larger cities. This suggests that density is a relevant factor for understanding the income structure of cities and might explain some of the differences observed between US and European cities.

preprint2015arXiv

Wandering in cities: a statistical physics approach to urban theory

The amount of data that is being gathered about cities is increasing in size and specificity. However, despite this wealth of information, we still have little understanding of what really drives the processes behind urbanisation. In this thesis we apply some ideas from statistical physics to the study of cities. We first present a stochastic, out-of-equilibrium model of city growth that describes the structure of the mobility pattern of individuals. The model explains the appearance of secondary subcenters as an effect of traffic congestion, and predicts a sublinear increase of the number of centers with population size. Within the framework of this model, we are further able to give a prediction for the scaling exponent of the total distance commuted daily, the total length of the road network, the total delay due to congestion, the quantity of CO2 emitted, and the surface area with the population size of cities. In the third part, we focus on the quantitative description of the patterns of residential segregation. We propose a unifying theoretical framework in which segregation can be empirically characterised. In the fourth and last part, we succinctly present the most important---theoretical and empirical---results of our studies on spatial networks. Throughout this thesis, we try to convey the idea that the complexity of cities is -- almost paradoxically -- better comprehended through simple approaches. Looking for structure in data, trying to isolate the most important processes, building simple models and only keeping those which agree with data, constitute a universal method that is also relevant to the study of urban systems.

preprint2014arXiv

A typology of street patterns

We propose a quantitative method to classify cities according to their street pattern. We use the conditional probability distribution of shape factor of blocks with a given area, and define what could constitute the `fingerprint' of a city. Using a simple hierarchical clustering method, these fingerprints can then serve as a basis for a typology of cities. We apply this method to a set of 131 cities in the world, and at an intermediate level of the dendrogram, we observe 4 large families of cities characterized by different abundances of blocks of a certain area and shape. At a lower level of the classification, we find that most European cities and American cities in our sample fall in their own sub-category, highlighting quantitatively the differences between the typical layouts of cities in both regions. We also show with the example of New York and its different Boroughs, that the fingerprint of a city can be seen as the sum of the ones characterising the different neighbourhoods inside a city. This method provides a quantitative comparison of urban street patterns, which could be helpful for a better understanding of the causes and mechanisms behind their distinct shapes.

preprint2014arXiv

How congestion shapes cities: from mobility patterns to scaling

The recent availability of data for cities has allowed scientists to exhibit scalings which present themselves in the form of a power-law dependence with population of various socio-economical and structural indicators. We propose here a dynamical, stochastic theory of urban growth which accounts for some of the observed scalings and we confirm these results on US and OECD empirical data. In particular, we show that the dependence with population size of the total number of miles driven daily, the total length of the road network, the total traffic delay, the total consumption of gasoline, the quantity of $CO_2$ emitted and the relation between area and population of cities, are all governed by a single parameter which characterizes the sensitivity to congestion. Finally, our results suggest that diseconomies associated with congestion scale superlinearly with population size, implying that, despite polycentrism, cities whose transportation infrastructure rely heavily on traffic sensitive modes are unsustainable.

preprint2014arXiv

Scaling in transportation networks

Subway systems span most large cities, and railway networks most countries in the world. These networks are fundamental in the development of countries and their cities, and it is therefore crucial to understand their formation and evolution. However, if the topological properties of these networks are fairly well understood, how they relate to population and socio-economical properties remains an open question. We propose here a general coarse-grained approach, based on a cost-benefit analysis that accounts for the scaling properties of the main quantities characterizing these systems (the number of stations, the total length, and the ridership) with the substrate's population, area and wealth. More precisely, we show that the length, number of stations and ridership of subways and rail networks can be estimated knowing the area, population and wealth of the underlying region. These predictions are in good agreement with data gathered for about $140$ subway systems and more than $50$ railway networks in the world. We also show that train networks and subway systems can be described within the same framework, but with a fundamental difference: while the interstation distance seems to be constant and determined by the typical walking distance for subways, the interstation distance for railways scales with the number of stations.

preprint2014arXiv

Scaling: Lost in the smog

In this commentary we discuss the validity of scaling laws and their relevance for understanding urban systems and helping policy makers. We show how the recent controversy about the scaling of CO2 transport-related emissions with population size, where different authors reach contradictory conclusions, is symptomatic of the lack of understanding of the underlying mechanisms. In particular, we highlight different sources of errors, ranging from incorrect estimate of CO2 to problems related with the definition of cities. We argue here that while data are necessary to build of a new science of cities, they are not enough: they have to go hand in hand with a theoretical understanding of the main processes. This effort of building models whose predictions agree with data is the prerequisite for a science of cities. In the meantime, policy advice are, at best, a shot in the dark.

preprint2014arXiv

Universal size effects for populations in group-outcome decision-making problems

Elections constitute a paradigm of decision-making problems that have puzzled experts of different disciplines for decades. We study two decision-making problems, where groups make decisions that impact only themselves as a group. In both studied cases, participation in local elections and the number of democratic representatives at different scales (from local to national), we observe a universal scaling with the constituency size. These results may be interpreted as constituencies having a hierarchical structure, where each group of $N$ agents, at each level of the hierarchy, is divided in about $N^δ$ subgroups with $δ\approx 1/3$. Following this interpretation, we propose a phenomenological model of vote participation where abstention is related to the perceived link of an agent to the rest of the constituency and which reproduces quantitatively the observed data.

preprint2013arXiv

Emergence of hierarchy in cost driven growth of spatial networks

One of the most important features of spatial networks such as transportation networks, power grids, Internet, neural networks, is the existence of a cost associated with the length of links. Such a cost has a profound influence on the global structure of these networks which usually display a hierarchical spatial organization. The link between local constraints and large-scale structure is however not elucidated and we introduce here a generic model for the growth of spatial networks based on the general concept of cost benefit analysis. This model depends essentially on one single scale and produces a family of networks which range from the star-graph to the minimum spanning tree and which are characterised by a continuously varying exponent. We show that spatial hierarchy emerges naturally, with structures composed of various hubs controlling geographically separated service areas, and appears as a large-scale consequence of local cost-benefit considerations. Our model thus provides the first building blocks for a better understanding of the evolution of spatial networks and their properties. We also find that, surprisingly, the average detour is minimal in the intermediate regime, as a result of a large diversity in link lengths. Finally, we estimate the important parameters for various world railway networks and find that --remarkably-- they all fall in this intermediate regime, suggesting that spatial hierarchy is a crucial feature for these systems and probably possesses an important evolutionary advantage.

preprint2013arXiv

Modeling the polycentric transition of cities

Empirical evidence suggest that most urban systems experience a transition from a monocentric to a polycentric organisation as they grow and expand. We propose here a stochastic, out-of-equilibrium model of the city which explains the appearance of subcenters as an effect of traffic congestion. We show that congestion triggers the unstability of the monocentric regime, and that the number of subcenters and the total commuting distance within a city scale sublinearly with its population, predictions which are in agreement with data gathered for around 9000 US cities between 1994 and 2010.