Source author record

Francois Rivest

Francois Rivest appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

3works
3topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

3 published item(s)

preprint2022arXiv

Interval Timing: Modeling the break-run-break pattern using start/stop threshold-less drift-diffusion model

Animal interval timing is often studied through the peak interval (PI) procedure. In this procedure, the animal is rewarded for the first response after a fixed delay from the stimulus onset, but on some trials, the stimulus remains and no reward is given. The common methods and models to analyse the response pattern describe it as break-run-break, a period of low rate response followed by rapid responding, followed by a low rate of response. The study of the pattern has found correlations between start, stop, and duration of the run period that hold across species and experiment. It is commonly assumed that in order to achieve the statistics with a pacemaker accumulator model it is necessary to have start and stop thresholds. In this paper we will develop a new model that varies response rate in relation to the likelihood of event occurrence, as opposed to a threshold, for changing the response rate. The new model reproduced the start and stop statistics that have been observed in 14 different PI experiments from 3 different papers. The developed model is also compared to the Time-adaptive Drift-diffusion Model (TDDM), the latest accumulator model subsuming the scalar expectancy theory (SET), on all 14 data-sets. The results show that it is unnecessary to have explicit start and stop thresholds or an internal equivalent to break-run-break states to reproduce the individual trials statistics and population behaviour and get the same break-run-break analysis results. The new model also produces more realistic individual trials compared to TDDM.

preprint2022arXiv

Towards Personalization of User Preferences in Partially Observable Smart Home Environments

The technologies used in smart homes have recently improved to learn the user preferences from feedback in order to enhance the user convenience and quality of experience. Most smart homes learn a uniform model to represent the thermal preferences of users, which generally fails when the pool of occupants includes people with different sensitivities to temperature, for instance due to age and physiological factors. Thus, a smart home with a single optimal policy may fail to provide comfort when a new user with a different preference is integrated into the home. In this paper, we propose a Bayesian Reinforcement learning framework that can approximate the current occupant state in a partially observable smart home environment using its thermal preference, and then identify the occupant as a new user or someone is already known to the system. Our proposed framework can be used to identify users based on the temperature and humidity preferences of the occupant when performing different activities to enable personalization and improve comfort. We then compare the proposed framework with a baseline long short-term memory learner that learns the thermal preference of the user from the sequence of actions which it takes. We perform these experiments with up to 5 simulated human models each based on hierarchical reinforcement learning. The results show that our framework can approximate the belief state of the current user just by its temperature and humidity preferences across different activities with a high degree of accuracy.

preprint2011arXiv

Adaptive Drift-Diffusion Process to Learn Time Intervals

Animals learn the timing between consecutive events very easily. Their precision is usually proportional to the interval to time (Weber's law for timing). Most current timing models either require a central clock and unbounded accumulator or whole pre-defined populations of delay lines, decaying traces or oscillators to represent elapsing time. Current adaptive recurrent neural networks fail at learning to predict the timing of future events (the 'when') in a realistic manner. In this paper, we present a new model of interval timing, based on simple temporal integrators, derived from drift-diffusion models. We develop a simple geometric rule to learn 'when' instead of 'what'. We provide an analytical proof that the model can learn inter-event intervals in a number of trials independent of the interval size and that the temporal precision of the system is proportional to the timed interval. This new model uses no clock, no gradient, no unbounded accumulators, no delay lines, and has internal noise allowing generations of individual trials. Three interesting predictions are made.