Source author record

Aaron Cohen

Aaron Cohen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2022arXiv

LaMDA: Language Models for Dialog Applications

We present LaMDA: Language Models for Dialog Applications. LaMDA is a family of Transformer-based neural language models specialized for dialog, which have up to 137B parameters and are pre-trained on 1.56T words of public dialog data and web text. While model scaling alone can improve quality, it shows less improvements on safety and factual grounding. We demonstrate that fine-tuning with annotated data and enabling the model to consult external knowledge sources can lead to significant improvements towards the two key challenges of safety and factual grounding. The first challenge, safety, involves ensuring that the model's responses are consistent with a set of human values, such as preventing harmful suggestions and unfair bias. We quantify safety using a metric based on an illustrative set of human values, and we find that filtering candidate responses using a LaMDA classifier fine-tuned with a small amount of crowdworker-annotated data offers a promising approach to improving model safety. The second challenge, factual grounding, involves enabling the model to consult external knowledge sources, such as an information retrieval system, a language translator, and a calculator. We quantify factuality using a groundedness metric, and we find that our approach enables the model to generate responses grounded in known sources, rather than responses that merely sound plausible. Finally, we explore the use of LaMDA in the domains of education and content recommendations, and analyze their helpfulness and role consistency.

preprint2016arXiv

Data Integration Model for Air Quality: A Hierarchical Approach to the Global Estimation of Exposures to Ambient Air Pollution

Air pollution is a major risk factor for global health, with both ambient and household air pollution contributing substantial components of the overall global disease burden. One of the key drivers of adverse health effects is fine particulate matter ambient pollution (PM$_{2.5}$) to which an estimated 3 million deaths can be attributed annually. The primary source of information for estimating exposures has been measurements from ground monitoring networks but, although coverage is increasing, there remain regions in which monitoring is limited. Ground monitoring data therefore needs to be supplemented with information from other sources, such as satellite retrievals of aerosol optical depth and chemical transport models. A hierarchical modelling approach for integrating data from multiple sources is proposed allowing spatially-varying relationships between ground measurements and other factors that estimate air quality. Set within a Bayesian framework, the resulting Data Integration Model for Air Quality (DIMAQ) is used to estimate exposures, together with associated measures of uncertainty, on a high resolution grid covering the entire world. Bayesian analysis on this scale can be computationally challenging and here approximate Bayesian inference is performed using Integrated Nested Laplace Approximations. Model selection and assessment is performed by cross-validation with the final model offering substantial increases in predictive accuracy, particularly in regions where there is sparse ground monitoring, when compared to current approaches: root mean square error (RMSE) reduced from 17.1 to 10.7, and population weighted RMSE from 23.1 to 12.1 $μ$gm$^{-3}$. Based on summaries of the posterior distributions for each grid cell, it is estimated that 92% of the world's population reside in areas exceeding the World Health Organization's Air Quality Guidelines.

preprint2011arXiv

An Ultra-Steep Spectrum Radio Relic in the Galaxy Cluster Abell 2443

We present newly discovered radio emission in the galaxy cluster Abell 2443 which is (1) diffuse, (2) extremely steep spectrum, (3) offset from the cluster center, (4) of irregular morphology and (5) not clearly associated with any of the galaxies within the cluster. The most likely explanation is that this emission is a cluster radio relic, associated with a cluster merger. We present deep observations of Abell 2443 at multiple low frequencies (1425, 325 and 74 MHz) which help characterize the spectrum and morphology of this relic. Based on the curved spectral shape of the relic emission and the presence of small scale structure, we suggest that this new source is likely a member of the radio phoenix class of radio relics.

preprint2010arXiv

The First Station of the Long Wavelength Array

The Long Wavelength Array (LWA) will be a new multi-purpose radio telescope operating in the frequency range 10-88 MHz. Upon completion, LWA will consist of 53 phased array "stations" distributed over a region about 400 km in diameter in the state of New Mexico. Each station will consist of 256 pairs of dipole-type antennas whose signals are formed into beams, with outputs transported to a central location for high-resolution aperture synthesis imaging. The resulting image sensitivity is estimated to be a few mJy (5 sigma, 8 MHz, 2 polarizations, 1 hr, zenith) in 20-80 MHz; with resolution and field of view of (8", 8 deg) and (2",2 deg) at 20 MHz and 80 MHz, respectively. All 256 dipole antennas are in place for the first station of the LWA (called LWA-1), and commissioning activities are well underway. The station is located near the core of the EVLA, and is expected to be fully operational in early 2011.