Source author record

Kyle Luther

Kyle Luther appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2022arXiv

Kernel similarity matching with Hebbian neural networks

Recent works have derived neural networks with online correlation-based learning rules to perform \textit{kernel similarity matching}. These works applied existing linear similarity matching algorithms to nonlinear features generated with random Fourier methods. In this paper attempt to perform kernel similarity matching by directly learning the nonlinear features. Our algorithm proceeds by deriving and then minimizing an upper bound for the sum of squared errors between output and input kernel similarities. The construction of our upper bound leads to online correlation-based learning rules which can be implemented with a 1 layer recurrent neural network. In addition to generating high-dimensional linearly separable representations, we show that our upper bound naturally yields representations which are sparse and selective for specific input patterns. We compare the approximation quality of our method to neural random Fourier method and variants of the popular but non-biological "Nystr{ö}m" method for approximating the kernel matrix. Our method appears to be comparable or better than randomly sampled Nystr{ö}m methods when the outputs are relatively low dimensional (although still potentially higher dimensional than the inputs) but less faithful when the outputs are very high dimensional.

preprint2022arXiv

Sensitivity of sparse codes to image distortions

Sparse coding has been proposed as a theory of visual cortex and as an unsupervised algorithm for learning representations. We show empirically with the MNIST dataset that sparse codes can be very sensitive to image distortions, a behavior that may hinder invariant object recognition. A locally linear analysis suggests that the sensitivity is due to the existence of linear combinations of active dictionary elements with high cancellation. A nearest neighbor classifier is shown to perform worse on sparse codes than original images. For a linear classifier with a sufficiently large number of labeled examples, sparse codes are shown to yield higher accuracy than original images, but no higher than a representation computed by a random feedforward net. Sensitivity to distortions seems to be a basic property of sparse codes, and one should be aware of this property when applying sparse codes to invariant object recognition.

preprint2022arXiv

Stacked unsupervised learning with a network architecture found by supervised meta-learning

Stacked unsupervised learning (SUL) seems more biologically plausible than backpropagation, because learning is local to each layer. But SUL has fallen far short of backpropagation in practical applications, undermining the idea that SUL can explain how brains learn. Here we show an SUL algorithm that can perform completely unsupervised clustering of MNIST digits with comparable accuracy relative to unsupervised algorithms based on backpropagation. Our algorithm is exceeded only by self-supervised methods requiring training data augmentation by geometric distortions. The only prior knowledge in our unsupervised algorithm is implicit in the network architecture. Multiple convolutional "energy layers" contain a sum-of-squares nonlinearity, inspired by "energy models" of primary visual cortex. Convolutional kernels are learned with a fast minibatch implementation of the K-Subspaces algorithm. High accuracy requires preprocessing with an initial whitening layer, representations that are less sparse during inference than learning, and rescaling for gain control. The hyperparameters of the network architecture are found by supervised meta-learning, which optimizes unsupervised clustering accuracy. We regard such dependence of unsupervised learning on prior knowledge implicit in network architecture as biologically plausible, and analogous to the dependence of brain architecture on evolutionary history.

preprint2015arXiv

Characterizing transiting exoplanet atmospheres with JWST

We explore how well James Webb Space Telescope (JWST) spectra will likely constrain bulk atmospheric properties of transiting exoplanets. We start by modeling the atmospheres of archetypal hot Jupiter, warm Neptune, warm sub-Neptune, and cool super-Earth planets with clear, cloudy, or high mean molecular weight atmospheres. Next we simulate the $λ= 1 - 11$ $μ$m transmission and emission spectra of these systems for several JWST instrument modes for single transit and eclipse events. We then perform retrievals to determine how well temperatures and molecular mixing ratios (CH$_4$, CO, CO$_2$, H$_2$O, NH$_3$) can be constrained. We find that $λ= 1 - 2.5$ $μ$m transmission spectra will often constrain the major molecular constituents of clear solar composition atmospheres well. Cloudy or high mean molecular weight atmospheres will often require full $1 - 11$ $μ$m spectra for good constraints, and emission data may be more useful in cases of sufficiently high $F_p$ and high $F_p/F_*$. Strong temperature inversions in the solar composition hot Jupiter atmosphere should be detectable with $1 - 2.5+$ $μ$m emission spectra, and $1 - 5+$ $μ$m emission spectra will constrain the temperature-pressure profiles of warm planets. Transmission spectra over $1 - 5+$ $μ$m will constrain [Fe/H] values to better than 0.5 dex for the clear atmospheres of the hot and warm planets studied. Carbon-to-oxygen ratios can be constrained to better than a factor of 2 in some systems. We expect that these results will provide useful predictions of the scientific value of single event JWST spectra until its on-orbit performance is known.