Source author record

Michael Schröder

Michael Schröder appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2022arXiv

Grammars for Free: Toward Grammar Inference for Ad Hoc Parsers

Ad hoc parsers are everywhere: they appear any time a string is split, looped over, interpreted, transformed, or otherwise processed. Every ad hoc parser gives rise to a language: the possibly infinite set of input strings that the program accepts without going wrong. Any language can be described by a formal grammar: a finite set of rules that can generate all strings of that language. But programmers do not write grammars for ad hoc parsers -- even though they would be eminently useful. Grammars can serve as documentation, aid program comprehension, generate test inputs, and allow reasoning about language-theoretic security. We propose an automatic grammar inference system for ad hoc parsers that would enable all of these use cases, in addition to opening up new possibilities in mining software repositories and bi-directional parser synthesis.

preprint2021arXiv

An Empirical Investigation of Command-Line Customization

The interactive command line, also known as the shell, is a prominent mechanism used extensively by a wide range of software professionals (engineers, system administrators, data scientists, etc.). Shell customizations can therefore provide insight into the tasks they repeatedly perform, how well the standard environment supports those tasks, and ways in which the environment could be productively extended or modified. To characterize the patterns and complexities of command-line customization, we mined the collective knowledge of command-line users by analyzing more than 2.2 million shell alias definitions found on GitHub. Shell aliases allow command-line users to customize their environment by defining arbitrarily complex command substitutions. Using inductive coding methods, we found three types of aliases that each enable a number of customization practices: Shortcuts (for nicknaming commands, abbreviating subcommands, and bookmarking locations), Modifications (for substituting commands, overriding defaults, colorizing output, and elevating privilege), and Scripts (for transforming data and chaining subcommands). We conjecture that identifying common customization practices can point to particular usability issues within command-line programs, and that a deeper understanding of these practices can support researchers and tool developers in designing better user experiences. In addition to our analysis, we provide an extensive reproducibility package in the form of a curated dataset together with well-documented computational notebooks enabling further knowledge discovery and a basis for learning approaches to improve command-line workflows.

preprint2021arXiv

Improving Efficient Neural Ranking Models with Cross-Architecture Knowledge Distillation

Retrieval and ranking models are the backbone of many applications such as web search, open domain QA, or text-based recommender systems. The latency of neural ranking models at query time is largely dependent on the architecture and deliberate choices by their designers to trade-off effectiveness for higher efficiency. This focus on low query latency of a rising number of efficient ranking architectures make them feasible for production deployment. In machine learning an increasingly common approach to close the effectiveness gap of more efficient models is to apply knowledge distillation from a large teacher model to a smaller student model. We find that different ranking architectures tend to produce output scores in different magnitudes. Based on this finding, we propose a cross-architecture training procedure with a margin focused loss (Margin-MSE), that adapts knowledge distillation to the varying score output distributions of different BERT and non-BERT passage ranking architectures. We apply the teachable information as additional fine-grained labels to existing training triples of the MSMARCO-Passage collection. We evaluate our procedure of distilling knowledge from state-of-the-art concatenated BERT models to four different efficient architectures (TK, ColBERT, PreTT, and a BERT CLS dot product model). We show that across our evaluated architectures our Margin-MSE knowledge distillation significantly improves re-ranking effectiveness without compromising their efficiency. Additionally, we show our general distillation method to improve nearest neighbor based index retrieval with the BERT dot product model, offering competitive results with specialized and much more costly training methods. To benefit the community, we publish the teacher-score training files in a ready-to-use package.

preprint2020arXiv

Fine-Grained Relevance Annotations for Multi-Task Document Ranking and Question Answering

There are many existing retrieval and question answering datasets. However, most of them either focus on ranked list evaluation or single-candidate question answering. This divide makes it challenging to properly evaluate approaches concerned with ranking documents and providing snippets or answers for a given query. In this work, we present FiRA: a novel dataset of Fine-Grained Relevance Annotations. We extend the ranked retrieval annotations of the Deep Learning track of TREC 2019 with passage and word level graded relevance annotations for all relevant documents. We use our newly created data to study the distribution of relevance in long documents, as well as the attention of annotators to specific positions of the text. As an example, we evaluate the recently introduced TKL document ranking model. We find that although TKL exhibits state-of-the-art retrieval results for long documents, it misses many relevant passages.

preprint2011arXiv

Patterson--Sullivan distributions in higher rank

For a compact locally symmetric space $\XG$ of non-positive curvature, we consider sequences of normalized joint eigenfunctions which belong to the principal spectrum of the algebra of invariant differential operators. Using an $h$-\psdiff\ calculus on $\XG$, we define and study lifted quantum limits as weak$^*$-limit points of Wigner distributions. The Helgason boundary values of the eigenfunctions allow us to construct Patterson--Sullivan distributions on the space of Weyl chambers. These distributions are asymptotic to lifted quantum limits and satisfy additional invariance properties, which makes them useful in the context of quantum ergodicity. Our results generalize results for compact hyperbolic surfaces obtained by Anantharaman and Zelditch.

preprint2002arXiv

Analytical ramifications of derivatives valuation: Asian options and special functions

Averaging problems are ubiquitous in Finance with the valuation of the so-called Asian options on arithmetic averages as their most conspicuous form. There is an abundance of numerical work on them, and their stochastic structure has been extensively studied by Yor and his school. However, the analytical structure of these problems is largely unstudied. Our philosophy now is that such valuation problems should be considered as an extension of the theory of special functions: they lead to new problems about new classes of special functions which should be studied in terms of and using of the methods of special functions and their theory. This is exemplified by deriving integral representations for the Black-Scholes prices based on Yor's Laplace transform ansatz to their valuation. They are obtained by analytic Laplace inversion using complex analytic methods. The analysis ultimately rests on the gamma function which in this sense is found to be at the base of Asian options. The results improve on those of Yor and have served us a as starting point for deriving first time benchmark prices for these options.

preprint2002arXiv

Brownian excursions an Parisian barrier options: a note

This note re-addresses the Paris barrier options proposed by Yor and collaborators and their valuation using the Laplace transform approach. The notion of Paris barrier options, based on excursion theory and using the Brownian meander, is extended such that their valuation is now possible at any point during their lifespan. The pertinent Laplace transforms are modified when necessary.

preprint2001arXiv

On the valuation of arithmetic-average Asian options: the Geman-Yor Laplace transform revisited

The 1993 Laplace transform approach of Geman and Yor is a celebrated advance in valuing Asian options. Its insights are fundamental from both a mathematical and a financial perspective. In this paper, we discuss two observations regarding the financial relevance of its results. First, we show that the Geman and Yor Laplace transform is not that of an Asian option price, as reported in Geman and Yor and other papers. We nonetheless show how the Geman and Yor Laplace transform can be used to obtain the price of an Asian option. Second, we find that following Geman and Yor these Laplace transfoms are available only if the risk-neutral drift is not less than half the squared volatility. Using complex analytic techniques, we lift this restriction, thus extending the financial applicability of the Laplace transform approach.

preprint2000arXiv

On the valuation of arithmetic-average Asian options: Laguerre series and Theta integrals

In a recent significant advance, using Laguerre series, the valuation of Asian options has been reduced by Dufresne to computing the negative moments of Yor's accumulation processes. For these he has given functional recursion rules whose probabilistic structure has been the object of intensive recent studies of Yor and co-workers. Stressing the role of Theta functions, this paper now solves these recursion rules and expresses these negative moments as linear combinations of certain Theta integrals. Using the Jacobi transformation formula, very rapidly and very stably convergent series for them are derived. In this way computable series for Black--Scholes price of the Asian option result which are numerically illustrated. Moreover, the Laguerre series approach of Dufresne is made rigorous, and extensions and modifications are discussed. The key for this is the analysis of the integrability and growth properties of Yor's 1992 Asia density, basic problems which seem to be addressed here for the first time.

preprint2000arXiv

On the valuation of Asian options: integral representations

This paper derives integral representations for the Black-Scholes price of arithmetic-average Asian options. Their proof is by Laplace inverting the 1992 Laplace transform of Geman-Yor using complex analytic methods. The analysis ultimately rests on the gamma function which in this sense is at the base of Asian options. The results of Geman-Yor are corrected and their validitity is extended.

preprint2000arXiv

On the valuation of Paris options: foundational results

This paper adresses the valuation of the Paris barrier options proposed by Yor, Jeanblanc-Picque, and Chesnay (Advances in Applied Probability, 29(1997), 165-184) using the Laplace transform approach. Based on suggestions by Pliska the notion of Paris options is extended such that their valuation is possible at any point during their lifespan. The Laplace transforms derived by Yor et al. are modified when necessary, and their basic analytic properties are discussed.