Source author record

Boris Bukh

Boris Bukh appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

20works
11topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

20 published item(s)

preprint2016arXiv

An improved bound on the fraction of correctable deletions

We consider codes over fixed alphabets against worst-case symbol deletions. For any fixed $k \ge 2$, we construct a family of codes over alphabet of size $k$ with positive rate, which allow efficient recovery from a worst-case deletion fraction approaching $1-\frac{2}{k+\sqrt k}$. In particular, for binary codes, we are able to recover a fraction of deletions approaching $1/(\sqrt 2 +1)=\sqrt 2-1 \approx 0.414$. Previously, even non-constructively the largest deletion fraction known to be correctable with positive rate was $1-Θ(1/\sqrt{k})$, and around $0.17$ for the binary case. Our result pins down the largest fraction of correctable deletions for $k$-ary codes as $1-Θ(1/k)$, since $1-1/k$ is an upper bound even for the simpler model of erasures where the locations of the missing symbols are known. Closing the gap between $(\sqrt 2 -1)$ and $1/2$ for the limit of worst-case deletions correctable by binary codes remains a tantalizing open question.

preprint2016arXiv

One-sided epsilon-approximants

Given a finite point set $P\subset\mathbb{R}^d$, we call a multiset $A$ a one-sided weak $\varepsilon$-approximant for $P$ (with respect to convex sets), if $|P\cap C|/|P|-|A\cap C|/|A|\leq\varepsilon$ for every convex set $C$. We show that, in contrast with the usual (two-sided) weak $\varepsilon$-approximants, for every set $P\subset \mathbb{R}^d$ there exists a one-sided weak $\varepsilon$-approximant of size bounded by a function of $\varepsilon$ and $d$.

preprint2016arXiv

Ranks of matrices with few distinct entries

An $L$-matrix is a matrix whose off-diagonal entries belong to a set $L$, and whose diagonal is zero. Let $N(r,L)$ be the maximum size of a square $L$-matrix of rank at most $r$. Many applications of linear algebra in extremal combinatorics involve a bound on $N(r,L)$. We review some of these applications, and prove several new results on $N(r,L)$. In particular, we classify the sets $L$ for which $N(r,L)$ is linear, and show that if $N(r,L)$ is superlinear and $L\subset \mathbb{Z}$, then $N(r,L)$ is at least quadratic. As a by-product of the work, we asymptotically determine the maximum multiplicity of an eigenvalue $λ$ in an adjacency matrix of a digraph of a given size.

preprint2015arXiv

Twins in words and long common subsequences in permutations

A large family of words must contain two words that are similar. We investigate several problems where the measure of similarity is the length of a common subsequence. We construct a family of n^{1/3} permutations on n letters, such that LCS of any two of them is only cn^{1/3}, improving a construction of Beame, Blais, and Huynh-Ngoc. We relate the problem of constructing many permutations with small LCS to the twin word problem of Axenovich, Person and Puzynina. In particular, we show that every word of length n over a k-letter alphabet contains two disjoint equal subsequences of length cnk^{-2/3}. Many problems are left open.

preprint2013arXiv

Erdos-Szekeres-type statements: Ramsey function and decidability in dimension 1

A classical and widely used lemma of Erdos and Szekeres asserts that for every n there exists N such that every N-term sequence a of real numbers contains an n-term increasing subsequence or an n-term nondecreasing subsequence; quantitatively, the smallest N with this property equals (n-1)^2+1. In the setting of the present paper, we express this lemma by saying that the set of predicates Phi={x_1<x_2,x_1\ge x_2}$ is Erdos-Szekeres with Ramsey function ES_Phi(n)=(n-1)^2+1. In general, we consider an arbitrary finite set Phi={Phi_1,...,Phi_m} of semialgebraic predicates, meaning that each Phi_j=Phi_j(x_1,...,x_k) is a Boolean combination of polynomial equations and inequalities in some number k of real variables. We define Phi to be Erdos-Szekeres if for every n there exists N such that each N-term sequence a of real numbers has an n-term subsequence b such that at least one of the Phi_j holds everywhere on b, which means that Phi_j(b_{i_1},...,b_{i_k}) holds for every choice of indices i_1,i_2,...,i_k, 1<=i_1<i_2<... <i_k<= n. We write ES_Phi(n) for the smallest N with the above property. We prove two main results. First, the Ramsey functions in this setting are at most doubly exponential (and sometimes they are indeed doubly exponential): for every Phi that is Erdős--Szekeres, there is a constant C such that ES_Phi(n) < exp(exp(Cn)). Second, there is an algorithm that, given Phi, decides whether it is Erdos-Szekeres; thus, one-dimensional Erdos-Szekeres-style theorems can in principle be proved automatically.

preprint2012arXiv

Turán numbers for $K_{s,t}$-free graphs: topological obstructions and algebraic constructions

We show that every hypersurface in $\R^s\times \R^s$ contains a large grid, i.e., the set of the form $S\times T$, with $S,T\subset \R^s$. We use this to deduce that the known constructions of extremal $K_{2,2}$-free and $K_{3,3}$-free graphs cannot be generalized to a similar construction of $K_{s,s}$-free graphs for any $s\geq 4$. We also give new constructions of extremal $K_{s,t}$-free graphs for large $t$.

preprint2012arXiv

Upper bounds for centerlines

In 2008, Bukh, Matousek, and Nivasch conjectured that for every n-point set S in R^d and every k, 0 <= k <= d-1, there exists a k-flat f in R^d (a "centerflat") that lies at "depth" (k+1) n / (k+d+1) - O(1) in S, in the sense that every halfspace that contains f contains at least that many points of S. This claim is true and tight for k=0 (this is Rado's centerpoint theorem), as well as for k = d-1 (trivial). Bukh et al. showed the existence of a (d-2)-flat at depth (d-1) n / (2d-1) - O(1) (the case k = d-2). In this paper we concentrate on the case k=1 (the case of "centerlines"), in which the conjectured value for the leading constant is 2/(d+2). We prove that 2/(d+2) is an *upper bound* for the leading constant. Specifically, we show that for every fixed d and every n there exists an n-point set in R^d for which no line in R^d lies at depth larger than 2n/(d+2) + o(n). This point set is the "stretched grid"---a set which has been previously used by Bukh et al. for other related purposes. Hence, in particular, the conjecture is now settled for R^3.

preprint2011arXiv

Sum-product estimates for rational functions

We establish several sum-product estimates over finite fields that involve polynomials and rational functions. First, |f(A)+f(A)|+|AA| is substantially larger than |A| for an arbitrary polynomial f over F_p. Second, a characterization is given for the rational functions f and g for which |f(A)+f(A)|+|g(A,A)| can be as small as |A|, for large |A|. Third, we show that under mild conditions on f, |f(A,A)| is substantially larger than |A|, provided |A| is large. We also present a conjecture on what the general sum-product result should be.

preprint2010arXiv

Radon partitions in convexity spaces

Tverberg's theorem asserts that every (k-1)(d+1)+1 points in R^d can be partitioned into k parts, so that the convex hulls of the parts have a common intersection. Calder and Eckhoff asked whether there is a purely combinatorial deduction of Tverberg's theorem from the special case k=2. We dash the hopes of a purely combinatorial deduction, but show that the case k=2 does imply that every set of O(k^2 log^2 k) points admits a Tverberg partition into k parts.

preprint2009arXiv

Lower bounds for weak epsilon-nets and stair-convexity

A set N is called a "weak epsilon-net" (with respect to convex sets) for a finite set X in R^d if N intersects every convex set that contains at least epsilon*|X| points of X. For every fixed d>=2 and every r>=1 we construct sets X in R^d for which every weak (1/r)-net has at least Omega(r log^{d-1} r) points; this is the first superlinear lower bound for weak epsilon-nets in a fixed dimension. The construction is a "stretched grid", i.e., the Cartesian product of d suitable fast-growing finite sequences, and convexity in this grid can be analyzed using "stair-convexity", a new variant of the usual notion of convexity. We also consider weak epsilon-nets for the diagonal of our stretched grid in R^d, d>=3, which is an "intrinsically 1-dimensional" point set. In this case we exhibit slightly superlinear lower bounds (involving the inverse Ackermann function), showing that upper bounds by Alon, Kaplan, Nivasch, Sharir, and Smorodinsky (2008) are not far from the truth in the worst case. Using the stretched grid we also improve the known upper bound for the so-called "second selection lemma" in the plane by a logarithmic factor: We obtain a set T of t triangles with vertices in an n-point set in the plane such that no point is contained in more than O(t^2 / (n^3 log (n^3/t))) triangles of T.

preprint2008arXiv

Stabbing simplices by points and flats

The following result was proved by Barany in 1982: For every d >= 1 there exists c_d > 0 such that for every n-point set S in R^d there is a point p in R^d contained in at least c_d n^{d+1} - O(n^d) of the simplices spanned by S. We investigate the largest possible value of c_d. It was known that c_d <= 1/(2^d(d+1)!) (this estimate actually holds for every point set S). We construct sets showing that c_d <= (d+1)^{-(d+1)}, and we conjecture this estimate to be tight. The best known lower bound, due to Wagner, is c_d >= gamma_d := (d^2+1)/((d+1)!(d+1)^{d+1}); in his method, p can be chosen as any centerpoint of S. We construct n-point sets with a centerpoint that is contained in no more than gamma_d n^{d+1}+O(n^d) simplices spanned by S, thus showing that the approach using an arbitrary centerpoint cannot be further improved. We also prove that for every n-point set S in R^d there exists a (d-2)-flat that stabs at least c_{d,d-2} n^3 - O(n^2) of the triangles spanned by S, with c_{d,d-2}>=(1/24)(1- 1/(2d-1)^2). To this end, we establish an equipartition result of independent interest (generalizing planar results of Buck and Buck and of Ceder): Every mass distribution in R^d can be divided into 4d-2 equal parts by 2d-1 hyperplanes intersecting in a common (d-2)-flat.