Source author record

Qianli Chen

Qianli Chen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2026arXiv

MiMo-V2-Flash Technical Report

We present MiMo-V2-Flash, a Mixture-of-Experts (MoE) model with 309B total parameters and 15B active parameters, designed for fast, strong reasoning and agentic capabilities. MiMo-V2-Flash adopts a hybrid attention architecture that interleaves Sliding Window Attention (SWA) with global attention, with a 128-token sliding window under a 5:1 hybrid ratio. The model is pre-trained on 27 trillion tokens with Multi-Token Prediction (MTP), employing a native 32k context length and subsequently extended to 256k. To efficiently scale post-training compute, MiMo-V2-Flash introduces a novel Multi-Teacher On-Policy Distillation (MOPD) paradigm. In this framework, domain-specialized teachers (e.g., trained via large-scale reinforcement learning) provide dense and token-level reward, enabling the student model to perfectly master teacher expertise. MiMo-V2-Flash rivals top-tier open-weight models such as DeepSeek-V3.2 and Kimi-K2, despite using only 1/2 and 1/3 of their total parameters, respectively. During inference, by repurposing MTP as a draft model for speculative decoding, MiMo-V2-Flash achieves up to 3.6 acceptance length and 2.6x decoding speedup with three MTP layers. We open-source both the model weights and the three-layer MTP weights to foster open research and community collaboration.

preprint2025arXiv

MiMo-Audio: Audio Language Models are Few-Shot Learners

Existing audio language models typically rely on task-specific fine-tuning to accomplish particular audio tasks. In contrast, humans are able to generalize to new audio tasks with only a few examples or simple instructions. GPT-3 has shown that scaling next-token prediction pretraining enables strong generalization capabilities in text, and we believe this paradigm is equally applicable to the audio domain. By scaling MiMo-Audio's pretraining data to over one hundred million of hours, we observe the emergence of few-shot learning capabilities across a diverse set of audio tasks. We develop a systematic evaluation of these capabilities and find that MiMo-Audio-7B-Base achieves SOTA performance on both speech intelligence and audio understanding benchmarks among open-source models. Beyond standard metrics, MiMo-Audio-7B-Base generalizes to tasks absent from its training data, such as voice conversion, style transfer, and speech editing. MiMo-Audio-7B-Base also demonstrates powerful speech continuation capabilities, capable of generating highly realistic talk shows, recitations, livestreaming and debates. At the post-training stage, we curate a diverse instruction-tuning corpus and introduce thinking mechanisms into both audio understanding and generation. MiMo-Audio-7B-Instruct achieves open-source SOTA on audio understanding benchmarks (MMSU, MMAU, MMAR, MMAU-Pro), spoken dialogue benchmarks (Big Bench Audio, MultiChallenge Audio) and instruct-TTS evaluations, approaching or surpassing closed-source models. Model checkpoints and full evaluation suite are available at https://github.com/XiaomiMiMo/MiMo-Audio.

preprint2012arXiv

Effect of lattice volume and strain on the conductivity of BaCeY-oxide ceramic proton conductors

In-situ electrochemical impedance spectroscopy was used to study the effect of lattice volume and strain on the proton conductivity of the yttrium-doped barium cerate proton conductor by applying the hydrostatic pressure up to 1.25 GPa. An increase from 0.62 eV to 0.73 eV in the activation energy of the bulk conductivity was found with increasing pressure during a unit cell volume change of 0.7%, confirming a previously suggested correlation between lattice volume and proton diffusivity in the crystal lattice. One strategy worth trying in the future development of the ceramic proton conductors could be to expand the lattice and potentially lower the activation energy under tensile strain.

preprint2011arXiv

Protons in lattice confinement: Static pressure on the Y-substituted, hydrated BaZrO3 ceramic proton conductor decreases proton mobility

Yttrium substituted BaZrO3, with nominal composition BaZr0.9Y0.1O3, a ceramic proton conductor, was subject to impedance spectroscopy for temperatures 300 K < T < 715 K at mechanical pressures 1 GPa < p < 2 GPa. The activation energies Ea of bulk and grain boundary conductivity from two perovskites synthesized by solid-state reaction and sol-gel method were determined under high pressures. At high temperature, the bulk activation energy increases with pressure by 5% for sol-gel derived sample and by 40% for solid-state derived sample. For the sample prepared by solid-state reaction, there is a large gap of 0.17 eV between the activation energy at 1.0 GPa and > 1.2 GPa. The grain boundary activation energy is around a factor two times as that of the bulk, and it reaches a maximum at 1.25 - 1.5 GPa, and then decrease as the pressure increases, indicating higher proton mobility in the grain boundaries at higher pressure. Since this effect is not reversible, it is suggested that the grain boundary resistance decreases as a result of pressure induced sintering. The steady increase of the bulk resistivity upon pressurizing suggests that the proton mobility depends on the space available in the lattice. In return, an expanded lattice with a/a0 > 1 should thus have a lower activation energy, suggesting that thin films expansive tensile strain could have a larger proton conductivity with desirable properties for applications.

preprint2011arXiv

The effect of compressive strain on the Raman modes of the dry and hydrated BaCe0.8Y0.2O3 proton conductor

The BaCe0.8Y0.2O3-δ proton conductor under hydration and under compressive strain has been analyzed with high pressure Raman spectroscopy and high pressure x-ray diffraction. The pressure dependent variation of the Ag and B2g bending modes from the O-Ce-O unit is suppressed when the proton conductor is hydrated, affecting directly the proton transfer by locally changing the electron density of the oxygen ions. Compressive strain causes a hardening of the Ce-O stretching bond. The activation barrier for proton conductivity is raised, in line with recent findings using high pressure and high temperature impedance spectroscopy. The increasing Raman frequency of the B1g and B3g modes thus implies that the phonons become hardened and increase the vibration energy in the a-c crystal plane upon compressive strain, whereas phonons are relaxed in the b-axis, and thus reveal softening of the Ag and B2g modes. Lattice toughening in the a-c crystal plane raises therefore a higher activation barrier for proton transfer and thus anisotropic conductivity. The experimental findings of the interaction of protons with the ceramic host lattice under external strain may provide a general guideline for yet to develop epitaxial strained proton conducting thin film systems with high proton mobility and low activation energy.