Source author record

Tian Pei

Tian Pei appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2026arXiv

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

General reasoning represents a long-standing and formidable challenge in artificial intelligence. Recent breakthroughs, exemplified by large language models (LLMs) and chain-of-thought prompting, have achieved considerable success on foundational reasoning tasks. However, this success is heavily contingent upon extensive human-annotated demonstrations, and models' capabilities are still insufficient for more complex problems. Here we show that the reasoning abilities of LLMs can be incentivized through pure reinforcement learning (RL), obviating the need for human-labeled reasoning trajectories. The proposed RL framework facilitates the emergent development of advanced reasoning patterns, such as self-reflection, verification, and dynamic strategy adaptation. Consequently, the trained model achieves superior performance on verifiable tasks such as mathematics, coding competitions, and STEM fields, surpassing its counterparts trained via conventional supervised learning on human demonstrations. Moreover, the emergent reasoning patterns exhibited by these large-scale models can be systematically harnessed to guide and enhance the reasoning capabilities of smaller models.

preprint2024arXiv

DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

The rapid development of open-source large language models (LLMs) has been truly remarkable. However, the scaling law described in previous literature presents varying conclusions, which casts a dark cloud over scaling LLMs. We delve into the study of scaling laws and present our distinctive findings that facilitate scaling of large scale models in two commonly used open-source configurations, 7B and 67B. Guided by the scaling laws, we introduce DeepSeek LLM, a project dedicated to advancing open-source language models with a long-term perspective. To support the pre-training phase, we have developed a dataset that currently consists of 2 trillion tokens and is continuously expanding. We further conduct supervised fine-tuning (SFT) and Direct Preference Optimization (DPO) on DeepSeek LLM Base models, resulting in the creation of DeepSeek Chat models. Our evaluation results demonstrate that DeepSeek LLM 67B surpasses LLaMA-2 70B on various benchmarks, particularly in the domains of code, mathematics, and reasoning. Furthermore, open-ended evaluations reveal that DeepSeek LLM 67B Chat exhibits superior performance compared to GPT-3.5.

preprint2016arXiv

Photon-assisted tunneling and charge dephasing in a carbon nanotube double quantum dot

We report microwave-driven photon-assisted tunneling in a suspended carbon nanotube double quantum dot. From the resonant linewidth at a temperature of 13 mK, the charge dephasing time is determined to be 280 +- 30 ps. The linewidth is independent of driving frequency, but increases with increasing temperature. The moderate temperature dependence is inconsistent with expectations from electron-phonon coupling alone, but consistent with charge noise arising in the device. The extracted level of charge noise is comparable with that expected from previous measurements of a valley-spin qubit, where it was hypothesized to be the main cause of qubit decoherence. Our results suggest a possible route towards improved valley-spin qubits.

preprint2016arXiv

Resonant optomechanics with a vibrating carbon nanotube and a radio-frequency cavity

In an optomechanical setup, the coupling between cavity and resonator can be increased by tuning them to the same frequency. We study this interaction between a carbon nanotube resonator and a radio-frequency circuit. In this resonant regime, the vacuum optomechanical coupling is enhanced by the DC voltage coupling the cavity and the mechanical resonator. Using the cavity to detect the nanotube's motion, we observe and simulate interference between mechanical and electrical oscillations. We measure the mechanical ring-down and show that further improvements to the system could enable measurement of mechanical motion at the quantum limit.