1. X
  2. Yuandong Tian
Log inSign up
Yuandong Tian
1,155 posts
user avatar
Yuandong Tian
@tydsh
Co-founder of @Recursive_SI. ex-Meta FAIR Director. ex-Google. Reasoning, Optimization and Understanding LLM. Novelist in spare time. PhD in @CMU_Robotics.
California, USA
yuandong-tian.com
Joined December 2009
950
Following
46.3K
Followers
RepliesRepliesMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    user avatar
    Yuandong Tian
    @tydsh
    Jun 11
    Early results from Recursive 🚀🚀 SotA results from our open-ended knowledge discovery system: 1️⃣NanoChat 5min pre-training (0.9372 bpb -> 0.9109 bpb, 2.8% lower Bits-Per-Byte than long-standing community SoTA) 2️⃣NanoGPT SpeedRun (79.7s -> 77.5s, 2.8% faster than long-standing
    user avatar
    Recursive
    @Recursive_SI
    Jun 11
    Article cover image
    Article
    First Steps Toward Automated AI Research
    Early results from Recursive’s automated AI research system on model training and GPU kernel benchmarks Today we are releasing early results from Recursive’s automated AI research system. Across three...
    65K
  • user avatar
    Yuandong Tian
    @tydsh
    Jul 25
    🚨A novel way to do RL in LLM post-training! Inspired by our previous path-not-taken work (arxiv.org/abs/2511.08567), we dig deep into the learning trajectory of RL and find that optimizing singular vectors (i.e., rotation) of weight matrices suffices for good performance in RL.
    user avatar
    Hanqing Zhu
    @zhu_hanqing666
    Jul 22
    People keep asking me: what's different about optimization in RL? Seemingly nothing — the pre-training stack just works (Adam, even SGD 👀 @saagnikkk). Bringing some answers from my last work (sorry for the delay — been cooking 🚀). We introduce ISO: Isospectral Optimization:
    67K
  • user avatar
    Yuandong Tian
    @tydsh
    Jul 25
    😂This has happened numerous times in the history. Human + smart phone = superhuman. Human + LLM = supersuperhuman. We are already way better than those miserable sapients 20 years ago.
    user avatar
    apefourone
    @ape41eth
    Jul 24
    Replying to @tydsh
    The other dystopia: Everyone has access to open-weight superhuman brains inside superhuman robots and can tell them what to do.
    9.4K
  • user avatar
    Yuandong Tian
    @tydsh
    Jul 24
    Strong support for that. The worst situation that could happen to human kinds is that a few elites control the best model (and their APIs), while other people treat them as "god" and pray for access. That would be the real dystopia...
    user avatar
    Jensen Huang
    NVIDIA
    @JensenHuang
    Jul 24
    For my first post, I’m sharing a letter @nvidia signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty.
    24K
  • user avatar
    Yuandong Tian
    @tydsh
    Jul 22
    In AMD AI Conference today.
    4.1K