Allie @endlessgit
voraciously vibing | prev @x algo + @xai posttraining anruigu.github.io San Francisco, CA Joined August 2015-
Tweets120
-
Followers744
-
Following534
-
Likes647
Data is one of the highest-leverage levers for AI safety. We're hiring an Applied Research Scientist, Safety & Alignment (SF / NYC) If you've done empirical work on alignment, evals, or model behavior and want to get your hands dirty with real model trajectories, feel free to reach out! lnkd.in/p/ghivyxiw fleetai.com/careers/resear…
Introducing Project Redwood 🚀🚀 @architectlabs is a frontier AI lab bringing together talent from Anthropic, xAI, Google DeepMind, and seasoned leaders across the hardware industry. We've raised a $24M seed round to build AI systems for chip design. 我們 Architect Labs 是一個 frontier AI lab ,由各家 Anthropic, xAI, Google DeepMind 還有各種硬體專家們組成,我們在做的是 ai system for chip design,目前募資 seed round $24M - Today, every major hardware company has its own chip design workflow—but these organizations and processes have become so large and entrenched that truly revolutionary change is difficult. We’re starting from first principles to create an AI-native chip design workflow—one that enables chip development to finally move at the speed of AI. 現在每家大硬體巨頭都有自己的晶片設計流程,但很多都大到不能做革命性的流程改動;我們在做的是從零思考,創造 ai native 的晶片設計流程,讓晶片設計真正跟上 AI 的速度。 - Project Redwood - from a single specification, our AI system designed, verified, and deployed a chip in under two weeks—delivering 3.4× better performance per watt than NVIDIA Jetson on billion-parameter models including Llama, Qwen, and Kimi. It autonomously generated the RTL, verification, firmware, drivers, and kernels, co-designing the model, software, and silicon in one optimization loop. We acknowledge that silicon is the ultimate ground-truth. We’re taking our approach all the way to GDS. We intend to tape-out multiple improved families of Redwood co-designed for various use-cases, on TSMC. Project Redwood - 一份規格書,我們的 AI 系統在兩週內完成晶片的設計、驗證與部署。在 Llama、Qwen 和 Kimi 等模型,其 performance per watt 比 NVIDIA Jetson 高出 3.4 倍。從 RTL、驗證、韌體、驅動到核心 kernels,全部由 AI 自己寫,並在同一個 loop 中自主設計模型、軟體與晶片。當然,流片才是最終的驗證標準。因此,我們會將這套方法一路推進至 GDS,並計畫採用台積電製程,針對不同應用場景協同設計多個持續改良的 Redwood 晶片系列並完成流片。 - Full report on Redwood architecture and its autonomous design. Follow @architectlabs on X 我們有公開 Redwood 架構及其自主設計流程,歡迎去看論文~
We gave our AI system a spec, and in under 2 weeks, it designed, verified, and deployed a chip that beats NVIDIA. It’s built for low-power physical AI workloads. We’re running live inference on >B+ parameter models like Llama, Qwen, and Kimi, serving at 3.4x better perf/watt
sharing my first blog! a meta harness for self-improving AI chip design… check it out if u wanna see how RL could work in chip design task, or how an ASIC design can become a hillclimbing task it introduced ASIC vs. GPU, RL Env setup, and the best generated design can run Kimi K3 at 87000 tps theoretically, which is 10x faster than the example shared by Kimi Team selected design files are open-sourced, happy to share the harness code with the enthusiasts! extremely enticing to bring RL capabilities to more complex system luoluo.ai/blog/kimi-k3
Through conversations with @andrewho03 and others at OpenAI and frontier labs, one thing has become clear: a data company lives or dies by its ability to understand what good data looks like. At Fleet, we’ve upstreamed computer-use and domain-specific capabilities into mainline models, and carefully studied perf gains through model post-training and scalable oversight. Designing good RL data is deeply nonobvious, and Andrew and I first connected over exactly this during late-night dead hangs at the gym (his hang time is crazy!). We are in the early innings of a new era of data (cc @willdepue), one where human ops doesn't scale across an expanding & uneven capability frontier. Better models create a Jevons-style effect — models make knowledge work cheaper, we attempt more of it, and the remaining work becomes harder and more contextual. The coding “slop” Andrew mentions comes as an evolution of how software eng happens now vs a year ago, and this pattern will repeat far beyond coding; moving the reliability gap from 30% to 90% creates harder and more creative data problems in every domain. The surface area of "general" in AGI is fractal. Most datasets miss economically useful work because real workflows are dynamic, deeply contextual, and difficult to capture faithfully in gradable, semi-synthetic environments. Andrew is tackling this problem first in scientific workflows, and I'm excited to see him bring this rigor to biology. At Fleet, we are tackling other frontier domains and have built the research foundation and platform needed to turn real workflows into reliable training signal. If you’re at a lab that cares deeply about data taste—or an engineer or researcher who wants to build it—come work with us. My DMs are open.
Today is my last day at @OpenAI. I'm glad to have spent the last eight months of my life working here! I'm starting a new company focused on the production of high-quality reinforcement learning datasets: 1. The generalization ability of LLMs is clearly very poor, with "spiky"
hiring research scientists in agentic {RL posttraining, scalable oversight, environment scaling} & high-taste applied researchers 🪄 links below, DM open fleet.ai/careers/resear… fleet.ai/careers/resear… fleet.ai/careers/resear… fleet.ai/careers/mts-ap…
@natolambert Referenced this a lot during my transition from general mle to llm posttraining. Thank you!
im sad to announce i passed the reverse turing test: today someone deleted some of my writing because they thought fable wrote it
I feel like this is a persona selection problem, like finding a way to tell the model “based on this inferred context, it’s safe to put your artist cap (beret or whatever) on now”. We see more creative flourishing in role-playing because it explicitly gives this permission. I’m wishing for a) model being able to autonomously modulate its assistant-ness internally without prompting/steering/verbalization; b) an opinionated default artist persona that is not just an interpolation of training data
Today, I’m excited to formally announce @mirendil with my amazing co-founders Harsh Mehta, Shayan Salehian, and Tara Rezaei! We’re fortunate to work with @a16z and @kleinerperkins, who led our seed round of $200M, followed by a major investment from NVIDIA, among others. Mirendil exists to accelerate science and technology, and through them, to help solve humanity's most pressing problems. Self-accelerating AI R&D is the most direct path to delivering on AI's broader promise, which is why we believe the most important application of AI is AI itself. Get this loop right, and it compounds. It fundamentally changes the rate of progress itself across all domains. We believe this capability should be democratized. It should be used to power all scientific efforts trying to innovate at the frontier. There are far more important problems—and broader ones—than any single lab can take on, so more groups should be able to pursue them. This pulls concentration of power away from a few labs: businesses and science labs can own their AI and infrastructure, keep their margins, and control their own destiny instead of ceding it all to a single AI lab. We’re a small team with a singular focus. Our founding team consists of 20 researchers and engineers from frontier institutions including Anthropic, xAI, Google DeepMind, and OpenAI, united by a passion for science and a drive to build the technologies that move it faster. If you want to build the system that builds systems, join us! @HarshMeh1a, @shayan_, @tararezaeikh
@PeterHndrsn would be curious about measuring the epistemic laziness from "explanation slop" making you think you understand something when you don't, via "this is the deeper insight" etc
flow state
We are incredibly excited to announce River AI. Our mission is to create personal AI that is owned and shaped by you. Today’s best AIs are controlled by a few large corporations. We are building the alternative: a new, personal stack for AI that works entirely for you, shares
We’re excited to release Agents’ Last Exam (ALE) 🚀 As agentic and long-horizon tasks become increasingly important for frontier LLMs, ALE provides the first comprehensive benchmark of its kind: 1,500+ real-world tasks developed with 300+ experts across 55 industries. See more at agents-last-exam.org HuggingFace: hf.co/papers/2606.05… Thanks @Xinyang_Han_ @YiyouSun and etc.
“AI agents will outperform humans at almost all jobs by 2026–2027.” - The forecast is everywhere. So we built the exam to test that claim, on real labor-market aligned work. On the hardest tier, top agents pass 2.6%. Meet Agents' Last Exam (ALE), a rolling benchmark measuring
Wrote this for fun: I/RL Algorithms to Live By, honoring Brian Christian's Algorithms to Live By that in part convinced me to switch from econ to cs. All life is an experiment and I am participant zero. Tried co-writing with AI for the first time, takeaway is to be the first-mover of thoughts and to choose the things worth being in slow-time for, like an idea that I can uniquely generate or concepts that take years to click. app.notion.com/p/RL-Algorithm…
been watching the cooking for a while. excited for their trajectory!
Today, @MichaelElabd, @QuantumArjun, and I are excited to announce Trajectory. We are a research lab and product company building the platform for Continual Learning. Our platform unlocks the signal already sitting in product usage, so companies can continuously post-train
@QuantumArjun congrats arjun! can't wait to see what ur up to next
Dylan Hailey @dylanshailey
28K Followers 2K Following securing agents & infra, previously @roblox, @atvi_ab, @FBI 🏳️🌈🍣🛫
B. Scot Rousse @bscotrousse
197 Followers 604 Following Philosophy + AI Drummer in Punk Bands Affiliate Researcher at UC Berkeley and Topos Institute
Michael @mgrczyk
1K Followers 527 Following Train bigger models, build more housing Big tent transhumanist @AnthropicAI anything I say is my own opinion and not Anthropic's
Gerard Sans | Axiom �... @gerardsans
39K Followers 12K Following Founder Axiom // Forging skills for the new era of AI. GDE in AI, Cloud & Angular. Building London's tech & art nexus @nextai_london. Speaker | MC | Trainer.
g @g9665926721001
73 Followers 4K Following
Jamie Kang @kang117577
3 Followers 63 Following
dario melappioni @DarioMelap40117
0 Followers 27 Following
Sijun Tan @sijun_tan
2K Followers 520 Following CS PhD @BerkeleySky | Project Lead @rllm_project @Agentica_ | Prev: @AIatMeta @Antgroup
Sunshine Chen @sunshine_cxn
184 Followers 1K Following stats & history @harvard | sometimes @colossusmag , @dormroomfund
Sam Leisure @tweeterSam3
612 Followers 176 Following MTS at Fleet, Early Eng @dydx, built GoldLink & Quickscope Beat Promised Consort Radahn w/o summons, Heolstor w/o friends & Lost Lace
Chase Kapler @ChaseKapler
9 Followers 43 Following
Xiyang Wu @IROS2026 (... @wu_xiyang
520 Followers 2K Following Ph.D. Student at @gammaumd @eceumd @umiacs @UofMaryland. Intern @NECLabsAmerica @Dolby. Previous: @EmoryUniversity @GeorgiaTech @TJU1895.
Vidhi Desai @_vidhidesai_
258 Followers 706 Following Building a space company. Prev @viasat @inmarsatglobal // Will show up for dancing, cocktails, musical theatre & board games.
Ameen Patel @Ameen_ml
3K Followers 3K Following @ stealth lab prev @PrimeIntellect, @togethercompute, @AmazonScience, @uwaterloo
Jayanth Chundru @jayanthchundru
12 Followers 552 Following LLMs & Multimodal agents M.S in CS @cincynlp . @uofcincy
Udari Madhushani Sehw... @UdariMadhu
175 Followers 514 Following Research Lead @ScaleAI, working on agentic AI, alignment, AI safety, and scalable oversight. Ex-intern @Deepmind @MetaAI
Hindol.exe @Hindollllll
65 Followers 495 Following Iterating manifolds ML research intern @lossfunk 🧑💻
雪枫 刘 @XueLiu78130
2 Followers 331 Following
shira @shiraeis
19K Followers 3K Following early childhood ai startup :) prev: ai @uchicago @mit @intel & a few other places. I personally think I’m quite funny.
Abhishek Ranjan @abhishek_rp2002
395 Followers 6K Following swe by heart, mle by practice @skyfallai pp: https://t.co/noqj4Rz9y9 applying machine intelligence to solve practical problems at scale.
Slava Akhmechet @spakhm
19K Followers 1K Following eng leader @azure // prev @stripe, cofounder @rethinkdb (yc s09). DM to say hi!
Tong Zheng @zhengtoong
186 Followers 232 Following PhD Student @umdcs; student researcher @Google ;ex-@NEU@TencentGlobal; Working on autonomous parallel thinking of LLMs. Parallel-R1; Parallel-Probe; AutoTTS;
keval shah @keval_sha
398 Followers 2K Following AI Researcher, systems stuff | @uchicago chasing 100% SoL
J. Mark Hou @JMarkHou
182 Followers 1K Following member of custodial staff @ my house, ex Meta / MIT
bethany @Elena9857860688
12 Followers 1K Following
Bhavin Jawade @BhavinJawade
943 Followers 5K Following Senior Research Scientist @ Netflix Ph.D in Computer Science @UBuffalo Research Intern @Netflix, @Yahoo, @Adobe https://t.co/dBXaNr0pq2 Post-Training
Marcel Hedman @hedman_marcel
93 Followers 689 Following AI | Strategy | Speaking | ML PhD candidate @UniofOxford Founder @nural_research --- Formerly Founding team @axion_ray DS @Harvard Phys @Cambridge_Uni
Juhil Modi @juhilmodihere
6 Followers 2K Following
gokul @gokulp01
614 Followers 4K Following PhD candidate @UofIllinois (UIUC) @CSL_Illinois | Current: robotics research intern @Nvidia | ex-research science intern @Adobe Research
Frosty40 @FrostForger
453 Followers 2K Following builder focused on tools and AI systems. Frequent open source contributions for access to knowledge. Suma Corn Loude @ Dunning Krugers University
Shan @ShanRizvi
8K Followers 4K Following AI + knowledge graphs + agentic systems. Building semantic memory, reasoning engines, & graph-based RL. Thinking about intelligence, structure, and the future.
Mushu @Narayanyengar89
23 Followers 1K Following
Aryan @Aryan2811DB
132 Followers 2K Following ex-Recherché Inc, ex-Build (https://t.co/0lrNKDDdSD), @hf0
ghyn ynv @GhynY86682
0 Followers 729 Following
Forethought @forethought_org
7K Followers 4 Following Research nonprofit exploring how to navigate explosive AI progress.
Dylan Hailey @dylanshailey
28K Followers 2K Following securing agents & infra, previously @roblox, @atvi_ab, @FBI 🏳️🌈🍣🛫
Michael @mgrczyk
1K Followers 527 Following Train bigger models, build more housing Big tent transhumanist @AnthropicAI anything I say is my own opinion and not Anthropic's
Udari Madhushani Sehw... @UdariMadhu
175 Followers 514 Following Research Lead @ScaleAI, working on agentic AI, alignment, AI safety, and scalable oversight. Ex-intern @Deepmind @MetaAI
Sijun Tan @sijun_tan
2K Followers 520 Following CS PhD @BerkeleySky | Project Lead @rllm_project @Agentica_ | Prev: @AIatMeta @Antgroup
Natasha Jaques @natashajaques
35K Followers 1K Following Assistant Professor leading the Social RL Lab https://t.co/ykwfJG84Bj @uwcse and Staff Research Scientist at @GoogleAI.
mei @multiply_matrix
3K Followers 220 Following AGI forecaster. In 2022 I predicted an AGI timeline of 2027 MIT dropout
shira @shiraeis
19K Followers 3K Following early childhood ai startup :) prev: ai @uchicago @mit @intel & a few other places. I personally think I’m quite funny.
Bhavin Jawade @BhavinJawade
943 Followers 5K Following Senior Research Scientist @ Netflix Ph.D in Computer Science @UBuffalo Research Intern @Netflix, @Yahoo, @Adobe https://t.co/dBXaNr0pq2 Post-Training
Zeinab Mohammadi @Zeinab___M
369 Followers 3K Following Machine learning for Neuroscience | Postdoc @NorthwesternU with @joshuaiglaser | Formerly Postdoc @princeton with @jpillowtime | PhD in Electrical Engineering
Liam @Liam862373
73 Followers 620 Following Robot Learning at @Stanford @ETH_en | Interned at @Microsoft @mimicrobotics @NASAJPL
Mimansa Jaiswal @ COL... @MimansaJ
6K Followers 6K Following Currently Senior RS @Netflix (prev @aiatmeta) | LLM/SLM Post Training | Data, Evals, RL and Agents https://t.co/Nv4aMHzPJa
Logan Thorneloe @loganthorneloe
10K Followers 1K Following Building research agents @Google. Pragmatic optimist. Teaching engineers AI: https://t.co/fEuvJt4TRK.
Hanqing Zhu @zhu_hanqing666
5K Followers 609 Following RL Science & Coding RL Bigrun @xai | Grok 4.5, 4.6, 4.7 4.🧑🍳 prev. @UTAustin @GoogleDeepMind @AIatMeta Theory-driven ML, efficient in practice Views my own
Alexis Ross @alexisjross
5K Followers 1K Following founding MTS @humansand | phd @MIT_CSAIL | formerly @allen_ai, @harvard '20
Aella @Aella_Girl
256K Followers 418 Following whorelord, survey artist, sample size queen, way too earnest. https://t.co/IcEgPhWaSW
zomglings @zomglings
446 Followers 579 Following mathematician, programmer, baboon, member of technical staff @fleet_ai
Haven Feng @HavenFeng
5K Followers 1K Following Postdoc @berkeley_ai, building Impossible. Prev. @MPI_IS @AdobeResearch @uni_tue Interested in recreating the physical world and the intelligence to do so.
Shuqi Dai @ShuqiDai1
871 Followers 172 Following Audio @xAI @SpaceX Grok Imagine, Train Voice mode @xAI, PhD in CS @ Carnegie Mellon University, Professional Pipa Player, Composer, Singer
Ludwig Schmidt @lschmidt3
7K Followers 425 Following Assistant professor at @Stanford and member of the technical staff at @AnthropicAI.
Su Park @sunotsue_
410 Followers 1K Following terminal bench · post-training & evals against global banality prev @ltiatcmu @columbia
Behnam Neyshabur @bneyshabur
45K Followers 1K Following Co-Founder & CEO @mirendil 💼 Past: co-led Discovery team @AnthropicAI & Blueshift team @GoogleDeepMind 🎒Traveling & Backpacking
Jiaxin Pei @jiaxin_pei
2K Followers 954 Following Assistant Professor @UTAustin, Research Fellow @StanfordHAI & @DigEconLab Researching AI Agents and Human-AI Interaction
Peter Henderson @PeterHndrsn
6K Followers 957 Following Assistant Professor @ Princeton (RL+strategic decision-making+Law). Prev: Stanford (JD+PhD); Mila; FAIR; Amazon; Cal Supreme Court. Views are all only my own.
Marwa Abdulhai @marwaabdulhai
661 Followers 254 Following PhD Student @berkeley_ai AI Safety, RL, and LLMs
Sanjay Kairam @skairam
3K Followers 1K Following Evals @Simile_AI | Ex-MTS @OpenAI | @Reddit @Twitch Alum | PhD CS @Stanford @StanfordHCI
Dan Zhang @ ICML @DZhang50
6K Followers 1K Following LLM Lead at Ricursive Intelligence | ex-Gemini @ Google DeepMind | Computer Architecture PhD @ UT Austin🤘 | Opinions stated here are my own.
Ze Liu @zeliu_
6K Followers 514 Following @xAI (Grok Video Gen, Grok Vision, Grok Voice Mode, Grok3). ICCV 21 Marr Prize.
Haotian Liu @imhaotian
13K Followers 606 Following @Meta Superintelligence Lab, prev. @xAI @grok imagine/omni/vision, #LLaVA, @MSFTResearch @UWMadison
Guanghan Ning @quietnning
644 Followers 272 Following Research @fleet_ai, formerly @ ByteDance-Seed (LLM) https://t.co/FL2xvayzCt · views my own
Cassidy Laidlaw @cassidy_laidlaw
1K Followers 247 Following PhD student at UC Berkeley studying RL and AI safety. Also at https://t.co/OrEPAiR8b0
Circleboom @circleboom
247K Followers 1K Following We create intuitive and easy-to-use social media products for brands, SMBs and users / Official X Enterprise Developer @circleboomhub @circleboomdev
Noah Ziems @NoahZiems
4K Followers 2K Following Resident @LaudeInstitute prev @MIT_CSAIL under @lateinteraction, PhD @NotreDame Pedagogical RL, GEPA, DSPy
Joshua Gu @astrogu_
330 Followers 261 Following PhD student @MIT_CSAIL @MIT, Recursive Research👨💻 | Previous: @LMCache, @tensormesh, undergrad @UChicago




































