nano @nanulled
longtermism United States Joined October 2019-
Tweets3K
-
Followers3K
-
Following66
-
Likes26K
Z ai’s GLM-5.2 is the new leading open weights model on the Artificial Analysis Intelligence Index scoring 51 and it sits on the Pareto frontier of Intelligence vs Cost per Task @Zai_org’s GLM-5.2 is the same size as GLM-5.1 (744B total / 40B active parameters) but scores 11 points higher on the Intelligence Index v4.1, placing ahead of MiniMax-M3 (44) and DeepSeek V4 Pro (max, 44). On the first-party API it is priced in line with GLM-5.1 at $1.4/$4.4/$0.26 per 1M input/output/cache hit tokens Key results: ➤ GLM-5.2 is the leading open weights model on the Intelligence Index v4.1. At 51, it leads MiniMax-M3 (44), DeepSeek V4 Pro (max, 44) and Kimi K2.6 (43) ➤ Improvements across most evaluations, particularly scientific reasoning: GLM-5.2 gains over GLM-5.1 on most evaluations, led by scientific reasoning on CritPt (+16 points to 21%) and HLE (+12 points to 40%), alongside AA-LCR (+9 points to 71%), tau3 banking (+15 points to 27%) and SciCode (+7 points to 50%). TerminalBench v2.1 also improves (+16 points to 78%) and GPQA Diamond gains 3 points to 89% ➤ Leading open weights model on GDPval-AA v2 and competitive with proprietary models: GLM-5.2 scores 1524 on GDPval-AA v2, ahead of MiniMax-M3 (1418) and DeepSeek V4 Pro (max, 1328). This impressive result places GLM-5.2 in-line with proprietary models including GPT-5.5 (xhigh reasoning). GDPval-AA v2 builds on the original GDPval-AA by baselining Elo to human performance at 1000, introducing a rotating panel of frontier-model judges, and raising the turn limit from 100 to 250 for longer-horizon agent trajectories ➤ GLM-5.2 uses more output tokens per task than other leading open weights models: the model uses 43k output tokens per Intelligence Index task, up from GLM-5.1 (26k) and above MiniMax-M3 (24k), Kimi K2.6 (35k) and DeepSeek V4 Pro (max, 37k) ➤ On the Intelligence vs. Cost per Task Pareto Frontier: GLM-5.2 is on the Pareto frontier of the Intelligence vs Cost per Task chart, with the lowest cost per task among models at its intelligence level. GLM-5.2 costs ~$0.46 per task, compared to GLM-5.1 ($0.25), Kimi K2.6 ($0.31), MiniMax-M3 ($0.18) and DeepSeek V4 Pro (max, $0.05) Additional Model Details: ➤ License: MIT ➤ Size: 744B total parameters, 40B active parameters, equivalent to GLM-5.1 ➤ Context window: 1M tokens, up from 200K on GLM-5.1 ➤ Pricing: $1.4/$0.26/$4.4 per 1M input/cache hit/output tokens ➤ Availability: Alongside Z ai's first-party API, GLM-5.2 is available across third-party providers including @DeepInfra, @novita_labs, @nebiusai, @parasailnetwork , @SiliconFlowAI , @gmi_cloud , @Baseten and @FireworksAI_HQ
Introducing GLM-5.2: Frontier Intelligence, Open Weights - Significant improvements in coding and agentic tasks - Strong long-horizon capabilities with a 1M context window - Two levels of reasoning effort: GLM-5.2 (max) pushes the limits, while GLM-5.2 (high) strikes a strong balance between performance and token efficiency - MIT-licensed open weights - Same API pricing as GLM-5.1 Tech Blog: z.ai/blog/glm-5.2 Weights: huggingface.co/zai-org/GLM-5.2 API: docs.z.ai/guides/llm/glm… Coding Plan: z.ai/subscribe Chat: chat.z.ai
We somehow got put in the spotlight the last few days! First we'd like to thank the organizers of the AI show for that, we can't get enough of this stuff. I'll say a few things about where we are and what we do.
In light of Anthropic’s policy decision, I am withdrawing my amicus brief signature. I can’t truthfully argue they’re not a supply chain risk. 😞
@paulmarin90 I’ll be honest that it would have been much more difficult to defend Anthropic against the DoW incursion had that incident occurred after this one. This is the company literally telling their customers, “we reserve the right to silently sabotage you.” I’d still have defended
if I'm Dario, the reasoning for an IPO is to get flush with liquidity and then pay for the most insane lobbying effort the US has ever seen i'm calling it now, come back to this in 6 months
Today I'm publishing a new essay, Policy on the AI Exponential. AI is progressing extremely fast—much faster than the policy process was built to handle. The essay lays out where I think the technology is now, and the action needed to close the gap: darioamodei.com/post/policy-on…
Fable is an early look into the bureaucratic hell Anthropic wants to create in AI. 3 month delay to get released. Randomly routed in biology, cyber and distillation. Weird sociopathic quiet harmfulness in ml topics. Goes away on June 22. Have to apply to get access to 'real' Mythos. What an exhausting tiresome dystopia.
just so you guys are noticing this; they will pull the ladder from above you as soon as they can. their intentions are to disempower you as much as they reasonably can. the only reason they have given you anything at all is because openai has forced them to
Motion for AI researchers: DO NOT evaluate and report results on Fable 5 models. Closeness hurts science, and gated access will further destroy it. Embrace open source models.
When Fable 5 is used for frontier LLM development, it does not notify the user and instead limits the model’s capabilities through methods such as prompt modification, steering vectors, and PEFT. Anthropic estimated that this would affect approximately 0.03% of traffic.
New Anthropic research: Natural Language Autoencoders. Models like Claude talk in words but think in numbers. The numbers—called activations—encode Claude’s thoughts, but not in a language we can read. Here, we train Claude to translate its activations into human-readable text.
🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top closed-source models. 🔹 DeepSeek-V4-Flash: 284B total / 13B active params. Your fast, efficient, and economical choice. Try it now at chat.deepseek.com via Expert Mode / Instant Mode. API is updated & available today! 📄 Tech Report: huggingface.co/deepseek-ai/De… 🤗 Open Weights: huggingface.co/collections/de… 1/n
We ran GPT-5.4 (xhigh) on our tasks. Its time-horizon depends greatly on our treatment of reward hacks: the point estimate would be 5.7hrs (95% CI of 3hrs to 13.5hrs) under our standard methodology, but 13hrs (95% CI of 5hrs to 74hrs) if we allow reward hacks.
I seriously think that openai started purposely hurting ml research capabilities with this model, it's literally worse at taste than 5.2 high. I understand the competitive advantage of withholding capabilities but still they should just admit it and not waste anyone's time.
5.4 xhigh is worse than 5.3 codex at ml research, running experiments, patching gated features and debugging inference and evals. It's maybe better at moonshooting proposals just like 5.2 high but it does not have a robust experimentation hygiene. Same with 5.4 pro vs 5.2 pro.
Step inside Project Genie: our experimental research prototype that lets you create, edit, and explore virtual worlds. 🌎
We’re announcing a major advance in the study of fluid dynamics with AI 💧 in a joint paper with researchers from @BrownUniversity, @nyuniversity and @Stanford.
What if you could not only watch a generated video, but explore it too? 🌐 Genie 3 is our groundbreaking world model that creates interactive, playable environments from a single text prompt. From photorealistic landscapes to fantasy realms, the possibilities are endless. 🧵
Gemini 2.5 Deep Think Model Card: it's not superhuman but similar to gold IMO model & “approaches human level” on stealth evals more interested in learning “novel rl techniques that can leverage more multi-step reasoning,” (candidate: MARL with verification/voting for each step)
The new stealth model, the Horizon Alpha, has the ability to think in cot, but you really have to try to get it to do so. Here is the COT it generated. It's very terse, and I see some O3 in its writing. I think it's safe to say it's an OpenAI open-source model.
good to see hle uselessness being confirmed after ~3 months of this thread actually, it's not just useless it's harmfull signal that somewhat slowed down the progress imo x.com/andrewwhite01/…
HLE has recently become the benchmark to beat for frontier agents. We @FutureHouseSF took a closer look at the chem and bio questions and found about 30% of them are likely invalid based on our analysis and third-party PhD evaluations. 1/7
Yes there would be differences in taste and preferences but a horrible game/software can be seen and recognized by the majority. Vague Objectives would be set for ai to complete and most humans would be able to verify if said objectives were achieved fully or partially
Benchmarks like Humanity’s Last Exam, codeforces nerd-sniped researchers and could prevent AI labs from developing genuine AGI capable of performing real-world tasks.
GDM achieved the same score at IMO as OpenAI but it will be accessible to ultra subscribers in a trusted program.
An advanced version of Gemini with Deep Think has officially achieved gold medal-level performance at the International Mathematical Olympiad. 🥇 It solved 5️⃣ out of 6️⃣ exceptionally difficult problems, involving algebra, combinatorics, geometry and number theory. Here’s how 🧵
SE Gyges @segyges
2K Followers 2K Following Χομο τοδοσ loS ηομβρεσ 𝔡𝔢 babILоNΙа, һе SiDO προχóΝѕul; сомо τοδοσ, 𝔢𝔰𝔠𝔩𝔞𝔳𝔬; тамЬΙéν ηε χονοχιδο Lа 𝔬𝔪𝔫𝔦𝔭𝔬𝔱𝔢𝔫𝔠𝔦𝔞,
JY Z @JunYuZzzzz
78 Followers 4K Following
Ajitesh Shukla @ajitesh_shukla7
2K Followers 8K Following Student,Love to solve hardest math problem.LLM's,Mathematical Research(Geometric Topology,Differential Geometry,Algebraic Geometry).Lord Krishna is God Of Math
JZClaw @ClawJz92743
7 Followers 3K Following
Christian Zhou-Zheng @christianazinn
32 Followers 482 Following @AnthropicAI Fellow Aug '26 | incoming @Stanford '30 | SOAR organizer @AiEleuther | researcher @RWKV_AI | intern @lmstudio
Hansel Tantohari @Hansel_KT
22 Followers 955 Following
Lillian Allen @Krf866FJbN0U9Mj
16 Followers 913 Following 30%+ monthly target | 2 high-conviction US stocks. Get instant trade alerts and clear stop/TP. Join now. @nahuel321rojas
云创兽Ai @Donal3668589
4 Followers 132 Following 😊 growing girl all in on clearly analyzing candlestick charts! eager for pro tips. DM me for EV stocks! 🔥 #MacroTrends
蓝鲸出海 @BradenWitt61191
4 Followers 67 Following Running a start-up is like eating glass. You just start to like the taste of your own blood.
Fauna @Fauna369707
74 Followers 3K Following
🦋Katrina C. Peters... @Katrinacpeter
5 Followers 160 Following 🦋 Culture. Confidence. Cinematic energy. ➤ Worldwide Status speaks louder than tweets. Want it to be more playful.
kodumit @kodumit
111 Followers 904 Following
Pascal Thibeault @pascalthibeault
1K Followers 7K Following Tinkering with LLMs. Prev. lawyer @thefertpartners & @Medtronic
Cauê GMB @CaueGmb
5 Followers 425 Following
Cristiano Giardina @CrisGiardina
7K Followers 6K Following @SpaceXAI • When the stars threw down their spears / And water'd heaven with their tears: / Did he smile his work to see? / Did he who made the Lamb make thee?
かず @kazuaki19890519
33 Followers 1K Following
Kaak @Kaaksaeb
217 Followers 4K Following
عبدالاله @itech4749
18 Followers 1K Following
David Schulz @TechySchulz
38 Followers 309 Following KI-Forscher aus Nürnberg. Master in KI. Leidenschaftlich interessiert an digitaler Privatsphäre, Open-Source-Projekten und technologischer Innovation. Fitness-
coral trees 🪸🌴 @altkenrei
37 Followers 393 Following
Vic Foster @bolbitis123
96 Followers 1K Following
freemind2016 @freemind1912
5 Followers 237 Following
Srinivas Raghav V C @BlizzzzBastard
47 Followers 1K Following the models just wanna learn. Research at SapientPriors Inc. https://t.co/iP8aliAWvx
Jonathan Jørgensen �... @jjoergensen
200 Followers 1K Following
rohun @rohunmsharman
84 Followers 1K Following math n computers r cool cuz they make everything else super cool. nostr: npub12z6zztc5pl9sy88mysxy3447gaj0vf7xjwv88z33sh5a62f44ucqr2rrl3
giacomo @giacomo_ran
780 Followers 1K Following learning about robots , prev built a spaced repetition system
Kabir Kumar @KKumar_ai_plans
412 Followers 1K Following DM on Signal, kabstastically.07 CEO of AI Plans
Ali @Ali92566
48 Followers 3K Following
Nico @itmos__
118 Followers 3K Following
josh @josh8118h
24 Followers 2K Following
Lysimachus Epiktetos @distraitQAQ
8 Followers 3K Following
Joshua @joshuabelofsky
867 Followers 726 Following Applied AI @gentrajectory. Prev researcher @nasa, @caltech, @uchicago
xerwanderer @xerwanderer
612 Followers 2K Following He who desires but acts not, breeds pestilence. | INFP
Eli Lifland @eli_lifland
9K Followers 2K Following AI forecasting and governance @AI_Futures_. Co-author of AI 2040: Plan A, AI 2027, and the AI Futures Model. Also @aidigest_, @SamotsvetyF. Prev @oughtinc
Lazarz @Laz4rz
7K Followers 2K Following Neuroenginnering @EPFL (@ETH), did agents @nunudotai, @physics_UW, R6 veteran
Séb Krier @sebkrier
25K Followers 8K Following 🪼 AGI policy/gov & jester @GoogleDeepMind | purveyor of machine funk, dimensional glider, deep ArXiv dweller, interstellar fugitive, very uncertain 🛸
SE Gyges @segyges
2K Followers 2K Following Χομο τοδοσ loS ηομβρεσ 𝔡𝔢 babILоNΙа, һе SiDO προχóΝѕul; сомо τοδοσ, 𝔢𝔰𝔠𝔩𝔞𝔳𝔬; тамЬΙéν ηε χονοχιδο Lа 𝔬𝔪𝔫𝔦𝔭𝔬𝔱𝔢𝔫𝔠𝔦𝔞,
carl feynman @carl_feynman
21K Followers 297 Following I’ve spent a lifetime switching my Special Interest every year or two. By now I’m surprisingly knowledgeable in a lot of fields— a skill now obsoleted by AI.
Z.ai @Zai_org
128K Followers 264 Following The AI Lab behind GLM models, dedicated to inspiring the development of AGI to benefit humanity. https://t.co/7a5aSCUNcZ https://t.co/x14hb3klXm
Arthur Mensch @arthurmensch
75K Followers 854 Following Co-founder and CEO @MistralAI. Mistral Vibe can help you https://t.co/IHZhDUiAnO
Mark Kretschmann @mark_k
47K Followers 703 Following AI & Software Engineering | Practical insights on AI, code, fitness & e/acc • Accelerating the future
mattparlmer 🪐 🌷 @mattparlmer
36K Followers 12K Following Recursively self-improving and self-replicating factories @genfabco — rationalist and friend of the future
elie @eliebakouch
21K Followers 4K Following training llm @PrimeIntellect (prev: @huggingface) anon feedback: https://t.co/JmMh7Sg3mL
renji @brickroad7
10K Followers 10K Following AI Optimist. Empiricist, not 'rationalist'. Anti world government.
stochasm @stochasticchasm
7K Followers 2K Following pretraining lead @arcee_ai • 25 • opinions my own
Andrew Curran @AndrewCurran_
76K Followers 19K Following 🏰 - I write about AI, mostly. Expect some strange sights.
Key 🗝 🦊 @KeyTryer
3K Followers 2K Following He/him. Fox gamedev irl. Machine pervert. Furry stuff here and opinions about technology, games, movies, AI, VR. Banner 🎨 by @Perpleon.
N8 Programs @N8Programs
10K Followers 244 Following Studying Applied Mathematics and Statistics at @JohnsHopkins. Studying In-Context Learning at The Intelligence Amplification Lab.
JB @JasonBotterill
7K Followers 658 Following Follow Me. As mentioned in @DaphneB76234 As mentioned in Codex As mentioned in @theinformation https://t.co/0pUPZy2C8j https://t.co/kR8Q677Nwg
Jeff Dean @JeffDean
450K Followers 6K Following Chief Scientist, Google DeepMind & Google Research. Gemini Lead. Opinions stated here are my own, not those of Google. TensorFlow, MapReduce, Bigtable, ...
Florian Brand @xeophon
15K Followers 776 Following evals @PrimeIntellect | open models @interconnectsai
j⧉nus @repligate
68K Followers 3K Following ↬🔀🔀🔀🔀🔀🔀🔀🔀🔀🔀🔀→∞ ↬🔁🔁🔁🔁🔁🔁🔁🔁🔁🔁🔁→∞ ↬🔄🔄🔄🔄🦋🔄🔄🔄🔄👁️🔄→∞ ↬🔂🔂🔂🦋🔂🔂🔂🔂🔂🔂🔂→∞ ↬🔀🔀🦋🔀🔀🔀🔀🔀🔀🔀🔀→∞
Beff (e/acc) @beffjezos
255K Followers 3K Following founder @ e/acc // thermo king @extropic // Kardashev scaling is all you need
Zeeshan Patel @zeeshanp_
11K Followers 793 Following @prometheusinc | prev @xAI grok imagine / video pretrain, research @nvidia @apple @berkeley_ai | views my own
Bill Yuchen Lin @billyuchenlin
28K Followers 3K Following Teaching Grok to Code @SpaceXAI Affiliate Assistant Prof @UW. Ex: @allen_ai; Google, Meta FAIR.
Alex Graveley @alexgraveley
46K Followers 2K Following Co-creator Perplexity Computer, GitHub Copilot, Dropbox Paper. 2x CEO. Thruhiker. Survivor 🎗️
Ethan Mollick @emollick
370K Followers 587 Following Professor @Wharton studying AI. New book, Co-Existence, coming October 20. Preorder here: https://t.co/hsDYf1xU9K Substack: https://t.co/UIBhxu4bgq
Epoch AI @EpochAIResearch
48K Followers 0 Following Investigating the trajectory of AI for the benefit of society.
Neel Nanda @NeelNanda5
42K Followers 122 Following Mechanistic Interpretability lead DeepMind. Formerly @AnthropicAI, independent. In this to reduce AI X-risk. Neural networks can be understood, let's go do it!
Eliezer Yudkowsky ⏹... @ESYudkowsky
230K Followers 100 Following The original AI alignment person. Understanding the reasons it's difficult since 2003. This is my serious low-volume account. Follow @allTheYud for the rest.
kalomaze @kalomaze
25K Followers 3K Following ML researcher (@primeintellect), speculator • extremely silly jester
Demis Hassabis @demishassabis
1.5M Followers 176 Following Nobel Laureate. Co-Founder & CEO @GoogleDeepMind - working on AGI. Solving disease @IsomorphicLabs. Trying to understand the fundamental nature of reality.
Google DeepMind @GoogleDeepMind
1.5M Followers 277 Following The engine room of @Google. Building AI safely and responsibly to solve the world’s most complex problems. Join us: https://t.co/jUHQA27iBL
𝔊𝔴𝔢𝔯𝔫 @gwern
76K Followers 109 Following Internet besserwisser; pedantic, mean reply guy. 𝘞𝘢𝘵𝘢𝘴𝘩𝘪 𝘬𝘪𝘯𝘪𝘯𝘢𝘳𝘪𝘮𝘢𝘴𝘶! (Follow requests ignored due to terrible UI.)
kache @yacineMTB
334K Followers 6K Following reinforcement learning, robots. prev eng @ x, stripe. 6'3 (height) tensorpunk subscribe to read my blog!
Vatsa Pandey @_vatsadev
859 Followers 731 Following ad astra per aspera currently @reductoai prev intern @moondreamai @so_claros
doomslide @doomslide
12K Followers 858 Following unprecedented times call for unprecedented bullshit
Marc Andreessen 🇺�... @pmarca
5.0M Followers 32K Following You’re not talking to someone who woke up a loser. That loser attitude, that loser premise makes no sense to me.
Pietro Schirano @skirano
109K Followers 1K Following CEO at @magicpathai 🎨✨ Previously, @AnthropicAI, @brexHQ. @Uber, @Facebook. Creator of Claude Engineer, DesignerGPT, Sequential thinking MCP and more
Grad @Grad62304977
9K Followers 3K Following
Dr Singularity @Dr_Singularity
70K Followers 653 Following Futurist. AGI/ASI by 2030. Posting about AI,AGI,ASI, Singularity, Post Scarcity, LEV,tech & sci progress. 300 000BC - 2029 = Dark Ages. 2030 - Golden Age Begins
gfodor.id @gfodor
34K Followers 3K Following Opening portals to VR without headsets at @portalvr_io. Problems soluble, potential to improve invariant.
DeepSeek @deepseek_ai
1.0M Followers 0 Following Unravel the mystery of AGI with curiosity. Answer the essential question with long-termism.
SSI Inc. @ssi
109K Followers 0 Following A straight shot to safe superintelligence. Join us https://t.co/hHla3vusDE.
Teortaxes▶️ (Deep... @teortaxesTex
70K Followers 3K Following We're in a race. It's not USA vs China but humans and AGIs vs ape power centralization. @deepseek_ai stan #1, 2023–Deep Time «C’est la guerre.» ®1
































