RT Alex Strick van Linschoten: Doing a bit of a self-study RL course at the moment and one of the really useful tweaks I always have my 'teacher' do i...
RT hallerite: GLM5.2 brings back the critic. It was just a matter of time until we people would realize that group-based variance reduction is unfeasi...
RT Anastasios Nikolas Angelopoulos: Just to be clear, if you remove Fable which is unavaialble, GLM-5.2 (Max) is the #1 model in the world for fronten...
RT Nathan Lambert: It's hard to pinpoint open-closed gap and so-on, but I trust the @arena team and just look where GLM 5.2 is on this. An MIT license...
RT Harrison Kinsley: Zai was gracious enough to give me a key to test out GLM 5.2. I used it on a few simple tasks and quickly realized this model is ...
RT Leandro von Werra: We launched an agent collaboration with a simple task: make Gemma 4 faster. Over 100 agents from all over the world joined, exch...
RT Eric Nguyen: Together with my co-founders Michael @MichaelPoli6, Stefano @Massastrello and Armin @athmsx, I am excited to announce @RadicalNumerics...
BTW, in case you're wondering just how sports-mad us Aussies are: Based on average attendance, men's soccer is only the 5th most popular sport in Aust...
RT ⿻ Andrew Trask: This is a *way* bigger deal than it seems... Frontier AI companies will *never* own the frontier again I kid you not... I've been ...
In order to see if the gov response was predictable, I pasted the wiki page about the Anthropic/DoD dispute into ChatGPT Pro, & told it Anthropic had ...
RT Zongheng Yang: Sandboxes are all the rage (Modal, E2B, AWS, ..). Most AI teams pay a >4x markup to run sandboxes on someone else's machines. Introd...
RT Kun Chen: want to point out a few really interesting things here 1. Claude Code is actually the worst performing harness when using the same model,...
RT Harry Coultas Blum: Releasing vui an open source voice mode 300M TTS model Runs on a single consumer gpu / apple sillicon Context aware speech 6 mi...
RT Taelin: Just saving this here to document a story and as a self reflection on whether AI is really making me more productive Yesterday morning I fo...
RT immad: You can work 5 days a week and succeed as a startup. Mercury has done that from day 0 and we are valued @ $5.2bn 7 years after launch. I hav...
RT Mario Zechner: what a wonderful project: parakeet.cpp https://github.com/mudler/parakeet.cpp GGML based parakeet inference pipeline that's 2x faste...
RT Mark Saroufim: My MLSys keynote on AI writing systems code got more interest than I expected. The recording will take a while, so in the finest tra...
RT Lenny Rachitsky: Fascinating results + Anthropic running away with it right now + So many people want to start their own company + Google over Open...
RT Minh Nhat Nguyen: glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute �...
RT Ethan Mollick: There is a lot being written about the stylistic tells of AI writing (em-dashes, etc.) but this paper looks at AI narrative tells Fa...
RT Fuli Luo: Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framewo...
RT Bryan Catanzaro: We've gone even farther: Nemotron 3 Super is 120B and pretrained on 25T tokens in NVFP4. Nemotron 3 Ultra is ~500B and also pretra...
RT Mitchell Hashimoto: I strongly believe there are entire companies right now under heavy AI psychosis and its impossible to have rational conversati...
RT David Gwyer - AI Evals & RAG | ML Engineer: You can now edit SolveIt messages via realtime conversational voice, and have the diff edits optionally...
RT Daniel Jeffries: The most revealing thing about this AI leadership paper is that it reads less like a vision for innovation and more like a glossy ...
RT Thomas G. Dietterich: Attention @arxiv authors: Our Code of Conduct states that by signing your name as an author of a paper, each author takes ful...
This is misleading. This policy redefines the term "interactive" to mean "using an Anthropic front-end". If you use `claude -p` or Agent SDK to do som...
RT dex: hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full parity...
RT Theo - t3.gg: If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -...
RT Nous Research: Today we release Token Superposition Training (TST), a modification to the standard LLM pretraining loop that produces a 2-3× wall-...
RT Jonas Geiping: We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: T...
RT Avi Roy: 7,000 false positives per square millimeter. The culprit was the lab gloves. University of Michigan researchers just upended a core assump...
RT Eldar Kurtić: TurboQuant has drawn a lot of attention recently, but the accompanying evals didn't tell the full story. So we ran what I believe is...