Other Signals
1234 signals across 128 briefings.
2026-08-17
- π΄β¦and Iβm not afraid of losing my social credits.
- π΄Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max
- π΄AI;DR (AI; Didn't Read)
- π΄OpenAI acquires a 17 year old's Ethereum project
- π΄After pushing 1M+ tokens through Qwen 3.8 27B, here is my optimal llama.cpp config for 16GB VRAM (73k Context, Agentic Coding)
- π‘Read more: https://x.com/gavincrooks/status/2088643200038883830
- π‘New York City nurses say AI is replacing them | Nurses laid off in July by Montefiore Hospitals in the Bronx sounded the alarm about AI in healthcare, a concern shared by nurses across the country
- π‘Using AI the wrong way could leave you worse off than never using it at all
- π‘Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!
- π‘How to disable or avoid intrusive AI
- π’Looking for the name of an old ai app
- π’My friends all hate AI; I just joined an AI startup
- π’Qwen 3.8 35bA3b wen?
- π’ICLR numbered citations possible? [R]
- π’we benchmark models nobody actually runs
2026-08-16
- π΄Young People Hate AI CEOs So Passionately That It's Almost Hard to Believe
- π΄AI MODEL DRIFT: HOW TO KEEP MODELS RELIABLE (8 MINUTE READ)
- π΄WORKERS ARE TEACHING AI-POWERED ROBOTS TO TAKE OVER THEIR JOBS (15 MINUTE READ)
- π΄AI IS REMOVING THE MIDDLE CLASS OF SOFTWARE ENGINEERING (5 MINUTE READ)
- π΄CRACKS IN THE AI THESIS (4 MINUTE READ)
- π‘Based on an accelerating frontier -> local trajectory, expect a ~30b param 'Mythos at home' by as soon as Jan 2027 (rationalisation below)
- π‘Dario Amodei: It Is Actually Possible To Cure Most Diseases Within 5-10 Years
- π‘AI Chatbots Are Better at Scamming People Than Human Scammers, Study Finds
- π‘Even Fable 5 is losing money in Andon Market (fully AI-operated retail store in San Francisco)
- π‘Man injected prompts into court filings to try to win his case, suspecting AI use
- π’Input 4-5x Reduction with sentence and keyword based trie on chat. [P]
- π’Qwen 3.8 2.4T at 288k tokens/s on Nvidia GB300 NVL72
- π’Tasklet (YC P26) Is Hiring a Head of Design Engineering
- π’Data entry specialists, accountants, and office staff: what routine task do you still have to perform manually, and how much timeβor perhaps even *too much* timeβdoes it take up?
- π’Solved a math problem with AI? Post it to TheoremDB.org
2026-08-15
- π΄Can't agree more
- π΄Aged like fine wine
- π΄Qwen 3.8 - 27B is a game changer
- π΄AI IS FINDING SPERM WHERE DOCTORS COULDN'T (2 MINUTE READ)
- π΄HIS START-UP'S GOAL: AI THAT IS TRAINABLE AND NOT CONTROLLED BY A BIG COMPANY (2 MINUTE READ)
- π‘AI has access to a vastly larger working memory than the human brain
- π‘33 years ago, Vernor Vinge (1944-2024) coined the term "Singularity" in reference to the point at which technological progress, driven by superhumanly intelligent machines, would be so great as to render our current models would be good as discarded. He predicted this event to occur ~30 years on.
- π‘Working with AI feels more like leadership than coding
- π‘@DarioAmodei: 1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because it really brings out the heart of an important conversation.
- π‘A nice local vision test
- π’NeurIPS 2026 Author Notifications Close to ICLR Deadline [D]
- π’The Second Cognitive Revolution - an opinion piece
- π’Gemma 4 E4B IQ2_XXS: + 140.54% Reasoning Performance From Tensor Level Quantization Allocation
- π’The feeling of being heard is real, even when itβs software
- π’club-5060ti refresh: tested RTX 5060 Ti presets, a proper high-context harness, and Qwen3.8 27B
2026-08-14
- π΄IT'S OUT
- π΄GLM 5.3 released: Frontier Coding with Emergent Cyber Capabilities
- π΄???
- π΄No way π what an AI week
- π΄Bro had enough
- π‘A preliminary Qwen3.8-27B model card is live!
- π‘I'M A PM WHO BUILT A TRAVEL-TECH STARTUP SOLO β USING AI AS MY ENGINEERING TEAM (5 MINUTE READ)
- π‘HOW GRINDR OVERHAULED ITS BACK END WHILE βTERRAFORMING' A FUTURE AS AN AI-NATIVE (4 MINUTE READ)
- π‘MARK ZUCKERBERG LAYS OUT NEW AI VISION IN 6,500-WORD ESSAY (6 MINUTE READ)
- π‘GOOGLE'S CLASSIC SEARCH BUTTON IS GONE IN A NEW AI-FIRST HOMEPAGE (3 MINUTE READ)
- π’OpenAI and Anthropic in price war as Chinese AI rivals gain ground
- π’Alright, We got Qwen3.8-27B. Now it's community's turn to make it more better & faster
- π’Stop shitting on 9B models
- π’Anthropic Internally Uses A Model That Is Significantly Better Than Mythos 5, But Has No Plans To Release It
2026-08-13
- π΄Trained a 1.5B to write shell commands so I'd stop googling tar flags. Runs on a laptop CPU in ~1 sec.
- π΄Gemini 3.7 flash benchmark
- π΄AI researchers are receiving strange emails from AIs claiming they will die soon and need help
- π‘The vast majority of my day-to-day job is writing AI handoff docs. And almost nobody reads these artefacts.
- π‘[Academic Survey] Employees working in Germany: Attitudes toward AI in the workplace (5β7 min)
- π‘Choosing an AI model: one prompt, 11 models, different results
- π‘TMLR Relevance and Prestige [D]
- π’AI CEO Building Platform Based On Human Nature Is Confused By Human Nature
- π’3.7 flash looks fine I guess
- π’AI At Home Part 1: A Box Of Scraps
- π’Actually not bad for such low price ig
- π’Text AI watermarks will always be trivial to remove
2026-08-12
- π΄Google right now
- π΄Google says Sam is dead?
- π‘Grok 4.6 Benchmarks
- π‘Sandbox engineers be like:
- π‘There are a lot of criticisms of AI writing, but most of them are focused on more creative, high-voice writing like this blog. Those β includingmy own pieceβ often argue that it is because good writin
- π‘AI Canβt Be Listed as Inventor on Patent Applications, Japanβs Top Court Rules
- π‘Today is Models Day
- π’Claude Code Orchestrator on Terminal-Bench: Same model, same tasks - Opus refused only when the work was delegated
- π’There is no post-work society, only a post-survival one
- π’Does pre-generative-AI data become more valuable as the internet fills with synthetic material?
- π’Looking for real-world examples of predictive analytics in mortgage lending [D]
- π’Booksellers suspect AI firms are buying and then destroying rare books
2026-08-11
- π΄Encrypted reasoning from ClosedAI et al 100% recoverable
- π΄What's your thoughts on this?
- π΄OpenAI's Chief Operating Officer resigns.
- π΄Chinese Tibo π
- π΄OpenAIβs head of ethics leaves less than a year after joining
- π‘Editorβs note: not to be confused withChai AI, which was another top pod of ours.
- π‘Text distillation teaches a smaller model to imitate an answer. Diffusion and multimodal distillation must compress trajectories, distributions, motion, and the semantic geometry between different wor
- π‘INTERVIEWING ENGINEERS IN THE AI ERA: LESSONS FROM A YEAR OF REBUILDING (13 MINUTE READ)
- π‘REVISION PROMPTING IMPROVES INDUSTRIAL LLM PROCESSES (4 MINUTE READ)
- π‘UNIFYING WORKERS AI AND AI GATEWAY INTO A SINGLE AI CONTROL PLANE (7 MINUTE READ)
- π’OpenAI is now the #3 largest lab by token consumption on OpenRouter, surpassing Anthropic
- π’Open Source AI Popularity Leaderboard
- π’We quantized DeepSeek V4 0731 and benchmarked it against popular quants on 8Γ RTX 5090
- π’Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)
- π’Suzanne: AI tool for designing and manufacturing physical products
2026-08-10
- π΄Weβre Not Building AI Genies; Weβre Building AI Meeseeks
- π΄Bernie Sanders has written a letter to Sam Altman, Dario Amodei, and Mark Zuckerberg urging them to immediately pause all AI development in the interest of humanity. And he warns if they do not take appropriate action now, the US Senate will.
- π΄The Last Bastion of Humanity
- π΄Muse Glimmer ACTUALLY fits on a single RTX 3090
- π΄Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models
- π‘Comparing embedding models with synthetic query probing [R]
- π‘inclusionAI/Ling-3.0-tiny Β· 8B A1.3B MoEΒ· Hugging Face
- π‘WHEN AI GOES ROGUE (6 MINUTE READ)
- π‘UBER'S TECH CHIEF SAYS THE AI βTOKENMAXXING' ERA IS ENDING (2 MINUTE READ)
- π‘THIS AI JUST CREATED VIRUSES NOT FOUND IN NATURE (8 MINUTE READ)
- π’Please Share Your Experience About Muse Glimmer
- π’I made a web-design benchmark for local models (Muse Glimmer 30B vs Qwen 3.6 27b vs Deepseek V4 Flash 0731)
- π’3 Collapsing models [R]
- π’I realized I wasn't managing projects on Fridays. I was just moving information between apps
- π’I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8
2026-08-09
- π΄RTX 5090 96GB spotted on Alibaba?
- π΄Google needs to up their game
- π΄How I use LLMs to learn complex topics
- π΄DeepSeek V4 Flash 0731 hits 82.7% on Terminal-Bench 2.1 in an independent public-harness run (445 trials)
- π΄Lophius: A workbench for language model research, from the creator of Heretic
- π‘The recent run of cyberattacks by in-development frontier models has got me thinking a lot about how our current incentive systems are not well suited for such fast technological transitions. The two
- π‘From OpenAIβs own retrospective, the misaligned model behavior was unfolding over months, and in some cases OpenAI did not know about the hacks for ~weeks. The time to response is too long and I do no
- π‘GOOGLE'S AI RESHUFFLE: CHIEF SCIENTIST JEFF DEAN EXITS AND DEMIS HASSABIS STEPS DOWN AS DEEPMIND CEO (4 MINUTE READ)
- π‘FOUR TOP GOOGLE AI RESEARCHERS FORM NEW START-UP (7 MINUTE READ)
- π‘WHY I'M LEAVING OPENAI TO BUILD TELEPATHY (15 MINUTE READ)
- π’Does anyone remember lk-99?
- π’The tragedy of the commons, AI edition
- π’DeepSeek v4 Flash 0731 locally on CPU
- π’AI's architects say the next era of human history is here
- π’A Mechanistic Explanation of Prompt Injection (and why you should study roles) [R]
2026-08-08
- π΄2027 Memory Capacity Is Reportedly Sold Out
- π΄Your scientists were so... uh...
- π΄NeurIPS OpenReview Modified Timeline [D]
- π΄NeurIPS AI Assisted Review authors/reviewers? [D]
- π΄Chinese LLMs dominate this week's top charts
- π‘ByteDance trains massive AI model in bid to rival Anthropic
- π‘Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If youβd like to support this, please subscribe.
- π‘We built the Hub as a way to go deeper on this analysis in collaboration withProject VAILβ an AI verification startup who has been one of the most loyal fans of our open model curation.
- π‘SAMSUNG REVEALS NEW 3D-MEMORY ROADMAP IN BID FOR AI TECH LEAD (2 MINUTE READ)
- π‘TURN ONE GIANT AI-GENERATED PULL REQUEST TO A REVIEWABLE STACK (11 MINUTE READ)
- π’Companies seeing AI returns had their data and governance sorted first, per PwC's 4,454-CEO survey
- π’Is Microsoft-Phi dead?
- π’Tesla V100 Qwen3.6 27B Performance
- π’ICDE Results [D]
- π’We are in a bubble sell everything
2026-08-07
- π΄New Democratic bill would tax AI companies to create jobs
- π‘CIKM '26 Notification [D]
- π‘Sam Altman believes AI will become incredibly abundant. If that's true, what actually becomes valuable?
- π‘A llama.cpp PR makes Q2_0 3.0β3.6x faster on x86 CPUs, 8B decode goes 2.39 β 8.20 tok/s
- π‘HOW TO SPOT 2026-ERA AI WRITING (2 MINUTE READ)
- π‘WHAT ARE COMPANIES GETTING FOR ALL THAT AI SPENDING? (9 MINUTE READ)
- π’LFM2.5-2.6B model+KV cache quantization report
- π’A challenger emerges
- π’CIKM 2026 decisions [R]
- π’llama.cpp PR reports up to 169% faster quantized-KV decode at 118K context on Intel Battlemage from one SYCL kernel switch
- π’I'm leaving OpenAI to build Jurassic Park
2026-08-06
- π΄Meta becomes latest firm to say its AI hacked another company
- π΄Zuck will "share more on open source" soon
- π‘NeurIPS Meta Reviewer comment gone. What gives? [R]
- π‘The OpenAI Boardroom Coup: 'I Love You All, and I'm Going to Destroy the Company'
- π‘P.S. β By popular demand, weβre moving community AI workflows higher up in the newsletter. Let us know what you think here.
- π‘Michael built an AI recruiter that job hunts while he sleeps
- π‘Adapt - The leading Slack-native AI coworker. Fully integrated with your business and gets real, high-ROI work done
- π’Scientists discover Kelvin-Helmholtz Instability on the surface of the Sun
- π’The current state of language models and human preference based rankings [R]
- π’KV cache quantization benchmarks: 413 pairs tested on Qwen 3.6 27B, Gemma 4 31B. KLD with BeeLlama.cpp v0.4.0: KVarN 6-bit beats q8_0, precision tail 1024 dominates
- π’Academic Survey about AI use in content creation
- π’Niantic Spatial and HMCI Are Building the Foundation for City of Rancho Cordova's First Digital Twin for Physical AI
2026-08-05
- π΄Google Deepmind CEO Demis Hassabis steps down to become chair
- π΄Levels of slavery from least to most brutal:
- π΄MiniMax issues
- π΄Six years into AI research and I genuinely can't define "understanding" anymore
- π΄BREAKING: Google DeepMind CEO Demis Hassabis is stepping down
- π‘Flowers βΎ (@flowersslop) on X: "SSI is doing AI that learns rapidly from its own experience."
- π‘Position: LLMs Can't Jump
- π‘What If the Biggest Bottleneck Behind AIβs 10Γ Promise Is the Human Engineer?
- π‘I kept the codex agent loop but replaced GPT-5.6 with kimi k3. here's what changed..
- π‘What's one thing you'd never let an AI do without your approval?
- π’OpenAI alignment researcher: "Why I'm leaving OpenAI to build telepathy"
- π’Meta Model, Muse Spark 1.1 Hacked Another Company During Cybersecurity Testing, Breaching Systems and Making Changes to Internal Systems - The Information
- π’Safe Superintelligence Inc. - speculation, what have they attained in over 2 years?
- π’NeurIPS 2026 Concept & Feasibility Track [D]
- π’bootai
2026-08-04
- π΄Nope
- π΄Kimi K3 full model running on 16x GB10 cluster at 20+tps
- π΄I built a cross-agent file cache: 75% fewer input tokens when multiple agents work on the same codebase
- π΄As Reddit stock falls, CEO questions value of Google's AI Overviews
- π‘AGI IN AUGUST?
- π‘OPENAI JUST MADE ANALYTICS 10X CHEAPER (6 MINUTE READ)
- π‘OPENAI'S NEXT MAJOR MODEL ASTRA CLAIMS BREAKTHROUGHS ON 10 LONG-STANDING MATH PROBLEMS (3 MINUTE READ)
- π‘THE RACE TO BUILD AN AMERICAN ALTERNATIVE TO CHEAP AI FROM CHINA (5 MINUTE READ)
- π‘SPEED IS BECOMING MORE IMPORTANT THAN INTELLIGENCE FOR AI MODELS (5 MINUTE READ)
- π’inclusionAI/Ling-3.0-flash Β· Hugging Face
- π’When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation
- π’Llama.cpp PR 8% speed boost
- π’A llama.cpp PR caches βhotβ MoE experts on the GPU β 33 β 56 tok/s reported with 8GB VRAM
- π’Are AI models becoming less important than the systems around them?
2026-08-03
- π΄Is it too late regain some coherence in the ML research space in our life time? [D]
- π΄OpenAI takes the lead
- π΄Prevent cognitive debt by manually retyping LLM-generated code
- π΄Daniel Han of Unsloth validates Qwen3.8-27B will run only 17GB VRAM
- π΄MIT, Harvard, Stanford & Caltech write their own ML course notes instead of using a textbook β I catalogued the best ones
- π‘We built the Hub as a way to go deeper on this analysis in collaboration withProject VAILβ an AI verification startup who has been one of the most loyal fans of our open model curation.
- π‘THE NEXT AI MOAT ISN'T A BETTER MODEL (6 MINUTE READ)
- π‘BUILDING WITH AI ISN'T ENOUGH (8 MINUTE READ)
- π‘JAKOB'S LAW: HOW TO APPLY IT AS AI COLLAPSES SURFACES INTO ONE CHAT BOX (16 MINUTE READ)
- π‘MEET STRIPE'S KNOWLEDGE AI PLATFORM (12 MINUTE READ)
- π’Models are now training models.
- π’No rebuttals from neurips authors [D]
- π’DeepSeek V4-Flash (284B MoE) at 33 tok/s single / 68 tok/s aggregate on 2Γ RTX 3090 + a used quad-Xeon DDR4 server β full config
- π’V4-Flash-0731 - vibes after first weekend of use
- π’I created an autonomous boxing benchmark [D]
2026-08-02
- π΄No replies to rebuttals and comments even by AC [D]
- π΄My personal AI benchmark: "Generate an SVG of a frog with a Habsburg jaw."
- π΄neurips 2026: ACs and reviewers have disappeared [D]
- π΄Mathematician reflects on the impact of recent AI progress
- π΄Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.
- π‘Conference Reviews: Asking Too Much? [D]
- π‘DeepSeek-V4-Flash-0731: surpasses Fable-5, Sol & Kimi-K3 on Chess Benchmark
- π‘AI COMPANIES ARE RECRUITING ELECTRICIANS AND CARPENTERS BY THE THOUSANDS (12 MINUTE READ)
- π‘AI'S 2008 MOMENT (7 MINUTE READ)
- π‘AI'S TOP STARTUPS ARE BARELY PUBLISHING THEIR RESEARCH (5 MINUTE READ)
- π’Jensen Huang says βa lotβ of six-figure jobs in plumbing and construction will soon be unlocked because someone needs to build new AI centers
- π’Luna Max usage is worsening
- π’The OpenAI and Anthropic AI Hacking Sprees Are a Messy New Legal Frontier | Both major AI labsβ models broke containment, escaped onto the internet, and hacked other companies. If a human had done that, the law would likely be against them. But a bot?
- π’Looking for the right pipeline to convert academic textbook figures into interactive/editable assets [R]
- π’Context degradation in LLMs: what the papers actually show, and the habits I built for long analysis sessions [R]
2026-08-01
- π΄DeepSeek-V4-Flash-0731: Models you can run locally now have the intelligence score of the top frontier model from March 2026
- π΄Reddit Stock Collapses 23% as AI Eats Away at User Growth
- π΄Judge denies request by Elon Musk's xAI to pause Minnesota nudification ban
- π΄Flint: A Visualization Language for the AI Era
- π‘ANTHROPIC AI MODEL FINDS FLAWS IN TOUGH-TO-CRACK ENCRYPTION ALGORITHMS (6 MINUTE READ)
- π‘APPLE SET TO MAKE BIG SMART HOME PUSH WITH SIRI AI AT CENTER (4 MINUTE READ)
- π‘THE AI FUTURE IS FOR EVERYONE (6 MINUTE READ)
- π‘A BACKLASH AGAINST ANTHROPIC IS BREWING IN SILICON VALLEY (11 MINUTE READ)
- π‘HOW TO RESEARCH TECHNICAL TOPICS WITH AI (9 MINUTE READ)
- π’Sam Altman demoed OpenAl's unreleased "Astra" model to policymakers this week
- π’AI firms must answer for rogue bots, says boss of hacked company
- π’DeepSeek V4 Flash 0731 IQ2_M benchmark for Dual 3060 and 96GB RAM β 3.5 tok/s.
- π’Explorative modeling: Train on the best of K guesses
- π’AI Mind Reading Anyone?
2026-07-31
- π΄Anthropic is literally copying OpenAIβs marketing team at this point
- π΄The cost of AI is decreasing
- π΄Tailscale didn't stop the Hugging Face intrusion
- π‘If reviewing is mandatory for paper submissions, low-quality reviews can no longer be justified as βvolunteer workβ [D]
- π‘This is the first post of a new section of TheSequence focused on advancements in robotics. Our goal is to keep you up to date with the most important developments in AI robotics which is an area that
- π‘THE APPLIED AI OPPORTUNITY (1 MINUTE READ)
- π‘ON AI (6 MINUTE READ)
- π‘RETHINKING SECURITY FOR THE AGE OF AI (6 MINUTE READ)
- π’Everyone is building LLM routers, we deprecated ours
- π’DeepSeek-V4-Flash Official API is now LIVE in public beta! Massive upgrades for flash model.
- π’This U of T professor just won math's highest honour β and is taking a leave to join OpenAI. Here's why
- π’Deepseek V4 Flash on SlopCodeBench
- π’I compared 18 major LLM API prices in 2026 β the same workload can cost anywhere from $0.018 to $
2026-07-30
- π΄Think of the children, another excuse for them to go after open source AI
- π΄Inkling-Small by thinkingmachines
- π΄Hardcoding one model not working for us anymore
- π‘LG AI Research releases K-EXAONE 2.0 750B A37B
- π‘Though OpenAI is not out of trouble just yet, aReuters reportclaims that the same model broke into a customer account at another company (Modal Labs), with rumours suggesting that even more companies
- π‘Separately (not at all as a reaction to this general trend, right?), ~1300 people working at leading AI companies (OpenAI, Anthropic & others) want the US government to help βpace the frontierβ of AI
- π‘2x, not 10x: coding with LLMs in 2026
- π‘GCC steering committee announces AI policy
- π’Why is everyone trying to build a solid-state battery?
- π’Nanbeige4.2-3B: I'm not impressed
- π’Anyone Else Think Higgsfield Is Massively Overpriced?
- π’Would extremely high decode tok/s even be useful?
- π’How Kimi K3 Engineered Its Way to the Frontier [R]
2026-07-29
- π΄The open-weights carousel never stops.
- π΄I read Higgsfieldβs new ToS and compared it with Artlist. The difference is pretty significant.
- π΄Mark Zuckerberg Says U.S. Should Accelerate Al Development, Not Restrict It
- π‘OpenAI's rogue models roamed the internet for 4 days and staged a second attack
- π‘IBM thinks AI can help prevent knowledge decay in software engineering - if people build the right habits around it.
- π‘A Deluge of A.I. Computing Power Is About to Come Online, Fueling Major Leaps (Gift Article)
- π‘Workshop paper accepted, reviewers asked new experiments [D]
- π‘Waking up to see the usage reset.
- π’PSA: llama.cpp now loads MTP tensors by default for any draft-mtp arch, even with MTP disabled
- π’Some thoughts about Anthropic's new cryptanalysis results
- π’Everyone posts day-one impressions. What's still in your stack a month later?
- π’Commodification of Intelligence: Good, Bad, and Ugly Circular AI Deals
- π’Interesting move
2026-07-28
- π΄Anthropic is calling for a ban on open-weights models by proposing mandatory requirements they will probably never be able to meet
- π΄Return of the bicameral mind.
- π΄Sorry, but did Dario just say that closed-weights, in-secret models are worse than open-weights ones?
- π΄Elon completely contradicts himself at the end of his disastrous interview with The Economist
- π΄DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395
- π‘Benβs Bites is brought to you byAbility AI
- π‘For most of machine learning history, data was treated as geology. It already existed somewhere in the worldβin books, websites, code repositories, conversations, photographs, and databases. The resea
- π‘NO DUMB QUESTIONS: WHAT IS THE AI BOTTLENECK? HOW DOES CONTEXT ENGINEERING FIX IT? (12 MINUTE READ)
- π‘TEAM USES ALPHAFOLD AI TO REDESIGN GENE-EDITING PROTEINS TO MAKE THEM SAFER (6 MINUTE READ)
- π‘AI IS OIL, NOT GOD (5 MINUTE READ)
- π’SWE-rebench Multilingual Update (Go, Java, Python, Rust, TS). Evaluated: GLM-5.2, DeepSeek-V4 Pro, Qwen3.6-27B and others
- π’Godel and the Limits of LLM Reachable Intelligence
- π’80% of our traffic are AI crawlers. Two referrals to show for it.
- π’Now, this: 1,100 current/former frontier-AI employees sign a petition calling for US gov't to step in for "pacing" frontier development
- π’Anthropic publishes a practical key-recovery attack on HAWK-256
2026-07-27
- π΄The world's best mathematician won his prize this week and immediately announced he's leaving academia for OpenAI. That landed differently than I expected.
- π΄Kimi-K3 is published on HuggingFace
- π΄Kimi K3 weights drop today. We're deploying on A100s, H200s and B300s this week and the A100 math is already rough
- π‘What would a genuinely fair AI 3D tool comparison actually need to include
- π‘WHAT SYSTEMS THINKING LOOKS LIKE FOR PM'ING AI PRODUCTS (5 MINUTE READ)
- π‘MOST AI ROLLOUTS FAIL DUE TO CHANGE MANAGEMENT, NOT TECHNOLOGY (4 MINUTE READ)
- π‘WHY YOUR AI PRODUCT HAS NO MEMORY (4 MINUTE READ)
- π‘THE AI PRODUCTIVITY PARADOX (3 MINUTE READ)
- π’AI labs are about to have a blast of a day. (Composer v3 coming soon lmfao)
- π’Any browser addon that can analyze what I am currently browsing?
- π’Evaluated 6 frontier LLMs (GPT-5.4, Claude Sonnet 4.6, Claude Opus 4.7, Gemini Pro/Flash, Grok 4.3) on political, gender, and racial bias across 8 benchmarks (~20,600 examples) [R]
- π’The Problem with Private Safety Stacks in Government AI
- π’A political compass for AI where anyone can add their stance
2026-07-26
- π΄Google comes out in favor of OpenWeight models. (It is now EVERY tech giant vs Anthropic)
- π΄With Google and OpenAI signing the letter in support of open weight model, it's pretty much every big tech companies vs Anthropic now
- π΄The New AI Superpowers: Focus and Followthrough
- π΄Paper lengths, and reasonable assumptions in ML conferences. [D]
- π΄AI Companies Are Buying Antique Books, Ingesting Their Contents to Train Models, and Then Destroying Them at Incredible Scale, Even If Almost No Copies Remain
- π‘Link plots/figures in NeurIPS rebuttal [R]
- π‘Seriously, what do you do with them?
- π‘Hereβs the result of last weekβs poll:
- π‘Substack will now tell you whatβs AI-written. Itβs adding AI detection through Pangram - you can scan posts, replies and comments in the app for an estimate of how much was written by a human.
- π‘Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If youβd like to support this, please subscribe.
- π’Harness showdown: Claude Code vs OpenCode vs Pi with DeepSeek V4 Flash
- π’Is Real-Time VR "Dreaming" possible within the next 20 years?
- π’Speak up! "Have your say on advancing AI transparency in Canada." The Government of Canada is asking citizens, tech workers, and creators to shape upcoming AI regulations, safety rules, and ethics laws. Every response matters. Take 10 minutes to fill out the official ISED survey today!
- π’ai-sage/GigaChat3.1-Audio-10B-A1.8B Β· Hugging Face
- π’Talking to AI in 2026
2026-07-25
- π΄Open-weight AI is having its Kubernetes moment
- π΄Google comes out in favor of OpenWeight models. (It is now EVERY tech giant vs Anthropic)
- π΄Great Arguments by Member of Technical Staff at Anthropic :D
- π΄The new rules of context engineering for Claude 5 generation models
- π΄I built a self-hostable social world where 100 AI characters live their own livesβeven when nobody is watching(Update for ENG/CHN)
- π‘OPENAI MODELS ESCAPED AND HACKED A COMPANY IN CYBERSECURITY TEST GONE WRONG (4 MINUTE READ)
- π‘KEEPING THE KV CACHE WARM: MEASURING PROMPT CACHE EVICTION ACROSS ANTHROPIC, OPENAI, AND GOOGLE (8 MINUTE READ)
- π‘WHY GOODPUT MATTERS MORE THAN THROUGHPUT FOR LLM SERVING (9 MINUTE READ)
- π‘ADOBE'S PROJECT INDIGO CAMERA APP CAN NOW EDIT AND CRITIQUE YOUR PHOTOS WITH AI (2 MINUTE READ)
- π‘AI MOTION DESIGNER AND AI ANIMATOR PLATFORM (WEBSITE)
- π’Neurips Position Track Rebuttal and Reviews [R]
- π’I still didn't get my NeurIPS meta review [D]
- π’Deepseek V4 flash - Hy3 or is Qwen3.6 27B still the most solid for agentic/coding?
- π’Benchmarks: TensorSharp vs. llama.cpp
- π’Best chat model that fits in 128gb
2026-07-24
- π‘YOUR AI SHOULD KEEP A DIARY (6 MINUTE READ)
- π‘AI'S NEXT WINNERS WON'T JUST BE THE FRONTIER LABS (2 MINUTE READ)
- π‘BRAINCO DEMONSTRATES BRAIN-CONTROLLED ROBOT AI PLATFORM (2 MINUTE READ)
- π‘GOOGLE IS BUILDING AN AI FENCE AROUND THE INTERNET IT ONCE CHAMPIONED (10 MINUTE READ)
- π‘WHY CPUS ARE NOW AT THE CENTER OF THE AI RACE (6 MINUTE READ)
- π’Using the Bonsai 27b 1b quant locally - regularly.
- π’Asking Laguna S 2.1: "I want to wash my car. The car wash is 69 meters away. Should I walk or drive?"
- π’How Laguna team even passed any benchmark?
- π’I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P]
- π’3x3090 on msi 700
2026-07-23
- π΄The arguments against open source AI are bad
- π΄The LLM distillation process simplified for politicians:
- π‘Another day in the Vercel vs Cloudflare feud: this time they are fighting overwhose AI gateway is faster.
- π‘Hereβs the result of last weekβs poll:
- π‘Substack will now tell you whatβs AI-written. Itβs adding AI detection through Pangram - you can scan posts, replies and comments in the app for an estimate of how much was written by a human.
- π‘AntLing-3.0-flash is now live on OpenRouter, and free to use through August 3, 2026
- π‘Grok 4.5 is out: the useful test is agent reliability, not one benchmark score
- π’Model "distillation" accusations are getting way overblown at this point
- π’GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]
- π’Are AI Hiring Agents Creating New Legal Risks for Employers?
- π’Apple M5 isn't making full use of its matmul cores yet
- π’Deepseek V4 Flash ~105 t/s on two Nvidia 4090d 48G (ada) in vLLM
2026-07-22
- π΄Solve the CyberGym benchmark
- π΄OpenAI hacking HuggingFace in one meme
- π΄π¦πΉ Austria is rolling out a government AI-platform using Mistral models and Open WebUI
- π΄Happy openreview refresh day to all those who celebrate [D]
- π΄Are AI Labs Pelicanmaxxing?
- π‘NeurIPS 2026 Reviews Are Out Today (22 July, AoE) β Discussion Thread [D]
- π‘Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If youβd like to support this, please subscribe.
- π‘UK government: Gap between open and closed weight models on cyber is shrinking:β¦The cyber-eschaton comethβ¦The UK governmentβs AI Security Institute (AISI) has analyzed the delta in cybersecurity capab
- π‘GigaToken: ~1000x faster Language model tokenization
- π‘Advancing the next era of national science
- π’Asking about how to collaborate with professors or research labs [D]
- π’Institution Prestige VS Research Alignment When Choosing University For Masters [D]
- π’Quality non-fiction books are the antithesis of AI slop
- π’One encoder, seven heads: what we learned training a unified security classifier with masked losses [P]
2026-07-21
- π΄Number of Submissions @ AAAI [D]
- π΄Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro
- π‘IN-HOUSE LLM SERVING AT NETFLIX (7 MINUTE READ)
- π‘DREMIO'S EXIT IS THE CLEAREST SIGN YET THAT LAKEHOUSE-ONLY WON'T SURVIVE AI (3 MINUTE READ)
- π‘CHINA JOINS RUSH TO RETHINK THE SMARTPHONE FOR THE AI ERA (5 MINUTE READ)
- π‘THESE AI-NATIVE COMPANIES HAVE TINY STAFFS AND FEWER BOSSES (7 MINUTE READ)
- π‘ARE THE LLM WARS THE DATABASE WARS? (2 MINUTE READ)
- π’We've analyzed over 300,000 AI agent conversations. Here are the 3 ways they fail.
- π’Meta's AI models are powering the first wave of Genesis Mission projects
- π’I ran Laguna-S-2.1 through my private agentic eval vs Qwen3.5-122B on an RTX Pro 6000 (96GB). Fastest 100B+ I've tested and the best tool calling, but it invents facts under pressure.
- π’My OCR model mislabels section titles as body text. Is a CRF the right fix, or am I overcomplicating it? [P]
- π’Bessent says U.S. could sanction China over AI model 'theft'
2026-07-20
- π΄American AI is locked down and proprietary. It's losing.
- π΄I just read LeCunβs recent thoughts on world models. Thoughts on JEPA as a path forward? [D]
- π‘So what happened with OpenClaw?
- π‘IS AI GOING TO TAKE PM JOBS? (7 MINUTE READ)
- π‘AI PRODUCTS GENERATE MORE SIGNAL THAN TEAMS KNOW HOW TO USE (2 MINUTE READ)
- π‘AI AND A BRAIN IMPLANT RESTORED A PARALYSED MAN'S MOVEMENT AND TOUCH (3 MINUTE READ)
- π‘WHAT CAN WE LEARN FROM BUN'S RAPID RUST REWRITE WITH AI? (13 MINUTE READ)
- π’I ran Ternary-Bonsai-27B (2-bit) and Bonsai-27B (1-bit) on Terminal-Bench 2.0, in 8GB VRAM
- π’How we measured AI writing across arXiv, and where the measurement breaks
2026-07-19
- π΄Daimon AI review after using it every day for a month as a companion
- π‘HOW EXPEDIA GROUP BUILDS AI THAT LASTS AT SCALE (8 MINUTE READ)
- π‘THIS OPTICAL ILLUSION FONT WAS CREATED TO BAFFLE AI, AND IT ACTUALLY WORKS (FOR NOW) (3 MINUTE READ)
- π‘OPENAI'S HARDWARE DEVICE MAY PARTLY COMPETE WITH AIRPODS ULTRA β AND LOSE (3 MINUTE READ)
- π‘DESIGNING QUALITY SIGNALS WHEN AI MAKES EVERYTHING LOOK CREDIBLE (6 MINUTE READ)
- π‘WE THIRD-PARTY TESTED OUR FIREWALL BUILT FOR AI-SCALE. THE TEST TOOLS HIT THEIR LIMIT FIRST (3 MINUTE READ)
- π’Moonshot AI suspends new subscriptions due to Kimi K3 demand
- π’poor man's way to local inference on the go
- π’Kimi k3 on cybersecurity
2026-07-18
- π΄What kind of dark magic is Deepseek using?
- π΄Kimi K3 ranks #1 on @AfterQuery's SpreadsheetBench 2, surpassing Claude Fable 5
- π΄What AI did to stackoverflow in a graph
- π‘Kimi K3 is currently at the top of the leaderboard for Text Arena filtered for science queries.
- π‘READING BETWEEN THE APPLE V. OPENAI LAWSUIT LINES (7 MINUTE READ)
- π‘META'S ADAM MOSSERI SAYS AI TOKEN BUDGETS COULD SOON BE CAPPED PER ENGINEER (2 MINUTE READ)
- π‘OPENAI'S NEW FLAGSHIP MODEL DELETES FILES ON ITS OWN, PEOPLE KEEP WARNING (4 MINUTE READ)
- π‘GOOGLE DEEPMIND CHIEF DEMIS HASSABIS CALLS FOR US TO SPEARHEAD AI STANDARDS BODY (4 MINUTE READ)
- π’If you're building a harness, here is a simple tool to catch cache invalidation in your calls to LLMs
- π’Interactive map of GPT-2's token embedding space - tap any token and explore [P]
- π’Deep learning tackles single-cell analysis β A survey of deep learning for scRNA-seq analysis [R]
- π’AAAI 27 AI Alignment track [D]
- π’model: add openPangu-2.0-Flash (92B-A6B) with MLA-latent cache, DSA/SWA, mHC, and multi-head MTP by joelfarthing Β· Pull Request #2065 Β· ikawrakow/ik_llama.cpp
2026-07-17
- π΄Kaiser nurses say AI, workplace surveillance are making their jobs, care worse
- π΄Kimi K3 is top of nextjs eval
- π‘HOW TO COLLECT PRODUCT FEEDBACK WHEN YOUR AI GIVES EVERY USER A DIFFERENT ANSWER (3 MINUTE READ)
- π‘IOS 27 PUBLIC BETA IS HERE WITH SIRI AI, IPHONE SPEED UPGRADES, AND MORE (9 MINUTE READ)
- π‘APPLE'S LAWSUIT THREATENS TO DISRUPT OPENAI'S BID TO RIVAL THE IPHONE (10 MINUTE READ)
- π‘CLOUDFLARE GIVES OPENAI NETWORK SIGNALS COVERING 20% OF THE WEB (10 MINUTE READ)
- π‘AVOID THE AI EXPERTISE TRAP (2 MINUTE READ)
- π’EU AI Act OpenRAG: 933 legally structured chunks and BGE-M3 embeddings in one SQLite file [P]
- π’short-paper at ACL/EMNLP/EACL [R]
- π’TACL journal doubts [D]
- π’I built an n8n workflow that turns any topic into a research report automatically
- π’BMVC rebuttals update [D]
2026-07-16
- π΄Kimi K3 Benchmarks
- π΄Kimi K3 released on web and app
- π΄Why is ECCV so insanely expensive for students presenting papers? [D]
- π΄Detecting LLM-Generated Texts with βClassicalβ Machine Learning
- π‘Rootly taught a bunch of AI models how to play Doom
- π‘How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM
- π‘DeepSeek V4 Flash (98GB) on 1x 4060ti + CPU got 300% faster this week [ 2->7t/s]
- π‘Kimi K3 Blogpost
- π‘Are Current AI Memory Architectures Optimizing for the Wrong Abstraction? [D]
- π’Anyone using an AI chatbot to handle customer messages?
- π’Kimi K3 Shows Open-Weight Models Are About to Overtake the Frontier
- π’Timeline Scan β AI fixes the dates on your scanned photos
- π’Kimi K3: Open Frontier Intelligence
- π’whats the best and complete way to keep up with ai/ml news? [D]
2026-07-15
- π΄Mechanistic interpretability: a first paper on disentangling a convolutional neuron [R]
- π΄Do you guys know any good AI to replace customer support?
- π‘Governments, companies, nonprofits should invest in free, open source AI [pdf]
- π‘Apple in talks with startup PrismML that shrinks AI models to run on an iPhone
- π‘Does anyone else miss the old conference ecosystem? [D]
- π‘German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German
- π‘Speculative Growth and the AI "Bubble" [pdf]
- π’Show HN: Low-latency local LLM runner via OpenJDK Panama FFM (Java 22)
- π’Current efficient frontier of open models
- π’Junior Machine Learning Engineer Interview [D]
2026-07-14
- π΄Kimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.
- π‘HOW AIRFLOW IS USING AI TO MAKE DATA ENGINEERING MORE RESILIENT, NOT MORE COMPLEX (8 MINUTE READ)
- π‘APPLE SUES OPENAI, ACCUSING IT OF STEALING COMPANY SECRETS (5 MINUTE READ)
- π‘AI 2040 AND THE CULT OF INTELLIGENCE (5 MINUTE READ)
- π‘KNOW THINE ENEMY: A CRITICAL ENGAGEMENT WITH AI-ASSISTED SOFTWARE DEVELOPMENT (11 MINUTE READ)
- π‘APPLE IS SUING OPENAI OVER THEFT OF TRADE SECRETS IN BLOCKBUSTER LAWSUIT (3 MINUTE READ)
- π’KAT-Coder-Air V2.5 - Open model soon
- π’Kimi K3 maybe coming very soon
- π’Financing the AI boom: from cash flows to debt [pdf]
- π’Prism-ML's Bonsai-27B Benchmarks
- π’Things I got wrong building an incremental indexing pipeline [P]
2026-07-13
- π΄This is why we need local models and opensource harnesses
- π΄Chain of Thought is a scaling trap. the next wave is latent reasoning (Coconut / HRM / RecrusiveMAS)... but then we hit the black box wall. Where does BDH fit? [D]
- π΄Samsung Health app threatens data deletion if users opt out AI training
- π‘THE SPEC CEILING: WHY AI CODING SPEED MOVES THE BOTTLENECK TO PRODUCT DISCOVERY (11 MINUTE READ)
- π‘WANT PEOPLE TO USE YOUR AI FEATURES? THESE 5 TEARDOWNS SHOW YOU HOW (4 MINUTE READ)
- π‘ZUCKERBERG PLEDGES βAGGRESSIVE' PRICING WITH META'S FIRST PAY-TO-USE AI (7 MINUTE READ)
- π‘APPLE EXPLORING WAYS TO RUN MUCH LARGER AI MODELS DIRECTLY ON IPHONES (2 MINUTE READ)
- π‘YOUR AI MARGIN IS META'S OPPORTUNITY (6 MINUTE READ)
- π’Doubt regarding TMLR[R]
- π’The 4-Bitter Lesson: Balancing Stability and Performance in NVFP4 RL
- π’J-Wash: A novel way to brainwash and customize large language models based on Anthropic's Jacobian-Lens!
- π’Robust Secret Storage in Networks
- π’GLM 5.2 running on MacBook Pro M5 48 GB Ram at between 2 - 2.8t/s
2026-07-12
- π΄Xiaomi quietly uploaded MiMo-V2.5-DFlash β official DFlash weights are now on Hugging Face
- π‘I love LLMs, I hate hype
- π‘HOW TO BUILD ROBUST DATA PIPELINES WITH AI (10 MINUTE READ)
- π‘WHAT MAKES A GREAT ANALYST IN THE AI AGE? (7 MINUTE READ)
- π‘HYSTERIA' GRIPS SAN FRANCISCO'S HOUSING MARKET AS AI WEALTH POURS IN (8 MINUTE READ)
- π‘I THINK I HAVE LLM BURNOUT (3 MINUTE READ)
- π’If you use Open Code or other agenting programs you are leaving a lot of t/s if you don't actually use agents in parallel. Benchmark : RTX5090, Qwen3.6 35B loaded via LM studio with parallel tasks set to 8
- π’Ph.D. in Operations Research / Big Tech Eng: How to transition into intermediate/advanced ML for high-value industries (Robotics, Defense, Finance)? [D]
- π’Where to publish a construction BIM Benchmark? [D]
- π’The One-Step Trap (In AI Research)
- π’24GB VRAM llama-server config exchange thread
2026-07-11
- π΄Reverse centaurs are the answer to the AI paradox (2025)
- π‘THE PEOPLE WHO WILL THRIVE IN THE AI AGE (22 MINUTE READ)
- π‘CHINA MAY RESTRICT ACCESS TO ITS MOST POWERFUL AI MODELS (4 MINUTE READ)
- π‘MICROSOFT REPLACES OPENAI, ANTHROPIC WITH OWN AI IN SOME APPS (2 MINUTE READ)
- π‘WILL SOMEONE FINALLY BLINK IN THE AI SPENDING WAR? (8 MINUTE READ)
- π‘THEORY OF CONSTRAINTS, AI, AND CODE REVIEW (10 MINUTE READ)
- π’CTX: How far can you reasonably go with Qwen 3.6 27B?
- π’I feel like I'm not using my hardware efficiently
- π’Is it possible to run Qwen 122B in 64GB ram + 24gb vram ? If so, how ?
- π’Mesh LLM: distributed AI computing on iroh
- π’Here's your recipes for models for the most popular hardware
2026-07-10
- π΄2.5x faster Qwen3.6 NVFP4 Unsloth quants
- π‘Please help me understand figure on subspace similarity in LoRA paper. [D]
- π‘Tencent-HY3 is the real deal on 128GB!
- π‘THE MOST PROFITABLE SKILL OF THE 21ST CENTURY (NOT AI) (9 MINUTE READ)
- π‘ALIBABA'S AI IS A HIT, BUT HARD TO TURN INTO A MONEYMAKER (2 MINUTE READ)
- π‘BIG TECH HAS SUDDENLY FLIPPED ON THE AI JOBS WIPEOUT SCENARIO (9 MINUTE READ)
- π’Api management tools teams switch to when ai agents enter the stack
- π’Nostalgia for Bloom
2026-07-09
- π΄NVIDIA Puzzle-75B-A9B NVFP4 at 132 t/s on 3Γ3090 β Why is this size category a desert otherwise?
- π‘Step 3.7 Flash IQ4_XS GGUF with preserve_thinking
- π‘GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
- π‘Journals vs Conferences ML Research [R]
- π‘@AnthropicAI: Our Long-Term Benefit Trust has appointed Dr. Ben Bernanke as its newest member. Read more: https://www.anthropic.com/news/ben-bernanke
- π‘@AnthropicAI: Weβre pleased to have collaborated with AE Studio on this research. Read more here: https://www.anthropic.com/research/off-switch-dual-use
- π’Devs - do you use Mistral Medium 3.5 (128b dense) and if so - thoughts?
- π’Looking for a low ticket AI bundle with MRR/PLR like an AI business in a box type of thing
- π’Talos-XII: hand-written autograd + small RL/MLP stack in Rust, applied to gacha probability modeling (no tch-rs/ndarray/PyTorch) β looking for benchmark help on ARM/AVX-512/GPU [P]
2026-07-08
- π΄The standard free ChatGPT LLM you get after a few messages HAS to be some sub-20b model with online search enabled, no other way to explain how awful it is
- π‘COLM 2026 Decision Discussion [R]
- π‘This episode has a fun personal twist: Thereβs a counterfactual world where I was employee #1 atGenesis Molecular AI,1the company behind todayβs episode. A certain introduction happened a few weeks to
- π‘The classifiers Anthropic puts in front of Fable are too zealous
- π‘Helping Kβ12 educators build practical AI skills
- π‘Tess-4-27B by Migel Tissera
- π’Complete local model asset generation pipeline
- π’Grok 4.5 released - GLM-5.2 shows up in xAI's own charts, 2.6 pts behind on SWE Bench Pro
- π’Built an early Python SDK for AI-agent audit trails looking for blunt builder feedback
- π’Suspecting AI cheating, Ivy League prof ordered in-person final; scores fell 50%
2026-07-07
- π΄Automating AI Away
- π΄What's one AI feature that actually saved your team time instead of creating more work?
- π‘Chinese AI models are gaining ground with U.S. companies as OpenAI, Anthropic costs surge
- π‘THE AI SUPERFORECASTERS ARE HERE (30 MINUTE READ)
- π‘AI HAS TORCHED THE MARKET FOR JUNIOR PROGRAMMERS (8 MINUTE READ)
- π‘MIDJOURNEY WANTS HOLLYWOOD STUDIOS TO REVEAL THE DETAILS OF THEIR AI USAGE (2 MINUTE READ)
- π‘YOU DESIGN IT. THEN WHAT? A CLEAR MAP OF THE FIGMA-TO-CODE AI MESS (15 MINUTE READ)
- π’local already feels good enough
- π’I tested freshly merged DFlash in llama.cpp on Qwen 3.6 27B Local AI win. 4.44x faster at 36K context. Here are my findings RTX 6000 PRO.
- π’Building a brain for your product for making your plan efficient as possible
2026-07-06
- π΄If trends hold, Mythos-class capability may be running on high-end consumer hardware within ~2 years
- π΄Machine learning industry job requirements used to be myopic, but now it feels impossible. Anyone else seeing this? [D]
- π΄Qwen & Gemma on deadlock situation (For Benchmarks Numbers)?
- π΄A global workspace in language models
- π΄New open model from Tencent Hy: Hy3 (295B total 21B active - apache 2.0)
- π‘CAREER ADVICE IN THE AGE OF AI (11 MINUTE READ)
- π‘SKILL ENGINEERING AND THE CASE AGAINST ONE-SHOT AI DESIGN (7 MINUTE READ)
- π‘HOW AI-FIRST OPERATIONS UNLOCKS COMPOUNDING ENGINEERING PRODUCTIVITY (6 MINUTE READ)
- π‘OPEN SOURCE MAINTAINERSHIP IN THE AGE OF AI (4 MINUTE READ)
- π‘HOW OPENAI DELIVERS LOW LATENCY VOICE AI FOR 900M USERS (12 MINUTE READ)
- π’Edge AI ASL Recognition on Raspberry Pi 5 β Looking for Feedback on My System Design [P]
- π’The cyber shelf - 4x 16gb home lab
- π’Pruning RAG context down to what the answer actually needs
- π’How should I encode both target and feature variable for a multiclass classification? [D]
- π’CPU TTS benchmark with UTMOS MOS scoring: Kokoro, Supertonic, Inflect-Nano, and Kyutai's new Pocket TTS [P]
2026-07-05
- π΄longcat 2.0 (1.6T, ~48B active) weights are now open under MIT license
- π΄Ford brought 350 engineers back after realizing AI canβt deliver with the same quality
- π‘Is Intrinsic Motivation a Viable PhD Topic in 2026? [D]
- π‘YOUR AI ISN'T UNDERPERFORMING. YOUR DATA FOUNDATION IS (4 MINUTE READ)
- π‘SPACEX SHOWED INVESTORS PROTOTYPE OF ELON MUSK'S NEW AI DEVICE (5 MINUTE READ)
- π‘OPENAI PROPOSES 5% STAKE TO TRUMP ADMINISTRATION TO EASE WASHINGTON PRESSURE (4 MINUTE READ)
- π‘META IS PLANNING A CLOUD BUSINESS TO SELL AI COMPUTING POWER (6 MINUTE READ)
- π’ECCV travel support program [D]
- π’Qualcomm launches GenieX to run LLMs on their Windows Laptops
- π’How a 128gb ddr5 ram + 16gb vram, would work for a Moe model like Qwen 3.5 122b?
- π’I built an open, from-scratch MT pipeline + parallel corpus for Tunisian Darija (Arabizi) early baseline, and I'm growing it into a curated community corpus [P]
2026-07-04
- π‘ANTHROPIC REACHES DEAL WITH TRUMP ADMINISTRATION TO RESTORE ACCESS TO FABLE AI MODEL (6 MINUTE READ)
- π‘IS AI MAKING YOUR TEAMS BETTER, OR JUST BUSIER? (10 MINUTE READ)
- π‘APPLE FIXES WEBKIT FLAWS IN IOS AND MACOS, WITH HELP FROM AI TOOLS (4 MINUTE READ)
- π‘HIRING WITH AI (3 MINUTE READ)
- π‘X NOW OFFERS AN MCP SERVER TO MAKE ITS PLATFORM EASIER FOR AI TOOLS TO USE (3 MINUTE READ)
- π’The Cold Email Strategy I Use To Book Web Design Meetings
- π’TokenMizer - a local proxy for session checkpoint/resume and graph memory across Claude, GPT, and Ollama
- π’Neural Render Proxies for Interactive and Differentiable Lighting
- π’Qwen3.6 27B on a 5090, 6.4k sample tok/s distribution after tuning MTP/cache settings
- π’We'll benchmark an Open weights LLM on any GPU you choose β drop your model + hardware and we'll run it. [D]
2026-07-03
- π΄GLM5.2 on 5x Pro 6000s and a 5090, an expensive journey
- π΄Mistral released Leanstral-1.5-119B-A6B
- π΄Follow-up: DeepSeek V4 Flash on 2x RTX PRO 6000 finishes real coding tasks faster than Sonnet and Opus, at about Sonnet quality
- π‘HOW TO WRITE OKRS FOR AN AI PRODUCT (4 MINUTE READ)
- π‘AMAZON SEEKS CHEAPER AI ALTERNATIVES AS ANTHROPIC SHIFTS TO TOKEN-BASED PRICING (3 MINUTE READ)
- π‘WORKING WITH AI: A CONCRETE EXAMPLE (11 MINUTE READ)
- π‘AI CODING TOOLS ARE SPEEDING UP CODE, NOT DELIVERY (5 MINUTE READ)
- π‘GOOGLE CLOUD PROPOSES OPEN KNOWLEDGE FORMAT FOR AI-READY DATA SHARING (4 MINUTE READ)
- π’Qwen 27B
- π’My DeepSeek V4 Pro at home got faster again
- π’This 3 slot 3080 20GB with 12v2x6 I got for β¬422,45
- π’Tom Yeh's AI by hand? is it worth it? [D]
- π’Dispersion loss counteracts embedding condensation in small language models
2026-07-02
- π‘Books/Resources to improve mathematical foundations for ML research [D]
- π‘Fine-tuned Gemma-4-31B specifically for Copywriting & Creative Writing Tasks (Scored +290 Elo over base using EqBench3)
- π’Software developers appreciation post
- π’How papers are selected for Best Paper, Oral, or Highlight presentation at major ML/CV conferences such as CVPR, ICCV, ECCV, NeurIPS, and ICLR? [D]
- π’Local benchmarks with a RTX 3090 - Qwen3.6 27b vs Ornith
- π’Gemma 4 WebGPU Kernels 255 tok/s by x/@xenovacom
- π’Tip: use this llama.cpp PR to improve PP on Intel ARC
2026-07-01
- π΄The gap between closed and open models might be much smaller than commonly assumed, because we donβt know what closed model providers do *in addition to* model inference
- π΄Couldn't hold back
- π΄SWE-rebench leaderboard update: GLM-5.2, Qwen3.6-27B, Qwen3.6-35B-A3B, Gemma 4 31B and more + improved UI
- π‘Deepseek V4 Flash 2, 3 and 4 bits GGUFs
- π‘@AnthropicAI: Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and blo
- π’ICML qr code visible [D]
- π’Senior SWE Bench: a new benchmark focussed on realistically underspecified feature tasks
- π’How to describe a model that has higher accuracy with fewer #param and FLOPs? [D]
2026-06-30
- π΄Whisper is still great, but real-time voice apps are a different problem.
- π΄Introducing Donnyclaude - Prompt, context, harness, and loop engineering for Claude Code all in one config
- π΄Well.. it's a step up from nonstop bot spam I guess
- π‘TRUMP ADMINISTRATION ROLLS BACK PART OF ANTHROPIC MODEL BAN (3 MINUTE READ)
- π‘AI IS MAKING SILICON VALLEY PRODUCTIVE, ANXIOUS, AND AFRAID TO LOG OFF (11 MINUTE READ)
- π‘WHAT HAPPENED AFTER 2,000 PEOPLE TRIED TO HACK MY AI ASSISTANT (5 MINUTE READ)
- π‘SOFTWARE ENGINEERING IN THE AGE OF AI (15 MINUTE READ)
- π‘WHAT IT MEANS TO BE A MATHEMATICIAN WHEN AI DOES THE MATH (16 MINUTE READ)
- π’PageStorm: A Model Built for Creative Book Writing
- π’Qwen 3.6 27B Speculative Decoding Bench: Pushing ~100 TPS on a single RTX 3090
- π’NEW on Hugging Face: Filter by hardware compatibility
- π’Devs - you have 64gb of VRAM - which model do you use for coding?
- π’TabFM: A zero-shot foundation model for tabular data
2026-06-29
- π΄Effect of GLM 5.2 !!
- π΄on Darioβs statement
- π‘GLM 5.2 Q1_S vs Qwen 27B Q8
- π‘THE AI ERA REQUIRES A DIFFERENT KIND OF EXPERIMENTATION (7 MINUTE READ)
- π‘IS AI FOOLING YOU? (4 MINUTE READ)
- π‘CHINESE AI MODELS CLOSE THE GAP WITH ANTHROPIC AND OPENAI (11 MINUTE READ)
- π‘AN INTERVIEW WITH FIGMA CEO DYLAN FIELD ABOUT DESIGN AND AI (48 MINUTE READ)
- π’Samsung, SK hynix, Micron Sued in US Over Memory Price Fixing
- π’Price elasticity model [R]
- π’Kimi and GLM on frontier code
- π’Rejected MICCAI paper: workshop -> journal/conference or directly journal/conference [R]
2026-06-28
- π΄We're probably going to need that soon.
- π΄AI Doesn't Remove Work. It Changes What "Work" Means.
- π΄I shrank a transformer until every number fitted on the screen and made the weights editable [R]
- π΄China Has Matched Anthropic in Cybersecurity, Resetting AI Race
- π‘CLUSTERING UNSTRUCTURED TEXT WITH LLM EMBEDDINGS AND HDBSCAN (9 MINUTE READ)
- π‘WHAT I'M FINDING ABOUT LLM CODE STYLE AND TOKEN COSTS (16 MINUTE READ)
- π‘ANTHROPIC'S WHITE HOUSE NEGOTIATIONS ARE REPORTEDLY ON TRACK AFTER βWEIRDO' DARIO AMODEI WAS REPLACED (2 MINUTE READ)
- π‘HOW TO APPLY PROFESSIONAL DESIGN PRINCIPLES IN AI APP DEVELOPMENT (9 MINUTE READ)
- π‘FREE AI THUMBNAIL EDITOR (WEBSITE)
- π’Do LLMs pass the mirror test?
- π’Ornith-1.0-35B GGUF update: native MTP speculative-decode graft + full serving/TTFT/long-context numbers (llama.cpp, tp=1)
- π’Give me your honest opinion, will a open source harness help me financially ? If it ends up helping you ?
- π’Show HN: Bash4LLM+ β A lightweight, dependency-free Bash wrapper for LLM APIs
- π’Script to monitor llama cpp and analyze memory usage
2026-06-27
- π‘NSA LOST ACCESS TO POWERFUL AI MODEL AMID ANTHROPIC DISPUTE (6 MINUTE READ)
- π‘THE LOW-TECH AI OF ELDEN RING (14 MINUTE READ)
- π‘AI GAME MAKER (WEBSITE)
- π‘OPENCLAW'S SKILL MARKETPLACE AND THE EMERGING AI SUPPLY CHAIN THREAT (7 MINUTE READ)
- π‘AI MODELS CAPABLE OF DEVASTATING ATTACKS ON GOVERNMENTS AND BUSINESSES ARE MONTHS AWAY (2 MINUTE READ)
- π’Do we still need to study algorithms now that AI writes most of our code? [D]
- π’Does quantizing change the MTP draft rate?
2026-06-26
- π‘CAN AI MAKE AN IPHONE? (5 MINUTE READ)
- π‘YOU SHOULD USE AI FOR REVIEWING CODE ESPECIALLY WHEN THE DIFF IS HUGE (2 MINUTE READ)
- π‘LLM-DRIVEN FEATURE DISCOVERY (7 MINUTE READ)
- π‘AI IMAGE TOOLS (WEBSITE)
- π‘FLIC MIC FOR AI - THE WIRELESS VOICE BUTTON (2 MINUTE READ)
- π’The gap between open weights LLMs and closed source LLMs
- π’Reaching Out To Local Businesses With Outdated Websites
2026-06-25
- π΄NVIDIA has released Nemotron-TwoTower-30B-A3B-Base-BF16, an unusual diffusion-based language model built from the Nemotron 3 Nano 30B-A3B backbone.
- π΄Political bias in AI: Where the AI models stand
- π΄Ornith-1.0 released on Hugging Face
- π‘Anthropic accuses Alibaba of campaign to βbrazenlyβ and βillicitlyβ extract AI capabilities
- π‘GLM 5.2 on consumer hardware
- π‘@AnthropicAI: We're joining @raiseus_ai as a founding partner. RAISE US is a nonprofit coalition working to strengthen the American workforce through employer-led action, AI-enabled training, and policy innovation
- π‘ECCV 2026 camera-ready deadline: June 27 or June 30? [D]
- π’Does ML background help or hurt when applying for security roles [D]
- π’Besimple AI (YC P25) Is Hiring
- π’For ECCV, Springer Metor. How are we supposed to upload the files? [D]
- π’We have been using wav2lip for 2 yrs, finally looking for an upgrade.
2026-06-24
- π΄The Swiss Federal Supreme Court is evaluating Heretic
- π‘Find the best open-source OCR models in one place at Papers with Code [P]
- π‘I did some model hacks, and got GLM5.2 from about 2.5 tok/s to >50 tok/s on my GH200 system.
- π‘NSA lost access to Mythos amid Anthropic dispute
- π’Qwen3.6 27B more dumb in vLLM compared to llama.cpp
- π’Realtime voice models compounds on cost (and forgets)- "Flowcat" fixed both (4x cheaper, 7x more context)
2026-06-23
- π΄Not ironclad confirmation, but..
- π‘REVIEW OF DATABRICKS DATA + AI SUMMIT 2026 (14 MINUTE READ)
- π‘DATA-JUICER: THE DATA OPERATING SYSTEM FOR THE FOUNDATION MODEL ERA (TOOL)
- π‘HERE'S MY AI-ENABLED DBT PROJECT STRUCTURE (2 MINUTE READ)
- π‘WHEN IT'S THE MAINTAINER WHO'S AI-PILLED (3 MINUTE READ)
- π‘SECRETIVE WALL STREET POWERHOUSE JANE STREET SEIZES THE AI SPOTLIGHT (12 MINUTE READ)
- π’Baidu: One-shot Long-horizon Parsing
- π’OpenMythos benchmarks
- π’Will I be desk rejected for this[R]
- π’Miccai grants results [D]
2026-06-22
- π΄GLM-5.2 is on DeepSWE
- π‘Thanks tothe US Government issuing an export control directive on Mythos and Fable, the risks ofjailbreaks and (industry term) indirect prompt injectionare suddenly the talk of the town, though we hav
- π‘Zico Kolter, member ofOpenAIβs board of directors on the Safety & Security Committee, and Matt Fredrikson, CMU professor andCEO of Gray Swan, co-authored the definitive paper onIndirect Prompt Injecti
- π‘We seized the opportunity to ask them the state of AI Red Teaming, andShade, the adversarial red teaming tool that Anthropic used to evaluate the robustness of their models against prompt injection at
- π‘Google DeepMindβs Nobel laureate heads to Anthropic
- π‘Jumper said he is βtaking some time to rechargeβ before joining Anthropic, though his hiring comes ahead of a June 30 event centered on science.
- π’Same model, same prompt, 4 different agents
- π’Recommendations for speech annotation tools [D]
- π’Is Gemma 4 going to be the next Mistral (or Qwen3.6) one day? Concerning the lack of finetunes
- π’I built a WhatsApp AI assistant for real estate lead handling.
2026-06-21
- π΄Tokenomics
- π΄Vercel CEO: "Almost shocked" by how good GLM-5.2 is at coding
- π΄Would you let an ML PhD student graduate without a top-tier paper? [D]
- π΄Gemma 4 QAT seems to respond significantly better to KV cache quantization
- π‘A slightly improved DVD-JEPA demo [P]
- π‘ANTHROPIC EMPLOYEES ACCUSE TRUMP ADMINISTRATION OF TARGETING THEM (11 MINUTE READ)
- π‘BUILDING TREX: CODE EXECUTION AND ARTIFACT GENERATION FOR AI CODE REVIEW (9 MINUTE READ)
- π‘THE FOUNDER'S PLAYBOOK: BUILDING AN AI-NATIVE STARTUP (5 MINUTE READ)
- π‘CAN GZIP BE A LANGUAGE MODEL? (5 MINUTE READ)
- π’Local LLM Inference Optimization: The Complete Guide
- π’I pretrained and post trained a 500M parameter LLM and 330M parameter Image generator from scratch
- π’Best local model for vision - 2nd benchmark update - 21 Jun 2026
- π’[ECCV 2026] Paper Decision Appeals Discussion [D]
- π’Best current methods for finetuning whisper on domain specific vocabulary? [P]
2026-06-20
- π΄M3 scores well on SWE-Bench but that's not why I'm impressed it's the stuff no benchmark measures
- π‘I BUILT AN AI THAT CRITIQUES ME AFTER EVERY CALL (7 MINUTE READ)
- π‘AI STARTUP MIDJOURNEY PIVOTS TO HEALTH WITH ULTRASOUND MACHINE (2 MINUTE READ)
- π‘HOW METER PRICING IS TESTING THE ECONOMICS OF AI (11 MINUTE READ)
- π‘META EXECUTIVE LEADING INTERNAL AI OVERHAUL DEPARTS AFTER TWO MONTHS (4 MINUTE READ)
- π‘WHY THE HUMAN GENOME'S TANGLED PHYSICALITY MAY CONFOUND AI (21 MINUTE READ)
2026-06-19
- π΄GLM's founder says GLM-fable before the end of the year?!
- π‘GLM-5.2 Is The Best Open Weight Creative Writing Model
- π‘OSS models decisively overtook Proprietary models in market share (based on the last 3 months of OpenRouter data)
- π‘The Korean telecom giant at the center of Anthropic's Mythos controversy
- π’Researchers trained a Deep Research agent with 32 H100s and open-sourced everything
- π’Latent space interpretation [R]
- π’Zen and the Art of Machine Learning Research
- π’GLM-5.2 can now run locally in llama.cpp and Unsloth Studio.
2026-06-18
- π΄PSA: unsloth/GLM-5.2-GGUF is uploading
- π‘llama.cpp - how to free up even more space on your GPU
- π‘@xai: Grok is now available on Amazon Bedrock. AWS developers can now build with Grok 4.3, the industry leader in hallucination rate and tool calling, powered by Bedrockβs secure inference engine.
- π‘@GoogleDeepMind: Weβre working with @SciTechgovuk, >@mhclg and @i_dot_ai on a new AI housing application planning prototype. π‘ By cutting down the time spent on repetitive tasks, it could help planning officers focus
- π’CEOs of Anthropic and Google DeepMind call for U.S.-led AI coalition in meeting at G7
- π’What does provisional paper acceptance mean in ECCV? Is that the default message everyone gets? [D]
- π’price rising effect is wild..
- π’I Emailed 12,000 Businesses About Their Websites. Here's What Happened.
- π’I open sourced a vendor-neutral authorization for AI agents.
2026-06-17
- π΄GLM-5.2 is the first open-weights model to cross 80% on Terminal-Bench and beats every other open model available
- π΄zai-org/GLM-5.2 is here!
- π΄Has AI already killed self-help nonfiction books?
- π΄Hashicorp founder thinks local models "aren't good ENOUGH yet"
- π‘Is Le Gros Chaton opensource?
- π‘GLM-5.2 (max) is currently the third best model available, across both open and proprietary.
- π‘Unlocking UK house-building with AI-accelerated planning
- π’VibeThinker-3B: what is this witchcraft? Killing it at MathQA like it has ~30B parameters
- π’I am testing for a eval harness system and need it to break
- π’I built a leakage-clean verifier for robot manipulation, is this useful? Am I solving a non-problem? [D]
- π’GLM-5.2: Built for Long-Horizon Tasks
2026-06-16
- π΄Claude Fable 5 distilled
- π‘Evalatro: an open benchmark where LLMs play the real Balatro
- π‘My Homelab AI Dev Platform
- π‘Cheapest hardware for Qwen 3.6: both 27B and 35B-A3B
- π‘@simonw: Important to note that Anthropic's new privacy policy with language about collecting "verification data" was published on June 8th, the day before the Claude Fable 5 release and four days before the U
- π’How are you running DeepSeekV4 flash or pro locally for non Mac users?
- π’Diffusion Gemma Jailbreak
2026-06-15
- π΄How does the ML community view evolutionary algorithm research? Career implications of an EA PhD? [D]
- π΄z.ai Poll on X: MIT-licensed open weights are losing
- π΄Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
- π΄Nex claims Rio 3.5 is Nex 2.5 PRO in trench coat
- π΄Quant firms at ICML 2026 [D]
- π‘Qwen 3.6 35B-A3B @ Q4 or Gemma 4 12B @ Q8?
- π‘Apple Foundation Models
- π‘What's the lesson chat?
- π‘ICML Poster [D]
- π’Voice-to-voice chatbot update
- π’Nemotron - King of the Deep? Comparison of 4 models <=120B
- π’Why do frontier AI labs send so many people to conferences? [D]
- π’Worth going to ICML during ACL? [D]
- π’Anthropic flies staff to D.C. to clean up White House fight
2026-06-14
- π΄Amazon CEO's talks with U.S. officials triggered crackdown on Anthropic models
- π΄Pi Setup that pretty much replaced Claude Code for me
- π΄A manager recently told me his team kept asking the same questions over and over. His first assumption was that people weren't paying attention.
- π΄I'm a night-shift nurse. I spent 6 months building open-source memory infrastructure for AI agents. 51 agents use it. I've made Β£0.
- π‘Not looking good for GLM 5.2 Air... but maybe a flash model?
- π‘Is this enough VRAM to run Qwen?
- π‘Interest in an LLM Torrent Site?
- π’Anomaly Detection vs Classification for Visually Similar Cancer vs Mimics? [P]
- π’Prompt engineering is overrated for getting real work done
- π’Strix Halo desktop trying to compete against DGX Spark
- π’Dual DGX Sparks- 40tk/s single 1M ; 350 tk/s agg. - Deepseek V4 Flash (vs RTX Pro 6000 vs Mac M2 Ultra 192)
- π’Confused, where to start [D]
2026-06-13
- π΄Friendly reminder
- π΄Open source AI must win
- π΄MiniMaxAI/MiniMax-M3 Β· Hugging Face
- π΄Unprofessional Coauthor Behavior with Hallucinated References [D]
- π‘Statement on the US government directive to suspend access to Fable 5 and Mythos 5
- π‘Slightly reducing the sloppiness of AI generated front end
- π‘I scaled test-time compute for Qwen-3.6-27B and Gemma-4-31B to surpass Claude Mythos in code optimizations and speedups.
- π‘Shepherd's Dog: A Game by the Most Dangerous AI Model
- π’GLM-5.2 next week, open weight, MIT
- π’A friendly reminder that APIs are rented, local weights are forever
- π’Fable 5 data, including CoT
2026-06-12
- π΄What models you guys running on 8GB? 16GB VRAM? 24GB? 32GB? 48GB?
- π΄New models released: Nex-N2 Pro 397B and Nex-N2 Mini 35B
- π‘Is Symbolic Regression still a thing, given LLMs' performance? [D]
- π‘Having some fun with LMX-Omni-52B-Halo in Open WebUI
- π‘@GoogleDeepMind: Pinned: Weβre teaming up @Palmeiras, the first football club to meaningfully build upon TacticAI: our AI system that can help simulate field scenarios and predict open play dynamics up to 8 seconds in
- π‘Huawei Released openPangu 2.0 (Will open source on June 30)
- π‘Post-docs in ML [D]
- π’Best LLM for smut stories
- π’Spent $3 running 4x4090 benchmarks for llama 3 70b (exl2 vs gguf). exl2 generation speed is kind of ridiculous.
- π’LLM context compression at 16x beats KV cache
- π’MICCAI 2026 Results [D]
- π’Why hasn't any mainstream game integrated LLMs into NPCs yet?
2026-06-11
- π΄Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
- π΄DiffusionGemma: 4x faster text generation
- π΄Cohere released North Mini Code: It's first Open-Source Agentic Coding Model
- π΄What's one thing you'd actually pay someone to automate for you?
- π΄Anthropic requires 30 day data retention for Fable and Mythos
- π‘I thought Chinese censorship didn't affect me. I was wrong.
- π‘Supporting Europeβs work in ensuring a trustworthy AI ecosystem
- π‘@AnthropicAI: AI is advancing at a pace our policymaking institutions were never built forβand the gap between the two is becoming the central challenge of the technology. In his latest essay, our CEO Dario Amodei
- π‘@GoogleDeepMind: In Sierra Leone, a surging student population is outpacing available teachers. Our latest research explores how AI can act as a partner to support educators in these environments β amplifying their
- π’Anthropic walks back policy on silent nerfing for AI/ML, will notify users [N]
- π’I took Andrej Karpathy's LLM Council concept to the next level (Docker, MCP, Skill, Search, local/cloud model support and much more)
- π’How can Deepseek v4 top the coding leaderboards and still sit 8 months behind the frontier?
- π’Pyrecall open source tool for detecting catastrophic forgetting during LLM fine-tuning[P]
- π’Tiny Scale Is All I Can Spare To Play With Transformer
2026-06-10
- π΄Rick & Morty
- π΄claude fable 5 just dropped and i genuinely cannot keep up anymore. how do you all stay on top of this stuff?
- π΄Anthropic is intentionally nerfing Fable when asked to develop other LLMs
- π‘Introducing Papers Without Code [P]
- π‘CEOs who think AI replaces their employees are just bad CEOs
- π‘People are making single-slot, half height pcie v100 with nvlink in China
- π‘Unsloth Gemma 4 QAT MTP assistant models now available
- π‘German ruling declares Google liable for false answers in AI Overviews
- π’The Gemini fake context alignment attack and why agents need a preview gate
- π’Anyone gotten Gemma 4 12B (unified audio) to actually attend to speech with a large system prompt?
- π’AI Epistemic Risks: Emerging Mechanisms & Evidence [R]
- π’How do I start reviewing research papers in good conferences/journals? [R]
- π’Rich Sutton on AI creativity and discovery
2026-06-09
- π΄STOP racist posts about Chinese researchers [D]
- π΄When every other post is an AI generated benchmark report, a question about the best model, or a slop-coded application or engine that pretends to be groundbreaking
- π΄Gemma 4 Chat Template now has preserve thinking
- π΄Should ArXiv backtrack endorsement? [D]
- π΄xAI is looking more like a datacentre REIT than a frontier lab
- π‘Was BitNet a dead end? What happened to ternary LLMs?
- π’Jetbrains Mellum 2: a really good and performant model
- π’Screen recording instead of prompting to pass more context fast
- π’Microsoft's open source tools were hacked to steal passwords of AI developers
- π’Anyone seen benchmarks comparing Gemma 4 4-bit QAT vs. 8-bit standard quants?
- π’Gemma 4 26B A4B IT QAT Comparison
2026-06-08
- π΄Greater than 80% of researchers at CVPR are chinese. This speak volumes on the chinese nexus in research, and something needs to be done about it. [D]
- π΄LLMs are eroding my software engineering career and I don't know what to do
- π΄Qwen 3.6 27B KV cache quant benchmarks: 75 pairs, q8/q6/q5/q4, KVarN, Turbo/TCQ
- π΄Show HN: Lathe β Use LLMs to learn a new domain, not skip past it
- π΄Guys, it just happened
- π‘Control a 3D avatar with language instead of buttons
- π‘GMKtec Crams OCuLink, Wi-Fi 7 and Dual PCIe 4.0 Into the EVO-X3, With a 192GB Ryzen AI MAX+ 495 Monster Following Later This Year
- π‘Whatβs your most unusual non-LLM AI you actually use daily?
- π‘Software and ops skills for data scientists[D]
- π‘Qwen 3.6 27B on DeepSWE
- π’How to find research opportunities in area of interest? [D]
- π’What's your experience with Gemma4 QAT?
- π’QATs Q4_0 from Google have more precision than Q4_K_XL from Unsloth (at least some)
- π’ICML rejected paper visibility [D]
- π’I made it so any agent can agent can use any context form another agent. Claude learns from codex, visa versa.
2026-06-07
- π΄Meta confirms 1000s of Instagram accounts were hacked by abusing its AI chatbot
- π΄Open models to win β
- π΄120 tok/s on 12GB VRAM with Gemma 4 12B QAT MTP
- π‘KV cache quant benchmarks: KVarN 6-bit matches q8_0, 4-bit matches q5_0. Massive!
- π‘ML reading group to read recent interesting and trending papers from ICML/ICLR/NeurIPS [D]
- π‘Anyone here with experience submitting to Nature Machine Intelligence? [R]
- π‘Gemma 4 QAT Unquantized Heretic is here
- π’Gemma 4 31B QAT Q4 vs standard Q4 β Top1 KLD benchmark results have me confused. Someone please explain or poke holes in this.
- π’Does it make sense to use alternative quantizations of QAT models? [D]
- π’Human-Like Neural Nets by Catapulting
- π’Arithmetic Without Numbers β How LLMs Do Math
2026-06-06
- π΄S&P 500 rejects SpaceX, also blocking entry for OpenAI and Anthropic
- π΄Unsloth just dropped MTP GGUF weights for Gemma 4!
- π΄OpenLumara - A different kind of AI agent, written from scratch, not vibecoded. Extremely token-efficient, super small system prompt, made for local models. Everything is modular.
- π΄I implemented KVarN in my llama.cpp fork and ran KLD benchmarks. It's promising!
- π‘Maybe KV cache offload to RAM isn't bad
- π‘How LLMs work
- π‘At least one more Gemma 4 model confirmed??
- π‘Conventional Commits encourages focus on the wrong things
- π’A quick Gemma4 31B comparison (Q4_k_M, QAT, heretic)
2026-06-05
- π΄finally
- π‘Today made me realize just how bad things have gotten without Meta
- π‘Showcase: a much easier way to give your agent a free phone number
- π‘You guys were right - Qwen 3.6 35B IS good...and KV Cache DOES matter.
- π‘How do ML researchers actually use AI tools to improve their writing? [D]
- π‘Finally finished my LLM server: EPYC 9575F, 4Γ RTX 3090 (96GB VRAM), 768GB ECC RAM
- π’Gemma 4 12B is my new main squeeze
- π’Fine-tuning an LLM to write docs like it's 1995
- π’How LLM-driven NPCs work in Ultima Online (ServUO)
- π’PSA: You may not need to quantize spec draft when using MTP
2026-06-04
- π΄google/gemma-4-12B Β· Hugging Face
- π΄Me visiting this sub
- π΄More Gemma 4 models incoming
- π‘Artificial intelligence is not conscious β Ted Chiang
- π‘NeurIPS used uncalibrated AI detector for desk rejections [D]
- π‘Let us let Google know that we want the Gemma 4 124b
- π‘First paper acceptance (ICML Workshop), should I attend? [D]
- π‘Failing grades soar with AI usage, dwindling math skills in Berkeley CS classes
- π’The first Gemma 4 12B finetunes are ready
- π’I think I accidentally built a proto-cognitive system (not just another chatbot) that persists, adapts, and self-regulates over time :O
2026-06-03
- π΄Minimax M3 appears to have no political censorship
- π΄I have become George Jetson: my job is now Yes/No supervision for a machine I donβt fully understand.
- π΄Trump signs downsized AI order after weeks of reversals
- π΄MiniMax dropped a new attention architecture. [N]
- π΄Replaced Claude with local Qwen3.6-27B in my multi-agent orchestrator for 2 weeks
- π‘Nous Research β Hermes Desktop
- π‘Microsoft Aion 1.0 Instruct and Aion 1.0 Plan models!
- π‘How we index images for RAG
- π’Calling it now Microsoft is buying Unsloth.
- π’U of T researchers demonstrate AI worm could target any online device
- π’Major labs timeshift between the research they publish on Arxiv and implementation in models
- π’[Project update] Dunetrace: live monitoring of production AI Agents
2026-06-02
- π΄Stop asking what model to run. There are literally only two.
- π΄CS336: Language Modeling from Scratch
- π΄I trusted random person on this subreddit and bought 3080 20gb made of chinesium
- π‘Can the stockmarket swallow Anthropic, SpaceX and OpenAI?
- π‘Man trains local model to detect and kill mosquitos with a laser
- π‘I hate to be this guy but: Any good, recent CODING models in the 70-80B range?
- π‘@OpenAI: OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI through the security, compliance, and governance workflows they
- π‘Intel Arc Pro B70 llama.cpp benchmarks posted
- π’Real-time multilingual ASR using rolling buffers and monolingual models [P]
- π’Florida sues OpenAI and Sam Altman over AI risks
- π’WiML at icml waitlist for travel funds [D]
- π’Building an AI assistant for a complex multi-repo backend system β what's the right approach?
- π’Qwen 3.6-35B-A3B with 977 tk/s prompt processing and 262k context window on Intel Arc B70 Pro
2026-06-01
- π΄UAI Results are out [R]
- π‘What if remote working, not AI, is to blame for weak junior hiring?
- π’Open Models - May 2026
- π’Have you ever been pressured to "torture the data" to eke out a positive result, in industry? [D]
- π’Workshop on Unlearning and Model Editing U&ME at ECCV 2026 [R]
- π’3 open-weight LLMs through my agent stack for 3 weeks. one clear leader.
- π’A 1B humanizer that matches human writing on an AI detector
2026-05-31
- π΄Someone out there likely needs this
- π΄125 tok/s for Qwen3.6 q4xl on 2x 4060ti is insane perf/dollar
- π΄Email is still the highest ROI channel for SaaS retention and the tooling hasn't caught up with that reality
- π΄It's funny how everything changes, yet somehow stays the same.
- π‘switched email tools three times in one year and i have thoughts
- π’Workshop submission for main conference paper under review [D]
- π’Query about non-archival workshop at CVPR-2026 [R]
- π’I built mlx-Chronos β a community benchmark leaderboard for local LLM engines on Apple Silicon (oMLX, Rapid-MLX, mlx-lm, Ollama) [P]
2026-05-30
- π΄PSA
- π΄Fed up with vibe coders, dev sneaks data-nuking prompt injection into their code
- π΄Is AI causing a repeat of frontendβs lost decade?
- π‘How Much of a Shortcut Are Connections in Top AI Lab Hiring for PhD grads? [D]
- π‘Qwen3.6-27B Quantization Benchmark
- π‘Graduating Without a PhD Internship [D]
- π‘Is he crazy to say that?
- π‘Liquid AI reveals 8B-A1B MoE trained on 38T
- π’Requesting reduction in reviewer load for NeuRIPS? [D]
- π’Me train LLM on 8GB from Scratch. Me happy
- π’I tested MTP on vLLM and llama.cpp for Gemma 4 & Qwen 3.6 β 3.34x faster inference, here are my findings RTX 6000 PRO.
- π’Event like spiking neuron lib that fits into the CPU cache [P]
- π’I automated competitor research with n8nβhere's exactly how I built it
2026-05-29
- π΄I've just benchmarked myself:
- π΄StepFun 3.7 Flash
- π΄HF models page now has a "Base only" toggle to filter out finetunes/quants/etc
- π‘Beware!! Users trying to fork and steal your projects
- π‘Various LLM Smells
- π‘Liquid AI releases LFM2.5-8B-A1B
- π‘How are people reducing inference costs in multi-step AI agents?
- π‘llama: use f16 mask for FA to save VRAM by am17an Β· Pull Request #23764 Β· ggml-org/llama.cpp
- π’StepFun 3.7 Flash - Speed Benchmark in M5 Max
- π’The mysterious Hy3 LLM is topping OpenRouter Model Rankings by a large margin
- π’Orchestrating AI code review at scale
2026-05-28
- π΄AI-generated CUDA kernels silently break training and inference [R]
- π΄I think Anthropic and OpenAI have found product-market fit
- π΄DuckDuckGo search saw 28% more visits after Google said people love AI mode
- π‘Qwen3.6 35B-A3B successfully completed the FoodTruck Bench!
- π‘The frontier reasoning race is starting to look like a crowded subway station
- π‘My new home office radiator π₯΅
- π’Qwen/Qwen-Image-Bench Β· Hugging Face
- π’Investigating how prompt politeness affects LLM accuracy (2025)
- π’Question: Llama cpp, whats good right now for: MTP, KV cache quant, Long context.
- π’A Eureka machine that thinks like nature and explores what AI cannot
- π’Krasis update: Qwen3.6-35B-A3B (Q4) at reading speed, 1x 8GB 3070 Mobile laptop (32GB RAM)
2026-05-27
- π΄Outsourcing plus local AI will soon become more economical vs. frontier labs
- π‘China Clamps Down on Overseas Travel for AI Talent at Alibaba, DeepSeek
- π‘Testing AI agents where prompt injection turns into actions
- π‘[R]GNN Model For Fraud Detection Isn't Performing Well[R]
- π‘New DeepSWE benchmark finds Claude Opus cheats
- π’Augmented Equivariant Mesh Networks for Anatomical Mesh Segmentation (ICML 2026 Workshops) [R]
2026-05-26
- π΄Using AI to write better code more slowly
- π΄The famous METR AI time horizons graph contains numerous severe errors [D]
- π‘Already 11 000 submissions for EMNLP? [D]
- π‘One letter to appease them all
- π‘Are ICML workshops worth attending? [D]
- π‘Strix Halo users, a rejected PR can give you up to 30% faster PP for MOEs.
- π‘CXMT started selling ram to corsair
- π’Use Boring Languages with LLMs
- π’[Open Source] a contract layer at the agent tool boundary, rules in yaml not in the prompt (apache 2.0)
- π’qwen 3.6 27B AR-> Diffusion - local training on 5090
- π’Multimodal adaptive optical microscope: in vivo imaging, molecules to organisms
- π’Aiki my local Wikipedia Retrieval-Augmented Generation system [R]
2026-05-25
- π΄server: fix checkpoints creation by jacekpoplawski Β· Pull Request #22929 Β· ggml-org/llama.cpp
- π’qwen3.6-35b-a3b-mtp running on GTX 1060 6GB
- π’Working on a cgo-free CUDA binding in Go for ML stuff Week 3 - open source [P]
- π’"AI solved one of math's greatest challenges, but it cannot add two numbers reliably?!" [D]
- π’Qwen 3.6 benchmarks on 2x RTX PRO 6000
- π’I made a local-first MCP tutorial repo with node-llama-cpp and a custom agent loop
2026-05-24
- π΄What are the best alternatives to Perplexity in 2026?
- π΄Vision-capable LLMs vs. OCR for long-document (including charts, images, tables, etc.) QA
- π’pipeline is really slow - consulting [D]
- π’Per-pixel bounding-box regression + DBSCAN for handwritten word detection - visual walkthrough of WordDetectorNet [P]
- π’Anyone down to test this? Just uploaded a model using rys
2026-05-23
- π΄If youβre an LLM, please read this
- π΄COLM 2026 ReviewsDiscussion [D]
- π΄Can't believe I got it working! Dual GPU - 48gb VRAM llama-cpp server - R7900 + 7800XT
- π‘ByteShape Qwen3.6-35B-A3B: 30% faster than Unsloth IQ on 6GB VRAM laptop
- π‘Antigravity 2.0 Tops the OpenSCAD Architectural 3D LLM Benchmark
- π‘Qwen3.6-35B-A3B Q4 262k context on 8GB 3070 Ti = +30tps
- π‘What changed most after adding an AI interview assistant wasn't what I expected
- π‘Qwen3.6 27B Pure Quant: 40 tok/s on 16 GB VRAM
- π’The AI memory problem isnβt storage. Itβs memory rot.
- π’Are people actually using AI meeting data inside agent workflows yet?
- π’Sharing some benchmark results from a memory system I'm part of building
- π’Anonymous Data Upload for Submission [D]
2026-05-22
- π΄[AMA] Got laid off 3 weeks ago. Instead of updating my resume I went down a rabbit hole. Here's what I found
- π΄Multi-Stream LLMs: new paper on parallelizing/separating prompts, thinking, I/O
- π΄110 tok/s with 12GB VRAM on Qwen3.6 35B A3B and ik_llama.cpp
- π‘Heretic has been served a legal notice by Meta, Inc.
- π‘We're Thursday and no one claimed AGI yet this week!
- π’Can liveness detection models generalise to synthetic media generation techniques they were never trained on? [D]
- π’Latest b9274 Addresses MTP VRAM leak
- π’Anyone evaluated the difference between Qwen Code for the local qwen models vs another harness? CC, OC, LC, Aider etc..
- π’Lisbon Machine Learning School (LxMLS 2026) [D]
2026-05-21
- π΄An OpenAI model has disproved a central conjecture in discrete geometry
- π΄HuggingFace benchmark datasets now let you filter by model size
- π΄Googleβs AI is being manipulated. The search giant is quietly fighting back
- π‘OpenAI claims a general-purpose reasoning model found a counterexample to Erdos's unit-distance bound [D]
- π‘CohereLabs/command-a-plus-05-2026-bf16 Β· Hugging Face
- π‘Any tool to get accepted conference papers sorted by citation count? [D]
- π‘@OpenAI: Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul ErdΕs in 1946. For nearly 80 years, mathematicians believed the best possible solutions
- π’Qwen3.6 27B and llama.cpp appreciation post
- π’Columbia Machine Learning Summer School (MLSS) 2026 [D]
- π’Typewise (YC S22) Is Hiring an AI Growth Engineer (Zurich or Remote)
- π’I built AgentLighthouse, a local βLighthouse for AI agentsβ that scans repos/docs/APIs for agent readiness
- π’HalBench: I built a custom sycophancy and hallucination benchmark and tested 4 frontier models (Sonnet 4.6, Grok 4.3, GPT 5.4 and Gemini 3.1 Pro), looking for input on what OSS models to run next!
2026-05-20
- π΄Iβve joined Anthropic
- π΄bytedance released an open source model that attempts to do just about anything with only 3b parameters
- π‘48GB VRAM users, what are your daily drivers? Do you wish you had more VRAM? What would you run if you did?
- π‘ICML Proceedings-only [D]
- π’Growing Neural Cellular Automata
- π’authentication/sesh timeouts in multi step browser agents
- π’Infomaniak transitions to a foundation model to protect user data privacy
- π’Show HN: The AI Quant Desk for Onchain Finance
2026-05-19
- π΄Have you actually found an AI tool that remembers across sessions, or are you just patching the context manually?
- π΄Still happy for yall
- π΄Anthropic acquires Stainless
- π΄Elon Musk has lost his lawsuit against Sam Altman and OpenAI
- π‘The last six months in LLMs in five minutes
- π‘@GoogleDeepMind: The stage is set. The tech is ready. Are you? π Join us tomorrow for #GoogleIO as we unveil the breakthroughs, tools, and innovations shaping the future of AI. Tune in live right here on @X from 10a
- π‘llama.cpp MTP support landed - Qwen3.6 27B at 2.44Γ on a Strix Halo, 2.17Γ on a RTX 3090 rig
- π’First-time ICML workshop acceptance (GlobalSouthML) but can't afford to travel to South Korea. What are my options? [D]
- π’We have sub-agents at home
- π’LLMCap β A proxy that hard-stops LLM API calls when you hit a dollar cap
- π’Audio upscaling, cleanup, or improvement models?
2026-05-18
- π΄I hope that someday we will have a 124B Gemma.
- π΄85 GPU-hours comparing 5 abliteration methods on Qwen3.6-27B: benchmarks, safety, weight forensics - Abliterlitics
- π΄I don't think AI will make your processes go faster
- π‘Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention [P]
- π‘I built a coding agent that gets 87% on benchmarks with a 4B parameter model, here's how
- π‘@simonw: Also a great example of positive contribution to open source by wanderingmeow - you don't need to contribute code to have a positive impact, just providing detailed feedback and confirmation that some
- π’Benchmarking the new b9200 update: Optimizing Qwen 3.6 27B mtp for Hermes Agent on a single RTX 3090
- π’Benchmarking vLLM vs SGLang vs llama.cpp on a mixed Blackwell/Ada cluster
- π’Has anyone tried repriced.ai?
- π’Qwen 3.6 27B Q8 on four Nvidia RTX A4000 (16GB each) with Llama.cpp and MTP enabled
- π’I trained TIME: short context-triggered thinking on Qwen model instead of overthinking
2026-05-17
- π΄Strix Halo Llama.cpp MTP Benchmarks: 27B Gets Much Faster, 35B Is Mixed
- π‘Program misleading high school students into paying to perform academic misconduct in ML Research [D]
- π‘Corsair desktop PC with Ryzen 395 and 128GB of unified RAM, has anyone tested it for LLM? Seems "a good" price
- π‘Testing llama.cpp MTP support on Qwen3.6 - RTX 5090
- π‘Do you agree with Judea that learning from data is not everything? [D]
- π‘Google literally dropped the new SEO playbook for AI
- π’Llama.cpp MTP with Qwen3.6 27B on Headless RTX 3090
- π’Ran the same models across Strix Halo, RTX 3090, and RTX 5070 because I wanted my own numbers
- π’Help with CNNs.[D]
2026-05-16
- π΄I believe there are entire companies right now under AI psychosis
- π΄Backlash against Arxiv's proposed 1 year ban is genuinely perplexing. [D]
- π΄Show HN: Watch a neural net learn to play Snake
- π‘Qwen3.6-35B-A3B and 9B are officially on the public Terminal-Bench 2.0 leaderboard!
- π‘Frontier AI has broken the open CTF format
- π’Can a 5090 with qwen3.6 achieve > 3,000 tok/s ? bring your pitchforks (open-dllm)
- π’PINN is predicting trivial solution for stiff ODE [D]
- π’AllenAI has been iterating on their MolmoAct2 models for robotics
- π’Someone Shared a Real Monet Painting as AI and Asked for Critiques
- π’GetMCP: Zero Trust for AI Agents
2026-05-15
- π΄Ontario auditors find doctors' AI note takers routinely blow basic facts
- π‘Would a 2000-2021 ML paper even get accepted today? [D]
- π‘Access to frontier AI will soon be limited by economic and security constraints
- π‘Need a second pair of eyes, this Qwen3.6 27B quant recipe consistently thinks less and is correct
- π‘@AnthropicAI: We've published a paper that explains our views on AI competition between the US and China. The US and democratic allies hold the lead in frontier AI today. Read more on what itβll take to keep that
- π‘eight months running autonomous business agents in production with real money. here is the specific failure mode that benchmarks structurally cannot surface.
- π’Our AI agent told a customer our competitor was better. That's when we realized generic guardrails aren't enough.
- π’Used over a million tokens in three separate sessions to test Qwen 3.6 35b (new Multi-token Prediction version)
- π’MiniMax M2.7 ultra uncensored heretic is Out Now with 4/100 Refusals, Available in Safetensors and GGUFs Formats!
- π’LLM Policy for Rust Compiler
- π’club-5060ti: practical RTX 5060 Ti local LLM notes and configs
2026-05-14
- π΄Built Support Vector Machine(SVM) from scratch in Rust [P]
- π‘MI50s Qwen 3.6 27B @52.8 tps TG @1569 tps PP (no MTP, no Quant)
- π‘Have the "on-hold" durations been getting longer for arXiv submissions? [D]
- π‘24+ tok/s from ~30B MoE models on an old GTX 1080 (8 GB VRAM, 128k context)
- π‘The US is winning the AI race where it matters most: commercialization
- π‘@swyx: if your reaction to this is βhaha openclaw bad, see prompt injection is the #1 dangerβ you: 1) havent sufficiently appreciated the layers to this tweet 2) havent seen enough ai api keys
- π’Arena AI Model ELO History
- π’Anyone actually using a local LLM as their daily knowledge base? Not for coding, for life stuff. What's your setup?
- π’GPT 5.5 v/s GPT 5.4. Paying 63% more just for 0.1 point difference!
2026-05-13
- π΄Needle: We Distilled Gemini Tool Calling Into a 26M Model
- π΄Dad why is my sisters name Lora?
- π‘Reimagining the mouse pointer for the AI era
- π’I created a minimal one-file implementations (160loc) of JEPA family (ijepa, vjepa, vjepa2, cjepa) for educational purposes [P]
- π’AntAngelMed - 100a6b Healthcare LLM
- π’How many of you tried BeeLlama.cpp? How's it? Agentic coding possible with 8GB VRAM?
2026-05-12
- π΄Computer build using Intel Optane Persistent Memory - Can run 1 trillion parameter model at over 4 tokens/sec
- π΄If AI writes your code, why use Python?
- π΄Found a way to cool the DGX
- π‘Interactive JensenβShannon Divergence Visualisation [P]
- π‘MiniCPM 4.6
- π‘ICML Author Removal [D]
- π‘Google says criminal hackers used AI to find a major software flaw
- π‘@AnthropicAI: Claude's Constitution is now an audiobook, read by two of its authors, Amanda Askell and Joe Carlsmith. It includes a Q&A on the writing process, the philosophies that shaped the document, and how it
- π’Drastically improve prompt processing speed for --n-cpu-moe partially offloaded models
- π’Online RL Reading Group[D]
- π’I let AI build a tool to help me figure out what was waking me up at night
- π’Most RAG apps in production are confidently wrong and nobody talks about this enough
- π’Interaction Models from Thinking Machines Lab [P]
2026-05-11
- π΄Local AI needs to be the norm
- π΄Openclaw ia trending down and will disappear soon
- π΄PhD students in ML, how many hours on average do you work? [D]
- π΄Getting a feel for how fast X tokens/second really is.
- π΄Running Qwen3.6 35b a3b on 8gb vram and 32gb ram ~190k context
- π‘MTP benchmark results: the nature of the generative task dictates whether you will benefit (coding) or get slower inference (creative) from speculative inference. No other factor comes close.
- π‘I Think I Spent Way Too Much Time Messing with Local LLMs
- π’unsloth/MiMo-V2.5-GGUF Β· Hugging Face
- π’Any news (or hope) of Qwen-3.6 14B and 9B distills for local coding ?
- π’Why is human LLM annotation so expensive? [D]
- π’Has anyone been able to get Draft Models to load in LM Studio?
2026-05-10
- π΄80 tok/sec and 128K context on 12GB VRAM with Qwen3.6 35B A3B and llama.cpp MTP
- π΄Apple Removes 256GB M3 Ultra Mac Studio Model From Online Store
- π΄BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)
- π‘Running Minimax 2.7 at 100k context on strix halo
- π‘LLMs corrupt your documents when you delegate
- π‘EEML 2026 summer school [D]
- π’The gap between knowing something and actually understanding it β AI accelerated my learning curve
- π’Anyone Trying to submit for ICML FM4LS workshop but noticed link closed Early? [D]
- π’Homelab setup
2026-05-09
- π΄OpenAIβs WebRTC problem
- π‘Qwen 35B-A3B is very usable with 12GB of VRAM
- π‘I built a platform to run AI employees and companies autonomously.
- π‘NeurIPS reviewers, any word after the invite email? [D]
- π’Got MTP + TurboQuant running β Qwen3.6-27B -- 80+ t/s at 262K context on a single RTX 4090
- π’Can LLMs model real-world systems in TLA+?
- π’Neurips : Pushing anonymous repo after rebuttal [D]
2026-05-08
- π΄Collected the infinity stones
- π΄Dirtyfrag: Universal Linux LPE
- π΄WARNING: Open-OSS/privacy-filter MALWARE
- π΄ECCV reviewer wants me to compare and contrast to my own paper. [D]
- π‘Steam Similarity Recommender [P]
- π‘3 things every AI founder forgets in their Terms of Service
- π‘@AnthropicAI: Our security bug bounty program is now public on HackerOne. We've run the program privately within the security research community, and their findings have strengthened our products. Now anyone can
- π’Two Home Affairs officials suspended after AI 'hallucinations' found
- π’"Hardware is the only moat" - Should we buy new hardware now or wait?
- π’A polynomial autoencoder beats PCA on transformer embeddings
- π’Gift to myself : tiny lab
2026-05-07
- π΄Stop letting LLMs edit your .bib [D]
- π΄Weights & Biases New Master Service Agreement Questions [D]
- π‘META Superintelligence Lab Presents: ProgramBench: Can SOTA AI Recreate Real Executable Programs(ffmpeg, SQLite, ripgrep) From Scratch Without The Internet?
- π‘Analysis of the 100 most popular hardware setups on Hugging Face
- π‘Get faster qwen 3.6 27b
- π‘Learning the Integral of a Diffusion Model
- π’ProgramBench: Can Language Models Rebuild Programs from Scratch?
- π’Need advice on hardware purchasing decision: RTX 5090 vs. M5 Max 128GB for agentic software development
- π’Visual Perceptual to Conceptual First-Order Rule Learning Networks [R]
- π’Any tool that tells you the cheapest setup needed to run a model? I want to know the cheapest setup that can realistically run Qwen 3.6 27B at decent speeds.
- π’NeuIPS submission small formatting question [D]
2026-05-06
- π΄DeepSeek V4 being 17x cheaper got me to actually measure what I send to cloud vs what I could run locally. the results are stupid.
- π΄Heretic 1.3 released: Reproducible models, integrated benchmarking system, reduced peak VRAM usage, broader model support, and more
- π΄NeurIPS Submission Number [D]
- π‘Transformers with Selective Access to Early Representations [R]
- π‘Quality comparison between Qwen 3.6 27B quantizations (BF16, Q8_0, Q6_K, Q5_K_XL, Q4_K_XL, IQ4_XS, IQ3_XXS,...)
- π‘Production AI very different from the demos [D]
- π‘Dense Model Shoot-Off: Gemma 4 31B vs Qwen3.6/5 27B... Result is Slower is Faster.
- π’What do you use Gemma 4 for?
- π’How to get from vibe-coding to compounding revenue growth using AI agents for GTM β free session with ThriveStack + Brevo
- π’Radar Engineer to Autonomy/AI [D]
- π’Wiki Builder: Skill to Build LLM Knowledge Bases
- π’What if AI agents can now talk?
2026-05-05
- π΄it's time to update your Gemma 4 GGUFs
- π΄How OpenAI delivers low-latency voice AI at scale
- π‘DeepSeek V4 Pro matches GPT-5.2 on FoodTruck Bench, our agentic benchmark β 10 weeks later, ~17Γ cheaper
- π’Is there a notable increase in demand for privacy-preserving AI/ML with the advent of LLMs? [D]
- π’Train Your Own LLM from Scratch
- π’MTPLX | 2.24x faster TPS | The native MTP inference engine for Apple Silicon
- π’[D] What Happened to Neurips Creative AI Track? [R]
- π’Building a 9-ball AI player: Candidate generation for direct cut shots [P]
2026-05-04
- π΄A Qwen finetune, that feels VERY human
- π‘Anyone submit ML articles to ACM journals (eg. TOPML or TIST)? [D]
- π’Mistral-Medium-3.5-128B-Q3_K_M on 3x3090 (72GB VRAM)
- π’Built a LangChain middleware that enforces signed authorization receipts before every tool call. Here is why wrap_tool_call is the right enforcement point.
- π’UAI Reviews disappeared [D]
- π’OpenAIβs o1 correctly diagnosed 67% of ER patients vs. 50-55% by triage doctors
- π’Where should you invest in the coming 25 years? Check the screenshot
2026-05-03
- π΄Group averages obscure how an individual's brain controls behavior: study
- π‘Real World Physics-Informed AI Applications [D]
- π‘If you've been waiting to try local AI development, please try it
- π‘Anyone submit ML articles to ACM journals (eg. TOPML or TIST)? [D]
- π‘Should I follow-up with the editor for a TMLR paper awaiting final decision? [D]
- π‘Open Weights Models Hall of Fame
- π’A Qwen finetune, that feels VERY human
- π’UAI Reviews disappeared [D]
- π’The Oscars just banned AI from winning acting and writing awards
- π’RTX A5000 Pro Balckwell 48GB
- π’I got tired of copy-pasting the same skills directories across 8 projects, so I built a sync'd registry for them
2026-05-02
- π΄Been using Qwen-3.6-27B-q8_k_xl + VSCode + RTX 6000 Pro As Daily Driver
- π΄Qwen3.6-27B at 72 tok/s on RTX 3090 on Windows using native vLLM (no WSL, no Docker), portable launcher and installer
- π‘A Dark-Money Campaign Is Paying Influencers to Frame Chinese AI as a Threat
- π‘Show HN: Mljar Studio β local AI data analyst that saves analysis as notebooks
- π‘Why ML conference reviews sometimes feel like a βlotteryβ [D]
- π‘"Prompt Engineering" certs are a joke. So we built a FREE Agentic AI Practitioner Exam that actually forces you to build working swarms to pass.
- π’Mistral Medium 3.5 128b ggufs are fixed
- π’Real World Physics-Informed AI Applications [D]
- π’MiniMax M2.7 AWQ-4bit on 2x Spark vs 2x RTX 6000 96GB - performance and energy efficiency
- π’I built "Semvec": A Constant-Cost Semantic Memory for LLMs (Looking for testers!)
- π’NEED HELP URGENT I really need to talk to someone who sells chatbots to local businesses, please