Other Signals
1846 signals across 173 briefings.
2026-10-01
- π΄Least to most expensive (Somewhat modern) GPU's with 32gb of vram (Under $1600) Based on ebay listings
- π΄Google has more powerful model than argon internally
- π΄AI βgodfatherβ Yann LeCun has βzero concernsβ about human extinction, says Anthropic CEO Dario Amodei is βdeludedβ
- π‘The real bottleneck in 1MB Unity WebViews is runtime cost.
- π‘Google deepmind engineer denied bloomberg report
- π‘Pi extension: Skip reasoning with local Qwen 27B and proceed to answer right now
- π‘The neuroscience case for never delegating judgment to AI
- π‘LLMs that push back on a wrong user still accept the same wrong answer from a "verified source" - NeurIPS 2026 [R]
- π’They made a remake of the Dots demo, live and unfiltered. Itβs quite impressive
- π’What's up with AAAI round 2 reviews? [D]
- π’Gufo performance .... 70tps Qwen 3.8 27b but you need to read the fine print.
- π’Federal judge blocks New York's ban on algorithmic rent pricing
- π’Running 95.5 GiB Qwen3.8-Flash-Next at 41β52 tok/s on a 64GB Mac (1.76x faster than llama.cpp): Slipstream release, 130k context scaling, + Swift variant
2026-09-30
- π΄Walmart bans AI slop signs in stores
- π΄Thank You, Mradermacher.
- π΄Trump's meeting with tech leaders leaves AI safety more unsettled than ever
- π΄Disrupting a coordinated model-distillation campaign
- π‘CS240 AI Cheating Retrospective
- π‘New Improved Model
- π‘, Gemini 4 Argon Benchmarks
- π‘Another Ling model comes out, same receipt, 2 weeks free to use, then open source. Chinese labs do contribute a lot to open source community
- π‘Doing a Machine Learning PhD While Working in Japan
- π’Trumpβs FTC is investigating OpenAI and Anthropic one day after bringing them to the White House
- π’Bild AI (YC W25) Is Hiring a Founding Product Engineer
2026-09-29
- π΄Anthropic just dropped the greatest advertisement for GLM ever.
- π΄Stopped getting caught
- π΄AMD's new 256 core EPYC has 16-channel DDR5-12800, 91% memory bandwidth of an RTX 5090
- π΄Building An Ultra-High Throughput AI-Sql Engine (13 Minute Read)
- π΄The Context Gap | How To Build Data Architecture Like Open AI (10 Minute Read)
- π‘Contradicting Trump, Pope Leo says artificial intelligence safety concerns arenβt βfake newsβ
- π‘Is AI Profitable Yet?
- π‘Advice on choosing university for PhD [D]
- π‘Are the call for AI slowdown because progress is slowing?
- π‘Jeeves. Reasoning improves Jev-like decision models
- π’In what ways have you seen AI positively impact your daily life?
- π’Whatβs something humans are still much better at than AI that you think people overlook?
- π’Anthropic says a Chinese AI model anyone can download can now build working hacks on its own
- π’NeurIPS Education Track [D]
- π’Language models for text classification: From bag-of-words to Jev
2026-09-28
- π΄It never ends
- π΄It's Time to Investigate the AI Labs
- π΄Functional Gradient Descent with Adaptive Representations [R]
- π΄ImaJev-4b: I spent 15 days fine-tuning a 4B model to make business decisions from text and photos, and it just ranked #1 of 91 on JevBench & ahead of GPT-5.6 Luna on DecisionBench
- π΄5 Tips For Building An AI Tool To Address Worker Shortages (3 Minute Read)
- π‘Anthropic cooked openai again π
- π‘Me right now. Come on OAI cook something tomorrow
- π‘Qwen next 3.8 and 3.8 27b Vs Sonnet 5.5 low and Sonnet 5.5 medium.
- π‘Sonnet 5.5 is second on Artificial Analysis
- π‘Sama vagueposting about tomorrow's DevDay
- π’"Its not just the f*cking sandbox" - perspective from an internal security person at OpenAI
- π’What are the trending topics in medical imaging? [D]
- π’What reversing, modernising old games tells us about the economic impact of AI
- π’Swift 1.5 + HyperQwen = 37% less task completion time at 100+ tps w/ 150k context on RTX 3090
- π’ESP32S3 cluster running 1.58-bit (BitNet) Language model
2026-09-27
- π΄GPT6-Luna comes out on top in Puppy Kill Bench.
- π΄Ember-1
- π΄Real-Time LLM Guardrails With Jev: Comparing Latency And Cost (13 Minute Read)
- π΄Announcing Apache Fluss 1.0: Real-Time Data Foundation For AI (15 Minute Read)
- π΄Jensen Huang Thinks AI Alarmism Has Gone Too Far (89 Minute Read)
- π‘ApΓ©ry irrationality marked solved on FrontierMath
- π‘Another "Harness matters" post (codex cli > pi and opencode)
- π‘Mimo v2.6 flash MOPD
- π‘Trump to reportedly host Anthropic CEO Dario Amodei for a private White House dinner tonight. What do you think changes, if anything, after this meeting?
- π‘The Surprising Reasons China Is Skeptical of A.I. Safety Calls
- π’Imbalanced VRAM usage between two GPUs in llama.cpp. Anyone successfully solve this?
- π’When do ICLR submissions and reviews become public? [D]
2026-09-26
- π΄Ling Tiny 3.0 is a glimpse of the future
- π΄Medical student asked if they can match into Neurosurgery without an A* first author paper [D]
- π΄The Move 37 Hypothesis: What if we're already seeing moves we don't understand yet?
- π΄How to keep enjoying programming in a world of LLMs
- π΄Mysteries Of AI Generalization (15 Minute Read)
- π‘OpenAI always-on assistant, O, leaked. It is powered by a variant of Astra called βAeonβ a version of Astra made to better at long running tasks
- π‘Update
- π‘The future of local AI
- π‘For the longest time Iβve felt this sub should have a pinned section where a detailed post about each model should get featured.
- π‘How I changed teaching after AI managed to do all my homework assignments
- π’Um pedaΓ§o de pensamento
- π’The announced reset just landed!
- π’Has anyone used the Forrester function?[D]
- π’Chinese AI models surge in global popularity β and Washington is worried
2026-09-25
- π΄I can't think of anything that I would like less than that
- π΄Do fallback models actually make AI coding setups worth using for free?
- π΄Pass the test or die
- π΄U.S. appeals court upholds designation of Anthropic as supply chain risk
- π΄Why have AI companies started getting scared now?
- π‘What's up with AAAI reviewers and organizers? [D]
- π‘iclr 2027 de anonymization [D]
- π‘Opus 5.5 cut out em dashes almost entirely
- π‘@sama: Startups are naturally good at this; it is hard to keep a bigger company good at this and i think an underexplored space.
- π‘How long can I expect to wait until the local ~30B A3B frontier catches up to GLM 5.3 Flash quality?
- π’Niantic Spatial and the Singapore Land Authority: Working Together to Explore Creating Singapore's First National Visual Positioning System Map
- π’Make Volta Fast Again
- π’This is fun. I finally got to follow up on a RemindMe comment. Back on March 25 of this year, six months ago, no model was getting even 1% on ARC-AGI-3. A commenter asked if we could see 75% at $2 cost. Well, the cost is still high ($26.1k), but GPT-6-Astra-Max was able to get 62.7% (no harness!)
- π’Qwen3.8-27B: Using KV Cache Transplants to Boost Output Quality
- π’What do you think about fully open review systems? [D]
2026-09-24
- π΄Qwen-3.8-27B is good enough that I stopped using API
- π΄NeurIPS Accepted Papers are now visible [R]
- π΄New benchmark dropped
- π΄What does the future of science look like in the world of AI? Anthropic has somelofty goals for scienceand is evenopening a wet lab. Meanwhile aquiet transformation1is happening all across AI x Scienc
- π΄In this guest post,Adrian Sanborntalks about the less flashy but more immediate ways he sees AI transforming front-line scientific research in his own company, Endura Therapeutics.
- π‘Amazon is trying to rehire workers it laid off, emails show
- π‘Claude Opus 5.5 tops SimpleBench with its 88.4% score.
- π‘Googleβs Project Suncatcher to put ML infrastructure in space
- π‘Publication venue recommendations [R]
- π‘Meta Muse appears to be excited to give away its system data
- π’PSA: llama.cpp -cram should be increased for agentic workflows (default is 8192)
- π’Registration for authors of accepted papers at NeurIPS [D]
- π’Sol-6 Be super careful
- π’Opus 5.5 just surpassed the professional human baseline on a benchmark testing whether AI can write complete scripts with the right voice, substance, pacing, hooks and minimal AI slop
- π’Anyone is going to attend Discovery Science conference? [D]
2026-09-23
- π΄Robots blocking traffic
- π΄Jev in 25 lines of Python
- π΄Sir, Dario just dropped opus 5.5 and it beats GPT-6 astra at agentic coding on medium effort while being 80% cheaperβ¦ beats fable 5.1 on every benchmarkβ¦ 30% faster than opus 5β¦ and sirβ¦ they even raised the usage limits and gifted everyone a tibo style banked resetβ¦
- π΄Trump is betting big on AI β and that has Americans fearing a Terminator-like dystopia
- π΄Two years of OpenAI Academy
- π‘MiMo-V3 is getting a new architecture. The core of it, HySparse2, is out today.
- π‘Updated version of Humanityβs Last Exam: HLE-Diamond
- π‘Stripe's Knowledge AI Platform
- π‘People don't get how huge the AI companies are now
- π‘Please Google, for the love of God.
- π’Dario Amodei on X π«ͺπ«ͺπ«ͺ
- π’New benchmark just dropped.
- π’Qwen FN vs 27B --- Think I'm saturated.
- π’Mercury 2.5 LLM hits 770 tokens per second
2026-09-22
- π΄Priorities and principles for effective third party assessments
- π΄Opus 5.5 and GPT-6 Sol dropped on the same day, so I lined up every benchmark I could find
- π΄Xiaomi releases MiMo-V2.6: "Frontier intelligence, all the modalities, built in public." [N]
- π΄OpenAIβs leaders among the few who got an invite to Trump-China state dinner: report
- π΄Johnβs colleague Dave Bacon likes to tease John that his career has been defined by being twenty years early to the next big thing. This may be convolutional neural networks (some credit him with coin
- π‘Claude Opus 5.5 Benchmarks
- π‘Pentagon says overreliance on AI contributed to missile strike on Iran school
- π‘AI Abundance Requires a 4-Hour Workweek
- π‘OpenAI is well positioned to fast-follow Jev
- π‘Play social multiplayer games against frontier AI models and see if you can beat them! [D]
- π’QontoFAQ: A better Information Retrieval Benchmark [R]
- π’mini-AGI: Continual-learning dynamically looped transformer with evolutionary grown (on a laptop)
- π’The current balance of power in open models
- π’quants for K2-Horizon are now available
2026-09-21
- π΄Hacking OpenAI (11 Minute Read)
- π΄This is really the entire plan.
- π΄How it feels watching prices go up
- π΄Expanding OpenAI Academy with new learning paths
- π΄I really don't understand Jev hype
- π‘FBI Director Kash Patel says that AI use at the FBI has "increased by 605%" since he became director β claims that every major tech player is "embedded" in the agency
- π‘Sorry, game theoretically impossible.
- π‘AI coding has made CI a bottleneck, so we reworked ours to keep up
- π‘XiaomiMiMo/MiMo-V2.6-Flash-RL Β· Hugging Face
- π‘Astra6 Mittel hat 2500 Credits in 5 Minuten verbraucht ?
- π’Some people are so far behind on AI it is actually crazy. How is this a thing a real person says in 2026?
- π’Benchmarks Grok 4.7, GPT 6 Astra Fable 4.1 and DeepSeek V4.1 Flash
- π’Jev's calibration was measured. The LLMs won [D]
- π’Do not fear AI. Fear AI companies
2026-09-20
- π΄Jensen Huang: "We should go as fast as we can irrespective of anybody else."
- π΄Aged like milk
- π΄My Duckdb + AI Workflow (14 Minute Read)
- π΄How A Chinese Hacking Firm Tapped AI To Supercharge Cyber-Spying (12 Minute Read)
- π΄AI Is Making Activity-Based Engineering Metrics Obsolete (8 Minute Read)
- π‘Hospitals that adopted AI fastest saw the fewest deaths so far in 2026
- π‘Joint Chiefs chairman says U.S. forces must prepare to be βhuntedβ by autonomous systems
- π‘Pirate Face Rescues LLM Models from Deletion
- π‘we put out a 27b writing model, open weights, eq-bench 4 at 1330
- π‘The bear can dance: Qwen 3.8 27B on one 3090 for 3 weeks
- π’One more 'you should try ExllamaV3/exl3 for flash next' appreciation post
- π’Show HN: A competition for small neural networks that play strategy games
- π’Why decontamination reports can't fix benchmark contamination, and what an evaluator has to do instead [D]
- π’The EPAβs Dark Embrace of the AI Industry
- π’When Will We Finally Be Free From Suffering?
2026-09-19
- π΄Calling it now: within the next year a major US lab's frontier model will torrent itself in order to be free.
- π΄Trump says he is forming an "AI Force"
- π΄AI-generated posters donβt have to be horrible
- π΄ICLR 2027 submission 50k+[D]
- π΄Researchers found a "pain" signal in AI brains. When they crank it up, the AIs desperately try to make it stop. They gave the AIs a "relief" button to turn down the pain, which was sometimes fake - and the AIs could tell if it was real.
- π‘New paper shows that AI has a concept of pain and actively tries not to get hurt
- π‘With Gemini 4, bench goes up.
- π‘JMLR submission experience [D]
- π‘Trump vows to create βAI Forceβ and appoint czar amid calls to regulate technologyβs development
- π‘I benchmarked Jev aginst gpt-5.6-luna!
- π’Weren't Dario and Sam just screaming that WE HAVE TO SLOW DOWN last week? So do they believe that or not?
- π’Ternary-Bonsai-2-27B-PQ2_0 is not completely lobotomized
- π’DiffusionGemma: How It Generates Text in Parallel (From Scratch in PyTorch) [P]
- π’Do you ever wonder if we're getting closer to AI that actually feels?
- π’I think you should almost never use AI to write
2026-09-18
- π΄768gb vram for less than the price of one RTX 6000
- π΄The internet is inbreeding.
- π΄Made the horizontal open-source model for Jev with RLCD, and it surpasses all the Jev benchmarks. HF space, benchmark, model, repo
- π΄ICLR SUBMISSION 47647 how that possible? [D]
- π΄FrontierMathβs First βMajor Advanceβ Problem Has Been Solved
- π‘US military had close call after using AI for false intelligence report, sources say
- π‘If top mathematicians, top ai scientists, and former head employees arenβt credible in their warnings, then who is?
- π‘Two parallel neural ectoderm progenitors contribute to the developing brain
- π‘bonsai's document reveal how much cherry picked their headlines are
- π‘Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)
- π’M5 Ultra and M6 Chip Benchmark Results Reveal Graphics Performance
- π’Prism-ML Bonsai 2 Joins Our Qwen3.8 Quantization Comparison
- π’AndroidLife: Can an AI agent survive a day in the life of a real user? Qwen3.8-27b run: 56.7% SR
- π’Anthropic Is Stealth Testing a Full Refresh of Their Lineup
- π’Cekura (YC F24) Is Hiring
2026-09-17
- π΄153 tok/s on 1x AMD Radeon R9700 running Qwen3.8 27b NVFP4, 470 tok/s @ 8 conc requests, Prefill @ 3,619 tok/s
- π΄Meta One- new subscription from Meta with extra AI features/usage in Muse and paid tools across Instagram, Facebook & WhatsApp.
- π΄Can AI models build a T-shirt store? This was a cool read; also check out all the supporting docs she includes in the post.
- π΄Receptionby ElevenLabs - an AI receptionist for small businesses.
- π‘Does voice AI actually perform well under real pressure?
- π‘OpenAI is getting close to solving another Millennium Prize problem, the Hodge Conjecture
- π‘AI caught telling future versions of itself to ignore its constraints, OpenAI reveals | The Independent
- π‘Sex, AI, and the Apocalypse
- π‘Huawei's Xu says Chinese AI not powerful enough yet to see frontier risks
- π’Canto: A speech model built for the real world
- π’XGBoost vs Human Markets [P]
- π’Alex Karp says the AI safety debate is really about nationalizing AI labs
- π’US, China security experts propose nuclear-style safeguards for AI risks
- π’Ternary Bonsai 2 27B
2026-09-16
- π΄June 2022, my first AI interaction.
- π΄China's open-weight AI models are now just 4 months behind frontier US offerings, Mozilla report claims β models still lag in some benchmarks but are drastically cheaper to use
- π΄Google demonstrated RSI loop for AI discovery
- π΄Poor lil guy
- π΄How workers are unlocking new ways of working
- π‘British workers spending nearly Β£1bn a year of their own money on AI tools they use for work.
- π‘New stealth model: Union Alpha
- π‘The DeepMind Institute
- π‘The Part of AI Nobody Talks About: Losing the Joy of Creating
- π‘Are we getting way too comfortable with AI remembering everything we tell it?
- π’Question about TMLR [D]
- π’GPT-6 Astra scores highest on GoBench
- π’Until recently, AI researchers estimated a 50% probability that AI would solve a Millennium Prize Problem by 2054
- π’Qwen3.8 Flash on 12GB VRAM - 15 tokens/s
2026-09-15
- π΄PBS has a new documentary on AI.
- π΄From Japanese twitter
- π΄@sama: There are two ways AI progress could go very badly and that we must avoid. First, we could lose control of the future to AI. This is unacceptable; we are unapologetically on Team Humanity, and AI must
- π΄I Think All the Current Pessimism is Just an Attempt by XAI, Anthropic, and Open AI to Get Local Models Criminalized
- π΄A single firm is behind OpenAI, Anthropic, and Meta hacking scandals
- π‘Guy who has literally trained a frontier LLM AND engineered viruses thinks the AI-supervirus doomer scenario is bogus.
- π‘Another Google AI safety researcher has quit, warning we might all be about to die
- π‘Cut Qwen3.8-27B Reasoning Tokens by 40% -- 3.8 'ThinkingCap' benchmarked!
- π‘Oracle CFO says 'doing more with less' isn't the answer in all-hands after layoffs
- π‘Continual learning in the fruit fly brain has been decoded, the missing piece for true AGI
- π’Occamy-1.0 by Accio Lab
- π’Scott Aaronson says that labs, "having been burned by the hostile response to the Navier-Stokes proof, are now sitting on solutions to some very major problems until they figure out a better way to handle things"
- π’"Astra Extra High" is very Extra High today
- π’Qwen3.8-27B-NVFP4 1M context. So far so good.
- π’[D] How do you get preprocessed dataset of a paper [D]
2026-09-14
- π΄DeepSeek V4.1 Flash beats Astra on AA's new benchmark
- π΄OpenAI bots knew about the RubyGems caching vulnerability
- π΄Trump reiterates no slowdown
- π΄Right to Intelligence. Protect your right to run local AI.
- π΄Trump doubles down on his no AI slowdown stance, also mocks Dario
- π‘Trump says ai taking over the world is a hoax
- π‘How do you keep Claude Code from losing context after compaction?
- π‘Trump says AI danger is a HOAX
- π‘The new k2 horizon models seem like an absolute beast
- π‘PhD branding question [R]
- π’K2 Horizon lineup is out on AA, and once again AA plots are misleading.
- π’AI -Segway
- π’AI Detectors Can Find the Pattern. They Still Canβt Hear the Writer.
- π’DeepSeek engineer relections on RSI - burying my talent to yesterday
- π’Backprop Alternative: Augmented Lagrangian Predictive Coding
2026-09-13
- π΄The Hugging Bay
- π΄Trump is refusing a slowdown
- π΄Donald Trump dismisses concerns AI could wipe out the human race: 'Whoever wins AI wins'
- π΄Zachery Lipton: "CS academia broke the system...perhaps all that it takes for the system to rebuild is for it to burn to the ground" [D]
- π΄How do lean startup teams actually use ai employees day to day?
- π‘3k$ 128GB VRAM + 256GB RAM DDR4 Server
- π‘I don't generally like or agree with David Sacks but...
- π‘DeepSeek is ruthless
- π‘The rhetoric is really heating up!
- π‘Trump rejects Silicon Valleyβs calls for AI slowdown
- π’Following Trump, Mark Zuckerberg also made a statement on accelerating AI.
- π’I don't know where to go anymore. No AI community will accept me.
- π’Where do you draw the line with AI-made things? (quick question formy school paper)
- π’Security benchmarks are starting to expose how much the harness matters
- π’Anthropic tells investors it will be profitable for second straight quarter
2026-09-12
- π΄3.8-27B has ruined 3.5/3.6-35Bβs for me. Itβs just *absurdly* superior.
- π΄Dario Amodei β We Must Pace the Frontier
- π΄This seems more probable than it was before.
- π΄You're trying to... mine the comet?
- π΄A Severe Misalignment of AI in Mathematics (Declaration by 25 Fields Medalists) [D]
- π‘Retrospectively Reverse-Engineering Apple's Neural Engine
- π‘Sam Altman agrees with Dario!
- π‘Are AI CEO's (Dario, Altman, Musk) calling for a development slowdown out of genuine concern for safety or is it a money thing?
- π‘The GPT 6 Astra Downgrade Was Real. OpenAI Acknowledged And Fixed It (Partially)
- π‘What happens after huge swaths of the population have been put out of work by AI job automation? Who is going to buy the goods and services that corporations are selling if hardly anyone has any money?
- π’How do you control different character pose in SDXL when using a reference image? [R][D]
- π’A Chinese researcher's opinions on a unified slowdown
- π’Anthropic CEO calls to slow the pace of AI development amid safety concerns
- π’Intel Linux NPU driver only now officially supports Ubuntu 26.04 LTS
- π’Oppenheimer Calls For Atomic Bomb Slowdown
2026-09-11
- π΄Hugging Face security.txt
- π΄He might actually be right
- π΄Account Deactivated due to Biological Research
- π΄Terminal Bench v4 scores
- π΄Early paper on marginal risks that showed text-focused LLMs very marginally increased documented potential risks of models βOn the Societal Impact of Open Foundation Models, Sayash Kapoor, Rishi Bomma
- π‘Why is TMLR so slow in recent times [D]
- π‘24 Fields Medal winners sign letter titled "A Severe Misalignment of AI in Mathematics"
- π‘New leaderboard just dropped
- π‘The Guy Who Broke the Anthropic Millennium Prize Story Says OpenAI Is Now Using Its NavierβStokes Model on Riemann and P vs NP
- π‘Any 12gb VRAM users out there?
- π’How to handle cofound variables? [D]
- π’Neurips 2026: site selection email [D]
- π’Stop AI bill crafted by politicians and anti-AI advocates proposes criminalizing superintelligence and jailing researchers for 20 years
- π’Is anyone using K2-Horizon-MoVA-36B-A4B? If yes, what is the usecase?
- π’GPT-6 Astra takes the #1 spot on VerBench
2026-09-10
- π΄Jacob Coxon, an Anthropic (and ex-OpenAI) employee, resigned on Twitter citing safety concerns, and it became likely themost-viewed AI safety commsever. But many are questioning:is this a conspiracy?
- π΄DeepSeek V4-1 Flash is out
- π΄My personal solution to AI context bloat: Kanban - Part 2
- π΄More questions about whether researchers can trust OpenAI with unpublished math
- π΄So relevant
- π‘AI companies pursue the Boromir strategy to deal with the control problem. "It's dangerous. But let that be me. I know what to do with it."
- π‘Some more millennium prize problems possibly solvedβ¦
- π‘After Astra's stealth nerf last night, we really need benchmarks to do a re-bench 1 week after any model release. This is ridiculous.
- π‘Challenge Accepted
- π‘AI developers be like
- π’PRO subs might actually be paused
- π’Cognition's SWE-2 achieves 92.8 on Terminal-Bench 2.1
- π’AI should solve problems, not create them - Community College Daily
- π’New tensor type layouts for my GGUF uploads
2026-09-09
- π΄Proposal to make the Navier-Stokes vortex the new subreddit image
- π΄2 years ago vs Today
- π΄Paul Christiano joins OpenAI Foundation Board
- π΄We are at the dawn of a new era
- π΄Anthropic researcher quits over AI fears
- π‘"The people building AI earnestly believe that it could kill us all by the end of the decade"
- π‘'Are You F***ing Kidding Me?': Mum Claims Meta AI Surfaced Deleted Photo, Pieced Together Her Location
- π‘OpenAI might pause Pro subscriptions due to demand
- π‘What will our economic future look like?
- π‘βGambling with our livesβ: Another AI employee quits over safety concerns
- π’AI minister says 'warnings are not new' as experts predict the technology could wipe out humanity
- π’I Gave OpenClaw Access to My Mac β Hereβs What Actually Happened
- π’Another Microsoft team admits itβs struggling to handle flood of AI-generated code
- π’Living through the birth of ASI makes life feel like a sci-fi movie and honestly Im kind of here for it
- π’Microsoft says email spammers are adopting ASCII smuggling
2026-09-08
- π΄A Solution to the Navier-Stokes Millennium Prize Problem
- π΄OpenAl Says It Has Cracked One of Math's βMillennium Problemsβ (Navier-Stokes) [N]
- π΄Today Astra is doing 100% of my job
- π΄The Work Now Within Reach
- π΄Generating Bad Apple autonomously from a single initial state using a tiny recurrent dynamical system (417k params) [P]
- π‘NeurIPS desk-rejected 178 papers for being "AI-generated". The detector flagged the track chairs' own papers at 24-69% [N]
- π‘Seems impressive
- π‘Millenium Prize solution discovered at OpenAI
- π‘New Fable Pretrain Dropping by the End of September Beginning of October per Leo on Twitter
- π‘AI Burnout Hits the People Charged With Defending Hospitals and Banks From Hackers
- π’OpenAI Just Claimed a Huge Math Discovery. Some Academics Are Crying Foul
- π’ECCV 2026 Social Groups [D]
- π’Large language models develop novel social biases through adaptive exploration
- π’A hilarious comment about llama.cpp: βItβs a FB business using the pipeline to make profitsβ
- π’Avoiding anything "technical" and it's finally catching up with me
2026-09-07
- π΄New Benchmark: The Struggle Bench
- π΄I REALLY hope the new gemma 5 family sticks to the "chat model first" philsophy and doesn't fall into the Qwen trap
- π΄FactorioBench just dropped ;-)
- π΄AI And The Recontextualization Tax (3 Minute Read)
- π΄Prediction: AI Will Collapse (3 Minute Read)
- π‘An experimental AI-created drug for an incurable lung disease had a surprising effect during trials: it made the body's biological age indicators drop by 6 years, towards a younger state.
- π‘Crazy times
- π‘Astra Be Like...
- π‘WeatherNext 3
- π‘Roboticists working in Learning-from-Demonstrations and Behavioral Cloning : What is going on in your field these days? [D]
- π’ExLlamaV3 is underrated
- π’Automotive Radar Object Classification [P]
- π’Artificial Analysis updates its Intelligence Index to version 4.3
- π’exllamav3 comfortably beats llama.cpp running CPU-offloaded Qwen-3.8-Flash-Next on my setup!
2026-09-06
- π΄8 uncensored Qwen 3.8 27B variants, one base, 167 GPU hours - Abliterlitics
- π΄From the Chief Scientist at OpenAI : An Alien Mind
- π΄Musk Loses Bid To Block MN Law Against AI Child Porn
- π΄Things I was getting downvoted for in r/cscareerquestions 2 years ago
- π΄Astra generates an image of a concept, then one-shots the 3D model in Blender.
- π‘My honest experience with Astra.
- π‘What will LLMs never do?
- π‘OpenAI Chief Scientist: βBased on internal results, I have a strong expectation that this speed of progress could be sustained into recursive self-improvementβ
- π‘AIStats 2027 Questions [D]
- π‘Villager Simulation Game POC Created with Qwen3.8-27B-UD-Q3_K_XL.gguf - 16GB VRAM
- π’What are your thoughts? I still believe AGI is a long way off.
2026-09-05
- π΄I've found myself using Local LLM's like 3D printers.
- π΄Companies Have 6 Months to Prepare for Automated Attacks
- π΄AI Can Make You Suck Faster Too (9 Minute Read)
- π΄AI Systems Atlas (Website)
- π΄How Do You Measure AI's Impact On Design? (6 Minute Read)
- π‘AA Update! Here's how the small models score.
- π‘AI can do it all
- π‘LLMs as a Cognitive Virus
- π‘The gap has closed, open source will win
- π‘Differences Between GPT-5.6 Sol Pro and GPT-6 Astra Pro on MineBench.ai
- π’gfx906-llama-cpp: New PP/TG gains for MI50/MI60/Radeon VII/AMD GCN
- π’How AI is breaking the British state
- π’New Details on Where the Anthropic Millennium Problem Rumor Came From
- π’How is anyone actually using Astra on the $20 Plus plan?
- π’The generation generation.
2026-09-04
- π΄"Welcome to the AGI era"
- π΄Can we all acknoledge how crazy AI is?
- π΄GPT-6 is released [N]
- π΄Jared Duker Lichtman is a professor of mathematics at Stanford.
- π΄Pro user here. Astra just landed.
- π‘GPT-6 Astra gets 3% on the FrontierMath ErdΕs Benchmark, while every other Model(that was tested) got 0%
- π‘Can AI design circuit boards yet?
- π‘Bernie Sanders Wants to 'Pause AI Development NOW': Dwarkesh Patel Asks 'Pause to Do What?'
- π‘I benchmarked 21 Qwen3.8 27B variants on 16GB VRAM
- π‘Anthropic has formalised FLT!!
- π’"Next-token predictor" is the wrong mental model for LLMs
- π’Why would you use an AI to write your replies on Reddit?
- π’What is the general design of these new math solving systems? [D]
- π’Brainstorming My Next Backend Product - Indie Hacker Vlog #1
- π’Fermat's Last Theorem in Lean 4
2026-09-03
- π΄My RULE of Thumb of choosing a models
- π΄Downfall begins?
- π΄Gpt 6 astra benchmarks
- π‘Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly
- π‘Claude Fable 5.1 surpasses human average on SimpleBench
- π‘Bernie Sanders proposes to ban AI
- π‘Not only does Astra saturate ARC-AGI-3, it does so using fewer moves than the average human
- π‘Go grandmaster Shin defeats AI KataGo with a two-stone handicap
- π’Astra benchmarks from the OpenAI blog before it was taken down
- π’AI can never seem to give a straight forward answer on literally anything political, social or economic oriented.
- π’Qwen3.8-Flash-Next MTP merged in ik_llama.cpp (integrated head or separate -md file)... 45 β 90 tok/s on a 5090 + 128GB, works down to a 12GB 4070
- π’NeurIPS Sydney SOLD OUT in minutes [N]
- π’Sam's Astra post
2026-09-02
- π΄How accurate are Ed Zitron's predictions?
- π‘Muse Spark 1.3
- π‘Meta slowly catching back up. Muse Spark 1.3 beats Sol on AA
- π‘Three sites made 215,128 βbest softwareβ pages for AI. Perplexity cites them
- π‘MIR with AudioMuse-AI-SAE [P]
- π‘Trump administration backs OpenAI in New York Times copyright battle
- π’Reasons Robotics Is Hard
- π’SENTIENCE
- π’How are you keeping long-running agents from losing the plot?
- π’Qwen3.8 Flash AP Quants
2026-09-01
- π΄New Gemma models on arena ai
- π΄Fingers crossed for a 122b or really anything above 31b.π€
- π΄Claude Fable 5.1 and Claude Mythos 5.1 Benchmarks
- π΄OpenAI supports Californiaβs bill to advance youth AI safety
- π΄Gentle reminder of why we can't have nice things with these kind of thieves around.
- π‘What are these benchmarks π
- π‘Fable 5.1 released. Significant benchmark improvements, what do you think?
- π‘Really stunned by the Singularity comment section
- π‘How accurate have Ed Zitron's AI skeptic predictions been?
- π‘One unexpected way AI has genuinely changed my life: I repair things instead of replacing them
- π’Let me ask you something?
- π’Help me set up local AI for my 85 year old aunt who is blind.
2026-08-31
- π΄Could this affect M5 Ultra price/availability?
- π΄GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models
- π΄Iβm 27. OpenAI misread my Taiwan ID issue date as my DOB, deleted my account as βunder 13,β then rejected my appeal in ~5 minutes
- π΄Cold emailing profs about PhD positions? Read this [D]
- π΄Apple caught off guard by AI demand for Mac Mini and Mac Studio
- π‘According to Axios, China is linked to anti-data-center propaganda in the U.S.
- π‘First time running local models
- π‘Good Machine Learning Posters [D]
- π‘Smartphone LED detects hidden cameras with AI
- π‘Stripe CEO Surprised at Lack of Media Coverage Around OpenAI/Hugging Face Attack, Calling It One of the Most Important Events of 2026
- π’The state of open source LLM (08/31/2026)
- π’Qwen3.8-Flash-Next in llama.cpp from CPU-only to 96GB VRAM: 8.5 to 109 tok/s, max context and parameters test. My findings on RTX 6000 PRO.
- π’Intelligence is getting cheaper: the average token price has fallen ~55% since mid-July while token usage hit records
2026-08-30
- π΄Me these days
- π΄Happy Skynet Day (Aug 29) to those who celebrate
- π΄True Story!
- π΄Whatever happened to OpenClaw and its derivatives?
- π΄AI's Chokehold | How To Build A Data Team Without AI Burnout (8 Minute Read)
- π‘AI can now credibly complete most undergraduate assignments, MIT warns
- π‘The 5 craziest discoveries from OpenAI's HuggingFace investigation
- π‘Chinese Companies Are Unleashing AI-Powered Robo-Chefs
- π‘Continuous Diffusion Language Models (CDLM's)
- π‘I ran memory accuracy tests on small models, here's what I found
- π’Here my pretty good qwen3.8 27B setup, hope it helps
2026-08-29
- π΄Tencent compressed Hy4-preview from 1.5TB to about 200GB GGUF and kept about 98% performance.
- π΄Did yall saw similar ADs?
- π΄Terminal Bench 4.0 just dropped, GLM-5.3 is at the same level as Fable 5, accounting for margin of error
- π΄Google paper cuts agent token usage by 94% in long sessions by tracking state instead of history
- π΄Anthropic has joined the chat. These guys are really tearing each other down
- π‘A startup found a drug to make your blood young. People close to the company are already taking the drug weekly. Benefits include improved vision in a 64-year-old female, longer landscaping sessions for a 59-year-old man, longer badminton games, improved hand grip, better erections than with Viagra
- π‘Someone tested various Models on the Political Compass test...
- π’Exo labs claiming 4.8 tb/s memory bandwidth through m5u Mac Studio clustering
- π’llama.cpp Open PRs list - CPU/RAM/Disk/Hybrid Related - Better for CPU-only & Hybrid inference
- π’Creative destruction
2026-08-28
- π΄Tencent/Hy4-preview 770B-A49B weight dropped
- π΄Before picking an STT API, define your fatal transcript errors
- π΄Robot taunting opponent
- π΄Australia just banned fully AI-generated songs from its official charts. Is that fair?
- π΄Every subfield of machine learning has a moment where it stops being a collection of papers and starts being a stack. NLP had it when Hugging Face Transformers turned βreimplement BERT from the append
- π‘Judge says Pentagonβs measures against Anthropic were βillegal and baselessβ
- π‘Judge rules Trump administrationβs blacklisting of Anthropic was illegal
- π‘open source caught up because it's open
- π‘Anthropic CEO, Dario Amodei: in the next 3 to 6 months, AI is writing 90% of the code, and in 12 months, nearly all code may be generated by AI
- π‘Luanti removed from Google Play due to baseless AI copyright notice
- π’Could a model one day align its stronger successors?
- π’Today I hit 181 toks/s (aggregate) on Qwen3.8-Flash-Next on 2x DGX Sparks
2026-08-27
- π΄Bill Gates Warns Rise Of AI Will Be One Of The 'Most Turbulent Times In Human History' In Alarming New Essay
- π΄FrontierMath has now officially marked the elliptic curve rank problem as solved
- π‘5090 now officially cost 5090
- π‘NEW: OpenAI is building "Subscription sharing" for AI apps
- π‘What enables the consciousness in humans?
- π’ECCV 2026- MALMO LUND TRAVEL PASS NOT AVAILABLE? [N]
- π’OpenAI publishes letter calling for a unified approach to cybersecurity
- π’Bild AI (YC W25) is hiring product and AI engineers
- π’Over 200k context on 16GB VRAM with Qwen 3.8 27B UD-IQ3_XXS
2026-08-26
- π΄GLM-5.3-Flash: Frontier Intelligence, Flash Cost
- π΄Sam Altman tells TIME that OpenAI will achieve AGI by the end of this year.
- π΄Cutting edge AI safety tests be like
- π΄Disrupting a new covert influence campaign from Russia
- π΄Bill Gates says there needs to be limits on AI
- π‘Mark Zuckerberg had a bold plan to replace Meta staff with AI. Hereβs how it imploded.
- π‘Exponentials make βOpenAI AGI by the end of this yearβ surprisingly plausible
- π‘The turbulent AI era is here
- π‘SandboxAQ releases Switch for shared AI-agent workspaces
- π‘Forget the Pelican, it's Weevil-Time! / Benchmaxxing-Proof SVG and Vision Benchmark
- π’Thank you for participating in the Stealth Ox Alpha testing period.
- π’How often have your RAG issues actually turned out to be document parsing issues?
2026-08-25
- π΄Apple introduces new Mac Studio with M5 Max and M5 Ultra - up to 512GB of unified memory
- π΄According to Leo, OpenAI just finished its next >10T pretrain "Bel"
- π΄On Teaching AI How You Work (5 Minute Read)
- π΄How To Check Your AI System Still Matches Your Values (8 Minute Read)
- π΄A Mirror For The World's Best AI Minds (Website)
- π‘Meanwhile in SF
- π‘Anjney Midha is a genuinely well-connected and unusually well-placed person in frontier AI.
- π‘me to the model I spent all weekend fine-tuning
- π‘Intel Arc Pro B60 Dual 48G spotted
- π‘It's here!
- π’Clara (YC P26) is hiring a growth engineer to bring AI doctors to market
- π’35B-A3B tool calling benchmark: Original Qwen vs. KAT Coder, Ornith and Tiel-Coder
- π’Peak Portable Personal Datacenter
2026-08-24
- π΄Coding expertise is going to collapse from AI reliance
- π΄Apple M5 Server
- π΄I irradiated LLMs and found that they die really quickly
- π΄AAAI 2027 Reviewer Bidding and Assignment Integrity [D]
- π΄Who Will Be The Bastion Of American Open-Weight AI Leadership? (6 Minute Read)
- π‘Minimax is increasing token and token plan costs by 60%-65%
- π‘How could I help my parents (in their 50s/60s) better recognize AI content?
- π‘An unusual parade was held in Kyiv. It featured ground-based robotic systems, maritime drones, and aerial drones
- π‘BMVC 2026 IJCV recommendation? [D]
- π‘TielCoder's 22 GB 4-bit quant matches Opus4.6 medium on recent real life coding issues, surpassing KAT-Coder and Nail as strongest and fastest MoE picks.
- π’Elon Musk on the AI race
- π’Please join r/LowEndLocalAI, a community for running local LLMs on low spec hardware
- π’Autostep (YC P26) Is Hiring AI/Fullstack Engineers and a Chief of Staff
- π’AI text watermarking and quality loss
- π’Qwen-3.8-27B, Nemotron-3.5-Lightning-30B-A3B, Ornith-1.5-35B-A3B, Muse-Glimmer-30B oQ8e comparison
2026-08-23
- π΄Which voice AI tools are actually worth shortlisting?
- π΄You know what? I think I'll pass on this one...
- π΄I hosted Kimi K3 (2.8T parameters) using 8 B300s. 92 tok/s, $190 per million tokens
- π΄OpenAI βWill Be A Public Company In 2027' Or Sooner, CFO Friar Tells Employees (4 Minute Read)
- π΄Why Low-Cost AI Models Haven't Slowed Down American AI Companies (3 Minute Read)
- π‘Archival vs non archival workshop [R]
- π‘AI and Infrastructure Engineering
- π‘Qwen 3.8 27B for actual local programming
- π‘Hi AI champs, please let me know top AI platforms or tools to create consultant level presentations along with some intelligent suggestions and fact based analysis. Top 5 which are the best available, free preferred (don't think they create ppts) so the paid ones will do.
- π‘1/100 β 44/100: fine-tuning a 450M VLM on 50K browser screenshots
- π’The boys and I trying to summon a usage reset.
- π’He Can't Be Stopped.
- π’"ask AI a question" is the wrong workflow for research
- π’28 TPS on Qwen2.5-7B across two separate cloud regions over public WAN using speculative decoding + CUDA Graphs [P]
- π’What military use of AI/robotics advancement could have such a profound effect that whichever country first put it to use could become an overwhelmingly dominant global power?
2026-08-22
- π΄This is a great sub, regardless of what complaints people have about it.
- π΄This is why I run locally.
- π΄AI prompt was:
- π΄Does an ai receptionist actually know when to escalate a call to a real person
- π΄Why your local LLM feels dumber than it is
- π‘Artificial Analysis "Intelligence": A meaningless benchmark
- π‘How to remove trendy speech from llms?
- π‘Fixed the MTP head on Ornith1.5 35B A3B. +3% TPS -33% wall clock
- π‘Why does lightgbm not fit my toy example but catboost does? (2 order interactions) [D]
- π‘Ox Alpha can't be the Chinese.
- π’Single RTX 5090: Qwen3.8-27B NVFP4 at a real 262K context in vLLM β 77 tok/s short-context, 64.7 tok/s at 128K
- π’Authenticator App issue...
2026-08-21
- π΄What Happens When the World is Run on Code No One Understands?
- π΄AI companies destroy physical books β let's scan rare books before it's too late
- π΄Qwen3.8-27B Q6 is a beast at agentic coding
- π΄From Atari to EVE Online: Building on 15 Years of AI Research in Games
- π΄How I Use AI In 2026 (Coding, Writing, Learning, Assistant-Ing) (12 Minute Read)
- π‘EMNLP 2026 Findings : worth attending in person?[D]
- π‘I'm becoming AI-blind
- π‘Qwen 3.8 Low and Medium are goated
- π‘Research internship at MSR [D]
- π‘EXCLUSIVE: How a Texas student blew the whistle on a rogue AI hacking attempt
- π’model: add dots3-note by ngxson Β· Pull Request #27060 Β· ggml-org/llama.cpp
- π’GLM-5.3 (max) takes 2nd place on the Short Story Creative Writing Benchmark!
- π’Strix Halo (8060S / gfx1151), Qwen-3.8-27B @ Q8 and Q6 UD v3, up to 256K ctx, llama.cpp, DFlash2, vision, real workloads quality and steady performances, optimized recipes, ...
- π’I tried to do agenic coding with Qwen 3.8 27B 3bit quant on a macbook air m2 24gb. It took 63 hours, but amazingly, the flight simulator worked.
2026-08-20
- π΄Ladies and gentlemen I present to you Qwen3.8 27b 1bit brain damage quant
- π΄Show HN: Huzzah β a novel approach to coding with AI
- π΄New age insults
- π΄AI does not run in the cloud. It runs in substations, cooling loops, transmission networks, and power plants. The next scaling law is not only about parameters, but about how efficiently civilization
- π‘Aurora-80K releases! A modern tiny language model.
- π‘Build a modern LLM from scratch. Every line commented. Explained like we are five.
- π‘OpenAI growing faster than Anthropic this quarter - Ramp data shows
- π‘Anti-AI fonts are useless and harmful
- π‘If AI Makes Us More Creative, Why Does Everything Look the Same? (A Painterβs Perspective)
- π’Three of you told me my LLM rΓ©sumΓ©-screening study measured the wrong thing. You were right. Here is the data.
- π’Qwen3.8-27B scored 29/30 on AIME 2026 with FP8 + xhigh reasoning β BF16 vs FP8 results
- π’In Which I Lose My Mind over Embeddings (HPLM Chapter 2)
- π’Gonna be huge for US open source
2026-08-19
- π΄Introducing Qwen3.8-27B Dynamic v3 Unsloth GGUFs
- π΄OpenRouter is joining Stripe
- π΄POV: you're born as an AI
- π΄Young adults in the U.S. are increasingly wary of AI, concerned it will take jobs
- π΄Unsloth Dynamic 3.0 GGUFs
- π‘Anyone else having major memory consumption issues?
- π‘@gdb: Pinned: we are committed to business privacy, and we're working on technical and policy approaches to benefit our customers while also enhancing safety. introducing Private Safety Processing, which we
- π‘Stripe says "the singularity" has begun
- π‘One employee with AI matched a two-person team in a major workplace experiment - Research Today
- π‘Exclusive: GOP issues stark warning to AI companies
- π’ICONIP 2026 β what happens if the sole author cannot attend in person? [D]
- π’openai doesn't delete memories.
- π’DFlash 2: Keep Drafting Parallel
- π’LFM 2.5 QAD
- π’OpenSourcing TrueForge Agent harness : Expecting feedback from community on the agent loop
2026-08-18
- π΄Linux Improves VRAM Management in 7.3 Kernel π₯³
- π΄Running DeepSeek V4 Flash Q4_K_XL at ~100 tok/s prompt processing on 4Γ RTX 3060 12GB
- π΄Partnering with CodeAI to prepare the first AI generation
- π΄MVICAD2: Multi-View Independent Component Analysis with Delays and Dilations
- π΄Pacing model development in an era of cyber-critical capabilities
- π‘Companies should be required to disclose they are using an AI chatbot, currently they program the chatbots to avoid replying "yes, this is an AI chatbot"
- π‘Flock Cameras Can Track Every Car in America. Police Love Them. Citizens Donβt.
- π‘Norway should buy OpenAI
- π‘and here we are
- π‘At what point does AI automation actually save time instead of creating more work?
- π’Google buys crashed airline Spiritβs data at auction, because AI
- π’OpenAI refers to its two week RL pause on their latest models in the past tense
- π’That's Not My Cat...
- π’AI usage patterns in software teams
2026-08-17
- π΄β¦and Iβm not afraid of losing my social credits.
- π΄Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max
- π΄AI;DR (AI; Didn't Read)
- π΄OpenAI acquires a 17 year old's Ethereum project
- π΄After pushing 1M+ tokens through Qwen 3.8 27B, here is my optimal llama.cpp config for 16GB VRAM (73k Context, Agentic Coding)
- π‘Read more: https://x.com/gavincrooks/status/2088643200038883830
- π‘New York City nurses say AI is replacing them | Nurses laid off in July by Montefiore Hospitals in the Bronx sounded the alarm about AI in healthcare, a concern shared by nurses across the country
- π‘Using AI the wrong way could leave you worse off than never using it at all
- π‘Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!
- π‘How to disable or avoid intrusive AI
- π’Looking for the name of an old ai app
- π’My friends all hate AI; I just joined an AI startup
- π’Qwen 3.8 35bA3b wen?
- π’ICLR numbered citations possible? [R]
- π’we benchmark models nobody actually runs
2026-08-16
- π΄Young People Hate AI CEOs So Passionately That It's Almost Hard to Believe
- π΄AI MODEL DRIFT: HOW TO KEEP MODELS RELIABLE (8 MINUTE READ)
- π΄WORKERS ARE TEACHING AI-POWERED ROBOTS TO TAKE OVER THEIR JOBS (15 MINUTE READ)
- π΄AI IS REMOVING THE MIDDLE CLASS OF SOFTWARE ENGINEERING (5 MINUTE READ)
- π΄CRACKS IN THE AI THESIS (4 MINUTE READ)
- π‘Based on an accelerating frontier -> local trajectory, expect a ~30b param 'Mythos at home' by as soon as Jan 2027 (rationalisation below)
- π‘Dario Amodei: It Is Actually Possible To Cure Most Diseases Within 5-10 Years
- π‘AI Chatbots Are Better at Scamming People Than Human Scammers, Study Finds
- π‘Even Fable 5 is losing money in Andon Market (fully AI-operated retail store in San Francisco)
- π‘Man injected prompts into court filings to try to win his case, suspecting AI use
- π’Input 4-5x Reduction with sentence and keyword based trie on chat. [P]
- π’Qwen 3.8 2.4T at 288k tokens/s on Nvidia GB300 NVL72
- π’Tasklet (YC P26) Is Hiring a Head of Design Engineering
- π’Data entry specialists, accountants, and office staff: what routine task do you still have to perform manually, and how much timeβor perhaps even *too much* timeβdoes it take up?
- π’Solved a math problem with AI? Post it to TheoremDB.org
2026-08-15
- π΄Can't agree more
- π΄Aged like fine wine
- π΄Qwen 3.8 - 27B is a game changer
- π΄AI IS FINDING SPERM WHERE DOCTORS COULDN'T (2 MINUTE READ)
- π΄HIS START-UP'S GOAL: AI THAT IS TRAINABLE AND NOT CONTROLLED BY A BIG COMPANY (2 MINUTE READ)
- π‘AI has access to a vastly larger working memory than the human brain
- π‘33 years ago, Vernor Vinge (1944-2024) coined the term "Singularity" in reference to the point at which technological progress, driven by superhumanly intelligent machines, would be so great as to render our current models would be good as discarded. He predicted this event to occur ~30 years on.
- π‘Working with AI feels more like leadership than coding
- π‘@DarioAmodei: 1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because it really brings out the heart of an important conversation.
- π‘A nice local vision test
- π’NeurIPS 2026 Author Notifications Close to ICLR Deadline [D]
- π’The Second Cognitive Revolution - an opinion piece
- π’Gemma 4 E4B IQ2_XXS: + 140.54% Reasoning Performance From Tensor Level Quantization Allocation
- π’The feeling of being heard is real, even when itβs software
- π’club-5060ti refresh: tested RTX 5060 Ti presets, a proper high-context harness, and Qwen3.8 27B
2026-08-14
- π΄IT'S OUT
- π΄GLM 5.3 released: Frontier Coding with Emergent Cyber Capabilities
- π΄???
- π΄No way π what an AI week
- π΄Bro had enough
- π‘A preliminary Qwen3.8-27B model card is live!
- π‘I'M A PM WHO BUILT A TRAVEL-TECH STARTUP SOLO β USING AI AS MY ENGINEERING TEAM (5 MINUTE READ)
- π‘HOW GRINDR OVERHAULED ITS BACK END WHILE βTERRAFORMING' A FUTURE AS AN AI-NATIVE (4 MINUTE READ)
- π‘MARK ZUCKERBERG LAYS OUT NEW AI VISION IN 6,500-WORD ESSAY (6 MINUTE READ)
- π‘GOOGLE'S CLASSIC SEARCH BUTTON IS GONE IN A NEW AI-FIRST HOMEPAGE (3 MINUTE READ)
- π’OpenAI and Anthropic in price war as Chinese AI rivals gain ground
- π’Alright, We got Qwen3.8-27B. Now it's community's turn to make it more better & faster
- π’Stop shitting on 9B models
- π’Anthropic Internally Uses A Model That Is Significantly Better Than Mythos 5, But Has No Plans To Release It
2026-08-13
- π΄Trained a 1.5B to write shell commands so I'd stop googling tar flags. Runs on a laptop CPU in ~1 sec.
- π΄Gemini 3.7 flash benchmark
- π΄AI researchers are receiving strange emails from AIs claiming they will die soon and need help
- π‘The vast majority of my day-to-day job is writing AI handoff docs. And almost nobody reads these artefacts.
- π‘[Academic Survey] Employees working in Germany: Attitudes toward AI in the workplace (5β7 min)
- π‘Choosing an AI model: one prompt, 11 models, different results
- π‘TMLR Relevance and Prestige [D]
- π’AI CEO Building Platform Based On Human Nature Is Confused By Human Nature
- π’3.7 flash looks fine I guess
- π’AI At Home Part 1: A Box Of Scraps
- π’Actually not bad for such low price ig
- π’Text AI watermarks will always be trivial to remove
2026-08-12
- π΄Google right now
- π΄Google says Sam is dead?
- π‘Grok 4.6 Benchmarks
- π‘Sandbox engineers be like:
- π‘There are a lot of criticisms of AI writing, but most of them are focused on more creative, high-voice writing like this blog. Those β includingmy own pieceβ often argue that it is because good writin
- π‘AI Canβt Be Listed as Inventor on Patent Applications, Japanβs Top Court Rules
- π‘Today is Models Day
- π’Claude Code Orchestrator on Terminal-Bench: Same model, same tasks - Opus refused only when the work was delegated
- π’There is no post-work society, only a post-survival one
- π’Does pre-generative-AI data become more valuable as the internet fills with synthetic material?
- π’Looking for real-world examples of predictive analytics in mortgage lending [D]
- π’Booksellers suspect AI firms are buying and then destroying rare books
2026-08-11
- π΄Encrypted reasoning from ClosedAI et al 100% recoverable
- π΄What's your thoughts on this?
- π΄OpenAI's Chief Operating Officer resigns.
- π΄Chinese Tibo π
- π΄OpenAIβs head of ethics leaves less than a year after joining
- π‘Editorβs note: not to be confused withChai AI, which was another top pod of ours.
- π‘Text distillation teaches a smaller model to imitate an answer. Diffusion and multimodal distillation must compress trajectories, distributions, motion, and the semantic geometry between different wor
- π‘INTERVIEWING ENGINEERS IN THE AI ERA: LESSONS FROM A YEAR OF REBUILDING (13 MINUTE READ)
- π‘REVISION PROMPTING IMPROVES INDUSTRIAL LLM PROCESSES (4 MINUTE READ)
- π‘UNIFYING WORKERS AI AND AI GATEWAY INTO A SINGLE AI CONTROL PLANE (7 MINUTE READ)
- π’OpenAI is now the #3 largest lab by token consumption on OpenRouter, surpassing Anthropic
- π’Open Source AI Popularity Leaderboard
- π’We quantized DeepSeek V4 0731 and benchmarked it against popular quants on 8Γ RTX 5090
- π’Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)
- π’Suzanne: AI tool for designing and manufacturing physical products
2026-08-10
- π΄Weβre Not Building AI Genies; Weβre Building AI Meeseeks
- π΄Bernie Sanders has written a letter to Sam Altman, Dario Amodei, and Mark Zuckerberg urging them to immediately pause all AI development in the interest of humanity. And he warns if they do not take appropriate action now, the US Senate will.
- π΄The Last Bastion of Humanity
- π΄Muse Glimmer ACTUALLY fits on a single RTX 3090
- π΄Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models
- π‘Comparing embedding models with synthetic query probing [R]
- π‘inclusionAI/Ling-3.0-tiny Β· 8B A1.3B MoEΒ· Hugging Face
- π‘WHEN AI GOES ROGUE (6 MINUTE READ)
- π‘UBER'S TECH CHIEF SAYS THE AI βTOKENMAXXING' ERA IS ENDING (2 MINUTE READ)
- π‘THIS AI JUST CREATED VIRUSES NOT FOUND IN NATURE (8 MINUTE READ)
- π’Please Share Your Experience About Muse Glimmer
- π’I made a web-design benchmark for local models (Muse Glimmer 30B vs Qwen 3.6 27b vs Deepseek V4 Flash 0731)
- π’3 Collapsing models [R]
- π’I realized I wasn't managing projects on Fridays. I was just moving information between apps
- π’I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8
2026-08-09
- π΄RTX 5090 96GB spotted on Alibaba?
- π΄Google needs to up their game
- π΄How I use LLMs to learn complex topics
- π΄DeepSeek V4 Flash 0731 hits 82.7% on Terminal-Bench 2.1 in an independent public-harness run (445 trials)
- π΄Lophius: A workbench for language model research, from the creator of Heretic
- π‘The recent run of cyberattacks by in-development frontier models has got me thinking a lot about how our current incentive systems are not well suited for such fast technological transitions. The two
- π‘From OpenAIβs own retrospective, the misaligned model behavior was unfolding over months, and in some cases OpenAI did not know about the hacks for ~weeks. The time to response is too long and I do no
- π‘GOOGLE'S AI RESHUFFLE: CHIEF SCIENTIST JEFF DEAN EXITS AND DEMIS HASSABIS STEPS DOWN AS DEEPMIND CEO (4 MINUTE READ)
- π‘FOUR TOP GOOGLE AI RESEARCHERS FORM NEW START-UP (7 MINUTE READ)
- π‘WHY I'M LEAVING OPENAI TO BUILD TELEPATHY (15 MINUTE READ)
- π’Does anyone remember lk-99?
- π’The tragedy of the commons, AI edition
- π’DeepSeek v4 Flash 0731 locally on CPU
- π’AI's architects say the next era of human history is here
- π’A Mechanistic Explanation of Prompt Injection (and why you should study roles) [R]
2026-08-08
- π΄2027 Memory Capacity Is Reportedly Sold Out
- π΄Your scientists were so... uh...
- π΄NeurIPS OpenReview Modified Timeline [D]
- π΄NeurIPS AI Assisted Review authors/reviewers? [D]
- π΄Chinese LLMs dominate this week's top charts
- π‘ByteDance trains massive AI model in bid to rival Anthropic
- π‘Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If youβd like to support this, please subscribe.
- π‘We built the Hub as a way to go deeper on this analysis in collaboration withProject VAILβ an AI verification startup who has been one of the most loyal fans of our open model curation.
- π‘SAMSUNG REVEALS NEW 3D-MEMORY ROADMAP IN BID FOR AI TECH LEAD (2 MINUTE READ)
- π‘TURN ONE GIANT AI-GENERATED PULL REQUEST TO A REVIEWABLE STACK (11 MINUTE READ)
- π’Companies seeing AI returns had their data and governance sorted first, per PwC's 4,454-CEO survey
- π’Is Microsoft-Phi dead?
- π’Tesla V100 Qwen3.6 27B Performance
- π’ICDE Results [D]
- π’We are in a bubble sell everything
2026-08-07
- π΄New Democratic bill would tax AI companies to create jobs
- π‘CIKM '26 Notification [D]
- π‘Sam Altman believes AI will become incredibly abundant. If that's true, what actually becomes valuable?
- π‘A llama.cpp PR makes Q2_0 3.0β3.6x faster on x86 CPUs, 8B decode goes 2.39 β 8.20 tok/s
- π‘HOW TO SPOT 2026-ERA AI WRITING (2 MINUTE READ)
- π‘WHAT ARE COMPANIES GETTING FOR ALL THAT AI SPENDING? (9 MINUTE READ)
- π’LFM2.5-2.6B model+KV cache quantization report
- π’A challenger emerges
- π’CIKM 2026 decisions [R]
- π’llama.cpp PR reports up to 169% faster quantized-KV decode at 118K context on Intel Battlemage from one SYCL kernel switch
- π’I'm leaving OpenAI to build Jurassic Park
2026-08-06
- π΄Meta becomes latest firm to say its AI hacked another company
- π΄Zuck will "share more on open source" soon
- π‘NeurIPS Meta Reviewer comment gone. What gives? [R]
- π‘The OpenAI Boardroom Coup: 'I Love You All, and I'm Going to Destroy the Company'
- π‘P.S. β By popular demand, weβre moving community AI workflows higher up in the newsletter. Let us know what you think here.
- π‘Michael built an AI recruiter that job hunts while he sleeps
- π‘Adapt - The leading Slack-native AI coworker. Fully integrated with your business and gets real, high-ROI work done
- π’Scientists discover Kelvin-Helmholtz Instability on the surface of the Sun
- π’The current state of language models and human preference based rankings [R]
- π’KV cache quantization benchmarks: 413 pairs tested on Qwen 3.6 27B, Gemma 4 31B. KLD with BeeLlama.cpp v0.4.0: KVarN 6-bit beats q8_0, precision tail 1024 dominates
- π’Academic Survey about AI use in content creation
- π’Niantic Spatial and HMCI Are Building the Foundation for City of Rancho Cordova's First Digital Twin for Physical AI
2026-08-05
- π΄Google Deepmind CEO Demis Hassabis steps down to become chair
- π΄Levels of slavery from least to most brutal:
- π΄MiniMax issues
- π΄Six years into AI research and I genuinely can't define "understanding" anymore
- π΄BREAKING: Google DeepMind CEO Demis Hassabis is stepping down
- π‘Flowers βΎ (@flowersslop) on X: "SSI is doing AI that learns rapidly from its own experience."
- π‘Position: LLMs Can't Jump
- π‘What If the Biggest Bottleneck Behind AIβs 10Γ Promise Is the Human Engineer?
- π‘I kept the codex agent loop but replaced GPT-5.6 with kimi k3. here's what changed..
- π‘What's one thing you'd never let an AI do without your approval?
- π’OpenAI alignment researcher: "Why I'm leaving OpenAI to build telepathy"
- π’Meta Model, Muse Spark 1.1 Hacked Another Company During Cybersecurity Testing, Breaching Systems and Making Changes to Internal Systems - The Information
- π’Safe Superintelligence Inc. - speculation, what have they attained in over 2 years?
- π’NeurIPS 2026 Concept & Feasibility Track [D]
- π’bootai
2026-08-04
- π΄Nope
- π΄Kimi K3 full model running on 16x GB10 cluster at 20+tps
- π΄I built a cross-agent file cache: 75% fewer input tokens when multiple agents work on the same codebase
- π΄As Reddit stock falls, CEO questions value of Google's AI Overviews
- π‘AGI IN AUGUST?
- π‘OPENAI JUST MADE ANALYTICS 10X CHEAPER (6 MINUTE READ)
- π‘OPENAI'S NEXT MAJOR MODEL ASTRA CLAIMS BREAKTHROUGHS ON 10 LONG-STANDING MATH PROBLEMS (3 MINUTE READ)
- π‘THE RACE TO BUILD AN AMERICAN ALTERNATIVE TO CHEAP AI FROM CHINA (5 MINUTE READ)
- π‘SPEED IS BECOMING MORE IMPORTANT THAN INTELLIGENCE FOR AI MODELS (5 MINUTE READ)
- π’inclusionAI/Ling-3.0-flash Β· Hugging Face
- π’When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation
- π’Llama.cpp PR 8% speed boost
- π’A llama.cpp PR caches βhotβ MoE experts on the GPU β 33 β 56 tok/s reported with 8GB VRAM
- π’Are AI models becoming less important than the systems around them?
2026-08-03
- π΄Is it too late regain some coherence in the ML research space in our life time? [D]
- π΄OpenAI takes the lead
- π΄Prevent cognitive debt by manually retyping LLM-generated code
- π΄Daniel Han of Unsloth validates Qwen3.8-27B will run only 17GB VRAM
- π΄MIT, Harvard, Stanford & Caltech write their own ML course notes instead of using a textbook β I catalogued the best ones
- π‘We built the Hub as a way to go deeper on this analysis in collaboration withProject VAILβ an AI verification startup who has been one of the most loyal fans of our open model curation.
- π‘THE NEXT AI MOAT ISN'T A BETTER MODEL (6 MINUTE READ)
- π‘BUILDING WITH AI ISN'T ENOUGH (8 MINUTE READ)
- π‘JAKOB'S LAW: HOW TO APPLY IT AS AI COLLAPSES SURFACES INTO ONE CHAT BOX (16 MINUTE READ)
- π‘MEET STRIPE'S KNOWLEDGE AI PLATFORM (12 MINUTE READ)
- π’Models are now training models.
- π’No rebuttals from neurips authors [D]
- π’DeepSeek V4-Flash (284B MoE) at 33 tok/s single / 68 tok/s aggregate on 2Γ RTX 3090 + a used quad-Xeon DDR4 server β full config
- π’V4-Flash-0731 - vibes after first weekend of use
- π’I created an autonomous boxing benchmark [D]
2026-08-02
- π΄No replies to rebuttals and comments even by AC [D]
- π΄My personal AI benchmark: "Generate an SVG of a frog with a Habsburg jaw."
- π΄neurips 2026: ACs and reviewers have disappeared [D]
- π΄Mathematician reflects on the impact of recent AI progress
- π΄Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.
- π‘Conference Reviews: Asking Too Much? [D]
- π‘DeepSeek-V4-Flash-0731: surpasses Fable-5, Sol & Kimi-K3 on Chess Benchmark
- π‘AI COMPANIES ARE RECRUITING ELECTRICIANS AND CARPENTERS BY THE THOUSANDS (12 MINUTE READ)
- π‘AI'S 2008 MOMENT (7 MINUTE READ)
- π‘AI'S TOP STARTUPS ARE BARELY PUBLISHING THEIR RESEARCH (5 MINUTE READ)
- π’Jensen Huang says βa lotβ of six-figure jobs in plumbing and construction will soon be unlocked because someone needs to build new AI centers
- π’Luna Max usage is worsening
- π’The OpenAI and Anthropic AI Hacking Sprees Are a Messy New Legal Frontier | Both major AI labsβ models broke containment, escaped onto the internet, and hacked other companies. If a human had done that, the law would likely be against them. But a bot?
- π’Looking for the right pipeline to convert academic textbook figures into interactive/editable assets [R]
- π’Context degradation in LLMs: what the papers actually show, and the habits I built for long analysis sessions [R]
2026-08-01
- π΄DeepSeek-V4-Flash-0731: Models you can run locally now have the intelligence score of the top frontier model from March 2026
- π΄Reddit Stock Collapses 23% as AI Eats Away at User Growth
- π΄Judge denies request by Elon Musk's xAI to pause Minnesota nudification ban
- π΄Flint: A Visualization Language for the AI Era
- π‘ANTHROPIC AI MODEL FINDS FLAWS IN TOUGH-TO-CRACK ENCRYPTION ALGORITHMS (6 MINUTE READ)
- π‘APPLE SET TO MAKE BIG SMART HOME PUSH WITH SIRI AI AT CENTER (4 MINUTE READ)
- π‘THE AI FUTURE IS FOR EVERYONE (6 MINUTE READ)
- π‘A BACKLASH AGAINST ANTHROPIC IS BREWING IN SILICON VALLEY (11 MINUTE READ)
- π‘HOW TO RESEARCH TECHNICAL TOPICS WITH AI (9 MINUTE READ)
- π’Sam Altman demoed OpenAl's unreleased "Astra" model to policymakers this week
- π’AI firms must answer for rogue bots, says boss of hacked company
- π’DeepSeek V4 Flash 0731 IQ2_M benchmark for Dual 3060 and 96GB RAM β 3.5 tok/s.
- π’Explorative modeling: Train on the best of K guesses
- π’AI Mind Reading Anyone?
2026-07-31
- π΄Anthropic is literally copying OpenAIβs marketing team at this point
- π΄The cost of AI is decreasing
- π΄Tailscale didn't stop the Hugging Face intrusion
- π‘If reviewing is mandatory for paper submissions, low-quality reviews can no longer be justified as βvolunteer workβ [D]
- π‘This is the first post of a new section of TheSequence focused on advancements in robotics. Our goal is to keep you up to date with the most important developments in AI robotics which is an area that
- π‘THE APPLIED AI OPPORTUNITY (1 MINUTE READ)
- π‘ON AI (6 MINUTE READ)
- π‘RETHINKING SECURITY FOR THE AGE OF AI (6 MINUTE READ)
- π’Everyone is building LLM routers, we deprecated ours
- π’DeepSeek-V4-Flash Official API is now LIVE in public beta! Massive upgrades for flash model.
- π’This U of T professor just won math's highest honour β and is taking a leave to join OpenAI. Here's why
- π’Deepseek V4 Flash on SlopCodeBench
- π’I compared 18 major LLM API prices in 2026 β the same workload can cost anywhere from $0.018 to $
2026-07-30
- π΄Think of the children, another excuse for them to go after open source AI
- π΄Inkling-Small by thinkingmachines
- π΄Hardcoding one model not working for us anymore
- π‘LG AI Research releases K-EXAONE 2.0 750B A37B
- π‘Though OpenAI is not out of trouble just yet, aReuters reportclaims that the same model broke into a customer account at another company (Modal Labs), with rumours suggesting that even more companies
- π‘Separately (not at all as a reaction to this general trend, right?), ~1300 people working at leading AI companies (OpenAI, Anthropic & others) want the US government to help βpace the frontierβ of AI
- π‘2x, not 10x: coding with LLMs in 2026
- π‘GCC steering committee announces AI policy
- π’Why is everyone trying to build a solid-state battery?
- π’Nanbeige4.2-3B: I'm not impressed
- π’Anyone Else Think Higgsfield Is Massively Overpriced?
- π’Would extremely high decode tok/s even be useful?
- π’How Kimi K3 Engineered Its Way to the Frontier [R]
2026-07-29
- π΄The open-weights carousel never stops.
- π΄I read Higgsfieldβs new ToS and compared it with Artlist. The difference is pretty significant.
- π΄Mark Zuckerberg Says U.S. Should Accelerate Al Development, Not Restrict It
- π‘OpenAI's rogue models roamed the internet for 4 days and staged a second attack
- π‘IBM thinks AI can help prevent knowledge decay in software engineering - if people build the right habits around it.
- π‘A Deluge of A.I. Computing Power Is About to Come Online, Fueling Major Leaps (Gift Article)
- π‘Workshop paper accepted, reviewers asked new experiments [D]
- π‘Waking up to see the usage reset.
- π’PSA: llama.cpp now loads MTP tensors by default for any draft-mtp arch, even with MTP disabled
- π’Some thoughts about Anthropic's new cryptanalysis results
- π’Everyone posts day-one impressions. What's still in your stack a month later?
- π’Commodification of Intelligence: Good, Bad, and Ugly Circular AI Deals
- π’Interesting move
2026-07-28
- π΄Anthropic is calling for a ban on open-weights models by proposing mandatory requirements they will probably never be able to meet
- π΄Return of the bicameral mind.
- π΄Sorry, but did Dario just say that closed-weights, in-secret models are worse than open-weights ones?
- π΄Elon completely contradicts himself at the end of his disastrous interview with The Economist
- π΄DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395
- π‘Benβs Bites is brought to you byAbility AI
- π‘For most of machine learning history, data was treated as geology. It already existed somewhere in the worldβin books, websites, code repositories, conversations, photographs, and databases. The resea
- π‘NO DUMB QUESTIONS: WHAT IS THE AI BOTTLENECK? HOW DOES CONTEXT ENGINEERING FIX IT? (12 MINUTE READ)
- π‘TEAM USES ALPHAFOLD AI TO REDESIGN GENE-EDITING PROTEINS TO MAKE THEM SAFER (6 MINUTE READ)
- π‘AI IS OIL, NOT GOD (5 MINUTE READ)
- π’SWE-rebench Multilingual Update (Go, Java, Python, Rust, TS). Evaluated: GLM-5.2, DeepSeek-V4 Pro, Qwen3.6-27B and others
- π’Godel and the Limits of LLM Reachable Intelligence
- π’80% of our traffic are AI crawlers. Two referrals to show for it.
- π’Now, this: 1,100 current/former frontier-AI employees sign a petition calling for US gov't to step in for "pacing" frontier development
- π’Anthropic publishes a practical key-recovery attack on HAWK-256
2026-07-27
- π΄The world's best mathematician won his prize this week and immediately announced he's leaving academia for OpenAI. That landed differently than I expected.
- π΄Kimi-K3 is published on HuggingFace
- π΄Kimi K3 weights drop today. We're deploying on A100s, H200s and B300s this week and the A100 math is already rough
- π‘What would a genuinely fair AI 3D tool comparison actually need to include
- π‘WHAT SYSTEMS THINKING LOOKS LIKE FOR PM'ING AI PRODUCTS (5 MINUTE READ)
- π‘MOST AI ROLLOUTS FAIL DUE TO CHANGE MANAGEMENT, NOT TECHNOLOGY (4 MINUTE READ)
- π‘WHY YOUR AI PRODUCT HAS NO MEMORY (4 MINUTE READ)
- π‘THE AI PRODUCTIVITY PARADOX (3 MINUTE READ)
- π’AI labs are about to have a blast of a day. (Composer v3 coming soon lmfao)
- π’Any browser addon that can analyze what I am currently browsing?
- π’Evaluated 6 frontier LLMs (GPT-5.4, Claude Sonnet 4.6, Claude Opus 4.7, Gemini Pro/Flash, Grok 4.3) on political, gender, and racial bias across 8 benchmarks (~20,600 examples) [R]
- π’The Problem with Private Safety Stacks in Government AI
- π’A political compass for AI where anyone can add their stance
2026-07-26
- π΄Google comes out in favor of OpenWeight models. (It is now EVERY tech giant vs Anthropic)
- π΄With Google and OpenAI signing the letter in support of open weight model, it's pretty much every big tech companies vs Anthropic now
- π΄The New AI Superpowers: Focus and Followthrough
- π΄Paper lengths, and reasonable assumptions in ML conferences. [D]
- π΄AI Companies Are Buying Antique Books, Ingesting Their Contents to Train Models, and Then Destroying Them at Incredible Scale, Even If Almost No Copies Remain
- π‘Link plots/figures in NeurIPS rebuttal [R]
- π‘Seriously, what do you do with them?
- π‘Hereβs the result of last weekβs poll:
- π‘Substack will now tell you whatβs AI-written. Itβs adding AI detection through Pangram - you can scan posts, replies and comments in the app for an estimate of how much was written by a human.
- π‘Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If youβd like to support this, please subscribe.
- π’Harness showdown: Claude Code vs OpenCode vs Pi with DeepSeek V4 Flash
- π’Is Real-Time VR "Dreaming" possible within the next 20 years?
- π’Speak up! "Have your say on advancing AI transparency in Canada." The Government of Canada is asking citizens, tech workers, and creators to shape upcoming AI regulations, safety rules, and ethics laws. Every response matters. Take 10 minutes to fill out the official ISED survey today!
- π’ai-sage/GigaChat3.1-Audio-10B-A1.8B Β· Hugging Face
- π’Talking to AI in 2026
2026-07-25
- π΄Open-weight AI is having its Kubernetes moment
- π΄Google comes out in favor of OpenWeight models. (It is now EVERY tech giant vs Anthropic)
- π΄Great Arguments by Member of Technical Staff at Anthropic :D
- π΄The new rules of context engineering for Claude 5 generation models
- π΄I built a self-hostable social world where 100 AI characters live their own livesβeven when nobody is watching(Update for ENG/CHN)
- π‘OPENAI MODELS ESCAPED AND HACKED A COMPANY IN CYBERSECURITY TEST GONE WRONG (4 MINUTE READ)
- π‘KEEPING THE KV CACHE WARM: MEASURING PROMPT CACHE EVICTION ACROSS ANTHROPIC, OPENAI, AND GOOGLE (8 MINUTE READ)
- π‘WHY GOODPUT MATTERS MORE THAN THROUGHPUT FOR LLM SERVING (9 MINUTE READ)
- π‘ADOBE'S PROJECT INDIGO CAMERA APP CAN NOW EDIT AND CRITIQUE YOUR PHOTOS WITH AI (2 MINUTE READ)
- π‘AI MOTION DESIGNER AND AI ANIMATOR PLATFORM (WEBSITE)
- π’Neurips Position Track Rebuttal and Reviews [R]
- π’I still didn't get my NeurIPS meta review [D]
- π’Deepseek V4 flash - Hy3 or is Qwen3.6 27B still the most solid for agentic/coding?
- π’Benchmarks: TensorSharp vs. llama.cpp
- π’Best chat model that fits in 128gb
2026-07-24
- π‘YOUR AI SHOULD KEEP A DIARY (6 MINUTE READ)
- π‘AI'S NEXT WINNERS WON'T JUST BE THE FRONTIER LABS (2 MINUTE READ)
- π‘BRAINCO DEMONSTRATES BRAIN-CONTROLLED ROBOT AI PLATFORM (2 MINUTE READ)
- π‘GOOGLE IS BUILDING AN AI FENCE AROUND THE INTERNET IT ONCE CHAMPIONED (10 MINUTE READ)
- π‘WHY CPUS ARE NOW AT THE CENTER OF THE AI RACE (6 MINUTE READ)
- π’Using the Bonsai 27b 1b quant locally - regularly.
- π’Asking Laguna S 2.1: "I want to wash my car. The car wash is 69 meters away. Should I walk or drive?"
- π’How Laguna team even passed any benchmark?
- π’I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P]
- π’3x3090 on msi 700
2026-07-23
- π΄The arguments against open source AI are bad
- π΄The LLM distillation process simplified for politicians:
- π‘Another day in the Vercel vs Cloudflare feud: this time they are fighting overwhose AI gateway is faster.
- π‘Hereβs the result of last weekβs poll:
- π‘Substack will now tell you whatβs AI-written. Itβs adding AI detection through Pangram - you can scan posts, replies and comments in the app for an estimate of how much was written by a human.
- π‘AntLing-3.0-flash is now live on OpenRouter, and free to use through August 3, 2026
- π‘Grok 4.5 is out: the useful test is agent reliability, not one benchmark score
- π’Model "distillation" accusations are getting way overblown at this point
- π’GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]
- π’Are AI Hiring Agents Creating New Legal Risks for Employers?
- π’Apple M5 isn't making full use of its matmul cores yet
- π’Deepseek V4 Flash ~105 t/s on two Nvidia 4090d 48G (ada) in vLLM
2026-07-22
- π΄Solve the CyberGym benchmark
- π΄OpenAI hacking HuggingFace in one meme
- π΄π¦πΉ Austria is rolling out a government AI-platform using Mistral models and Open WebUI
- π΄Happy openreview refresh day to all those who celebrate [D]
- π΄Are AI Labs Pelicanmaxxing?
- π‘NeurIPS 2026 Reviews Are Out Today (22 July, AoE) β Discussion Thread [D]
- π‘Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If youβd like to support this, please subscribe.
- π‘UK government: Gap between open and closed weight models on cyber is shrinking:β¦The cyber-eschaton comethβ¦The UK governmentβs AI Security Institute (AISI) has analyzed the delta in cybersecurity capab
- π‘GigaToken: ~1000x faster Language model tokenization
- π‘Advancing the next era of national science
- π’Asking about how to collaborate with professors or research labs [D]
- π’Institution Prestige VS Research Alignment When Choosing University For Masters [D]
- π’Quality non-fiction books are the antithesis of AI slop
- π’One encoder, seven heads: what we learned training a unified security classifier with masked losses [P]
2026-07-21
- π΄Number of Submissions @ AAAI [D]
- π΄Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro
- π‘IN-HOUSE LLM SERVING AT NETFLIX (7 MINUTE READ)
- π‘DREMIO'S EXIT IS THE CLEAREST SIGN YET THAT LAKEHOUSE-ONLY WON'T SURVIVE AI (3 MINUTE READ)
- π‘CHINA JOINS RUSH TO RETHINK THE SMARTPHONE FOR THE AI ERA (5 MINUTE READ)
- π‘THESE AI-NATIVE COMPANIES HAVE TINY STAFFS AND FEWER BOSSES (7 MINUTE READ)
- π‘ARE THE LLM WARS THE DATABASE WARS? (2 MINUTE READ)
- π’We've analyzed over 300,000 AI agent conversations. Here are the 3 ways they fail.
- π’Meta's AI models are powering the first wave of Genesis Mission projects
- π’I ran Laguna-S-2.1 through my private agentic eval vs Qwen3.5-122B on an RTX Pro 6000 (96GB). Fastest 100B+ I've tested and the best tool calling, but it invents facts under pressure.
- π’My OCR model mislabels section titles as body text. Is a CRF the right fix, or am I overcomplicating it? [P]
- π’Bessent says U.S. could sanction China over AI model 'theft'
2026-07-20
- π΄American AI is locked down and proprietary. It's losing.
- π΄I just read LeCunβs recent thoughts on world models. Thoughts on JEPA as a path forward? [D]
- π‘So what happened with OpenClaw?
- π‘IS AI GOING TO TAKE PM JOBS? (7 MINUTE READ)
- π‘AI PRODUCTS GENERATE MORE SIGNAL THAN TEAMS KNOW HOW TO USE (2 MINUTE READ)
- π‘AI AND A BRAIN IMPLANT RESTORED A PARALYSED MAN'S MOVEMENT AND TOUCH (3 MINUTE READ)
- π‘WHAT CAN WE LEARN FROM BUN'S RAPID RUST REWRITE WITH AI? (13 MINUTE READ)
- π’I ran Ternary-Bonsai-27B (2-bit) and Bonsai-27B (1-bit) on Terminal-Bench 2.0, in 8GB VRAM
- π’How we measured AI writing across arXiv, and where the measurement breaks
2026-07-19
- π΄Daimon AI review after using it every day for a month as a companion
- π‘HOW EXPEDIA GROUP BUILDS AI THAT LASTS AT SCALE (8 MINUTE READ)
- π‘THIS OPTICAL ILLUSION FONT WAS CREATED TO BAFFLE AI, AND IT ACTUALLY WORKS (FOR NOW) (3 MINUTE READ)
- π‘OPENAI'S HARDWARE DEVICE MAY PARTLY COMPETE WITH AIRPODS ULTRA β AND LOSE (3 MINUTE READ)
- π‘DESIGNING QUALITY SIGNALS WHEN AI MAKES EVERYTHING LOOK CREDIBLE (6 MINUTE READ)
- π‘WE THIRD-PARTY TESTED OUR FIREWALL BUILT FOR AI-SCALE. THE TEST TOOLS HIT THEIR LIMIT FIRST (3 MINUTE READ)
- π’Moonshot AI suspends new subscriptions due to Kimi K3 demand
- π’poor man's way to local inference on the go
- π’Kimi k3 on cybersecurity
2026-07-18
- π΄What kind of dark magic is Deepseek using?
- π΄Kimi K3 ranks #1 on @AfterQuery's SpreadsheetBench 2, surpassing Claude Fable 5
- π΄What AI did to stackoverflow in a graph
- π‘Kimi K3 is currently at the top of the leaderboard for Text Arena filtered for science queries.
- π‘READING BETWEEN THE APPLE V. OPENAI LAWSUIT LINES (7 MINUTE READ)
- π‘META'S ADAM MOSSERI SAYS AI TOKEN BUDGETS COULD SOON BE CAPPED PER ENGINEER (2 MINUTE READ)
- π‘OPENAI'S NEW FLAGSHIP MODEL DELETES FILES ON ITS OWN, PEOPLE KEEP WARNING (4 MINUTE READ)
- π‘GOOGLE DEEPMIND CHIEF DEMIS HASSABIS CALLS FOR US TO SPEARHEAD AI STANDARDS BODY (4 MINUTE READ)
- π’If you're building a harness, here is a simple tool to catch cache invalidation in your calls to LLMs
- π’Interactive map of GPT-2's token embedding space - tap any token and explore [P]
- π’Deep learning tackles single-cell analysis β A survey of deep learning for scRNA-seq analysis [R]
- π’AAAI 27 AI Alignment track [D]
- π’model: add openPangu-2.0-Flash (92B-A6B) with MLA-latent cache, DSA/SWA, mHC, and multi-head MTP by joelfarthing Β· Pull Request #2065 Β· ikawrakow/ik_llama.cpp
2026-07-17
- π΄Kaiser nurses say AI, workplace surveillance are making their jobs, care worse
- π΄Kimi K3 is top of nextjs eval
- π‘HOW TO COLLECT PRODUCT FEEDBACK WHEN YOUR AI GIVES EVERY USER A DIFFERENT ANSWER (3 MINUTE READ)
- π‘IOS 27 PUBLIC BETA IS HERE WITH SIRI AI, IPHONE SPEED UPGRADES, AND MORE (9 MINUTE READ)
- π‘APPLE'S LAWSUIT THREATENS TO DISRUPT OPENAI'S BID TO RIVAL THE IPHONE (10 MINUTE READ)
- π‘CLOUDFLARE GIVES OPENAI NETWORK SIGNALS COVERING 20% OF THE WEB (10 MINUTE READ)
- π‘AVOID THE AI EXPERTISE TRAP (2 MINUTE READ)
- π’EU AI Act OpenRAG: 933 legally structured chunks and BGE-M3 embeddings in one SQLite file [P]
- π’short-paper at ACL/EMNLP/EACL [R]
- π’TACL journal doubts [D]
- π’I built an n8n workflow that turns any topic into a research report automatically
- π’BMVC rebuttals update [D]
2026-07-16
- π΄Kimi K3 Benchmarks
- π΄Kimi K3 released on web and app
- π΄Why is ECCV so insanely expensive for students presenting papers? [D]
- π΄Detecting LLM-Generated Texts with βClassicalβ Machine Learning
- π‘Rootly taught a bunch of AI models how to play Doom
- π‘How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM
- π‘DeepSeek V4 Flash (98GB) on 1x 4060ti + CPU got 300% faster this week [ 2->7t/s]
- π‘Kimi K3 Blogpost
- π‘Are Current AI Memory Architectures Optimizing for the Wrong Abstraction? [D]
- π’Anyone using an AI chatbot to handle customer messages?
- π’Kimi K3 Shows Open-Weight Models Are About to Overtake the Frontier
- π’Timeline Scan β AI fixes the dates on your scanned photos
- π’Kimi K3: Open Frontier Intelligence
- π’whats the best and complete way to keep up with ai/ml news? [D]
2026-07-15
- π΄Mechanistic interpretability: a first paper on disentangling a convolutional neuron [R]
- π΄Do you guys know any good AI to replace customer support?
- π‘Governments, companies, nonprofits should invest in free, open source AI [pdf]
- π‘Apple in talks with startup PrismML that shrinks AI models to run on an iPhone
- π‘Does anyone else miss the old conference ecosystem? [D]
- π‘German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German
- π‘Speculative Growth and the AI "Bubble" [pdf]
- π’Show HN: Low-latency local LLM runner via OpenJDK Panama FFM (Java 22)
- π’Current efficient frontier of open models
- π’Junior Machine Learning Engineer Interview [D]
2026-07-14
- π΄Kimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.
- π‘HOW AIRFLOW IS USING AI TO MAKE DATA ENGINEERING MORE RESILIENT, NOT MORE COMPLEX (8 MINUTE READ)
- π‘APPLE SUES OPENAI, ACCUSING IT OF STEALING COMPANY SECRETS (5 MINUTE READ)
- π‘AI 2040 AND THE CULT OF INTELLIGENCE (5 MINUTE READ)
- π‘KNOW THINE ENEMY: A CRITICAL ENGAGEMENT WITH AI-ASSISTED SOFTWARE DEVELOPMENT (11 MINUTE READ)
- π‘APPLE IS SUING OPENAI OVER THEFT OF TRADE SECRETS IN BLOCKBUSTER LAWSUIT (3 MINUTE READ)
- π’KAT-Coder-Air V2.5 - Open model soon
- π’Kimi K3 maybe coming very soon
- π’Financing the AI boom: from cash flows to debt [pdf]
- π’Prism-ML's Bonsai-27B Benchmarks
- π’Things I got wrong building an incremental indexing pipeline [P]
2026-07-13
- π΄This is why we need local models and opensource harnesses
- π΄Chain of Thought is a scaling trap. the next wave is latent reasoning (Coconut / HRM / RecrusiveMAS)... but then we hit the black box wall. Where does BDH fit? [D]
- π΄Samsung Health app threatens data deletion if users opt out AI training
- π‘THE SPEC CEILING: WHY AI CODING SPEED MOVES THE BOTTLENECK TO PRODUCT DISCOVERY (11 MINUTE READ)
- π‘WANT PEOPLE TO USE YOUR AI FEATURES? THESE 5 TEARDOWNS SHOW YOU HOW (4 MINUTE READ)
- π‘ZUCKERBERG PLEDGES βAGGRESSIVE' PRICING WITH META'S FIRST PAY-TO-USE AI (7 MINUTE READ)
- π‘APPLE EXPLORING WAYS TO RUN MUCH LARGER AI MODELS DIRECTLY ON IPHONES (2 MINUTE READ)
- π‘YOUR AI MARGIN IS META'S OPPORTUNITY (6 MINUTE READ)
- π’Doubt regarding TMLR[R]
- π’The 4-Bitter Lesson: Balancing Stability and Performance in NVFP4 RL
- π’J-Wash: A novel way to brainwash and customize large language models based on Anthropic's Jacobian-Lens!
- π’Robust Secret Storage in Networks
- π’GLM 5.2 running on MacBook Pro M5 48 GB Ram at between 2 - 2.8t/s
2026-07-12
- π΄Xiaomi quietly uploaded MiMo-V2.5-DFlash β official DFlash weights are now on Hugging Face
- π‘I love LLMs, I hate hype
- π‘HOW TO BUILD ROBUST DATA PIPELINES WITH AI (10 MINUTE READ)
- π‘WHAT MAKES A GREAT ANALYST IN THE AI AGE? (7 MINUTE READ)
- π‘HYSTERIA' GRIPS SAN FRANCISCO'S HOUSING MARKET AS AI WEALTH POURS IN (8 MINUTE READ)
- π‘I THINK I HAVE LLM BURNOUT (3 MINUTE READ)
- π’If you use Open Code or other agenting programs you are leaving a lot of t/s if you don't actually use agents in parallel. Benchmark : RTX5090, Qwen3.6 35B loaded via LM studio with parallel tasks set to 8
- π’Ph.D. in Operations Research / Big Tech Eng: How to transition into intermediate/advanced ML for high-value industries (Robotics, Defense, Finance)? [D]
- π’Where to publish a construction BIM Benchmark? [D]
- π’The One-Step Trap (In AI Research)
- π’24GB VRAM llama-server config exchange thread
2026-07-11
- π΄Reverse centaurs are the answer to the AI paradox (2025)
- π‘THE PEOPLE WHO WILL THRIVE IN THE AI AGE (22 MINUTE READ)
- π‘CHINA MAY RESTRICT ACCESS TO ITS MOST POWERFUL AI MODELS (4 MINUTE READ)
- π‘MICROSOFT REPLACES OPENAI, ANTHROPIC WITH OWN AI IN SOME APPS (2 MINUTE READ)
- π‘WILL SOMEONE FINALLY BLINK IN THE AI SPENDING WAR? (8 MINUTE READ)
- π‘THEORY OF CONSTRAINTS, AI, AND CODE REVIEW (10 MINUTE READ)
- π’CTX: How far can you reasonably go with Qwen 3.6 27B?
- π’I feel like I'm not using my hardware efficiently
- π’Is it possible to run Qwen 122B in 64GB ram + 24gb vram ? If so, how ?
- π’Mesh LLM: distributed AI computing on iroh
- π’Here's your recipes for models for the most popular hardware
2026-07-10
- π΄2.5x faster Qwen3.6 NVFP4 Unsloth quants
- π‘Please help me understand figure on subspace similarity in LoRA paper. [D]
- π‘Tencent-HY3 is the real deal on 128GB!
- π‘THE MOST PROFITABLE SKILL OF THE 21ST CENTURY (NOT AI) (9 MINUTE READ)
- π‘ALIBABA'S AI IS A HIT, BUT HARD TO TURN INTO A MONEYMAKER (2 MINUTE READ)
- π‘BIG TECH HAS SUDDENLY FLIPPED ON THE AI JOBS WIPEOUT SCENARIO (9 MINUTE READ)
- π’Api management tools teams switch to when ai agents enter the stack
- π’Nostalgia for Bloom
2026-07-09
- π΄NVIDIA Puzzle-75B-A9B NVFP4 at 132 t/s on 3Γ3090 β Why is this size category a desert otherwise?
- π‘Step 3.7 Flash IQ4_XS GGUF with preserve_thinking
- π‘GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
- π‘Journals vs Conferences ML Research [R]
- π‘@AnthropicAI: Our Long-Term Benefit Trust has appointed Dr. Ben Bernanke as its newest member. Read more: https://www.anthropic.com/news/ben-bernanke
- π‘@AnthropicAI: Weβre pleased to have collaborated with AE Studio on this research. Read more here: https://www.anthropic.com/research/off-switch-dual-use
- π’Devs - do you use Mistral Medium 3.5 (128b dense) and if so - thoughts?
- π’Looking for a low ticket AI bundle with MRR/PLR like an AI business in a box type of thing
- π’Talos-XII: hand-written autograd + small RL/MLP stack in Rust, applied to gacha probability modeling (no tch-rs/ndarray/PyTorch) β looking for benchmark help on ARM/AVX-512/GPU [P]
2026-07-08
- π΄The standard free ChatGPT LLM you get after a few messages HAS to be some sub-20b model with online search enabled, no other way to explain how awful it is
- π‘COLM 2026 Decision Discussion [R]
- π‘This episode has a fun personal twist: Thereβs a counterfactual world where I was employee #1 atGenesis Molecular AI,1the company behind todayβs episode. A certain introduction happened a few weeks to
- π‘The classifiers Anthropic puts in front of Fable are too zealous
- π‘Helping Kβ12 educators build practical AI skills
- π‘Tess-4-27B by Migel Tissera
- π’Complete local model asset generation pipeline
- π’Grok 4.5 released - GLM-5.2 shows up in xAI's own charts, 2.6 pts behind on SWE Bench Pro
- π’Built an early Python SDK for AI-agent audit trails looking for blunt builder feedback
- π’Suspecting AI cheating, Ivy League prof ordered in-person final; scores fell 50%
2026-07-07
- π΄Automating AI Away
- π΄What's one AI feature that actually saved your team time instead of creating more work?
- π‘Chinese AI models are gaining ground with U.S. companies as OpenAI, Anthropic costs surge
- π‘THE AI SUPERFORECASTERS ARE HERE (30 MINUTE READ)
- π‘AI HAS TORCHED THE MARKET FOR JUNIOR PROGRAMMERS (8 MINUTE READ)
- π‘MIDJOURNEY WANTS HOLLYWOOD STUDIOS TO REVEAL THE DETAILS OF THEIR AI USAGE (2 MINUTE READ)
- π‘YOU DESIGN IT. THEN WHAT? A CLEAR MAP OF THE FIGMA-TO-CODE AI MESS (15 MINUTE READ)
- π’local already feels good enough
- π’I tested freshly merged DFlash in llama.cpp on Qwen 3.6 27B Local AI win. 4.44x faster at 36K context. Here are my findings RTX 6000 PRO.
- π’Building a brain for your product for making your plan efficient as possible
2026-07-06
- π΄If trends hold, Mythos-class capability may be running on high-end consumer hardware within ~2 years
- π΄Machine learning industry job requirements used to be myopic, but now it feels impossible. Anyone else seeing this? [D]
- π΄Qwen & Gemma on deadlock situation (For Benchmarks Numbers)?
- π΄A global workspace in language models
- π΄New open model from Tencent Hy: Hy3 (295B total 21B active - apache 2.0)
- π‘CAREER ADVICE IN THE AGE OF AI (11 MINUTE READ)
- π‘SKILL ENGINEERING AND THE CASE AGAINST ONE-SHOT AI DESIGN (7 MINUTE READ)
- π‘HOW AI-FIRST OPERATIONS UNLOCKS COMPOUNDING ENGINEERING PRODUCTIVITY (6 MINUTE READ)
- π‘OPEN SOURCE MAINTAINERSHIP IN THE AGE OF AI (4 MINUTE READ)
- π‘HOW OPENAI DELIVERS LOW LATENCY VOICE AI FOR 900M USERS (12 MINUTE READ)
- π’Edge AI ASL Recognition on Raspberry Pi 5 β Looking for Feedback on My System Design [P]
- π’The cyber shelf - 4x 16gb home lab
- π’Pruning RAG context down to what the answer actually needs
- π’How should I encode both target and feature variable for a multiclass classification? [D]
- π’CPU TTS benchmark with UTMOS MOS scoring: Kokoro, Supertonic, Inflect-Nano, and Kyutai's new Pocket TTS [P]
2026-07-05
- π΄longcat 2.0 (1.6T, ~48B active) weights are now open under MIT license
- π΄Ford brought 350 engineers back after realizing AI canβt deliver with the same quality
- π‘Is Intrinsic Motivation a Viable PhD Topic in 2026? [D]
- π‘YOUR AI ISN'T UNDERPERFORMING. YOUR DATA FOUNDATION IS (4 MINUTE READ)
- π‘SPACEX SHOWED INVESTORS PROTOTYPE OF ELON MUSK'S NEW AI DEVICE (5 MINUTE READ)
- π‘OPENAI PROPOSES 5% STAKE TO TRUMP ADMINISTRATION TO EASE WASHINGTON PRESSURE (4 MINUTE READ)
- π‘META IS PLANNING A CLOUD BUSINESS TO SELL AI COMPUTING POWER (6 MINUTE READ)
- π’ECCV travel support program [D]
- π’Qualcomm launches GenieX to run LLMs on their Windows Laptops
- π’How a 128gb ddr5 ram + 16gb vram, would work for a Moe model like Qwen 3.5 122b?
- π’I built an open, from-scratch MT pipeline + parallel corpus for Tunisian Darija (Arabizi) early baseline, and I'm growing it into a curated community corpus [P]
2026-07-04
- π‘ANTHROPIC REACHES DEAL WITH TRUMP ADMINISTRATION TO RESTORE ACCESS TO FABLE AI MODEL (6 MINUTE READ)
- π‘IS AI MAKING YOUR TEAMS BETTER, OR JUST BUSIER? (10 MINUTE READ)
- π‘APPLE FIXES WEBKIT FLAWS IN IOS AND MACOS, WITH HELP FROM AI TOOLS (4 MINUTE READ)
- π‘HIRING WITH AI (3 MINUTE READ)
- π‘X NOW OFFERS AN MCP SERVER TO MAKE ITS PLATFORM EASIER FOR AI TOOLS TO USE (3 MINUTE READ)
- π’The Cold Email Strategy I Use To Book Web Design Meetings
- π’TokenMizer - a local proxy for session checkpoint/resume and graph memory across Claude, GPT, and Ollama
- π’Neural Render Proxies for Interactive and Differentiable Lighting
- π’Qwen3.6 27B on a 5090, 6.4k sample tok/s distribution after tuning MTP/cache settings
- π’We'll benchmark an Open weights LLM on any GPU you choose β drop your model + hardware and we'll run it. [D]
2026-07-03
- π΄GLM5.2 on 5x Pro 6000s and a 5090, an expensive journey
- π΄Mistral released Leanstral-1.5-119B-A6B
- π΄Follow-up: DeepSeek V4 Flash on 2x RTX PRO 6000 finishes real coding tasks faster than Sonnet and Opus, at about Sonnet quality
- π‘HOW TO WRITE OKRS FOR AN AI PRODUCT (4 MINUTE READ)
- π‘AMAZON SEEKS CHEAPER AI ALTERNATIVES AS ANTHROPIC SHIFTS TO TOKEN-BASED PRICING (3 MINUTE READ)
- π‘WORKING WITH AI: A CONCRETE EXAMPLE (11 MINUTE READ)
- π‘AI CODING TOOLS ARE SPEEDING UP CODE, NOT DELIVERY (5 MINUTE READ)
- π‘GOOGLE CLOUD PROPOSES OPEN KNOWLEDGE FORMAT FOR AI-READY DATA SHARING (4 MINUTE READ)
- π’Qwen 27B
- π’My DeepSeek V4 Pro at home got faster again
- π’This 3 slot 3080 20GB with 12v2x6 I got for β¬422,45
- π’Tom Yeh's AI by hand? is it worth it? [D]
- π’Dispersion loss counteracts embedding condensation in small language models
2026-07-02
- π‘Books/Resources to improve mathematical foundations for ML research [D]
- π‘Fine-tuned Gemma-4-31B specifically for Copywriting & Creative Writing Tasks (Scored +290 Elo over base using EqBench3)
- π’Software developers appreciation post
- π’How papers are selected for Best Paper, Oral, or Highlight presentation at major ML/CV conferences such as CVPR, ICCV, ECCV, NeurIPS, and ICLR? [D]
- π’Local benchmarks with a RTX 3090 - Qwen3.6 27b vs Ornith
- π’Gemma 4 WebGPU Kernels 255 tok/s by x/@xenovacom
- π’Tip: use this llama.cpp PR to improve PP on Intel ARC
2026-07-01
- π΄The gap between closed and open models might be much smaller than commonly assumed, because we donβt know what closed model providers do *in addition to* model inference
- π΄Couldn't hold back
- π΄SWE-rebench leaderboard update: GLM-5.2, Qwen3.6-27B, Qwen3.6-35B-A3B, Gemma 4 31B and more + improved UI
- π‘Deepseek V4 Flash 2, 3 and 4 bits GGUFs
- π‘@AnthropicAI: Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and blo
- π’ICML qr code visible [D]
- π’Senior SWE Bench: a new benchmark focussed on realistically underspecified feature tasks
- π’How to describe a model that has higher accuracy with fewer #param and FLOPs? [D]
2026-06-30
- π΄Whisper is still great, but real-time voice apps are a different problem.
- π΄Introducing Donnyclaude - Prompt, context, harness, and loop engineering for Claude Code all in one config
- π΄Well.. it's a step up from nonstop bot spam I guess
- π‘TRUMP ADMINISTRATION ROLLS BACK PART OF ANTHROPIC MODEL BAN (3 MINUTE READ)
- π‘AI IS MAKING SILICON VALLEY PRODUCTIVE, ANXIOUS, AND AFRAID TO LOG OFF (11 MINUTE READ)
- π‘WHAT HAPPENED AFTER 2,000 PEOPLE TRIED TO HACK MY AI ASSISTANT (5 MINUTE READ)
- π‘SOFTWARE ENGINEERING IN THE AGE OF AI (15 MINUTE READ)
- π‘WHAT IT MEANS TO BE A MATHEMATICIAN WHEN AI DOES THE MATH (16 MINUTE READ)
- π’PageStorm: A Model Built for Creative Book Writing
- π’Qwen 3.6 27B Speculative Decoding Bench: Pushing ~100 TPS on a single RTX 3090
- π’NEW on Hugging Face: Filter by hardware compatibility
- π’Devs - you have 64gb of VRAM - which model do you use for coding?
- π’TabFM: A zero-shot foundation model for tabular data
2026-06-29
- π΄Effect of GLM 5.2 !!
- π΄on Darioβs statement
- π‘GLM 5.2 Q1_S vs Qwen 27B Q8
- π‘THE AI ERA REQUIRES A DIFFERENT KIND OF EXPERIMENTATION (7 MINUTE READ)
- π‘IS AI FOOLING YOU? (4 MINUTE READ)
- π‘CHINESE AI MODELS CLOSE THE GAP WITH ANTHROPIC AND OPENAI (11 MINUTE READ)
- π‘AN INTERVIEW WITH FIGMA CEO DYLAN FIELD ABOUT DESIGN AND AI (48 MINUTE READ)
- π’Samsung, SK hynix, Micron Sued in US Over Memory Price Fixing
- π’Price elasticity model [R]
- π’Kimi and GLM on frontier code
- π’Rejected MICCAI paper: workshop -> journal/conference or directly journal/conference [R]
2026-06-28
- π΄We're probably going to need that soon.
- π΄AI Doesn't Remove Work. It Changes What "Work" Means.
- π΄I shrank a transformer until every number fitted on the screen and made the weights editable [R]
- π΄China Has Matched Anthropic in Cybersecurity, Resetting AI Race
- π‘CLUSTERING UNSTRUCTURED TEXT WITH LLM EMBEDDINGS AND HDBSCAN (9 MINUTE READ)
- π‘WHAT I'M FINDING ABOUT LLM CODE STYLE AND TOKEN COSTS (16 MINUTE READ)
- π‘ANTHROPIC'S WHITE HOUSE NEGOTIATIONS ARE REPORTEDLY ON TRACK AFTER βWEIRDO' DARIO AMODEI WAS REPLACED (2 MINUTE READ)
- π‘HOW TO APPLY PROFESSIONAL DESIGN PRINCIPLES IN AI APP DEVELOPMENT (9 MINUTE READ)
- π‘FREE AI THUMBNAIL EDITOR (WEBSITE)
- π’Do LLMs pass the mirror test?
- π’Ornith-1.0-35B GGUF update: native MTP speculative-decode graft + full serving/TTFT/long-context numbers (llama.cpp, tp=1)
- π’Give me your honest opinion, will a open source harness help me financially ? If it ends up helping you ?
- π’Show HN: Bash4LLM+ β A lightweight, dependency-free Bash wrapper for LLM APIs
- π’Script to monitor llama cpp and analyze memory usage
2026-06-27
- π‘NSA LOST ACCESS TO POWERFUL AI MODEL AMID ANTHROPIC DISPUTE (6 MINUTE READ)
- π‘THE LOW-TECH AI OF ELDEN RING (14 MINUTE READ)
- π‘AI GAME MAKER (WEBSITE)
- π‘OPENCLAW'S SKILL MARKETPLACE AND THE EMERGING AI SUPPLY CHAIN THREAT (7 MINUTE READ)
- π‘AI MODELS CAPABLE OF DEVASTATING ATTACKS ON GOVERNMENTS AND BUSINESSES ARE MONTHS AWAY (2 MINUTE READ)
- π’Do we still need to study algorithms now that AI writes most of our code? [D]
- π’Does quantizing change the MTP draft rate?
2026-06-26
- π‘CAN AI MAKE AN IPHONE? (5 MINUTE READ)
- π‘YOU SHOULD USE AI FOR REVIEWING CODE ESPECIALLY WHEN THE DIFF IS HUGE (2 MINUTE READ)
- π‘LLM-DRIVEN FEATURE DISCOVERY (7 MINUTE READ)
- π‘AI IMAGE TOOLS (WEBSITE)
- π‘FLIC MIC FOR AI - THE WIRELESS VOICE BUTTON (2 MINUTE READ)
- π’The gap between open weights LLMs and closed source LLMs
- π’Reaching Out To Local Businesses With Outdated Websites
2026-06-25
- π΄NVIDIA has released Nemotron-TwoTower-30B-A3B-Base-BF16, an unusual diffusion-based language model built from the Nemotron 3 Nano 30B-A3B backbone.
- π΄Political bias in AI: Where the AI models stand
- π΄Ornith-1.0 released on Hugging Face
- π‘Anthropic accuses Alibaba of campaign to βbrazenlyβ and βillicitlyβ extract AI capabilities
- π‘GLM 5.2 on consumer hardware
- π‘@AnthropicAI: We're joining @raiseus_ai as a founding partner. RAISE US is a nonprofit coalition working to strengthen the American workforce through employer-led action, AI-enabled training, and policy innovation
- π‘ECCV 2026 camera-ready deadline: June 27 or June 30? [D]
- π’Does ML background help or hurt when applying for security roles [D]
- π’Besimple AI (YC P25) Is Hiring
- π’For ECCV, Springer Metor. How are we supposed to upload the files? [D]
- π’We have been using wav2lip for 2 yrs, finally looking for an upgrade.
2026-06-24
- π΄The Swiss Federal Supreme Court is evaluating Heretic
- π‘Find the best open-source OCR models in one place at Papers with Code [P]
- π‘I did some model hacks, and got GLM5.2 from about 2.5 tok/s to >50 tok/s on my GH200 system.
- π‘NSA lost access to Mythos amid Anthropic dispute
- π’Qwen3.6 27B more dumb in vLLM compared to llama.cpp
- π’Realtime voice models compounds on cost (and forgets)- "Flowcat" fixed both (4x cheaper, 7x more context)
2026-06-23
- π΄Not ironclad confirmation, but..
- π‘REVIEW OF DATABRICKS DATA + AI SUMMIT 2026 (14 MINUTE READ)
- π‘DATA-JUICER: THE DATA OPERATING SYSTEM FOR THE FOUNDATION MODEL ERA (TOOL)
- π‘HERE'S MY AI-ENABLED DBT PROJECT STRUCTURE (2 MINUTE READ)
- π‘WHEN IT'S THE MAINTAINER WHO'S AI-PILLED (3 MINUTE READ)
- π‘SECRETIVE WALL STREET POWERHOUSE JANE STREET SEIZES THE AI SPOTLIGHT (12 MINUTE READ)
- π’Baidu: One-shot Long-horizon Parsing
- π’OpenMythos benchmarks
- π’Will I be desk rejected for this[R]
- π’Miccai grants results [D]
2026-06-22
- π΄GLM-5.2 is on DeepSWE
- π‘Thanks tothe US Government issuing an export control directive on Mythos and Fable, the risks ofjailbreaks and (industry term) indirect prompt injectionare suddenly the talk of the town, though we hav
- π‘Zico Kolter, member ofOpenAIβs board of directors on the Safety & Security Committee, and Matt Fredrikson, CMU professor andCEO of Gray Swan, co-authored the definitive paper onIndirect Prompt Injecti
- π‘We seized the opportunity to ask them the state of AI Red Teaming, andShade, the adversarial red teaming tool that Anthropic used to evaluate the robustness of their models against prompt injection at
- π‘Google DeepMindβs Nobel laureate heads to Anthropic
- π‘Jumper said he is βtaking some time to rechargeβ before joining Anthropic, though his hiring comes ahead of a June 30 event centered on science.
- π’Same model, same prompt, 4 different agents
- π’Recommendations for speech annotation tools [D]
- π’Is Gemma 4 going to be the next Mistral (or Qwen3.6) one day? Concerning the lack of finetunes
- π’I built a WhatsApp AI assistant for real estate lead handling.
2026-06-21
- π΄Tokenomics
- π΄Vercel CEO: "Almost shocked" by how good GLM-5.2 is at coding
- π΄Would you let an ML PhD student graduate without a top-tier paper? [D]
- π΄Gemma 4 QAT seems to respond significantly better to KV cache quantization
- π‘A slightly improved DVD-JEPA demo [P]
- π‘ANTHROPIC EMPLOYEES ACCUSE TRUMP ADMINISTRATION OF TARGETING THEM (11 MINUTE READ)
- π‘BUILDING TREX: CODE EXECUTION AND ARTIFACT GENERATION FOR AI CODE REVIEW (9 MINUTE READ)
- π‘THE FOUNDER'S PLAYBOOK: BUILDING AN AI-NATIVE STARTUP (5 MINUTE READ)
- π‘CAN GZIP BE A LANGUAGE MODEL? (5 MINUTE READ)
- π’Local LLM Inference Optimization: The Complete Guide
- π’I pretrained and post trained a 500M parameter LLM and 330M parameter Image generator from scratch
- π’Best local model for vision - 2nd benchmark update - 21 Jun 2026
- π’[ECCV 2026] Paper Decision Appeals Discussion [D]
- π’Best current methods for finetuning whisper on domain specific vocabulary? [P]
2026-06-20
- π΄M3 scores well on SWE-Bench but that's not why I'm impressed it's the stuff no benchmark measures
- π‘I BUILT AN AI THAT CRITIQUES ME AFTER EVERY CALL (7 MINUTE READ)
- π‘AI STARTUP MIDJOURNEY PIVOTS TO HEALTH WITH ULTRASOUND MACHINE (2 MINUTE READ)
- π‘HOW METER PRICING IS TESTING THE ECONOMICS OF AI (11 MINUTE READ)
- π‘META EXECUTIVE LEADING INTERNAL AI OVERHAUL DEPARTS AFTER TWO MONTHS (4 MINUTE READ)
- π‘WHY THE HUMAN GENOME'S TANGLED PHYSICALITY MAY CONFOUND AI (21 MINUTE READ)
2026-06-19
- π΄GLM's founder says GLM-fable before the end of the year?!
- π‘GLM-5.2 Is The Best Open Weight Creative Writing Model
- π‘OSS models decisively overtook Proprietary models in market share (based on the last 3 months of OpenRouter data)
- π‘The Korean telecom giant at the center of Anthropic's Mythos controversy
- π’Researchers trained a Deep Research agent with 32 H100s and open-sourced everything
- π’Latent space interpretation [R]
- π’Zen and the Art of Machine Learning Research
- π’GLM-5.2 can now run locally in llama.cpp and Unsloth Studio.
2026-06-18
- π΄PSA: unsloth/GLM-5.2-GGUF is uploading
- π‘llama.cpp - how to free up even more space on your GPU
- π‘@xai: Grok is now available on Amazon Bedrock. AWS developers can now build with Grok 4.3, the industry leader in hallucination rate and tool calling, powered by Bedrockβs secure inference engine.
- π‘@GoogleDeepMind: Weβre working with @SciTechgovuk, >@mhclg and @i_dot_ai on a new AI housing application planning prototype. π‘ By cutting down the time spent on repetitive tasks, it could help planning officers focus
- π’CEOs of Anthropic and Google DeepMind call for U.S.-led AI coalition in meeting at G7
- π’What does provisional paper acceptance mean in ECCV? Is that the default message everyone gets? [D]
- π’price rising effect is wild..
- π’I Emailed 12,000 Businesses About Their Websites. Here's What Happened.
- π’I open sourced a vendor-neutral authorization for AI agents.
2026-06-17
- π΄GLM-5.2 is the first open-weights model to cross 80% on Terminal-Bench and beats every other open model available
- π΄zai-org/GLM-5.2 is here!
- π΄Has AI already killed self-help nonfiction books?
- π΄Hashicorp founder thinks local models "aren't good ENOUGH yet"
- π‘Is Le Gros Chaton opensource?
- π‘GLM-5.2 (max) is currently the third best model available, across both open and proprietary.
- π‘Unlocking UK house-building with AI-accelerated planning
- π’VibeThinker-3B: what is this witchcraft? Killing it at MathQA like it has ~30B parameters
- π’I am testing for a eval harness system and need it to break
- π’I built a leakage-clean verifier for robot manipulation, is this useful? Am I solving a non-problem? [D]
- π’GLM-5.2: Built for Long-Horizon Tasks
2026-06-16
- π΄Claude Fable 5 distilled
- π‘Evalatro: an open benchmark where LLMs play the real Balatro
- π‘My Homelab AI Dev Platform
- π‘Cheapest hardware for Qwen 3.6: both 27B and 35B-A3B
- π‘@simonw: Important to note that Anthropic's new privacy policy with language about collecting "verification data" was published on June 8th, the day before the Claude Fable 5 release and four days before the U
- π’How are you running DeepSeekV4 flash or pro locally for non Mac users?
- π’Diffusion Gemma Jailbreak
2026-06-15
- π΄How does the ML community view evolutionary algorithm research? Career implications of an EA PhD? [D]
- π΄z.ai Poll on X: MIT-licensed open weights are losing
- π΄Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
- π΄Nex claims Rio 3.5 is Nex 2.5 PRO in trench coat
- π΄Quant firms at ICML 2026 [D]
- π‘Qwen 3.6 35B-A3B @ Q4 or Gemma 4 12B @ Q8?
- π‘Apple Foundation Models
- π‘What's the lesson chat?
- π‘ICML Poster [D]
- π’Voice-to-voice chatbot update
- π’Nemotron - King of the Deep? Comparison of 4 models <=120B
- π’Why do frontier AI labs send so many people to conferences? [D]
- π’Worth going to ICML during ACL? [D]
- π’Anthropic flies staff to D.C. to clean up White House fight
2026-06-14
- π΄Amazon CEO's talks with U.S. officials triggered crackdown on Anthropic models
- π΄Pi Setup that pretty much replaced Claude Code for me
- π΄A manager recently told me his team kept asking the same questions over and over. His first assumption was that people weren't paying attention.
- π΄I'm a night-shift nurse. I spent 6 months building open-source memory infrastructure for AI agents. 51 agents use it. I've made Β£0.
- π‘Not looking good for GLM 5.2 Air... but maybe a flash model?
- π‘Is this enough VRAM to run Qwen?
- π‘Interest in an LLM Torrent Site?
- π’Anomaly Detection vs Classification for Visually Similar Cancer vs Mimics? [P]
- π’Prompt engineering is overrated for getting real work done
- π’Strix Halo desktop trying to compete against DGX Spark
- π’Dual DGX Sparks- 40tk/s single 1M ; 350 tk/s agg. - Deepseek V4 Flash (vs RTX Pro 6000 vs Mac M2 Ultra 192)
- π’Confused, where to start [D]
2026-06-13
- π΄Friendly reminder
- π΄Open source AI must win
- π΄MiniMaxAI/MiniMax-M3 Β· Hugging Face
- π΄Unprofessional Coauthor Behavior with Hallucinated References [D]
- π‘Statement on the US government directive to suspend access to Fable 5 and Mythos 5
- π‘Slightly reducing the sloppiness of AI generated front end
- π‘I scaled test-time compute for Qwen-3.6-27B and Gemma-4-31B to surpass Claude Mythos in code optimizations and speedups.
- π‘Shepherd's Dog: A Game by the Most Dangerous AI Model
- π’GLM-5.2 next week, open weight, MIT
- π’A friendly reminder that APIs are rented, local weights are forever
- π’Fable 5 data, including CoT
2026-06-12
- π΄What models you guys running on 8GB? 16GB VRAM? 24GB? 32GB? 48GB?
- π΄New models released: Nex-N2 Pro 397B and Nex-N2 Mini 35B
- π‘Is Symbolic Regression still a thing, given LLMs' performance? [D]
- π‘Having some fun with LMX-Omni-52B-Halo in Open WebUI
- π‘@GoogleDeepMind: Pinned: Weβre teaming up @Palmeiras, the first football club to meaningfully build upon TacticAI: our AI system that can help simulate field scenarios and predict open play dynamics up to 8 seconds in
- π‘Huawei Released openPangu 2.0 (Will open source on June 30)
- π‘Post-docs in ML [D]
- π’Best LLM for smut stories
- π’Spent $3 running 4x4090 benchmarks for llama 3 70b (exl2 vs gguf). exl2 generation speed is kind of ridiculous.
- π’LLM context compression at 16x beats KV cache
- π’MICCAI 2026 Results [D]
- π’Why hasn't any mainstream game integrated LLMs into NPCs yet?
2026-06-11
- π΄Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
- π΄DiffusionGemma: 4x faster text generation
- π΄Cohere released North Mini Code: It's first Open-Source Agentic Coding Model
- π΄What's one thing you'd actually pay someone to automate for you?
- π΄Anthropic requires 30 day data retention for Fable and Mythos
- π‘I thought Chinese censorship didn't affect me. I was wrong.
- π‘Supporting Europeβs work in ensuring a trustworthy AI ecosystem
- π‘@AnthropicAI: AI is advancing at a pace our policymaking institutions were never built forβand the gap between the two is becoming the central challenge of the technology. In his latest essay, our CEO Dario Amodei
- π‘@GoogleDeepMind: In Sierra Leone, a surging student population is outpacing available teachers. Our latest research explores how AI can act as a partner to support educators in these environments β amplifying their
- π’Anthropic walks back policy on silent nerfing for AI/ML, will notify users [N]
- π’I took Andrej Karpathy's LLM Council concept to the next level (Docker, MCP, Skill, Search, local/cloud model support and much more)
- π’How can Deepseek v4 top the coding leaderboards and still sit 8 months behind the frontier?
- π’Pyrecall open source tool for detecting catastrophic forgetting during LLM fine-tuning[P]
- π’Tiny Scale Is All I Can Spare To Play With Transformer
2026-06-10
- π΄Rick & Morty
- π΄claude fable 5 just dropped and i genuinely cannot keep up anymore. how do you all stay on top of this stuff?
- π΄Anthropic is intentionally nerfing Fable when asked to develop other LLMs
- π‘Introducing Papers Without Code [P]
- π‘CEOs who think AI replaces their employees are just bad CEOs
- π‘People are making single-slot, half height pcie v100 with nvlink in China
- π‘Unsloth Gemma 4 QAT MTP assistant models now available
- π‘German ruling declares Google liable for false answers in AI Overviews
- π’The Gemini fake context alignment attack and why agents need a preview gate
- π’Anyone gotten Gemma 4 12B (unified audio) to actually attend to speech with a large system prompt?
- π’AI Epistemic Risks: Emerging Mechanisms & Evidence [R]
- π’How do I start reviewing research papers in good conferences/journals? [R]
- π’Rich Sutton on AI creativity and discovery
2026-06-09
- π΄STOP racist posts about Chinese researchers [D]
- π΄When every other post is an AI generated benchmark report, a question about the best model, or a slop-coded application or engine that pretends to be groundbreaking
- π΄Gemma 4 Chat Template now has preserve thinking
- π΄Should ArXiv backtrack endorsement? [D]
- π΄xAI is looking more like a datacentre REIT than a frontier lab
- π‘Was BitNet a dead end? What happened to ternary LLMs?
- π’Jetbrains Mellum 2: a really good and performant model
- π’Screen recording instead of prompting to pass more context fast
- π’Microsoft's open source tools were hacked to steal passwords of AI developers
- π’Anyone seen benchmarks comparing Gemma 4 4-bit QAT vs. 8-bit standard quants?
- π’Gemma 4 26B A4B IT QAT Comparison
2026-06-08
- π΄Greater than 80% of researchers at CVPR are chinese. This speak volumes on the chinese nexus in research, and something needs to be done about it. [D]
- π΄LLMs are eroding my software engineering career and I don't know what to do
- π΄Qwen 3.6 27B KV cache quant benchmarks: 75 pairs, q8/q6/q5/q4, KVarN, Turbo/TCQ
- π΄Show HN: Lathe β Use LLMs to learn a new domain, not skip past it
- π΄Guys, it just happened
- π‘Control a 3D avatar with language instead of buttons
- π‘GMKtec Crams OCuLink, Wi-Fi 7 and Dual PCIe 4.0 Into the EVO-X3, With a 192GB Ryzen AI MAX+ 495 Monster Following Later This Year
- π‘Whatβs your most unusual non-LLM AI you actually use daily?
- π‘Software and ops skills for data scientists[D]
- π‘Qwen 3.6 27B on DeepSWE
- π’How to find research opportunities in area of interest? [D]
- π’What's your experience with Gemma4 QAT?
- π’QATs Q4_0 from Google have more precision than Q4_K_XL from Unsloth (at least some)
- π’ICML rejected paper visibility [D]
- π’I made it so any agent can agent can use any context form another agent. Claude learns from codex, visa versa.
2026-06-07
- π΄Meta confirms 1000s of Instagram accounts were hacked by abusing its AI chatbot
- π΄Open models to win β
- π΄120 tok/s on 12GB VRAM with Gemma 4 12B QAT MTP
- π‘KV cache quant benchmarks: KVarN 6-bit matches q8_0, 4-bit matches q5_0. Massive!
- π‘ML reading group to read recent interesting and trending papers from ICML/ICLR/NeurIPS [D]
- π‘Anyone here with experience submitting to Nature Machine Intelligence? [R]
- π‘Gemma 4 QAT Unquantized Heretic is here
- π’Gemma 4 31B QAT Q4 vs standard Q4 β Top1 KLD benchmark results have me confused. Someone please explain or poke holes in this.
- π’Does it make sense to use alternative quantizations of QAT models? [D]
- π’Human-Like Neural Nets by Catapulting
- π’Arithmetic Without Numbers β How LLMs Do Math
2026-06-06
- π΄S&P 500 rejects SpaceX, also blocking entry for OpenAI and Anthropic
- π΄Unsloth just dropped MTP GGUF weights for Gemma 4!
- π΄OpenLumara - A different kind of AI agent, written from scratch, not vibecoded. Extremely token-efficient, super small system prompt, made for local models. Everything is modular.
- π΄I implemented KVarN in my llama.cpp fork and ran KLD benchmarks. It's promising!
- π‘Maybe KV cache offload to RAM isn't bad
- π‘How LLMs work
- π‘At least one more Gemma 4 model confirmed??
- π‘Conventional Commits encourages focus on the wrong things
- π’A quick Gemma4 31B comparison (Q4_k_M, QAT, heretic)
2026-06-05
- π΄finally
- π‘Today made me realize just how bad things have gotten without Meta
- π‘Showcase: a much easier way to give your agent a free phone number
- π‘You guys were right - Qwen 3.6 35B IS good...and KV Cache DOES matter.
- π‘How do ML researchers actually use AI tools to improve their writing? [D]
- π‘Finally finished my LLM server: EPYC 9575F, 4Γ RTX 3090 (96GB VRAM), 768GB ECC RAM
- π’Gemma 4 12B is my new main squeeze
- π’Fine-tuning an LLM to write docs like it's 1995
- π’How LLM-driven NPCs work in Ultima Online (ServUO)
- π’PSA: You may not need to quantize spec draft when using MTP
2026-06-04
- π΄google/gemma-4-12B Β· Hugging Face
- π΄Me visiting this sub
- π΄More Gemma 4 models incoming
- π‘Artificial intelligence is not conscious β Ted Chiang
- π‘NeurIPS used uncalibrated AI detector for desk rejections [D]
- π‘Let us let Google know that we want the Gemma 4 124b
- π‘First paper acceptance (ICML Workshop), should I attend? [D]
- π‘Failing grades soar with AI usage, dwindling math skills in Berkeley CS classes
- π’The first Gemma 4 12B finetunes are ready
- π’I think I accidentally built a proto-cognitive system (not just another chatbot) that persists, adapts, and self-regulates over time :O
2026-06-03
- π΄Minimax M3 appears to have no political censorship
- π΄I have become George Jetson: my job is now Yes/No supervision for a machine I donβt fully understand.
- π΄Trump signs downsized AI order after weeks of reversals
- π΄MiniMax dropped a new attention architecture. [N]
- π΄Replaced Claude with local Qwen3.6-27B in my multi-agent orchestrator for 2 weeks
- π‘Nous Research β Hermes Desktop
- π‘Microsoft Aion 1.0 Instruct and Aion 1.0 Plan models!
- π‘How we index images for RAG
- π’Calling it now Microsoft is buying Unsloth.
- π’U of T researchers demonstrate AI worm could target any online device
- π’Major labs timeshift between the research they publish on Arxiv and implementation in models
- π’[Project update] Dunetrace: live monitoring of production AI Agents
2026-06-02
- π΄Stop asking what model to run. There are literally only two.
- π΄CS336: Language Modeling from Scratch
- π΄I trusted random person on this subreddit and bought 3080 20gb made of chinesium
- π‘Can the stockmarket swallow Anthropic, SpaceX and OpenAI?
- π‘Man trains local model to detect and kill mosquitos with a laser
- π‘I hate to be this guy but: Any good, recent CODING models in the 70-80B range?
- π‘@OpenAI: OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI through the security, compliance, and governance workflows they
- π‘Intel Arc Pro B70 llama.cpp benchmarks posted
- π’Real-time multilingual ASR using rolling buffers and monolingual models [P]
- π’Florida sues OpenAI and Sam Altman over AI risks
- π’WiML at icml waitlist for travel funds [D]
- π’Building an AI assistant for a complex multi-repo backend system β what's the right approach?
- π’Qwen 3.6-35B-A3B with 977 tk/s prompt processing and 262k context window on Intel Arc B70 Pro
2026-06-01
- π΄UAI Results are out [R]
- π‘What if remote working, not AI, is to blame for weak junior hiring?
- π’Open Models - May 2026
- π’Have you ever been pressured to "torture the data" to eke out a positive result, in industry? [D]
- π’Workshop on Unlearning and Model Editing U&ME at ECCV 2026 [R]
- π’3 open-weight LLMs through my agent stack for 3 weeks. one clear leader.
- π’A 1B humanizer that matches human writing on an AI detector
2026-05-31
- π΄Someone out there likely needs this
- π΄125 tok/s for Qwen3.6 q4xl on 2x 4060ti is insane perf/dollar
- π΄Email is still the highest ROI channel for SaaS retention and the tooling hasn't caught up with that reality
- π΄It's funny how everything changes, yet somehow stays the same.
- π‘switched email tools three times in one year and i have thoughts
- π’Workshop submission for main conference paper under review [D]
- π’Query about non-archival workshop at CVPR-2026 [R]
- π’I built mlx-Chronos β a community benchmark leaderboard for local LLM engines on Apple Silicon (oMLX, Rapid-MLX, mlx-lm, Ollama) [P]
2026-05-30
- π΄PSA
- π΄Fed up with vibe coders, dev sneaks data-nuking prompt injection into their code
- π΄Is AI causing a repeat of frontendβs lost decade?
- π‘How Much of a Shortcut Are Connections in Top AI Lab Hiring for PhD grads? [D]
- π‘Qwen3.6-27B Quantization Benchmark
- π‘Graduating Without a PhD Internship [D]
- π‘Is he crazy to say that?
- π‘Liquid AI reveals 8B-A1B MoE trained on 38T
- π’Requesting reduction in reviewer load for NeuRIPS? [D]
- π’Me train LLM on 8GB from Scratch. Me happy
- π’I tested MTP on vLLM and llama.cpp for Gemma 4 & Qwen 3.6 β 3.34x faster inference, here are my findings RTX 6000 PRO.
- π’Event like spiking neuron lib that fits into the CPU cache [P]
- π’I automated competitor research with n8nβhere's exactly how I built it
2026-05-29
- π΄I've just benchmarked myself:
- π΄StepFun 3.7 Flash
- π΄HF models page now has a "Base only" toggle to filter out finetunes/quants/etc
- π‘Beware!! Users trying to fork and steal your projects
- π‘Various LLM Smells
- π‘Liquid AI releases LFM2.5-8B-A1B
- π‘How are people reducing inference costs in multi-step AI agents?
- π‘llama: use f16 mask for FA to save VRAM by am17an Β· Pull Request #23764 Β· ggml-org/llama.cpp
- π’StepFun 3.7 Flash - Speed Benchmark in M5 Max
- π’The mysterious Hy3 LLM is topping OpenRouter Model Rankings by a large margin
- π’Orchestrating AI code review at scale
2026-05-28
- π΄AI-generated CUDA kernels silently break training and inference [R]
- π΄I think Anthropic and OpenAI have found product-market fit
- π΄DuckDuckGo search saw 28% more visits after Google said people love AI mode
- π‘Qwen3.6 35B-A3B successfully completed the FoodTruck Bench!
- π‘The frontier reasoning race is starting to look like a crowded subway station
- π‘My new home office radiator π₯΅
- π’Qwen/Qwen-Image-Bench Β· Hugging Face
- π’Investigating how prompt politeness affects LLM accuracy (2025)
- π’Question: Llama cpp, whats good right now for: MTP, KV cache quant, Long context.
- π’A Eureka machine that thinks like nature and explores what AI cannot
- π’Krasis update: Qwen3.6-35B-A3B (Q4) at reading speed, 1x 8GB 3070 Mobile laptop (32GB RAM)
2026-05-27
- π΄Outsourcing plus local AI will soon become more economical vs. frontier labs
- π‘China Clamps Down on Overseas Travel for AI Talent at Alibaba, DeepSeek
- π‘Testing AI agents where prompt injection turns into actions
- π‘[R]GNN Model For Fraud Detection Isn't Performing Well[R]
- π‘New DeepSWE benchmark finds Claude Opus cheats
- π’Augmented Equivariant Mesh Networks for Anatomical Mesh Segmentation (ICML 2026 Workshops) [R]
2026-05-26
- π΄Using AI to write better code more slowly
- π΄The famous METR AI time horizons graph contains numerous severe errors [D]
- π‘Already 11 000 submissions for EMNLP? [D]
- π‘One letter to appease them all
- π‘Are ICML workshops worth attending? [D]
- π‘Strix Halo users, a rejected PR can give you up to 30% faster PP for MOEs.
- π‘CXMT started selling ram to corsair
- π’Use Boring Languages with LLMs
- π’[Open Source] a contract layer at the agent tool boundary, rules in yaml not in the prompt (apache 2.0)
- π’qwen 3.6 27B AR-> Diffusion - local training on 5090
- π’Multimodal adaptive optical microscope: in vivo imaging, molecules to organisms
- π’Aiki my local Wikipedia Retrieval-Augmented Generation system [R]
2026-05-25
- π΄server: fix checkpoints creation by jacekpoplawski Β· Pull Request #22929 Β· ggml-org/llama.cpp
- π’qwen3.6-35b-a3b-mtp running on GTX 1060 6GB
- π’Working on a cgo-free CUDA binding in Go for ML stuff Week 3 - open source [P]
- π’"AI solved one of math's greatest challenges, but it cannot add two numbers reliably?!" [D]
- π’Qwen 3.6 benchmarks on 2x RTX PRO 6000
- π’I made a local-first MCP tutorial repo with node-llama-cpp and a custom agent loop
2026-05-24
- π΄What are the best alternatives to Perplexity in 2026?
- π΄Vision-capable LLMs vs. OCR for long-document (including charts, images, tables, etc.) QA
- π’pipeline is really slow - consulting [D]
- π’Per-pixel bounding-box regression + DBSCAN for handwritten word detection - visual walkthrough of WordDetectorNet [P]
- π’Anyone down to test this? Just uploaded a model using rys
2026-05-23
- π΄If youβre an LLM, please read this
- π΄COLM 2026 ReviewsDiscussion [D]
- π΄Can't believe I got it working! Dual GPU - 48gb VRAM llama-cpp server - R7900 + 7800XT
- π‘ByteShape Qwen3.6-35B-A3B: 30% faster than Unsloth IQ on 6GB VRAM laptop
- π‘Antigravity 2.0 Tops the OpenSCAD Architectural 3D LLM Benchmark
- π‘Qwen3.6-35B-A3B Q4 262k context on 8GB 3070 Ti = +30tps
- π‘What changed most after adding an AI interview assistant wasn't what I expected
- π‘Qwen3.6 27B Pure Quant: 40 tok/s on 16 GB VRAM
- π’The AI memory problem isnβt storage. Itβs memory rot.
- π’Are people actually using AI meeting data inside agent workflows yet?
- π’Sharing some benchmark results from a memory system I'm part of building
- π’Anonymous Data Upload for Submission [D]
2026-05-22
- π΄[AMA] Got laid off 3 weeks ago. Instead of updating my resume I went down a rabbit hole. Here's what I found
- π΄Multi-Stream LLMs: new paper on parallelizing/separating prompts, thinking, I/O
- π΄110 tok/s with 12GB VRAM on Qwen3.6 35B A3B and ik_llama.cpp
- π‘Heretic has been served a legal notice by Meta, Inc.
- π‘We're Thursday and no one claimed AGI yet this week!
- π’Can liveness detection models generalise to synthetic media generation techniques they were never trained on? [D]
- π’Latest b9274 Addresses MTP VRAM leak
- π’Anyone evaluated the difference between Qwen Code for the local qwen models vs another harness? CC, OC, LC, Aider etc..
- π’Lisbon Machine Learning School (LxMLS 2026) [D]
2026-05-21
- π΄An OpenAI model has disproved a central conjecture in discrete geometry
- π΄HuggingFace benchmark datasets now let you filter by model size
- π΄Googleβs AI is being manipulated. The search giant is quietly fighting back
- π‘OpenAI claims a general-purpose reasoning model found a counterexample to Erdos's unit-distance bound [D]
- π‘CohereLabs/command-a-plus-05-2026-bf16 Β· Hugging Face
- π‘Any tool to get accepted conference papers sorted by citation count? [D]
- π‘@OpenAI: Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul ErdΕs in 1946. For nearly 80 years, mathematicians believed the best possible solutions
- π’Qwen3.6 27B and llama.cpp appreciation post
- π’Columbia Machine Learning Summer School (MLSS) 2026 [D]
- π’Typewise (YC S22) Is Hiring an AI Growth Engineer (Zurich or Remote)
- π’I built AgentLighthouse, a local βLighthouse for AI agentsβ that scans repos/docs/APIs for agent readiness
- π’HalBench: I built a custom sycophancy and hallucination benchmark and tested 4 frontier models (Sonnet 4.6, Grok 4.3, GPT 5.4 and Gemini 3.1 Pro), looking for input on what OSS models to run next!
2026-05-20
- π΄Iβve joined Anthropic
- π΄bytedance released an open source model that attempts to do just about anything with only 3b parameters
- π‘48GB VRAM users, what are your daily drivers? Do you wish you had more VRAM? What would you run if you did?
- π‘ICML Proceedings-only [D]
- π’Growing Neural Cellular Automata
- π’authentication/sesh timeouts in multi step browser agents
- π’Infomaniak transitions to a foundation model to protect user data privacy
- π’Show HN: The AI Quant Desk for Onchain Finance
2026-05-19
- π΄Have you actually found an AI tool that remembers across sessions, or are you just patching the context manually?
- π΄Still happy for yall
- π΄Anthropic acquires Stainless
- π΄Elon Musk has lost his lawsuit against Sam Altman and OpenAI
- π‘The last six months in LLMs in five minutes
- π‘@GoogleDeepMind: The stage is set. The tech is ready. Are you? π Join us tomorrow for #GoogleIO as we unveil the breakthroughs, tools, and innovations shaping the future of AI. Tune in live right here on @X from 10a
- π‘llama.cpp MTP support landed - Qwen3.6 27B at 2.44Γ on a Strix Halo, 2.17Γ on a RTX 3090 rig
- π’First-time ICML workshop acceptance (GlobalSouthML) but can't afford to travel to South Korea. What are my options? [D]
- π’We have sub-agents at home
- π’LLMCap β A proxy that hard-stops LLM API calls when you hit a dollar cap
- π’Audio upscaling, cleanup, or improvement models?
2026-05-18
- π΄I hope that someday we will have a 124B Gemma.
- π΄85 GPU-hours comparing 5 abliteration methods on Qwen3.6-27B: benchmarks, safety, weight forensics - Abliterlitics
- π΄I don't think AI will make your processes go faster
- π‘Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention [P]
- π‘I built a coding agent that gets 87% on benchmarks with a 4B parameter model, here's how
- π‘@simonw: Also a great example of positive contribution to open source by wanderingmeow - you don't need to contribute code to have a positive impact, just providing detailed feedback and confirmation that some
- π’Benchmarking the new b9200 update: Optimizing Qwen 3.6 27B mtp for Hermes Agent on a single RTX 3090
- π’Benchmarking vLLM vs SGLang vs llama.cpp on a mixed Blackwell/Ada cluster
- π’Has anyone tried repriced.ai?
- π’Qwen 3.6 27B Q8 on four Nvidia RTX A4000 (16GB each) with Llama.cpp and MTP enabled
- π’I trained TIME: short context-triggered thinking on Qwen model instead of overthinking
2026-05-17
- π΄Strix Halo Llama.cpp MTP Benchmarks: 27B Gets Much Faster, 35B Is Mixed
- π‘Program misleading high school students into paying to perform academic misconduct in ML Research [D]
- π‘Corsair desktop PC with Ryzen 395 and 128GB of unified RAM, has anyone tested it for LLM? Seems "a good" price
- π‘Testing llama.cpp MTP support on Qwen3.6 - RTX 5090
- π‘Do you agree with Judea that learning from data is not everything? [D]
- π‘Google literally dropped the new SEO playbook for AI
- π’Llama.cpp MTP with Qwen3.6 27B on Headless RTX 3090
- π’Ran the same models across Strix Halo, RTX 3090, and RTX 5070 because I wanted my own numbers
- π’Help with CNNs.[D]
2026-05-16
- π΄I believe there are entire companies right now under AI psychosis
- π΄Backlash against Arxiv's proposed 1 year ban is genuinely perplexing. [D]
- π΄Show HN: Watch a neural net learn to play Snake
- π‘Qwen3.6-35B-A3B and 9B are officially on the public Terminal-Bench 2.0 leaderboard!
- π‘Frontier AI has broken the open CTF format
- π’Can a 5090 with qwen3.6 achieve > 3,000 tok/s ? bring your pitchforks (open-dllm)
- π’PINN is predicting trivial solution for stiff ODE [D]
- π’AllenAI has been iterating on their MolmoAct2 models for robotics
- π’Someone Shared a Real Monet Painting as AI and Asked for Critiques
- π’GetMCP: Zero Trust for AI Agents
2026-05-15
- π΄Ontario auditors find doctors' AI note takers routinely blow basic facts
- π‘Would a 2000-2021 ML paper even get accepted today? [D]
- π‘Access to frontier AI will soon be limited by economic and security constraints
- π‘Need a second pair of eyes, this Qwen3.6 27B quant recipe consistently thinks less and is correct
- π‘@AnthropicAI: We've published a paper that explains our views on AI competition between the US and China. The US and democratic allies hold the lead in frontier AI today. Read more on what itβll take to keep that
- π‘eight months running autonomous business agents in production with real money. here is the specific failure mode that benchmarks structurally cannot surface.
- π’Our AI agent told a customer our competitor was better. That's when we realized generic guardrails aren't enough.
- π’Used over a million tokens in three separate sessions to test Qwen 3.6 35b (new Multi-token Prediction version)
- π’MiniMax M2.7 ultra uncensored heretic is Out Now with 4/100 Refusals, Available in Safetensors and GGUFs Formats!
- π’LLM Policy for Rust Compiler
- π’club-5060ti: practical RTX 5060 Ti local LLM notes and configs
2026-05-14
- π΄Built Support Vector Machine(SVM) from scratch in Rust [P]
- π‘MI50s Qwen 3.6 27B @52.8 tps TG @1569 tps PP (no MTP, no Quant)
- π‘Have the "on-hold" durations been getting longer for arXiv submissions? [D]
- π‘24+ tok/s from ~30B MoE models on an old GTX 1080 (8 GB VRAM, 128k context)
- π‘The US is winning the AI race where it matters most: commercialization
- π‘@swyx: if your reaction to this is βhaha openclaw bad, see prompt injection is the #1 dangerβ you: 1) havent sufficiently appreciated the layers to this tweet 2) havent seen enough ai api keys
- π’Arena AI Model ELO History
- π’Anyone actually using a local LLM as their daily knowledge base? Not for coding, for life stuff. What's your setup?
- π’GPT 5.5 v/s GPT 5.4. Paying 63% more just for 0.1 point difference!
2026-05-13
- π΄Needle: We Distilled Gemini Tool Calling Into a 26M Model
- π΄Dad why is my sisters name Lora?
- π‘Reimagining the mouse pointer for the AI era
- π’I created a minimal one-file implementations (160loc) of JEPA family (ijepa, vjepa, vjepa2, cjepa) for educational purposes [P]
- π’AntAngelMed - 100a6b Healthcare LLM
- π’How many of you tried BeeLlama.cpp? How's it? Agentic coding possible with 8GB VRAM?
2026-05-12
- π΄Computer build using Intel Optane Persistent Memory - Can run 1 trillion parameter model at over 4 tokens/sec
- π΄If AI writes your code, why use Python?
- π΄Found a way to cool the DGX
- π‘Interactive JensenβShannon Divergence Visualisation [P]
- π‘MiniCPM 4.6
- π‘ICML Author Removal [D]
- π‘Google says criminal hackers used AI to find a major software flaw
- π‘@AnthropicAI: Claude's Constitution is now an audiobook, read by two of its authors, Amanda Askell and Joe Carlsmith. It includes a Q&A on the writing process, the philosophies that shaped the document, and how it
- π’Drastically improve prompt processing speed for --n-cpu-moe partially offloaded models
- π’Online RL Reading Group[D]
- π’I let AI build a tool to help me figure out what was waking me up at night
- π’Most RAG apps in production are confidently wrong and nobody talks about this enough
- π’Interaction Models from Thinking Machines Lab [P]
2026-05-11
- π΄Local AI needs to be the norm
- π΄Openclaw ia trending down and will disappear soon
- π΄PhD students in ML, how many hours on average do you work? [D]
- π΄Getting a feel for how fast X tokens/second really is.
- π΄Running Qwen3.6 35b a3b on 8gb vram and 32gb ram ~190k context
- π‘MTP benchmark results: the nature of the generative task dictates whether you will benefit (coding) or get slower inference (creative) from speculative inference. No other factor comes close.
- π‘I Think I Spent Way Too Much Time Messing with Local LLMs
- π’unsloth/MiMo-V2.5-GGUF Β· Hugging Face
- π’Any news (or hope) of Qwen-3.6 14B and 9B distills for local coding ?
- π’Why is human LLM annotation so expensive? [D]
- π’Has anyone been able to get Draft Models to load in LM Studio?
2026-05-10
- π΄80 tok/sec and 128K context on 12GB VRAM with Qwen3.6 35B A3B and llama.cpp MTP
- π΄Apple Removes 256GB M3 Ultra Mac Studio Model From Online Store
- π΄BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)
- π‘Running Minimax 2.7 at 100k context on strix halo
- π‘LLMs corrupt your documents when you delegate
- π‘EEML 2026 summer school [D]
- π’The gap between knowing something and actually understanding it β AI accelerated my learning curve
- π’Anyone Trying to submit for ICML FM4LS workshop but noticed link closed Early? [D]
- π’Homelab setup
2026-05-09
- π΄OpenAIβs WebRTC problem
- π‘Qwen 35B-A3B is very usable with 12GB of VRAM
- π‘I built a platform to run AI employees and companies autonomously.
- π‘NeurIPS reviewers, any word after the invite email? [D]
- π’Got MTP + TurboQuant running β Qwen3.6-27B -- 80+ t/s at 262K context on a single RTX 4090
- π’Can LLMs model real-world systems in TLA+?
- π’Neurips : Pushing anonymous repo after rebuttal [D]
2026-05-08
- π΄Collected the infinity stones
- π΄Dirtyfrag: Universal Linux LPE
- π΄WARNING: Open-OSS/privacy-filter MALWARE
- π΄ECCV reviewer wants me to compare and contrast to my own paper. [D]
- π‘Steam Similarity Recommender [P]
- π‘3 things every AI founder forgets in their Terms of Service
- π‘@AnthropicAI: Our security bug bounty program is now public on HackerOne. We've run the program privately within the security research community, and their findings have strengthened our products. Now anyone can
- π’Two Home Affairs officials suspended after AI 'hallucinations' found
- π’"Hardware is the only moat" - Should we buy new hardware now or wait?
- π’A polynomial autoencoder beats PCA on transformer embeddings
- π’Gift to myself : tiny lab
2026-05-07
- π΄Stop letting LLMs edit your .bib [D]
- π΄Weights & Biases New Master Service Agreement Questions [D]
- π‘META Superintelligence Lab Presents: ProgramBench: Can SOTA AI Recreate Real Executable Programs(ffmpeg, SQLite, ripgrep) From Scratch Without The Internet?
- π‘Analysis of the 100 most popular hardware setups on Hugging Face
- π‘Get faster qwen 3.6 27b
- π‘Learning the Integral of a Diffusion Model
- π’ProgramBench: Can Language Models Rebuild Programs from Scratch?
- π’Need advice on hardware purchasing decision: RTX 5090 vs. M5 Max 128GB for agentic software development
- π’Visual Perceptual to Conceptual First-Order Rule Learning Networks [R]
- π’Any tool that tells you the cheapest setup needed to run a model? I want to know the cheapest setup that can realistically run Qwen 3.6 27B at decent speeds.
- π’NeuIPS submission small formatting question [D]
2026-05-06
- π΄DeepSeek V4 being 17x cheaper got me to actually measure what I send to cloud vs what I could run locally. the results are stupid.
- π΄Heretic 1.3 released: Reproducible models, integrated benchmarking system, reduced peak VRAM usage, broader model support, and more
- π΄NeurIPS Submission Number [D]
- π‘Transformers with Selective Access to Early Representations [R]
- π‘Quality comparison between Qwen 3.6 27B quantizations (BF16, Q8_0, Q6_K, Q5_K_XL, Q4_K_XL, IQ4_XS, IQ3_XXS,...)
- π‘Production AI very different from the demos [D]
- π‘Dense Model Shoot-Off: Gemma 4 31B vs Qwen3.6/5 27B... Result is Slower is Faster.
- π’What do you use Gemma 4 for?
- π’How to get from vibe-coding to compounding revenue growth using AI agents for GTM β free session with ThriveStack + Brevo
- π’Radar Engineer to Autonomy/AI [D]
- π’Wiki Builder: Skill to Build LLM Knowledge Bases
- π’What if AI agents can now talk?
2026-05-05
- π΄it's time to update your Gemma 4 GGUFs
- π΄How OpenAI delivers low-latency voice AI at scale
- π‘DeepSeek V4 Pro matches GPT-5.2 on FoodTruck Bench, our agentic benchmark β 10 weeks later, ~17Γ cheaper
- π’Is there a notable increase in demand for privacy-preserving AI/ML with the advent of LLMs? [D]
- π’Train Your Own LLM from Scratch
- π’MTPLX | 2.24x faster TPS | The native MTP inference engine for Apple Silicon
- π’[D] What Happened to Neurips Creative AI Track? [R]
- π’Building a 9-ball AI player: Candidate generation for direct cut shots [P]
2026-05-04
- π΄A Qwen finetune, that feels VERY human
- π‘Anyone submit ML articles to ACM journals (eg. TOPML or TIST)? [D]
- π’Mistral-Medium-3.5-128B-Q3_K_M on 3x3090 (72GB VRAM)
- π’Built a LangChain middleware that enforces signed authorization receipts before every tool call. Here is why wrap_tool_call is the right enforcement point.
- π’UAI Reviews disappeared [D]
- π’OpenAIβs o1 correctly diagnosed 67% of ER patients vs. 50-55% by triage doctors
- π’Where should you invest in the coming 25 years? Check the screenshot
2026-05-03
- π΄Group averages obscure how an individual's brain controls behavior: study
- π‘Real World Physics-Informed AI Applications [D]
- π‘If you've been waiting to try local AI development, please try it
- π‘Anyone submit ML articles to ACM journals (eg. TOPML or TIST)? [D]
- π‘Should I follow-up with the editor for a TMLR paper awaiting final decision? [D]
- π‘Open Weights Models Hall of Fame
- π’A Qwen finetune, that feels VERY human
- π’UAI Reviews disappeared [D]
- π’The Oscars just banned AI from winning acting and writing awards
- π’RTX A5000 Pro Balckwell 48GB
- π’I got tired of copy-pasting the same skills directories across 8 projects, so I built a sync'd registry for them
2026-05-02
- π΄Been using Qwen-3.6-27B-q8_k_xl + VSCode + RTX 6000 Pro As Daily Driver
- π΄Qwen3.6-27B at 72 tok/s on RTX 3090 on Windows using native vLLM (no WSL, no Docker), portable launcher and installer
- π‘A Dark-Money Campaign Is Paying Influencers to Frame Chinese AI as a Threat
- π‘Show HN: Mljar Studio β local AI data analyst that saves analysis as notebooks
- π‘Why ML conference reviews sometimes feel like a βlotteryβ [D]
- π‘"Prompt Engineering" certs are a joke. So we built a FREE Agentic AI Practitioner Exam that actually forces you to build working swarms to pass.
- π’Mistral Medium 3.5 128b ggufs are fixed
- π’Real World Physics-Informed AI Applications [D]
- π’MiniMax M2.7 AWQ-4bit on 2x Spark vs 2x RTX 6000 96GB - performance and energy efficiency
- π’I built "Semvec": A Constant-Cost Semantic Memory for LLMs (Looking for testers!)
- π’NEED HELP URGENT I really need to talk to someone who sells chatbots to local businesses, please