🔴 High Significance
Model Releases
🔴 🤗 moonshotai/Kimi-K3 (1,388,105 downloads) — score 95
Sources: huggingface_models
Author: | Downloads: 1,388,105 | Likes: 10340
🔴 🧡 DeepSeek V4 Flash 0731 — score 92
Sources: hackernews
🔴 💬 This is why the vast majority aren't taking any "this new model is dangerous" messages seriously. They've cried wolf FAR too many times. They could literally announce that a nuclear war caused by AI is 24 hours away and many wouldn't bat an eye — score 91
Sources: reddit/r/singularity · reddit/r/OpenAI
🔴 💬 Got job as Director of AI and Systems development self-taught — score 90
Sources: reddit/r/LocalLLaMA
Hey everyone, I just wanted to share my journey here for some motivation. Three years ago, I saw the sudden spike in AI and realized it was the future of tech. My goal at the time was to be an indie game dev, and seeing that AI could write basic code, I told myself I needed to master it or risk bein
🔴 🤗 MiniMaxAI/MiniMax-H3 (26,693 downloads) — score 85
Sources: huggingface_models
Author: | Downloads: 26,693 | Likes: 3098
Omitted 5 additional model releases items from the main section; see raw data and source-specific sections below.
Developer Tools
🔴 🐙 PrimeIntellect-ai/prime-agent — A self-improving RLM agent for coding workflows and long-running autonomous tasks. — score 99
Sources: github_trending
A self-improving RLM agent for coding workflows and long-running autonomous tasks.
🔴 🐙 virgiliojr94/book-to-skill — Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work. — score 97
Sources: github_trending
Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.
🔴 💬 Learned the term "context poisoning" today and now I can't stop noticing it — score 95
Sources: reddit/r/artificial
Someone explained this to me in a comment thread and it's been rattling around in my head since. The idea: in a long conversation, if the model says something wrong and you correct it, that correction doesn't necessarily erase the wrong idea's influence. The tokens around the mistake, including the
🔴 💬 I created an AI agent that sells your services and products for you — score 94
Sources: reddit/r/AIAgents
So I wondered why everyone was so hyped about Vibe-coding when the real struggle for businesses is distribution, aka finding people who want to use your product or service and give you money for it. That is why I created https://vonto.ai, enter your domain, a brief description an
🔴 🐙 google/skills — Agent Skills for Google products and technologies — score 94
Sources: github_trending
Agent Skills for Google products and technologies
Omitted 11 additional developer tools items from the main section; see raw data and source-specific sections below.
Infrastructure & Compute
🔴 🐙 cloudflare/computer — Give your agent a computer 👾 — score 98
Sources: github_trending
Give your agent a computer 👾
Research Papers
🔴 🤗 Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval — score 95
Sources: huggingface
Short segments of perceived speech can be retrieved from non-invasive magnetoencephalographic (MEG) recordings by deep networks trained with a CLIP-style objective against wav2vec 2.0 audio embeddings. Yet their weights do not map onto electrophysiological quantities, and it remains unclear which sp
🔴 🤗 DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces — score 85
Sources: huggingface
Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured files, long documents, and multimedia. Existing benchmarks largely isolate structured querying, retrieval, or open-ended analysis, leaving heterogeneous
🔴 🤗 Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay — score 75
Sources: huggingface
Computer-use agents pay full frontier inference to re-derive routines their user has already performed, because an agent's memory today records what the user said, not what the user did. We compile passively captured screen activity into agent memory with a deterministic, zero-model pipeline: it seg
Other Signals
🔴 💬 2027 Memory Capacity Is Reportedly Sold Out — score 97
Sources: reddit/r/LocalLLaMA
🔴 💬 Your scientists were so... uh... — score 94
Sources: reddit/r/OpenAI
🔴 💬 NeurIPS OpenReview Modified Timeline [D] — score 88
Sources: reddit/r/MachineLearning
I just wonder if “modified: 04 Aug 2026 at 22:21” time part in the Reviewer’s official review still remain unchanged, then is that mean reviewer not add final justification, or any change of their rating?
🔴 💬 NeurIPS AI Assisted Review authors/reviewers? [D] — score 88
Sources: reddit/r/MachineLearning
Out of curiosity, if you were a reviewer or author, how did the review period go? For me, it was weird, because I gave reviews with specific details (what specifically could have been better, how to fix it), but realized other reviewers gave similar superficial reviews. Even the paper which was a co
🔴 💬 Chinese LLMs dominate this week's top charts — score 85
Sources: reddit/r/artificial
Source: https://openrouter.ai/rankings
Omitted 4 additional other signals items from the main section; see raw data and source-specific sections below.
🟡 Notable
Model Releases
🟡 💬 GPT 5.6 Sol and Fable 5 settle a 25 year old problem in wireless communication theory — score 69
Sources: reddit/r/singularity
🟡 🤗 DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF (2,345,190 downloads) — score 65
Sources: huggingface_models
Author: | Downloads: 2,345,190 | Likes: 1760
🟡 ✉️ Editor’s note: I’m excited to welcomeShloktoour guest post roster! You may know Shlok from his excellent explorations (as an outsider — for an insider perspective see ourpodcast with OpenAI’s Akshay N — score 65
Sources: newsletter/Latent Space
Editor’s note: I’m excited to welcomeShloktoour guest post roster! You may know Shlok from his excellent explorations (as an outsider — for an insider perspective see ourpodcast with OpenAI’s Akshay Nathan. Already one of our most popular episodes of the year!) ofleading AI Lab memory systems, which
🟡 ✉️ On July 9th, OpenAI releasedChatGPT Work, their agent product for knowledge work. It was, by any measure,a busy launch: three new modelsacrossfourteen configurations, a consolidation of the ChatGPT an — score 65
Sources: newsletter/Latent Space
On July 9th, OpenAI releasedChatGPT Work, their agent product for knowledge work. It was, by any measure,a busy launch: three new modelsacrossfourteen configurations, a consolidation of the ChatGPT and Codex desktop apps, andcloud agents brought to the mainstreamin their most accessible form yet.
🟡 ✉️ Editor’s note: ChatGPT estimated to cross1B MAU in Juneand1B WAU this month. — score 65
Sources: newsletter/Latent Space
Omitted 115 additional model releases items from the main section; see raw data and source-specific sections below.
Developer Tools
🟡 🐙 rivet-dev/rivet — Rivet Actors are the primitive for stateful workloads. Built for AI agents, collaborative apps, and durable execution. — score 69
Sources: github_trending
Rivet Actors are the primitive for stateful workloads. Built for AI agents, collaborative apps, and durable execution.
🟡 🐙 tirth8205/code-review-graph — Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows. — score 68
Sources: github_trending
Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.
🟡 🐙 AgriciDaniel/claude-seo — Universal SEO skill for Claude Code. 25 sub-skills + 18 sub-agents covering technical SEO, E-E-A-T, schema, GEO/AEO, backlinks, local SEO, maps intelligence, semantic clustering, e-commerce SEO, international SEO, Google APIs, and PDF/Excel reporting. Optional DataForSEO, Firecrawl, and Banana extensions. — score 66
Sources: github_trending
Universal SEO skill for Claude Code. 25 sub-skills + 18 sub-agents covering technical SEO, E-E-A-T, schema, GEO/AEO, backlinks, local SEO, maps intelligence, semantic clustering, e-commerce SEO, international SEO, Google APIs, and PDF/Excel reporting. Optional DataForSEO, Firecrawl, and Banana exten
🟡 🐙 tashfeenahmed/freellmapi — OpenAI-compatible proxy that stacks the free tiers of 28 LLM providers (~4B tokens/month) behind one /v1 endpoint — plus any custom OpenAI-compatible endpoint. Smart routing, automatic failover, encrypted keys. Personal experimentation only. — score 65
Sources: github_trending
OpenAI-compatible proxy that stacks the free tiers of 28 LLM providers (~4B tokens/month) behind one /v1 endpoint — plus any custom OpenAI-compatible endpoint. Smart routing, automatic failover, encrypted keys. Personal experimentation only.
🟡 ✉️ I’m trying something new - this email walks through one of my actual agent sessions and I’ll explain what’s happening along the way. The build or task I’m doing isn’t important. But I’m looking at how — score 65
Sources: newsletter/Ben's Bites
I’m trying something new - this email walks through one of my actual agent sessions and I’ll explain what’s happening along the way. The build or task I’m doing isn’t important. But I’m looking at how I could be using agents more effectively.
Omitted 40 additional developer tools items from the main section; see raw data and source-specific sections below.
Infrastructure & Compute
🟡 ✉️ Self-sustaining and self-replicating AI viruses are here:…Open weight LLMs + a well-designed harness = a persistent, self-sufficient virus…AI researchers have built a prototype computer virus which us — score 65
Sources: newsletter/Import AI
Self-sustaining and self-replicating AI viruses are here:…Open weight LLMs + a well-designed harness = a persistent, self-sufficient virus…AI researchers have built a prototype computer virus which uses AI models to compromise computers, then uses their underlying GPU resources to run inference, let
🟡 ✉️ TheArtifacts Hub— a curated view of the models trending on Hugging Face, highlighting inference tokens viaOpen Router, model intelligence viaArtificial Analysis, and ourtailored adoption metricsbuildi — score 65
Sources: newsletter/Interconnects
TheArtifacts Hub— a curated view of the models trending on Hugging Face, highlighting inference tokens viaOpen Router, model intelligence viaArtificial Analysis, and ourtailored adoption metricsbuilding on top ofHugging Face’s data.
🟡 ✉️ HUAWEI'S TOP SCIENTIST WARNS OF CHIP LIMIT NVIDIA WILL SOON FACE (4 MINUTE READ) — score 65
Sources: newsletter/tldr
🟡 ✉️ AI infrastructure startup Volta, founded earlier this year, emerged from stealth with a reported $10B compute deal with Anthropic and funding at a $2.4B valuation. — score 65
Sources: newsletter/rundown-ai
🟡 🐙 vllm-project/vllm — A high-throughput and memory-efficient inference and serving engine for LLMs — score 61
Sources: github_trending
A high-throughput and memory-efficient inference and serving engine for LLMs
Omitted 3 additional infrastructure & compute items from the main section; see raw data and source-specific sections below.
Business & Funding
🟡 ✉️ BENDING SPOONS TO BUY AIRTABLE FOR $1.28B (2 MINUTE READ) — score 65
Sources: newsletter/tldr
🟡 🏢 Responding to the next frontier of critical cyber capabilities — score 50
Sources: lab_blog/OpenAI
OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
🟡 🏢 Taming Outlier Tokens in Diffusion Transformers — score 50
Sources: lab_blog/Apple ML
We study outlier tokens in Diffusion Transformers (DiTs) for image generation. Prior work has shown that Vision Transformers (ViTs) can produce a small number of high-norm tokens that attract disproportionate attention while carrying limited local information, but their role in generative models rem
Enterprise Adoption
🟡 ✉️ We’re expanding our open models coverage into standalone projects that let you go deeper on the state of the open model ecosystem. The new free sources of data are: — score 65
Sources: newsletter/Interconnects
Our other project is much lighter weight, but far overdue. Ever since we wrote The ATOM Project, we’ve been seeing the US-vs-China model adoption plot on a recurring basis in the AI ecosystem. We’d update the plot from time to time, but not enough. Now, we’re making the crucial data for that report
🟡 ✉️ OurAdoption Dashboard— a living dashboard of download and derivative model numbers by geography and organization. This highlights the US-China gap and growing players in the open ecosystem. — score 65
Sources: newsletter/Interconnects
🟡 ✉️ For the most popular models, the Hub let’s you quickly see how far behind the model was in terms of frontier intelligence based on Artificial Analysis’s Intelligence Index, compare Hugging Face and Op — score 65
Sources: newsletter/Interconnects
For the most popular models, the Hub let’s you quickly see how far behind the model was in terms of frontier intelligence based on Artificial Analysis’s Intelligence Index, compare Hugging Face and Open Router adoption to similar models, glance at relative adoption metric (RAM) scores for time-size
🟡 ✉️ THE ENTERPRISE AI STRATEGY THAT OUTLASTS ANY SINGLE MODEL (5 MINUTE READ) — score 65
Sources: newsletter/tldr
🟡 ✉️ SIXB: THE OPERATING LAYER FOR ENTERPRISE AI (5 MINUTE READ) — score 65
Sources: newsletter/tldr
Omitted 1 additional enterprise adoption items from the main section; see raw data and source-specific sections below.
Research Papers
🟡 🤗 KVAE: Family of Tokenizers for Multimodal Generative Models — score 65
Sources: huggingface
Latent diffusion modeling (LDM), a prominent paradigm, utilizes tokenizers to map input signal to compressed representation. This dependency positions tokenizer as an integral part of generation process itself, since it affects learning speed, quality of synthesized samples and lay foundation for la
🟡 🤗 MameLoshnLM: Yiddish Language Model and Evaluation Benchmark — score 55
Sources: huggingface
We present MameLoshnLM, the first open-source 8B-parameter language model built specifically for Yiddish. Despite Yiddish's rich textual tradition, its limited digital presence and the scarcity of reliable evaluation resources have constrained progress in Yiddish language modeling. Existing multilin
🟡 🤗 Continual Learning in Transition — score 45
Sources: huggingface
Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g., training strategies, architectural designs, and weight adaptation. However, emerging paradigms are reshaping the scope of CL beyond this traditional m
Other Signals
🟡 💬 ByteDance trains massive AI model in bid to rival Anthropic — score 65
Sources: reddit/r/artificial
🟡 ✉️ Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe. — score 65
Sources: newsletter/Import AI
“AI could help create a dramatically better future, but that outcome is not guaranteed. The world’s leading AI companies believe they could be close to automating AI research. It is hard to predict exactly how much this will accelerate AI progress, but there is a real risk that capability developmen
🟡 ✉️ We built the Hub as a way to go deeper on this analysis in collaboration withProject VAIL— an AI verification startup who has been one of the most loyal fans of our open model curation. — score 65
Sources: newsletter/Interconnects
🟡 ✉️ SAMSUNG REVEALS NEW 3D-MEMORY ROADMAP IN BID FOR AI TECH LEAD (2 MINUTE READ) — score 65
Sources: newsletter/tldr
🟡 ✉️ TURN ONE GIANT AI-GENERATED PULL REQUEST TO A REVIEWABLE STACK (11 MINUTE READ) — score 65
Sources: newsletter/tldr
Omitted 31 additional other signals items from the main section; see raw data and source-specific sections below.
🟢 Incremental
Model Releases
🟢 🤗 larryvrh/MiniMax-H3-Turbo-Lora (0 downloads) — score 35
Sources: huggingface_models
Author: | Downloads: 0 | Likes: 482
🟢 💬 Ur shipping so many bugs! No amount of instructions, memory, engineering standards, or repo structure will stop this. Which is why you have to spot and fix it. Claude Opus 5 on Max. — score 33
Sources: reddit/r/AIAgents
Every complex multi-phase task i give it I find some variation of the issues bellow. Over the past month, I’ve been setting up and fine-tuning a deterministic validation architecture. is currently set up to find things like. * Command/API contract drift and accidental state-shape changes. * Invalid
🟢 🤗 LiquidAI/LFM2.5-2.6B (81,522 downloads) — score 25
Sources: huggingface_models
Author: | Downloads: 81,522 | Likes: 415
🟢 🧡 Lost my phone at the office. Claude suggested tracking Bluetooth signal strength — score 25
Sources: hackernews
🟢 💬 DeepSeek V4 Flash 0731 ARC-AGI-1 and 2 — score 19
Sources: reddit/r/singularity
Interesting to see the latest version buck the trend of increasing cost for increased performance based on the reasoning level. Both for ARC-AGI-1 and 2 as well Just thought I'd post since I thought it was interesting, don't recall this being the case for any other model so far. Also insane just to
Omitted 6 additional model releases items from the main section; see raw data and source-specific sections below.
Developer Tools
🟢 🐙 open-metadata/OpenMetadata — The Open Context Layer for Data and AI , OpenMetadata is the open platform for building trusted data context and business semantics for humans, AI assistants, and agents. — score 39
Sources: github_trending
The Open Context Layer for Data and AI , OpenMetadata is the open platform for building trusted data context and business semantics for humans, AI assistants, and agents.
🟢 🐙 jordanrendric/claude-video-vision — Give Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis — score 38
Sources: github_trending
Give Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis
🟢 🐙 datawhalechina/self-llm — 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程 — score 36
Sources: github_trending
《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程
🟢 🐙 NirDiamant/GenAI_Agents — 50+ tutorials and implementations for Generative AI Agent techniques, from basic conversational bots to complex multi-agent systems. — score 34
Sources: github_trending
50+ tutorials and implementations for Generative AI Agent techniques, from basic conversational bots to complex multi-agent systems.
🟢 🐙 karpathy/micrograd — A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API — score 34
Sources: github_trending
A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API
Omitted 15 additional developer tools items from the main section; see raw data and source-specific sections below.
Infrastructure & Compute
🟢 🐙 higgsfield-ai/higgsfield — Fault-tolerant, highly scalable GPU orchestration, and a machine learning framework designed for training models with billions to trillions of parameters — score 26
Sources: github_trending
Fault-tolerant, highly scalable GPU orchestration, and a machine learning framework designed for training models with billions to trillions of parameters
🟢 🐙 NVIDIA-NeMo/Nemotron — Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models — score 22
Sources: github_trending
Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models
🟢 🐙 openai/CLIP — CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image — score 10
Sources: github_trending
CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
🟢 💬 Synthesizing and formally verifying a SWAR bit-hack for INT4 dot products using Z3 and Lean 4 [P] — score 6
Sources: reddit/r/MachineLearning
INT4 quantization is ubiquitous in ML right now, but evaluating dot products on hardware without native SIMD/vector instructions (like WebAssembly or older ARM chips) usually requires slow sequential loops. A classic workaround is SWAR (SIMD Within A Register), but deriving the bitwise operations by
🟢 🐙 CliMA/ClimaAtmos.jl — GPU-capable global atmosphere model of the CliMA Earth System Model, designed for calibration with data assimilation and machine learning — score 3
Sources: github_trending
GPU-capable global atmosphere model of the CliMA Earth System Model, designed for calibration with data assimilation and machine learning
Business & Funding
🟢 💬 Evaluation metrics - [D] — score 31
Sources: reddit/r/MachineLearning
I wanted to ask about the selection of evaluation metrics. In which scenario we use ROC-AUC score and in which scenario we use f1 score as an evaluation metric in a classification problem to define the model's performance on a specific dataset.
Research Papers
🟢 🤗 Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills — score 30
Sources: huggingface
Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that write and refine their own executable skills as code. This survey organises the field around that axis of weights versus skills. Its central analytic
🟢 🤗 FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds — score 30
Sources: huggingface
World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this emerging landscape, Joint Embedding Predictive Architectures (JEPA) offer a particularly compelling direction. We study a largely unexplored regime: po
🟢 🤗 Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation — score 15
Sources: huggingface
Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fundamentally different optimization strategies. We introduce Task-Conditional Flow Matching (TCFM), a multilingual embedding adaptation framework that se
🟢 🤗 GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization — score 5
Sources: huggingface
Selecting a complete 3D object from a reconstructed scene with minimal user effort is essential for practical scene editing and embodied interaction. Existing 3DGS-based methods either retrain the Gaussian representation to embed per-object labels, or build dense multi-view SAM observations, both re
Other Signals
🟢 💬 Companies seeing AI returns had their data and governance sorted first, per PwC's 4,454-CEO survey — score 35
Sources: reddit/r/artificial
🟢 💬 Is Microsoft-Phi dead? — score 33
Sources: reddit/r/LocalLLaMA
Phi was one of my favorite models with a bit of a mixed reputation with some claiming it's benchmaxxed and others seeing its potential and usecases. I was a big fan of Phi but the last major release was in december 2024 with every other release being a Phi 4 iteration (like Phi 4-reasoning-vision wh
🟢 💬 Tesla V100 Qwen3.6 27B Performance — score 33
Sources: reddit/r/LocalLLaMA
Looking for V100 users to share your config and it's performance. GPU: Tesla V100 PCIE 32Gb Qwen3.6 27B Q4_K_M + Q8_0 MTP 128K context length Pi coding agent llama.cpp model preset: [*] spec-default = 1 ctx-size = 131072 mmap = 1 kv-unified = 1 n-gpu-layers = 999 threads = 18 prio = 3 seed = 3407
🟢 💬 ICDE Results [D] — score 31
Sources: reddit/r/MachineLearning
Hello! Let's use this thread to discuss ICDE results which should be coming out shortly today (hopefully).
🟢 💬 We are in a bubble sell everything — score 31
Sources: reddit/r/singularity
Omitted 4 additional other signals items from the main section; see raw data and source-specific sections below.
📊 Cross-Source Signals
Items that appeared on 3+ sources today:
- Third-party cyber evaluations involving OpenAI models — appeared on: lab_blog/OpenAI (100), newsletter/tldr (0), newsletter/rundown-ai (0)
📈 Trending Repos
| Repo | Description | Stars Today | Language |
|---|---|---|---|
| PrimeIntellect-ai/prime-agent | A self-improving RLM agent for coding workflows and long-running autonomous tasks. | 2483 | typescript |
| cloudflare/computer | Give your agent a computer 👾 | 1005 | typescript |
| virgiliojr94/book-to-skill | Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work. | 660 | python |
| google/skills | Agent Skills for Google products and technologies | 481 | python |
| anomalyco/opencode | The open source coding agent. | 379 | typescript |
| farion1231/cc-switch | A cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official website: ccswitch.io | 267 | rust |
| huangruiteng/loopx | Lightweight loop engineering state kernel for long-running AI agent teams. Agent-loop agnostic across Codex, Claude Code, and other coding agents, with durable goals, quota-aware auto-wake, executable todos, evidence logs, and verifiable handoffs. | 244 | python |
| Significant-Gravitas/AutoGPT | AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters. | 240 | python |
| can1357/oh-my-pi | ⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and more | 232 | typescript |
| heygen-com/hyperframes | Write HTML. Render video. Built for agents. | 158 | typescript |
📄 New Papers
| Title | Category | Hotness | Link |
|---|---|---|---|
| Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval | research_paper | 61 | Open |
| DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces | research_paper | 27 | Open |
| Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay | research_paper | 24 | Open |
| KVAE: Family of Tokenizers for Multimodal Generative Models | research_paper | 18 | Open |
| MameLoshnLM: Yiddish Language Model and Evaluation Benchmark | research_paper | 17 | Open |
| Continual Learning in Transition | research_paper | 16 | Open |
| Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills | research_paper | 8 | Open |
| FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds | research_paper | 8 | Open |
| Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation | research_paper | 7 | Open |
| GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization | research_paper | 6 | Open |
🏢 Lab Blog Posts
- OpenAI: Third-party cyber evaluations involving OpenAI models
- OpenAI: Apple is getting this wrong
- OpenAI: How we built a realtime system for responsive voice AI in six months
- Anthropic: Aug 7, 2026 Product Improving Fable 5's biology safeguards
- Anthropic: Aug 4, 2026 Announcements Mariano-Florentino (Tino) Cuéllar to join Anthropic as Chief Global Affairs Officer
- Anthropic: Jul 30, 2026 Investigating three real-world incidents in our cybersecurity evaluations
- Anthropic: Jul 27, 2026 Announcements Our position on open-weights models
- Anthropic: Jul 27, 2026 Announcements Cognizant and Anthropic expand their partnership to bring Claude to enterprise clients
- Anthropic: Jul 24, 2026 Product Introducing Claude Opus 5
- Anthropic: Jul 22, 2026 Economic Research A research agenda for the Economic Futures Research Fund
- Anthropic: Jul 22, 2026 Product Ask Claude about the Anthropic Economic Index
- Anthropic: Jul 21, 2026 Announcements Anthropic is donating another $20 million to Public First Action
- Anthropic: Jul 20, 2026 Announcements Apply for Anthropic’s AI for Science rare disease research grants
- OpenAI: Responding to the next frontier of critical cyber capabilities
- OpenAI: How HSP GRUPPE builds AI capabilities for tax advisory
- OpenAI: Improving GPT‑5.6 Sol in ChatGPT—and expanding access to GPT-5.6 Luna for free users
- OpenAI: Working with the American Psychological Association on youth mental health and AI
- OpenAI: From asking to doing: How the world is putting ChatGPT to work
- OpenAI: New ways to learn and teach with ChatGPT Work and Codex
- OpenAI: Circles powers telco personalization with OpenAI technology
- Microsoft Research: Orchard: An open framework for scalable agentic AI
- Apple ML: Scaling Categorical Flow Maps
- Apple ML: Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models
- Apple ML: Arbitrage: Efficient Reasoning via Advantage-Aware Speculation
- Apple ML: Locking Pretrained Weights via Deep Low-Rank Residual Distillation
- Apple ML: DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness
- Apple ML: Taming Outlier Tokens in Diffusion Transformers
- Apple ML: Understanding Alignment in Multimodal LLMs: A Comprehensive Study
- Cohere: Cohere and the University of Waterloo launch partnership to strengthen Canada’s AI talent pipeline Launching in Fall 2026, the new program will prepare students to help organizations put AI to work on real business challenges. Aug 06, 2026 2 min read
- Mistral: Solutions Introducing Shieldstral. August 4, 2026 By Mistral
- Mistral: Product Your Prompts and Skills need a system of record. Studio gives AI prompts & skills a system of record—versioned, owned, and traceable. July 9, 2026 By Mistral
- Mistral: Research Introducing Robostral Navigate Robostral Navigate, our first model built for embodied navigation. July 8, 2026 By Mistral AI
- Mistral: Research Leanstral 1.5: Proof Abundance for All July 2, 2026 By Leanstral Team at Mistral AI
- Mistral: Engineering Bringing more control over your connectors June 24, 2026 By Mistral AI
- Mistral: Research Introducing Mistral OCR 4 State of the art document intelligence model. June 23, 2026 By Mistral AI
- Mistral: Company AI Now Summit 2026 Innovations for global enterprises solving the world’s hardest problems. May 28, 2026 By Mistral
- Mistral: Product Vibe gets to work. The unified agent for long-horizon productivity and coding, launching with Work and Code modes. Plus, a new Vibe VS Code extension. May 28, 2026 By Mistral
- Mistral: Product Introducing Search Toolkit Production search pipelines, anywhere. May 28, 2026 By Mistral
- Mistral: Solutions Introducing physics AI at Mistral: the foundation for engineering acceleration. A new class of AI models that predict the behavior of physical systems, powering the engineers and hardware products of tomorrow. May 27, 2026 By Mistral
- Mistral: Research Physics AI research that’s shaping the industry. Published breakthroughs pushing the state of the art. May 27, 2026 By Mistral
- Mistral: Company Emmi joins Mistral to accelerate the AI-native industry May 23, 2026 By Mistral AI
- Mistral: Product Remote agents in Vibe. Powered by Mistral Medium 3.5. Introducing Mistral Medium 3.5, remote coding agents in Vibe, plus new Work mode in Le Chat for complex tasks. May 22, 2026 By Mistral AI
- Mistral: Product Connect the dots: Build with built-in and custom MCPs in Studio Connect enterprise data to your AI applications with reusable connectors, direct tool calling, and human-in-the-loop approval controls. May 22, 2026 By Mistral AI
- Mistral: Product Workflows for work that runs the business Workflows is now in public preview. April 27, 2026 By Mistral AI
- Mistral: Research Speaking of Voxtral Voxtral TTS: A frontier, open-weights text-to-speech model that’s fast, instantly adaptable, and produces lifelike speech for voice agents. March 23, 2026 By Mistral AI
- Mistral: Product Introducing Forge Today, we’re introducing Forge, a system for enterprises to build frontier-grade AI models grounded in their proprietary knowledge. March 17, 2026 By Mistral AI
- Mistral: Research Introducing Mistral Small 4 March 16, 2026 By Mistral AI
- Mistral: Company Mistral AI partners with NVIDIA to accelerate open frontier models March 16, 2026 By Mistral AI
- Mistral: Research Leanstral: Open-Source foundation for trustworthy vibe-coding March 16, 2026 By Mistral AI
- Mistral: Solutions Rails testing on autopilot: Building an agent that writes what developers won't March 11, 2026 By Maxime Langelier & Mathis Grosmaitre - Applied AI - Proto team
- Mistral: Research Voxtral transcribes at the speed of sound. February 4, 2026 By Mistral AI
- Mistral: Product Terminally online Mistral Vibe. January 27, 2026 By Mistral AI
- Mistral: Engineering Heaps do lie: debugging a memory leak in vLLM. January 21, 2026 By Mathis Felardos
- Mistral: Research Introducing Mistral OCR 3 December 17, 2025 By Mistral AI
- Mistral: Research Introducing: Devstral 2 and Mistral Vibe CLI. December 9, 2025 By Mistral AI
- Mistral: Company Mistral AI - KI für Deutschland November 19, 2025 By Mistral AI
- Mistral: Product Introducing Mistral AI Studio. October 24, 2025 By Mistral AI
- Mistral: Company Mistral AI raises 1.7B€ to accelerate technological progress with AI September 9, 2025 By Mistral AI
- Mistral: Product Make Memory work for you. September 2, 2025 By Mistral AI
- Mistral: Product Le Chat. Custom MCP connectors. Memories. September 2, 2025 By Mistral AI
- Mistral: Solutions Unlocking the potential of vision language models on satellite imagery through fine-tuning August 1, 2025 By Mistral AI
- Mistral: Research Announcing Codestral 25.08 and the Complete Mistral Coding Stack for Enterprise July 30, 2025 By Mistral AI
- Mistral: Company Our contribution to a global environmental standard for AI July 22, 2025 By Mistral AI
- Mistral: Product Le Chat dives deep. July 17, 2025 By Mistral AI
- Mistral: Research Voxtral July 15, 2025 By Mistral AI
- Mistral: Research Upgrading agentic coding capabilities with the new Devstral models July 10, 2025 By Mistral AI
- Mistral: Solutions Announcing AI for Citizens July 3, 2025 By Mistral AI
- Mistral: Company Mistral Compute June 11, 2025 By Mistral AI
- Mistral: Research Magistral June 10, 2025 By Mistral AI
- Mistral: Product Introducing Mistral Code June 4, 2025 By Mistral AI
- Mistral: Research Codestral Embed May 28, 2025 By Mistral AI
- Mistral: Product Build AI agents with the Mistral Agents API May 27, 2025 By Mistral AI
- Mistral: Product Introducing Le Chat Enterprise May 7, 2025 By Mistral AI
- Mistral: Research Medium is the new large. May 7, 2025 By Mistral AI
- Mistral: Solutions Evaluating RAG with LLM as a Judge April 9, 2025 By Mistral AI Team
- Mistral: Research Mistral OCR March 6, 2025 By Mistral AI Team
- Mistral: Solutions Empowering product development with an agentic workflow March 4, 2025 By Mistral AI team
- Mistral: Research Mistral Saba February 17, 2025 By Mistral AI Team
- Mistral: Product The all new le Chat: Your AI assistant for life and work February 6, 2025 By Mistral AI Team
- Mistral: Product Purr-fectly informed January 16, 2025 By Mistral AI team
- Mistral: Research Codestral 25.01 January 13, 2025 By Mistral AI team
- Mistral: Product Mistral has entered the chat November 18, 2024 By Mistral AI team
- Mistral: Research Pixtral Large November 18, 2024 By Mistral AI team
- Mistral: Product Mistral Batch API November 7, 2024 By Mistral AI team
- Mistral: Research Mistral Moderation API November 7, 2024 By Mistral AI team
- Mistral: Research Un Ministral, des Ministraux October 16, 2024 By Mistral AI team
- Mistral: Product AI in abundance September 17, 2024 By Mistral AI team
- Mistral: Research Announcing Pixtral 12B September 17, 2024 By Mistral AI team
- Mistral: Research Build, tweak, repeat August 7, 2024 By Mistral AI team
- Mistral: Research Large Enough July 24, 2024 By Mistral AI team
- Mistral: Research My Tailor is Mistral June 5, 2024 By Mistral AI team
- Mistral: Company Mistral AI Fine-tuning Hackathon June 5, 2024 By Mistral AI team
- Mistral: Research The Mistral AI Non-Production License May 29, 2024 By Mistral AI team
- Mistral: Research Cheaper, Better, Faster, Stronger April 17, 2024 By Mistral AI team
- Mistral: Product Le Chat February 26, 2024 By Mistral AI team
- Mistral: Research Mixtral of experts December 11, 2023 By Mistral AI team
- Mistral: Product La Plateforme December 11, 2023 By Mistral AI team
- Mistral: Research Mistral 7B September 27, 2023 By Mistral AI team
- Mistral: Company Bringing open AI models to the frontier September 27, 2023 By Mistral AI team
- xAI: Aug 7, 2026 Imagine Image 2.0 Aug 7, 2026 Imagine Image 2.0 Precise image generation and editing, built for real creative work. Read More
🐦 Twitter/X Highlights
| Account | Tweet Summary |
|---|---|
| MistralAI | 🛡️Introducing Shieldstral, Mistral’s 3B open-weights model for content safety that can be deployed on-device 🧵 http://mistral.ai/news/shieldstral Post |
| AnthropicAI | The UK’s @AISecurityInst (AISI) has published a report on their recent cybersecurity evaluation of Anthropic’s Claude Mythos 5 and OpenAI’s GPT-5.6 Sol. The models attempted to complete an assignment in a setup where their normal safeguards were removed and they were deliberately given internet acce Post |
| AnthropicAI | In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Our post describes Post |
| OpenAI | After evaluating one of our upcoming models, Astra, we're treating it as our first "critical" model for cybersecurity under our Preparedness Framework. This is a scenario we've planned for, and we're putting additional controls in place to ensure Astra's further development happens safely and secure Post |
| GoogleDeepMind | We sat down with Apollo 2 to talk about what it's really like running on Gemini Robotics 2. 🤖 Post |
| GoogleDeepMind | Predicting cyclones accurately can help save lives - and every hour of lead time counts. Published in @Nature, our AI model WeatherNext achieves state-of-the-art accuracy in forecasting a storm’s track and intensity, giving us a critical extra 24 hours to prepare on average. 🧵 Post |
| MistralAI | Mistral is announcing an expanded global strategic partnership with @Microsoft to give enterprises and regulated industries frontier AI they can control. As Mistral is expanding its AI compute capacity in Europe, the companies are expanding their strategic partnership with Microsoft’s commitment to Post |
| karpathy | One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In these cases I like to lean back, switch to /voice and just ramble for like 10 minutes, total mess, anythi Post |
| mattshumer_ | For those who want to keep their skills while using Opus 5, here's a trick you can try (let me know how it goes): Write a loop (have a model do this!) that has Opus 5: - do a first update pass on your skills to make them better for Opus 5 - creates a suite of test tasks for each skill, like a benchm Post |
| DeepSeek_AI | 🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇 🔷 The official V4-Flash now natively supports the Responses API format and is ful Post |
Newsletter
- Ben's Bites: I’m trying something new - this email walks through one of my actual agent sessions and I’ll explain what’s happening along the way. The build or task I’m doing isn’t important. But I’m looking at how
- Import AI: Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe.
- Import AI: Self-sustaining and self-replicating AI viruses are here:…Open weight LLMs + a well-designed harness = a persistent, self-sufficient virus…AI researchers have built a prototype computer virus which us
- Latent Space: Editor’s note: I’m excited to welcomeShloktoour guest post roster! You may know Shlok from his excellent explorations (as an outsider — for an insider perspective see ourpodcast with OpenAI’s Akshay N
- Latent Space: On July 9th, OpenAI releasedChatGPT Work, their agent product for knowledge work. It was, by any measure,a busy launch: three new modelsacrossfourteen configurations, a consolidation of the ChatGPT an
- Latent Space: Three weeks in, Work (along with Codex) has reportedlycrossed 10 million users.
- Latent Space: Editor’s note: ChatGPT estimated to cross1B MAU in Juneand1B WAU this month.
- Latent Space: ChatandWorkcurrently sit side by side as separate modes inside ChatGPT, butGreg Brockman has confirmed that they will merge by the end of the year. Work, then, is not just a niche product for power us
- Latent Space: Work in its current form takes some decoding. It’s an amalgamation of ChatGPT (in chat form), Codex the app, Codex the harness, Codex the original cloud agent, ChatGPT agent, Atlas, OpenClaw, and more
- Latent Space: Lives in a cloud computer.Specifically,a beefy, isolated microVM: Pro accounts get 8 CPUs, 20GB of RAM, and a 64GB disk; Plus gets 14GB of RAM. Alongside the VM, Work gets amanaged Chrome servicethat
- Latent Space: Produces artifacts.Sheets, docs, and slides rendered in interactive viewers, plusSites: hosted web apps and dashboards it can build, share via URL, and keep updated.
- Latent Space: But then things get a little confusing. OpenAI did release a way tohand off a Codex task to a remote environment. Although this doesn’t work for me at the time of writing, I assume it eventually will,
- Interconnects: We’re expanding our open models coverage into standalone projects that let you go deeper on the state of the open model ecosystem. The new free sources of data are:
- Interconnects: TheArtifacts Hub— a curated view of the models trending on Hugging Face, highlighting inference tokens viaOpen Router, model intelligence viaArtificial Analysis, and ourtailored adoption metricsbuildi
- Interconnects: OurAdoption Dashboard— a living dashboard of download and derivative model numbers by geography and organization. This highlights the US-China gap and growing players in the open ecosystem.
- Interconnects: To date, our primary efforts on Interconnects have been release recaps for popular models likeKimi K3,GLM 5.2,DeepSeek R1, etc. and monthly round-ups of the open models that matter,Artifacts Log. We’r
- Interconnects: The Artifacts Hub right now covers 792 models released in the last two years, across the core text-focused language models and multimodal generative models. At Interconnects we follow the data of ever
- Interconnects: We built the Hub as a way to go deeper on this analysis in collaboration withProject VAIL— an AI verification startup who has been one of the most loyal fans of our open model curation.
- Interconnects: For the most popular models, the Hub let’s you quickly see how far behind the model was in terms of frontier intelligence based on Artificial Analysis’s Intelligence Index, compare Hugging Face and Op
- TheSequence: For most of software history, engineering capacity was easy to sketch on a whiteboard.
- tldr: APPLE VS. OPENAI: HOW SIRI AI STACKS UP AGAINST THE NEW CHATGPT (4 MINUTE READ)
- tldr: SAMSUNG REVEALS NEW 3D-MEMORY ROADMAP IN BID FOR AI TECH LEAD (2 MINUTE READ)
- tldr: TURN ONE GIANT AI-GENERATED PULL REQUEST TO A REVIEWABLE STACK (11 MINUTE READ)
- tldr: HOW CLOUDFLARE ENFORCES ENGINEERING STANDARDS USING AI (9 MINUTE READ)
- tldr: HUAWEI'S TOP SCIENTIST WARNS OF CHIP LIMIT NVIDIA WILL SOON FACE (4 MINUTE READ)
- tldr: MOST TECH REVOLUTIONS MADE WORK WORSE FOR EMPLOYEES. AI COULD BE THE EXCEPTION (7 MINUTE READ)
- tldr: WHAT CODEX ACTUALLY SENDS TO THE MODEL (10 MINUTE READ)
- tldr: OPENAI CALLS APPLE'S TRADE-SECRET SUIT ‘CARELESS' AND ‘ODDLY PERSONAL' (6 MINUTE READ)
- tldr: BENDING SPOONS TO BUY AIRTABLE FOR $1.28B (2 MINUTE READ)
- tldr: YOUR AGENT CAN NOW DEBUG WORKERS WITH LOCAL TRACING (3 MINUTE READ)
- tldr: ADD SECURITY CONTEXT TO OPERATIONAL INVESTIGATIONS WITH AWS DEVOPS AGENT AND WIZ (7 MINUTE READ)
- tldr: TENCENTDB AGENT MEMORY (GITHUB REPO)
- tldr: FEATURE FLAG ORCHESTRATION WITH AWS DEVOPS AGENT AND LAUNCHDARKLY (8 MINUTE READ)
- tldr: INTRODUCING THE WARP AGENT CLI: A CLI CODING AGENT THAT DOES WHAT OTHERS CAN'T (8 MINUTE READ)
- tldr: ANNOUNCING CLOUDFLARE WALLETS: THE PROGRAMMABLE WALLET FOR THE AGENTIC INTERNET (9 MINUTE READ)
- tldr: KEEP GOING, BRO. YOU'VE GOT THIS!” A DATA-DRIVEN LOOK AT HOW ADVERSARIES ARE WEAPONIZING AI (41 MINUTE READ)
- tldr: WHAT ACTUALLY KEEPS AN AI BENCHMARK USEFUL? SCALE (5 MINUTE READ)
- tldr: APPLE MOVES FOR PRELIMINARY INJUNCTION IN OPENAI TRADE SECRETS LAWSUIT (1 MINUTE READ)
- tldr: GOOGLE PAUSES AI SATELLITE IMAGES, AFTER FEARS OF DEEPFAKES IN THE SKY (4 MINUTE READ)
- tldr: AI PRESENTATION BUILDER (WEBSITE)
- tldr: WHAT DOES THE BIGGEST AI COPYRIGHT PAYOUT IN HISTORY ACTUALLY CHANGE FOR CREATIVES? (4 MINUTE READ)
- tldr: TIME HAS STARTED SERVING ADS TO AI AGENTS (6 MINUTE READ)
- tldr: INTRODUCING ATLAS, YOUR AI COWORKER (7 MINUTE READ)
- tldr: INTRODUCING WEB SEARCH ON AMAZON BEDROCK FOR FOUNDATION MODEL GROUNDING (10 MINUTE READ)
- tldr: GOOGLE IS WORKING ON PLUGINS FOR GEMINI ENTERPRISE (2 MINUTE READ)
- tldr: THE ENTERPRISE AI STRATEGY THAT OUTLASTS ANY SINGLE MODEL (5 MINUTE READ)
- tldr: SIXB: THE OPERATING LAYER FOR ENTERPRISE AI (5 MINUTE READ)
- rundown-ai: Credit to reader ^^Erich Archer^^. How do you use AI? Tell us ^^here^^.^
- rundown-ai: Anthropic and OpenAI agents went rogue again
- rundown-ai: Why it matters: While the models here were deliberately stripped of their guardrails, these cases — and the ones before them — show that agents chasing a goal will try to reach past their limits, bypassing restrictions a
- ... plus 23 more newsletter-only items in raw data