🔴 High Significance

Model Releases

🔴 🤗 moonshotai/Kimi-K3 (1,388,105 downloads) — score 95 Sources: huggingface_models

Author: | Downloads: 1,388,105 | Likes: 10340

🔴 🧡 DeepSeek V4 Flash 0731 — score 92 Sources: hackernews

🔴 💬 This is why the vast majority aren't taking any "this new model is dangerous" messages seriously. They've cried wolf FAR too many times. They could literally announce that a nuclear war caused by AI is 24 hours away and many wouldn't bat an eye — score 91 Sources: reddit/r/singularity · reddit/r/OpenAI

🔴 💬 Got job as Director of AI and Systems development self-taught — score 90 Sources: reddit/r/LocalLLaMA

Hey everyone, I just wanted to share my journey here for some motivation. Three years ago, I saw the sudden spike in AI and realized it was the future of tech. My goal at the time was to be an indie game dev, and seeing that AI could write basic code, I told myself I needed to master it or risk bein

🔴 🤗 MiniMaxAI/MiniMax-H3 (26,693 downloads) — score 85 Sources: huggingface_models

Author: | Downloads: 26,693 | Likes: 3098

Omitted 5 additional model releases items from the main section; see raw data and source-specific sections below.

Developer Tools

🔴 🐙 PrimeIntellect-ai/prime-agent — A self-improving RLM agent for coding workflows and long-running autonomous tasks. — score 99 Sources: github_trending

A self-improving RLM agent for coding workflows and long-running autonomous tasks.

🔴 🐙 virgiliojr94/book-to-skill — Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work. — score 97 Sources: github_trending

Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.

🔴 💬 Learned the term "context poisoning" today and now I can't stop noticing it — score 95 Sources: reddit/r/artificial

Someone explained this to me in a comment thread and it's been rattling around in my head since. The idea: in a long conversation, if the model says something wrong and you correct it, that correction doesn't necessarily erase the wrong idea's influence. The tokens around the mistake, including the

🔴 💬 I created an AI agent that sells your services and products for you — score 94 Sources: reddit/r/AIAgents

So I wondered why everyone was so hyped about Vibe-coding when the real struggle for businesses is distribution, aka finding people who want to use your product or service and give you money for it. That is why I created https://vonto.ai, enter your domain, a brief description an

🔴 🐙 google/skills — Agent Skills for Google products and technologies — score 94 Sources: github_trending

Agent Skills for Google products and technologies

Omitted 11 additional developer tools items from the main section; see raw data and source-specific sections below.

Infrastructure & Compute

🔴 🐙 cloudflare/computer — Give your agent a computer 👾 — score 98 Sources: github_trending

Give your agent a computer 👾

Research Papers

🔴 🤗 Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval — score 95 Sources: huggingface

Short segments of perceived speech can be retrieved from non-invasive magnetoencephalographic (MEG) recordings by deep networks trained with a CLIP-style objective against wav2vec 2.0 audio embeddings. Yet their weights do not map onto electrophysiological quantities, and it remains unclear which sp

🔴 🤗 DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces — score 85 Sources: huggingface

Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured files, long documents, and multimedia. Existing benchmarks largely isolate structured querying, retrieval, or open-ended analysis, leaving heterogeneous

🔴 🤗 Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay — score 75 Sources: huggingface

Computer-use agents pay full frontier inference to re-derive routines their user has already performed, because an agent's memory today records what the user said, not what the user did. We compile passively captured screen activity into agent memory with a deterministic, zero-model pipeline: it seg

Other Signals

🔴 💬 2027 Memory Capacity Is Reportedly Sold Out — score 97 Sources: reddit/r/LocalLLaMA

🔴 💬 Your scientists were so... uh... — score 94 Sources: reddit/r/OpenAI

🔴 💬 NeurIPS OpenReview Modified Timeline [D] — score 88 Sources: reddit/r/MachineLearning

I just wonder if “modified: 04 Aug 2026 at 22:21” time part in the Reviewer’s official review still remain unchanged, then is that mean reviewer not add final justification, or any change of their rating?

🔴 💬 NeurIPS AI Assisted Review authors/reviewers? [D] — score 88 Sources: reddit/r/MachineLearning

Out of curiosity, if you were a reviewer or author, how did the review period go? For me, it was weird, because I gave reviews with specific details (what specifically could have been better, how to fix it), but realized other reviewers gave similar superficial reviews. Even the paper which was a co

🔴 💬 Chinese LLMs dominate this week's top charts — score 85 Sources: reddit/r/artificial

Source: https://openrouter.ai/rankings

Omitted 4 additional other signals items from the main section; see raw data and source-specific sections below.

🟡 Notable

Model Releases

🟡 💬 GPT 5.6 Sol and Fable 5 settle a 25 year old problem in wireless communication theory — score 69 Sources: reddit/r/singularity

🟡 🤗 DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF (2,345,190 downloads) — score 65 Sources: huggingface_models

Author: | Downloads: 2,345,190 | Likes: 1760

🟡 ✉️ Editor’s note: I’m excited to welcomeShloktoour guest post roster! You may know Shlok from his excellent explorations (as an outsider — for an insider perspective see ourpodcast with OpenAI’s Akshay N — score 65 Sources: newsletter/Latent Space

Editor’s note: I’m excited to welcomeShloktoour guest post roster! You may know Shlok from his excellent explorations (as an outsider — for an insider perspective see ourpodcast with OpenAI’s Akshay Nathan. Already one of our most popular episodes of the year!) ofleading AI Lab memory systems, which

🟡 ✉️ On July 9th, OpenAI releasedChatGPT Work, their agent product for knowledge work. It was, by any measure,a busy launch: three new modelsacrossfourteen configurations, a consolidation of the ChatGPT an — score 65 Sources: newsletter/Latent Space

On July 9th, OpenAI releasedChatGPT Work, their agent product for knowledge work. It was, by any measure,a busy launch: three new modelsacrossfourteen configurations, a consolidation of the ChatGPT and Codex desktop apps, andcloud agents brought to the mainstreamin their most accessible form yet.

🟡 ✉️ Editor’s note: ChatGPT estimated to cross1B MAU in Juneand1B WAU this month. — score 65 Sources: newsletter/Latent Space

Omitted 115 additional model releases items from the main section; see raw data and source-specific sections below.

Developer Tools

🟡 🐙 rivet-dev/rivet — Rivet Actors are the primitive for stateful workloads. Built for AI agents, collaborative apps, and durable execution. — score 69 Sources: github_trending

Rivet Actors are the primitive for stateful workloads. Built for AI agents, collaborative apps, and durable execution.

🟡 🐙 tirth8205/code-review-graph — Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows. — score 68 Sources: github_trending

Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.

🟡 🐙 AgriciDaniel/claude-seo — Universal SEO skill for Claude Code. 25 sub-skills + 18 sub-agents covering technical SEO, E-E-A-T, schema, GEO/AEO, backlinks, local SEO, maps intelligence, semantic clustering, e-commerce SEO, international SEO, Google APIs, and PDF/Excel reporting. Optional DataForSEO, Firecrawl, and Banana extensions. — score 66 Sources: github_trending

Universal SEO skill for Claude Code. 25 sub-skills + 18 sub-agents covering technical SEO, E-E-A-T, schema, GEO/AEO, backlinks, local SEO, maps intelligence, semantic clustering, e-commerce SEO, international SEO, Google APIs, and PDF/Excel reporting. Optional DataForSEO, Firecrawl, and Banana exten

🟡 🐙 tashfeenahmed/freellmapi — OpenAI-compatible proxy that stacks the free tiers of 28 LLM providers (~4B tokens/month) behind one /v1 endpoint — plus any custom OpenAI-compatible endpoint. Smart routing, automatic failover, encrypted keys. Personal experimentation only. — score 65 Sources: github_trending

OpenAI-compatible proxy that stacks the free tiers of 28 LLM providers (~4B tokens/month) behind one /v1 endpoint — plus any custom OpenAI-compatible endpoint. Smart routing, automatic failover, encrypted keys. Personal experimentation only.

🟡 ✉️ I’m trying something new - this email walks through one of my actual agent sessions and I’ll explain what’s happening along the way. The build or task I’m doing isn’t important. But I’m looking at how — score 65 Sources: newsletter/Ben's Bites

I’m trying something new - this email walks through one of my actual agent sessions and I’ll explain what’s happening along the way. The build or task I’m doing isn’t important. But I’m looking at how I could be using agents more effectively.

Omitted 40 additional developer tools items from the main section; see raw data and source-specific sections below.

Infrastructure & Compute

🟡 ✉️ Self-sustaining and self-replicating AI viruses are here:…Open weight LLMs + a well-designed harness = a persistent, self-sufficient virus…AI researchers have built a prototype computer virus which us — score 65 Sources: newsletter/Import AI

Self-sustaining and self-replicating AI viruses are here:…Open weight LLMs + a well-designed harness = a persistent, self-sufficient virus…AI researchers have built a prototype computer virus which uses AI models to compromise computers, then uses their underlying GPU resources to run inference, let

🟡 ✉️ TheArtifacts Hub— a curated view of the models trending on Hugging Face, highlighting inference tokens viaOpen Router, model intelligence viaArtificial Analysis, and ourtailored adoption metricsbuildi — score 65 Sources: newsletter/Interconnects

TheArtifacts Hub— a curated view of the models trending on Hugging Face, highlighting inference tokens viaOpen Router, model intelligence viaArtificial Analysis, and ourtailored adoption metricsbuilding on top ofHugging Face’s data.

🟡 ✉️ HUAWEI'S TOP SCIENTIST WARNS OF CHIP LIMIT NVIDIA WILL SOON FACE (4 MINUTE READ) — score 65 Sources: newsletter/tldr

🟡 ✉️ AI infrastructure startup Volta, founded earlier this year, emerged from stealth with a reported $10B compute deal with Anthropic and funding at a $2.4B valuation. — score 65 Sources: newsletter/rundown-ai

🟡 🐙 vllm-project/vllm — A high-throughput and memory-efficient inference and serving engine for LLMs — score 61 Sources: github_trending

A high-throughput and memory-efficient inference and serving engine for LLMs

Omitted 3 additional infrastructure & compute items from the main section; see raw data and source-specific sections below.

Business & Funding

🟡 ✉️ BENDING SPOONS TO BUY AIRTABLE FOR $1.28B (2 MINUTE READ) — score 65 Sources: newsletter/tldr

🟡 🏢 Responding to the next frontier of critical cyber capabilities — score 50 Sources: lab_blog/OpenAI

OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.

🟡 🏢 Taming Outlier Tokens in Diffusion Transformers — score 50 Sources: lab_blog/Apple ML

We study outlier tokens in Diffusion Transformers (DiTs) for image generation. Prior work has shown that Vision Transformers (ViTs) can produce a small number of high-norm tokens that attract disproportionate attention while carrying limited local information, but their role in generative models rem

Enterprise Adoption

🟡 ✉️ We’re expanding our open models coverage into standalone projects that let you go deeper on the state of the open model ecosystem. The new free sources of data are: — score 65 Sources: newsletter/Interconnects

Our other project is much lighter weight, but far overdue. Ever since we wrote The ATOM Project, we’ve been seeing the US-vs-China model adoption plot on a recurring basis in the AI ecosystem. We’d update the plot from time to time, but not enough. Now, we’re making the crucial data for that report

🟡 ✉️ OurAdoption Dashboard— a living dashboard of download and derivative model numbers by geography and organization. This highlights the US-China gap and growing players in the open ecosystem. — score 65 Sources: newsletter/Interconnects

🟡 ✉️ For the most popular models, the Hub let’s you quickly see how far behind the model was in terms of frontier intelligence based on Artificial Analysis’s Intelligence Index, compare Hugging Face and Op — score 65 Sources: newsletter/Interconnects

For the most popular models, the Hub let’s you quickly see how far behind the model was in terms of frontier intelligence based on Artificial Analysis’s Intelligence Index, compare Hugging Face and Open Router adoption to similar models, glance at relative adoption metric (RAM) scores for time-size

🟡 ✉️ THE ENTERPRISE AI STRATEGY THAT OUTLASTS ANY SINGLE MODEL (5 MINUTE READ) — score 65 Sources: newsletter/tldr

🟡 ✉️ SIXB: THE OPERATING LAYER FOR ENTERPRISE AI (5 MINUTE READ) — score 65 Sources: newsletter/tldr

Omitted 1 additional enterprise adoption items from the main section; see raw data and source-specific sections below.

Research Papers

🟡 🤗 KVAE: Family of Tokenizers for Multimodal Generative Models — score 65 Sources: huggingface

Latent diffusion modeling (LDM), a prominent paradigm, utilizes tokenizers to map input signal to compressed representation. This dependency positions tokenizer as an integral part of generation process itself, since it affects learning speed, quality of synthesized samples and lay foundation for la

🟡 🤗 MameLoshnLM: Yiddish Language Model and Evaluation Benchmark — score 55 Sources: huggingface

We present MameLoshnLM, the first open-source 8B-parameter language model built specifically for Yiddish. Despite Yiddish's rich textual tradition, its limited digital presence and the scarcity of reliable evaluation resources have constrained progress in Yiddish language modeling. Existing multilin

🟡 🤗 Continual Learning in Transition — score 45 Sources: huggingface

Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g., training strategies, architectural designs, and weight adaptation. However, emerging paradigms are reshaping the scope of CL beyond this traditional m

Other Signals

🟡 💬 ByteDance trains massive AI model in bid to rival Anthropic — score 65 Sources: reddit/r/artificial

🟡 ✉️ Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe. — score 65 Sources: newsletter/Import AI

“AI could help create a dramatically better future, but that outcome is not guaranteed. The world’s leading AI companies believe they could be close to automating AI research. It is hard to predict exactly how much this will accelerate AI progress, but there is a real risk that capability developmen

🟡 ✉️ We built the Hub as a way to go deeper on this analysis in collaboration withProject VAIL— an AI verification startup who has been one of the most loyal fans of our open model curation. — score 65 Sources: newsletter/Interconnects

🟡 ✉️ SAMSUNG REVEALS NEW 3D-MEMORY ROADMAP IN BID FOR AI TECH LEAD (2 MINUTE READ) — score 65 Sources: newsletter/tldr

🟡 ✉️ TURN ONE GIANT AI-GENERATED PULL REQUEST TO A REVIEWABLE STACK (11 MINUTE READ) — score 65 Sources: newsletter/tldr

Omitted 31 additional other signals items from the main section; see raw data and source-specific sections below.

🟢 Incremental

Model Releases

🟢 🤗 larryvrh/MiniMax-H3-Turbo-Lora (0 downloads) — score 35 Sources: huggingface_models

Author: | Downloads: 0 | Likes: 482

🟢 💬 Ur shipping so many bugs! No amount of instructions, memory, engineering standards, or repo structure will stop this. Which is why you have to spot and fix it. Claude Opus 5 on Max. — score 33 Sources: reddit/r/AIAgents

Every complex multi-phase task i give it I find some variation of the issues bellow. Over the past month, I’ve been setting up and fine-tuning a deterministic validation architecture. is currently set up to find things like. * Command/API contract drift and accidental state-shape changes. * Invalid

🟢 🤗 LiquidAI/LFM2.5-2.6B (81,522 downloads) — score 25 Sources: huggingface_models

Author: | Downloads: 81,522 | Likes: 415

🟢 🧡 Lost my phone at the office. Claude suggested tracking Bluetooth signal strength — score 25 Sources: hackernews

🟢 💬 DeepSeek V4 Flash 0731 ARC-AGI-1 and 2 — score 19 Sources: reddit/r/singularity

Interesting to see the latest version buck the trend of increasing cost for increased performance based on the reasoning level. Both for ARC-AGI-1 and 2 as well Just thought I'd post since I thought it was interesting, don't recall this being the case for any other model so far. Also insane just to

Omitted 6 additional model releases items from the main section; see raw data and source-specific sections below.

Developer Tools

🟢 🐙 open-metadata/OpenMetadata — The Open Context Layer for Data and AI , OpenMetadata is the open platform for building trusted data context and business semantics for humans, AI assistants, and agents. — score 39 Sources: github_trending

The Open Context Layer for Data and AI , OpenMetadata is the open platform for building trusted data context and business semantics for humans, AI assistants, and agents.

🟢 🐙 jordanrendric/claude-video-vision — Give Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis — score 38 Sources: github_trending

Give Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis

🟢 🐙 datawhalechina/self-llm — 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程 — score 36 Sources: github_trending

《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程

🟢 🐙 NirDiamant/GenAI_Agents — 50+ tutorials and implementations for Generative AI Agent techniques, from basic conversational bots to complex multi-agent systems. — score 34 Sources: github_trending

50+ tutorials and implementations for Generative AI Agent techniques, from basic conversational bots to complex multi-agent systems.

🟢 🐙 karpathy/micrograd — A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API — score 34 Sources: github_trending

A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API

Omitted 15 additional developer tools items from the main section; see raw data and source-specific sections below.

Infrastructure & Compute

🟢 🐙 higgsfield-ai/higgsfield — Fault-tolerant, highly scalable GPU orchestration, and a machine learning framework designed for training models with billions to trillions of parameters — score 26 Sources: github_trending

Fault-tolerant, highly scalable GPU orchestration, and a machine learning framework designed for training models with billions to trillions of parameters

🟢 🐙 NVIDIA-NeMo/Nemotron — Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models — score 22 Sources: github_trending

Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models

🟢 🐙 openai/CLIP — CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image — score 10 Sources: github_trending

CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image

🟢 💬 Synthesizing and formally verifying a SWAR bit-hack for INT4 dot products using Z3 and Lean 4 [P] — score 6 Sources: reddit/r/MachineLearning

INT4 quantization is ubiquitous in ML right now, but evaluating dot products on hardware without native SIMD/vector instructions (like WebAssembly or older ARM chips) usually requires slow sequential loops. A classic workaround is SWAR (SIMD Within A Register), but deriving the bitwise operations by

🟢 🐙 CliMA/ClimaAtmos.jl — GPU-capable global atmosphere model of the CliMA Earth System Model, designed for calibration with data assimilation and machine learning — score 3 Sources: github_trending

GPU-capable global atmosphere model of the CliMA Earth System Model, designed for calibration with data assimilation and machine learning

Business & Funding

🟢 💬 Evaluation metrics - [D] — score 31 Sources: reddit/r/MachineLearning

I wanted to ask about the selection of evaluation metrics. In which scenario we use ROC-AUC score and in which scenario we use f1 score as an evaluation metric in a classification problem to define the model's performance on a specific dataset.

Research Papers

🟢 🤗 Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills — score 30 Sources: huggingface

Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that write and refine their own executable skills as code. This survey organises the field around that axis of weights versus skills. Its central analytic

🟢 🤗 FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds — score 30 Sources: huggingface

World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this emerging landscape, Joint Embedding Predictive Architectures (JEPA) offer a particularly compelling direction. We study a largely unexplored regime: po

🟢 🤗 Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation — score 15 Sources: huggingface

Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fundamentally different optimization strategies. We introduce Task-Conditional Flow Matching (TCFM), a multilingual embedding adaptation framework that se

🟢 🤗 GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization — score 5 Sources: huggingface

Selecting a complete 3D object from a reconstructed scene with minimal user effort is essential for practical scene editing and embodied interaction. Existing 3DGS-based methods either retrain the Gaussian representation to embed per-object labels, or build dense multi-view SAM observations, both re

Other Signals

🟢 💬 Companies seeing AI returns had their data and governance sorted first, per PwC's 4,454-CEO survey — score 35 Sources: reddit/r/artificial

🟢 💬 Is Microsoft-Phi dead? — score 33 Sources: reddit/r/LocalLLaMA

Phi was one of my favorite models with a bit of a mixed reputation with some claiming it's benchmaxxed and others seeing its potential and usecases. I was a big fan of Phi but the last major release was in december 2024 with every other release being a Phi 4 iteration (like Phi 4-reasoning-vision wh

🟢 💬 Tesla V100 Qwen3.6 27B Performance — score 33 Sources: reddit/r/LocalLLaMA

Looking for V100 users to share your config and it's performance. GPU: Tesla V100 PCIE 32Gb Qwen3.6 27B Q4_K_M + Q8_0 MTP 128K context length Pi coding agent llama.cpp model preset: [*] spec-default = 1 ctx-size = 131072 mmap = 1 kv-unified = 1 n-gpu-layers = 999 threads = 18 prio = 3 seed = 3407

🟢 💬 ICDE Results [D] — score 31 Sources: reddit/r/MachineLearning

Hello! Let's use this thread to discuss ICDE results which should be coming out shortly today (hopefully).

🟢 💬 We are in a bubble sell everything — score 31 Sources: reddit/r/singularity

Omitted 4 additional other signals items from the main section; see raw data and source-specific sections below.

📊 Cross-Source Signals

Items that appeared on 3+ sources today:

RepoDescriptionStars TodayLanguage
PrimeIntellect-ai/prime-agentA self-improving RLM agent for coding workflows and long-running autonomous tasks.2483typescript
cloudflare/computerGive your agent a computer 👾1005typescript
virgiliojr94/book-to-skillTurn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.660python
google/skillsAgent Skills for Google products and technologies481python
anomalyco/opencodeThe open source coding agent.379typescript
farion1231/cc-switchA cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official website: ccswitch.io267rust
huangruiteng/loopxLightweight loop engineering state kernel for long-running AI agent teams. Agent-loop agnostic across Codex, Claude Code, and other coding agents, with durable goals, quota-aware auto-wake, executable todos, evidence logs, and verifiable handoffs.244python
Significant-Gravitas/AutoGPTAutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.240python
can1357/oh-my-pi⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and more232typescript
heygen-com/hyperframesWrite HTML. Render video. Built for agents.158typescript

📄 New Papers

TitleCategoryHotnessLink
Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrievalresearch_paper61Open
DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspacesresearch_paper27Open
Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replayresearch_paper24Open
KVAE: Family of Tokenizers for Multimodal Generative Modelsresearch_paper18Open
MameLoshnLM: Yiddish Language Model and Evaluation Benchmarkresearch_paper17Open
Continual Learning in Transitionresearch_paper16Open
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skillsresearch_paper8Open
FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worldsresearch_paper8Open
Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptationresearch_paper7Open
GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimizationresearch_paper6Open

🏢 Lab Blog Posts

🐦 Twitter/X Highlights

AccountTweet Summary
MistralAI🛡️Introducing Shieldstral, Mistral’s 3B open-weights model for content safety that can be deployed on-device 🧵 http://mistral.ai/news/shieldstral Post
AnthropicAIThe UK’s @AISecurityInst (AISI) has published a report on their recent cybersecurity evaluation of Anthropic’s Claude Mythos 5 and OpenAI’s GPT-5.6 Sol. The models attempted to complete an assignment in a setup where their normal safeguards were removed and they were deliberately given internet acce Post
AnthropicAIIn a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Our post describes Post
OpenAIAfter evaluating one of our upcoming models, Astra, we're treating it as our first "critical" model for cybersecurity under our Preparedness Framework. This is a scenario we've planned for, and we're putting additional controls in place to ensure Astra's further development happens safely and secure Post
GoogleDeepMindWe sat down with Apollo 2 to talk about what it's really like running on Gemini Robotics 2. 🤖 Post
GoogleDeepMindPredicting cyclones accurately can help save lives - and every hour of lead time counts. Published in @Nature, our AI model WeatherNext achieves state-of-the-art accuracy in forecasting a storm’s track and intensity, giving us a critical extra 24 hours to prepare on average. 🧵 Post
MistralAIMistral is announcing an expanded global strategic partnership with @Microsoft to give enterprises and regulated industries frontier AI they can control. As Mistral is expanding its AI compute capacity in Europe, the companies are expanding their strategic partnership with Microsoft’s commitment to Post
karpathyOne pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In these cases I like to lean back, switch to /voice and just ramble for like 10 minutes, total mess, anythi Post
mattshumer_For those who want to keep their skills while using Opus 5, here's a trick you can try (let me know how it goes): Write a loop (have a model do this!) that has Opus 5: - do a first update pass on your skills to make them better for Opus 5 - creates a suite of test tasks for each skill, like a benchm Post
DeepSeek_AI🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇 🔷 The official V4-Flash now natively supports the Responses API format and is ful Post

Newsletter