πŸ”΄ High Significance

Model Releases

πŸ”΄ πŸ’¬ Non Us Ally should be afraid. β€” score 82 Sources: reddit/r/LocalLLaMA

Spyware-like code in Claude Code that covertly targets Chinese users.

Developer Tools

πŸ”΄ πŸ’¬ Looking for feedback from people building AI agents that use browsers β€” score 94 Sources: reddit/r/AIAgents

I'm interested in getting technical feedback from people building AI agents that interact with websites. One issue I kept running into was Playwright sessions getting identified by anti-bot systems, so I started experimenting with a Firefox fork that modifies fingerprinting at the browser level inst

πŸ”΄ πŸ’¬ How do you use AI agent working for social media data analysis? β€” score 83 Sources: reddit/r/AIAgents

Just want to know which ai agent can do or how to use these diverse agents to do analysis on social media platforms and give insigts on content creation?

πŸ”΄ πŸ™ virgiliojr94/book-to-skill β€” Turn any technical book PDF into a Claude Code skill β€” ready to study, reference, and use while you work. β€” score 76 Sources: github_trending

Turn any technical book PDF into a Claude Code skill β€” ready to study, reference, and use while you work.

πŸ”΄ πŸ™ open-webui/open-webui β€” User-friendly AI Interface (Supports Ollama, OpenAI API, ...) β€” score 70 Sources: github_trending

User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

Infrastructure & Compute

πŸ”΄ πŸ™ allenai/olmocr β€” Toolkit for linearizing PDFs for LLM datasets/training β€” score 80 Sources: github_trending

Toolkit for linearizing PDFs for LLM datasets/training

Business & Funding

πŸ”΄ πŸ’¬ On July 1, 2026, arXiv will spin out from Cornell University, its home for the past 25 years, to become an independent nonprofit organization. Major funding support from Simons Foundation and Schmidt Sciences. Ditching the red for their website. [N] β€” score 94 Sources: reddit/r/MachineLearning

arXiv’s next chapter: Updates on our spin out from Cornell University: https://blog.arxiv.org/2026/06/30/arxivs-next-chapter/

Other Signals

πŸ”΄ πŸ’¬ The gap between closed and open models might be much smaller than commonly assumed, because we don’t know what closed model providers do in addition to model inference β€” score 96 Sources: reddit/r/LocalLLaMA

When Claude dominates GLM-5.2 in benchmarks, it’s usually assumed that Anthropic has superior model architectures, superior training pipelines, and other advanced machine learning techniques that make their models better than the competition. But actually, this doesn’t follow. Because the benchmarks

πŸ”΄ πŸ’¬ Couldn't hold back β€” score 89 Sources: reddit/r/LocalLLaMA

Had been waiting for months and the cards finally got delivered today. No one at my workplace was excited, maybe because no one cares for AI stuff that i work on. But I just wanted to share it with you guys. Can't wait to build the server and start working on them.

πŸ”΄ πŸ’¬ SWE-rebench leaderboard update: GLM-5.2, Qwen3.6-27B, Qwen3.6-35B-A3B, Gemma 4 31B and more + improved UI β€” score 75 Sources: reddit/r/LocalLLaMA

Hi all, We made several updates to the SWE-rebench leaderboard: added new models, refreshed recent results, and reworked the leaderboard UI to make results easier to read, compare, and understand. New Models: * Claude Opus 4.8 xhigh: 56.5% β€” 2.48M tokens * GLM-5.2: 51.1% β€” 2.62M tokens * Gemini 3.5

🟑 Notable

Model Releases

🟑 πŸ’¬ Open Models - June 2026 β€” score 61 Sources: reddit/r/LocalLLaMA

After overwhelming April, OK May, here's June. Yeah, Graph has only less items. Because we got other items here last month. Finetunes: * Nex-N2 * Ornith-1.0 * Agents-A1 * Holo3.1 * Tmax-27b *

🟑 𝕏 @GoogleDeepMind: We’re shipping 2 major releases: πŸ”˜ Nano Banana 2 Lite: our fastest and cheapest Gemini Image model πŸ”˜ Gemini Omni Flash: now available via the Gemini API and in @GoogleAIStudio to help developers gene β€” score 60 Sources: twitter_rss

We’re shipping 2 major releases: πŸ”˜ Nano Banana 2 Lite: our fastest and cheapest Gemini Image model πŸ”˜ Gemini Omni Flash: now available via the Gemini API and in @GoogleAIStudio to help developers generate and edit high-quality videos.

🟑 𝕏 @xai: Introducing Voice Agent Builder: a no-code platform to create human-like voice agents with Grok Voice. Available today at $0.05 / min. http://x.ai/voice β€” score 60 Sources: twitter_rss

Introducing Voice Agent Builder: a no-code platform to create human-like voice agents with Grok Voice. Available today at $0.05 / min. http://x.ai/voice

🟑 πŸ’¬ New PyMuPDF release, supports Markdown [N] β€” score 56 Sources: reddit/r/MachineLearning

https://pymupdf.io/blog/markdown-in-pymupdf-1-28 PyMuPDF 1.28 release, introduces Markdown as a first class document in PyMuPDF. Seems useful for a variety of workflows. You can create PDFs from Markdown text with control over appearance using CSS

🟑 🏒 Redeploying Fable 5 Announcements Jun 30, 2026 Fable 5 returns globally July 1. We're also proposing an industry-wide framework for scoring jailbreak severity, together with Amazon, Microsoft, Google, and other Glasswing partners. β€” score 50 Sources: lab_blog/Anthropic

Product Jun 30, 2026 Introducing Claude Sonnet 5 Sonnet 5 delivers frontier performance across coding, agents, and professional work at scale. Announcements Jun 30, 2026 Claude Science, an AI workbench for scientists, is now available Claude Science is a customizable app that integrates the tools an

Omitted 2 additional model releases items from the main section; see raw data and source-specific sections below.

Developer Tools

🟑 πŸ’¬ P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P] β€” score 69 Sources: reddit/r/MachineLearning

We just open-sourced MOTHRAG, a multi-hop RAG framework that skips the knowledge graph entirely. We kept hitting the same wall building multi-hop RAG: the systems with the best accuracy (GraphRAG, HippoRAG, RAPTOR) all lean on a knowledge graph built offline, and that’s great numbers, until the mome

🟑 πŸ™ HKUDS/VideoAgent β€” "VideoAgent: All-in-One Agentic Framework for Video Understanding, Editing, and Remaking" β€” score 66 Sources: github_trending

"VideoAgent: All-in-One Agentic Framework for Video Understanding, Editing, and Remaking"

🟑 πŸ’¬ how many parameters can i run in my machine to make it as an ai agent assistant in my system β€” score 61 Sources: reddit/r/AIAgents

i wanted to do is to have an ai assistant that can do basic to moderate stuff, like to read whats on folders, create and structure them, maybe write some basic to intermediate code and commit and push it to github, maybe some shell commands. i am thinking of building another 2 or 3 agents to to act

🟑 πŸ’¬ I extended Gemma4-31B to 44B (88 layers) β€” since Google won't give us anything bigger than 31B β€” score 54 Sources: reddit/r/LocalLLaMA

I've been just sit on this thread for a while now, both as a reader and occasional poster, so I figured it was finally time to share something I've been working on last weekends. Google hasn't shipped a dense Gemma4 bigger than 31B, so I decided to just build one myself. Heads up though β€” I'm not a

🟑 πŸ™ getmaxun/maxun β€” πŸ”₯ The open-source no-code platform for web scraping, crawling, search and AI data extraction β€’ Turn websites into structured APIs in minutes πŸ”₯ β€” score 51 Sources: github_trending

πŸ”₯ The open-source no-code platform for web scraping, crawling, search and AI data extraction β€’ Turn websites into structured APIs in minutes πŸ”₯

Omitted 2 additional developer tools items from the main section; see raw data and source-specific sections below.

Business & Funding

🟑 πŸ’¬ ACL ARR May 2026[D] β€” score 44 Sources: reddit/r/MachineLearning

Hi everyone. Do the ACL arr may 2026 reviews come out of July 2nd or do they come out on July 7 th?? How much does one need to get into Main or Findings? I am a bit new to this. Thanks a lot folks.

Research Papers

🟑 πŸ€— TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning β€” score 68 Sources: huggingface Β· arxiv/cs.AI

Agentic reinforcement learning requires assigning credit to environment-facing actions such as searches, clicks, edits, navigation commands, and object interactions. Standard GRPO uses the final verifier outcome as a uniform advantage over all action tokens. This outcome signal is useful but structu

Other Signals

🟑 πŸ’¬ Deepseek V4 Flash 2, 3 and 4 bits GGUFs β€” score 68 Sources: reddit/r/LocalLLaMA

🟑 𝕏 @AnthropicAI: Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and blo β€” score 50 Sources: twitter_rss

Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and block more cybersecurity tasks. In the near term, some routine tasks like coding and debugging will fal

🟒 Incremental

Model Releases

🟒 πŸ’¬ End of an Agony. Real production service that uses LLM to earn money my team had made and now we are so happy that it will die. Here are some of my final "experiences". β€” score 39 Sources: reddit/r/LocalLLaMA

Hello everyone. I had posted in this sub about making a production service about 8 months ago. Here the link of my previous post . The idea was the same. We wanted to make a real production ser

🟒 πŸ’¬ Looking for real experiences with AI chat agents in the NDIS space β€” score 39 Sources: reddit/r/AIAgents

Quick question for anyone working with an NDIS provider. We're seriously considering adding an ai chat agent this quarter because enquiries keep growing and our admin team is stretched. Nothing crazy, we just want people to get answers without waiting forever. I've narrowed it down to AusGPT and Exp

🟒 πŸ’¬ gemma-4-31B on Cerebras is better than ChatGPT voice mode β€” score 29 Sources: reddit/r/LocalLLaMA

open models will win on inference too πŸš€

🟒 πŸ’¬ A system-level approach to prompt injection: separating instruction and data channels in LLM agents [P] β€” score 12 Sources: reddit/r/MachineLearning

Prompt injection has emerged as one of the most persistent failure modes in tool-using LLM systems, particularly in agentic workflows where models interact with external data sources. Most mitigation strategies focus on input filtering or model-side alignment, but these approaches struggle because t

Developer Tools

🟒 πŸ™ vas3k/TaxHacker β€” Self-hosted AI accounting app. LLM analyzer for receipts, invoices, transactions with custom prompts and categories β€” score 39 Sources: github_trending

Self-hosted AI accounting app. LLM analyzer for receipts, invoices, transactions with custom prompts and categories

🟒 πŸ’¬ ZCode: New Agentic Code Editor from the Makers of GLM β€” score 29 Sources: reddit/r/LocalLLaMA

🟒 πŸ™ RedPlanetHQ/core β€” Your Personal AI OS β€” score 24 Sources: github_trending

Your Personal AI OS

🟒 πŸ’¬ Ways to reduce token cost in AI agents β€” score 22 Sources: reddit/r/AIAgents

Ive been building AI agents for a while and noticed a few patterns that help cut token usage without wrecking the workflow. A few things that seem to matter most: - Set hard token budgets per task or step. - Stop runaway loops early with guardrails or circuit breakers. - Use smaller models for ch

🟒 πŸ™ togatoga/karukan β€” Japanese Input Method System for Linux, macOS, Neural Kana-Kanji Conversion Engine β€” score 21 Sources: github_trending

Japanese Input Method System for Linux, macOS, Neural Kana-Kanji Conversion Engine

Omitted 3 additional developer tools items from the main section; see raw data and source-specific sections below.

Infrastructure & Compute

🟒 πŸ’¬ Anyone using TensTorrent gpus for your local ai? What's been your experience? β€” score 4 Sources: reddit/r/LocalLLaMA

I'm always keeping an eye on competitive hardware and was looking at tenstorrent cards, particularly the p150a which while its memory bandwidth is only 512GB/s, it does have 32 GB of GDDR6 and a high-speed Ethernet fabric (4Γ—800 GbE) so multi-card systems don't rely on PCIe alone. something like thi

Research Papers

🟒 πŸ€— SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE β€” score 20 Sources: huggingface

We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical priors into pre-trained diffusion transformers. Existing methods either rely on costly fine-tuning on scarce panoramic data that limits generalization,

🟒 πŸ€— Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing β€” score 20 Sources: huggingface

Existing instruction-based video editing datasets commonly focus on single-task appearance editing, failing to meet the complex creative demands of real-world scenarios. To bridge this gap, we present Goku, a large-scale dataset featuring 2 million high-quality, instruction-aligned video editing pai

Other Signals

🟒 πŸ’¬ ICML qr code visible [D] β€” score 31 Sources: reddit/r/MachineLearning

Hi everyone, The check in QR code is visible at my profile despite that my card isn’t accepting the payment transaction. What does that even mean? Thanks!

🟒 πŸ’¬ Senior SWE Bench: a new benchmark focussed on realistically underspecified feature tasks β€” score 18 Sources: reddit/r/LocalLLaMA

🟒 πŸ’¬ How to describe a model that has higher accuracy with fewer #param and FLOPs? [D] β€” score 12 Sources: reddit/r/MachineLearning

Hello, My supervisor is nowhere to be found so I am turning to the internet for my naive questions.

RepoDescriptionStars TodayLanguage
allenai/olmocrToolkit for linearizing PDFs for LLM datasets/training295python
virgiliojr94/book-to-skillTurn any technical book PDF into a Claude Code skill β€” ready to study, reference, and use while you work.205python
open-webui/open-webuiUser-friendly AI Interface (Supports Ollama, OpenAI API, ...)171python
HKUDS/VideoAgent"VideoAgent: All-in-One Agentic Framework for Video Understanding, Editing, and Remaking"150python
getmaxun/maxunπŸ”₯ The open-source no-code platform for web scraping, crawling, search and AI data extraction β€’ Turn websites into structured APIs in minutes πŸ”₯88typescript
vas3k/TaxHackerSelf-hosted AI accounting app. LLM analyzer for receipts, invoices, transactions with custom prompts and categories60typescript
RedPlanetHQ/coreYour Personal AI OS32typescript
togatoga/karukanJapanese Input Method System for Linux, macOS, Neural Kana-Kanji Conversion Engine29rust

πŸ“„ New Papers

TitleCategoryHotnessLink
TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learningresearch_paper7Open
What Drives Interactive Improvement from Feedback?cs.AI0Open
Contrastive Reflection for Iterative Prompt Optimizationcs.AI0Open
How Can AI Find My Model? A Model-Finding Experimental Study Considering Data Formats, Embeddings, and Retrieval Strategiescs.AI0Open
BayesBench: Evaluating LLM Belief Trajectories Under Multi-Turn Evidence Accumulationcs.AI0Open
When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Modelscs.AI0Open
Beyond expert users: agents should help users construct preferences, not just elicit themcs.AI0Open
Investigating Multi-Agent Deliberation in Lawcs.AI0Open
Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineeringcs.AI0Open
RoPoLL: Robust Panel of LLM Judgescs.AI0Open
AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performancecs.AI0Open
Neuro-Bayesian-Symbolic Residual Attention Shallow Network: Explainable Deep Learning for Cybersecurity Risk Assessmentcs.AI0Open
HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observationcs.AI0Open
AgentBound: Verifiable Behavioral Governance for Autonomous AI Agentscs.AI0Open
When Regulation Has Memory: Hysteresis and Control Burden in Artificial Agencycs.AI0Open

🏒 Lab Blog Posts

🐦 Twitter/X Highlights

AccountTweet Summary
GoogleDeepMindWe’re shipping 2 major releases: πŸ”˜ Nano Banana 2 Lite: our fastest and cheapest Gemini Image model πŸ”˜ Gemini Omni Flash: now available via the Gemini API and in @GoogleAIStudio to help developers generate and edit high-quality videos. Post
xaiIntroducing Voice Agent Builder: a no-code platform to create human-like voice agents with Grok Voice. Available today at $0.05 / min. http://x.ai/voice Post
AnthropicAIClaude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and block more cybersecurity tasks. In the near term, some routine tasks like coding and debugging will fal Post
AnthropicAIWe’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring access tomorrow, and will share an update soon. We’re grateful to our users for their patience, and to everyone who worked with us on redeploying the models. Post
samaGood new first: Sol is a smart, efficient, and a significant step forward. It is the same price as GPT-5.5. Also launching in the GPT-5.6 family is Terra, with 5.5-level performance at half the price. Bad news: at the request of the US government, it is launching today in limited preview instead of Post

Newsletter

Repeated From Recent Briefings