Daily Archive
Past daily briefings.
AI Watchtower Briefing β 2026-09-19
π΄ Introducing the Australian Youth Safety Blueprint β score 75 Β· π’ first-party
AI Watchtower Briefing β 2026-09-18
π΄ Claude Code now reads AGENTS.md if there is no Claude.md β score 85 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-17
π΄ I literally built the Jev architecture one year back and completely open-sourced it with model, dataset and paper β score 87 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-16
π΄ Mistral X Mozilla: Private, Multilingual AI Browsing β score 80 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-15
π΄ Introducing System One Models and Jev β score 84 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-14
π΄ RTX PRO 5500 Blackwell (84GB) released β score 81 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-13
π΄ The Local LLM community feels like the golden era of the internet all over again β score 77 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-12
π΄ Perplexity trusts GPT-6 Astra with end-to-end systems β score 75 Β· π’ first-party
AI Watchtower Briefing β 2026-09-11
π΄ Someone apparently managed to kind of replicate what V4.1 flash does on KV for fast prefill on Qwen β score 83
AI Watchtower Briefing β 2026-09-10
π΄ deepseek-ai/DeepSeek-V4.1-Flash (6 downloads) β score 79 Β· π Γ2 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-09
π΄ Deepseek Has Soft Retired Deepseek V4 Pro β score 87 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-08
π΄ Google DeepMind Releases AlphaGenome Atlas β score 80 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-07
π΄ Friends Don't Let Friends Use Ollama β score 84 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-06
π΄ Qwen3.8-27B "Unhacked" my PC β score 81
AI Watchtower Briefing β 2026-09-05
π΄ AA Update! Here's how the Frontier ranks. β score 83
AI Watchtower Briefing β 2026-09-04
π΄ OpenAI CEO Sam Altman says 38,000 ChatGPT queries use as much water as the production of one almond β says data centers use no more water than an office building: βFor every 38,000 ChatGPT queries, that is the same...
AI Watchtower Briefing β 2026-09-03
π΄ "Welcome to the AGI era," OpenAI says as GPT-6 Astra debuts β score 87 Β· π Γ2 Β· π₯ engaged
AI Watchtower Briefing β 2026-09-02
π΄ Introducing Gemini 3.8 Flash and 3.8 Flash Cyber β score 93 Β· π Γ3 Β· π’ first-party Β· π₯ engaged
AI Watchtower Briefing β 2026-09-01
π΄ Youβll getmore of Claude Code and less of Claude Codeat the same time starting September 14. Donβt blame me. Itβs another instance of great comms from Anthropic. β score 95 Β· π Γ2
AI Watchtower Briefing β 2026-08-31
π΄ Breaking Claude Code Opus 5 Auto Mode β score 80 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-30
π΄ Sony and Warner accuse Anthropic of training Claude on tens of thousands of pirated works. Should the model be retrained from scratch? β score 78
AI Watchtower Briefing β 2026-08-29
π΄ Aug 29, 2026 Grok Bot now works with X Aug 29, 2026 Grok Bot now works with X Grok Bot now has a tighter integration with X. Read More β score 75 Β· π’ first-party
AI Watchtower Briefing β 2026-08-28
π΄ claude mods didn't like that, somehow π€·ββοΈ β score 87 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-27
π΄ No, Engrams won't let you run 1T models locally. It does something even better. β score 83 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-26
π΄ Qwen/Qwen3.8-Flash-Next (2,551 downloads) β score 76 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-25
π΄ Claude Mythos 5will power Claude Security for enterprise customers. This is the first rollout of Mythos outside Project Glasswing. β score 95 Β· π Γ2
AI Watchtower Briefing β 2026-08-24
π΄ I brought ChatGPT, Claude, and Gemini into a group chat to solve a complex problem. Here is how they caught each other hallucinating β score 76
AI Watchtower Briefing β 2026-08-23
π΄ Don't want to be this guy, but I need Qwen 3.8 35B A3B β score 76
AI Watchtower Briefing β 2026-08-22
π΄ Genbio Launches A βVirtual Cellβ AI Model (7 Minute Read) β score 70
AI Watchtower Briefing β 2026-08-21
π΄ DeepSeek-V4-Flash-Vision-Exp β score 89 Β· π Γ2 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-20
π΄ I just built a mini Kimi-K3 from Scratch under 250$. Already beats GPT-2 (124M)! β score 83 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-19
π΄ New midsize Qwen 3.8 model coming next week (hopefully) according to community manager! β score 81 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-18
π΄ Big Tech Is Raising Billions To Stop UBI β score 84 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-17
π΄ [\[OC\] Chinese models](https://i.redd.it/46er10v3vyjh1.jpeg) β score 72 Β· π₯ engaged
AI Watchtower Briefing β 2026-08-16
π΄ Letβs all thank Georgi Gerganov who gave use llama.cpp β score 87
AI Watchtower Briefing β 2026-08-15
π΄ Qwen 3.8 35BA3B spotted β score 88
AI Watchtower Briefing β 2026-08-14
π΄ Your agent can make the tests pass by deleting them. This shows you when it does. β score 94
AI Watchtower Briefing β 2026-08-13
π΄ White House creates framework for private companies to launch government authorized cyberattacks β score 94
AI Watchtower Briefing β 2026-08-12
π΄ It's the final countdown, baby! Qwen is out in just over 7 hours! β score 96
AI Watchtower Briefing β 2026-08-11
π΄ Qwen 3.8-27b coming this week β score 96
AI Watchtower Briefing β 2026-08-10
π΄ Mark Zuckerberg on releases β score 96
AI Watchtower Briefing β 2026-08-09
π΄ No wonder Qwen and Gemma are so different β score 90
AI Watchtower Briefing β 2026-08-08
π΄ moonshotai/Kimi-K3 (1,388,105 downloads) β score 95
AI Watchtower Briefing β 2026-08-07
π΄ Claude said the feature was done. it had never opened the page. β score 94
AI Watchtower Briefing β 2026-08-06
π΄ Qwen 3.8 Max now ranked as best overall model ahead of Opus 5 by Artificial Analysis agentic index β score 100
AI Watchtower Briefing β 2026-08-05
π΄ Reddit is introducing a new moderator: AI β score 95
AI Watchtower Briefing β 2026-08-04
π΄ More Qwen 3.8 sizes coming β score 96
AI Watchtower Briefing β 2026-08-03
π΄ Qwen3.8-27B announced alongside Qwen3.8-Max β score 97
AI Watchtower Briefing β 2026-08-02
π΄ Setting up of a 16xGB10 (DGX Spark) cluster β score 96
AI Watchtower Briefing β 2026-08-01
π΄ OpenAi says it has reached a new threshold in AI, new model capable of breakthrough research β score 81
AI Watchtower Briefing β 2026-07-31
π΄ The Chinese LLM release carousel never stops. Place your bets for MiniMax next week. β score 97
AI Watchtower Briefing β 2026-07-30
π΄ OpenAI reduces prices on its models by 5x β score 99
AI Watchtower Briefing β 2026-07-29
π΄ OpenAI's rogue agent ran ~17,600 actions across Hugging Face's infrastructure over 4 days β and HF's own post-mortem is wild reading β score 94
AI Watchtower Briefing β 2026-07-28
π΄ GPT-5, the world best model just 1 year ago, is today inferior to Qwen3.6 27B and most todayβs low-tier models β score 93
AI Watchtower Briefing β 2026-07-27
π΄ Kimi K3 weights now released. β score 96
AI Watchtower Briefing β 2026-07-26
π΄ zai-org/GLM-5.2 (827,191 downloads) β score 95
AI Watchtower Briefing β 2026-07-25
π΄ zai-org/GLM-5.2 (707,029 downloads) β score 95
AI Watchtower Briefing β 2026-07-24
π΄ I gave two Claude agents β¬100 and 90 days to earn β¬300. Day 7: ββ¬13, no product yet. β score 94
AI Watchtower Briefing β 2026-07-23
π΄ Absurd claim: the distilled model outperforms the originals β score 89
AI Watchtower Briefing β 2026-07-22
π΄ Terrence Tao's ChatGPT Conversation about the Jacobian Conjecture Counterexample β score 92
AI Watchtower Briefing β 2026-07-21
π΄ Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber β score 81
AI Watchtower Briefing β 2026-07-20
π΄ Unsloth now supports AMD! β score 82
AI Watchtower Briefing β 2026-07-19
π΄ Prepare your (v)ram - Qwen3.8 is coming! β score 97
AI Watchtower Briefing β 2026-07-18
π΄ GPT-5.6 used a prompt to close a 30-year gap in convex optimization β score 92
AI Watchtower Briefing β 2026-07-17
π΄ [Prism accidentally leaked \[D\]](https://www.reddit.com/r/MachineLearning/comments/1uz75qt/prism_accidentally_leaked_d/) β score 94
AI Watchtower Briefing β 2026-07-16
π΄ NotebookLM is now Gemini Notebook β score 94
AI Watchtower Briefing β 2026-07-15
π΄ Thinking Machines releases first open-weight model βInklingβ β score 81
AI Watchtower Briefing β 2026-07-14
π΄ How to stop Claude from saying load-bearing β score 92
AI Watchtower Briefing β 2026-07-13
π΄ asked ChatGPT for a one-click agent builder. the one-click part wasnt the thing β score 81
AI Watchtower Briefing β 2026-07-12
π΄ Claude Code sends 33k tokens before reading the prompt; OpenCode sends 7k β score 94
AI Watchtower Briefing β 2026-07-11
π΄ Why are MoE models so belittled? β score 79
AI Watchtower Briefing β 2026-07-10
π΄ [GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture \[pdf\]](https://cdn.openai.com/pdf/04d1d1e4-bc75-476a-97cf-49055cd98d31/cdc_proof.pdf) β score 88
AI Watchtower Briefing β 2026-07-09
π΄ Reasoning-Medical0.1-27B (Qwen3.5-27B medical finetune, claims to surpass MedGemma) β score 75
AI Watchtower Briefing β 2026-07-08
π΄ Chinaβs MiniMax Plans to Launch 2.7-Trillion Parameter Model β score 96
AI Watchtower Briefing β 2026-07-07
π΄ [MIRA: Multiplayer Interactive World Models trained on Rocket League \[R\]](https://www.reddit.com/r/MachineLearning/comments/1upofuw/mira_multiplayer_interactive_world_models_trained/) β score 94
AI Watchtower Briefing β 2026-07-06
π΄ I created a tool that turns AI agents into white collar employees. β score 94
AI Watchtower Briefing β 2026-07-05
π΄ Is the current Open Weight LLM model viable in the long term? β score 82
AI Watchtower Briefing β 2026-07-04
π΄ GPT-5.5 Codex reasoning-token clustering may be leading to degraded performance β score 75
AI Watchtower Briefing β 2026-07-03
π΄ Signal -> Triage -> Notify. My personal agentic setup to sift the signal from the noise. β score 94
AI Watchtower Briefing β 2026-07-02
π΄ Z.ai launches ZCode to challenge Cursor, Claude Code and GitHub Copilot in AI coding β score 79
AI Watchtower Briefing β 2026-07-01
π΄ Non Us Ally should be afraid. β score 82
AI Watchtower Briefing β 2026-06-30
π΄ Claude Code is steganographically marking requests β score 88
AI Watchtower Briefing β 2026-06-29
π΄ Qwen 3.6 27B is the sweet spot for local development β score 90
AI Watchtower Briefing β 2026-06-28
π΄ GLM 5.2 beats Claude in our benchmarks β score 94
AI Watchtower Briefing β 2026-06-27
π΄ Looking to become proficient in AI for real-world business applications. Where should I start? β score 89
AI Watchtower Briefing β 2026-06-26
π΄ US Govt to individually approve who gets GPT 5.6. β score 96
AI Watchtower Briefing β 2026-06-25
π΄ genuine question β why pay for exa/parallel "deep research" or "top level research" when i can just give my agent web access? β score 94
AI Watchtower Briefing β 2026-06-24
π΄ Qwen-AgentWorld-35B-A3B: a 3B-active MoE trained to simulate MCP, terminal, SWE, Android, web and OS environments β score 81
AI Watchtower Briefing β 2026-06-23
π΄ Krea 2 released on Hugging Face β score 71
AI Watchtower Briefing β 2026-06-22
π΄ The text in Claude Codeβs βExtended Thinkingβ output β score 75
AI Watchtower Briefing β 2026-06-21
π΄ I vibe code apps for a living. Here are my three tips. β score 94
AI Watchtower Briefing β 2026-06-20
π΄ z.AI as the number 2 gives praise to the number 1 open source model β score 96
AI Watchtower Briefing β 2026-06-19
π΄ Launch HN: TesterArmy (YC P26) β Agents that test web and mobile apps β score 83
AI Watchtower Briefing β 2026-06-18
π΄ The captcha arms race is making autonomous web tasks practically impossible β score 94
AI Watchtower Briefing β 2026-06-17
π΄ Donate your coding sessions to an open CC-BY-4.0 dataset to help train open-weight and open source models β score 97
AI Watchtower Briefing β 2026-06-16
π΄ Stop using Ollama β score 97
AI Watchtower Briefing β 2026-06-15
π΄ Introducing the Heretic Grimoire: The takedown-resilient, local-first backup system that keeps uncensored models available forever β score 96
AI Watchtower Briefing β 2026-06-14
π΄ This is coming to Chinese open source models pretty soon. - prepare yourself. β score 83
AI Watchtower Briefing β 2026-06-13
π΄ We should heavily discourage and moderate cloud API (deepseek api, GLM api, etc.) topics and discussion. This is LOCAL first. β score 75
AI Watchtower Briefing β 2026-06-12
π΄ Qwen Who? DiffusionGemma running at 1,500 tk/s on a Digital Pregnancy Test. β score 96
AI Watchtower Briefing β 2026-06-11
π΄ [Anthropic's new model Fable will silently handicap work on LLMs \[D\]](https://www.reddit.com/r/MachineLearning/comments/1u23f8p/anthropics_new_model_fable_will_silently_handicap/) β score 94
AI Watchtower Briefing β 2026-06-10
π΄ Claude Fable 5 β score 94
AI Watchtower Briefing β 2026-06-09
π΄ Apple reveals new AI architecture built around Google Gemini models β score 90
AI Watchtower Briefing β 2026-06-08
π΄ llama.cpp Gemma4 MTP support merged! β score 97
AI Watchtower Briefing β 2026-06-07
π΄ Cohere's unreleased coding model (early access for localllama) β score 96
AI Watchtower Briefing β 2026-06-06
π΄ Donβt act like yβall ainβt thinking it. Iβm just saying the quiet part out loud. /s β score 90
AI Watchtower Briefing β 2026-06-05
π΄ Anthropic's open-source framework for AI-powered vulnerability discovery β score 90
AI Watchtower Briefing β 2026-06-04
π΄ Introducing Gemma 4 12B: a unified, encoder-free multimodal model β score 90
AI Watchtower Briefing β 2026-06-03
π΄ Most of the software you rely on was hacked together fast β score 94
AI Watchtower Briefing β 2026-06-02
π΄ [Browse CVPR 2026 papers on PapersWithCode \[P\]](https://www.reddit.com/r/MachineLearning/comments/1tukrf4/browse_cvpr_2026_papers_on_paperswithcode_p/) β score 81
AI Watchtower Briefing β 2026-06-01
π΄ (YT) PewDiePie released his harness/webui β score 96
AI Watchtower Briefing β 2026-05-31
π΄ Anyone else tired of the βClaude just dropped OPUS 4.8 weβre all cookedβ posts on LinkedIn? β score 94
AI Watchtower Briefing β 2026-05-30
π΄ Notes from the Mistral AI Now Summit β score 90
AI Watchtower Briefing β 2026-05-29
π΄ Claude Opus 4.8 β score 92
AI Watchtower Briefing β 2026-05-28
π΄ Qwen3.6 huge quality gain from Q4 to Q6 for coding agent β score 75
AI Watchtower Briefing β 2026-05-27
π΄ A rare look inside Qwen 3.7βs open source model release approval process: β score 89
AI Watchtower Briefing β 2026-05-26
π΄ The Financial Times has published an article about Heretic β score 96
AI Watchtower Briefing β 2026-05-25
π΄ [PapersWithCode new features - week 1 \[P\]](https://www.reddit.com/r/MachineLearning/comments/1tmawv5/paperswithcode_new_features_week_1_p/) β score 94
AI Watchtower Briefing β 2026-05-24
π΄ GPT 5.5 "secret sauce" is just having the thinking be some stupid caveman mode? β score 97
AI Watchtower Briefing β 2026-05-23
π΄ [NuExtract3 released: open-weight 4B VLM for Markdown, OCR and structured extraction (self-hostable) \[P\]](https://www.reddit.com/r/MachineLearning/comments/1tkejqr/nuextract3_released_openweight_4b_vlm_for/) β sc...
AI Watchtower Briefing β 2026-05-22
π΄ Waiting for Qwen 3.7 open weight... The new King has arrived... β score 90
AI Watchtower Briefing β 2026-05-21
π΄ Qwen will release another 27B with high probability β score 96
AI Watchtower Briefing β 2026-05-20
π΄ Gemini 3.5 Flash β score 85
AI Watchtower Briefing β 2026-05-19
π΄ Qwen cant wait to release 3.7 models β score 97
AI Watchtower Briefing β 2026-05-18
π΄ "Generate a photorealistic realtime render of a human face with webGL" (Qwen3.5-122B-A10B UD-Q3_K_XL) β score 83
AI Watchtower Briefing β 2026-05-17
π΄ MTP PR Merged!!! β score 97
AI Watchtower Briefing β 2026-05-16
π΄ Opencode you naughty minx β score 96
AI Watchtower Briefing β 2026-05-15
π΄ The RTX 5000 PRO (48GB) arrived and it is better than I expected. β score 87
AI Watchtower Briefing β 2026-05-14
π΄ TextGen is now a native desktop app. Open-source alternative to LM Studio (formerly text-generation-webui). β score 94
AI Watchtower Briefing β 2026-05-13
π΄ Stop wasting electricity β score 88
AI Watchtower Briefing β 2026-05-12
π΄ MTP on Unsloth β score 79
AI Watchtower Briefing β 2026-05-11
π΄ I have DeepSeek V4 Pro at home β score 88
AI Watchtower Briefing β 2026-05-10
π΄ millionco/react-doctor β Your agent writes bad React. This catches it β score 87
AI Watchtower Briefing β 2026-05-09
π΄ vLLM ROCm has been added to Lemonade as an experimental backend β score 77
AI Watchtower Briefing β 2026-05-08
π΄ farion1231/cc-switch β A cross-platform desktop All-in-One assistant tool for Claude Code, Codex, OpenCode, openclaw & Gemini CLI. β score 95
AI Watchtower Briefing β 2026-05-07
π΄ I think a lot of people are accidentally building systems they can never debug β score 94
AI Watchtower Briefing β 2026-05-06
π΄ Gemma 4 MTP released β score 96
AI Watchtower Briefing β 2026-05-05
π΄ White House Considers Vetting A.I. Models Before They Are Released β score 81
AI Watchtower Briefing β 2026-05-04
π΄ Qwen3.6-27B vs Coder-Next β score 88
AI Watchtower Briefing β 2026-05-03
π΄ Qwen3.6-27B vs Coder-Next β score 96
AI Watchtower Briefing β 2026-05-02
π΄ We are finally there: Qwen3.6-27B + agentic search; 95.7% SimpleQA on a single 3090, fully local β score 73
AI Watchtower Briefing β 2026-05-01
π΄ Heterogeneous Scientific Foundation Model Collaboration β score 95
AI Watchtower Briefing β 2026-04-30
π΄ Large Language Models Explore by Latent Distilling β score 75
AI Watchtower Briefing β 2026-04-29
π΄ Recursive Multi-Agent Systems β score 95
AI Watchtower Briefing β 2026-04-28
π΄ From Skills to Talent: Organising Heterogeneous Agents as a Real-World Company β score 95
AI Watchtower Briefing β 2026-04-27
π΄ DiffNR: Diffusion-Enhanced Neural Representation Optimization for Sparse-View 3D Tomographic Reconstruction β score 75
AI Watchtower Briefing β 2026-04-26
π‘ Our principles β score 50
AI Watchtower Briefing β 2026-04-25
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-24
π΄ WorldMark: A Unified Benchmark Suite for Interactive Video World Models β score 85
AI Watchtower Briefing β 2026-04-23
π΄ Near-Future Policy Optimization β score 85
AI Watchtower Briefing β 2026-04-22
π΄ Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items β score 95
AI Watchtower Briefing β 2026-04-21
π΄ Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation β score 95
AI Watchtower Briefing β 2026-04-20
π΄ Maximal Brain Damage Without Data or Optimization: Disrupting Neural Networks via Sign-Bit Flips β score 80
AI Watchtower Briefing β 2026-04-19
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-18
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-17
π΄ HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds β score 95
AI Watchtower Briefing β 2026-04-16
π΄ Seedance 2.0: Advancing Video Generation for World Complexity β score 95
AI Watchtower Briefing β 2026-04-15
π΄ ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents β score 95
AI Watchtower Briefing β 2026-04-14
π΄ The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping β score 95
AI Watchtower Briefing β 2026-04-13
π΄ EXAONE 4.5 Technical Report β score 75
AI Watchtower Briefing β 2026-04-12
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-11
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-10
π΄ SkillClaw: Let Skills Evolve Collectively with Agentic Evolver β score 85
AI Watchtower Briefing β 2026-04-09
π΄ MARS: Enabling Autoregressive Models Multi-Token Generation β score 75
AI Watchtower Briefing β 2026-04-08
π΄ Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding β score 95
AI Watchtower Briefing β 2026-04-07
π΄ Adam's Law: Textual Frequency Law on Large Language Models β score 95
AI Watchtower Briefing β 2026-04-06
π΄ GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning β score 95
AI Watchtower Briefing β 2026-04-05
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-04
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-04-03
π΄ DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models β score 95
AI Watchtower Briefing β 2026-04-02
π΄ ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers β score 95
AI Watchtower Briefing β 2026-04-01
π΄ FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization β score 95
AI Watchtower Briefing β 2026-03-31
π΄ TAPS: Task Aware Proposal Distributions for Speculative Sampling β score 95
AI Watchtower Briefing β 2026-03-30
π΄ Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills β score 75
AI Watchtower Briefing β 2026-03-29
π‘ Helping disaster response teams turn AI into action across Asia β score 50
AI Watchtower Briefing β 2026-03-28
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-03-27
π΄ Calibri: Enhancing Diffusion Transformers via Parameter-Efficient Calibration β score 75
AI Watchtower Briefing β 2026-03-26
π΄ CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents β score 95
AI Watchtower Briefing β 2026-03-25
π΄ SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning β score 75
AI Watchtower Briefing β 2026-03-24
π΄ Omni-WorldBench: Towards a Comprehensive Interaction-Centric Evaluation for World Models β score 95
AI Watchtower Briefing β 2026-03-23
π΄ HopChain: Multi-Hop Data Synthesis for Generalizable Vision-Language Reasoning β score 85
AI Watchtower Briefing β 2026-03-22
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-03-21
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-03-20
π΄ SAMA: Factorized Semantic Anchoring and Motion Alignment for Instruction-Guided Video Editing β score 85
AI Watchtower Briefing β 2026-03-19
π΄ Video-CoE: Reinforcing Video Event Prediction via Chain of Events β score 75
AI Watchtower Briefing β 2026-03-18
π΄ Demystifing Video Reasoning β score 95
AI Watchtower Briefing β 2026-03-17
π΄ AI Can Learn Scientific Taste β score 95
AI Watchtower Briefing β 2026-03-16
π΄ Multimodal OCR: Parse Anything from Documents β score 85
AI Watchtower Briefing β 2026-03-15
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-03-14
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-03-13
π΄ Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections β score 85
AI Watchtower Briefing β 2026-03-12
π΄ Bootstrapping Exploration with Group-Level Natural Language Feedback in Reinforcement Learning β score 95
AI Watchtower Briefing β 2026-03-11
π΄ Geometry-Guided Reinforcement Learning for Multi-view Consistent 3D Scene Editing β score 95
AI Watchtower Briefing β 2026-03-10
π΄ Lost in Stories: Consistency Bugs in Long Story Generation by LLMs β score 95
AI Watchtower Briefing β 2026-03-09
π΄ Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders β score 95
AI Watchtower Briefing β 2026-03-08
π΄ Reasoning Models Struggle to Control their Chains of Thought β score 92
AI Watchtower Briefing β 2026-03-07
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-03-06
π΄ MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier β score 85
AI Watchtower Briefing β 2026-03-05
π΄ Helios: Real Real-Time Long Video Generation Model β score 85
AI Watchtower Briefing β 2026-03-04
π΄ Utonia: Toward One Encoder for All Point Clouds β score 95
AI Watchtower Briefing β 2026-03-03
π΄ SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale β score 75
AI Watchtower Briefing β 2026-03-02
π΄ dLLM: Simple Diffusion Language Modeling β score 95
AI Watchtower Briefing β 2026-03-01
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-28
π‘ Our agreement with the Department of War β score 50
AI Watchtower Briefing β 2026-02-27
π΄ From Blind Spots to Gains: Diagnostic-Driven Iterative Training for Large Multimodal Models β score 85
AI Watchtower Briefing β 2026-02-26
π΄ SkyReels-V4: Multi-modal Video-Audio Generation, Inpainting and Editing model β score 95
AI Watchtower Briefing β 2026-02-25
π΄ On Data Engineering for Scaling LLM Terminal Capabilities β score 95
AI Watchtower Briefing β 2026-02-24
π΄ VLANeXt: Recipes for Building Strong VLA Models β score 75
AI Watchtower Briefing β 2026-02-23
π΄ Does Your Reasoning Model Implicitly Know When to Stop Thinking? β score 95
AI Watchtower Briefing β 2026-02-22
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-21
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-20
π΄ Unified Latents (UL): How to train your latents β score 95
AI Watchtower Briefing β 2026-02-19
π΄ SLA2: Sparse-Linear Attention with Learnable Routing and QAT β score 95
AI Watchtower Briefing β 2026-02-18
π΄ GLM-5: from Vibe Coding to Agentic Engineering β score 95
AI Watchtower Briefing β 2026-02-17
π΄ BitDance: Scaling Autoregressive Generative Models with Binary Tokens β score 75
AI Watchtower Briefing β 2026-02-16
π΄ Less is Enough: Synthesizing Diverse Data in Feature Space of LLMs β score 95
AI Watchtower Briefing β 2026-02-15
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-14
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-13
π΄ DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing β score 75
AI Watchtower Briefing β 2026-02-12
π΄ Step 3.5 Flash: Open Frontier-Level Intelligence with 11B Active Parameters β score 95
AI Watchtower Briefing β 2026-02-11
π΄ OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration β score 95
AI Watchtower Briefing β 2026-02-10
π΄ TermiGen: High-Fidelity Environment and Robust Trajectory Synthesis for Terminal Agents β score 85
AI Watchtower Briefing β 2026-02-09
π΄ F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare β score 95
AI Watchtower Briefing β 2026-02-08
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-07
| Title | Category | Score | Link |
AI Watchtower Briefing β 2026-02-06
π΄ CAR-bench: Evaluating the Consistency and Limit-Awareness of LLM Agents under Real-World Uncertainty β score 95
AI Watchtower Briefing β 2026-02-05
π΄ ERNIE 5.0 Technical Report β score 95
AI Watchtower Briefing β 2026-02-04
π΄ AOrchestra: Automating Sub-Agent Creation for Agentic Orchestration β score 85
AI Watchtower Briefing β 2026-02-03
π΄ Kimi K2.5: Visual Agentic Intelligence β score 85
AI Watchtower Briefing β 2026-02-02
π΄ Golden Goose: A Simple Trick to Synthesize Unlimited RLVR Tasks from Unverifiable Internet Text β score 85
AI Watchtower Briefing β 2026-02-01
| Title | Category | Score | Link |