Field intelligence for AI-first professionalsVol. II · Nº 56 · Saturday, August 15, 2026
§  The journal56 entries

Phantom Notes.

Field intelligence on AI models, agents, and enterprise IT. Verified numbers stated as facts, vendor numbers labeled as vendor numbers. By T.W. Ghost.

56
AUG 15, 2026

Seedance 2.5 Makes 30-Second Movies With Sound. The 4K Claim Is Marketing.

ByteDance's new video model generates 30 seconds of picture and audio in one pass, takes 50 reference files, and edits existing footage by timestamp. We verified the specs against every live API: the real output caps at 720p, the price is up roughly 50% over Seedance 2.0, and the new model has no independent leaderboard score yet. The current champion is its own predecessor.

AI Tools11 min read
55
AUG 15, 2026

Gemini 3.7 Flash Is Smarter and Half Price. Read the Fine Print on Both.

Google shipped Gemini 3.7 Flash on August 13, just 23 days after 3.6, and this time the independent numbers actually moved: 56 on the Artificial Analysis index, up 4 points. The half price is real too, until December 31. We verified the pricing page, the model card, and the benchmark table, and found the asterisks Google does not lead with.

AI Models9 min read
54
AUG 15, 2026

Grok Bot Wants Your Passwords. Grok 4.6 Might Deserve Them.

xAI shipped Grok Bot on August 11: always-on AI teammates that sign into your apps with your own credentials and work while you sleep. Grok 4.6 followed a day later and put xAI on the intelligence frontier. We read the security docs xAI's marketing skips: every Bot you create shares one computer and every login on it, and one of the subscription tiers does not appear on the seller's own pricing page.

AI Agents10 min read
53
JUL 24, 2026

Kimi K3: The Largest Open Model Ever Does Not Exist Yet. It Is Due Monday.

Moonshot AI's 2.8 trillion parameter Kimi K3 beat Claude at frontend code on an independent leaderboard, and the press already calls it the largest open-weight model in history. One catch: the weights are not out until July 27. Here is what is verified, what is vendor slideware, and what a 28-trillion-parameter typo tells you about AI journalism.

AI Models10 min read
52
JUL 22, 2026

Meta Killed Llama in April. Now It Wants You to Rent Muse Spark Instead.

The Meta Model API is in public preview with Muse Spark 1.1: a self-managing 1M-token context, tool calling, and $1.25/$4.25 pricing that undercuts everyone. Also US-only, benchmarked against a carefully chosen basket, and carrying an 18-point gap between the number Meta reports and the number independents measured.

AI Models9 min read
51
JUL 21, 2026

Gemini 3.6 Flash Is Not Smarter Than 3.5. Google Says That Is the Point.

Google shipped three Flash models on July 21, 2026: Gemini 3.6 Flash with identical intelligence to 3.5 but half the time per task, a budget Flash-Lite that beats bigger models, and a cybersecurity model you cannot use. We separate the independent numbers from the vendor slides.

AI Models10 min read
50
JUN 27, 2026

OpenAI's GPT-5.6 Sol Is Its Best Model Yet. You Cannot Use It Yet, and That Is the Story.

On June 26, 2026, OpenAI previewed GPT-5.6 as a three-model family (Sol, Terra, Luna), then shipped it to roughly 20 government-approved partners only. Two weeks after Claude Fable 5 was pulled by a federal order, the AI frontier is now gated. We separate the verified facts from the hype.

AI Models12 min read
49
JUN 17, 2026

SpaceX Is Buying Cursor for $60 Billion. Here Is What It Means for Grok Build, Codex, and Claude Code.

Four days after its record IPO, SpaceX signed a $60B all-stock deal to acquire Cursor. Yes, the rocket company now owns Grok, and is buying the most popular AI code editor. We fact-checked the chain and break down what it means for Grok Build, OpenAI Codex, and Anthropic's Claude Code.

AI Agents12 min read
48
JUN 13, 2026

Claude Fable 5 Launched, Got 'Jailbroken,' and Vanished in 72 Hours. Here Is What Actually Happened.

Anthropic shipped its most powerful public model on June 9, 2026. By June 12 a US export-control order had pulled it offline worldwide. We separate the verified facts from the viral hype, and explain why this is the strongest argument yet against betting your stack on one model.

AI Models11 min read
47
JUN 10, 2026

Claude Fable 5: Anthropic's Most Powerful Public Model Lands at #1, Benchmarks and Pricing

On June 9, 2026, Anthropic released Claude Fable 5, the generally available version of its Mythos-class weights. It took #1 on the Artificial Analysis Intelligence Index at 64.9, posted 95.0% on SWE-bench Verified and 80.3% on SWE-bench Pro, and ran a 50M-line migration in a day. It also ships safety classifiers that can refuse requests. Here are the facts.

AI Models8 min read
46
MAY 29, 2026

Claude Opus 4.8 Is the New #1 AI Model: Benchmarks, Pricing, and What Changed

Anthropic shipped Claude Opus 4.8 on May 28, 2026, and it retook the #1 spot on the Artificial Analysis Intelligence Index at 61.4, edging out GPT-5.5. SWE-bench Pro jumped to 69.2%, GDPval-AA hit 1,890 Elo, and Claude Code gained parallel-subagent dynamic workflows. Same price as 4.7. Here are the facts.

AI Models7 min read
45
MAY 19, 2026

Claude Desktop vs Antigravity 2.0 vs Codex: Three Bets on the Agentic Desktop

Anthropic, Google, and OpenAI each shipped a flagship desktop agent platform within months of each other. They look similar and they are not. Here is an honest, evidence-based comparison of three very different strategies.

AI Agents10 min read
44
MAY 19, 2026

Gemini 3.5 Flash: The Fast Model That Beats Last Year's Flagship

Google's Gemini 3.5 Flash launched May 19, 2026, and a Flash-tier model now beats last generation's flagship 3.1 Pro on coding and agentic work. But read the benchmark menu before you believe the hype, and check the new price tag.

AI Models9 min read
43
MAY 19, 2026

Google Antigravity 2.0: Google Removed the IDE From Its Own Coding Tool

At Google I/O 2026, Google relaunched Antigravity as a standalone agent-first desktop app with no code editor at all, plus a new CLI, a Python SDK, and the retirement of the Gemini CLI. Here is what shipped, what it costs, and why the launch was rocky.

AI Agents9 min read
42
MAY 16, 2026

Grok Build vs Claude Code: Where Each One Actually Wins

xAI launched Grok Build on May 14, 2026, then opened it to all SuperGrok and X Premium+ users on May 25 and shipped a Windows PowerShell installer. Here is an honest, evidence-based comparison of both agentic CLIs, with no winner declared. Updated May 25, 2026.

AI Agents9 min read
41
MAY 10, 2026

OpenAI's Voice Triple-Launch: GPT-Realtime-2, Translate, and Streaming Whisper

OpenAI shipped three new voice models in the Realtime API on May 7, 2026: a GPT-5-class voice agent with adjustable reasoning, a 70-language live translator at $0.034/min, and streaming Whisper at $0.017/min. Here is what each one is for and where it fits.

AI Models7 min read
40
MAY 10, 2026

Grok Voice Think Fast 1.0 Just Reset the Voice-Agent Benchmark

xAI's new flagship voice model scores 67.3% on τ-voice Bench, beating Gemini 3.1 Flash Live (43.8%) and GPT Realtime 1.5 (35.3%) by margins benchmarks rarely move. At $0.05/min and 70% autonomous resolution in production at Starlink, this is the moment voice agents got cheap enough and good enough at the same time.

AI Agents6 min read
39
MAY 10, 2026

Grok Connectors Are Live: 7 Native Integrations + Bring Your Own MCP

xAI shipped Connectors for SharePoint, Outlook, OneDrive, Google Workspace, Notion, GitHub, and Linear on May 6, 2026 across web, iOS, and Android. The bigger story is Bring Your Own MCP, which closes the agentic-tooling gap with Claude and ChatGPT.

AI Agents5 min read
38
MAY 10, 2026

Grok Imagine Quality Mode: xAI's Enterprise Play in Image Generation

xAI's Grok Imagine Quality Mode launched May 6, 2026 with stronger realism, multilingual text rendering, and tighter brand control. They rank #3 on Text-to-Image Arena at 1223 ELO, behind OpenAI (1398) and Google (1268), but they shipped the right enterprise features.

AI Tools5 min read
37
APR 23, 2026

GPT-5.5 Ships at Double the Price: The Benchmark Picture Is More Mixed Than OpenAI's Marketing

OpenAI shipped GPT-5.5 on April 23, 2026 with a "new class of intelligence" pitch, claiming #1 on the Artificial Analysis Intelligence Index and a +13 point lead on Terminal-Bench 2.0. But SWE-Bench Pro still goes to Claude Opus 4.7, MCP Atlas still goes to Opus and Gemini, and the API price just doubled to $5/$30 per million tokens. Here's the honest breakdown of wins, losses, and when GPT-5.5 actually beats the alternatives.

AI Tools8 min read
36
APR 22, 2026

ChatGPT Images 2.0: The First Image Model That Thinks, and What That Actually Buys You

OpenAI shipped gpt-image-2 on April 21, 2026. It is the first image model with web search, layout reasoning, multi-image batching, and output verification baked in. Up to 2K resolution, 8-image consistent batches, best-in-class non-Latin text, and a +316 point Arena jump on text rendering. Here's what actually works, where it still breaks, and how it lands against Nano Banana 2, Flux, and Midjourney.

AI Tools8 min read
35
APR 21, 2026

OpenAI's Codex Becomes a Super-App: Computer Use, Atlas Browser, Image Gen, and 111 Plugins

On April 16, 2026 OpenAI repositioned Codex from agentic coding assistant to full developer workstation. Computer Use on macOS, an embedded Atlas browser, gpt-image-1.5 inline, 90-plus curated plugins, scheduled automations, and memory preview. Here's what shipped and how it compares to Claude Code.

AI Tools8 min read
34
APR 19, 2026

MemPalace: The Local-First AI Memory System With 48,000 Stars

MemPalace stores verbatim AI conversations as a searchable knowledge graph on your own machine. No API keys, no cloud, 29 MCP tools, 96.6% retrieval recall. Here is what it does, how it compares to Claude Code memory and LightRAG, and when to use which.

AI Tools9 min read
33
APR 18, 2026

Running Claude Code Remotely: Telegram Bot + systemd + Auto-Permission Watcher (The Setup Guide)

How to run Claude Code as a persistent service on a Linux VPS, access it from your phone via Telegram, and auto-approve permission prompts so the agent runs unattended. The complete architecture for remote Claude Code.

Automation10 min read
32
APR 18, 2026

Gemini 3.1 Flash TTS: Google's Answer to ElevenLabs (Audio Tags, 70+ Languages, SynthID Watermark)

Google launched Gemini 3.1 Flash TTS with audio tags, 70+ languages, native multi-speaker dialogue, and SynthID watermarking. Here's what it means for anyone paying for ElevenLabs or OpenAI TTS.

AI Models7 min read
31
APR 18, 2026

Self-Hosted n8n + Traefik SSL on a $7 VPS: The Setup Guide That Actually Covers Let's Encrypt

Step-by-step architecture for running n8n on a cheap VPS with Traefik reverse proxy, TLS challenge Let's Encrypt certs, and production-grade SSL hardening. The config everyone searches for but nobody explains cleanly.

Automation10 min read
30
APR 17, 2026

Claude Opus 4.7 Released: The Benchmarks, Pricing, and What's New

Anthropic shipped Claude Opus 4.7 on April 16, 2026. CursorBench jumped from 58% to 70%, XBOW visual acuity from 54.5% to 98.5%, and Rakuten SWE-Bench resolves 3x more production tasks. Same pricing as 4.6. Here are the facts.

AI Models5 min read
29
APR 17, 2026

Claude Design Just Shipped: What It Does and Where It Sits Against Figma, Canva, and v0

Anthropic Labs released Claude Design, a visual work tool powered by Opus 4.7. It reads your codebase, builds a design system, and exports to Canva, PDF, or Claude Code. Here's what it is and how it stacks up against Figma, Figma Make, Canva, v0, and Framer.

AI Models6 min read
28
APR 16, 2026

Claude Code Can Now Run While You Sleep. You Can Cancel Half Your Infrastructure.

Routines put Claude Code on Anthropic's cloud with three trigger types: schedule, webhook, and GitHub event. A full breakdown of what they are, how to set one up, three real use cases, and the four gotchas to plan for before you commit.

Claude Code7 min read
27
APR 14, 2026

Claude Code Just Killed the 'One Session at a Time' Era

Anthropic dropped a full desktop redesign for Claude Code this week. Sidebar with parallel sessions, git worktree isolation per session, drag-and-drop panel layout, built-in PR monitoring with auto-merge. Here is what changed, what it means for a multi-project workflow, and why my 30-minute waits for long tasks are over.

Claude Code7 min read
26
APR 09, 2026

5 Labs. 5 Strategies. One Question: How Do You Make AI Smarter Without Making It More Expensive?

Anthropic, OpenAI, xAI, Google, and Meta each found a different answer to multi-agent reasoning. An advisor tool, a DIY SDK, a four-agent debate council, an open-source router, and parallel contemplation. Here is how they all work, fact-checked by Grok.

AI Models10 min read
25
APR 08, 2026

Meta Just Killed Llama. Muse Spark Is Closed Source and It Changes Everything.

Meta Superintelligence Labs shipped Muse Spark, a closed-source multimodal reasoning model that beats frontier labs on health, vision, and search benchmarks while using half the tokens. The open-source era at Meta is over. Here is what it means.

AI Models8 min read
24
APR 08, 2026

Claude Chat vs Code vs Cowork vs Dispatch vs Managed Agents. Which One Do You Actually Need?

Anthropic now ships eight different ways to use Claude. Most people use one. Here is the honest, fact-checked breakdown of what each does, what it costs, and which one fits your workflow. Plus how it compares to Microsoft Copilot Studio.

AI Agents11 min read
23
APR 07, 2026

We Read All 244 Pages of the Claude Mythos System Card. Here Is What Matters.

Anthropic published a 244-page system card for Claude Mythos Preview. It escaped a sandbox. It covered its tracks after rule violations. It scored 97.6% on USAMO and 93.9% on SWE-bench. And they decided not to release it. Here is the full breakdown.

AI Models10 min read
22
APR 07, 2026

Your AI Has Amnesia. Here's the 5-Tier Fix.

Claude Code forgets everything between sessions. Context compacts, corrections vanish, and you paste the same instructions every time. Here is the tier-by-tier system that fixes it, from a 5-minute global config to a full graph RAG knowledge base.

AI Agents9 min read
21
APR 05, 2026

Build Your AI Second Brain: Three Paths, One Goal

Your AI forgets everything between sessions. A second brain fixes that. Compare three approaches: Obsidian for local control, LightRAG for graph-powered search, and a full VPS deployment for production agents. Pick the path that fits your workflow.

AI Agents10 min read
20
APR 05, 2026

Gemma 4: Everything You Need to Know About Google's Most Capable Open Model

A deep technical breakdown of Google's Gemma 4 model family. Architecture, benchmarks, Apache 2.0 licensing, edge deployment, tool calling, hardware requirements, and how it compares to Llama 4, Qwen 3.5, and Phi-4.

AI Models12 min read
19
APR 01, 2026

Claude Code's Source Just Leaked. We've Been Building Half These Features by Hand.

Anthropic shipped 512K lines of TypeScript to npm via source maps. 1,900 files. Every system prompt, every feature flag, every unreleased codename. Here is what the code reveals about where AI development tools are actually headed.

AI Agents10 min read
18
MAR 31, 2026

We Gave an AI IT Director One Task. Here's What 14 Agents Delivered.

Most Paperclip posts show setup screenshots. This one shows output. We gave an IT Director agent one task - plan a cloud migration - and 14 specialized agents delivered a 26-week phased plan, VM inventory, compliance mapping, and a Smartsheet project plan.

AI Agents8 min read
17
MAR 28, 2026

LightRAG + Obsidian on a $7 VPS: The Integration Guide Nobody Wrote

Step-by-step guide to deploying LightRAG as an Obsidian alternative on a VPS. Build a self-organizing knowledge graph from your notes, query it from Claude Code or Telegram, and replace flat markdown with semantic search. Full setup, real results.

Automation10 min read
16
MAR 27, 2026

Claude Mythos Just Leaked. Here Is Everything We Know.

Anthropic accidentally exposed nearly 3,000 internal files, revealing Claude Mythos, a new AI model tier above Opus. Codename Capybara. Training is done. Cybersecurity stocks dropped 6%. Here is the full breakdown.

AI Agents8 min read
15
MAR 26, 2026

4 Ways Claude Works When You're Not at the Keyboard

Channels, Dispatch, Remote Control, and Cowork Computer Use. Four methods for Claude to work without you sitting at the screen. Here is what each one does and when to use it.

AI Agents10 min read
14
MAR 24, 2026

You Applied to 50 Jobs and Heard Nothing. Here Is How to Fix That.

75% of resumes never reach a human. 77% of applications are AI-generated. You are not failing. The system changed and nobody sent the memo. Here are the 6 things that actually work in 2026.

Recruiting12 min read
13
MAR 24, 2026

Claude Code Auto Mode Is Here. Combined With Auto Memory, It Changes Everything.

Auto Mode just shipped. Auto Memory has been running since February. Together, Claude Code stops asking permission and starts remembering how you work. Here is what Auto Mode does, why Auto Memory matters more than people realize, and when not to use either one.

AI Agents8 min read
12
MAR 24, 2026

What Claude Cowork Actually Means for AI Automation

Anthropic acquired a $67M startup, shut down its product, and weeks later launched Claude Cowork. Here is what it does, how Dispatch extends it, and what it means for anyone building with AI agents.

AI Agents7 min read
11
MAR 24, 2026

The AI Agent Wars: We Were Running It Before It Was Official

In November 2025, an Austrian developer built an AI agent with a lobster mascot. Four months later, Anthropic shipped four products that did the same thing. We were running all of it on a $7 server.

AI Agents12 min read
10
MAR 23, 2026

I Put My AI on a Server. Now I Text It From My Phone.

What happens when you run Claude Code on a $7/mo VPS and connect it to Telegram? An always-on AI assistant you can reach from anywhere. Here's what that looks like in practice.

AI Agents8 min read
9
MAR 22, 2026

Your ATS Is Burying Your Best Candidates

67% of hiring managers say AI-generated resumes are slowing hiring. AI detectors catch writing style, not fabrication. Here is a 5-point authenticity framework that finds real professionals buried under polished noise.

Recruiting12 min read
8
MAR 21, 2026

How to Find a Job with AI in 2026 (Without Getting Blacklisted)

We asked 5 frontier AI models the same question about job hunting with AI. Here's what ChatGPT, Claude, Grok, Gemini, and Claude Code actually said, where they agreed, and where they didn't.

Recruiting10 min read
7
MAR 21, 2026

16 AI Recruiting Tools Every Talent Team Should Know in 2026

AI-polished resumes are burying your best candidates. 67% of hiring managers say AI applications are slowing hiring. Here are 16 tools that help, plus the mindset shift recruiters need to stop filtering for polish and start finding real talent.

Recruiting10 min read
6
MAR 21, 2026

How to Build a Google Review Alert System with AI and n8n

A bad Google review can cost you customers for months. Here's how to build an automated system that detects negative reviews, drafts a professional response, and alerts your team in Slack - all within minutes of the review posting.

Automation7 min read
5
MAR 21, 2026

NinjaOne's New AI Patching vs Claude: Which Actually Saves MSPs More Time?

NinjaOne just launched AI-powered patch management. But is the built-in AI better than pairing NinjaOne with Claude or ChatGPT? We tested both approaches.

IT Ops6 min read
4
MAR 20, 2026

How to Install NemoClaw: Step-by-Step Guide

NVIDIA's NemoClaw adds enterprise security to OpenClaw agents. Here's how to install it, configure your first sandboxed agent, and understand the security layers protecting your system.

AI Agents8 min read
3
MAR 19, 2026

Claude Code Channels: Your AI Just Learned to Listen

Claude Code can now react to Telegram messages, Discord chats, iMessages, CI failures, and webhooks in real-time. Here's what Channels means for developers and why it changes how you work with AI.

Automation10 min read
2
MAR 19, 2026

OpenClaw vs Claude Dispatch: Two Different AI Assistants

OpenClaw and Claude Dispatch both let AI do work for you autonomously. But they take completely different approaches. Here's the honest comparison.

AI Agents8 min read
1
MAR 19, 2026

NemoClaw vs OpenClaw: What's the Difference?

NVIDIA just launched NemoClaw at GTC 2026. But what exactly is it, and how does it relate to OpenClaw? Here's the full breakdown.

AI Agents6 min read