Live AI developer news, ranked and linked to original sources.
> ▌
Markdown sits near the point where human readability and machine readability meet. HTML adds a rendering layer where humans and agents can stop seeing the same artifact.

AI LABS

The PrimeTime

Prompt Engineering

Income stream surfers

WorldofAI

Two Minute Papers

DIY Smart Code

OpenAI
Grok 4.6 reportedly worked autonomously for 48 hours to build a playable shooter using Matt Shumer’s Gauntlet Loop method. The demo suggests frontier models are moving beyond one-shot code generation toward persistent, self-critiquing development workflows.
TypingMind sells a one-time license while customers pay model providers directly through their own API keys. That separation lets the product offer lifetime access without absorbing unpredictable inference costs.
super.engineering now lets developers review and respond to GitHub pull-request comments without leaving its native GPUI workspace. A configurable one-click action can trigger an automated review flow alongside existing agent sessions, diffs, checks, and merge actions.
Alibaba’s new 27B Qwen3.8 model is now available for local deployment, with early benchmarks showing it trading blows with or surpassing Claude Opus 4.6 on several coding tasks. The results could make frontier-level open-weight coding practical on consumer hardware.
Qwen released the open weights for Qwen3.8-27B, a 27B multimodal model built for coding and office workflows. It offers 262K native context, up to 1M tokens with YaRN, and an Apache 2.0 license.
Anthropic is embedding imperceptible, machine-detectable watermarks into text generated by new Claude models launched from August 2, 2026, while adding signed provenance metadata to supported files. The move responds to the EU AI Act’s transparency requirements.
Mercury Cloud is offering LongCat-2.0 at a limited-time 60% discount, giving developers cheaper access to Meituan’s 1.6T-parameter MoE model built for coding, long-context reasoning, and agentic workflows.
Cohere says its open-source North Mini Code model has surpassed 150,000 downloads as developers continue experimenting with agentic coding workflows. With OpenCode’s free tier ending, developers can access it through Model Vault, OpenRouter, or local deployment.
Gemini users can now disable visible watermarks on AI-generated images and videos through Settings > Media watermark. The rollout appears limited to the desktop web experience for now, while invisible SynthID and C2PA provenance data remain intact.
X released a substantially expanded open-source For You recommendation stack on August 13, including ranking weights, visibility filtering, labeling systems, Phoenix model training code, and synthetic data. The release also introduces an Under the Hood transparency tool showing aggregate visibility-impacting labels on accounts and posts.
Higgsfield’s Seedance 2.5 upgrade enables high-resolution, one-pass commercial generation with synchronized audio, timestamped action, dialogue, and consistent visual references. The workflow turns a detailed shot list into a near-finished ad from a single prompt.
Rudrank Riyam’s ASC CLI turns App Store Connect into a scriptable workflow for builds, TestFlight, submissions, signing, metadata, screenshots, and Apple Ads. Its agent-ready design also lets coding assistants automate much of the iOS release process.
DeepSeek Harness is an open-source developer-preview agent framework where models, tools, memory, sandboxes, workflows, and UI are composable plugins powered by Cordis. Its companion paper frames the architecture as infrastructure for dynamic composition, including self-evolving agent harnesses.
Merge and PostHog will host an evening of technical talks in New York’s Flatiron district on August 27, featuring demos of model routing, AI observability, and safeguards for agents acting on production data.
Infisical refreshed its homepage and positioning around open-source security infrastructure for developers and AI agents. The platform now brings secrets, certificates, and privileged access management under one identity-security stack.
François Chollet clarifies that ARC-AGI-3’s public games are a demonstration set, not training data or an evaluation set. Scores on them should not be interpreted as evidence of progress on the benchmark’s private tests.
ScientiaCapital’s open-source Claude Code skills library packages repeatable engineering, sales, research, and trading workflows into installable Markdown-based skills. The repository gives teams a versioned home for prompts that would otherwise disappear in chat threads.
SwapnanilDhol joins ASC CLI as a new contributor with a feature for managing Developer Portal App Groups through its web interface. The contribution also hardens CSRF handling and redirect validation.
dots3-note Preview is a 280B-parameter mixture-of-experts model with 16B active parameters and up to 512K context. The open-weight multimodal model targets reasoning, tool use, coding, and long-horizon agent workflows across text, images, video, and audio.
Cursor’s in-editor alert warns when an agent is about to start expensive processes, offering controls to manage them or suppress future notifications. It adds a useful layer of transparency as agents increasingly run commands autonomously.
Prem says its inference router will soon offer an uncensored version of DeepSeek V4 Flash 0731. The move targets developers seeking fewer content restrictions while retaining DeepSeek’s latest agent and coding capabilities.
Pi v0.84.2 improves its terminal coding-agent workflow with fullscreen transcript search, configurable default tools, per-run themes, and experimental strict JSON-schema sampling. The release also fixes streaming, TUI startup, tool-call namespace, and SDK messaging issues.
A developer argues that Opus 5 is more capable on benchmarks yet less pleasant in real coding work because it makes bold assumptions, rewrites plans, and asks fewer clarifying questions. The gap highlights how benchmark optimization can conflict with the judgment and restraint developers want from coding agents.
DeepSeek has released V4-Pro generally with stronger agent performance, adjustable reasoning effort, and native OpenAI Responses API support. New API pricing begins August 16, with peak-hour rates doubling off-peak prices.
Grok Bot’s early beta brings always-on AI agents, persistent cloud computers, browser access, terminals, and filesystem tools to Windows, macOS, and iPhone—but Android users are currently left out. The platform gap is especially notable because the underlying Grok assistant already supports Android.
Google Antigravity now supports Gemini 3.7 Flash for faster, lower-cost autonomous development workflows across its IDE, CLI, SDK, browser automation, and parallel subagents. The update targets multi-step coding, MCP tool use, and automated Lighthouse audits.
Anthropic reports that an unreleased Claude research model improved the known lower bound for Riemann zeta zeros on the critical line from 41.6% to 67.2%. It did not solve the Riemann hypothesis, but produced a formally verifiable mathematical result.
Z.ai’s GLM Coding Plan expands access to its latest coding models through ZCode and popular tools including Claude Code, OpenCode, Cline, and Cursor. The subscription combines discounted usage, caching benefits, and off-peak capacity for developers seeking a lower-cost alternative to premium coding models.
Truvyx argues that swappable models, MCP servers, skills, and prompts turn AI products into moving targets whose behavior can change without a conventional release. Its evaluation platform addresses this risk with constraint testing, formal verification, root-cause analysis, and production monitoring.
SpatialAxiom is an open-weight vision-language model family focused on 3D relational inference, perspective taking, multi-view correspondence, and embodied video understanding. Its 9B dense and 35B-A3B MoE models build on Qwen3.5 and support Transformers and vLLM.
Rajasthan Royals use ChatGPT across player analysis, match preparation, ticketing, HR, marketing, sponsorships, and finance. The case study shows conversational AI moving from experimentation into everyday sports operations.
Rajasthan Royals use OpenAI Codex to build and improve internal tools across their organization. The video shows AI-assisted software development extending beyond cricket operations into broader analytics and efficiency workflows.
Outcome turns creator expertise into personalized funnels that generate tailored action plans, audits, scores, roadmaps, or recommendations for each lead. It replaces generic lead magnets and rigid quiz buckets with AI-generated outcomes based on individual responses.
oxpecker monitors third-party API changes across 26 vendors and identifies the exact files and lines they could break. It runs in your CI, opens reviewable pull requests, and flags changes before deprecation deadlines.
OpenMotion is a free macOS motion-design studio that turns prompts, screenshots, and brand assets into editable launch videos, explainers, and social clips. Users can refine scenes on a real canvas and timeline, then export video, transparent WebM, or HTML.
isolate.video turns raw screen recordings into polished product videos with crop spotlighting, motion zooms, editable timeline effects, and AI-generated background music. It targets founders and marketers who need launch-ready demos without manual keyframing.
min. builds relationship capsules automatically from emails, calendars, and meetings, preserving context, commitments, and follow-ups without manual data entry. Teams can query, maintain, and share each relationship through a single living record.
Suno’s browser-based generative DAW adds MIDI, built-in synths, audio effects, automation, stem separation, and an AI chat bar for creating custom plugins and presets. The update is available to Premier subscribers.
Hoplite transfers local coding-agent sessions, MCP servers, dependencies, and CLIs into isolated cloud sandboxes. It enables parallel agents, instant previews, automated workflows, and remote task execution without laptop or port-management overhead.
Orca lets developers move active Claude sessions to Codex or other agents when Claude becomes unavailable. Its parallel-agent workspace keeps coding work moving across isolated environments.
BitsLab AI showcases three real-world applications of its Web3 security platform, combining smart-contract analysis, exploit validation, and threat intelligence. The product targets developers who need actionable vulnerability evidence rather than noisy automated alerts.
Riley Brown previewed GLM 5.3 alongside upcoming Grok, Claude Code, GPT, DeepSeek, and Gemini updates for an Agent Native weekly roundup. However, Z.ai has not yet published an official GLM 5.3 release announcement, so this remains a tease rather than a confirmed launch.
GPT Image 2 Skill bundles a curated image-prompt gallery, runnable agent skill, and CLI for generating and editing images through Claude Code, Codex, OpenClaw, and similar runtimes. Its examples span research figures, UI mockups, anime, photography, typography, and reference-image workflows.
Z.ai’s GLM-5.3 improves coding performance by 50% over GLM-5.2 through post-training, while reaching open-weights leadership on Terminal Bench 3.0 and Agents’ Last Exam. Its unexpectedly strong vulnerability-discovery and exploitation results prompted Z.ai to delay weight release for safety hardening.
AICodeKing’s early-access test presents GLM-5.3 as a security-oriented open coding model, highlighting frontend generation, backend engineering, code auditing, vulnerability discovery, and agentic workflows. The video reports a 73/80 KingBench 3 score, though Z.ai has not publicly documented an official GLM-5.3 release.

AICodeKing

Every

Github Awesome

Eric Michaud

OpenAI

Every

Eric Michaud

Rob The AI Guy

Prompt Engineering

Discover AI

Bijan Bowen

Theo - t3․gg