Live AI developer news, ranked and linked to original sources.
> ▌
Markdown sits near the point where human readability and machine readability meet. HTML adds a rendering layer where humans and agents can stop seeing the same artifact.

Discover AI

Bijan Bowen

Theo - t3․gg

OpenAI

OpenAI

AICodeKing

Income stream surfers

Better Stack

WorldofAI
Cogility’s Cogynt.ai v2.22 adds HTTP API data sources, Amazon Bedrock inference-profile selection, Markdown chatbot responses, and richer geospatial visualization. The update strengthens its continuous decision-intelligence platform for turning live enterprise data into auditable insights.
Cogility’s Cogynt.ai v2.22 adds real-time API data connectivity, Amazon Bedrock model selection, Markdown-ready chatbot responses, and new heatmap and geospatial controls. The update strengthens its decision-intelligence platform for analyzing high-volume operational and risk data.
TencentARC’s SCoPE adds camera sightline coordinates to Wan2.2-I2V-A14B, enabling image-to-video generation guided by explicit camera trajectories. The Apache-2.0 release includes self-contained weights, inference code, and support for custom camera poses.
ByteDance’s video model is being tested on an anime-style food sequence combining a broth pour, serving, first bite, and reaction line in one generation with audio included. The demo highlights Seedance 2.5’s ability to maintain narrative continuity across multiple beats.
Exa’s agentic web search is now natively available through Vercel AI Gateway and serves as the default web_search tool for eve agents. Usage is free through August 31.
Merge Gateway now supports image, video, and audio inputs through one API, letting developers access multimodal models across providers with intelligent routing and built-in reliability controls.
Anna is positioning itself as an AI-native app platform combining UI, persistent runtime, sandboxed execution, and built-in AI infrastructure. Its developer pitch pairs easier app distribution with a claimed 70% share of usage profit.
A new paper introduces SkillTriage, a differential framework for attributing agent failures and cost regressions to loaded skills. Across SkillsBench and SWE-Skills-Bench, the authors identify 125 functional failures and 182 efficiency regressions.
Netlify is showcasing Agent Runners in a live RenderATL demo, showing how developers can use Claude, Codex, or Gemini to build, preview, and ship applications from a prompt. The workflow connects AI-generated code to production infrastructure, deployment previews, and approval controls.
Wix’s new Grok Build plugin connects AI-built sites to Wix’s managed backend, including commerce, bookings, CMS, payments, and business operations. Developers can build frontends conversationally while managing production infrastructure through Wix.
Oxide details how customer workloads drove integrations with Rancher, Omni, Cluster API, and its cloud controller manager. The company is also developing native storage support while using customer friction to improve the underlying infrastructure stack.
Christopher Domas’s open-source research project rewires AMD DRAM address mappings to read and modify memory regions normally hidden from the operating system, including PSP, SMM, C6 state, and microcode. Tested on AMD Family 16h CPUs, it demonstrates how memory-controller assumptions underpin higher-level security boundaries.
Claude Code can connect to a UniFi controller through its API, inspect network health, identify WiFi and security problems, and propose fixes. DHH’s tip highlights a practical expansion of coding agents into real-world infrastructure operations.
Cursor will host a free online workshop on Thursday, August 20, from 5–6pm UTC. Kiara Polychroniadi will demonstrate how to tailor agents to a team’s codebase, conventions, and technology stack. Register: https://luma.com/6yh0erpt
Gloomberb is an open-source, keyboard-driven finance terminal available as a desktop app or TUI, combining market research, portfolios, charts, filings, alerts, and AI tools. Its plugin architecture makes the Bloomberg-style workflow extensible without requiring a Bloomberg subscription.
Vercel now includes one free domain for Pro customers across select extensions, including .online, .site, .space, .store, .tech, and .website. The offer gives paid users a simpler path from deployment to branded production URL.
Temporal and MongoDB have published a reference architecture for production AI agents, combining Temporal’s durable execution with MongoDB Atlas and Voyage AI for persistent data, retrieval, and agent memory.
Anthropic and Redwood Research introduce a benchmark suite measuring AI reasoning on philosophical, decision-theoretic, and AI-safety questions lacking reliable empirical answers. The index combines LMCA, ACCoRD, and DTBench, with Claude Opus 5 scoring 73.6 out of an estimated ceiling of 91.
AINFT is positioning itself as TRON’s AI infrastructure layer, combining multi-agent development tools, model access, agent deployment, and on-chain execution. The project aims to move TRON from a payments-focused blockchain toward an operating layer for autonomous AI agents.
Meta’s open-weight 30B multimodal model targets local agentic coding and tool-use workflows, with a 131K context window and quantized versions small enough for consumer GPUs. Its Apache 2.0 release makes advanced local agents more practical for developers without cloud inference costs.
NVIDIA’s open-weight Nemotron 3.5 Lightning is a 30B-parameter MoE model with roughly 3B active parameters, built for fast, always-on agent workloads. Its efficiency and local-deployment potential are compelling, though benchmark performance is comparatively modest.
Netlify tested 11 AI models on identical coffee-shop website prompts, revealing sharp differences in visual quality, consistency, and credit usage. The comparison follows Netlify’s OpenRouter integration, which expands Agent Runners beyond Claude, Codex, and Gemini.
DeepSeek is opening its agent harness to developers worldwide as an MIT-licensed open-source project. Built on the Cordis meta-framework, it targets developers creating production-grade agent experiences.
An analysis of tens of thousands of Arena outputs reportedly finds Claude Opus 5 uses roughly three times more em dashes than earlier Opus models. Anthropic’s own prompting guidance acknowledges that Opus 5 produces longer default user-facing responses.
Vercel’s AI SDK now supports any Agent Client Protocol-compatible harness through the new @ai-sdk/harness-acp meta adapter. Developers can define custom adapters with createACP and run them through HarnessAgent.
IBM is joining OpenAI’s elite partner tier to deploy GPT-5.6, Codex, and ChatGPT Work across regulated enterprise operations. The alliance pairs OpenAI’s models with IBM’s consulting, governance, cybersecurity, and systems-integration expertise.
Google DeepMind’s SL2T model brings sign-language-to-text input to Gboard and Live Transcribe on Pixel 11, initially translating American Sign Language into English. Developed with Deaf-community input, it lets users sign to search, write, and interact with Gemini.
OpenAI’s ChatGPT Sites lets users turn natural-language instructions into hosted, shareable websites and lightweight apps. Its forecasting demo shows account-level controls updating an interactive financial model and aggregated results.
Grok 4.6 is now available in Grok Build and Cursor, with users receiving double included usage for seven days. The model targets long-running coding agents, complex engineering tasks, and interactive app development.
Tesana demonstrates a complete multiplayer game supporting up to 8v8 players, reportedly built in 1.5 hours for $35 in platform credits. The AI game-building platform generates playable Godot projects from natural-language prompts.
A new paper argues that CLAUDE.md and AGENTS.md files grow indefinitely because adding rules is cheap, while deleting them risks regressions when their original rationale is forgotten. It finds that comments preserving rationale can reduce excess instructions by 99.3% and improve agent instruction-following by up to 23.1%.
B.AI has upgraded its Web Chat and API offerings to DeepSeek-V4-Flash-0731 and DeepSeek-V4-Pro-0813, with no migration required. The rollout gives developers immediate access to DeepSeek’s latest long-context, open-weight models through an existing platform.
Higgsfield’s ChatGPT-connected workflow can generate UGC-style videos, combine multiple AI video models, and score finished clips for virality. It compresses scripting, production, and creative testing into one subscription.
xAI’s Grok 4.6 matches GPT-5.6 Sol on composite intelligence benchmarks and delivers strong coding performance at $2 per million input tokens. A hands-on OpenCode test, however, exposes weaker visual judgment when building production-style frontend experiences.
ASC CLI maintainer Rudrank Riyam is preparing versions 4.1.1 and 4.2.0, with work specifically aimed at improving Rork’s App Store publishing workflow. The CLI automates App Store Connect operations from the terminal, including builds, metadata, TestFlight, signing, and submissions.
Hermes Agent reportedly leads OpenRouter’s agent rankings with 34.9T tokens and first-place positions across four categories. Its strongest differentiator is persistent, always-on orchestration that combines memory, reusable skills, multi-provider routing, and remote messaging.
The App Store Connect CLI’s maintainer is asking users whether they prefer daily releases or a slower cadence of several releases per week. The question highlights the tradeoff between rapid fixes and dependable automation for release pipelines.
A conversation highlights how developers can get more from asc, an open-source command-line tool for automating App Store Connect workflows. It covers releases, TestFlight, metadata, signing, screenshots, and AI-agent integrations.
FromFlow adds an AI agent layer to no-code Discord bots, letting communities support members, automate server tasks, and coordinate events without writing code. Its visual workflows can schedule announcements, open channels, assign roles, and trigger follow-up actions.
MongoDB.local Build Fest brings developers together in San Francisco today for talks on agent infrastructure and the latest approaches to embeddings and reranking. The agenda features Latent Space and Voyage AI speakers focused on practical AI application architecture.
Human Behavior uses AI to analyze session replays, identify product friction, and trigger actions across Slack, Linear, CRMs, and code repositories. Its SDK and autonomous browser agents aim to replace manual dashboard monitoring with a continuous product-improvement loop.
Mem Agent turns notes, meetings, messages, and calendar context into proactive follow-ups. It identifies unfinished commitments, resurfaces the relevant details, and nudges users when there is still time to act.
FluidDocs CLI turns prompts into interactive documents that answer reader questions, generate summaries, and report engagement back to the author. A single command publishes the document, while prompt-based edits update the same link.
Scrimba Explain generates narrated video tutorials from questions, files, links, or code. Its DOM-based playback enables faster explainers with visual aids, code walkthroughs, diagrams, captions, and voiceover.
Execlave is a runtime governance platform that evaluates AI-agent actions before they reach production systems. It combines policy enforcement, human approvals, kill switches, cryptographically signed audit trails, and MCP tool-integrity checks.
Nuphos is an AI-native DevOps workspace where agents learn infrastructure context, investigate incidents, monitor costs, plan migrations, and operate cloud systems. Fine-grained IAM, approvals, memory, and audit trails keep human oversight in the loop.
Skilldocs turns Markdown skills into collaborative workspaces with live cursors, inline comments, real-time rendering, and agent handoff for conversations and diffs.
Qencode MCP lets AI assistants start, monitor, and retrieve video transcoding and processing jobs through natural-language requests. It connects Qencode’s cloud video platform to Claude, ChatGPT, Gemini, Cursor, and other MCP-compatible clients.
Ito runs an isolated copy of your app for every pull request, exercising affected user flows in a real browser before code merges. It posts failures with video, logs, reproduction steps, severity ratings, and likely responsible code.
Cursor’s Compass predictor scores each turn from 0 to 1 based on its predicted likelihood of satisfying the user, helping Cursor Router choose the appropriate model. The author says turns rated most likely to succeed receive a positive performance signal 96% of the time.
Playyy combines AI image generation with an editable canvas where creators can select individual elements, move layers, and apply targeted changes through natural-language prompts. It aims to close the gap between fast AI generation and production-ready visual editing.

Better Stack

Github Awesome

Better Stack

Bijan Bowen

Cole Medin

Better Stack

AI Revolution

Wes Roth

Theo - t3․gg

Rob The AI Guy

Every