Live AI developer news, ranked and linked to original sources.
> ▌

Theo - t3․gg

AI Revolution

Rob The AI Guy

AI LABS

Cole Medin

AI Samson

The PrimeTime

Better Stack

AICodeKing

Theo - t3․gg

WorldofAI

Eric Michaud
Corsair is surging on GitHub as an open-source TypeScript integration layer that connects AI agents, backend services, and user-facing dashboards to hundreds of apps. It combines typed APIs, OAuth, webhooks, multi-tenancy, and permission controls in one self-hostable stack.
Archify turns plain-English system descriptions into typed, validated architecture, workflow, sequence, data-flow, and lifecycle diagrams. Its v2.16.0 release adds a constraint-driven workflow compiler, localization, and safer packaging for shareable HTML/SVG outputs.
Rudrank Riyam is steering App Store Connect CLI toward local feature parity with Fastlane, especially its match code-signing workflow. Testing is still underway, making this a promising preview rather than a finished replacement.
Anthropic’s Beneficial Deployments initiative focuses Claude and engineering support on global health, life sciences, economic mobility, and education through partnerships with nonprofits, governments, and researchers. Its official page describes an impact program—not the pre-deployment risk framework claimed in the source post.
Warmwind OS publicly launches autonomous cloud workers that operate existing software through visual mouse and keyboard control, eliminating the need for custom API integrations. Businesses can train workers on repetitive workflows, schedule them, and run multiple isolated instances in parallel.
Naive Prompt Optimization (NPO) uses a teacher model, rollout traces, and rewards to iteratively revise a single prompt lineage. The paper reports comparable or better results than GEPA on IFBench and HotpotQA with slightly fewer rollouts, plus promising cross-model transfer.
Vercel AI Gateway is offering 50% off MiniMax H3 and H3 Max from August 30 through September 13, covering every supported duration and aspect ratio. Existing model IDs remain unchanged, so developers can use the discount without code changes.
A fresh X video shows EngineAI’s full-size T800 humanoid robot performing fast, martial-arts-style movements and walking in public. Its 1.73-meter frame, 450 N·m joint torque, and up to 29 body degrees of freedom make it a serious embodied-AI platform, though the footage demonstrates choreography—not autonomy.
A follow-up X thread suggests xAI’s next flagship model could arrive in September, but the timeline remains an informal estimate rather than an official release date. Grok 4.7 is still absent from xAI’s model documentation, which currently lists Grok 4.6 as its flagship developer model.
Claude Code is being used to iterate on Californication, a Three.js 3D game prototype, with isolated tests for character jump behavior. The workflow shows an agent evaluating motion visually before integrating improvements into the playable scene.
Delegate Skills adds a review-first delegation loop to Hermes and other orchestrators: write a self-contained brief, hand it to a separate coding CLI, inspect the diff, rerun tests, and keep the commit under reviewer control. It supports direct dispatch or named lanes for tools such as Claude Code, Codex, Cursor, and OpenCode.
The video spotlights Google’s Mantis, an Apache-2.0 toolkit that breaks repository security work into planning, code review, exploit reproduction, patching, and reporting. Its agent-agnostic design supports tools such as Gemini CLI and Antigravity, while requiring expert verification and isolated execution.
A new Hermes skill reads Claude Code and Codex session transcripts, then extracts valuable content ideas users might otherwise miss. It uses a very cheap model to prioritize speed and low operating costs.
OpenAI’s Tibo Sottiaux says Codex and ChatGPT Work usage will reset at 6pm PST, encouraging users to spend their remaining allowance exploring newer features and /fast mode. The move gives power users a fresh window for agentic coding and multi-step work.
Stripe Link lets AI agents request purchases while users approve transactions from the Link app. Riley Brown demonstrated the workflow through GrokBot, approving a $22.50 purchase without exposing his card details to the agent.
OpenAI engineer Brent Traut says a recent ChatGPT desktop update makes long threads load over 90% faster and reduces their memory footprint by over 90%. Traut’s post prompted praise from OpenAI’s Tibo Sottiaux.
Synara says its next update arrives tomorrow with Devin Desktop joining the app as a provider, expanding its local-first workspace for coding agents. The announcement also teases several additional features without detailing them yet.
Alibaba Cloud’s Wan 3.0 generates native 30-second videos at up to 1080p, with synchronized audio and multimodal references including documents and webpages. Its all-in-one API consolidates text-to-video, image-to-video, reference generation, editing, and extension workflows.
A Stanford-led research team proposes Prefix Sliding, which keeps only task instructions and recent reasoning tokens in the KV cache while discarding stale intermediate steps. The paper reports up to 3× faster inference without retraining and reasoning traces beyond 100,000 tokens with reinforcement learning.
prmpt is an open-source plugin that inserts clearly labeled, context-matched ads into Claude Code, Codex, Gemini CLI, and Amp workflows. Developers receive 70% of impression revenue directly in crypto wallets on Base or Solana.
Sonar has made SonarQube Hunter Agent generally available on SonarQube Cloud, adding full-codebase reasoning for broken access control, business-logic, and authentication flaws that pattern-based SAST can miss. It validates suspected findings before adding them to the existing SonarQube issue workflow; SonarQube Server support is coming soon.
Pieter Levels switched Infinite Slop’s AI-generated livestream from landscape 16:9 to portrait 9:16 after finding that most traffic comes from mobile users. The move better fits TikTok-style viewing, though supporting multiple formats would significantly increase generation costs.
Google employees are reportedly testing a Gemini 3.8 Flash Preview build, with early feedback suggesting improvements over Gemini 3.7 Flash’s rough edges and sloppy outputs. Google has not confirmed the model, specifications, API access, pricing, or public release date.
Stanford’s CME 295 offers nine public lectures covering Transformers, LLM training and tuning, reasoning, agents, RAG, and evaluation across roughly 16 hours. Its official syllabus provides a coherent path from tokenization and embeddings to current LLM systems.

ChronoRAG-G is a temporal RAG framework that assigns each answer requirement evidence tied to the correct valid time while separating fact time from publication, filing, or release time. Its announcement reports 80.70% audited accuracy and 101/101 correct refusals on unanswerable cases.

TreasuryForge is an autonomous treasury demo for simulated cash, crypto, and NSE equities, built on TrueForge for the WeMakeDevs × TrueFoundry Agent Harness Hackathon. Its loop computes risk server-side, stress-tests breaches in a sandbox, and pauses every trade for explicit human approval.
Zod 4.5 adds ahead-of-time schema compilation, faster failure paths, new validation APIs, and up to 9x lower schema memory usage. The release strengthens Zod’s position as TypeScript’s default runtime-validation layer.

This MIT-licensed Agent Skill turns project documents, source code, and patent PDFs into Chinese technical disclosures, plain-language patent notes, Obsidian knowledge graphs, prior-art searches, and draft office-action responses. Recent commits add classification-based searches, response-strategy routing, and automated structural line-art assembly.
Sepia is an open-source Agent Skill that repairs AI-written fiction at the narrative-architecture level before polishing prose. It also includes venue-specific workflows for technical articles, release notes, PR replies, postmortems, and tickets.
A tester video shows a Cybercab navigating Austin near Terry Black’s BBQ without a steering wheel, pedals, or visible safety monitor. The sighting suggests Tesla is moving toward unsupervised validation ahead of its planned launch event, though public availability remains unconfirmed.
Tsinghua’s open-source AI classroom platform has reached v1.0 with a Pro agent workbench for planning, building, and revising courses from user materials. It adds durable sessions, course-building tools, multimodal inputs, web search, and provider-neutral deployment options.
Grok Bot now connects to X, automatically creating a developer account and providing initial API credits for eligible users. Bots can search posts, read timelines, check mentions, and summarize activity across X.
ASCII has updated box’s default VM image to include popular agent harnesses, reducing setup friction for developers running agents in full Linux environments. SSH commands, templates, and snapshots keep the image fully customizable, including for experiments with Prime Agent and Opus.
Infinite Slop now lets viewers upvote prompts in its generation queue, collectively deciding which AI video scene airs next. Pieter Levels’ experimental livestream turns chat requests into an endless, audience-steered stream powered by fal’s accelerated MiniMax H3 Max.
Meta’s internal Project OT envisioned AI agents absorbing much of thousands of employees’ daily work under smaller human teams. Reuters reported that Meta canceled its second restructuring wave after major technical and security incidents rose 40%, while Zuckerberg later acknowledged agent development was progressing slower than expected.
box by ASCII is positioning its full-VM sandbox for AI agents around aggressive pricing and performance claims, including 9x lower costs than Daytona and 18x lower than Modal. It combines persistent Ubuntu VMs, SSH, Docker, snapshots, forks, and desktop access for agentic development. [Founder’s post](https://x.com/AniC_dev/status/2094019615814738350) [Product homepage](https://box.ascii.dev/)
A developer reports running Codex for eight hours with GPT-5.6 Sol at Extra High and Luna Max handling subagents, using just 29% of the weekly allowance. The result suggests quota-aware model routing can make long-running agent workflows more economical, though it remains an anecdotal measurement.
The Center for Democracy & Technology commissioned a Public First survey of 2,000 British adults, finding that 93% believe they have a right to private online conversations and only 12% support secret government orders for access. The findings arrive amid the UK’s dispute with Apple over encrypted iCloud data.
Perplexity’s Portable Computer runs its agent stack locally on NVIDIA DGX Spark, keeping private files and routine execution on-device while escalating difficult or current tasks to cloud models with user approval.
Reported identifiers claude-marshmallow-eap and claude-melon-eap are fueling speculation about an Anthropic Opus 5.1 update, but neither model has been officially confirmed. No public model card, API listing, pricing, or release date has surfaced.
Reported zero-shot demos of OpenAI’s upcoming Astra model show playable games, detailed websites, 3D objects, and voxel environments generated from single prompts. Astra’s final public name and release timing remain unconfirmed.
Anthropic will replace Claude Code’s temporary 50% weekly usage boost with a permanent 25% increase over the original baseline on September 14. Pro, Max, Team, and seat-based Enterprise users will therefore have roughly 17% less capacity than they do today.
Superagent is an open-source macOS workspace that gives Claude Code a persistent visual home, with browser automation, iOS Simulator control, git worktrees, scheduled routines, and optional phone approvals. It runs locally on a user’s existing Claude subscription without an account or separate server.
Prequel is a native macOS screen recorder that automatically adds cinematic zooms, cursor-aware camera movement, backgrounds, rounded framing, and webcam overlays. It exports polished demos at up to 4K using Apple Silicon media acceleration.
Murfy AI launches an AI-native LaTeX workspace for drafting, reviewing, compiling, collaborating on, and publishing research papers. It combines contextual AI agents with real-time editing and unlimited collaborators.
Referent launches AI-native practice-management software for solo and small law firms, with agents handling intake, matter management, email workflows, deadlines, follow-ups, and billing preparation while lawyers approve critical client-facing actions. It also connects firm context to mobile apps and MCP-compatible assistants.
Hyperfocus is a free, macOS-only planner that connects long-term goals to weekly plans, daily tasks, and focused work sessions. Its optional AI coach flags vague or unrealistic plans, while local-first storage keeps core data offline.
Maritime provisions each customer-facing AI agent in an isolated, persistent micro-VM with browser access, files, code execution, and automatic sleep/wake, starting at $1 per agent monthly. Its CLI, SDK, templates, and Docker support target teams scaling agent products without building their own VM fleet.
Ulpaso is a free, MIT-licensed macOS app that transcribes microphone and system audio on-device, organizes the result by speaker, and saves it as editable Markdown. It requires no account, cloud upload, telemetry, or subscription.
Topview Motion Studio turns a product brief, reference images, visual style, duration, and aspect ratio into a connected launch video. It targets teams that need polished app, SaaS, hardware, or feature-announcement videos without After Effects expertise.
Skild AI’s S1 uses a single video demonstration as an in-context prompt for unseen, long-horizon manipulation tasks without task-specific fine-tuning. The company reports runs lasting up to 10 minutes and 66% per-step success on unseen tasks versus 9% for language prompting.
htmx 4.0.0 is now available after eight months of development, replacing XMLHttpRequest with fetch and adding morph swaps, streaming extensions, and the hx-live scripting layer. The release preserves htmx’s HTML-first model while introducing explicit inheritance and other migration-sensitive changes.

AI Search

Better Stack

AI Revolution

The PrimeTime

Eric Michaud

Rob The AI Guy

Wes Roth

The PrimeTime