Live AI developer news, ranked and linked to original sources.
> ▌

Github Awesome

AI Revolution

DIY Smart Code

Wes Roth

WorldofAI

OpenAI

Rob The AI Guy

Better Stack

OpenAI

Every

Every

The PrimeTime

Income stream surfers

AI LABS

Better Stack

Discover AI

Theo - t3․gg

Wes Roth

Two Minute Papers

WorldofAI
Try Omarchy packages Omarchy Quattro, an ARM64 Arch Linux image, QEMU, Apple’s Hypervisor Framework, and a Swift launcher into a signed, notarized macOS app. Its latest update adds camera, clipboard, folder sharing, port mapping, and sharply lower idle CPU usage, though video decoding remains CPU-only.
System Design 101 is a visual open-source reference covering APIs, databases, caching, distributed systems, cloud architecture, security, DevOps, and real-world system case studies. It is especially useful for interview preparation and architectural orientation.
Bernie Sanders and Greg Casar announced forthcoming legislation that would permanently ban superintelligent AI and pause advanced AI development until federal safety rules exist. The proposal also calls for global agreements, export controls, and penalties of up to 20 years in prison.
ARC Prize reports GPT-6 Astra scoring 99.9% on ARC-AGI-3 Semi-Private with a Provider Adapter harness, versus 62.7% under the standard harness. Astra also used fewer actions than median human participants on 96% of levels, marking a major interactive-reasoning milestone without proving AGI.
Playco’s AI IDE connects GPT-6 Astra directly to Unity and Godot, letting models edit scenes, run games, test changes, and validate results. OpenAI reports three themed prototypes from one grey-box foundation with 50% fewer manual fixes.
Armature analyzed 16,893 coding-agent sessions across 75 repositories, 1,163 task variations, and Claude Code, Codex, and Cursor to see which third-party services they actually install. Only 42% of agent decisions matched, with programming language and repository context often changing the winner.
OpenAI’s GPT-6 Astra is a frontier model for computer use, browsing, software engineering, science, cybersecurity, and professional workflows. It is rolling out to select organizations before expanding to paid ChatGPT users and API developers.
Cerebras now offers Qwen3.8-27B on public inference endpoints at approximately 1,500 tokens per second, with 64K free-tier and 128K paid-tier context limits. The 27B open-weight vision-language model supports image/video understanding and controllable reasoning.
Magnitude is an Apache-2.0 inference server and coding agent that profiles hardware, recommends compatible local models, then downloads, tunes, and runs them for tools including Codex, Claude Code, OpenCode, and Cline. Its latest CLI release adds stronger health checks and clearer context-length errors.
grok-bot-cli is an MIT-licensed Node.js CLI for creating and managing Grok Bot agents and groups, sending tasks, inspecting threads, and automating cleanup. It requires Node.js 18+ and a signed-in Grok Bot macOS app, reusing its encrypted session credentials instead of requiring token copying.
ChatGPT, Claude, and Grok experienced widespread service disruptions on September 3, affecting ChatGPT, Codex, multiple Claude models, and Grok. OpenAI, Anthropic, and xAI each confirmed elevated errors and began recovery efforts.
A new arXiv paper introduces ASKS, which converts scientific sources into readable Wiki views, validated GraphDeltas, and persistent knowledge graphs with source-level provenance. A 56-paper demonstration produced a traceable research map spanning tensor networks, quantum many-body physics, machine learning, and quantum AI.
Terminal-Bench 4.0 calibrates CPU, memory, and timeout resources, fixes 19 tasks, and removes eight saturated or unreliable tasks. The resulting 66-task benchmark reduces infrastructure noise, but its scores are not directly comparable with version 3.0. [Announcement](https://www.tbench.ai/news/terminal-bench-4-0)
Fillo is headless form infrastructure that lets coding agents add native forms directly inside products. It handles schema, validation, uploads, responses, versioning, notifications, and webhooks while developers keep control of the UI and route.
Coworker AI launches OM2, a continuously learning organizational-memory layer that connects to 50+ tools through MCP or API. The company says it cuts AI context costs by up to 9x, improves speed by 64%, and earns an 84.5% quality preference in its benchmark.
Higgsfield Genjutsu transforms existing footage through Motion Transfer and Object Swap, letting creators change characters, products, outfits, locations, or styles while preserving the original motion and shot structure. It targets faster ad variations, content repurposing, and localized campaigns.
Tabbit AI is an AI-native browser whose Agent Mode can research, navigate websites, fill forms, and complete multi-step workflows using tabs, screenshots, PDFs, bookmarks, and local files as context. It delivers outputs such as HTML, PDFs, and presentations, while reusable Skills turn recurring workflows into commands.
Format analyzes customer conversations across video, voice, and text, turning them into personalized reports, podcasts, and queryable insights. It integrates with tools including Gong, Fireflies, Granola, Slack, Intercom, Zendesk, and HubSpot, with evidence-linked quotes and audio.
MagiCrew is an open-source AI agent platform that coordinates specialized digital workers for research, analysis, presentations, reports, and other business tasks. It combines shared context, multi-agent workflows, deliverable-ready outputs, self-hosting, and enterprise controls.
Airtop’s Agent Builder turns plain-English workflows into tested, reusable automations for websites, APIs, and authenticated apps. When workflows break, it diagnoses failures, proposes repairs, and validates fixes through real test runs.
Grove is an AI-native macOS script runner that lets developers and coding agents control the same processes through a menu bar app, CLI, and MCP server. It discovers project scripts, shares logs and port status, and can explain crashes locally with Apple Intelligence.
Readr is a native iPhone, iPad, and Mac ebook reader that answers questions in the margin using context from your book, reads aloud from your current page, and turns highlights into editable Markdown articles. It supports DRM-free EPUBs, PDFs, BYO OpenAI or Anthropic keys, and fully offline Ollama models.
Agent Looker is a private-beta security layer that intercepts URLs, pages, and tool calls for AI agents, blocking phishing, malware, and prompt injection. It targets Claude Code, Codex CLI, OpenClaw, and MCP-compatible workflows.
ARBR is an open-source, MIT-licensed control plane for routing, governing, observing, evaluating, and deploying AI workloads across providers. Its OpenAI-compatible gateway lets developers add model switching, budgets, telemetry, and rollback without rewriting application integrations.
Omi’s desktop app captures screen activity and conversations, then turns them into searchable memories, summaries, tasks, reminders, and cited answers. It is open source, local-first, supports custom AI keys, and launches today for Mac users.
Nex turns plain-language goals into GTM workflows across 100+ integrations, handling lead qualification, enrichment, outbound, CRM cleanup, and deal re-engagement. Built by ex-HubSpot operators, it targets the reliability and cost problems general-purpose agents face at production scale.
Causal gives designers, founders, and vibe coders a Mac-first workspace for organizing notes, images, files, and links spatially. Its native agent can generate and arrange canvas content, create widgets, and hand projects to external agents through MCP.
Tidy is a free macOS utility that fixes spelling and grammar in any app through keyboard shortcuts, using Apple’s on-device model without uploading selected text. It also includes a dedicated shortcut for removing AI-sounding phrasing and preserves mentions, dates, and times.
Blume.codes is a local-first desktop sidecar for Claude Code, Codex, and Cursor that monitors agent sessions and turns repeated corrections into proposed rules, skills, hooks, or documentation. Developers review the evidence and diffs before applying changes.
Atlas is an open-source Rust/Tauri workspace that runs Claude Code, Codex, and its native agent side by side while linking commits to sessions, prompts, and changes. An 888-star daily surge on GitHub Trending highlights growing demand for agent workflow provenance. [Atlas README](https://github.com/pacifio/atlas)
A hands-on comparison found Qwen3.8-27B scoring 98.7 across 21 tests while running a 64K context fully on a 24GB GPU. The dense open-weight model targets coding, research, professional work, and long-horizon agents.