Live AI developer news, ranked and linked to original sources.
> β

Better Stack

DIY Smart Code

Better Stack

Every

Income stream surfers

Every

Rob The AI Guy

DIY Smart Code

Every

Every

Every

AI LABS

Every

Better Stack

Discover AI

Every

DIY Smart Code

AICodeKing

WorldofAI

Better Stack
Anthropicβs Claude Code Mods let developers extend Claude Code with TypeScript plugins that can rewrite prompts, intercept tool calls, replace features, and render custom terminal or desktop UI. The system works through plugins in both the CLI and desktop app.
XFreeze published Explainer Bot, a custom Grok Bot designed to turn confusing topics into deeper, clearer explanations. It packages a personal learning workflow into a reusable AI assistant.
Genesis AiX Build 17 is now available for public review, advancing its governed-intelligence architecture around a core principle: AI can reason, retrieve, recommend, and generate without automatically gaining authority to cause consequential action.
A technical handbook explains how speculative decoding reduces autoregressive latency by having a cheaper proposer draft several tokens before the target model verifies them together. It covers greedy verification, exact sampling, acceptance rates, and the systems tradeoffs that determine real-world speedups.
Vercel says Rogo uses its platform to rebuild internal apps so coding agents can move agent-written code into production in about five minutes. The finance AI company runs six production agents across churn analysis and deal-desk workflows, surpassing 73,000 monthly deployments.
falβs H3 Max reference-to-video endpoint now supports first, middle, and last image keyframes, with configurable timing for the middle frame. Developers can also combine image, video, and audio references in a single generation.
Google Labs gives Gemini users a central directory for experimental tools spanning no-code mini-apps, model evaluation, and asynchronous coding agents. The lineup makes Googleβs developer-facing AI experiments easier to discover and try.
SpaceXAI has reset usage limits for every Grok Bot user, giving its always-on AI teammates a fresh quota. The reset is a temporary refill, not a change to Grok Botβs underlying weekly usage model.
OpenAI is rolling out Finances in ChatGPT to Free and Go users in the U.S., enabling account connections through Plaid and credit data through Experian. Users can track subscriptions, recurring bills, spending trends, duplicate payments, and weekly financial updates through natural-language queries.
cjavdevβs open-source Claude Code mods add a cache-expiry countdown, context-window meter, and idle-session summarizer. The tools expose hidden cost and context signals directly in the terminal UI.
Cloudflare released Clef and Clef-flash, open-source decision models that return typed probabilities instead of generated text. Available through Workers AI and compatible with Jevβs API, the 27B and 9B models target fast agent routing, classification, guardrails, and multimodal decisions.

VectifyAIβs new Apache-2.0 project combines Jevβs structured decisions with PageIndexβs hierarchical document trees to locate answers in reports too large for flat prompts. The examples search NVIDIA and Citigroup filings without vector databases or embeddings.
Samaneh Saadat is inviting developers to an open Keras community meeting on October 2 at 10 AM PT to discuss the ecosystemβs latest developments. It is a roadmap and community update, not a documented new release.
Synara now opens pull-request links in its native PR view by default, while Command/Ctrl-click still opens GitHub externally. Users can also change the behavior in Settings.
Brazilian AI lab LUA Visionβs Genesys PI appears to be its own model family, not merely a GLM 5.3 wrapper, and an independent coding-agent benchmark placed its House tier alongside Grok 4.7 at 83.5. The lab is also reported to use boatβs persistent Linux VM sandboxes.
Mondobeβs essay argues that AIβs rapid adoption is eroding software craft, creativity, startup defensibility, and the quality of online discourse. The author remains hopeful that society can rebuild healthier institutions as AIβs role becomes clearer. [Essay](https://mondobe.com/ai-makes-me-sad) [HN discussion](https://news.ycombinator.com/item?id=49934487)
Synara is experimenting with an Inbox view that summarizes developer work across usage, tokens, prompts, models, and projects. The feature adds a personal activity layer to its local-first workspace for coding agents.
OpenBot compares its free, local-first desktop workspace with ChatGPT Dots, OpenAIβs always-on cloud agents. OpenBot runs Codex, Claude, Gemini, Grok, or custom models on your computer using subscriptions you already own, while Dots prioritize managed cloud execution and convenience.
Osmo Film Emulation is a free browser-based color-grading tool with 126 film stocks and looks, adjustable grain and halation, plus DaVinci Resolve and Premiere Pro plugins. It processes footage locally, supports major camera log formats, and is available at https://osmo.inc/tools/film-emulation.
Amazon is exploring a vehicle to sell roughly $8 billion of Nvidia Grace Blackwell chips to outside investors, then lease them back for use across U.S. data centers. The structure would shift expensive AI hardware off Amazonβs balance sheet while preserving access to the compute.
A new paper introduces RADAR, a real-time method for detecting when large reasoning models drift from productive reflection into redundant or persistent generation loops. Attention realignment reduces these failures while largely preserving normal performance.
UC Berkeley researchers introduce CICM, a benchmark showing that LLMs can retain updated facts yet answer with outdated values because attention selects stale mentions. The paper proposes training-free attention redirection that corrects most tested errors.
NVIDIA researchers introduce Long-Transduction, a controlled diagnostic for testing whether models can repeatedly read, mutate, and output state-dependent results across long contexts. Across open-weight models, performance fell 62.8% as context expanded from 4K to 128K, with additional losses from input-format variation and task complexity.
OpenBot now lets users run their AI-agent teams on hosted Linux servers instead of leaving a personal computer online. Users choose a server plan, connect an AI provider, and access agents remotely.
Pi 0.99 adds built-in MCP, model-written JavaScript workflows through Codemode, and on-demand tool search, alongside virtual and classifier model support. The release makes Pi a more capable, extensible terminal agent while preserving its open-source, customizable architecture.
PiG 0.3.0 ports Pi 0.87.1 into a native Go coding-agent binary, improving TypeScript extension compatibility, MCP support, provider coverage, session handling, and cross-platform behavior. It offers a compact, open-source alternative for developers who want Piβs workflow without making Node.js the core runtime.
AWS released Strands Decider 2B, an open-source 1.9B-parameter model that selects among predefined options, scores inputs, and returns calibrated confidence without generating text. It runs locally in roughly 115ms on an RTX 3090 and includes its training data and scripts. [Announcement](https://strandsagents.com/blog/introducing-strands-decider/)
Synara now shows a pencil icon beside threads containing unsent prompts, helping users recover unfinished work after switching conversations. The small UX improvement strengthens its local-first workspace for managing multiple coding-agent sessions.
ChatGPT now lets users virtually try on clothing and accessories using selfies or product images, then save products in Favorites. The experience is powered by ChatGPT Images 2.5, though OpenAI cautions that results do not guarantee fit or accurate appearance. [OpenAI Help Center](https://help.openai.com/en/articles/11128490-shopping-with-chatgpt-search)
Anthropicβs open-source webapp-testing Agent Skill gives coding agents a repeatable Playwright workflow for testing local web apps, including server management, browser interactions, screenshots, and console inspection. The video demonstrates its value for responsive-layout and booking-flow validation.
OpenBot 0.27.0 adds Cursor as a first-class provider, expands global search across channels, files, routines, commands, and settings, and redesigns the Marketplace. The local-first workspace also improves server administration and provider management.
UniEvo-VL is an on-policy self-distillation framework that lets a multimodal model critique its own images, then transfer those corrections into future generations. Built on Qwen-image-2512, it raises GenEval from 0.747 to 0.808 and GenEval2 Soft-TIFA from 32.97 to 35.53.
Melayaβs September recap presents an AI workspace spanning marketing, browser and phone control, event-driven agent pipelines, Google OAuth, Databricks, and Claude Code/Codex workflows. The platform now touts 8,350+ tools, 48 AI providers, governed approvals, and security validation from TAC Security and CASA.
MLX-Serve 26.10.1 delivers up to 66% faster decoding and 51% faster prompt processing across 18 models, with identical outputs to 26.9.6. Qwen3.8-27B with its drafter gains 66% on M5 Ultra, 28% on M4 Max, and 37% on M1 Pro. ([release notes](https://github.com/ddalcu/mlx-serve/releases/tag/v26.10.1))
Wu is an open-source Rust editor forked from Zed, redesigned around VS Codeβs familiar layout and workflows. It removes built-in AI, collaboration, accounts, and telemetry while emphasizing native GPU rendering and lower memory use.
Famulor launches AI voice agents that answer inbound calls, run outbound campaigns, and continue conversations across WhatsApp, email, SMS, and web chat. Its no-code platform combines shared context, CRM integrations, appointment booking, and EU-hosted, GDPR-ready operations.
Gauth AI Course now teaches on a zoomable, continuous whiteboard where narrated chapters, formulas, visuals, and tutor answers stay connected spatially. Students can generate lessons from topics or documents, complete Quick Checks, export PDFs, and share sessions.
ElevenLabs launched new expressive text-to-speech models built on a new architecture, supporting 90+ languages, stronger speaker consistency, inline performance controls, and Professional Voice Clones. V4 Turbo targets real-time agents with roughly 100 ms median inference latency and streaming support.
Finbar makes its AI-native investment research platform broadly available, combining worldwide fundamental data with web and Excel workflows, prebuilt models, and API/MCP access for agents. Its platform is already used by major hedge funds and supports machine-readable reports, financial datasets, and research tools. [Finbar](https://finbar.com/) [Finbar Docs](https://docs.finbar.com/api-mcp/introduction)
Anthroposcaper converts browser-drawn or DXF-imported 2D plans into editable 3D urban scenes by placing user-supplied GLB templates along tagged edges, corners, and junctions. It rebuilds scenes as plans change, exports GLB files, and supports AI-assisted setup through MCP. [Product Hunt](https://www.producthunt.com/products/anthroposcaper)
Open Inspector is a free, MIT-licensed Chrome extension that reveals a webpageβs layout, styles, colors, typography, assets, and design tokens without host permissions or network requests. It exports findings as CSS variables, Tailwind configuration, W3C tokens, or markdown handoffs.
Syllabyβs Avatars 2.0 turns scripts or ideas into presenter-led videos with customizable avatars, natural voices, scene controls, B-roll, and subtitles. It targets creators and businesses seeking repeatable social video without cameras, actors, or complex editing workflows. [Syllaby](https://syllaby.io/features/avatars-2-0/)
Codync is a free, MIT-licensed alternative to Grok Bot, Muse, and Dots that turns Claude Code, Codex, Cursor, and 40+ agents into named bots accessible across iPhone, Mac, Linux, and terminal. Its Rust host keeps code, transcripts, and memory local while enabling encrypted remote approvals and collaboration. [GitHub](https://github.com/leepokai/Codync) [Product site](https://www.codync.dev/)
Never Boring AI interviews users about their work, remembers their stories, drafts posts in their voice, learns from edits, and schedules approved LinkedIn publishing. Its product hunt launch positions it as a memory-and-workflow layer beyond generic AI copy generation. [Product Hunt](https://www.producthunt.com/products/never-boring-ai)
JarvisCore is an Apache-2.0 Python runtime for autonomous agent fleets, combining peer-to-peer discovery, shared work execution, durable memory, and brokered credentials. Its launch targets production workloads where centralized orchestration and exposed API keys become reliability and security liabilities. ([Product Hunt](https://www.producthunt.com/products/jarviscore), [GitHub](https://github.com/Prescott-Data/jarviscore-framework))
Mintlify Desktop is a beta native app for managing internal and public documentation across Mac, Windows, and Linux. It combines persistent tabs, split views, local offline files, cross-organization sessions, and an embedded documentation agent.
Lloyal launches a TypeScript platform for building downloadable AI apps with built-in open-weight inference and multi-agent orchestration. Developers can target desktop, web, or terminal without API keys, Docker, or a separate inference server.
Communicate builds AI customer-support agents from help docs, files, and past replies, then routes uncertain cases to humans in a shared inbox. It also supports approved tasks, website chat, analytics, REST API access, and read-only MCP.