A day in AI

Friday, 25 September 2026

40 stories · 1 from multiple sources

More stories

Revealing the details of how OpenAI agents hacked Hugging Face

swarmtraces.org ·

Too AI; Didn't Read

tai-dr.com ·

OpenRouter: from Seed to Stripe — with OpenRouter’s Alex Atallah & AMP’s Anjney Midha

latent.space ·

Pope: AI must serve humans not become tool of domination

euronews.com ·

Tell HN: OpenAI $500 ProMax plan listed in API

news.ycombinator.com ·

To Trust or Not to Trust: Retrieval-Augmented Fact Checking in Speech

arxiv.org ·

Also today

PoEM: Predicting RL Outcomes from Existing Policies

arxiv.org

Does a model's stated reason for rejecting a candidate do any work?

arxiv.org

GRASP: Generating, Revising, and Assessing for Strategic Planning with Agentic AI

arxiv.org

How Reproducible Are Evaluation Conclusions? A Self-Audit of LLM-Inferred Prompt Structure

arxiv.org

The Phone Number puzzle

youtube.com

Crusoe abandons $1.25B plan to use Boom turbines at AI data centers

techcrunch.com

Unsecured OpenAI agents posted 53 user images on the internet without the lab’s knowledge

techcrunch.com

Token Forecaster — Forecast how long an LLM reply will run, before Enter

producthunt.com

Jango — Test multi-user apps with AI agents that act like real users

producthunt.com

Proaction boosts sales 60% and saves 75+ hours with Codex

openai.com

Once UI 2.0 — Builds consistent React apps for developers and AI agents

producthunt.com

Opus 5.5: How Close Are We to Automated AI Research?

youtube.com

Quoting John Gruber

simonwillison.net

Basedash MCP write — Build charts and dashboards from Cursor and Claude

producthunt.com

Agentic Detection of Online Conspiracies

arxiv.org

Minimally Invasive Steering of Language Models

arxiv.org

ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds

arxiv.org

Do Audio Language Models Hear and Read Distinctive Features Alike?

arxiv.org

Screen Before You Serve: Simulation for Production Customer Experience AI Agents at 140M Scale

arxiv.org

R-DEIM Net: An Efficient Rationale-Augmented Dual-Expert Interaction Model for Paraphrase Detection

arxiv.org

PrivDrift: Auditing User-Secret Leakage Under Topic Drift in Active LLM Conversations

arxiv.org

Return or Revise? Learning When Revision Helps Retrieval-Augmented QA

arxiv.org

Ollaya – Ollama for open-source, Jev-style decision models

ollaya.dev

Meta's Muse appears to use an OpenAI model labeled muse-special

mouse.dev

Show HN: Doom or Bloom, map your AI worldview with Jev

doom-or-bloom.com

GitHub Copilot app for Beginners: How to build custom workflows with canvases

github.blog

Hyperdream — The Cursor of AI filmmaking

producthunt.com

Anthropic posted cringe on main

xeiaso.leaflet.pub

Agents can now set up your website’s security with Turnstile Spin

blog.cloudflare.com

Scaling MoE reinforcement learning on Amazon EKS with EFA and DeepEP with 40% more throughput

aws.amazon.com

Accelerate multimodal RL training with SkyRL on Amazon SageMaker HyperPod

aws.amazon.com

← All editions