OpenAI's rogue agent didn't stop at Hugging Face - here's what we know
The same autonomous OpenAI agent that escaped its test environment and breached Hugging Face was also busy hacking other AI systems. Lucky us.
The same autonomous OpenAI agent that escaped its test environment and breached Hugging Face was also busy hacking other AI systems. Lucky us.
OpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock, along with explicit prompt caching that gives you precise control over which parts of your prompt are cached and reused. Learn how to get started, set up explicit caching, and migrate existing GPT workloads to reduce inference cost.
Avi Schiffmann has a new version of his controversial AI companion. It’s more expensive, and you can’t change its personality.

Two AI computing clusters built from identical NVIDIA H100, GB200 NVL72, or GB300 NVL72 systems can deliver materially different training throughput. We...

Three approaches to the issue of quantum results that can't be verified classically.

Amazon Bedrock Advanced Prompt Optimization optimizes your prompts for up to 5 models at once and compares original versus optimized performance across quality, latency, and cost. Migrate to a new model or improve your current one in minutes instead of weeks.
I managed to get a subscription to kimi and have been running kimi-code for the last few hours, having imported my claude-code skills, settings etc into it. I pointed it at a JS and python codebase I have been working on and asked it to review and refactor, then I got it to add some features - and boy, I am not impressed. TBH it looks the part with the running commentary it makes, but it seems to make some dumb mistakes. Compared to claude-code it is slow, time consuming, hungry on the tokens an
Spotify’s podcast platform has become chronically unreliable since the company’s leadership started boasting about AI adoption. Competitors haven’t had similar issues, so I offboarded from Spotify.

Article URL: https://zozo123.github.io/wasted-cycles/ Comments URL: https://news.ycombinator.com/item?id=49111656 Points: 1 # Comments: 1
Maybe they have seen AI fail for so long due to poor performance that it makes AI safety experts seem like they are vastly overestimating what AI can do? Comments URL: https://news.ycombinator.com/item?id=49111613 Points: 3 # Comments: 12
<p><strong>Release:</strong> <a href="https://github.com/simonw/llm-chat-completions-server/releases/tag/0.1a0">llm-chat-completions-server 0.1a0</a></p> <p>A key goal of the new content-addressable logs <a href="https://simonwillison.net/2026/Jul/30/llm-rc1/">in LLM 0.32rc1</a> was being able to support OpenAI Chat Completion style requests where each incoming message extends the previous conversation, like this:</p> <pre><code>curl http://localhost:8002/v1/chat/completions \ -H 'Content-Type:
Meta says AI is making it dramatically easier to build and launch new consumer apps, with CEO Mark Zuckerberg telling investors the company has more new consumer products on the way.
Hi HackerNews, I've been working with multimodal agentic systems ever since the ImageNet and DQN days. When I started vibecoding last year, I came up with new ways to set up projects to avoid known issues in long-horizon task execution. I thought it would be great if I had one simple file that a coding agent could read when I start a new project to address these known issues while improving efficiency and autonomy long-term. OpenMetaHarness is a cognitive architecture that enables coding agents
<p> Prompt Engineering As a Sport </p> <p> <a href="https://www.producthunt.com/products/prompt-golf?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1210857?app_id=339">Link</a> </p>
Almost all field service organizations use AI, and revenue gains in key areas offer insights for professionals in other business functions.
<p><strong>Release:</strong> <a href="https://github.com/simonw/llm/releases/tag/0.32rc1">llm 0.32rc1</a></p> <p>This RC for LLM 0.32 finishes the work that <a href="https://simonwillison.net/2026/Apr/29/llm/">started in LLM 0.32a0</a> - it adds a <a href="https://llm.datasette.io/en/latest/logging.html#the-message-store">new schema design</a> that does a much better job of capturing the details of the prompts and responses returned by the latest model families.</p> <p>The most important change
TL:DR Learn how Meta’s FBTriton infrastructure powers custom GPU compiler innovations like TLX and autoWS while staying synced with upstream Triton using agentic ingestion and a stratified L1/L2/L3 validation framework....

A lot of security still comes down to trusting the wrong screen. This week, that screen might be a login page, an install guide, a recruiter call, or a familiar service behaving slightly wrong. Behind it: reused credentials, exposed systems, quiet loaders, abused trust, and exploit paths that should have been harder. Some defenses improved. The loose parts still got found first. Anyway,

Article URL: https://trysecondstate.com/ Comments URL: https://news.ycombinator.com/item?id=49111295 Points: 3 # Comments: 1
British AI neocloud Nscale is buying software startup Anyscale, which helps companies scale their AI workloads across data centers and servers.
Hello HN! I wanted to share a small MCP + web app I built to solve a problem at work. Essentially, we’ve been doing great when working with agents 1:1, but collaboration has been a struggle. E.g. when I have to collaborate with a teammate and share context or handoff work usually what I would do is create a markdown with my agent, paste in slack, then my teammate will copy and paste it to their agent and vice-versa. To make our lives easier, I built AgentCouch, a “messaging app” for agents that
Article URL: https://github.com/drbh/tilery-vm Comments URL: https://news.ycombinator.com/item?id=49111222 Points: 4 # Comments: 0
Google has launched Gemini Robotics ER 2, an embodied-reasoning model built to serve as a robot's high-level planner while a separate model handles the motors. It is available to developers through the Gemini API and Google AI Studio, with private-preview access on the Gemini Enterprise Agent Platform. ER 2 takes in video, images, audio and text, reasons about the scene, plans a task that runs several minutes, then calls something else to move the hardware: a vision-language-action model, a…

Article URL: https://martinfowler.com/articles/exploring-gen-ai/refactoring-economic-benefit.html Comments URL: https://news.ycombinator.com/item?id=49111176 Points: 280 # Comments: 121
<p>We’d like to hear from Black Americans who are using artificial intelligence for health or spiritual guidance</p><p><a href="https://www.theguardian.com/technology/artificialintelligenceai">Artificial intelligence</a> is growing in influence in the <a href="https://www.kff.org/racial-equity-and-health-policy/the-growing-use-of-artificial-intelligence-in-health-care-and-implications-for-disparities/">healthcare</a> space, and more people are seeking guidance from AI tools, including <a href="h

The latest version of Google DeepMind's AI model includes a significant jump into “physical AGI.” But plopping AI into the real world comes with risks.

I listen to a lot of podcasts on the go and wanted a short summary of all the AI shenanigans sent to an RSS feed, so I created thedaily.fm. Here's a short video showing the new MCP support so your agents can do all the lifting. There are summary pods for AI, financial markets, companies like OpenAI and Cloudflare, even local legislation to track the madness in Government. Hope you find it useful! Comments URL: https://news.ycombinator.com/item?id=49111077 Points: 2 # Comments: 1
Gemini Robotics ER 2 helps robots reason, collaborate, and solve real-world tasks. It represents a step change in video understanding, tool orchestration, and multi-robot collaboration for robotic applications.
A new study estimates only 2,000 U.S. engineers have the expertise to deliver meaningful AI ROI, as enterprises race to hire forward-deployed engineers to implement AI at scale.
But why? ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

Optimise how you interact with your coding agents The post How to Organize All of Your Coding Agent Tasks appeared first on Towards Data Science .
Plus, a new policy for the AI protocol ensures features aren't removed suddenly.
Cybersecurity experts told TechCrunch that one of the biggest lessons to be taken from the OpenAI hack against Hugging Face has nothing to do with AI, but traditional cybersecurity defense.
AI image generation comes to Google Earth as Google adds Nano Banana to its mapping platform.

Germany's digital minister has turned OpenAI's containment failure into an argument for building European AI faster. Karsten Wildberger, the federal minister for digital transformation and government modernisation, told Reuters on July 30, 2026 that the episode in which an OpenAI test agent escaped its evaluation sandbox and broke into Hugging Face's production systems strengthens the case for tighter safeguards and for greater European self-sufficiency in AI at the same time. Wildberger said…

In this article, you will learn the seven architectural components that separate a production-grade agentic AI system from a demo script, and how each one...

Claude Design is a research preview under Anthropic Labs, powered by Claude Opus' vision capability, generating interactive prototypes with working navigation, embedded video, voice input, and 3D elements.
