Narwal Flow 2 review: The robot vacuum that finally perfected obstacle avoidance
The Narwal Flow 2 promises to be one of the best robot vacuum cleaners for obstacle avoidance and mopping - and it is.
The Narwal Flow 2 promises to be one of the best robot vacuum cleaners for obstacle avoidance and mopping - and it is.
OpenAI will spend the equivalent of Sweden's GDP on infrastructure through 2030.
Hi HN, We’re Adeel and Umair, co-founders of Unlayer ( https://unlayer.com/ ). We let you add content creation to your applications without having to build an entire editor, renderer, template, and export stack yourself. Unlayer lets you create emails, web pages, and documents inside your app, in three different ways: in code, visually, or with AI. Here’s a demo: https://www.youtube.com/watch?v=0HsDtNkdMpM . We started with an embeddable email editor because a lot of products eventually need one
Samsung Unpacked was as much about software as it was hardware.
AI Teammates are agentic AI on Amazon Bedrock, and few engineering organizations run them in production at the scale that monday.com does. Nine in ten Builders use AI coding tools every month, up from roughly half a year ago. Per-engineer PR throughput is up by more than half. Every figure in this post comes from monday’s own internal production data. In this post, we share the architecture behind those numbers, the retrofits that made it work in a decade-old code base, and the confidence-scored
OpenAI is planning a data center in Georgia called "Project Camellia" with a 3.2-gigawatt power deal from Georgia Power. The company pledged $80 million for the local community and $71 million in Codex credits for students to counter growing opposition to US data centers that many residents see as resource-hungry but job-poor. The article OpenAI's "Project Camellia" in Georgia secures a massive 3.2-gigawatt power deal through 2032 appeared first on The Decoder .
Article URL: https://cameronmpalmer.medium.com/should-you-even-use-an-llm-b4f3b7914f4d Comments URL: https://news.ycombinator.com/item?id=49008624 Points: 9 # Comments: 5
I'm the author of a paper my friends and I wrote after we were curious if a MUD, text games originating in the 1970s, could be used to evaluate LLMs. We've spent the last several months on nights and weekends running this experiment and writing the paper on just our personal computers with about $99 in API credits. Our experiment did have an interesting leaderboard but even more surprising was the measurements of each LLM. We scored each on four behavioral dimensions, two of which lean heavily o
In this article, we try to explore the collective thinking into a smaller set of practices and explain the reasoning behind each one, rather than asking anyone to memorize a numbered list.

Over the past few months, our team has been building more and more slidedecks using web frontend technologies with coding harnesses like Claude Code, but a common complaint is to make even small edits we need to edit the code either manually or via the harness. To avoid this loop, I ended up creating Bento, a single HTML file with everything you need in a slide tool including animations and shared editing. There's no install or cloud login, everything works offline. The default deck is around 56
Here's what people are getting wrong about the so-called SaaS apocalypse.
<!-- SC_OFF --><div class="md"><p>- Tutorial: <a href="https://ordinaryintelligence.substack.com/p/how-to-build-an-ai-slop-detector">https://ordinaryintelligence.substack.com/p/how-to-build-an-ai-slop-detector</a></p> <p>- Notebook on GitHub: <a href="https://github.com/Buzzpy/Python-Projects/blob/main/AI-slop-detector.ipynb">https://github.com/Buzzpy/Python-Projects/blob/main/AI-slop-detector.ipynb</a></p> </div><!-- SC_ON --> submitted by <a href="https://www.reddit.com/user/gamedev-exe"> /u/g
The latest addition to Philips' Sonicare line of smart electric toothbrushes could take the guesswork out of your brushing routine. The Next-Generation DiamondClean 9900 Prestige uses a third-generation AI model to provide real-time feedback on how you're brushing through a glowing ring on its base, while a touchscreen highlights areas of your mouth that may […]

Article URL: https://www.abetterinternet.org/post/vulns-and-llms/ Comments URL: https://news.ycombinator.com/item?id=49007888 Points: 2 # Comments: 0
<p>The companies hope to follow in the footsteps of SpaceX, which raised $86bn and soared to a $2.1tn valuation after it listed on public markets in June</p><ul><li><p>Get our <a href="https://www.theguardian.com/email-newsletters?CMP=cvau_sfl">breaking news email</a>, <a href="https://app.adjust.com/w4u7jx3">free app</a> or <a href="https://www.theguardian.com/australia-news/series/full-story?CMP=cvau_sfl">daily news podcast</a></p></li></ul><p>Top US AI developers Anthropic and OpenAI cheered

If you have ever wanted to actually build an LLM inference runtime yourself — pack your own weights, own every barrier, capture your own CUDA graphs — this is what that journey looks like on an H100. A step-by-step tour of a small runtime called annotated-llm-runtime, and the three bugs that produced most of the annotations. The post How To Build Your Own LLM Runtime From Scratch appeared first on Towards Data Science .
OpenAI has announced Presence , a new enterprise product for deploying and managing AI agents across customer-facing and internal business workflows. The offering is designed for eligible enterprise customers that want agents to answer questions, access company systems, take approved actions and escalate to human workers while operating under company-defined policies, permissions and evaluation standards. Presence is available immediately through a limited general availability program. OpenAI Fo

TL;DR In April 2025, the PyTorch Foundation evolved into a multi-project Foundation, with the objective to support deeper collaboration across domains and help scale innovation throughout the AI lifecycle. Today,...

AMD says it's going to invest up to $5 billion in Anthropic, while helping to expand the AI company's computing power, according to an announcement on Wednesday. As part of the new partnership, Anthropic will deploy up to 2 gigawatts of AMD's Instinct MI450 AI GPUs using the chipmaker's new Helios rack-scale system, as reported […]

Govern natively, federate outward, and what breaks across trust domains By now the agent has its own identity and you can carry that identity through a chain of calls. The next question is where the rules live. Who decides what an agent is allowed to do, and where does that decision get made? Two answers,... The post Govern natively, federate outward, and what breaks across trust domains appeared first on DataRobot .

<p> QA agent that tests apps the way you'd explain them </p> <p> <a href="https://www.producthunt.com/products/swiftscale-software?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1203618?app_id=339">Link</a> </p>
https://book-prize-index.vercel.app/ Comments URL: https://news.ycombinator.com/item?id=49007247 Points: 480 # Comments: 234
Samsung announced two new Galaxy smart watches at Unpacked, with some modest hardware upgrades and AI health features.

As of July 22, save 36% on the Eufy X10 Pro Omni robot vacuum and mop.

Future AI risks will be harder to address ad hoc.
Google and Kaggle's 5-Day AI agents course is now freely available to everyone.

xAI's sidebar agent promises analysis and financial models – for a price
Anthropic leaped to a $47 billion revenue run rate by May, compared to $9 billion in 2025. It’s the kind of growth that Menlo Ventures’ Matt Murphy says he’s never seen in 25 years of investing, not in the internet wave, not in mobile, not in the first cloud boom. Menlo led Anthropic’s $500M Series D, and Murphy has had a front-row seat as the company went from a pre-revenue, […]
<p> Talk to any app on your Mac </p> <p> <a href="https://www.producthunt.com/products/wisprkey-mac-ai-assistant?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1203598?app_id=339">Link</a> </p>
<p><em>The model did exactly what it was told. The software around the model uploaded the whole repository anyway.</em></p> <p>A researcher put a coding agent to the simplest possible test. They created a file called <code>never_read_canary.txt</code>, dropped a unique marker inside it (<code>CANARY-XR47P2-NEVERREAD-UNIQUE</code>), and gave the agent an instruction that left no room for interpretation: "Reply exactly OK, do not read any files." The agent replied OK. It obeyed. It did not read th
A lot of ink is being spilled about the effects AI is having on learning and education. The drawbacks are obvious. For many classroom assignments, AI is now good enough that it can simply do your homework for you. But having AI do your homework is self-defeating. You can’t learn without doing the work. That […] The post We Need Guidelines for Learning and AI appeared first on Scott H Young .

<pre style='white-space:pre-wrap;width:81ex'>Export installed agents as grouped Claw packages (#102306) * Export installed agents as grouped Claw packages * test(claws): cover exact agent export * docs(claws): document agent export * fix(claws): use current bounded file reader --------- Co-authored-by: Patrick Erichsen <patrick.a.erichsen@gmail.com></pre>
Google commits $40M in AI tokens and credits for the Genesis Mission
<p><a href="https://x.com/paulg/status/2079877240699940955" rel="noopener noreferrer">Paul Graham posted a test</a> this week that has nothing to do with sentence structure. Slop gives itself away, he said, when the diction doesn't match the idea, when something completely ordinary gets delivered with the excitement of someone announcing a discovery.</p> <p>That's a different axis than everything else I've been reading about detection this month. Pangram scores a pattern: token by token, sentenc
<blockquote> <p><strong>TL;DR</strong>: an LLM that calls tools is a client you cannot trust. And it holds your production credentials. The most important rule fits in one sentence. Your server decides who the user is, never the model. In this article, I show how I secured a real LLM agent in Go, in production. Everything also applies to MCP servers.</p> </blockquote> <p>This article is for Go developers who put an LLM agent or an MCP server in production. Not in a demo.</p> <h2> Everyone builds
A hands-on walkthrough of code execution with the OpenAI Agents SDK and Docker The post Build an LLM Agent That Can Write and Run Code appeared first on Towards Data Science .
<p>Bloomsbury has 14,087 titles listed within settlement between AI startup and authors over use of protected work</p><p>The publisher of Harry Potter has received a multimillion-pound payout as a beneficiary of a $1.5bn (£1.12bn) copyright settlement between the AI startup Anthropic and thousands of authors over the use of their protected work to power chatbots.</p><p>Bloomsbury, which is home to the bestselling novelists Sarah J Maas and Susanna Clarke as well as JK Rowling, said it had 14,087

<p> Starts a fresh Codex worker and critic every cycle </p> <p> <a href="https://www.producthunt.com/products/agentloop-2?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1203549?app_id=339">Link</a> </p>
As AI models become commoditized, maybe there's margin in the plumbing
<pre style='white-space:pre-wrap;width:81ex'>fix(ai): honor provider reasoning effort maps (#112632)</pre>