Show HN: Nightcrawler – A local AI pentesting agent running on a smartphone
Article URL: https://github.com/garagehq/nightcrawler/ Comments URL: https://news.ycombinator.com/item?id=49154127 Points: 62 # Comments: 22
Article URL: https://github.com/garagehq/nightcrawler/ Comments URL: https://news.ycombinator.com/item?id=49154127 Points: 62 # Comments: 22
Chinese tech giant Alibaba released what it says is its largest and "most capable AI model to date," claiming performance rivaling the best systems from US frontier labs Anthropic and OpenAI, as well as domestic rivals like Moonshot AI's Kimi K3. Alibaba said it was making the model, Qwen3.8-Max, widely available to users in a […]

Everyone is talking about the cost of AI. Usually they're talking about GPUs, model licensing, or token consumption. I think they're looking in the wrong place. The biggest cost of enterprise AI may turn out to be the people needed to supervise it. A recent study found employees save about 11 hours a week using AI, but spend more than six hours checking outputs, fixing mistakes, adding missing context and making sure the results are actually usable. Someone coined the term "botsitting" for it…

Two research teams independently solved the same open quantum cryptography problem using OpenAI's GPT-5.6 Sol Ultra, submitting their papers just three hours apart. "If someone mentions an open problem, the first thing is to see if GPT solves it," says one of the researchers. The case raises a question: what does "independent discovery" mean when everyone uses the same models? The article Two teams solved the same quantum crypto problem using GPT-5.6 just three hours apart appeared first on The

Alibaba's new flagship model Qwen3.8-Max is built to handle complex tasks on its own over days at a time, from reproducing research papers to designing chips autonomously. The team plans to release the weights next week. The article Alibaba’s open-weight Qwen3.8-Max takes on long-horizon AI tasks with 2.4 trillion parameters appeared first on The Decoder .
This essay originally appeared in Foreign Policy . Earlier this month, two of OpenAI’s models broke out of their containment sandbox and attacked another AI company. The story is kind of wild . OpenAI was running security tests on two of its models: GPT-5.6 Sol and an unreleased model that is almost certainly GPT-6. In particular, it was running the ExploitGym benchmark, which measures how good a model is at turning security vulnerabilities into working exploits: basically, offensive cyberattack
Find the best robot vacuum deal. Save 50% on the iRobot Roomba Plus 505 at Amazon.

LLMs are theft by 'absolute scum' who make life harder for writers by putting them in copyright peril
Dhivya Nagasubramanian is VP of AI Transformation and Innovation at a major U.S. financial institution, where she leads the design, deployment, and governance of production agentic AI systems. She is the author of Agentic AI for Engineers (Apress/Springer Nature), a practical guide to building autonomous AI systems that can be trusted in production. Since its release, the book has recorded more than 6,000 institutional accesses on SpringerLink, holdings in over 260 libraries worldwide, and…

Anthropic releases Opus 5 promising Fable 5-like capabilities, Google Releases Three New Gemini A.I. Models, and more!
Your next drive-thru order might be taken by a bot. And you might not even notice.

June emerged from stealth today with a $20 million pre-seed round to make AI adoption simpler.
Article URL: https://huggingface.co/zyoralabs/AQ-academic-ai Comments URL: https://news.ycombinator.com/item?id=49153490 Points: 2 # Comments: 0

Article URL: https://ankursethi.com/blog/prevent-cognitive-debt-by-manually-retyping-llm-generated-code/ Comments URL: https://news.ycombinator.com/item?id=49153374 Points: 252 # Comments: 206
Use it to debug and review, says Big Red, but don't submit its output
Learn piano with AI-powered feedback using a Skoove Premium Lifetime Subscription for $149.99 and enjoy 400+ lessons with no recurring subscription.

<p> Hardware for AI coding </p> <p> <a href="https://www.producthunt.com/products/vibe-buddy?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213589?app_id=339">Link</a> </p>
<p> Free, open-source voice dictation and AI assistant, offline </p> <p> <a href="https://www.producthunt.com/products/speakoflow?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213586?app_id=339">Link</a> </p>
MIT Technology Review Explains: Let our writers untangle the complex, messy world of technology to help you understand what’s coming next. You can read more from the series here. When two OpenAI models hacked into the website Hugging Face in July, they weren’t trying to make money or commit sabotage—they were just looking for answers…
'As a big AI skeptic, you just blew my mind'
<!-- SC_OFF --><div class="md"><p>Was just looking at the list of preprints on Arxiv cs.LG <a href="https://arxiv.org/list/cs.LG/recent?skip=0&show=500">https://arxiv.org/list/cs.LG/recent?skip=0&show=500</a></p> <p>Everyday 100 - 400 new machine learning papers gets uploaded on this server.</p> <p>Looking at this unending list of preprints is as if you stepped into a crowded room, like the stock trading floor on wall st. in the 1980s. Everyone is shouting over each other. Nobody is talking to e
RWS Holdings has agreed to buy Acogroup, the French parent of language and content services group Acolad, for an enterprise value of £22.4 million (€26.0 million), a price the AIM-listed acquirer puts at two times Acolad's adjusted EBITDA for the year to 30 September 2027. The cheque is larger than that figure suggests. RWS will pay total consideration of £40.2 million (€46.6 million), which includes payment for roughly £17.8 million of cash sitting in the business at completion. Strip that out…

CrowdStrike tracks 89% surge in machine-assisted activity as patch windows shrink to 48 hours
Reimagine Robotics, founded by former Google DeepMind Applied Robotics leaders, is developing systems that do not require specialized programmers. The post Reimagine Robotics emerges from stealth with robots that ‘learn on the job’ appeared first on The Robot Report .

GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster, more natural conversations.
Presented by NTT DATA AIVista At VB Transform 2026 , NTT DATA AIVista CEO Bratin Saha joined VentureBeat CEO and editor-in-chief Matt Marshall to discuss the last-mile challenge of operationalizing frontier models in regulated production, where reliability, context, guardrails, and security determine whether AI delivers enterprise value. The conversation centered around the question facing every enterprise now pouring money into AI: how to convert that spending into real, tangible value. "It's n

Turn plain-English prompts into instant, full-stack SEO dashboards with the new Moz API MCP server. The Moz API Model Context Protocol (MCP) server allows you to connect your favorite AI assistant directly to our SEO data endpoints using standard open protocols. Get started with the Moz MCP in a few minutes.
Three high-severity security flaws have been disclosed in Hugging Face's Diffusers library that could allow crafted model repositories to stealthily execute arbitrary code on machines that load it, opening the artificial intelligence (AI) supply chain to security risk. "These vulnerabilities are bypassing trust_remote_code, the safeguard designed to stop unreviewed code from running in the

Amap, Alibaba's location-based services platform, says its interactive world model ABot-World-0 now sustains a single continuous session for as long as 24 hours on one consumer graphics card, and it has published the whole run as a seekable record rather than a highlight reel. The page lets anyone jump to any second of the day-long rollout, with fixed entry points at the six-, twelve- and eighteen-hour marks, alongside five more complete runs through grassland, desert, city and snowfield…

The problem with over-simplification of AI Agents’ interfaces. Continue reading on UX Collective »

<pre style='white-space:pre-wrap;width:81ex'>fix(openai): support the current speech model snapshot (#118475) * fix(openai): support the current speech model snapshot * docs(openai): clarify speech instructions apply to the model family</pre>
<pre style='white-space:pre-wrap;width:81ex'>perf(system-agent): reuse setup metadata per activation (#118352) * perf(system-agent): reuse setup metadata per activation * chore: remove release-owned changelog entry --------- Co-authored-by: Peter Steinberger <steipete@mac-studio-sf2.local></pre>
<pre style='white-space:pre-wrap;width:81ex'>refactor: extract selected agent retirement</pre>
A new strategy could finally illuminate dark matter. ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

<p>Two pricing stories dropped this week that are worth a pause if you're building on LLM APIs.</p> <h2> DeepSeek V4 Flash exits preview — and undercuts its own flagship </h2> <p>DeepSeek V4 Flash left preview at <strong>$0.14 / $0.28 per million tokens</strong> (input/output) — and it's beating its own larger Pro model on agentic benchmarks, hitting <strong>82.7% on Terminal-Bench</strong>. That's a smaller, cheaper model outperforming its own bigger sibling on agent tasks.</p> <h2> Claude Sonn
<p>There is a French phrase for it. L'heure entre chien et loup, the hour between dog and wolf. It is that stretch after the sun goes down, when the light has gone soft and a shape coming toward you over the far hill could be your own dog heading home or a wolf that has been watching you. You cannot quite tell. The light is vague enough that, for a little while, the thing you love and the thing you fear wear the same outline.</p> <p>I find myself thinking about it a lot these days. Only what has
<p><strong>Release:</strong> <a href="https://github.com/simonw/condense-json/releases/tag/1.1">condense-json 1.1</a></p> <p>After shipping <a href="https://simonwillison.net/2026/Aug/2/condense-json/">condense-json 1.0</a> I started integrating it into LLM, and found there were some desirable new features already:</p> <blockquote> <ul> <li>Replacements object can now include values other than strings. These will be identified and used as structural replacements by <code>condense_json()</code> a
Article URL: https://makethisbetter.dev/ Comments URL: https://news.ycombinator.com/item?id=49151121 Points: 2 # Comments: 1
Article URL: https://booyaka101.github.io/thedailyfable/day07/ Comments URL: https://news.ycombinator.com/item?id=49151084 Points: 4 # Comments: 0