I tested MSI's Windows handheld PC, and it beats the Legion Go in a major way
MSI's Claw 8 EX AI+ is a worthy sequel, with stronger performance, better ergonomics, and highly effective cooling.
MSI's Claw 8 EX AI+ is a worthy sequel, with stronger performance, better ergonomics, and highly effective cooling.
Article URL: https://github.com/pkjaslam/Cambium_AI Comments URL: https://news.ycombinator.com/item?id=48706088 Points: 2 # Comments: 0
Researchers at Princeton University built CEO-Bench, a test where AI agents have to run a fictional software company for 500 simulated days. Most current models go broke, and a simple rule-based heuristic with no AI beats nearly all of them. The article Only three AI models finished above starting capital in a 500-day startup survival test appeared first on The Decoder .

Google's AI offers a lot of convenience in your car, but you're offering up a lot of sensitive information.
The forecast calls for salty skies and pink haze. ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

This is a weekend project I created due to my frustration with the classic terminal applications that have the issue that if I restart my Mac, then my Claude Code instances or my shell instances—are stopped, and I have to manually run them again. Plus, they are not quite optimized for running multiple Claude Code instances, so I created this project to address this issue. Comments URL: https://news.ycombinator.com/item?id=48705880 Points: 2 # Comments: 1
<pre style='white-space:pre-wrap;width:81ex'>fix(agents): preserve structured tool result visible text (#97268) * fix(agents): preserve structured tool result visible text * fix(agents): address structured tool result review blockers * fix(agents): cover real structured tool result shapes * fix(agents): preserve canonical tool error statuses * fix(agents): harden structured result rendering * fix(agents): preserve typed structured results * fix(agents): bound provider result media --------- Co-a
Article URL: https://www.smolhub.com/posts/jetson-nano-super-benchmark-non-reasoning/ Comments URL: https://news.ycombinator.com/item?id=48705815 Points: 2 # Comments: 0
360 founder Zhou Hongyi presents two AI security tools designed to compete with Anthropic's Mythos. One has already flagged 3,432 vulnerabilities. Zhou admits Chinese models trail Western ones by 20 to 30 percent, but compares Mythos to "cyber nuclear weapons" and calls for China to build its own strategic deterrent. The article Chinese cybersecurity firm builds AI tools to rival Mythos and frames the race as cyber-nuclear deterrence appeared first on The Decoder .


<pre style='white-space:pre-wrap;width:81ex'>[AI] fix(plugins): recognize document-extractors as a capability kind… (#91597) * [AI] fix(plugins): recognize document-extractors as a capability kind in inspect-shape PluginCapabilityKind did not include "document-extractors", causing plugins that declare contracts.documentExtractors (like document-extract) to show capabilityCount=0 and shape="non-capability" in plugins inspect. Add "document-extractors" to PluginCapabilityKind and read from plugin.
AGIBOT has rolled out 15,000 wheeled semi-humanoid robots as it moves from embodied AI from development and production to deployment. The post AGIBOT produces 15,000th robot, marking a milestone in embodied AI deployment appeared first on The Robot Report .

A guide to hookups apps with no AI features. Want a break from endless AI? These are the hookup apps for you.

<pre style='white-space:pre-wrap;width:81ex'>fix(cron): propagate cleanupCliLiveSessionOnRunEnd to isolated cron CLI branch (#97227) * fix(cron): propagate cleanupCliLiveSessionOnRunEnd to isolated cron CLI branch * test(cron): add CLI interim retry coverage for isolated cron cleanup flag Verify cleanupCliLiveSessionOnRunEnd is passed on both the initial and retry CLI runs during isolated cron interim-ack retry loops. Proves the inner boundary is safe: each runCliAgent call creates a fresh conte
There has been help with our productivity, but I felt like it's taking the joy of coding away. Like I review AI-composed code, but it doesn't necessarily mean I know what every piece of code is doing there. Do you guys have the same issue? Comments URL: https://news.ycombinator.com/item?id=48705434 Points: 2 # Comments: 1
Article URL: https://bharad.dev/blog/from-transformer-to-llm Comments URL: https://news.ycombinator.com/item?id=48705350 Points: 1 # Comments: 0
Is there another way? ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

Sina Weibo's VibeThinker-3B has just three billion parameters but matches models like DeepSeek V3.2 and Kimi K2.5 on math and coding benchmarks. Those models are up to 333 times larger. The secret isn't size but multi-stage post-training. The researchers propose a hypothesis based on their findings: logical reasoning compresses well into small models, but broad world knowledge does not. The article Sina's open model VibeThinker-3B aims to show reasoning compresses well but factual knowledge does

Be the first to try the #Proton #lumo image generation and give me your feedback #ai #api #image_generation https://carlostkd.ch/api-tracker Comments URL: https://news.ycombinator.com/item?id=48705201 Points: 1 # Comments: 0
https://github.com/temataro/better-graphs I want to teach good Matplotlib taste to agents and humans. This repo contains: 1. Agent instructions + design motifs (Claude Code skills + a CLAUDE.md). 2. An online "blog" tutorial of the same skills, for people. 3. minerva.mplstyle, my opinionated sane matplotlib defaults. Inspired by Tufte's book: The Visual Display of Quantitative Information (a gift from my boss!), plus data-to-viz.com and the python-graph-gallery. Aiming for halfway to the gorgeou
We do have in place proper PRDs and system designs and implementation plans that we feed to the agent. I'm an MLE and SWE, so it's not like I can't read code or don't know any proper practices. What is different though are that Since we are not spending 6-8 hours a day working with the code base unlike pre-coding agents, we lack the indepth knowledge and intuition that helps us narrow down root cause when bugs arise, or know how to best extend the current code base to support a new feature Becau
Article URL: https://www.bloomberg.com/news/articles/2026-06-25/pentagon-sees-broader-role-for-ai-in-setting-military-targets Comments URL: https://news.ycombinator.com/item?id=48704962 Points: 3 # Comments: 0
Article URL: https://chromewebstore.google.com/detail/ai-assistant-for-amazon/ohpekhndmbmkpdoikmphbmdpailacjeo Comments URL: https://news.ycombinator.com/item?id=48704938 Points: 2 # Comments: 0
Meta was asked to use AI tokens more efficiently due to limited Google Gemini AI capacity.

<h2> What are MCP Servers? </h2> <p>The Model Context Protocol (MCP) is an open standard that lets AI agents use external tools through a unified interface. Think of it as USB-C for AI — one protocol connects any AI client (Claude Desktop, Cursor, VS Code with Cline) to any tool or data source.</p> <p>I built three production-ready MCP servers and published them to PyPI and GitHub. Here's what they do and how to use them.</p> <h2> 1. Web Search MCP Server </h2> <p><code>uvx crewai-web-search-mcp
Article URL: https://aileadgenr.com/en Comments URL: https://news.ycombinator.com/item?id=48704870 Points: 2 # Comments: 0
<p>A product page is not a contract.</p> <p>It is a presentation surface.</p> <p>That distinction matters more once AI agents start interacting with commerce systems.</p> <p>Traditional ecommerce platforms can rely on human interpretation. A human can read a product title, inspect images, compare delivery notes, scan a return policy, notice uncertainty, and decide whether to continue. A product page can be visually useful even when the underlying commercial state is incomplete, stale, or spread
Article URL: https://www.ft.com/content/c5d52f72-71ef-40bc-bad3-61afdba8b378 Comments URL: https://news.ycombinator.com/item?id=48704836 Points: 3 # Comments: 0
<p>Most CLAUDE.md files are 500-line monoliths. When you switch LLMs, you rewrite everything. After the third rewrite, I built a three-layer architecture that makes model swaps trivial.</p> <h2> The Problem </h2> <p>I run DeepSeek V4 Pro as my daily driver for Claude Code. But sometimes I need Claude Opus for complex reasoning, or Sonnet for fast iterations.</p> <p>Every time I swapped, I rewrote my entire CLAUDE.md. DeepSeek needs tighter tool-call discipline. Claude Opus needs less output spli
Article URL: https://thehill.com/opinion/technology/5942757-ai-demands-new-social-norms/ Comments URL: https://news.ycombinator.com/item?id=48704759 Points: 2 # Comments: 1
Article URL: https://www.cnbc.com/2026/06/26/oracle-stock-ends-worst-week-since-2001-as-investors-dwell-on-finances.html Comments URL: https://news.ycombinator.com/item?id=48704720 Points: 4 # Comments: 1
<p><em>The Model Context Protocol ecosystem exploded to nearly 20,000 servers. Most are noise. I installed, wired up, and stress-tested 100 of them — mostly inside Claude Code — to find the handful that actually earn a permanent slot in your config. Here are the 12 that survived, the ones I uninstalled, and the uncomfortable 2026 truth nobody selling you MCP servers wants to admit.</em></p> <h2> Why I Went Down This Rabbit Hole </h2> <p>When Anthropic open-sourced the <strong>Model Context Proto
<pre style='white-space:pre-wrap;width:81ex'>test: promote OpenAI HTTP QA coverage (#97369)</pre>
<p>There is a workflow inside your company that everyone quietly works around.</p> <p>Nobody officially owns fixing it.</p> <p>Everyone knows it is painful.</p> <p>New hires learn it through screenshots, Slack threads, and “ask Priya, she knows how this works.”</p> <p>A spreadsheet sits in the middle of it.</p> <p>A manager checks it manually every Friday.</p> <p>A customer probably feels the delay, even if they never see the process.</p> <p>That workflow is not just annoying.</p> <p>It is a tax
Dilip Asbe said that newer UPI apps could be more competitive with a viable commercial model
Article URL: https://warmapper.substack.com/p/how-to-spot-ai-maps Comments URL: https://news.ycombinator.com/item?id=48704490 Points: 3 # Comments: 0
Article URL: https://github.com/Tylersuard/Synapse_neural_net_training_game/tree/main Comments URL: https://news.ycombinator.com/item?id=48704463 Points: 2 # Comments: 1
Article URL: https://github.com/Adirdabush1/cerberus Comments URL: https://news.ycombinator.com/item?id=48704458 Points: 3 # Comments: 0
Article URL: https://github.com/suhasbhairav/ai-chief-of-staff Comments URL: https://news.ycombinator.com/item?id=48704381 Points: 3 # Comments: 0
Article URL: https://code.intellios.ai/photo/ Comments URL: https://news.ycombinator.com/item?id=48704343 Points: 2 # Comments: 1